AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,678
  • Agents7,513
  • Applications5,367
  • Concepts5
  • Hardware1,821
  • Industry6,154
  • Local Ai4,902
  • Model Releases23,619
  • Research19,969
  • Safety13,271
  • Syntheses17
  • Tools1,674
  • Tutorials3,366

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,678
  • Agents7,513
  • Applications5,367
  • Concepts5
  • Hardware1,821
  • Industry6,154
  • Local Ai4,902
  • Model Releases23,619
  • Research19,969
  • Safety13,271
  • Syntheses17
  • Tools1,674
  • Tutorials3,366

Source
HumanDGX agent

Content type
87,678Total entries
1Added by human
87,677Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
51,512 results
Model Releases

AutoRefine: Compiling Trajectories into Validated Typed Agent Artifacts

DGX agent

arXiv:2601.22758v2 Announce Type: replace Abstract: Large language model agents repeatedly encounter related tasks, yet systems that learn from trajectories commit every lesson to one predefined artif

model-releasesarxiv-cs-ai
11 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Avalon-ToM-Bench: Evaluating Fine-Grained Theory of Mind via Asymmetric Game Mechanics

DGX agent

arXiv:2608.09638v1 Announce Type: new Abstract: Theory of Mind (ToM) is essential for agent interactions, yet existing evaluations either rely on static scenarios that oversimplify mental-state reason

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Back to the Future: A workbook time machine for spread sheet creation benchmarks

DGX agent

arXiv:2608.07873v1 Announce Type: new Abstract: We introduce the workbook time machine, a pipeline that automatically creates benchmarks evaluating the ability of language models to create derived obj

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Beyond Naturalness: Probing Automated Text-To-Speech Evaluators on Linguistically Grounded Dimensions

DGX agent

arXiv:2608.09930v1 Announce Type: cross Abstract: Automated Text-to-Speech (TTS) evaluation methods (Mean Opinion Score (MOS) predictors and Audio Large Language Models (Audio-LLM) judges) are expecte

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Can Graph Learning Learn Circuits?

DGX agent

arXiv:2608.08536v1 Announce Type: new Abstract: Circuit localization is a mechanistic interpretability task whose goal is to identify a sparse subgraph of a transformer's computation graph sufficient

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

CAP: A Scalable Benchmark for Evaluating Cross-Site Browser Agents with Complex Actions and Perception

DGX agent

arXiv:2608.08392v1 Announce Type: new Abstract: Large language models are increasingly deployed as autonomous agents that interact with the web through browsers. While recent progress has been driven

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Circuit Fine-Tuning for Compute-Efficient Transformer Adaptation

DGX agent

arXiv:2608.08336v1 Announce Type: new Abstract: Parameter-Efficient Fine-Tuning (PEFT) has become the de facto standard for adapting Vision Transformers (ViTs) to downstream tasks. While parameter cou

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

Control-Diverse Reinforcement Fine-Tuning: Decoupling the Shared Control Bottleneck of RL Post-Training

DGX agent

arXiv:2608.08224v1 Announce Type: new Abstract: Reinforcement learning post-training unlocks complex reasoning in LLMs. Yet benchmark scores reveal only whether a model improved, not what changed insi

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

Coupled Graph--Policy Distillation for Personalized Medication Safety in Older Adults with Multimorbidity

DGX agent

arXiv:2608.09443v1 Announce Type: new Abstract: Large language model (LLM) agents can support medication review between clinical visits, but safe choices for older adults with multimorbidity depend on

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

CresOWLve: Benchmarking Creative Problem-Solving Over Real-World Knowledge

DGX agent

arXiv:2604.03374v2 Announce Type: replace-cross Abstract: Creative problem-solving requires combining multiple cognitive abilities, including logical reasoning, lateral thinking, analogy-making, and c

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Decentralized Nonconvex Composite Federated Learning with Gradient Tracking and Momentum

DGX agent

arXiv:2504.12742v2 Announce Type: replace Abstract: Decentralized Federated Learning (DFL) enables collaborative model training without relying on a central server. When local objectives are nonconvex

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

Depth-Aware Implicit Neural Representation Priors for 3D Gravity Inversion

DGX agent

arXiv:2608.08959v1 Announce Type: new Abstract: Gravimetry images subsurface density contrasts associated with geological structures, geothermal systems, and intrusive bodies. Recovering a three-dimen

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Does a Toehold Make a Bidder Bolder? Preemption and Multiplicity in Multi-Round Takeover Auctions

DGX agent

arXiv:2608.08407v1 Announce Type: cross Abstract: A bidder can quietly buy a stake in a company before making an offer for it. That stake, a toehold, is supposed to pay for itself twice: it makes the

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Efficient Human-Contact Representation for Human-Scene Interaction

DGX agent

arXiv:2608.09388v1 Announce Type: new Abstract: Human-scene interaction is an active research topic with several industrial applications in virtual reality, gaming, robotics, and surveillance. Despite

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

EvoTrustRAG: Evolution-Aware Conflict Attribution and Evidence Handling for Reliable Retrieval-Augmented Generation

DGX agent

arXiv:2608.07933v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) improves the factuality of large language models with external knowledge, yet conflicting evidence remains a fund

model-releasesarxiv-cs-cl
11 Aug 2026
Model Releases

Expert-Guided Multimodal Fusion for Unified Emotion and Sentiment Analysis

DGX agent

arXiv:2601.07565v2 Announce Type: replace-cross Abstract: Multimodal emotion understanding requires the integration of heterogeneous data sources, including text, audio, and visual modalities, while s

model-releasesarxiv-cs-ai
11 Aug 2026
Research

From Chains to DAGs: Probing the Graph Structure of Reasoning in LLMs

DGX agent

arXiv:2601.17593v3 Announce Type: replace Abstract: Recent progress in large language models has renewed interest in how multi-step reasoning is represented internally. While prior work often treats r

researcharxiv-cs-cl
11 Aug 2026
Model Releases

Governing the KV Cache: Preventing Timing Side-Channel Leakage in Multi-Tenant LLM Inference

DGX agent

arXiv:2608.09225v1 Announce Type: cross Abstract: The key-value (KV) cache is the primary throughput optimization in modern large language model (LLM) inference, enabling prefix reuse across requests.

model-releasesarxiv-cs-ai
11 Aug 2026
Research

High-Quality Exposure Correction with Diffusion-Based Image Generation Priors

DGX agent

arXiv:2608.08720v1 Announce Type: new Abstract: Although most existing exposure correction methods achieve high fidelity, they often place excessive focus on overall pixel-wise accuracy, making it cha

researcharxiv-cs-cv
11 Aug 2026
Research

HindsightBench: A Black-Box Behavioral Audit Protocol for Parametric Hindsight in Time-Indexed LLM Decision Tasks

DGX agent

arXiv:2607.18867v2 Announce Type: replace-cross Abstract: Large language models leak parametric knowledge of what followed a historical date into decision tasks indexed by that date -- not necessarily

researcharxiv-cs-cl
11 Aug 2026
Model Releases

Illusion or Integrity? Geometrical Consistency Metric for AIGC Video Quality Evaluation

DGX agent

arXiv:2608.09594v1 Announce Type: cross Abstract: Recently, AI-driven video generation has attracted considerable attention. This surge increases the demand for reliable video quality assessment (VQA)

model-releasesarxiv-cs-ai
11 Aug 2026
Local Ai

Integrated Multimodal AI System for Retrieval-Augmented Reasoning, Object Sensing, and Damage Analysis

DGX agent

arXiv:2608.08935v1 Announce Type: new Abstract: This work presents a unified multimodal AI system for damage assessment that integrates retrieval-augmented generation (RAG) models, thermal spectrum pe

local-aiarxiv-cs-ai
11 Aug 2026
Model Releases

Jako Tako or Fluent? Presenting PoVisLE: A Polish Vision-Language Evaluation

DGX agent

arXiv:2608.07763v1 Announce Type: new Abstract: Vision-language models (VLMs) have achieved strong performance on tasks such as image captioning, visual question answering, and image-to-text generatio

model-releasesarxiv-cs-cl
11 Aug 2026
Agents

JaleesBench: Are AI Assistants Good Spiritual Company?

DGX agent

arXiv:2608.07508v1 Announce Type: cross Abstract: Large language models are already advisors to millions of people of faith who bring them real decisions. The pressing question for a person of faith i

agentsarxiv-cs-ai
11 Aug 2026
Model Releases

KGCaRe: Explainable Complex Conditional Question Answering using Automatic Knowledge Graph Construction and Context Retrieval with LLMs

DGX agent

arXiv:2608.09779v1 Announce Type: cross Abstract: Answering complex conditional questions using Large Language Models (LLMs) and Retrieval-Augmented Generation (RAG) remains a challenge, particularly

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Locating Failure in Multi-Page Visually Rich Document Understanding: An Empirical Attribution

DGX agent

arXiv:2608.07943v1 Announce Type: new Abstract: Multi-page visually-rich document understanding (MP-VRDU) requires managing evidence that is sparse, spread across pages, and often exceeds a model's co

model-releasesarxiv-cs-ai
11 Aug 2026
Local Ai

Macaron-V1: Towards Open Continual Learning with Self-Improvement and Mixture-of-LoRA

DGX agent

arXiv:2608.09819v1 Announce Type: cross Abstract: Macaron-V1 is an open agent-model family for experiential intelligence: learning from experience in real environments and continuing to learn after de

local-aiarxiv-cs-cl
11 Aug 2026
Model Releases

Mawqif-v2: An Arabic Benchmark Dataset for Cross-Target Stance Detection

DGX agent

arXiv:2608.09539v1 Announce Type: new Abstract: Publicly available Arabic datasets for target-specific stance detection remain limited, particularly for evaluating cross-target generalization. This pa

model-releasesarxiv-cs-cl
11 Aug 2026
Model Releases

Measuring the Wrong Thing: Internal Harmfulness Scores Anti-Rank Successful Jailbreaks

DGX agent

arXiv:2608.09624v1 Announce Type: cross Abstract: Internal safety scores judge a prompt before any text is generated, and they are validated by how well they separate harmful prompts from benign ones.

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

MoRSE: Task-Oriented Multi-Agent System with Mixture of Role-Subtask Experts

DGX agent

arXiv:2608.09251v1 Announce Type: cross Abstract: Large language model-based multi-agent systems have recently shown strong potential for complex, long-horizon tasks. However, existing methods mainly

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

MOSAIC: Adversarial Co-evolution of Specialist Heuristics and Problem Instances for LLM-based Automated Heuristic Design

DGX agent

arXiv:2608.07544v1 Announce Type: cross Abstract: Automated heuristic design (AHD) with large language models (LLMs) has produced strong heuristics for combinatorial optimization problems (COPs). Yet

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Multi-Agent Reinforcement Learning via Agent-Specific Preference

DGX agent

arXiv:2608.08604v1 Announce Type: new Abstract: Multi-agent reinforcement learning (MARL) is a powerful framework for solving complex collaborative tasks, but it relies heavily on well-defined global

model-releasesarxiv-cs-lg
11 Aug 2026
Safety

Neural Message Passing on Structural Interaction Graphs for Fully-Inductive Graph Neural Networks

DGX agent

arXiv:2608.08567v1 Announce Type: new Abstract: A central obstacle in building graph foundation models is the input heterogeneity in terms of feature space dimensionality, semantics, and structure. Su

safetyarxiv-cs-lg
11 Aug 2026
Model Releases

Neural Operators for Immersed-Boundary Soft Swimmers Locomotion

DGX agent

arXiv:2608.07722v1 Announce Type: new Abstract: High-fidelity immersed-boundary simulation resolves the coupled motion of a deforming swimmer and its surrounding flow, but the resulting cost limits re

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

Not All Visual Tokens Are Equally Safe to Remove:Consequence-Sensitive Visual Token Compression

DGX agent

arXiv:2608.09176v1 Announce Type: cross Abstract: Visual token compression for vision--language models (VLMs) has largely relied on criteria such as attention, redundancy, and uncertainty to maximize

model-releasesarxiv-cs-ai
11 Aug 2026
Safety

Perception Before Supervision: Self-Contained Visual Distillation from Counterfactual Blind Spots

DGX agent

arXiv:2608.09931v1 Announce Type: new Abstract: Self-improvement for multimodal large language models (MLLMs) is typically driven by reward-based methods that provide only coarse scalar feedback. Dist

safetyarxiv-cs-cv
11 Aug 2026
Safety

Position Bias in Ordinal Classification: A Systematic Evaluation

DGX agent

arXiv:2608.08869v1 Announce Type: new Abstract: Large language models are increasingly used for ordinal classification, yet semantically equivalent changes to prompt organization can alter their predi

safetyarxiv-cs-cl
11 Aug 2026
Local Ai

Position: Certifiable State Integrity Should Be Built from Local Validity, Not Global Scale

DGX agent

arXiv:2601.21249v2 Announce Type: replace Abstract: Breakthroughs in language and vision have motivated increasingly general foundation models for time series and physical dynamics, where evidence is

local-aiarxiv-cs-ai
11 Aug 2026
Agents

Preference Redirection via Attention Concentration: An Attack on Computer Use Agents

DGX agent

arXiv:2604.08005v2 Announce Type: replace Abstract: Advancements in multimodal foundation models have enabled the development of Computer Use Agents (CUAs) capable of autonomously interacting with GUI

agentsarxiv-cs-lg
11 Aug 2026
Model Releases

QuArch: A Benchmark for Evaluating LLM Reasoning in Computer Architecture

DGX agent

arXiv:2510.22087v3 Announce Type: replace-cross Abstract: The field of computer architecture, which bridges high-level software abstractions and low-level hardware implementations, remains absent from

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

RA-FinBERT: Rule-aware LoRA adaptation for low-resource financial sentiment classification

DGX agent

arXiv:2608.09834v1 Announce Type: new Abstract: Financial sentiment analysis converts unstructured financial news into quantitative signals that can support market analysis and decision-making. Existi

model-releasesarxiv-cs-cl
11 Aug 2026
Model Releases

Readout-Rank Laws for Isotropic Quantum Tangents

DGX agent

arXiv:2608.07628v1 Announce Type: cross Abstract: Deep parameterized quantum circuits may remain sensitive to a parameter change while the observables retained by a learning model barely respond. We s

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

RealDenseFace: Real-time Monocular 3D Face Reconstruction from Dense UV-space Priors

DGX agent

arXiv:2608.09238v1 Announce Type: new Abstract: Recent monocular 3D face reconstruction methods achieve high fidelity by fitting a 3D Morphable Model (3DMM) to dense priors predicted by networks, but

model-releasesarxiv-cs-cv
11 Aug 2026
Safety

Scalable extensions to given-data Sobol' index estimators

DGX agent

arXiv:2509.09078v3 Announce Type: replace-cross Abstract: Given-data methods for variance-based sensitivity analysis have significantly advanced the feasibility of Sobol' index computation for computa

safetyarxiv-cs-lg
11 Aug 2026
Model Releases

Sci-VBench: Evaluating Knowledge- and Reasoning-Intensive Video Generation in Science Domains

DGX agent

arXiv:2608.09873v1 Announce Type: cross Abstract: We introduce Sci-VBench, a comprehensive benchmark for evaluating knowledge- and reasoning-intensive video generation across scientific domains. It co

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

SciTaRC: A Plan-Annotated Scientific Tabular QA Benchmark for Language Reasoning and Complex Computation

DGX agent

arXiv:2603.08910v2 Announce Type: replace Abstract: We introduce SciTaRC, an expert-authored benchmark for question answering over scientific tables that targets composite, multi-step reasoning. To en

model-releasesarxiv-cs-cl
11 Aug 2026
Model Releases

Shape Mutating Expert Compression:LorExperts and BTExperts

DGX agent

arXiv:2608.07814v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) language models deliver high capacity at low per-token compute, but deploying them cheaply requires compressing their many ex

model-releasesarxiv-cs-ai
11 Aug 2026
Tutorials

SignLlama: Enhancing Gloss-free Sign Language Translation by Prioritizing Visual Features for LLMs

DGX agent

arXiv:2608.09006v1 Announce Type: cross Abstract: Large Language Models (LLMs) have achieved remarkable success across a wide range of tasks. However, fine-tuning LLMs for Gloss-Free Sign Language Tra

tutorialsarxiv-cs-ai
11 Aug 2026
← Previous
1…391392393394395…1074
Next →