AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,457
  • Agents7,399
  • Applications5,302
  • Concepts5
  • Hardware1,786
  • Industry6,117
  • Local Ai4,835
  • Model Releases23,193
  • Research19,715
  • Safety13,094
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,457
  • Agents7,399
  • Applications5,302
  • Concepts5
  • Hardware1,786
  • Industry6,117
  • Local Ai4,835
  • Model Releases23,193
  • Research19,715
  • Safety13,094
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
86,457Total entries
1Added by human
86,456Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
50,764 results
Model Releases

BDH-CQ: In-Context Learning with Recurrent Latent Reasoning

DGX agent

arXiv:2608.09888v1 Announce Type: cross Abstract: We introduce BDH-CQ, a reasoning model that combines in-context learning with recurrent latent reasoning. Inputs presented at inference time continuou

model-releasesarxiv-cs-ai
11 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Build it, Break it, Repeat: Benchmarking and improving LLM-manipulated disinformation detection in social media posts

DGX agent

arXiv:2608.09510v1 Announce Type: cross Abstract: Detecting machine-generated disinformation on social media is increasingly difficult as large language models (LLMs) make it easier to generate and re

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Can LLM Agents Stick to the Script? A Benchmark for Long-Horizon Consistency in Interactive Narratives

DGX agent

arXiv:2608.08160v1 Announce Type: cross Abstract: The rapid advancement of Large Language Models (LLMs) is revolutionizing AI for Games by enabling open-ended and fluid interactive storytelling. Howev

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

DarwinX: Evolving Agent Harnesses Through Natural Selection

DGX agent

arXiv:2608.07545v1 Announce Type: cross Abstract: An LLM agent's capability depends not only on model weights but on its harness: prompts, tools, skills, and control flow. Self-improvement loops alrea

model-releasesarxiv-cs-ai
11 Aug 2026
Research

Different Feedback, Different Updates: Selective Self-Learning from User Interactions for Large Language Models

DGX agent

arXiv:2608.09109v1 Announce Type: new Abstract: User feedback offers natural supervision for persistent LLM improvement, but a single message may support multiple behavioral changes with different sco

researcharxiv-cs-ai
11 Aug 2026
Research

DiffSafeMerge: Mitigating Backdoor Inheritance in Diffusion Model Merging

DGX agent

arXiv:2608.09445v1 Announce Type: cross Abstract: Unconditional diffusion checkpoint merging assumes benign sources, yet a compromised public checkpoint can transfer a dormant backdoor while clean gen

researcharxiv-cs-cv
11 Aug 2026
Safety

From Fake to Real: Pretraining on Balanced Synthetic Images to Prevent Spurious Correlations in Image Recognition

DGX agent

arXiv:2308.04553v4 Announce Type: replace Abstract: Visual recognition models are prone to learning spurious correlations induced by a biased training set where certain conditions B (eg, Indoors) are

safetyarxiv-cs-cv
11 Aug 2026
Local Ai

Hallucination-Free GUI Grounding via Regression-Free Layout-Aware Matching

DGX agent

arXiv:2608.09654v1 Announce Type: new Abstract: GUI agents are shifting from metadata-dependent large language models to purely visual multimodal large language models (MLLMs) that operate directly on

local-aiarxiv-cs-ai
11 Aug 2026
Research

How Far Do Foundation Models Transfer to Infant Signals? A Cross-Dataset Transfer Audit with a Unified Need Ontology

DGX agent

arXiv:2608.08989v1 Announce Type: cross Abstract: Public infant cry corpora are small, label-incompatible, and almost always evaluated one corpus at a time. We ask what this practice hides and what fi

researcharxiv-cs-ai
11 Aug 2026
Agents

IntelliAudit: Using Large Language Models to Evaluate Audit Controls

DGX agent

arXiv:2608.07688v1 Announce Type: new Abstract: IT audits require auditors to judge whether heterogeneous organizational evidence satisfies semantic security and compliance controls. This judgment is

agentsarxiv-cs-ai
11 Aug 2026
Research

LHSDet: High-Resolution AI-Generated Image Detection via Visual Question Answering

DGX agent

arXiv:2608.07863v1 Announce Type: new Abstract: Driven by advances in diffusion models and autoregressive models, the fidelity and resolution of AI-generated images now rival those of real images. How

researcharxiv-cs-cv
11 Aug 2026
Model Releases

LIRA: Local Cross-Layer Information Routing for Vision-Language-Action Decoding

DGX agent

arXiv:2608.07596v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models transform representations from pretrained vision-language models (VLMs) into robot actions, yet the interface that

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

Long SKILL Compliance as Logical Reasoning: Closure-Grounded Detection with Scaling-Guided On-Policy Distillation

DGX agent

arXiv:2608.08146v1 Announce Type: new Abstract: The increasing complexity of enterprise business scenarios has promoted the widespread adoption of long SKILL documents in agent systems, posing new cha

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

MCIF: Multimodal Crosslingual Instruction-Following Benchmark from Scientific Talks

DGX agent

arXiv:2507.19634v4 Announce Type: replace-cross Abstract: Recent advances in large language models have laid the foundation for multimodal LLMs (MLLMs), which unify text, speech, and vision within a s

model-releasesarxiv-cs-ai
11 Aug 2026
Research

Rethinking Reasoning with MDLMs: Early Exits, Post-hoc Reasoning, and Beyond

DGX agent

arXiv:2510.19990v2 Announce Type: replace Abstract: The reasoning paradigm, where language models reason before answering, has enabled breakthroughs on tasks such as mathematical problem-solving. Whil

researcharxiv-cs-lg
11 Aug 2026
Model Releases

UNMASK: Discovering and Causally Verifying Spurious Shortcuts in Text Classifiers

DGX agent

arXiv:2608.09209v1 Announce Type: new Abstract: Neural language models trained on large crowdsourced corpora frequently exploit spurious surface patterns tied to target labels without true linguistic

model-releasesarxiv-cs-cl
11 Aug 2026
Safety

VANE: Reliable Test-Time Training for Vision-Language-Action Models via Future Visual Representation Prediction

DGX agent

arXiv:2608.09448v1 Announce Type: cross Abstract: Test-time training (TTT) offers a lightweight way to adapt vision--language--action (VLA) policies from unlabeled deployment streams, but it remains d

safetyarxiv-cs-cv
11 Aug 2026
Model Releases

VeinCast: Physics-Guided Dynamic Field Graphs with Graph-Conditioned Fusion for Global Medium-Range Weather Forecasting

DGX agent

arXiv:2608.09286v1 Announce Type: cross Abstract: Global medium-range weather forecasting requires modeling structured yet state-dependent interactions among heterogeneous atmospheric fields. Existing

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

When Grammar Guides the Attack: Uncovering Control-Plane Vulnerabilities in LLMs with Structured Output

DGX agent

arXiv:2503.24191v4 Announce Type: replace-cross Abstract: Content Warning: This paper may contain unsafe or harmful content generated by LLMs that may be offensive to readers. Large Language Models (L

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

An Exploratory Evaluation of LLM-Assisted Rewriting of Moderate-Complexity Financial Sentences for DisCoCat-Based Sentiment Analysis

DGX agent

arXiv:2608.07439v1 Announce Type: new Abstract: Quantum natural language processing (QNLP) provides a grammar-aware framework for text modeling, and Distributional Compositional Categorical (DisCoCat)

model-releasesarxiv-cs-cl
10 Aug 2026
Research

Are Visual Place Recognition Models Recognizing Places or Conditions? Distractor-Augmented Evaluation and Condition Suppression

DGX agent

arXiv:2608.06847v1 Announce Type: cross Abstract: Long-term Visual Place Recognition (VPR) is typically evaluated by matching queries from one condition against a database from another. Crowdsourced m

researcharxiv-cs-cv
10 Aug 2026
Research

Beyond 'AI Language': The case for the idiolectal nature of LLM output

DGX agent

arXiv:2608.06589v1 Announce Type: cross Abstract: While large language model outputs are frequently analysed as a collective super variety termed 'AI language,' this chapter argues that this perspecti

researcharxiv-cs-ai
10 Aug 2026
Safety

CreativeInstruct: Scalably Teaching LLMs to Balance Quality, Creativity, and Diversity

DGX agent

arXiv:2608.07460v1 Announce Type: cross Abstract: While post-training improves the capabilities of large language models (LLMs), it generally lowers their output diversity and creativity, negatively i

safetyarxiv-cs-ai
10 Aug 2026
Model Releases

DATAREEL: Automated Data-Driven Video Story Generation with Animations

DGX agent

arXiv:2604.25220v2 Announce Type: replace Abstract: Data videos combine animated visualizations with synchronized narration to communicate quantitative information and are widely used in journalism, e

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Diffusion LLMs as Targets and Adversaries: Mechanistic Safety Exploits

DGX agent

arXiv:2608.07430v1 Announce Type: cross Abstract: Diffusion Large Language Models (DLLMs) replace autoregressive next-token prediction with iterative parallel denoising, yet their internal safety mech

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Do AI Personas Grow? Analyzing and Benchmarking Personality Evolution in LLM Agents After Life Events

DGX agent

arXiv:2608.06485v1 Announce Type: cross Abstract: Personality-conditioned LLM agents (PC-Agents) are increasingly used in emotional support, social simulation, and role-playing, motivating the develop

model-releasesarxiv-cs-ai
10 Aug 2026
Safety

From Optimal Actions to World Models: Identifiability of Transition Kernels in Discounted MDPs

DGX agent

arXiv:2608.07301v1 Announce Type: new Abstract: We study what can be recovered about the transition probabilities of a Markov decision process from optimal actions alone. This is closely related to th

safetyarxiv-cs-lg
10 Aug 2026
Model Releases

Learning in Deep Networks under Dale's Constraint

DGX agent

arXiv:2608.06963v1 Announce Type: new Abstract: Biologically plausible learning models aim to explain how neural circuits can implement effective learning under the constraints of real neurons. Althou

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Multi-Agent Forensic Reasoning for Generalizable Deepfake Video Detection

DGX agent

arXiv:2608.06865v1 Announce Type: cross Abstract: The malicious use of generative artificial intelligence to create highly realistic deepfake videos raises serious ethical concerns and poses substanti

model-releasesarxiv-cs-ai
10 Aug 2026
Research

Pre-Inference Routing for Cost-Efficient Document Field Extraction

DGX agent

arXiv:2608.06607v1 Announce Type: new Abstract: Most document-extraction systems use a single model for all documents. This is simple but can be costly for easy cases and less effective for difficult

researcharxiv-cs-cl
10 Aug 2026
Research

ReQuant: Fixed-Grid Discrete Refinement for Post-Training Quantization

DGX agent

arXiv:2608.07019v1 Announce Type: new Abstract: Post-training quantization (PTQ) is widely used to reduce the memory and computational cost of large language models. Existing PTQ methods typically obt

researcharxiv-cs-ai
10 Aug 2026
Model Releases

Stoicheia: Character-Level Masked Diffusion for Ancient Greek Textual Restoration, Parsing, and Metrical Scansion

DGX agent

arXiv:2608.07249v1 Announce Type: new Abstract: We introduce Stoicheia, a 405M-parameter character-level masked-diffusion encoder for Ancient Greek whose input factors into five aligned, independently

model-releasesarxiv-cs-cl
10 Aug 2026
Model Releases

Tensor Network Kernel Machines: A JAX Framework for Machine Learning and Nonlinear System Identification

DGX agent

arXiv:2608.07043v1 Announce Type: cross Abstract: Developing nonlinear models that are both expressive and computationally efficient remains a challenge in machine learning and nonlinear system identi

model-releasesarxiv-cs-lg
10 Aug 2026
Research

BendTwin: Robust Dense-to-Sparse Physical Reconstruction with Bending-Aware Differentiable Spring-Mass Models

DGX agent

arXiv:2608.06164v1 Announce Type: new Abstract: Reconstructing objects with mechanical properties from video observations enables physically consistent dynamic prediction, benefiting robotics planning

researcharxiv-cs-cv
7 Aug 2026
Model Releases

Decoupling Perception from Description: Computation-Grounded Representation Alignment between Multivariate Time Series and Language

DGX agent

arXiv:2608.05238v1 Announce Type: new Abstract: Training multimodal models to align time series with language runs into a self-supervision trap. The usual recipe asks an LLM to read a series and write

model-releasesarxiv-cs-lg
7 Aug 2026
Model Releases

Evaluating and Improving Pedagogical Fit in LLM-Based AI Tutors with the Pedagogical Suitability Index

DGX agent

arXiv:2608.05411v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used as AI tutors, but a correct answer is not always a pedagogically appropriate one. In classroom learni

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

FormBharo: Designing and Evaluating a Voice Agent for Conversational Form Filling in Rural India

DGX agent

arXiv:2608.06027v1 Announce Type: cross Abstract: In India, almost every social benefit starts with a form, yet the people who need these benefits most are often unable to read or write. Reaching them

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Hijacking Robots with a Piece of Paper: A Systematic Study of Physical Prompt Injection in VLM-Controlled Robots

DGX agent

arXiv:2608.05715v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) are increasingly deployed as planners in robotic systems, where they translate natural-language commands into executable

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Kastor: An efficient fine-tuning strategy for generative emulation of PDE simulations

DGX agent

arXiv:2608.06107v1 Announce Type: new Abstract: Machine learning offers a promising avenue to accelerate physical simulations by replacing computationally expensive traditional Partial Differential Eq

model-releasesarxiv-cs-lg
7 Aug 2026
Model Releases

LangChoiceBench: Measuring and Explaining Programming-Language Choice in LLMs

DGX agent

arXiv:2608.06041v1 Announce Type: cross Abstract: Large language models (LLMs) have been shown to exhibit strong Python preferences when generating project-level code, but there is currently no system

model-releasesarxiv-cs-cl
7 Aug 2026
Model Releases

Look Twice: Training-Free Evidence Highlighting for Knowledge-based Visual Question Answering

DGX agent

arXiv:2604.01280v2 Announce Type: replace-cross Abstract: Knowledge-based Visual Question Answering (KB-VQA) requires Multimodal Large Language Models (MLLMs) to identify and combine fine-grained visu

model-releasesarxiv-cs-ai
7 Aug 2026
Tutorials

Measuring and Detecting Harmful AI Sycophancy

DGX agent

arXiv:2608.05624v1 Announce Type: new Abstract: Sycophantic responses are becoming pervasive in large language models (LLMs), and prior work has pointed out that some of them could be harmful. This pa

tutorialsarxiv-cs-ai
7 Aug 2026
Safety

On-Policy Self-Distillation without Any Supervision

DGX agent

arXiv:2608.06296v1 Announce Type: new Abstract: On-policy (Self-)Distillation (OPD / OPSD) has shown strong potential for post-training large language models (LLMs). However, existing methods still re

safetyarxiv-cs-lg
7 Aug 2026
Agents

OPERA: Operator-residual feedback for reliable autonomous optical experiments with language-model agents

DGX agent

arXiv:2608.05990v1 Announce Type: new Abstract: Autonomous agents choose actions using scores that may not reflect experimental success. We developed OPERA, an operator-residual framework for optical

agentsarxiv-cs-ai
7 Aug 2026
Model Releases

Seeing Is Not Deciding: Can Multimodal LLMs Act as Effective CEOs?

DGX agent

arXiv:2608.05864v1 Announce Type: new Abstract: Large language models are increasingly applied as autonomous decision-making agents. However, in executive business decisions, existing benchmarks are l

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Shrinking the Generation-Verification Gap with Weak Verifiers

DGX agent

arXiv:2506.18203v3 Announce Type: replace Abstract: Verifiers can improve language model capabilities by scoring and ranking responses from generated candidates. Currently, high-quality verifiers are

model-releasesarxiv-cs-cl
7 Aug 2026
Tutorials

Spectral Distillation: From Nonlinear Dynamics to Linear State-Space Models

DGX agent

arXiv:2608.05416v1 Announce Type: new Abstract: Can nonlinear dynamical systems be learned through a compact linear state-space representation, without directly solving a non-convex system-identificat

tutorialsarxiv-cs-lg
7 Aug 2026
Model Releases

Subliminal Learning is Non-Semantic Distillation

DGX agent

arXiv:2608.05734v1 Announce Type: new Abstract: Subliminal Learning (SL) is a surprising type of generalization displayed by modern language models. It allows the transfer of a bias or behavior from a

model-releasesarxiv-cs-ai
7 Aug 2026
← Previous
1…274275276277278…1058
Next →