AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,435 results
Safety

Character Beyond Speech: Leveraging Role-Playing Evaluation in Audio Large Language Models via Reinforcement Learning

DGX agent

arXiv:2604.13804v1 Announce Type: new Abstract: The rapid evolution of multimodal large models has revolutionized the simulation of diverse characters in speech dialogue systems, enabling a novel inte

safetyarxiv-cs-lg
16 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

Co-FactChecker: A Framework for Human-AI Collaborative Claim Verification Using Large Reasoning Models

DGX agent

arXiv:2604.13706v1 Announce Type: new Abstract: Professional fact-checkers rely on domain knowledge and deep contextual understanding to verify claims. Large language models (LLMs) and large reasoning

agentsarxiv-cs-cl
16 Apr 2026
Research

Don't Let the Video Speak: Audio-Contrastive Preference Optimization for Audio-Visual Language Models

DGX agent

arXiv:2604.14129v1 Announce Type: new Abstract: While Audio-Visual Language Models (AVLMs) have achieved remarkable progress over recent years, their reliability is bottlenecked by cross-modal halluci

researcharxiv-cs-cv
16 Apr 2026
Research

Empirical Evidence of Complexity-Induced Limits in Large Language Models on Finite Discrete State-Space Problems with Explicit Validity Constraints

DGX agent

arXiv:2604.13371v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly described as possessing strong reasoning capabilities, supported by high performance on mathematical, logi

researcharxiv-cs-cl
16 Apr 2026
Safety

Foresight Optimization for Strategic Reasoning in Large Language Models

DGX agent

arXiv:2604.13592v1 Announce Type: new Abstract: Reasoning capabilities in large language models (LLMs) have generally advanced significantly. However, it is still challenging for existing reasoning-ba

safetyarxiv-cs-cl
16 Apr 2026
Local Ai

From Anchors to Supervision: Memory-Graph Guided Corpus-Free Unlearning for Large Language Models

DGX agent

arXiv:2604.13777v1 Announce Type: new Abstract: Large language models (LLMs) may memorize sensitive or copyrighted content, raising significant privacy and legal concerns. While machine unlearning has

local-aiarxiv-cs-cl
16 Apr 2026
Model Releases

GeoBridge: A Semantic-Anchored Multi-View Foundation Model Bridging Images and Text for Geo-Localization

DGX agent

arXiv:2512.02697v3 Announce Type: replace Abstract: Cross-view geo-localization infers a location by retrieving geo-tagged reference images that visually correspond to a query image. However, the trad

model-releasesarxiv-cs-cv
16 Apr 2026
Research

How Can We Synthesize High-Quality Pretraining Data? A Systematic Study of Prompt Design, Generator Model, and Source Data

DGX agent

arXiv:2604.13977v1 Announce Type: new Abstract: Synthetic data is a standard component in training large language models, yet systematic comparisons across design dimensions, including rephrasing stra

researcharxiv-cs-cl
16 Apr 2026
Research

Linear Probe Accuracy Scales with Model Size and Benefits from Multi-Layer Ensembling

DGX agent

arXiv:2604.13386v1 Announce Type: new Abstract: Linear probes can detect when language models produce outputs they 'know' are wrong, a capability relevant to both deception and reward hacking. However

researcharxiv-cs-lg
16 Apr 2026
Tutorials

Modeling Student Learning with 3.8 Million Program Traces

DGX agent

arXiv:2510.05056v2 Announce Type: replace Abstract: As programmers write code, they often edit and retry multiple times, creating rich 'interaction traces' that reveal how they approach coding tasks a

tutorialsarxiv-cs-lg
16 Apr 2026
Research

Native Hybrid Attention for Efficient Sequence Modeling

DGX agent

arXiv:2510.07019v3 Announce Type: replace Abstract: Transformers excel at sequence modeling but face quadratic complexity, while linear attention offers improved efficiency but often compromises recal

researcharxiv-cs-cl
16 Apr 2026
Research

Quantifying and Understanding Uncertainty in Large Reasoning Models

DGX agent

arXiv:2604.13395v1 Announce Type: cross Abstract: Large Reasoning Models (LRMs) have recently demonstrated significant improvements in complex reasoning. While quantifying generation uncertainty in LR

researcharxiv-cs-lg
16 Apr 2026
Research

CLASP: Class-Adaptive Layer Fusion and Dual-Stage Pruning for Multimodal Large Language Models

DGX agent

arXiv:2604.12767v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) suffer from substantial computational overhead due to the high redundancy in visual token sequences. Existing

researcharxiv-cs-ai
15 Apr 2026
Research

Cognition-Inspired Dual-Stream Semantic Enhancement for Vision-Based Dynamic Emotion Modeling

DGX agent

arXiv:2604.12777v1 Announce Type: cross Abstract: The human brain constructs emotional percepts not by processing facial expressions in isolation, but through a dynamic, hierarchical integration of se

researcharxiv-cs-ai
15 Apr 2026
Research

CycloneMAE: A Scalable Multi-Task Learning Model for Global Tropical Cyclone Probabilistic Forecasting

DGX agent

arXiv:2604.12180v1 Announce Type: cross Abstract: Tropical cyclones (TCs) rank among the most destructive natural hazards, yet their forecasting faces fundamental trade-offs: numerical weather predict

researcharxiv-cs-ai
15 Apr 2026
Research

Efficient Inference for Large Vision-Language Models: Bottlenecks, Techniques, and Prospects

DGX agent

arXiv:2604.05546v2 Announce Type: replace Abstract: Large Vision-Language Models (LVLMs) enable sophisticated reasoning over images and videos, yet their inference is hindered by a systemic efficiency

researcharxiv-cs-cl
15 Apr 2026
Local Ai

Evaluating Language Models for Harmful Manipulation

DGX agent

arXiv:2603.25326v4 Announce Type: replace Abstract: Interest in the concept of AI-driven harmful manipulation is growing, yet current approaches to evaluating it are limited. This paper introduces a f

local-aiarxiv-cs-ai
15 Apr 2026
Model Releases

How memory can affect collective and cooperative behaviors in an LLM-Based Social Particle Swarm

DGX agent

arXiv:2604.12250v1 Announce Type: new Abstract: This study examines how model-specific characteristics of Large Language Model (LLM) agents, including internal alignment, shape the effect of memory on

model-releasesarxiv-cs-ai
15 Apr 2026
Safety

InsightFlow: LLM-Driven Synthesis of Patient Narratives for Mental Health into Causal Models

DGX agent

arXiv:2604.12721v1 Announce Type: new Abstract: Clinical case formulation organizes patient symptoms and psychosocial factors into causal models, often using the 5P framework. However, constructing su

safetyarxiv-cs-cl
15 Apr 2026
Model Releases

LLM-Enhanced Log Anomaly Detection: A Comprehensive Benchmark of Large Language Models for Automated System Diagnostics

DGX agent

arXiv:2604.12218v1 Announce Type: new Abstract: System log anomaly detection is critical for maintaining the reliability of large-scale software systems, yet traditional methods struggle with the hete

model-releasesarxiv-cs-lg
15 Apr 2026
Local Ai

Locate, Steer, and Improve: A Practical Survey of Actionable Mechanistic Interpretability in Large Language Models

DGX agent

arXiv:2601.14004v4 Announce Type: replace Abstract: Mechanistic Interpretability (MI) has emerged as a vital approach to demystify the opaque decision-making of Large Language Models (LLMs). However,

local-aiarxiv-cs-cl
15 Apr 2026
Applications

Mantis: A Foundation Model for Mechanistic Disease Forecasting

DGX agent

arXiv:2508.12260v5 Announce Type: replace Abstract: Infectious disease forecasting in novel outbreaks or low-resource settings is hampered by the need for large disease and covariate data sets, bespok

applicationsarxiv-cs-ai
15 Apr 2026
Model Releases

ParetoBandit: Budget-Paced Adaptive Routing for Non-Stationary LLM Serving

DGX agent

arXiv:2604.00136v2 Announce Type: replace-cross Abstract: Multi-model LLM serving operates in a non-stationary, noisy environment: providers revise pricing, model quality can shift or regress without

model-releasesarxiv-cs-cl
15 Apr 2026
Safety

Preventing Safety Drift in Large Language Models via Coupled Weight and Activation Constraints

DGX agent

arXiv:2604.12384v1 Announce Type: new Abstract: Safety alignment in Large Language Models (LLMs) remains highly fragile during fine-tuning, where even benign adaptation can degrade pre-trained refusal

safetyarxiv-cs-ai
15 Apr 2026
Model Releases

Revisiting the Reliability of Language Models in Instruction-Following

DGX agent

arXiv:2512.14754v2 Announce Type: replace-cross Abstract: Advanced LLMs have achieved near-ceiling instruction-following accuracy on benchmarks such as IFEval. However, these impressive scores do not

model-releasesarxiv-cs-ai
15 Apr 2026
Applications

Robotic Manipulation is Vision-to-Geometry Mapping (f(v) rightarrow G): Vision-Geometry Backbones over Language and Video Models

DGX agent

arXiv:2604.12908v1 Announce Type: new Abstract: At its core, robotic manipulation is a problem of vision-to-geometry mapping (f(v) rightarrow G). Physical actions are fundamentally defined by geometri

applicationsarxiv-cs-ro
15 Apr 2026
Research

Speaker effects in language comprehension: An integrative model of language and speaker processing

DGX agent

arXiv:2412.07238v3 Announce Type: replace Abstract: The identity of a speaker influences language comprehension through modulating perception and expectation. This review explores speaker effects and

researcharxiv-cs-cl
15 Apr 2026
Model Releases

TeRA: Vector-based Random Tensor Network for High-Rank Adaptation of Large Language Models

DGX agent

arXiv:2509.03234v2 Announce Type: replace Abstract: Parameter-Efficient Fine-Tuning (PEFT) methods, such as Low-Rank Adaptation (LoRA), have significantly reduced the number of trainable parameters ne

model-releasesarxiv-cs-lg
15 Apr 2026
Safety

Thinking Sparks!: Emergent Attention Heads in Reasoning Models During Post Training

DGX agent

arXiv:2509.25758v2 Announce Type: replace Abstract: The remarkable capabilities of modern large reasoning models are largely unlocked through post-training techniques such as supervised fine-tuning (S

safetyarxiv-cs-ai
15 Apr 2026
Research

A Survey of Inductive Reasoning for Large Language Models

DGX agent

arXiv:2510.10182v2 Announce Type: replace-cross Abstract: Reasoning is an important task for large language models (LLMs). Among all the reasoning paradigms, inductive reasoning is one of the fundamen

researcharxiv-cs-ai
14 Apr 2026
Model Releases

AdaQE-CG: Adaptive Query Expansion for Web-Scale Generative AI Model and Data Card Generation

DGX agent

arXiv:2604.09617v1 Announce Type: new Abstract: Transparent and standardized documentation is essential for building trustworthy generative AI (GAI) systems. However, existing automated methods for ge

model-releasesarxiv-cs-ai
14 Apr 2026
Tutorials

Asymptotic Learning Curves for Diffusion Models with Random Features Score and Manifold Data

DGX agent

arXiv:2603.22962v2 Announce Type: replace Abstract: We study the theoretical behavior of denoising score matching--the learning task associated to diffusion models--when the data distribution is suppo

tutorialsarxiv-cs-lg
14 Apr 2026
Applications

Bootstrapping Video Semantic Segmentation Model via Distillation-assisted Test-Time Adaptation

DGX agent

arXiv:2604.10950v1 Announce Type: new Abstract: Fully supervised Video Semantic Segmentation (VSS) relies heavily on densely annotated video data, limiting practical applicability. Alternatively, appl

applicationsarxiv-cs-cv
14 Apr 2026
Research

Breaking the KV Cache Bottleneck: Fan Duality Model Achieves O(1) Decode Memory with Superior Associative Recall

DGX agent

arXiv:2604.07716v2 Announce Type: replace Abstract: We present FDM (Fan Duality Model), a linear sequence architecture that resolves the fundamental tension between memory efficiency and associative r

researcharxiv-cs-lg
14 Apr 2026
Research

Cognitive Training for Language Models: Towards General Capabilities via Cross-Entropy Games

DGX agent

arXiv:2603.22479v3 Announce Type: replace-cross Abstract: Defining a constructive process to build general capabilities for language models in an automatic manner is considered an open problem in arti

researcharxiv-cs-ai
14 Apr 2026
Research

Computational Implementation of a Model of Category-Theoretic Metaphor Comprehension

DGX agent

arXiv:2604.10035v1 Announce Type: cross Abstract: In this study, we developed a computational implementation for a model of metaphor comprehension based on the theory of indeterminate natural transfor

researcharxiv-cs-ai
14 Apr 2026
Model Releases

CounterBench: Evaluating and Improving Counterfactual Reasoning in Large Language Models

DGX agent

arXiv:2502.11008v2 Announce Type: replace Abstract: Counterfactual reasoning is widely recognized as one of the most challenging and intricate aspects of causality in artificial intelligence. In this

model-releasesarxiv-cs-cl
14 Apr 2026
Research

Critical-CoT: A Robust Defense Framework against Reasoning-Level Backdoor Attacks in Large Language Models

DGX agent

arXiv:2604.10681v1 Announce Type: cross Abstract: Large Language Models (LLMs), despite their impressive capabilities across domains, have been shown to be vulnerable to backdoor attacks. Prior backdo

researcharxiv-cs-ai
14 Apr 2026
Safety

CROP: Conservative Reward for Model-based Offline Policy Optimization

DGX agent

arXiv:2310.17245v2 Announce Type: replace-cross Abstract: Offline reinforcement learning (RL) aims to optimize a policy using collected data without online interactions. Model-based approaches are par

safetyarxiv-cs-ai
14 Apr 2026
Hardware

Deep Optimizer States: Towards Scalable Training of Transformer Models Using Interleaved Offloading

DGX agent

arXiv:2410.21316v2 Announce Type: replace-cross Abstract: Transformers and large language models~(LLMs) have seen rapid adoption in all domains. Their sizes have exploded to hundreds of billions of pa

hardwarearxiv-cs-ai
14 Apr 2026
Research

Delving Aleatoric Uncertainty in Medical Image Segmentation via Vision Foundation Models

DGX agent

arXiv:2604.10963v1 Announce Type: new Abstract: Medical image segmentation supports clinical workflows by precisely delineating anatomical structures and lesions. However, medical image datasets medic

researcharxiv-cs-ai
14 Apr 2026
Local Ai

Different types of syntactic agreement recruit the same units within large language models

DGX agent

arXiv:2512.03676v2 Announce Type: replace Abstract: Large language models (LLMs) can reliably distinguish grammatical from ungrammatical sentences, but how grammatical knowledge is represented within

local-aiarxiv-cs-cl
14 Apr 2026
Local Ai

Do vision models perceive illusory motion in static images like humans?

DGX agent

arXiv:2604.09853v1 Announce Type: new Abstract: Understanding human motion processing is essential for building reliable, human-centered computer vision systems. Although deep neural networks (DNNs) a

local-aiarxiv-cs-cv
14 Apr 2026
Research

Efficient Process Reward Modeling via Contrastive Mutual Information

DGX agent

arXiv:2604.10660v1 Announce Type: cross Abstract: Recent research has devoted considerable effort to verifying the intermediate reasoning steps of chain-of-thought (CoT) trajectories using process rew

researcharxiv-cs-ai
14 Apr 2026
Safety

Efficient Training for Cross-lingual Speech Language Models

DGX agent

arXiv:2604.11096v1 Announce Type: cross Abstract: Currently, large language models (LLMs) predominantly focus on the text modality. To enable more natural human-AI interaction, speech LLMs are emergin

safetyarxiv-cs-ai
14 Apr 2026
Safety

Empowering Video Translation using Multimodal Large Language Models

DGX agent

arXiv:2604.11283v1 Announce Type: new Abstract: Recent developments in video translation have further enhanced cross-lingual access to video content, with multimodal large language models (MLLMs) play

safetyarxiv-cs-cv
14 Apr 2026
Tutorials

Energy-oriented Diffusion Bridge for Image Restoration with Foundational Diffusion Models

DGX agent

arXiv:2604.10983v1 Announce Type: new Abstract: Diffusion bridge models have shown great promise in image restoration by explicitly connecting clean and degraded image distributions. However, they oft

tutorialsarxiv-cs-cv
14 Apr 2026
Tutorials

EviCare: Enhancing Diagnosis Prediction with Deep Model-Guided Evidence for In-Context Reasoning

DGX agent

arXiv:2604.10455v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) have enabled promising progress in diagnosis prediction from electronic health records (EHRs). However,

tutorialsarxiv-cs-cl
14 Apr 2026
← Previous
1…143144145146147…1030
Next →