AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,617
  • Agents7,497
  • Applications5,365
  • Concepts5
  • Hardware1,816
  • Industry6,151
  • Local Ai4,900
  • Model Releases23,593
  • Research19,967
  • Safety13,267
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,617
  • Agents7,497
  • Applications5,365
  • Concepts5
  • Hardware1,816
  • Industry6,151
  • Local Ai4,900
  • Model Releases23,593
  • Research19,967
  • Safety13,267
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

Content type
87,617Total entries
1Added by human
87,616Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
51,512 results
Model Releases

Take Out Your Calculators: Estimating the Real Difficulty of Question Items with LLM Student Simulations

DGX agent

arXiv:2601.09953v2 Announce Type: replace Abstract: Standardized math assessments require expensive human pilot studies to establish the difficulty of test items. We investigate the predictive value o

model-releasesarxiv-cs-cl
22 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Talking to a Know-It-All GPT or a Second-Guesser Claude? How Repair reveals unreliable Multi-Turn Behavior in LLMs

DGX agent

arXiv:2604.19245v1 Announce Type: cross Abstract: Repair, an important resource for resolving trouble in human-human conversation, remains underexplored in human-LLM interaction. In this study, we inv

model-releasesarxiv-cs-ai
22 Apr 2026
Research

TrEEStealer: Stealing Decision Trees via Enclave Side Channels

DGX agent

arXiv:2604.18716v1 Announce Type: cross Abstract: Today, machine learning is widely applied in sensitive, security-related, and financially lucrative applications. Model extraction attacks undermine c

researcharxiv-cs-lg
22 Apr 2026
Research

Understanding LLM Performance Degradation in Multi-Instance Processing: The Roles of Instance Count and Context Length

DGX agent

arXiv:2603.22608v2 Announce Type: replace Abstract: Users often rely on Large Language Models (LLMs) for processing multiple documents or performing analysis over a number of instances. For example, a

researcharxiv-cs-ai
22 Apr 2026
Model Releases

Unveiling Fine-Grained Visual Traces: Evaluating Multimodal Interleaved Reasoning Chains in Multimodal STEM Tasks

DGX agent

arXiv:2604.19697v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have shown promising reasoning abilities, yet evaluating their performance in specialized domains remains chall

model-releasesarxiv-cs-cv
22 Apr 2026
Model Releases

VDPP: Video Depth Post-Processing for Speed and Scalability

DGX agent

arXiv:2604.06665v2 Announce Type: replace Abstract: Video depth estimation is essential for providing 3D scene structure in applications ranging from autonomous driving to mixed reality. Current end-t

model-releasesarxiv-cs-cv
22 Apr 2026
Model Releases

Visual Reasoning Agent: Robust Vision Systems in Remote Sensing via Inference-Time Scaling

DGX agent

arXiv:2509.16343v2 Announce Type: replace-cross Abstract: Building robust vision systems for high-stakes domains such as remote sensing requires stronger visual reasoning than what single-pass inferen

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

When and What to Ask: AskBench and Rubric-Guided RLVR for LLM Clarification

DGX agent

arXiv:2602.11199v2 Announce Type: replace Abstract: Large language models (LLMs) often respond even when prompts omit critical details or include misleading information, leading to hallucinations or r

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

Adaptive Local Frequency Filtering for Fourier-Encoded Implicit Neural Representations

DGX agent

arXiv:2604.02846v2 Announce Type: replace Abstract: Fourier-encoded implicit neural representations (INRs) have shown strong capability in modeling continuous signals from discrete samples. However, c

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

AeroRAG: Structured Multimodal Retrieval-Augmented LLM for Fine-Grained Aerial Visual Reasoning

DGX agent

arXiv:2604.17889v1 Announce Type: new Abstract: Despite recent progress in multimodal large language models (MLLMs), reliable visual question answering in aerial scenes remains challenging. In such sc

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

AeroScene: Progressive Scene Synthesis for Aerial Robotics

DGX agent

arXiv:2603.23224v2 Announce Type: replace Abstract: Generative models have shown substantial impact across multiple domains, their potential for scene synthesis remains underexplored in robotics. This

model-releasesarxiv-cs-ro
21 Apr 2026
Model Releases

ArgBench: Benchmarking LLMs on Computational Argumentation Tasks

DGX agent

arXiv:2604.17366v1 Announce Type: new Abstract: Argumentation skills are an essential toolkit for large language models (LLMs). These skills are crucial in various use cases, including self-reflection

model-releasesarxiv-cs-cl
21 Apr 2026
Local Ai

Beyond Memorization: Extending Reasoning Depth with Recurrence, Memory and Test-Time Compute Scaling

DGX agent

arXiv:2508.16745v2 Announce Type: replace Abstract: Reasoning is a core capability of large language models, yet how multi-step reasoning is learned and executed remains unclear. We study this questio

local-aiarxiv-cs-lg
21 Apr 2026
Research

Bridging Coarse and Fine Recognition: A Hybrid Approach for Open-Ended Multi-Granularity Object Recognition in Interactive Educational Games

DGX agent

arXiv:2604.16785v1 Announce Type: new Abstract: Recent advances in Multimodal Large Language Models (MLLMs) have enabled open-ended object recognition, yet they struggle with fine-grained tasks. In co

researcharxiv-cs-cv
21 Apr 2026
Applications

Calibrated? Not for Everyone: How Sexual Orientation and Religious Markers Distort LLM Accuracy and Confidence in Medical QA

DGX agent

arXiv:2604.17316v1 Announce Type: new Abstract: Safe clinical deployment of Large Language Models (LLMs) requires not only high accuracy but also robust uncertainty calibration to ensure models defer

applicationsarxiv-cs-cl
21 Apr 2026
Applications

Can we generate portable representations for clinical time series data using LLMs?

DGX agent

arXiv:2603.23987v2 Announce Type: replace Abstract: Deploying clinical ML is slow and brittle: models that work at one hospital often degrade under distribution shifts at the next. In this work, we st

applicationsarxiv-cs-lg
21 Apr 2026
Model Releases

CaseFacts: A Benchmark for Legal Fact-Checking and Precedent Retrieval

DGX agent

arXiv:2601.17230v2 Announce Type: replace Abstract: Automated Fact-Checking has largely focused on verifying general knowledge against static corpora, overlooking high-stakes domains like law where tr

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

CBRS: Cognitive Blood Request System with Bilingual Dataset and Dual-Layer Filtering for Multi-Platform Social Streams

DGX agent

arXiv:2604.16665v1 Announce Type: new Abstract: Urgent blood donation seeking posts and messages on social media often go unnoticed due to the overwhelming volume of daily communications. Traditional

model-releasesarxiv-cs-cl
21 Apr 2026
Research

Coevolving Representations in Joint Image-Feature Diffusion

DGX agent

arXiv:2604.17492v1 Announce Type: new Abstract: Joint image-feature generative modeling has recently emerged as an effective strategy for improving diffusion training by coupling low-level VAE latents

researcharxiv-cs-cv
21 Apr 2026
Agents

CogDriver: Integrating Cognitive Inertia for Temporally Coherent Planning in Autonomous Driving

DGX agent

arXiv:2509.00789v2 Announce Type: replace Abstract: The pursuit of autonomous agents capable of temporally coherent planning is hindered by a fundamental flaw in current vision-language models (VLMs):

agentsarxiv-cs-cv
21 Apr 2026
Research

Compressing then Matching: An Efficient Pre-training Paradigm for Multimodal Embedding

DGX agent

arXiv:2511.08480v3 Announce Type: replace Abstract: Multimodal Large Language Models advance multimodal representation learning by acquiring transferable semantic embeddings, thereby substantially enh

researcharxiv-cs-cv
21 Apr 2026
Research

DisCa: Accelerating Video Diffusion Transformers with Distillation-Compatible Learnable Feature Caching

DGX agent

arXiv:2602.05449v3 Announce Type: replace Abstract: While diffusion models have achieved great success in the field of video generation, this progress is accompanied by a rapidly escalating computatio

researcharxiv-cs-cv
21 Apr 2026
Safety

DMax: Aggressive Parallel Decoding for dLLMs

DGX agent

arXiv:2604.08302v2 Announce Type: replace Abstract: We present DMax, a new paradigm for efficient diffusion language models (dLLMs). It mitigates error accumulation in parallel decoding, enabling aggr

safetyarxiv-cs-lg
21 Apr 2026
Model Releases

Do LLM-derived graph priors improve multi-agent coordination?

DGX agent

arXiv:2604.17191v1 Announce Type: new Abstract: Multi-agent reinforcement learning (MARL) is crucial for AI systems that operate collaboratively in distributed and adversarial settings, particularly i

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Document-as-Image Representations Fall Short for Scientific Retrieval

DGX agent

arXiv:2604.18508v1 Announce Type: cross Abstract: Many recent document embedding models are trained on document-as-image representations, embedding rendered pages as images rather than the underlying

model-releasesarxiv-cs-cl
21 Apr 2026
Safety

Downgrade to Upgrade: Optimizer Simplification Enhances Robustness in LLM Unlearning

DGX agent

arXiv:2510.00761v5 Announce Type: replace Abstract: Large language model (LLM) unlearning aims to surgically remove the influence of undesired data or knowledge from an existing model while preserving

safetyarxiv-cs-lg
21 Apr 2026
Model Releases

EgoSound: Benchmarking Sound Understanding in Egocentric Videos

DGX agent

arXiv:2602.14122v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have recently achieved remarkable progress in vision-language understanding. Yet, human perception is inher

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Enhancing Glass Surface Reconstruction via Depth Prior for Robot Navigation

DGX agent

arXiv:2604.18336v1 Announce Type: cross Abstract: Indoor robot navigation is often compromised by glass surfaces, which severely corrupt depth sensor measurements. While foundation models like Depth A

model-releasesarxiv-cs-cv
21 Apr 2026
Safety

EVE: Verifiable Self-Evolution of MLLMs via Executable Visual Transformations

DGX agent

arXiv:2604.18320v1 Announce Type: new Abstract: Self-evolution of multimodal large language models (MLLMs) remains a critical challenge: pseudo-label-based methods suffer from progressive quality degr

safetyarxiv-cs-cv
21 Apr 2026
Model Releases

EvoCoT: Overcoming the Exploration Bottleneck in Reinforcement Learning

DGX agent

arXiv:2508.07809v5 Announce Type: replace Abstract: Reinforcement learning with verifiable reward (RLVR) has become a promising paradigm for post-training large language models (LLMs) to improve their

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

FireScope: Wildfire Risk Prediction with a Chain-of-Thought Oracle

DGX agent

arXiv:2511.17171v4 Announce Type: replace Abstract: Predicting wildfire risk is a reasoning-intensive spatial problem that requires the integration of visual, climatic, and geographic factors to infer

model-releasesarxiv-cs-cv
21 Apr 2026
Research

Forget What Matters, Keep the Rest: Selective Unlearning of Informative Tokens

DGX agent

arXiv:2604.17785v1 Announce Type: new Abstract: Unlearning in large language models (LLMs) has emerged as a promising safeguard against adversarial behaviors. When the forgetting loss is applied unifo

researcharxiv-cs-cl
21 Apr 2026
Model Releases

FRIGID: Scaling Diffusion-Based Molecular Generation from Mass Spectra at Training and Inference Time

DGX agent

arXiv:2604.16648v1 Announce Type: new Abstract: In this work, we present FRIGID, a framework with a novel diffusion language model that generates molecular structures conditioned on mass spectra via i

model-releasesarxiv-cs-lg
21 Apr 2026
Research

From Attribution to Abstention: Training-Free Attention-Based Auditing for Clinical Summarization

DGX agent

arXiv:2601.16397v2 Announce Type: replace Abstract: Deploying multimodal large language models (MLLMs) for clinical summarization demands not only fluent generation but also transparency about where e

researcharxiv-cs-cl
21 Apr 2026
Model Releases

From log pi to pi: Taming Divergence in Soft Clipping via Bilateral Decoupled Decay of Probability Gradient Weight

DGX agent

arXiv:2603.14389v2 Announce Type: replace Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has catalyzed a leap in Large Language Model (LLM) reasoning, yet its optimization dynamics re

model-releasesarxiv-cs-lg
21 Apr 2026
Safety

Geometric Stability: The Missing Axis of Representations

DGX agent

arXiv:2601.09173v4 Announce Type: replace-cross Abstract: Representational similarity analysis and related methods have become standard tools for comparing the internal geometries of neural networks a

safetyarxiv-cs-cl
21 Apr 2026
Safety

GeometryZero: Advancing Geometry Solving via Group Contrastive Policy Optimization

DGX agent

arXiv:2506.07160v3 Announce Type: replace Abstract: Recent progress in large language models (LLMs) has boosted mathematical reasoning, yet geometry remains challenging where auxiliary construction is

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

GR4CIL: Gap-compensated Routing for CLIP-based Class Incremental Learning

DGX agent

arXiv:2604.17822v1 Announce Type: new Abstract: Class-Incremental Learning (CIL) aims to continuously acquire new categories while preserving previously learned knowledge. Recently, Contrastive Langua

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Harness as an Asset: Enforcing Determinism via the Convergent AI Agent Framework (CAAF)

DGX agent

arXiv:2604.17025v1 Announce Type: cross Abstract: Large Language Models (LLMs) produce a controllability gap in safety-critical engineering: even low rates of undetected constraint violations render a

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

HiGMem: A Hierarchical and LLM-Guided Memory System for Long-Term Conversational Agents

DGX agent

arXiv:2604.18349v1 Announce Type: new Abstract: Long-term conversational large language model (LLM) agents require memory systems that can recover relevant evidence from historical interactions withou

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

iDocV2: Leveraging Self-Supervision and Open-Set Detection for Improving Pattern Spotting in Historical Documents

DGX agent

arXiv:2604.16726v1 Announce Type: new Abstract: Considering the imminent massification of digital books, it has become critical to facilitate searching collections through graphical patterns. Current

model-releasesarxiv-cs-cv
21 Apr 2026
Research

iPhoneme: Brain-to-Text Communication for ALS Using ConformerXL Decoding

DGX agent

arXiv:2604.16441v1 Announce Type: cross Abstract: Brain-computer interfaces (BCIs) for speech restoration hold transformative potential for the approximately 173,000--232,500 individuals worldwide wit

researcharxiv-cs-cl
21 Apr 2026
Model Releases

Knowing When to Quit: A Principled Framework for Dynamic Abstention in LLM Reasoning

DGX agent

arXiv:2604.18419v1 Announce Type: cross Abstract: Large language models (LLMs) using chain-of-thought reasoning often waste substantial compute by producing long, incorrect responses. Abstention can m

model-releasesarxiv-cs-cl
21 Apr 2026
Research

Latent Phase-Shift Rollback: Inference-Time Error Correction via Residual Stream Monitoring and KV-Cache Steering

DGX agent

arXiv:2604.18567v1 Announce Type: cross Abstract: Large language models frequently commit unrecoverable reasoning errors mid-generation: once a wrong step is taken, subsequent tokens compound the mist

researcharxiv-cs-cl
21 Apr 2026
Research

LLaVA-Octopus: Unlocking Instruction-Driven Adaptive Projector Fusion for Video Understanding

DGX agent

arXiv:2501.05067v3 Announce Type: replace Abstract: In this paper, we introduce LLaVA-Octopus, a novel video multimodal large language model. LLaVA-Octopus adaptively weights features from different v

researcharxiv-cs-cv
21 Apr 2026
Model Releases

Mind the Way You Select Negative Texts: Pursuing the Distance Consistency in OOD Detection with VLMs

DGX agent

arXiv:2603.02618v3 Announce Type: replace Abstract: Out-of-distribution (OOD) detection seeks to identify samples from unknown classes, a critical capability for deploying machine learning models in o

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

MM-JudgeBias: A Benchmark for Evaluating Compositional Biases in MLLM-as-a-Judge

DGX agent

arXiv:2604.18164v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have been increasingly used as automatic evaluators-a paradigm known as MLLM-as-a-Judge. However, their reliabi

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

MobileAgeNet: Lightweight Facial Age Estimation for Mobile Deployment

DGX agent

arXiv:2604.17007v1 Announce Type: new Abstract: Mobile deployment of facial age estimation requires models that balance predictive accuracy with low latency and compact size. In this work, we present

model-releasesarxiv-cs-cv
21 Apr 2026
← Previous
1…380381382383384…1074
Next →