AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,617
  • Agents7,497
  • Applications5,365
  • Concepts5
  • Hardware1,816
  • Industry6,151
  • Local Ai4,900
  • Model Releases23,593
  • Research19,967
  • Safety13,267
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,617
  • Agents7,497
  • Applications5,365
  • Concepts5
  • Hardware1,816
  • Industry6,151
  • Local Ai4,900
  • Model Releases23,593
  • Research19,967
  • Safety13,267
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

Content type
87,617Total entries
1Added by human
87,616Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
51,512 results
Applications

MVIGER: Multi-View Variational Integration of Complementary Knowledge for Generative Recommender

DGX agent

arXiv:2408.08686v4 Announce Type: replace-cross Abstract: Language Models (LMs) have been widely used in recommender systems to incorporate textual information of items into item IDs, leveraging their

applicationsarxiv-cs-ai
28 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

No Test Cases, No Problem: Distillation-Driven Code Generation for Scientific Workflows

DGX agent

arXiv:2604.23106v1 Announce Type: cross Abstract: Existing multi-agent Large Language Model (LLM) frameworks for code generation typically use execution feedback and improve iteratively using Input/Ou

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Patterns vs. Patients: Evaluating LLMs against Mental Health Professionals on Personality Disorder Diagnosis through First-Person Narratives

DGX agent

arXiv:2512.20298v2 Announce Type: replace-cross Abstract: Growing reliance on LLMs for psychiatric self-assessment raises questions about their ability to interpret qualitative patient narratives. Thi

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

PoseX: AI Defeats Physics Approaches on Protein-Ligand Cross Docking

DGX agent

arXiv:2505.01700v3 Announce Type: replace Abstract: Existing protein-ligand docking studies typically focus on the self-docking scenario, which is less practical in real applications. Moreover, some s

model-releasesarxiv-cs-lg
28 Apr 2026
Model Releases

Progressive Approximation in Deep Residual Networks: Theory and Validation

DGX agent

arXiv:2604.24154v1 Announce Type: cross Abstract: The Universal Approximation Theorem (UAT) guarantees universal function approximation but does not explain how residual models distribute approximatio

model-releasesarxiv-cs-ai
28 Apr 2026
Safety

Protecting the Trace: A Principled Black-Box Approach Against Distillation Attacks

DGX agent

arXiv:2604.23238v1 Announce Type: cross Abstract: Frontier models push the boundaries of what is learnable at extreme computational costs, yet distillation via sampling reasoning traces exposes closed

safetyarxiv-cs-ai
28 Apr 2026
Safety

Quantifying Divergence in Inter-LLM Communication Through API Retrieval and Ranking

DGX agent

arXiv:2604.22760v1 Announce Type: cross Abstract: Large language models (LLMs) increasingly operate as autonomous agents that reason over external APIs to perform complex tasks. However, their reliabi

safetyarxiv-cs-ai
28 Apr 2026
Model Releases

RealFin: How Well Do LLMs Reason About Finance When Users Leave Things Unsaid?

DGX agent

arXiv:2602.07096v2 Announce Type: replace-cross Abstract: Reliable financial reasoning requires knowing not only how to answer, but also when an answer cannot be justified. In real financial practice,

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Resource-Lean Lexicon Induction for German Dialects

DGX agent

arXiv:2604.23824v1 Announce Type: new Abstract: Automatic induction of high-quality dictionaries is essential for building lexical resources, yet low-resource languages and dialects pose several chall

model-releasesarxiv-cs-cl
28 Apr 2026
Local Ai

Self Knowledge Re-expression: A Fully Local Method for Adapting LLMs to Tasks Using Intrinsic Knowledge

DGX agent

arXiv:2604.22939v1 Announce Type: cross Abstract: While the next-token prediction (NTP) paradigm enables large language models (LLMs) to express their intrinsic knowledge, its sequential nature constr

local-aiarxiv-cs-ai
28 Apr 2026
Research

SemiSAM-O1: How far can we push the boundary of annotation-efficient medical image segmentation?

DGX agent

arXiv:2604.24109v1 Announce Type: new Abstract: Semi-supervised learning (SSL) has become a promising solution to alleviate the annotation burden of deep learning-based medical image segmentation mode

researcharxiv-cs-cv
28 Apr 2026
Model Releases

SEVerA: Verified Synthesis of Self-Evolving Agents

DGX agent

arXiv:2603.25111v2 Announce Type: replace Abstract: Recent advances have shown the effectiveness of self-evolving LLM agents on tasks such as program repair and scientific discovery. In this paradigm,

model-releasesarxiv-cs-lg
28 Apr 2026
Model Releases

Skill Retrieval Augmentation for Agentic AI

DGX agent

arXiv:2604.24594v1 Announce Type: cross Abstract: As large language models (LLMs) evolve into agentic problem solvers, they increasingly rely on external, reusable skills to handle tasks beyond their

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Stabilizing Efficient Reasoning with Step-Level Advantage Selection

DGX agent

arXiv:2604.24003v1 Announce Type: new Abstract: Large language models (LLMs) achieve strong reasoning performance by allocating substantial computation at inference time, often generating long and ver

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

Switch Attention: Towards Dynamic and Fine-grained Hybrid Transformers

DGX agent

arXiv:2603.26380v2 Announce Type: replace Abstract: The attention mechanism has been the core component in modern transformer architectures. However, the computation of standard full attention scales

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

Symmetric Equilibrium Propagation for Thermodynamic Diffusion Training

DGX agent

arXiv:2604.23806v1 Announce Type: cross Abstract: The reverse process in score-based diffusion models is formally equivalent to overdamped Langevin dynamics in a time-dependent energy landscape. In ou

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

SynthPert: Enhancing LLM Biological Reasoning via Synthetic Reasoning Traces for Cellular Perturbation Prediction

DGX agent

arXiv:2509.25346v2 Announce Type: replace Abstract: Predicting cellular responses to genetic perturbations represents a fundamental challenge in systems biology, critical for advancing therapeutic dis

model-releasesarxiv-cs-ai
28 Apr 2026
Local Ai

Test-Time Adaptation for Unsupervised Combinatorial Optimization

DGX agent

arXiv:2601.21048v2 Announce Type: replace Abstract: Unsupervised neural combinatorial optimization (NCO) enables learning powerful solvers without access to ground-truth solutions. Existing approaches

local-aiarxiv-cs-lg
28 Apr 2026
Model Releases

The Price of Agreement: Measuring LLM Sycophancy in Agentic Financial Applications

DGX agent

arXiv:2604.24668v1 Announce Type: new Abstract: Given the increased use of LLMs in financial systems today, it becomes important to evaluate the safety and robustness of such systems. One failure mode

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

When VLMs 'Fix' Students: Identifying and Penalizing Over-Correction in the Evaluation of Multi-line Handwritten Math OCR

DGX agent

arXiv:2604.22774v1 Announce Type: cross Abstract: Accurate transcription of handwritten mathematics is crucial for educational AI systems, yet current benchmarks fail to evaluate this capability prope

model-releasesarxiv-cs-ai
28 Apr 2026
Safety

World-R1: Reinforcing 3D Constraints for Text-to-Video Generation

DGX agent

arXiv:2604.24764v1 Announce Type: new Abstract: Recent video foundation models demonstrate impressive visual synthesis but frequently suffer from geometric inconsistencies. While existing methods atte

safetyarxiv-cs-cv
28 Apr 2026
Model Releases

Bridging the Long-Tail Gap: Robust Retrieval-Augmented Relation Completion via Multi-Stage Paraphrase Infusion

DGX agent

arXiv:2604.22261v1 Announce Type: new Abstract: Large language models (LLMs) struggle with relation completion (RC), both with and without retrieval-augmented generation (RAG), particularly when the r

model-releasesarxiv-cs-cl
27 Apr 2026
Model Releases

Call-Chain-Aware LLM-Based Test Generation for Java Projects

DGX agent

arXiv:2604.22046v1 Announce Type: cross Abstract: Large language models (LLMs) have recently shown strong potential for generating project-level unit tests. However, existing state-of-the-art approach

model-releasesarxiv-cs-ai
27 Apr 2026
Research

Causal Concept Graphs in LLM Latent Space for Stepwise Reasoning

DGX agent

arXiv:2603.10377v2 Announce Type: replace-cross Abstract: Sparse autoencoders can localize where concepts live in language models, but not how they interact during multi-step reasoning. We propose Cau

researcharxiv-cs-ai
27 Apr 2026
Research

DocPrune:Efficient Document Question Answering via Background, Question, and Comprehension-aware Token Pruning

DGX agent

arXiv:2604.22281v1 Announce Type: new Abstract: Recent advances in vision-language models have demonstrated remarkable performance across diverse multi-modal tasks, including document question answeri

researcharxiv-cs-cv
27 Apr 2026
Model Releases

How Do AI Agents Spend Your Money? Analyzing and Predicting Token Consumption in Agentic Coding Tasks

DGX agent

arXiv:2604.22750v1 Announce Type: new Abstract: The wide adoption of AI agents in complex human workflows is driving rapid growth in LLM token consumption. When agents are deployed on tasks that requi

model-releasesarxiv-cs-cl
27 Apr 2026
Model Releases

Learning Coverage- and Power-Optimal Transmitter Placement from Building Maps: A Comparative Study of Direct and Indirect Neural Approaches

DGX agent

arXiv:2604.22056v1 Announce Type: new Abstract: Optimal wireless transmitter placement is a central task in radio-network planning, yet exhaustive search becomes prohibitively expensive at scale. This

model-releasesarxiv-cs-lg
27 Apr 2026
Tutorials

Multimodal Diffusion to Mutually Enhance Polarized Light and Low Resolution EBSD Data

DGX agent

arXiv:2604.22212v1 Announce Type: cross Abstract: In spite of the utility of 3-D electron back-scattered diffraction (EBSD) microscopy, the data collection process can be time-consuming with serial-se

tutorialsarxiv-cs-cv
27 Apr 2026
Model Releases

Optimal sequential decision-making for error propagation mitigation in digital twins

DGX agent

arXiv:2604.22168v1 Announce Type: new Abstract: Here, we explore the problem of error propagation mitigation in modular digital twins as a sequential decision process. Building on a companion study th

model-releasesarxiv-cs-lg
27 Apr 2026
Research

Outcome Rewards Do Not Guarantee Verifiable or Causally Important Reasoning

DGX agent

arXiv:2604.22074v1 Announce Type: new Abstract: Reinforcement Learning from Verifiable Rewards (RLVR) on chain-of-thought reasoning has become a standard part of language model post-training recipes.

researcharxiv-cs-cl
27 Apr 2026
Research

Rethinking Token Pruning for Historical Screenshots in GUI Visual Agents: Semantic, Spatial, and Temporal Perspectives

DGX agent

arXiv:2603.26041v3 Announce Type: replace Abstract: In recent years, GUI visual agents built upon Multimodal Large Language Models (MLLMs) have demonstrated strong potential in navigation tasks. Howev

researcharxiv-cs-cv
27 Apr 2026
Research

Selective Rotary Position Embedding

DGX agent

arXiv:2511.17388v2 Announce Type: replace Abstract: Position information is essential for language modeling. In softmax transformers, Rotary Position Embeddings (extit{RoPE}) encode positions through

researcharxiv-cs-cl
27 Apr 2026
Applications

Shared Lexical Task Representations Explain Behavioral Variability In LLMs

DGX agent

arXiv:2604.22027v1 Announce Type: cross Abstract: One of the most common complaints about large language models (LLMs) is their prompt sensitivity -- that is, the fact that their ability to perform a

applicationsarxiv-cs-ai
27 Apr 2026
Model Releases

SpaMEM: Benchmarking Dynamic Spatial Reasoning via Perception-Memory Integration in Embodied Environments

DGX agent

arXiv:2604.22409v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have advanced static visual--spatial reasoning, yet they often fail to preserve long-horizon spatial coherence

model-releasesarxiv-cs-cv
27 Apr 2026
Model Releases

Spend Less, Fit Better: Budget-Efficient Scaling Law Fitting via Active Experiment Selection

DGX agent

arXiv:2604.22753v1 Announce Type: new Abstract: Scaling laws are used to plan multi-million-dollar training runs, but fitting those laws can itself cost millions. In modern large-scale workflows, asse

model-releasesarxiv-cs-lg
27 Apr 2026
Research

StateX: Enhancing RNN Recall via Post-training State Expansion

DGX agent

arXiv:2509.22630v3 Announce Type: replace-cross Abstract: Recurrent neural networks (RNNs), such as linear attention and state-space models, have gained popularity due to their constant per-token comp

researcharxiv-cs-ai
27 Apr 2026
Safety

TabSCM: A practical Framework for Generating Realistic Tabular Data

DGX agent

arXiv:2604.22337v1 Announce Type: new Abstract: Most tabular-data generators match marginal statistics yet ignore causal structure, leading downstream models to learn spurious or unfair patterns. We p

safetyarxiv-cs-lg
27 Apr 2026
Research

The Shape of Adversarial Influence: Characterizing LLM Latent Spaces with Persistent Homology

DGX agent

arXiv:2505.20435v3 Announce Type: replace-cross Abstract: Existing interpretability methods for Large Language Models (LLMs) predominantly capture linear directions or isolated features. This overlook

researcharxiv-cs-ai
27 Apr 2026
Model Releases

When Does LLM Self-Correction Help? A Control-Theoretic Markov Diagnostic and Verify-First Intervention

DGX agent

arXiv:2604.22273v1 Announce Type: new Abstract: Iterative self-correction is widely used in agentic LLM systems, but when repeated refinement helps versus hurts remains unclear. We frame self-correcti

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

Wiggle and Go! System Identification for Zero-Shot Dynamic Rope Manipulation

DGX agent

arXiv:2604.22102v1 Announce Type: cross Abstract: Many robotic tasks are unforgiving; a single mistake in a dynamic throw can lead to unacceptable delays or unrecoverable failure. To mitigate this, we

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

APCoTTA: Continual Test-Time Adaptation for Semantic Segmentation of Airborne LiDAR Point Clouds

DGX agent

arXiv:2505.09971v3 Announce Type: replace Abstract: Airborne laser scanning (ALS) point cloud semantic segmentation is a fundamental task for large-scale 3D scene understanding. Fixed models deployed

model-releasesarxiv-cs-cv
24 Apr 2026
Safety

ATATA: One Algorithm to Align Them All

DGX agent

arXiv:2601.11194v2 Announce Type: replace Abstract: We suggest a new multi-modal algorithm for joint inference of paired structurally aligned samples with Rectified Flow models. While some existing me

safetyarxiv-cs-cv
24 Apr 2026
Model Releases

Beyond N-gram: Data-Aware X-GRAM Extraction for Efficient Embedding Parameter Scaling

DGX agent

arXiv:2604.21724v1 Announce Type: new Abstract: Large token-indexed lookup tables provide a compute-decoupled scaling path, but their practical gains are often limited by poor parameter efficiency and

model-releasesarxiv-cs-cl
24 Apr 2026
Research

Calibeating Prediction-Powered Inference

DGX agent

arXiv:2604.21260v1 Announce Type: cross Abstract: We study semisupervised mean estimation with a small labeled sample, a large unlabeled sample, and a black-box prediction model whose output may be mi

researcharxiv-cs-ai
24 Apr 2026
Model Releases

Counterfactual Segmentation Reasoning: Diagnosing and Mitigating Pixel-Grounding Hallucination

DGX agent

arXiv:2506.21546v4 Announce Type: replace-cross Abstract: Segmentation Vision-Language Models (VLMs) have significantly advanced grounded visual understanding, yet they remain prone to pixel-grounding

model-releasesarxiv-cs-ai
24 Apr 2026
Applications

Cross-Domain Data Selection and Augmentation for Automatic Compliance Detection

DGX agent

arXiv:2604.21469v1 Announce Type: new Abstract: Automating the detection of regulatory compliance remains a challenging task due to the complexity and variability of legal texts. Models trained on one

applicationsarxiv-cs-cl
24 Apr 2026
Model Releases

Decoupled DiLoCo for Resilient Distributed Pre-training

DGX agent

arXiv:2604.21428v1 Announce Type: new Abstract: Modern large-scale language model pre-training relies heavily on the single program multiple data (SPMD) paradigm, which requires tight coupling across

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

Dr. Assistant: Enhancing Clinical Diagnostic Inquiry via Structured Diagnostic Reasoning Data and Reinforcement Learning

DGX agent

arXiv:2601.13690v2 Announce Type: replace Abstract: Clinical Decision Support Systems (CDSSs) provide reasoning and inquiry guidance for physicians, yet they face notable challenges, including high ma

model-releasesarxiv-cs-cl
24 Apr 2026
← Previous
1…377378379380381…1074
Next →