AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,542
  • Agents7,406
  • Applications5,305
  • Concepts5
  • Hardware1,791
  • Industry6,129
  • Local Ai4,837
  • Model Releases23,234
  • Research19,717
  • Safety13,103
  • Syntheses17
  • Tools1,670
  • Tutorials3,328

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,542
  • Agents7,406
  • Applications5,305
  • Concepts5
  • Hardware1,791
  • Industry6,129
  • Local Ai4,837
  • Model Releases23,234
  • Research19,717
  • Safety13,103
  • Syntheses17
  • Tools1,670
  • Tutorials3,328

Source
HumanDGX agent

Content type
86,542Total entries
1Added by human
86,541Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
50,764 results
Model Releases

A Parameter-Efficient Transfer Learning Approach through Multitask Prompt Distillation and Decomposition for Clinical NLP

DGX agent

arXiv:2604.06650v1 Announce Type: cross Abstract: Existing prompt-based fine-tuning methods typically learn task-specific prompts independently, imposing significant computing and storage overhead at

model-releasesarxiv-cs-ai
10 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Alloc-MoE: Budget-Aware Expert Activation Allocation for Efficient Mixture-of-Experts Inference

DGX agent

arXiv:2604.08133v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) has become a dominant architecture for scaling large language models due to their sparse activation mechanism. However, the s

model-releasesarxiv-cs-cl
10 Apr 2026
Local Ai

Auditing Black-Box LLM APIs with a Rank-Based Uniformity Test

DGX agent

arXiv:2506.06975v5 Announce Type: replace-cross Abstract: As API access becomes a primary interface to large language models (LLMs), users often interact with black-box systems that offer little trans

local-aiarxiv-cs-cl
10 Apr 2026
Model Releases

Bayesian E(3)-Equivariant Interatomic Potential with Iterative Restratification of Many-body Message Passing

DGX agent

arXiv:2510.03046v2 Announce Type: replace Abstract: Machine learning potentials (MLPs) have become essential for large-scale atomistic simulations, enabling ab initio-level accuracy with computational

model-releasesarxiv-cs-lg
10 Apr 2026
Model Releases

Beyond Case Law: Evaluating Structure-Aware Retrieval and Safety in Statute-Centric Legal QA

DGX agent

arXiv:2604.06173v1 Announce Type: cross Abstract: Legal QA benchmarks have predominantly focused on case law, overlooking the unique challenges of statute-centric regulatory reasoning. In statutory do

model-releasesarxiv-cs-ai
10 Apr 2026
Research

Cram Less to Fit More: Training Data Pruning Improves Memorization of Facts

DGX agent

arXiv:2604.08519v1 Announce Type: new Abstract: Large language models (LLMs) can struggle to memorize factual knowledge in their parameters, often leading to hallucinations and poor performance on kno

researcharxiv-cs-cl
10 Apr 2026
Model Releases

CycleChart: A Unified Consistency-Based Learning Framework for Bidirectional Chart Understanding and Generation

DGX agent

arXiv:2512.19173v2 Announce Type: replace Abstract: Current chart-related tasks, such as chart generation (NL2Chart), chart schema parsing, chart data parsing, and chart question answering (ChartQA),

model-releasesarxiv-cs-cl
10 Apr 2026
Research

Distilling Specialized Orders for Visual Generation

DGX agent

arXiv:2504.17069v2 Announce Type: replace Abstract: Autoregressive (AR) image generators are becoming increasingly popular due to their ability to produce high-quality images and their scalability. Ty

researcharxiv-cs-cv
10 Apr 2026
Model Releases

Don't Overthink It: Inter-Rollout Action Agreement as a Free Adaptive-Compute Signal for LLM Agents

DGX agent

arXiv:2604.08369v1 Announce Type: cross Abstract: Inference-time compute scaling has emerged as a powerful technique for improving the reliability of large language model (LLM) agents, but existing me

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

DROP: Distributional and Regular Optimism and Pessimism for Reinforcement Learning

DGX agent

arXiv:2410.17473v2 Announce Type: replace Abstract: In reinforcement learning (RL), temporal difference (TD) error is known to be related to the firing rate of dopamine neurons. It has been observed t

model-releasesarxiv-cs-lg
10 Apr 2026
Model Releases

Dual-level Modality Debiasing Learning for Unsupervised Visible-Infrared Person Re-Identification

DGX agent

arXiv:2512.03745v2 Announce Type: replace Abstract: Two-stage learning pipeline has achieved promising results in unsupervised visible-infrared person re-identification (USL-VI-ReID). It first perform

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

ESOM: Efficiently Understanding Streaming Video Anomalies with Open-world Dynamic Definitions

DGX agent

arXiv:2604.07772v1 Announce Type: new Abstract: Open-world video anomaly detection (OWVAD) aims to detect and explain abnormal events under different anomaly definitions, which is important for applic

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

Evaluating LLM-Based 0-to-1 Software Generation in End-to-End CLI Tool Scenarios

DGX agent

arXiv:2604.06742v1 Announce Type: cross Abstract: Large Language Models (LLMs) are driving a shift towards intent-driven development, where agents build complete software from scratch. However, existi

model-releasesarxiv-cs-ai
10 Apr 2026
Tutorials

Face2Scene: Using Facial Degradation as an Oracle for Diffusion-Based Scene Restoration

DGX agent

arXiv:2603.16570v2 Announce Type: replace Abstract: Recent advances in image restoration have enabled high-fidelity recovery of faces from degraded inputs using reference-based face restoration models

tutorialsarxiv-cs-cv
10 Apr 2026
Model Releases

FLeX: Fourier-based Low-rank EXpansion for multilingual transfer

DGX agent

arXiv:2604.06253v1 Announce Type: cross Abstract: Cross-lingual code generation is critical in enterprise environments where multiple programming languages coexist. However, fine-tuning large language

model-releasesarxiv-cs-ai
10 Apr 2026
Safety

From Ground Truth to Measurement: A Statistical Framework for Human Labeling

DGX agent

arXiv:2604.07591v1 Announce Type: cross Abstract: Supervised machine learning assumes that labeled data provide accurate measurements of the concepts models are meant to learn. Yet in practice, human

safetyarxiv-cs-cl
10 Apr 2026
Safety

FVD: Inference-Time Alignment of Diffusion Models via Fleming-Viot Resampling

DGX agent

arXiv:2604.06779v1 Announce Type: new Abstract: We introduce Fleming-Viot Diffusion (FVD), an inference-time alignment method that resolves the diversity collapse commonly observed in Sequential Monte

safetyarxiv-cs-ai
10 Apr 2026
Model Releases

HiCI: Hierarchical Construction-Integration for Long-Context Attention

DGX agent

arXiv:2603.20843v2 Announce Type: replace Abstract: Long-context language modeling is commonly framed as a scalability challenge of token-level attention, yet local-to-global information structuring r

model-releasesarxiv-cs-cl
10 Apr 2026
Tutorials

How to sketch a learning algorithm

DGX agent

arXiv:2604.07328v1 Announce Type: new Abstract: How does the choice of training data influence an AI model? This question is of central importance to interpretability, privacy, and basic science. At i

tutorialsarxiv-cs-lg
10 Apr 2026
Model Releases

In-Context Decision Making for Optimizing Complex AutoML Pipelines

DGX agent

arXiv:2508.13657v2 Announce Type: replace-cross Abstract: Combined Algorithm Selection and Hyperparameter Optimization (CASH) has been fundamental to traditional AutoML systems. However, with the adva

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

KEO: Knowledge Extraction on OMIn via Knowledge Graphs and RAG for Safety-Critical Aviation Maintenance

DGX agent

arXiv:2510.05524v2 Announce Type: replace Abstract: We present Knowledge Extraction on OMIn (KEO), a domain-specific knowledge extraction and reasoning framework with large language models (LLMs) in s

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Mitigating Spurious Background Bias in Multimedia Recognition with Disentangled Concept Bottlenecks

DGX agent

arXiv:2510.15770v3 Announce Type: replace Abstract: Concept Bottleneck Models (CBMs) enhance interpretability by predicting human-understandable concepts as intermediate representations. However, exis

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

Noise Immunity in In-Context Tabular Learning: An Empirical Robustness Analysis of TabPFN's Attention Mechanisms

DGX agent

arXiv:2604.04868v2 Announce Type: replace-cross Abstract: Tabular foundation models (TFMs) such as TabPFN (Tabular Prior-Data Fitted Network) are designed to generalize across heterogeneous tabular da

model-releasesarxiv-cs-ai
10 Apr 2026
Research

On the Step Length Confounding in LLM Reasoning Data Selection

DGX agent

arXiv:2604.06834v1 Announce Type: cross Abstract: Large reasoning models have recently demonstrated strong performance on complex tasks that require long chain-of-thought reasoning, through supervised

researcharxiv-cs-ai
10 Apr 2026
Safety

Quality-preserving Model for Electronics Production Quality Tests Reduction

DGX agent

arXiv:2604.06451v1 Announce Type: new Abstract: Manufacturing test flows in high-volume electronics production are typically fixed during product development and executed unchanged on every unit, even

safetyarxiv-cs-lg
10 Apr 2026
Agents

Scaling-Aware Data Selection for End-to-End Autonomous Driving Systems

DGX agent

arXiv:2604.08366v1 Announce Type: cross Abstract: Large-scale deep learning models for physical AI applications depend on diverse training data collection efforts. These models and correspondingly, th

agentsarxiv-cs-cv
10 Apr 2026
Model Releases

Sell More, Play Less: Benchmarking LLM Realistic Selling Skill

DGX agent

arXiv:2604.07054v2 Announce Type: replace Abstract: Sales dialogues require multi-turn, goal-directed persuasion under asymmetric incentives, which makes them a challenging setting for large language

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

SpatialMosaic: A Multiview VLM Dataset for Partial Visibility

DGX agent

arXiv:2512.23365v3 Announce Type: replace Abstract: The rapid progress of Multimodal Large Language Models (MLLMs) has unlocked the potential for enhanced 3D scene understanding and spatial reasoning.

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

Strategic Persuasion with Trait-Conditioned Multi-Agent Systems for Iterative Legal Argumentation

DGX agent

arXiv:2604.07028v1 Announce Type: cross Abstract: Strategic interaction in adversarial domains such as law, diplomacy, and negotiation is mediated by language, yet most game-theoretic models abstract

model-releasesarxiv-cs-ai
10 Apr 2026
Agents

SubSearch: Intermediate Rewards for Unsupervised Guided Reasoning in Complex Retrieval

DGX agent

arXiv:2604.07415v1 Announce Type: cross Abstract: Large language models (LLMs) are probabilistic in nature and perform more reliably when augmented with external information. As complex queries often

agentsarxiv-cs-cl
10 Apr 2026
Model Releases

Symbiotic-MoE: Unlocking the Synergy between Generation and Understanding

DGX agent

arXiv:2604.07753v1 Announce Type: cross Abstract: Empowering Large Multimodal Models (LMMs) with image generation often leads to catastrophic forgetting in understanding tasks due to severe gradient c

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Tensor-Augmented Convolutional Neural Networks: Enhancing Expressivity with Generic Tensor Kernels

DGX agent

arXiv:2604.08072v1 Announce Type: new Abstract: Convolutional Neural Networks (CNNs) excel at extracting local features hierarchically, but their performance in capturing complex correlations hinges h

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

The Stepwise Informativeness Assumption: Why are Entropy Dynamics and Reasoning Correlated in LLMs?

DGX agent

arXiv:2604.06192v1 Announce Type: cross Abstract: Recent work uses entropy-based signals at multiple representation levels to study reasoning in large language models, but the field remains largely em

model-releasesarxiv-cs-ai
10 Apr 2026
Research

Transformer See, Transformer Do: Copying as an Intermediate Step in Learning Analogical Reasoning

DGX agent

arXiv:2604.06501v1 Announce Type: new Abstract: Analogical reasoning is a hallmark of human intelligence, enabling us to solve new problems by transferring knowledge from one situation to another. Yet

researcharxiv-cs-lg
10 Apr 2026
Model Releases

TSUBASA: Improving Long-Horizon Personalization via Evolving Memory and Self-Learning with Context Distillation

DGX agent

arXiv:2604.07894v1 Announce Type: new Abstract: Personalized large language models (PLLMs) have garnered significant attention for their ability to align outputs with individual's needs and preference

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

VisCoder2: Building Multi-Language Visualization Coding Agents

DGX agent

arXiv:2510.23642v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have recently enabled coding agents capable of generating, executing, and revising visualization code. However, e

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

ChainSpace: A Chained-Reasoning Paradigm for Spatial Intelligence

DGX agent

arXiv:2608.15788v1 Announce Type: new Abstract: Spatial intelligence requires foundation models to maintain coherent spatial state across interactions with the physical world. However, existing data-c

model-releasesarxiv-cs-cv
18 Aug 2026
Research

Do Uncertainty Signals Help? A Systematic Study of Uncertainty-Aware Decoding with Rollback Mechanisms

DGX agent

arXiv:2608.14653v1 Announce Type: cross Abstract: Prediction uncertainty is a widely adopted metric for quantifying model confidence, with downstream applications spanning model explanation, data sele

researcharxiv-cs-ai
18 Aug 2026
Model Releases

Does a Tool Result Carry More Authority Than Plain Text? Three Prospective Studies of False-Claim Adoption in a Synthetic Assignment Task with Claude Opus 5

DGX agent

arXiv:2608.14992v1 Announce Type: new Abstract: Language-model systems increasingly read from stores they also write to, so a claim that was merely written earlier can return looking retrieved. We tes

model-releasesarxiv-cs-ai
18 Aug 2026
Model Releases

Does the LM Head Create a Harmful Gradient Bottleneck? A Causal Test

DGX agent

arXiv:2608.16671v1 Announce Type: new Abstract: The language-model head maps a hidden state of width D to a vocabulary of size V, so its transpose can return at most D independent directions to the Tr

model-releasesarxiv-cs-cl
18 Aug 2026
Model Releases

Enhancing the Non-Functional Quality Compliance of LLM-Generated Code through Quality-Aware Preference Learning

DGX agent

arXiv:2503.09020v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have been widely adopted in commercial code completion engines, significantly enhancing coding efficiency and pro

model-releasesarxiv-cs-ai
18 Aug 2026
Model Releases

FirstDiff: One-Step Diffusion-Based Anomaly Detection for Multivariate Time Series via Initial Noise Prediction

DGX agent

arXiv:2608.15727v1 Announce Type: cross Abstract: Diffusion models have recently shown strong potential for multivariate time-series anomaly detection by learning the distribution of normal data throu

model-releasesarxiv-cs-ai
18 Aug 2026
Research

Forward Pass Domain Adaptation (Without Cross-Layer Backpropagation)

DGX agent

arXiv:2608.14563v1 Announce Type: cross Abstract: Forward-Pass-Only MLP training (FPO) adapts large language models without a backward pass through the model body, achieving 2.7--3.2x the throughput o

researcharxiv-cs-ai
18 Aug 2026
Model Releases

Gathered, Not Admitted: How Attention Brings a Latent Variable into Verbalizable Form

DGX agent

arXiv:2608.15022v1 Announce Type: new Abstract: Language models hold latent quantities in a form they can report on, and more of a quantity is present in that form when the task requires reusing it fl

model-releasesarxiv-cs-ai
18 Aug 2026
Tutorials

Human Pose Estimation in Trampoline Gymnastics: How to Improve Performance on Extreme Poses

DGX agent

arXiv:2604.01322v2 Announce Type: replace Abstract: Trampoline gymnastics involves extreme human poses and uncommon viewpoints, on which state-of-the art pose estimation models tend to under-perform.

tutorialsarxiv-cs-cv
18 Aug 2026
Model Releases

HyMem: Hierarchical Context Management for Long-Horizon Agents via Information Isolation

DGX agent

arXiv:2608.15703v1 Announce Type: new Abstract: Large language model (LLM) agents often perform poorly on complex, long-horizon tasks because their context becomes increasingly cluttered over time. As

model-releasesarxiv-cs-ai
18 Aug 2026
Model Releases

Improving Influence-based Instruction Tuning Data Selection for Balanced Learning of Diverse Capabilities

DGX agent

arXiv:2501.12147v2 Announce Type: replace-cross Abstract: Selecting appropriate training data is crucial for instruction fine-tuning of large language models (LLMs), which aims to (1) elicit strong ca

model-releasesarxiv-cs-ai
18 Aug 2026
Model Releases

Listen, Reason, and Segment: Aligning LALMs with Editorial Judgment for Media Chapterization

DGX agent

arXiv:2608.16539v1 Announce Type: cross Abstract: Large Audio Language Models (LALMs) have made rapid progress on standardized benchmarks, yet their deployment in practical media workflows, curation,

model-releasesarxiv-cs-ai
18 Aug 2026
← Previous
1…299300301302303…1058
Next →