AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,435 results
Research

Can Nano Banana 2 Replace Traditional Image Restoration Models? An Evaluation of Its Performance on Image Restoration Tasks

DGX agent

arXiv:2604.03061v2 Announce Type: replace Abstract: Recent advances in generative AI raise the question of whether general-purpose image editing models can serve as unified solutions for image restora

researcharxiv-cs-cv
13 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

CATS: Cascaded Adaptive Tree Speculation for Memory-Limited LLM Inference Acceleration

DGX agent

arXiv:2605.11186v1 Announce Type: new Abstract: Auto-regressive decoding in Large Language Models (LLMs) is inherently memory-bound: every generation step requires loading the model weights and interm

model-releasesarxiv-cs-lg
13 May 2026
Safety

Combining On-Policy Optimization and Distillation for Long-Context Reasoning in Large Language Models

DGX agent

arXiv:2605.12227v1 Announce Type: new Abstract: Adapting large language models (LLMs) to long-context tasks requires post-training methods that remain accurate and coherent over thousands of tokens. E

safetyarxiv-cs-cl
13 May 2026
Research

Concepts in Motion: Temporal Concept Bottleneck Model for Interpretable Video Classification

DGX agent

arXiv:2509.20899v3 Announce Type: replace Abstract: Concept Bottleneck Models (CBMs) enable interpretable image classification by structuring predictions around human-understandable concepts, but exte

researcharxiv-cs-cv
13 May 2026
Model Releases

Correcting Selection Bias in Sparse User Feedback for Large Language Model Quality Estimation: A Multi-Agent Hierarchical Bayesian Approach

DGX agent

arXiv:2605.12177v1 Announce Type: new Abstract: [Abridged] Production LLM deployments receive feedback from a non-random fraction of users: thumbs sit mostly in the tails of the satisfaction distribut

model-releasesarxiv-cs-cl
13 May 2026
Applications

Dynamic Execution Commitment of Vision-Language-Action Models

DGX agent

arXiv:2605.11567v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models predominantly adopt action chunking, i.e., predicting and committing to a short horizon of consecutive low-level act

applicationsarxiv-cs-cv
13 May 2026
Research

From Model Uncertainty to Human Attention: Localization-Aware Visual Cues for Scalable Annotation Review

DGX agent

arXiv:2605.12303v1 Announce Type: cross Abstract: High-quality labeled data is essential for training robust machine learning models, yet obtaining annotations at scale remains expensive. AI-assisted

researcharxiv-cs-cv
13 May 2026
Research

Is Monotonic Sampling Necessary in Diffusion Models?

DGX agent

arXiv:2605.11773v1 Announce Type: new Abstract: Diffusion models generate samples by iteratively denoising a Gaussian prior, traversing a sequence of noise levels that, in every published sampler, dec

researcharxiv-cs-lg
13 May 2026
Research

One-Step Generative Modeling via Wasserstein Gradient Flows

DGX agent

arXiv:2605.11755v1 Announce Type: cross Abstract: Diffusion models and flow-based methods have shown impressive generative capability, especially for images, but their sampling is expensive because it

researcharxiv-cs-cv
13 May 2026
Safety

ORCE: Order-Aware Alignment of Verbalized Confidence in Large Language Models

DGX agent

arXiv:2605.12446v1 Announce Type: cross Abstract: Large language models (LLMs) often produce answers with high certainty even when they are incorrect, making reliable confidence estimation essential f

safetyarxiv-cs-cl
13 May 2026
Research

Overparametrized models with posterior drift

DGX agent

arXiv:2506.23619v2 Announce Type: replace-cross Abstract: This paper investigates the impact of posterior drift on out-of-sample forecasting accuracy in overparametrized machine learning models. We do

researcharxiv-cs-lg
13 May 2026
Agents

Predicting Decisions of AI Agents from Limited Interaction through Text-Tabular Modeling

DGX agent

arXiv:2605.12411v1 Announce Type: cross Abstract: AI agents negotiate and transact in natural language with unfamiliar counterparts: a buyer bot facing an unknown seller, or a procurement assistant ne

agentsarxiv-cs-cl
13 May 2026
Safety

Pretraining Exposure Explains Popularity Judgments in Large Language Models

DGX agent

arXiv:2605.12382v1 Announce Type: new Abstract: Large language models (LLMs) exhibit systematic preferences for well-known entities, a phenomenon often attributed to popularity bias. However, the exte

safetyarxiv-cs-cl
13 May 2026
Safety

Simulation Distillation: Pretraining World Models in Simulation for Rapid Real-World Adaptation

DGX agent

arXiv:2603.15759v2 Announce Type: replace-cross Abstract: Robot learning requires adaptation methods that improve reliably from limited, mixed-quality interaction data. This is especially challenging

safetyarxiv-cs-lg
13 May 2026
Research

Steering Without Breaking: Mechanistically Informed Interventions for Discrete Diffusion Language Models

DGX agent

arXiv:2605.10971v1 Announce Type: cross Abstract: Discrete diffusion language models (DLMs) generate text by iteratively denoising all positions in parallel, offering an alternative to autoregressive

researcharxiv-cs-cl
13 May 2026
Tutorials

What makes a word hard to learn? Modeling L1 influence on English vocabulary difficulty

DGX agent

arXiv:2605.12281v1 Announce Type: new Abstract: What makes a word difficult to learn, and how does the difficulty depend on the learner's native language? We computationally model vocabulary difficult

tutorialsarxiv-cs-cl
13 May 2026
Safety

A Single Neuron Is Sufficient to Bypass Safety Alignment in Large Language Models

DGX agent

arXiv:2605.08513v1 Announce Type: cross Abstract: Safety alignment in language models operates through two mechanistically distinct systems: refusal neurons that gate whether harmful knowledge is expr

safetyarxiv-cs-ai
12 May 2026
Research

Any2Any 3D Diffusion Models with Knowledge Transfer: A Radiotherapy Planning Study

DGX agent

arXiv:2605.09622v1 Announce Type: cross Abstract: Voxel-wise dose prediction is a critical yet challenging task in practical radiotherapy (RT) planning, as bespoke models trained from scratch often st

researcharxiv-cs-ai
12 May 2026
Research

BaLoRA: Bayesian Low-Rank Adaptation of Large Scale Models

DGX agent

arXiv:2605.08110v1 Announce Type: cross Abstract: Low-Rank Adaptation (LoRA) has become the standard for fine-tuning large pre-trained models at reduced computational cost. However, its low-rank point

researcharxiv-cs-ai
12 May 2026
Research

BetaEdit: Null-Space Constrained Sequential Model Editing

DGX agent

arXiv:2605.09285v1 Announce Type: new Abstract: Null-space-based methods have garnered considerable attention in model editing by constraining updates to the null space of the pre-existing knowledge r

researcharxiv-cs-cl
12 May 2026
Local Ai

Beyond ViT Tokens: Masked-Diffusion Pretrained Convolutional Pathology Foundation Model for Cell-Level Dense Prediction

DGX agent

arXiv:2605.08276v1 Announce Type: new Abstract: Cell-level dense prediction is central to computational pathology, but remains challenging due to fine-grained histological structures, strong domain sh

local-aiarxiv-cs-cv
12 May 2026
Model Releases

BGG: Bridging the Geometric Gap between Cross-View images by Vision Foundation Model Adaptation for Geo-Localization

DGX agent

arXiv:2605.10345v1 Announce Type: new Abstract: Geometric differences between cross-view images, such as drone and satellite views, significantly increase the challenge of Cross-View Geo-Localization

model-releasesarxiv-cs-cv
12 May 2026
Safety

Compute Where it Counts: Self Optimizing Language Models

DGX agent

arXiv:2605.10875v1 Announce Type: cross Abstract: Efficient LLM inference research has largely focused on reducing the cost of each decoding step (e.g., using quantization, pruning, or sparse attentio

safetyarxiv-cs-cl
12 May 2026
Tutorials

Continuum Robot Modeling with Action Conditioned Flow Matching

DGX agent

arXiv:2605.09216v1 Announce Type: new Abstract: Predicting the shape of tendon driven continuum robots (TDCRs) at steady state from actuation remains challenging due to continuous deformation, complex

tutorialsarxiv-cs-ro
12 May 2026
Model Releases

CORTEG: Foundation Models Enable Cross-Modality Representation Transfer from Scalp to Intracranial Brain Recordings

DGX agent

arXiv:2605.10337v1 Announce Type: new Abstract: Intracranial electrocorticography (ECoG) offers high-signal-to-noise access to cortical activity for brain-computer interfaces, yet limited per-patient

model-releasesarxiv-cs-ai
12 May 2026
Local Ai

Data-driven Circuit Discovery for Interpretability of Language Models

DGX agent

arXiv:2605.09129v1 Announce Type: new Abstract: Circuit discovery aims to explain how language models (LMs) implement a specific task by localizing and interpreting a circuit, a computational subgraph

local-aiarxiv-cs-ai
12 May 2026
Local Ai

DetRefiner: Model-Agnostic Detection Refinement with Feature Fusion Transformer

DGX agent

arXiv:2605.10190v1 Announce Type: new Abstract: Open-vocabulary object detection (OVOD) aims to detect both seen and unseen categories, yet existing methods often struggle to generalize to novel objec

local-aiarxiv-cs-cv
12 May 2026
Model Releases

Do Foundation Model Embeddings Improve Cross-Country Crop Yield Generalisation? A Leave-One-Country-Out Evaluation in Sub-Saharan Africa

DGX agent

arXiv:2605.08113v1 Announce Type: cross Abstract: Accurate predictions of smallholder maize yields across national boundaries are critical for food security planning in sub-Saharan Africa, yet most pu

model-releasesarxiv-cs-cv
12 May 2026
Agents

EvoDriveVLA: Evolving Driving VLA Models via Collaborative Perception-Planning Distillation

DGX agent

arXiv:2603.09465v3 Announce Type: replace-cross Abstract: Vision-Language-Action models have shown great promise for autonomous driving, yet they suffer from degraded perception after unfreezing the v

agentsarxiv-cs-ai
12 May 2026
Model Releases

Explicit Reasoning Makes Better Judges: A Systematic Study on Accuracy, Efficiency, and Robustness

DGX agent

arXiv:2509.13332v2 Announce Type: replace Abstract: As Large Language Models (LLMs) are increasingly adopted as automated judges in benchmarking and reward modeling, ensuring their reliability, effici

model-releasesarxiv-cs-ai
12 May 2026
Hardware

Forcing-KV: Hybrid KV Cache Compression for Efficient Autoregressive Video Diffusion Models

DGX agent

arXiv:2605.09681v1 Announce Type: new Abstract: Autoregressive (AR) video diffusion models adopt a streaming generation framework, enabling long-horizon video generation with real-time responsiveness,

hardwarearxiv-cs-cv
12 May 2026
Tutorials

HairGPT: Strand-as-Language Autoregressive Modeling for Realistic 3D Hairstyle Synthesis

DGX agent

arXiv:2605.08824v1 Announce Type: cross Abstract: Hair is a rich medium of visual and cultural expression, yet its digital modeling remains challenging due to the duality of fluidity and structure. Ma

tutorialsarxiv-cs-cv
12 May 2026
Safety

HapticLDM: A Diffusion Model for Text-to-Vibrotactile Generation

DGX agent

arXiv:2605.09971v1 Announce Type: cross Abstract: Text-to-vibration generation converts natural language into haptic feedback, enabling vibration-effect designers to get scenarios-fitted vibrations mo

safetyarxiv-cs-ai
12 May 2026
Agents

Heteroscedastic Diffusion for Multi-Agent Trajectory Modeling

DGX agent

arXiv:2605.10717v1 Announce Type: cross Abstract: Multi-agent trajectory modeling traditionally focuses on forecasting, often neglecting more general tasks like trajectory completion, which is essenti

agentsarxiv-cs-cv
12 May 2026
Research

How Much Do Circuits Tell Us? Measuring the Consistency and Specificity of Language Model Circuits

DGX agent

arXiv:2605.08348v1 Announce Type: new Abstract: The circuits framework in mechanistic interpretability aims to identify causally important sparse subgraphs of model components, typically evaluated by

researcharxiv-cs-cl
12 May 2026
Research

Improved Mean Flows: On the Challenges of Fastforward Generative Models

DGX agent

arXiv:2512.02012v2 Announce Type: replace Abstract: MeanFlow (MF) has recently been established as a framework for one-step generative modeling. However, its ``fastforward'' nature introduces key chal

researcharxiv-cs-cv
12 May 2026
Safety

Internalizing Safety Understanding in Large Reasoning Models via Verification

DGX agent

arXiv:2605.08930v1 Announce Type: new Abstract: While explicit Chain-of-Thought (CoT) empowers large reasoning models (LRMs), it enables the generation of riskier final answers. Current alignment para

safetyarxiv-cs-ai
12 May 2026
Research

Learning Graph Foundation Models on Riemannian Graph-of-Graphs

DGX agent

arXiv:2605.09993v1 Announce Type: new Abstract: Graph foundation models (GFMs), pretrained on massive graph data, have transformed graph machine learning by supporting general-purpose reasoning across

researcharxiv-cs-lg
12 May 2026
Applications

LLaVA-CKD: Bottom-Up Cascaded Knowledge Distillation for Vision-Language Models

DGX agent

arXiv:2605.10641v1 Announce Type: cross Abstract: Large Vision-Language Models (VLMs) are successful in addressing a multitude of vision-language understanding tasks, such as Visual Question Answering

applicationsarxiv-cs-ai
12 May 2026
Research

Measuring Embedding Sensitivity to Authorial Style in French: Comparing Literary Texts with Language Model Rewritings

DGX agent

arXiv:2605.10606v1 Announce Type: cross Abstract: Large language models (LLMs) can convincingly imitate human writing styles, yet it remains unclear how much stylistic information is encoded in embedd

researcharxiv-cs-ai
12 May 2026
Research

Metacognitive Behavioral Tuning of Large Language Models for Multi-Hop Question Answering

DGX agent

arXiv:2602.22508v2 Announce Type: replace Abstract: Large Language Models (LLMs) often produce incorrect answers on multi-hop question answering even when the reasoning trace already contains a correc

researcharxiv-cs-ai
12 May 2026
Safety

Mid-Training with Self-Generated Data Improves Reinforcement Learning in Language Models

DGX agent

arXiv:2605.08472v1 Announce Type: new Abstract: The effectiveness of Reinforcement Learning (RL) in Large Language Models (LLMs) depends on the nature and diversity of the data used before and during

safetyarxiv-cs-ai
12 May 2026
Research

Mitigating Watermark Forgery in Generative Models via Randomized Key Selection

DGX agent

arXiv:2507.07871v4 Announce Type: replace-cross Abstract: Watermarking enables GenAI providers to verify whether content was generated by their models. A watermark is a hidden signal in the content, w

researcharxiv-cs-ai
12 May 2026
Research

Network-Efficient World Model Token Streaming

DGX agent

arXiv:2605.09886v1 Announce Type: new Abstract: Generative driving world models rely on compact latent state representations that must be efficiently transmitted and synchronized across distributed co

researcharxiv-cs-ro
12 May 2026
Research

NoiseRater: Meta-Learned Noise Valuation for Diffusion Model Training

DGX agent

arXiv:2605.08144v1 Announce Type: cross Abstract: Diffusion models have achieved remarkable success across a wide range of generative tasks, yet their training paradigm largely treats injected noise a

researcharxiv-cs-ai
12 May 2026
Research

Restoration-Aligned Generative Flow Models for Blind Motion Deblurring

DGX agent

arXiv:2605.08854v1 Announce Type: new Abstract: Generative flow models offer powerful priors learned from large-scale natural images, but directly adapting them to restoration tasks such as motion deb

researcharxiv-cs-cv
12 May 2026
Model Releases

SDiaReward: Modeling and Benchmarking Spoken Dialogue Rewards with Modality and Colloquialness

DGX agent

arXiv:2603.14889v2 Announce Type: replace-cross Abstract: The rapid evolution of end-to-end spoken dialogue systems demands transcending mere textual semantics to incorporate paralinguistic nuances an

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

SenseBench: A Benchmark for Remote Sensing Low-Level Visual Perception and Description in Large Vision-Language Models

DGX agent

arXiv:2605.10576v1 Announce Type: cross Abstract: Low-level visual perception underpins reliable remote sensing (RS) image analysis, yet current image quality assessment (IQA) methods output uninterpr

model-releasesarxiv-cs-ai
12 May 2026
← Previous
1…162163164165166…1030
Next →