AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,573
  • Agents7,495
  • Applications5,364
  • Concepts5
  • Hardware1,813
  • Industry6,149
  • Local Ai4,892
  • Model Releases23,569
  • Research19,967
  • Safety13,263
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,573
  • Agents7,495
  • Applications5,364
  • Concepts5
  • Hardware1,813
  • Industry6,149
  • Local Ai4,892
  • Model Releases23,569
  • Research19,967
  • Safety13,263
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

Content type
87,573Total entries
1Added by human
87,572Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
51,512 results
Model Releases

SpaceDrive: Infusing Spatial Awareness into VLM-based Autonomous Driving

DGX agent

arXiv:2512.10719v2 Announce Type: replace Abstract: End-to-end autonomous driving methods built on vision language models (VLMs) have undergone rapid development driven by their universal visual under

model-releasesarxiv-cs-cv
22 May 2026
Tutorials
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

TextTeacher: What Can Language Teach About Images?

DGX agent

arXiv:2605.22098v1 Announce Type: new Abstract: The platonic representation hypothesis suggests that sufficiently large models converge to a shared representation geometry, even across modalities. Mot

tutorialsarxiv-cs-cv
22 May 2026
Research

Token-weighted Direct Preference Optimization with Attention

DGX agent

arXiv:2605.21883v1 Announce Type: new Abstract: Direct Preference Optimization (DPO) aligns Large Language Models with human preferences without the need for a separate reward model. However, DPO trea

researcharxiv-cs-cl
22 May 2026
Model Releases

VGenST-Bench: A Benchmark for Spatio-Temporal Reasoning via Active Video Synthesis

DGX agent

arXiv:2605.22570v1 Announce Type: new Abstract: Spatio-temporal reasoning is a core capability for Multimodal Large Language Models (MLLMs) operating in the real world. As such, evaluating it precisel

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

When Cases Get Rare: A Retrieval Benchmark for Off-Guideline Clinical Question Answering

DGX agent

arXiv:2605.21807v1 Announce Type: new Abstract: Across medical specialties, clinical practice is anchored in evidence-based guidelines that codify best studied diagnostic and treatment pathways. These

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

A Free Lunch in LLM Compression: Revisiting Retraining after Pruning

DGX agent

arXiv:2510.14444v3 Announce Type: replace Abstract: Post-training pruning can substantially reduce LLM inference costs, but it often degrades quality unless the remaining weights are adapted. Since gl

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Anatomy of Agentic Memory: Taxonomy and Empirical Analysis of Evaluation and System Limitations

DGX agent

arXiv:2602.19320v2 Announce Type: replace Abstract: Agentic memory systems enable large language model (LLM) agents to maintain state across long interactions, supporting long-horizon reasoning and pe

model-releasesarxiv-cs-cl
21 May 2026
Research

Can Conversational XAI Improve User Performance? An Experimental Study

DGX agent

arXiv:2605.20439v1 Announce Type: new Abstract: Explainable AI (XAI) techniques aim to provide insights into predictive models and enhance user performance, yet they often fall short of these expectat

researcharxiv-cs-lg
21 May 2026
Model Releases

Causal Path Alignment: Anchoring the Optimization Trajectory for Controllable In-Parameter Knowledge Editing

DGX agent

arXiv:2506.04042v2 Announce Type: replace Abstract: Knowledge editing is pivotal for efficiently updating the parametric memory of Large Language Models (LLMs), enabling them to function as evolving a

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Dynamic Video Generation: Shaping Video Generation Across Time and Space

DGX agent

arXiv:2605.21042v1 Announce Type: new Abstract: Diffusion models have achieved impressive performance in video generation, but their iterative denoising process remains computationally expensive due t

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

Exploring Deep Learning and Ultra-Widefield Imaging for Diabetic Retinopathy and Macular Edema

DGX agent

arXiv:2603.08235v2 Announce Type: replace Abstract: Diabetic retinopathy (DR) and diabetic macular edema (DME) are leading causes of preventable blindness among working-age adults. Traditional approac

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy

DGX agent

arXiv:2605.20965v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) have shown remarkable performance on a wide range of vision-language tasks. Despite this progress, they are still p

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

FT-Dojo: Towards Autonomous LLM Fine-Tuning with Language Agents

DGX agent

arXiv:2603.01712v2 Announce Type: replace-cross Abstract: Fine-tuning large language models for vertical domains remains labor-intensive, requiring practitioners to curate data, configure training, an

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

How Much Online RL is Enough? Informative Rollouts for Offline Preference Optimization in RLVR

DGX agent

arXiv:2605.21266v1 Announce Type: new Abstract: Reinforcement Learning from Verifiable Rewards (RLVR) has emerged as a powerful paradigm for reasoning in language models, with GRPO as its primary exam

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Learning Query-Aware Budget-Tier Routing for Runtime Agent Memory

DGX agent

arXiv:2602.06025v2 Announce Type: replace Abstract: Memory is increasingly central to Large Language Model (LLM) agents operating beyond a single context window, yet most existing systems rely on offl

model-releasesarxiv-cs-cl
21 May 2026
Local Ai

LoCar: Localization-Aware Evaluation of In-Vehicle Assistants through Fine-Grained Sociolinguistic Control

DGX agent

arXiv:2605.21086v1 Announce Type: new Abstract: While Large Language Models (LLMs) are increasingly integrated into in-vehicle conversational systems, identifying the optimal model remains challenging

local-aiarxiv-cs-cl
21 May 2026
Model Releases

Mechanics of Bias and Reasoning: Interpreting the Impact of Chain-of-Thought Prompting on Gender Bias in LLMs

DGX agent

arXiv:2605.20410v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed in socially sensitive settings despite substantial documentation that they encode gender biases.

model-releasesarxiv-cs-cl
21 May 2026
Safety

Mechanistic Interpretability for Learning Assurance of a Vision-Based Landing System

DGX agent

arXiv:2605.20607v1 Announce Type: cross Abstract: EASA's learning-assurance guidance requires data-driven aviation systems to build and monitor their own situation representation, yet for neural netwo

safetyarxiv-cs-cv
21 May 2026
Model Releases

PGC: Peak-Guided Calibration for Generalizable AI-Generated Image Detection

DGX agent

arXiv:2605.21207v1 Announce Type: new Abstract: The rapid evolution of generative AI, from GANs to modern diffusion models, has resulted in increasingly subtle discriminative clues. These fine-grained

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

QwenSafe: Multimodal Content Rating Description Identification via Preference-Aligned VLMs

DGX agent

arXiv:2605.20584v1 Announce Type: new Abstract: Mobile app marketplaces require developers to disclose standardized content rating descriptors (CRDs) to inform users about potentially sensitive or res

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution

DGX agent

arXiv:2605.21195v1 Announce Type: new Abstract: Discrete autoregressive (AR) text-to-image (T2I) models pair a VQ tokenizer with an AR policy, and current post-training pipelines optimize only the pol

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

Refining and Reusing Annotation Guidelines for LLM Annotation

DGX agent

arXiv:2605.20809v1 Announce Type: new Abstract: While Large Language Models (LLMs) demonstrate remarkable performance on zero-shot annotation tasks, they often struggle with the specialized convention

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

SOLAR: A Self-Optimizing Open-Ended Autonomous Agent for Lifelong Learning and Continual Adaptation

DGX agent

arXiv:2605.20189v1 Announce Type: cross Abstract: Despite the remarkable success of large language models (LLMs), they still face bottlenecks while deploying in dynamic, real-world settings with prima

model-releasesarxiv-cs-lg
21 May 2026
Local Ai

Sustainability Is Not Linear: Quantifying Performance, Energy, and Privacy Trade-offs in On-Device Intelligence

DGX agent

arXiv:2603.26603v2 Announce Type: replace-cross Abstract: The migration of Large Language Models (LLMs) from cloud clusters to edge devices promises enhanced privacy and offline accessibility, but thi

local-aiarxiv-cs-lg
21 May 2026
Model Releases

TASTE: A Designer-Annotated Multi-Dimensional Preference Dataset for AI-Generated Graphic Design

DGX agent

arXiv:2605.20731v1 Announce Type: new Abstract: Text-to-image models produce graphic design at production scale, but their supervision comes from photo-style preference data with a single overall verd

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

Terminal-World: Scaling Terminal-Agent Environments via Agent Skills

DGX agent

arXiv:2605.20876v1 Announce Type: new Abstract: Terminal agents extend Large Language Models with the ability to execute tasks directly in command-line environments, but their progress is bottlenecked

model-releasesarxiv-cs-cl
21 May 2026
Research

TreeText-CTS: Compact, Source-Traceable Tree-Path Evidence for Irregular Clinical Time-Series Prediction

DGX agent

arXiv:2605.20292v1 Announce Type: new Abstract: Numerical time-series models can effectively process irregular electronic health record (EHR) trajectories, but they do not naturally expose the measure

researcharxiv-cs-lg
21 May 2026
Safety

What Semantics Survive the Connector? Diagnosing VLM-to-DiT Alignment in Video Editing

DGX agent

arXiv:2605.20795v1 Announce Type: new Abstract: Flow matching based video generative models have been increasingly relying on prepended Vision-Language Models (VLMs) to handle complex, instruction-bas

safetyarxiv-cs-cv
21 May 2026
Model Releases

A Bitter Lesson for Data Filtering

DGX agent

arXiv:2605.19407v1 Announce Type: cross Abstract: We investigate data filtering for large model pretraining via new scaling studies that target the high compute, data-scarce regime. In spite of an app

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

A Reproducibility Analysis of PO4ISR: Diagnosing and Mitigating Semantic Drift in LLM-Based Session Recommendation

DGX agent

arXiv:2605.18780v1 Announce Type: cross Abstract: Reasoning-based Large Language Models (LLMs) like PO4ISR have set new benchmarks in session-based recommendation. However, the reproducibility of thei

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Aero-World: Action-Conditioned Aerial Video Generation from Inertial Controls

DGX agent

arXiv:2605.19728v1 Announce Type: new Abstract: Foundation video models produce visually impressive results, but their use in embodied AI remains limited because they are primarily trained on natural

model-releasesarxiv-cs-cv
20 May 2026
Agents

AQuaUI: Visual Token Reduction for GUI Agents with Adaptive Quadtrees

DGX agent

arXiv:2605.19260v1 Announce Type: new Abstract: Large Multimodal Models (LMMs) have recently emerged as promising backbones for GUI-agent models, where high-resolution GUI screenshots are introduced t

agentsarxiv-cs-ai
20 May 2026
Research

Block-Based Double Decoders

DGX agent

arXiv:2605.18807v1 Announce Type: cross Abstract: Encoder-decoder models offer substantial inference-time savings over decoder-only models, but their pretraining objectives suffer from sparse supervis

researcharxiv-cs-ai
20 May 2026
Research

Cross-Paradigm Knowledge Distillation: A Comprehensive Study of Bidirectional Transfer Between Random Forests and Deep Neural Networks for Big Data Applications

DGX agent

arXiv:2605.19299v1 Announce Type: new Abstract: The exponential growth of big data has intensified the need for efficient and interpretable machine learning models that can handle diverse data charact

researcharxiv-cs-lg
20 May 2026
Tutorials

Diffusion and Flow-based Copulas: Forgetting and Remembering Dependencies

DGX agent

arXiv:2509.19707v2 Announce Type: replace-cross Abstract: Copulas are a fundamental tool for modelling multivariate dependencies in data, forming the method of choice in diverse fields and application

tutorialsarxiv-cs-lg
20 May 2026
Model Releases

EgoBabyVLM: Benchmarking Cross-Modal Learning from Naturalistic Egocentric Video Data

DGX agent

arXiv:2605.19130v1 Announce Type: cross Abstract: Children acquire language grounding with remarkable robustness from limited visuo-linguistic input in ways that surpass today's best large multimodal

model-releasesarxiv-cs-ai
20 May 2026
Safety

ESLD (External Surrogate Latent Defense): A Latent-Space Architecture for Faster, Stronger Prompt-Injection Defense

DGX agent

arXiv:2605.18918v1 Announce Type: cross Abstract: Modern AI assistants are agentic. To answer a single user request, the underlying language model pulls in information from many sources, such as web s

safetyarxiv-cs-ai
20 May 2026
Model Releases

FedMental: Evaluating Federated Learning for Mental Health Detection from Social Media Data

DGX agent

arXiv:2605.18936v1 Announce Type: cross Abstract: Social media text data are often used to train Machine Learning (ML) models to identify users exhibiting high-risk mental health behaviors. However, s

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

From Prompts to Pavement Through Time: Temporal Grounding in Agentic Scene-to-Plan Reasoning

DGX agent

arXiv:2605.19824v1 Announce Type: new Abstract: Recent attempts to support high-level scene interpretation and planning in Autonomous Vehicles (AVs) using ensembles of Large Language Models (LLMs) and

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

GRAB: A Risk Taxonomy--Grounded Benchmark for Unsupervised Topic Discovery in Financial Disclosures

DGX agent

arXiv:2509.21698v2 Announce Type: replace Abstract: Risk categorization in 10-K risk disclosures matters for oversight and investment, yet no public benchmark evaluates unsupervised topic models for t

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

Hallucination as Exploit: Evidence-Carrying Multimodal Agents

DGX agent

arXiv:2605.19192v1 Announce Type: new Abstract: Multimodal agents use screenshots, documents, and webpages to choose tool calls. When a false visual claim triggers a click, email, extraction, or trans

model-releasesarxiv-cs-ai
20 May 2026
Safety

How Does Overparameterization Affect Machine Unlearning of Deep Neural Networks?

DGX agent

arXiv:2503.08633v2 Announce Type: replace Abstract: Machine unlearning is the task of updating a trained model to forget specific training data without retraining from scratch. In this paper, we inves

safetyarxiv-cs-lg
20 May 2026
Local Ai

INSIGHTS: Demonstration-Based Summaries of Time Series Predictors

DGX agent

arXiv:2605.18849v1 Announce Type: cross Abstract: Explainability methods have progressed rapidly, but global explanations for time-series models remain underdeveloped, with most approaches focusing on

local-aiarxiv-cs-ai
20 May 2026
Model Releases

Learn-by-Wire Training Control Governance: Bounded Autonomous Training Under Stress for Stability and Efficiency

DGX agent

arXiv:2605.19008v1 Announce Type: new Abstract: Modern language-model training is increasingly exposed to instability, degraded runs, and wasted compute, especially under aggressive learning-rate, sca

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Measuring Safety Alignment Effects in Autonomous Security Agents

DGX agent

arXiv:2605.19722v1 Announce Type: cross Abstract: Do stock safety-aligned language models and their uncensored or abliterated derivatives behave differently when run as autonomous security agents? Sin

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

MSAVBench: Towards Comprehensive and Reliable Evaluation of Multi-Shot Audio-Video Generation

DGX agent

arXiv:2605.20183v1 Announce Type: new Abstract: Video generation is rapidly evolving from single-shot synthesis to complex multi-shot audio-video (MSAV) narratives to meet real-world demands. However,

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

Not All Tokens Are Worth Caching: Learning Semantic-Aware Eviction for LLM Prefix Caches

DGX agent

arXiv:2605.18825v1 Announce Type: new Abstract: Prefix caching is a key optimization in Large Language Model (LLM) serving, reusing attention Key-Value (KV) states across requests with shared prompt p

model-releasesarxiv-cs-lg
20 May 2026
Local Ai

Perceptual misalignment of texture representations in convolutional neural networks

DGX agent

arXiv:2604.01341v2 Announce Type: replace Abstract: Mathematical modeling of visual textures traces back to Julesz's intuition that texture perception in humans is based on local correlations between

local-aiarxiv-cs-cv
20 May 2026
← Previous
1…366367368369370…1074
Next →