AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,020
  • Agents7,759
  • Applications5,540
  • Concepts5
  • Hardware1,925
  • Industry6,204
  • Local Ai5,102
  • Model Releases24,783
  • Research20,783
  • Safety13,742
  • Syntheses17
  • Tools1,680
  • Tutorials3,480

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,020
  • Agents7,759
  • Applications5,540
  • Concepts5
  • Hardware1,925
  • Industry6,204
  • Local Ai5,102
  • Model Releases24,783
  • Research20,783
  • Safety13,742
  • Syntheses17
  • Tools1,680
  • Tutorials3,480

Source
HumanDGX agent

Content type
91,020Total entries
1Added by human
91,019Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
53,690 results
Model Releases

MMDG-Bench: A Benchmark for Multimodal Domain Generalization

DGX agent

arXiv:2606.00891v1 Announce Type: new Abstract: Multi-modal Domain Generalization (MMDG) seeks to leverage complementary modalities to enhance model robustness on unseen domains. Despite extensive pro

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

MOC: Multi-Order Communication in LLM-based Multi-Agent Systems

DGX agent

arXiv:2606.02359v1 Announce Type: new Abstract: Despite the remarkable progress of Large Language Model (LLM) based Multi-Agent Systems, most research focuses on optimizing coordination topology while

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

MomentKV: Closing the Directional Gap in KV Cache Eviction for Long-Context Inference

DGX agent

arXiv:2606.01563v1 Announce Type: new Abstract: Autoregressive decoding in Transformer-based language models relies on the KV cache, whose memory footprint grows linearly with sequence length and beco

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

Move the Query, Not the Cache: Characterizing Cross-Instance Latent Attention Redistribution Across GPU Fabrics

DGX agent

arXiv:2606.01502v1 Announce Type: cross Abstract: Frontier LLMs increasingly decide what a query attends to with a sparse-attention indexer that picks a few KV-cache blocks per query: attention's unit

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Multimodal Action Diffusion for Robust End-to-End Autonomous Driving

DGX agent

arXiv:2606.02105v1 Announce Type: new Abstract: End-to-End Autonomous Driving (E2E-AD) systems have largely converged on predicting intermediate trajectory waypoints, delegating final control to hand-

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Multimodal Approaches for Visually-Rich Document Type Classification: A Comparative Analysis

DGX agent

arXiv:2606.02162v1 Announce Type: cross Abstract: Document type classification in visually rich documents remains challenging, as relevant information is distributed across textual, visual, and layout

model-releasesarxiv-cs-ai
2 Jun 2026
Local Ai

Multimodal Function Vectors for Visual Relations

DGX agent

arXiv:2510.02528v2 Announce Type: replace Abstract: Large Multimodal Models (LMMs) demonstrate impressive in-context learning abilities from few multimodal demonstrations, yet the internal mechanisms

local-aiarxiv-cs-ai
2 Jun 2026
Research

Not All Explanations Simulate Equally: Comparing Verbalized Feature Attributions and Self-Generated Rationales

DGX agent

arXiv:2606.01148v1 Announce Type: new Abstract: Natural-language explanations are often treated as a unified interface for understanding model behavior, but different explanation sources may support s

researcharxiv-cs-cl
2 Jun 2026
Model Releases

OneVLA: A Unified Framework for Embodied Tasks

DGX agent

arXiv:2606.01241v1 Announce Type: new Abstract: Navigation and manipulation are fundamental capabilities of embodied intelligence, enabling robots to interpret natural language commands and interact p

model-releasesarxiv-cs-ro
2 Jun 2026
Safety

OPD+: Rethinking the Advantage Design for On-Policy Distillation

DGX agent

arXiv:2606.01039v1 Announce Type: cross Abstract: On-policy distillation (OPD) is a widely used technique to transfer capabilities from capable teacher language models to the base student models, and

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

Parameter-efficient Dual-encoder Architecture with Differentiable Choquet Integral Fusion for Underwater Acoustic Classification

DGX agent

arXiv:2606.02341v1 Announce Type: cross Abstract: Underwater acoustic classification has a wide array of oceanic applications, but faces challenges due to an increasingly complex acoustic environment.

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

Personalized 3D Myocardial Infarct Geometry Reconstruction from Cine MRI for Cardiac Digital Twins

DGX agent

arXiv:2606.01808v1 Announce Type: new Abstract: Accurate 3D geometric characterization of myocardial infarction (MI) is essential for building cardiac digital twins (CDTs) to precisely simulate infarc

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Physics from Video: Identifiability of Time-Invariant Second-Order ODEs under Minimal Trajectory Conditions

DGX agent

arXiv:2606.00115v1 Announce Type: new Abstract: Bridging the gap between visual realism and physical understanding is a core challenge for video-based world models. We study the structural identifiabi

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

POIROT: Interrogating Agents for Failure Detection in Multi-Agent Systems

DGX agent

arXiv:2606.02282v1 Announce Type: new Abstract: Orchestrating Large Language Models into Multi-Agent Systems (LLM-MAS) has unlocked remarkable reasoning capabilities, yet emergent failures and halluci

model-releasesarxiv-cs-ai
2 Jun 2026
Safety

Post-Deterministic Distributed Systems: A New Foundation for Trustworthy Autonomous Infrastructure

DGX agent

arXiv:2606.01722v1 Announce Type: cross Abstract: For decades, distributed systems have typically assumed that correct participants execute protocol-specified behavior with stable, externally defined,

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

ProtStructQA: A Denotation Threshold in Protein Structural Reasoning

DGX agent

arXiv:2606.00451v1 Announce Type: new Abstract: Protein-language systems are often evaluated by whether they generate plausible biological text, but a structural question has a sharper semantics: it d

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Quality-Diversity Evolution for Discovering Diverse Vulnerabilities in LLM Safety

DGX agent

arXiv:2606.00801v1 Announce Type: cross Abstract: Current approaches to LLM adversarial testing suffer from coverage gaps: manual red-teaming does not scale, LLM-as-attacker methods exhibit mode colla

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

ResNet-34 with Lightweight Decoder for Accurate and Efficient Segmentation of Fetal Brain MRI

DGX agent

arXiv:2606.01293v1 Announce Type: cross Abstract: Accurate segmentation of fetal brain tissues in Magnetic Resonance Imaging (MRI) is critical for early diagnosis of congenital abnormalities and impro

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Rethinking RL Evaluation: Can Benchmarks Truly Reveal Failures of RL Methods?

DGX agent

arXiv:2510.10541v2 Announce Type: replace-cross Abstract: Current benchmarks are inadequate for evaluating progress in reinforcement learning (RL) for large language models (LLMs).Despite recent bench

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

RoboStressBench: Benchmarking VLM Robustness to Physical Visual Stress in Embodied Scenes

DGX agent

arXiv:2606.00828v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have shown strong visual understanding and are increasingly deployed in embodied AI systems, where reliable perception und

model-releasesarxiv-cs-cv
2 Jun 2026
Tutorials

Robust Reasoning via Dynamic Token Selection for Distribution-Aligned Self-Distillation

DGX agent

arXiv:2606.00628v1 Announce Type: new Abstract: Self-distillation improves learning efficiency by rewriting reference answers as training data that better matches the model's own distribution. However

tutorialsarxiv-cs-cl
2 Jun 2026
Model Releases

ROGLE: Robust Global-Local Alignment with Automated Region Supervision for Text-Based Person Search

DGX agent

arXiv:2606.01825v1 Announce Type: new Abstract: Text-Based Person Search (TBPS) aims to retrieve pedestrian images using natural language queries. However, existing TBPS models, especially those based

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

ROGUE: Misaligned Agent Behavior Arising from Ordinary Computer Use

DGX agent

arXiv:2606.00341v1 Announce Type: cross Abstract: As AI agents are increasingly deployed in real personal and corporate settings (email accounts, development workflows, company databases, etc.), safet

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

RoleCDE:Benchmarking and Mitigating Role-Alignment Trade-offs in Role-Playing Agents

DGX agent

arXiv:2606.01552v1 Announce Type: new Abstract: Role-playing agents(RPAs) are widely used to steer large language models(LLMs) toward role-consistent behavior, yet existing benchmarks mainly evaluate

model-releasesarxiv-cs-ai
2 Jun 2026
Safety

SafeMCP: Proactive Power Regulation for LLM Agent Defense via Environment-Grounded Look-Ahead Reasoning

DGX agent

arXiv:2606.01991v1 Announce Type: new Abstract: As Large Language Model (LLM) agents increasingly leverage the Model Context Protocol (MCP) to operate in complex environments, the expansion of their a

safetyarxiv-cs-ai
2 Jun 2026
Tutorials

SCL: Towards Domain Generalization via Single-Temporal Multimodal Contrastive Learning for Remote Sensing Change Detection

DGX agent

arXiv:2404.11326v5 Announce Type: replace Abstract: In recent years, change detection and anomaly detection models based on CNN and transformer have achieved remarkable success across various datasets

tutorialsarxiv-cs-cv
2 Jun 2026
Research

Score imes Decoder: A Unified View of Unsupervised Inference-Time Scaling for Hallucination Mitigation

DGX agent

arXiv:2606.00739v1 Announce Type: new Abstract: Large language models hallucinate even when the answer lies within their parameters. While inference-time scaling can surface this latent knowledge, the

researcharxiv-cs-lg
2 Jun 2026
Agents

Seq-DeepIPC: Sequential Sensing for End-to-End Control in Legged Robot Navigation

DGX agent

arXiv:2510.23057v2 Announce Type: replace-cross Abstract: We present Seq-DeepIPC, a sequential end-to-end perception-to-control model for legged robot navigation in real-world environments. Seq-DeepIP

agentsarxiv-cs-cv
2 Jun 2026
Model Releases

SHARP: Sleep-based Hierarchical Accelerated Replay for Long Range Non-Stationary Temporal Pattern Recognition

DGX agent

arXiv:2606.00732v1 Announce Type: new Abstract: Learning long-range non-stationary temporal patterns remains a core challenge for modern sequence models, particularly in strict streaming settings. In

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Short-form Text Rewriting with Phi Silica

DGX agent

arXiv:2606.00462v1 Announce Type: cross Abstract: Short-form text rewriting is a constrained variant of paraphrasing in which limited context and high semantic density leave little room for variation.

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

SkillAdaptor: Self-Adapting Skills for LLM Agents from Trajectories

DGX agent

arXiv:2606.01311v1 Announce Type: cross Abstract: Large language model (LLM) agents increasingly rely on reusable external skills to solve long-horizon interactive tasks. Existing training-free skill

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

SMH-Bench: Benchmarking LLM Agents for Environment-Grounded Reasoning and Action in Smart Homes

DGX agent

arXiv:2606.01912v1 Announce Type: new Abstract: Smart homes are evolving toward complex state-dependent living environments, requiring Large Language Models (LLMs) to reason over user intent, preferen

model-releasesarxiv-cs-ai
2 Jun 2026
Agents

SpaceTools: Tool-Augmented Spatial Reasoning via Double Interactive RL

DGX agent

arXiv:2512.04069v2 Announce Type: replace Abstract: Vision Language Models (VLMs) demonstrate strong qualitative visual understanding, but struggle with metrically precise spatial reasoning required f

agentsarxiv-cs-cv
2 Jun 2026
Model Releases

Temporally-Aligned Evaluation for Audio-Driven Talking Head Generation

DGX agent

arXiv:2606.01031v1 Announce Type: cross Abstract: Audio-driven talking-head generation has advanced rapidly, yet existing evaluation protocols mainly rely on frame-wise metrics that assume strict temp

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

TempRet: Temporal Enhancement and Two-Stage Reranking for CVPR 2026 EPIC-KITCHENS-100 Multi-Instance Retrieval Challenge

DGX agent

arXiv:2605.24470v2 Announce Type: replace Abstract: Video-text retrieval has witnessed remarkable progress driven by large-scale vision-language pretraining, yet most existing approaches inherit an im

model-releasesarxiv-cs-cv
2 Jun 2026
Research

The Role of Ambiguity in Error Prediction via Uncertainty Quantification

DGX agent

arXiv:2606.02093v1 Announce Type: cross Abstract: The task of Error Prediction, namely predicting whether a model output is correct, is commonly tackled with Uncertainty Quantification (UQ). However,

researcharxiv-cs-ai
2 Jun 2026
Model Releases

TimeSage-MT: A Multi-Turn Benchmark for Evaluating Agentic Time Series Reasoning

DGX agent

arXiv:2606.01498v1 Announce Type: cross Abstract: Time series data inform critical decisions across many real-world domains. While large language model (LLM) agents can analyze data through natural la

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Towards Simple and Provable Parameter-Free Adaptive Gradient Methods

DGX agent

arXiv:2412.19444v2 Announce Type: replace Abstract: Optimization algorithms such as AdaGrad and Adam have significantly advanced the training of deep models by dynamically adjusting the learning rate

model-releasesarxiv-cs-lg
2 Jun 2026
Research

TriLens: Per-Layer Logit-Lens Entropy for White-Box Hallucination Detection

DGX agent

arXiv:2606.01033v1 Announce Type: new Abstract: When a language model hallucinates, the final answer is wrong, but the mistake is not necessarily invisible inside the model. Different internal pathway

researcharxiv-cs-ai
2 Jun 2026
Model Releases

URDF-Anything+: End-to-End Generation for Simulation-Ready Articulated Assets

DGX agent

arXiv:2603.14010v2 Announce Type: replace Abstract: Articulated objects are fundamental for robotics, simulation of physics, and interactive virtual environments. However, recovering them from visual

model-releasesarxiv-cs-ro
2 Jun 2026
Model Releases

Value Flows

DGX agent

arXiv:2510.07650v4 Announce Type: replace-cross Abstract: While most reinforcement learning methods today flatten the distribution of future returns to a single scalar value, distributional RL methods

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

What Do LLMs Know About Alzheimer's Disease? Multi-loss Fine-Tuning and Probing for AD Detection

DGX agent

arXiv:2602.11177v2 Announce Type: replace-cross Abstract: Reliable early detection of Alzheimer's disease (AD) is challenging, particularly due to the limited availability of labeled data. While large

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

When Parallelism Pays Off: Cohesion-Aware Task Partitioning for Multi-Agent Coding

DGX agent

arXiv:2606.00953v1 Announce Type: new Abstract: Multi-agent Large Language Model (LLM) systems offer a way to decompose complex tasks, such as coding, through parallelization and context isolation. Ho

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

When Safe Skills Collide: Measuring Compositional Risk in Agent Skill Ecosystems

DGX agent

arXiv:2606.00448v1 Announce Type: cross Abstract: LLM agents increasingly rely on community-contributed skills that expand an agent's operational capability set. We study a core safety problem in agen

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

You Can Learn Tokenization End-to-End with Reinforcement Learning

DGX agent

arXiv:2602.13940v2 Announce Type: replace-cross Abstract: Tokenization is a hardcoded compression step which remains in the training pipeline of Large Language Models (LLMs), despite a general trend t

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

A Kinetic Energy Perspective of Flow Matching

DGX agent

arXiv:2602.07928v2 Announce Type: replace-cross Abstract: Flow-based generative models can be viewed through a physics lens: sampling transports a particle from noise to data by integrating a learned

model-releasesarxiv-cs-ai
1 Jun 2026
Research

A Padding Method for Enhanced Encoding of Inorganic Structures with Varying Chemical Compositions

DGX agent

arXiv:2605.30743v1 Announce Type: cross Abstract: Designing novel inorganic materials through generative models remains an important challenge for material science, driven by the complexity and divers

researcharxiv-cs-cl
1 Jun 2026
Model Releases

A Visually Impaired Assistance Benchmark for VLM-as-a-Judge Evaluation

DGX agent

arXiv:2605.31351v1 Announce Type: new Abstract: AI-based Visually Impaired Assistance (VIA) remains challenging, largely due to the high cost of human evaluation. The VLM-as-a-Judge paradigm may offer

model-releasesarxiv-cs-cl
1 Jun 2026
← Previous
1…435436437438439…1119
Next →