AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,483
  • Agents7,562
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,175
  • Local Ai4,939
  • Model Releases23,953
  • Research20,128
  • Safety13,373
  • Syntheses17
  • Tools1,677
  • Tutorials3,405

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,483
  • Agents7,562
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,175
  • Local Ai4,939
  • Model Releases23,953
  • Research20,128
  • Safety13,373
  • Syntheses17
  • Tools1,677
  • Tutorials3,405

Source
HumanDGX agent

Content type
88,483Total entries
1Added by human
88,482Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
51,935 results
Safety

A KL-regularization Framework for Learning to Plan with Adaptive Priors

DGX agent

arXiv:2510.04280v2 Announce Type: replace-cross Abstract: Effective exploration remains a central challenge in model-based reinforcement learning (MBRL), particularly in high-dimensional continuous co

safetyarxiv-cs-ro
22 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

A Task-Agnostic Algebraic Integrity Metric for Event-Camera Streams Toward SOTIF-Compliant Perception using Pearson Correlation Coefficient

DGX agent

arXiv:2605.21500v1 Announce Type: cross Abstract: Event cameras have emerged as a high-bandwidth, low-latency sensing modality for safety-critical perception in automated driving systems (ADS), offeri

model-releasesarxiv-cs-cv
22 May 2026
Applications

AVI-HT: Adaptive Vision-IMU Fusion for 3D Hand Tracking

DGX agent

arXiv:2605.21714v1 Announce Type: new Abstract: We present AVI-HT, an adaptive visual-IMU fusion approach for tracking 3D hand poses by jointly modeling the egocentric image with on-glove 6-DoF IMU si

applicationsarxiv-cs-cv
22 May 2026
Model Releases

ChronoMedKG: A Temporally-Grounded Biomedical Knowledge Graph and Benchmark for Clinical Reasoning

DGX agent

arXiv:2605.22734v1 Announce Type: new Abstract: Biomedical knowledge graphs (KGs) treat disease associations as static facts, but temporal information is crucial for clinical reasoning, e.g., a sympto

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

COCOTree: A Dataset and Benchmark for Open Tree-Structured Visual Decomposition

DGX agent

arXiv:2605.22068v1 Announce Type: new Abstract: We formalize and enable the task of open tree decomposition, which segments an image into hierarchical trees of visual components with unconstrained gra

model-releasesarxiv-cs-cv
22 May 2026
Applications

Codec-Robust Attacks on Audio LLMs

DGX agent

arXiv:2605.20519v1 Announce Type: cross Abstract: Prior attacks on Audio Large Language Models (Audio LLMs) demonstrated that carefully crafted waveform-domain perturbations can force targeted adversa

applicationsarxiv-cs-ai
22 May 2026
Model Releases

Cohesion-6K: An Arabic Dataset for Analyzing Social Cohesion and Conflict in Online Discourse

DGX agent

arXiv:2605.22447v1 Announce Type: new Abstract: The study of online discourse has become central to understanding societal polarization. While much research has focused on detecting overt toxicity, th

model-releasesarxiv-cs-cl
22 May 2026
Research

D3Seg: Dependency-Aware Diffusion for Brain Tumor Segmentation with Missing Modalities

DGX agent

arXiv:2605.22249v1 Announce Type: new Abstract: Accurate brain tumor segmentation using multiparametric MRI is critical for effective treatment planning. However, in clinical settings, complete acquis

researcharxiv-cs-cv
22 May 2026
Research

Entropy-Guided Self-Supervised Learning for Medical Image Classification

DGX agent

arXiv:2605.21970v1 Announce Type: cross Abstract: Accurate and robust medical image classification is paramount for early disease diagnosis and treatment planning. However, challenges such as limited

researcharxiv-cs-cv
22 May 2026
Model Releases

FashionLens: Toward Versatile Fashion Image Retrieval via Task-Adaptive Learning

DGX agent

arXiv:2605.22552v1 Announce Type: new Abstract: Fashion image retrieval is a cornerstone of modern e-commerce systems. A unified framework that supports diverse query formats and search intentions is

model-releasesarxiv-cs-cv
22 May 2026
Research

Geometry-Adaptive Explainer for Faithful Dictionary-Based Interpretability under Distribution Shift

DGX agent

arXiv:2605.21849v1 Announce Type: cross Abstract: Mechanistic interpretability aims to explain a model's behavior by identifying causally responsible internal structures. Dictionary-based explainers s

researcharxiv-cs-cl
22 May 2026
Model Releases

HyLoVQA: Dynamic Hypernetwork-Generated Low-Rank Adaptation for Continual Visual Question Answering

DGX agent

arXiv:2605.22035v1 Announce Type: cross Abstract: Continual Visual Question Answering (VQA) requires learning from non-stationary streams of visual inputs and questions while preserving past knowledge

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

InfVSR: Breaking Length Limits of Generic Video Super-Resolution

DGX agent

arXiv:2510.00948v2 Announce Type: replace Abstract: Real-world videos often extend over thousands of frames. Existing generative video super-resolution (VSR) approaches, however, face two persistent c

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

MLLMs Know When Before Speaking: Revealing and Recovering Temporal Grounding via Attention Cues

DGX agent

arXiv:2605.21954v1 Announce Type: new Abstract: Video temporal grounding (VTG), which localizes the start and end times of a queried event in an untrimmed video, is a key test of whether multimodal la

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

MM-Conv: A Multimodal Dataset and Benchmark for Context-Aware Grounding in 3D Dialogue

DGX agent

arXiv:2605.21796v1 Announce Type: cross Abstract: Grounding language in the physical world requires AI systems to interpret references that emerge dynamically during conversation. While current vision

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

MOTOR: A Multimodal Dataset for Two-Wheeler Rider Behavior Understanding

DGX agent

arXiv:2605.22550v1 Announce Type: new Abstract: Two-wheelers account for a disproportionately high share of road fatalities in the Global South. Research on two-wheeler rider behavior, however, lags f

model-releasesarxiv-cs-cv
22 May 2026
Safety

NaviAgent: Graph-Driven Bilevel Planning for Scalable Tool Orchestration

DGX agent

arXiv:2506.19500v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) increasingly act as function-call agents that invoke external tools to tackle tasks beyond their static knowledge

safetyarxiv-cs-cl
22 May 2026
Safety

Open-source LLMs administer maximum electric shocks in a Milgram-like obedience experiment

DGX agent

arXiv:2605.21401v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed as autonomous agents that make sequences of decisions over extended interactions in high-stakes

safetyarxiv-cs-ai
22 May 2026
Model Releases

PromptNCE: Pointwise Mutual Information Predictions Using Only LLMs and Contrastive Estimation Prompts

DGX agent

arXiv:2605.21776v1 Announce Type: new Abstract: Estimating mutual information from text usually requires training a task-specific critic, which limits its use in low-data settings. We ask whether larg

model-releasesarxiv-cs-cl
22 May 2026
Applications

Quantizing Whisper-small: How design choices affect ASR performance

DGX agent

arXiv:2511.08093v2 Announce Type: replace-cross Abstract: Large speech recognition models like Whisper-small achieve high accuracy but are difficult to deploy on edge devices due to their high computa

applicationsarxiv-cs-cl
22 May 2026
Local Ai

SegCompass: Exploring Interpretable Alignment with Sparse Autoencoders for Enhanced Reasoning Segmentation

DGX agent

arXiv:2605.22658v1 Announce Type: new Abstract: While large language models provide strong compositional reasoning, existing reasoning segmentation pipelines fail to transparently connect this reasoni

local-aiarxiv-cs-cv
22 May 2026
Model Releases

SFN-YOLO: Towards Free-Range Poultry Detection via Scale-aware Fusion Networks

DGX agent

arXiv:2509.17086v2 Announce Type: replace Abstract: Detecting and localizing poultry is essential for advancing smart poultry farming. Despite the progress of detection-centric methods, challenges per

model-releasesarxiv-cs-cv
22 May 2026
Agents

SpecHop: Continuous Speculation for Accelerating Multi-Hop Retrieval Agents

DGX agent

arXiv:2605.21965v1 Announce Type: new Abstract: Large language models increasingly use external tools such as web search and document retrieval to solve information-intensive tasks. However, multi-hop

agentsarxiv-cs-cl
22 May 2026
Research

ST-SimDiff: Balancing Spatiotemporal Similarity and Difference for Efficient Video Understanding with MLLMs

DGX agent

arXiv:2605.22158v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) face significant computational overhead when processing long videos due to the massive number of visual token

researcharxiv-cs-cv
22 May 2026
Model Releases

Steins;Gate Drive: Semantic Safety Arbitration over Structured Futures for Latency-Decoupled LLM Planning

DGX agent

arXiv:2605.22456v1 Announce Type: new Abstract: Cloud-hosted LLM driver agents provide useful semantic judgments, but their inference latency exceeds stepwise vehicle-control windows. Learned world mo

model-releasesarxiv-cs-ro
22 May 2026
Model Releases

SURGE: An Event-Centric Social Media Sentiment Time Series Benchmark with Interaction Structure

DGX agent

arXiv:2605.21198v1 Announce Type: cross Abstract: Public events on social media generate large volumes of discussion whose collective dynamics carry direct value for opinion forecasting and crisis res

model-releasesarxiv-cs-ai
22 May 2026
Model Releases

Training-Trajectory-Aware Token Selection

DGX agent

arXiv:2601.10348v2 Announce Type: replace Abstract: Efficient distillation is a key pathway for converting expensive reasoning capability into deployable efficiency, yet in the frontier regime where t

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

TransitLM: A Large-Scale Dataset and Benchmark for Map-Free Transit Route Generation

DGX agent

arXiv:2605.22355v1 Announce Type: new Abstract: Public transit route planning traditionally depends on structured map infrastructure and complex routing engines, and no existing dataset supports train

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

Two is better than one: A Collapse-free Multi-Reward RLIF Training Framework

DGX agent

arXiv:2605.22620v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has substantially improved the reasoning ability of LLMs, but often depends on external supervis

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

UniVL: Unified Vision-Language Embedding for Spatially Grounded Contextual Image Generation

DGX agent

arXiv:2605.21611v1 Announce Type: new Abstract: We introduce spatially grounded contextual image generation, a controllable image generation task that reframes the conditioning paradigm. Instead of su

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

APEX: Autonomous Policy Exploration for Self-Evolving LLM Agents

DGX agent

arXiv:2605.21240v1 Announce Type: new Abstract: LLM agents have shown strong performance across a wide range of complex tasks, including interactive environments that require long-horizon decision mak

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

APM: Evaluating Style Personalization in LLMs with Arbitrary Preference Mappings

DGX agent

arXiv:2605.21063v1 Announce Type: new Abstract: Typical LLM responses tend to follow a default style, even though users often have distinct preferences regarding tone, verbosity, and formality that th

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Batched Single-Index Global Multi-Armed Bandits with Covariates

DGX agent

arXiv:2503.00565v3 Announce Type: replace-cross Abstract: The multi-armed bandits (MAB) framework is a widely used approach for sequential decision-making, where a decision-maker selects an arm in eac

model-releasesarxiv-cs-lg
21 May 2026
Research

Beyond Routing: Characterising Expert Tuning and Representation in Vision Mixture-of-Experts

DGX agent

arXiv:2605.20610v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) models are often interpreted by analysing which categories are routed to which experts. However, routing alone does not reveal

researcharxiv-cs-cv
21 May 2026
Research

Block-Sparse Global Attention for Efficient Multi-View Geometry Transformers

DGX agent

arXiv:2509.07120v2 Announce Type: replace Abstract: Efficient and accurate feed-forward multi-view reconstruction has long been an important task in computer vision. Recent transformer-based models li

researcharxiv-cs-cv
21 May 2026
Safety

Choose Wisely and Privately: Proactive Client Selection for Fair and Efficient Federated Learning

DGX agent

arXiv:2605.20975v1 Announce Type: new Abstract: Federated Learning enables collaborative model training across decentralized data sources without data transfer. Averaging-based FL is limited by the pr

safetyarxiv-cs-lg
21 May 2026
Model Releases

ChunkFT: Byte-Streamed Optimization for Memory-Efficient Full Fine-Tuning

DGX agent

arXiv:2605.21177v1 Announce Type: cross Abstract: This work presents extsc{ChunkFT}, a memory-efficient fine-tuning framework that reformulates full-parameter fine-tuning around a dynamically activate

model-releasesarxiv-cs-cl
21 May 2026
Agents

Code Generation by Differential Test Time Scaling

DGX agent

arXiv:2605.20473v1 Announce Type: cross Abstract: Test-time scaling has emerged as a promising approach for improving code generation by exploring large solution spaces at inference time. However, exi

agentsarxiv-cs-lg
21 May 2026
Safety

Cross-lingual robustness of LLM-brain alignment and its computational roots

DGX agent

arXiv:2605.21049v1 Announce Type: new Abstract: Large language models (LLMs) reliably predict neural activity during language comprehension and transformer depth has been interpreted as mirroring hier

safetyarxiv-cs-cl
21 May 2026
Safety

Cumulative Meta-Learning from Active Learning Queries for Robustness to Spurious Correlations

DGX agent

arXiv:2605.20771v1 Announce Type: new Abstract: Spurious correlations in real-world datasets cause machine learning models to rely on irrelevant patterns, undermining reliability, generalization, and

safetyarxiv-cs-lg
21 May 2026
Model Releases

Deltaynamics: Language-Based Representation for Inferring Rigid-Body Dynamics From Videos

DGX agent

arXiv:2605.20576v1 Announce Type: new Abstract: Inferring rigid-body physical states and properties from monocular videos is a fundamental step toward physics-based perception and simulation. Existing

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

Diagnosing Overhead in Dispatch Operations: Cross-architecture Observatory

DGX agent

arXiv:2605.20982v1 Announce Type: cross Abstract: AlltoAll dispatch is the dominant bottleneck of MoE expert parallelism, and the interconnect community has responded with four families of mitigations

model-releasesarxiv-cs-lg
21 May 2026
Agents

Draw2Think: Harnessing Geometry Reasoning through Constraint Engine Interaction

DGX agent

arXiv:2605.20743v1 Announce Type: cross Abstract: Vision-language models solve geometry problems with rising accuracy, yet their intermediate states remain latent and unverifiable: a relation expresse

agentsarxiv-cs-cl
21 May 2026
Model Releases

DySink: Dynamic Frame Sinks for Autoregressive Long Video Generation

DGX agent

arXiv:2605.21028v1 Announce Type: new Abstract: Autoregressive long video generation often adopts bounded-memory streaming for efficiency, typically combining local windows for short-term continuity w

model-releasesarxiv-cs-cv
21 May 2026
Research

Efficient Table QA via TableGrid Navigation and Progressive Inference Prompting

DGX agent

arXiv:2605.20254v1 Announce Type: cross Abstract: Large Language Models (LLMs) have shown promising results on NLP tasks, however, their performance on tabular data still needs research attention, bec

researcharxiv-cs-cv
21 May 2026
Applications

Findings of the Counter Turing Test: AI-Generated Image Detection

DGX agent

arXiv:2605.20787v1 Announce Type: new Abstract: The rapid advancements in generative AI technologies, such as Stable Diffusion, DALL-E, and Midjourney, have significantly transformed the creation of s

applicationsarxiv-cs-cv
21 May 2026
Model Releases

Hack-Verifiable Environments: Towards Evaluating Reward Hacking at Scale

DGX agent

arXiv:2605.20744v1 Announce Type: new Abstract: Aligning autonomous agents with human intent remains a central challenge in modern AI. A key manifestation of this challenge is reward hacking, whereby

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Hyper-V2X: Hypernetworks for Estimating Epistemic and Aleatoric Uncertainty in Cooperative Bird's-Eye-View Semantic Segmentation

DGX agent

arXiv:2605.21309v1 Announce Type: new Abstract: Cooperative perception enabled by Vehicle-to-Everything (V2X) communication enhances autonomous driving safety by creating a unified environmental repre

model-releasesarxiv-cs-cv
21 May 2026
← Previous
1…578579580581582…1082
Next →