AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,271
  • Agents7,543
  • Applications5,408
  • Concepts5
  • Hardware1,828
  • Industry6,162
  • Local Ai4,927
  • Model Releases23,818
  • Research20,122
  • Safety13,367
  • Syntheses17
  • Tools1,674
  • Tutorials3,400

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,271
  • Agents7,543
  • Applications5,408
  • Concepts5
  • Hardware1,828
  • Industry6,162
  • Local Ai4,927
  • Model Releases23,818
  • Research20,122
  • Safety13,367
  • Syntheses17
  • Tools1,674
  • Tutorials3,400

Source
Human
88,271Total entries
1Added by human
88,270Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
62,897 results
Model Releases

Transcription and Recognition of Italian Parliamentary Speeches Using Vision-Language Models

DGX agent

arXiv:2603.28103v2 Announce Type: replace-cross Abstract: Parliamentary proceedings represent a rich yet challenging resource for computational analysis, particularly when preserved only as scanned hi

model-releasesarxiv-cs-ai
22 May 2026
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

TransitLM: A Large-Scale Dataset and Benchmark for Map-Free Transit Route Generation

DGX agent

arXiv:2605.22355v1 Announce Type: new Abstract: Public transit route planning traditionally depends on structured map infrastructure and complex routing engines, and no existing dataset supports train

model-releasesarxiv-cs-cl
22 May 2026
Research

Translating Signals to Languages for sEMG-Based Activity Recognition

DGX agent

arXiv:2605.22403v1 Announce Type: new Abstract: Surface electromyography (sEMG) signal-based activity recognition has attracted increasing research attention in recent years. To develop accurate sEMG

researcharxiv-cs-cv
22 May 2026
Model Releases

Transporting Task Vectors across Different Architectures without Training

DGX agent

arXiv:2602.12952v2 Announce Type: replace-cross Abstract: Adapting large pre-trained models to downstream tasks often produces task-specific parameter updates that are expensive to relearn for every m

model-releasesarxiv-cs-cv
22 May 2026
Safety

TriSweep: A Four-Drone Swarm Framework for Electromagnetic Side-Channel Analysis

DGX agent

arXiv:2605.22709v1 Announce Type: cross Abstract: Electromagnetic (EM) side-channel analysis traditionally assumes a stationary, close-proximity probe - a threat model that underestimates aerial adver

safetyarxiv-cs-ro
22 May 2026
Research

TWINGS: Thin Plate Splines Warp-aligned Initialization for Sparse-View Gaussian Splatting

DGX agent

arXiv:2605.22069v1 Announce Type: new Abstract: Novel view synthesis from sparse-view inputs poses a significant challenge in 3D computer vision, particularly for achieving high-quality scene reconstr

researcharxiv-cs-cv
22 May 2026
Model Releases

Two is better than one: A Collapse-free Multi-Reward RLIF Training Framework

DGX agent

arXiv:2605.22620v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has substantially improved the reasoning ability of LLMs, but often depends on external supervis

model-releasesarxiv-cs-cl
22 May 2026
Research

Two-Stage Multimodal Framework for Emotion Mimicry Intensity Prediction

DGX agent

arXiv:2605.21869v1 Announce Type: new Abstract: We present our submission to the Hume-ABAW10 Emotional Mimicry Intensity (EMI) Challenge, which aims to predict six continuous emotion intensity dimensi

researcharxiv-cs-cv
22 May 2026
Research

UIKA: Fast Universal Head Avatar from Pose-Free Images

DGX agent

arXiv:2601.07603v3 Announce Type: replace Abstract: We present UIKA, a feed-forward animatable Gaussian head model from an arbitrary number of pose-free inputs, including a single image, multi-view ca

researcharxiv-cs-cv
22 May 2026
Model Releases

Ultra-High-Definition Image Quality Assessment via Graph Representation Learning

DGX agent

arXiv:2605.22192v1 Announce Type: new Abstract: Blind image quality assessment (BIQA) for ultrahighdefinition (UHD) images remains challenging because native-resolution inference is computationally ex

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

Understanding Data Temporality Impact on Large Language Models Pre-training

DGX agent

arXiv:2605.22769v1 Announce Type: new Abstract: Large language models (LLMs) are typically trained on shuffled corpora, yielding models whose knowledge is frozen at train time and whose temporal groun

model-releasesarxiv-cs-cl
22 May 2026
Research

Understanding Multimodal Failure in Action-Chunking Behavioral Cloning

DGX agent

arXiv:2605.22493v1 Announce Type: cross Abstract: Behavioral cloning becomes difficult when the same observation admits several valid actions. We study this problem for action-chunking policies and sh

researcharxiv-cs-ro
22 May 2026
Tutorials

Unified Data Selection for LLM Reasoning

DGX agent

arXiv:2605.22389v1 Announce Type: new Abstract: Effectively training Large Language Models (LLMs) for complex, long-CoT reasoning is often bottlenecked by the need for massive high-quality reasoning d

tutorialsarxiv-cs-cl
22 May 2026
Safety

Unifying Masked Diffusion Models with Various Generation Orders and Beyond

DGX agent

arXiv:2602.02112v2 Announce Type: replace-cross Abstract: Masked diffusion models (MDMs) are a potential alternative to autoregressive models (ARMs) for language generation, but generation quality dep

safetyarxiv-cs-cl
22 May 2026
Safety

UniSD: Towards a Unified Self-Distillation Framework for Large Language Models

DGX agent

arXiv:2605.06597v2 Announce Type: replace Abstract: Self-distillation (SD) offers a promising path for adapting large language models (LLMs) without relying on stronger external teachers. However, SD

safetyarxiv-cs-cl
22 May 2026
Safety

Universal CT Representations from Anatomy to Disease Phenotype through Agglomerative Pretraining

DGX agent

arXiv:2605.21906v1 Announce Type: new Abstract: Computed tomography (CT) is a central to three-dimensional medical imaging, yet CT-based artificial intelligence remains fragmented across task-specific

safetyarxiv-cs-cv
22 May 2026
Model Releases

UniVL: Unified Vision-Language Embedding for Spatially Grounded Contextual Image Generation

DGX agent

arXiv:2605.21611v1 Announce Type: new Abstract: We introduce spatially grounded contextual image generation, a controllable image generation task that reframes the conditioning paradigm. Instead of su

model-releasesarxiv-cs-cv
22 May 2026
Safety

Value-Gradient Hypothesis of RL for LLMs

DGX agent

arXiv:2605.21654v1 Announce Type: cross Abstract: Reinforcement learning substantially improves pretrained language models, but it remains understudied why critic-free methods such as PPO and GRPO wor

safetyarxiv-cs-cl
22 May 2026
Local Ai

VBFDD-Agent for Electric Vehicle Battery Fault Detection and Diagnosis: Descriptive Text Modeling of Battery Digital Signals

DGX agent

arXiv:2605.20742v1 Announce Type: new Abstract: With the rapid proliferation of electric vehicles, the safety and reliability of lithium-ion batteries have become critical concerns. Effective anomaly

local-aiarxiv-cs-ai
22 May 2026
Tutorials

VChain: Chain-of-Visual-Thought for Reasoning in Video Generation

DGX agent

arXiv:2510.05094v2 Announce Type: replace Abstract: Recent video generation models can produce smooth and visually appealing clips, but they often struggle to synthesize complex dynamics with a cohere

tutorialsarxiv-cs-cv
22 May 2026
Model Releases

VDE Bench: Evaluating The Capability of Image Editing Models to Modify Visual Documents

DGX agent

arXiv:2602.00122v2 Announce Type: replace Abstract: In recent years, image editing models have made significant progress, enabling users to manipulate visual content in a flexible and interactive mann

model-releasesarxiv-cs-cv
22 May 2026
Safety

Vector Policy Optimization: Training for Diversity Improves Test-Time Search

DGX agent

arXiv:2605.22817v1 Announce Type: cross Abstract: Language models must now generalize out of the box to novel environments and work inside inference-scaling search procedures, such as AlphaEvolve, tha

safetyarxiv-cs-cl
22 May 2026
Model Releases

VEELA: A Clinically-Constrained Benchmark for Liver Vessel Segmentation in Computed Tomography Angiography

DGX agent

arXiv:2605.22357v1 Announce Type: new Abstract: Accurate segmentation of hepatic and portal vessels in contrast-enhanced computed tomography angiography (CTA) remains challenging due to complex vascul

model-releasesarxiv-cs-cv
22 May 2026
Research

Vendi Novelty Scores for Out-of-Distribution Detection

DGX agent

arXiv:2602.10062v2 Announce Type: replace-cross Abstract: Out-of-distribution (OOD) detection is critical for the safe deployment of machine learning systems. Existing post-hoc detectors typically rel

researcharxiv-cs-cv
22 May 2026
Model Releases

VGenST-Bench: A Benchmark for Spatio-Temporal Reasoning via Active Video Synthesis

DGX agent

arXiv:2605.22570v1 Announce Type: new Abstract: Spatio-temporal reasoning is a core capability for Multimodal Large Language Models (MLLMs) operating in the real world. As such, evaluating it precisel

model-releasesarxiv-cs-cv
22 May 2026
Research

Video as Natural Augmentation: Towards Unified AI-Generated Image and Video Detection

DGX agent

arXiv:2605.21977v1 Announce Type: new Abstract: AI-generated content (AIGC) is rapidly improving, creating an urgent need for detectors that generalize across data sources, deployment pipelines, and v

researcharxiv-cs-cv
22 May 2026
Research

Video-o3: Native Interleaved Clue Seeking for Long Video Multi-Hop Reasoning

DGX agent

arXiv:2601.23224v2 Announce Type: replace Abstract: Existing multimodal large language models for long-video understanding predominantly rely on uniform sampling and single-turn inference, limiting th

researcharxiv-cs-cv
22 May 2026
Research

Virtual 3D H&E Staining from Phase-contrast Back-illumination Interference Tomography

DGX agent

arXiv:2605.22000v1 Announce Type: new Abstract: Three-dimensional (3D) histopathology of unprocessed tissues has the potential to transform disease management by enabling volumetric characterization o

researcharxiv-cs-cv
22 May 2026
Model Releases

VisPhyWorld: Probing Physical Reasoning via Code-Driven Video Reconstruction

DGX agent

arXiv:2602.13294v3 Announce Type: replace Abstract: Evaluating whether Multimodal Large Language Models (MLLMs) genuinely reason about physical dynamics remains challenging. Most existing benchmarks r

model-releasesarxiv-cs-cv
22 May 2026
Research

VISTA: Validation-Guided Integration of Spatial and Temporal Foundation Models with Anatomical Decoding for Rare-Pathology VCE Event Detection -- after competition results

DGX agent

arXiv:2605.22096v1 Announce Type: new Abstract: Capsule endoscopy event detection is challenging because clinically relevant findings are sparse, visually heterogeneous, and evaluated at the event lev

researcharxiv-cs-cv
22 May 2026
Model Releases

Visual-Advantage On-Policy Distillation for Vision-Language Models

DGX agent

arXiv:2605.21924v1 Announce Type: new Abstract: On-policy knowledge distillation has proven effective for language models, yet its application to vision-language models (VLMs) remains underexplored. W

model-releasesarxiv-cs-cv
22 May 2026
Local Ai

VRXU-net: A Deep Learning Approach for Brain Ischemic Stroke Lesion Detection and Segmentation in T1W MRI

DGX agent

arXiv:2605.21633v1 Announce Type: cross Abstract: When the blood supply to the brain is obstructed by a clot, oxygen delivery to brain tissues becomes insufficient, leading to cellular necrosis. In he

local-aiarxiv-cs-cv
22 May 2026
Safety

What Does the Caption Really Say? Counterfactual Phrase Intervention for Compositional Data Selection in Vision-Language Pretraining

DGX agent

arXiv:2605.22651v1 Announce Type: new Abstract: CLIP-style contrastive pretraining typically curates web-scale image-text pairs using sample-level filtering signals, often based on pair-level alignmen

safetyarxiv-cs-cv
22 May 2026
Agents

What Does Vision Tool-Use Reinforcement Learning Really Learn? Disentangling Tool-Induced and Intrinsic Effects for Crop-and-Zoom

DGX agent

arXiv:2602.01334v2 Announce Type: replace Abstract: Vision tool-use reinforcement learning (RL) can equip vision language models with visual operators such as crop-and-zoom and achieves strong perform

agentsarxiv-cs-cv
22 May 2026
Model Releases

When Cases Get Rare: A Retrieval Benchmark for Off-Guideline Clinical Question Answering

DGX agent

arXiv:2605.21807v1 Announce Type: new Abstract: Across medical specialties, clinical practice is anchored in evidence-based guidelines that codify best studied diagnostic and treatment pathways. These

model-releasesarxiv-cs-cl
22 May 2026
Research

When Shared Knowledge Hurts: Spectral Over-Accumulation in Model Merging

DGX agent

arXiv:2602.05536v2 Announce Type: replace-cross Abstract: Model merging combines multiple fine-tuned models into a single model by adding their weight updates, providing a lightweight alternative to r

researcharxiv-cs-cl
22 May 2026
Agents

When Simultaneous Localization and Mapping Meets Wireless Communications: A Survey

DGX agent

arXiv:2602.06995v2 Announce Type: replace-cross Abstract: This paper surveys the state-of-the-art in the nexus of SLAM and Wireless Communications, attributing the bidirectional impact of each with a

agentsarxiv-cs-cv
22 May 2026
Local Ai

Which Way Did It Move? Diagnosing and Overcoming Directional Motion Blindness in Video-LLMs

DGX agent

arXiv:2605.22823v1 Announce Type: new Abstract: Video Large Language Models (Video-LLMs) have made rapid progress on temporal video understanding, yet many fail at a basic perceptual primitive: signed

local-aiarxiv-cs-cv
22 May 2026
Research

Whose Voice Counts? Mapping Stakeholder Perspectives on AI Through Public Submissions to the U.S. Government

DGX agent

arXiv:2605.22650v1 Announce Type: new Abstract: As artificial intelligence (AI) systems become more common in our daily lives, it is important to understand how different stakeholders comprehend and e

researcharxiv-cs-cl
22 May 2026
Safety

Why Semantic Entropy Fails: Geometry-Aware and Calibrated Uncertainty for Policy Optimization

DGX agent

arXiv:2605.21801v1 Announce Type: cross Abstract: Post-training has become central to improving reasoning and alignment in large language models, where critic-free models enable scalable learning from

safetyarxiv-cs-cl
22 May 2026
Hardware

WorldKV: Efficient World Memory with World Retrieval and Compression

DGX agent

arXiv:2605.22718v1 Announce Type: new Abstract: Autoregressive video diffusion models have enabled real-time, action-conditioned world generation. However, sustaining a persistent world, where revisit

hardwarearxiv-cs-cv
22 May 2026
Safety

'Would You Want an AI Tutor?' Understanding Stakeholder Perceptions of LLM-based Systems in the Classroom

DGX agent

arXiv:2503.02885v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have gained traction in educational settings, often framed as virtual tutors or teaching assistants. Following ea

safetyarxiv-cs-cl
22 May 2026
Safety

X-OmniClaw Technical Report: A Unified Mobile Agent for Multimodal Understanding and Interaction

DGX agent

arXiv:2605.05765v2 Announce Type: replace Abstract: Inspired by the development of OpenClaw, there is a growing demand for mobile-based personal agents capable of handling complex and intuitive intera

safetyarxiv-cs-cv
22 May 2026
Model Releases

X-Token: Projection-Guided Cross-Tokenizer Knowledge Distillation

DGX agent

arXiv:2605.21699v1 Announce Type: cross Abstract: Cross-tokenizer knowledge distillation allows a student model to learn from teachers with incompatible vocabularies. Prior work operates on hidden sta

model-releasesarxiv-cs-cl
22 May 2026
Research

Zero-Shot Temporal Action Localization Through Textual Guidance

DGX agent

arXiv:2605.22201v1 Announce Type: new Abstract: Zero-shot temporal action localization (ZS-TAL) consists of classifying and localizing actions in untrimmed videos, where action classes are unseen at t

researcharxiv-cs-cv
22 May 2026
Safety

3D Reconstruction and Knowledge Distillation to Improve Multi-View Image Models to Explore Spike Volume Estimation in Wheat

DGX agent

arXiv:2605.20940v1 Announce Type: new Abstract: Accurate estimation of wheat spike volume is important for yield component analysis and stress resilience assessment, yet field-based measurement remain

safetyarxiv-cs-cv
21 May 2026
Safety

A 10,000-Year Global Stochastic Tropical Cyclone Catalog with Wind-Dependent Track Transitions (WHITS)

DGX agent

arXiv:2605.20494v1 Announce Type: new Abstract: Reliable assessment of tropical cyclone (TC) risk is limited by the brevity and spatial sparsity of the historical record, particularly for the rare, hi

safetyarxiv-cs-lg
21 May 2026
Research

A Comprehensive Comparison of Deep Learning Architectures for COVID-19 Classification on CT & X-ray Imagery

DGX agent

arXiv:2605.20445v1 Announce Type: new Abstract: COVID-19 was a significant challenge that led to the loss of numerous lives daily. Not only a certain country was involved in this outbreak, but even th

researcharxiv-cs-cv
21 May 2026
← Previous
1…798799800801802…1311
Next →