AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,171
  • Agents7,461
  • Applications5,337
  • Concepts5
  • Hardware1,806
  • Industry6,146
  • Local Ai4,871
  • Model Releases23,435
  • Research19,874
  • Safety13,191
  • Syntheses17
  • Tools1,673
  • Tutorials3,355

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,171
  • Agents7,461
  • Applications5,337
  • Concepts5
  • Hardware1,806
  • Industry6,146
  • Local Ai4,871
  • Model Releases23,435
  • Research19,874
  • Safety13,191
  • Syntheses17
  • Tools1,673
  • Tutorials3,355

Source
Human
87,171Total entries
1Added by human
87,170Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
62,006 results
22 May 2026

TacO: Benchmarking Tactile Sensors for Object Manipulation

SafetyDGX agent

arXiv:2605.21976v1 Announce Type: new Abstract: Vision-based learning from demonstrations has achieved remarkable success in enabling robots to perform manipulation tasks and high-level semantic reaso

Targeting Clause Type Distributions: a Picklock for Random Satisfiability Problems

Model ReleasesDGX agent

arXiv:2605.20328v1 Announce Type: cross Abstract: Optimization problems such as the NP-complete 3-SAT provide an important benchmark for the difficult task of finding ground-states in strongly correla

Teaching AI Through Benchmark Construction: QuestBench as a Course-Based Practice for Accountable Knowledge Work

Model ReleasesDGX agent

arXiv:2605.21413v2 Announce Type: new Abstract: As AI becomes part of everyday learning, many courses teach students to use it mainly as a productivity tool: how to prompt, search, summarize, write, c

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Teaching Language Models to Forecast Research Success Through Comparative Idea Evaluation

Model ReleasesDGX agent

arXiv:2605.21491v1 Announce Type: cross Abstract: As language models accelerate scientific research by automating hypothesis generation and implementation, a new bottleneck emerges: evaluating and fil

Temporal Coding as a Substrate for Sensorimotor Object Inference: A Spiking Reinterpretation of Thousand Brains Architecture

Model ReleasesDGX agent

arXiv:2605.22206v1 Announce Type: cross Abstract: The Thousand Brains Theory (TBT) and its open-source Monty framework model object recognition through sensorimotor inference -- identifying objects by

Terminal Constraint Model Predictive Control for Image-Based Visual Servoing of UAVs with Kalman Filter-Based Moment Loss Compensation

ResearchDGX agent

arXiv:2605.22443v1 Announce Type: new Abstract: Image-Based Visual Servoing (IBVS) provides an efficient vision-guided control paradigm for unmanned aerial vehicles (UAVs) by directly regulating image

TextTeacher: What Can Language Teach About Images?

TutorialsDGX agent

arXiv:2605.22098v1 Announce Type: new Abstract: The platonic representation hypothesis suggests that sufficiently large models converge to a shared representation geometry, even across modalities. Mot

The Double Dilemma in Multi-Task Radiology Report Generation: A Gradient Dynamics Analysis and Solution

SafetyDGX agent

arXiv:2605.22635v1 Announce Type: cross Abstract: While multi-task learning based automatic radiology report generation (RRG) is widely adopted to ensure clinical consistency, most focus on architectu

The Neglected Baseline in Model Interpretation

ResearchDGX agent

arXiv:2605.22417v1 Announce Type: new Abstract: We observe that existing model interpretation methods generally ignore the baseline, and such neglect often results in imprecise or even incorrect inter

Thermo-VL: Extending Vision-Language Models to Thermal Infrared Perception

Model ReleasesDGX agent

arXiv:2605.21882v1 Announce Type: new Abstract: Vision-language models (VLMs) often fail under low illumination because their visual grounding is learned predominantly from RGB imagery, whereas therma

Time-varying rPPG signal separation via block-sparse signal model

ResearchDGX agent

arXiv:2605.22425v1 Announce Type: cross Abstract: Remote photoplethysmography (rPPG) enables non-contact measurement of cardiac pulse signals by analyzing subtle color changes in facial videos. Nevert

Token-Level LLM Collaboration via FusionRoute

Model ReleasesDGX agent

arXiv:2601.05106v4 Announce Type: replace-cross Abstract: Large language models (LLMs) exhibit strengths across diverse domains. However, achieving strong performance across these domains with a singl

Token-weighted Direct Preference Optimization with Attention

ResearchDGX agent

arXiv:2605.21883v1 Announce Type: new Abstract: Direct Preference Optimization (DPO) aligns Large Language Models with human preferences without the need for a separate reward model. However, DPO trea

Tokenisation via Convex Relaxations

ResearchDGX agent

arXiv:2605.22821v1 Announce Type: new Abstract: Tokenisation is an integral part of the current NLP pipeline. Current tokenisation algorithms such as BPE and Unigram are greedy algorithms -- they make

Tokenization with Split Trees

Model ReleasesDGX agent

arXiv:2605.22705v1 Announce Type: new Abstract: We introduce Tokenization with Split Trees (ToaST), a subword tokenization method that directly optimizes compression under a new recursive inference pr

Tool-Augmented Agent for Closed-loop Optimization,Simulation,and Modeling Orchestration

AgentsDGX agent

arXiv:2605.20190v1 Announce Type: new Abstract: Iterative industrial design-simulation optimization is bottlenecked by the CAD-CAE semantic gap: translating simulation feedback into valid geometric ed

Towards Clinically Interpretable Ophthalmic VQA via Spatially-Grounded Lesion Evidence

Model ReleasesDGX agent

arXiv:2605.22414v1 Announce Type: new Abstract: Visual Question Answering (VQA) holds great promise for clinical support, particularly in ophthalmology, where retinal fundus photography is essential f

Towards Initialization-free Calibrated Bundle Adjustment

ResearchDGX agent

arXiv:2506.23808v2 Announce Type: replace Abstract: A recent series of works has shown that initialization-free BA can be achieved using pseudo Object Space Error (pOSE) as a surrogate objective. The

Towards Selection of Large Multimodal Models as Engines for Burned-in Protected Health Information Detection in Medical Images

Model ReleasesDGX agent

arXiv:2511.02014v2 Announce Type: replace Abstract: The detection of Protected Health Information (PHI) in medical imaging is critical for safeguarding patient privacy and ensuring compliance with reg

Training-Free Fine-Grained Semantic Segmentations in Low Data Regimes: A FungiTastic Baseline

ResearchDGX agent

arXiv:2605.22492v1 Announce Type: new Abstract: Fine-grained semantic segmentation requires both precise localization and discrimination between visually similar classes. In FungiTastic, this problem

Training-Trajectory-Aware Token Selection

Model ReleasesDGX agent

arXiv:2601.10348v2 Announce Type: replace Abstract: Efficient distillation is a key pathway for converting expensive reasoning capability into deployable efficiency, yet in the frontier regime where t

Transcription and Recognition of Italian Parliamentary Speeches Using Vision-Language Models

Model ReleasesDGX agent

arXiv:2603.28103v2 Announce Type: replace-cross Abstract: Parliamentary proceedings represent a rich yet challenging resource for computational analysis, particularly when preserved only as scanned hi

TransitLM: A Large-Scale Dataset and Benchmark for Map-Free Transit Route Generation

Model ReleasesDGX agent

arXiv:2605.22355v1 Announce Type: new Abstract: Public transit route planning traditionally depends on structured map infrastructure and complex routing engines, and no existing dataset supports train

Translating Signals to Languages for sEMG-Based Activity Recognition

ResearchDGX agent

arXiv:2605.22403v1 Announce Type: new Abstract: Surface electromyography (sEMG) signal-based activity recognition has attracted increasing research attention in recent years. To develop accurate sEMG

Transporting Task Vectors across Different Architectures without Training

Model ReleasesDGX agent

arXiv:2602.12952v2 Announce Type: replace-cross Abstract: Adapting large pre-trained models to downstream tasks often produces task-specific parameter updates that are expensive to relearn for every m

TriSweep: A Four-Drone Swarm Framework for Electromagnetic Side-Channel Analysis

SafetyDGX agent

arXiv:2605.22709v1 Announce Type: cross Abstract: Electromagnetic (EM) side-channel analysis traditionally assumes a stationary, close-proximity probe - a threat model that underestimates aerial adver

TWINGS: Thin Plate Splines Warp-aligned Initialization for Sparse-View Gaussian Splatting

ResearchDGX agent

arXiv:2605.22069v1 Announce Type: new Abstract: Novel view synthesis from sparse-view inputs poses a significant challenge in 3D computer vision, particularly for achieving high-quality scene reconstr

Two is better than one: A Collapse-free Multi-Reward RLIF Training Framework

Model ReleasesDGX agent

arXiv:2605.22620v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has substantially improved the reasoning ability of LLMs, but often depends on external supervis

Two-Stage Multimodal Framework for Emotion Mimicry Intensity Prediction

ResearchDGX agent

arXiv:2605.21869v1 Announce Type: new Abstract: We present our submission to the Hume-ABAW10 Emotional Mimicry Intensity (EMI) Challenge, which aims to predict six continuous emotion intensity dimensi

UIKA: Fast Universal Head Avatar from Pose-Free Images

ResearchDGX agent

arXiv:2601.07603v3 Announce Type: replace Abstract: We present UIKA, a feed-forward animatable Gaussian head model from an arbitrary number of pose-free inputs, including a single image, multi-view ca

Ultra-High-Definition Image Quality Assessment via Graph Representation Learning

Model ReleasesDGX agent

arXiv:2605.22192v1 Announce Type: new Abstract: Blind image quality assessment (BIQA) for ultrahighdefinition (UHD) images remains challenging because native-resolution inference is computationally ex

Understanding Data Temporality Impact on Large Language Models Pre-training

Model ReleasesDGX agent

arXiv:2605.22769v1 Announce Type: new Abstract: Large language models (LLMs) are typically trained on shuffled corpora, yielding models whose knowledge is frozen at train time and whose temporal groun

Understanding Multimodal Failure in Action-Chunking Behavioral Cloning

ResearchDGX agent

arXiv:2605.22493v1 Announce Type: cross Abstract: Behavioral cloning becomes difficult when the same observation admits several valid actions. We study this problem for action-chunking policies and sh

Unified Data Selection for LLM Reasoning

TutorialsDGX agent

arXiv:2605.22389v1 Announce Type: new Abstract: Effectively training Large Language Models (LLMs) for complex, long-CoT reasoning is often bottlenecked by the need for massive high-quality reasoning d

Unifying Masked Diffusion Models with Various Generation Orders and Beyond

SafetyDGX agent

arXiv:2602.02112v2 Announce Type: replace-cross Abstract: Masked diffusion models (MDMs) are a potential alternative to autoregressive models (ARMs) for language generation, but generation quality dep

UniSD: Towards a Unified Self-Distillation Framework for Large Language Models

SafetyDGX agent

arXiv:2605.06597v2 Announce Type: replace Abstract: Self-distillation (SD) offers a promising path for adapting large language models (LLMs) without relying on stronger external teachers. However, SD

Universal CT Representations from Anatomy to Disease Phenotype through Agglomerative Pretraining

SafetyDGX agent

arXiv:2605.21906v1 Announce Type: new Abstract: Computed tomography (CT) is a central to three-dimensional medical imaging, yet CT-based artificial intelligence remains fragmented across task-specific

UniVL: Unified Vision-Language Embedding for Spatially Grounded Contextual Image Generation

Model ReleasesDGX agent

arXiv:2605.21611v1 Announce Type: new Abstract: We introduce spatially grounded contextual image generation, a controllable image generation task that reframes the conditioning paradigm. Instead of su

Value-Gradient Hypothesis of RL for LLMs

SafetyDGX agent

arXiv:2605.21654v1 Announce Type: cross Abstract: Reinforcement learning substantially improves pretrained language models, but it remains understudied why critic-free methods such as PPO and GRPO wor

VBFDD-Agent for Electric Vehicle Battery Fault Detection and Diagnosis: Descriptive Text Modeling of Battery Digital Signals

Local AiDGX agent

arXiv:2605.20742v1 Announce Type: new Abstract: With the rapid proliferation of electric vehicles, the safety and reliability of lithium-ion batteries have become critical concerns. Effective anomaly

VChain: Chain-of-Visual-Thought for Reasoning in Video Generation

TutorialsDGX agent

arXiv:2510.05094v2 Announce Type: replace Abstract: Recent video generation models can produce smooth and visually appealing clips, but they often struggle to synthesize complex dynamics with a cohere

VDE Bench: Evaluating The Capability of Image Editing Models to Modify Visual Documents

Model ReleasesDGX agent

arXiv:2602.00122v2 Announce Type: replace Abstract: In recent years, image editing models have made significant progress, enabling users to manipulate visual content in a flexible and interactive mann

Vector Policy Optimization: Training for Diversity Improves Test-Time Search

SafetyDGX agent

arXiv:2605.22817v1 Announce Type: cross Abstract: Language models must now generalize out of the box to novel environments and work inside inference-scaling search procedures, such as AlphaEvolve, tha

VEELA: A Clinically-Constrained Benchmark for Liver Vessel Segmentation in Computed Tomography Angiography

Model ReleasesDGX agent

arXiv:2605.22357v1 Announce Type: new Abstract: Accurate segmentation of hepatic and portal vessels in contrast-enhanced computed tomography angiography (CTA) remains challenging due to complex vascul

Vendi Novelty Scores for Out-of-Distribution Detection

ResearchDGX agent

arXiv:2602.10062v2 Announce Type: replace-cross Abstract: Out-of-distribution (OOD) detection is critical for the safe deployment of machine learning systems. Existing post-hoc detectors typically rel

VGenST-Bench: A Benchmark for Spatio-Temporal Reasoning via Active Video Synthesis

Model ReleasesDGX agent

arXiv:2605.22570v1 Announce Type: new Abstract: Spatio-temporal reasoning is a core capability for Multimodal Large Language Models (MLLMs) operating in the real world. As such, evaluating it precisel

Video as Natural Augmentation: Towards Unified AI-Generated Image and Video Detection

ResearchDGX agent

arXiv:2605.21977v1 Announce Type: new Abstract: AI-generated content (AIGC) is rapidly improving, creating an urgent need for detectors that generalize across data sources, deployment pipelines, and v

Video-o3: Native Interleaved Clue Seeking for Long Video Multi-Hop Reasoning

ResearchDGX agent

arXiv:2601.23224v2 Announce Type: replace Abstract: Existing multimodal large language models for long-video understanding predominantly rely on uniform sampling and single-turn inference, limiting th

Virtual 3D H&E Staining from Phase-contrast Back-illumination Interference Tomography

ResearchDGX agent

arXiv:2605.22000v1 Announce Type: new Abstract: Three-dimensional (3D) histopathology of unprocessed tissues has the potential to transform disease management by enabling volumetric characterization o

VisPhyWorld: Probing Physical Reasoning via Code-Driven Video Reconstruction

Model ReleasesDGX agent

arXiv:2602.13294v3 Announce Type: replace Abstract: Evaluating whether Multimodal Large Language Models (MLLMs) genuinely reason about physical dynamics remains challenging. Most existing benchmarks r

VISTA: Validation-Guided Integration of Spatial and Temporal Foundation Models with Anatomical Decoding for Rare-Pathology VCE Event Detection -- after competition results

ResearchDGX agent

arXiv:2605.22096v1 Announce Type: new Abstract: Capsule endoscopy event detection is challenging because clinically relevant findings are sparse, visually heterogeneous, and evaluated at the event lev

Visual-Advantage On-Policy Distillation for Vision-Language Models

Model ReleasesDGX agent

arXiv:2605.21924v1 Announce Type: new Abstract: On-policy knowledge distillation has proven effective for language models, yet its application to vision-language models (VLMs) remains underexplored. W

VRXU-net: A Deep Learning Approach for Brain Ischemic Stroke Lesion Detection and Segmentation in T1W MRI

Local AiDGX agent

arXiv:2605.21633v1 Announce Type: cross Abstract: When the blood supply to the brain is obstructed by a clot, oxygen delivery to brain tissues becomes insufficient, leading to cellular necrosis. In he

What Does the Caption Really Say? Counterfactual Phrase Intervention for Compositional Data Selection in Vision-Language Pretraining

SafetyDGX agent

arXiv:2605.22651v1 Announce Type: new Abstract: CLIP-style contrastive pretraining typically curates web-scale image-text pairs using sample-level filtering signals, often based on pair-level alignmen

What Does Vision Tool-Use Reinforcement Learning Really Learn? Disentangling Tool-Induced and Intrinsic Effects for Crop-and-Zoom

AgentsDGX agent

arXiv:2602.01334v2 Announce Type: replace Abstract: Vision tool-use reinforcement learning (RL) can equip vision language models with visual operators such as crop-and-zoom and achieves strong perform

When Cases Get Rare: A Retrieval Benchmark for Off-Guideline Clinical Question Answering

Model ReleasesDGX agent

arXiv:2605.21807v1 Announce Type: new Abstract: Across medical specialties, clinical practice is anchored in evidence-based guidelines that codify best studied diagnostic and treatment pathways. These

When Shared Knowledge Hurts: Spectral Over-Accumulation in Model Merging

ResearchDGX agent

arXiv:2602.05536v2 Announce Type: replace-cross Abstract: Model merging combines multiple fine-tuned models into a single model by adding their weight updates, providing a lightweight alternative to r

When Simultaneous Localization and Mapping Meets Wireless Communications: A Survey

AgentsDGX agent

arXiv:2602.06995v2 Announce Type: replace-cross Abstract: This paper surveys the state-of-the-art in the nexus of SLAM and Wireless Communications, attributing the bidirectional impact of each with a

Which Way Did It Move? Diagnosing and Overcoming Directional Motion Blindness in Video-LLMs

Local AiDGX agent

arXiv:2605.22823v1 Announce Type: new Abstract: Video Large Language Models (Video-LLMs) have made rapid progress on temporal video understanding, yet many fail at a basic perceptual primitive: signed

Whose Voice Counts? Mapping Stakeholder Perspectives on AI Through Public Submissions to the U.S. Government

ResearchDGX agent

arXiv:2605.22650v1 Announce Type: new Abstract: As artificial intelligence (AI) systems become more common in our daily lives, it is important to understand how different stakeholders comprehend and e

← Previous
1…623624625626627…1034
Next →