AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,171
  • Agents7,461
  • Applications5,337
  • Concepts5
  • Hardware1,806
  • Industry6,146
  • Local Ai4,871
  • Model Releases23,435
  • Research19,874
  • Safety13,191
  • Syntheses17
  • Tools1,673
  • Tutorials3,355

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,171
  • Agents7,461
  • Applications5,337
  • Concepts5
  • Hardware1,806
  • Industry6,146
  • Local Ai4,871
  • Model Releases23,435
  • Research19,874
  • Safety13,191
  • Syntheses17
  • Tools1,673
  • Tutorials3,355

Source
Human
87,171Total entries
1Added by human
87,170Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
62,006 results
22 May 2026

Representability-Aware Neural Networks for Reduced Density Matrices: Application to Fractional Chern Insulators

ResearchDGX agent

arXiv:2605.20326v1 Announce Type: cross Abstract: We develop a representability-aware and interpolable neural network (NN) framework for predicting two-particle reduced density matrices (2-RDMs). The

Residual Skill Optimization for Text-to-SQL Ensembles

Model ReleasesDGX agent

arXiv:2605.21792v1 Announce Type: new Abstract: Text-to-SQL ensembles improve over single-candidate generation by drawing multiple SQL candidates and selecting one, but their effectiveness is bounded

Retaining Suboptimal Actions to Follow Shifting Optima in Multi-Agent Reinforcement Learning

SafetyDGX agent

arXiv:2602.17062v2 Announce Type: replace Abstract: Value decomposition is a core approach for cooperative multi-agent reinforcement learning (MARL). However, existing methods still rely on a single o

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Rethinking Noise-Robust Training for Frozen Vision Foundation Models: A Cross-Dataset Benchmark with a Case Study of Small-Loss Failure

Model ReleasesDGX agent

arXiv:2605.22591v1 Announce Type: new Abstract: Frozen Vision Foundation Models (VFMs) with lightweight classification heads are increasingly used in medical imaging because they offer efficient and r

Rethinking Token Reduction for Diffusion Models via Output-Similarity-Awareness

ResearchDGX agent

arXiv:2605.22011v1 Announce Type: new Abstract: Diffusion Transformers (DiTs) achieve superior image generation quality but suffer from quadratic computational complexity relative to token count. Whil

Revisiting Integration of Image and Metadata for DICOM Series Classification: Cross-Attention and Dictionary Learning

ResearchDGX agent

arXiv:2602.23833v2 Announce Type: replace-cross Abstract: Automated identification of DICOM image series is essential for large-scale medical image analysis, quality control, protocol harmonization, a

RiT: Vanilla Diffusion Transformers Suffice in Representation Space

ResearchDGX agent

arXiv:2605.21981v1 Announce Type: new Abstract: Flow matching with x-prediction -- regressing the clean data point rather than the ambient velocity -- is known to exploit low-dimensional manifold stru

RobuQ: Pushing DiTs to W1.58A2 via Robust Activation Quantization

ResearchDGX agent

arXiv:2509.23582v2 Announce Type: replace Abstract: Diffusion Transformers (DiTs) have recently emerged as a powerful backbone for image generation, demonstrating superior scalability and performance

Robustness of breast lesion segmentation under MRI undersampling improves with k-space-aware deep learning

ResearchDGX agent

arXiv:2605.22327v1 Announce Type: new Abstract: Purpose: To assess whether breast lesion segmentation can be learned directly from acquired MRI k-space, and whether doing so improves robustness when d

SADGE: Structure and Appearance Domain Gap Estimation of Synthetic and Real Data

Model ReleasesDGX agent

arXiv:2605.22467v1 Announce Type: new Abstract: We propose SADGE, a quantitative similarity metric that predicts the performance of synthetic image datasets for common computer vision tasks without do

Safe and Steerable Geometric Motion Policies for Robotic Dexterous Manipulation

SafetyDGX agent

arXiv:2605.21811v1 Announce Type: new Abstract: Robotic dexterous manipulation requires continuously reconciling objectives and constraints defined on heterogeneous geometric spaces: a robot controlle

Scene Abstraction for Lexical Semantics: Structured Representations of Situated Meaning

ResearchDGX agent

arXiv:2605.22542v1 Announce Type: new Abstract: Coffee and tea share many properties, yet they evoke strikingly different situations, atmospheres, and affective associations. These situated dimensions

SceneAligner: 3D-Grounded Floorplan Localization in the Wild

Local AiDGX agent

arXiv:2605.22581v1 Announce Type: new Abstract: Many public buildings provide floorplans with a 'you are here' indicator to help visitors orient themselves. Floorplan localization seeks to computation

SceneGraphGrounder: Zero-Shot 3D Visual Grounding via Structured Scene Graph Matching

Model ReleasesDGX agent

arXiv:2605.21788v1 Announce Type: new Abstract: Zero-shot 3D visual grounding requires localizing objects in unstructured environments from free-form natural language. Recent vision-language model (VL

ScenePilot: Controllable Boundary-Driven Critical Scenario Generation for Autonomous Driving

SafetyDGX agent

arXiv:2605.21168v1 Announce Type: new Abstract: Safety-critical scenarios are central to evaluating autonomous driving systems, yet their rarity in naturalistic logs makes simulation-based stress test

Scout-Assisted Planning for Heterogeneous Robot Teams under Partially Known Environments

AgentsDGX agent

arXiv:2605.22693v1 Announce Type: new Abstract: Autonomous robot teams navigating partially known environments face costly backtracking when ground robots encounter blocked roads that are only reveale

SDGBiasBench: Benchmarking and Mitigating Vision--Language Models' Biases in Sustainable Development Goals

Model ReleasesDGX agent

arXiv:2605.21919v1 Announce Type: new Abstract: Assessing progress toward the Sustainable Development Goals (SDGs) requires multi-step reasoning over visual cues, contextual knowledge, and development

SE3Kit: A Lightweight Python Library for Specialized Geometric Primitives in Robotics

ApplicationsDGX agent

arXiv:2605.22633v1 Announce Type: new Abstract: The Python robotics ecosystem faces a challenge: while many libraries exist for rigid body transformations, few are both lightweight and mathematically

Search-E1: Self-Distillation Drives Self-Evolution in Search-Augmented Reasoning

SafetyDGX agent

arXiv:2605.22511v1 Announce Type: cross Abstract: Post-training has become the dominant recipe for turning a language model into a competent search-augmented reasoning agent. A line of recent work pus

Security Document Classification with a Fine-Tuned Local Large Language Model: Benchmark Data and an Open-Source System

Model ReleasesDGX agent

arXiv:2605.20368v1 Announce Type: cross Abstract: Organizations that scan documents for sensitive information face a practical problem. Cloud services require data to be sent to external infrastructur

Seeing the Poem: Image-Semantic Detection of AI-Generated Modern Chinese Poetry with MLLMs

Model ReleasesDGX agent

arXiv:2605.22654v1 Announce Type: new Abstract: Previous detection studies have shown that LLMs cannot be effectively used as detectors, but these studies have not addressed modern Chinese poetry. Mor

SEGA: Spectral-Energy Guided Attention for Resolution Extrapolation in Diffusion Transformers

ResearchDGX agent

arXiv:2605.22668v1 Announce Type: new Abstract: Diffusion transformers (DiTs) have emerged as a dominant architecture for text-to-image generation, yet their performance drops when generating at resol

SegCompass: Exploring Interpretable Alignment with Sparse Autoencoders for Enhanced Reasoning Segmentation

Local AiDGX agent

arXiv:2605.22658v1 Announce Type: new Abstract: While large language models provide strong compositional reasoning, existing reasoning segmentation pipelines fail to transparently connect this reasoni

SegGuidedNet: Sub-Region-Aware Attention Supervision for Interpretable Brain Tumor Segmentation

Model ReleasesDGX agent

arXiv:2605.22572v1 Announce Type: new Abstract: Accurate segmentation of brain tumour sub-regions from multi-parametric MRI is critical for treatment planning yet remains challenging due to morphologi

Segment Anything with Motion, Geometry, and Semantic Adaptation for Complex Nonlinear Visual Object Tracking

TutorialsDGX agent

arXiv:2605.22538v1 Announce Type: new Abstract: Traditional visual object tracking (VOT) methods typically rely on task-specific supervised training, limiting their generalization to unseen objects an

Seizure-Semiology-Suite (S3): A Clinically Multimodal Dataset, Benchmark, and Models for Seizure Semiology Understanding

Model ReleasesDGX agent

arXiv:2605.21852v1 Announce Type: new Abstract: While Multimodal Large Language Models (MLLMs) have demonstrated remarkable proficiency in general video understanding, their capacity to interpret invo

Self-Policy Distillation via Capability-Selective Subspace Projection

SafetyDGX agent

arXiv:2605.22675v1 Announce Type: new Abstract: Self-distillation bootstraps large language models (LLMs) by training on their own generations. However, existing methods either rely on external signal

Sem-Detect: Semantic Level Detection of AI Generated Peer-Reviews

ResearchDGX agent

arXiv:2605.21713v1 Announce Type: new Abstract: How can we distinguish whether a peer review was written by a human or generated by an AI model? We argue that, in this setting, authorship should not b

SENIOR: Efficient Query Selection and Preference-Guided Exploration in Preference-based Reinforcement Learning

SafetyDGX agent

arXiv:2506.14648v2 Announce Type: replace Abstract: Preference-based Reinforcement Learning (PbRL) methods provide a solution to avoid reward engineering by learning reward models based on human prefe

Sensor2Sensor: Cross-Embodiment Sensor Conversion for Autonomous Driving

AgentsDGX agent

arXiv:2605.22809v1 Announce Type: new Abstract: Robust training and validation of Autonomous Driving Systems (ADS) require massive, diverse datasets. Proprietary data collected by Autonomous Vehicle (

SFN-YOLO: Towards Free-Range Poultry Detection via Scale-aware Fusion Networks

Model ReleasesDGX agent

arXiv:2509.17086v2 Announce Type: replace Abstract: Detecting and localizing poultry is essential for advancing smart poultry farming. Despite the progress of detection-centric methods, challenges per

SiameseNorm: Breaking the Barrier to Reconciling Pre/Post-Norm

Model ReleasesDGX agent

arXiv:2602.08064v2 Announce Type: replace-cross Abstract: The long-standing tension between Pre- and Post-Norm remains an open problem in Transformer architecture, reflecting a fundamental trade-off b

Skarimva: Skeleton-based Action Recognition is a Multi-view Application

ResearchDGX agent

arXiv:2602.23231v2 Announce Type: replace Abstract: Human action recognition plays an important role when developing intelligent interactions between humans and machines. While there is a lot of activ

Slimmable ConvNeXt: Width-Adaptive Inference for Efficient Multi-Device Deployment

ResearchDGX agent

arXiv:2605.22677v1 Announce Type: new Abstract: Deploying vision models across devices with varying resource constraints, or even on a single device where available compute fluctuates due to battery s

SO-Mamba: State-Ownership Mamba for Unrolled MRI Reconstruction

ResearchDGX agent

arXiv:2605.22031v1 Announce Type: new Abstract: Accelerated MRI reconstruction requires recovering missing details while preserving anatomically coherent structures across large spatial regions. State

SONIC: Supersizing Motion Tracking for Natural Humanoid Whole-Body Control

Model ReleasesDGX agent

arXiv:2511.07820v3 Announce Type: replace-cross Abstract: Despite the rise of billion-parameter foundation models trained across thousands of GPUs, similar scaling gains have not been shown for humano

SpaceDG: Benchmarking Spatial Intelligence under Visual Degradation

Model ReleasesDGX agent

arXiv:2605.22536v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have made rapid progress in spatial intelligence, yet existing spatial reasoning benchmarks largely assume pr

SpaceDrive: Infusing Spatial Awareness into VLM-based Autonomous Driving

Model ReleasesDGX agent

arXiv:2512.10719v2 Announce Type: replace Abstract: End-to-end autonomous driving methods built on vision language models (VLMs) have undergone rapid development driven by their universal visual under

Spatial Memory for Out-of-Vision Manipulation in Vision-Language-Action

Model ReleasesDGX agent

arXiv:2605.22283v1 Announce Type: new Abstract: We introduce SOMA, the Spatial Memory framework for Out-of-Vision Manipulation in Vision-Language-Action (VLA) models. Most existing VLAs implicitly ass

SpecHop: Continuous Speculation for Accelerating Multi-Hop Retrieval Agents

AgentsDGX agent

arXiv:2605.21965v1 Announce Type: new Abstract: Large language models increasingly use external tools such as web search and document retrieval to solve information-intensive tasks. However, multi-hop

Spectral Tail Auxiliary Learning for AI-Generated Image Detection

ApplicationsDGX agent

arXiv:2605.22751v1 Announce Type: new Abstract: As generative image models evolve rapidly, the perceptual gap between generated and real images continues to narrow, making AI-generated image detection

SPIRAL: Self-Evolving Action-Conditioned Video Generation via Reflective Planning Agents

ResearchDGX agent

arXiv:2603.08403v3 Announce Type: replace Abstract: Long-horizon action-conditioned video generation aims to synthesize temporally coherent videos that follow complex action instructions over extended

ST-SimDiff: Balancing Spatiotemporal Similarity and Difference for Efficient Video Understanding with MLLMs

ResearchDGX agent

arXiv:2605.22158v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) face significant computational overhead when processing long videos due to the massive number of visual token

stable-worldmodel: A Platform for Reproducible World Modeling Research and Evaluation

ResearchDGX agent

arXiv:2605.21800v1 Announce Type: cross Abstract: World models are central to building agents that can reason, plan, and generalize beyond their training data. However, research on world models is cur

Stdlib or Third-Party? Empirical Performance and Correctness of LLM-Assisted Zero-Dependency Python Libraries

AgentsDGX agent

arXiv:2605.21405v1 Announce Type: cross Abstract: Third-party Python libraries introduce dependency management overhead, supply chain risk, and deployment friction in constrained environments. A natur

Steins;Gate Drive: Semantic Safety Arbitration over Structured Futures for Latency-Decoupled LLM Planning

Model ReleasesDGX agent

arXiv:2605.22456v1 Announce Type: new Abstract: Cloud-hosted LLM driver agents provide useful semantic judgments, but their inference latency exceeds stepwise vehicle-control windows. Learned world mo

STRUCTSENSE: A Task-Agnostic Agentic Framework for Structured Information Extraction with Human-In-The-Loop Evaluation and Benchmarking

AgentsDGX agent

arXiv:2507.03674v3 Announce Type: replace Abstract: Extracting structured information from scientific literature is critical for accelerating discovery, yet Large Language Models (LLMs) often struggle

Structural Anchor Pruning: Training-Free Multi-Vector Compression for Visual Document Retrieval

Model ReleasesDGX agent

arXiv:2601.20107v2 Announce Type: replace-cross Abstract: Recent Vision-Language Models (e.g., ColPali) enable fine-grained Visual Document Retrieval (VDR) but incur prohibitive multi-vector index sto

Structure Retention in Embedding Spaces as a Predictor of Benchmark Performance

Model ReleasesDGX agent

arXiv:2605.22202v1 Announce Type: new Abstract: In this paper, we show that high-performing embedding models organize their embedding spaces in a consistent way. We evaluate 25 contemporary embedding

Structured-Sparse Attention for Entity Tracking with Subquadratic Sequence Complexity

Local AiDGX agent

arXiv:2605.22476v1 Announce Type: cross Abstract: Entity tracking requires maintaining and updating latent states for entities and attributes over long sequences. Recent task-specific attention operat

Sub-exponential Growth Dynamics in Complex Systems: A Piecewise Power-Law Model for the Diffusion of New Words and Names

Model ReleasesDGX agent

arXiv:2511.04106v5 Announce Type: replace-cross Abstract: The diffusion of ideas and language in society has conventionally been described by S-shaped models, such as the logistic curve. However, the

Superhuman Safe and Agile Racing through Multi-Agent Reinforcement Learning

SafetyDGX agent

arXiv:2605.22748v1 Announce Type: new Abstract: Autonomous systems have achieved superhuman performance in isolation or simulation, yet they remain brittle in shared, dynamic real-world spaces. This f

Supervised Classification Heads as Semantic Prototypes: Unlocking Vision-Language Alignment via Weight Recycling

SafetyDGX agent

arXiv:2605.22484v1 Announce Type: new Abstract: Vision-Language Models (VLMs) excel at tasks like zero-shot classification and cross-modal retrieval by mapping images and text to a shared space, but t

SURGE: An Event-Centric Social Media Sentiment Time Series Benchmark with Interaction Structure

Model ReleasesDGX agent

arXiv:2605.21198v1 Announce Type: cross Abstract: Public events on social media generate large volumes of discussion whose collective dynamics carry direct value for opinion forecasting and crisis res

Survive or Collapse: The Asymmetric Roles of Data Gating and Reward Grounding in Self-Play RL

Model ReleasesDGX agent

arXiv:2605.22217v1 Announce Type: cross Abstract: Self-play reinforcement learning trains language models on their own generated tasks, co-evolving a proposer and solver without human labels. Recent s

Swift Sampling: Selecting Temporal Surprises via Taylor Series

ResearchDGX agent

arXiv:2605.22678v1 Announce Type: new Abstract: While most frames in long-form video are redundant, the critical information resides in temporal surprises: moments where the actual visual features dev

Symmetries Here and There, Combined Everywhere: Cross-space Symmetry Compositions in Robotics

ApplicationsDGX agent

arXiv:2605.22639v1 Announce Type: new Abstract: Robots exhibit a rich variety of symmetries arising from their mechanical structure and the properties of their tasks. Although many robotics problems e

SynAE: A Framework for Measuring the Quality of Synthetic Data for Tool-Calling Agent Evaluations

AgentsDGX agent

arXiv:2605.22564v1 Announce Type: new Abstract: Today, tool-calling agents are commonly evaluated or tested on static datasets of execution traces, including input commands, agent responses, and assoc

Synthetic Data Alone is Enough? Rethinking Data Scarcity in Pediatric Rare Disease Recognition

ApplicationsDGX agent

arXiv:2605.22767v1 Announce Type: new Abstract: Children with rare genetic diseases often exhibit distinctive facial phenotypes, yet developing computer vision systems for early diagnosis remains chal

Tackle CSM in JPEG Steganalysis with Data Adaptation

Model ReleasesDGX agent

arXiv:2605.21523v1 Announce Type: cross Abstract: Steganalysis models excel on benchmark datasets but struggle in the wild when analyzed images are produced by a processing pipeline unseen during trai

← Previous
1…622623624625626…1034
Next →