AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
Human
84,661Total entries
1Added by human
84,660Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,851 results
15 Jul 2026

Streamlining stereo differentiable rendering for marker-free real-time tracking of surgical robots

SafetyDGX agent

arXiv:2607.12604v1 Announce Type: new Abstract: Purpose: Marker-based tracking of surgical robots is occlusion-prone in cluttered operating rooms. We evaluate stereo differentiable rendering for marke

Structure-Semantic Co-optimized Latent Diffusion Model for Fast Visual Anagram Synthesis

SafetyDGX agent

arXiv:2606.16241v3 Announce Type: replace Abstract: Visual anagram is an intriguing form of art creation wherein a single image presents different conceptual interpretations under transformations such

SymbOmni: Evolving Agentic Omni Models via Symbolic Concept Learning

AgentsDGX agent

arXiv:2607.12042v1 Announce Type: new Abstract: Visual generation is increasingly ubiquitous in diverse domains, from text-to-image/video synthesis to multimodal interactive creation. Yet prevailing m

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

TADPO: Reinforcement Learning Goes Off-road

SafetyDGX agent

arXiv:2603.05995v2 Announce Type: replace-cross Abstract: Off-road autonomous driving poses significant challenges such as navigating unmapped, variable terrain with uncertain and diverse dynamics. Ad

TAKE: Trajectory-Aware Knowledge Estimation for Text Dataset Distillation

ResearchDGX agent

arXiv:2607.11898v1 Announce Type: new Abstract: Large-scale text corpora have become a quiet bottleneck in modern NLP, not just in storage, but in the accumulated cost of training, fine-tuning, and co

TerraLogic: A Benchmark for Hierarchical Geospatial Reasoning in Earth Observation

Model ReleasesDGX agent

arXiv:2607.12497v1 Announce Type: new Abstract: Beyond perception, reasoning is essential in remote sensing for advanced interpretation, inference, and decision-making. Recent advances in large langua

TerraZero: Procedural Driving Simulation for Zero-Demonstration Self-Play at Scale

Model ReleasesDGX agent

arXiv:2607.13028v1 Announce Type: cross Abstract: Training robust autonomous driving agents requires a simulator that is fast enough for reinforcement learning at scale, realistic enough to ground beh

Text-Aided Multi-Modal Panoptic Symbol Spotting for CAD Floor Plan Drawings

SafetyDGX agent

arXiv:2607.12678v1 Announce Type: cross Abstract: Computer-Aided Design (CAD) floor plan drawings contain both graphical primitives and textual annotations, which provide complementary geometric and s

The Benjamini--Hochberg Procedure Can Fail to Control the FDR for Correlated Two-Sided Gaussian Tests

Model ReleasesDGX agent

arXiv:2607.12208v1 Announce Type: cross Abstract: We show that the Benjamini--Hochberg procedure can fail to control the false discovery rate (FDR) at its nominal level for correlated two-sided Gaussi

The Capacity of Thought: Benchmarking Llama 3.2 in Semantic fMRI Neural Language Decoding and Improving the Huth Encoding-Model Baseline

Model ReleasesDGX agent

arXiv:2607.12079v1 Announce Type: new Abstract: Decoding continuous language from fMRI signals remains a core challenge in non-invasive brain-computer interface research. We present two complementary

The Computational Basis of Confidence in Large Language Models

ResearchDGX agent

arXiv:2607.12447v1 Announce Type: cross Abstract: Reliable confidence -- the probability that a model's own answer is correct -- is essential for the trustworthy deployment of language models. Existin

The Emerging Paradigm of Geospatial Foundation Models: From Pre-Training to Agentic Reasoning

AgentsDGX agent

arXiv:2607.12177v1 Announce Type: new Abstract: The analysis of satellite and aerial imagery has entered a new era with the advent of foundation models. This paper describes the concept of Geospatial

The Geometry of Memorization: Finite-Time Spectral Sensitivity as a Diagnostic for Flow Matching Models

ResearchDGX agent

arXiv:2607.12616v1 Announce Type: new Abstract: Continuous-time generative frameworks construct probability paths between base and target domains by optimizing time-dependent velocity fields. While th

The GEST-Engine: From Event Graphs to Synthetic Video. A Full Technical Report

AgentsDGX agent

arXiv:2607.12231v1 Announce Type: new Abstract: We present the GEST-Engine, a complete system that goes from natural-language text to fully-annotated multi-actor video. At its core is an explicit worl

The Illusion of Robustness: Aggregate Accuracy Hides Prediction Flips under Task-Irrelevant Context

Model ReleasesDGX agent

arXiv:2607.12963v1 Announce Type: new Abstract: As large language models (LLMs) grow more capable, they are increasingly deployed in context-rich settings where task inputs are often accompanied by lo

The Limits and Potentials of Local SGD for Distributed Heterogeneous Learning with Intermittent Communication

ResearchDGX agent

arXiv:2405.11667v2 Announce Type: replace Abstract: Local SGD is a popular optimization method in distributed learning, often outperforming other algorithms in practice, including mini-batch SGD. Desp

The Model Knows Your Project, Not You: Measuring Recognition in LLMs with NameRank

ResearchDGX agent

arXiv:2607.12520v1 Announce Type: new Abstract: What a frontier model recalls about a person or tool from its own weights -- before any retrieval step -- often shapes the first description a human see

The One-Word Census: Answer-Choice Conformity Across 44 Language Models

Model ReleasesDGX agent

arXiv:2607.12796v1 Announce Type: cross Abstract: When a language model must pick one answer from a large space of equally valid options, which does it pick -- and how often is it the same answer ever

The Seriality Gap in Video Diffusion Models

ResearchDGX agent

arXiv:2607.13031v1 Announce Type: cross Abstract: When one ball strikes another, then another, video models should predict the consequences of each bounce. In controlled experiments on multi-ball hard

The Sound of Absence: Audio-Language Embedding Models Struggle with Negation

SafetyDGX agent

arXiv:2607.12290v1 Announce Type: cross Abstract: Audio-language embedding models such as CLAP are widely evaluated on matching present sound events, but rarely on negation. We show this affirmation-o

The Spectrum Is Not Enough: When Context Helps Time-Series Forecasting

ResearchDGX agent

arXiv:2607.13006v1 Announce Type: new Abstract: A growing family of indices scores how predictable a series is from its spectrum. Practitioners increasingly read these scores as answering a different

The TopCoW Challenge -- Topology-Aware Circle of Willis Segmentation for CT and MR Angiography

Model ReleasesDGX agent

arXiv:2312.17670v5 Announce Type: replace Abstract: The Circle of Willis (CoW) is an important network of arteries connecting major circulations of the brain. Its vascular architecture is believed to

Thompson Sampling Is 2-Competitive for Mistakes

SafetyDGX agent

arXiv:2607.12389v1 Announce Type: cross Abstract: We consider Bayesian bandit models and prove that Thompson sampling makes at most twice the expected number of mistakes (selections of a suboptimal ar

Tilt-Ropter: A Fully Actuated Hybrid Aerial-Terrestrial Vehicle with Tilt Rotors and Passive Wheels

ApplicationsDGX agent

arXiv:2602.01700v3 Announce Type: replace Abstract: In this work, we present Tilt-Ropter, a fully actuated hybrid aerial-terrestrial vehicle (HATV) that integrates tilt rotors with passive wheels to e

Together, Then Apart: Balancing Alignment and Distinctiveness for Multimodal Survival Analysis

SafetyDGX agent

arXiv:2511.18089v2 Announce Type: replace Abstract: Multimodal survival analysis aims to improve cancer prognosis using heterogeneous biomedical data, such as histopathology images and genomic profile

Token Reduction Is Not Cost Reduction

Model ReleasesDGX agent

arXiv:2607.12161v1 Announce Type: new Abstract: Context-reduction layers for API-based coding agents, including command-output compressors, retrieval rankers, and payload-optimizing proxies, are usual

Toward Localizing and Repairing Bias in Transformer Attention Heads

Local AiDGX agent

arXiv:2607.12863v1 Announce Type: cross Abstract: Transformer language models are increasingly used as software components, yet biased outputs remain difficult to localize and repair inside the model.

Toward Metaphor-Fluid Conversation Design for Voice User Interfaces

ResearchDGX agent

arXiv:2502.11554v3 Announce Type: replace-cross Abstract: Metaphors play a critical role in shaping user experiences with Voice User Interfaces (VUIs), yet existing designs often rely on static, human

Toward Production-Ready Federated Learning in Healthcare: Privacy, Orchestration, and Governance in MLOps

Local AiDGX agent

arXiv:2607.10467v2 Announce Type: replace-cross Abstract: Healthcare organizations often cannot freely centralize patient data because medical records are sensitive, regulated, and institutionally con

Toward Trustworthy Autonomous Science: A Two-Year Community Roadmap

SafetyDGX agent

arXiv:2607.12113v1 Announce Type: cross Abstract: One year ago, the AISLE roadmap argued that autonomous laboratories operated as isolated islands and proposed a grassroots network organized around fi

Towards Self-Evolving Agents: A Human-Inspired Adaptive Exploration-Exploitation Framework for Genetic Network Programming

Model ReleasesDGX agent

arXiv:2607.11913v1 Announce Type: cross Abstract: Recent advancements in agentic AI have increasingly moved toward graph-based methods, driven by the demand for explainable, human-centered, and non-li

Towards Vision-Free CIR: Attribute-Augmented Scoring and LLM-Based Reranking for Zero-Shot Composed Image Retrieval

ResearchDGX agent

arXiv:2607.12621v1 Announce Type: new Abstract: Recent work has shown that 'Vision-Free'' approaches (representing images as text) can be effective for standard image retrieval tasks. However, it rema

TRACE: An Operational Reasoning Schema for Auditable Agentic Commitments

Model ReleasesDGX agent

arXiv:2607.12480v1 Announce Type: new Abstract: This paper defines TRACE (Typed Reasoning And Commitment Evidence): a typed, versioned schema for recording reasoning traces, a reference procedure for

Traceback Translators Against Forgetting in Continual Fake Speech Detection

ResearchDGX agent

arXiv:2607.12569v1 Announce Type: cross Abstract: Fake speech detectors are increasingly challenged by the development of new and more accurate generative models. To cope with this problem, continual

TraceSynth: Generating Production-Quality Kernel Traces with Constraint-Guided Diffusion Models

ApplicationsDGX agent

arXiv:2607.12104v1 Announce Type: cross Abstract: Machine learning models for system diagnostics rely on kernel execution traces to capture fine-grained system behavior, but collecting production trac

Tracing Agentic Failure from the Flow of Success

AgentsDGX agent

arXiv:2607.12747v1 Announce Type: new Abstract: Failure attribution for LLM-based agentic systems, i.e., identifying which steps in a failure trajectory caused the task to fail, is critical for debugg

Track, Rank, Crack: Epistemic Working Memory Scales Multi-Hop Reasoning in Language Agents

Model ReleasesDGX agent

arXiv:2607.12267v1 Announce Type: cross Abstract: Language agents that interleave reasoning and tool use degrade sharply as reasoning chains lengthen, even when each individual step is easy. We trace

TRAIL: A Platform for Configurable Human--AI Teaming Experiments

SafetyDGX agent

arXiv:2607.12180v1 Announce Type: cross Abstract: An AI teammate's design properties (personality, communication style, when it speaks) can shape a team's trust, coordination, and decisions. Studying

Traj-VLN: Learning Pixel-Space Interaction via Autoregressive Trajectory Generation

AgentsDGX agent

arXiv:2607.10744v2 Announce Type: replace Abstract: Benefiting from the powerful priors embedded in large-scale pre-training data and the emerging commonsense reasoning ability, large language models

Transforming LLMs into Efficient Cross-Encoders via Knowledge Distillation for RAG Reranking

Model ReleasesDGX agent

arXiv:2607.11933v1 Announce Type: new Abstract: Cross-encoders achieve high reranking accuracy in Retrieval-Augmented Generation (RAG) pipelines but impose quadratic inference costs that limit real-ti

Translation as a Computationally Efficient Bridge: Feasibility of English BERT for Low-Resource Languages

ResearchDGX agent

arXiv:2607.12612v1 Announce Type: new Abstract: BERT models have revolutionised Natural Language Processing (NLP) through their ability to process unstructured text across diverse domains. However, de

TrustVLA: Mechanism-Guided Inference-Time Defense Against Vision-Language-Action Backdoors

SafetyDGX agent

arXiv:2607.12571v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models are deployed through pipelines that end users cannot audit, and a poisoned VLA can behave normally on clean observat

TSCA-Net: Temporal-Spatial Clique Attention for Interpretable Multimodal Pedestrian Trajectory Prediction

Local AiDGX agent

arXiv:2607.11939v1 Announce Type: new Abstract: Accurate pedestrian trajectory prediction in crowded environments remains challenging due to the multimodal uncertainty of human motion and the variable

UMSS: Towards Unsupervised Multi-modal Semantic Segmentation

TutorialsDGX agent

arXiv:2607.12372v1 Announce Type: new Abstract: Multimodal semantic segmentation (MSS) is essential for robust perception in complex environments, yet its potential remains largely untapped because of

Uncertainty-Aware Multi-Source Retinal Fluid Segmentation in OCT

ResearchDGX agent

arXiv:2607.12212v1 Announce Type: cross Abstract: Measuring retinal fluid from optical coherence tomography (OCT) drives treatment decisions in macular disease, but manual annotation is slow and segme

Understanding Sources of Demographic Predictability in Brain MRI via Disentangling Anatomy and Contrast

SafetyDGX agent

arXiv:2603.04113v2 Announce Type: replace-cross Abstract: Demographic attributes can be predicted from medical images, raising concerns about bias in clinical AI systems. In X-ray imaging, acquisition

Understanding Structured Health Data through Interaction-Aware Mixture-of-Experts

ResearchDGX agent

arXiv:2607.12255v1 Announce Type: new Abstract: We study interaction-aware mixture-of-experts for post-stroke rigidity prediction using multi-level views of structured health records. Despite minimal

UniMedSeg: Unified In-Context Learning for Multi-Paradigm 2D/3D Medical Image Segmentation

ResearchDGX agent

arXiv:2607.12896v1 Announce Type: new Abstract: Medical image segmentation foundation models are expected to generalize across diverse clinical scenarios, yet existing universal methods remain fragmen

UniVR: Thinking in Visual Space for Unified Visual Reasoning

Model ReleasesDGX agent

arXiv:2607.12800v1 Announce Type: new Abstract: Learning broad world knowledge directly from raw visual data is a fundamental capability of intelligence. We introduce UniVR, the first investigation in

Unveiling Complex Collective Behaviors from Simple Rewards

AgentsDGX agent

arXiv:2607.12861v1 Announce Type: cross Abstract: Multi-agent Reinforcement Learning (MARL) holds great potential for robot swarms, but the black-box nature of neural policies complicates strategic an

UR-VC: Unsupervised Robotic Value Correction for Time-Derived Progress Proxies

Local AiDGX agent

arXiv:2607.12892v1 Announce Type: cross Abstract: Modern robot learning systems increasingly rely on dense progress or value signals to evaluate intermediate states, guide policy learning, and detect

VanillaBench: The Hidden Accuracy Cost of Adversarial Robustness

Model ReleasesDGX agent

arXiv:2607.12545v1 Announce Type: cross Abstract: Adversarial robustness research has produced hundreds of defended models over the past decade, yet the literature almost universally reports robustnes

Variational Mixture of Graph Neural Experts for Alzheimer's Disease Recognition across Frequency Bands in EEG Brain Networks

TutorialsDGX agent

arXiv:2510.11917v2 Announce Type: replace Abstract: Dementia disorders such as Alzheimer's disease (AD) and frontotemporal dementia (FTD) exhibit overlapping electrophysiological signatures in EEG tha

Verifier-Based Reinforcement Fine-Tuning of Reasoning Models for Thermal Energy Storage Control

Model ReleasesDGX agent

arXiv:2607.12856v1 Announce Type: new Abstract: Buildings are expected to shift cooling loads in response to grid conditions. Thermal energy storage (TES) enables this shift, but scheduling it well re

Vertical Standardisation for High-Risk AI Systems under the EU AI Act: A Domain-Specific Framework for Algorithmic Hiring

SafetyDGX agent

arXiv:2607.12588v1 Announce Type: new Abstract: According to the recent European legislation, high-risk AI systems will have to adapt in order to comply with requirements related to specific areas, li

ViCo3D: Empowering LiDAR-based Collaborative 3D Object Detection with Vision Foundation Models

AgentsDGX agent

arXiv:2607.12959v1 Announce Type: new Abstract: LiDAR-based collaborative 3D perception in Vehicle-to-Everything (V2X) systems typically relies on fusing bird's-eye-view (BEV) features across agents.

ViHoRec: A Quality-Controlled Vietnamese Hotel Recommendation Dataset and Cold-Start Benchmark

Model ReleasesDGX agent

arXiv:2607.12946v1 Announce Type: cross Abstract: Recommender-system research for Vietnamese remains limited by the absence of a public, well-documented hotel interaction resource. Building such a res

Virtual Chromoendscopy with Tunable Visibility Enhancement

ResearchDGX agent

arXiv:2607.12416v1 Announce Type: new Abstract: Chromoendoscopy (CE) is a common clinical practice that sprays indigo carmine blue dye onto the gastric surface to improve the visibility of gastric les

VISA: VLM-Guided Instance Semantic Auditing for 3D Occupancy World Models

AgentsDGX agent

arXiv:2606.13460v2 Announce Type: replace Abstract: Semantic 3D occupancy provides a voxelized world state for autonomous driving and robot decision making, but object and rare-class errors can affect

VisCo: Leveraging Large Language Models as Intrinsic Encoders for Visual Token Compression

Model ReleasesDGX agent

arXiv:2607.12756v1 Announce Type: new Abstract: Vision-language models (VLMs) process large numbers of visual tokens, resulting in substantial inference latency and memory overhead. This has motivated

← Previous
1…198199200201202…998
Next →