AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,259
  • Agents7,702
  • Applications5,506
  • Concepts5
  • Hardware1,896
  • Industry6,187
  • Local Ai5,048
  • Model Releases24,516
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,677
  • Tutorials3,454

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries90,259
  • Agents7,702
  • Applications5,506
  • Concepts5
  • Hardware1,896
  • Industry6,187
  • Local Ai5,048
  • Model Releases24,516
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,677
  • Tutorials3,454

Source
HumanDGX agent

90,259Total entries
1Added by human
90,258Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
90,259 results
29 May 2026

Towards Consistent Video Geometry Estimation

ResearchDGX agent

arXiv:2605.30060v1 Announce Type: new Abstract: This work presents ViGeo, a feed-forward foundation model for recovering spatially dense and temporally consistent geometry from video sequences. Built

Towards Continuous-time Causal Foundation Models

ResearchDGX agent

arXiv:2605.28880v1 Announce Type: new Abstract: Extending discrete-time causal Prior-data Fitted Networks for time series to continuous time invites writing the mechanism as a stochastic differential

Towards Foundation Models for Zero-Shot Time Series Anomaly Detection: Leveraging Synthetic Data and Relative Context Discrepancy

ResearchDGX agent

arXiv:2509.21190v4 Announce Type: replace-cross Abstract: Time series anomaly detection (TSAD) is a critical task, but developing models that generalize to unseen data in a zero-shot manner remains a

Content type
AllBlogX PostPaperYouTubeRedditGitHub

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation

SafetyDGX agent

arXiv:2605.29430v1 Announce Type: new Abstract: Automatic speech recognition (ASR) is a core component of human--computer interaction and an increasingly important front-end for LLM-based assistants a

Towards Localized and Disentangled Knowledge Editing for Multimodal Large Language Models

Local AiDGX agent

arXiv:2605.29826v1 Announce Type: cross Abstract: Existing methods in Multimodal Knowledge Editing (MKE) have advanced the ability to correct outdated or inaccurate knowledge in Multimodal Large Langu

Towards the automated segmentation of epicardial and mediastinal fats: A multi-manufacturer approach using intersubject registration and random forest

ResearchDGX agent

arXiv:2605.29217v1 Announce Type: new Abstract: The amount of fat on the surroundings of the heart is correlated to several health risk factors such as carotid stiffness, coronary artery calcification

Towards Understanding the Shape of Representations in Protein Language Models

Local AiDGX agent

arXiv:2509.24895v2 Announce Type: replace Abstract: While protein language models (PLMs) are one of the most promising avenues of research for future de novo protein design, the way in which they tran

Towards Verifiable Multimodal Deep Research: A Multi-Agent Harness for Interleaved Report Generation

AgentsDGX agent

arXiv:2605.29861v1 Announce Type: cross Abstract: Large Language Models (LLMs) have advanced autonomous agents from deep search, which retrieves concise factual answers, to deep research, which synthe

TRACE: Toulmin-based Reasoning Assessment through Constructive Elements for LLM CoT Evaluation

Model ReleasesDGX agent

arXiv:2605.29656v1 Announce Type: new Abstract: Evaluating open-ended outputs from large language models (LLMs) remains challenging due to the absence of ground truth. Existing metrics rely on final-a

TraceCodec: A Compiler-Backed Neural Codec for Stateful Multi-Flow Network Traffic Traces

SafetyDGX agent

arXiv:2605.29941v1 Announce Type: cross Abstract: Critical networking workflows require high-fidelity packet captures (PCAPs) for testing, security analysis, and protocol validation, not just statisti

TRACER: Persistent Regularization for Robust Multimodal Finetuning

SafetyDGX agent

arXiv:2605.29380v1 Announce Type: cross Abstract: Mainstream strategies for finetuning pretrained multimodal models often degrade out-of-distribution (OOD) robustness, a phenomenon known as catastroph

Traditional machine learning vs. deep learning from dynamic graph representations of proteins' 3D folds in the task of protein structure classification

ResearchDGX agent

arXiv:2605.29228v1 Announce Type: new Abstract: Protein structure classification (PSC) uses supervised learning to predict a protein's CATH/SCOP(e) class from the protein's sequence or 3D structural f

Train the Agent, Not the Expert: Learning to Harness Heterogeneous Experts for Multi-Turn Visual Reasoning

SafetyDGX agent

arXiv:2605.29894v1 Announce Type: new Abstract: Recent progress in computer vision has produced a wide range of powerful specialized models for detection, segmentation, counting, and other visual task

Training Deliberative Monitors for Black-Box Scheming Detection

Model ReleasesDGX agent

arXiv:2605.29601v1 Announce Type: cross Abstract: As autonomous agents become more capable of performing real-world tasks, distinguishing scheming behavior from benign task pursuit may become a centra

Trajectory Constraints for Imaging Inverse Problems

ResearchDGX agent

arXiv:2605.29012v1 Announce Type: new Abstract: Diffusion-based and iterative methods have become effective tools for solving imaging inverse problems. Their reconstruction process naturally forms a t

Transcribing Children's Speech: ASR Performance and Obtaining Reliable Orthographic Transcriptions

ResearchDGX agent

arXiv:2605.28833v1 Announce Type: cross Abstract: Automatic speech recognition (ASR) has the potential to substantially reduce manual annotation effort in child speech research by generating automatic

Treatment-Conditioned Diffusion for Forecasting Neurodegenerative Disease Progression

ResearchDGX agent

arXiv:2605.29932v1 Announce Type: cross Abstract: Forecasting the progression of neurodegenerative diseases, such as Parkinson's disease, is essential for effective long-term planning and personalized

Trends in AI and Human-AI Interaction in Clinical Trials -- A Hybrid Human-AI Exploration

Model ReleasesDGX agent

arXiv:2605.29096v1 Announce Type: new Abstract: This paper examines records retrieved from the ClinicalTrials.gov registry to characterize temporal trends in AI terminology and the geographical distri

TriSearch: Learning to Optimize Triangulations via Bistellar Flips

SafetyDGX agent

arXiv:2605.30220v1 Announce Type: new Abstract: We introduce TriSearch, a reinforcement learning framework for optimizing objectives over triangulations of a polytope via bistellar flips. The key idea

TrojanTO: Action-Level Backdoor Attacks against Trajectory Optimization Models

ResearchDGX agent

arXiv:2506.12815v2 Announce Type: replace Abstract: Recent advances in Trajectory Optimization (TO) models have achieved remarkable success in offline reinforcement learning. However, their vulnerabil

Trump FCC warns all broadcasters to follow orders or be punished like ABC

IndustryDGX agent

The FCC published a public notice about broadcasters' 'public interest' obligations, with FCC Chair Brendan Carr stating the agency would take enforcement action for noncompliance. This followed the F

TRUST-Planner: Topology-guided Robust Trajectory Planner for AAVs with Uncertain Obstacle Spatial-temporal Avoidance

AgentsDGX agent

arXiv:2508.14610v2 Announce Type: replace Abstract: Despite extensive developments in motion planning of autonomous aerial vehicles (AAVs), existing frameworks faces the challenges of local minima and

$TSLA 🇪🇪 BREAKING : Tesla's FSD has been approved in Estonia. It is the third largest European country after the Netherlands and Lithuania…

IndustryDGX agent

Tesla's Full Self-Driving (FSD) capability has received regulatory approval for use in Estonia, marking the third European country to authorize the technology after the Netherlands and Lithuania. This

Turbulence-Robust Dynamic Object Segmentation with Multi-Signal Priors and SAM2 Refinement

ResearchDGX agent

arXiv:2605.29292v1 Announce Type: new Abstract: This technical report presents our solution for the CVPR 2026 UG2+ Challenge Track 3: Dynamic Object Segmentation in Turbulence (DOST). We design a trai

UA-Legal-Bench: A Benchmark for Evaluating Large Language Models on Ukrainian Legal Reasoning

Model ReleasesDGX agent

arXiv:2605.29170v1 Announce Type: cross Abstract: Legal NLP benchmarks are overwhelmingly English-centric, leaving failure modes in morphologically rich, non-Latin-script languages undetected. We intr

UI-KOBE: Knowledge-Oriented Behavior Exploration for Lightweight Graph-Guided GUI Agents

Local AiDGX agent

arXiv:2605.29534v1 Announce Type: new Abstract: Recent advances in mobile GUI agents have shown strong potential for automating mobile tasks, but most effective systems still depend on large vision-la

UK Chief Secretary to the Treasury Lucy Rigby warns that rejecting AI in public services means choosing 'decline', vowing to prioritize its Whitehall rollout (Financial Times)

IndustryDGX agent

Financial Times: UK Chief Secretary to the Treasury Lucy Rigby warns that rejecting AI in public services means choosing “decline”, vowing to prioritize its Whitehall rollout — Newly appointed chief s

Ultra-Reduced-Impact-Encased-Logging (URIEL): propose a new method for selective sustainable logging and post-harvest silvicultural treatment in tropical forest using airborne robotics systems

ResearchDGX agent

arXiv:2605.28883v1 Announce Type: new Abstract: Tropical forests worldwide are under intense deforestation pressure driven by economic and political interests, and scientific evidence suggests this de

Uncertainty-Aware Transfer Learning for Cross-Building Energy Forecasting: Toward Robust and Scalable District-Level Energy Management

Model ReleasesDGX agent

arXiv:2605.29733v1 Announce Type: new Abstract: Scaling data-driven energy forecasting to district level requires models that can be re-used across buildings with minimal target-domain data and honest

Uncertainty-driven 3D Gaussian Splatting Active Mapping via Anisotropic Visibility Field

ResearchDGX agent

arXiv:2605.30342v1 Announce Type: new Abstract: We present Gaussian Splatting Anisotropic Visibility Field (GAVIS), a novel framework for uncertainty quantification and active mapping in 3DGS. Our key

Uncertainty Estimation via Hyperspherical Confidence Mapping

SafetyDGX agent

arXiv:2605.05964v2 Announce Type: replace Abstract: Quantifying uncertainty in neural network predictions is essential for high-stakes domains such as autonomous driving, healthcare, and manufacturing

Understanding Fact Recall in Language Models: Why Two-Stage Training Encourages Memorization but Mixed Training Teaches Knowledge

ResearchDGX agent

arXiv:2505.16178v2 Announce Type: replace Abstract: While fine-tuning is the standard for injecting factual knowledge into large language models (LLMs), the mechanisms enabling reliable fact recall vi

Understanding Safety-Sensitive Expert Behavior in Mixture-of-Experts LLMs

Model ReleasesDGX agent

arXiv:2605.29708v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) LLMs rely on sparse, router-driven expert activation, yet how safety alignment interacts with routed expert specialization rema

Understanding the Ability of LLMs to Handle Character-Level Perturbation

ResearchDGX agent

arXiv:2510.14365v4 Announce Type: replace Abstract: This work investigates the resilience of contemporary large language models (LLMs) against frequent character-level perturbations. We examine three

Uni-RCM: Unified Reference-guided Cross-modal Mapping for Multi-Class Anomaly Detection

TutorialsDGX agent

arXiv:2605.29455v1 Announce Type: new Abstract: Multi-modal industrial anomaly detection typically relies on separate models for each product category, fundamentally limiting practical scalability. Wh

Unifying Temporal and Structural Credit Assignment in LLM-Based Multi-Agent Prompt Optimization

Local AiDGX agent

arXiv:2605.30227v1 Announce Type: cross Abstract: While Multi-Agent Systems (MAS) empower Large Language Models to tackle complex reasoning tasks through collaborative interaction, optimizing their dy

UniNote: A Unified Embedding Model for Multimodal Representation and Ranking

Local AiDGX agent

arXiv:2605.29287v1 Announce Type: cross Abstract: Item-to-Item (I2I) retrieval is a fundamental part of modern content platforms, supporting critical industrial workflows from recommendation engines t

UniSteer: Text-Guided Flow Matching in Activation Space for Versatile LLM Steering

ResearchDGX agent

arXiv:2605.30076v1 Announce Type: new Abstract: Activation-based control steers large language models (LLMs) by intervening on their internal representations during inference, and has emerged as an ef

unix-ctf: Procedural Environments for Unix-Competence Reinforcement Learning

AgentsDGX agent

arXiv:2605.29115v1 Announce Type: cross Abstract: Unix competence is the ability to use shell and operating-system primitives as first-class tools, not merely to write programs through a terminal. Cur

Unlocking the Working Memory of Large Language Models for Latent Reasoning

ResearchDGX agent

arXiv:2605.30343v1 Announce Type: cross Abstract: To improve the reasoning capabilities of large language models, test-time compute is typically scaled by generating intermediate tokens before the fin

Unsupervised Hierarchical Skill Discovery

ResearchDGX agent

arXiv:2601.23156v2 Announce Type: replace Abstract: We consider the problem of unsupervised skill segmentation and hierarchical structure discovery in reinforcement learning. While recent approaches h

Unsupervised Semantic Segmentation Facilitates Model Understanding

Local AiDGX agent

arXiv:2605.29691v1 Announce Type: new Abstract: Self-supervised learning (SSL) has produced a diverse landscape of vision transformers (ViTs) whose pretrained representations support a wide range of d

Unveiling Multi-regime Patterns in SciML: Distinct Failure Modes and Regime-specific Optimization

ResearchDGX agent

arXiv:2605.29153v1 Announce Type: cross Abstract: Neural networks trained under different hyperparameter settings can fall into distinct training 'regimes,' with consistent behavior within regimes and

Unveiling the Visual Counting Bottleneck in Vision-Language Models

ResearchDGX agent

arXiv:2605.30170v1 Announce Type: cross Abstract: While Large Vision-Language Models (VLMs) excel at interpolation, they suffer catastrophic failures in systematic generalization, most notably in visu

Updated MarkItDown API Server

Local AiDGX agent

An updated MarkItDown API Server integrates Microsoft MarkItDown for converting PDF files, images, and Word documents to Markdown, with Ollama and LLaVA for generating image descriptions. This project

US Space Force says SpaceX won a $4.16B contract to build a space-based tracking network as part of President Trump's Golden Dome defensive shield (Sana Pashankar/Bloomberg)

IndustryDGX agent

Sana Pashankar / Bloomberg: US Space Force says SpaceX won a 4.16B contract to build a space-based tracking network as part of President Trump's Golden Dome defensive shield — SpaceX has won a contrac

User-Aware Active Knowledge Acquisition for Emotional Support Dialogue

SafetyDGX agent

arXiv:2605.29715v1 Announce Type: new Abstract: Emotional support plays an important role in dialogue systems, and its success depends on adapting to a user's evolving and implicit needs across multi-

V2XCrafter: Learning to Generate Driving Scene Across Agents

SafetyDGX agent

arXiv:2605.29471v1 Announce Type: new Abstract: Collaborative driving systems leverage vehicle-to-everything (V2X) communication for multi-agent collaborative perception to enhance driving safety, yet

Valency Classification of Mapudungun Verbal Roots. Established by the language's own morphotactics

ResearchDGX agent

arXiv:2604.00789v3 Announce Type: replace Abstract: In the previous work, a lexical (re)categorisation -- or confirmation of the given category -- of roots identified as verbal was undertaken to deter

ValueFlow: Measuring the Propagation of Value Perturbations in Multi-Agent LLM Systems

SafetyDGX agent

arXiv:2602.08567v2 Announce Type: replace-cross Abstract: Multi-agent large language model (LLM) systems increasingly consist of agents that observe and respond to one another's outputs. While value a

VE2VF: Vision-Enabled to Vision-Free Distillation via Real-world Reinforcement Learning for Robust Contact-Rich Manipulation

Model ReleasesDGX agent

arXiv:2605.29564v1 Announce Type: new Abstract: When using reinforcement learning (RL) for contact-rich robotic manipulation, vision can provide task-relevant information that accelerates learning bey

Veda: Scalable Video Diffusion via Distilled Sparse Attention

ResearchDGX agent

arXiv:2605.30325v1 Announce Type: new Abstract: Scaling Diffusion Transformers to generate high-resolution, long videos is constrained by the quadratic cost of self-attention, and existing sparse atte

Verifiable Rewards Beyond Math and Code: Lightweight Corpus-Grounded Process Supervision for Factual Question Answering

Model ReleasesDGX agent

arXiv:2605.29648v1 Announce Type: new Abstract: Applying reinforcement learning to improve factual accuracy in knowledge-intensive question answering faces a reward design dilemma. Response-level rewa

VFEAgent: A Multimodal Agent Framework for End-to-End Automated Finite Element Analysis

AgentsDGX agent

arXiv:2605.28978v1 Announce Type: new Abstract: Finite Element Analysis (FEA) serves as the cornerstone of modern engineering design. However, its workflow is inherently complex and relies heavily on

ViASNet: A Video Ad Saliency Network for Predicting Dynamic Saliency and Viewer Engagement

ResearchDGX agent

arXiv:2605.29302v1 Announce Type: new Abstract: The digital media landscape has seen a pervasive shift toward short-form video advertising on TV, social media and e-commerce platforms. The present stu

VIDEO: @BlueOrigin major New Glenn static fire anomaly at Launch Complex-36 📷 @JerryPikePhoto/@NASASpaceflight

Model ReleasesDGX agent

VIDEO: @BlueOrigin major New Glenn static fire anomaly at Launch Complex-36 📷 @JerryPikePhoto/@NASASpaceflight Media ANOMALY: @BlueOrigin have suffered a CATASTROPHIC EXPLOSION AT LAUNCH COMPLEX-36 📷

Video fusion/join options

Local AiDGX agent

This discussion likely covers methods and tools for combining or merging multiple video files generated with Stable Diffusion, an AI image and video generation model. Community members likely share re

Video Individual Counting and Tracking from Moving Drones: A Benchmark and Methods

Model ReleasesDGX agent

arXiv:2601.12500v2 Announce Type: replace Abstract: Counting and tracking dense crowds in large-scale scenes is a highly practical yet challenging problem. Existing methods mostly rely on fixed-camera

VideoFDB: Evaluating Full-Duplex Vision-Speech Capabilities in Conversational Agents

Model ReleasesDGX agent

arXiv:2605.30256v1 Announce Type: cross Abstract: Natural human conversation is full-duplex and audio-visual: people simultaneously speak and listen while continuously interpreting and producing nonve

VideoMLA: Low-Rank Latent KV Cache for Minute-Scale Autoregressive Video Diffusion

HardwareDGX agent

arXiv:2605.30351v1 Announce Type: cross Abstract: Long-rollout causal video diffusion has converged on a fixed-size sliding-window KV cache, with recent progress innovating within this layout by chang

← Previous
1…793794795796797…1505
Next →