AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
Human
84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,851 results
16 Jul 2026

FixItFlow: Automated Troubleshooting Guide Generation from Cloud Incidents

TutorialsDGX agent

arXiv:2607.13035v1 Announce Type: cross Abstract: Cloud services experience frequent incidents that require rapid diagnosis and resolution. Troubleshooting guides help engineers respond consistently,

Flow-aware Optimal Navigation in Unsteady Flows through Reinforcement Learning

AgentsDGX agent

arXiv:2607.13553v1 Announce Type: cross Abstract: Autonomous robotic navigation in nonstationary time-varying fluid flows remains a fundamental challenge due to partial observability and the unpredict

FM^2: Unified Federated Foundation Models for Heterogeneous Multimodal Medical Imaging

Model ReleasesDGX agent

arXiv:2607.13386v1 Announce Type: new Abstract: Building foundation models for medical imaging requires pooling data across institutions, yet privacy regulations prohibit centralized aggregation. Exis

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

FOLIO: Focused Semantic Memory for Streaming Video Understanding

TutorialsDGX agent

arXiv:2607.13298v1 Announce Type: new Abstract: In online streaming video understanding, a video stream continues to arrive and queries may be issued at any time. Because streaming frames grow without

FreeLit: Paired-Free Indoor Relighting via Physics-Guided Diffusion

SafetyDGX agent

arXiv:2607.13656v1 Announce Type: new Abstract: Image-based indoor scene relighting remains challenging due to the complex interplay between cluttered geometry and local illumination, requiring precis

From Language to Navigation Goals: A Vision-Language Approach for Semantic Navigation of Mobile Robots Using RGB-D Perception

Model ReleasesDGX agent

arXiv:2607.13624v1 Announce Type: cross Abstract: Natural language interaction provides an intuitive way for non-expert users to communicate with robotic platforms. However, transforming user requests

From Novice to Expert: Cost-Aware Bandits for Evolving Worker Performance in Crowdsensing

TutorialsDGX agent

arXiv:2607.13546v1 Announce Type: new Abstract: Mobile crowdsensing (MC) recruits mobile users to perform sensing tasks using their smartphones, enabling large-scale applications such as traffic monit

From Pixels to States: Rethinking Interactive World Models as Game Engines

ResearchDGX agent

arXiv:2607.14076v1 Announce Type: new Abstract: Building interactive worlds that respond coherently to player actions has long been a shared goal of computer graphics, games, and artificial intelligen

From Prediction to Collaboration: Interactive Symbolic Music Analysis

Model ReleasesDGX agent

arXiv:2607.13587v1 Announce Type: cross Abstract: Automatic symbolic music analysis has made substantial progress, yet existing systems are typically designed for a single mode of use, such as full-sc

From Surface Forecasting to Observability Forecasting: A Latent World Model for Cloud-Aware EO Monitoring

Model ReleasesDGX agent

arXiv:2607.13651v1 Announce Type: new Abstract: The bottleneck of Earth Observation processing chains is not the arrival of new imagery but whether the surface is actually visible when the image arriv

Full-Pipeline Inference Optimization for MiMo-V2.5 Series: Pushing Hybrid SWA Efficiency to the Limit

HardwareDGX agent

arXiv:2607.13095v1 Announce Type: cross Abstract: We present a full-pipeline inference optimization for the MiMo-V2.5 model family, which combines Hybrid Sliding Window Attention (Hybrid SWA), sparse

Gauge-Invariant, Parameter-Insensitive Regularization for Potential Recovery from Flow on Directed Graphs

Model ReleasesDGX agent

arXiv:2607.13609v1 Announce Type: new Abstract: Recovering a latent potential from observed flow on a directed graph (a discrete Poisson problem with Dirichlet boundaries) is ill-posed, and the standa

Generalizable VLA Finetuning via Representation Anchoring and Language-Action Alignment

SafetyDGX agent

arXiv:2607.13429v1 Announce Type: cross Abstract: Finetuning a pretrained vision-language model (VLM) on robot demonstrations via behavior cloning (BC) has become the standard recipe for vision-langua

Generative Compilation: On-the-Fly Compiler Feedback as AI Generates Code

TutorialsDGX agent

arXiv:2607.13921v1 Announce Type: cross Abstract: Languages with rich static semantics, such as Rust, provide stronger guarantees for AI-generated code, but their strictness makes generation more diff

GeoAnchor: Collaborative Reasoning via Latent Decomposition for 3D Spatial Understanding

Local AiDGX agent

arXiv:2607.13454v1 Announce Type: cross Abstract: Although multimodal large language models (MLLMs) have achieved remarkable progress, understanding 3D spatial relationships from 2D images remains a c

GFlowRL: Scaling Distribution-Matching RL to Large Language Models

ResearchDGX agent

arXiv:2607.13394v1 Announce Type: cross Abstract: Generative Flow Networks (GFlowNets) offer a promising alternative to reward-maximizing reinforcement learning (RL) for large reasoning models, encour

GHR-VLM: Making Zero-Shot Transit Video Analytics Realizable with Grounded Hybrid Reasoning

ApplicationsDGX agent

arXiv:2607.13569v1 Announce Type: cross Abstract: Transit video understanding can provide valuable fine-grained data that conventional passenger counters and fare systems cannot capture. However, supe

GigaWorld-Policy-0.5: A Faster and Stronger WAM Empowered by AutoResearch

Local AiDGX agent

arXiv:2607.13960v1 Announce Type: new Abstract: World Action Models (WAMs) improve robot policy learning by jointly modeling actions and future visual observations, using future scene evolution as den

GNN-DIP: Neural Corridor Selection for Decomposition-Based Motion Planning

SafetyDGX agent

arXiv:2603.12361v2 Announce Type: replace Abstract: Motion planning through narrow passages remains a core challenge: sampling-based planners rarely place samples inside these narrow but critical regi

GPOcc++: Unified Sparse Gaussian Occupancy Prediction with Visual Geometry Priors

Model ReleasesDGX agent

arXiv:2607.13481v1 Announce Type: new Abstract: Accurate 3D scene understanding is fundamental to embodied intelligence and autonomous driving, where 3D occupancy provides a unified representation of

GPUSimBench: Towards Scalable and Reliable GPU-Accelerated Simulators in Embodied AI

Model ReleasesDGX agent

arXiv:2607.13059v1 Announce Type: new Abstract: Data-driven embodied AI is rapidly transitioning into a paradigm that scales training through massively parallel simulation, where GPU-accelerated simul

Graded Entity-Familiarity Readouts in Language Models: Polish Adaptation, Cross-Language Robustness, and Refusal Steering

Model ReleasesDGX agent

arXiv:2607.13568v1 Announce Type: cross Abstract: Can a language model estimate its familiarity with an entity before generating an answer? We study activations at the final prompt token in twelve ins

Graph Partitioning with Demands: Generalized Conductance and its Applications

ResearchDGX agent

arXiv:2607.13218v1 Announce Type: cross Abstract: In this work, we study various graph partitioning problems under a general demand model. In each such task, we are given a graph G=(V,E,c,w) with a ca

Greedy Volume Maximization of Gradient Embeddings for Long-Tailed Frame-Level Bioacoustic Active Learning

ResearchDGX agent

arXiv:2607.13555v1 Announce Type: cross Abstract: Bioacoustic call-type classification relies on costly expert annotation. Active learning can reduce this burden by selecting a small batch of segments

Groc-PO: Grounded Context Preference Optimization for Truthful Multimodal LLMs

SafetyDGX agent

arXiv:2607.13712v1 Announce Type: cross Abstract: Despite the rapid progress of Multimodal Large Language Models (MLLMs), they still suffer from untruthfulness issues, such as visual hallucinations, c

Grounded world models in biological organisms and future embodied AI

AgentsDGX agent

arXiv:2607.13560v1 Announce Type: cross Abstract: Recent advances in generative and embodied AI have been driven by large-scale predictive learning over multimodal data. However, the resulting systems

Harness Handbook: Making Evolving Agent Harnesses Readable,Navigable, and Editable

Local AiDGX agent

arXiv:2607.13285v1 Announce Type: new Abstract: The capability of a modern AI agent depends not only on its foundation model but also on its harness, which constructs prompts, manages state, invokes t

Heavy-Tailed Flow Matching via Random Clocks

Model ReleasesDGX agent

arXiv:2607.13841v1 Announce Type: new Abstract: Heavy-tailed data arise in many domains where rare events carry disproportionate importance, such as imbalanced image datasets, financial returns, and w

HEDGEHOG: Hierarchical Evaluation of Drug Generators Through Rigorous Filtration

Model ReleasesDGX agent

arXiv:2607.13155v1 Announce Type: new Abstract: Generative molecular models can support early drug discovery by proposing new candidate compounds de novo. In practice, useful candidates must balance t

HELP: Human-Efficient Large-Scale Robot Post-Training with Rollout Segmentation

SafetyDGX agent

arXiv:2607.09776v2 Announce Type: replace-cross Abstract: When adapting Vision Language Action (VLA) models to downstream tasks, multiple rounds of post-training are often required to progressively ad

Hierarchical F-Clustering: Approximation and Hardness of Clustering into Trees and Bounded Diameter Graphs

ResearchDGX agent

arXiv:2607.13217v1 Announce Type: cross Abstract: Consider the following variation on the Hierarchical Clustering problem: Usually, while building a hierarchical clustering, one recursively partitions

HIVE-3D: Hierarchical Voxel Enhancement for High-Quality 3D Scene Generation

ResearchDGX agent

arXiv:2607.13468v1 Announce Type: new Abstract: Recently, a line of works can generate impressive 3D objects from a single image, but they are limited by restricted representation resolution, making t

How Agents Ask for Permission: User Permissions for AI Agents, from Interfaces to Enforcement

AgentsDGX agent

arXiv:2607.13718v1 Announce Type: cross Abstract: As AI agents gain prevalance, users are increasingly exposed to the risks such systems entail. Prompt injection attacks, as well as hallucination, can

How Far Can Root Cause Analysis Go on Real-World Telemetry Data?

Model ReleasesDGX agent

arXiv:2607.13548v1 Announce Type: new Abstract: Identifying root causes in production microservice failures requires reasoning over large-scale, multimodal telemetry spanning metrics, logs, and traces

How the Hessian-Spectrum of Neural Networks Depends on Data

ResearchDGX agent

arXiv:2607.13631v1 Announce Type: new Abstract: The Hessian matrix is an important quantity of interest when it comes to studying the loss landscape and optimization dynamics in deep learning, as well

HRIBench: Benchmarking Interaction-Centric Human-Robot Collaboration

Model ReleasesDGX agent

arXiv:2607.13056v1 Announce Type: cross Abstract: Current vision-language-action (VLA) benchmarks primarily evaluate isolated manipulation skills while leaving human-robot interaction structure largel

HRO: Hierarchical Room-to-Object Framework for Zero-Shot Object Goal Navigation with Large Language Models

Local AiDGX agent

arXiv:2607.13072v1 Announce Type: cross Abstract: Zero-shot object-goal navigation aims to enable an intelligent agent to explore and navigate to objects of unknown categories in an unfamiliar environ

Human4K: A Large-Scale 4K Multi-View Mocap Dataset for Whole-Body 3D Human Reconstruction

SafetyDGX agent

arXiv:2607.13646v1 Announce Type: cross Abstract: Recent advances in 3D human reconstruction have improved overall performance, yet current models still fail in the most challenging real-world scenari

IMMNet: Hybrid Fusion of Model-based and Data-driven Approaches for Maneuvering Target Tracking

TutorialsDGX agent

arXiv:2607.13573v1 Announce Type: cross Abstract: Maneuvering target tracking in three-dimensional space remains a challenging problem due to complex motion dynamics and model mismatch. To address thi

Implementations of Quantum and Classical Topology-Aligned Architectures for Molecular Property Prediction

Model ReleasesDGX agent

arXiv:2607.13737v1 Announce Type: new Abstract: For low-data and resource-constrained regimes typical of quantum chemistry, parameter-efficient learning is a key objective. Here, we propose a topology

Improving Map Consistency in Graph-Based LiDAR SLAM Through Information-Aware Odometry and Retroactive Loop Closure

ResearchDGX agent

arXiv:2607.13516v1 Announce Type: new Abstract: High-quality maps are fundamental for robotics tasks such as navigation and planning. Although modern graph-based LiDAR SLAM systems achieve good trajec

Improving Medical Image Generative Models with Frechet Distance Loss

ResearchDGX agent

arXiv:2607.13300v1 Announce Type: new Abstract: Diffusion generative models have demonstrated immense potential for synthetic medical image generation. However, these models often struggle to capture

Improving Molecular Property Prediction in Small Language Models Using Graph-based Tools

AgentsDGX agent

arXiv:2607.13115v1 Announce Type: new Abstract: Small language models (SLMs) have shown promise for zero-shot molecular property prediction from SMILES strings, yet they often suffer from structural b

Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models

Model ReleasesDGX agent

arXiv:2607.13408v1 Announce Type: cross Abstract: Recent text-to-audio models generate high-quality audio, but often fail to follow instructions involving multiple sound events and temporal order. Thi

Improving Wind and Solar Power Prediction with Efficient Wrapper-based Feature Selection: An Empirical Study

ApplicationsDGX agent

arXiv:2607.14024v1 Announce Type: cross Abstract: With rising global energy demand and growing awareness of climate change and its impacts, the share of renewable energies in the global energy mix con

Industrial Dexterity Benchmark: A Hardware-Software Benchmarking Platform for Industrial Dexterous Manipulation

Model ReleasesDGX agent

arXiv:2607.14021v1 Announce Type: new Abstract: Dexterous manipulation remains a critical bottleneck in industrial automation; tasks such as cable routing, connector insertion, and precision assembly

Inference Economics of Enterprise Coding Agents: A Case Study of Cloud vs. On-Premise LLMs

Model ReleasesDGX agent

arXiv:2607.13080v1 Announce Type: cross Abstract: Autonomous coding agents force engineering organizations to choose between API-based frontier models -- strong reasoning at high token cost -- and on-

Interaction Protocol Shapes Moral Judgment in Multi-Agent Debate

Model ReleasesDGX agent

arXiv:2510.10002v3 Announce Type: replace Abstract: As large language models (LLMs) are increasingly deployed in sensitive everyday contexts -- offering personal advice, mental health support, and mor

Interventional Grounding Audits: Black-Box Premise-Dependency Tests for LLM Chain-of-Thought via Predicate Substitution

Model ReleasesDGX agent

arXiv:2607.13069v1 Announce Type: new Abstract: Large language models produce chain-of-thought (CoT) reasoning that appears logically sound yet may not genuinely depend on its stated premises. We intr

Introducing Human-Centeredness in AI-Assisted Lexicography

SafetyDGX agent

arXiv:2607.11808v2 Announce Type: replace-cross Abstract: This paper proposes a human-centered artificial intelligence (HCAI) framework for AI-assisted lexicography. While generative AI offers signifi

Inverse-LLaVA: Rethinking Multimodal Alignment via Text-to-Vision Mapping

SafetyDGX agent

arXiv:2508.12466v2 Announce Type: replace-cross Abstract: Traditional multimodal learning approaches rely on alignment pre-training to bridge vision and language modalities, typically by projecting vi

Is the Statistical Advantage Worth the Cost? An Empirical Comparison of KANs and MLPs for Structured Data Classification

Model ReleasesDGX agent

arXiv:2607.13413v1 Announce Type: cross Abstract: This study presents an empirical benchmarking comparison between Kolmogorov-Arnold Networks (KANs) and Multi-Layer Perceptrons (MLPs) on structured ta

Joint On-and-Off Policy Learning for Vision-and-Language Navigation

SafetyDGX agent

arXiv:2607.13461v1 Announce Type: new Abstract: Vision-and-Language Navigation (VLN) necessitates an embodied agent to navigate in the physical world by adhering to natural language instructions. Rece

Just-In-Time Scene Graph Growth: Combating Perceptual Saturation in Long-Horizon Robotics

ResearchDGX agent

arXiv:2607.13245v1 Announce Type: new Abstract: While 3D Scene Graphs (3DSGs) provide crucial structured representations for embodied agents, conventional Ahead-of-Time, build-everything-then-filter p

Kaleido: Algorithm-Hardware Co-Design for Video Diffusion Transformers by Exploiting Latent Space Correlations

ResearchDGX agent

arXiv:2607.13770v1 Announce Type: cross Abstract: Video diffusion transformers (vDiTs) generate high quality video but introduce extremely high compute cost due to the long diffusion timesteps and sel

Kepler-Encoder-v0.1: Towards a Multimodal Embedding Model for Robots

ResearchDGX agent

arXiv:2607.13522v1 Announce Type: cross Abstract: A robot must understand the state of its own body, but a camera sees only part of it. Force and contact leave almost no trace in a single frame, and r

Koopman-driven grip force prediction through EMG sensing

ResearchDGX agent

arXiv:2409.17340v2 Announce Type: replace-cross Abstract: Loss of hand function due to conditions like stroke or multiple sclerosis significantly impacts daily activities. Robotic rehabilitation provi

Lag Operator SSMs: A Geometric Framework for Structured State Space Modeling

ResearchDGX agent

arXiv:2512.18965v2 Announce Type: replace Abstract: Structured State Space Models (SSMs), which are at the heart of the recently popular Mamba architecture, are powerful tools for sequence modeling. H

LaME: Learning to Think in Latent Space for Multimodal Embedding via Information Bottleneck

ResearchDGX agent

arXiv:2606.13061v2 Announce Type: replace Abstract: Reasoning-driven universal multimodal embedding has advanced rapidly by introducing Chain-of-Thought (CoT) reasoning into the embedding pipeline. De

LAPO: Leave-One-Turn Attribution for Self-Generated Process Rewards in Multi-Turn Search Reasoning

SafetyDGX agent

arXiv:2607.13501v1 Announce Type: new Abstract: Reinforcement learning for multi-turn search reasoning typically relies on terminal outcome rewards, which cannot distinguish useful, redundant, and har

← Previous
1…186187188189190…998
Next →