AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
Human
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,295 results
6 Aug 2026

Sample Complexity of Multicalibration for Multilevel Properties

ResearchDGX agent

arXiv:2608.04288v1 Announce Type: new Abstract: Calibration requires a predictor to be unbiased after conditioning on its own predictions. Multicalibration asks for this guarantee simultaneously acros

SciCode-Verified: How Benchmark Defects Underestimated the Scientific-Coding Ability of Language Models

Model ReleasesDGX agent

arXiv:2608.04975v1 Announce Type: cross Abstract: SciCode is the standard measure of the scientific-coding ability of language models: research-level problems that demand both frontier scientific theo

SCOPE: Field-of-View-Aware Path Planning in Unknown 3D Environments via Safety-Volume Certification

SafetyDGX agent

arXiv:2608.04420v1 Announce Type: new Abstract: Safe navigation with a body-mounted limited-field-of-view sensor requires the complete robot-inflated volume of an intended motion to be observed and ve

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Scrouting: Cost-Aware Routing of Coding Agents by Scouting the Repository First

Model ReleasesDGX agent

arXiv:2608.04804v1 Announce Type: cross Abstract: Frontier language models can resolve repository-level software issues, but each attempt is expensive, and existing routers select a model from the iss

SEAR: Simple and Efficient Adaptation of Visual Geometric Transformers for Unpaired RGB+Thermal 3D Reconstruction

Model ReleasesDGX agent

arXiv:2603.18774v2 Announce Type: replace Abstract: Foundational feed-forward visual geometry models enable accurate and efficient camera pose estimation and scene reconstruction by learning strong sc

Searching for Sound-Meaning Collisions: Graph-Based Affordance Retrieval and Multi-Evaluator Ranking for Pun Translation at CLEF 2026 JOKER Task 2

ResearchDGX agent

arXiv:2608.04299v1 Announce Type: new Abstract: Fifteen years ago, Low proposed that pun translators should stop searching for equivalent words and instead search for new points of contact between sou

Season: Spectrum-Aware Orthogonal Gradient Refinement for Transfer-Based Adversarial Attacks

ResearchDGX agent

arXiv:2608.04441v1 Announce Type: new Abstract: Transfer-based adversarial attacks often transfer poorly across heterogeneous architectures because CNNs favor local textures while Vision Transformers

Seeking Physics in Diffusion Noise

ResearchDGX agent

arXiv:2603.14294v3 Announce Type: replace-cross Abstract: Do video diffusion models encode signals predictive of physical plausibility? We probe intermediate denoising representations of pretrained Di

Segmentation Pre-training for Label-Efficient Lumbar Spine Degeneration Grading

ResearchDGX agent

arXiv:2608.04810v1 Announce Type: new Abstract: Automated assessment of degenerative pathology in the lumbar spine on magnetic resonance imaging (MRI) requires access to large-scale datasets of expert

Semantic Frame Interpolation

Model ReleasesDGX agent

arXiv:2507.05173v2 Announce Type: replace Abstract: Generating intermediate video content of varying lengths based on given first and last frames, along with text prompt information, offers significan

Sequential Kernel-based Conditional Independence Testing via Adaptive Betting

SafetyDGX agent

arXiv:2606.18993v2 Announce Type: replace-cross Abstract: Testing conditional independence is fundamental yet intrinsically difficult: without additional assumptions, Type I error control is impossibl

SHIELD: A Segmented Hierarchical Memory Architecture for Energy-Efficient LLM Inference on Edge NPUs

ResearchDGX agent

arXiv:2604.07396v2 Announce Type: replace-cross Abstract: Large Language Model (LLM) inference on edge Neural Processing Units (NPUs) is fundamentally constrained by limited on-chip memory capacity. A

Short-term load forecasting under EU-AI Act Requirements in Safety-Critical Environments: Results from a 41-day live challenge on the aggregated German transmission-grid load

Model ReleasesDGX agent

arXiv:2608.05018v1 Announce Type: new Abstract: Short-term load forecasting (STLF) play a vital role in the electric power industry. It serves infrastructure that European and German law designate as

SIGNPOST-Bench: Benchmarking Text-Vision Conflict Resolution in Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2608.04244v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) make grounded predictions in real-world scenes by combining visual and textual cues, yet existing benchmarks

SiMDex: Mining Similar Egocentric Videos for Cross-Embodiment Dexterous Manipulation

ResearchDGX agent

arXiv:2608.04196v1 Announce Type: cross Abstract: Recent years have witnessed an explosive trend of scaling ego-centric human videos for robot manipulation, yet it remains unclear which data actually

Simile Understanding in Text-to-Image Models: An Evaluation Framework

ResearchDGX agent

arXiv:2608.04750v1 Announce Type: cross Abstract: Similes provide a compact and expressive way to describe visual characteristics in text prompts. Recent text-to-image models (t2i models) can produce

SimMOF: AI agent for Automated MOF Simulations

Model ReleasesDGX agent

arXiv:2603.29152v2 Announce Type: replace Abstract: Metal-organic frameworks (MOFs) offer a vast design space, and as such, computational simulations play a critical role in predicting their structura

Simultaneous estimation of multiple discrete unimodal distributions under stochastic order constraints

ApplicationsDGX agent

arXiv:2603.11532v2 Announce Type: replace-cross Abstract: We study the problem of estimating multiple discrete unimodal distributions, motivated by search behavior analysis on a real-world platform. T

SJEPA: Learning Elegant Latent Dynamics with Hybrid Symbolic-Neural Predictors

TutorialsDGX agent

arXiv:2608.04060v1 Announce Type: cross Abstract: Joint-embedding predictive architectures learn abstract states by predicting target embeddings from context embeddings, but their transition models ar

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses?

Model ReleasesDGX agent

arXiv:2608.04828v1 Announce Type: new Abstract: Large language model (LLM) agents increasingly rely on skills, structured documents that specify when to act, which procedure to follow, and which tools

SmartMage: Dynamic Modality Orchestration for 3D Scene Understanding

Model ReleasesDGX agent

arXiv:2608.05137v1 Announce Type: new Abstract: Understanding 3D scenes is fundamental to embodied intelligence, requiring joint reasoning over heterogeneous information from multiple modalities, incl

Social Pressure Breaks Majority Voting in LLM Safety Panels

SafetyDGX agent

arXiv:2608.04415v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used to detect unsafe content. A common approach is to combine judgments from a panel of models to correct

SparseDitto: Customizing GPU Kernels for Different Sparsity Patterns with LLM-Based Agentic System

HardwareDGX agent

arXiv:2608.05033v1 Announce Type: cross Abstract: Sparse matrix kernels are fundamental to scientific computing, graph analytics, and machine learning. Their GPU performance depends strongly on the in

Spatiotemporal Graph Transformer for Traffic Intelligence in Edge Computing

TutorialsDGX agent

arXiv:2608.04075v1 Announce Type: cross Abstract: Accurate traffic forecasting is essential for proactive resource management in edge computing, where service demand evolves dynamically across both sp

SpecDrop: Parameter-Free Category-Conditioned Routing for Modular Specialization

Model ReleasesDGX agent

arXiv:2608.04084v1 Announce Type: cross Abstract: Mixture-of-experts (MoE) networks pursue specialization through learned routers, gates, and load-balancing losses, yet at matched total-parameter budg

SpecRoll: Fast-Slow Verifier-Feedback Adaptation for Speculative Reinforcement Learning Rollouts

SafetyDGX agent

arXiv:2608.04962v1 Announce Type: cross Abstract: Reinforcement learning (RL) post-training improves the reasoning capabilities of large language models, but autoregressive rollout generation remains

Spend Bits Where Queries Look: KV Cache Vector Quantization with Attention-Preserving Transforms

ResearchDGX agent

arXiv:2608.04074v1 Announce Type: cross Abstract: Long-context LLM decoding reads the key-value (KV) cache at every step. Loading it takes longer than computing attention over it, so throughput is ban

SpikingNav: Robust Embodied Navigation with Spiking Neural Policies

SafetyDGX agent

arXiv:2608.05078v1 Announce Type: new Abstract: Embodied navigation requires an agent to make sequential decisions from egocentric observations in a physical environment. Existing Artificial Neural Ne

Splat-Based Metal Artifact Reduction in Cone-Beam CT via Compact Attenuation Modeling

SafetyDGX agent

arXiv:2608.04764v1 Announce Type: new Abstract: X-ray computed tomography (CT) suffers from severe metal artifacts when high-attenuation objects such as dental fillings or orthopedic implants are pres

Spoken Function Calling: A New Perspective on Spoken Language Understanding for Large Audio Language Models

AgentsDGX agent

arXiv:2608.05126v1 Announce Type: new Abstract: Spoken Language Understanding (SLU) is the core component of task-oriented dialogue systems and a pivotal link in achieving seamless human-agent interac

SPOT: Sparse Probing and Outcome Calibration for On-Policy Distillation

SafetyDGX agent

arXiv:2608.04419v1 Announce Type: cross Abstract: On-policy distillation (OPD) provides dense teacher supervision on student-generated trajectories, but standard reverse-KL training can assign insuffi

SSC: A Verifiable Structured Representation for Bimanual Manipulation Labelling

SafetyDGX agent

arXiv:2608.04425v1 Announce Type: new Abstract: Subtask labels decompose a long-horizon manipulation demonstration into shorter semantic segments for policy training and evaluation. Natural language d

SSTQ:Privacy-Preserving Vector Quantization via Subsampled Stochastic TurboQuant

ResearchDGX agent

arXiv:2608.05127v1 Announce Type: cross Abstract: Achieving local differential privacy in distributed optimization while maintaining low communication cost remains challenging. Existing vector quantiz

Stabilizing Multi-Attack Adversarial Training via Bandit Optimization

Model ReleasesDGX agent

arXiv:2511.12265v2 Announce Type: replace-cross Abstract: Deep Neural Networks (DNNs) remain vulnerable to diverse adversarial perturbations, motivating multi-attack adversarial training (AT) for impr

Stable Density Ridges: Consistency and Convergence of Subspace Constrained Mean Shift

ResearchDGX agent

arXiv:2608.05112v1 Announce Type: cross Abstract: The Subspace Constrained Mean Shift (SCMS) algorithm is a popular nonparametric method for extracting density ridges, which serve as a low-dimensional

State2State: Environment-Derived Mid-Training for LLM Agents

AgentsDGX agent

arXiv:2608.04934v1 Announce Type: new Abstract: Training LLM agents commonly relies on supervised fine-tuning from expert trajectories or online reinforcement learning over human-specified tasks with

Static Timing Orchestration for Tree-Structured Robot Control Firmware

ResearchDGX agent

arXiv:2608.04600v1 Announce Type: new Abstract: As robotic systems become increasingly complex, generating control firmware from structural description files has emerged as a promising paradigm for re

StaticSegFormer: An Efficient High-Performance Semantic Segmentation Based on Static Structured Pruning

HardwareDGX agent

arXiv:2608.04811v1 Announce Type: new Abstract: Structured pruning enhances the efficiency of deep neural networks (DNNs) by eliminating groups of parameters during inference. Previous methods mostly

Statistical learning theory and Occam's razor: Regularization

ResearchDGX agent

arXiv:2608.04049v1 Announce Type: cross Abstract: The principle of Occam's razor, which instructs us to prefer simplicity in inductive inference, has attracted much scrutiny both in the philosophy of

STEP-OPD: Rethinking Output Targets and Internal Dynamics in On-Policy Distillation for Diffusion Models

SafetyDGX agent

arXiv:2608.04887v1 Announce Type: new Abstract: On-policy distillation (OPD) has become an effective approach for consolidating multiple task-specialized image generation models into a single student.

Stochastic Emulation using Generalized Stratified Sampling for Performance-Based Risk Optimization of Structures

ResearchDGX agent

arXiv:2608.05006v1 Announce Type: new Abstract: Metamodels are instrumental in reducing the computational burden associated with nested reliability analyses and optimization loops in Performance-Based

Strategic Evaluation of Planning Strategies for LLM Agents in Cyber-Physical Systems

Model ReleasesDGX agent

arXiv:2608.04265v1 Announce Type: cross Abstract: Evaluations of LLM planning agents largely ask whether a task succeeds or a declared plan is followed. In strategic cyber-physical systems, a stronger

stratum: A System Infrastructure for Massive Agent-Centric ML Workloads

AgentsDGX agent

arXiv:2603.03589v3 Announce Type: replace-cross Abstract: Recent advances in large language models (LLMs) transform how machine learning (ML) pipelines are developed and evaluated. LLMs enable a new t

Strengthening Target-Language Features: SAE-Based Steering for Multilingual Inference

Model ReleasesDGX agent

arXiv:2608.04904v1 Announce Type: new Abstract: Multilingual large language models exhibit substantial performance differences across languages, while existing adaptation methods often require paramet

STRIVE: Probing Reasoning Limits in Graded Plausibility Generation and Evaluation

Model ReleasesDGX agent

arXiv:2608.04567v1 Announce Type: new Abstract: Event knowledge concerns who does what to whom. Psycholinguists use event-plausibility judgments to examine how this knowledge supports human language p

Structured LLM Reasoning for Zero-Shot Human--Robot Coordination Under Hidden Goals

SafetyDGX agent

arXiv:2608.04309v1 Announce Type: new Abstract: We present a structured large-language-model (LLM) architecture for zero-shot human--robot coordination in a cooperative construction task with private

Sublogarithmic Swap Regret in Multiplayer General-Sum Games via Hybrid Regularization

ResearchDGX agent

arXiv:2608.04149v1 Announce Type: cross Abstract: Swap regret governs the rate at which uncoupled learning dynamics converge to correlated equilibria in multiplayer general-sum games. Under full-infor

Suppression Sticks, Locality Is Fragile: A Closed-Loop Target-and-Control Audit of Task-Vector Negation in VLA Policies

Local AiDGX agent

arXiv:2608.04692v1 Announce Type: cross Abstract: Task-vector arithmetic offers a closed-form way to modify a model, yet its behavioral locality remains unclear in closed-loop robot control. We presen

SurgNarrator: A Generative Retrieval Framework for Surgical Video Understanding

TutorialsDGX agent

arXiv:2608.04676v1 Announce Type: new Abstract: Surgical procedures unfold as structured and recurring clinical events, whose real-time understanding via intraoperative surgical videos is critical for

SVI-DAG: A Structured Variational Inference Approach to Bayesian Causal Discovery

ResearchDGX agent

arXiv:2608.04930v1 Announce Type: cross Abstract: Bayesian causal discovery seeks to determine the posterior distribution of causal theories, which are interpreted as directed acyclic graphs (DAGs) th

Tactus: Open-Vocabulary Object Recognition from Low-Cost Pressure Arrays

Model ReleasesDGX agent

arXiv:2608.04043v1 Announce Type: new Abstract: Resistive pressure arrays are the cheapest and most widely shipped tactile sensors, yet tactile representation learning has concentrated on optical sens

Talk2Sensors: 3D Visual Grounding in Autonomous Driving via Sensor-Adaptive Physical Cue Matching

Model ReleasesDGX agent

arXiv:2608.04568v1 Announce Type: new Abstract: As a key capability for embodied intelligence, 3D visual grounding (3DVG) has been predominantly studied in indoor scenes with RGB-D or point-cloud inpu

Teaching Foundation Models to Read mmWave: Pose-Guided Kinematic Representation for Human Behavior Understanding

Model ReleasesDGX agent

arXiv:2608.04127v1 Announce Type: new Abstract: Large language model agents need to perceive human behavior in physical environments. Millimeter-wave (mmWave) radar provides a privacy-friendly and con

Teaching MLLMs to Say No: Generalized Referring Expression Comprehension via Refusal Calibrated GRPO

Local AiDGX agent

arXiv:2608.04698v1 Announce Type: cross Abstract: We tackle the challenging yet underexplored task of Generalized Referring Expression Comprehension (GREC), which requires a model to localize the obje

Teaching Nemotron Greek: Mining a Corpus, Adapting Retrieval, and Grounding Generation for Modern Greek across Specialist Domains

Model ReleasesDGX agent

arXiv:2608.05138v1 Announce Type: cross Abstract: Modern Greek is absent from NVIDIA's Nemotron retrieval models and from major multilingual retrieval benchmarks, despite being important for retrieval

Temporal Context Awareness: A Defense Framework Against Multi-turn Manipulation Attacks on Large Language Models

SafetyDGX agent

arXiv:2503.15560v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly vulnerable to sophisticated multi-turn manipulation attacks, where adversaries strategically build conte

Terminal Agents Suffice for Enterprise Automation

AgentsDGX agent

arXiv:2604.00073v3 Announce Type: replace-cross Abstract: There has been growing interest in building agents that can interact with digital platforms to execute meaningful enterprise tasks autonomousl

Test, then Route: How Language Models Execute In-Context Conditional Rules Across Models and Languages

Model ReleasesDGX agent

arXiv:2608.04183v1 Announce Type: new Abstract: When a language model follows an in-context conditional rule such as 'if P(x) then A else B,' does it assemble a runtime circuit with one module that te

Text2GraphQuery-Bench: A Text to Graph Query Benchmark

Model ReleasesDGX agent

arXiv:2602.11745v2 Announce Type: replace Abstract: Graph models are fundamental to data analysis in domains rich with complex relationships. Unlike SQL, which benefits from a rel- atively unified sta

The Calibration Floor: Format Repair Can Masquerade as Self-Correction at Small-to-Mid Scale

Model ReleasesDGX agent

arXiv:2608.04355v1 Announce Type: new Abstract: Accuracy changes after language-model self-revision are usually interpreted as changes in reasoning. We show this can fail at the answer-extraction boun

← Previous
1…6465666768…989
Next →