AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
59,840 results
Safety

Inverse Reinforcement Learning with Just Classification and a Few Regressions

DGX agent

arXiv:2509.21172v2 Announce Type: replace Abstract: Inverse reinforcement learning (IRL) aims to infer rewards from observed behavior, but rewards are not identified from the policy alone: many reward

safetyarxiv-cs-lg
11 May 2026
Research
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

It Just Takes Two: Scaling Amortized Inference to Large Sets

DGX agent

arXiv:2605.07972v1 Announce Type: cross Abstract: Neural posterior estimation has emerged as a powerful tool for amortized inference, with growing adoption across scientific and applied domains. In ma

researcharxiv-cs-ai
11 May 2026
Research

Latent Order Bandits

DGX agent

arXiv:2605.07304v1 Announce Type: new Abstract: Bandit algorithms solve diverse sequential decision-making problems, but are often too sample-inefficient for from-scratch personalization. To substanti

researcharxiv-cs-lg
11 May 2026
Safety

Learned Lyapunov Shielding for Adaptive Control

DGX agent

arXiv:2605.06934v1 Announce Type: new Abstract: We augment the Slotine--Li adaptive controller for Euler--Lagrange systems with three learned components: a structured-quadratic Lyapunov function (V_ps

safetyarxiv-cs-lg
11 May 2026
Local Ai

Learning to Communicate Locally for Large-Scale Multi-Agent Pathfinding

DGX agent

arXiv:2605.07637v1 Announce Type: new Abstract: Multi-agent pathfinding (MAPF) is a widely used abstraction for multi-robot trajectory planning problems, where multiple homogeneous agents move simulta

local-aiarxiv-cs-ai
11 May 2026
Safety

Learning to Track Instance from Single Nature Language Description

DGX agent

arXiv:2605.07064v1 Announce Type: new Abstract: How to achieve vision-language (VL) tracking using natural language descriptions from a video sequence extbf{without relying on any bounding-box ground

safetyarxiv-cs-cv
11 May 2026
Research

LLMs are not (consistently) Bayesian: Quantifying internal (in)consistencies of LLMs' probabilistic beliefs

DGX agent

arXiv:2605.06915v1 Announce Type: new Abstract: Modern AI systems are being deployed in complex domains such as medicine, science, and law, where it is important that they not only produce correct ans

researcharxiv-cs-lg
11 May 2026
Tutorials

// LLMs Improving LLMs // Interesting progress the past of couple of weeks around self-improving AI agents. If autoresearch was interesting,…

DGX agent

// LLMs Improving LLMs // Interesting progress the past of couple of weeks around self-improving AI agents. If autoresearch was interesting, you will like this read. (bookmark it) We've been hand-tuni

tutorialsdair-ai--x
11 May 2026
Tutorials

LookWhen? Fast Video Recognition by Learning When, Where, and What to Compute

DGX agent

arXiv:2605.06809v1 Announce Type: new Abstract: Transformers dominate video recognition. They split videos into tokens, and processing them has expensive superlinear computational cost. Yet videos are

tutorialsarxiv-cs-cv
11 May 2026
Safety

MAGIQ: A Post-Quantum Multi-Agentic AI Governance System with Provable Security

DGX agent

arXiv:2605.06933v1 Announce Type: new Abstract: Our computing ecosystem is being transformed by two emerging paradigms: the increased deployment of agentic AI systems and advancements in quantum compu

safetyarxiv-cs-lg
11 May 2026
Agents

mathsf{VISTA}: Decentralized Machine Learning in Adversary Dominated Environments

DGX agent

arXiv:2605.07841v1 Announce Type: cross Abstract: Decentralized machine learning often relies on outsourcing computations, such as gradient evaluations, to untrusted worker nodes. Existing robust aggr

agentsarxiv-cs-ai
11 May 2026
Research

Measuring and Mitigating the Distributional Gap Between Real and Simulated User Behaviors

DGX agent

arXiv:2605.07847v1 Announce Type: new Abstract: As user simulators are increasingly used for interactive training and evaluation of AI assistants, it is essential that they represent the diverse behav

researcharxiv-cs-cl
11 May 2026
Research

Mechanistic Interpretability with Sparse Autoencoder Neural Operators

DGX agent

arXiv:2509.03738v4 Announce Type: replace-cross Abstract: We introduce sparse autoencoder neural operators (SAE-NOs), a new class of sparse autoencoders that operate in function spaces rather than fix

researcharxiv-cs-ai
11 May 2026
Applications

Medical Imaging Classification with Cold-Atom Reservoir Computing using Auto-Encoders and Surrogate-Driven Training

DGX agent

arXiv:2605.06727v1 Announce Type: new Abstract: We introduce a hybrid quantum-classical pipeline, based on neutral-atom reservoir computing, for medical image classification, focusing on the binary cl

applicationsarxiv-cs-lg
11 May 2026
Hardware

MemSearcher: Training LLMs to Reason, Search and Manage Memory via End-to-End Reinforcement Learning

DGX agent

arXiv:2511.02805v2 Announce Type: replace-cross Abstract: LLM-based search agents often concatenate the full interaction history into the context, producing long and noisy inputs, and increasing compu

hardwarearxiv-cs-ai
11 May 2026
Research

Multimodal Diffusion Transformer with Memory Bank for Scalable Long-Duration Talking Video Generation

DGX agent

arXiv:2411.16748v5 Announce Type: replace Abstract: Long-duration talking video synthesis faces enduring challenges in achieving high video quality, portrait consistency, temporal coherence, and compu

researcharxiv-cs-cv
11 May 2026
Safety

No Forgetting Learning: Buffer-free Continual Learning Classification

DGX agent

arXiv:2503.04638v3 Announce Type: replace Abstract: Most Continual Learning (CL) methods maintain performance on earlier tasks by storing exemplars in a replay buffer, introducing memory overhead that

safetyarxiv-cs-lg
11 May 2026
Research

Normalized Maximum Likelihood Code-Length on Riemannian Data Spaces

DGX agent

arXiv:2508.21466v2 Announce Type: replace Abstract: In recent years, with the large-scale expansion of graph data, there has been an increased focus on Riemannian manifold data spaces other than Eucli

researcharxiv-cs-lg
11 May 2026
Safety

Openclaw token consumption fell by half in a month, per openrouter data What happened?

DGX agent

Openclaw token consumption dropped 50% within a month according to OpenRouter usage data, as reported by Gary Marcus. The post likely discusses potential causes for this significant decline, such as c

safetygary-marcus--x
11 May 2026
Research

Optimising MFCC parameters for the automatic detection of respiratory diseases

DGX agent

arXiv:2408.07522v2 Announce Type: replace-cross Abstract: Voice signals originating from the respiratory tract are utilized as valuable acoustic biomarkers for the diagnosis and assessment of respirat

researcharxiv-cs-lg
11 May 2026
Tutorials

PAMPOS: Causal Transformer-based Trajectory Prediction for Attack-Agnostic Misbehavior Detection in V2X Networks

DGX agent

arXiv:2605.06833v1 Announce Type: cross Abstract: Misbehavior detection in Vehicle-to-Everything (V2X) networks is a second line of defense against insider falsification attacks that cryptographic mec

tutorialsarxiv-cs-ai
11 May 2026
Safety

Persistent-Transient Policy Evaluation for Markov Chains via Minimal Peripheral Quotients

DGX agent

arXiv:2602.00474v2 Announce Type: replace-cross Abstract: We study fixed-policy evaluation for finite Markov chains that may be reducible and periodic. Classical evaluation methods with gain and bias

safetyarxiv-cs-lg
11 May 2026
Research

PET-Adapter: Test-Time Domain Adaptation for Full and Limited-Angle PET Image Reconstruction

DGX agent

arXiv:2605.08030v1 Announce Type: new Abstract: Positron Emission Tomography (PET) image reconstruction is inherently challenged by Poisson noise and physical degradation factors, which are further ex

researcharxiv-cs-cv
11 May 2026
Hardware

Physics-Based Flow Matching for Full-Field Prediction of Silicon Photonic Devices

DGX agent

arXiv:2605.06929v1 Announce Type: cross Abstract: Designing photonic integrated circuits requires accurate electromagnetic field simulations, which remain computationally expensive even for simple dev

hardwarearxiv-cs-lg
11 May 2026
Research

Pre-training Enables Extraordinary All-optical Image Denoising

DGX agent

arXiv:2605.07810v1 Announce Type: cross Abstract: Optical neural networks are emerging as powerful machine learning and information processing tools because of their potential advantages in speed and

researcharxiv-cs-cv
11 May 2026
Applications

Predictive quality starts where defect detection stops

DGX agent

Predictive quality extends beyond traditional defect detection by using data analytics and machine learning to anticipate quality issues before they occur rather than simply identifying existing defec

applicationsdatabricks
11 May 2026
Safety

PRPO: Paragraph-level Policy Optimization for Vision-Language Deepfake Detection

DGX agent

arXiv:2509.26272v3 Announce Type: replace Abstract: The rapid rise of synthetic media has made deepfake detection a critical challenge for online safety and trust. Progress remains constrained by the

safetyarxiv-cs-cv
11 May 2026
Research

R^3L: Reasoning 3D Layouts from Relative Spatial Relations

DGX agent

arXiv:2605.06758v1 Announce Type: cross Abstract: Relative spatial relations provide a compact representation of spatial structure and are fundamental to relative spatial reasoning in 3D layout genera

researcharxiv-cs-ai
11 May 2026
Applications

Really doubt what Hinton says here. Self-play for games like Go is not like the open-ended real world.

DGX agent

Really doubt what Hinton says here. Self-play for games like Go is not like the open-ended real world. Geoffrey Hinton says AI systems can keep improving when they are not limited by human data Like A

applicationsgary-marcus--x
11 May 2026
Applications

RECON: Robust symmetry discovery via Explicit Canonical Orientation Normalization

DGX agent

arXiv:2505.13289v5 Announce Type: replace-cross Abstract: Real world data often exhibits unknown, instance-specific symmetries that rarely exactly match a transformation group G fixed a priori. Class-

applicationsarxiv-cs-cv
11 May 2026
Safety

Reinforcement Learning for Exponential Utility: Algorithms and Convergence in Discounted MDPs

DGX agent

arXiv:2605.08053v1 Announce Type: new Abstract: Reinforcement learning (RL) for exponential-utility optimization in discounted Markov decision processes (MDPs) lacks principled value-based algorithms.

safetyarxiv-cs-lg
11 May 2026
Research

Replicating Human Motivated Reasoning Studies with LLMs

DGX agent

arXiv:2601.16130v2 Announce Type: replace-cross Abstract: Motivated reasoning - the idea that individuals processing information may be motivated to either arrive at accurate beliefs or arrive at desi

researcharxiv-cs-ai
11 May 2026
Safety

Rethinking Importance Sampling in LLM Policy Optimization: A Cumulative Token Perspective

DGX agent

arXiv:2605.07331v1 Announce Type: cross Abstract: Reinforcement learning, including reinforcement learning with verifiable rewards (RLVR), has emerged as a powerful approach for LLM post-training. Cen

safetyarxiv-cs-ai
11 May 2026
Safety

Risk-Consistent Multiclass Learning from Random Label-Subset Membership Queries

DGX agent

arXiv:2605.07413v1 Announce Type: new Abstract: Obtaining accurate class labels is often costly or unreliable, and may also be limited by privacy or other practical conditions. Compared with asking an

safetyarxiv-cs-lg
11 May 2026
Safety

Robustness of Refugee-Matching Gains to Off-Policy Evaluation Choices

DGX agent

arXiv:2605.06686v1 Announce Type: new Abstract: Previous research has investigated the potential of refugee matching for boosting refugee outcomes, first considered by Bansak et al. (2018). This paper

safetyarxiv-cs-lg
11 May 2026
Safety

Same Signal, Opposite Meaning: Direction-Informed Adaptive Learning for LLM Agents

DGX agent

arXiv:2605.06908v1 Announce Type: cross Abstract: Adaptive test-time compute for LLM agents aims to invoke extra computation only when it improves performance. Existing methods typically use confidenc

safetyarxiv-cs-ai
11 May 2026
Safety

SB-TRPO: Towards Safe Reinforcement Learning with Hard Constraints

DGX agent

arXiv:2512.23770v3 Announce Type: replace-cross Abstract: In safety-critical domains, reinforcement learning (RL) agents must often satisfy strict, zero-cost safety constraints while accomplishing tas

safetyarxiv-cs-ai
11 May 2026
Agents

SCOUT: Closed-Loop in-vivo System for Continuous Methane Concentration Monitoring in Cattle

DGX agent

arXiv:2508.04056v2 Announce Type: replace Abstract: Enteric methane measurement from ruminant livestock faces fundamental trade-offs between accuracy and operational feasibility. Existing methods quan

agentsarxiv-cs-ro
11 May 2026
Agents

Searching for Privacy Risks in LLM Agents via Simulation

DGX agent

arXiv:2508.10880v3 Announce Type: replace-cross Abstract: The widespread deployment of LLM-based agents is likely to introduce a critical privacy threat: malicious agents that proactively engage other

agentsarxiv-cs-ai
11 May 2026
Local Ai

Seeing Across Skies and Streets: Feedforward 3D Reconstruction from Satellite, Drone, and Ground Images

DGX agent

arXiv:2605.07978v1 Announce Type: new Abstract: Cross-view localization classically asks: where does this ground image lie on the satellite tile? Existing methods are typically limited to 3-DoF estima

local-aiarxiv-cs-cv
11 May 2026
Local Ai

Self Driving Datasets: From 20 Million Papers to Nuanced Biomedical Knowledge at Scale

DGX agent

arXiv:2605.07022v1 Announce Type: new Abstract: Manually curated biomedical repositories -- spanning bioactivity, genomics, and chemistry -- are expensive to maintain, lag behind primary literature, a

local-aiarxiv-cs-lg
11 May 2026
Local Ai

Signal Reshaping for GRPO in Weak-Feedback Agentic Code Repair

DGX agent

arXiv:2605.07276v1 Announce Type: new Abstract: Code-agent RL often receives weak feedback: rollout-time signals are reliable and executable, but capture only necessary or surface conditions for task

local-aiarxiv-cs-ai
11 May 2026
Safety

SimCT: Recovering Lost Supervision for Cross-Tokenizer On-Policy Distillation

DGX agent

arXiv:2605.07711v1 Announce Type: new Abstract: On-policy distillation (OPD) is a standard tool for transferring teacher behavior to a smaller student, but it implicitly assumes that teacher and stude

safetyarxiv-cs-cl
11 May 2026
Research

SoLAR: Error-Resilient Streamable Long-Horizon Free-Viewpoint Video Reconstruction with Anchor Activation and Latent Recalibration

DGX agent

arXiv:2605.07346v1 Announce Type: new Abstract: Free-Viewpoint Video (FVV) has emerged as a cornerstone of next-generation immersive media systems and attracted widespread attention. Previous methods

researcharxiv-cs-cv
11 May 2026
Hardware

Sparse Attention as a Range Searching Problem: Towards an Inference-Efficient Index for KV Cache

DGX agent

arXiv:2605.06763v1 Announce Type: new Abstract: Sparse attention improves LLM inference efficiency by selecting a subset of key-value entries, but at the cost of potential accuracy degradation. In par

hardwarearxiv-cs-lg
11 May 2026
Research

Sparse Attention as Compact Kernel Regression

DGX agent

arXiv:2601.22766v3 Announce Type: replace Abstract: Recent work has revealed a link between self-attention mechanisms in transformers and test-time kernel regression via the Nadaraya-Watson estimator,

researcharxiv-cs-lg
11 May 2026
Research

Spectral Filtering for Complex Linear Dynamical Systems

DGX agent

arXiv:2601.22400v2 Announce Type: replace-cross Abstract: We study the problem of learning complex-valued linear dynamical systems (CLDS) with sector-bounded spectrum. This class captures oscillatory

researcharxiv-cs-ai
11 May 2026
Research

Spectral Surgery: Class-Targeted Post-Hoc Rebalancing via Hessian Spike Perturbation

DGX agent

arXiv:2605.07790v1 Announce Type: cross Abstract: The Hessian spectrum of trained deep networks exhibits a characteristic structure: a continuous bulk of near-zero eigenvalues and a small number of la

researcharxiv-cs-cv
11 May 2026
← Previous
1…12001201120212031204…1247
Next →