AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,965
  • Agents7,446
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,131
  • Local Ai4,857
  • Model Releases23,360
  • Research19,834
  • Safety13,174
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,965
  • Agents7,446
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,131
  • Local Ai4,857
  • Model Releases23,360
  • Research19,834
  • Safety13,174
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

Content type
AllBlog
86,965Total entries
1Added by human
86,964Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,458 results
Safety

Not only where, But when: Temporal Scheduling for RLVR

DGX agent

arXiv:2605.25381v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has become a core technique for post-training of Large Language Models (LLMs). While policy optimi

safetyarxiv-cs-lg
26 May 2026
Safety

On Reliability of Efficient Membership Inference Vulnerability Evaluation

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2605.25819v1 Announce Type: new Abstract: Membership inference attacks (MIAs) are popular methods for empirically assessing the leakage of sensitive information in the training data through mode

safetyarxiv-cs-lg
26 May 2026
Agents

OPAL: Omnidirectional Path-efficient Aerial 3D expLoration

DGX agent

arXiv:2605.25423v1 Announce Type: new Abstract: Autonomous exploration is critical for robot mapping unknown environments. Desirable characteristics of exploration algorithms include compute efficienc

agentsarxiv-cs-ro
26 May 2026
Research

PASA: A Principled Embedding-Space Watermarking Approach for LLM-Generated Text under Semantic-Invariant Attacks

DGX agent

arXiv:2605.10977v2 Announce Type: replace-cross Abstract: Watermarking for large language models (LLMs) is a promising approach for detecting LLM-generated text and enabling responsible deployment. Ho

researcharxiv-cs-ai
26 May 2026
Safety

PID-Guided Partial Alignment for Multimodal Decentralized Federated Learning

DGX agent

arXiv:2601.10012v2 Announce Type: replace Abstract: Multimodal decentralized federated learning (DFL) must support collaboration among agents that hold different modality subsets and often different m

safetyarxiv-cs-lg
26 May 2026
Applications

Prism: A Plug-in Reproducible Infrastructure for Scalable Multimodal Continual Instruction Tuning

DGX agent

arXiv:2605.26110v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) achieve versatility by reformulating diverse tasks into a unified instruction-following framework via instruc

applicationsarxiv-cs-cl
26 May 2026
Research

Proactive for Uncertainty: Cause-Aware Error Diagnosis and Interactive Clarification for Spoken Dialogue Systems

DGX agent

arXiv:2605.25404v1 Announce Type: new Abstract: Cascaded Automatic Speech Recognition -- Large Language Model (ASR-LLM) pipelines remain popular for industrial Spoken Dialogue Systems (SDS), primarily

researcharxiv-cs-cl
26 May 2026
Research

Probability Distributions Computed by Autoregressive Transformers

DGX agent

arXiv:2510.27118v4 Announce Type: replace Abstract: Most expressivity results for transformers treat them as language recognizers -- devices that accept or reject strings -- rather than as they are us

researcharxiv-cs-cl
26 May 2026
Agents

Proper Scoring Rules for Agentic Uncertainty Quantification

DGX agent

arXiv:2605.24756v1 Announce Type: new Abstract: Language-model agents increasingly emit uncertainty signals throughout a trajectory, but existing agentic UQ evaluations often conflate ranking usefulne

agentsarxiv-cs-ai
26 May 2026
Research

Psychometric Item Validation Using Virtual Respondents with Trait-Response Mediators

DGX agent

arXiv:2507.05890v4 Announce Type: replace-cross Abstract: As psychometric surveys are increasingly used to assess the traits of large language models (LLMs), the need for scalable survey item generati

researcharxiv-cs-ai
26 May 2026
Research

Quantifying Empirical Compute-Supervision Tradeoffs in RLVR

DGX agent

arXiv:2605.25252v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has become a standard paradigm for post-training language models, but in practice, verifiers are

researcharxiv-cs-ai
26 May 2026
Safety

Quantitative Evaluation of the Severity of Posttraumatic Stress Disorder through Transfer Learning from Specific Phobia Data

DGX agent

arXiv:2605.25933v1 Announce Type: cross Abstract: Posttraumatic stress disorder (PTSD) is a prevalent and debilitating mental health condition with significant personal and societal impacts. Current c

safetyarxiv-cs-ai
26 May 2026
Safety

RecGOAT: Graph Optimal Adaptive Transport for LLM-Enhanced Multimodal Recommendation with Dual Semantic Alignment

DGX agent

arXiv:2602.00682v2 Announce Type: replace-cross Abstract: Integrating large language model (LLM) representations into multimodal recommendation has shown promise, yet a fundamental challenge remains l

safetyarxiv-cs-ai
26 May 2026
Safety

Referential Security as a New Paradigm for AI Evaluations

DGX agent

arXiv:2605.25673v1 Announce Type: cross Abstract: Security evaluations inherently depend on stable identifiers. Any finding, audit, or regulatory decision must remain attached to the specific artifact

safetyarxiv-cs-ai
26 May 2026
Research

Representation-Guided Discrete Molecular Graph Retrosynthesis

DGX agent

arXiv:2605.24428v1 Announce Type: new Abstract: Stochastic process-based molecular graph generators have become the state of the art for template-free single-step retrosynthesis. However, these models

researcharxiv-cs-lg
26 May 2026
Local Ai

Retrieval-Augmented Detection of Potentially Abusive Clauses in Chilean Terms of Service

DGX agent

arXiv:2605.26019v1 Announce Type: cross Abstract: Online Terms of Service often function as contracts of adhesion, creating asymmetries that may expose consumers to potentially abusive clauses. In Chi

local-aiarxiv-cs-ai
26 May 2026
Research

Routing by Analogy: kNN-Augmented Expert Assignment for Mixture-of-Experts

DGX agent

arXiv:2601.02144v2 Announce Type: replace-cross Abstract: Mixture-of-Experts (MoE) architectures scale large language models efficiently by employing a parametric ``router'' to dispatch tokens to a sp

researcharxiv-cs-ai
26 May 2026
Safety

Safety-Critical Whole-Body Control for Humanoid Robots via Input-to-State Safe Control Barrier Functions

DGX agent

arXiv:2605.25546v1 Announce Type: new Abstract: Safety-critical control is essential for humanoid robots operating in complex human-centered environments, where physical safety constraints such as joi

safetyarxiv-cs-ro
26 May 2026
Research

Scaling Natural-Language Graph-Based Test Time Compute for Automated Theorem Proving

DGX agent

arXiv:2503.11657v3 Announce Type: replace Abstract: Large language models have demonstrated remarkable capabilities in natural language processing tasks requiring multi-step logical reasoning capabili

researcharxiv-cs-cl
26 May 2026
Agents

Security of OpenClaw Agents: Fundamentals, Attacks, and Countermeasures

DGX agent

arXiv:2605.25435v1 Announce Type: new Abstract: The rapid evolution of large language model (LLM)-driven autonomous agents has given rise to OpenClaw, a new class of open-source agent frameworks that

agentsarxiv-cs-ai
26 May 2026
Safety

Selective Latent Thinking: Adaptive Compression of LLM Reasoning Chains

DGX agent

arXiv:2605.25745v1 Announce Type: new Abstract: Explicit chain-of-thought (CoT) reasoning substantially improves the reasoning ability of large language models (LLMs), but incurs high inference cost d

safetyarxiv-cs-cl
26 May 2026
Research

Some Robustness Properties of Label Cleaning

DGX agent

arXiv:2509.11379v3 Announce Type: replace-cross Abstract: We demonstrate that learning procedures that rely on aggregated labels, e.g., label information distilled from noisy responses, enjoy robustne

researcharxiv-cs-lg
26 May 2026
Applications

SPACE: Unifying Symmetric and Asymmetric Routing Problems for Generalist Neural Solver

DGX agent

arXiv:2605.24484v1 Announce Type: new Abstract: Generalist neural routing solvers have shown great potential in solving diverse vehicle routing problems (VRPs) with a unified model. However, existing

applicationsarxiv-cs-ai
26 May 2026
Local Ai

Spatio-temporal, multi-field deep learning of shock propagation in meso-structured media

DGX agent

arXiv:2509.16139v5 Announce Type: replace Abstract: Predicting the extreme hydrodynamic response of porous and architected lattice materials is a fundamental challenge in high energy density physics,

local-aiarxiv-cs-lg
26 May 2026
Safety

SpecAlign: A Semantic Alignment Framework for SystemVerilog Assertion Generation

DGX agent

arXiv:2605.25181v1 Announce Type: new Abstract: Existing Large Language Model (LLM) approaches to SystemVerilog Assertion (SVA) generation primarily focus on syntactic validity and formal verification

safetyarxiv-cs-ai
26 May 2026
Safety

STaT: Resolving Shape Distortion in Non-Stationary Time Series via Tri-Modal Synergy

DGX agent

arXiv:2605.25943v1 Announce Type: new Abstract: Recent research in time series forecasting frequently investigates the integration of textual and visual modalities with numerical models to better navi

safetyarxiv-cs-lg
26 May 2026
Research

Step-TP: A Grounded, Step-Level Dataset with Chain-of-Thought Reasoning for LLM-Guided Tensor Program Optimization

DGX agent

arXiv:2605.25954v1 Announce Type: cross Abstract: Despite the strong reasoning capabilities of large language models (LLMs), optimizing the execution efficiency of tensor programs remains challenging

researcharxiv-cs-ai
26 May 2026
Safety

Strat-Reasoner: Reinforcing Strategic Reasoning of LLMs in Multi-Agent Games

DGX agent

arXiv:2605.04906v2 Announce Type: replace Abstract: While Large Language Models (LLMs) excel in certain reasoning tasks, they struggle in multi-agent games where the final outcome depends on the joint

safetyarxiv-cs-ai
26 May 2026
Applications

StrTransformer: Source-Wise Structured Transformers for Unsupervised Blind Source Recovery

DGX agent

arXiv:2605.25648v1 Announce Type: cross Abstract: This paper proposes StrTransformer, a source-wise structured Transformer framework for blind source recovery and branch-wise latent modeling. Instead

applicationsarxiv-cs-lg
26 May 2026
Research

Sum of Costs Diffusion with Dynamic Guidance for Motion Planning

DGX agent

arXiv:2605.24690v1 Announce Type: cross Abstract: The motion planning problem for robotic manipulation can be addressed through classical or deep learning approaches. Existing methods face significant

researcharxiv-cs-lg
26 May 2026
Research

The Multilingual Curse at the Retrieval Layer: Evidence from Amharic

DGX agent

arXiv:2605.24556v1 Announce Type: cross Abstract: Multilingual retrieval increasingly underpins cross-lingual question answering and retrieval-augmented generation. Strong zero-shot scores on multilin

researcharxiv-cs-cl
26 May 2026
Research

The Quantization Benefits of Residual-Free Transformers

DGX agent

arXiv:2605.25880v1 Announce Type: new Abstract: Large-scale transformer training and deployment are increasingly constrained by the transfer of activations, gradients, and optimizer states across acce

researcharxiv-cs-lg
26 May 2026
Tutorials

Today's Training Data episode takes us BTS on the infrastructure challenges required to do large RL runs at scale, featuring @ellev3n11 (Com…

DGX agent

Today's Training Data episode takes us BTS on the infrastructure challenges required to do large RL runs at scale, featuring @ellev3n11 (Composer Lead at @cursor_ai) and @dzhulgakov (Co-Founder at @Fi

tutorialssonya-huang--x
26 May 2026
Safety

TorchLean: Formalizing Neural Networks in Lean

DGX agent

arXiv:2602.22631v2 Announce Type: replace-cross Abstract: Neural networks are increasingly deployed in scientific, safety critical, and mission critical pipelines, yet verification and analysis are of

safetyarxiv-cs-lg
26 May 2026
Research

Towards Long-Horizon Interpretability: Efficient and Faithful Multi-Token Attribution for Reasoning LLMs

DGX agent

arXiv:2602.01914v2 Announce Type: replace Abstract: Token attribution methods provide intuitive explanations for language model outputs by identifying causally important input tokens. However, as mode

researcharxiv-cs-lg
26 May 2026
Safety

Towards trustworthy agentic AI: a comprehensive survey of safety, robustness, privacy, and system security

DGX agent

arXiv:2605.23989v1 Announce Type: new Abstract: Agentic AI systems -- Large Language Models (LLMs) augmented with planning, tool use, memory, and long-horizon interactions -- can execute complex tasks

safetyarxiv-cs-ai
26 May 2026
Safety

Trait-Aware Policy Optimization for Autoregressive Multi-Trait Essay Scoring

DGX agent

arXiv:2605.25731v1 Announce Type: new Abstract: Multi-trait essay scoring aims to provide fine-grained evaluation of writing quality across multiple dimensions. However, how to effectively post-train

safetyarxiv-cs-cl
26 May 2026
Research

Understanding, Accelerating, and Improving MeanFlow Training

DGX agent

arXiv:2511.19065v2 Announce Type: replace-cross Abstract: MeanFlow promises high-quality generative modeling in few steps, by jointly learning instantaneous and average velocity fields. Yet, the under

researcharxiv-cs-ai
26 May 2026
Research

Variable Clustering via Distributionally Robust Nodewise Regression

DGX agent

arXiv:2212.07944v3 Announce Type: replace Abstract: We study a multi-factor block model for variable clustering and connect it to regularized subspace clustering through a distributionally robust vers

researcharxiv-cs-lg
26 May 2026
Research

Voting with the Graph: Stable RLAIF via Topological Consistency Maximization

DGX agent

arXiv:2510.15514v3 Announce Type: replace Abstract: Reinforcement Learning from AI Feedback (RLAIF) relies on LLM judges as preference measurement instruments, yet these instruments are fundamentally

researcharxiv-cs-ai
26 May 2026
Applications

What Happens Next? Anticipating Future Motion by Generating Point Trajectories

DGX agent

arXiv:2509.21592v2 Announce Type: replace-cross Abstract: We consider the problem of forecasting motion from a single image, i.e., predicting how objects in the world are likely to move, without the a

applicationsarxiv-cs-ai
26 May 2026
Research

What Questions Should Robots Be Able to Answer? A Dataset of User Questions for Explainable Robotics

DGX agent

arXiv:2510.16435v2 Announce Type: replace-cross Abstract: With the growing use of large language models and conversational interfaces in human-robot interaction, robots' ability to answer user questio

researcharxiv-cs-cl
26 May 2026
Safety

When Does Multi-Agent RL Improve LLM Workflows? Workflow, Scale, and Policy-Sharing Tradeoffs

DGX agent

arXiv:2605.24202v1 Announce Type: new Abstract: Multi-agent LLM workflows route inference through specialized roles to lift end-task accuracy, but jointly training those roles with reinforcement learn

safetyarxiv-cs-ai
26 May 2026
Agents

You can’t build an autonomous agent on a static RAG pipeline. Because goals mutate dynamically, agents require an advanced knowledge infrast…

DGX agent

You can’t build an autonomous agent on a static RAG pipeline. Because goals mutate dynamically, agents require an advanced knowledge infrastructure, rather than forcing the model to hunt through a raw

agentspinecone--x
26 May 2026
Safety

4DThinker: Thinking with 4D Imagery for Dynamic Spatial Understanding

DGX agent

arXiv:2605.05997v2 Announce Type: replace Abstract: Dynamic spatial reasoning from monocular video is essential for bridging visual intelligence and the physical world, yet remains challenging for vis

safetyarxiv-cs-cv
25 May 2026
Research

A Simple Plug-in for Improving Eviction-Based KV Cache Compression

DGX agent

arXiv:2605.23258v1 Announce Type: new Abstract: KV cache growth is a major bottleneck for long-context inference in large language models. Existing methods are often dominated by binary eviction or re

researcharxiv-cs-lg
25 May 2026
Tutorials

Accelerating Divisible Load Processing Through Machine Learning: A Practical Framework for Large-Scale Workloads

DGX agent

arXiv:2605.23247v1 Announce Type: new Abstract: In this paper, we introduce the first machine learning framework for predicting optimal processing times in Single-Level Tree Network (SLTN) architectur

tutorialsarxiv-cs-lg
25 May 2026
Research

Amortized Simulation-Based Inference in Generalized Bayes via Neural Posterior Estimation

DGX agent

arXiv:2601.22367v2 Announce Type: replace-cross Abstract: Generalized Bayesian Inference (GBI) tempers a loss with a temperature eta > 0 to mitigate overconfidence and improve robustness under model m

researcharxiv-cs-lg
25 May 2026
← Previous
1…980981982983984…1302
Next →