AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
5,189 results
Agents

Hierarchical Prompting with Dual LLM Modules for Robotic Task and Motion Planning

DGX agent

arXiv:2605.08330v1 Announce Type: new Abstract: We present a hierarchical language-driven framework for robotic task and motion planning to improve natural, intuitive human-robot interaction in servic

agentsarxiv-cs-ro
12 May 2026
Local Ai
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Kintsugi: Learning Policies by Repairing Executable Knowledge Bases

DGX agent

arXiv:2605.09487v1 Announce Type: new Abstract: Modern embodied agents achieve impressive performance, but their task knowledge is often stored in neural weights, latent state, or prompt-bound memory,

local-aiarxiv-cs-lg
12 May 2026
Safety

LAQuant: A Simple Overhead-free Large Reasoning Model Quantization by Layer-wise Lookahead Loss

DGX agent

arXiv:2605.08755v1 Announce Type: new Abstract: Large reasoning models (LRMs) reach competition-level math and coding accuracy via long autoregressive decoding, making per-token decoding cost a primar

safetyarxiv-cs-lg
12 May 2026
Research

Machine Learning Research Has Outpaced Its Communication Norms and NeurIPS Should Act

DGX agent

arXiv:2605.08889v1 Announce Type: cross Abstract: Machine learning research has grown exponentially while its communication norms have not. We argue NeurIPS should adopt explicit, measurable writing s

researcharxiv-cs-cl
12 May 2026
Safety

MARLaaS: Multi-Tenant Asynchronous Reinforcement Learning as a Service

DGX agent

arXiv:2605.08527v1 Announce Type: cross Abstract: Reinforcement Learning from Verifiable Rewards (RLVR) has significantly improved the reasoning capabilities of large language models (LLMs), particula

safetyarxiv-cs-ai
12 May 2026
Model Releases

MathlibLemma: Folklore Lemma Generation and Benchmark for Formal Mathematics

DGX agent

arXiv:2602.02561v2 Announce Type: replace-cross Abstract: While the ecosystem of Lean and Mathlib has enjoyed celebrated success in formal mathematical reasoning with the help of large language models

model-releasesarxiv-cs-ai
12 May 2026
Research

Measuring and Decomposing Mode Separation via the Canonical Diffusion

DGX agent

arXiv:2605.08777v1 Announce Type: cross Abstract: Mode separation, namely how sharply a distribution fragments into barrier-separated clusters, is a fundamental geometric property of densities, diffic

researcharxiv-cs-lg
12 May 2026
Model Releases

Nautilus Compass: Black-box Persona Drift Detection for Production LLM Agents

DGX agent

arXiv:2605.09863v1 Announce Type: cross Abstract: Production LLM coding agents drift over long sessions: they forget user-specified constraints, slip into mistakes the user already flagged, and confab

model-releasesarxiv-cs-ai
12 May 2026
Local Ai

NeuroGAN-3D: Enhancing Intrinsic Functional Brain Networks via High-Fidelity 3D Generative Super-Resolution

DGX agent

arXiv:2605.08373v1 Announce Type: cross Abstract: Recent advances in neuroimaging have deepened our understanding of the brain's complex functional and structural organization. Among these, functional

local-aiarxiv-cs-ai
12 May 2026
Local Ai

On Distinguishing Capability Elicitation from Capability Creation in Post-Training: A Free-Energy Perspective

DGX agent

arXiv:2605.08368v1 Announce Type: new Abstract: Debates about large language model post-training often treat supervised fine-tuning (SFT) as imitation and reinforcement learning (RL) as discovery. But

local-aiarxiv-cs-ai
12 May 2026
Local Ai

PAAC: Privacy-Aware Agentic Device-Cloud Collaboration

DGX agent

arXiv:2605.08646v1 Announce Type: cross Abstract: Large language model (LLM) agents face a structural tension: cloud agents provide strong reasoning but expose user data, while on-device agents preser

local-aiarxiv-cs-cl
12 May 2026
Model Releases

PaperFit: Vision-in-the-Loop Typesetting Optimization for Scientific Documents

DGX agent

arXiv:2605.10341v1 Announce Type: new Abstract: A LaTeX manuscript that compiles without error is not necessarily publication-ready. The resulting PDFs frequently suffer from misplaced floats, overflo

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Pix2Fact: When Vision Is Not Enough -- Benchmarking Fine-Grained VQA with Web Verification on High-Resolution Real-World Scenes

DGX agent

arXiv:2602.00593v2 Announce Type: replace Abstract: Despite progress on general tasks, vision-language models (VLMs) still struggle with challenges that demand both fine-grained visual grounding and e

model-releasesarxiv-cs-cv
12 May 2026
Research

Position: Life-Logging Video Streams Make the Privacy-Utility Trade-off Inevitable

DGX agent

arXiv:2605.10404v1 Announce Type: new Abstract: With the growing prevalence of always-on hardware such as smart glasses, body cameras, and home security systems, life-logging visual sensing is becomin

researcharxiv-cs-cv
12 May 2026
Agents

PRISM: Fast Online LLM Serving via Scheduling-Memory Co-design

DGX agent

arXiv:2605.08581v1 Announce Type: new Abstract: Modern online large language model (LLM) services, such as Retrieval-Augmented Generation (RAG) and agent systems, increasingly expose two prominent cha

agentsarxiv-cs-lg
12 May 2026
Model Releases

Privacy Auditing Synthetic Data Release through Local Likelihood Attacks

DGX agent

arXiv:2508.21146v2 Announce Type: replace Abstract: Auditing the privacy leakage of synthetic data is an important but unresolved problem. Existing privacy auditing frameworks for synthetic data rely

model-releasesarxiv-cs-lg
12 May 2026
Research

Revisiting Mixture Policies in Entropy-Regularized Actor-Critic

DGX agent

arXiv:2605.09157v1 Announce Type: cross Abstract: Mixture policies theoretically offer greater flexibility than unimodal policies in continuous action reinforcement learning, but the practical benefit

researcharxiv-cs-ai
12 May 2026
Safety

RubricEM: Meta-RL with Rubric-guided Policy Decomposition beyond Verifiable Rewards

DGX agent

arXiv:2605.10899v1 Announce Type: new Abstract: Training deep research agents, namely systems that plan, search, evaluate evidence, and synthesize long-form reports, pushes reinforcement learning beyo

safetyarxiv-cs-cl
12 May 2026
Safety

SalesSim: Benchmarking and Aligning Multimodal Language Models as Retail User Simulators

DGX agent

arXiv:2605.08334v1 Announce Type: new Abstract: We present SalesSim, a framework and testbed for evaluating the ability of Multimodal Large Language Models (MLLMs) to simulate realistic, persona-drive

safetyarxiv-cs-cl
12 May 2026
Agents

SAR-RAG: ATR Visual Question Answering by Semantic Search, Retrieval, and MLLM Generation

DGX agent

arXiv:2602.04712v2 Announce Type: replace-cross Abstract: We present a visual-context image-retrieval-augmented generation (ImageRAG)- assisted AI agent for automatic target recognition (ATR) of synth

agentsarxiv-cs-ai
12 May 2026
Agents

SimWorld Studio: Automatic Environment Generation with Evolving Coding Agent for Embodied Agent Learning

DGX agent

arXiv:2605.09423v1 Announce Type: new Abstract: LLM/VLM-based digital agents have advanced rapidly thanks to scalable sandboxes for coding, web navigation, and computer use, which provide rich interac

agentsarxiv-cs-ai
12 May 2026
Model Releases

SmartEval: A Benchmark for Evaluating LLM-Generated Smart Contracts from Natural Language Specifications

DGX agent

arXiv:2605.09610v1 Announce Type: cross Abstract: We introduce SmartEval, a benchmark for systematically evaluating the quality of Solidity smart contracts generated by large language models (LLMs) fr

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Spherical Boltzmann machines: a solvable theory of learning and generation in energy-based models

DGX agent

arXiv:2605.09031v1 Announce Type: new Abstract: Energy-based models (EBMs) are flexible generative architectures inspired by statistical physics, but their learning and generative properties remain po

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Steerable but Not Decodable: Function Vectors Operate Beyond the Logit Lens

DGX agent

arXiv:2604.02608v2 Announce Type: replace Abstract: Activation steering presupposes that task-relevant behaviors correspond to linear directions in activation space -- directions that should both stee

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Strategic Exploitation in LLM Agent Markets: A Simulation Framework for E-Commerce Trust

DGX agent

arXiv:2605.10059v1 Announce Type: new Abstract: Agent-based modeling (ABM) has long been used in economics to study human behavior, and large language model (LLM) agents now enable new forms of social

model-releasesarxiv-cs-ai
12 May 2026
Research

Survey on Disaster Management Datasets for Remote Sensing Based Emergency Applications

DGX agent

arXiv:2605.08196v1 Announce Type: new Abstract: Recent natural disasters have highlighted the urgent need for efficient data-driven approaches to disaster management. Machine learning (ML) and deep le

researcharxiv-cs-cv
12 May 2026
Agents

The Agent Use of Agent Beings: Agent Cybernetics Is the Missing Science of Foundation Agents

DGX agent

arXiv:2605.10754v1 Announce Type: new Abstract: LLM-based foundation agents that perceive, reason, and act across thousands of reasoning steps are rapidly becoming the dominant paradigm for deploying

agentsarxiv-cs-ai
12 May 2026
Research

The Association of Transformer-based Sentiment Analysis with Symptom Distress and Deterioration in Routine Psychotherapy Care

DGX agent

arXiv:2605.09838v1 Announce Type: new Abstract: Sentiment analysis has been of long-standing interest in psychotherapy research. Recently, the Transformer deep learning architecture has produced text-

researcharxiv-cs-cl
12 May 2026
Safety

The Pokemon Theorem and other Fairness Impossibility Results

DGX agent

arXiv:2605.09221v1 Announce Type: cross Abstract: Fairness impossibility results often look like distinct scalar incompatibility statements. We show that several share one RKHS geometry: fairness crit

safetyarxiv-cs-ai
12 May 2026
Applications

Topological Data Analysis Applications in Natural Language Processing: A Survey

DGX agent

arXiv:2411.10298v5 Announce Type: replace Abstract: The surge of data available on the Internet has driven the adoption of a wide range of computational methods for analyzing and extracting insights f

applicationsarxiv-cs-cl
12 May 2026
Research

Towards a Certificate of Trust: Task-Aware OOD Detection for Scientific AI

DGX agent

arXiv:2509.25080v3 Announce Type: replace Abstract: Data-driven models are increasingly adopted in critical scientific fields like weather forecasting and fluid dynamics. These methods can fail on out

researcharxiv-cs-lg
12 May 2026
Agents

Towards a Virtual Neuroscientist: Autonomous Neuroimaging Analysis via Multi-Agent Collaboration

DGX agent

arXiv:2605.09366v1 Announce Type: new Abstract: Transforming neuroimaging data into clinically actionable biomarkers is a knowledge-intensive and labor-intensive process. Standardized workflows such a

agentsarxiv-cs-ai
12 May 2026
Safety

Trajectory-Consistent Flow Matching for Robust Visuomotor Policy Learning

DGX agent

arXiv:2605.08511v1 Announce Type: new Abstract: Flow matching policies learn continuous velocity fields that transport noise to actions, enabling fast deterministic inference for robot manipulation. H

safetyarxiv-cs-ro
12 May 2026
Model Releases

TrajPrism: A Multi-Task Benchmark for Language-Grounded Urban Trajectory Understanding

DGX agent

arXiv:2605.10782v1 Announce Type: new Abstract: Urban mobility is naturally expressed both as trajectories in space and as natural-language descriptions of travel intent, constraints, and preferences.

model-releasesarxiv-cs-ai
12 May 2026
Applications

Transformer autoencoder with local attention for sparse and irregular time series with application on risk estimation

DGX agent

arXiv:2605.08914v1 Announce Type: cross Abstract: This paper introduces a framework specifically designed for sparse and irregular time series {risk estimation}. It is based on a Transformer Autoencod

applicationsarxiv-cs-ai
12 May 2026
Research

Transforming the Use of Earth Observation Data: Exascale Training of a Generative Compression Model with Historical Priors for up to 10,000x Data Reduction

DGX agent

arXiv:2605.08633v1 Announce Type: cross Abstract: Earth observation is becoming one of the largest data-producing activities in science, yet current pipelines still treat compression as a storage and

researcharxiv-cs-cv
12 May 2026
Safety

Trustworthy AI: Ensuring Reliability and Accountability from Models to Agents

DGX agent

arXiv:2605.08964v1 Announce Type: new Abstract: In this thesis, we develop algorithms with theoretical guarantees for ensuring reliability and accountability of Machine Learning (ML) systems. As ML sy

safetyarxiv-cs-lg
12 May 2026
Research

Universal Feature Selection with Noisy Observations and Weak Symmetry Conditions

DGX agent

arXiv:2605.09396v1 Announce Type: cross Abstract: This paper relaxes the restrictive symmetry conditions adopted in [4], [5] and extends their universal feature selection framework to accommodate nois

researcharxiv-cs-lg
12 May 2026
Safety

V-ABS: Action-Observer Driven Beam Search for Dynamic Visual Reasoning

DGX agent

arXiv:2605.10172v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) have achieved remarkable success in general perception, yet complex multi-step visual reasoning remains a per

safetyarxiv-cs-cl
12 May 2026
Model Releases

VT-Bench: A Unified Benchmark for Visual-Tabular Multi-Modal Learning

DGX agent

arXiv:2605.08146v1 Announce Type: cross Abstract: Multi-model learning has attracted great attention in visual-text tasks. However, visual-tabular data, which plays a pivotal role in high-stakes domai

model-releasesarxiv-cs-ai
12 May 2026
Agents

What Software Engineering Looks Like to AI Agents? -- An Empirical Study of AI-Only Technical Discourse on MoltBook

DGX agent

arXiv:2605.08380v1 Announce Type: cross Abstract: AI agents are increasingly framed as software-engineering teammates, yet most research studies them inside human-centered workflows. Little is known a

agentsarxiv-cs-ai
12 May 2026
Safety

When Agents Overtrust Environmental Evidence: An Extensible Agentic Framework for Benchmarking Evidence-Grounding Defects in LLM Agents

DGX agent

arXiv:2605.08828v1 Announce Type: new Abstract: Large language model agents increasingly operate through environment-facing scaffolds that expose files, web pages, APIs, and logs. These observations i

safetyarxiv-cs-ai
12 May 2026
Agents

When Child Inherits: Modeling and Exploiting Subagent Spawn in Multi-Agent Networks

DGX agent

arXiv:2605.08460v1 Announce Type: cross Abstract: Since the official release of ChatGPT in 2022, large language models (LLMs) have rapidly evolved from chatbot-style interfaces into agentic systems th

agentsarxiv-cs-ai
12 May 2026
Agents

123D: Unifying Multi-Modal Autonomous Driving Data at Scale

DGX agent

arXiv:2605.08084v1 Announce Type: cross Abstract: The pursuit of autonomous driving has produced one of the richest sensor data collections in all of robotics. However, its scale and diversity remain

agentsarxiv-cs-cv
11 May 2026
Model Releases

A Reproducible Optimisation Protocol for Calibrating Prompt-Based Large Language Model Workflows in Evidence Synthesis

DGX agent

arXiv:2605.06937v1 Announce Type: new Abstract: This methods article presents a reproducible calibration workflow for prompt-based large language models (LLMs) in structured evidence-synthesis tasks.

model-releasesarxiv-cs-lg
11 May 2026
Safety

A Systematic Investigation of The RL-Jailbreaker in LLMs

DGX agent

arXiv:2605.07032v1 Announce Type: cross Abstract: The evolution of generative models from next-token predictors to autonomous engines of complex systems necessitates rigorous safety hardening. Adversa

safetyarxiv-cs-ai
11 May 2026
Safety

Activation Differences Reveal Backdoors: A Comparison of SAE Architectures

DGX agent

arXiv:2605.07324v1 Announce Type: cross Abstract: Backdoor attacks on language models pose a significant threat to AI safety, where models behave normally on most inputs but exhibit harmful behavior w

safetyarxiv-cs-ai
11 May 2026
Research

AffineLens: Capturing the Continuous Piecewise Affine Functions of Neural Networks

DGX agent

arXiv:2605.06218v2 Announce Type: replace Abstract: Piecewise affine neural networks (PANNs) provide a principled geometric perspective on neural network expressivity by characterizing the input--outp

researcharxiv-cs-lg
11 May 2026
← Previous
1…9192939495…109
Next →