AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
5,189 results
Safety

LiteOdyssey: A Lightweight Reasoning AI Agent for Interpretable Rare-Disease Diagnosis

DGX agent

arXiv:2606.16149v2 Announce Type: replace Abstract: Rare disease diagnosis involves interpreting clinical and genetic findings through complex diagnostic reasoning. We investigated whether this reason

safetyarxiv-cs-ai
10 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

RetailBench: Evaluating Long-Horizon Autonomous Decision-Making and Strategy Stability of LLM Agents in Realistic Retail Environments

DGX agent

arXiv:2603.16453v3 Announce Type: replace Abstract: Large language model (LLM) agents have made rapid progress on short-horizon, well-scoped tasks, yet their ability to sustain coherent decisions in d

model-releasesarxiv-cs-ai
10 Jul 2026
Local Ai

SPL: Orchestrating Workflows with Declarative Deterministic-Probabilistic Composition

DGX agent

arXiv:2607.07727v1 Announce Type: cross Abstract: We present SPL (Structured Prompt Language), a declarative language that composes deterministic and probabilistic computation modes in a single specif

local-aiarxiv-cs-cl
10 Jul 2026
Tutorials

Switch-Reasoner: Learn When to Think in Multitask Mixtures via Reinforcement Learning

DGX agent

arXiv:2607.08572v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) often follow a fixed Think-then-Answer paradigm, which is inefficient in heterogeneous multitask settings becau

tutorialsarxiv-cs-cv
10 Jul 2026
Safety

TNODEV: Toolbox for Neural ODE Verification

DGX agent

arXiv:2606.16567v2 Announce Type: replace Abstract: Neural ordinary differential equations (neural ODE) gained attention in safety critical settings such as continuous-time controllers for cyber-physi

safetyarxiv-cs-ai
10 Jul 2026
Model Releases

TOPO-Bench: An Open-Source Topological Mapping Evaluation Framework with Quantifiable Perceptual Aliasing

DGX agent

arXiv:2510.04100v2 Announce Type: replace-cross Abstract: Topological mapping offers a compact and robust representation for navigation, but progress in the field is hindered by the lack of standardiz

model-releasesarxiv-cs-ai
10 Jul 2026
Local Ai

TRACE: A Two-Channel Robust Attribution Watermark via Complementary Embeddings for LLM-Agent Trajectories

DGX agent

arXiv:2607.08400v1 Announce Type: cross Abstract: LLM agents reach users through resellers, who may rebrand a developer's agent or substitute a cheaper model. When provenance is disputed, attribution

local-aiarxiv-cs-ai
10 Jul 2026
Model Releases

UniClawBench: A Universal Benchmark for Proactive Agents on Real-World Tasks

DGX agent

arXiv:2607.08768v1 Announce Type: new Abstract: The rapid development of large language models and multimodal large language models has accelerated the emergence of proactive agents capable of operati

model-releasesarxiv-cs-cl
10 Jul 2026
Local Ai

WebSwarm: Recursive Multi-Agent Orchestration for Deep-and-Wide Web Search

DGX agent

arXiv:2607.08662v1 Announce Type: cross Abstract: Large language model (LLM)-based web search agents are transforming information seeking from simple factoid question answering into complex, deep-and-

local-aiarxiv-cs-ai
10 Jul 2026
Research

What LLM Forecasters Know but Don't Say: Probing Internal Representations for Calibration and Faithfulness

DGX agent

arXiv:2607.08046v1 Announce Type: cross Abstract: Large language models fine-tuned for forecasting can be accurate yet poorly calibrated, and their chain-of-thought (CoT) reasoning may not faithfully

researcharxiv-cs-ai
10 Jul 2026
Agents

A Closed-Loop Multi-Agent Framework for Robust Multi-Robot Manipulation

DGX agent

arXiv:2607.06990v1 Announce Type: new Abstract: Multi-robot systems provide the parallelism and redundancy necessary for long-horizon tasks, while Large Language Models (LLMs) offer the reasoning capa

agentsarxiv-cs-ro
9 Jul 2026
Agents

Agent-Exploitation Affordances: From Basic to Complex Representation Patterns

DGX agent

arXiv:2607.07475v1 Announce Type: new Abstract: In robotics, the capability of an artificial agent to represent the range of its action possibilities, i.e. affordances, is crucial to understand how it

agentsarxiv-cs-ro
9 Jul 2026
Model Releases

AgentLens: Production-Assessed Trajectory Reviews for Coding Agent Evaluation

DGX agent

arXiv:2607.06624v1 Announce Type: new Abstract: We present AgentLens, a production-assessed benchmark for interactive code agents. Most code-agent benchmarks reduce a run to a single bit -- did the ta

model-releasesarxiv-cs-ai
9 Jul 2026
Model Releases

AI Chatbot Suicide Risk Detection and Response: Human Validation Study of the Open-Source VERA-MH Safety Evaluation

DGX agent

arXiv:2602.05088v4 Announce Type: replace Abstract: Millions of people now use generative AI chatbots for psychological support. Despite their promise, the most pressing question in AI for mental heal

model-releasesarxiv-cs-ai
9 Jul 2026
Research

Compass: Prostate Cancer Detection Needs Multi-View Context

DGX agent

arXiv:2607.06919v1 Announce Type: new Abstract: Artificial intelligence (AI) analysis of micro-ultrasound (muUS) has shown promise for prostate cancer (PCa) detection. However, most existing AI method

researcharxiv-cs-cv
9 Jul 2026
Agents

ContestTrade: A Multi-Agent Trading System Based on Internal Contest Mechanism

DGX agent

arXiv:2508.00554v4 Announce Type: replace-cross Abstract: In financial trading, large language model (LLM)-based agents demonstrate significant potential, but their decisions can be sensitive to noisy

agentsarxiv-cs-cl
9 Jul 2026
Model Releases

Cost-Effective Agent Harnesses for Abstract Reasoning and Generalization on ARC-AGI-1

DGX agent

arXiv:2607.06764v1 Announce Type: new Abstract: Recent progress on ARC-AGI-1 from disclosed architectures has come broadly from two regimes: heavy test-time compute over frontier models (evolutionary

model-releasesarxiv-cs-ai
9 Jul 2026
Applications

Deployment Risk Assessment Using Diff-Aware Features: A Case Study at Prime Video

DGX agent

arXiv:2607.06766v1 Announce Type: cross Abstract: At Amazon Prime Video, we face the critical operational challenge of managing code deployments during live events and rapid feature releases without c

applicationsarxiv-cs-lg
9 Jul 2026
Agents

End-to-End LLM Flight Planning with RAG-based Memory and Multi-modal Coach Agent

DGX agent

arXiv:2607.06964v1 Announce Type: cross Abstract: Bridging the gap between human pilot intent and autonomous flight operation is critical for real-world electric vertical takeoff and landing (eVTOL) a

agentsarxiv-cs-ai
9 Jul 2026
Model Releases

Evaluating SageMath-Augmented LLM Agents for Computational and Experimental Mathematics

DGX agent

arXiv:2607.06820v1 Announce Type: new Abstract: Recent advances in AI for Mathematics have focused largely on autoformalization and theorem proving, leaving the role of Computer Algebra Systems (CAS)

model-releasesarxiv-cs-ai
9 Jul 2026
Research

Fast determinantal sampling on general spaces and diffusion geometry

DGX agent

arXiv:2607.06644v1 Announce Type: cross Abstract: Determinantal point processes have recently emerged as a kernel-based alternative to standard independent sampling for constructing efficient minibatc

researcharxiv-cs-lg
9 Jul 2026
Model Releases

Fast segmentation of watermarked texts from large language models through an epidemic change-point framework

DGX agent

arXiv:2509.21160v2 Announce Type: replace-cross Abstract: With the growing use of large language models, concerns over content authenticity have spurred a variety of watermarking schemes. These scheme

model-releasesarxiv-cs-lg
9 Jul 2026
Research

Fingerprint, Not Blueprint: How Positional Schemes Set the Default Spectral Algebra of Attention

DGX agent

arXiv:2607.06621v1 Announce Type: new Abstract: The pre-softmax score of an attention head is a bilinear form score(i,j) = x_i^T M x_j in a learned operator M = W_q^T W_k. Because M is generally non-s

researcharxiv-cs-lg
9 Jul 2026
Applications

From Jumps to Signatures: a Generative Method for Temporal Point Processes

DGX agent

arXiv:2607.06652v1 Announce Type: new Abstract: Rough path signatures are a universal feature map for continuous paths and, via the expected signature, characterise path distributions. These guarantee

applicationsarxiv-cs-lg
9 Jul 2026
Agents

Future Confidence Distillation in Large Language Models

DGX agent

arXiv:2607.07626v1 Announce Type: cross Abstract: Reliable confidence estimation is essential for deploying large language models (LLMs) in confidence-aware systems, where downstream decisions such as

agentsarxiv-cs-ai
9 Jul 2026
Model Releases

GrandTour: A Legged Robotics Dataset in the Wild for Multi-Modal Perception and State Estimation

DGX agent

arXiv:2602.18164v3 Announce Type: replace Abstract: Accurate state estimation and multi-modal perception are prerequisites for autonomous legged robots in complex, large-scale environments. To date, n

model-releasesarxiv-cs-ro
9 Jul 2026
Safety

Mechanistic Interpretability for Neural Networks: Circuits, Sparse Features and Symbolic Reasoning

DGX agent

arXiv:2607.07316v1 Announce Type: new Abstract: This article offers a comprehensive overview of mechanistic interpretability, an emerging field that seeks to reverse-engineer the internal algorithms o

safetyarxiv-cs-lg
9 Jul 2026
Model Releases

Reliable mechanistic operator recovery with biologically-informed neural networks: principles for architecture and optimisation design

DGX agent

arXiv:2607.07425v1 Announce Type: cross Abstract: Many biological processes are governed by complex dynamical mechanisms that remain incompletely understood despite increasing volumes of experimental

model-releasesarxiv-cs-lg
9 Jul 2026
Model Releases

SmartHomeSecure: Automated Detection and Repair of Smart Home Configuration Errors Using Large Language Models

DGX agent

arXiv:2607.06748v1 Announce Type: cross Abstract: Smart home automation platforms increasingly rely on user-authored YAML configuration files to define device behaviors, but these files are prone to s

model-releasesarxiv-cs-ai
9 Jul 2026
Model Releases

What Predicts Correctness in Text-to-SQL? A Selective-Prediction Study

DGX agent

arXiv:2607.06799v1 Announce Type: cross Abstract: Evaluating uncertainty in AI-generated SQL queries requires estimating whether a query is correct, where correct means it executes to the same result

model-releasesarxiv-cs-ai
9 Jul 2026
Applications

AlayaWorld: Long-Horizon and Playable Video World Generation

DGX agent

arXiv:2607.06291v1 Announce Type: new Abstract: Game worlds have traditionally been built through labor-intensive production pipelines, making them costly to develop, difficult to customization, and e

applicationsarxiv-cs-cv
8 Jul 2026
Model Releases

ArtisanCAD: An Industrial-Level CAD Agent with Expert-Grounded Knowledge Distillation

DGX agent

arXiv:2607.05750v1 Announce Type: new Abstract: Computer-aided design (CAD) for industrial components requires long-horizon procedural modeling, robust feature dependencies, editable parametric geomet

model-releasesarxiv-cs-ai
8 Jul 2026
Agents

Assessing the Operational Impact of Poisoning Attacks over Augmented 3D Point Cloud Public Datasets for Connected and Autonomous Vehicles

DGX agent

arXiv:2607.06484v1 Announce Type: cross Abstract: Poisoning attacks against public datasets lead to major concerns, such as (i) misclassification of perceived objects when the poisoned data is used fo

agentsarxiv-cs-cv
8 Jul 2026
Model Releases

Auditing of Unlearning Algorithms

DGX agent

arXiv:2607.05898v1 Announce Type: new Abstract: Evaluating whether unlearning algorithms truly remove training data influence remains an open challenge. We propose a practical auditor that computes da

model-releasesarxiv-cs-lg
8 Jul 2026
Agents

Beyond Static Evaluation: Building Simulation Environments for Scalable Agentic Reinforcement Learning

DGX agent

arXiv:2607.05773v1 Announce Type: new Abstract: As Large Language Models (LLMs) evolve into autonomous agents, traditional static evaluation fails to capture multi-step decision-making. We introduce A

agentsarxiv-cs-ai
8 Jul 2026
Safety

Bridging Physical Reasoning and Task Generalization via Visual Action Outcome Reasoning Alignment

DGX agent

arXiv:2607.06522v1 Announce Type: new Abstract: Vision-language models (VLMs) struggle to generalize in interactive physical reasoning, particularly under unseen tasks and environments. Two key failur

safetyarxiv-cs-ai
8 Jul 2026
Model Releases

Evaluating calibrated refusal and safe usefulness in dual-use biology settings

DGX agent

arXiv:2607.05462v1 Announce Type: cross Abstract: As AI agents are incorporated into life science workflows, the capabilities that speed discovery might also enable misuse. We present BioSecBench-Refu

model-releasesarxiv-cs-ai
8 Jul 2026
Research

From Textural Counterpoint to Feature Encoding: A Multi-Dimensional Machine Representation Study of Haydn's 'The Lark' Integrating Electroacoustic Analysis

DGX agent

arXiv:2607.05902v1 Announce Type: cross Abstract: Chamber music, as a highly precise multi-part interactive system, contains a logic of 'role assignment and dynamic interaction' that provides an extre

researcharxiv-cs-ai
8 Jul 2026
Safety

GraspIT: A Dataset Bridging the Sim-to-Real gap and back for Validated Grasping SE(3) Pose Generation

DGX agent

arXiv:2607.05869v1 Announce Type: cross Abstract: Robust robotic grasping of novel objects requires datasets that simultaneously provide photorealistic RGB-D observations, physically validated grasp q

safetyarxiv-cs-cv
8 Jul 2026
Model Releases

Is Your NPU Ready for LLMs? Dissecting the Hidden Efficiency Bottlenecks in Mobile LLM Inference

DGX agent

arXiv:2607.05475v1 Announce Type: cross Abstract: Deploying Large Language Models (LLMs) on mobile devices enhances privacy and reduces latency, but is severely bottlenecked by hardware inefficiency.

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

LLM Agents for Deliberative Collaboration: A Study on Joint Decision Making Under Partial Observability

DGX agent

arXiv:2607.06157v1 Announce Type: cross Abstract: Deliberation plays a crucial role in collaboration; when humans work together, they naturally engage in communication to align information and reach a

model-releasesarxiv-cs-ai
8 Jul 2026
Safety

MAME: Multidimensional Adaptive Metamer Exploration with Human Perceptual Feedback

DGX agent

arXiv:2503.13212v3 Announce Type: replace Abstract: Alignment between human brain networks and artificial models has become an active research area in vision science and machine learning. A widely ado

safetyarxiv-cs-lg
8 Jul 2026
Model Releases

Memory in the Loop: In-Process Retrieval as ExtendedWorking Memory for Language Agents

DGX agent

arXiv:2607.05690v1 Announce Type: new Abstract: Language agents run a loop - observe, reason, act - but the memory they reason over sits outside it: a store queried at most once per turn. We study the

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

PCBWorld: A Benchmark Environment for Engine-Grounded PCB Design Automation

DGX agent

arXiv:2607.05915v1 Announce Type: new Abstract: PCB routing is the task of connecting the nets of a board with copper traces under strict design rules, yet learning-based methods still lag behind rule

model-releasesarxiv-cs-ai
8 Jul 2026
Safety

Position: EU AI Act's Research Exemptions Can Break the Publication Norms of Major AI Conferences

DGX agent

arXiv:2506.03218v2 Announce Type: replace-cross Abstract: The EU has become one of the vanguards in regulating the digital age. A particularly important regulation in the Artificial Intelligence (AI)

safetyarxiv-cs-ai
8 Jul 2026
Safety

Property-Driven Synthetic Data Engineering for Data-Scarce Software Systems: Reflections from the Breast Cancer Domain

DGX agent

arXiv:2607.06133v1 Announce Type: cross Abstract: Modern software systems increasingly depend on data for analysis, prediction, testing, and decision-making. Yet many important domains, including medi

safetyarxiv-cs-ai
8 Jul 2026
Research

Regularity and Stability Properties of Selective SSMs with Discontinuous Gating

DGX agent

arXiv:2505.11602v3 Announce Type: replace Abstract: Selective State-Space Models (SSMs) such as Mamba have become central to long-sequence modeling. Still, their stability is poorly understood: their

researcharxiv-cs-lg
8 Jul 2026
Research

A Random Matrix Theory Perspective on the Consistency of Diffusion Models

DGX agent

arXiv:2602.02908v2 Announce Type: replace-cross Abstract: Diffusion models trained on different, non-overlapping subsets of a dataset often produce strikingly similar outputs when given the same noise

researcharxiv-cs-ai
7 Jul 2026
← Previous
1…6364656667…109
Next →