AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
5,189 results
Safety

Libra: Efficient Resource Management for Agentic RL Post-Training

DGX agent

arXiv:2606.03077v1 Announce Type: cross Abstract: Reinforcement learning (RL) has become a standard post-training paradigm for large language models (LLMs), extending beyond preference alignment to co

safetyarxiv-cs-ai
3 Jun 2026
Model Releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

SCOPE: Real-Time Natural Language Camera Agent at the Edge

DGX agent

arXiv:2606.02951v1 Announce Type: cross Abstract: Deploying language-driven agents in robotics requires evaluations that reflect real-world task demands: natural-language instructions with reproducibl

model-releasesarxiv-cs-ai
3 Jun 2026
Agents

Uncertainty-Aware Clarification in LLM Agents with Information Gain

DGX agent

arXiv:2606.03135v1 Announce Type: new Abstract: Large Language Model (LLM) agents often operate under underspecified user instructions, where latent uncertainty over user intent leads to erroneous too

agentsarxiv-cs-ai
3 Jun 2026
Safety

Addressing Longstanding Challenges in Cognitive Science with Language Models

DGX agent

arXiv:2511.00206v3 Announce Type: replace Abstract: Cognitive science faces ongoing challenges in research integration, formalization, conceptual clarity, and other areas, in part due to its multiface

safetyarxiv-cs-ai
2 Jun 2026
Safety

Agent Operating Systems (AOS): Integrating Agentic Control Planes into, and Beyond, Traditional Operating Systems

DGX agent

arXiv:2606.01508v1 Announce Type: cross Abstract: Traditional operating systems were designed around deterministic programs, explicit control flow, and human initiated workflows. Their core abstractio

safetyarxiv-cs-ai
2 Jun 2026
Agents

Agent-R1: A Unified and Modular Framework for Agentic Reinforcement Learning

DGX agent

arXiv:2511.14460v2 Announce Type: replace Abstract: Large language models (LLMs) have rapidly evolved from single-turn text generators into the foundation of increasingly capable agents. As these agen

agentsarxiv-cs-cl
2 Jun 2026
Agents

Agentic Authoring of Interactive Multiview Visualizations in Genomics

DGX agent

arXiv:2606.00370v1 Announce Type: cross Abstract: Diverse genomics data, scientific questions, and analysis tasks typically demand highly specialized visualizations. Therefore, users often must custom

agentsarxiv-cs-ai
2 Jun 2026
Agents

Automated Conjecture Resolution with Formal Verification

DGX agent

arXiv:2604.03789v2 Announce Type: replace-cross Abstract: Recent advances in large language models have significantly improved their ability to perform mathematical reasoning, extending from elementar

agentsarxiv-cs-ai
2 Jun 2026
Research

DeepLatent: Think with Images via Parallel Latent Visual Reasoning

DGX agent

arXiv:2606.00562v1 Announce Type: new Abstract: The emerging paradigm of 'thinking with images' embeds visual states into intermediate reasoning steps, defining a new frontier for Vision-Language Mode

researcharxiv-cs-cv
2 Jun 2026
Safety

IstGPT: LLM-based Anomaly Detection for Spatial-Temporal Graph in Industrial Systems

DGX agent

arXiv:2606.01691v1 Announce Type: cross Abstract: Industrial Internet systems face increasing threats from sophisticated industrial control system (ICS) attacks, resulting in critical safety incidents

safetyarxiv-cs-lg
2 Jun 2026
Agents

Learning to Construct Practical Agentic Systems

DGX agent

arXiv:2606.00189v1 Announce Type: cross Abstract: Automated design and optimization of agentic LLM-based systems leads to sophisticated systems that substantially improve result quality over off-the-s

agentsarxiv-cs-ai
2 Jun 2026
Agents

OctoT2I: A Self-Evolving Agentic Text-to-Image Router

DGX agent

arXiv:2606.01803v1 Announce Type: new Abstract: The explosive growth of Text-to-Image (T2I) models, from large-scale versions to lightweight, real-time ones, now faces diminishing marginal returns fro

agentsarxiv-cs-ai
2 Jun 2026
Applications

Predicting the risk of colorectal anastomotic leak based on preoperative mapping of the blood supply of the bowel

DGX agent

arXiv:2606.02156v1 Announce Type: cross Abstract: Anastomotic leak remains one of the most serious complications following colorectal cancer surgery, substantially affecting patient outcomes, recovery

applicationsarxiv-cs-ai
2 Jun 2026
Model Releases

ProtStructQA: A Denotation Threshold in Protein Structural Reasoning

DGX agent

arXiv:2606.00451v1 Announce Type: new Abstract: Protein-language systems are often evaluated by whether they generate plausible biological text, but a structural question has a sharper semantics: it d

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Quality Audio Prototyping: a prototype system for unified sound retrieval and procedural generation

DGX agent

arXiv:2606.00629v1 Announce Type: cross Abstract: Sound design workflows frequently oscillate between time-consuming library searches and the complexity of procedural synthesis, with practitioners typ

model-releasesarxiv-cs-lg
2 Jun 2026
Safety

SafeMCP: Proactive Power Regulation for LLM Agent Defense via Environment-Grounded Look-Ahead Reasoning

DGX agent

arXiv:2606.01991v1 Announce Type: new Abstract: As Large Language Model (LLM) agents increasingly leverage the Model Context Protocol (MCP) to operate in complex environments, the expansion of their a

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

Sandboxed Coding Agents are Competitive Omni-modal Task Solvers

DGX agent

arXiv:2606.00579v1 Announce Type: new Abstract: As multimodal LLMs increasingly target video and audio, it is often assumed that such tasks require native omnimodal models. We show that this is not al

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

SPADE-Bench: Evaluating Spontaneous Strategic Deception in Agents via Plan-Action Divergence

DGX agent

arXiv:2606.02380v1 Announce Type: cross Abstract: As LLM-based agents expand their operational scope, reliability becomes a prerequisite for real-world deployment. However, in practical applications,

model-releasesarxiv-cs-ai
2 Jun 2026
Research

Step-Level Sparse Autoencoder for Reasoning Process Interpretation

DGX agent

arXiv:2603.03031v2 Announce Type: replace Abstract: Large Language Models (LLMs) have achieved strong complex reasoning capabilities through Chain-of-Thought (CoT) reasoning. However, their reasoning

researcharxiv-cs-lg
2 Jun 2026
Model Releases

When Safe Skills Collide: Measuring Compositional Risk in Agent Skill Ecosystems

DGX agent

arXiv:2606.00448v1 Announce Type: cross Abstract: LLM agents increasingly rely on community-contributed skills that expand an agent's operational capability set. We study a core safety problem in agen

model-releasesarxiv-cs-ai
2 Jun 2026
Safety

Counterfactual Evaluation Reveals Hidden Capability Profiles in Clinical LLMs and Agents

DGX agent

arXiv:2605.30590v1 Announce Type: cross Abstract: Two clinical AI systems can score nearly identically on coverage-based rubrics yet behave radically differently when their patient inputs change: one

safetyarxiv-cs-ai
1 Jun 2026
Research

EBuddy: a workflow orchestrator for industrial human-machine collaboration

DGX agent

arXiv:2603.28579v2 Announce Type: replace Abstract: This paper presents EBuddy, a voice-guided workflow orchestrator for natural human-machine collaboration in industrial environments. EBuddy targets

researcharxiv-cs-ro
1 Jun 2026
Model Releases

LegSegNet: A Public Deep Learning System for Lower Extremity CT Tissue Segmentation and Quantification

DGX agent

arXiv:2605.30829v1 Announce Type: new Abstract: Lower extremity computed tomography (CT) contains clinically relevant information for body composition analysis, sarcopenia assessment, and musculoskele

model-releasesarxiv-cs-cv
1 Jun 2026
Model Releases

Mellum2 Technical Report

DGX agent

arXiv:2605.31268v1 Announce Type: new Abstract: We present Mellum 2, an open-weight 12B-parameter Mixture-of-Experts (MoE) language model with 2.5B active parameters per token. Mellum 2 is a general-p

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

Brain-IT-VQA: From Brain Signals to Answers

DGX agent

arXiv:2605.29588v1 Announce Type: cross Abstract: Decoding visual content from fMRI signals recorded while a person views images, and specifically answering questions about the seen images, is a long-

model-releasesarxiv-cs-ai
29 May 2026
Research

Eigen-Spike Emergence and Quadratic Equivalents for Conjugate Kernels on Nonlinearly Separable Data

DGX agent

arXiv:2605.29669v1 Announce Type: cross Abstract: Recent work in random matrix theory (RMT) has developed the notion of deterministic equivalents: typically linear surrogate models that approximate th

researcharxiv-cs-lg
29 May 2026
Model Releases

GroundAct: Can LLM Agents Ground Actions in Environmental States?

DGX agent

arXiv:2508.05614v2 Announce Type: replace-cross Abstract: LLM agents achieve 85-96% success on tasks where instructions fully specify the action, but drop to 29-53% when action feasibility depends on

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

TaxDistill: Improving Metagenomic Taxonomic Annotation via Distilled Genomic Foundation Models

DGX agent

arXiv:2605.28868v1 Announce Type: cross Abstract: Metagenomic taxonomic annotation aims to identify the microbial origins of DNA fragments in environmental samples. Traditional methods that rely on se

model-releasesarxiv-cs-ai
29 May 2026
Local Ai

Towards Understanding the Shape of Representations in Protein Language Models

DGX agent

arXiv:2509.24895v2 Announce Type: replace Abstract: While protein language models (PLMs) are one of the most promising avenues of research for future de novo protein design, the way in which they tran

local-aiarxiv-cs-lg
29 May 2026
Local Ai

Decoupling Skeleton and Flesh: Efficient Multimodal Table Reasoning with Disentangled Alignment and Structure-aware Guidance

DGX agent

arXiv:2602.03491v2 Announce Type: replace-cross Abstract: Reasoning over table images remains challenging for Large Vision-Language Models (LVLMs) due to complex layouts and tightly coupled structure-

local-aiarxiv-cs-cl
28 May 2026
Agents

DeltaMCP: Incremental Regeneration via Spec-Aware Transformation for MCP servers

DGX agent

arXiv:2605.28148v1 Announce Type: cross Abstract: The rapid development of LLMs coupled with the introduction of Model Context Protocol (MCP) has revolutionized how intelligent agents interact with AP

agentsarxiv-cs-ai
28 May 2026
Safety

Diagnosing Live Within-Policy Instruction Conflicts in LLM Agents with Witnessed Resolution Profiles

DGX agent

arXiv:2605.27784v1 Announce Type: new Abstract: LLM agents are governed by long-lived natural-language prompt policies, but individually reasonable standing rules can interact in uninspected ways. We

safetyarxiv-cs-ai
28 May 2026
Safety

LACUNA: Safe Agents as Recursive Program Holes

DGX agent

arXiv:2605.28617v1 Announce Type: new Abstract: LLM agents increasingly act by writing code, yet a split persists between the runtime that drives the agent and the code the model writes. The runtime o

safetyarxiv-cs-ai
28 May 2026
Model Releases

LiveBrowseComp: Are Search Agents Searching, or Just Verifying What They Already Know?

DGX agent

arXiv:2605.28721v1 Announce Type: new Abstract: Are LLM-based search agents genuinely searching, or using the web to verify what they already know? We study this question on BrowseComp with three diag

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

MolLingo: Molecule-Native Representations for LLM-Powered Scientific Agents

DGX agent

arXiv:2605.27853v1 Announce Type: new Abstract: We present MolLingo, a multi-agent system that emulates the reasoning process of a chemist to automate molecular design. Existing LLM-based approaches e

model-releasesarxiv-cs-ai
28 May 2026
Research

RCM Constraint-Consistent Dynamic Control in Surgical Robots

DGX agent

arXiv:2509.14075v2 Announce Type: replace Abstract: Robotic-assisted minimally invasive surgery (RAMIS) requires accurate enforcement of the remote center of motion (RCM) constraint to ensure safe too

researcharxiv-cs-ro
28 May 2026
Model Releases

Thermodynamic properties of chemically disordered compounds via AI-driven estimation of partition function with the PULSE method

DGX agent

arXiv:2605.28594v1 Announce Type: cross Abstract: In this article, we present an improved version of the PULSE method (Partition function Unsupervised Learning Sampling and Evaluation) for estimating

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Advancing Creative Physical Intelligence in Large Multimodal Models

DGX agent

arXiv:2605.26396v1 Announce Type: new Abstract: Large multimodal models (LMMs) have rapidly advanced in perception and reasoning; however, it remains unclear whether these capabilities generalize to d

model-releasesarxiv-cs-ai
27 May 2026
Safety

AI-Driven Contribution Evaluation and Conflict Resolution: A Framework & Design for Group Workload Investigation

DGX agent

arXiv:2511.07667v2 Announce Type: replace Abstract: The equitable assessment of individual contribution in teams remains a persistent challenge, where conflict and disparity in workload can result in

safetyarxiv-cs-ai
27 May 2026
Research

Guiding LLM Post-training Data Engineering with Model Internals from Sparse Autoencoders

DGX agent

arXiv:2605.27354v1 Announce Type: cross Abstract: Model internals encode rich information about how a large language model (LLM) processes its training data; however, post-training data engineering la

researcharxiv-cs-ai
27 May 2026
Local Ai

Persistent AI Agents in Academic Research: A Single-Investigator Implementation Case Study

DGX agent

arXiv:2605.26870v1 Announce Type: cross Abstract: Background: Large language models are typically evaluated as models, benchmarks, or short conversational episodes. Less is known about what happens wh

local-aiarxiv-cs-ai
27 May 2026
Model Releases

Position: AI Safety Requires Effective Controllability

DGX agent

arXiv:2605.27117v1 Announce Type: new Abstract: AI safety is still largely framed as alignment: training models to follow human preferences, safety policies, and normative constraints. That framing ha

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Shopping Companion: A Memory-Augmented LLM Agent for Real-World E-Commerce Tasks

DGX agent

arXiv:2603.14864v2 Announce Type: replace Abstract: In e-commerce, LLM agents show promise for shopping tasks such as recommendations, budget management, and bundle deals, where accurately capturing u

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

SWE-Adept: An LLM-Based Agentic Framework for Deep Codebase Analysis and Structured Issue Resolution

DGX agent

arXiv:2603.01327v2 Announce Type: replace-cross Abstract: Large language models (LLMs) exhibit strong performance on self-contained programming tasks. However, they still struggle with repository-leve

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

Towards Error-Free EHRs: Reasoning-Intensive Consistency Verification Between Clinical Notes and Structured Tables in Electronic Health Records

DGX agent

arXiv:2605.26463v1 Announce Type: cross Abstract: Data consistency between unstructured clinical notes and structured tables in Electronic Health Records (EHRs) is essential for patient safety and cli

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

AgentGrounder: Zero-Shot 3D Visual Pointcloud Grounding using Multimodal Language Models

DGX agent

arXiv:2605.25901v1 Announce Type: cross Abstract: 3D Visual Grounding (3DVG) is an essential capability for embodied AI, requiring agents to localize objects in 3D scenes based on natural language des

model-releasesarxiv-cs-ro
26 May 2026
Research

Deep Learning-Enabled Prediction of Geoeffective CMEs Using SOHO and SDO Observations

DGX agent

arXiv:2605.24748v1 Announce Type: cross Abstract: Understanding and forecasting the geoeffectiveness of a coronal mass ejection (CME) is crucial for protecting infrastructure in the near-Earth space e

researcharxiv-cs-lg
26 May 2026
Model Releases

From Model Scaling to System Scaling: Scaling the Harness in Agentic AI

DGX agent

arXiv:2605.26112v1 Announce Type: new Abstract: This paper studies the next major bottleneck in agentic AI as system scaling, not only model scaling: the design of auditable, persistent, modular, and

model-releasesarxiv-cs-ai
26 May 2026
← Previous
1…4647484950…109
Next →