AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
5,189 results
Model Releases

Coordination as an Architectural Layer for LLM-Based Multi-Agent Systems

DGX agent

arXiv:2605.03310v1 Announce Type: cross Abstract: Multi-agent LLM systems fail in production at rates between 41% and 87%, mostly due to coordination defects rather than base-model capability. Existin

model-releasesarxiv-cs-lg
6 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

EO-Gym: A Multimodal, Interactive Environment for Earth Observation Agents

DGX agent

arXiv:2605.01250v1 Announce Type: new Abstract: Earth Observation (EO) analysis is inherently interactive: resolving uncertainty often requires expanding the region of interest, retrieving historical

model-releasesarxiv-cs-ai
6 May 2026
Research

psifx -- Psychological and Social Interactions Feature Extraction Package

DGX agent

arXiv:2407.10266v5 Announce Type: replace Abstract: psifx is a plug-and-play multi-modal feature extraction toolkit, aiming to facilitate and democratize the use of state-of-the-art machine learning t

researcharxiv-cs-cl
6 May 2026
Agents

Say the Mission, Execute the Swarm: Agent-Enhanced LLM Reasoning in the Web-of-Drones

DGX agent

arXiv:2605.03788v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly explored as high-level reasoning engines for cyber-physical systems, yet their application to real-time

agentsarxiv-cs-ro
6 May 2026
Model Releases

Trojan Hippo: Weaponizing Agent Memory for Data Exfiltration

DGX agent

arXiv:2605.01970v2 Announce Type: cross Abstract: Memory systems enable otherwise-stateless LLM agents to persist user information across sessions, but also introduce a new attack surface. We characte

model-releasesarxiv-cs-ai
6 May 2026
Model Releases

InfantAgent-Next: A Multimodal Generalist Agent for Automated Computer Interaction

DGX agent

arXiv:2505.10887v3 Announce Type: replace Abstract: This paper introduces extsc{InfantAgent-Next}, a generalist agent capable of interacting with computers in a multimodal manner, encompassing text, i

model-releasesarxiv-cs-ai
5 May 2026
Safety

SynPAIN: A Synthetic Dataset of Pain and Non-Pain Facial Expressions

DGX agent

arXiv:2507.19673v3 Announce Type: replace Abstract: Accurate pain assessment in patients with limited ability to communicate, such as older adults with severe dementia, represents a critical healthcar

safetyarxiv-cs-cv
5 May 2026
Research

Directed Social Regard: Surfacing Targeted Advocacy, Opposition, Aid, Harms, and Victimization in Online Media

DGX agent

arXiv:2605.00776v1 Announce Type: new Abstract: The language in online platforms, influence operations, and political rhetoric frequently directs a mix of pro-social sentiment (e.g., advocacy, helpful

researcharxiv-cs-cl
4 May 2026
Agents

Position: agentic AI orchestration should be Bayes-consistent

DGX agent

arXiv:2605.00742v1 Announce Type: cross Abstract: LLMs excel at predictive tasks and complex reasoning tasks, but many high-value deployments rely on decisions under uncertainty, for example, which to

agentsarxiv-cs-lg
4 May 2026
Safety

Reinforcement Learning with LLM-Guided Action Spaces for Synthesizable Lead Optimization

DGX agent

arXiv:2604.07669v2 Announce Type: replace Abstract: Lead optimization in drug discovery requires improving therapeutic properties while ensuring that molecular modifications correspond to feasible syn

safetyarxiv-cs-lg
4 May 2026
Model Releases

When RAG Chatbots Expose Their Backend: An Anonymized Case Study of Privacy and Security Risks in Patient-Facing Medical AI

DGX agent

arXiv:2605.00796v1 Announce Type: cross Abstract: Background: Patient-facing medical chatbots based on retrieval-augmented generation (RAG) are increasingly promoted to deliver accessible, grounded he

model-releasesarxiv-cs-cl
4 May 2026
Applications

A Unified Framework of Hyperbolic Graph Representation Learning Methods

DGX agent

arXiv:2604.28070v1 Announce Type: new Abstract: Hyperbolic geometry has emerged as an effective latent space for representing complex networks, owing to its ability to capture hierarchical organizatio

applicationsarxiv-cs-lg
1 May 2026
Model Releases

ChipLingo: A Systematic Training Framework for Large Language Models in EDA

DGX agent

arXiv:2604.27415v1 Announce Type: new Abstract: With the rapid advancement of semiconductor technology, Electronic Design Automation (EDA) has become an increasingly knowledge-intensive and document-d

model-releasesarxiv-cs-lg
1 May 2026
Research

Language Ideologies in a Multilingual Society: An LLM-based Analysis of Luxembourgish News Comments

DGX agent

arXiv:2604.27661v1 Announce Type: new Abstract: Detecting language ideologies is a valuable yet complex task for understanding how identities are constructed through discourse. In Luxembourg's multicu

researcharxiv-cs-cl
1 May 2026
Model Releases

LLM-Guided Runtime Parameter Optimization for Energy-Efficient Model Inference

DGX agent

arXiv:2604.27032v1 Announce Type: cross Abstract: Large Language Models (LLMs) have become an integral part of many real-world workflows. However, LLMs consume a lot of energy, which becomes a large c

model-releasesarxiv-cs-lg
1 May 2026
Hardware

MARS: Efficient, Adaptive Co-Scheduling for Heterogeneous Agentic Systems

DGX agent

arXiv:2604.26963v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed as the execution core of autonomous agents rather than as standalone text generators. Agentic w

hardwarearxiv-cs-lg
1 May 2026
Agents

Pragmos: A Process Agentic Modeling System

DGX agent

arXiv:2604.27311v1 Announce Type: cross Abstract: The advent of Large Language Models (LLMs) has significantly transformed tasks across Software Engineering. In the context of Business Process Managem

agentsarxiv-cs-ai
1 May 2026
Research

Auto-ARGUE: LLM-Based Report Generation Evaluation

DGX agent

arXiv:2509.26184v5 Announce Type: replace-cross Abstract: Generation of citation-backed reports is a primary use case for retrieval-augmented generation (RAG) systems. While open-source evaluation too

researcharxiv-cs-ai
30 Apr 2026
Model Releases

Benchmarks for Trajectory Safety Evaluation and Diagnosis in OpenClaw and Codex: ATBench-Claw and ATBench-Codex

DGX agent

arXiv:2604.14858v2 Announce Type: replace Abstract: As agent systems move into increasingly diverse execution settings, trajectory-level safety evaluation and diagnosis require benchmarks that evolve

model-releasesarxiv-cs-ai
30 Apr 2026
Agents

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents

DGX agent

arXiv:2604.26752v1 Announce Type: new Abstract: We present GLM-5V-Turbo, a step toward native foundation models for multimodal agents. As foundation models are increasingly deployed in real environmen

agentsarxiv-cs-cv
30 Apr 2026
Research

NeuralEmu: in situ Measurement-Driven, ML-based, High-Fidelity 5G Network Emulation

DGX agent

arXiv:2604.26080v1 Announce Type: cross Abstract: Current and future applications demand ultra-low latency and consistent throughput, yet frequently traverse 5G cellular networks, so cope with volatil

researcharxiv-cs-lg
30 Apr 2026
Model Releases

The Prompt Engineering Report Distilled: Quick Start Guide for Life Sciences

DGX agent

arXiv:2509.11295v2 Announce Type: replace Abstract: Developing effective prompts demands significant cognitive investment to generate reliable, high-quality responses from Large Language Models (LLMs)

model-releasesarxiv-cs-cl
30 Apr 2026
Model Releases

Cutscene Agent: An LLM Agent Framework for Automated 3D Cutscene Generation

DGX agent

arXiv:2604.25318v1 Announce Type: cross Abstract: Cutscenes are carefully choreographed cinematic sequences embedded in video games and interactive media, serving as the primary vehicle for narrative

model-releasesarxiv-cs-cl
29 Apr 2026
Safety

DGLight: DQN-Guided GRPO Fine-Tuning of Large Language Models for Traffic Signal Control

DGX agent

arXiv:2604.25259v1 Announce Type: new Abstract: Traffic signal control (TSC) plays a central role in reducing congestion and maintaining urban mobility. This dissertation introduces DGLight, a critic-

safetyarxiv-cs-lg
29 Apr 2026
Safety

Generative AI Carries Non-Democratic Biases and Stereotypes: Representation of Women, Black Individuals, Age Groups, and People with Disability in AI-Generated Images across Occupations

DGX agent

arXiv:2409.13869v2 Announce Type: replace-cross Abstract: In this study, I investigate how generative artificial intelligence (AI) systems reproduce and reinforce societal biases, with a specific focu

safetyarxiv-cs-cl
29 Apr 2026
Research

Measuring the Sensitivity of Classification Models with the Error Sensitivity Profile

DGX agent

arXiv:2604.25765v1 Announce Type: new Abstract: The quality of training data is critical to the performance of machine learning models. In this paper, the Error Sensitivity Profile (ESP) is proposed.

researcharxiv-cs-lg
29 Apr 2026
Agents

MiMo-Embodied: X-Embodied Foundation Model Technical Report

DGX agent

arXiv:2511.16518v2 Announce Type: replace-cross Abstract: We open-source MiMo-Embodied, the first cross-embodied foundation model to successfully integrate and achieve state-of-the-art performance in

agentsarxiv-cs-cl
29 Apr 2026
Research

Why Do LLM-based Web Agents Fail? A Hierarchical Planning Perspective

DGX agent

arXiv:2603.14248v2 Announce Type: replace-cross Abstract: Large language model (LLM) web agents are increasingly used for web navigation but remain far from human reliability on realistic, long-horizo

researcharxiv-cs-cl
29 Apr 2026
Model Releases

A Limit Theory of Foundation Models: A Mathematical Approach to Understanding Emergent Intelligence and Scaling Laws

DGX agent

arXiv:2604.24037v1 Announce Type: new Abstract: Emergent intelligence have played a major role in the modern AI development. While existing studies primarily rely on empirical observations to characte

model-releasesarxiv-cs-lg
28 Apr 2026
Agents

Agent-Aided Design for Dynamic CAD Models

DGX agent

arXiv:2604.15184v2 Announce Type: replace Abstract: In the past year, researchers have created agentic systems that can design real-world CAD-style objects in a training-free setting, a new variety of

agentsarxiv-cs-ai
28 Apr 2026
Local Ai

AgenticCache: Cache-Driven Asynchronous Planning for Embodied AI Agents

DGX agent

arXiv:2604.24039v1 Announce Type: cross Abstract: Embodied AI agents increasingly rely on large language models (LLMs) for planning, yet per-step LLM calls impose severe latency and cost. In this pape

local-aiarxiv-cs-ai
28 Apr 2026
Model Releases

Audio2Tool: Bridging Spoken Language Understanding and Function Calling

DGX agent

arXiv:2604.22821v1 Announce Type: cross Abstract: Voice assistants increasingly rely on Speech Language Models (SpeechLMs) to interpret spoken queries and execute complex tasks, yet existing benchmark

model-releasesarxiv-cs-lg
28 Apr 2026
Safety

Bridging Reasoning and Action: Hybrid LLM-RL Framework for Efficient Cross-Domain Task-Oriented Dialogue

DGX agent

arXiv:2604.23345v1 Announce Type: new Abstract: Cross-domain task-oriented dialogue requires reasoning over implicit and explicit feasibility constraints while planning long-horizon, multi-turn action

safetyarxiv-cs-cl
28 Apr 2026
Model Releases

CiteAudit: You Cited It, But Did You Read It? A Benchmark for Verifying Scientific References in the LLM Era

DGX agent

arXiv:2602.23452v2 Announce Type: replace Abstract: Scientific research relies on accurate citation for attribution and integrity, yet large language models (LLMs) introduce a new risk: fabricated ref

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

ClimAgent: LLM as Agents for Autonomous Open-ended Climate Science Analysis

DGX agent

arXiv:2604.16922v2 Announce Type: replace Abstract: Climate research is pivotal for mitigating global environmental crises, yet the accelerating volume of multi-scale datasets and the complexity of an

model-releasesarxiv-cs-ai
28 Apr 2026
Safety

Efficient Rationale-based Retrieval: On-policy Distillation from Generative Rerankers based on JEPA

DGX agent

arXiv:2604.23336v1 Announce Type: cross Abstract: Unlike traditional fact-based retrieval, rationale-based retrieval typically necessitates cross-encoding of query-document pairs using large language

safetyarxiv-cs-cl
28 Apr 2026
Research

Exploring the Impact of Dataset Statistical Effect Size on Model Performance and Data Sample Size Sufficiency

DGX agent

arXiv:2501.02673v4 Announce Type: replace Abstract: Having a sufficient quantity of quality data is a critical enabler of training effective machine learning models. Being able to effectively determin

researcharxiv-cs-lg
28 Apr 2026
Safety

Fix Initial Codes and Iteratively Refine Textual Directions Toward Safe Multi-Turn Code Correction

DGX agent

arXiv:2604.23989v1 Announce Type: cross Abstract: Recent work on large language models (LLMs) has emphasized the importance of scaling inference compute. From this perspective, the state-of-the-art me

safetyarxiv-cs-ai
28 Apr 2026
Agents

FormalScience: Scalable Human-in-the-Loop Autoformalisation of Science with Agentic Code Generation in Lean

DGX agent

arXiv:2604.23002v1 Announce Type: new Abstract: Formalising informal mathematical reasoning into formally verifiable code is a significant challenge for large language models. In scientific fields suc

agentsarxiv-cs-ai
28 Apr 2026
Model Releases

NeuroClaw Technical Report

DGX agent

arXiv:2604.24696v1 Announce Type: new Abstract: Agentic artificial intelligence systems promise to accelerate scientific workflows, but neuroimaging poses unique challenges: heterogeneous modalities (

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

OmniSch: A Multimodal PCB Schematic Benchmark For Structured Diagram Visual Reasoning

DGX agent

arXiv:2604.00270v2 Announce Type: replace Abstract: Recent large multimodal models (LMMs) have made rapid progress in visual grounding, document understanding, and diagram reasoning tasks. However, th

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

See Further, Think Deeper: Advancing VLM's Reasoning Ability with Low-level Visual Cues and Reflection

DGX agent

arXiv:2604.24339v1 Announce Type: cross Abstract: Recent advances in Vision-Language Models (VLMs) have benefited from Reinforcement Learning (RL) for enhanced reasoning. However, existing methods sti

model-releasesarxiv-cs-ai
28 Apr 2026
Applications

What Understanding Means in AI-Laden Astronomy

DGX agent

arXiv:2601.10038v2 Announce Type: replace-cross Abstract: Artificial intelligence is rapidly transforming astronomical research, yet the scientific community has largely treated this transformation as

applicationsarxiv-cs-ai
28 Apr 2026
Local Ai

Agentic World Modeling: Foundations, Capabilities, Laws, and Beyond

DGX agent

arXiv:2604.22748v1 Announce Type: new Abstract: As AI systems move from generating text to accomplishing goals through sustained interaction, the ability to model environment dynamics becomes a centra

local-aiarxiv-cs-ai
27 Apr 2026
Agents

AgentMark: Utility-Preserving Behavioral Watermarking for Agents

DGX agent

arXiv:2601.03294v2 Announce Type: replace-cross Abstract: LLM-based agents are increasingly deployed to autonomously solve complex tasks, raising urgent needs for IP protection and regulatory provenan

agentsarxiv-cs-ai
27 Apr 2026
Local Ai

An LLM-Driven Closed-Loop Autonomous Learning Framework for Robots Facing Uncovered Tasks in Open Environments

DGX agent

arXiv:2604.22199v1 Announce Type: cross Abstract: Autonomous robots operating in open environments need the ability to continuously handle tasks that are not covered by predefined local methods. Howev

local-aiarxiv-cs-ai
27 Apr 2026
Research

Explanation of Dynamic Physical Field Predictions using WassersteinGrad: Application to Autoregressive Weather Forecasting

DGX agent

arXiv:2604.22580v1 Announce Type: cross Abstract: As the demand to integrate Artificial Intelligence into high-stakes environments continues to grow, explaining the reasoning behind neural-network pre

researcharxiv-cs-lg
27 Apr 2026
Research

Fast, close, non-singular and property-preserving approximations of entropic measures

DGX agent

arXiv:2505.14234v2 Announce Type: replace-cross Abstract: Entropic measures like Shannon entropy (SE), its quantum mechanical analogue von Neumann entropy, and Kullback-Leibler divergence (KL) are key

researcharxiv-cs-ai
27 Apr 2026
← Previous
1…4950515253…109
Next →