AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,289 results
Model Releases

AutoRISE: Agent-Driven Strategy Evolution for Red-Teaming Large Language Models

DGX agent

arXiv:2604.22871v1 Announce Type: cross Abstract: Automated red-teaming methods for large language models typically optimize attack prompts within a fixed, human-designed strategy, leaving the attack

model-releasesarxiv-cs-ai
28 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

Enabling Transparent Cyber Threat Intelligence Combining Large Language Models and Domain Ontologies

DGX agent

arXiv:2509.00081v2 Announce Type: replace-cross Abstract: Effective Cyber Threat Intelligence (CTI) relies upon accurately structured and semantically enriched information extracted from cybersecurity

agentsarxiv-cs-ai
28 Apr 2026
Agents

PARASITE: Conditional System Prompt Poisoning to Hijack LLMs

DGX agent

arXiv:2505.16888v4 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly deployed via third-party system prompts downloaded from public marketplaces. We identify a criti

agentsarxiv-cs-ai
28 Apr 2026
Agents

PhysNote: Self-Knowledge Notes for Evolvable Physical Reasoning in Vision-Language Model

DGX agent

arXiv:2604.24443v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have demonstrated strong performance on textbook-style physics problems, yet they frequently fail when confronted with dyn

agentsarxiv-cs-ai
28 Apr 2026
Model Releases

QED: An Open-Source Multi-Agent System for Generating Mathematical Proofs on Open Problems

DGX agent

arXiv:2604.24021v1 Announce Type: new Abstract: We explore a central question in AI for mathematics: can AI systems produce original, nontrivial proofs for open research problems? Despite strong bench

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

RefEvo: Agentic Design with Co-Evolutionary Verification for Agile Reference Model Generation

DGX agent

arXiv:2604.24218v1 Announce Type: cross Abstract: As the complexity of System-on-Chip (SoC) designs grows, the shift-left paradigm necessitates the rapid development of high-fidelity reference models

model-releasesarxiv-cs-ai
28 Apr 2026
Safety

Rewarding the Scientific Process: Process-Level Reward Modeling for Agentic Data Analysis

DGX agent

arXiv:2604.24198v1 Announce Type: cross Abstract: Process Reward Models (PRMs) have achieved remarkable success in augmenting the reasoning capabilities of Large Language Models (LLMs) within static d

safetyarxiv-cs-ai
28 Apr 2026
Agents

Architectures for Robust Self-Organizing Energy Systems under Information and Control Constraints

DGX agent

arXiv:2604.21529v1 Announce Type: cross Abstract: Applying the concept of controlled self-organization in agent-based Cyber-Physical Energy Systems (CPES) is a promising approach to ensure system robu

agentsarxiv-cs-ai
24 Apr 2026
Agents

Clinically Interpretable Sepsis Early Warning via LLM-Guided Simulation of Temporal Physiological Dynamics

DGX agent

arXiv:2604.20924v1 Announce Type: new Abstract: Timely and interpretable early warning of sepsis remains a major clinical challenge due to the complex temporal dynamics of physiological deterioration.

agentsarxiv-cs-lg
24 Apr 2026
Safety

FairQE: Multi-Agent Framework for Mitigating Gender Bias in Translation Quality Estimation

DGX agent

arXiv:2604.21420v1 Announce Type: new Abstract: Quality Estimation (QE) aims to assess machine translation quality without reference translations, but recent studies have shown that existing QE models

safetyarxiv-cs-ai
24 Apr 2026
Model Releases

From Research Question to Scientific Workflow: Leveraging Agentic AI for Science Automation

DGX agent

arXiv:2604.21910v1 Announce Type: new Abstract: Scientific workflow systems automate execution -- scheduling, fault tolerance, resource management -- but not the semantic translation that precedes it.

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

GeoMind: An Agentic Workflow for Lithology Classification with Reasoned Tool Invocation

DGX agent

arXiv:2604.21501v1 Announce Type: new Abstract: Lithology classification in well logs is a fundamental geoscience data mining task that aims to infer rock types from multi dimensional geophysical sequ

model-releasesarxiv-cs-ai
24 Apr 2026
Applications

Promoting Simple Agents: Ensemble Methods for Event-Log Prediction

DGX agent

arXiv:2604.21629v1 Announce Type: cross Abstract: We compare lightweight automata-based models (n-grams) with neural architectures (LSTM, Transformer) for next-activity prediction in streaming event l

applicationsarxiv-cs-ai
24 Apr 2026
Agents

Task-Driven Co-Design of Heterogeneous Multi-Robot Systems

DGX agent

arXiv:2604.21894v1 Announce Type: new Abstract: Designing multi-agent robotic systems requires reasoning across tightly coupled decisions spanning heterogeneous domains, including robot design, fleet

agentsarxiv-cs-ro
24 Apr 2026
Agents

Unveiling Unicode's Unseen Underpinnings in Undermining Authorship Attribution

DGX agent

arXiv:2508.15840v5 Announce Type: replace-cross Abstract: When using a public communication channel--whether formal or informal, such as commenting or posting on social media--end users have no expect

agentsarxiv-cs-cl
24 Apr 2026
Model Releases

Bimanual Robot Manipulation via Multi-Agent In-Context Learning

DGX agent

arXiv:2604.20348v1 Announce Type: cross Abstract: Language Models (LLMs) have emerged as powerful reasoning engines for embodied control. In particular, In-Context Learning (ICL) enables off-the-shelf

model-releasesarxiv-cs-ai
23 Apr 2026
Agents

FlexServe: A Fast and Secure LLM Serving System for Mobile Devices with Flexible Resource Isolation

DGX agent

arXiv:2603.09046v2 Announce Type: replace-cross Abstract: Device-side Large Language Models (LLMs) have witnessed explosive growth, offering higher privacy and availability compared to cloud-side LLMs

agentsarxiv-cs-lg
23 Apr 2026
Safety

LLM-Guided Safety Agent for Edge Robotics with an ISO-Compliant Perception-Compute-Control Architecture

DGX agent

arXiv:2604.20193v1 Announce Type: new Abstract: Ensuring functional safety in human-robot interaction is challenging because AI perception is inherently probabilistic, whereas industrial standards req

safetyarxiv-cs-ro
23 Apr 2026
Agents

Neural Bandit Based Optimal LLM Selection for a Pipeline of Subtasks

DGX agent

arXiv:2508.09958v3 Announce Type: replace Abstract: As large language models (LLMs) become increasingly popular, there is a growing need to predict which out of a set of LLMs will yield a successful a

agentsarxiv-cs-cl
23 Apr 2026
Hardware

OpenCLAW-P2P v6.0: Resilient Multi-Layer Persistence, Live Reference Verification, and Production-Scale Evaluation of Decentralized AI Peer Review

DGX agent

arXiv:2604.19792v1 Announce Type: new Abstract: This paper presents OpenCLAW-P2P v6.0, a comprehensive evolution of the decentralized collective-intelligence platform in which autonomous AI agents pub

hardwarearxiv-cs-ai
23 Apr 2026
Agents

Statistics, Not Scale: Modular Medical Dialogue with Bayesian Belief Engine

DGX agent

arXiv:2604.20022v1 Announce Type: cross Abstract: Large language models are increasingly deployed as autonomous diagnostic agents, yet they conflate two fundamentally different capabilities: natural-l

agentsarxiv-cs-ai
23 Apr 2026
Model Releases

To Know is to Construct: Schema-Constrained Generation for Agent Memory

DGX agent

arXiv:2604.20117v1 Announce Type: new Abstract: Constructivist epistemology argues that knowledge is actively constructed rather than passively copied. Despite the generative nature of Large Language

model-releasesarxiv-cs-cl
23 Apr 2026
Agents

A Survey on MLLM-based Visually Rich Document Understanding: Methods, Challenges, and Emerging Trends

DGX agent

arXiv:2507.09861v2 Announce Type: replace-cross Abstract: Visually Rich Document Understanding (VRDU) has become a pivotal area of research, driven by the need to automatically interpret documents tha

agentsarxiv-cs-ai
22 Apr 2026
Model Releases

Council Mode: Mitigating Hallucination and Bias in LLMs via Multi-Agent Consensus

DGX agent

arXiv:2604.02923v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs), particularly those employing Mixture-of-Experts (MoE) architectures, have achieved remarkable capabilities acros

model-releasesarxiv-cs-ai
22 Apr 2026
Safety

MRS: Multi-Resolution Skills for HRL Agents

DGX agent

arXiv:2505.21410v2 Announce Type: replace Abstract: Hierarchical reinforcement learning (HRL) decomposes the policy into a manager and a worker, enabling long-horizon planning but introducing a perfor

safetyarxiv-cs-ai
22 Apr 2026
Model Releases

ORCA: An Agentic Reasoning Framework for Hallucination and Adversarial Robustness in Vision-Language Models

DGX agent

arXiv:2509.15435v2 Announce Type: replace-cross Abstract: Large Vision-Language Models (LVLMs) exhibit strong multimodal capabilities but remain vulnerable to hallucinations from intrinsic errors and

model-releasesarxiv-cs-ai
22 Apr 2026
Agents

Personalized Embodied Navigation for Portable Object Finding

DGX agent

arXiv:2403.09905v5 Announce Type: replace-cross Abstract: Embodied navigation methods commonly operate in static environments with stationary objects. In this work, we present approaches for tackling

agentsarxiv-cs-cv
22 Apr 2026
Agents

Human Cognition in Machines: A Unified Perspective of World Models

DGX agent

arXiv:2604.16592v1 Announce Type: cross Abstract: This comprehensive report distinguishes prior works by the cognitive functions they innovate. Many works claim an almost 'human-like' cognitive capabi

agentsarxiv-cs-cv
21 Apr 2026
Agents

Learning to Trade Like an Expert: Cognitive Fine-Tuning for Stable Financial Reasoning in Language Models

DGX agent

arXiv:2604.16862v1 Announce Type: new Abstract: Recent deployments of large language models (LLMs) as autonomous trading agents raise questions about whether financial decision-making competence gener

agentsarxiv-cs-lg
21 Apr 2026
Model Releases

SeekerGym: A Benchmark for Reliable Information Seeking

DGX agent

arXiv:2604.17143v1 Announce Type: new Abstract: Despite their substantial successes, AI agents continue to face fundamental challenges in terms of trustworthiness. Consider deep research agents, taske

model-releasesarxiv-cs-lg
21 Apr 2026
Agents

Sensorimotor Self-Recognition in Multimodal Large Language Model-Driven Robots

DGX agent

arXiv:2505.19237v2 Announce Type: replace-cross Abstract: Self-recognition -- the ability to maintain an internal representation of one's own body within the environment -- underpins intelligent, auto

agentsarxiv-cs-ro
21 Apr 2026
Agents

Evaluating LLM Simulators as Differentially Private Data Generators

DGX agent

arXiv:2604.15461v1 Announce Type: cross Abstract: LLM-based simulators offer a promising path for generating complex synthetic data where traditional differentially private (DP) methods struggle with

agentsarxiv-cs-cl
20 Apr 2026
Model Releases

Making Image Editing Easier via Adaptive Task Reformulation with Agentic Executions

DGX agent

arXiv:2604.15917v1 Announce Type: new Abstract: Instruction guided image editing has advanced substantially with recent generative models, yet it still fails to produce reliable results across many se

model-releasesarxiv-cs-cv
20 Apr 2026
Safety

The Informational Cost of Agency: A Bounded Measure of Interaction Efficiency for Deployed Reinforcement Learning

DGX agent

arXiv:2603.01283v2 Announce Type: replace Abstract: Deployed RL agents operate in closed-loop systems where reliable performance depends on maintaining coherent coupling between observations, actions,

safetyarxiv-cs-ai
20 Apr 2026
Model Releases

ADAPT: Benchmarking Commonsense Planning under Unspecified Affordance Constraints

DGX agent

arXiv:2604.14902v1 Announce Type: cross Abstract: Intelligent embodied agents should not simply follow instructions, as real-world environments often involve unexpected conditions and exceptions. Howe

model-releasesarxiv-cs-cl
17 Apr 2026
Agents

CROP: Token-Efficient Reasoning in Large Language Models via Regularized Prompt Optimization

DGX agent

arXiv:2604.14214v1 Announce Type: new Abstract: Large Language Models utilizing reasoning techniques improve task performance but incur significant latency and token costs due to verbose generation. E

agentsarxiv-cs-cl
17 Apr 2026
Agents

MIND: AI Co-Scientist for Material Research

DGX agent

arXiv:2604.13699v1 Announce Type: cross Abstract: Large language models (LLMs) have enabled agentic AI systems for scientific discovery, but most approaches remain limited to textbased reasoning witho

agentsarxiv-cs-ai
17 Apr 2026
Model Releases

ReviewGrounder: Improving Review Substantiveness with Rubric-Guided, Tool-Integrated Agents

DGX agent

arXiv:2604.14261v1 Announce Type: new Abstract: The rapid rise in AI conference submissions has driven increasing exploration of large language models (LLMs) for peer review support. However, LLM-base

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

VeriGraphi: A Multi-Agent Framework of Hierarchical RTL Generation for Large Hardware Designs

DGX agent

arXiv:2604.14550v1 Announce Type: cross Abstract: Generating synthesizable Verilog for large, hierarchical hardware designs remains a significant challenge for large language models (LLMs), which stru

model-releasesarxiv-cs-lg
17 Apr 2026
Agents

ACDC: Adaptive Curriculum Planning with Dynamic Contrastive Control for Goal-Conditioned Reinforcement Learning in Robotic Manipulation

DGX agent

arXiv:2603.02104v2 Announce Type: replace Abstract: Goal-conditioned reinforcement learning has shown considerable potential in robotic manipulation; however, existing approaches remain limited by the

agentsarxiv-cs-ro
15 Apr 2026
Safety

Advancing Multi-Agent RAG Systems with Minimalist Reinforcement Learning

DGX agent

arXiv:2505.17086v4 Announce Type: replace Abstract: Large Language Models (LLMs) equipped with modern Retrieval-Augmented Generation (RAG) systems often employ multi-turn interaction pipelines to inte

safetyarxiv-cs-cl
15 Apr 2026
Agents

AgenticAI-DialogGen: Topic-Guided Conversation Generation for Fine-Tuning and Evaluating Short- and Long-Term Memories of LLMs

DGX agent

arXiv:2604.12179v1 Announce Type: new Abstract: Recent advancements in Large Language Models (LLMs) have improved their ability to process extended conversational contexts, yet fine-tuning and evaluat

agentsarxiv-cs-cl
15 Apr 2026
Agents

Instructions are all you need: Self-supervised Reinforcement Learning for Instruction Following

DGX agent

arXiv:2510.14420v4 Announce Type: replace-cross Abstract: Language models often struggle to follow multi-constraint instructions that are crucial for real-world applications. Existing reinforcement le

agentsarxiv-cs-ai
15 Apr 2026
Agents

OVAL: Open-Vocabulary Augmented Memory Model for Lifelong Object Goal Navigation

DGX agent

arXiv:2604.12872v1 Announce Type: new Abstract: Object Goal Navigation (ObjectNav) refers to an agent navigating to an object in an unseen environment, which is an ability often required in the accomp

agentsarxiv-cs-ro
15 Apr 2026
Model Releases

A Benchmark and Multi-Agent System for Instruction-driven Cinematic Video Compilation

DGX agent

arXiv:2604.10456v1 Announce Type: new Abstract: The surging demand for adapting long-form cinematic content into short videos has motivated the need for versatile automatic video compilation systems.

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Both Ends Count! Just How Good are LLM Agents at 'Text-to-Big SQL'?

DGX agent

arXiv:2602.21480v4 Announce Type: replace-cross Abstract: Text-to-SQL and Big Data are both extensively benchmarked fields, yet there is limited research that evaluates them jointly. In the real world

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

CARE-ECG: Causal Agent-based Reasoning for Explainable and Counterfactual ECG Interpretation

DGX agent

arXiv:2604.10420v1 Announce Type: new Abstract: Large language models (LLMs) enable waveform-to-text ECG interpretation and interactive clinical questioning, yet most ECG-LLM systems still rely on wea

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

DeepReviewer 2.0: A Traceable Agentic System for Auditable Scientific Peer Review

DGX agent

arXiv:2604.09590v1 Announce Type: new Abstract: Automated peer review is often framed as generating fluent critique, yet reviewers and area chairs need judgments they can audit: where a concern applie

model-releasesarxiv-cs-ai
14 Apr 2026
← Previous
1…115116117118119…236
Next →