AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,289 results
Safety

UniIntervene: Agentic Intervention for Efficient Real-World Reinforcement Learning

DGX agent

arXiv:2606.12372v1 Announce Type: cross Abstract: Human-in-the-loop reinforcement learning (HiL-RL) has emerged as an effective paradigm for real-world robotic manipulation, enabling online policy imp

safetyarxiv-cs-lg
11 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Local Ai

AnomaMind: Agentic Time Series Anomaly Detection with Tool-Augmented Reasoning

DGX agent

arXiv:2602.13807v2 Announce Type: replace Abstract: Time series anomaly detection is critical in many real-world applications, where effective solutions must localize anomalous regions and support rel

local-aiarxiv-cs-lg
10 Jun 2026
Safety

Causal Ensemble Agent: Hierarchical Causal Discovery with LLM-guided Expert Reweighting

DGX agent

arXiv:2606.10607v1 Announce Type: cross Abstract: Causal discovery aims to uncover causal structures from observational data, which is crucial for real-world decision-making. However, different causal

safetyarxiv-cs-ai
10 Jun 2026
Research

Soul Computing: A Theoretical Framework and Technical Architecture for Intelligent Agents with Independent Consciousness

DGX agent

arXiv:2606.10413v1 Announce Type: new Abstract: Breakthroughs in large language models and multimodal generation technologies have propelled the digital reconstruction of human mental traits, emotiona

researcharxiv-cs-ai
10 Jun 2026
Agents

Byzantine Cheap Talk: Adversarial Resilience and Topology Effects in LLM Coordination Games

DGX agent

arXiv:2606.07790v1 Announce Type: new Abstract: Multi-agent LLM systems increasingly rely on communication protocols for coordination, yet their robustness under adversarial and structural constraints

agentsarxiv-cs-lg
9 Jun 2026
Model Releases

FineGen: A VLM-based Multi-Agent Framework for Fine-Grained Image-Text Dataset Construction

DGX agent

arXiv:2606.07645v1 Announce Type: cross Abstract: The scarcity of hard negative samples in current vision-language datasets significantly hinders fine-grained perception. To address this, we propose F

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

HDSL: A Hierarchical Domain-Specific Language for Structured 3D Indoor Scene Generation and Localized Editing with LLM Agents

DGX agent

arXiv:2606.09738v1 Announce Type: new Abstract: Text-driven indoor scene generation and editing require an intermediate representation that language models can both produce and revise. Existing LLM-ba

model-releasesarxiv-cs-cv
9 Jun 2026
Safety

IEA: Amateur-Friendly Conversational Image Editing Agent via Three Stages of Multitask Alignment

DGX agent

arXiv:2606.08016v1 Announce Type: cross Abstract: Current image editing software often hinges on fixed filters or expert tuning, leaving a gap between amateur users' intent and outcomes. Creations by

safetyarxiv-cs-ai
9 Jun 2026
Agents

In-Context Reinforcement Learning via Communicative World Models

DGX agent

arXiv:2508.06659v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) agents often struggle to generalize to new tasks and contexts without updating their parameters, mainly because th

agentsarxiv-cs-ai
9 Jun 2026
Model Releases

MAGIS: Evidence-Based Multi-Agent Reasoning for Interpretable Strabismus Clinical Decision-Making

DGX agent

arXiv:2606.09249v1 Announce Type: new Abstract: Strabismus is a common ocular disorder that requires fine-grained subtype diagnosis for individualized treatment planning. However, existing deep learni

model-releasesarxiv-cs-cv
9 Jun 2026
Safety

Agentic World Modeling for 6G: Near-Real-Time Generative State-Space Reasoning

DGX agent

arXiv:2511.02748v2 Announce Type: replace-cross Abstract: We argue that sixth-generation (6G) intelligence is not fluent token prediction but the capacity to imagine and choose -- to simulate future s

safetyarxiv-cs-lg
8 Jun 2026
Agents

CRAFT: Coaching Reinforcement Learning Autonomously using Foundation Models for Multi-Robot Coordination Tasks

DGX agent

arXiv:2509.14380v3 Announce Type: replace Abstract: Multi-Agent Reinforcement Learning (MARL) provides a powerful framework for learning coordination in multi-agent systems. However, applying MARL to

agentsarxiv-cs-ro
8 Jun 2026
Model Releases

What Your Posts Reveal: A Benchmark and Agentic Framework for User-Level Privacy Leakage on Social Media

DGX agent

arXiv:2606.06784v1 Announce Type: cross Abstract: Public social media posts can reveal private information through weak cues scattered across text, images, or metadata. Such leakage is often cumulativ

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

Evaluating Agentic Configuration Repair for Computer Networks

DGX agent

arXiv:2606.06212v1 Announce Type: new Abstract: Misconfigurations in computer networks remain a major source of critical Internet outages. Research is turning to Large Language Models (LLMs) to automa

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

How Far Did They Go? The Persuasive Tactics of Covert LLM Agents in a Discontinued Field Experiment

DGX agent

arXiv:2606.05256v1 Announce Type: new Abstract: This study analyzes a publicly released dataset from a discontinued field experiment on Reddit's r/ChangeMyView. The intervention, conducted by unknown,

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

ProfiliTable: Profiling-Driven Tabular Data Processing via Agentic Workflows

DGX agent

arXiv:2605.12376v2 Announce Type: replace Abstract: Table processing-including cleaning, transformation, augmentation, and matching-is a foundational yet error-prone stage in real-world data pipelines

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

AURA: Intent-Directed Probing for Implicit-Need Surfacing in Situated LLM Agents

DGX agent

arXiv:2606.05557v1 Announce Type: new Abstract: A situated query like 'where is Lin Wei?' often encodes more than its literal content: the user may also want to know whether Lin Wei is free, in a good

model-releasesarxiv-cs-cl
5 Jun 2026
Safety

HANDOFF: Humanoid Agentic Task-Space Whole-Body Control via Distilled Complementary Teachers

DGX agent

arXiv:2606.06493v1 Announce Type: new Abstract: For a humanoid robot to be deployed in the real world, the choice of command space (i.e., the interface between task planning and whole-body control) is

safetyarxiv-cs-ro
5 Jun 2026
Safety

Membrane: A Self-Evolving Contrastive Safety Memory for LLM Agent Defense

DGX agent

arXiv:2606.05743v1 Announce Type: cross Abstract: Despite advances in safety alignment, large language models remain vulnerable to continuously evolving jailbreaks. Existing fine-tuned safety classifi

safetyarxiv-cs-cl
5 Jun 2026
Model Releases

ProSPy: A Profiling-Driven SQL-Python Agentic Framework for Enterprise Text-to-SQL

DGX agent

arXiv:2606.05836v1 Announce Type: new Abstract: Large language models have substantially advanced Text-to-SQL systems, yet applying them to enterprise-scale databases remains challenging. Real-world d

model-releasesarxiv-cs-cl
5 Jun 2026
Agents

Consensus is Strategically Insufficient: Reasoning-Trace Disagreement as a Knowledge-Representation Signal

DGX agent

arXiv:2606.04223v1 Announce Type: new Abstract: Multi-agent systems are commonly designed to reduce disagreement through voting, consensus protocols, debate, or fault-tolerant aggregation. We argue th

agentsarxiv-cs-ai
4 Jun 2026
Model Releases

HighTide: An Agent-Curated Open-Source VLSI Benchmark Suite

DGX agent

arXiv:2606.04126v1 Announce Type: cross Abstract: We introduce HighTide, an evolving AI-assisted benchmark suite. Specifically, the contributions are: (i) a diverse open-source suite spanning multiple

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Impostor: An Agent-Curated Benchmark for Realistic AIGC Manipulation Localization

DGX agent

arXiv:2606.04545v1 Announce Type: new Abstract: Recent advances in generative image editing have improved the realism and controllability of localized image manipulation, raising new challenges for im

model-releasesarxiv-cs-cv
4 Jun 2026
Model Releases

LifeSide: Benchmarking Agents as Lifelong Digital Companions

DGX agent

arXiv:2606.04660v1 Announce Type: new Abstract: Lifelong digital companions must integrate cross-session cues, continually update their understanding of users, and adapt to shifting privacy boundaries

model-releasesarxiv-cs-cl
4 Jun 2026
Safety

Neetyabhas: A Framework for Uncertainty-Aware Public Policy Optimization in Rational Agent-Based Models

DGX agent

arXiv:2606.04562v1 Announce Type: new Abstract: Purpose The WHO's COVID-19 non-pharmaceutical interventions (e.g., lockdowns, vaccinations) effectively curb transmission but impose heavy economic stra

safetyarxiv-cs-ai
4 Jun 2026
Safety

Simplicial Embeddings Improve Sample Efficiency in Actor-Critic Agents

DGX agent

arXiv:2510.13704v2 Announce Type: replace-cross Abstract: Recent works have proposed accelerating the wall-clock training time of actor-critic methods via the use of large-scale environment paralleliz

safetyarxiv-cs-ai
4 Jun 2026
Model Releases

Towards Efficient and Evidence-grounded Mobility Prediction with LLM-Driven Agent

DGX agent

arXiv:2606.05130v1 Announce Type: cross Abstract: Individual-level mobility prediction is central to urban simulation, transportation planning, and policy analysis. Supervised sequence models achieve

model-releasesarxiv-cs-ai
4 Jun 2026
Safety

AI Agents Enable Adaptive Computer Worms

DGX agent

arXiv:2606.03811v1 Announce Type: cross Abstract: A computer worm is malware that spreads on a network by replicating itself from one machine to another. Traditional worms, like WannaCry, exploited pr

safetyarxiv-cs-ai
3 Jun 2026
Model Releases

BigFinanceBench: A Workflow-Grounded Benchmark for Financial-Research Agents

DGX agent

arXiv:2606.03829v1 Announce Type: new Abstract: Financial-research answers are decision-relevant only when another analyst can audit how they were produced: which source was chosen, which period and a

model-releasesarxiv-cs-ai
3 Jun 2026
Agents

Human-Like Goalkeeping in a Realistic Football Simulation: a Sample-Efficient Reinforcement Learning Approach

DGX agent

arXiv:2510.23216v4 Announce Type: replace Abstract: While several high profile video games have served as testbeds for Deep Reinforcement Learning (DRL), this technique has rarely been employed by the

agentsarxiv-cs-ai
3 Jun 2026
Model Releases

JAVEDIT: Joint Audio-Visual Instruction-Guided Video Editing with Agentic Data Curation

DGX agent

arXiv:2606.03168v1 Announce Type: new Abstract: While instruction-based video editing has seen significant progress, joint audio-visual editing remains constrained by the absence of dedicated datasets

model-releasesarxiv-cs-cv
3 Jun 2026
Model Releases

LEAP: Supercharging LLMs for Formal Mathematics with Agentic Frameworks

DGX agent

arXiv:2606.03303v1 Announce Type: new Abstract: Large Language Models (LLMs) exhibit strong informal mathematical reasoning but struggle to generate mechanically verifiable proofs in formal languages

model-releasesarxiv-cs-ai
3 Jun 2026
Hardware

Scaling Multi Agent Reinforcement Learning for Underwater Acoustic Tracking via Autonomous Vehicles

DGX agent

arXiv:2505.08222v3 Announce Type: replace-cross Abstract: Autonomous vehicles (AVs) offer a cost-effective solution for scientific missions such as underwater tracking. Reinforcement learning (RL) has

hardwarearxiv-cs-ai
3 Jun 2026
Model Releases

Skill-RM: Unifying Heterogeneous Evaluation Criteria via Agent Skill

DGX agent

arXiv:2606.03980v1 Announce Type: cross Abstract: Reward models (RMs) provide critical feedback signals for LLM post-training, notably in reinforced fine-tuning (RFT) and reinforcement learning (RL) p

model-releasesarxiv-cs-cl
3 Jun 2026
Model Releases

3DCodeBench: Benchmarking Agentic Procedural 3D Modeling Via Code

DGX agent

arXiv:2606.01057v1 Announce Type: cross Abstract: Procedural 3D modeling through code is emerging as a versatile paradigm, offering deterministic, engine-ready, and precisely editable assets that neur

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Active Exploring like a Pigeon: Reinforcing Spatial Reasoning via Agentic Vision-Language Models

DGX agent

arXiv:2606.02459v1 Announce Type: new Abstract: Enabling Vision-Language Models (VLMs) to perform spatial reasoning remains challenging. Existing approaches treat VLMs as passive observers, which is d

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

AgentPLM: Agentic Protein Language Models with Reasoning-Augmented Decoding for Protein Sequence Design

DGX agent

arXiv:2606.02386v1 Announce Type: new Abstract: Protein language models (PLMs) are passive oracles: they generate sequences in a single forward pass with no mechanism to consult external biophysical f

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Are LLMs Ready for Neural-integrated Mechanistic Modeling? A Benchmark and Agentic Framework

DGX agent

arXiv:2602.18008v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have shown promise in constructing mechanistic models from data. However, existing evaluations largely focus on s

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

ATLAS: Agentic Test-time Learning-to-Allocate Scaling

DGX agent

arXiv:2606.01667v1 Announce Type: new Abstract: Test-time scaling has become a major way to improve large language model reasoning, but its orchestration has remained designer-engineered: a fixed samp

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

Better with Experience: Self-Evolving LLM Agents for Evidence-Grounded Health Community Notes

DGX agent

arXiv:2606.02215v1 Announce Type: new Abstract: Large Language Model (LLM)-augmented Community Notes offer a scalable path for timely, evidence-grounded correction of health misinformation on social p

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

ClinEnv: An Interactive Multi-Stage Long Horizon EHR Environment for Agents

DGX agent

arXiv:2606.02568v1 Announce Type: new Abstract: Clinical practice is not the selection of an answer from enumerated options: a physician gathers heterogeneous information incrementally and commits to

model-releasesarxiv-cs-ai
2 Jun 2026
Safety

Dive into Ambiguity: A*-Inspired Multi-Agents Commonsense Obfuscation Attack on LLM Prompts

DGX agent

arXiv:2606.01441v1 Announce Type: new Abstract: Large language models (LLMs) excel in reasoning and knowledge-intensive tasks but remain vulnerable to prompt-level adversarial attacks that preserve in

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

DrugClaw and DrugAudit: A Primary-Source-Grounded Agent and Authority-Aware Benchmark for Drug-Information Question Answering

DGX agent

arXiv:2606.01434v1 Announce Type: new Abstract: Drug-information question answering is a high-stakes setting where hallucinated facts can mislead clinical decision-making and the provenance of each ci

model-releasesarxiv-cs-cl
2 Jun 2026
Agents

Easier to Mislead Than to Correct: Harmful and Beneficial Revision in LLM Conformity

DGX agent

arXiv:2606.01637v1 Announce Type: cross Abstract: Large language models are increasingly used in multi-agent systems, where they see and respond to other agents' answers. A key risk is conformity: a m

agentsarxiv-cs-ai
2 Jun 2026
Local Ai

PEACE: A Planner-Executor Agent with Constraint Enforcement for UAVs

DGX agent

arXiv:2606.00104v1 Announce Type: cross Abstract: Foundation models are increasingly used to drive autonomous systems, yet existing approaches either keep the model in a tight control loop, raising la

local-aiarxiv-cs-ai
2 Jun 2026
Model Releases

SMH-Bench: Benchmarking LLM Agents for Environment-Grounded Reasoning and Action in Smart Homes

DGX agent

arXiv:2606.01912v1 Announce Type: new Abstract: Smart homes are evolving toward complex state-dependent living environments, requiring Large Language Models (LLMs) to reason over user intent, preferen

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

TravelEval: A Comprehensive Benchmarking Framework for Evaluating LLM-Powered Travel Planning Agents

DGX agent

arXiv:2606.01046v1 Announce Type: new Abstract: The development of Large Language Models (LLMs) has significantly improved travel planning applications, yet evaluating such models is limited by existi

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Worlds Within Words: Translating Culture in Ancient Chinese Texts with Multi-Agent Coordination

DGX agent

arXiv:2606.01276v1 Announce Type: new Abstract: Large language model (LLM)-based machine translation has advanced cross-cultural communication, yet it still struggles with culture-loaded words (CLWs)

model-releasesarxiv-cs-cl
2 Jun 2026
← Previous
1…112113114115116…236
Next →