AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,964 results
11 Jun 2026

TreeSeeker: Tree-Structured Trial, Error, and Return in Deep Search

AgentsDGX agent

arXiv:2606.11662v1 Announce Type: new Abstract: Deep search requires agents to answer complex questions through multi-step web search, browsing, evidence comparison, and synthesis. A central challenge

10 Jun 2026

ActiveMem: Distributed Active Memory for Long-Horizon LLM Reasoning

AgentsDGX agent

arXiv:2606.10532v1 Announce Type: new Abstract: Memory is essential for enabling large language model (LLM) agents to handle long-horizon reasoning tasks. Existing memory mechanisms are largely centra

Robust Deep Reinforcement Learning Through Adversarial Attacks and Training : A Survey

AgentsDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2403.00420v3 Announce Type: replace-cross Abstract: Deep Reinforcement Learning (DRL) is a subfield of machine learning for training autonomous agents that take sequential actions across complex

Understanding and mitigating the risks of OpenClaw for non-technical users: A practical guide with Skill

AgentsDGX agent

arXiv:2606.11007v1 Announce Type: cross Abstract: OpenClaw has rapidly emerged as a transformative artificial intelligence (AI) agent framework, and its ability to autonomously execute complex, multi-

9 Jun 2026

MASS: Deep Research for Social Sciences with Memory-Augmented Social Simulation

AgentsDGX agent

arXiv:2606.09198v1 Announce Type: new Abstract: Deep Research agents powered by Large Language Models (LLMs) have exhibited extraordinary potential in automated paper writing tasks. However, existing

PRISM: Recovering Instruction Sets from Language Model Activations

AgentsDGX agent

arXiv:2606.09563v1 Announce Type: new Abstract: As LLMs are deployed as agents, reliable monitoring requires knowing not only what they output, but which instructions are steering their behavior. This

8 Jun 2026

deepagents in 90 seconds!

AgentsDGX agent

DeepAgents is a framework or tool designed to enable the creation of intelligent agents, likely demonstrated or explained in a brief 90-second format by Harrison Chase, a co-founder of LangChain. The

@photon_hq More on the Photon integration: https://hermes-agent.nousresearch.com/docs/user-guide/messaging/photon

AgentsDGX agent

Nous Research shared documentation about their Photon integration, providing users with guidance on implementing Photon messaging functionality within the Hermes Agent platform. The resource likely co

7 Jun 2026

The Top AI Papers of the Week (May 31 - June 7) - LEAP - AutoLab - Learn From Your Own Latents - Reusable Context Engineering - Self-Revisin…

AgentsDGX agent

The Top AI Papers of the Week (May 31 - June 7) - LEAP - AutoLab - Learn From Your Own Latents - Reusable Context Engineering - Self-Revising Discovery Systems - Scaling Laws for Agent Harnesses - Dis

6 Jun 2026

Microskill Architecture: A Modular Skill-Driven Framework for AI-Native Code Generation

AgentsDGX agent

arXiv:2606.05720v1 Announce Type: cross Abstract: Large language models and AI coding agents have reshaped software development, but the path to fully AI-native systems faces structural challenges. Ch

5 Jun 2026

Enhancing Multi-Robot Exploration Using Probabilistic Frontier Prioritization with Dirichlet Process Gaussian Mixtures

AgentsDGX agent

arXiv:2604.03042v2 Announce Type: replace Abstract: Multi-agent autonomous exploration is essential for applications such as environmental monitoring, search and rescue, and industrial-scale surveilla

4 Jun 2026

FALSIFYBENCH: Evaluating Inductive Reasoning in LLMs with Rule Discovery Games

AgentsDGX agent

arXiv:2606.04751v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed as autonomous agents in scientific tasks. Yet whether these systems can effectively engage in for

From idea to live store, in minutes. 🛍️ Tomorrow on the Friday Showcase: → Replit's new Shopify partnership, with @Davidizek . Tell Replit …

AgentsDGX agent

From idea to live store, in minutes. 🛍️ Tomorrow on the Friday Showcase: → Replit's new Shopify partnership, with @Davidizek . Tell Replit Agent what you want to sell and it builds your storefront, cr

INTACT: Ego-Guided Typed Sparse Evidence Retrieval for Heterogeneous Collaborative Perception

AgentsDGX agent

arXiv:2606.04437v1 Announce Type: new Abstract: Collaborative perception extends the perceptual range of autonomous vehicles by sharing information across agents, but heterogeneous sensors and percept

TIDE: Proactive Multi-Problem Discovery via Template-Guided Iteration

AgentsDGX agent

arXiv:2606.04743v1 Announce Type: cross Abstract: Agents are widely deployed as assistants over documents, tools, and code. However, they typically act only on explicit user requests, which surface on

3 Jun 2026

Closed-Loop Molecular Design with Calibrated Deference

AgentsDGX agent

arXiv:2606.02618v1 Announce Type: cross Abstract: We present Cognitive Loop via In-Situ Optimization (CLIO), an agent that couples a continuously-updated belief-state graph with a recursive plan-then-

Learning to See via Epiretinal Implant Stimulation in silico with Model-Based Deep Reinforcement Learning

AgentsDGX agent

arXiv:2606.03118v1 Announce Type: cross Abstract: Objective: Diseases such as age-related macular degeneration and retinitis pigmentosa cause the degradation of the photoreceptor layer. One approach t

2 Jun 2026

ClawHub Security Signals: When VirusTotal, Static Analysis, and SkillSpector Disagree

Model ReleasesDGX agent

arXiv:2606.01494v1 Announce Type: cross Abstract: Agent skills extend AI agents with reusable instructions, tools, scripts, references, and workflows, establishing a security boundary distinct from bo

Interaction-Centered Intelligence: Toward Interaction as the Primary Unit of Analysis in Co-Creative AI and Human-AI Systems

AgentsDGX agent

arXiv:2606.00807v1 Announce Type: new Abstract: Traditional artificial intelligence has largely conceptualized intelligence as isolated computation occurring within bounded agents. Across classical AI

MindGames Arena Generalization Track: In2AI Solution with Delayed Per-Step Reward Attribution

Model ReleasesDGX agent

arXiv:2606.00017v1 Announce Type: new Abstract: Training language model agents for multi-agent strategic interaction presents a core difficulty: the quality of any action may depend on future events t

Strong AI data foundations turn enterprise chaos into competitive advantage

AgentsDGX agent

Enterprises are learning that the race to AI-native operations is not won by deploying the most tools or agents — it is won by the organizations that build the most trusted AI data foundations underne

1 Jun 2026

Auto-Discovery-Bench: Diagnosing Structured State Tracking in Oracle-Guided Discovery

Model ReleasesDGX agent

arXiv:2502.15224v2 Announce Type: replace-cross Abstract: Interactive discovery requires agents to maintain and update structured beliefs over many rounds of feedback. Before evaluating agents in nois

Provably Convergent Actor-Critic for MARL through Risk-aversion

AgentsDGX agent

arXiv:2602.12386v2 Announce Type: replace-cross Abstract: Learning stationary policies in infinite-horizon general-sum Markov games (MGs) remains a fundamental open problem in Multi-Agent Reinforcemen

// Reusable Context Engineering // Context bloat quietly kills long-horizon runs, but you can fix it from the outside without fine-tuning th…

SafetyDGX agent

// Reusable Context Engineering // Context bloat quietly kills long-horizon runs, but you can fix it from the outside without fine-tuning the underlying agent. (bookmark this) Context management is us

slowly we're all realizing that tools should be called from code, not from within the llm api

AgentsDGX agent

slowly we're all realizing that tools should be called from code, not from within the llm api Introducing Search as Code, our new search architecture for AI agents. It writes Python that calls our sea

The Surface You Test Is Not the Surface That Breaks

Model ReleasesDGX agent

arXiv:2605.30454v1 Announce Type: cross Abstract: Tool-augmented LLM agents are vulnerable to prompt injection: a third party who controls part of the agent's context can plant instructions that the a

@Windows More info: https://hermes-agent.nousresearch.com/docs/user-guide/windows-native

AgentsDGX agent

This entry likely covers Nous Research's documentation for Windows native integration or functionality within their Hermes agent system. The resource appears to be a user guide section explaining how

31 May 2026

From reactive operations to autonomous infrastructure: What IT leaders must do next

AgentsDGX agent

As artificial intelligence agents begin to proliferate across information technology infrastructure, IT leaders are moving away from asking, “How do we monitor every alert?” to “How do we design infra

29 May 2026

As always, 'hermes update' More info: https://hermes-agent.nousresearch.com/docs/user-guide/features/tool-search

AgentsDGX agent

Nous Research announced an update to Hermes, their AI agent framework, with details available in their user guide documentation covering tool-search features. The update likely enhances Hermes' capabi

Envy-Free Allocation of Indivisible Goods via Noisy Queries

AgentsDGX agent

arXiv:2602.06361v2 Announce Type: replace-cross Abstract: We introduce a problem of fairly allocating indivisible goods (items) in which the agents' valuations cannot be observed directly, but instead

GUITestScape: Towards Open-set Evaluation on Exploratory GUI Testing

Model ReleasesDGX agent

arXiv:2605.29532v1 Announce Type: cross Abstract: Exploratory GUI testing is a particularly demanding setting for MLLM agents: without predefined test scripts, an agent must autonomously navigate an a

models underestimate how much work it takes (token usage) to accomplish a task, just like us

Model ReleasesDGX agent

models underestimate how much work it takes (token usage) to accomplish a task, just like us 🧵 Claude-Opus-4.8 takes you too much tokens - but is this issue general across agents? Do agents know how m

28 May 2026

Calibrating Conservatism for Scalable Oversight

AgentsDGX agent

arXiv:2605.28807v1 Announce Type: new Abstract: Agentic AI systems capable of autonomous planning and extended environmental interaction pose a fundamental control problem: how can humans maintain mea

slide from

AgentsDGX agent

slide from The Redpoint InfraRed 100 is now live. These are the companies building the infrastructure that powers everything happening in AI right now, from world models and agent runtimes to the sand

TRACER: Turn-level Regret Matching with Inner Reinforcement Credit for Cooperative Multi-LLM Reasoning

Model ReleasesDGX agent

arXiv:2605.28699v1 Announce Type: new Abstract: Large language models increasingly rely on either reinforcement learning or multi-agent prompting to improve reasoning, yet these two paradigms remain d

Uni-LaViRA: Language-Vision-Robot Actions Translation for Unified Embodied Navigation

HardwareDGX agent

arXiv:2605.27582v1 Announce Type: cross Abstract: Embodied navigation requires an agent to map language and visual observations to a stream of spatial actions that drive a real robot through environme

27 May 2026

Decoupled Delay Compensation: Enhancing Pre-trained MARL Policies via Learned Dynamics Filtering

AgentsDGX agent

arXiv:2605.26286v1 Announce Type: cross Abstract: Real-world multi-agent reinforcement learning (MARL) systems must often operate under stale observations, stochastic communication delays, and intermi

The Redpoint InfraRed 100 is now live. These are the companies building the infrastructure that powers everything happening in AI right now,…

AgentsDGX agent

The Redpoint InfraRed 100 is now live. These are the companies building the infrastructure that powers everything happening in AI right now, from world models and agent runtimes to the sandboxes, data

Vital Trace: Protocol-Constrained Patient-State Reasoning for Longitudinal Clinical Trajectories

AgentsDGX agent

arXiv:2602.12833v2 Announce Type: replace-cross Abstract: Longitudinal clinical reasoning over electronic health records requires tracking evolving physiological measurements, laboratory results, and

26 May 2026

Act or Clarify? Modeling Sensitivity to Uncertainty and Cost in Communication

AgentsDGX agent

arXiv:2602.02843v3 Announce Type: replace Abstract: When deciding how to act under uncertainty, agents may choose to act to reduce uncertainty or they may act despite that uncertainty. In communicativ

Attested Tool-Server Admission: A Security Extension to the Model Context Protocol

AgentsDGX agent

arXiv:2605.24248v1 Announce Type: cross Abstract: The Model Context Protocol (MCP) standardizes how a large-language-model (LLM) agent and an external tool server exchange messages, but not trust: a h

Back to Parsimonious Latents: Learning Task-Centric World Models from Visual Foundations

AgentsDGX agent

arXiv:2605.25620v1 Announce Type: new Abstract: World models enable agents to predict future dynamics conditioned on actions, making the choice of latent representation central to planning and control

Evo-Attacker: Memory-Augmented Reinforcement Learning for Long-Horizon Tool Attacks on LLM-MAS

AgentsDGX agent

arXiv:2605.25389v1 Announce Type: cross Abstract: While Large Language Model-based Multi-Agent Systems (LLM-MAS) demonstrate remarkable capabilities in solving complex tasks by orchestrating specializ

haha sweet, gbrain meets activegraph

AgentsDGX agent

haha sweet, gbrain meets activegraph Gbrain is what an agent *knows* — a durable markdown/git knowledge substrate with hybrid search, typed links, facts, timelines, and a real ingestion pipeline. @Act

Hybrid Deep Searcher: Scalable Parallel and Sequential Search Reasoning

AgentsDGX agent

arXiv:2508.19113v3 Announce Type: replace Abstract: Large reasoning models (LRMs) combined with retrieval-augmented generation (RAG) have enabled deep research agents capable of multi-step reasoning w

Partner-Aware Hierarchical Skill Discovery for Robust Human-AI Collaboration

Model ReleasesDGX agent

arXiv:2605.24352v1 Announce Type: new Abstract: Multi-agent collaboration, especially in human-AI teaming, requires agents that can adapt to novel partners with diverse and dynamic behaviors. Conventi

Quantum Frog: Emergent Cooperation and Difficulty Scaling in a Quantized-Time Cooperative Game

SafetyDGX agent

arXiv:2605.23930v1 Announce Type: new Abstract: We introduce Quantum Frog, a two-player cooperative game built on a novel quantized-time mechanic in which the environment advances only when a player a

the importance of traces!

AgentsDGX agent

the importance of traces! Trace data is literally worth its weight in gold these days, if you know what to do with it! As has been established, creating effective agents requires shipping early, obser

Toward Enactive Artificial Intelligence

AgentsDGX agent

arXiv:2605.24238v1 Announce Type: new Abstract: In this paper, we advocate for incorporating enactive approaches to perception and cognition into artificial intelligence (AI). Enactive approaches view

25 May 2026

As token costs shoot up, Dell doubles down on desktop AI

AgentsDGX agent

As enterprises confront spiraling cloud inference bills, on-prem AI computing is emerging as the decisive lever for making agentic AI economically viable at scale. Dell Technologies Inc.’s announcemen

FastKernels: Benchmarking GPU Kernel Generation in Production

Model ReleasesDGX agent

arXiv:2605.23215v1 Announce Type: cross Abstract: LLM-based agents for GPU kernel generation are advancing rapidly, yet their progress is fundamentally constrained by the benchmarks they optimize agai

Latent Cache Flow: Model-to-Model Communication Without Text

AgentsDGX agent

arXiv:2605.22863v1 Announce Type: new Abstract: LLM agents today communicate via text, which incurs considerable latency and information loss due to the need to autoregressively decode the sharer mode

22 May 2026

Always enjoy chatting with @hwchase17 . Excited to sit down with him at our @traversal_ai HQ in a couple of weeks and discuss all things age…

AgentsDGX agent

Always enjoy chatting with @hwchase17 . Excited to sit down with him at our @traversal_ai HQ in a couple of weeks and discuss all things agents. im going to be in NYC in ~1 week, and am doing a firesi

21 May 2026

You can try out the interactive demo at https://agent.datasette.io/

AgentsDGX agent

Simon Willison announced an interactive demo for an AI agent built with Datasette, available at https://agent.datasette.io/. The demo allows users to experience hands-on functionality of the Datasette

20 May 2026

Beyond Majority Voting: LLM Aggregation by Leveraging Higher-Order Information

AgentsDGX agent

arXiv:2510.01499v2 Announce Type: replace-cross Abstract: With the rapid progress of multi-agent large language model (LLM) reasoning, how to effectively aggregate answers from multiple LLMs has emerg

Impetus shifts Leap AI toward enterprise context engineering

AgentsDGX agent

Enterprise AI needs agentic transformation to turn scattered experiments into measurable business outcomes. The next phase of AI adoption is less about picking models and more about making enterprise

Stoked to work closely with @Vtrivedy10 on this. I've spoken about it before but an emerging trend i'm seeing with many of the ai-native com…

AgentsDGX agent

Stoked to work closely with @Vtrivedy10 on this. I've spoken about it before but an emerging trend i'm seeing with many of the ai-native companies I work with (shoutout @larsen_weigle_ ) is a focus on

the first step to building evals is often times figuring out what you want to eval (sounds simple, in practice it has a lot of nuance)

AgentsDGX agent

the first step to building evals is often times figuring out what you want to eval (sounds simple, in practice it has a lot of nuance) you can condense long horizon evals with agents into smaller subs

When Individually Calibrated Models Become Collectively Miscalibrated

AgentsDGX agent

arXiv:2605.18858v1 Announce Type: cross Abstract: Probabilistic prediction systems often aggregate probability estimates from multiple models into a single decision. A common assumption is that if eac

19 May 2026

DexHoldem: Playing Texas Hold'em with Dexterous Embodied System

Model ReleasesDGX agent

arXiv:2605.18727v1 Announce Type: cross Abstract: Evaluating embodied systems on real dexterous hardware requires more than isolated primitive skills: an agent must perceive a changing tabletop scene,

← Previous
1…161162163164165…300
Next →