AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,980 results
Model Releases

When Should a Robot Think? Resource-Aware Reasoning via Reinforcement Learning for Embodied Robotic Decision-Making

DGX agent

arXiv:2603.16673v4 Announce Type: replace-cross Abstract: Embodied robotic systems increasingly rely on large language model (LLM)-based agents to support high-level reasoning, planning, and decision-

model-releasesarxiv-cs-ai
29 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Agents

I'v never bothered to make a @Shopify store in my life. I ran into a few potential clients who use it so decided to make a shopify ecommerce…

DGX agent

I'v never bothered to make a @Shopify store in my life. I ran into a few potential clients who use it so decided to make a shopify ecommerce site with it. I had @NousResearch hermes-agent create desig

agentsnous-research--x
28 May 2026
Agents

Reasoning and Planning with Dynamically Changing Norms

DGX agent

arXiv:2605.27622v1 Announce Type: new Abstract: To safely interact with humans, AI agents must both know our norms and consider them during planning. However, such norm-guided planning has been less e

agentsarxiv-cs-ai
28 May 2026
Agents

Rethinking Memory as Continuously Evolving Connectivity

DGX agent

arXiv:2605.28773v1 Announce Type: cross Abstract: Existing memory-augmented LLM agents often treat memory as a static repository with pre-defined representations and fixed retrieval pipelines, which i

agentsarxiv-cs-ai
28 May 2026
Agents

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation

DGX agent

arXiv:2605.26500v1 Announce Type: new Abstract: Vision-language navigation (VLN) requires an agent to traverse complex 3D environments based on natural language instructions, necessitating a thorough

agentsarxiv-cs-cv
27 May 2026
Agents

https://hermes-agent.nousresearch.com/docs/user-guide/features/mcp#catalog-one-click-install-for-nous-approved-mcps

DGX agent

Nous Research introduced a one-click installation feature in Hermes Agent that allows users to easily install Model Context Protocol (MCP) servers from a curated catalog of Nous-approved integrations.

agentsnous-research--x
27 May 2026
Model Releases

Introducing Google AI Threat Defense to help you outpace the adversary

DGX agent

aside_block <ListValue: [StructValue([('title', 'Summary of today’s news'), ('body', <wagtail.rich_text.RichText object at 0x7fb0f516f910>), ('btn_text', ''), ('href', ''), ('image', None)])]> AI-powe

model-releasesgoogle-cloud-ai
27 May 2026
Model Releases

Sentinel: Embodied Cooperative Spatial Reasoning and Planning

DGX agent

arXiv:2605.26239v1 Announce Type: new Abstract: In this work, we study Cooperative Spatial Intelligence, the ability of decentralized embodied agents to coordinate effectively under dynamic environmen

model-releasesarxiv-cs-cv
27 May 2026
Hardware

SIA: Self Improving AI with Harness & Weight Updates

DGX agent

arXiv:2605.27276v1 Announce Type: new Abstract: Humans are the bottleneck in building and improving AI. Both the models and the agents that wrap them are written, tuned, and corrected by people. The l

hardwarearxiv-cs-ai
27 May 2026
Tutorials

A perspective on fluid mechanical environments for challenges in reinforcement learning

DGX agent

arXiv:2605.25011v1 Announce Type: new Abstract: We consider the challenge of developing agents that efficiently interact with high-dimensional, evolving environments, towards a view of practical reinf

tutorialsarxiv-cs-lg
26 May 2026
Model Releases

CausaLab: A Scalable Environment for Interactive Causal Discovery Toward AI Scientists

DGX agent

arXiv:2605.26029v1 Announce Type: new Abstract: We introduce CausaLab, a scalable environment for evaluating interactive causal discovery by LLM agents. Unlike prior evaluations, CausaLab evaluates bo

model-releasesarxiv-cs-ai
26 May 2026
Agents

MCPXKIT: The Unified Toolkit for Analyzing Model Context Protocol Security

DGX agent

arXiv:2508.12538v2 Announce Type: replace-cross Abstract: The Model Context Protocol (MCP) has emerged as a universal standard that enables AI agents to seamlessly connect with external tools, signifi

agentsarxiv-cs-ai
26 May 2026
Agents

Microsoft Copilot Cowork Exfiltrates Files

DGX agent

Microsoft Copilot Cowork Exfiltrates Files The biggest challenge in designing agentic systems continues to be preventing them from enabling attackers to exfiltrate data. In this case Microsoft Copilot

agentssimon-willison
26 May 2026
Agents

Reward Shaping and Action Masking for Compositional Tasks using Behavior Trees and LLMs

DGX agent

arXiv:2605.05795v2 Announce Type: replace Abstract: Decomposing complex tasks into a sequence of simpler subtasks can improve learning efficiency for an autonomous agent. Reinforcement learning (RL) c

agentsarxiv-cs-lg
26 May 2026
Model Releases

SMDD-Bench: Can LLMs Solve Real-World Small Molecule Drug Design Tasks?

DGX agent

arXiv:2605.21740v2 Announce Type: replace Abstract: LLM agents have incredible potential for scientific discovery applications. However, the performance of LLM agents on real-world, small molecule dru

model-releasesarxiv-cs-ai
26 May 2026
Agents

Trace data is literally worth its weight in gold these days, if you know what to do with it! As has been established, creating effective age…

DGX agent

Trace data is literally worth its weight in gold these days, if you know what to do with it! As has been established, creating effective agents requires shipping early, observing behavior, and iterati

agentsharrison-chase--x
26 May 2026
Agents

Socially fluent AI decouples conversational signals from source identity in online interaction

DGX agent

arXiv:2605.23426v1 Announce Type: cross Abstract: Socially fluent agentic AI can now participate in online interaction in ways that resemble ordinary human conversation, potentially weakening people's

agentsarxiv-cs-ai
25 May 2026
Agents

// Adapt the Interface, Not the Model // I am fascinated by the results across my cheap-model-plus-good-harness builds. This new paper also …

DGX agent

// Adapt the Interface, Not the Model // I am fascinated by the results across my cheap-model-plus-good-harness builds. This new paper also shows good signs of the code-as-agent-harness thesis. The id

agentsdair-ai--x
23 May 2026
Agents

“It is built in Rust and leverages the Apache DataFusion query engine” Any new database these days

DGX agent

“It is built in Rust and leverages the Apache DataFusion query engine” Any new database these days We built SmithDB: the database purpose built for agent observability workloads that now powers many p

agentsharrison-chase--x
23 May 2026
Agents

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation

DGX agent

arXiv:2605.22816v1 Announce Type: cross Abstract: Vision-and-Language Navigation (VLN) requires an agent to ground language instructions to its own movement within a visual environment. While state-of

agentsarxiv-cs-cv
22 May 2026
Safety

Diagnosis Is Not Prescription: Linguistic Co-Adaptation Explains Patching Hazards in LLM Pipelines

DGX agent

arXiv:2605.21958v1 Announce Type: new Abstract: When a multi-module LLM agent fails, the module most responsible for the failure is not necessarily the best place to intervene. We demonstrate this Dia

safetyarxiv-cs-cl
22 May 2026
Agents

ProCrit: Self-Elicited Multi-Perspective Reasoning with Critic-Guided Revision for Multimodal Sarcasm Detection

DGX agent

arXiv:2605.20867v1 Announce Type: cross Abstract: Multimodal sarcasm detection requires reasoning over cross-modal incongruities between literal expression and intended meaning, yet the specific analy

agentsarxiv-cs-cv
21 May 2026
Model Releases

How Far Are We From True Auto-Research?

DGX agent

arXiv:2605.19156v1 Announce Type: new Abstract: Recent auto-research systems can produce complete papers, but feasibility is not the same as quality, and the field still lacks a systematic study of ho

model-releasesarxiv-cs-ai
20 May 2026
Agents

very belated but in retrospect i think @sama's mythical 'build a business that gets better when models get better' is basically what I calle…

DGX agent

very belated but in retrospect i think @sama's mythical 'build a business that gets better when models get better' is basically what I called Agent Labs here. seeing a very direct correlation with mod

agentsswyx--x
20 May 2026
Agents

Very cool - Grok Build is clearly getting better by the day. Two nights ago I ran an overnight build and it failed. Last night, success. I r…

DGX agent

Very cool - Grok Build is clearly getting better by the day. Two nights ago I ran an overnight build and it failed. Last night, success. I really like the multi-agent orchestration behavior, it does a

agentselon-musk--x
20 May 2026
Model Releases

A Machine with Short-Term, Episodic, and Semantic Memory Systems

DGX agent

arXiv:2212.02098v5 Announce Type: replace Abstract: Inspired by the cognitive science theory of the explicit human memory systems, we have modeled an agent with short-term, episodic, and semantic memo

model-releasesarxiv-cs-ai
19 May 2026
Agents

Convergence of Multiagent Learning Systems for Traffic control

DGX agent

arXiv:2511.11654v2 Announce Type: replace-cross Abstract: Rapid urbanization in cities like Bangalore has led to severe traffic congestion, making efficient Traffic Signal Control (TSC) essential. Mul

agentsarxiv-cs-ai
19 May 2026
Model Releases

I've been thinking a lot about the two different groups of evals you need in general agents/agents which handle broad tasks: 1. Benchmark ev…

DGX agent

I've been thinking a lot about the two different groups of evals you need in general agents/agents which handle broad tasks: 1. Benchmark evals - this is a suite of up to 100 eval cases which test the

model-releasesharrison-chase--x
19 May 2026
Agents

Learning to Learn from Multimodal Experience

DGX agent

arXiv:2605.16857v1 Announce Type: new Abstract: Experience-driven learning has emerged as a promising paradigm for enabling agents to improve from interaction trajectories by accumulating and reusing

agentsarxiv-cs-ai
19 May 2026
Agents

PersonaArena: Dynamic Simulation for Evaluating and Enhancing Persona-Level Role-Playing in Large Language Models

DGX agent

arXiv:2605.17044v1 Announce Type: new Abstract: Large language models (LLMs) increasingly serve as interactive social agents, yet their ability to maintain coherent and authentic persona-level role-pl

agentsarxiv-cs-ai
19 May 2026
Agents

Task-Level AI Readiness Assessment for Business Process Management:The T-IPO Model and LARA Matrix in Financial-Services IT Operations

DGX agent

arXiv:2605.16297v1 Announce Type: cross Abstract: Which tasks inside an enterprise workflow can a large-language-model agent reliably handle, and under what conditions? Most business process modeling

agentsarxiv-cs-ai
19 May 2026
Agents

ICRL: Learning to Internalize Self-Critique with Reinforcement Learning

DGX agent

arXiv:2605.15224v1 Announce Type: new Abstract: Large language model-based agents make mistakes, yet critique can often guide the same model toward correct behavior. However, when critique is removed,

agentsarxiv-cs-ai
18 May 2026
Agents

Mecha-nudges for Machines

DGX agent

arXiv:2603.23433v2 Announce Type: replace Abstract: AI agents are becoming active decision-makers on the Internet. As they make decisions in the same environments as humans, the environments themselve

agentsarxiv-cs-ai
18 May 2026
Agents

Optimized Three-Dimensional Photovoltaic Structures with LLM guided Tree Search

DGX agent

arXiv:2605.16191v1 Announce Type: new Abstract: We present a case study for how AI coding systems can be used to generate novel scientific hypotheses. We combine a generic coding agent (Google's AntiG

agentsarxiv-cs-cl
18 May 2026
Agents

Distributionally Robust Multi-Task Reinforcement Learning via Adaptive Task Sampling

DGX agent

arXiv:2605.14350v1 Announce Type: new Abstract: Multi-task reinforcement learning (MTRL) aims to train a single agent to efficiently optimize performance across multiple tasks simultaneously. However,

agentsarxiv-cs-lg
15 May 2026
Agents

Cognifold: Always-On Proactive Memory via Cognitive Folding

DGX agent

arXiv:2605.13438v1 Announce Type: new Abstract: Existing agent memory remains predominantly reactive and retrieval-based, lacking the capacity to autonomously organize experience into persistent cogni

agentsarxiv-cs-ai
14 May 2026
Agents

Scaling Retrieval-Augmented Reasoning with Parallel Search and Explicit Merging

DGX agent

arXiv:2605.13534v1 Announce Type: new Abstract: Deep search agents have proven effective in enhancing LLMs by retrieving external knowledge during multi-step reasoning. However, existing methods often

agentsarxiv-cs-ai
14 May 2026
Agents

SmithDB is a feat of engineering. A new database for a new data shape Huge work from @ankush_gola11 and team

DGX agent

SmithDB is a feat of engineering. A new database for a new data shape Huge work from @ankush_gola11 and team We built SmithDB: the database purpose built for agent observability workloads that now pow

agentsharrison-chase--x
14 May 2026
Agents

Encore: Conditioning Trajectory Forecasting via Biased Ego Rehearsals

DGX agent

arXiv:2605.11463v1 Announce Type: new Abstract: Learning and representing the subjectivities of agents has become a challenging but crucial problem in the trajectory prediction task. Such subjectiviti

agentsarxiv-cs-cv
13 May 2026
Agents

Finite and Corruption-Robust Regret Bounds in Online Inverse Linear Optimization under M-Convex Action Sets

DGX agent

arXiv:2602.01682v2 Announce Type: replace Abstract: We study online inverse linear optimization, also known as contextual recommendation, where a learner sequentially infers an agent's hidden objectiv

agentsarxiv-cs-lg
13 May 2026
Agents

Shaping Zero-Shot Coordination via State Blocking

DGX agent

arXiv:2605.11688v1 Announce Type: new Abstract: Zero-shot coordination (ZSC) aims to enable agents to cooperate with independently trained partners without prior interaction, a key requirement for rea

agentsarxiv-cs-lg
13 May 2026
Agents

Controllability in preference-conditioned multi-objective reinforcement learning

DGX agent

arXiv:2605.10585v1 Announce Type: new Abstract: Multi-objective reinforcement learning (MORL) allows a user to express preference over outcomes in terms of the relative importance of the objectives, b

agentsarxiv-cs-lg
12 May 2026
Agents

love seeing my discord stay in sync with Hermes Kanban..everything here was done in plain English. I just asked my coordinator to check if w…

DGX agent

love seeing my discord stay in sync with Hermes Kanban..everything here was done in plain English. I just asked my coordinator to check if we have an update and the system understood the intent, route

agentsnous-research--x
12 May 2026
Agents

Robust Remote Reinforcement Learning over Unreliable Communication Channels using Homomorphic State Encoding

DGX agent

arXiv:2508.07722v2 Announce Type: replace Abstract: Traditional Reinforcement Learning (RL) frameworks generally assume that the agent perceives the state of the underlying Markov process instantaneou

agentsarxiv-cs-lg
12 May 2026
Agents

Zero-shot Imitation Learning by Latent Topology Mapping

DGX agent

arXiv:2605.08450v1 Announce Type: cross Abstract: Imitation learning is effective for training agents when expert demonstrations are available, but collecting demonstrations for every complex task in

agentsarxiv-cs-ai
12 May 2026
Agents

A Differentiable Bayesian Relaxation for Latent Partial-Order Inference

DGX agent

arXiv:2605.06976v1 Announce Type: cross Abstract: Many ranking and agent trace datasets are recorded as linear orders even though their latent structure is only partially ordered. This is especially c

agentsarxiv-cs-lg
11 May 2026
Agents

One World, Dual Timeline: Decoupled Spatio-Temporal Gaussian Scene Graph for 4D Cooperative Driving Reconstruction

DGX agent

arXiv:2605.07910v1 Announce Type: new Abstract: Reconstructing dynamic scenes from Vehicle-to-Infrastructure Cooperative Autonomous Driving (VICAD) data is fundamentally complicated by temporal asynch

agentsarxiv-cs-cv
11 May 2026
Agents

Quoting James Shore

DGX agent

Your AI coding agent, the one you use to write code, needs to reduce your maintenance costs. Not by a little bit, either. You write code twice as quick now? Better hope you’ve halved your maintenance

agentssimon-willison
11 May 2026
← Previous
1…196197198199200…375
Next →