AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,958 results
Agents

From manual to autonomous: how AI agents are transforming electric grid operations

DGX agent

AI agents are automating traditional manual processes in electric grid operations, improving efficiency and response times for grid management tasks. These autonomous systems leverage machine learning

agentsdatabricks
14 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Agents

if you're already thinking about hacking around on agents this weekend, why not play around with @videodb_io api and potentially win a prize…

DGX agent

if you're already thinking about hacking around on agents this weekend, why not play around with @videodb_io api and potentially win a prize? videodb is an API for turning streaming video into text in

agentsyohei-nakajima--x
14 May 2026
Agents

Introducing LangSmith LLM Gateway: The runtime governance layer for your agents. 💸 Enforce cost limits 🔒 Detect PII ✅ Act on violations …A…

DGX agent

Introducing LangSmith LLM Gateway: The runtime governance layer for your agents. 💸 Enforce cost limits 🔒 Detect PII ✅ Act on violations …All without leaving LangSmith. Now in Private Beta https://www.

agentsharrison-chase--x
14 May 2026
Agents

our engineers had an awesome time attending Interrupt:2026 by @LangChain learning about observability, agent evals, and best practices from …

DGX agent

our engineers had an awesome time attending Interrupt:2026 by @LangChain learning about observability, agent evals, and best practices from the industry leaders! can't wait to see what they cook up wi

agentsharrison-chase--x
14 May 2026
Safety

Quantitative Certification of Agentic Tool Selection

DGX agent

arXiv:2510.03992v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly deployed in agentic systems, where a fundamental task is mapping user intents to relevant extern

safetyarxiv-cs-ai
14 May 2026
Agents

Such a cool hackathon submission for the Hermes Agent @Kimi_Moonshot track!

DGX agent

Such a cool hackathon submission for the Hermes Agent @Kimi_Moonshot track! I mean just look at this beauty! seriously well deserved finalist @evvaaannnn in the @NousResearch x @Kimi_Moonshot creative

agentsnous-research--x
14 May 2026
Agents

suuuuper excited to be collaborating with the excellent LangChain Labs team on this effort prod agent tracing is the seed that lets you clos…

DGX agent

suuuuper excited to be collaborating with the excellent LangChain Labs team on this effort prod agent tracing is the seed that lets you close the loop for continual learning. too much data gets collec

agentsharrison-chase--x
14 May 2026
Model Releases

Systematic Failures in Collective Reasoning under Distributed Information in Multi-Agent LLMs

DGX agent

arXiv:2505.11556v4 Announce Type: replace-cross Abstract: Multi-agent systems built on large language models (LLMs) are expected to enhance decision-making by pooling distributed information, yet syst

model-releasesarxiv-cs-ai
14 May 2026
Agents

There's a ton on unexplored space in building enterprise-grade agent harnesses that continuously improve over time. Congrats to @hwchase17 a…

DGX agent

There's a ton on unexplored space in building enterprise-grade agent harnesses that continuously improve over time. Congrats to @hwchase17 and @LangChain on the launch of LangChain Labs and excited to

agentsharrison-chase--x
14 May 2026
Agents

ToolMol: Evolutionary Agentic Framework for Multi-objective Drug Discovery

DGX agent

arXiv:2605.12784v1 Announce Type: new Abstract: Advances in large language models (LLMs) have recently opened new and promising avenues for small-molecule drug discovery. Yet existing LLM-based approa

agentsarxiv-cs-lg
14 May 2026
Agents

You can now import your project from Lovable, Base44, V0 into @Replit for free. After importing, Replit Agent will build a free mobile app f…

DGX agent

You can now import your project from Lovable, Base44, V0 into @Replit for free. After importing, Replit Agent will build a free mobile app for it and get it onto the App Store in minutes. All free for

agentsreplit--x
14 May 2026
Agents

Your AI agent can create an entire Google Form for you by chatting, and it will use the browser to type and build the survey automatically.

DGX agent

An AI agent can autonomously create a complete Google Form through natural conversation, using browser automation to directly interact with Google Forms' interface by typing and building the survey wi

agentskimi-moonshot--x
14 May 2026
Agents

yst on stream codex wrote a skill for my hermes agent to make music it can take in feedback and improve its output i clipped the best music …

DGX agent

yst on stream codex wrote a skill for my hermes agent to make music it can take in feedback and improve its output i clipped the best music it made with minimax. still far from good music but im prett

agentsnous-research--x
14 May 2026
Agents

4/5 Best-of-N: Run one agent config N times in parallel → select the best trajectory. Leverages LLMs’ non-determinism - but hinges on a good…

DGX agent

4/5 Best-of-N: Run one agent config N times in parallel → select the best trajectory. Leverages LLMs’ non-determinism - but hinges on a good eval mechanism (we use an LLM-as-a-Judge). Can increase acc

agentsai21-labs--x
13 May 2026
Agents

5/5 Ensemble: Run multiple distinct agent configs in parallel → select the best trajectory. Leverages success of the portfolio vs single var…

DGX agent

5/5 Ensemble: Run multiple distinct agent configs in parallel → select the best trajectory. Leverages success of the portfolio vs single variants (see also: @/karpathy’s LLM Council). Can outperform b

agentsai21-labs--x
13 May 2026
Model Releases

An Empirical Study of Automating Agent Evaluation

DGX agent

arXiv:2605.11378v1 Announce Type: new Abstract: Agent evaluation requires assessing complex multi-step behaviors involving tool use and intermediate reasoning, making it costly and expertise-intensive

model-releasesarxiv-cs-cl
13 May 2026
Agents

Deep Reasoning in General Purpose Agents via Structured Meta-Cognition

DGX agent

arXiv:2605.11388v1 Announce Type: new Abstract: Humans intuitively solve complex problems by flexibly shifting among reasoning modes: they plan, execute, revise intermediate goals, resolve ambiguity t

agentsarxiv-cs-cl
13 May 2026
Safety

Do multi-agent systems make LLM reasoning better? Most AI devs assume that it should. But this new paper shows that this is often not the ca…

DGX agent

Do multi-agent systems make LLM reasoning better? Most AI devs assume that it should. But this new paper shows that this is often not the case. It ran 22,500 deterministic trajectories across GAIA, SW

safetydair-ai--x
13 May 2026
Agents

DORA: Dynamic Online Reinforcement Agent for Token Merging in Vision Transformers

DGX agent

arXiv:2605.11683v1 Announce Type: new Abstract: Vision Transformers (ViTs) incur significant computational overhead due to the quadratic complexity of self-attention relative to the token sequence len

agentsarxiv-cs-cv
13 May 2026
Agents

Dynamic Full-body Motion Agent with Object Interaction via Blending Pre-trained Modular Controllers

DGX agent

arXiv:2605.11369v1 Announce Type: new Abstract: Generating physically plausible dynamic motions of human-object interaction (HOI) remains challenging, mainly due to existing HOI datasets limited to st

agentsarxiv-cs-cv
13 May 2026
Model Releases

ExploitGym: Can AI Agents Turn Security Vulnerabilities into Real Attacks?

DGX agent

arXiv:2605.11086v1 Announce Type: cross Abstract: AI agents are rapidly gaining capabilities that could significantly reshape cybersecurity, making rigorous evaluation urgent. A critical capability is

model-releasesarxiv-cs-lg
13 May 2026
Agents

Great example of why you should 1. Run your agent on a separate machine from the sandbox it uses (e.g. sandbox as a tool) 2. Never set env v…

DGX agent

Great example of why you should 1. Run your agent on a separate machine from the sandbox it uses (e.g. sandbox as a tool) 2. Never set env vars in your sandbox. Instead, use something like LangSmith’s

agentsharrison-chase--x
13 May 2026
Agents

Introducing SWE-ZERO-12M-trajectories: the largest agentic trace dataset in the open, 5.7x larger than the previous largest. 112B tokens · 1…

DGX agent

Introducing SWE-ZERO-12M-trajectories: the largest agentic trace dataset in the open, 5.7x larger than the previous largest. 112B tokens · 12M trajectories · 122K PRs · 3K repos · 16 languages https:/

agentsclem-delangue--x
13 May 2026
Agents

🚀Launching: LangSmith Engine LangSmith Engine is an agent that sits on top of your traces It runs in the background and automatically ident…

DGX agent

🚀Launching: LangSmith Engine LangSmith Engine is an agent that sits on top of your traces It runs in the background and automatically identifies issues It then proactively suggests action items (code

agentsharrison-chase--x
13 May 2026
Model Releases

No More, No Less: Task Alignment in Terminal Agents

DGX agent

arXiv:2605.12233v1 Announce Type: new Abstract: Terminal agents are increasingly capable of executing complex, long-horizon tasks autonomously from a single user prompt. To do so, they must interpret

model-releasesarxiv-cs-lg
13 May 2026
Agents

Not to mention 7 blog posts dropped today, including a new Deep Agents version with significant improvements… Especially excited about the C…

DGX agent

Not to mention 7 blog posts dropped today, including a new Deep Agents version with significant improvements… Especially excited about the Code Interpreter feature which is a sneaky powerful feature e

agentsharrison-chase--x
13 May 2026
Agents

Securing AI agents: How AWS and Cisco AI Defense scale MCP and A2A deployments

DGX agent

The Cisco and AWS partnership addresses three challenges enterprises face when scaling AI agents: visibility gaps, security bottlenecks, and compliance risks. In this post, we explore how you can over

agentsaws-ml-blog
13 May 2026
Model Releases

SkillSafetyBench: Evaluating Agent Safety under Skill-Facing Attack Surfaces

DGX agent

arXiv:2605.12015v1 Announce Type: cross Abstract: Reusable skills are becoming a common interface for extending large language model agents, packaging procedural guidance with access to files, tools,

model-releasesarxiv-cs-cl
13 May 2026
Agents

The Hermes Agent Creative Hackathon sponsored by @Kimi_Moonshot has ended! Finalists were selected by Nous and Kimi staff out of 227 submiss…

DGX agent

The Hermes Agent Creative Hackathon sponsored by @Kimi_Moonshot has ended! Finalists were selected by Nous and Kimi staff out of 227 submissions on creativity, usefulness and presentation. We were abs

agentsnous-research--x
13 May 2026
Agents

Yann LeCun says you cannot build a reliable agentic system without a world model LLMs don't have world models. They can't predict the conseq…

DGX agent

Yann LeCun says you cannot build a reliable agentic system without a world model LLMs don't have world models. They can't predict the consequences of their actions before taking them 'they just act, a

agentsyann-lecun--x
13 May 2026
Local Ai

AdaSwitch: Adaptive Switching between Small and Large Agents for Effective Cloud-Local Collaborative Learning

DGX agent

arXiv:2410.13181v2 Announce Type: replace Abstract: Recent advancements in large language models (LLMs) have been remarkable. Users face a choice between using cloud-based LLMs for generation quality

local-aiarxiv-cs-cl
12 May 2026
Hardware

CellDX AI Autopilot: Agent-Guided Training and Deployment of Pathology Classifiers

DGX agent

arXiv:2605.10362v1 Announce Type: new Abstract: Training AI models for computational pathology currently requires access to expensive whole-slide-image datasets, GPU infrastructure, deep expertise in

hardwarearxiv-cs-cv
12 May 2026
Model Releases

Collective Alignment in LLM Multi-Agent Systems: Disentangling Bias from Cooperation via Statistical Physics

DGX agent

arXiv:2605.10528v1 Announce Type: cross Abstract: We investigate the emergent collective dynamics of LLM-based multi-agent systems on a 2D square lattice and present a model-agnostic statistical-physi

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

EnactToM: An Evolving Benchmark for Functional Theory of Mind in Embodied Agents

DGX agent

arXiv:2605.09826v1 Announce Type: new Abstract: Theory of Mind (ToM), the ability to track others epistemic state, makes humans efficient collaborators. AI agents need the same capacity in multi agent

model-releasesarxiv-cs-ai
12 May 2026
Agents

Enhancing Consistency Models for Multi-Agent Trajectory Prediction

DGX agent

arXiv:2605.08572v1 Announce Type: new Abstract: Diffusion models for multi-agent trajectory prediction are limited by iterative denoising, which causes inference latency that hinders their use in time

agentsarxiv-cs-cv
12 May 2026
Model Releases

M2A: Synergizing Mathematical and Agentic Reasoning in Large Language Models

DGX agent

arXiv:2605.09879v1 Announce Type: new Abstract: While reasoning has become a central capability of large language models (LLMs), the reasoning patterns required for different scenarios are often misal

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

MAGS-SLAM: Monocular Multi-Agent Gaussian Splatting SLAM for Geometrically and Photometrically Consistent Reconstruction

DGX agent

arXiv:2605.10760v1 Announce Type: new Abstract: Collaborative photorealistic 3D reconstruction from multiple agents enables rapid large-scale scene capture for virtual production and cooperative multi

model-releasesarxiv-cs-ro
12 May 2026
Model Releases

MDrive: Benchmarking Closed-Loop Cooperative Driving for End-to-End Multi-agent Systems

DGX agent

arXiv:2605.10904v1 Announce Type: new Abstract: Vehicle-to-Everything (V2X) communication has emerged as a promising paradigm for autonomous driving, enabling connected agents to share complementary p

model-releasesarxiv-cs-ro
12 May 2026
Safety

Mem-W: Latent Memory-Native GUI Agents

DGX agent

arXiv:2605.09317v1 Announce Type: new Abstract: GUI agents are beginning to operate the web, mobile, and desktop as interactive worlds, where successful control depends on carrying forward visual, pro

safetyarxiv-cs-cl
12 May 2026
Agents

PiCA: Pivot-Based Credit Assignment for Search Agentic Reinforcement Learning

DGX agent

arXiv:2605.09287v1 Announce Type: new Abstract: Large Language Model (LLM)-based search agents trained with reinforcement learning (RL) have significantly improved the performance of knowledge-intensi

agentsarxiv-cs-ai
12 May 2026
Agents

SAGE: Agentic Framework for Interpretable and Clinically Translatable Computational Pathology Biomarker Discovery

DGX agent

arXiv:2602.00953v2 Announce Type: replace Abstract: Engineered image-based biomarkers offer a clinically interpretable alternative to black-box AI in computational pathology, yet their discovery remai

agentsarxiv-cs-lg
12 May 2026
Local Ai

Scaling Mobile Agent Systems: From Capability Density to Collective Intelligence

DGX agent

arXiv:2605.08124v1 Announce Type: cross Abstract: Mobile agent systems are emerging as a key paradigm for enabling intelligent applications on edge devices and in AIoT ecosystems. However, their scala

local-aiarxiv-cs-cl
12 May 2026
Agents

SkillRAE: Agent Skill-Based Context Compilation for Retrieval-Augmented Execution

DGX agent

arXiv:2605.10114v1 Announce Type: new Abstract: Large Language Model (LLM)-based agents (e.g., OpenClaw) increasingly rely on reusable skill libraries to solve artifact-rich tasks such as document-cen

agentsarxiv-cs-cl
12 May 2026
Agents

The agent also reimplemented the “Blues Improvisation” experiment by @douglas_eck and @SchmidhuberAI in 2002 which show that LSTMs can learn…

DGX agent

The agent also reimplemented the “Blues Improvisation” experiment by @douglas_eck and @SchmidhuberAI in 2002 which show that LSTMs can learn temporal structure in music. Finding temporal structure in

agentsdavid-ha--x
12 May 2026
Agents

the Mini Shai-Hulud attack is scary because it attacks new AI coding workflows like CI, editor hooks, agent configs, etc

DGX agent

The Mini Shai-Hulud attack targets emerging AI-assisted development workflows by compromising multiple integration points including continuous integration systems, code editor hooks, and AI agent conf

agentsyohei-nakajima--x
12 May 2026
Agents

The scale of the infra on HF is insane. If you're still hosting models, datasets, agent memory,... in S3 or R2, talk to use and we can help …

DGX agent

Hugging Face offers substantial infrastructure capabilities for hosting machine learning models, datasets, and agent memory systems. The statement suggests that organizations currently using alternati

agentsclem-delangue--x
12 May 2026
Agents

TMAS: Scaling Test-Time Compute via Multi-Agent Synergy

DGX agent

arXiv:2605.10344v1 Announce Type: new Abstract: Test-time scaling has become an effective paradigm for improving the reasoning ability of large language models by allocating additional computation dur

agentsarxiv-cs-ai
12 May 2026
Agents

🚨 Today: OpenMed Agent ships in preview. Built on @huggingface: → HF endpoints power clinical extraction + terminology → MCP for your own s…

DGX agent

🚨 Today: OpenMed Agent ships in preview. Built on @huggingface: → HF endpoints power clinical extraction + terminology → MCP for your own services → Every tool call, every plan, fully visible 1,000+ O

agentsclem-delangue--x
12 May 2026
← Previous
1…107108109110111…375
Next →