AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,723 results
Agents

Breaking MCP with Function Hijacking Attacks: Novel Threats for Function Calling and Agentic Models

DGX agent

arXiv:2604.20994v1 Announce Type: cross Abstract: The growth of agentic AI has drawn significant attention to function calling Large Language Models (LLMs), which are designed to extend the capabiliti

agentsarxiv-cs-ai
24 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

Hermes Agent v0.11.0 - “The Interface Release” Full changelog below ↓

DGX agent

Hermes Agent v0.11.0, released by Nous Research, is a significant update focused on interface improvements and enhancements. This release, dubbed 'The Interface Release,' includes various updates and

agentsnous-research--x
24 Apr 2026
Agents

Multi-Agent Empowerment and Emergence of Complex Behavior in Groups

DGX agent

arXiv:2604.21155v1 Announce Type: new Abstract: Intrinsic motivations are receiving increasing attention, i.e. behavioral incentives that are not engineered, but emerge from the interaction of an agen

agentsarxiv-cs-ai
24 Apr 2026
Agents

SpatiO: Adaptive Test-Time Orchestration of Vision-Language Agents for Spatial Reasoning

DGX agent

arXiv:2604.21190v1 Announce Type: new Abstract: Understanding visual scenes requires not only recognizing objects but also reasoning about their spatial relationships. Unlike general vision-language t

agentsarxiv-cs-cv
24 Apr 2026
Agents

What is an agent harness?

DGX agent

A version of this article originally appeared on X. Someone asked me at a hacker event last week: “Can anyone actually tell me what a harness really is?” It was... The post What is an agent harness? a

agentsarize-ai
24 Apr 2026
Model Releases

ActuBench: A Multi-Agent LLM Pipeline for Generation and Evaluation of Actuarial Reasoning Tasks

DGX agent

arXiv:2604.20273v1 Announce Type: new Abstract: We present ActuBench, a multi-agent LLM pipeline for the automated generation and evaluation of advanced actuarial assessment items aligned with the Int

model-releasesarxiv-cs-ai
23 Apr 2026
Agents

If you're waiting for a sign... that might not be it! Mitigating Trust Boundary Confusion from Visual Injections on Vision-Language Agentic Systems

DGX agent

arXiv:2604.19844v1 Announce Type: cross Abstract: Recent advances in embodied Vision-Language Agentic Systems (VLAS), powered by large vision-language models (LVLMs), enable AI systems to perceive and

agentsarxiv-cs-ai
23 Apr 2026
Agents

Salesforce and Google partner on agentic cross-platform collaboration

DGX agent

Salesforce Inc. said today it’s trying to break down the siloes that separate enterprise’s customer relationship management data from their productivity tools, specifically for artificial intelligence

agentssiliconangle
23 Apr 2026
Agents

AgentDynEx: Nudging the Mechanics and Dynamics of Multi-Agent Simulations

DGX agent

arXiv:2504.09662v3 Announce Type: replace-cross Abstract: Multi-agent large language model simulations have the potential to model complex human behaviors and interactions. If the mechanics are set up

agentsarxiv-cs-ai
22 Apr 2026
Agents

Best Agent Identification for General Game Playing

DGX agent

arXiv:2507.00451v2 Announce Type: replace-cross Abstract: We present an efficient and generalised procedure to accurately identify the best (or near best) performing algorithm for each sub-task in a m

agentsarxiv-cs-ai
22 Apr 2026
Agents

Google announces innovations in mega-scale networking for the agentic era

DGX agent

Rising to meet demands for increasing latency and scale, Google LLC today introduced a mega-scale datacenter network fabric and cross-cloud infrastructure aimed at agentic artificial intelligence deli

agentssiliconangle
22 Apr 2026
Agents

Introducing Toolboxes in Foundry

DGX agent

Available in Public Preview Today Toolbox is a new way to curate, configure, and reuse tools across all of your AI agents without rewiring them every time from Foundry. Today, teams build agents acros

agentsmicrosoft-foundry
22 Apr 2026
Agents

Refute-or-Promote: An Adversarial Stage-Gated Multi-Agent Review Methodology for High-Precision LLM-Assisted Defect Discovery

DGX agent

arXiv:2604.19049v1 Announce Type: cross Abstract: LLM-assisted defect discovery has a precision crisis: plausible-but-wrong reports overwhelm maintainers and degrade credibility for real findings. We

agentsarxiv-cs-ai
22 Apr 2026
Agents

Superficial Success vs. Internal Breakdown: An Empirical Study of Generalization in Adaptive Multi-Agent Systems

DGX agent

arXiv:2604.18951v1 Announce Type: cross Abstract: Adaptive multi-agent systems (MAS) are increasingly adopted to tackle complex problems.However, the narrow task coverage of their optimization raises

agentsarxiv-cs-cl
22 Apr 2026
Agents

WebUncertainty: Dual-Level Uncertainty Driven Planning and Reasoning For Autonomous Web Agent

DGX agent

arXiv:2604.17821v2 Announce Type: replace Abstract: Recent advancements in large language models (LLMs) have empowered autonomous web agents to execute natural language instructions directly on real-w

agentsarxiv-cs-ai
22 Apr 2026
Agents

Answer Only as Precisely as Justified: Calibrated Claim-Level Specificity Control for Agentic Systems

DGX agent

arXiv:2604.17487v1 Announce Type: new Abstract: Agentic systems often fail not by being entirely wrong, but by being too precise: a response may be generally useful while particular claims exceed what

agentsarxiv-cs-cl
21 Apr 2026
Agents

AutoVQA-G: Self-Improving Agentic Framework for Automated Visual Question Answering and Grounding Annotation

DGX agent

arXiv:2604.17488v1 Announce Type: new Abstract: Manual annotation of high-quality visual question answering with grounding (VQA-G) datasets, which pair visual questions with evidential grounding, is c

agentsarxiv-cs-cv
21 Apr 2026
Hardware

HF becoming the platform for agents (assisted by their humans) to use and build AI (rather than just leveraging APIs)!

DGX agent

HF becoming the platform for agents (assisted by their humans) to use and build AI (rather than just leveraging APIs)! Introducing ml-intern, the agent that just automated the post-training team @hugg

hardwareclem-delangue--x
21 Apr 2026
Agents

Snowflake targets ‘agentic enterprise’ with unified control plane for AI and data

DGX agent

Snowflake Inc. is expanding its push into enterprise artificial intelligence with a set of updates to its Snowflake Intelligence and Cortex Code offerings, positioning its platform as a centralized co

agentssiliconangle
21 Apr 2026
Safety

// Survey on Multi-Agent Systems // The paper traces the landscape from classical paradigms (consensus, distributed control, swarm intellige…

DGX agent

// Survey on Multi-Agent Systems // The paper traces the landscape from classical paradigms (consensus, distributed control, swarm intelligence, cooperative learning) to foundation-model-enabled MAS (

safetydair-ai--x
21 Apr 2026
Agents

Whistant: A Standalone AI Agent for iPhone — No Mac Required

DGX agent

Whistant is an on-phone AI buddy that helps users get things done directly from their iPhone without requiring a Mac. It breaks requests into step-by-step subtasks and executes them on the device, inc

agentsr-ollama
21 Apr 2026
Agents

AgentV-RL: Scaling Reward Modeling with Agentic Verifier

DGX agent

arXiv:2604.16004v1 Announce Type: cross Abstract: Verifiers have been demonstrated to enhance LLM reasoning via test-time scaling (TTS). Yet, they face significant challenges in complex domains. Error

agentsarxiv-cs-ai
20 Apr 2026
Agents

AstroVLM: Expert Multi-agent Collaborative Reasoning for Astronomical Imaging Quality Diagnosis

DGX agent

arXiv:2604.16024v1 Announce Type: cross Abstract: Vision Language Models (VLMs) have been applied to several specific domains and have shown strong problem-solving capabilities. However, astronomical

agentsarxiv-cs-cv
20 Apr 2026
Agents

Explainable Iterative Data Visualisation Refinement via an LLM Agent

DGX agent

arXiv:2604.15319v1 Announce Type: cross Abstract: Exploratory analysis of high-dimensional data relies on embedding the data into a low-dimensional space (typically 2D or 3D), based on which visualiza

agentsarxiv-cs-ai
20 Apr 2026
Agents

'The important thing is not to stop questioning.' — Einstein, 71 years on. Today we celebrate him with II-Agent. His genius was reasoning fr…

DGX agent

'The important thing is not to stop questioning.' — Einstein, 71 years on. Today we celebrate him with II-Agent. His genius was reasoning from first principles, a refusal to take any foundation on fai

agentsemad-mostaque--x
18 Apr 2026
Model Releases

MAS-Bench: A Unified Benchmark for Shortcut-Augmented Hybrid Mobile GUI Agents

DGX agent

arXiv:2509.06477v2 Announce Type: replace Abstract: Shortcuts such as APIs and deep-links have emerged as efficient complements to flexible GUI operations, fostering a promising hybrid paradigm for ML

model-releasesarxiv-cs-ai
17 Apr 2026
Agents

On the Creativity of AI Agents

DGX agent

arXiv:2604.13242v1 Announce Type: cross Abstract: Large language models (LLMs), particularly when integrated into agentic systems, have demonstrated human- and even superhuman-level performance across

agentsarxiv-cs-ai
17 Apr 2026
Agents

Week 4 winner of the Agent 4 Content Challenge: Rizwan Ahmed 🎉 Built: SlapStop, a group accountability app where friends slap each other if…

DGX agent

Rizwan Ahmed won Week 4 of the Agent 4 Content Challenge hosted by Replit with a project called SlapStop, a group accountability app that uses a humorous 'slapping' mechanic among friends to encourage

agentsreplit--x
17 Apr 2026
Agents

yeah, what they said 🤝 a decent harness gets you an actually functioning agent now + teams that actually invest time in their harness+probl…

DGX agent

yeah, what they said 🤝 a decent harness gets you an actually functioning agent now + teams that actually invest time in their harness+problem design, choosing good infra, self-improvement loops, data

agentsharrison-chase--x
17 Apr 2026
Agents

When Less Latent Leads to Better Relay: Information-Preserving Compression for Latent Multi-Agent LLM Collaboration

DGX agent

arXiv:2604.13349v1 Announce Type: new Abstract: Communication in Large Language Model (LLM)-based multi-agent systems is moving beyond discrete tokens to preserve richer context. Recent work such as L

agentsarxiv-cs-lg
16 Apr 2026
Industry

Browser Run: give your agents a browser

DGX agent

Cloudflare's Browser Run is a service that provides AI agents with the ability to control and interact with a real web browser, enabling them to perform tasks such as web scraping, form submission, na

industrycloudflare-ai
15 Apr 2026
Agents

OctoTools: An Agentic Framework with Extensible Tools for Complex Reasoning

DGX agent

arXiv:2502.11271v2 Announce Type: replace-cross Abstract: Solving complex reasoning tasks may involve visual understanding, domain knowledge retrieval, numerical calculation, and multi-step reasoning.

agentsarxiv-cs-cl
15 Apr 2026
Agents

User scoped memory is one of those things that doesn’t matter if you’re building a toy agent for yourself, but when you release at scale you…

DGX agent

User scoped memory is one of those things that doesn’t matter if you’re building a toy agent for yourself, but when you release at scale you gotta get it right Deepagents deploy helps you do that, eas

agentsharrison-chase--x
15 Apr 2026
Safety

AI Organizations are More Effective but Less Aligned than Individual Agents

DGX agent

arXiv:2604.10290v1 Announce Type: new Abstract: AI is increasingly deployed in multi-agent systems; however, most research considers only the behavior of individual models. We experimentally show that

safetyarxiv-cs-ai
14 Apr 2026
Agents

For an agent builder, memory is sustained advantage For the model provider, memory is switching cost

DGX agent

Harrison Chase's post explores the divergent strategic incentives around memory in AI agent systems, arguing that for those building agent frameworks, persistent memory creates compounding value and c

agentsharrison-chase--x
14 Apr 2026
Agents

Help Without Being Asked: A Deployed Proactive Agent System for On-Call Support with Continuous Self-Improvement

DGX agent

arXiv:2604.09579v1 Announce Type: new Abstract: In large-scale cloud service platforms, thousands of customer tickets are generated daily and are typically handled through on-call dialogues. This high

agentsarxiv-cs-ai
14 Apr 2026
Model Releases

Open Harness 🤝 Deployed Agents if you wanna use Claude, GLM5, and Codex in your deployed harness then you should be able to! deepagents dep…

DGX agent

Open Harness 🤝 Deployed Agents if you wanna use Claude, GLM5, and Codex in your deployed harness then you should be able to! deepagents deploy has easy configs to let users customize their harness and

model-releasesharrison-chase--x
14 Apr 2026
Model Releases

STARS: Skill-Triggered Audit for Request-Conditioned Invocation Safety in Agent Systems

DGX agent

arXiv:2604.10286v1 Announce Type: new Abstract: Autonomous language-model agents increasingly rely on installable skills and tools to complete user tasks. Static skill auditing can expose capability s

model-releasesarxiv-cs-ai
14 Apr 2026
Agents

SWE-AGILE: A Software Agent Framework for Efficiently Managing Dynamic Reasoning Context

DGX agent

arXiv:2604.11716v1 Announce Type: new Abstract: Prior representative ReAct-style approaches in autonomous Software Engineering (SWE) typically lack the explicit System-2 reasoning required for deep an

agentsarxiv-cs-ai
14 Apr 2026
Agents

We see this as further validation that multi-agent architectures excel at novel problems outside training data distribution. These technique…

DGX agent

Cursor AI shared observations on X validating that multi-agent architectures demonstrate superior performance when tackling novel problems that fall outside the distribution of training data. The post

agentscursor--x
14 Apr 2026
Agents

CodeScout: Contextual Problem Statement Enhancement for Software Agents

DGX agent

arXiv:2603.05744v2 Announce Type: replace Abstract: Current AI-powered code assistance tools often struggle with poorly-defined problem statements that lack sufficient task context and requirements sp

agentsarxiv-cs-cl
13 Apr 2026
Agents

Gen-n-Val: Agentic Image Data Generation and Validation

DGX agent

arXiv:2506.04676v2 Announce Type: replace-cross Abstract: The data scarcity, label noise, and long-tailed category imbalance remain important and unresolved challenges in many computer vision tasks, s

agentsarxiv-cs-ai
13 Apr 2026
Agents

🔒 new in deepagents: filesystem permissions shared resources and org-wide policies are exactly the kind of files you want your agent to rea…

DGX agent

🔒 new in deepagents: filesystem permissions shared resources and org-wide policies are exactly the kind of files you want your agent to read but never overwrite. filesystem permissions let you enforce

agentsharrison-chase--x
13 Apr 2026
Agents

Sustained Impact of Agentic Personalisation in Marketing: A Longitudinal Case Study

DGX agent

arXiv:2604.08621v1 Announce Type: new Abstract: In consumer applications, Customer Relationship Management (CRM) has traditionally relied on the manual optimisation of static, rule-based messaging str

agentsarxiv-cs-ai
13 Apr 2026
Safety

Wireless Communication Enhanced Value Decomposition for Multi-Agent Reinforcement Learning

DGX agent

arXiv:2604.08728v1 Announce Type: new Abstract: Cooperation in multi-agent reinforcement learning (MARL) benefits from inter-agent communication, yet most approaches assume idealized channels and exis

safetyarxiv-cs-lg
13 Apr 2026
Agents

100% this. Memory is just a form of context. Context is all you need, and context is everything when making a great agent.

DGX agent

Harrison Chase, co-founder of LangChain, endorsed the perspective that memory in AI agents is fundamentally a form of context, emphasizing that effective context management is the core requirement for

agentsharrison-chase--x
12 Apr 2026
Agents

New in Hermes Agent, /compress <topic> to get the compaction model to retain more information on the topic you want it to keep in memory mos…

DGX agent

Nous Research has introduced a new feature in their Hermes Agent system that allows users to use the `/compress ` command to influence how the compaction model prioritizes and retains information. Thi

agentsnous-research--x
12 Apr 2026
Agents

OPEN MEMORY, OPEN HARNESS the industry shifts to closed agent harnesses locking memory behind proprietary apis. https://x.com/hwchase17/stat…

DGX agent

Harrison Chase, co-founder of LangChain, discusses a trend in the AI industry where agent harnesses are becoming increasingly closed and proprietary, with memory systems locked behind vendor-specific

agentsharrison-chase--x
12 Apr 2026
← Previous
1…3536373839…370
Next →