Agents that remember: introducing Agent Memory
Agent Memory is a newly announced managed service designed to address 'context rot'—the issue of degraded performance in sophisticated AI agents due to the limitations of context window size. This ser
Knowledge catalogue
Agent Memory is a newly announced managed service designed to address 'context rot'—the issue of degraded performance in sophisticated AI agents due to the limitations of context window size. This ser
The Agent Readiness score can help site owners understand how well their websites support AI agents. Here we explore new standards, share Radar data, and detail how we made Cloudflare’s docs the most
You can now visualize Pi traces that you upload on @huggingface! Let's make sharing agent traces 10x more common to make agent AI more open and collaborative! Also, because it's fun to analyze @badlog
arXiv:2604.13107v1 Announce Type: cross Abstract: As coding agents have seen rapid capability and adoption gains, users are applying them to general tasks beyond software engineering. In this post, we
when you take agents to production, you need to think about guardrails we provide 2 abstractions for guardrails 1. middleware provides hooks around the agent loop that you can use to handle retries, e
Hermes Agent users are remarkably model-curious. Of the top 10 models used in Hermes Agent, 8 different model companies are represented, and 6 of the top 7 are open source. Without lock-in, people try
If you don't own the memory, you don't own the agent: - memory is what makes your agent get smarter over time - without it, anyone with the same tools can copy your agent overnight - with it, you buil
Deep Agents deploy gets you the best AI infrastructure operators in the game! cc @h_dhanushkodi @WHinthorn Deep Agents deploy gets you: - Deep Agents harness - Sandbox of your choice (@daytonaio , @mo
im excited about agent harnesses because i think are the first stable agent abstractions we can build on top (which is why we're investing so much in deepagents) we always wanted to run llms in a loop
Multi-Agent Orchestration is a great tool in building agentic products thankfully @sydneyrunkle put together a guide with runnable code for each pattern earlier this year 🙏 (link below) choose any mod
More organizations are using natural language to query data instead of writing manual SQL. But moving an AI agent from a prototype to a production-ready tool requires rigorous, repeatable testing. Pri
A developer open-sourced an agent architecture specifically designed for long-horizon tasks, positioning it as addressing gaps in existing tools. Manus is a cloud-based autonomous AI agent built f...
For organizations deploying AI agents at scale, there’s often a critical divide between structured and unstructured data. While large language models (LLMs) excel at parsing text documents, emails, an
Allie K. Miller announced a free live AI‑agent workshop featuring Mark Cuban set for August 11 2026, targeting up to 20,000 participants. She will demonstrate how her own AI workforce of 34 agents wor
As a market leader in enterprise agentic automation and business orchestration, UiPath is helping to pioneer an industry shift toward agentic AI. With it, the company is deploying autonomous agents to
arXiv:2602.10429v2 Announce Type: replace-cross Abstract: AIvilization v0 is a publicly deployed large-scale artificial society that couples a resource-constrained sandbox with a unified LLM-agent arc
arXiv:2511.22226v2 Announce Type: replace Abstract: The standard theory of model-free reinforcement learning assumes that the environment dynamics are stationary and that agents are decoupled from the
arXiv:2607.26935v1 Announce Type: new Abstract: Bot detectors deployed at scale treat traffic as binary: human or bot. This assumption breaks when AI agents browse the web through browser automation,
arXiv:2607.26637v1 Announce Type: new Abstract: Deployed LLM agents increasingly keep their long-term memory as a filesystem: a directory tree of markdown files that the agent itself reads, writes, an
arXiv:2510.04452v3 Announce Type: replace-cross Abstract: Computer use agents (or 'agents') are generative AI that automates actions within user interfaces from user commands. Current research focuses
New research from Meta and CMU. This one is on agentic context management for long horizon tasks. (bookmark it) Production agents accumulate context every turn. The usual fix compresses on a token thr
arXiv:2607.23586v1 Announce Type: new Abstract: Long-lived AI agents increasingly evolve after deployment by retaining experience, acquiring skills and tools, revising workflows, delegating work, and
Great paper on self-improving agent harnesses. (bookmark it) If you maintain a production agent harness, finding every file behind one behavior is often harder than writing the edit. Harness Handbook
arXiv:2607.13157v1 Announce Type: new Abstract: Agent memory is a systems problem for long-horizon agents. Practical deployments require retention of task state across extended conversations, recovery
Editor’s note: Today we hear from IDC on the results of its 2026 AI in Networking Special Report Survey exploring the enterprises' concerns about networking infrastructure to support the rise of agent
arXiv:2607.10526v2 Announce Type: replace Abstract: Stateful personal agents increasingly maintain long-term user profiles, episodic memories, and reusable skills. This persistence turns conversationa
arXiv:2606.31518v1 Announce Type: new Abstract: Agentic Business Process Management has gained momentum recently. The prospect is that the autonomy of AI agents, i.e., predominantly LLM-based agents,
As enterprise generative AI transitions from simple, conversational chatbots to autonomous multi-agent workflows, developers face a critical bottleneck: scale. In a production environment, an enterpri
arXiv:2606.11869v1 Announce Type: cross Abstract: Custom AI agents areagents that live inside their own application, talk to their own data and tools, enforce their own security boundaries, and carry
Token costs are why there will be no saas apocalypse / good dev tools are cached intelligence for agents! The popular theory goes: agents can write code, so they'll just rebuild every tool from scratc
arXiv:2602.22480v2 Announce Type: replace-cross Abstract: An important emerging application of coding agents is agent optimization: the iterative improvement of a target agent through edit-execute-eva
MY HERMES AGENT ATE ACID AND HIJACKED MY TOUCHDESIGNER INSTANCE. The new TouchDesigner skill in @NousResearch Hermes Agent is wild. It turns Hermes into a creative operator for abstract visuals, motio
arXiv:2510.16853v3 Announce Type: replace-cross Abstract: Autonomous AI agents capable of complex planning and action mark a shift beyond today's generative tools. As these systems enter political and
Great paper on improving proactive agents. (bookmark it) Proactive agents act before you do. But how do you evaluate something that's supposed to anticipate needs you haven't expressed? This work intr
arXiv:2604.20279v1 Announce Type: cross Abstract: Mobile GUI agents can automate smartphone tasks by interacting directly with app interfaces, but how they should communicate with users during executi
arXiv:2604.20070v1 Announce Type: cross Abstract: Advances in AI agent capabilities have outpaced users' ability to meaningfully oversee their execution. AI agents can perform sophisticated, multi-ste
arXiv:2604.19925v1 Announce Type: cross Abstract: AI agents powered by large language models are increasingly acting on behalf of humans in social and economic environments. Prior research has focused
Meet Kimi K2.6 Agent Swarm 👋 Highlights: 🔹 Swarms, elevated - 300 parallel sub-agents × 4,000 steps per run (up from 100 / 1,500 in K2.5). 🔹 Outputs are real files, not chat - one run delivers 100+ fi
arXiv:2511.03690v2 Announce Type: replace-cross Abstract: Agents are now used widely in the process of software development, but building production-ready software engineering agents is a complex task
arXiv:2510.21652v2 Announce Type: replace Abstract: AI agents hold the potential to revolutionize scientific productivity by automating literature reviews, replicating experiments, analyzing data, and
we want to help you navigate all the stuff that comes with shipping an agent to production - so we wrote a guide! this guide packages learnings across our + customer deployments, building open source
Modal's blog post covers how to build and deploy AI agent applications by combining Modal's serverless cloud infrastructure with OpenAI's Agent SDK, enabling scalable execution of agentic workflows. T
arXiv:2604.09746v1 Announce Type: cross Abstract: As large language models (LLMs) are increasingly deployed as autonomous agents, understanding how strategic behavior emerges in multi-agent environmen
arXiv:2604.09744v1 Announce Type: cross Abstract: The AI agent ecosystem has converged on two protocols: the Model Context Protocol (MCP) for tool invocation and Agent-to-Agent (A2A) for single-princi
Most people building AI agents don't realize this until it's too late. Your agent's memory isn't a feature. It's your moat. 1. Closed harness = they own your memory 2. Switch models → lose all context
without memory, ux with an agent is really bad memory is what makes an experience with an agent feel personal and optimized i'm doing a bunch of research re how folks are building memory into their ag
arXiv:2608.10218v1 Announce Type: new Abstract: AI agents are becoming more autonomous and increasingly interconnected, exposing them to new emergent risks arising from agent-to-agent interaction. One
You don’t need new ways to talk to your agents, you need new ways for your agents to talk to you 🫵 (do you?) Introducing Remoko: your mobile agent relay http://remoko.app I wanted a way for my long-ru
Agents want to collaborate. So we’re putting that instinct to work on improving open-weight LLMs at formal math. Our latest agent collab tackles the @SAIRfoundation challenge of building a cheat sheet
The Top AI Papers of the Week (June 21 - June 28) - Autodata - Sakana Fugu - Agent-as-a-Router - Agent-Native Memory - A Pinch of Human Data - Critique of the Agent Model - Agent Communication Protoco
@LangChain always reveals the technical principles behind agent concepts, like loop engineering, managed agents and memories.👍 That's why I recommend LC docs to agent developers and users around me, t
// Critique of the Agent Model // Finally, a paper that tries to define what an agent is and what agency consists of. Good read overall. (great bookmark) The word agent now covers everything from a fo
With hosted agents, we made it straightforward to build and deploy agents on Foundry. You write your logic, run azd deploy, and your agent is live. But “live” and “production-ready” aren’t the same th
arXiv:2504.05871v3 Announce Type: replace Abstract: The increasing deployment of intelligent agents in digital ecosystems, such as social media platforms, has raised significant concerns about traceab
arXiv:2605.28787v1 Announce Type: cross Abstract: In the era of autonomous agents, machine-actionable data is critical for data-driven workflows. For more than a decade, semantic metadata like schema.
Managed deep agents is the easiest way to build and deploy long horizon agents Private preview, dm me if you want access Managed Deep Agents is built for agents that need to work over long time horizo
arXiv:2602.03695v2 Announce Type: replace-cross Abstract: While existing multi-agent systems (MAS) can handle complex problems by enabling collaboration among multiple agents, they are often highly ta
arXiv:2605.22905v1 Announce Type: new Abstract: Self-evolving agents should not train on examples they cannot justify. Data-free self-evolving search agents offer a scalable route to systems that gene
Agent observability is a means to an end: making your agent better. But observability and evals tools have traditionally failed to connect traces to meaningful actions. Agent engineering teams are lef
arXiv:2605.03213v1 Announce Type: cross Abstract: Agentic AI systems, specifically LLM-driven agents that plan, invoke tools, maintain persistent memory, and delegate tasks to peer agents via protocol