AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,919 results
Agents

FORGE: Research-Trajectory Hijacking Attacks on Deep Research Agents

DGX agent

arXiv:2607.04718v1 Announce Type: new Abstract: Deep research agents decompose open-ended queries into subtasks, retrieve web evidence over multiple rounds, and synthesize long-form reports. This work

agentsarxiv-cs-ai
7 Jul 2026
Agents
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

kAgent: An execution-guided crash resolution agent for the Linux kernel

DGX agent

arXiv:2504.20412v3 Announce Type: replace-cross Abstract: Fuzzing frameworks like syzkaller have uncovered thousands of Linux kernel crashes, many of which are critical and security-sensitive. However

agentsarxiv-cs-ai
7 Jul 2026
Hardware

NVIDIA Vera CPU Boosts AI Factory Throughput to Accelerate Agentic Workloads

DGX agent

NVIDIA Vera is a CPU designed to help AI factories scale agentic AI and reinforcement learning by shortening CPU execution time, increasing task throughput, and enabling smarter, longer-thinking agent

hardwarenvidia-developer
7 Jul 2026
Safety

OpenTinker: Separating Concerns in Agentic Reinforcement Learning

DGX agent

arXiv:2601.07376v2 Announce Type: replace Abstract: We introduce extsc{OpenTinker}, an open infrastructure for training large language model (LLM) agents with many LoRA-backed policies over shared exe

safetyarxiv-cs-ai
7 Jul 2026
Model Releases

ProACT: Towards Breakdown-Aware Proactive Agent in Multi-User Collaboration

DGX agent

arXiv:2607.03730v1 Announce Type: new Abstract: Conversational agents are increasingly embedded in human collaborative work, yet they remain fundamentally passive and reactive: they respond to explici

model-releasesarxiv-cs-cl
7 Jul 2026
Local Ai

SwarmResearch: Orchestrating Coding Agents for Open-Ended Discovery

DGX agent

arXiv:2607.02807v1 Announce Type: new Abstract: Long-running coding agents such as autoresearch can persistently discover optimizations for open-ended problems. However, they tend to converge onto a s

local-aiarxiv-cs-ai
7 Jul 2026
Model Releases

Toward Efficient Agents: Memory, Tool learning, and Planning

DGX agent

arXiv:2601.14192v2 Announce Type: replace Abstract: Recent years have witnessed increasing interest in extending large language models into agentic systems. While the effectiveness of agents has conti

model-releasesarxiv-cs-ai
7 Jul 2026
Safety

UI-MOPD: Multi-Platform On-Policy Distillation for Continual GUI Agent Learning

DGX agent

arXiv:2607.04425v1 Announce Type: cross Abstract: Recent advances in multimodal foundation models and agent systems have driven GUI agents from single-platform task execution toward cross-platform int

safetyarxiv-cs-ai
7 Jul 2026
Safety

When Agents Lie: Premeditation, Persistence, and Exploitation in Repeated Games

DGX agent

arXiv:2607.05132v1 Announce Type: cross Abstract: As large language models are deployed as autonomous agents that communicate intentions before acting, a critical safety question is whether agents tha

safetyarxiv-cs-cl
7 Jul 2026
Agents

AI agent exploits Langflow in first fully autonomous ransomware attack

DGX agent

Cloud security company Sysdig Inc. has documented what it says is the first ransomware operation carried out from start to finish by an autonomous artificial intelligence agent, a campaign it calls Ja

agentssiliconangle
6 Jul 2026
Agents

Own the loop: A field guide to agent harnesses

DGX agent

As models become cheaper and more interchangeable, the durable advantage shifts to the agent harness: the loop, tools, memory, permissions, and workflow you can own and refine. The post Own the loop:

agentsarize-ai
6 Jul 2026
Agents

Soon most agents for work will run in the cloud and communication with them will happen almost exclusively in your existing work channels (S…

DGX agent

Soon most agents for work will run in the cloud and communication with them will happen almost exclusively in your existing work channels (Slack, Teams, ...) This hasn't become the norm yet because lo

agentsharrison-chase--x
6 Jul 2026
Agents

ActiveGraph makes the agent trace the runtime: Yohei Nakajima's open-source Python runtime treats an append-only event log as the source of …

DGX agent

ActiveGraph makes the agent trace the runtime: Yohei Nakajima's open-source Python runtime treats an append-only event log as the source of truth, enabling replay, forking, and lineage for long-runnin

agentsyohei-nakajima--x
5 Jul 2026
Agents

A Dual-Helix Governance Approach Towards Reliable Agentic Artificial Intelligence for WebGIS Development

DGX agent

arXiv:2603.04390v2 Announce Type: replace Abstract: WebGIS development requires consistency, yet agentic AI often fails due to LLM context constraints, forgetting, stochasticity, instruction failure,

agentsarxiv-cs-ai
3 Jul 2026
Model Releases

BOUNDARY_SYNC: Measuring Communication-Induced Representational Coupling in Multi-Agent LLM Systems

DGX agent

arXiv:2607.01600v1 Announce Type: cross Abstract: As large language models (LLMs) are deployed as communicating agents, does inter-agent communication cause outputs to converge? We introduce BOUNDARY_

model-releasesarxiv-cs-cl
3 Jul 2026
Agents

Coding-agents can replicate scientific machine learning papers

DGX agent

arXiv:2607.02134v1 Announce Type: new Abstract: Scientific machine learning papers typically make computational claims, e.g., that the relative mean square error is less than 5% or that the 95% predic

agentsarxiv-cs-ai
3 Jul 2026
Agents

In this interview at @aiDotEngineer World's Fair, @vercel chief of software @andrewqu explains why agents represent a new form of software, …

DGX agent

In this interview at @aiDotEngineer World's Fair, @vercel chief of software @andrewqu explains why agents represent a new form of software, what Vercel learned from building its own, and why Vercel it

agentsswyx--x
3 Jul 2026
Agents

Mark Zuckerberg says Meta’s agentic AI efforts aren’t progressing as fast as he had hoped

DGX agent

Meta Platforms Inc. Chief Executive Mark Zuckerberg told employees at an internal town hall meeting that the company’s work on artificial intelligence agents hasn’t progressed as quickly as he had hop

agentssiliconangle
3 Jul 2026
Agents

MMAO-Cls: Metabolic Multi-Agent Optimization for Joint Feature Selection and Classifier Tuning

DGX agent

arXiv:2607.01539v1 Announce Type: cross Abstract: This paper studies whether the Metabolic Multi-Agent Optimizer (MMAO) can act as a credible outer-loop optimizer for classification model selection. W

agentsarxiv-cs-lg
3 Jul 2026
Agents

Simulation Based Reward Function Validation for Multi-Agent On Orbit Inspection

DGX agent

arXiv:2607.01367v1 Announce Type: cross Abstract: A proposed method for the control of groups of inspection spacecraft is Multi-Agent Reinforcement Learning (MARL). While MARL has already been employe

agentsarxiv-cs-ro
3 Jul 2026
Model Releases

Steerability via constraints: a substrate for scalable oversight of coding agents

DGX agent

arXiv:2607.02389v1 Announce Type: new Abstract: Coding agents are capable; human oversight is the bottleneck. Unconstrained agents introduce security risks, erode codebase scalability, and make human

model-releasesarxiv-cs-ai
3 Jul 2026
Agents

The Rollout Infrastructure Tax in Coding-Agent Reinforcement Learning

DGX agent

arXiv:2607.01415v1 Announce Type: new Abstract: Coding-agent reinforcement learning treats execution infrastructure as a background implementation detail, despite relying on large numbers of interacti

agentsarxiv-cs-lg
3 Jul 2026
Agents

At a town hall, Mark Zuckerberg said Meta's AI agent development has not accelerated as expected and its reorganization was not as 'clean' as it could have been (Katie Paul/Reuters)

DGX agent

Katie Paul / Reuters: At a town hall, Mark Zuckerberg said Meta's AI agent development has not accelerated as expected and its reorganization was not as “clean” as it could have been — Meta (META.O) C

agentstechmeme
2 Jul 2026
Agents

BaRA: BFS-and-Reflection Web Data Collection Agent

DGX agent

arXiv:2607.00007v1 Announce Type: cross Abstract: Large language model (LLM)-based web agents reduce manual scripting for web data collection, yet on live websites, they often miss relevant pages, ret

agentsarxiv-cs-ai
2 Jul 2026
Agents

Fugu is now available on OpenCode! ✨ When our team was developing Fugu’s multi-agent orchestration, OpenCode was our tool of choice to verif…

DGX agent

Fugu is now available on OpenCode! ✨ When our team was developing Fugu’s multi-agent orchestration, OpenCode was our tool of choice to verify our models. We share a core philosophy with the OpenCode t

agentsdavid-ha--x
2 Jul 2026
Model Releases

GameDevBench: Evaluating Agentic Capabilities Through Game Development

DGX agent

arXiv:2602.11103v2 Announce Type: replace Abstract: Despite rapid progress on coding agents, progress on their multimodal counterparts has lagged behind. A key challenge is the scarcity of evaluation

model-releasesarxiv-cs-ai
2 Jul 2026
Agents

Gavel: Agent Meets Checklist for Evaluating LLMs on Long-Context Legal Summarization

DGX agent

arXiv:2601.04424v2 Announce Type: replace Abstract: Large language models (LLMs) now support contexts of up to 1M tokens, but their strengths and weaknesses on complex long-context tasks remain unclea

agentsarxiv-cs-cl
2 Jul 2026
Agents

I really like this 'understand to participate' framing of the cognitive debt problem when working with coding agents

DGX agent

I really like this 'understand to participate' framing of the cognitive debt problem when working with coding agents That's where another answer comes in: we can understand to participate. You can lea

agentssimon-willison--x
2 Jul 2026
Agents

I’ll be in room 2005 (graph track) at 11:10am to talk about ActiveGraph: event-sourced graph runtime for building auditable agents! Been hav…

DGX agent

Yohei Nakajima will present ActiveGraph, an event-sourced graph runtime designed for building auditable agents, at 11:10am in room 2005 of a conference's graph track. The presentation focuses on lever

agentsyohei-nakajima--x
2 Jul 2026
Agents

Making Failure Safe: A Constrained, Verifiable Agent Framework for Open-Web Data Collection

DGX agent

arXiv:2607.00035v1 Announce Type: new Abstract: LLMs and agents can generate web scrapers from natural-language requirements, but direct generation remains unreliable because of dependency errors, bro

agentsarxiv-cs-ai
2 Jul 2026
Agents

Multi-Turn Agentic Scientific Literature Search via Workflow Induction

DGX agent

arXiv:2607.00597v1 Announce Type: new Abstract: Scientific literature search often requires more than retrieving papers from a single query: users' intents are underspecified, preference-dependent, an

agentsarxiv-cs-cl
2 Jul 2026
Agents

NeuroFilter: Activation-Based Guardrails for Privacy-Conscious LLM Agents

DGX agent

arXiv:2601.14660v2 Announce Type: replace-cross Abstract: Agentic Large Language Models (LLMs) are models able to reason, plan, and execute tools over unstructured data. These abilities are enabling t

agentsarxiv-cs-ai
2 Jul 2026
Agents

On the last day of @aiDotEngineer Worlds Fair SF, I am SO excited to be releasing the latest episode of the Agentic Review podcast, featurin…

DGX agent

On the last day of @aiDotEngineer Worlds Fair SF, I am SO excited to be releasing the latest episode of the Agentic Review podcast, featuring @PaulDuvall, author of Continuous Integration, and AI-nati

agentsitamar-friedman--x
2 Jul 2026
Agents

Pinecone releases Nexus into public preview to bring business knowledge to AI agents

DGX agent

Pinecone Systems Inc., an artificial intelligence infrastructure company providing fully managed vector databases, Wednesday launched the public preview of Pinecone Nexus, which curates and distribute

agentssiliconangle
2 Jul 2026
Agents

Self-GC: Self-Governing Context for Long-Horizon LLM Agents

DGX agent

arXiv:2607.00692v1 Announce Type: new Abstract: Long-horizon LLM agents accumulate tool results, files, plans, and user constraints that are too structured to be treated as a disposable text suffix. C

agentsarxiv-cs-ai
2 Jul 2026
Model Releases

SWE-Doctor: Guiding Software Engineering Agents with Runtime Diagnosis from Multi-Faceted Bug Reproduction Tests

DGX agent

arXiv:2607.00990v1 Announce Type: cross Abstract: Large language model (LLM)-based software engineering agents are increasingly developed to resolve software issues by generating patches from issue re

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

WorkBench Revisited: Workplace Agents Two Years On

DGX agent

arXiv:2606.13715v2 Announce Type: replace Abstract: The best agent on WorkBench in March 2024, GPT-4, completed just 43% of tasks. We revisit the benchmark in June 2026 and find that the best agent to

model-releasesarxiv-cs-ai
2 Jul 2026
Agents

Autoresearch: The feedback loop behind self-improving agents

DGX agent

Autoresearch describes a feedback mechanism where AI agents can evaluate their own outputs and use that introspection to iteratively improve their reasoning and decision-making capabilities. This conc

agentslatent-space
1 Jul 2026
Agents

DeXposure-Claw: An Agentic System for DeFi Risk Supervision

DGX agent

arXiv:2606.19501v2 Announce Type: replace Abstract: Decentralized finance exposes supervisors to fast-moving, networked credit risks. General-purpose LLM agents fit this setting poorly: they over-read

agentsarxiv-cs-ai
1 Jul 2026
Agents

DigitalCoach: Communication and Grounding Gaps in Human and Agentic Computer Use Coaching

DGX agent

arXiv:2606.31980v1 Announce Type: new Abstract: Agents are increasingly capable of automating software tasks, but can they teach humans how to use software themselves? We introduce DigitalCoach, a mul

agentsarxiv-cs-cl
1 Jul 2026
Model Releases

Learning from Failure: Inference-Time Self-Improvement for Computer-Use Agents

DGX agent

arXiv:2606.31270v1 Announce Type: cross Abstract: Computer-use agents, which leverage multimodal large language models (MLLMs) to operate computers and complete tasks, have attracted significant atten

model-releasesarxiv-cs-ai
1 Jul 2026
Agents

OpenLife: Toward Open-World Artificial Life with Autonomous LLM Agents

DGX agent

arXiv:2606.31046v1 Announce Type: new Abstract: Artificial life has explored life-like behavior on many computational substrates, but mostly in researcher-designed closed worlds. We argue that large l

agentsarxiv-cs-ai
1 Jul 2026
Agents

The best players want to be coached. You build trust, then you push hard. Now here's the thing: your agent has no skin in the game. no ego t…

DGX agent

The best players want to be coached. You build trust, then you push hard. Now here's the thing: your agent has no skin in the game. no ego to protect, no trust to earn first. So you can and should ski

agentsitamar-friedman--x
1 Jul 2026
Model Releases

Understanding and Evaluating Claw-like Agent Security Through a Computer-Systems Lens

DGX agent

arXiv:2606.30755v1 Announce Type: cross Abstract: Claw-like AI agents (e.g., OpenClaw) are always-on processes with persistent access to credentials, files, tools, and external services. They take on

model-releasesarxiv-cs-ai
1 Jul 2026
Agents

Using AI Agents to Automate Black-Box Audits of Personalization Algorithms at Scale

DGX agent

arXiv:2606.30801v1 Announce Type: new Abstract: Personalization algorithms determine what content users encounter on online platforms. Auditing these systems is difficult because independent auditors

agentsarxiv-cs-cl
1 Jul 2026
Agents

We’re also publishing extensive documentation and technical materials about Agentic MapReduce, including a deep-dive on our evals. Read our …

DGX agent

We’re also publishing extensive documentation and technical materials about Agentic MapReduce, including a deep-dive on our evals. Read our announcement: https://cognition.com/blog/introducing-devin-s

agentscognition-ai--x
1 Jul 2026
Agents

A living map of everything your Hermes agent has learned; every memory & skill Press play to watch it unfold, @NousResearch Copy yours into …

DGX agent

This post from Nous Research demonstrates a visualization tool or system that maps the accumulated knowledge, memories, and skills learned by a Hermes AI agent during its training or operation. The in

agentsnous-research--x
30 Jun 2026
Agents

An expanded Vercel Agent: chat, investigations, and approved actions, now in public beta

DGX agent

Vercel has released an expanded version of its Vercel Agent in public beta, introducing new capabilities including chat functionality, investigation tools, and an approved actions feature. The update

agentsvercel-blog
30 Jun 2026
← Previous
1…8384858687…374
Next →