AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,761 results
Agents

Neural Procedural Memory: Empowering LLM Agents with Implicit Activation Steering

DGX agent

arXiv:2606.29824v1 Announce Type: cross Abstract: While Large Language Models (LLMs) excel as static solvers, transforming them into autonomous agents remains challenging. This transition requires con

agentsarxiv-cs-ai
30 Jun 2026
Agents
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Self-Evolving World Models for LLM Agent Planning

DGX agent

arXiv:2606.30639v1 Announce Type: new Abstract: World models offer a principled way to equip long-horizon LLM agents with foresight: predictions of action consequences before execution. However, unrel

agentsarxiv-cs-ai
30 Jun 2026
Agents

SkillOpt: Agent skills as trainable parameters

DGX agent

AI agents often fail because their instructions, or skills, are manually modified with no guarantee of improvement. Learn how SkillOpt turns skill editing into a training process, making agent behavio

agentsmicrosoft-research
30 Jun 2026
Agents

Agentic Publication Protocol: An Attempt to Modernize Scientific Publication

DGX agent

arXiv:2606.27386v1 Announce Type: cross Abstract: Scientific publication is still organized primarily around static manuscripts, even though much of scientific progress depends on tacit know-how: how

agentsarxiv-cs-ai
29 Jun 2026
Agents

Delayed Verification Destabilizes Multi-Agent LLM Belief: Instability Thresholds and Optimal Corrector Placement

DGX agent

arXiv:2606.27409v1 Announce Type: cross Abstract: Multi-agent large language model (LLM) systems often rely on verifier and critic agents to suppress hallucinations, but verification is delayed. Durin

agentsarxiv-cs-cl
29 Jun 2026
Model Releases

Govern the Repository, Not the Agent: Measuring Ecosystem-Level Risk in AI-Native Software

DGX agent

arXiv:2606.28235v1 Announce Type: cross Abstract: Autonomous coding agents now open and merge pull requests in shared repositories at scale, and the field evaluates them the way it has always evaluate

model-releasesarxiv-cs-ai
29 Jun 2026
Agents

Straiker, which develops tech for securing enterprise AI agents, raised a 64M Series A, bringing its total funding to 85M (Chris Metinko/Axios)

DGX agent

Chris Metinko / Axios: Straiker, which develops tech for securing enterprise AI agents, raised a 64M Series A, bringing its total funding to 85M — Straiker, which secures AI agents, raised a $64 milli

agentstechmeme
29 Jun 2026
Safety

A Process Harness for Uplifting Legacy Workflows to Agentic BPM: Design and Realization in CUGA FLO

DGX agent

arXiv:2606.27188v1 Announce Type: new Abstract: We introduce the process harness, a new mechanism for uplifting legacy workflows into Agentic Business Process Management (Agentic BPM) without replacin

safetyarxiv-cs-ai
26 Jun 2026
Agents

Excited to see @cohere open-source how they use AI coding agents to maintain their vLLM fork. 🙌 Keeping a long-lived fork in sync is tricky…

DGX agent

Excited to see @cohere open-source how they use AI coding agents to maintain their vLLM fork. 🙌 Keeping a long-lived fork in sync is tricky. Their agent treats it as a control loop: rebase onto each u

agentscohere--x
26 Jun 2026
Agents

What happens when AI agents collaborate on open science? At @aiDotEngineer World’s Fair, @james_y_zou will share work on EinsteinArena and D…

DGX agent

What happens when AI agents collaborate on open science? At @aiDotEngineer World’s Fair, @james_y_zou will share work on EinsteinArena and DSGym, from multi-agent math discovery to better evaluation f

agentstogether-ai--x
26 Jun 2026
Agents

Agentic System as Compressor: Quantifying System Intelligence in Bits

DGX agent

arXiv:2606.25960v1 Announce Type: new Abstract: Large language models are turning from isolated predictors into agentic systems: they call tools, retrieve evidence, obey environment constraints, use v

agentsarxiv-cs-ai
25 Jun 2026
Agents

Autodata: An agentic data scientist to create high quality synthetic data

DGX agent

arXiv:2606.25996v1 Announce Type: cross Abstract: We introduce Autodata, a general method that enables AI agents to act as data scientists who build high quality training and evaluation data. We show

agentsarxiv-cs-cl
25 Jun 2026
Agents

Domain-Specific Agents for Cherenkov Telescope Array Control Software and Gamma-Ray Data Analysis

DGX agent

arXiv:2510.01299v3 Announce Type: replace-cross Abstract: We present domain-adapted large language model agents designed to support Cherenkov Telescope Array operation and data analysis. The agents co

agentsarxiv-cs-ai
25 Jun 2026
Agents

Evaluating AGENTS.md: Are Repository-Level Context Files Helpful for Coding Agents?

DGX agent

arXiv:2602.11988v2 Announce Type: replace-cross Abstract: A widespread practice in software development is to tailor coding agents to repositories using context files, such as AGENTS.md. Although this

agentsarxiv-cs-ai
25 Jun 2026
Safety

Low Variance Trust Region Optimization with Independent Actors and Sequential Updates in Cooperative Multi-agent Reinforcement Learning

DGX agent

arXiv:2606.25526v1 Announce Type: new Abstract: Cooperative multi-agent reinforcement learning assumes each agent shares the same reward function and can be trained effectively using the Trust Region

safetyarxiv-cs-lg
25 Jun 2026
Model Releases

Salesforce launches Help Agent to simplify AI customer service deployment

DGX agent

Salesforce Inc. is launching a new prepackaged artificial intelligence agent for customer service, enabling organizations to quickly build and deploy AI agents. Today Salesforce announced Help Agent,

model-releasessiliconangle
25 Jun 2026
Agents

Agent architecture is essentially a solved problem. What's not: memory. Great article by @jakebroekhuizen who's a true expert on this

DGX agent

This post highlights that while agent architecture design has become well-established and solved, memory systems remain an unsolved challenge in AI agent development. The tweet references an article b

agentsharrison-chase--x
24 Jun 2026
Agents

🧠LangSmith Engine as Sleep Time Compute Memory for agents is often described as “sleep time compute” or “dreaming” This involves running a …

DGX agent

🧠LangSmith Engine as Sleep Time Compute Memory for agents is often described as “sleep time compute” or “dreaming” This involves running a background process to analyze agent trajectories and update a

agentsharrison-chase--x
24 Jun 2026
Agents

RIFT-Bench: Dynamic Red-teaming For Agentic AI Systems

DGX agent

arXiv:2606.23927v1 Announce Type: new Abstract: Agentic AI systems powered by large language models (LLMs) are rapidly evolving into autonomous decision-making systems, exposing attack vectors beyond

agentsarxiv-cs-ai
24 Jun 2026
Model Releases

AgentMisalignment: Measuring the Propensity for Misaligned Behaviour in LLM-Based Agents

DGX agent

arXiv:2506.04018v3 Announce Type: replace-cross Abstract: As Large Language Model (LLM) agents become more widespread, associated misalignment risks increase. While prior research has studied agents'

model-releasesarxiv-cs-lg
23 Jun 2026
Agents

Build real agentic apps using CUGA: two dozen working examples on a lightweight harness

DGX agent

CUGA is a lightweight framework from IBM Research for building agentic AI applications, featuring two dozen practical working examples. The framework enables developers to create autonomous AI agents

agentshugging-face
23 Jun 2026
Agents

ClickHouse brings real-time analytics to agentic AI

DGX agent

The growing use of AI agents throughout the enterprise is forcing a thorough reevaluation of the data layer. This shift is driven by the need for millisecond responses that enable agents to make decis

agentssiliconangle
23 Jun 2026
Safety

Counsel: A Meta-Evaluation Dataset for Agentic Tasks

DGX agent

arXiv:2606.21627v1 Announce Type: cross Abstract: As agentic systems tackle increasingly complex multi-step tasks, evaluating their trajectories presents a major bottleneck - human annotation of a sin

safetyarxiv-cs-lg
23 Jun 2026
Agents

Learn to use the new eve agentic framework from Vercel. Go try out the hands-on labs now.

DGX agent

Learn to use the new eve agentic framework from Vercel. Go try out the hands-on labs now. I'm digging the eve agentic framework from Vercel. I like that everything is files, from the tools to the skil

agentsdair-ai--x
23 Jun 2026
Agents

you need docs built for agents

DGX agent

you need docs built for agents Docs are the eyes and ears of Agents. But we're moving so fast that they are always outdated. So hence an outdated doc also confuses the models. But it's also hard keep

agentsharrison-chase--x
23 Jun 2026
Agents

context engineering docs for agentic engineering - plans, research, etc SHOULD NOT be stored in version control: A good docs management syst…

DGX agent

context engineering docs for agentic engineering - plans, research, etc SHOULD NOT be stored in version control: A good docs management system keeps them: > outside your repo > accesible to agent via

agentsharrison-chase--x
22 Jun 2026
Agents

How does it work? Sakana Fugu is itself an LLM, trained to call various LLMs in an agent pool, including instances of itself recursively. Fu…

DGX agent

How does it work? Sakana Fugu is itself an LLM, trained to call various LLMs in an agent pool, including instances of itself recursively. Fugu dynamically orchestrates the world's best models to tackl

agentsdavid-ha--x
22 Jun 2026
Agents

Agentic Environment Engineering for Large Language Models: A Survey of Environment Modeling, Synthesis, Evaluation, and Application

DGX agent

arXiv:2606.12191v1 Announce Type: cross Abstract: Environments serve as interactive systems for large language model (LLM) based agents across diverse scenarios and play a crucial role in driving the

agentsarxiv-cs-ai
11 Jun 2026
Model Releases

AI Coding Agents in Social Science: Methodologically Diverse, Empirically Consistent, Interpretively Vulnerable

DGX agent

arXiv:2606.11456v1 Announce Type: cross Abstract: The deployment of LLM-based agents in scientific analysis raises opposing concerns: that agents may reduce methodological diversity, or that they may

model-releasesarxiv-cs-ai
11 Jun 2026
Safety

An Ethical eValuation Agent (EeVA): Results of a Proof-of-Concept Test on a Prototype Agentic-like Workflow to Assist Ethical Deliberations

DGX agent

arXiv:2606.11218v1 Announce Type: cross Abstract: Ethical deliberation is often misunderstood as a search for single right or wrong answers, creating difficulties for non-ethically trained personnel w

safetyarxiv-cs-ai
11 Jun 2026
Agents

FlowBank: Query-Adaptive Agentic Workflows Optimization through Precompute-and-Reuse

DGX agent

arXiv:2606.11290v1 Announce Type: cross Abstract: Large Language Model (LLM)-based multi-agent systems are increasingly powerful, but current agentic workflow optimization paradigms make an unsatisfyi

agentsarxiv-cs-ai
11 Jun 2026
Agents

Runtime Skill Audit: Targeted Runtime Probing for Agent Skill Security

DGX agent

arXiv:2606.11671v1 Announce Type: cross Abstract: Agent skills let LLM agents reuse instructions, resources, tools, and workflows, but they also create a new place for malicious behavior to hide. A sk

agentsarxiv-cs-ai
11 Jun 2026
Model Releases

STAGE-Claw: Automated State-based Agent Benchmarking for Realistic Scenarios

DGX agent

arXiv:2606.10394v1 Announce Type: new Abstract: Large language models are increasingly used to power personal agents for everyday applications, but evaluating these agents remains a challenge. Existin

model-releasesarxiv-cs-ai
10 Jun 2026
Agents

Stop hand-tuning kernels: How Neuron Agentic Development accelerates AWS Trainium optimizations

DGX agent

Today, we’re announcing the Neuron Agentic Development capabilities: a collection of AI agents and skills that make this possible for developers building on AWS Trainium and AWS Inferentia. In this po

agentsaws-ml-blog
10 Jun 2026
Agents

Visa partners with OpenAI to let AI agents make payments for users

DGX agent

Visa Inc. has struck a deal with OpenAI Group PBC to let artificial intelligence agents make payments for users, bringing one of the world’s largest payment networks into ChatGPT’s push toward agentic

agentssiliconangle
10 Jun 2026
Safety

Claw-R1: A Step-Level Data Middleware System for Agentic Reinforcement Learning

DGX agent

arXiv:2606.09138v1 Announce Type: new Abstract: Agentic reinforcement learning (RL) has become an important post-training paradigm for turning LLMs from static chatbots into interactive agents, giving

safetyarxiv-cs-lg
9 Jun 2026
Agents

DeltaBox: Scaling Stateful AI Agents with Millisecond-Level Sandbox Checkpoint/Rollback

DGX agent

arXiv:2605.22781v2 Announce Type: replace-cross Abstract: LLM-powered AI agents require high-frequency state exploration (e.g., test-time tree search and reinforcement learning), relying on rapid chec

agentsarxiv-cs-ai
9 Jun 2026
Agents

Multi-Turn Evaluation of Deep Research Agents Under Process-Level Feedback

DGX agent

arXiv:2606.09748v1 Announce Type: new Abstract: Existing benchmarks for deep research agents (DRAs) assess only single-shot outputs, ignoring a key question: can DRAs improve their reports when guided

agentsarxiv-cs-ai
9 Jun 2026
Model Releases

Act As a Real Researcher: A Suite of Benchmarks Evaluating Frontier LLMs and Agentic Harnesses in Research Lifecycle

DGX agent

arXiv:2606.07462v1 Announce Type: new Abstract: As foundation models advance and agent scaffolding becomes increasingly sophisticated, agents have demonstrated remarkable proficiency in complex, long-

model-releasesarxiv-cs-ai
8 Jun 2026
Agents

Deep Agents explained in <90 seconds by @sydneyrunkle

DGX agent

Deep Agents are an AI concept that Sydney Runkle explains concisely in under 90 seconds, likely covering how agents can be designed to operate with deeper reasoning and decision-making capabilities. T

agentsharrison-chase--x
8 Jun 2026
Agents

Now you can text Hermes on iMessage 💬 It's just that simple. Photon brings the BEST agents to everyone where they already are -> http://try…

DGX agent

Now you can text Hermes on iMessage 💬 It's just that simple. Photon brings the BEST agents to everyone where they already are -> http://tryphoton.ai Your Hermes Agent now lives in iMessage via @photon

agentsnous-research--x
8 Jun 2026
Agents

The Three-Ring Architecture: Governing Agents in the Era of On-Platform Organisations

DGX agent

arXiv:2606.07119v1 Announce Type: cross Abstract: The current phase of enterprise AI deployment faces a structural failure: organisations are acquiring agentic capability without the infrastructure to

agentsarxiv-cs-ai
8 Jun 2026
Agents

When Does Multi-Agent Collaboration Help? An Entropy Perspective

DGX agent

arXiv:2602.04234v6 Announce Type: cross Abstract: Multi-agent systems (MAS) have emerged as a prominent paradigm for leveraging large language models (LLMs) to tackle complex tasks. However, the mecha

agentsarxiv-cs-ai
8 Jun 2026
Agents

Snowflake, Databricks and the model makers: The battle for the agentic client and AI back end

DGX agent

Agentic artificial intelligence is being misread as a set of separate battles – for example, Snowflake Inc. versus Databricks Inc., copilots versus agents, model makers versus application vendors. We

agentssiliconangle
7 Jun 2026
Safety

From Risk Classification to Action Plan Remediation: A Guardrail Feedback Driven Framework for LLM Agents

DGX agent

arXiv:2606.05805v1 Announce Type: new Abstract: LLM-based guardrails typically safeguard agents by evaluating proposed actions or inputs before execution, producing safety signals such as binary allow

safetyarxiv-cs-ai
6 Jun 2026
Agents

Hermes Agent is for the Artists

DGX agent

Hermes Agent, developed by Nous Research, is an AI agent designed specifically for creative professionals and artists. The tool likely focuses on automating tasks, enhancing workflows, or providing sp

agentsnous-research--x
6 Jun 2026
Agents

Hermes Agent v0.16.0 - “The Surface Release” Changelog below:

DGX agent

Hermes Agent v0.16.0, released by Nous Research and dubbed 'The Surface Release,' represents an update to their Hermes Agent framework. This version likely includes bug fixes, performance improvements

agentsnous-research--x
6 Jun 2026
Model Releases

SentinelBench: A Benchmark for Long-Running Monitoring Agents

DGX agent

arXiv:2606.05342v1 Announce Type: new Abstract: AI agents are increasingly asked to carry out work that spans minutes, hours, or longer. Yet the default model of agent behavior is continuous action: i

model-releasesarxiv-cs-ai
6 Jun 2026
← Previous
1…5152535455…371
Next →