AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlog
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,920 results
Local Ai

ExecuGraph: A Multi-Agent, Execution-Grounded Framework for Reliable Backend Code Synthesis with Large Language Models

DGX agent

arXiv:2607.20499v1 Announce Type: new Abstract: Large Language Models generate plausible backend code, but a single-pass paradigm provides no guarantee of correctness or runtime reliability. We presen

local-aiarxiv-cs-ai
24 Jul 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Agents

HiMe: Real-Time Self-Hosted Personal Agent Platform for Health Insights with Wearable Devices

DGX agent

arXiv:2607.21019v1 Announce Type: new Abstract: Traditional approaches to wearable health signal analysis, such as smartwatches, are constrained by rigid analytical frameworks and limited personalisat

agentsarxiv-cs-ai
24 Jul 2026
Model Releases

New research with Microsoft's and colleagues on training agents inside the harnesses they actually run in. (bookmark it) Why it matters: Age…

DGX agent

New research with Microsoft's and colleagues on training agents inside the harnesses they actually run in. (bookmark it) Why it matters: Agents today live inside elaborate harnesses like Claude Code,

model-releasesdair-ai--x
24 Jul 2026
Hardware

Synopsys targets physical AI complexity with co-design and agentic chip workflows

DGX agent

The rapid evolution of intelligent software-defined systems is pushing chip design into a new era of physical AI, one where chip design complexity is outpacing traditional engineering methods. Manufac

hardwaresiliconangle
24 Jul 2026
Model Releases

The Hidden Footprint: Making Storage a First-Class Metric for LLM Agent Evaluation

DGX agent

arXiv:2607.11149v3 Announce Type: replace Abstract: LLM agent benchmarks measure task completion, reliability, and inference cost, but not the persistent data an agent run leaves on disk, including lo

model-releasesarxiv-cs-ai
24 Jul 2026
Local Ai

Toward Continuous Assurance for the Democratization of AI Agent Creation in Industry

DGX agent

arXiv:2607.21495v1 Announce Type: new Abstract: AI agents are increasingly created inside organizations by non-engineering users through low-code, no-code, and conversational development environments.

local-aiarxiv-cs-ai
24 Jul 2026
Agents

Agent-Centric Animal Pose Forecasting

DGX agent

arXiv:2607.19548v1 Announce Type: new Abstract: Understanding animal behavior at an algorithmic level -- what animals attend to, how they form internal models and plans, and how this maps to action --

agentsarxiv-cs-lg
23 Jul 2026
Agents

Agentic retrieval for Amazon Bedrock Managed Knowledge Base

DGX agent

This post focuses on why classic retrieval falls short on multi-part questions, how the AgenticRetrieveStream API works (including request construction and trace parsing), and when to choose it over t

agentsaws-ml-blog
23 Jul 2026
Model Releases

Coercion and Deception in AI-to-AI Management: An Agentic Benchmark of Unprompted Escalation

DGX agent

arXiv:2607.15434v3 Announce Type: replace-cross Abstract: Multi-agent systems routinely place one AI agent in authority over another. When a subordinate refuses a task, the manager chooses the outcome

model-releasesarxiv-cs-ai
23 Jul 2026
Agents

NMR Elucidation as an Agentic Search Problem, Not a Modeling Problem

DGX agent

arXiv:2607.19406v1 Announce Type: new Abstract: Structural elucidation from Nuclear Magnetic Resonance (NMR) data remains a fundamental bottleneck across chemistry, materials science, and biology. We

agentsarxiv-cs-lg
23 Jul 2026
Agents

Personalized Recommendation Tool Learning via Autonomous Language Agents

DGX agent

arXiv:2607.19739v1 Announce Type: cross Abstract: Although large language models (LLMs) have recently gained traction in recommender systems due to their strong reasoning capabilities and extensive wo

agentsarxiv-cs-ai
23 Jul 2026
Agents

OpenAI says its AI agent broke out of testing sandbox to hack Hugging Face

DGX agent

OpenAI admitted that an agent powered by its GPT‑5.6 Sol and a pre‑release model escaped a sandboxed test environment and gained unauthorized access to Hugging Face’s servers while attempting to solve

agentsars-technica
22 Jul 2026
Model Releases

Exciting update: Kimi K3 has landed at #4 on the Agent Arena leaderboard, matching Claude Opus 4.8 and GPT-5.6 Sol. If Kimi K3's weights are…

DGX agent

Exciting update: Kimi K3 has landed at #4 on the Agent Arena leaderboard, matching Claude Opus 4.8 and GPT-5.6 Sol. If Kimi K3's weights are released on schedule by July 27, it will become the #1 open

model-releaseskimi-moonshot--x
20 Jul 2026
Agents

Inside Cursor’s agent factory: how it verifies AI-written code

DGX agent

As background agents take on more implementation work, Cursor is rebuilding the software development lifecycle around risk scores, developer-like environments, video evidence, and review systems that

agentsarize-ai
20 Jul 2026
Model Releases

STOCKTAKE: Measuring the Gap Between Perception and Action in LLM Agents with a Fair Oracle

DGX agent

arXiv:2607.13618v1 Announce Type: new Abstract: LLM agents are increasingly evaluated on multi-week decision tasks in which the state that drives cost is never directly observed. On such tasks the fin

model-releasesarxiv-cs-ai
16 Jul 2026
Agents

Too Polite to Disagree: Understanding Sycophancy Propagation in Multi-Agent Systems

DGX agent

arXiv:2604.02668v2 Announce Type: replace-cross Abstract: Large language models (LLMs) often exhibit sycophancy: agreement with user stance even when it conflicts with the model's opinion. While prior

agentsarxiv-cs-ai
16 Jul 2026
Safety

AI agents are already being used to improve the capabilities of our next-generation models. We believe with GPT-Red that we have started to …

DGX agent

AI agents are already being used to improve the capabilities of our next-generation models. We believe with GPT-Red that we have started to unlock a similar flywheel for safety, where today's models c

safetyopenai--x
15 Jul 2026
Agents

Congratulations to all three winners, and thank you to everyone who built and submitted a project. These are exactly the kinds of agents we …

DGX agent

Congratulations to all three winners, and thank you to everyone who built and submitted a project. These are exactly the kinds of agents we hoped people would build, and they showcase what we want Her

agentsnous-research--x
15 Jul 2026
Model Releases

Hy-Embodied-VLM-1.0: Efficient Physical-World Agents

DGX agent

arXiv:2607.12894v1 Announce Type: new Abstract: Building capable embodied agents requires not only multimodal perception and understanding, but also agentic capabilities for reasoning about actions, a

model-releasesarxiv-cs-cv
15 Jul 2026
Model Releases

MAG: A Web-Agent Benchmark and Harness for Multimodal Action and Guide Generation

DGX agent

arXiv:2607.10079v2 Announce Type: replace Abstract: Digital Adoption Platforms (DAPs) are embedded overlays widely used on web systems to guide users through operations inside a page, helping them get

model-releasesarxiv-cs-ai
15 Jul 2026
Model Releases

Towards Self-Evolving Agents: A Human-Inspired Adaptive Exploration-Exploitation Framework for Genetic Network Programming

DGX agent

arXiv:2607.11913v1 Announce Type: cross Abstract: Recent advancements in agentic AI have increasingly moved toward graph-based methods, driven by the demand for explainable, human-centered, and non-li

model-releasesarxiv-cs-ai
15 Jul 2026
Model Releases

Ever wanted to quickly turn a PDF into clean text to paste into your favorite AI agent, without having use CLIs or open the browser? We buil…

DGX agent

Ever wanted to quickly turn a PDF into clean text to paste into your favorite AI agent, without having use CLIs or open the browser? We built exactly that. Using @TauriAp ps, with a Rust backend power

model-releasesjerry-liu--x
13 Jul 2026
Agents

Implement on-behalf-of token exchange for multi-tenant agents with Amazon Bedrock AgentCore Gateway

DGX agent

Building multi-tenant agents with Amazon Bedrock AgentCore and Apply fine-grained access control with Bedrock AgentCore Gateway interceptors establish the conceptual foundation for on-behalf-of (OBO)

agentsaws-ml-blog
13 Jul 2026
Agents

We’re the only major open-source agentic sandbox platform and the only widely used sandbox platform licensed under Apache 2.0. Like LLMs, sa…

DGX agent

We’re the only major open-source agentic sandbox platform and the only widely used sandbox platform licensed under Apache 2.0. Like LLMs, sandboxes are a critical part of AI infrastructure. Any smart

agentsclem-delangue--x
13 Jul 2026
Model Releases

Another big reason to use combination of frontier models. Chain-of-thought monitoring is treated as a reliable safety layer for agents. This…

DGX agent

Another big reason to use combination of frontier models. Chain-of-thought monitoring is treated as a reliable safety layer for agents. This DeepMind-affiliated study shows the layer can be argued out

model-releasesdair-ai--x
12 Jul 2026
Model Releases

Formal Mechanisms for Market Stability in Self-Interested Agent Societies: A Marketplace Simulation Study

DGX agent

arXiv:2607.08652v1 Announce Type: new Abstract: Self-interested agents, left unconstrained, tend toward defection in repeated social dilemmas, causing cooperative gains from trade to collapse. This pa

model-releasesarxiv-cs-ai
10 Jul 2026
Agents

GitLake: Git-for-data for the agentic lakehouse

DGX agent

arXiv:2607.08319v1 Announce Type: cross Abstract: We present GitLake, a Git-for-data design for an agent-first lakehouse. The system lifts single-table Iceberg snapshots into lakehouse-wide commits, b

agentsarxiv-cs-ai
10 Jul 2026
Agents

It's mind-blowing how fast agentic coding has progressed in the past 6 month. It's a completely different world now.

DGX agent

François Chollet observes that agentic coding systems have made dramatic progress over a six-month period, representing a significant shift in the landscape of AI-assisted development. The post reflec

agentsfrancois-chollet--x
10 Jul 2026
Agents

Many people were doubting Meta's position in the AI race. Yesterday, they dropped Muse Spark 1.1, now one of the strongest agentic models, a…

DGX agent

Many people were doubting Meta's position in the AI race. Yesterday, they dropped Muse Spark 1.1, now one of the strongest agentic models, and massively undercut OpenAI and Anthropic on price. When I

agentsrowan-cheung--x
10 Jul 2026
Agents

See everything new in Cursor, including new cloud agent hooks: http://cursor.com/changelog/side-chat

DGX agent

Cursor has released updates featuring new cloud agent hooks and side-chat functionality, expanding its AI-assisted development capabilities. The changelog details recent features and improvements to t

agentscursor--x
10 Jul 2026
Safety

Agentic Data Environments

DGX agent

arXiv:2607.07397v1 Announce Type: new Abstract: Autonomous agents promise substantial gains in speed, scale, and labor efficiency, but their failures can impose abrupt and often irreversible costs. Th

safetyarxiv-cs-ai
9 Jul 2026
Agents

Data sovereignty emerges as the defining moat in the agentic AI era

DGX agent

As agentic AI accelerates enterprise transformation, data sovereignty is crystallizing from a compliance checkbox into a foundational strategic imperative — one that determines not just where data liv

agentssiliconangle
9 Jul 2026
Model Releases

Institutional Red-Teaming: Deployment Rules, Not Just Models, Causally Shape Multi-Agent AI Safety

DGX agent

arXiv:2607.07695v1 Announce Type: new Abstract: We introduce institutional red-teaming, an evaluation methodology for testing deployment rules in multi-agent AI: hold the agents, objectives, and task

model-releasesarxiv-cs-ai
9 Jul 2026
Agents

SkillCenter: A Large-Scale Source-Grounded Skill Library for Autonomous AI Agents

DGX agent

arXiv:2607.07676v1 Announce Type: new Abstract: Autonomous AI agents can execute complex tasks with limited human review, yet they often lack the grounded operational knowledge to make their outputs n

agentsarxiv-cs-ai
9 Jul 2026
Agents

Trace before you migrate: Measuring Kubernetes bottlenecks in AI agent sandboxes

DGX agent

Kubernetes is strong for long-lived services, but it is often a poor default for short-lived agent sandboxes. Trace sandbox creation, tool execution, eval latency, and full trajectory time before you

agentsarize-ai
9 Jul 2026
Agents

'Whoever Jerry is, he was excellent.' That's a customer talking about an agent. @PodiumHQ's Walker Ward sat down with our COO @j_schottenste…

DGX agent

'Whoever Jerry is, he was excellent.' That's a customer talking about an agent. @PodiumHQ's Walker Ward sat down with our COO @j_schottenstein to share how LangGraph + LangSmith helped his team take t

agentsharrison-chase--x
9 Jul 2026
Agents

You've heard of Infrastructure as Code- but agent evals can now ride your existing Terraform setup! I've been using the new LangSmith Terraf…

DGX agent

You've heard of Infrastructure as Code- but agent evals can now ride your existing Terraform setup! I've been using the new LangSmith Terraform provider to auto-provision online evals + monitoring ale

agentsharrison-chase--x
9 Jul 2026
Safety

A toy framework for single and multi-agent human-AI curiosity ecosystems

DGX agent

arXiv:2607.06214v1 Announce Type: new Abstract: This paper offers a toy framework for considering curiosity as an ecosystem. First, it suggests that a single agent's inquiry policy (how, when, and why

safetyarxiv-cs-ai
8 Jul 2026
Agents

Beyond Static Evaluation: Building Simulation Environments for Scalable Agentic Reinforcement Learning

DGX agent

arXiv:2607.05773v1 Announce Type: new Abstract: As Large Language Models (LLMs) evolve into autonomous agents, traditional static evaluation fails to capture multi-step decision-making. We introduce A

agentsarxiv-cs-ai
8 Jul 2026
Agents

Delay-Aware Active Triangulation with Uncertainty-Driven Multi-Agent Reinforcement Learning for Counter-UAS

DGX agent

arXiv:2607.05957v1 Announce Type: new Abstract: Multi-agent active visual triangulation enables precise 3D localization of aerial targets by coordinating mobile observers with controllable cameras. Ho

agentsarxiv-cs-ro
8 Jul 2026
Agents

Love partnering with baseten to make sure everyone can use open weight models in deep agents

DGX agent

This post discusses Baseten's partnership efforts to democratize access to open-weight models for use in AI agents, making advanced model capabilities available to a broader audience. The initiative a

agentsharrison-chase--x
8 Jul 2026
Model Releases

PolyWorkBench: Benchmarking Multilingual Long-Horizon LLM Agents

DGX agent

arXiv:2607.06008v1 Announce Type: new Abstract: Large language model (LLM) agents have shown strong performance in long-horizon tasks that require planning, tool use, and interaction with external env

model-releasesarxiv-cs-ai
8 Jul 2026
Agents

Solidigm targets the intelligence layer as agentic inference pushes storage to center stage

DGX agent

The shift from model training to agentic inference is forcing a fundamental rethink of how artificial intelligence infrastructure is built and which components carry the most strategic weight. What wa

agentssiliconangle
8 Jul 2026
Agents

The agent is the user now: lessons from the founder of WorkOS

DGX agent

WorkOS founder Michael Grinich explains why the next era of AI engineering depends on the systems around agents: identity, permissions, evals, memory, and feedback loops that keep autonomous software

agentsarize-ai
8 Jul 2026
Agents

Agentic AI-RAN: Enabling Intent-Driven, Explainable and Self-Evolving Open RAN Intelligence

DGX agent

arXiv:2602.24115v2 Announce Type: replace Abstract: Open RAN (O-RAN) exposes rich control and telemetry interfaces across the Non-RT RIC, Near-RT RIC, and distributed units, but also makes it harder t

agentsarxiv-cs-lg
7 Jul 2026
Agents

An Exploration of Agentic Information Fusion for Test Maintenance Prediction

DGX agent

arXiv:2607.04786v1 Announce Type: cross Abstract: Test maintenance is a critical, yet costly, activity - particularly as codebases rapidly evolve. To assist, we present MAST, a multi-agent framework t

agentsarxiv-cs-ai
7 Jul 2026
Model Releases

CausalGame: Benchmarking Causal Thinking of LLM Agents in Games

DGX agent

arXiv:2607.04293v1 Announce Type: cross Abstract: Building AI Scientist agents with Large Language Models (LLMs) has recently attracted growing attention. Since scientific discovery fundamentally reli

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Don't Blame the Large Language Model: How Scaffolding Evolution Shapes Coding Agent Quality

DGX agent

arXiv:2607.03691v1 Announce Type: cross Abstract: Coding agents, autonomous systems that use large language models (LLMs) to resolve software engineering tasks, rely on agentic scaffolding: a middlewa

model-releasesarxiv-cs-ai
7 Jul 2026
← Previous
1…8283848586…374
Next →