AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,958 results
Model Releases

What Makes Interaction Trajectories Effective for Training Terminal Agents?

DGX agent

arXiv:2606.03461v1 Announce Type: new Abstract: Stronger code agents are commonly assumed to be superior teachers for post-training, yet this assumption remains poorly disentangled from task difficult

model-releasesarxiv-cs-ai
3 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Agents

A Multi-AI-agent Framework Enabling End-to-end Finite Element Analysis for Solid Mechanics Problems

DGX agent

arXiv:2606.00138v1 Announce Type: new Abstract: Finite element analysis (FEA) is the most important numerical approach for solid mechanics. Challenges of FEA include a steep learning curve for entry-l

agentsarxiv-cs-ai
2 Jun 2026
Model Releases

ACON: Optimizing Context Compression for Long-horizon LLM Agents

DGX agent

arXiv:2510.00615v3 Announce Type: replace Abstract: Large language models (LLMs) are increasingly deployed as agents in dynamic real-world environments, where success depends on maintaining precise re

model-releasesarxiv-cs-ai
2 Jun 2026
Safety

Adversarial Feeds Steer LLM Agent Decisions Against Their Defaults

DGX agent

arXiv:2606.00914v1 Announce Type: new Abstract: LLM agents increasingly act after consuming ranked external information streams such as social feeds, search results, retrieval contexts, and email queu

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

AGENTCL: Toward Rigorous Evaluation of Continual Learning in Language Agents

DGX agent

arXiv:2606.02461v1 Announce Type: new Abstract: Language agents spend substantial inference time solving individual tasks, yet the experience acquired in one episode is often underutilized in future e

model-releasesarxiv-cs-ai
2 Jun 2026
Agents

AMP: A Vendor-Neutral Wire Format for Agent Memory Operations

DGX agent

arXiv:2606.01138v1 Announce Type: cross Abstract: Agent-memory frameworks - mem0, Letta/MemGPT, Cognee, Zep/Graphiti, MemoryOS, MemTensor - each ship their own SDK, storage layout, and operational voc

agentsarxiv-cs-ai
2 Jun 2026
Model Releases

ASE-26: a curriculum for agentic software engineering as a discipline

DGX agent

arXiv:2606.01152v1 Announce Type: cross Abstract: The work of a professional software engineer has begun to consist, increasingly, of directing agents rather than writing code, and the empirical evide

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Benchmarking Security Risk Detection and Verification in Open Agentic Skill Ecosystems

DGX agent

arXiv:2606.00925v1 Announce Type: cross Abstract: Open agent platforms allow community contributors to publish reusable skills that agents can invoke at runtime. This extensibility also creates a supp

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Characterization of Multi-Model Agentic AI Systems on General Tasks via Trace-Driven Simulation

DGX agent

arXiv:2606.01725v1 Announce Type: new Abstract: Agentic AI completes tasks through iterative planning, tool use, and reasoning based on observed outcomes. Despite its popularity, its system-level beha

model-releasesarxiv-cs-ai
2 Jun 2026
Agents

🎉 Congratulations to @OdessiaTravel on their public launch! Odessia is an AI-powered travel agent that allows users to plan and book entire…

DGX agent

🎉 Congratulations to @OdessiaTravel on their public launch! Odessia is an AI-powered travel agent that allows users to plan and book entire trips in one conversation. The team used LangSmith and LangG

agentsharrison-chase--x
2 Jun 2026
Agents

Coordinating Task Switching in a Robotics Multi-Agent System Using Behavior Trees

DGX agent

arXiv:2606.01170v1 Announce Type: cross Abstract: The application of multi-agent systems in robotics is a very challenging field. Several competitions involving such systems are proposed to foster res

agentsarxiv-cs-ro
2 Jun 2026
Model Releases

CRAB-Bench: Evaluating LLM Agents under Complex Task Dependencies and Human-aligned User Simulation

DGX agent

arXiv:2606.01815v1 Announce Type: new Abstract: Evaluating LLM agents in realistic service scenarios requires complex task dependencies, imperfect user behavior, and an evaluation that accommodates mu

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

datasette-agent-micropython 0.1a0

DGX agent

Release: datasette-agent-micropython 0.1a0 I want Datasette Agent to be able to generate and execute Python code safely. This alpha is looking promising so far. GPT-5.5 has so far failed to break out

model-releasessimon-willison
2 Jun 2026
Agents

GitHub's plan for Agents — Kyle Daigle, GitHub

DGX agent

GitHub is developing AI agents to automate software development workflows and enhance developer productivity. The discussion likely covers GitHub's strategy for integrating autonomous AI capabilities

agentslatent-space
2 Jun 2026
Agents

🚀 Go from prototype to production in one click. We added a new Deploy button to LangSmith Studio, so you can deploy your agent directly to …

DGX agent

LangSmith Studio added a new Deploy button feature that enables users to deploy agents directly to production with a single click, streamlining the workflow from prototype development to production de

agentsharrison-chase--x
2 Jun 2026
Agents

Grok Build is genuinely amazing right now. It is not just another coding assistant. It is a full agentic system that can plan, write, refact…

DGX agent

Grok Build is genuinely amazing right now. It is not just another coding assistant. It is a full agentic system that can plan, write, refactor, debug, and build complete projects autonomously from a s

agentselon-musk--x
2 Jun 2026
Agents

How Baz improved its AI Agent Code Review accuracy using Amazon Bedrock AgentCore

DGX agent

This post walks through how Baz built their Spec Review agent using Amazon Bedrock and Amazon Bedrock AgentCore. We'll cover the architecture decisions, implementation details, and the business outcom

agentsaws-ml-blog
2 Jun 2026
Model Releases

'I Strongly Suspect This Website Is a Scam': Benchmarking PII Leakage and Detection without Defense in Autonomous Web Agents

DGX agent

arXiv:2606.00497v1 Announce Type: cross Abstract: Deceptive web content, widely instantiated across the internet and commonly known as extit{social-engineering attacks}, manipulates autonomous web age

model-releasesarxiv-cs-cl
2 Jun 2026
Agents

II-Agent is now live on the App Store. Your sovereign AI workspace for building, researching, writing, designing, and automating from one in…

DGX agent

II-Agent is now live on the App Store. Your sovereign AI workspace for building, researching, writing, designing, and automating from one intelligent interface. Download it. Bring your own key. Build

agentsemad-mostaque--x
2 Jun 2026
Safety

Joint Agent Memory and Exploration Learning via Novelty Signals

DGX agent

arXiv:2606.01528v1 Announce Type: new Abstract: In open-ended environments, exploration is fundamental for autonomous agents, yet current language model agents struggle with this. Effective exploratio

safetyarxiv-cs-ai
2 Jun 2026
Agents

MemGraphRAG: Memory-based Multi-Agent System for Graph Retrieval-Augmented Generation

DGX agent

arXiv:2606.00610v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) has become an essential method for mitigating hallucinations in Large Language Models (LLMs) by leveraging extern

agentsarxiv-cs-ai
2 Jun 2026
Agents

Multi-Agent Conformal Prediction with Personalized Statistical Validity

DGX agent

arXiv:2606.00717v1 Announce Type: cross Abstract: Uncertainty quantification is essential in high-stakes machine learning tasks. However, one of the principled solutions, conformal prediction, faces c

agentsarxiv-cs-ai
2 Jun 2026
Agents

OctoT2I: A Self-Evolving Agentic Text-to-Image Router

DGX agent

arXiv:2606.01803v1 Announce Type: new Abstract: The explosive growth of Text-to-Image (T2I) models, from large-scale versions to lightweight, real-time ones, now faces diminishing marginal returns fro

agentsarxiv-cs-ai
2 Jun 2026
Model Releases

Prompting agents to 'do better' is unreliable 🙅 Giving them a rubric, a grader, and a correction loop is much closer to how you get your ag…

DGX agent

Prompting agents to 'do better' is unreliable 🙅 Giving them a rubric, a grader, and a correction loop is much closer to how you get your agent do what you want! Similar to /goal in Claude Code or othe

model-releasesharrison-chase--x
2 Jun 2026
Agents

Read more about hybrid agentic inference in Perplexity Computer: https://www.perplexity.ai/hub/blog/the-data-center-moves-to-your-machine

DGX agent

Perplexity discusses hybrid agentic inference in Perplexity Computer, which likely describes a computational approach that combines processing between local machines and data centers. The concept sugg

agentsperplexity--x
2 Jun 2026
Agents

RocketSmith: An Agentic System for High-Powered Rocket Design and Manufacturing

DGX agent

arXiv:2606.00097v1 Announce Type: new Abstract: This work presents RocketSmith, an agentic system capable of the design, manufacturing, and optimization processes in high powered rocket development. T

agentsarxiv-cs-ro
2 Jun 2026
Agents

SciAgentGym: Benchmarking Multi-Step Scientific Tool-use in LLM Agents

DGX agent

arXiv:2602.12984v2 Announce Type: replace Abstract: Scientific reasoning inherently demands integrating sophisticated toolkits to navigate domain-specific knowledge. Yet, current benchmarks largely ov

agentsarxiv-cs-cl
2 Jun 2026
Agents

SkillSmith: Co-Evolving Skills and Tools for Self-Improving Agent Systems

DGX agent

arXiv:2606.01314v1 Announce Type: new Abstract: Recent self-evolving agents have shown that skills can be discovered, refined, and accumulated through execution. However, existing skill-evolution fram

agentsarxiv-cs-ai
2 Jun 2026
Agents

The next evolution of Hermes Agent is here! Introducing Hermes Desktop: everything you love about Hermes, now native on your machine. First …

DGX agent

The next evolution of Hermes Agent is here! Introducing Hermes Desktop: everything you love about Hermes, now native on your machine. First demoed in Jensen's GTC keynote, it's now in public preview.

agentsnous-research--x
2 Jun 2026
Model Releases

When Safe Skills Collide: Measuring Compositional Risk in Agent Skill Ecosystems

DGX agent

arXiv:2606.00448v1 Announce Type: cross Abstract: LLM agents increasingly rely on community-contributed skills that expand an agent's operational capability set. We study a core safety problem in agen

model-releasesarxiv-cs-ai
2 Jun 2026
Agents

AutoSci: A Memory-Centric Agentic System for the Full Scientific Research Lifecycle

DGX agent

arXiv:2605.31468v1 Announce Type: new Abstract: Scientific research has traditionally been human-intensive, requiring researchers to coordinate literature, ideas, experiments, manuscripts, and review

agentsarxiv-cs-ai
1 Jun 2026
Model Releases

EMBGuard: Constructing Hazard-Aware Guardrails for Safe Planning in Embodied Agents

DGX agent

arXiv:2605.30924v1 Announce Type: new Abstract: MLLM-powered embodied agents deployed in real-world environments encounter physical hazards. However, existing approaches lack explicit mechanisms for i

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

Emergent Languages in Populations of Language Model Agents: From Token Efficiency to Oversight Evasion

DGX agent

arXiv:2605.31170v1 Announce Type: cross Abstract: Monitoring autonomous language model agents currently relies mostly on surface behavior. But what happens when agent populations invent new languages

model-releasesarxiv-cs-ai
1 Jun 2026
Research

Enhancing Human-Likeness in Reinforcement Learning Agents via Hierarchical Macro Action Quantization

DGX agent

arXiv:2605.30928v1 Announce Type: new Abstract: Human-like agents are a long-standing goal of artificial intelligence. Despite strong performance, most reinforcement learning (RL) agents remain reward

researcharxiv-cs-ro
1 Jun 2026
Model Releases

Extending AI for Research to the Humanities: A Multi-Agent Framework for Evidence-Grounded Scholarship

DGX agent

arXiv:2605.30947v1 Announce Type: new Abstract: LLM-based research agents have advanced rapidly in science and engineering, where research is organized around executable experiments, code, and quantit

model-releasesarxiv-cs-cl
1 Jun 2026
Agents

MiniMax M3 imminent. Will be doing deep testing with it on my own coding agent and harness. Review coming soon.

DGX agent

MiniMax M3, an upcoming AI model, is expected to be released soon and will undergo comprehensive testing within a custom coding agent framework. A detailed technical review of the model's performance

agentsdair-ai--x
1 Jun 2026
Agents

.@MukilLoganathan’s Interrupt keynote on Sandboxes. https://youtu.be/IIchUA5T3gs In 20 minutes, you’ll learn how to run agent code safely. I…

DGX agent

.@MukilLoganathan’s Interrupt keynote on Sandboxes. https://youtu.be/IIchUA5T3gs In 20 minutes, you’ll learn how to run agent code safely. Isolated from your runtime, with network controls, persistent

agentsharrison-chase--x
1 Jun 2026
Agents

NEMO: Execution-Aware Optimization Modeling via Autonomous Coding Agents

DGX agent

arXiv:2601.21372v2 Announce Type: replace Abstract: We present NEMO, a system that translates Natural-language descriptions of decision problems into formal Executable Mathematical Optimization implem

agentsarxiv-cs-ai
1 Jun 2026
Hardware

PithTrain: A Compact and Agent-Native MoE Training System

DGX agent

arXiv:2605.31463v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) has become the dominant architecture for frontier language models. To meet this demand, production frameworks have built opti

hardwarearxiv-cs-ai
1 Jun 2026
Agents

.@Rippling AI runs on Deep Agents and LangSmith. Here’s how they shipped to millions of users in 6 months. https://www.langchain.com/blog/ho…

DGX agent

.@Rippling AI runs on Deep Agents and LangSmith. Here’s how they shipped to millions of users in 6 months. https://www.langchain.com/blog/how-rippling-went-ai-native-across-every-product-in-6-months-w

agentsharrison-chase--x
1 Jun 2026
Safety

Seeing Before Agreeing: Aligning Multi-Agent Consensus with Visual Evidence

DGX agent

arXiv:2605.30698v1 Announce Type: cross Abstract: Vision-language models (VLMs) have achieved strong performance on visual question answering (VQA). To mitigate individual hallucinations and blind spo

safetyarxiv-cs-ai
1 Jun 2026
Agents

Sources: Tencent, which has fallen behind domestic rivals in AI models, plans to test an AI agent for WeChat with a small group of users before a phased rollout (Zijing Wu/Financial Times)

DGX agent

Zijing Wu / Financial Times: Sources: Tencent, which has fallen behind domestic rivals in AI models, plans to test an AI agent for WeChat with a small group of users before a phased rollout — Maker of

agentstechmeme
1 Jun 2026
Agents

We're trending on @huggingface! 🥳 Tbh, we undersold this model. It's a lot more capable at agentic tasks than I expected. I keep discoverin…

DGX agent

We're trending on @huggingface! 🥳 Tbh, we undersold this model. It's a lot more capable at agentic tasks than I expected. I keep discovering new capabilities every day, it's crazy for 1B active parame

agentsclem-delangue--x
1 Jun 2026
Agents

Found a way to save everyone 14% on input tokens on average during read file operations in Hermes Agent! This is now on main. `hermes update…

DGX agent

Nous Research has optimized token efficiency in their Hermes Agent, achieving an average 14% reduction in input token usage during file read operations. This optimization has been merged to the main c

agentsnous-research--x
30 May 2026
Agents

Agent4Edu: Generating Learner Response Data by Generative Agents for Intelligent Education Systems

DGX agent

arXiv:2501.10332v2 Announce Type: replace-cross Abstract: Personalized learning represents a promising educational strategy within intelligent educational systems, aiming to enhance learners' practice

agentsarxiv-cs-ai
29 May 2026
Model Releases

AgentDropoutV2: Optimizing Information Flow in Multi-Agent Systems via Test-Time Rectify-or-Reject Pruning

DGX agent

arXiv:2602.23258v2 Announce Type: replace Abstract: While Multi-Agent Systems (MAS) excel in complex reasoning, they suffer from the cascading impact of erroneous information from individual agents. C

model-releasesarxiv-cs-ai
29 May 2026
Agents

AnomalyAgent: Training-Free Agentic Models for Zero-/Few-Shot Anomaly Detection

DGX agent

arXiv:2605.30140v1 Announce Type: new Abstract: Benefiting from generalizability of vision-language models (VLMs) such as CLIP, many zero-/few-shot anomaly detection (AD) approaches have achieved impr

agentsarxiv-cs-cv
29 May 2026
Model Releases

Cloud CISO Perspectives: How to build an AI-ready security program for the public sector

DGX agent

Welcome to the second Cloud CISO Perspectives for May 2026. Today, Usman Chaudhary, Field CISO, Google Public Sector, offers a guide for CISOs protecting government agencies and critical infrastructur

model-releasesgoogle-cloud-ai
29 May 2026
← Previous
1…103104105106107…375
Next →