AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
All
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

agents

GridTimelineEvolution
7,215 results
Agents

Let engine help you build better agents

DGX agent

This post likely discusses how the LangChain framework (which Harrison Chase co-founded) can assist developers in constructing more effective AI agents by providing tools, abstractions, and patterns f

agentsharrison-chase--x
1 Jun 2026
Agents

LH-Bench: Skill-Grounded Evaluation of Long-Horizon Agents on Subjective Enterprise Tasks

Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2603.22744v2 Announce Type: replace Abstract: Large language models excel on objectively verifiable tasks such as math and programming, where evaluation reduces to unit tests or a single correct

agentsarxiv-cs-ai
1 Jun 2026
Agents

LLM Anonymization Against Agentic Re-Identificatio

DGX agent

arXiv:2605.30848v1 Announce Type: cross Abstract: Agentic LLMs with web search change the threat model for text anonymization: weak contextual cues can become cross-referenceable evidence for re-ident

agentsarxiv-cs-cl
1 Jun 2026
Agents

longmemeval experiment arch: 1) deterministic ingestion/extraction (85.6% accuracy, 86.2% retrieval) 2) semantic ingestion/extraction (84.8%…

DGX agent

longmemeval experiment arch: 1) deterministic ingestion/extraction (85.6% accuracy, 86.2% retrieval) 2) semantic ingestion/extraction (84.8% accuracy, 94.9% retention) 3) semantic ingestion/determinis

agentsyohei-nakajima--x
1 Jun 2026
Agents

LongTraceRL: Learning Long-Context Reasoning from Search Agent Trajectories with Rubric Rewards

DGX agent

arXiv:2605.31584v1 Announce Type: cross Abstract: Long-context reasoning remains a central challenge for large language models, which often fail to locate and integrate key information in extensive di

agentsarxiv-cs-ai
1 Jun 2026
Agents

Managed Deep Agents keeps the project shape you already know: ↳ AGENTS.md, skills/, subagents/, + tools.json Context Hub gives your agent a …

DGX agent

Managed Deep Agents keeps the project shape you already know: ↳ AGENTS.md, skills/, subagents/, + tools.json Context Hub gives your agent a managed place to retain and update this context across sessi

agentsharrison-chase--x
1 Jun 2026
Agents

MASPOB: Bandit-Based Prompt Optimization for Multi-Agent Systems with Graph Neural Networks

DGX agent

arXiv:2603.02630v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have achieved great success in many real-world applications, especially the one serving as the cognitive backbone

agentsarxiv-cs-ai
1 Jun 2026
Agents

MatchFixAgent: Language-Agnostic Autonomous Repository-Level Code Translation Validation and Repair

DGX agent

arXiv:2509.16187v3 Announce Type: replace-cross Abstract: Code translation transforms source code from one programming language (PL) to another. Validating the functional equivalence of translation an

agentsarxiv-cs-lg
1 Jun 2026
Agents

May 2026 newsletter

DGX agent

I just sent out the May edition of my sponsors-only monthly newsletter. If you are a sponsor (or if you start a sponsorship now) you can access it here. This month: Al got expensive, and Anthropic had

agentssimon-willison
1 Jun 2026
Agents

MedCoG: Maximizing LLM Inference Density in Medical Reasoning via Meta-Cognitive Regulation

DGX agent

arXiv:2602.07905v2 Announce Type: replace Abstract: Large Language Models (LLMs) have shown strong potential in complex medical reasoning yet face diminishing gains under inference scaling laws. While

agentsarxiv-cs-ai
1 Jun 2026
Agents

MiniMax M3 imminent. Will be doing deep testing with it on my own coding agent and harness. Review coming soon.

DGX agent

MiniMax M3, an upcoming AI model, is expected to be released soon and will undergo comprehensive testing within a custom coding agent framework. A detailed technical review of the model's performance

agentsdair-ai--x
1 Jun 2026
Agents

MiniMax-M3 will by arrive on HuggingFace openweight at next week!

DGX agent

MiniMax-M3 will by arrive on HuggingFace openweight at next week! Introducing MiniMax M3: The First Open-Weights Model to Combine Three Frontier Capabilities - Coding & Agentic Frontier: 59.0% SWE-Ben

agentsclem-delangue--x
1 Jun 2026
Agents

More info about Search as Code in the Perplexity Agent API docs: https://docs.perplexity.ai/docs/agent-api/tools/sandbox

DGX agent

The Perplexity Agent API documentation includes a 'Search as Code' feature accessible through the sandbox tools section, enabling developers to integrate search functionality programmatically within a

agentsperplexity--x
1 Jun 2026
Agents

More on the Hermes Skills Hub: https://hermes-agent.nousresearch.com/docs/guides/work-with-skills#the-skills-hub

DGX agent

The Hermes Skills Hub is a feature that allows users to discover, manage, and integrate skills within the Hermes agent framework, enabling extended functionality and customization of agent capabilitie

agentsnous-research--x
1 Jun 2026
Agents

.@MukilLoganathan’s Interrupt keynote on Sandboxes. https://youtu.be/IIchUA5T3gs In 20 minutes, you’ll learn how to run agent code safely. I…

DGX agent

.@MukilLoganathan’s Interrupt keynote on Sandboxes. https://youtu.be/IIchUA5T3gs In 20 minutes, you’ll learn how to run agent code safely. Isolated from your runtime, with network controls, persistent

agentsharrison-chase--x
1 Jun 2026
Agents

Multi-Turn Multi-Agent Dialogue for Collaborative Reconstruction Improves VLM Performance on Spatial Reasoning, But Only Barely

DGX agent

arXiv:2605.31387v1 Announce Type: new Abstract: Robots operating in diverse environments rely on visual input to interpret objects and spatial layouts. In human-collaborative tasks, they are expected

agentsarxiv-cs-cl
1 Jun 2026
Agents

NEMO: Execution-Aware Optimization Modeling via Autonomous Coding Agents

DGX agent

arXiv:2601.21372v2 Announce Type: replace Abstract: We present NEMO, a system that translates Natural-language descriptions of decision problems into formal Executable Mathematical Optimization implem

agentsarxiv-cs-ai
1 Jun 2026
Agents

NTR: Neural Token Reconstruction for Scene Token Bottleneck in End-to-End Driving

DGX agent

arXiv:2605.31116v1 Announce Type: new Abstract: Recent perception-free end-to-end (E2E) autonomous driving methods bypass explicit perception outputs by compressing dense image patch tokens into compa

agentsarxiv-cs-cv
1 Jun 2026
Agents

Open models!

DGX agent

Open models! Introducing MiniMax M3: The First Open-Weights Model to Combine Three Frontier Capabilities - Coding & Agentic Frontier: 59.0% SWE-Bench Pro, 66.0% Terminal Bench 2.1, 34.8% SWE-fficiency

agentsharrison-chase--x
1 Jun 2026
Agents

PictSure: Pretraining Embeddings Matters for In-Context Learning Image Classifiers

DGX agent

arXiv:2506.14842v2 Announce Type: replace-cross Abstract: Building image classification models remains cumbersome in data-scarce domains, where collecting large labeled datasets is impractical. In-con

agentsarxiv-cs-ai
1 Jun 2026
Agents

Provably Convergent Actor-Critic for MARL through Risk-aversion

DGX agent

arXiv:2602.12386v2 Announce Type: replace-cross Abstract: Learning stationary policies in infinite-horizon general-sum Markov games (MGs) remains a fundamental open problem in Multi-Agent Reinforcemen

agentsarxiv-cs-lg
1 Jun 2026
Agents

Pull Requests as a Training Signal for Repo-Level Code Editing

DGX agent

arXiv:2602.07457v2 Announce Type: replace-cross Abstract: Repository-level code editing requires models to understand complex dependencies and execute precise multi-file modifications across a large c

agentsarxiv-cs-ai
1 Jun 2026
Agents

.@Rippling AI runs on Deep Agents and LangSmith. Here’s how they shipped to millions of users in 6 months. https://www.langchain.com/blog/ho…

DGX agent

.@Rippling AI runs on Deep Agents and LangSmith. Here’s how they shipped to millions of users in 6 months. https://www.langchain.com/blog/how-rippling-went-ai-native-across-every-product-in-6-months-w

agentsharrison-chase--x
1 Jun 2026
Agents

RT @jayfarei: I think this idea a lot!

DGX agent

I cannot provide an accurate summary of this post as the tweet content itself is not provided, only a reference indicating Jay Farei's engagement with an idea shared by Harrison Chase. Without access

agentsharrison-chase--x
1 Jun 2026
Agents

S^3LDBO: A Snapshot Single-Loop Algorithm for Decentralized Bilevel Optimization

DGX agent

arXiv:2605.31311v1 Announce Type: cross Abstract: Networked AI systems increasingly rely on multiple agents that collaboratively learn and adapt models over communication networks. In such systems, bi

agentsarxiv-cs-lg
1 Jun 2026
Agents

SAGE: A Novelty Gate for Efficient Memory Evolution in Agentic LLMs

DGX agent

arXiv:2605.30711v1 Announce Type: cross Abstract: Agentic LLMs must continuously decide whether newly extracted facts should be added, merged with existing memories, or ignored, yet prior work has foc

agentsarxiv-cs-ai
1 Jun 2026
Agents

Sent out the May edition of my sponsors-only newsletter, for people who don't have time to read my blog every day and want to pay me money t…

DGX agent

Simon Willison announced the release of a May edition of his sponsors-only newsletter, which is designed for supporters who prefer a curated summary format instead of following his daily blog posts. T

agentssimon-willison--x
1 Jun 2026
Agents

Skill Reuse as Compression in Agentic RL

DGX agent

arXiv:2605.31509v1 Announce Type: cross Abstract: Large language model agents trained with reinforcement learning (RL) often learn brittle, task-specific shortcuts. We hypothesize that agents generali

agentsarxiv-cs-ai
1 Jun 2026
Agents

slowly we're all realizing that tools should be called from code, not from within the llm api

DGX agent

slowly we're all realizing that tools should be called from code, not from within the llm api Introducing Search as Code, our new search architecture for AI agents. It writes Python that calls our sea

agentsyohei-nakajima--x
1 Jun 2026
Agents

Social Reasoning in Machines: Investigating Collective Truth-Seeking Dynamics in Large Language Model Debate

DGX agent

arXiv:2605.30391v1 Announce Type: cross Abstract: Human reasoning has long been theorised to operate socially, not through isolated individual cognition, but through collective adversarial discourse,

agentsarxiv-cs-ai
1 Jun 2026
Agents

Sophrosyne: Agentic Exploration of Relational Data Systems Needs Moderation

DGX agent

arXiv:2605.30862v1 Announce Type: cross Abstract: Text2SQL agents powered by LLMs translate natural language intent into SQL by exploring the data system through tool calls before formulating the quer

agentsarxiv-cs-ai
1 Jun 2026
Agents

Sources: Tencent, which has fallen behind domestic rivals in AI models, plans to test an AI agent for WeChat with a small group of users before a phased rollout (Zijing Wu/Financial Times)

DGX agent

Zijing Wu / Financial Times: Sources: Tencent, which has fallen behind domestic rivals in AI models, plans to test an AI agent for WeChat with a small group of users before a phased rollout — Maker of

agentstechmeme
1 Jun 2026
Agents

SpecDB: LLM-Generated Customized Databases via Feature-Oriented Decomposition

DGX agent

arXiv:2605.31097v1 Announce Type: cross Abstract: Mainstream relational databases ship a uniform feature set across deployments, although individual workloads exercise only a fraction of the available

agentsarxiv-cs-ai
1 Jun 2026
Agents

Stop manually triaging agent failures. Let LangSmith Engine fix it.

DGX agent

LangSmith Engine is a tool designed to automatically diagnose and resolve agent failures, eliminating the need for manual troubleshooting and triage. The feature appears to leverage automated analysis

agentsharrison-chase--x
1 Jun 2026
Agents

Subspace-Decomposed JEPAs: Disentangling Progression and Content in Latent World Models

DGX agent

arXiv:2605.31111v1 Announce Type: new Abstract: Joint-Embedding Predictive Architectures (JEPAs) learn compact latent world models by predicting future embeddings, but no single coordinate of the late

agentsarxiv-cs-lg
1 Jun 2026
Agents

Surprised by Attention: Predictable Query Dynamics for Time Series Anomaly Detection

DGX agent

arXiv:2603.12916v3 Announce Type: replace-cross Abstract: Multivariate time series anomalies often manifest as shifts in cross-channel dependencies rather than simple amplitude excursions. In autonomo

agentsarxiv-cs-ai
1 Jun 2026
Agents

Survival Reinforcement Learning: Toward Scalable Self-Supervised RL

DGX agent

arXiv:2605.31273v1 Announce Type: new Abstract: While self-supervised Contrastive Reinforcement Learning (CRL) has shown remarkable depth-scaling capabilities, successfully using networks over 64 laye

agentsarxiv-cs-lg
1 Jun 2026
Agents

The best eval harness for production AI and agents: A comparison

DGX agent

A practical comparison of production AI evaluation harnesses, including what to look for across instrumentation, evaluators, online evals, CI gates, and agent workflows. The post The best eval harness

agentsarize-ai
1 Jun 2026
Agents

This is great @hwchase17 @bryonkuchML Seeing this update, I’m building a tutorial repo around LangChain + Groq + GEPA. The idea is simple: •…

DGX agent

This is great @hwchase17 @bryonkuchML Seeing this update, I’m building a tutorial repo around LangChain + Groq + GEPA. The idea is simple: •LangChain builds the RAG/agent workflow •Groq gives fast inf

agentsharrison-chase--x
1 Jun 2026
Agents

This macroeconomic research agent powered by Deep Agents, LangSmith, and the @youdotcom Finance Research API: ✅ Analyzes GDP data ✅ Detects …

DGX agent

This macroeconomic research agent powered by Deep Agents, LangSmith, and the @youdotcom Finance Research API: ✅ Analyzes GDP data ✅ Detects anomalies ✅ Investigates structural & cyclical drivers at th

agentsharrison-chase--x
1 Jun 2026
Agents

This one doesn't fit in a wave either! See you tomorrow

DGX agent

This post from Cognition AI's Windsurf account appears to be a casual, informal message likely referencing a specific project, feature, or update that doesn't conform to an expected pattern or schedul

agentscognition-ai--x
1 Jun 2026
Agents

We're trending on @huggingface! 🥳 Tbh, we undersold this model. It's a lot more capable at agentic tasks than I expected. I keep discoverin…

DGX agent

We're trending on @huggingface! 🥳 Tbh, we undersold this model. It's a lot more capable at agentic tasks than I expected. I keep discovering new capabilities every day, it's crazy for 1B active parame

agentsclem-delangue--x
1 Jun 2026
Agents

Why Video Agent models are next — Ethan He, xAI Grok Imagine

DGX agent

Video agent models represent the next frontier in AI by extending language model capabilities to process, understand, and act on video content in real-time, enabling autonomous agents to perceive and

agentslatent-space
1 Jun 2026
Agents

@Windows More info: https://hermes-agent.nousresearch.com/docs/user-guide/windows-native

DGX agent

This entry likely covers Nous Research's documentation for Windows native integration or functionality within their Hermes agent system. The resource appears to be a user guide section explaining how

agentsnous-research--x
1 Jun 2026
Agents

🧑‍⚖️Evaluating Deep Agents with LangSmith on AWS Great deep dive blog with our friends at AWS on evaluating DeepAgents with LangSmith Cover…

DGX agent

🧑‍⚖️Evaluating Deep Agents with LangSmith on AWS Great deep dive blog with our friends at AWS on evaluating DeepAgents with LangSmith Covers datapoint and evaluator design for longer horizon agents ht

agentsharrison-chase--x
31 May 2026
Agents

From reactive operations to autonomous infrastructure: What IT leaders must do next

DGX agent

As artificial intelligence agents begin to proliferate across information technology infrastructure, IT leaders are moving away from asking, “How do we monitor every alert?” to “How do we design infra

agentssiliconangle
31 May 2026
Agents

Great article on harness engineering. https://www.langchain.com/blog/the-anatomy-of-an-agent-harness

DGX agent

This article from LangChain explores the architectural components and design principles of agent harnesses, which are systems that manage the execution and behavior of AI agents. The piece likely cove

agentsharrison-chase--x
31 May 2026
Agents

It's a big week for Hermes Agent.

DGX agent

Nous Research announced significant developments or milestones for Hermes Agent, their AI agent framework, indicating multiple new capabilities, updates, or releases during that particular week. The a

agentsnous-research--x
31 May 2026
← Previous
1…6768697071…151
Next →