AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

agents

GridTimelineEvolution
7,214 results
1 Jun 2026

The best eval harness for production AI and agents: A comparison

AgentsDGX agent

A practical comparison of production AI evaluation harnesses, including what to look for across instrumentation, evaluators, online evals, CI gates, and agent workflows. The post The best eval harness

This is great @hwchase17 @bryonkuchML Seeing this update, I’m building a tutorial repo around LangChain + Groq + GEPA. The idea is simple: •…

AgentsDGX agent

This is great @hwchase17 @bryonkuchML Seeing this update, I’m building a tutorial repo around LangChain + Groq + GEPA. The idea is simple: •LangChain builds the RAG/agent workflow •Groq gives fast inf

This macroeconomic research agent powered by Deep Agents, LangSmith, and the @youdotcom Finance Research API: ✅ Analyzes GDP data ✅ Detects …

AgentsDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

This macroeconomic research agent powered by Deep Agents, LangSmith, and the @youdotcom Finance Research API: ✅ Analyzes GDP data ✅ Detects anomalies ✅ Investigates structural & cyclical drivers at th

This one doesn't fit in a wave either! See you tomorrow

AgentsDGX agent

This post from Cognition AI's Windsurf account appears to be a casual, informal message likely referencing a specific project, feature, or update that doesn't conform to an expected pattern or schedul

We're trending on @huggingface! 🥳 Tbh, we undersold this model. It's a lot more capable at agentic tasks than I expected. I keep discoverin…

AgentsDGX agent

We're trending on @huggingface! 🥳 Tbh, we undersold this model. It's a lot more capable at agentic tasks than I expected. I keep discovering new capabilities every day, it's crazy for 1B active parame

Why Video Agent models are next — Ethan He, xAI Grok Imagine

AgentsDGX agent

Video agent models represent the next frontier in AI by extending language model capabilities to process, understand, and act on video content in real-time, enabling autonomous agents to perceive and

@Windows More info: https://hermes-agent.nousresearch.com/docs/user-guide/windows-native

AgentsDGX agent

This entry likely covers Nous Research's documentation for Windows native integration or functionality within their Hermes agent system. The resource appears to be a user guide section explaining how

31 May 2026

🧑‍⚖️Evaluating Deep Agents with LangSmith on AWS Great deep dive blog with our friends at AWS on evaluating DeepAgents with LangSmith Cover…

AgentsDGX agent

🧑‍⚖️Evaluating Deep Agents with LangSmith on AWS Great deep dive blog with our friends at AWS on evaluating DeepAgents with LangSmith Covers datapoint and evaluator design for longer horizon agents ht

From reactive operations to autonomous infrastructure: What IT leaders must do next

AgentsDGX agent

As artificial intelligence agents begin to proliferate across information technology infrastructure, IT leaders are moving away from asking, “How do we monitor every alert?” to “How do we design infra

Great article on harness engineering. https://www.langchain.com/blog/the-anatomy-of-an-agent-harness

AgentsDGX agent

This article from LangChain explores the architectural components and design principles of agent harnesses, which are systems that manage the execution and behavior of AI agents. The piece likely cove

It's a big week for Hermes Agent.

AgentsDGX agent

Nous Research announced significant developments or milestones for Hermes Agent, their AI agent framework, indicating multiple new capabilities, updates, or releases during that particular week. The a

love seeing tldraw's canvas sdk get absolutely maxed by replit team. Proud infra dad

AgentsDGX agent

love seeing tldraw's canvas sdk get absolutely maxed by replit team. Proud infra dad The best design work doesn't happen in a chat box. You need space to explore ideas, create variants, and iterate Me

[New Technical Blog Post] In our second longmemeval experiment, we introduce semantic ingestion into recall leveraging the ActiveGraph runti…

AgentsDGX agent

[New Technical Blog Post] In our second longmemeval experiment, we introduce semantic ingestion into recall leveraging the ActiveGraph runtime. We started from a 60.6% baseline and improved to 83.4%/8

😂PewDiePie building his own agent orchestrator and releasing it was not on my 2026 bingo card. Own the agent. Own the harness. It's not tha…

AgentsDGX agent

PewDiePie has developed and released his own AI agent orchestrator, representing a trend toward independent creators building custom AI infrastructure rather than relying on existing platforms. The po

The Top AI Papers of the Week (May 24 - May 31) - SkillOpt - AutoScientists - The Efficiency Frontier - Language Models Need Sleep - Adaptin…

AgentsDGX agent

The Top AI Papers of the Week (May 24 - May 31) - SkillOpt - AutoScientists - The Efficiency Frontier - Language Models Need Sleep - Adapting the Interface, Not the Model - Forecasting Scientific Prog

Why ‘human in the loop’ falls short – and what to do about it

AgentsDGX agent

Agentic artificial intelligence governance depends upon humans to keep agentic AI from going off the rails. However, putting humans in the loop is woefully insufficient. Here are the problems – and pe

30 May 2026

1/10 - PDF parsing at browser speed Jerry Liu showed LiteParse v2 using Rust to WebAssembly for sub-second extraction of messy PDFs, without…

AgentsDGX agent

1/10 - PDF parsing at browser speed Jerry Liu showed LiteParse v2 using Rust to WebAssembly for sub-second extraction of messy PDFs, without calling a model. It can sit as a default step in agent and

everyone from openai like my tweet

AgentsDGX agent

Jerry Liu shared a tweet that received engagement from multiple OpenAI team members, highlighting interest from the organization in his content or perspective. The tweet likely relates to developments

Found a way to save everyone 14% on input tokens on average during read file operations in Hermes Agent! This is now on main. `hermes update…

AgentsDGX agent

Nous Research has optimized token efficiency in their Hermes Agent, achieving an average 14% reduction in input token usage during file read operations. This optimization has been merged to the main c

In a few months, people will start to realize how fundamentally important MCP for agents is. It's not even about connecting tools. There are…

AgentsDGX agent

In a few months, people will start to realize how fundamentally important MCP for agents is. It's not even about connecting tools. There are many ways to do that. It's about the types of abstraction i

Increasingly, HTML Artifacts are becoming a core part of how I work with AI agents. Long-horizon agent sessions need a better way to surface…

AgentsDGX agent

Increasingly, HTML Artifacts are becoming a core part of how I work with AI agents. Long-horizon agent sessions need a better way to surface insights about what work it has done. This may not be obvio

LangChain🤝GEPA shout out to @bryonkuchML for contributing a PR to the GEPA repo to make it work for LangChain! You can now optimize your La…

AgentsDGX agent

LangChain🤝GEPA shout out to @bryonkuchML for contributing a PR to the GEPA repo to make it work for LangChain! You can now optimize your LangChain chains Docs: https://gepa-ai.github.io/gepa/tutorials

Learn more about the latest from @james_y_zou and our Frontier Agents Research team!

AgentsDGX agent

Learn more about the latest from @james_y_zou and our Frontier Agents Research team! To evaluate frontier AI agents, we need more complex tasks. But such tasks are also more prone to have design mista

Step 3.7 Flash, free for 30 days for Hermes Agent users. What could possibly go wrong? 🍿 Thanks @NousResearch for making it happen. Can’t w…

AgentsDGX agent

Step 3.7 Flash, free for 30 days for Hermes Agent users. What could possibly go wrong? 🍿 Thanks @NousResearch for making it happen. Can’t wait to see what Hermes users build! Step 3.7 Flash is now fre

Step 3.7 Flash is now free for 30 days via Nous Portal It is a new MoE vision-language model focused on agent efficiency, coding, search, an…

AgentsDGX agent

Step 3.7 Flash is now free for 30 days via Nous Portal It is a new MoE vision-language model focused on agent efficiency, coding, search, and multimodal workflows — and Hermes Agent users have been lo

the market is speaking

AgentsDGX agent

the market is speaking The latest finding in the LangSmith Signal: Open Models are having a moment. 1 in 3 AI teams ran an open-weights model in April 2026, up from 1 in 5 nine months ago. The overall

The secret to LiteParse lies in the grid projection algorithm. We project a complex page layout with text and tables into well-structured te…

AgentsDGX agent

The secret to LiteParse lies in the grid projection algorithm. We project a complex page layout with text and tables into well-structured text, that humans can read and agents can understanding. This

This recent @latentspacepod pod was a good one (they usually all are). One specific piece resonated: They were talking about AI agents writi…

AgentsDGX agent

This recent @latentspacepod pod was a good one (they usually all are). One specific piece resonated: They were talking about AI agents writing code, and the line was basically that without caution / d

29 May 2026

A Deep Learning Model of Mental Rotation Informed by Interactive VR Experiments

AgentsDGX agent

arXiv:2512.13517v2 Announce Type: replace-cross Abstract: Mental rotation -- the ability to compare objects seen from different viewpoints -- is a fundamental example of mental simulation and spatial

A Domain-Informed Multi-Objective Framework for EEG Channel Selection in Motor Imagery BCIs

AgentsDGX agent

arXiv:2605.29943v1 Announce Type: cross Abstract: Motor imagery (MI) classification using electroencephalography (EEG) signals is essential for advancing brain-computer interfaces (BCIs). Traditional

Adobe’s conversational AI agent is a mediocre design intern

AgentsDGX agent

AI image tools rarely make me feel like I'm part of the creative process. They are, after all, mostly designed so that people with no design experience can type in a few words and get back a usable re

Agent actions that aren't on your allowlist or can't be sandboxed go to a classifier subagent. This separate agent decides whether to allow …

AgentsDGX agent

Agent actions that aren't on your allowlist or can't be sandboxed go to a classifier subagent. This separate agent decides whether to allow the tool call, try a different approach, or ask you for appr

agent builder!

AgentsDGX agent

agent builder! Start creating agents using everyday language with LangSmith Fleet. Learn how to build no-code agents for real work. Take our free LangChain Academy course today: https://academy.langch

Agent4Edu: Generating Learner Response Data by Generative Agents for Intelligent Education Systems

AgentsDGX agent

arXiv:2501.10332v2 Announce Type: replace-cross Abstract: Personalized learning represents a promising educational strategy within intelligent educational systems, aiming to enhance learners' practice

Agentic AI success helps UiPath swing to a profit, but investors weren’t impressed

AgentsDGX agent

Business automation software company UiPath Inc. delivered mixed results in its latest quarter, posting a solid revenue beat but falling short on earnings — but it did at least manage to return to pro

AgentSchool: An LLM-Powered Multi-Agent Simulation for Education

AgentsDGX agent

arXiv:2605.30144v1 Announce Type: new Abstract: Despite the rapid deployment of LLMs into classrooms, validating educational AI remains uniquely intractable: interventions act on developing learners w

An Approach for Thyroid Nodule Analysis Using Thermographic Images

AgentsDGX agent

arXiv:2605.29221v1 Announce Type: new Abstract: Thyroid cancer is said to be the second most common type of cancer in female individuals and the third in males by 2030, according to projections. In ge

AnomalyAgent: Training-Free Agentic Models for Zero-/Few-Shot Anomaly Detection

AgentsDGX agent

arXiv:2605.30140v1 Announce Type: new Abstract: Benefiting from generalizability of vision-language models (VLMs) such as CLIP, many zero-/few-shot anomaly detection (AD) approaches have achieved impr

As always, 'hermes update' More info: https://hermes-agent.nousresearch.com/docs/user-guide/features/tool-search

AgentsDGX agent

Nous Research announced an update to Hermes, their AI agent framework, with details available in their user guide documentation covering tool-search features. The update likely enhances Hermes' capabi

Beyond Consensus: Trace-Level Synthesis in Mixture of Agents

AgentsDGX agent

arXiv:2605.29116v1 Announce Type: new Abstract: When multiple LLM agents solve the same problem, standard practice compresses each agent's reasoning into a majority vote or layered synthesis, treating

BitTP: The Lightweight Trajectory Prediction Model with BitLLM for Edge-Devices

AgentsDGX agent

arXiv:2605.29705v1 Announce Type: new Abstract: Trajectory prediction is a fundamental task for autonomous systems, requiring complex reasoning about multi-agent interactions and intents. Large langua

Bosses, Kings, and the Commons: Cooperation Under Power Asymmetry in LLM Societies

AgentsDGX agent

arXiv:2605.29062v1 Announce Type: new Abstract: Communities can sustainably manage shared resources (commons) through self-governance and cooperative norms, a central finding of Ostrom's theory of sel

Catalyst-Agent: Autonomous heterogeneous catalyst screening with an LLM Agent

AgentsDGX agent

arXiv:2603.01311v2 Announce Type: replace Abstract: The discovery of novel catalysts tailored for particular applications is a major challenge for the twenty-first century. Traditional methods for thi

Code-QA-Bench: Separating Code Reasoning from Documentation Memorization in Repository-Level QA

AgentsDGX agent

arXiv:2605.29277v1 Announce Type: cross Abstract: We present Code-QA-Bench, a fully automated framework for synthesizing repository-level code understanding benchmarks that separates genuine code comp

Compass: Navigating Global Marine Lead Data Integration through Expert-Guided LLM Agent

AgentsDGX agent

arXiv:2605.29966v1 Announce Type: new Abstract: Marine lead (Pb) and its isotopes are critical tracers for ocean circulation and anthropogenic pollution, yet in-situ observations remain costly and spa

CompilerDream: Learning a Compiler World Model for General Code Optimization

AgentsDGX agent

arXiv:2404.16077v4 Announce Type: replace-cross Abstract: Effective code optimization in compilers is crucial for computer and software engineering. The success of these optimizations primarily depend

CONCAT: Consensus- and Confidence-Driven Ad Hoc Teaming for Efficient LLM-Based Multi-Agent Systems

AgentsDGX agent

arXiv:2605.29612v1 Announce Type: cross Abstract: Although large language model (LLM) based multi-agent systems (MAS) show their capability to solve complex tasks and achieve higher performance over s

Croissant Tasks: A Metadata Format for Reproducible Machine Learning Evaluations

AgentsDGX agent

arXiv:2605.29786v1 Announce Type: new Abstract: Reproducibility is fundamental to the scientific method, yet remains a critical challenge in machine learning. Contributing factors include underspecifi

Designing Active Tether-Net Systems for Space Debris Capture with Graph-Learning-Aided Mixed-Combinatorial Optimization

AgentsDGX agent

arXiv:2605.29021v1 Announce Type: new Abstract: Active tether-net systems are a promising solution for capturing large non-cooperative targets, such as space debris, by deploying a flexible net manipu

DGSG-Mind: Dynamic 3D Gaussian Scene Graphs for Long-Term Scene Understanding and Grounding

AgentsDGX agent

arXiv:2605.29879v1 Announce Type: new Abstract: Integrating open-vocabulary semantic information into dynamic 3D scene representations is essential for long-term embodied scene understanding. However,

Dissociative Identity: Language Model Agents Lack Grounding for Reputation Mechanisms

AgentsDGX agent

arXiv:2605.30169v1 Announce Type: cross Abstract: As autonomous language model agents proliferate, forming an emerging agentic web with real-world consequences, what credibility signals can you use to

Distributed Non-Uniform Scaling Control of Multi-Agent Formation with Dynamic Agent Joining

AgentsDGX agent

arXiv:2605.29191v1 Announce Type: cross Abstract: Non-uniform scaling control of formation enables multi-agent systems to adjust their shape by scaling with different ratios along different coordinate

Does The Way You Plan Matter? An Empirical Study of Planning Representations for LLM Web Agents

AgentsDGX agent

arXiv:2605.29927v1 Announce Type: cross Abstract: Despite recent advances, LLM-based web agents still struggle with limited exploration, omission of critical steps, and sensitivity to task constraints

⏸️ Don’t pay for resources that aren’t doing anything. LangSmith Sandboxes pause automatically when idle.

AgentsDGX agent

LangSmith Sandboxes include an automatic pause feature that activates when the environment is idle, preventing users from incurring costs for unused resources. This functionality helps optimize cloud

E-valuator: Reliable Agent Verifiers with Sequential Hypothesis Testing

AgentsDGX agent

arXiv:2512.03109v2 Announce Type: replace-cross Abstract: Agentic AI systems execute a sequence of actions, such as reasoning steps or tool calls, in response to a user prompt. To evaluate the success

Enhancing Multi-Agent Communication through Attention Steering with Context Relevance

AgentsDGX agent

arXiv:2605.30136v1 Announce Type: new Abstract: LLM-based multi-agent systems have demonstrated remarkable performance on complex tasks through collaborative reasoning. However, these systems tend to

Envy-Free Allocation of Indivisible Goods via Noisy Queries

AgentsDGX agent

arXiv:2602.06361v2 Announce Type: replace-cross Abstract: We introduce a problem of fairly allocating indivisible goods (items) in which the agents' valuations cannot be observed directly, but instead

Error as a Lens: Probing LLM Reasoning through Synthetic Misconception Generation

AgentsDGX agent

arXiv:2605.29007v1 Announce Type: new Abstract: Personalized tutoring, teacher training, and education research need access to targeted synthetic misconceptions, but privacy and IRB constraints make l

Estimating the Empowerment of Language Model Agents

AgentsDGX agent

arXiv:2509.22504v3 Announce Type: replace Abstract: As language model (LM) agents become increasingly capable and adopted in real-world applications, there is a growing need for scalable evaluation fr

Evolve as a Team: Collaborative Self-Evolution for LLM-based Multi-Agent Systems

AgentsDGX agent

arXiv:2605.29790v1 Announce Type: cross Abstract: LLM-based multi-agent systems (MAS) have emerged as an effective paradigm for complex and long-horizon tasks. However, in real-world tasks, MAS often

← Previous
1…5455565758…121
Next →