AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,964 results
Model Releases

InfantAgent-Next: A Multimodal Generalist Agent for Automated Computer Interaction

DGX agent

arXiv:2505.10887v3 Announce Type: replace Abstract: This paper introduces extsc{InfantAgent-Next}, a generalist agent capable of interacting with computers in a multimodal manner, encompassing text, i

model-releasesarxiv-cs-ai
5 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Agents

LinkAnchor: An Autonomous LLM-Based Agent for Issue-to-Commit Link Recovery

DGX agent

arXiv:2508.12232v3 Announce Type: replace-cross Abstract: Issue-to-commit link recovery in software repositories is fundamental to software traceability and project management, yet it remains a challe

agentsarxiv-cs-ai
5 May 2026
Model Releases

MOSAIC: Multi-agent Orchestration for Task-Intelligent Scientific Coding

DGX agent

arXiv:2510.08804v3 Announce Type: replace Abstract: We present MOSAIC, a multi-agent Large Language Model (LLM) framework for solving challenging scientific coding tasks. Unlike general-purpose coding

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Planner Matters! An Efficient and Unbalanced Multi-agent Collaboration Framework for Long-horizon Planning

DGX agent

arXiv:2605.02168v1 Announce Type: cross Abstract: Language model (LM)-based agents have demonstrated promising capabilities in automating complex tasks from natural language instructions, yet they con

model-releasesarxiv-cs-lg
5 May 2026
Safety

Who Decides What Is Harmful? Content Moderation Policy Through A Multi-Agent Personalised Inference Framework

DGX agent

arXiv:2605.01416v1 Announce Type: cross Abstract: The increasing scale and complexity of online platforms raises critical policy questions around harmful content, digital well-being, and user autonomy

safetyarxiv-cs-cl
5 May 2026
Tools

Pinecone Nexus: The Knowledge Engine for Agents

DGX agent

Pinecone Nexus is a knowledge infrastructure platform designed to enable AI agents to access, retrieve, and reason over enterprise data at scale. The system addresses the challenge of integrating larg

toolspinecone
4 May 2026
Agents

[AINews] AI Engineer World's Fair — Autoresearch, Memory, World Models, Tokenmaxxing, Agentic Commerce, and Vertical AI Call for Speakers

DGX agent

The AI Engineer World's Fair is a conference event featuring talks on emerging AI research and engineering topics including autoresearch systems, advanced memory architectures, world models, tokenizat

agentslatent-space
2 May 2026
Agents

AgentEconomist: An End-to-end Agentic System Translating Economic Intuitions into Executable Computational Experiments

DGX agent

arXiv:2604.27725v1 Announce Type: cross Abstract: A long-standing challenge in economics lies not in the lack of intuition, but in the difficulty of translating intuitive insights into verifiable rese

agentsarxiv-cs-ai
1 May 2026
Applications

Context as Prior: Bayesian-Inspired Intent Inference for Non-Speaking Agents with a Household Cat Testbed

DGX agent

arXiv:2604.27445v1 Announce Type: new Abstract: Many agents in real-world environments cannot reliably communicate their goals through language, including household pets, pre-verbal infants, and other

applicationsarxiv-cs-cv
1 May 2026
Model Releases

DeepTutor: Towards Agentic Personalized Tutoring

DGX agent

arXiv:2604.26962v1 Announce Type: cross Abstract: Education represents one of the most promising real-world applications for Large Language Models (LLMs). However, conventional tutoring systems rely o

model-releasesarxiv-cs-ai
1 May 2026
Safety

Detecting Clinical Discrepancies in Health Coaching Agents: A Dual-Stream Memory and Reconciliation Architecture

DGX agent

arXiv:2604.27045v1 Announce Type: cross Abstract: As Large Language Model (LLM) agents transition from single-session tools to persistent systems managing longitudinal healthcare journeys, their memor

safetyarxiv-cs-ai
1 May 2026
Agents

End-to-End Evaluation and Governance of an EHR-Embedded AI Agent for Clinicians

DGX agent

arXiv:2604.27309v1 Announce Type: new Abstract: Clinical AI systems require not just point-in-time evaluation but continuous governance: the ongoing practice of monitoring, evaluating, iterating, and

agentsarxiv-cs-ai
1 May 2026
Agents

Heterogeneous Scientific Foundation Model Collaboration

DGX agent

arXiv:2604.27351v1 Announce Type: new Abstract: Agentic large language model systems have demonstrated strong capabilities. However, their reliance on language as the universal interface fundamentally

agentsarxiv-cs-ai
1 May 2026
Safety

Learning When to Remember: Risk-Sensitive Contextual Bandits for Abstention-Aware Memory Retrieval in LLM-Based Coding Agents

DGX agent

arXiv:2604.27283v1 Announce Type: cross Abstract: Large language model (LLM)-based coding agents increasingly rely on external memory to reuse prior debugging experience, repair traces, and repository

safetyarxiv-cs-ai
1 May 2026
Agents

MCP vs. CLI Skills for agents: what our eval found (and which you should use)

DGX agent

Twitter said pick a side. The eval said the question was wrong. Six months ago, MCP (model context protocol) was the hot new thing: tool usage with a built-in discovery... The post MCP vs. CLI Skills

agentsarize-ai
1 May 2026
Agents

Optimal Stop-Loss and Take-Profit Parameterization for Autonomous Trading Agent Swarm

DGX agent

arXiv:2604.27150v1 Announce Type: new Abstract: Autonomous crypto trading systems often spend most of their design effort on finding entries, while exits are left to fixed rules that are rarely tested

agentsarxiv-cs-ai
1 May 2026
Model Releases

Progressive Multi-Agent Reasoning for Biological Perturbation Prediction

DGX agent

arXiv:2602.07408v2 Announce Type: replace Abstract: Predicting gene regulation responses to biological perturbations requires reasoning about underlying biological causalities. While large language mo

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

What Makes a Good Terminal-Agent Benchmark Task: A Guideline for Adversarial, Difficult, and Legible Evaluation Design

DGX agent

arXiv:2604.28093v1 Announce Type: new Abstract: Terminal-agent benchmarks have become a primary signal for measuring the coding and system-administration capabilities of large language models. As the

model-releasesarxiv-cs-ai
1 May 2026
Agents

Agentic opportunity: OpenAI and Stripe build for rising tide of new entrepreneurial firms

DGX agent

In corporate parlance, any metric on a chart that shows rapid, significant growth over a short period of time is known as a “hockey stick.” On Wednesday, Stripe Inc. Chief Executive Patrick Collison s

agentssiliconangle
30 Apr 2026
Model Releases

Creating highly efficient agents: 450M tool-calling tokens distilled for post-training from top open-source models

DGX agent

Harnesses If you've used Claude Code or Codex, you've used a harness. A harness is the infrastructure layer that wraps an AI coding agent and decides how it operates, what it can touch, and how you me

model-releaseslambda-labs
30 Apr 2026
Safety

Evaluating Strategic Reasoning in Forecasting Agents

DGX agent

arXiv:2604.26106v1 Announce Type: new Abstract: Forecasting benchmarks produce accuracy leaderboards but little insight into why some forecasters are more accurate than others. We introduce Bench to t

safetyarxiv-cs-ai
30 Apr 2026
Model Releases

EvoDev: An Iterative Feature-Driven Framework for End-to-End Software Development with LLM-based Agents

DGX agent

arXiv:2511.02399v2 Announce Type: replace-cross Abstract: Recent advances in large language model agents offer the promise of automating end-to-end software development from natural language requireme

model-releasesarxiv-cs-ai
30 Apr 2026
Agents

Lightweight Quantum Agent for Edge Systems: Joint PQC and NOMA Resource Allocation

DGX agent

arXiv:2604.25980v1 Announce Type: cross Abstract: In the context of quantum secure scenarios, existing research on mobile edge devices and intelligent computing and edge (ICE) systems based on the Non

agentsarxiv-cs-ai
30 Apr 2026
Local Ai

Provable Coordination for LLM Agents via Message Sequence Charts

DGX agent

arXiv:2604.17612v2 Announce Type: replace-cross Abstract: Multi-agent systems built on large language models (LLMs) are difficult to reason about. Coordination errors such as deadlocks or type-mismatc

local-aiarxiv-cs-ai
30 Apr 2026
Agents

StreamAgent: Towards Anticipatory Agents for Streaming Video Understanding

DGX agent

arXiv:2508.01875v4 Announce Type: replace Abstract: Real-time streaming video understanding in domains such as autonomous driving and intelligent surveillance poses challenges beyond conventional offl

agentsarxiv-cs-cv
30 Apr 2026
Model Releases

Auvik launches Aurora AI agents to speed ticket resolution and prevent outages

DGX agent

Information technology management software provider Auvik Networks Inc. today announced the launch of Auvik Aurora: artificial intelligence-powered IT agents that are designed to help IT professionals

model-releasessiliconangle
29 Apr 2026
Model Releases

Aviatrix launches AI agent containment platform for cloud workloads

DGX agent

Aviatrix Inc. today announced the launch of a new platform designed to contain artificial intelligence agents and enforce security controls and communications across AI workloads without changing AI a

model-releasessiliconangle
29 Apr 2026
Applications

Databricks and Stripe Projects: Infrastructure Built for Agents

DGX agent

Databricks and Stripe have collaborated on infrastructure projects designed to support AI agents, likely focusing on data processing and payment integration capabilities. The initiative demonstrates h

applicationsdatabricks
29 Apr 2026
Model Releases

Frontier Coding Agents Can Now Implement an AlphaZero Self-Play Machine Learning Pipeline For Connect Four That Performs Comparably to an External Solver

DGX agent

arXiv:2604.25067v1 Announce Type: cross Abstract: Forecasting when AI systems will become capable of meaningfully accelerating AI research is a central challenge for AI safety. Existing benchmarks mea

model-releasesarxiv-cs-lg
29 Apr 2026
Model Releases

GAIA-v2-LILT: Multilingual Adaptation of Agent Benchmark beyond Translation

DGX agent

arXiv:2604.24929v1 Announce Type: new Abstract: Agent benchmarks remain largely English-centric, while their multilingual versions are often built with machine translation (MT) and limited post-editin

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

Evaluating the Search Agent in a Parallel World

DGX agent

arXiv:2603.04751v2 Announce Type: replace Abstract: Integrating web search tools has significantly extended the capability of LLMs to address open-world, real-time, and long-tail problems. However, ev

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

GradMAP: Gradient-Based Multi-Agent Proximal Learning for Grid-Edge Flexibility

DGX agent

arXiv:2604.24549v1 Announce Type: cross Abstract: Coordinating large populations of grid-edge devices requires learning methods that remain fully decentralised in deployment while still respecting thr

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

InquireMobile: Teaching VLM-based Mobile Agent to Request Human Assistance via Reinforcement Fine-Tuning

DGX agent

arXiv:2508.19679v2 Announce Type: replace Abstract: Recent advances in Vision-Language Models (VLMs) have enabled mobile agents to perceive and interact with real-world mobile environments based on hu

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Mobile-R1: Towards Interactive Capability for VLM-Based Mobile Agent via Systematic Training

DGX agent

arXiv:2506.20332v4 Announce Type: replace Abstract: Vision-language model-based mobile agents have gained the ability to understand complex instructions and mobile screenshots, benefiting from reinfor

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Scalable Agentic Reasoning for Designing Biologics Targeting Intrinsically Disordered Proteins

DGX agent

arXiv:2512.15930v2 Announce Type: replace-cross Abstract: Intrinsically disordered proteins (IDPs) represent crucial therapeutic targets due to their significant role in disease -- approximately 80% o

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

xOffense: An Autonomous Multi-Agent Framework for Penetration Testing with Domain-Adapted Large Language Models

DGX agent

arXiv:2509.13021v2 Announce Type: replace-cross Abstract: This work introduces xOffense, an AI-driven, multi-agent penetration testing framework that shifts the process from labor-intensive, expert-dr

model-releasesarxiv-cs-ai
28 Apr 2026
Agents

A Probabilistic Framework for Hierarchical Goal Recognition

DGX agent

arXiv:2604.22256v1 Announce Type: cross Abstract: Goal recognition aims to infer an agent's goal from observations of its behaviour. In realistic settings, recognition can benefit from exploiting hier

agentsarxiv-cs-ai
27 Apr 2026
Agents

Behavioral Canaries: Auditing Private Retrieved Context Usage in RL Fine-Tuning

DGX agent

arXiv:2604.22191v1 Announce Type: cross Abstract: In agentic workflows, LLMs frequently process retrieved contexts that are legally protected from further training. However, auditors currently lack a

agentsarxiv-cs-cl
27 Apr 2026
Model Releases

Happy to announce that Hermes Agent's repo just surpassed Anthropic's Claude Code repo

DGX agent

Nous Research announced that their Hermes Agent repository has surpassed Anthropic's Claude Code repository in popularity metrics, likely referring to GitHub stars or similar engagement measures. This

model-releasesnous-research--x
27 Apr 2026
Research

Rethinking Token Pruning for Historical Screenshots in GUI Visual Agents: Semantic, Spatial, and Temporal Perspectives

DGX agent

arXiv:2603.26041v3 Announce Type: replace Abstract: In recent years, GUI visual agents built upon Multimodal Large Language Models (MLLMs) have demonstrated strong potential in navigation tasks. Howev

researcharxiv-cs-cv
27 Apr 2026
Agents

SOLAR-RL: Semi-Online Long-horizon Assignment Reinforcement Learning

DGX agent

arXiv:2604.22558v1 Announce Type: cross Abstract: As Multimodal Large Language Models (MLLMs) mature, GUI agents are evolving from static interactions to complex navigation. While Reinforcement Learni

agentsarxiv-cs-ai
27 Apr 2026
Agents

The greatest con of the decade was calling autocomplete “AI”. The second greatest is calling autocomplete-in-a-loop an “agent”.

DGX agent

Gary Marcus critiques the marketing of large language model autocomplete systems as 'artificial intelligence,' arguing this terminology misrepresents their actual capabilities. He extends this critici

agentsgary-marcus--x
27 Apr 2026
Agents

The terminal hasn’t changed much since the 1970s. What you do with it has. Introducing Devin for Terminal: everything we learned building De…

DGX agent

The terminal hasn’t changed much since the 1970s. What you do with it has. Introducing Devin for Terminal: everything we learned building Devin, now as a local agent, available right in your shell. An

agentscognition-ai--x
27 Apr 2026
Model Releases

7) Multi-agent design. I loved the design of Cove (they were the acquired by Microsoft). It was more of a whiteboard than tabs or chat threa…

DGX agent

7) Multi-agent design. I loved the design of Cove (they were the acquired by Microsoft). It was more of a whiteboard than tabs or chat threads. I don’t think we’ve cracked the right UI for managing ag

model-releasesallie-k--miller--x
26 Apr 2026
Model Releases

DeepSeek-V4: a million-token context that agents can actually use

DGX agent

DeepSeek-V4 is an advanced language model featuring a million-token context window that enables practical agentic applications beyond simple retrieval. The model demonstrates improved efficiency and u

model-releaseshugging-face
24 Apr 2026
Model Releases

GPT-5.5 is now available in Devin as an Agent Preview! GPT-5.5 has set a new bar for what's possible with Devin. It runs longer and more aut…

DGX agent

GPT-5.5 is now available in Devin as an Agent Preview! GPT-5.5 has set a new bar for what's possible with Devin. It runs longer and more autonomously than any GPT model we've tested, surfacing bugs no

model-releasescognition-ai--x
24 Apr 2026
Agents

Structural Quality Gaps in Practitioner AI Governance Prompts: An Empirical Study Using a Five-Principle Evaluation Framework

DGX agent

arXiv:2604.21090v1 Announce Type: cross Abstract: AI governance programmes increasingly rely on natural language prompts to constrain and direct AI agent behaviour. These prompts function as executabl

agentsarxiv-cs-ai
24 Apr 2026
Agents

Tool Attention Is All You Need

DGX agent

Tool Attention Is All You Need // Tool Attention Is All You Need // New research proposes a practical fix for the hidden 'MCP tax.' The work introduces a dynamic tool gating mechanism built on an Inte

agentsdair-ai--x
24 Apr 2026
← Previous
1…137138139140141…375
Next →