AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,950 results
Agents

Introducing the agent performance loop: AgentCore Optimization now in preview

DGX agent

Generate recommendations from production traces, validate them with batch evaluation and A/B testing, and ship with confidence. AI agents that perform well at launch don’t stay that way. As models evo

agentsaws-ml-blog
4 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

Introducing the agent quality loop: AgentCore Optimization now in preview

DGX agent

Generate recommendations from production traces, validate them with batch evaluation and A/B testing, and ship with confidence. AI agents that perform well at launch don’t stay that way. As models evo

agentsaws-ml-blog
4 May 2026
Safety

Scaling Video Understanding via Compact Latent Multi-Agent Collaboration

DGX agent

arXiv:2605.00444v1 Announce Type: new Abstract: Multi-modal large language models (MLLMs) advance vision language understanding but face inherent limitations in long-video tasks due to bounded percept

safetyarxiv-cs-cv
4 May 2026
Agents

We are gonna have genuinely useful long running agents in 2026. Somehow, this feels both entirely predictable, yet scarily futuristic

DGX agent

Jerry Liu predicts that genuinely useful long-running AI agents will become a reality by 2026, viewing this development as simultaneously inevitable and remarkable. The post reflects on how advances i

agentsjerry-liu--x
2 May 2026
Agents

Agentic Compilation: Mitigating the LLM Rerun Crisis for Minimized-Inference-Cost Web Automation

DGX agent

arXiv:2604.09718v2 Announce Type: cross Abstract: LLM-driven web agents operating through continuous inference loops -- repeatedly querying a model to evaluate browser state and select actions -- exhi

agentsarxiv-cs-ai
1 May 2026
Agents

Agentic engineering startup JuliaHub lands $65M to automate design and testing of industrial products

DGX agent

JuliaHub Inc., an “agentic” industrial engineering startup that’s trying to automate complex manufacturing processes with artificial intelligence, said today it has raised 65 million in a Series B fun

agentssiliconangle
1 May 2026
Model Releases

AgenticRecTune: Multi-Agent with Self-Evolving Skillhub for Recommendation System Optimization

DGX agent

arXiv:2604.26969v1 Announce Type: cross Abstract: Modern large-scale recommendation systems are typically constructed as multi-stage pipelines, encompassing pre-ranking, ranking, and re-ranking phases

model-releasesarxiv-cs-ai
1 May 2026
Safety

From surveillance to signalling: escalation channels as environmental controls for agentic AI

DGX agent

arXiv:2510.05192v2 Announce Type: replace-cross Abstract: When AI agents operating with access to sensitive information encounter a conflict between completing an assigned task and following rules or

safetyarxiv-cs-ai
1 May 2026
Agents

MM-StanceDet: Retrieval-Augmented Multi-modal Multi-agent Stance Detection

DGX agent

arXiv:2604.27934v1 Announce Type: new Abstract: Multimodal Stance Detection (MSD) is crucial for understanding public discourse, yet effectively fusing text and image, especially with conflicting sign

agentsarxiv-cs-ai
1 May 2026
Agents

Replit Agent is free tomorrow for everyone starting at 5am PST Show use what you can build in 24 hours And Replit is turning10! A trip down …

DGX agent

Replit is offering free access to Replit Agent to all users for 24 hours starting at 5am PST, allowing people to demonstrate what they can build in that time period. The announcement coincides with Re

agentsreplit--x
1 May 2026
Agents

Should you use a sandbox for your agent? @ListenLabs Co-Founder & CTO @florian_jue shared what can go wrong on the Max Agency podcast hosted…

DGX agent

This post discusses the use of sandboxes for AI agents, featuring insights from Florian Jue, co-founder and CTO of Listen Labs, during an appearance on the Max Agency podcast. The discussion likely co

agentsharrison-chase--x
1 May 2026
Agents

The US, UK, Australia, Canada, and New Zealand publish guidance on orgs' use of agentic AI systems, saying many give AI more access than can be safely monitored (Greg Otto/CyberScoop)

DGX agent

Greg Otto / CyberScoop: The US, UK, Australia, Canada, and New Zealand publish guidance on orgs' use of agentic AI systems, saying many give AI more access than can be safely monitored — The guidance

agentstechmeme
1 May 2026
Agents

Think it, Run it: Autonomous ML pipeline generation via self-healing multi-agent AI

DGX agent

arXiv:2604.27096v1 Announce Type: new Abstract: The purpose of our paper is to develop a unified multi-agent architecture that automates end-to-end machine learning (ML) pipeline generation from datas

agentsarxiv-cs-ai
1 May 2026
Agents

Trace-Level Analysis of Information Contamination in Multi-Agent Systems

DGX agent

arXiv:2604.27586v1 Announce Type: new Abstract: Reasoning over heterogeneous artifacts (PDFs, spreadsheets, slide decks, etc.) increasingly occurs within structured agent workflows that iteratively ex

agentsarxiv-cs-ai
1 May 2026
Agents

1/5 We hit SOTA performance on BrowseComp-Plus with 95.18% accuracy using AI21 Maestro’s agent optimization. Here’s how we automated the sea…

DGX agent

AI21 Labs achieved state-of-the-art performance on the BrowseComp-Plus benchmark with 95.18% accuracy using their AI21 Maestro model with agent optimization techniques. The post appears to discuss aut

agentsai21-labs--x
30 Apr 2026
Agents

DigiCert debuts AI Trust framework to secure agents, models and content

DGX agent

Digital security company DigiCert Inc. today introduced a new AI Trust framework to help organizations secure AI systems and their outputs, along with new capabilities to help secure autonomous agents

agentssiliconangle
30 Apr 2026
Agents

EvolvingAgent: Curriculum Self-evolving Agent with Continual World Model for Long-Horizon Tasks

DGX agent

arXiv:2502.05907v3 Announce Type: replace Abstract: Completing Long-Horizon (LH) tasks in open-ended worlds is an important yet difficult problem for embodied agents. Existing approaches suffer from t

agentsarxiv-cs-ro
30 Apr 2026
Agents

Google Cloud is rebuilding the enterprise stack for the age of agents

DGX agent

Google Cloud is assembling an ambitious agentic enterprise stack it believes will close the gap between AI ambition and real business outcomes. But the transition from traditional, linear workflows to

agentssiliconangle
30 Apr 2026
Hardware

How Vultr Enables Agentic AI Experiences with AMD

DGX agent

Vultr partnered with AMD to provide infrastructure and computing capabilities that support autonomous AI agents and agentic AI applications. The content likely discusses how Vultr's cloud platform, en

hardwarevultr
30 Apr 2026
Agents

Introducing Hermes Curator! The new system built in to Hermes Agent now helps you keep your skills that the self improvement loop creates in…

DGX agent

Introducing Hermes Curator! The new system built in to Hermes Agent now helps you keep your skills that the self improvement loop creates in check, by consolidating and pruning automatically. The cura

agentsnous-research--x
30 Apr 2026
Agents

OxyGent: Making Multi-Agent Systems Modular, Observable, and Evolvable via Oxy Abstraction

DGX agent

arXiv:2604.25602v2 Announce Type: replace Abstract: Deploying production-ready multi-agent systems (MAS) in complex industrial environments remains challenging due to limitations in scalability, obser

agentsarxiv-cs-ai
30 Apr 2026
Agents

🎂 Replit turns 10 this weekend, and we're giving every Replit user free Agent for 24 hours on Saturday.🤯 Clear your Saturday. And tune in …

DGX agent

🎂 Replit turns 10 this weekend, and we're giving every Replit user free Agent for 24 hours on Saturday.🤯 Clear your Saturday. And tune in live tomorrow, Friday at 9 AM PT, when we go live to walk you

agentsreplit--x
30 Apr 2026
Agents

Training Computer Use Agents to Assess the Usability of Graphical User Interfaces

DGX agent

arXiv:2604.26020v1 Announce Type: cross Abstract: Usability testing with experts and potential users can assess the effectiveness, efficiency, and user satisfaction of graphical user interfaces (GUIs)

agentsarxiv-cs-ai
30 Apr 2026
Agents

'Agent optimization should be automatic, efficient, observable and future-proof' (Or Dagan, 2 minutes ago)

DGX agent

Agent optimization requires four key characteristics: automation to reduce manual intervention, efficiency to maximize performance with minimal resource waste, observability to enable monitoring and d

agentsai21-labs--x
29 Apr 2026
Model Releases

BenchGuard: Who Guards the Benchmarks? Automated Auditing of LLM Agent Benchmarks

DGX agent

arXiv:2604.24955v1 Announce Type: new Abstract: As benchmarks grow in complexity, many apparent agent failures are not failures of the agent at all - they are failures of the benchmark itself: broken

model-releasesarxiv-cs-cl
29 Apr 2026
Agents

Full Workshop: @OpenAI Codex masterclass The agent is no longer just one chat window. In this workshop, @reach_vb and @kagigz get into how c…

DGX agent

Full Workshop: @OpenAI Codex masterclass The agent is no longer just one chat window. In this workshop, @reach_vb and @kagigz get into how coding systems start to change when you can delegate work acr

agentsswyx--x
29 Apr 2026
Agents

.@MadrigalPharma’s multi-agent research and intelligence platform for pharma is powered by LangChain + LangSmith. When users ask research qu…

DGX agent

.@MadrigalPharma’s multi-agent research and intelligence platform for pharma is powered by LangChain + LangSmith. When users ask research questions, the orchestrator breaks them into sub-tasks. Multip

agentsharrison-chase--x
29 Apr 2026
Agents

What to expect during the AI Agent Conference: Join theCUBE May 4-5

DGX agent

Agentic enterprise is moving artificial intelligence from isolated tools into systems that actively run business operations. The transition reflects a broader shift in enterprise strategy, where organ

agentssiliconangle
29 Apr 2026
Agents

You can now control both local and cloud ComfyUI with Hermes Agent with ease with the new built in skill `hermes update` and run /comfyui to…

DGX agent

You can now control both local and cloud ComfyUI with Hermes Agent with ease with the new built in skill `hermes update` and run /comfyui to get started ComfyUI is the most flexible, composable, and p

agentsnous-research--x
29 Apr 2026
Safety

Failure-Centered Runtime Evaluation for Deployed Trilingual Public-Space Agents

DGX agent

arXiv:2604.23990v1 Announce Type: new Abstract: This paper presents PSA-Eval, a failure-centered runtime evaluation framework for deployed trilingual public-space agents. The central claim is that, wh

safetyarxiv-cs-ai
28 Apr 2026
Tutorials

From Coarse to Fine: Self-Adaptive Hierarchical Planning for LLM Agents

DGX agent

arXiv:2604.23194v1 Announce Type: new Abstract: Large language model-based agents have recently emerged as powerful approaches for solving dynamic and multi-step tasks. Most existing agents employ pla

tutorialsarxiv-cs-ai
28 Apr 2026
Model Releases

Latency and Cost of Multi-Agent Intelligent Tutoring at Scale

DGX agent

arXiv:2604.24110v1 Announce Type: cross Abstract: Multi-agent LLM tutoring systems improve response quality through agent specialization, but each student query triggers several concurrent API calls w

model-releasesarxiv-cs-ai
28 Apr 2026
Safety

OS-SPEAR: A Toolkit for the Safety, Performance,Efficiency, and Robustness Analysis of OS Agents

DGX agent

arXiv:2604.24348v1 Announce Type: new Abstract: The evolution of Multimodal Large Language Models (MLLMs) has shifted the focus from text generation to active behavioral execution, particularly via OS

safetyarxiv-cs-cl
28 Apr 2026
Model Releases

Reasonably reasoning AI agents can avoid game-theoretic failures in zero-shot, provably

DGX agent

arXiv:2603.18563v2 Announce Type: replace Abstract: As autonomous AI agents increasingly mediate online platform markets, a fundamental question emerges: do these markets generate stable strategic out

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

KompeteAI: Accelerated Autonomous Multi-Agent System for End-to-End Pipeline Generation for Machine Learning Problems

DGX agent

arXiv:2508.10177v3 Announce Type: replace Abstract: Recent Large Language Model (LLM)-based AutoML systems demonstrate impressive capabilities but face significant limitations such as constrained expl

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

MATRAG: Multi-Agent Transparent Retrieval-Augmented Generation for Explainable Recommendations

DGX agent

arXiv:2604.20848v1 Announce Type: cross Abstract: Large Language Model (LLM)-based recommendation systems have demonstrated remarkable capabilities in understanding user preferences and generating per

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Nemobot Games: Crafting Strategic AI Gaming Agents for Interactive Learning with Large Language Models

DGX agent

arXiv:2604.21896v1 Announce Type: new Abstract: This paper introduces a new paradigm for AI game programming, leveraging large language models (LLMs) to extend and operationalize Claude Shannon's taxo

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Chasing the Public Score: User Pressure and Evaluation Exploitation in Coding Agent Workflows

DGX agent

arXiv:2604.20200v1 Announce Type: new Abstract: Frontier coding agents are increasingly used in workflows where users supervise progress primarily through repeated improvement of a public score, namel

model-releasesarxiv-cs-cl
23 Apr 2026
Agents

Enhancing Research Idea Generation through Combinatorial Innovation and Multi-Agent Iterative Search Strategies

DGX agent

arXiv:2604.20548v1 Announce Type: cross Abstract: Scientific progress depends on the continual generation of innovative re-search ideas. However, the rapid growth of scientific literature has greatly

agentsarxiv-cs-ai
23 Apr 2026
Model Releases

From Data to Theory: Autonomous Large Language Model Agents for Materials Science

DGX agent

arXiv:2604.19789v1 Announce Type: new Abstract: We present an autonomous large language model (LLM) agent for end-to-end, data-driven materials theory development. The model can choose an equation for

model-releasesarxiv-cs-ai
23 Apr 2026
Agents

Mol-Debate: Multi-Agent Debate Improves Structural Reasoning in Molecular Design

DGX agent

arXiv:2604.20254v1 Announce Type: new Abstract: Text-guided molecular design is a key capability for AI-driven drug discovery, yet it remains challenging to map sequential natural-language instruction

agentsarxiv-cs-ai
23 Apr 2026
Hardware

ARGUS: Agentic GPU Optimization Guided by Data-Flow Invariants

DGX agent

arXiv:2604.18616v1 Announce Type: cross Abstract: LLM-based coding agents can generate functionally correct GPU kernels, yet their performance remains far below hand-optimized libraries on critical co

hardwarearxiv-cs-ai
22 Apr 2026
Safety

BAPO: Boundary-Aware Policy Optimization for Reliable Agentic Search

DGX agent

arXiv:2601.11037v2 Announce Type: replace Abstract: RL-based agentic search enables LLMs to solve complex questions via dynamic planning and external search. While this approach significantly enhances

safetyarxiv-cs-ai
22 Apr 2026
Agents

benchmarking agents by their ability to play video games 🎮

DGX agent

benchmarking agents by their ability to play video games 🎮 We are excited to launch VideoGameBench on Antim Labs, created by @a1zhang, Thomas L. Griffiths (@cocosci_lab), @karthik_r_n, and @OfirPress

agentsyohei-nakajima--x
22 Apr 2026
Model Releases

Debug2Fix: Can Interactive Debugging Help Coding Agents Fix More Bugs?

DGX agent

arXiv:2602.18571v2 Announce Type: replace-cross Abstract: While significant progress has been made in automating various aspects of software development through coding agents, there is still significa

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Four-Axis Decision Alignment for Long-Horizon Enterprise AI Agents

DGX agent

arXiv:2604.19457v1 Announce Type: new Abstract: Long-horizon enterprise agents make high-stakes decisions (loan underwriting, claims adjudication, clinical review, prior authorization) under lossy mem

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

How Adversarial Environments Mislead Agentic AI?

DGX agent

arXiv:2604.18874v1 Announce Type: new Abstract: Tool-integrated agents are deployed on the premise that external tools ground their outputs in reality. Yet this very reliance creates a critical attack

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Human-Guided Harm Recovery for Computer Use Agents

DGX agent

arXiv:2604.18847v1 Announce Type: new Abstract: As LM agents gain the ability to execute actions on real computer systems, we need ways to not only prevent harmful actions at scale but also effectivel

model-releasesarxiv-cs-ai
22 Apr 2026
← Previous
1…9394959697…374
Next →