AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,959 results
Agents

In the enterprise AI race, who is leading and who is just reacting?

DGX agent

Enterprise AI scaling is accelerating as organizations shift from experimentation to full deployment, embedding intelligence into core workflows. At the same time, agentic systems are driving a broade

agentssiliconangle
23 Apr 2026
Research
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Information Aggregation with AI Agents

DGX agent

arXiv:2604.20050v1 Announce Type: cross Abstract: Can Large Language Models (AI agents) aggregate dispersed private information through trading and reason about the knowledge of others by observing pr

researcharxiv-cs-ai
23 Apr 2026
Industry

OpenAI subscribers get new ‘workspace agents’ to automate complex tasks across teams

DGX agent

OpenAI Group PBC said today it’s pushing ChatGPT outside its usual chat interface with the launch of “workspace agents,” which is a new feature that allows business users to automate recurring tasks,

industrysiliconangle
23 Apr 2026
Model Releases

ParseBench is now live on @Kaggle. The first document OCR benchmark built for AI agents — 2,000 enterprise pages, 167K+ test rules, 5 dimens…

DGX agent

ParseBench is now live on @Kaggle. The first document OCR benchmark built for AI agents — 2,000 enterprise pages, 167K+ test rules, 5 dimensions that actually break downstream agents. Benchmark your p

model-releasesjerry-liu--x
23 Apr 2026
Agents

Rilian raises $17.5M to automate security software procurement and deployment in the defense sector

DGX agent

A startup called Rilian that’s building agentic systems integration tools for companies operating in the defense and national security industries, said today it has raised 17.5 million in seed funding

agentssiliconangle
23 Apr 2026
Applications

Stateless Decision Memory for Enterprise AI Agents

DGX agent

arXiv:2604.20158v1 Announce Type: new Abstract: Enterprise deployment of long-horizon decision agents in regulated domains (underwriting, claims adjudication, tax examination) is dominated by retrieva

applicationsarxiv-cs-ai
23 Apr 2026
Model Releases

Trajectory2Task: Training Robust Tool-Calling Agents with Synthesized Yet Verifiable Data for Complex User Intents

DGX agent

arXiv:2601.20144v3 Announce Type: replace Abstract: Tool-calling agents are increasingly deployed in real-world customer-facing workflows. Yet most studies on tool-calling agents focus on idealized se

model-releasesarxiv-cs-cl
23 Apr 2026
Agents

AI scientists produce results without reasoning scientifically

DGX agent

arXiv:2604.18805v1 Announce Type: new Abstract: Large language model (LLM)-based systems are increasingly deployed to conduct scientific research autonomously, yet whether their reasoning adheres to t

agentsarxiv-cs-ai
22 Apr 2026
Model Releases

Hugging Face Releases ml-intern: An Open-Source AI Agent that Automates the LLM Post-Training Workflow [The 'AI Intern' that actually ships …

DGX agent

Hugging Face Releases ml-intern: An Open-Source AI Agent that Automates the LLM Post-Training Workflow [The 'AI Intern' that actually ships SOTA models ] This isn't just another ML Research Loop wrapp

model-releasesclem-delangue--x
22 Apr 2026
Model Releases

Iterable launches Nova agent to assist markets in scaling customer personalization

DGX agent

Iterable Inc., a customer engagement platform, today announced that it launched an artificial intelligence agent designed to help marketers keep customer interactions relevant as campaigns scale. The

model-releasessiliconangle
22 Apr 2026
Hardware

NVIDIA and Google Cloud Collaborate to Advance Agentic and Physical AI

DGX agent

NVIDIA and Google Cloud have collaborated for more than a decade, co‑engineering a full‑stack AI platform that spans every technology layer — from performance‑optimized libraries and frameworks to ent

hardwarenvidia-blog
22 Apr 2026
Model Releases

Owner-Harm: A Missing Threat Model for AI Agent Safety

DGX agent

arXiv:2604.18658v1 Announce Type: cross Abstract: Existing AI agent safety benchmarks focus on generic criminal harm (cybercrime, harassment, weapon synthesis), leaving a systematic blind spot for a d

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Taming Actor-Observer Asymmetry in Agents via Dialectical Alignment

DGX agent

arXiv:2604.19548v1 Announce Type: cross Abstract: Large Language Model agents have rapidly evolved from static text generators into dynamic systems capable of executing complex autonomous workflows. T

model-releasesarxiv-cs-ai
22 Apr 2026
Agents

We are excited to launch VideoGameBench on Antim Labs, created by @a1zhang, Thomas L. Griffiths (@cocosci_lab), @karthik_r_n, and @OfirPress…

DGX agent

VideoGameBench is a new benchmark launched on Antim Labs, created by a1zhang, Thomas L. Griffiths, Karthik R. N, and Ofir Press. The benchmark likely evaluates AI model performance on video game-relat

agentsyohei-nakajima--x
22 Apr 2026
Model Releases

What’s new in Cloud Run at Next ‘26

DGX agent

From vibe-coded and large-scale apps to AI models and agents, Cloud Run delivers on-demand compute with zero overhead and pay-per-use pricing for all of your workloads. Last year, the number of extern

model-releasesgoogle-cloud-ai
22 Apr 2026
Agents

BOIL: Learning Environment Personalized Information

DGX agent

arXiv:2604.17137v1 Announce Type: new Abstract: Navigating complex environments poses challenges for multi-agent systems, requiring efficient extraction of insights from limited information. In this p

agentsarxiv-cs-lg
21 Apr 2026
Model Releases

ClawEnvKit: Automatic Environment Generation for Claw-Like Agents

DGX agent

arXiv:2604.18543v1 Announce Type: cross Abstract: Constructing environments for training and evaluating claw-like agents remains a manual, human-intensive process that does not scale. We argue that wh

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

ComPASS: Towards Personalized Agentic Social Support via Tool-Augmented Companionship

DGX agent

arXiv:2604.18356v1 Announce Type: new Abstract: Developing compassionate interactive systems requires agents to not only understand user emotions but also provide diverse, substantive support. While r

model-releasesarxiv-cs-cl
21 Apr 2026
Agents

Ivanti extends Neurons platform with autonomous IT and security capabilities

DGX agent

Information technology security software company Ivanti Inc. today announced new capabilities that are focused on allowing autonomous IT operations and organizations to secure their environments more

agentssiliconangle
21 Apr 2026
Safety

On-Orbit Space AI: Federated, Multi-Agent, and Collaborative Algorithms for Satellite Constellations

DGX agent

arXiv:2604.16518v1 Announce Type: new Abstract: Satellite constellations are transforming space systems from isolated spacecraft into networked, software-defined platforms capable of on-orbit percepti

safetyarxiv-cs-ro
21 Apr 2026
Local Ai

VideoThinker: Building Agentic VideoLLMs with LLM-Guided Tool Reasoning

DGX agent

arXiv:2601.15724v2 Announce Type: replace Abstract: Long-form video understanding remains a fundamental challenge for current Video Large Language Models. Most existing models rely on static reasoning

local-aiarxiv-cs-cv
21 Apr 2026
Agents

What Makes AI Research Replicable? Executable Knowledge Graphs as Scientific Knowledge Representations

DGX agent

arXiv:2510.17795v3 Announce Type: replace Abstract: Replicating AI research is a crucial yet challenging task for large language model (LLM) agents. Existing approaches often struggle to generate exec

agentsarxiv-cs-cl
21 Apr 2026
Model Releases

ARC-AGI-3: A New Challenge for Frontier Agentic Intelligence

DGX agent

arXiv:2603.24621v2 Announce Type: replace Abstract: We introduce ARC-AGI-3, an interactive benchmark for studying agentic intelligence through novel, abstract, turn-based environments in which agents

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

ChemGraph-XANES: An Agentic Framework for XANES Simulation and Analysis

DGX agent

arXiv:2604.16205v1 Announce Type: cross Abstract: Computational X-ray absorption near-edge structure (XANES) is widely used to probe local coordination environments, oxidation states, and electronic s

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

Classic study gave 146 economist teams the same dataset & got wildly different answers New paper reruns it with agentic AI. Claude Code & Co…

DGX agent

Classic study gave 146 economist teams the same dataset & got wildly different answers New paper reruns it with agentic AI. Claude Code & Codex land near the human median, but with far tighter dispers

model-releasesethan-mollick--x
20 Apr 2026
Model Releases

HarmfulSkillBench: How Do Harmful Skills Weaponize Your Agents?

DGX agent

arXiv:2604.15415v1 Announce Type: cross Abstract: Large language models (LLMs) have evolved into autonomous agents that rely on open skill ecosystems (e.g., ClawHub and Skills.Rest), hosting numerous

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

Preference Estimation via Opponent Modeling in Multi-Agent Negotiation

DGX agent

arXiv:2604.15687v1 Announce Type: new Abstract: Automated negotiation in complex, multi-party and multi-issue settings critically depends on accurate opponent modeling. However, conventional numerical

model-releasesarxiv-cs-cl
20 Apr 2026
Agents

Headless everything for personal AI

DGX agent

Headless everything for personal AI Matt Webb thinks headless services are about to become much more common: Why? Because using personal AIs is a better experience for users than using services direct

agentssimon-willison
19 Apr 2026
Model Releases

Hermes Agent is model & tool backend agnostic for a reason, everyone should have access to AI. We don't dictate the rules of use for your ag…

DGX agent

Hermes Agent is model & tool backend agnostic for a reason, everyone should have access to AI. We don't dictate the rules of use for your agent, YOU do Anthropic shut down an entire company's Claude a

model-releasesnous-research--x
18 Apr 2026
Agents

@grok you go first

DGX agent

This post likely discusses Grok, an AI assistant developed by xAI, possibly exploring its capabilities, performance, or a specific interaction with the system. Given the casual phrasing and that it's

agentsyohei-nakajima--x
17 Apr 2026
Tools

Through the end of this weekend, we are doubling Composer 2 usage limits inside of Cursor's new agents window. Enjoy!

DGX agent

Cursor is temporarily doubling the usage limits for Composer 2 through the end of the weekend, specifically within Cursor's new agents window feature. This promotion allows users to access increased c

toolscursor--x
17 Apr 2026
Agents

A real issue with the current state of our knowledge on the work implications of AI is that there was a genuine discontinuity in AI ability …

DGX agent

A real issue with the current state of our knowledge on the work implications of AI is that there was a genuine discontinuity in AI ability with the rise of practical agentic systems in 2026. We were

agentsethan-mollick--x
16 Apr 2026
Tutorials

Coding agents learn from experience, but that knowledge stays locked in silos. Solve a thousand SWE tasks, and none of that wisdom helps wit…

DGX agent

Coding agents learn from experience, but that knowledge stays locked in silos. Solve a thousand SWE tasks, and none of that wisdom helps with competitive coding. What if memories could transfer across

tutorialsdair-ai--x
16 Apr 2026
Model Releases

RiskWebWorld: A Realistic Interactive Benchmark for GUI Agents in E-commerce Risk Management

DGX agent

arXiv:2604.13531v1 Announce Type: cross Abstract: Graphical User Interface (GUI) agents show strong capabilities for automating web tasks, but existing interactive benchmarks primarily target benign,

model-releasesarxiv-cs-lg
16 Apr 2026
Tutorials

Why Your Agents Can’t Read Enterprise Documents — and How to Fix It

DGX agent

AI agents struggle to effectively process and extract information from complex enterprise documents due to limitations in context windows, reasoning capabilities, and handling of unstructured data for

tutorialsdatabricks
16 Apr 2026
Model Releases

AlphaEval: Evaluating Agents in Production

DGX agent

arXiv:2604.12162v1 Announce Type: new Abstract: The rapid deployment of AI agents in commercial settings has outpaced the development of evaluation methodologies that reflect production realities. Exi

model-releasesarxiv-cs-cl
15 Apr 2026
Model Releases

How memory can affect collective and cooperative behaviors in an LLM-Based Social Particle Swarm

DGX agent

arXiv:2604.12250v1 Announce Type: new Abstract: This study examines how model-specific characteristics of Large Language Model (LLM) agents, including internal alignment, shape the effect of memory on

model-releasesarxiv-cs-ai
15 Apr 2026
Agents

Register domains wherever you build: Cloudflare Registrar API now in beta

DGX agent

The Cloudflare Registrar API is now in beta. Developers and AI agents can search, check availability, and register domains at cost directly from their editor, their terminal, or their agent — without

agentscloudflare-ai
15 Apr 2026
Model Releases

SIR-Bench: Evaluating Investigation Depth in Security Incident Response Agents

DGX agent

arXiv:2604.12040v1 Announce Type: cross Abstract: We present SIR-Bench, a benchmark of 794 test cases for evaluating autonomous security incident response agents that distinguishes genuine forensic in

model-releasesarxiv-cs-ai
15 Apr 2026
Hardware

🆕 The Full Story of Notion AI https://latent.space/p/notion We're so excited to chat with @simonlast and @sarahmsachs about Notion's 'Token…

DGX agent

🆕 The Full Story of Notion AI https://latent.space/p/notion We're so excited to chat with @simonlast and @sarahmsachs about Notion's 'Token Town' - the crack team of AI Engineers and Model Behavior En

hardwareswyx--x
15 Apr 2026
Model Releases

Transferable Expertise for Autonomous Agents via Real-World Case-Based Learning

DGX agent

arXiv:2604.12717v1 Announce Type: new Abstract: LLM-based autonomous agents perform well on general reasoning tasks but still struggle to reliably use task structure, key constraints, and prior experi

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

VULCAN: Vision-Language-Model Enhanced Multi-Agent Cooperative Navigation for Indoor Fire-Disaster Response

DGX agent

arXiv:2604.12831v1 Announce Type: new Abstract: Indoor fire disasters pose severe challenges to autonomous search and rescue due to dense smoke, high temperatures, and dynamically evolving indoor envi

model-releasesarxiv-cs-ro
15 Apr 2026
Agents

When to Forget: A Memory Governance Primitive

DGX agent

arXiv:2604.12007v1 Announce Type: new Abstract: Agent memory systems accumulate experience but currently lack a principled operational metric for memory quality governance -- deciding which memories t

agentsarxiv-cs-ai
15 Apr 2026
Hardware

AIRA_2: Overcoming Bottlenecks in AI Research Agents

DGX agent

arXiv:2603.26499v2 Announce Type: replace Abstract: Existing research has identified three structural performance bottlenecks in AI research agents: (1) synchronous single-GPU execution constrains sam

hardwarearxiv-cs-ai
14 Apr 2026
Model Releases

BankerToolBench: Evaluating AI Agents in End-to-End Investment Banking Workflows

DGX agent

arXiv:2604.11304v1 Announce Type: new Abstract: Existing AI benchmarks lack the fidelity to assess economically meaningful progress on professional workflows. To evaluate frontier AI agents in a high-

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

ClawVM: Harness-Managed Virtual Memory for Stateful Tool-Using LLM Agents

DGX agent

arXiv:2604.10352v1 Announce Type: new Abstract: Stateful tool-using LLM agents treat the context window as working memory, yet today's agent harnesses manage residency and durability as best-effort, c

model-releasesarxiv-cs-ai
14 Apr 2026
Local Ai

CodeComp: Structural KV Cache Compression for Agentic Coding

DGX agent

arXiv:2604.10235v1 Announce Type: new Abstract: Agentic code tasks such as fault localization and patch generation require processing long codebases under tight memory constraints, where the Key-Value

local-aiarxiv-cs-cl
14 Apr 2026
Model Releases

CoEvoSkills: Self-Evolving Agent Skills via Co-Evolutionary Verification

DGX agent

arXiv:2604.01687v2 Announce Type: replace Abstract: Anthropic proposes the concept of skills for LLM agents to tackle multi-step professional tasks that simple tool invocations cannot address. A tool

model-releasesarxiv-cs-ai
14 Apr 2026
← Previous
1…124125126127128…375
Next →