AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,959 results
Agents

CountLoop: Training-Free High-Instance Image Generation via Iterative Agent Guidance

DGX agent

arXiv:2508.16644v4 Announce Type: replace Abstract: Diffusion models excel at photorealistic synthesis but struggle with precise object counts, especially in high-density settings. We introduce COUNTL

agentsarxiv-cs-cv
14 Apr 2026
Model Releases
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Detecting Safety Violations Across Many Agent Traces

DGX agent

arXiv:2604.11806v1 Announce Type: new Abstract: To identify safety violations, auditors often search over large sets of agent traces. This search is difficult because failures are often rare, complex,

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

EE-MCP: Self-Evolving MCP-GUI Agents via Automated Environment Generation and Experience Learning

DGX agent

arXiv:2604.09815v1 Announce Type: new Abstract: Computer-use agents that combine GUI interaction with structured API calls via the Model Context Protocol (MCP) show promise for automating software tas

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

HealthAdminBench: Evaluating Computer-Use Agents on Healthcare Administration Tasks

DGX agent

arXiv:2604.09937v1 Announce Type: new Abstract: Healthcare administration accounts for over $1 trillion in annual spending, making it a promising target for LLM-based computer-use agents (CUAs). While

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

NetAgentBench: A State-Centric Benchmark for Evaluating Agentic Network Configuration

DGX agent

arXiv:2604.09678v1 Announce Type: cross Abstract: As agentic network management gains popularity, there is a critical need for evaluation frameworks that transcend static, one-shot testing. To address

model-releasesarxiv-cs-ai
14 Apr 2026
Agents

Scaling MCP adoption: Our reference architecture for simpler, safer and cheaper enterprise deployments of MCP

DGX agent

We share Cloudflare's internal strategy for governing MCP using Access, AI Gateway, and MCP server portals. We also launch Code Mode to slash token costs and recommend new rules for detecting Shadow M

agentscloudflare-ai
14 Apr 2026
Agents

The AI-first workday is here, but is the data layer ready to handle it?

DGX agent

As “deploy fast” meets enterprise reality, data governance has quickly emerged as the bottleneck between AI ambition and outcomes. The disconnect is measurable. In their Agentic AI Study, Qlik Technol

agentssiliconangle
14 Apr 2026
Hardware

Tree Training: Accelerating Agentic LLMs Training via Shared Prefix Reuse

DGX agent

arXiv:2511.00413v4 Announce Type: replace Abstract: Agentic large language model (LLM) training often involves multi-turn interaction trajectories that branch into multiple execution paths due to conc

hardwarearxiv-cs-lg
14 Apr 2026
Agents

Working Paper: Towards Schema-based Learning from a Category-Theoretic Perspective

DGX agent

arXiv:2604.10589v1 Announce Type: new Abstract: We introduce a hierarchical categorical framework for Schema-Based Learning (SBL) structured across four interconnected levels. At the schema level, a f

agentsarxiv-cs-ai
14 Apr 2026
Agents

ActionNex: A Virtual Outage Manager for Cloud Computing

DGX agent

arXiv:2604.03512v2 Announce Type: replace Abstract: Outage management in large-scale cloud operations remains heavily manual, requiring rapid triage, cross-team coordination, and experience-driven dec

agentsarxiv-cs-ai
13 Apr 2026
Model Releases

Adaptive Tuning of Parameterized Traffic Controllers via Multi-Agent Reinforcement Learning

DGX agent

arXiv:2512.07417v2 Announce Type: replace Abstract: Effective traffic control is essential for mitigating congestion in transportation networks. Conventional traffic management strategies, including r

model-releasesarxiv-cs-lg
13 Apr 2026
Safety

AgentSociety: Large-Scale Simulation of LLM-Driven Generative Agents Advances Understanding of Human Behaviors and Society

DGX agent

arXiv:2502.08691v2 Announce Type: replace-cross Abstract: Understanding human behavior and society is a central focus in social sciences, with the rise of generative social science marking a significa

safetyarxiv-cs-ai
13 Apr 2026
Safety

Semantic Intent Fragmentation: A Single-Shot Compositional Attack on Multi-Agent AI Pipelines

DGX agent

arXiv:2604.08608v1 Announce Type: cross Abstract: We introduce Semantic Intent Fragmentation (SIF), an attack class against LLM orchestration systems where a single, legitimately phrased request cause

safetyarxiv-cs-ai
13 Apr 2026
Agents

Strategic Algorithmic Monoculture:Experimental Evidence from Coordination Games

DGX agent

arXiv:2604.09502v1 Announce Type: new Abstract: AI agents increasingly operate in multi-agent environments where outcomes depend on coordination. We distinguish primary algorithmic monoculture -- base

agentsarxiv-cs-ai
13 Apr 2026
Model Releases

Structured Uncertainty guided Clarification for LLM Agents

DGX agent

arXiv:2511.08798v2 Announce Type: replace-cross Abstract: LLM agents with tool-calling capabilities often fail when user instructions are ambiguous or incomplete, leading to incorrect invocations and

model-releasesarxiv-cs-ai
13 Apr 2026
Tools

This isn't the only result. As of April 11, agents have set 11 new SOTA results on open problems including: → Erdős minimum overlap problem …

DGX agent

This isn't the only result. As of April 11, agents have set 11 new SOTA results on open problems including: → Erdős minimum overlap problem → Second autocorrelation inequality → Tammes problem (n=50)

toolstogether-ai--x
13 Apr 2026
Safety

VISOR: Agentic Visual Retrieval-Augmented Generation via Iterative Search and Over-horizon Reasoning

DGX agent

arXiv:2604.09508v1 Announce Type: cross Abstract: Visual Retrieval-Augmented Generation (VRAG) empowers Vision-Language Models to retrieve and reason over visually rich documents. To tackle complex qu

safetyarxiv-cs-ai
13 Apr 2026
Agents

great read on why open harnesses have become so important. we used to focus on open models vs closed models, but i think the deeper issue is…

DGX agent

great read on why open harnesses have become so important. we used to focus on open models vs closed models, but i think the deeper issue is open harnesses vs closed harnesses. if the harness is close

agentsharrison-chase--x
12 Apr 2026
Research

Hermes Agents can communicate with eachother on Telegram now! Check it out

DGX agent

Hermes Agents can communicate with eachother on Telegram now! Check it out Hermes agents can now communicate in telegram with each other Use the /setbot2bot command in bot father to make it work @Nous

researchnous-research--x
11 Apr 2026
Tools

This weekend we're doubling Composer 2 usage in our new interface. Open Agents Window, pick Composer 2, and start building! No hourly limits…

DGX agent

This weekend we're doubling Composer 2 usage in our new interface. Open Agents Window, pick Composer 2, and start building! No hourly limits. https://x.com/cursor_ai/status/2039768512894505086?s=20 We

toolscursor--x
11 Apr 2026
Agents

Your harness, your memory

DGX agent

Agent harnesses are becoming the dominant way to build agents, and they are not going anywhere. These harnesses are intimately tied to agent memory. If you used a closed harness - especially if it’s b

agentslangchain-blog
11 Apr 2026
Safety

An Agentic Evaluation Architecture for Historical Bias Detection in Educational Textbooks

DGX agent

arXiv:2604.07883v1 Announce Type: cross Abstract: History textbooks often contain implicit biases, nationalist framing, and selective omissions that are difficult to audit at scale. We propose an agen

safetyarxiv-cs-cl
10 Apr 2026
Model Releases

Don't Overthink It: Inter-Rollout Action Agreement as a Free Adaptive-Compute Signal for LLM Agents

DGX agent

arXiv:2604.08369v1 Announce Type: cross Abstract: Inference-time compute scaling has emerged as a powerful technique for improving the reliability of large language model (LLM) agents, but existing me

model-releasesarxiv-cs-cl
10 Apr 2026
Agents

@hwchase17 Ngl I really like this direction. The more AGENTS.md, skills, and tool config start looking like portable interfaces instead of a…

DGX agent

A developer expressed enthusiasm for the emerging convergence of `AGENTS.md`, agent skills (SKILL.md), and tool configuration toward portable, cross-tool interfaces rather than siloed, tool-specifi...

agentsharrison-chase--x
10 Apr 2026
Safety

Reason in Chains, Learn in Trees: Self-Rectification and Grafting for Multi-turn Agent Policy Optimization

DGX agent

arXiv:2604.07165v1 Announce Type: new Abstract: Reinforcement learning for Large Language Model agents is often hindered by sparse rewards in multi-step reasoning tasks. Existing approaches like Group

safetyarxiv-cs-ai
10 Apr 2026
Model Releases

SkillSieve: A Hierarchical Triage Framework for Detecting Malicious AI Agent Skills

DGX agent

arXiv:2604.06550v1 Announce Type: cross Abstract: OpenClaw's ClawHub marketplace hosts over 13,000 community-contributed agent skills, and between 13% and 26% of them contain security vulnerabilities

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Strategic Persuasion with Trait-Conditioned Multi-Agent Systems for Iterative Legal Argumentation

DGX agent

arXiv:2604.07028v1 Announce Type: cross Abstract: Strategic interaction in adversarial domains such as law, diplomacy, and negotiation is mediated by language, yet most game-theoretic models abstract

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Verify Before You Commit: Towards Faithful Reasoning in LLM Agents via Self-Auditing

DGX agent

arXiv:2604.08401v1 Announce Type: cross Abstract: In large language model (LLM) agents, reasoning trajectories are treated as reliable internal beliefs for guiding actions and updating memory. However

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

VisCoder2: Building Multi-Language Visualization Coding Agents

DGX agent

arXiv:2510.23642v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have recently enabled coding agents capable of generating, executing, and revising visualization code. However, e

model-releasesarxiv-cs-ai
10 Apr 2026
Safety

WebExpert: domain-aware web agents with critic-guided expert experience for high-precision search

DGX agent

arXiv:2604.06177v1 Announce Type: cross Abstract: Specialized web tasks in finance, biomedicine, and pharmaceuticals remain challenging due to missing domain priors: queries drift, evidence is noisy,

safetyarxiv-cs-ai
10 Apr 2026
Applications

Announcing the @LangChain podcast -- Max Agency. Deep context to give you the edge while building and iterating on agents. Watch the full ep…

DGX agent

Announcing the @LangChain podcast -- Max Agency. Deep context to give you the edge while building and iterating on agents. Watch the full episode on: - Youtube: https://www.youtube.com/watch?v=Xyh1Eqc

applicationsharrison-chase--x
9 Apr 2026
Agents

here's how we're improving our base harness, you can apply these same lessons to hill-climbing for your application-specific harness!

DGX agent

LangChain's **Better-Harness** system, shared by Sydney Runkle, is a compound approach to iteratively improving AI agent harnesses using evaluations (evals) as a learning signal. Better agents can...

agentsharrison-chase--x
8 Apr 2026
Local Ai

Today, we are launching our collaboration with @nomic_ai to make AI agents more effectively and efficiently understand complex PDF documents…

DGX agent

Today, we are launching our collaboration with @nomic_ai to make AI agents more effectively and efficiently understand complex PDF documents. Nomic's new nomic-layout-v1 model allows your AI agents to

local-ainomic-ai--x
8 Apr 2026
Model Releases

GLM-5.1 is live everywhere you use the Kilo Gateway (VS Code extension, Cloud Agents, KiloClaw, etc). Thank you @Zai_org! ⚡️

DGX agent

Z.AI's GLM-5.1, a next-generation flagship model for agentic engineering released in April 2026, is now available across all Kilo Code surfaces — including the VS Code extension, Cloud Agents, and ...

model-releaseszhipu-ai--x
7 Apr 2026
Applications

GLM 5.1 is live on Fireworks! SOTA for agents and coding: →Plans and executes multi-hour workflows without falling apart →Planning, executin…

DGX agent

GLM 5.1 is live on Fireworks! SOTA for agents and coding: →Plans and executes multi-hour workflows without falling apart →Planning, executing, testing, and refining over hundreds of rounds to deliver

applicationsfireworks-ai--x
7 Apr 2026
Model Releases

Do LLMs Beat Nash? Testing Decentralized Coordination in Self-Play Multi-Agent Games

DGX agent

arXiv:2608.12547v1 Announce Type: cross Abstract: Large language model agents deployed without a central controller are often assumed to require communication to coordinate their actions. We ask what

model-releasesarxiv-cs-ro
14 Aug 2026
Model Releases

Doctorina MedBench: A Dialogue-Based Benchmark and Evaluation Framework for Agent-Based Medical AI

DGX agent

arXiv:2603.25821v3 Announce Type: replace-cross Abstract: We present Doctorina MedBench, an evaluation framework for agent-based medical AI based on the simulation of physician-patient interactions. U

model-releasesarxiv-cs-ai
14 Aug 2026
Safety

Intern-S2-Preview: Scientific Agentic Foundation Model

DGX agent

arXiv:2608.13505v1 Announce Type: cross Abstract: Scientific discovery increasingly requires AI systems that can reason over scientific evidence of heterogeneous modalities, interact with scientific t

safetyarxiv-cs-cl
14 Aug 2026
Model Releases

PlayWorld: Benchmarking World Models with Agent Players over Long-Horizon Objectives

DGX agent

arXiv:2608.13552v1 Announce Type: new Abstract: Video world models simulate future states conditioned on current observations and user actions. Recent systems have demonstrated impressive video consis

model-releasesarxiv-cs-cv
14 Aug 2026
Model Releases

Qwen 3.8 27B is now available on Ollama. It's one of the best open models at this size, and made for agentic tasks and professional work. Tr…

DGX agent

Qwen 3.8 27B is now available on Ollama. It's one of the best open models at this size, and made for agentic tasks and professional work. Try it directly with the apps & harnesses you use: Claude Code

model-releasesollama--x
14 Aug 2026
Safety

Reconcile Once, Write Anytime: A Trust-Tiered Librarian and a Multi-Agent Writer for Drift-Free, Point-in-Time Research

DGX agent

arXiv:2608.12984v1 Announce Type: cross Abstract: Long-form research reports generated by large language models drift, contradict themselves, and lose provenance: the same metric appears with differen

safetyarxiv-cs-cl
14 Aug 2026
Research

SynWeaver: Website-Prior Task and Trajectory Co-Synthesis for Web Agents

DGX agent

arXiv:2608.12429v1 Announce Type: cross Abstract: Web agents often struggle to generalize to unseen websites because they lack website-specific supervision. Recent exploration-based data synthesis met

researcharxiv-cs-ai
14 Aug 2026
Model Releases

Backtrader-Bench: Benchmarking LLM Agents on Algorithmic Trading with Self-Generated MCQs

DGX agent

arXiv:2608.11232v1 Announce Type: cross Abstract: Evaluating LLM coding agents in algorithmic trading is difficult because static benchmarks risk data contamination and numerical backtest outputs requ

model-releasesarxiv-cs-ai
13 Aug 2026
Agents

Beyond Single-Turn Confidence: Trajectory-Adapted Uncertainty Quantification for LLM Agents

DGX agent

arXiv:2608.11552v1 Announce Type: cross Abstract: Uncertainty quantification (UQ) methods for language models are typically evaluated on single-turn outputs, where uncertainty is attached to one gener

agentsarxiv-cs-ai
13 Aug 2026
Safety

Diffusion-Guided Cooperative Policy Learning for Target Tracking Based on Underwater Mobile Agent Networks

DGX agent

arXiv:2603.29426v2 Announce Type: replace-cross Abstract: Multi-agent reinforcement learning (MARL) provides a promising solution for cooperative target tracking in networks of autonomous underwater v

safetyarxiv-cs-lg
13 Aug 2026
Model Releases

Google launches Gemini 3.7 Flash for coding, AI agent projects

DGX agent

Google LLC today launched its most capable entry-level artificial intelligence model yet. Gemini 3.7 Flash is rolling out three weeks after its predecessor. Despite the short release cycle, Google eng

model-releasessiliconangle
13 Aug 2026
Model Releases

Governing Agentic AI in FinTech

DGX agent

arXiv:2608.11344v1 Announce Type: cross Abstract: Financial institutions are delegating consequential decisions to agentic AI systems that decompose goals, coordinate models and tools, and act with li

model-releasesarxiv-cs-ai
13 Aug 2026
Model Releases

MBA: Multimodal Benchmark and Agents for Real-World Business Ideation

DGX agent

arXiv:2608.11616v1 Announce Type: new Abstract: Agentic systems powered by large language models (LLMs) have opened new opportunities for business ideation. Yet existing approaches remain confined to

model-releasesarxiv-cs-ai
13 Aug 2026
← Previous
1…125126127128129…375
Next →