AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,963 results
Local Ai

Love seeing this open-sourced. Had a great chat with @nicoalbanese10 some weeks ago where he hinted to something like this. Great reference …

DGX agent

Love seeing this open-sourced. Had a great chat with @nicoalbanese10 some weeks ago where he hinted to something like this. Great reference architecture for cloud coding agents. Open Agents gives you

local-aiharrison-chase--x
14 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

MADQRL: Distributed Quantum Reinforcement Learning Framework for Multi-Agent Environments

DGX agent

arXiv:2604.11131v1 Announce Type: new Abstract: Reinforcement learning (RL) is one of the most practical ways to learn from real-life use-cases. Motivated from the cognitive methods used by humans mak

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

Multi-ORFT: Stable Online Reinforcement Fine-Tuning for Multi-Agent Diffusion Planning in Cooperative Driving

DGX agent

arXiv:2604.11734v1 Announce Type: cross Abstract: Closed-loop cooperative driving requires planners that generate realistic multimodal multi-agent trajectories while improving safety and traffic effic

model-releasesarxiv-cs-ai
14 Apr 2026
Agents

OpeFlo: Automated UX Evaluation via Simulated Human Web Interaction with GUI Grounding

DGX agent

arXiv:2604.09581v1 Announce Type: new Abstract: Evaluating web usability typically requires time-consuming user studies and expert reviews, which often limits iteration speed during product developmen

agentsarxiv-cs-ai
14 Apr 2026
Model Releases

Playing Along: Learning a Double-Agent Defender for Belief Steering via Theory of Mind

DGX agent

arXiv:2604.11666v1 Announce Type: cross Abstract: As large language models (LLMs) become the engine behind conversational systems, their ability to reason about the intentions and states of their dial

model-releasesarxiv-cs-ai
14 Apr 2026
Agents

The Devil is in the Details -- From OCR for Old Church Slavonic to Purely Visual Stemma Reconstruction

DGX agent

arXiv:2604.11724v1 Announce Type: new Abstract: The age of artificial intelligence has brought many new possibilities and pitfalls in many fields and tasks. The devil is in the details, and those come

agentsarxiv-cs-cv
14 Apr 2026
Model Releases

Time is Not a Label: Continuous Phase Rotation for Temporal Knowledge Graphs and Agentic Memory

DGX agent

arXiv:2604.11544v1 Announce Type: cross Abstract: Structured memory representations such as knowledge graphs are central to autonomous agents and other long-lived systems. However, most existing appro

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

AgentCE-Bench: Agent Configurable Evaluation with Scalable Horizons and Controllable Difficulty under Lightweight Environments

DGX agent

arXiv:2604.06111v2 Announce Type: replace Abstract: Existing Agent benchmarks suffer from two critical limitations: high environment interaction overhead (up to 41% of total evaluation time) and imbal

model-releasesarxiv-cs-ai
13 Apr 2026
Applications

Microsoft says it is 'exploring the potential of technologies like OpenClaw in an enterprise context', including a team of always-on agents within Microsoft 365 (Aaron Holmes/The Information)

DGX agent

Aaron Holmes / The Information: Microsoft says it is “exploring the potential of technologies like OpenClaw in an enterprise context”, including a team of always-on agents within Microsoft 365 — As Mi

applicationstechmeme
13 Apr 2026
Safety

Plasticity-Enhanced Multi-Agent Mixture of Experts for Dynamic Objective Adaptation in UAVs-Assisted Emergency Communication Networks

DGX agent

arXiv:2604.09028v1 Announce Type: cross Abstract: Unmanned aerial vehicles serving as aerial base stations can rapidly restore connectivity after disasters, yet abrupt changes in user mobility and tra

safetyarxiv-cs-lg
13 Apr 2026
Model Releases

SEA-Eval: A Benchmark for Evaluating Self-Evolving Agents Beyond Episodic Assessment

DGX agent

arXiv:2604.08988v1 Announce Type: new Abstract: Current LLM-based agents demonstrate strong performance in episodic task execution but remain constrained by static toolsets and episodic amnesia, faili

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

SPASM: Stable Persona-driven Agent Simulation for Multi-turn Dialogue Generation

DGX agent

arXiv:2604.09212v1 Announce Type: new Abstract: Large language models are increasingly deployed in multi-turn settings such as tutoring, support, and counseling, where reliability depends on preservin

model-releasesarxiv-cs-cl
13 Apr 2026
Agents

'every memory system is choosing a position on the raw/derived spectrum. and neither extreme works.' the axis nobody talks about: ownership …

DGX agent

'every memory system is choosing a position on the raw/derived spectrum. and neither extreme works.' the axis nobody talks about: ownership > memory is your agent's compounding advantage > derived fac

agentsharrison-chase--x
12 Apr 2026
Agents

Frameworks For Supporting LLM/Agentic Benchmarking [P]

DGX agent

This r/MachineLearning post discusses the landscape of frameworks and tools used to support benchmarking of LLMs and agentic AI systems, covering how to systematically evaluate model capabilities beyo

agentsr-machinelearning
12 Apr 2026
Agents

Memory is where the harness stops being a wrapper and becomes an ownership layer. Once it controls what gets remembered, retrieved, compress…

DGX agent

Memory is where the harness stops being a wrapper and becomes an ownership layer. Once it controls what gets remembered, retrieved, compressed, and acted on, it starts shaping the agent’s judgment, no

agentsharrison-chase--x
12 Apr 2026
Agents

I built modern AI client for Mac with agentic tools, elegant UI, interactive charts and maps, sortable tables, Slack-like threads and access to local and cloud models

DGX agent

A Reddit post on r/ollama showcasing a community-built, feature-rich macOS AI client designed for both local and cloud model access, including support for Ollama. The application emphasizes a modern,

agentsr-ollama
11 Apr 2026
Concepts

All People

DGX agent

Auto-generated index of all people mentioned across the wiki.

conceptspeopleindex
11 Apr 2026
Safety

Are GUI Agents Focused Enough? Automated Distraction via Semantic-level UI Element Injection

DGX agent

arXiv:2604.07831v1 Announce Type: cross Abstract: Existing red-teaming studies on GUI agents have important limitations. Adversarial perturbations typically require white-box access, which is unavaila

safetyarxiv-cs-cl
10 Apr 2026
Agents

Exploring Plan Space through Conversation: An Agentic Framework for LLM-Mediated Explanations in Planning

DGX agent

arXiv:2603.02070v2 Announce Type: replace-cross Abstract: When automating plan generation for a real-world sequential decision problem, the goal is often not to replace the human planner, but to facil

agentsarxiv-cs-cl
10 Apr 2026
Safety

Learning to Negotiate: Multi-Agent Deliberation for Collective Value Alignment in LLMs

DGX agent

arXiv:2603.10476v2 Announce Type: replace Abstract: LLM alignment has progressed in single-agent settings through paradigms such as RL with human feedback (RLHF), while recent work explores scalable a

safetyarxiv-cs-cl
10 Apr 2026
Agents

Lighting-grounded Video Generation with Renderer-based Agent Reasoning

DGX agent

arXiv:2604.07966v1 Announce Type: new Abstract: Diffusion models have achieved remarkable progress in video generation, but their controllability remains a major limitation. Key scene factors such as

agentsarxiv-cs-cv
10 Apr 2026
Applications

we are doing this podcast specifically to highlight the innermost details of building agents in production may be a little niche, but its a …

DGX agent

we are doing this podcast specifically to highlight the innermost details of building agents in production may be a little niche, but its a niche i like 🤷‍♂️ @hwchase17 Great content, really appreciat

applicationsharrison-chase--x
10 Apr 2026
Model Releases

what he said 🗣️ the very best agents today obsessively tailor the harness layer around the model I’m looking at you “5 things I learned fro…

DGX agent

what he said 🗣️ the very best agents today obsessively tailor the harness layer around the model I’m looking at you “5 things I learned from the Claude code leak” bros 👀 orchestration patterns, tool d

model-releasesharrison-chase--x
10 Apr 2026
Model Releases

Is a backlash brewing? Rapid innovation in AI coding and agents may force push for enterprise order and control

DGX agent

Artificial intelligence is proving to be a big bet for many companies across the enterprise landscape, and the gamblers are getting nervous. A survey of 2,400 global employees and C-suite leaders rele

model-releasessiliconangle
9 Apr 2026
Agents

its a directionally correct form factor but way too much lock in

DGX agent

its a directionally correct form factor but way too much lock in @hwchase17 Hey Harrison, We use Deep Agents pretty heavily and love it. Curious what you think about the new managed agents from Anthro

agentsharrison-chase--x
9 Apr 2026
Agents

AgonAlpha: Autonomous Alpha Discovery via Prompt Economy and Scalable Agentic Search

DGX agent

arXiv:2608.11250v1 Announce Type: new Abstract: Language models can propose many plausible trading factors, but an autonomous research system must also allocate its evaluation budget, verify its own e

agentsarxiv-cs-ai
13 Aug 2026
Model Releases

An Agentic Workflow for Legacy HPC Modernization: Converting the Two-Electron-Integral Core of GAMESS

DGX agent

arXiv:2608.12249v1 Announce Type: new Abstract: Modernizing legacy Fortran is a problem of volume: the transformations are individually routine, but the codebases can be enormous, and across much of c

model-releasesarxiv-cs-ai
13 Aug 2026
Research

Backdoor Decontamination Dynamics in LLM Agents

DGX agent

arXiv:2608.11295v1 Announce Type: cross Abstract: Open-weight LLM agents are vulnerable to backdoors installed during fine-tuning, which may be undetectable if the trigger conditions are never met dur

researcharxiv-cs-ai
13 Aug 2026
Model Releases

🧩 DeepSeek Harness v0.1 is now available in Developer Preview! 🔹 We’re opening it up to developers building agent harnesses worldwide and …

DGX agent

🧩 DeepSeek Harness v0.1 is now available in Developer Preview! 🔹 We’re opening it up to developers building agent harnesses worldwide and open-sourcing the codebase in MIT license. 🔹 Powered by the Co

model-releasesdeepseek--x
13 Aug 2026
Model Releases

FrontierFinance: A Challenging Benchmark for Measuring Frontier Intelligence of Finance Agents

DGX agent

arXiv:2608.11683v1 Announce Type: new Abstract: AI agents are increasingly deployed for professional investment research, yet no benchmark captures the complexity of the full investor workflow. Existi

model-releasesarxiv-cs-ai
13 Aug 2026
Model Releases

Harnessing agent memory to build lifelong AI partners for materials scientists

DGX agent

arXiv:2608.11224v1 Announce Type: new Abstract: Materials research advances through accumulated experience - scripts that work, protocols that are trusted, warnings attached to failed calculations or

model-releasesarxiv-cs-ai
13 Aug 2026
Safety

LoongReflect: Boosting Long-Horizon Reflection in Search Agents via Global Perspective Distillation

DGX agent

arXiv:2608.11967v1 Announce Type: cross Abstract: Large language model agents increasingly rely on long-horizon reasoning to solve complex tasks involving planning, tool use, and memory. A critical ca

safetyarxiv-cs-ai
13 Aug 2026
Safety

One Frozen Simulator Is Not Enough: Simulator Collapse in Multi-Agent RL

DGX agent

arXiv:2608.12253v1 Announce Type: cross Abstract: Multi-agent reinforcement learning for human-AI interaction typically relies on a single large language model to simulate user behavior. We show that

safetyarxiv-cs-ai
13 Aug 2026
Hardware

Ready Cohorts: Bounding GPU Opportunity and Avoiding Host Round Trips in LLM-Agent Control

DGX agent

arXiv:2608.12123v1 Announce Type: cross Abstract: LLM-agent services repeatedly execute small deterministic transitions between model and tool calls: route an outcome, update state, and emit the next

hardwarearxiv-cs-ai
13 Aug 2026
Model Releases

RecSys Factory: Bounding LLM Agent Autonomy to Decision Points in the Industrial Recommender Lifecycle

DGX agent

arXiv:2608.11241v1 Announce Type: new Abstract: Deploying LLM agents into industrial recommender operations exposes a three-way tension we frame as the autonomy-determinism-efficiency trilemma: genera

model-releasesarxiv-cs-ai
13 Aug 2026
Model Releases

From Sync APIs to support for the GPT-5 model series and agentic workflows: What’s new in Azure Content Understanding – August 2026

DGX agent

Enterprise content is no longer just something people read. AI apps and agents are only as useful as the information they can understand, yet much of the world’s enterprise knowledge is locked in docu

model-releasesmicrosoft-foundry
12 Aug 2026
Model Releases

Qwen3.8-2.4T-A95B is now live on Together AI. The Qwen Team’s latest flagship model is built for coding and long-horizon agent workflows, wi…

DGX agent

The Qwen Team has released its flagship model, Qwen3.8‑2.4T‑A95B, on the Together AI platform (togethercompute) as of August 12 2026. This 2.4‑trillion‑parameter model is engineered for coding tasks a

model-releasestogether-ai--x
12 Aug 2026
Model Releases

360CityArena: A Realistic Virtual Urban Navigation Benchmark for Embodied Agents

DGX agent

arXiv:2608.08814v1 Announce Type: cross Abstract: We present 360CityArena, a benchmark for evaluating the urban exploration capabilities of embodied agents within a photorealistic environment construc

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

An Agentic AI Framework Overcomes Fundamental Limitations of Large Language Models for Glaucoma Detection from Fundus Photography

DGX agent

arXiv:2608.07651v1 Announce Type: new Abstract: Large language models (LLMs) show promise in medical image interpretation but suffer from hallucination, limited accuracy, and run-to-run inconsistency.

model-releasesarxiv-cs-ai
11 Aug 2026
Safety

An Agentic Generative Large Language Model for Treatment Planning of Colorectal Cancer

DGX agent

arXiv:2608.09142v1 Announce Type: new Abstract: Treatment planning in precision oncology requires synthesizing heterogeneous patient information with rapidly evolving clinical guidelines to ensure gui

safetyarxiv-cs-cl
11 Aug 2026
Model Releases

BibTeX Citation Errors in Scientific Publishing Agents: Evaluation and Mitigation

DGX agent

arXiv:2604.03159v2 Announce Type: replace-cross Abstract: Large language models with web search are increasingly used in scientific publishing agents, yet they produce BibTeX entries with pervasive fi

model-releasesarxiv-cs-cl
11 Aug 2026
Agents

CARD: Controlled Agentic Reddit Discussions for Credit Card Simulation

DGX agent

arXiv:2608.09790v1 Announce Type: new Abstract: Online credit card discussions provide a natural setting for studying how consumers communicate about financial products. Simulating these discussions r

agentsarxiv-cs-ai
11 Aug 2026
Model Releases

ComboShoppingBench: Evaluating LLM Agents for Budget-Constrained Basket Shopping with Coupons

DGX agent

arXiv:2608.09282v1 Announce Type: new Abstract: Real-world shopping often requires constructing a basket of complementary items rather than retrieving a single product. Such combo-shopping tasks arise

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Hierarchical Self-Improvement: A Framework for Task-Specific Evolvable Agent Harnesses

DGX agent

arXiv:2608.08466v1 Announce Type: new Abstract: Modern LLM agents are often improved by modifying prompts, tools, or workflows manually, while the executable scaffold surrounding the model---the harne

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

NVIDIA Nemotron 3.5 Lightning is now live on Together AI. The fastest open model in its class is built for always-on agents that need to com…

DGX agent

NVIDIA’s Nemotron 3.5 Lightning—a fast open AI model designed for always‑on agents that perform high‑volume, specialized work—has been launched on the Together AI platform as of 11 August 2026. NVIDIA

model-releasestogether-ai--x
11 Aug 2026
Safety

Search-G1: Grounded Search Agents via Representation-Based Intrinsic Rewards

DGX agent

arXiv:2608.07531v1 Announce Type: cross Abstract: Search-augmented language agents should retrieve external information only when necessary and ground their answers in retrieved evidence. Existing ext

safetyarxiv-cs-ai
11 Aug 2026
Model Releases

Towards Researcher Agents for Knowledge-Graph Question Answering

DGX agent

arXiv:2608.07700v1 Announce Type: new Abstract: Translating a natural-language question into a SPARQL query that can be executed against a large knowledge graph requires resolving lexical ambiguity, g

model-releasesarxiv-cs-ai
11 Aug 2026
Research

Tree-of-Experience: Hierarchical Experience Management for Self-Evolving Agents

DGX agent

arXiv:2608.09044v1 Announce Type: new Abstract: Continual self-evolution requires LLM agents to transform environmental interactions into reliable and reusable experience. Existing methods typically r

researcharxiv-cs-cl
11 Aug 2026
← Previous
1…139140141142143…375
Next →