AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,986 results
Model Releases

Thought-Retriever: Don't Just Retrieve Raw Data, Retrieve Thoughts for Memory-Augmented Agentic Systems

DGX agent

arXiv:2604.12231v1 Announce Type: new Abstract: Large language models (LLMs) have transformed AI research thanks to their powerful internal capabilities and knowledge. However, existing LLMs still fai

model-releasesarxiv-cs-cl
15 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Towards Long-horizon Agentic Multimodal Search

DGX agent

arXiv:2604.12890v1 Announce Type: cross Abstract: Multimodal deep search agents have shown great potential in solving complex tasks by iteratively collecting textual and visual evidence. However, mana

model-releasesarxiv-cs-ai
15 Apr 2026
Safety

A Dual-Positive Monotone Parameterization for Multi-Segment Bids and a Validity Assessment Framework for Reinforcement Learning Agent-based Simulation of Electricity Markets

DGX agent

arXiv:2604.10252v1 Announce Type: new Abstract: Reinforcement learning agent-based simulation (RL-ABS) has become an important tool for electricity market mechanism analysis and evaluation. In the mod

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

From Translation to Superset: Benchmark-Driven Evolution of a Production AI Agent from Rust to Python

DGX agent

arXiv:2604.11518v1 Announce Type: cross Abstract: Cross-language migration of large software systems is a persistent engineering challenge, particularly when the source codebase evolves rapidly. We pr

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

Hubble: An LLM-Driven Agentic Framework for Safe and Automated Alpha Factor Discovery

DGX agent

arXiv:2604.09601v1 Announce Type: new Abstract: Discovering predictive alpha factors in quantitative finance remains a formidable challenge due to the vast combinatorial search space and inherently lo

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

I have a Macbook AIR M5 Base and I want to run an Agentic Coding program, similar to Claude Code or Codex. Besides the model, how do I do it? I've already tried with Ollama, VS Code, Opencode, and haven't been able to. (I'm not a developer, sorry)

DGX agent

This Reddit thread addresses a common challenge for non-developers trying to run a local agentic coding assistant on a MacBook Air M5: while tools like Ollama, VS Code, and OpenCode are the right piec

model-releasesr-ollama
14 Apr 2026
Agents

Improving Layout Representation Learning Across Inconsistently Annotated Datasets via Agentic Harmonization

DGX agent

arXiv:2604.11042v1 Announce Type: new Abstract: Fine-tuning object detection (OD) models on combined datasets assumes annotation compatibility, yet datasets often encode conflicting spatial definition

agentsarxiv-cs-cv
14 Apr 2026
Agents

Instructing LLMs to Negotiate using Reinforcement Learning with Verifiable Rewards

DGX agent

arXiv:2604.09855v1 Announce Type: new Abstract: The recent advancement of Large Language Models (LLMs) has established their potential as autonomous interactive agents. However, they often struggle in

agentsarxiv-cs-ai
14 Apr 2026
Local Ai

Love seeing this open-sourced. Had a great chat with @nicoalbanese10 some weeks ago where he hinted to something like this. Great reference …

DGX agent

Love seeing this open-sourced. Had a great chat with @nicoalbanese10 some weeks ago where he hinted to something like this. Great reference architecture for cloud coding agents. Open Agents gives you

local-aiharrison-chase--x
14 Apr 2026
Safety

MADQRL: Distributed Quantum Reinforcement Learning Framework for Multi-Agent Environments

DGX agent

arXiv:2604.11131v1 Announce Type: new Abstract: Reinforcement learning (RL) is one of the most practical ways to learn from real-life use-cases. Motivated from the cognitive methods used by humans mak

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

Multi-ORFT: Stable Online Reinforcement Fine-Tuning for Multi-Agent Diffusion Planning in Cooperative Driving

DGX agent

arXiv:2604.11734v1 Announce Type: cross Abstract: Closed-loop cooperative driving requires planners that generate realistic multimodal multi-agent trajectories while improving safety and traffic effic

model-releasesarxiv-cs-ai
14 Apr 2026
Agents

OpeFlo: Automated UX Evaluation via Simulated Human Web Interaction with GUI Grounding

DGX agent

arXiv:2604.09581v1 Announce Type: new Abstract: Evaluating web usability typically requires time-consuming user studies and expert reviews, which often limits iteration speed during product developmen

agentsarxiv-cs-ai
14 Apr 2026
Model Releases

Playing Along: Learning a Double-Agent Defender for Belief Steering via Theory of Mind

DGX agent

arXiv:2604.11666v1 Announce Type: cross Abstract: As large language models (LLMs) become the engine behind conversational systems, their ability to reason about the intentions and states of their dial

model-releasesarxiv-cs-ai
14 Apr 2026
Agents

The Devil is in the Details -- From OCR for Old Church Slavonic to Purely Visual Stemma Reconstruction

DGX agent

arXiv:2604.11724v1 Announce Type: new Abstract: The age of artificial intelligence has brought many new possibilities and pitfalls in many fields and tasks. The devil is in the details, and those come

agentsarxiv-cs-cv
14 Apr 2026
Model Releases

Time is Not a Label: Continuous Phase Rotation for Temporal Knowledge Graphs and Agentic Memory

DGX agent

arXiv:2604.11544v1 Announce Type: cross Abstract: Structured memory representations such as knowledge graphs are central to autonomous agents and other long-lived systems. However, most existing appro

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

AgentCE-Bench: Agent Configurable Evaluation with Scalable Horizons and Controllable Difficulty under Lightweight Environments

DGX agent

arXiv:2604.06111v2 Announce Type: replace Abstract: Existing Agent benchmarks suffer from two critical limitations: high environment interaction overhead (up to 41% of total evaluation time) and imbal

model-releasesarxiv-cs-ai
13 Apr 2026
Applications

Microsoft says it is 'exploring the potential of technologies like OpenClaw in an enterprise context', including a team of always-on agents within Microsoft 365 (Aaron Holmes/The Information)

DGX agent

Aaron Holmes / The Information: Microsoft says it is “exploring the potential of technologies like OpenClaw in an enterprise context”, including a team of always-on agents within Microsoft 365 — As Mi

applicationstechmeme
13 Apr 2026
Safety

Plasticity-Enhanced Multi-Agent Mixture of Experts for Dynamic Objective Adaptation in UAVs-Assisted Emergency Communication Networks

DGX agent

arXiv:2604.09028v1 Announce Type: cross Abstract: Unmanned aerial vehicles serving as aerial base stations can rapidly restore connectivity after disasters, yet abrupt changes in user mobility and tra

safetyarxiv-cs-lg
13 Apr 2026
Model Releases

SEA-Eval: A Benchmark for Evaluating Self-Evolving Agents Beyond Episodic Assessment

DGX agent

arXiv:2604.08988v1 Announce Type: new Abstract: Current LLM-based agents demonstrate strong performance in episodic task execution but remain constrained by static toolsets and episodic amnesia, faili

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

SPASM: Stable Persona-driven Agent Simulation for Multi-turn Dialogue Generation

DGX agent

arXiv:2604.09212v1 Announce Type: new Abstract: Large language models are increasingly deployed in multi-turn settings such as tutoring, support, and counseling, where reliability depends on preservin

model-releasesarxiv-cs-cl
13 Apr 2026
Agents

'every memory system is choosing a position on the raw/derived spectrum. and neither extreme works.' the axis nobody talks about: ownership …

DGX agent

'every memory system is choosing a position on the raw/derived spectrum. and neither extreme works.' the axis nobody talks about: ownership > memory is your agent's compounding advantage > derived fac

agentsharrison-chase--x
12 Apr 2026
Agents

Frameworks For Supporting LLM/Agentic Benchmarking [P]

DGX agent

This r/MachineLearning post discusses the landscape of frameworks and tools used to support benchmarking of LLMs and agentic AI systems, covering how to systematically evaluate model capabilities beyo

agentsr-machinelearning
12 Apr 2026
Agents

Memory is where the harness stops being a wrapper and becomes an ownership layer. Once it controls what gets remembered, retrieved, compress…

DGX agent

Memory is where the harness stops being a wrapper and becomes an ownership layer. Once it controls what gets remembered, retrieved, compressed, and acted on, it starts shaping the agent’s judgment, no

agentsharrison-chase--x
12 Apr 2026
Agents

I built modern AI client for Mac with agentic tools, elegant UI, interactive charts and maps, sortable tables, Slack-like threads and access to local and cloud models

DGX agent

A Reddit post on r/ollama showcasing a community-built, feature-rich macOS AI client designed for both local and cloud model access, including support for Ollama. The application emphasizes a modern,

agentsr-ollama
11 Apr 2026
Concepts

All People

DGX agent

Auto-generated index of all people mentioned across the wiki.

conceptspeopleindex
11 Apr 2026
Safety

Are GUI Agents Focused Enough? Automated Distraction via Semantic-level UI Element Injection

DGX agent

arXiv:2604.07831v1 Announce Type: cross Abstract: Existing red-teaming studies on GUI agents have important limitations. Adversarial perturbations typically require white-box access, which is unavaila

safetyarxiv-cs-cl
10 Apr 2026
Agents

Exploring Plan Space through Conversation: An Agentic Framework for LLM-Mediated Explanations in Planning

DGX agent

arXiv:2603.02070v2 Announce Type: replace-cross Abstract: When automating plan generation for a real-world sequential decision problem, the goal is often not to replace the human planner, but to facil

agentsarxiv-cs-cl
10 Apr 2026
Safety

Learning to Negotiate: Multi-Agent Deliberation for Collective Value Alignment in LLMs

DGX agent

arXiv:2603.10476v2 Announce Type: replace Abstract: LLM alignment has progressed in single-agent settings through paradigms such as RL with human feedback (RLHF), while recent work explores scalable a

safetyarxiv-cs-cl
10 Apr 2026
Agents

Lighting-grounded Video Generation with Renderer-based Agent Reasoning

DGX agent

arXiv:2604.07966v1 Announce Type: new Abstract: Diffusion models have achieved remarkable progress in video generation, but their controllability remains a major limitation. Key scene factors such as

agentsarxiv-cs-cv
10 Apr 2026
Applications

we are doing this podcast specifically to highlight the innermost details of building agents in production may be a little niche, but its a …

DGX agent

we are doing this podcast specifically to highlight the innermost details of building agents in production may be a little niche, but its a niche i like 🤷‍♂️ @hwchase17 Great content, really appreciat

applicationsharrison-chase--x
10 Apr 2026
Model Releases

what he said 🗣️ the very best agents today obsessively tailor the harness layer around the model I’m looking at you “5 things I learned fro…

DGX agent

what he said 🗣️ the very best agents today obsessively tailor the harness layer around the model I’m looking at you “5 things I learned from the Claude code leak” bros 👀 orchestration patterns, tool d

model-releasesharrison-chase--x
10 Apr 2026
Model Releases

Is a backlash brewing? Rapid innovation in AI coding and agents may force push for enterprise order and control

DGX agent

Artificial intelligence is proving to be a big bet for many companies across the enterprise landscape, and the gamblers are getting nervous. A survey of 2,400 global employees and C-suite leaders rele

model-releasessiliconangle
9 Apr 2026
Agents

its a directionally correct form factor but way too much lock in

DGX agent

its a directionally correct form factor but way too much lock in @hwchase17 Hey Harrison, We use Deep Agents pretty heavily and love it. Curious what you think about the new managed agents from Anthro

agentsharrison-chase--x
9 Apr 2026
Agents

AgonAlpha: Autonomous Alpha Discovery via Prompt Economy and Scalable Agentic Search

DGX agent

arXiv:2608.11250v1 Announce Type: new Abstract: Language models can propose many plausible trading factors, but an autonomous research system must also allocate its evaluation budget, verify its own e

agentsarxiv-cs-ai
13 Aug 2026
Model Releases

An Agentic Workflow for Legacy HPC Modernization: Converting the Two-Electron-Integral Core of GAMESS

DGX agent

arXiv:2608.12249v1 Announce Type: new Abstract: Modernizing legacy Fortran is a problem of volume: the transformations are individually routine, but the codebases can be enormous, and across much of c

model-releasesarxiv-cs-ai
13 Aug 2026
Research

Backdoor Decontamination Dynamics in LLM Agents

DGX agent

arXiv:2608.11295v1 Announce Type: cross Abstract: Open-weight LLM agents are vulnerable to backdoors installed during fine-tuning, which may be undetectable if the trigger conditions are never met dur

researcharxiv-cs-ai
13 Aug 2026
Model Releases

🧩 DeepSeek Harness v0.1 is now available in Developer Preview! 🔹 We’re opening it up to developers building agent harnesses worldwide and …

DGX agent

🧩 DeepSeek Harness v0.1 is now available in Developer Preview! 🔹 We’re opening it up to developers building agent harnesses worldwide and open-sourcing the codebase in MIT license. 🔹 Powered by the Co

model-releasesdeepseek--x
13 Aug 2026
Model Releases

FrontierFinance: A Challenging Benchmark for Measuring Frontier Intelligence of Finance Agents

DGX agent

arXiv:2608.11683v1 Announce Type: new Abstract: AI agents are increasingly deployed for professional investment research, yet no benchmark captures the complexity of the full investor workflow. Existi

model-releasesarxiv-cs-ai
13 Aug 2026
Model Releases

Harnessing agent memory to build lifelong AI partners for materials scientists

DGX agent

arXiv:2608.11224v1 Announce Type: new Abstract: Materials research advances through accumulated experience - scripts that work, protocols that are trusted, warnings attached to failed calculations or

model-releasesarxiv-cs-ai
13 Aug 2026
Safety

LoongReflect: Boosting Long-Horizon Reflection in Search Agents via Global Perspective Distillation

DGX agent

arXiv:2608.11967v1 Announce Type: cross Abstract: Large language model agents increasingly rely on long-horizon reasoning to solve complex tasks involving planning, tool use, and memory. A critical ca

safetyarxiv-cs-ai
13 Aug 2026
Safety

One Frozen Simulator Is Not Enough: Simulator Collapse in Multi-Agent RL

DGX agent

arXiv:2608.12253v1 Announce Type: cross Abstract: Multi-agent reinforcement learning for human-AI interaction typically relies on a single large language model to simulate user behavior. We show that

safetyarxiv-cs-ai
13 Aug 2026
Hardware

Ready Cohorts: Bounding GPU Opportunity and Avoiding Host Round Trips in LLM-Agent Control

DGX agent

arXiv:2608.12123v1 Announce Type: cross Abstract: LLM-agent services repeatedly execute small deterministic transitions between model and tool calls: route an outcome, update state, and emit the next

hardwarearxiv-cs-ai
13 Aug 2026
Model Releases

RecSys Factory: Bounding LLM Agent Autonomy to Decision Points in the Industrial Recommender Lifecycle

DGX agent

arXiv:2608.11241v1 Announce Type: new Abstract: Deploying LLM agents into industrial recommender operations exposes a three-way tension we frame as the autonomy-determinism-efficiency trilemma: genera

model-releasesarxiv-cs-ai
13 Aug 2026
Model Releases

From Sync APIs to support for the GPT-5 model series and agentic workflows: What’s new in Azure Content Understanding – August 2026

DGX agent

Enterprise content is no longer just something people read. AI apps and agents are only as useful as the information they can understand, yet much of the world’s enterprise knowledge is locked in docu

model-releasesmicrosoft-foundry
12 Aug 2026
Model Releases

Qwen3.8-2.4T-A95B is now live on Together AI. The Qwen Team’s latest flagship model is built for coding and long-horizon agent workflows, wi…

DGX agent

The Qwen Team has released its flagship model, Qwen3.8‑2.4T‑A95B, on the Together AI platform (togethercompute) as of August 12 2026. This 2.4‑trillion‑parameter model is engineered for coding tasks a

model-releasestogether-ai--x
12 Aug 2026
Model Releases

360CityArena: A Realistic Virtual Urban Navigation Benchmark for Embodied Agents

DGX agent

arXiv:2608.08814v1 Announce Type: cross Abstract: We present 360CityArena, a benchmark for evaluating the urban exploration capabilities of embodied agents within a photorealistic environment construc

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

An Agentic AI Framework Overcomes Fundamental Limitations of Large Language Models for Glaucoma Detection from Fundus Photography

DGX agent

arXiv:2608.07651v1 Announce Type: new Abstract: Large language models (LLMs) show promise in medical image interpretation but suffer from hallucination, limited accuracy, and run-to-run inconsistency.

model-releasesarxiv-cs-ai
11 Aug 2026
Safety

An Agentic Generative Large Language Model for Treatment Planning of Colorectal Cancer

DGX agent

arXiv:2608.09142v1 Announce Type: new Abstract: Treatment planning in precision oncology requires synthesizing heterogeneous patient information with rapidly evolving clinical guidelines to ensure gui

safetyarxiv-cs-cl
11 Aug 2026
← Previous
1…139140141142143…375
Next →