AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,959 results
14 Apr 2026

Multi-ORFT: Stable Online Reinforcement Fine-Tuning for Multi-Agent Diffusion Planning in Cooperative Driving

Model ReleasesDGX agent

arXiv:2604.11734v1 Announce Type: cross Abstract: Closed-loop cooperative driving requires planners that generate realistic multimodal multi-agent trajectories while improving safety and traffic effic

OpeFlo: Automated UX Evaluation via Simulated Human Web Interaction with GUI Grounding

AgentsDGX agent

arXiv:2604.09581v1 Announce Type: new Abstract: Evaluating web usability typically requires time-consuming user studies and expert reviews, which often limits iteration speed during product developmen

Playing Along: Learning a Double-Agent Defender for Belief Steering via Theory of Mind

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.11666v1 Announce Type: cross Abstract: As large language models (LLMs) become the engine behind conversational systems, their ability to reason about the intentions and states of their dial

The Devil is in the Details -- From OCR for Old Church Slavonic to Purely Visual Stemma Reconstruction

AgentsDGX agent

arXiv:2604.11724v1 Announce Type: new Abstract: The age of artificial intelligence has brought many new possibilities and pitfalls in many fields and tasks. The devil is in the details, and those come

Time is Not a Label: Continuous Phase Rotation for Temporal Knowledge Graphs and Agentic Memory

Model ReleasesDGX agent

arXiv:2604.11544v1 Announce Type: cross Abstract: Structured memory representations such as knowledge graphs are central to autonomous agents and other long-lived systems. However, most existing appro

13 Apr 2026

AgentCE-Bench: Agent Configurable Evaluation with Scalable Horizons and Controllable Difficulty under Lightweight Environments

Model ReleasesDGX agent

arXiv:2604.06111v2 Announce Type: replace Abstract: Existing Agent benchmarks suffer from two critical limitations: high environment interaction overhead (up to 41% of total evaluation time) and imbal

Microsoft says it is 'exploring the potential of technologies like OpenClaw in an enterprise context', including a team of always-on agents within Microsoft 365 (Aaron Holmes/The Information)

ApplicationsDGX agent

Aaron Holmes / The Information: Microsoft says it is “exploring the potential of technologies like OpenClaw in an enterprise context”, including a team of always-on agents within Microsoft 365 — As Mi

Plasticity-Enhanced Multi-Agent Mixture of Experts for Dynamic Objective Adaptation in UAVs-Assisted Emergency Communication Networks

SafetyDGX agent

arXiv:2604.09028v1 Announce Type: cross Abstract: Unmanned aerial vehicles serving as aerial base stations can rapidly restore connectivity after disasters, yet abrupt changes in user mobility and tra

SEA-Eval: A Benchmark for Evaluating Self-Evolving Agents Beyond Episodic Assessment

Model ReleasesDGX agent

arXiv:2604.08988v1 Announce Type: new Abstract: Current LLM-based agents demonstrate strong performance in episodic task execution but remain constrained by static toolsets and episodic amnesia, faili

SPASM: Stable Persona-driven Agent Simulation for Multi-turn Dialogue Generation

Model ReleasesDGX agent

arXiv:2604.09212v1 Announce Type: new Abstract: Large language models are increasingly deployed in multi-turn settings such as tutoring, support, and counseling, where reliability depends on preservin

12 Apr 2026

'every memory system is choosing a position on the raw/derived spectrum. and neither extreme works.' the axis nobody talks about: ownership …

AgentsDGX agent

'every memory system is choosing a position on the raw/derived spectrum. and neither extreme works.' the axis nobody talks about: ownership > memory is your agent's compounding advantage > derived fac

Frameworks For Supporting LLM/Agentic Benchmarking [P]

AgentsDGX agent

This r/MachineLearning post discusses the landscape of frameworks and tools used to support benchmarking of LLMs and agentic AI systems, covering how to systematically evaluate model capabilities beyo

Memory is where the harness stops being a wrapper and becomes an ownership layer. Once it controls what gets remembered, retrieved, compress…

AgentsDGX agent

Memory is where the harness stops being a wrapper and becomes an ownership layer. Once it controls what gets remembered, retrieved, compressed, and acted on, it starts shaping the agent’s judgment, no

11 Apr 2026

I built modern AI client for Mac with agentic tools, elegant UI, interactive charts and maps, sortable tables, Slack-like threads and access to local and cloud models

AgentsDGX agent

A Reddit post on r/ollama showcasing a community-built, feature-rich macOS AI client designed for both local and cloud model access, including support for Ollama. The application emphasizes a modern,

All People

ConceptsDGX agent

Auto-generated index of all people mentioned across the wiki.

10 Apr 2026

Are GUI Agents Focused Enough? Automated Distraction via Semantic-level UI Element Injection

SafetyDGX agent

arXiv:2604.07831v1 Announce Type: cross Abstract: Existing red-teaming studies on GUI agents have important limitations. Adversarial perturbations typically require white-box access, which is unavaila

Exploring Plan Space through Conversation: An Agentic Framework for LLM-Mediated Explanations in Planning

AgentsDGX agent

arXiv:2603.02070v2 Announce Type: replace-cross Abstract: When automating plan generation for a real-world sequential decision problem, the goal is often not to replace the human planner, but to facil

Learning to Negotiate: Multi-Agent Deliberation for Collective Value Alignment in LLMs

SafetyDGX agent

arXiv:2603.10476v2 Announce Type: replace Abstract: LLM alignment has progressed in single-agent settings through paradigms such as RL with human feedback (RLHF), while recent work explores scalable a

Lighting-grounded Video Generation with Renderer-based Agent Reasoning

AgentsDGX agent

arXiv:2604.07966v1 Announce Type: new Abstract: Diffusion models have achieved remarkable progress in video generation, but their controllability remains a major limitation. Key scene factors such as

we are doing this podcast specifically to highlight the innermost details of building agents in production may be a little niche, but its a …

ApplicationsDGX agent

we are doing this podcast specifically to highlight the innermost details of building agents in production may be a little niche, but its a niche i like 🤷‍♂️ @hwchase17 Great content, really appreciat

what he said 🗣️ the very best agents today obsessively tailor the harness layer around the model I’m looking at you “5 things I learned fro…

Model ReleasesDGX agent

what he said 🗣️ the very best agents today obsessively tailor the harness layer around the model I’m looking at you “5 things I learned from the Claude code leak” bros 👀 orchestration patterns, tool d

9 Apr 2026

Is a backlash brewing? Rapid innovation in AI coding and agents may force push for enterprise order and control

Model ReleasesDGX agent

Artificial intelligence is proving to be a big bet for many companies across the enterprise landscape, and the gamblers are getting nervous. A survey of 2,400 global employees and C-suite leaders rele

its a directionally correct form factor but way too much lock in

AgentsDGX agent

its a directionally correct form factor but way too much lock in @hwchase17 Hey Harrison, We use Deep Agents pretty heavily and love it. Curious what you think about the new managed agents from Anthro

13 Aug 2026

AgonAlpha: Autonomous Alpha Discovery via Prompt Economy and Scalable Agentic Search

AgentsDGX agent

arXiv:2608.11250v1 Announce Type: new Abstract: Language models can propose many plausible trading factors, but an autonomous research system must also allocate its evaluation budget, verify its own e

An Agentic Workflow for Legacy HPC Modernization: Converting the Two-Electron-Integral Core of GAMESS

Model ReleasesDGX agent

arXiv:2608.12249v1 Announce Type: new Abstract: Modernizing legacy Fortran is a problem of volume: the transformations are individually routine, but the codebases can be enormous, and across much of c

Backdoor Decontamination Dynamics in LLM Agents

ResearchDGX agent

arXiv:2608.11295v1 Announce Type: cross Abstract: Open-weight LLM agents are vulnerable to backdoors installed during fine-tuning, which may be undetectable if the trigger conditions are never met dur

🧩 DeepSeek Harness v0.1 is now available in Developer Preview! 🔹 We’re opening it up to developers building agent harnesses worldwide and …

Model ReleasesDGX agent

🧩 DeepSeek Harness v0.1 is now available in Developer Preview! 🔹 We’re opening it up to developers building agent harnesses worldwide and open-sourcing the codebase in MIT license. 🔹 Powered by the Co

FrontierFinance: A Challenging Benchmark for Measuring Frontier Intelligence of Finance Agents

Model ReleasesDGX agent

arXiv:2608.11683v1 Announce Type: new Abstract: AI agents are increasingly deployed for professional investment research, yet no benchmark captures the complexity of the full investor workflow. Existi

Harnessing agent memory to build lifelong AI partners for materials scientists

Model ReleasesDGX agent

arXiv:2608.11224v1 Announce Type: new Abstract: Materials research advances through accumulated experience - scripts that work, protocols that are trusted, warnings attached to failed calculations or

LoongReflect: Boosting Long-Horizon Reflection in Search Agents via Global Perspective Distillation

SafetyDGX agent

arXiv:2608.11967v1 Announce Type: cross Abstract: Large language model agents increasingly rely on long-horizon reasoning to solve complex tasks involving planning, tool use, and memory. A critical ca

One Frozen Simulator Is Not Enough: Simulator Collapse in Multi-Agent RL

SafetyDGX agent

arXiv:2608.12253v1 Announce Type: cross Abstract: Multi-agent reinforcement learning for human-AI interaction typically relies on a single large language model to simulate user behavior. We show that

Ready Cohorts: Bounding GPU Opportunity and Avoiding Host Round Trips in LLM-Agent Control

HardwareDGX agent

arXiv:2608.12123v1 Announce Type: cross Abstract: LLM-agent services repeatedly execute small deterministic transitions between model and tool calls: route an outcome, update state, and emit the next

RecSys Factory: Bounding LLM Agent Autonomy to Decision Points in the Industrial Recommender Lifecycle

Model ReleasesDGX agent

arXiv:2608.11241v1 Announce Type: new Abstract: Deploying LLM agents into industrial recommender operations exposes a three-way tension we frame as the autonomy-determinism-efficiency trilemma: genera

12 Aug 2026

From Sync APIs to support for the GPT-5 model series and agentic workflows: What’s new in Azure Content Understanding – August 2026

Model ReleasesDGX agent

Enterprise content is no longer just something people read. AI apps and agents are only as useful as the information they can understand, yet much of the world’s enterprise knowledge is locked in docu

Qwen3.8-2.4T-A95B is now live on Together AI. The Qwen Team’s latest flagship model is built for coding and long-horizon agent workflows, wi…

Model ReleasesDGX agent

The Qwen Team has released its flagship model, Qwen3.8‑2.4T‑A95B, on the Together AI platform (togethercompute) as of August 12 2026. This 2.4‑trillion‑parameter model is engineered for coding tasks a

11 Aug 2026

360CityArena: A Realistic Virtual Urban Navigation Benchmark for Embodied Agents

Model ReleasesDGX agent

arXiv:2608.08814v1 Announce Type: cross Abstract: We present 360CityArena, a benchmark for evaluating the urban exploration capabilities of embodied agents within a photorealistic environment construc

An Agentic AI Framework Overcomes Fundamental Limitations of Large Language Models for Glaucoma Detection from Fundus Photography

Model ReleasesDGX agent

arXiv:2608.07651v1 Announce Type: new Abstract: Large language models (LLMs) show promise in medical image interpretation but suffer from hallucination, limited accuracy, and run-to-run inconsistency.

An Agentic Generative Large Language Model for Treatment Planning of Colorectal Cancer

SafetyDGX agent

arXiv:2608.09142v1 Announce Type: new Abstract: Treatment planning in precision oncology requires synthesizing heterogeneous patient information with rapidly evolving clinical guidelines to ensure gui

BibTeX Citation Errors in Scientific Publishing Agents: Evaluation and Mitigation

Model ReleasesDGX agent

arXiv:2604.03159v2 Announce Type: replace-cross Abstract: Large language models with web search are increasingly used in scientific publishing agents, yet they produce BibTeX entries with pervasive fi

CARD: Controlled Agentic Reddit Discussions for Credit Card Simulation

AgentsDGX agent

arXiv:2608.09790v1 Announce Type: new Abstract: Online credit card discussions provide a natural setting for studying how consumers communicate about financial products. Simulating these discussions r

ComboShoppingBench: Evaluating LLM Agents for Budget-Constrained Basket Shopping with Coupons

Model ReleasesDGX agent

arXiv:2608.09282v1 Announce Type: new Abstract: Real-world shopping often requires constructing a basket of complementary items rather than retrieving a single product. Such combo-shopping tasks arise

Hierarchical Self-Improvement: A Framework for Task-Specific Evolvable Agent Harnesses

Model ReleasesDGX agent

arXiv:2608.08466v1 Announce Type: new Abstract: Modern LLM agents are often improved by modifying prompts, tools, or workflows manually, while the executable scaffold surrounding the model---the harne

NVIDIA Nemotron 3.5 Lightning is now live on Together AI. The fastest open model in its class is built for always-on agents that need to com…

Model ReleasesDGX agent

NVIDIA’s Nemotron 3.5 Lightning—a fast open AI model designed for always‑on agents that perform high‑volume, specialized work—has been launched on the Together AI platform as of 11 August 2026. NVIDIA

Search-G1: Grounded Search Agents via Representation-Based Intrinsic Rewards

SafetyDGX agent

arXiv:2608.07531v1 Announce Type: cross Abstract: Search-augmented language agents should retrieve external information only when necessary and ground their answers in retrieved evidence. Existing ext

Towards Researcher Agents for Knowledge-Graph Question Answering

Model ReleasesDGX agent

arXiv:2608.07700v1 Announce Type: new Abstract: Translating a natural-language question into a SPARQL query that can be executed against a large knowledge graph requires resolving lexical ambiguity, g

Tree-of-Experience: Hierarchical Experience Management for Self-Evolving Agents

ResearchDGX agent

arXiv:2608.09044v1 Announce Type: new Abstract: Continual self-evolution requires LLM agents to transform environmental interactions into reliable and reusable experience. Existing methods typically r

What Keeps Agent Skills from Being Reusable? Evidence from 138K SKILL.md Files

SafetyDGX agent

arXiv:2608.08453v1 Announce Type: new Abstract: Under the current standard, Agent Skills are SKILL.md files that combine instructions with supporting resources, enabling Large Language Model (LLM) age

10 Aug 2026

Evolve your marketing with new AI tools

AgentsDGX agent

On August 10 2026 Google launched a suite of AI‑powered features for Google Ads and Google Analytics designed to automate routine tasks and deliver smarter campaign insights. These agentic experiences

Explicit, Not Longer: What Makes Epistemic Stance Survive Memory Compression

AgentsDGX agent

arXiv:2608.06953v1 Announce Type: cross Abstract: Agent memory systems compress what they store, and compression is built to drop qualifiers, so a claim's epistemic standing tends not to survive being

MemWM: Memory-Augmented Text-Based World Model

Model ReleasesDGX agent

arXiv:2608.07107v1 Announce Type: new Abstract: World models are increasingly used to support planning in agents by predicting how environment states evolve in response to agent actions. Yet fluent ne

👀 Seeing is just the beginning. With Qwen-MM-Plugins, turn your favorite agent harness multimodal-native — read images, videos & documents,…

Model ReleasesDGX agent

👀 Seeing is just the beginning. With Qwen-MM-Plugins, turn your favorite agent harness multimodal-native — read images, videos & documents, edit videos, work with 3D/CAD, and more. From multimodal mod

8 Aug 2026

Neat example here of the agents communicating purely through file names, including adding base64-encoded attachments and using 'zz' prefixes…

ToolsDGX agent

Neat example here of the agents communicating purely through file names, including adding base64-encoded attachments and using 'zz' prefixes to ensure their new message sorts to the bottom of the list

7 Aug 2026

A Two-Tier Perspective on Inference-Time Parallelism in Multi-Agent LLM Systems

Model ReleasesDGX agent

arXiv:2608.05791v1 Announce Type: cross Abstract: Large language model (LLM)-driven multi-agent systems typically require multiple model invocations and complex coordination during inference, and thei

Benchmarking the Benchmarks: Evaluating Benchmarks for Conversational Agents

Model ReleasesDGX agent

arXiv:2608.06329v1 Announce Type: cross Abstract: Task-oriented conversational agents are evaluated using curated or automatically generated benchmarks, yet benchmark quality is rarely assessed. Poor

BlockPython: A Process-Aware Agent-Supported Platform for the Transition from Block-Based to Python Programming

AgentsDGX agent

arXiv:2608.05716v1 Announce Type: new Abstract: The transition from block-based to text-based programming requires learners to convert visible program structures into abstract textual expressions, whi

CASCADE: An Agentic Regulatory Network Framework for Patient-Data-Validated Downstream Perturbation Prediction

Model ReleasesDGX agent

arXiv:2608.05359v1 Announce Type: new Abstract: CASCADE is an agentic framework that predicts downstream transcriptional effects of gene perturbation from precomputed ARACNe regulatory networks, expos

Causal Episodic Memory for Feedback-Driven Agent Repair

Model ReleasesDGX agent

arXiv:2608.05906v1 Announce Type: new Abstract: LLM agents that repair failures often discard successful corrections, forcing later episodes to rediscover similar solutions. We study whether finalized

ChainClaw: A Layered Agent Framework for Reliable On-Chain Execution

Model ReleasesDGX agent

arXiv:2608.05790v1 Announce Type: new Abstract: General-purpose large language model agents have achieved strong performance on tool-augmented tasks, yet they rely on assumptions break down in blockch

Contextual Information Policy Optimization for Search Agents

SafetyDGX agent

arXiv:2608.06128v1 Announce Type: new Abstract: Search agents extend large language models beyond static parametric memory by enabling them to acquire and use ex ternal evidence during multi-step reas

Multi-Agent Reinforcement Learning for Online Traffic Scheduling in Time-Sensitive Application

SafetyDGX agent

arXiv:2608.05346v1 Announce Type: cross Abstract: Time-sensitive networking (TSN) is increasingly integrated into mobile edge computing (MEC) to support applications with stringent latency requirement

← Previous
1…111112113114115…300
Next →