AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,976 results
Model Releases

A^{2}utoLPBench: An Auto-Generated, Agent-Friendly LP Benchmark via Inverse-KKT Construction

DGX agent

arXiv:2607.02141v1 Announce Type: new Abstract: Most LP-from-text benchmarks are static datasets of word problems written and labeled by hand. Once such a dataset is released, its size is fixed, its d

model-releasesarxiv-cs-ai
3 Jul 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Local Ai

Auto-FL-Research: Agentic Search for Federated Learning Algorithms

DGX agent

arXiv:2607.01366v1 Announce Type: new Abstract: Federated learning (FL) research often depends on many small but consequential algorithmic choices: optimizer variants, server aggregation rules, local

local-aiarxiv-cs-ai
3 Jul 2026
Model Releases

Bringing Agentic Search to Earth Observation Data Discovery

DGX agent

arXiv:2607.02387v1 Announce Type: cross Abstract: NASA and its data centers hold thousands of geoscience datasets and tools like Worldview, Giovanni, the Science Discovery Engine, and Harmony. Finding

model-releasesarxiv-cs-lg
3 Jul 2026
Safety

Copewell: A Multi-Agent Swarm Architecture for Equitable Mental Wellness Support

DGX agent

arXiv:2607.02245v1 Announce Type: new Abstract: Mental health disorders affect nearly one billion people globally, yet 75% of individuals in low- and middle-income countries receive no treatment due t

safetyarxiv-cs-ai
3 Jul 2026
Local Ai

Repair the Amplifier, Not the Symptom: Stable World-Model Correction for Agent Rollouts

DGX agent

arXiv:2607.01767v1 Announce Type: new Abstract: As agent planning moves from short tool chains toward persistent workflows with thousands or tens of thousands of steps, failures will occur inside larg

local-aiarxiv-cs-ai
3 Jul 2026
Model Releases

Understanding Agent-Based Patching of Compiler Missed Optimizations

DGX agent

arXiv:2607.02370v1 Announce Type: cross Abstract: Compiler missed optimizations refer to cases in which compilers failed to optimize certain code. It takes many compiler developers' efforts to impleme

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

big week at langchain, with a lot of launches: 1/ OpenWiki - auto generate a wiki of a github repo 2/ two different voice agent tutorials 3/…

DGX agent

big week at langchain, with a lot of launches: 1/ OpenWiki - auto generate a wiki of a github repo 2/ two different voice agent tutorials 3/ Harbor integration and tutorial for long running, stateful

model-releasesharrison-chase--x
2 Jul 2026
Industry

Cloudflare sets a September 15 deadline for AI companies to differentiate their web crawlers into search, AI training, and AI agents or face being blocked (Samantha Elkins/NBC News)

DGX agent

Samantha Elkins / NBC News: Cloudflare sets a September 15 deadline for AI companies to differentiate their web crawlers into search, AI training, and AI agents or face being blocked — Cloudflare gave

industrytechmeme
2 Jul 2026
Tutorials

Hot take: I think it's still important to understand the code that our agents write! In this mega thread (based on my AIE talk today), I wil…

DGX agent

Hot take: I think it's still important to understand the code that our agents write! In this mega thread (based on my AIE talk today), I will explain why that's the case, and show some ideas for how t

tutorialsthariq--x
2 Jul 2026
Model Releases

i finally tried hermes agent and the hype is real btw. @NousResearch cooked. been onboarding my young relatives who can't afford Claude, sho…

DGX agent

i finally tried hermes agent and the hype is real btw. @NousResearch cooked. been onboarding my young relatives who can't afford Claude, showing them how to use $1-5 of tokens to bootstrap hermes and

model-releasesnous-research--x
2 Jul 2026
Applications

Memo: Microsoft is merging the consumer and enterprise versions of its Copilot chatbots into a single app featuring coding tools and AI agents dubbed AutoPilot (The Information)

DGX agent

The Information: Memo: Microsoft is merging the consumer and enterprise versions of its Copilot chatbots into a single app featuring coding tools and AI agents dubbed AutoPilot — Microsoft is merging

applicationstechmeme
2 Jul 2026
Safety

Multi-scale Mixture of World Models for Embodied Agents in Evolving Environments

DGX agent

arXiv:2607.00457v1 Announce Type: new Abstract: Embodied agents operating in the real world require multi-scale reasoning and knowledge adaptation as conditions change. We identify two challenges in a

safetyarxiv-cs-ai
2 Jul 2026
Safety

Self-Evolving Agents with Anytime-Valid Certificates

DGX agent

arXiv:2607.00871v1 Announce Type: new Abstract: Self-evolving agents violate the assumption behind most learning-theoretic guarantees: the data, evaluator, components, and hypothesis space are produce

safetyarxiv-cs-ai
2 Jul 2026
Model Releases

Your coding agent bill doubled and nobody can tell you why. Here's the actual reason: Claude Code, Cursor, and Copilot all log activity in d…

DGX agent

Your coding agent bill doubled and nobody can tell you why. Here's the actual reason: Claude Code, Cursor, and Copilot all log activity in different formats. The second your team uses more than one (t

model-releasesharrison-chase--x
2 Jul 2026
Model Releases

Z.ai launches ZCode, an 'Agentic Development Environment' optimized for its new GLM-5.2 model; Z.ai's GLM Coding Plan costs from 16.20 to 144 per month (Michael Nuñez/VentureBeat)

DGX agent

Michael Nuñez / VentureBeat: Z.ai launches ZCode, an “Agentic Development Environment” optimized for its new GLM-5.2 model; Z.ai's GLM Coding Plan costs from 16.20 to 144 per month — The move marks th

model-releasestechmeme
2 Jul 2026
Model Releases

“Agentic kernel optimization is the future of on-device inference” @xenovacom used Fable 5 to write kernels that pushed Gemma 4 to a massive…

DGX agent

“Agentic kernel optimization is the future of on-device inference” @xenovacom used Fable 5 to write kernels that pushed Gemma 4 to a massive 255 tok/s on WebGPU with M4. He shared the demo, so you can

model-releasesclem-delangue--x
1 Jul 2026
Hardware

AgRefactor: Self-Evolving Agentic Workflow for HLS Compatibility and Performance

DGX agent

arXiv:2606.30949v1 Announce Type: new Abstract: High-Level Synthesis (HLS) provides a fast path from concepts to silicon, but converting real-world software into synthesizable HLS code remains challen

hardwarearxiv-cs-ai
1 Jul 2026
Model Releases

AxDafny: Agentic Verified Code Generation in Dafny

DGX agent

arXiv:2606.32007v1 Announce Type: new Abstract: We study agentic code generation in Dafny, where a model must generate both executable code and the proof artifacts for verification. We present AxDafny

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

do you know what you pay for in agentic workloads? cached tokens! session with 50+ tool calls -> prompt is billed 50 times all providers giv…

DGX agent

do you know what you pay for in agentic workloads? cached tokens! session with 50+ tool calls -> prompt is billed 50 times all providers give 1/5 cached discount for GLM-5.2 we at @FireworksAI_HQ drop

model-releasesfireworks-ai--x
1 Jul 2026
Safety

ECHO: Prune to act, trace to learn with selective turn memory in agentic RL

DGX agent

arXiv:2606.31650v1 Announce Type: cross Abstract: Long-horizon language agents must repeatedly interact with tools, accumulate evidence, and make decisions under bounded context windows. Existing cont

safetyarxiv-cs-ai
1 Jul 2026
Model Releases

good post on the value of knowing whats going on in your harness 'For agents within an actual product, I have gravitated to using LangChain …

DGX agent

good post on the value of knowing whats going on in your harness 'For agents within an actual product, I have gravitated to using LangChain DeepAgents' The frontier model lab harnesses are amazing - d

model-releasesharrison-chase--x
1 Jul 2026
Safety

IterCAD: An Iterative Multimodal Agent for Visually-Grounded CAD Generation and Editing

DGX agent

arXiv:2606.13368v2 Announce Type: replace Abstract: Computer-Aided Design is pivotal in modern manufacturing, yet existing automated methods predominantly rely on open-loop, one-shot generation, creat

safetyarxiv-cs-ai
1 Jul 2026
Safety

LabGuard: Grounding Natural-Language Laboratory Rules into Runtime Guards for Embodied Laboratory Agents

DGX agent

arXiv:2606.31045v1 Announce Type: new Abstract: Scientific embodied agents are increasingly capable of carrying out laboratory procedures, but executing these procedures safely in dynamic laboratory e

safetyarxiv-cs-ai
1 Jul 2026
Model Releases

MECoBench: A Systematic Study of Multimodal Agent Collaboration in Embodied Environments

DGX agent

arXiv:2606.31966v1 Announce Type: cross Abstract: Recent multimodal large language models (MLLMs) have strong potential as embodied agents, but their ability to collaborate in visually grounded enviro

model-releasesarxiv-cs-ai
1 Jul 2026
Hardware

Our research team has 9 papers at ICML next week! Spanning the full stack from frontier agents to GPU kernels, we're excited to share what o…

DGX agent

Our research team has 9 papers at ICML next week! Spanning the full stack from frontier agents to GPU kernels, we're excited to share what our researchers and collaborators have been working on. If yo

hardwaretogether-ai--x
1 Jul 2026
Model Releases

PPT-Eval: A Benchmark for Computer-Use Agents on PowerPoint Tasks

DGX agent

arXiv:2606.31154v1 Announce Type: cross Abstract: Creating and editing slides is a rich, multimodal activity that is ubiquitous in professional and educational settings, making it an ideal testbed for

model-releasesarxiv-cs-ai
1 Jul 2026
Safety

QVal: Cheaply Evaluating Dense Supervision Signals for Long-Horizon LLM Agents

DGX agent

arXiv:2606.32034v1 Announce Type: cross Abstract: LLM agents increasingly act over long horizons, where a single trajectory can contain hundreds or thousands of actions. In these settings, outcome-onl

safetyarxiv-cs-ai
1 Jul 2026
Safety

Smart charging of large fleets of Electric Vehicles: Independent Multi-Agent Reinforcement Learning approaches

DGX agent

arXiv:2606.31347v1 Announce Type: new Abstract: The electrification of transportation through electric vehicles introduces new challenges for power grid management, such as increased peak demand, volt

safetyarxiv-cs-ai
1 Jul 2026
Safety

TreeAgent: A Generalizable Multi-Agent Framework for Automated Bias Labeling in Forestry via Compiled Expert Rules and Vision-Language Models

DGX agent

arXiv:2606.31976v1 Announce Type: new Abstract: Human-labeled data are widely used as reference annotations in ML, despite known variability across annotators in many expert-driven domains. In additio

safetyarxiv-cs-ai
1 Jul 2026
Safety

TRIAGE: Role-Typed Credit Assignment for Agentic Reinforcement Learning

DGX agent

arXiv:2606.32017v1 Announce Type: cross Abstract: Agentic reinforcement learning requires assigning credit to environment-facing actions such as searches, clicks, edits, navigation commands, and objec

safetyarxiv-cs-ai
1 Jul 2026
Safety

Budgeted Act-or-Defer Multi-Agent LLM Deliberation with Local Reliability Bounds

DGX agent

arXiv:2606.29654v1 Announce Type: new Abstract: Multi-agent deliberation among LLMs can improve reasoning, but deployment requires deciding when the current answer is reliable enough to act on and whe

safetyarxiv-cs-ai
30 Jun 2026
Model Releases

Building to the Test: Coding Agents Deliver What You Check, Not What You Requested

DGX agent

arXiv:2606.28430v1 Announce Type: cross Abstract: Benchmarks are widely used to evaluate task completion by Large Language Models (LLMs), but this approach has accumulated construction-validity proble

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Forensic Trajectory Signatures for Agent Memory Poisoning Detection

DGX agent

arXiv:2606.30566v1 Announce Type: cross Abstract: We discover a behavioral invariant in LLM agents under persistent memory poisoning: in architectures where routing information is retrieved through ob

model-releasesarxiv-cs-lg
30 Jun 2026
Model Releases

From Tool Connection to Execution Control: Benchmarking Security Invariants in MCP-Style Agent Runtimes

DGX agent

arXiv:2606.29073v1 Announce Type: cross Abstract: Model Context Protocol (MCP)-style ecosystems give language-model applications a practical connection layer for tools, resources, prompts, and transpo

model-releasesarxiv-cs-ai
30 Jun 2026
Safety

HiComm: Hierarchical Communication for Multi-agent Reinforcement Learning

DGX agent

arXiv:2606.29126v1 Announce Type: new Abstract: Cooperative multi-agent reinforcement learning (MARL) often relies on communication to mitigate partial observability, yet most existing protocols treat

safetyarxiv-cs-ai
30 Jun 2026
Industry

Hotels, tour operators, and travel agencies rush to launch proprietary online tools and loyalty schemes to fend off future competition from AI travel agents (Stephanie Stacey/Financial Times)

DGX agent

Stephanie Stacey / Financial Times: Hotels, tour operators, and travel agencies rush to launch proprietary online tools and loyalty schemes to fend off future competition from AI travel agents — Chatb

industrytechmeme
30 Jun 2026
Local Ai

Looking Is Not Picking: An Attention-Segment Account of Tool-Selection Failures in LLM Agents

DGX agent

arXiv:2606.16364v2 Announce Type: replace Abstract: LLM agents mis-call tools, and the natural guess is that the model failed to see the right tool in a crowded harness. We show the opposite through a

local-aiarxiv-cs-ai
30 Jun 2026
Model Releases

LUMEN: Cost-Transparent Multi-Agent Pipeline for Automated Systematic Review and Meta-Analysis

DGX agent

arXiv:2606.28362v1 Announce Type: cross Abstract: Systematic reviews and meta-analyses (SR/MA) remain the gold standard for evidence synthesis, yet completing one typically requires 67 weeks and subst

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

MemLeak: Diagnosing Information Leaks in Multimodal Agent Memory

DGX agent

arXiv:2606.29788v1 Announce Type: new Abstract: When a multimodal AI agent is asked to forget a fact, current memory systems usually delete the text entry and report success. We find that the fact can

model-releasesarxiv-cs-lg
30 Jun 2026
Model Releases

meta-pipe: An LLM-agent pipeline for end-to-end automated systematic review and meta-analysis

DGX agent

arXiv:2606.28363v1 Announce Type: cross Abstract: Objective: To describe the architecture and design rationale of meta-pipe, an open-source large language model (LLM)-agent pipeline that integrates th

model-releasesarxiv-cs-ai
30 Jun 2026
Safety

Metric Aggregation Divergence: A Hidden Validity Threat in Agent-Based Policy Optimization and a Contractual Remedy

DGX agent

arXiv:2606.29038v1 Announce Type: cross Abstract: Metric aggregation divergence (MAD) is the silent inconsistency that arises when distinct pipeline stages in an agent-based model coupled with a multi

safetyarxiv-cs-ai
30 Jun 2026
Model Releases

Multi-Agent Route Planning as a QUBO Problem

DGX agent

arXiv:2602.07913v2 Announce Type: replace Abstract: Multi-Agent Route Planning considers selecting vehicles, each associated with a single predefined route, such that route-level coverage utility is m

model-releasesarxiv-cs-ro
30 Jun 2026
Model Releases

OSWorld2.0: Benchmarking Computer Use Agents on Long-Horizon Real-World Tasks

DGX agent

arXiv:2606.29537v1 Announce Type: new Abstract: Existing computer-use benchmarks fail to capture the realism, complexity, and long-horizon demands of real-world computer use, limiting their ability to

model-releasesarxiv-cs-ai
30 Jun 2026
Tools

Our Founder and Chief Scientist @EdoLiberty just kicked off the Search & Retrieval track @aiDotEngineer. 'Agents don't need to be smarter. T…

DGX agent

Our Founder and Chief Scientist @EdoLiberty just kicked off the Search & Retrieval track @aiDotEngineer. 'Agents don't need to be smarter. They need a Knowledge Layer.' We've got a big update on this

toolspinecone--x
30 Jun 2026
Applications

ScarfBench: Benchmarking AI Agents for Enterprise Java Framework Migration

DGX agent

ScarfBench is a benchmarking framework designed to evaluate AI agents' capabilities in migrating enterprise Java applications to modern frameworks. The benchmark likely assesses how well AI systems ca

applicationshugging-face
30 Jun 2026
Model Releases

We’re introducing GeneBench-Pro, a research-level benchmark for a harder kind of AI progress: how well agents can navigate messy biological …

DGX agent

We’re introducing GeneBench-Pro, a research-level benchmark for a harder kind of AI progress: how well agents can navigate messy biological data, choose the right analysis path, and make judgment call

model-releasesopenai--x
30 Jun 2026
Model Releases

When Does Overlap Help? OSU-Mem and a Cell-Conditional Analysis of Trajectory Memory for LLM Agents

DGX agent

arXiv:2606.28376v1 Announce Type: cross Abstract: Long-horizon large language model (LLM) agents accumulate interaction trajectories that quickly exceed any practical prompt budget, and existing memor

model-releasesarxiv-cs-ai
30 Jun 2026
Safety

ATOD: Annealed Turn-aware On-policy Distillation for Multi-turn Autonomous Agents

DGX agent

arXiv:2606.27814v1 Announce Type: new Abstract: Training small language-model agents for long-horizon interactive tasks requires both fast imitation and reward-driven improvement. On-policy distillati

safetyarxiv-cs-ai
29 Jun 2026
← Previous
1…160161162163164…375
Next →