AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,980 results
Agents

Fun conversation with @swyx on our journey building the cloud for true elastic inference, sandboxes, and more. And of course, how we're evol…

DGX agent

Fun conversation with @swyx on our journey building the cloud for true elastic inference, sandboxes, and more. And of course, how we're evolving Modal's dev experience to be better for agents. Modal's

agentsswyx--x
9 Jul 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Agents

Join us for a LangChain + Clay meetup with @palashshah, @jeffbarg, Vyshu Khota, and Soroush Khadem. https://luma.com/jqif2hti Palash will br…

DGX agent

Join us for a LangChain + Clay meetup with @palashshah, @jeffbarg, Vyshu Khota, and Soroush Khadem. https://luma.com/jqif2hti Palash will break down how he built a self-improving agent at LangChain, L

agentsharrison-chase--x
9 Jul 2026
Agents

RLVP: Penalize the Path, Reward the Outcome

DGX agent

arXiv:2607.07435v1 Announce Type: cross Abstract: Agents acting on our behalf in the real world (e.g. placing phone calls) must learn online from costly, often irreversible interactions rather than ch

agentsarxiv-cs-ai
9 Jul 2026
Model Releases

AutoCedar: An Agentic Framework for Verifier-Guided Access Control Policy Synthesis

DGX agent

arXiv:2607.03656v1 Announce Type: cross Abstract: Large Language Models are increasingly used to turn natural-language requirements into code. In access control, that shortcut is dangerous: a generate

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

GLM-5 Serving Parameter Tuning for OpenClaw: Single-Deployment MaaS Inference Optimization for Long-Context Agent Workloads

DGX agent

arXiv:2607.02518v1 Announce Type: cross Abstract: OpenClaw requests are dominated by long, tool-augmented prefixes, including system prompts, conversation history, and tool outputs fed back into the c

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

VCB Bench: An Evaluation Benchmark for Audio-Grounded Large Language Model Conversational Agents

DGX agent

arXiv:2510.11098v5 Announce Type: replace-cross Abstract: Recent advances in large audio language models (LALMs) have greatly enhanced multimodal conversational systems. However, existing benchmarks r

model-releasesarxiv-cs-cl
7 Jul 2026
Agents

Come join us on Thursday! It'll be a great session you'll want to add to your MEMORY.md 😄

DGX agent

Come join us on Thursday! It'll be a great session you'll want to add to your MEMORY.md 😄 does your agent need a wiki? should humans and agent use the same wiki? should you have one wiki or many wikis

agentsharrison-chase--x
6 Jul 2026
Agents

devin has a hot take that wikis are NOT the right abstraction should be a fun webinar!!

DGX agent

devin has a hot take that wikis are NOT the right abstraction should be a fun webinar!! does your agent need a wiki? should humans and agent use the same wiki? should you have one wiki or many wikis?

agentsharrison-chase--x
6 Jul 2026
Model Releases

sqlite-utils 4.0rc2, mostly written by Claude Fable (for about $149.25)

DGX agent

I wrote about the sqlite-utils 4.0rc1 release a couple of weeks ago. Since we only have Claude Fable on our Max subscriptions for a few more days, I decided to see if it could help me get to a 4.0 sta

model-releasessimon-willison
5 Jul 2026
Safety

Beyond Next-Token Prediction: An RLVR Proof of Concept for Tool-Use Agents on Atlassian Workflows

DGX agent

arXiv:2607.01465v1 Announce Type: new Abstract: Large language models are trained to predict the next token, not to act inside a specific API. In niche enterprise SaaS workflows -- where success means

safetyarxiv-cs-ai
3 Jul 2026
Model Releases

COMFYCLAW: Self-Evolving Skill Harnesses for Image Generation Workflows

DGX agent

arXiv:2607.01709v1 Announce Type: new Abstract: Agents are increasingly used to construct workflows and assist humans in completing recurring tasks more efficiently. As these workflows become repeated

model-releasesarxiv-cs-ai
3 Jul 2026
Safety

DiPS: Dialogue Policy Selection for High-Stakes Persuasion Agents

DGX agent

arXiv:2607.01557v1 Announce Type: cross Abstract: Large Language Models (LLMs) often struggle with persuasion in high-stakes scenarios. People's individual personalities and concerns require tailored

safetyarxiv-cs-ai
3 Jul 2026
Safety

Episodic-to-Semantic Consolidation Without Identity Drift

DGX agent

arXiv:2607.01988v1 Announce Type: new Abstract: Long-running adaptive intelligent agents face a structural tension between knowledge consolidation and information integrity. Memory consolidation is co

safetyarxiv-cs-ai
3 Jul 2026
Agents

SkillFuzz: Fuzzing Skill Composition for Implicit Intents Discovery in Open Skill Marketplaces

DGX agent

arXiv:2607.02345v1 Announce Type: cross Abstract: Large Language Model (LLM)-based agents increasingly automate software engineering tasks through reusable skills, natural-language instruction documen

agentsarxiv-cs-ai
3 Jul 2026
Tutorials

World Feedback for Clinical Agents: Diagnosing RL in FHIR Environments

DGX agent

arXiv:2607.01470v1 Announce Type: new Abstract: Clinical protocol-execution tasks -- checking a lab value, applying a threshold, placing a correctly structured FHIR order -- are natural candidates for

tutorialsarxiv-cs-ai
3 Jul 2026
Model Releases

Mapping the Evaluation Frontier: An Empirical Survey of the Bias-Reliability Tradeoff Across Eleven Evaluator-Agent Conditions

DGX agent

arXiv:2607.00304v1 Announce Type: cross Abstract: The bias-reliability tradeoff conjectures that LLM evaluation systems are constrained in (gamma, H, CV) space, where evaluator coupling (gamma), strat

model-releasesarxiv-cs-ai
2 Jul 2026
Safety

OpenReward: Learning to Reward Long-form Agentic Tasks via Reinforcement Learning

DGX agent

arXiv:2510.24636v3 Announce Type: replace Abstract: Reward models (RMs) have become essential for aligning large language models (LLMs), serving as scalable proxies for human evaluation in both traini

safetyarxiv-cs-cl
2 Jul 2026
Agents

How Inscribe uses Amazon Bedrock to stop document fraud in seconds

DGX agent

In this post, you will learn how Inscribe developed an agentic AI system using Amazon Bedrock that reasons across documents the way an expert fraud analyst would. With this new agentic AI system, Insc

agentsaws-ml-blog
1 Jul 2026
Model Releases

MIRTH: Mutual-Information Reasoning with Temporal Hubs for Vision-Language-Action Agents

DGX agent

arXiv:2606.31167v1 Announce Type: cross Abstract: VLA models have emerged as a powerful paradigm for transferring semantic knowledge from web-scale data to physical robotic control. However, current s

model-releasesarxiv-cs-ai
1 Jul 2026
Safety

Training Therapeutic Judges and Multi-Agent Systems for Human-Aligned Mental Health Support

DGX agent

arXiv:2606.30887v1 Announce Type: cross Abstract: Large language models show promise for mental health support, yet therapeutic quality improves only when evaluation functions as an actionable control

safetyarxiv-cs-ai
1 Jul 2026
Model Releases

A Diagnostic Framework and Multi-Evaluator Audit of Evaluator-Driven Preference Dynamics in Self-Adapting LLM Agents

DGX agent

arXiv:2606.29719v1 Announce Type: cross Abstract: Measurements of proprietary LLM evaluators can become invalid within weeks -- we document one case and provide the diagnostic framework to detect it.

model-releasesarxiv-cs-cl
30 Jun 2026
Safety

Agentic Safety is an Epistemic Property, Not a Behavioral One

DGX agent

arXiv:2606.28347v1 Announce Type: cross Abstract: Contemporary AI safety spans pre-training interventions, post-training alignment, deployment-time controls, monitoring, and red-teaming. These methods

safetyarxiv-cs-ai
30 Jun 2026
Agents

Improved Multi-Dimensional Forecasting for Swap Regret

DGX agent

arXiv:2606.29533v1 Announce Type: cross Abstract: We study the problem of forecasting for an arbitrary number of downstream agents with unknown objectives, each of whom best responds to the forecaster

agentsarxiv-cs-lg
30 Jun 2026
Agents

Manufactured Confidence: How Memory Consolidation Turns Hearsay into Confident Facts

DGX agent

arXiv:2606.29279v1 Announce Type: cross Abstract: LLM agents carry conclusions across steps and sessions in compressed memory, and memory products (e.g., mem0, LangMem) rewrite conversation into store

agentsarxiv-cs-ai
30 Jun 2026
Model Releases

NVIDIA BioNeMo Agent Toolkit Brings Accelerated AI to Life Sciences Researchers in Claude Science

DGX agent

Life sciences has entered an era of computational scale, and for more than a decade, NVIDIA has built the full GPU-accelerated computing stack — spanning hardware, frameworks, libraries, models, micro

model-releasesnvidia-genai-blog
30 Jun 2026
Agents

On the Necessity of a Liquid Substrate for Mesh Intelligence

DGX agent

arXiv:2606.28413v1 Announce Type: cross Abstract: A mesh of sovereign agents has no center: no shared clock, no shared model, and no coordinator to gather data or retrain. Its competence rests on each

agentsarxiv-cs-ai
30 Jun 2026
Safety

The Two Genie Game: Adoption and Welfare in Audit-Grounded AI Governance

DGX agent

arXiv:2606.28710v1 Announce Type: new Abstract: We ask under what conditions an agent with a harm-minimizing policy can displace an approval-seeking (RLHF) agent in a competitive market, and when that

safetyarxiv-cs-ai
30 Jun 2026
Tutorials

AI agents are not your “coworkers”

DGX agent

This story originally appeared in The Algorithm, our weekly newsletter on AI. To get stories like this in your inbox first, sign up here. Imagine coming in to work to learn that a new underling will r

tutorialsmit-tech-review
29 Jun 2026
Safety

Chai: Agentic Discovery of Cryptographic Misuse Vulnerabilities

DGX agent

arXiv:2606.26933v1 Announce Type: cross Abstract: AI-assisted vulnerability discovery has proven effective for bug classes like memory safety, where instrumentation confirms memory violations and effi

safetyarxiv-cs-ai
26 Jun 2026
Agents

Open + closed models = better together. Our previous research with @harvey showed the benefits of combining a frontier closed model as an ad…

DGX agent

Open + closed models = better together. Our previous research with @harvey showed the benefits of combining a frontier closed model as an advisor agent with fine-tuned, open-source worker agents. Thre

agentsfireworks-ai--x
25 Jun 2026
Agents

SoK: AI Secure Code Generation: Progress, Pitfalls, and Paths Forward

DGX agent

arXiv:2606.25195v1 Announce Type: cross Abstract: The increasing use of AI systems for code generation raises a central security question: what can today's models and coding agents actually do to prod

agentsarxiv-cs-ai
25 Jun 2026
Model Releases

Introducing Claude for Music. You can now create songs from Claude Code, Hermes, Codex, or any agent you’re using. SOTA music model @MiniMax…

DGX agent

I cannot provide an accurate summary for this entry. The URL and source attribution appear inconsistent (title credits Yohei Nakajima but URL references a different user), and the post references prod

model-releasesyohei-nakajima--x
24 Jun 2026
Model Releases

When Retrieval Metrics Mislead: Measuring Policy Signal in Long-Horizon Tool-Use Agents

DGX agent

arXiv:2606.23937v1 Announce Type: cross Abstract: Exact-match retrieval recall is often used as a proxy for whether a retriever supplies useful policy context to a downstream decision model. We test t

model-releasesarxiv-cs-ai
24 Jun 2026
Hardware

NVIDIA Brings Trusted, 24/7 AI Agents to Telecom Operations

DGX agent

Telecom operators have seen remarkable returns from using generative AI to automate network management, customer care and back-office operations. Most of that impact has been task‑based: automation th

hardwarenvidia-blog
23 Jun 2026
Agents

Position: Correct Answer, Wrong Mechanism -- When AI Scientists Defend General Claims Their Own Data Contradicts

DGX agent

arXiv:2606.23175v1 Announce Type: new Abstract: AI scientist systems are described as tools, coauthors, or founders, but we evaluate them as if only the final answer matters. This position paper argue

agentsarxiv-cs-lg
23 Jun 2026
Safety

Fourier Features Let Agents Learn High Precision Policies with Imitation Learning

DGX agent

arXiv:2606.12334v1 Announce Type: new Abstract: High-precision robotic manipulation requires fine-grained spatial reasoning that is often difficult to achieve with RGB-only policies due to depth ambig

safetyarxiv-cs-lg
11 Jun 2026
Model Releases

BadRobot: Jailbreaking Embodied LLM Agents in the Physical World

DGX agent

arXiv:2407.20242v5 Announce Type: replace-cross Abstract: Embodied AI represents systems where AI is integrated into physical entities. Large Language Model (LLM), which exhibits powerful language und

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Constructing coherent spatial memory in LLM agents through graph rectification

DGX agent

arXiv:2510.04195v2 Announce Type: replace Abstract: Given a map description through global traversal navigation instructions, an LLM can often infer the implicit spatial layout and answer user queries

model-releasesarxiv-cs-ai
10 Jun 2026
Agents

it’s actually so cool to work at LangChain the…database company (??) yup, the cracked team that built SmithDB is doing a cool blog series on…

DGX agent

it’s actually so cool to work at LangChain the…database company (??) yup, the cracked team that built SmithDB is doing a cool blog series on “How to build the internals of a database” —> for agent sca

agentsharrison-chase--x
10 Jun 2026
Agents

Regimes: An Auditable, Held-Out-Gated Improvement Loop Demonstrated on LongMemEval with ActiveGraph

DGX agent

arXiv:2606.10241v1 Announce Type: new Abstract: Autonomous improvement loops are hard to trust because the improvement process is usually external scaffolding bolted onto the agent: failures go unlogg

agentsarxiv-cs-ai
10 Jun 2026
Applications

Trace2Policy: From Expert Behavior Traces to Self-Evolving Decision Agents

DGX agent

arXiv:2606.10457v1 Announce Type: new Abstract: Decision rules that enterprise experts apply tacitly -- in auditing, compliance, and contract review -- can be systematically recovered and improved thr

applicationsarxiv-cs-ai
10 Jun 2026
Agents

On the Hardness of Optimal Motion on Trees

DGX agent

arXiv:2606.06686v1 Announce Type: new Abstract: This paper presents a simple framework that settles the complexity of Multi-Agent Path Finding (MAPF) on trees across standard objectives--distance, mak

agentsarxiv-cs-ro
8 Jun 2026
Agents

New MIT study. Code volume surges by 300%, but output increases by only 30%: The AI dividend meets an awkward reality Autonomous AI coding a…

DGX agent

New MIT study. Code volume surges by 300%, but output increases by only 30%: The AI dividend meets an awkward reality Autonomous AI coding agents raised commits by 180%, but releases rose only 30%. Th

agentsgary-marcus--x
7 Jun 2026
Research

PerceptUI: LLM Agents as Human-Aligned Synthetic Users for UI/UX Evaluation

DGX agent

arXiv:2606.05697v1 Announce Type: new Abstract: User interface (UI) and user experience (UX) evaluation is central to product development, yet reliable feedback still relies on recruiting human partic

researcharxiv-cs-ai
6 Jun 2026
Model Releases

Running Python code in a sandbox with MicroPython and WASM

DGX agent

I've been experimenting with different approaches to running code in a sandbox for several years now, but my latest attempt feels like it might finally have all of the characteristics I've been lookin

model-releasessimon-willison
6 Jun 2026
Agents

this looks cool

DGX agent

this looks cool An experimental programming language from Vercel Labs that is truly made for AI agents! Not just a new syntax. Zero is a graph-first language where agents can Read & edit program struc

agentsyohei-nakajima--x
6 Jun 2026
Agents

MLEvolve: A Self-Evolving Framework for Automated Machine Learning Algorithm Discovery

DGX agent

arXiv:2606.06473v1 Announce Type: cross Abstract: Large language model (LLM) agents are increasingly applied to long-horizon tasks such as scientific discovery and machine learning engineering (MLE),

agentsarxiv-cs-cl
5 Jun 2026
Local Ai

AgenticDiffusion: Agentic Diffusion-based Path Planning for Vision-Based UAV Navigation

DGX agent

arXiv:2606.04111v1 Announce Type: cross Abstract: Indoor UAV navigation requires efficient exploration, scene understanding, and reliable trajectory execution under limited field-of-view observations.

local-aiarxiv-cs-ai
4 Jun 2026
← Previous
1…189190191192193…375
Next →