AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

agents

GridTimelineEvolution
7,195 results
3 Jul 2026

CLAP: Closed-Loop Training, Evaluation, and Release Control for Domain Agent Post-training

AgentsDGX agent

arXiv:2607.01846v1 Announce Type: new Abstract: Domain agents often face noisy business data, uncertain post-training gains, offline/application mismatch, and adapter-release risk. This paper presents

Coding-agents can replicate scientific machine learning papers

AgentsDGX agent

arXiv:2607.02134v1 Announce Type: new Abstract: Scientific machine learning papers typically make computational claims, e.g., that the relative mean square error is less than 5% or that the 95% predic

CommonRoad-Game: A Human-in-the-Loop Simulation Framework for Autonomous Driving

AgentsDGX agent

arXiv:2607.01382v1 Announce Type: new Abstract: Motion planning algorithms should be evaluated in human-in-the-loop environments to ensure they produce safe and efficient behaviors during interactions


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

ContextNest: Verifiable Context Governance for Autonomous AI Agent

AgentsDGX agent

arXiv:2607.02116v1 Announce Type: new Abstract: Autonomous AI agents increasingly depend on external knowledge stores, yet most retrieval pipelines provide relevance without durable guarantees of prov

Decentralized Stochastic Subgradient-type Methods with Communication Compression for Nonsmooth Nonconvex Optimization

AgentsDGX agent

arXiv:2607.01755v1 Announce Type: cross Abstract: In this paper, we consider the nonsmooth nonconvex decentralized optimization problem, where inter-agent communication is compressed. We propose a gen

FaithMed: Training LLMs For Faithful Evidence-Based Medical Reasoning

AgentsDGX agent

arXiv:2607.01440v1 Announce Type: new Abstract: Faithful reasoning is essential in medicine, where clinical decisions require transparent justification grounded in reliable evidence. Current medical L

Grounded autonomous research: a fault-tolerant LLM pipeline from corpus to manuscript in frontier computational physics

AgentsDGX agent

arXiv:2607.02329v1 Announce Type: new Abstract: Autonomous-research agents have demonstrated end-to-end LLM automation in machine-learning sandboxes where execution provides calibration. Frontier phys

Grounded Optimization: A Layered Engineering Framework for Reducing LLM Hallucination in Automated Personal Document Rewriting

AgentsDGX agent

arXiv:2607.01457v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly applied to resume optimization for applicant tracking systems, introducing hallucination failures distin

Had an amazing time talking to some of the most energetic builders in AI at @aiDotEngineer thanks to @swyx for organizing and thanks to @alt…

AgentsDGX agent

Had an amazing time talking to some of the most energetic builders in AI at @aiDotEngineer thanks to @swyx for organizing and thanks to @altryne and friends for hallway chats and podcasts to share abo

Hello brother

AgentsDGX agent

Hello brother We've been thinking about the best way to use screen data for a while now. After 500k hours of beta use, we're launching Dayflow: the open source automatic work journal. One of my favori

hosted three side events, a lunch, and did a talk this week. pretty tired, but worth it!

AgentsDGX agent

hosted three side events, a lunch, and did a talk this week. pretty tired, but worth it! Heaps of fun watching last night’s USA v BIH World Cup match at @yoheinakajima’s watch party after @aiDotEngine

In this interview at @aiDotEngineer World's Fair, @vercel chief of software @andrewqu explains why agents represent a new form of software, …

AgentsDGX agent

In this interview at @aiDotEngineer World's Fair, @vercel chief of software @andrewqu explains why agents represent a new form of software, what Vercel learned from building its own, and why Vercel it

Is 'loop engineering' peak linkedin tech influencer slop?

AgentsDGX agent

# Summary This post likely critiques the trend of 'loop engineering' as an example of superficial, trend-chasing content popular among tech influencers on LinkedIn, questioning whether it represents g

Janus: a Playground for User-Involved Agentic Permission Management

AgentsDGX agent

arXiv:2607.01510v1 Announce Type: new Abstract: AI agents that autonomously execute tool calls on a user's behalf raise pressing questions about permission management: what role could users play, and

Language Models as Measurement Apparatus for Culture

AgentsDGX agent

arXiv:2607.02459v1 Announce Type: new Abstract: Language models are increasingly used to quantify cultural phenomena, but what makes such measurement distinctively cultural? This paper argues that NLP

Leveraging Metamemory Agent for Enhanced Data-Free Code Generation in Large Language Models

AgentsDGX agent

arXiv:2501.07892v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have shown strong performance in automated code generation, with few-shot prompting widely used for its simplicit

LIME: Learning Intent-aware Camera Motion from Egocentric Video

AgentsDGX agent

arXiv:2607.02417v1 Announce Type: cross Abstract: Autonomous robots often need to move their camera before they can act: to inspect an object, reveal an occluded region, or obtain a view that responds

Lynx: Progressive Speculative Quantization for accelerating KV Transfer in Long-Context Inference

AgentsDGX agent

arXiv:2607.01831v1 Announce Type: cross Abstract: Long-context inference is increasingly common in large language model (LLM) serving, driven by retrieval-augmented generation and agentic systems. In

Mark Zuckerberg says Meta’s agentic AI efforts aren’t progressing as fast as he had hoped

AgentsDGX agent

Meta Platforms Inc. Chief Executive Mark Zuckerberg told employees at an internal town hall meeting that the company’s work on artificial intelligence agents hasn’t progressed as quickly as he had hop

MMAO-Cls: Metabolic Multi-Agent Optimization for Joint Feature Selection and Classifier Tuning

AgentsDGX agent

arXiv:2607.01539v1 Announce Type: cross Abstract: This paper studies whether the Metabolic Multi-Agent Optimizer (MMAO) can act as a credible outer-loop optimizer for classification model selection. W

Multimodal prompting is clearly the future. I love experimenting with new ways to interact with agents. As a researcher and engineer, I've f…

AgentsDGX agent

Multimodal prompting is clearly the future. I love experimenting with new ways to interact with agents. As a researcher and engineer, I've found that the richer the inputs to the agent and the richer

'My belief is that, three years down the road, people will not own PCs. You will stop buying PCs and laptops. You will just have a phone, or…

AgentsDGX agent

'My belief is that, three years down the road, people will not own PCs. You will stop buying PCs and laptops. You will just have a phone, or maybe you might have a tablet.' -- AGI Inc.'s Div Garg on w

OpenWiki is at 1.7k stars in just 3 days! Right now it's just for codebases, but we're working to expand it to everything for memory. What d…

AgentsDGX agent

OpenWiki is at 1.7k stars in just 3 days! Right now it's just for codebases, but we're working to expand it to everything for memory. What do you want to see in a general purpose memory wiki agent? ht

OpenWiki up to 1.7k GitHub stars in two days Most common ask it to make it more general purpose (not just coding) What sources do people wan…

AgentsDGX agent

OpenWiki up to 1.7k GitHub stars in two days Most common ask it to make it more general purpose (not just coding) What sources do people want to see? Notion? GMail? Slack? GDrive? Internet search? Oth

OpenWiki 📈📈 What do you want to see from a general purpose wiki?

AgentsDGX agent

OpenWiki 📈📈 What do you want to see from a general purpose wiki? OpenWiki up to 1.7k GitHub stars in two days Most common ask it to make it more general purpose (not just coding) What sources do peopl

Ophiuchus: Incentivizing Tool-augmented 'Think with Images' for Joint Medical Segmentation, Understanding and Reasoning

AgentsDGX agent

arXiv:2512.14157v2 Announce Type: replace Abstract: Recent medical MLLMs have made significant progress in generating step-by-step textual reasoning chains. However, they still struggle with complex c

Prompt Coverage Adequacy

AgentsDGX agent

arXiv:2607.02057v1 Announce Type: cross Abstract: In recent years, it has become increasingly evident that large language models (LLMs) and autonomous agents raise the level of abstraction in software

Prompt engineering is costing you money. Learn how to fine-tune your models. Stuffing a lot of text into every single API call slows down yo…

AgentsDGX agent

Prompt engineering is costing you money. Learn how to fine-tune your models. Stuffing a lot of text into every single API call slows down your app because the model has to process all those tokens bef

Real-Time Visual Intelligence on Low-Cost UAVs: A Modular Approach for Tracking, Scanning, and Navigation

AgentsDGX agent

arXiv:2607.02298v1 Announce Type: new Abstract: Autonomous drones are rapidly transforming modern warfare and civil applications alike. This paper presents the development of an integrated intelligent

Recorded an impromptu podcast episode with @swyx for @latentspacepod last month at @aiDotEngineer SG. Covered good ground including: - Why '…

AgentsDGX agent

Recorded an impromptu podcast episode with @swyx for @latentspacepod last month at @aiDotEngineer SG. Covered good ground including: - Why 'second brain' is the killer agent use case - Messaging platf

Recursive Models for Long-Horizon Reasoning

AgentsDGX agent

arXiv:2603.02112v2 Announce Type: replace-cross Abstract: Modern language models reason within bounded context, an inherent constraint that poses a fundamental barrier to long-horizon reasoning. We id

Simulation Based Reward Function Validation for Multi-Agent On Orbit Inspection

AgentsDGX agent

arXiv:2607.01367v1 Announce Type: cross Abstract: A proposed method for the control of groups of inspection spacecraft is Multi-Agent Reinforcement Learning (MARL). While MARL has already been employe

SkillCoach: Self-Evolving Rubrics for Evaluating and Enhancing Agentic Skill-Use

AgentsDGX agent

arXiv:2607.01874v1 Announce Type: new Abstract: Skills are becoming a reusable operational layer for LLM agents, encoding SOPs, domain rules, tool workflows, scripts, and validation routines. In reali

SkillFuzz: Fuzzing Skill Composition for Implicit Intents Discovery in Open Skill Marketplaces

AgentsDGX agent

arXiv:2607.02345v1 Announce Type: cross Abstract: Large Language Model (LLM)-based agents increasingly automate software engineering tasks through reusable skills, natural-language instruction documen

So i built LangGraph Sync. it’s an interactive web tool that visualizes langgraph workflows & lets you edit & bidirectionally synchronize th…

AgentsDGX agent

So i built LangGraph Sync. it’s an interactive web tool that visualizes langgraph workflows & lets you edit & bidirectionally synchronize the graph & code in real time. you edit the visual graph, your

The Agentic Garden of Forking Paths

AgentsDGX agent

arXiv:2607.01507v1 Announce Type: new Abstract: Empirical research rarely admits a unique analysis. Different analytical choices can lead to different conclusions from the same data, yet these hidden

The Rollout Infrastructure Tax in Coding-Agent Reinforcement Learning

AgentsDGX agent

arXiv:2607.01415v1 Announce Type: new Abstract: Coding-agent reinforcement learning treats execution infrastructure as a background implementation detail, despite relying on large numbers of interacti

Traceable Fault Diagnosis for Battery Energy Storage Systems via Retrieval-Augmented Multi-Agent O&M Assistant

AgentsDGX agent

arXiv:2607.01992v1 Announce Type: new Abstract: Large-scale battery energy storage systems (BESSs) require O&M decisions that combine alarms, cell-level measurements, device topology, diagnostic table

Vercel's Andrew Qu on why agents are a new kind of software

AgentsDGX agent

Andrew Qu from Vercel discusses how AI agents represent a fundamentally different category of software compared to traditional applications, exploring their unique characteristics and implications for

When brainstorming new AI agent uses, I always remind myself what used to be expensive and is now nearly zero… 1) cost of reading everything…

AgentsDGX agent

When brainstorming new AI agent uses, I always remind myself what used to be expensive and is now nearly zero… 1) cost of reading everything fell → you can now watch 100% instead of a sample (every ex

WorkPods agent memory as wikis. @langchain #openwiki @cognition #deepwiki @karpathy #llmwiki @Factory #autowiki @hwchase17 @BraceSproul

AgentsDGX agent

This post discusses using WorkPods agent memory systems organized as wikis, likely exploring how LLM agents can maintain and access persistent knowledge repositories in wiki format for improved contex

You are going to end up running hallucinated code

AgentsDGX agent

You are going to end up running hallucinated code ~60% Fable cost cut by transparently turning the code into an image and having the model OCR it. WILD idea. also hilarious. https://github.com/teamcho

2 Jul 2026

1/ OpenWiki First, we wrote about why 'wikis' are interesting as a form of memory: https://www.langchain.com/blog/wiki-memory Cites examples…

AgentsDGX agent

1/ OpenWiki First, we wrote about why 'wikis' are interesting as a form of memory: https://www.langchain.com/blog/wiki-memory Cites examples from @FactoryAI, @cognition, @karpathy Then @BraceSproul op

15 examples of real-world challenges: Insights from the AWS Summit Washington, D.C. event

AgentsDGX agent

As organizations race to operationalize generative and agentic artificial intelligence, the conversation is shifting from pilots and proofs of concept to real-world AI deployments that deliver measura

3/ Harbor integration Harbor is a framework for running long running, stateful agent evals We integrate with it in multiple different ways. …

AgentsDGX agent

3/ Harbor integration Harbor is a framework for running long running, stateful agent evals We integrate with it in multiple different ways. We wrote a blog on those integrations: https://www.langchain

4/ programatic subagents in deepagents Deep Agents is our open source, model agnostic agent harness: https://github.com/langchain-ai/deepage…

AgentsDGX agent

4/ programatic subagents in deepagents Deep Agents is our open source, model agnostic agent harness: https://github.com/langchain-ai/deepagents We added the ability to programatically call subagents.

A Contextual-Bandit Oversight Game with Two-Sided Informational Asymmetry

AgentsDGX agent

arXiv:2607.00155v1 Announce Type: new Abstract: We study runtime human oversight of an AI agent when private information runs in both directions: the human privately knows her reward function, while t

🎧A look at how we built LangSmith Engine with @hwchase17 + @bentannyhill

AgentsDGX agent

🎧A look at how we built LangSmith Engine with @hwchase17 + @bentannyhill Special episode with @bentannyhill on the Max Agency podcast. A month ago, his team shipped LangSmith Engine, our agent that hu

A Task-State Representation for Long-Horizon Mobile GUI Agents

AgentsDGX agent

arXiv:2607.00502v1 Announce Type: new Abstract: While long-horizon mobile GUI agents typically rely on thought-action-observation loops, they struggle to separate persistent task states from transient

Active Sensing for RIS-Aided Tracking and Power Control: A Hybrid Neuroevolution and Supervised Learning Approach

AgentsDGX agent

arXiv:2607.00056v1 Announce Type: cross Abstract: This paper studies energy efficient tracking of power-limited mobile users with the assistance of a Reconfigurable Intelligent Surface (RIS). Since lo

Agentic generation of verifiable rules for deterministic, self-expanding reaction classification

AgentsDGX agent

arXiv:2607.01061v1 Announce Type: new Abstract: Computer-assisted synthesis planning breaks target molecules into accessible precursors using large libraries of reaction rules that assign each transfo

Agri-SAGE: Simulation-Grounded Multi-Agent LLM for Context-Aware Agricultural Advisory Generation

AgentsDGX agent

arXiv:2607.00454v1 Announce Type: new Abstract: Agricultural advisory systems face a fundamental tension: static agronomic guidelines offer consistent, evidence-based recommendations, yet remain blind

AI, Trust, and Teaming: The Humans-as-Handlers Approach for Autonomous and Opaque AI Systems

AgentsDGX agent

arXiv:2607.00523v1 Announce Type: cross Abstract: Artificial intelligence (AI) is becoming ubiquitous, and across domains, increasingly autonomous systems are carrying out tasks which raise significan

as expected by whom? 🤔

AgentsDGX agent

as expected by whom? 🤔 META: ZUCKERBERG SAYS AI PROGRESS HAS BEEN SLOWER THAN EXPECTED • ZUCKERBERG SAID AI AGENT DEVELOPMENT HAS NOT ACCELERATED AS EXPECTED OVER THE PAST FOUR MONTHS • SAID META'S 20

At a town hall, Mark Zuckerberg said Meta's AI agent development has not accelerated as expected and its reorganization was not as 'clean' as it could have been (Katie Paul/Reuters)

AgentsDGX agent

Katie Paul / Reuters: At a town hall, Mark Zuckerberg said Meta's AI agent development has not accelerated as expected and its reorganization was not as “clean” as it could have been — Meta (META.O) C

BaRA: BFS-and-Reflection Web Data Collection Agent

AgentsDGX agent

arXiv:2607.00007v1 Announce Type: cross Abstract: Large language model (LLM)-based web agents reduce manual scripting for web data collection, yet on live websites, they often miss relevant pages, ret

Behavior-Adaptive Conversational Agents: Toward a Fluid Personality Framework

AgentsDGX agent

arXiv:2607.01034v1 Announce Type: cross Abstract: Large language model (LLM)-based conversational agents (CAs) are now ubiquitous, creating new opportunities for AI-mediated behavior change. Their cap

Best practices for multi-turn reinforcement learning in Amazon SageMaker AI

AgentsDGX agent

In this post, we share best practices for reliable multi-turn RL training. We cover how to build a training environment you can trust, set up an external evaluation, design a reward aligned with the e

Beyond Line of Sight: Hybrid Validation of V2X Collective Perception in Complex Scenarios

AgentsDGX agent

arXiv:2607.00874v1 Announce Type: new Abstract: This paper introduces a probabilistic framework and hybrid validation methodology for V2X-enabled Collective Perception (CP) in complex traffic scenario

Building on Devin's new capabilities, today we're announcing the Security Vulnerability Remediation Program: a tailored engagement to take y…

AgentsDGX agent

Building on Devin's new capabilities, today we're announcing the Security Vulnerability Remediation Program: a tailored engagement to take your security backlog towards zero in six weeks. Remediation

← Previous
1…2526272829…120
Next →