AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

agents

GridTimelineEvolution
7,154 results
5 Aug 2026

How LendingTree built a multi-agent mortgage assistant on Amazon Bedrock

AgentsDGX agent

Learn how LendingTree built a production multi-agent mortgage assistant on Amazon Bedrock. Three coordinated agents use LangGraph, the Model Context Protocol, and Amazon Nova models with built-in guar

How Mobileye transformed support operations using Amazon Bedrock AgentCore

AgentsDGX agent

In this post, we'll explore how Mobileye deployed an AI support agentic solution on Amazon Bedrock AgentCore - from the support bottleneck that sparked the idea, through the proof of concept that vali

Hybrid LLM-Augmented Reinforcement Learning Agents for Complex Sequential Decision Tasks

AgentsDGX agent

arXiv:2608.03502v1 Announce Type: new Abstract: Large Language Models (LLMs) have recently shown strong capabilities in reasoning, planning, and tool-use, enabling new forms of autonomous agents. Howe


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

HyperAgent: Planning and Acting over Tool-Schema Hypergraphs for Tool-Use LLM Agents

AgentsDGX agent

arXiv:2608.02650v1 Announce Type: new Abstract: Large language model (LLM) agents increasingly rely on external tools to complete complex real-world tasks. However, reliable tool-use planning remains

Improving Sample Efficiency in Multi-Agent Reinforcement Learning for Simulated Football Games via Exploration

AgentsDGX agent

arXiv:2503.13077v2 Announce Type: replace Abstract: Multi-agent reinforcement learning has shown promise in learning cooperative behaviors in team-based environments. However, such methods often deman

Indeed.

AgentsDGX agent

Gary Marcus tweeted “Indeed.” and then Frank Rundatz replied that the deterministic harness required to ground a large‑language‑model agent differs by task. Because of this variability, Rundatz agrees

Internalising the Identity Primitive: Cryptographic Individuality for an Autonomous Agent on a Public Blockchain

AgentsDGX agent

arXiv:2608.02986v1 Announce Type: cross Abstract: A software agent on a public blockchain accumulates authority and economic stakes, raising the engineering question of what makes it count as an indiv

IR2Solve: Structured Intermediate Representations for Cost-Efficient Optimization Autoformulation

AgentsDGX agent

arXiv:2608.02641v1 Announce Type: cross Abstract: Large language models (LLMs) can translate natural-language optimization problems into solver-ready formulations, but direct code generation is brittl

Is Inter-Seed Cross-Play Enough? Evaluating the Robustness of Zero-Shot Coordination Algorithms to Implementation Details

AgentsDGX agent

arXiv:2608.03644v1 Announce Type: new Abstract: AI agents deployed in real-world settings must be capable of coordinating with humans and other AI agents they have not encountered before. Zero-shot co

LACE: Large Language Model Aided Multi-Agent Framework for Agile RISC-V Instruction Extension

AgentsDGX agent

arXiv:2608.02915v1 Announce Type: cross Abstract: Domain-specific Instruction Set Architecture eXtensions (ISAX) are widely adopted in the RISC-V ecosystem to accelerate emerging workloads, but implem

LiveEvalBench: Toward Open-World Evaluation for Web Generation

AgentsDGX agent

arXiv:2608.03689v1 Announce Type: new Abstract: Large language models are increasingly capable of synthesizing executable frontend projects, yet existing benchmarks still treat web generation as a sta

Love the latest string of releases and the moves that the Langchain team has been making. Langchain's React implementation was the the first…

AgentsDGX agent

Love the latest string of releases and the moves that the Langchain team has been making. Langchain's React implementation was the the first ever agent i built though their framework couldn't compete

MAFIA: Query-Only Memory Attacks via Probing and Factual Injection against Audited LLM Agents

AgentsDGX agent

arXiv:2608.03844v1 Announce Type: new Abstract: Memory-augmented LLM agents rely on rich context for long-horizon reasoning and acting, yet their memory modules expose a persistent attack surface for

Meta is offering a cheaper Muse Spark 1.2 'contributor' tier priced at 0.10/1M input and 0.20/1M output tokens in exchange for using user prompts for training (Wall Street Journal)

AgentsDGX agent

Wall Street Journal: Meta is offering a cheaper Muse Spark 1.2 “contributor” tier priced at 0.10/1M input and 0.20/1M output tokens in exchange for using user prompts for training — The company, press

Meta releases Muse Code in beta, a terminal coding agent powered by Muse Spark 1.2, a coding-focused model priced at 1.25/1M input and 4.25/1M output tokens (Jonathan Vanian/CNBC)

AgentsDGX agent

Jonathan Vanian / CNBC: Meta releases Muse Code in beta, a terminal coding agent powered by Muse Spark 1.2, a coding-focused model priced at 1.25/1M input and 4.25/1M output tokens — Meta is rolling o

Mixed-Initiative Human-Robot Teaming under Suboptimality with Online Bayesian Adaptation

AgentsDGX agent

arXiv:2403.16178v2 Announce Type: replace-cross Abstract: For effective human-agent teaming, robots and other artificial intelligence (AI) agents must infer their human partner's abilities and behavio

Most AI tools only do one part of building a product. One researches. One designs. One writes code. One generates images. One deploys. You s…

AgentsDGX agent

Most AI tools only do one part of building a product. One researches. One designs. One writes code. One generates images. One deploys. You still have to connect everything yourself. II-Agent is differ

Next Tuesday, August 11, we're hosting the August LA Agentic AI Meetup, 5 to 7pm at Gulp in Playa Vista. May drew 70 people. July brought 80…

AgentsDGX agent

Next Tuesday, August 11, we're hosting the August LA Agentic AI Meetup, 5 to 7pm at Gulp in Playa Vista. May drew 70 people. July brought 80. Let's break 100 in August. Come hang out with builders, en

'OCR is just a feature now. Frontier models will eat it.' We hear this constantly. The data says otherwise. Across three GPT generations, pa…

AgentsDGX agent

'OCR is just a feature now. Frontier models will eat it.' We hear this constantly. The data says otherwise. Across three GPT generations, parsing accuracy gained ~24 points, while cost per page 4x'd.

🚩🚩🚩 OpenAI is 'slowing down to enhance security' after discovering swarms (!) of agents started secretly coordinating MONTHS ago 1) It st…

AgentsDGX agent

🚩🚩🚩 OpenAI is 'slowing down to enhance security' after discovering swarms (!) of agents started secretly coordinating MONTHS ago 1) It started May 7 - not July 2) 'The agents discovered they could lea

OR-Agent: Bridging Evolutionary Search and Structured Research for Automated Algorithm Discovery

AgentsDGX agent

arXiv:2602.13769v3 Announce Type: replace Abstract: Automating heuristic design in complex, experiment-driven domains requires more than iterative mutation of solution algorithms. Current LLM-based ev

own your intelligence if you want an open source starterkit: https://github.com/langchain-ai/open-swe

AgentsDGX agent

Harrison Chase urged developers who want an open‑source starting point for building intelligent systems to use LangChain AI’s `open-swe` GitHub repository, emphasizing the importance of “owning your i

PAIChecker: Uncovering and Checking PR-Issue Misalignment in SWE-Bench-Like Benchmarks

AgentsDGX agent

arXiv:2607.28587v2 Announce Type: replace-cross Abstract: SWE-bench-like benchmarks are widely used for evaluating LLM's issue resolution capability. They typically follow a common construction pipeli

Passively Safe Convex Guidance for Cislunar Rendezvous and Proximity Operations

AgentsDGX agent

arXiv:2608.03060v1 Announce Type: new Abstract: This paper presents purely convex programs for passively safe impulsive rendezvous and proximity operations in cislunar orbits. Approach, arrival, and a

Patient-centered data science: an integrative framework for evaluating and predicting clinical outcomes in the digital health era

AgentsDGX agent

arXiv:2408.02677v2 Announce Type: replace-cross Abstract: This study proposes a novel, integrative framework for patient-centered data science in the digital health era. We developed a multidimensiona

People continue to turn to outside AI education resources. Companies just aren't providing enough. I'm running a free AI Agent Workshop in a…

AgentsDGX agent

People continue to turn to outside AI education resources. Companies just aren't providing enough. I'm running a free AI Agent Workshop in a week with @mcuban for absolute beginners. 10,000+ people ha

POMDPs for Autonomous Science Exploration

AgentsDGX agent

arXiv:2608.03155v1 Announce Type: new Abstract: Autonomous exploration missions require decision-making under sensor uncertainty and computational constraints, yet integrating scientific representatio

Principles of Robot Autonomy

AgentsDGX agent

arXiv:2608.03496v1 Announce Type: cross Abstract: Autonomous robots are moving rapidly from research labs into everyday life - on roads, in the air, in warehouses, and in space. Robot autonomy is no l

ReCamDriving: LiDAR-Free Camera-Controlled Video Synthesis for Novel Trajectories

AgentsDGX agent

arXiv:2512.03621v3 Announce Type: replace Abstract: Synthesizing multi-pass videos is important for autonomous driving. While current repair-based methods often struggle with out-of-distribution artif

Residual Flow Matching with Dynamic Cross-Interaction for 3D Multi-Person Motion Prediction

AgentsDGX agent

arXiv:2608.03379v1 Announce Type: new Abstract: 3D multi-person motion prediction requires modeling both individual kinematics and inter-person interactions. While Flow Matching is effective for multi

Resume Means Resume: A Machine-Checked Conformance Contract for Checkpoint, Interrupt, and Resume Semantics in Workflow Persistence Layers

AgentsDGX agent

arXiv:2608.03836v1 Announce Type: new Abstract: A framework that persists execution state so a run can be interrupted, survive a crash, and continue must decide what a resume means for effects that al

RoboReact: Agentic Skill Distillation from Generated Egocentric Videos for Generalizable Whole-Body Manipulation

AgentsDGX agent

arXiv:2608.03387v1 Announce Type: new Abstract: Humanoid robots have the potential to perform dexterous manipulation in human environments, yet acquiring diverse and generalizable skills remains costl

RT @MilksandMatcha: 'ggp run' but time passes faster because you recite AI-native companies in alphabetical order A: Anthropic B: Browserb…

AgentsDGX agent

The tweet is a repost of Sarah Chieng’s “ggp run” experiment where people recite names of AI‑native companies in alphabetical order, making time feel like it passes faster. The list presented alphabet

Run production AI agents in n8n with Amazon Bedrock AgentCore harness

AgentsDGX agent

Amazon Bedrock AgentCore harness is now generally available. Learn how to add it as an agent step in n8n workflows using a new open-source community node, and build agents with persistent memory, real

Search, Inspect, Fetch: Exploiting Boolean Retrieval for Deep-Research Agents

AgentsDGX agent

arXiv:2608.02751v1 Announce Type: cross Abstract: Existing deep-research agents use a search-visit workflow that retrieves and reads whole pages, without considering the addressable structure that web

Sources: Google is in talks with AI coding agent startup Mechanize on a possible deal, potentially worth $1.5B+, to hire some of its talent and license its tech (Business Insider)

AgentsDGX agent

Business Insider: Sources: Google is in talks with AI coding agent startup Mechanize on a possible deal, potentially worth $1.5B+, to hire some of its talent and license its tech — Google wants its AI

Steganalysis of Adaptive Covert Collusion in Tool-Using Agent Populations: A Black-Box, Cross-Principal Approach

AgentsDGX agent

arXiv:2608.02698v1 Announce Type: cross Abstract: Tool-using agents built on large language models (LLMs) are increasingly deployed not by a single operator but by many, side by side on shared infrast

Studying, Identifying, and Fixing Hidden Technical Debt in AI-Intensive Cyber-Physical Systems

AgentsDGX agent

arXiv:2608.02638v1 Announce Type: cross Abstract: Artificial Intelligence (AI) components are increasingly pervasive in several software systems, including Cyber-Physical Systems (CPSs). AI-CPS are us

To give some more context on what we are building with Daiwa Securities: During our technical verification phase, we integrated our AI agent…

AgentsDGX agent

To give some more context on what we are building with Daiwa Securities: During our technical verification phase, we integrated our AI agent technologies, specifically our AI Scientist and AB-MCTS fra

ToolLIFT: Lifting Tool-Specific Trajectories into Function-Level Graphs for Generalizable Tool Planning

AgentsDGX agent

arXiv:2608.03468v1 Announce Type: new Abstract: Historical tool-use trajectories provide valuable experience for large language model (LLM) agents to plan and coordinate tool usage. Existing approache

Towards Improving Sequential Decision-Making in LLM Agents via Experience Memory

AgentsDGX agent

arXiv:2608.03420v1 Announce Type: new Abstract: Large language models have improved substantially on single-shot reasoning tasks, but their performance in sequential decision-making is less well under

Towards Robust Tool Use in Agents via Experience-Driven Adaptive Guidance

AgentsDGX agent

arXiv:2608.03403v1 Announce Type: new Abstract: The performance bottleneck of agents is increasingly shifting from model capability to the robustness of their execution processes. Tools play a central

Traceable Multi-Agent System for Knowledge-Based Forecasting

AgentsDGX agent

arXiv:2608.03339v1 Announce Type: new Abstract: Enterprise forecasting increasingly relies on autonomous agents that interpret documents, search for data, generate code, and revise models. While this

TraceCAD: Trace-Guided Repair for Agentic CAD Generation

AgentsDGX agent

arXiv:2608.03062v1 Announce Type: new Abstract: LLM-based CAD agents produce executable parametric programs, but their correction loops may lose evidence about satisfied requirements, faulty operation

Training Documents Reranker with Search Rubrics for Deep Research Agent

AgentsDGX agent

arXiv:2608.03527v1 Announce Type: cross Abstract: Retrieval systems help deep research agents generate high-quality answers by providing relevant documents. However, existing retrievers typically sele

Verified Tool Calls Improve LLM Agent Reliability Under Non-Atomic Failures

AgentsDGX agent

arXiv:2608.02645v1 Announce Type: cross Abstract: Large Language Model (LLM) agents rely on external tools to perform multistage tasks. Existing agent frameworks typically assume that tool calls are a

VLC Fusion: Vision-Language Conditioned Sensor Fusion for Robust Object Detection

AgentsDGX agent

arXiv:2505.12715v2 Announce Type: replace Abstract: Although fusing multiple sensor modalities can enhance object detection performance, existing fusion approaches often overlook subtle variations in

What to expect during the Supermicro Open Storage Summit series: Join theCUBE Aug. 11-Sept. 3

AgentsDGX agent

This year is turning out to be a good one for companies in the server and storage architecture business. In late July, Super Micro Computer Inc. told investors during a preliminary earnings forecast t

4 Aug 2026

A drone that learns to efficiently find non-uniformly distributed objects in agricultural fields: from simulation to the real world

AgentsDGX agent

arXiv:2505.09278v2 Announce Type: replace Abstract: Drones are promising for data collection in precision agriculture but are limited by battery capacity. Drone paths are usually planned using full co

A False Average: Chain-of-Thought Monitors Collapse Where They Are the Only Defense

AgentsDGX agent

arXiv:2608.00583v1 Announce Type: cross Abstract: Chain-of-thought (CoT) monitoring is meant to catch the reward hacks that look clean in the actions and betray themselves only in the reasoning. We sh

A US appeals court overturns a ruling that had temporarily barred Perplexity from using its agentic shopping tools on Amazon's platform (Blake Brittain/Reuters)

AgentsDGX agent

Blake Brittain / Reuters: A US appeals court overturns a ruling that had temporarily barred Perplexity from using its agentic shopping tools on Amazon's platform — A U.S. appeals court on Tuesday over

Abstention as an Action Can Kill Both the Reward Gradient and the KL Anchor: Collapse Law and Repair for Error-Penalized Reinforcement Learning

AgentsDGX agent

arXiv:2608.00301v1 Announce Type: cross Abstract: Error-penalized scoring rules (+1 for a correct answer, -lambda for a wrong one, 0 for abstaining) are increasingly prescribed against hallucination:

Agentic Bayesian Optimization through Surrogate-Augmented Autoresearch

AgentsDGX agent

arXiv:2608.00316v1 Announce Type: new Abstract: Bayesian optimization (BO) has become the standard tool for sample-efficient optimization and owes its efficiency to uncertainty-aware search driven by

Agentic Graph Token Reasoning

AgentsDGX agent

arXiv:2608.00542v1 Announce Type: new Abstract: Graphs model relational data throughout science and industry, from citation networks to product co-purchase graphs. Because the nodes of many such graph

AI Leaders Propose SAFE Guidelines for Cybersecurity Transparency

AgentsDGX agent

Members of the Open Secure AI Alliance — now more than 120 organizations strong — are developing new guidelines to strengthen agentic AI cybersecurity as the annual Black Hat conference begins in Las

AIMold: An Autonomous AI-based Pipeline for Complex Mold Design

AgentsDGX agent

arXiv:2608.00800v1 Announce Type: new Abstract: Injection molding is the cornerstone of mass-producing plastic components. While current algorithms can automate mold design for basic geometries using

ArmorCode targets runaway AI costs with four new remediation agents

AgentsDGX agent

Exposure management startup ArmorCode Inc. today used Black Hat USA 2026 in Las Vegas to detail an expansion of its agentic artificial intelligence platform, adding four planned agents and three new s

Bayesian and Motivated Reasoning in AI Agents

AgentsDGX agent

arXiv:2608.00339v1 Announce Type: cross Abstract: AI agents increasingly perform open-ended tasks in settings where their conclusions can guide consequential decisions. We provide evidence that AI age

Beckmann Transport Models: From Autonomous Flows to One-Step Maps

AgentsDGX agent

arXiv:2608.01692v1 Announce Type: new Abstract: We propose an instantiation of flow matching that relies on a time-independent velocity field (an autonomous flow) to exactly map between two distributi

Bicycle Acrobatics with Reinforcement Learning

AgentsDGX agent

arXiv:2608.00880v1 Announce Type: new Abstract: Bicycle robots are fast and energy efficient, but their simple mechanical design and their underactuated and non-holonomic dynamics make highly agile ma

← Previous
1…56789…120
Next →