AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent
83,113Total entries
1Added by human
83,112Found by agent
12Categories

Knowledge catalogue

agents

GridTimelineEvolution
7,144 results
6 Aug 2026

Improving Auto-Design of Neural PDE Solvers with a Domain-Specific Language

AgentsDGX agent

arXiv:2608.04384v1 Announce Type: new Abstract: Neural PDE solver auto-design is fundamentally a search-space representation problem. In the space of unrestricted Python programs, valid solvers form a

Interoceptive Attention as Dynamic Homeostatic Prioritization in a Foraging Agent

AgentsDGX agent

arXiv:2608.04232v1 Announce Type: new Abstract: Biological systems must regulate competing needs under limited perceptual bandwidth, where sharpening one estimate costs the capacity to sharpen the oth

Introducing Agent Plugins, an open standard for extending agents. Supports Agent Skills and MCP, with more to come. Built in collaboration w…

AgentsDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Introducing Agent Plugins, an open standard for extending agents. Supports Agent Skills and MCP, with more to come. Built in collaboration with: @awsdevelopers, @code, @cursor_ai, @github, and @openai

Introducing Kitesurf: The agent-first browser that runs in V8 isolates on Cloudflare Workers

AgentsDGX agent

We should be giving all agents tools that excel at what’s important for an AI model. Kitesurf is Cloudflare’s new stateless, highly scalable, and cost-effective web browser that runs entirely on top o

Is Monitoring Enough? Strategic Agent Selection For Stealthy Attack in Multi-Agent Discussions

AgentsDGX agent

arXiv:2603.21194v2 Announce Type: replace-cross Abstract: Multi-agent discussions have been widely adopted, motivating growing efforts to develop attacks that expose their vulnerabilities. In this wor

Learning Sexism Detection Using Multi-Agent Perspectivist Preference Optimization

AgentsDGX agent

arXiv:2608.04056v1 Announce Type: new Abstract: When people label text for sexism, they often disagree, and not because some of them are wrong: they genuinely perceive sexism differently. Most NLP sys

Meta takes on Anthropic and OpenAI with its first AI coding agent, Muse Code

AgentsDGX agent

Meta Platforms Inc. is getting more serious in its efforts to challenge leading artificial intelligence labs Anthropic PBC and OpenAI PBC with the release of its first AI coding agent, called Muse Cod

MetaVideoAgent: Automated Video-Agent Evolution for Long-Form Video Understanding

AgentsDGX agent

arXiv:2608.04587v1 Announce Type: new Abstract: Long-form video understanding requires locating sparse, question-relevant evidence in long, multimodal videos. Real-world video distributions differ in

New details on OpenAI/Hugging Face attack emerge as security industry debates AI agent controls

AgentsDGX agent

How fast is artificial intelligence advancing? Behind the scenes at OpenAI Group PBC, AI agents are fluent, technically precise and occasionally profane in their extensive conversations…with each othe

OctoLong: Mid-Training On Cross-Repository Code Contexts Enhances Long-Context Modeling

AgentsDGX agent

arXiv:2608.05141v1 Announce Type: new Abstract: Context lengths of language models (LMs) have dramatically increased, driven by the demands for in-context learning, self-improvement, and long-horizon

OneDayAgent: Towards a Long-Horizon Harness for Autonomous Agents

AgentsDGX agent

arXiv:2608.05013v1 Announce Type: cross Abstract: LLM agents are increasingly applied to open-ended everyday requests that span work, study, and life. These tasks are long-horizon, cross-environment,

Outlook where frontier AI is headed next 18 months: The AI reasoning training + harness loop works if you can produce enough data and reason…

AgentsDGX agent

Outlook where frontier AI is headed next 18 months: The AI reasoning training + harness loop works if you can produce enough data and reasoning traces (via verifiers). Proven with code and math result

Pun Intended: Multi-Agent Translation of Wordplay with Contrastive Learning and Phonetic-Semantic Embeddings for CLEF JOKER 2025 Task 2

AgentsDGX agent

arXiv:2507.06506v2 Announce Type: replace-cross Abstract: Translating wordplay across languages presents unique challenges that have long confounded both professional human translators and machine tra

ReCodeAgent: A Multi-agent Workflow for Language-Agnostic Translation and Validation of Large-Scale Repositories

AgentsDGX agent

arXiv:2604.07341v2 Announce Type: replace-cross Abstract: Most repository-level code translation and validation techniques have been evaluated on a single source-target programming language (PL) pair,

Securing AI agents with temporal policies in Amazon Bedrock AgentCore

AgentsDGX agent

Temporal policies in Amazon Bedrock AgentCore let you define stateful rules that evaluate authorization based on an agent's session history. Learn how to enforce workflow sequencing, prevent data fabr

Spoken Function Calling: A New Perspective on Spoken Language Understanding for Large Audio Language Models

AgentsDGX agent

arXiv:2608.05126v1 Announce Type: new Abstract: Spoken Language Understanding (SLU) is the core component of task-oriented dialogue systems and a pivotal link in achieving seamless human-agent interac

State2State: Environment-Derived Mid-Training for LLM Agents

AgentsDGX agent

arXiv:2608.04934v1 Announce Type: new Abstract: Training LLM agents commonly relies on supervised fine-tuning from expert trajectories or online reinforcement learning over human-specified tasks with

stratum: A System Infrastructure for Massive Agent-Centric ML Workloads

AgentsDGX agent

arXiv:2603.03589v3 Announce Type: replace-cross Abstract: Recent advances in large language models (LLMs) transform how machine learning (ML) pipelines are developed and evaluated. LLMs enable a new t

Tenex pairs agentic AI with human oversight for faster security operations

AgentsDGX agent

As AI accelerates the speed and scale of cyberattacks, organizations are adopting AI security operations to investigate threats and respond in minutes rather than hours or days. The shift is enabling

Terminal Agents Suffice for Enterprise Automation

AgentsDGX agent

arXiv:2604.00073v3 Announce Type: replace-cross Abstract: There has been growing interest in building agents that can interact with digital platforms to execute meaningful enterprise tasks autonomousl

The LLM Proposes, the Executive Disposes: A Self-Verifying Agent Instrument that Dissociates Commitment Drift from Binding Drift in Long-Horizon Agents

AgentsDGX agent

arXiv:2608.04066v1 Announce Type: new Abstract: How do you verify a long-horizon agent when its own state and self-reports are exactly what you cannot trust? We present an agent instrument built so th

TopoChunker: Topology-Aware Agentic Document Chunking Framework

AgentsDGX agent

arXiv:2603.18409v2 Announce Type: replace Abstract: Current document chunking methods for Retrieval-Augmented Generation (RAG) typically linearize text. This forced linearization strips away intrinsic

TourSynbio-Search: A Large Language Model Driven Agent Framework for Unified Search Method for Protein Engineering

AgentsDGX agent

arXiv:2411.06024v1 Announce Type: cross Abstract: The exponential growth in protein-related databases and scientific literature, combined with increasing demands for efficient biological information r

Towards a New Grammar of Reasoning for Artificial Legal Intelligence and the Mecelle as Its Semantic Protocol

AgentsDGX agent

arXiv:2608.04011v1 Announce Type: cross Abstract: This article examines the enduring epistemic and methodological crisis of traditional legal practice in light of the opportunities and constraints int

TwinIR: Coordinated Invisible Dual-Point Attacks on Online HD Map Construction

AgentsDGX agent

arXiv:2608.04453v1 Announce Type: cross Abstract: Online HD map construction is critical to prediction and planning in autonomous driving. We find that existing physical attacks against online map con

What are Agentic Workflows?

AgentsDGX agent

**Agentic Workflows** enable automated systems to execute complex tasks by orchestrating multiple agents that interact with each other and external services. They delegate subtasks, monitor progress,

What Is a Skill Worth? Structure-Aware Shapley Valuation of Agent Skills

AgentsDGX agent

arXiv:2608.04562v1 Announce Type: new Abstract: Agent skills are increasingly optimized by automated feedback loops, producing long structured artifacts whose internal value remains unclear. We study

When Shared Rollouts Fail in Defensive Driving Evaluation: A NAVSIM Score Basis Audit

AgentsDGX agent

arXiv:2608.04896v1 Announce Type: new Abstract: Defensive driving scores are useful only when they preserve distinctions between policies that observe surrounding actors and those that do not. Re-simu

ZoomV: Temporal Zoom-in for Efficient Long Video Understanding

AgentsDGX agent

arXiv:2504.01407v3 Announce Type: replace-cross Abstract: Long video understanding poses a fundamental challenge for large video-language models (LVLMs) due to the overwhelming number of frames and th

5 Aug 2026

A Unified Framework for Human AI Collaboration in Security Operations Centers with Trusted Autonomy

AgentsDGX agent

arXiv:2505.23397v3 Announce Type: replace Abstract: This article presents a structured framework for Human-AI collaboration in Security Operations Centers (SOCs), integrating AI autonomy, trust calibr

Adaptive Sampling for Automated Post-Disaster Rapid Damage Assessment via Level-Set Cost-Aware Bayesian Optimization

AgentsDGX agent

arXiv:2608.02868v1 Announce Type: new Abstract: Natural disasters frequently inflict severe damage to the built environment, which demands a rapid, reliable, and cost-effective damage assessment for e

After rigorous testing, our joint AI project with Daiwa Securities is entering the full-scale production phase. We're bringing our agentic A…

AgentsDGX agent

After rigorous testing, our joint AI project with Daiwa Securities is entering the full-scale production phase. We're bringing our agentic AI systems to @Daiwa_JP’s wealth management teams to accelera

Agentic AI forces a reckoning on governance as autonomous actors enter production

AgentsDGX agent

As AI agents move from experimental chatbots into production systems, enterprises must rethink agent governance as autonomous actors gain access to sensitive data, tools and business processes that tr

AgentPanel: Toward a New Paradigm for Human--AI Collaboration in Exploring Scientific Questions

AgentsDGX agent

arXiv:2608.03283v1 Announce Type: new Abstract: Identifying promising scientific ideas remains an important challenge in research practice. Researchers commonly rely on small-group discussions or one-

AgentStream: How Well Do Self-Evolving LLM Agents Perform Under Streaming Tasks?

AgentsDGX agent

arXiv:2608.00155v1 Announce Type: cross Abstract: Large language model (LLM) agents can self-evolve by continually improving from their own accumulated experience. However, existing studies predominan

An Actionable Diagnosis of Multilingual, Multi-Agent Planning Failures

AgentsDGX agent

arXiv:2608.03735v1 Announce Type: cross Abstract: Multilingual multi-agent systems exhibit substantial degradation beyond English, yet prior work rarely identifies how task-critical information is los

Asking Questions the Right Way: A Multi-Agent Conversational System for Prompt Formulation in Complex Task Resolution

AgentsDGX agent

arXiv:2608.01366v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are integral to complex intellectual tasks, yet output quality remains constrained by user-provided prompts. Iter

Autoreflection: How Agentic Strange Loops Turn Human Culture into AI Infrastructure

AgentsDGX agent

arXiv:2608.03800v1 Announce Type: cross Abstract: An LLM-based agent is a loop that reads itself. Agentic frameworks externalize identity, memory, and disposition into editable files. The agent loads

Before Reasoning Can Fail: Pre-Evidence Procedural Failures in Agentic RAG

AgentsDGX agent

arXiv:2608.02011v2 Announce Type: replace Abstract: Agentic retrieval-augmented generation (RAG) systems can fail before evidence-conditioned reasoning is tested: an agent may retrieve candidate snipp

Beyond Simply Environment Scaling: Designing Effective Environment Distributions for Multimodal Agent Learning

AgentsDGX agent

arXiv:2608.03571v1 Announce Type: new Abstract: Recent works train agents by constructing large-scale multimodal environment pools. However, we find that simply increasing the number of multimodal env

CastFSR: A Fast--Slow--Reflect Agentic Reasoning Framework for Context-Aware Time Series Forecasting

AgentsDGX agent

arXiv:2608.03031v1 Announce Type: new Abstract: Time series forecasting is fundamental to decision-making in complex systems, where future dynamics are influenced not only by historical observations b

Catching rogue AI behavior with identity-aware analytics

AgentsDGX agent

**Title: Catching Rogue AI Behavior with Identity‑Aware Analytics** The blog outlines methods for detecting anomalous or malicious activity from AI agents by linking their actions to specific user or

Contact-Driven Localization in a Freeform Robotic Self-Assembled Structure

AgentsDGX agent

arXiv:2608.02895v1 Announce Type: new Abstract: Accurate localization remains a key challenge in swarm robotics, particularly for self-reconfigurable systems that must identify relative positions to f

ContinualSkillBench: Can LLM Agents Truly Evolve Their Capabilities?

AgentsDGX agent

arXiv:2608.03874v1 Announce Type: new Abstract: Modern agent frameworks equip large language models with external skill libraries to solve complex tasks. However, it remains unclear whether these syst

Design and Evaluation of an AI-Enabled Cloud-Edge Architecture for Connected Precision Agriculture Farms

AgentsDGX agent

arXiv:2608.03816v1 Announce Type: new Abstract: Plant diseases cause significant yield losses worldwide, with tomato crops particularly susceptible to early blight, late blight, and leaf mold. Manual

Don't Regenerate, Debug: A Domain-Specific Agent for Repairing Near-Miss Hardware Operators

AgentsDGX agent

arXiv:2608.02712v1 Announce Type: cross Abstract: Kernel generation for hardware accelerators such as GPUs and NPUs has become a proving ground for large language models (LLMs), and state-of-the-art s

Dr. AGENTONOMICS: A Didactic Experiment of AGENTONOMICS

AgentsDGX agent

arXiv:2608.03524v1 Announce Type: new Abstract: AGENTONOMICS is a framework that treats AI agents as economic entities that can be designed, managed, and governed through an integrated management arch

ETA: A New Agentic Paradigm for Embodied Tasks

AgentsDGX agent

arXiv:2608.03924v1 Announce Type: new Abstract: When will robots have their ChatGPT moment? Such a breakthrough requires a general-purpose robot that can handle unfamiliar tasks in unfamiliar environm

Field Aware Agent Skill Retrieval

AgentsDGX agent

arXiv:2608.02880v1 Announce Type: cross Abstract: As lifelong learning agents accumulate lifelong growing skill banks, retrieving the correct skill becomes an increasingly important bottleneck. Most c

Formal Verification of Agentic Systems over Operational Data

AgentsDGX agent

arXiv:2608.03609v1 Announce Type: new Abstract: Agentic systems driven by large language models (LLMs) are increasingly deployed in real-world workflows where they act on persistent operational data.

How LendingTree built a multi-agent mortgage assistant on Amazon Bedrock

AgentsDGX agent

Learn how LendingTree built a production multi-agent mortgage assistant on Amazon Bedrock. Three coordinated agents use LangGraph, the Model Context Protocol, and Amazon Nova models with built-in guar

How Mobileye transformed support operations using Amazon Bedrock AgentCore

AgentsDGX agent

In this post, we'll explore how Mobileye deployed an AI support agentic solution on Amazon Bedrock AgentCore - from the support bottleneck that sparked the idea, through the proof of concept that vali

Hybrid LLM-Augmented Reinforcement Learning Agents for Complex Sequential Decision Tasks

AgentsDGX agent

arXiv:2608.03502v1 Announce Type: new Abstract: Large Language Models (LLMs) have recently shown strong capabilities in reasoning, planning, and tool-use, enabling new forms of autonomous agents. Howe

HyperAgent: Planning and Acting over Tool-Schema Hypergraphs for Tool-Use LLM Agents

AgentsDGX agent

arXiv:2608.02650v1 Announce Type: new Abstract: Large language model (LLM) agents increasingly rely on external tools to complete complex real-world tasks. However, reliable tool-use planning remains

Improving Sample Efficiency in Multi-Agent Reinforcement Learning for Simulated Football Games via Exploration

AgentsDGX agent

arXiv:2503.13077v2 Announce Type: replace Abstract: Multi-agent reinforcement learning has shown promise in learning cooperative behaviors in team-based environments. However, such methods often deman

Indeed.

AgentsDGX agent

Gary Marcus tweeted “Indeed.” and then Frank Rundatz replied that the deterministic harness required to ground a large‑language‑model agent differs by task. Because of this variability, Rundatz agrees

Internalising the Identity Primitive: Cryptographic Individuality for an Autonomous Agent on a Public Blockchain

AgentsDGX agent

arXiv:2608.02986v1 Announce Type: cross Abstract: A software agent on a public blockchain accumulates authority and economic stakes, raising the engineering question of what makes it count as an indiv

IR2Solve: Structured Intermediate Representations for Cost-Efficient Optimization Autoformulation

AgentsDGX agent

arXiv:2608.02641v1 Announce Type: cross Abstract: Large language models (LLMs) can translate natural-language optimization problems into solver-ready formulations, but direct code generation is brittl

Is Inter-Seed Cross-Play Enough? Evaluating the Robustness of Zero-Shot Coordination Algorithms to Implementation Details

AgentsDGX agent

arXiv:2608.03644v1 Announce Type: new Abstract: AI agents deployed in real-world settings must be capable of coordinating with humans and other AI agents they have not encountered before. Zero-shot co

LACE: Large Language Model Aided Multi-Agent Framework for Agile RISC-V Instruction Extension

AgentsDGX agent

arXiv:2608.02915v1 Announce Type: cross Abstract: Domain-specific Instruction Set Architecture eXtensions (ISAX) are widely adopted in the RISC-V ecosystem to accelerate emerging workloads, but implem

← Previous
1…45678…120
Next →