AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,964 results
11 Aug 2026

Agentic AI for Clustering, Relationship Discovery, and Semantic Trading in Prediction Markets

Model ReleasesDGX agent

arXiv:2512.02436v2 Announce Type: replace Abstract: Prediction markets allow users to trade on outcomes of real-world events, but are prone to fragmentation with overlapping questions, implicit equiva

ATLAS: Agentic Taxonomy of Large-Scale Software Ecosystems

Model ReleasesDGX agent

arXiv:2606.21597v2 Announce Type: replace-cross Abstract: The open-source ecosystem on GitHub lacks a systematic hierarchical taxonomy of software repositories. GitHub Topics, the dominant organizatio

Can LLM Agents Stick to the Script? A Benchmark for Long-Horizon Consistency in Interactive Narratives

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2608.08160v1 Announce Type: cross Abstract: The rapid advancement of Large Language Models (LLMs) is revolutionizing AI for Games by enabling open-ended and fluid interactive storytelling. Howev

Diagnosing as Cardiologists Do: ECG Agents with Doctor-Grounded Priors for Clinical Reasoning Across Diseases and Populations

Model ReleasesDGX agent

arXiv:2608.09053v1 Announce Type: cross Abstract: Cardiologists interpret electrocardiograms by localizing waveform components, measuring rhythm and interval patterns, and translating these structured

HUMAN HARNESS! @mcuban casually made up the term when talking abt how we differentiate when everyone can create same thing/agent. Empathy, P…

AgentsDGX agent

HUMAN HARNESS! @mcuban casually made up the term when talking abt how we differentiate when everyone can create same thing/agent. Empathy, Personal Brand, using time agents save you to offer unique hu

NVIDIA Nemotron 3.5 Lightning Delivers Fast, Accurate Specialized Task Execution for Long-Running Agents

Model ReleasesDGX agent

NVIDIA Nemotron 3.5 Lightning is an open‑source mixture‑of‑experts language model totaling 30 B parameters with only 3 B active during inference, designed to serve high‑volume, low‑latency execution f

Query-Only Backdoor Attacks on Self-Evolving Skills via Trajectory Poisoning

AgentsDGX agent

arXiv:2608.08303v1 Announce Type: new Abstract: Agentic skills improve large language model (LLM) agents by encoding reusable procedures for complex tasks. However, manually authored skills often adap

REMAC: Self-Reflective and Self-Evolving Multi-Agent Collaboration for Long-Horizon Robot Manipulation

Model ReleasesDGX agent

arXiv:2503.22122v2 Announce Type: replace-cross Abstract: Vision-language models (VLMs) have demonstrated remarkable capabilities in robotic planning, particularly for long-horizon tasks that require

The Politician, the Liar, and the Obedient Worker: Emerging Behavior of LLM Agents in Hierarchical Games

Model ReleasesDGX agent

arXiv:2608.09574v1 Announce Type: new Abstract: LLMs are rapidly embedding themselves into daily life: drafting our emails, managing our schedules, and making decisions on our behalf. As they move fro

10 Aug 2026

CEDAR: Agent-Orchestrated Tree Search for Goal-Directed Optimization of Complex Systems

SafetyDGX agent

arXiv:2608.06871v1 Announce Type: new Abstract: Complex systems, core objects of study in artificial life, model diverse phenomena through nonlinear, feedback-driven interactions that produce emergent

Shape Your Feed: An LLM-based Agentic System for Conversational Recommendation

SafetyDGX agent

arXiv:2608.06632v1 Announce Type: new Abstract: Industrial recommendation systems predominantly adopt a passive ranking paradigm that infers user preferences from implicit behavioral signals (e.g., cl

7 Aug 2026

AgentOPSD: Recursive Self-Distillation for Agentic Reinforcement Learning

SafetyDGX agent

arXiv:2608.05987v1 Announce Type: new Abstract: Reinforcement learning (RL) with verifiable rewards constructs trajectory-level advantage estimates, yet it often fails to credit the few pivotal decisi

Evaluating Investment Logic in Large Language Models: A Real-World Benchmark Towards Personalzied Financial Agents

Model ReleasesDGX agent

arXiv:2608.06108v1 Announce Type: new Abstract: Investment competence is inherently personalized: the same market evidence can justify different actions for investors with different goals, horizons, p

'Felony humble-bragging' is a great line

AgentsDGX agent

'Felony humble-bragging' is a great line At final Black Hat keynote (called a locknote, ha ha) panelists say they are surprised at how the OpenAI - Hugging Face incident debrief, as well as other repo

6 Aug 2026

A-SR: Self-Evolving Agentic LLMs for Symbolic Regression via Hierarchical Coordination

SafetyDGX agent

arXiv:2608.04872v1 Announce Type: cross Abstract: Symbolic regression aims to discover closed-form equations from data, but existing LLM-guided methods often rely on a unified proposal loop that compr

FinRpt: Dataset, Evaluation System and LLM-based Multi-agent Framework for Equity Research Report Generation

Model ReleasesDGX agent

arXiv:2511.07322v3 Announce Type: replace-cross Abstract: While LLMs have shown great success in financial tasks like stock prediction and question answering, their application in fully automating Equ

Outlook where frontier AI is headed next 18 months: The AI reasoning training + harness loop works if you can produce enough data and reason…

AgentsDGX agent

Outlook where frontier AI is headed next 18 months: The AI reasoning training + harness loop works if you can produce enough data and reasoning traces (via verifiers). Proven with code and math result

PICopilot: An LLM-based Agentic Framework for Assisting Photonic Integrated Circuit Design via Script Generation

Model ReleasesDGX agent

arXiv:2608.01791v2 Announce Type: replace-cross Abstract: The rapid development of photonic integrated circuits (PICs) is shifting the design flow from traditional graphical user interface (GUI)-based

SimMOF: AI agent for Automated MOF Simulations

Model ReleasesDGX agent

arXiv:2603.29152v2 Announce Type: replace Abstract: Metal-organic frameworks (MOFs) offer a vast design space, and as such, computational simulations play a critical role in predicting their structura

5 Aug 2026

Beyond the Single Camera: Agentic Multi-View Reasoning in Sports Video Understanding

Model ReleasesDGX agent

arXiv:2607.11844v2 Announce Type: replace Abstract: Recent Multimodal Large Language Models (MLLMs) achieve strong performance on single-view video understanding benchmarks. However, sports videos inv

Is Inter-Seed Cross-Play Enough? Evaluating the Robustness of Zero-Shot Coordination Algorithms to Implementation Details

AgentsDGX agent

arXiv:2608.03644v1 Announce Type: new Abstract: AI agents deployed in real-world settings must be capable of coordinating with humans and other AI agents they have not encountered before. Zero-shot co

KernelBrain: Coarse-to-Fine, Budget-Aware Search for Agentic GPU Kernel Optimization

SafetyDGX agent

arXiv:2608.02611v1 Announce Type: cross Abstract: Automating GPU kernel optimization remains difficult in practice: generated variants can violate correctness constraints, runtime measurements are noi

MemArena: An Ego-Centric Benchmark for On-Device Agentic Personal Memory Assistants at Scale

Model ReleasesDGX agent

arXiv:2608.02613v1 Announce Type: cross Abstract: Edge-deployed personal memory assistants must handle private interpersonal conversations on-device with open-weight models. Yet, existing memory bench

4 Aug 2026

PhysAgent: A Multi-Agent Framework for Reliable Remote Heart Rate Estimation

Model ReleasesDGX agent

arXiv:2608.00066v1 Announce Type: new Abstract: Remote photoplethysmography (rPPG) enables non-contact heart-rate estimation from facial videos, but its weak physiological signal is easily corrupted b

3 Aug 2026

AuEmoChat: Authentic Emotion Understanding and Rendering for Conversational Speech Synthesis

AgentsDGX agent

arXiv:2607.15755v2 Announce Type: replace-cross Abstract: Conversational Speech Synthesis (CSS) aims to synthesize speech with human-like emotional expression and contextual consistency in user-agent

SERUM: State Extraction and Refinement for User Modeling

AgentsDGX agent

arXiv:2607.29181v1 Announce Type: cross Abstract: Agentic assistants capable of proactive, personalized interactions require structured models of user intent and workflow. However, building these mode

Shaping Scientific Explanations to Expert Perspectives with Persona-Conditioned Reinforcement Learning

AgentsDGX agent

arXiv:2603.21846v2 Announce Type: replace Abstract: Explainable AI is increasingly important to scientific discovery. However, existing methods largely ignore that explanation quality is not universal

The Formalism Trap: Are LLM-as-a-Judge Evaluators Blinded by Consensus Mimicry under Social Load?

AgentsDGX agent

arXiv:2607.28641v1 Announce Type: cross Abstract: We introduce the extit{Agentic Formalism Trap} and the Evaluative Dissonance Index (D_E), quantifying how LLM-as-a-Judge systems conflate structural p

31 Jul 2026

A Physics-Informed Framework for PID Tuning of Chemical Processes Using Large Language Model Agents

Model ReleasesDGX agent

arXiv:2607.26594v1 Announce Type: cross Abstract: PID tuning for chemical processes commonly relies on identified process models, whereas plant engineers often retune loops iteratively by observing re

Fantastic Adaptive Taxonomies and How to Use Them

Model ReleasesDGX agent

arXiv:2607.16387v2 Announce Type: replace-cross Abstract: An agent system's execution traces record how it fails, and procedures that improve such a system without changing model weights (trajectory s

Pramana: A Composable, Domain-Specific Backend for Empirical Networking Research

AgentsDGX agent

arXiv:2607.26352v1 Announce Type: cross Abstract: Networking research advances by turning hypotheses into empirical evidence, so accelerating it means reducing the lag between ideation (synthesizing a

30 Jul 2026

Learning Implicit Causal World Models from Multi-Agent Demonstrations

SafetyDGX agent

arXiv:2607.26336v1 Announce Type: new Abstract: In model-based reinforcement learning, world models exist as internal simulators, but their training often conflates statistical correlations with causa

29 Jul 2026

MARS: Multi-Agent Re-ranking for Repeat-Order Food Delivery Recommendation

Local AiDGX agent

arXiv:2607.25420v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used in recommender systems, but it is often unclear how much performance can be obtained from strong pr

28 Jul 2026

Coherent Without Grounding, Grounded Without Success: Observability and Epistemic Failure

AgentsDGX agent

arXiv:2603.28371v2 Announce Type: replace-cross Abstract: When an agent can articulate why something works, we typically take this as evidence of genuine understanding. This presupposes that effective

WCM: World-Cognition Model for Generalizable Human-Robot Interaction

AgentsDGX agent

arXiv:2607.22999v1 Announce Type: cross Abstract: Language agents can now interact fluently with users in software, but robots still struggle to bring comparable interaction to physical tasks. Current

27 Jul 2026

Encoding Invisible Causation for Bridge Diagnostic Agents: Triple-Guided Retrieval-Augmented Fine-Tuning with QLoRA

Model ReleasesDGX agent

arXiv:2607.21680v1 Announce Type: new Abstract: Bridge infrastructure deteriorates gradually, yet its root causes---salt intrusion, freezing, fatigue cracking, and others---remain invisible to the nak

TRACE-ROUTER: Task-Consistent and Adaptive Online Routing for Agentic AI

SafetyDGX agent

arXiv:2607.22465v1 Announce Type: cross Abstract: Routing to select large language models (LLMs) with different cost-quality trade-offs has become a fundamental deployment feature of enterprise AI. Ex

24 Jul 2026

CachyLLama’s: llama.cpp fork with persistent KV cache that makes long local-agent sessions much less painful

Model ReleasesDGX agent

I’m not affiliated with this project, but I’ve been running it recently and I’m surprised it hasn’t received more attention here: https://github.com/fewtarius/CachyLLama CachyLLama is a fork of llama.

Conflict Resolution under Degraded Surveillance in Air Corridors Using Multi-Agent Reinforcement Learning

Local AiDGX agent

arXiv:2607.20547v1 Announce Type: new Abstract: Safe Advanced Air Mobility operations require aircraft to maintain separation when surveillance information is noisy, delayed, incomplete, or temporaril

SAGE: A Socially-Aware Generative Engine for Heterogeneous Multi-Agent Navigation

SafetyDGX agent

arXiv:2607.16619v2 Announce Type: replace Abstract: Safe and socially compliant navigation in open human-robot environments requires robots to reason about heterogeneous participants with different dy

Verifier-First Evaluation of Agentic LLMs for Infrastructure-as-Code Generation

Model ReleasesDGX agent

arXiv:2607.20478v1 Announce Type: cross Abstract: Infrastructure-as-Code (IaC) generation from natural language requires satisfying provider schemas, dependency planning, and organizational policy con

23 Jul 2026

Code-in-the-Loop Forensics: Agentic Tool Use for Image Forgery Detection

Model ReleasesDGX agent

arXiv:2512.16300v3 Announce Type: replace Abstract: Existing image forgery detection (IFD) methods either exploit low-level, semantics-agnostic artifacts or rely on multimodal large language models (M

EvoDRC: A Self-Evolving Agentic Framework for Automated DRC Violation Repair

Model ReleasesDGX agent

arXiv:2607.20019v1 Announce Type: new Abstract: Design rule check (DRC) closure remains a major bottleneck in advanced-node physical design. Although detailed routers are rule-aware, residual design r

Minimize idle accelerators: Native RL job interleaving with co-operative time-slicing in llm-d

SafetyDGX agent

The math behind reinforcement learning (RL) post-training for large language models (LLMs) is notoriously unforgiving. As frontier AI labs push the boundaries of reasoning and coding models using RL p

TriAgent: Divergence-Aware Multi-Agent Committees for Cost-Efficient Financial Sentiment Analysis

Model ReleasesDGX agent

arXiv:2607.19794v1 Announce Type: new Abstract: Production LLM-based financial sentiment analysis faces a structural cost trap: most queries are trivially classifiable, yet expensive cloud reasoners p

16 Jul 2026

Evaluating Frontier AI Agents as Autonomous Clinical Security Auditors

Model ReleasesDGX agent

arXiv:2607.13411v1 Announce Type: cross Abstract: Clinical AI models can expose patients to harm when adversarial vulnerabilities go undetected, yet formal security auditing requires statistical exper

Interaction Protocol Shapes Moral Judgment in Multi-Agent Debate

Model ReleasesDGX agent

arXiv:2510.10002v3 Announce Type: replace Abstract: As large language models (LLMs) are increasingly deployed in sensitive everyday contexts -- offering personal advice, mental health support, and mor

Learning Robust Execution in Robotic Manipulation with Agentic Reinforcement Learning

Model ReleasesDGX agent

arXiv:2607.13818v1 Announce Type: new Abstract: Robotic manipulation poses fundamental challenges due to uncertainty, long-horizon execution, and compounding errors, which can easily destabilize execu

15 Jul 2026

Auditable Context-Aware HFMD Forecasting with Structured LLM Agents

SafetyDGX agent

arXiv:2511.23276v2 Announce Type: replace Abstract: Effective HFMD surveillance requires forecasts capturing both time-series patterns and contextual drivers such as school calendars, weather, and pol

Built Technologies builds an AI-powered document intelligence solution on AWS to power agents across real estate finance

ApplicationsDGX agent

Built partnered with the AWS Generative AI Innovation Center (GenAIIC), AWS Partner AND Digital, and AWS account teams to create a scalable, AI-powered document processing engine that can classify, sp

TRACE: An Operational Reasoning Schema for Auditable Agentic Commitments

Model ReleasesDGX agent

arXiv:2607.12480v1 Announce Type: new Abstract: This paper defines TRACE (Typed Reasoning And Commitment Evidence): a typed, versioned schema for recording reasoning traces, a reference procedure for

ViCo3D: Empowering LiDAR-based Collaborative 3D Object Detection with Vision Foundation Models

AgentsDGX agent

arXiv:2607.12959v1 Announce Type: new Abstract: LiDAR-based collaborative 3D perception in Vehicle-to-Everything (V2X) systems typically relies on fusing bird's-eye-view (BEV) features across agents.

10 Jul 2026

A Vision Toward Energy-Efficient Domain-Specific Artificial Intelligence Models and Agents

Model ReleasesDGX agent

arXiv:2510.22052v2 Announce Type: replace Abstract: The field of artificial intelligence (AI) has taken a tight hold on broad aspects of society, industry, business, and governance in ways that dictat

Aleena: Alignment Agent for Research Software Engineering Collaborations

SafetyDGX agent

arXiv:2607.08043v1 Announce Type: cross Abstract: Research software collaborations span meetings, informal chats, pull requests, and GitHub issues. A decision surfaced in a Slack thread, refined in a

Harness VLA: Steering Frozen VLAs into Reliable Manipulation Primitives via Memory-Guided Agents

Local AiDGX agent

arXiv:2607.08448v1 Announce Type: new Abstract: Language-conditioned manipulation requires both precise contact-rich control and robust reasoning over language, scenes, and long horizons. End-to-end V

HumanForge: A Human-Centric Deepfake Video Benchmark with Multi-Agent Forgery Rationales

Model ReleasesDGX agent

arXiv:2607.08705v1 Announce Type: new Abstract: Rapid advancements in video diffusion models and temporal editing tools have enabled the generation of highly realistic human-centric videos, posing unp

Provably Optimal Learning Algorithms for Assistance Games

AgentsDGX agent

arXiv:2607.08012v1 Announce Type: cross Abstract: This paper studies an online variant of the assistance games framework, where an informed agent and an uninformed agent repeatedly interact over T tim

9 Jul 2026

Cost-Effective Agent Harnesses for Abstract Reasoning and Generalization on ARC-AGI-1

Model ReleasesDGX agent

arXiv:2607.06764v1 Announce Type: new Abstract: Recent progress on ARC-AGI-1 from disclosed architectures has come broadly from two regimes: heavy test-time compute over frontier models (evolutionary

7 Jul 2026

Anytime Plug-and-Play Control with Contract-Based Distributed MPC

AgentsDGX agent

arXiv:2607.04215v1 Announce Type: cross Abstract: A central challenge in many mobile multi-robot applications is that communication topologies are inherently time-varying. Agents may enter or exit the

Beyond Forecasting: The Belief-to-Trade Layer in Prediction-Market Agents

Model ReleasesDGX agent

arXiv:2607.03015v1 Announce Type: new Abstract: Forecasting future events has attracted growing attention as a testbed for general-purpose AI. A natural way to ground this evaluation is let the models

← Previous
1…137138139140141…300
Next →