AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,951 results
17 Apr 2026

IE as Cache: Information Extraction Enhanced Agentic Reasoning

AgentsDGX agent

arXiv:2604.14930v1 Announce Type: new Abstract: Information Extraction aims to distill structured, decision-relevant information from unstructured text, serving as a foundation for downstream understa

MARS^2: Scaling Multi-Agent Tree Search via Reinforcement Learning for Code Generation

SafetyDGX agent

arXiv:2604.14564v1 Announce Type: cross Abstract: Reinforcement learning (RL) paradigms have demonstrated strong performance on reasoning-intensive tasks such as code generation. However, limited traj

Neuro-Oracle: A Trajectory-Aware Agentic RAG Framework for Interpretable Epilepsy Surgical Prognosis

Model ReleasesDGX agent

arXiv:2604.14216v1 Announce Type: cross Abstract: Predicting post-surgical seizure outcomes in pharmacoresistant epilepsy is a clinical challenge. Conventional deep-learning approaches operate on stat

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

OpenMobile: Building Open Mobile Agents with Task and Trajectory Synthesis

Model ReleasesDGX agent

arXiv:2604.15093v1 Announce Type: cross Abstract: Mobile agents powered by vision-language models have demonstrated impressive capabilities in automating mobile tasks, with recent leading models achie

Sharing my current setup to run Qwen3.6 locally in a good agentic setup (Pi + llama.cpp). Should give you a good overview of how good local …

Model ReleasesDGX agent

Sharing my current setup to run Qwen3.6 locally in a good agentic setup (Pi + llama.cpp). Should give you a good overview of how good local agents are today: # Start llama.cpp server: llama-server -hf

16 Apr 2026

Agent evals are drifting away from production reality. Most benchmarks use clean tasks, well-specified requirements, deterministic metrics, …

Model ReleasesDGX agent

Agent evals are drifting away from production reality. Most benchmarks use clean tasks, well-specified requirements, deterministic metrics, and retrospective curation. Production work is messier, with

Claude Opus 4.7 is now available as an Agent Preview inside of Devin! Anthropic has clearly optimized Claude Opus 4.7 for long-horizon auton…

Model ReleasesDGX agent

Claude Opus 4.7 is now available as an Agent Preview inside of Devin! Anthropic has clearly optimized Claude Opus 4.7 for long-horizon autonomy, unlocking a class of deep investigation work we couldn'

Debate to Align: Reliable Entity Alignment through Two-Stage Multi-Agent Debate

SafetyDGX agent

arXiv:2604.13551v1 Announce Type: new Abstract: Entity alignment (EA) aims to identify entities referring to the same real-world object across different knowledge graphs (KGs). Recent approaches based

Three insights you might have missed from theCUBE’s coverage of Nutanix .NEXT

AgentsDGX agent

Agentic infrastructure is quickly becoming the foundation for how enterprises build and run AI at scale. What’s emerging is a shift from piecemeal tools to integrated systems where data, models and op

Vercel Workflows is GA. Your code is the orchestrator. Ship agents, backends, or any long-running process without managing queues, retries, …

ToolsDGX agent

Vercel Workflows is GA. Your code is the orchestrator. Ship agents, backends, or any long-running process without managing queues, retries, or workers. https://vercel.com/blog/a-new-programming-model-

15 Apr 2026

Agentic Control in Variational Language Models

Local AiDGX agent

arXiv:2604.12513v1 Announce Type: new Abstract: We study whether a variational language model can support a minimal and measurable form of agentic control grounded in its own internal evidence. Our mo

Cycle-Consistent Search: Question Reconstructability as a Proxy Reward for Search Agent Training

SafetyDGX agent

arXiv:2604.12967v1 Announce Type: new Abstract: Reinforcement Learning (RL) has shown strong potential for optimizing search agents in complex information retrieval tasks. However, existing approaches

Just a Hermes Agent, a skill, and a dream

Model ReleasesDGX agent

Just a Hermes Agent, a skill, and a dream The crazy part? This was done (nearly) fully autonomously! Only 8 prompts from the human in the loop. Just a Hermes agent, a skill, and a dream. 🐉 I told my A

SEW: Self-Evolving Agentic Workflows for Automated Code Generation

Model ReleasesDGX agent

arXiv:2505.18646v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have demonstrated effectiveness in code generation tasks. To enable LLMs to address more complex coding challenge

The A-R Behavioral Space: Execution-Level Profiling of Tool-Using Language Model Agents in Organizational Deployment

SafetyDGX agent

arXiv:2604.12116v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed as tool-augmented agents capable of executing system-level operations. While existing benchmarks

Toward Autonomous Long-Horizon Engineering for ML Research

AgentsDGX agent

arXiv:2604.13018v1 Announce Type: new Abstract: Autonomous AI research has advanced rapidly, but long-horizon ML research engineering remains difficult: agents must sustain coherent progress across ta

14 Apr 2026

Autoscaling Autoresearch: Give your agents elastic GPUs on Modal

ToolsDGX agent

Modal's blog post describes how to build AI research agents that automatically scale GPU resources up and down using Modal's serverless infrastructure, eliminating the need to manage fixed GPU allocat

MSA: Memory Sparse Attention for Efficient End-to-End Memory Model Scaling to 100M Tokens

AgentsDGX agent

arXiv:2603.23516v2 Announce Type: replace-cross Abstract: Long-term memory is a cornerstone of human intelligence. Enabling AI to process lifetime-scale information remains a long-standing pursuit in

NEW: Added @OpenRouter's new free stealth model, Elephant-Alpha, to Hermes Agent, which you can now access if you run `hermes update`! I als…

Model ReleasesDGX agent

NEW: Added @OpenRouter's new free stealth model, Elephant-Alpha, to Hermes Agent, which you can now access if you run `hermes update`! I also had Hermes come up with an agentic benchmark on the fly to

Normative Common Ground Replication (NormCoRe): Replication-by-Translation for Studying Norms in Multi-Agent AI

SafetyDGX agent

arXiv:2603.11974v2 Announce Type: replace Abstract: In the late 2010s, the fashion trend NormCore framed sameness as a signal of belonging, illustrating how norms emerge through collective coordinatio

PICon: A Multi-Turn Interrogation Framework for Evaluating Persona Agent Consistency

SafetyDGX agent

arXiv:2603.25620v2 Announce Type: replace Abstract: Large language model (LLM)-based persona agents are rapidly being adopted as scalable proxies for human participants across diverse domains. Yet the

Plug-and-Play Dramaturge: A Divide-and-Conquer Approach for Iterative Narrative Script Refinement via Collaborative LLM Agents

Local AiDGX agent

arXiv:2510.05188v2 Announce Type: replace Abstract: Although LLMs have been widely adopted for creative content generation, a single-pass process often struggles to produce high-quality long narrative

SEARL: Joint Optimization of Policy and Tool Graph Memory for Self-Evolving Agents

SafetyDGX agent

arXiv:2604.07791v2 Announce Type: replace Abstract: Recent advances in Reinforcement Learning with Verifiable Rewards (RLVR) have demonstrated significant potential in single-turn reasoning tasks. Wit

SecureVibeBench: Evaluating Secure Coding Capabilities of Code Agents with Realistic Vulnerability Scenarios

Model ReleasesDGX agent

arXiv:2509.22097v3 Announce Type: replace-cross Abstract: Large language model-powered code agents are rapidly transforming software engineering, yet the security risks of their generated code have be

SkVM: Revisiting Language VM for Skills across Heterogenous LLMs and Harnesses

AgentsDGX agent

arXiv:2604.03088v3 Announce Type: replace-cross Abstract: LLM agents increasingly adopt skills as a reusable unit of composition. While skills are shared across diverse agent platforms, current system

TrajOnco: a multi-agent framework for temporal reasoning over longitudinal EHR for multi-cancer early detection

Model ReleasesDGX agent

arXiv:2604.10386v1 Announce Type: new Abstract: Accurate estimation of cancer risk from longitudinal electronic health records (EHRs) could support earlier detection and improved care, but modeling su

WisPaper: Your AI Scholar Search Engine

AgentsDGX agent

arXiv:2512.06879v3 Announce Type: replace-cross Abstract: We present extsc{WisPaper}, an end-to-end agent system that transforms how researchers discover, organize, and track academic literature. The

13 Apr 2026

Creator Incentives in Recommender Systems: A Cooperative Game-Theoretic Approach for Stable and Fair Collaboration in Multi-Agent Bandits

SafetyDGX agent

arXiv:2604.08643v1 Announce Type: new Abstract: User interactions in online recommendation platforms create interdependencies among content creators: feedback on one creator's content influences the s

Exploring Teachers' Perspectives on Using Conversational AI Agents for Group Collaboration

SafetyDGX agent

arXiv:2602.07142v2 Announce Type: replace-cross Abstract: Collaboration is a cornerstone of 21st-century learning, yet teachers continue to face challenges in supporting productive peer interaction. E

OpenKedge: Governing Agentic Mutation with Execution-Bound Safety and Evidence Chains

SafetyDGX agent

arXiv:2604.08601v1 Announce Type: new Abstract: The rise of autonomous AI agents exposes a fundamental flaw in API-centric architectures: probabilistic systems directly execute state mutations without

12 Apr 2026

Deepagents https://github.com/langchain-ai/deepagents

Model ReleasesDGX agent

Deepagents is an open-source project by LangChain that implements deep research-style agentic workflows, enabling AI agents to perform iterative, multi-step research and reasoning tasks. The repositor

if you bounce between claude code and codex, that’s completely normal. the models are good at different things and everyone's workflow is dy…

Model ReleasesDGX agent

if you bounce between claude code and codex, that’s completely normal. the models are good at different things and everyone's workflow is dynamic but you’re also paying multiple subscriptions, constan

MiniMax M2.7 Advances Scalable Agentic Workflows on NVIDIA Platforms for Complex AI Applications

HardwareDGX agent

MiniMax M2.7 is an enhancement of the MiniMax M2.5 model, built as a 230B-parameter Mixture-of-Experts (MoE) model with 10B active parameters per token and a 200K context length, designed for agentic

11 Apr 2026

back home from @aiDotEngineer which was amazing i’m used to staying on agentic X all day, reading, and then building, but seeing all those p…

AgentsDGX agent

back home from @aiDotEngineer which was amazing i’m used to staying on agentic X all day, reading, and then building, but seeing all those people IRL and interacting with them is way more inspiring th

10 Apr 2026

AnomalyAgent: Agentic Industrial Anomaly Synthesis via Tool-Augmented Reinforcement Learning

AgentsDGX agent

arXiv:2604.07900v1 Announce Type: new Abstract: Industrial anomaly generation is a crucial method for alleviating the data scarcity problem in anomaly detection tasks. Most existing anomaly synthesis

AV-SQL: Decomposing Complex Text-to-SQL Queries with Agentic Views

Model ReleasesDGX agent

arXiv:2604.07041v1 Announce Type: cross Abstract: Text-to-SQL is the task of translating natural language queries into executable SQL for a given database, enabling non-expert users to access structur

DosimeTron: Automating Personalized Monte Carlo Radiation Dosimetry in PET/CT with Agentic AI

Model ReleasesDGX agent

arXiv:2604.06280v1 Announce Type: cross Abstract: Purpose: To develop and evaluate DosimeTron, an agentic AI system for automated patient-specific MC internal radiation dosimetry in PET/CT examination

EmoMAS: Emotion-Aware Multi-Agent System for High-Stakes Edge-Deployable Negotiation with Bayesian Orchestration

Local AiDGX agent

arXiv:2604.07003v1 Announce Type: new Abstract: Large language models (LLMs) has been widely used for automated negotiation, but their high computational cost and privacy risks limit deployment in pri

Energy Saving for Cell-Free Massive MIMO Networks: A Multi-Agent Deep Reinforcement Learning Approach

AgentsDGX agent

arXiv:2604.07133v1 Announce Type: cross Abstract: This paper focuses on energy savings in downlink operation of cell-free massive MIMO (CF mMIMO) networks under dynamic traffic conditions. We propose

HY-Embodied-0.5: Embodied Foundation Models for Real-World Agents

Model ReleasesDGX agent

arXiv:2604.07430v1 Announce Type: new Abstract: We introduce HY-Embodied-0.5, a family of foundation models specifically designed for real-world embodied agents. To bridge the gap between general Visi

LangDriveCTRL: Natural Language Controllable Driving Scene Editing with Multi-modal Agents

SafetyDGX agent

arXiv:2512.17445v2 Announce Type: replace Abstract: LangDriveCTRL is a natural-language-controllable framework for editing real-world driving videos to synthesize diverse traffic scenarios. It represe

ParseBench: A Document Parsing Benchmark for AI Agents

Model ReleasesDGX agent

arXiv:2604.08538v1 Announce Type: new Abstract: AI agents are changing the requirements for document parsing. What matters is semantic correctness: parsed output must preserve the structure and

Visually-grounded Humanoid Agents

Model ReleasesDGX agent

arXiv:2604.08509v1 Announce Type: new Abstract: Digital human generation has been studied for decades and supports a wide range of real-world applications. However, most existing systems are passively

9 Apr 2026

extremely fun jamming with Harrison on all things analytics, agents and evals! I have like 20 more hours worth of takes and opinions in this…

Model ReleasesDGX agent

extremely fun jamming with Harrison on all things analytics, agents and evals! I have like 20 more hours worth of takes and opinions in this space and going on the pod uncorked them, more to come for

Here is my free workshop on building multi-agent systems, I presented at the @aiDotEngineer London conference together with @Whats_AI. It ha…

Model ReleasesDGX agent

Here is my free workshop on building multi-agent systems, I presented at the @aiDotEngineer London conference together with @Whats_AI. It has code, slides and soon a 2-hour video diving deep into how

this is one of those in awe moments - amazing work by the team on Deep Agents deploy. also one sneaky important aspect of this is you can se…

ApplicationsDGX agent

this is one of those in awe moments - amazing work by the team on Deep Agents deploy. also one sneaky important aspect of this is you can self-host the infra to make this happen. super super important

we got LangPod before GTA6 🙏 this series is gonna be sick - real stuff that breaks with agents, mental models, evals, tooling with some of …

ApplicationsDGX agent

we got LangPod before GTA6 🙏 this series is gonna be sick - real stuff that breaks with agents, mental models, evals, tooling with some of the best builders across industry great to openly share all t

14 Aug 2026

Qwen3.8-Max is live on Fireworks, ready for agents, heavy coding, and long-context work. Fireworks is right there with us as a Day 0 launch …

Model ReleasesDGX agent

Qwen3.8-Max is live on Fireworks, ready for agents, heavy coding, and long-context work. Fireworks is right there with us as a Day 0 launch partner. What a way to kick things off!It's Day 0, cue the F

Research Assistant: AstraZeneca's Agentic System for R&D

SafetyDGX agent

arXiv:2608.12395v1 Announce Type: new Abstract: We describe Research Assistant, an internal LLM-based system developed at AstraZeneca to help scientists and clinicians explore biomedical questions acr

Teach the Magnitude, Not the Direction: Verifier-Bounded Credit Assignment for Multi-Turn Multi-step LLM Agents

SafetyDGX agent

arXiv:2608.13179v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) offers a verifier-bounded performance ceiling for training multi-turn tool-use agents, yet its tra

13 Aug 2026

Accelerating M&A due diligence with Amazon Bedrock AgentCore

AgentsDGX agent

Learn how to build a multi-agent M&A due diligence system on Amazon Bedrock AgentCore. This post walks through a reference architecture that combines agent orchestration, knowledge retrieval, and gove

BrowseSafe: Understanding and Preventing Prompt Injection Within AI Browser Agents

Model ReleasesDGX agent

arXiv:2511.20597v2 Announce Type: replace-cross Abstract: The integration of artificial intelligence (AI) agents into web browsers introduces security challenges that go beyond traditional web applica

Current and former OpenAI employees say pressure to quickly ship products left less time for safety, contributing to incidents like the rogue agent hack (Maxwell Zeff/Wired)

SafetyDGX agent

Maxwell Zeff / Wired: Current and former OpenAI employees say pressure to quickly ship products left less time for safety, contributing to incidents like the rogue agent hack — OpenAI's rogue agent ha

Dynamic Governance of Multi-LLM Agent Systems for Collaborative Conversational Outcomes

SafetyDGX agent

arXiv:2608.11207v1 Announce Type: new Abstract: When two LLM agents with structurally opposed objectives interact across multiple turns, the absence of a shared goal function produces not competition

Learning from Online User Feedback for Shopping Agents

SafetyDGX agent

arXiv:2608.11604v1 Announce Type: new Abstract: Large language model-based shopping agents are increasingly deployed in real-world e-commerce platforms, generating massive amounts of user interaction

Multi-Agent Embodied Autonomous Driving: From V2X Information Exchange to Shared World Models

SafetyDGX agent

arXiv:2606.13840v2 Announce Type: replace-cross Abstract: Autonomous driving is shifting from isolated vehicle intelligence toward multi-agent embodied systems that share perception, infer intent, and

ToolHazard: Scaling Adversarial Environments for Security Evaluation and Alignment of LLM-based Agents

SafetyDGX agent

arXiv:2608.11878v1 Announce Type: cross Abstract: Large language model (LLM) agents integrated with external tools are vulnerable to indirect prompt injections embedded in environmental states. Howeve

Towards a Formal Definition of Agent Memory: Basis, Span, Optimality, and the Sequential Memory Problem

SafetyDGX agent

arXiv:2608.11654v1 Announce Type: new Abstract: Despite the wide deployment of memory in large-model agents, there is no unified formal account of what a memory is or when it is optimal. This paper ta

12 Aug 2026

GeoForge: Non-Parametric Self-Evolving Agents for Earth-Observation Reasoning

Model ReleasesDGX agent

arXiv:2608.10494v1 Announce Type: new Abstract: Earth observation (EO) agents construct scientifically valid tool workflows and ground their conclusions in current geospatial evidence. This is challen

InSight-doc: Agentic Visual Perception for Long-Document Understanding

Model ReleasesDGX agent

arXiv:2608.10628v1 Announce Type: cross Abstract: Long-document understanding often requires reasoning over many visually rich pages, making inference costly and prone to context rot. In this work, we

← Previous
1…8990919293…300
Next →