AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,959 results
24 Apr 2026

Try out Devin with the GPT-5.5 Agent Preview today: https://devin.ai

Model ReleasesDGX agent

Cognition AI announced an early preview opportunity for Devin integrated with GPT-5.5 Agent capabilities, inviting users to test the combination at devin.ai. This likely represents an update to Devin'

23 Apr 2026

SkillGraph: Graph Foundation Priors for LLM Agent Tool Sequence Recommendation

Model ReleasesDGX agent

arXiv:2604.19793v1 Announce Type: new Abstract: LLM agents must select tools from large API libraries and order them correctly. Existing methods use semantic similarity for both retrieval and ordering

Supplement Generation Training for Enhancing Agentic Task Performance

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases
DGX agent

arXiv:2604.20727v1 Announce Type: cross Abstract: Training large foundation models for agentic tasks is increasingly impractical due to the high computational costs, long iteration cycles, and rapid o

You’re about to feel the AI money squeeze

AgentsDGX agent

Earlier this month, millions of OpenClaw users woke up to a sweeping mandate: The viral AI agent tool, which this year took the worldwide tech industry by storm, had been severely restricted by Anthro

22 Apr 2026

A-MAR: Agent-based Multimodal Art Retrieval for Fine-Grained Artwork Understanding

Model ReleasesDGX agent

arXiv:2604.19689v1 Announce Type: new Abstract: Understanding artworks requires multi-step reasoning over visual content and cultural, historical, and stylistic context. While recent multimodal large

Excited to gather a group of technical builders Tuesday May 12 at @cognition's beautiful SF rooftop Join us for a Q&A with @ScottWu46 & Cogn…

AgentsDGX agent

Excited to gather a group of technical builders Tuesday May 12 at @cognition's beautiful SF rooftop Join us for a Q&A with @ScottWu46 & Cognition employees on scaling up AI agents within F500 companie

Google puts Gemini Enterprise at the heart of the new agentic taskforce for enterprise automation

Model ReleasesDGX agent

Google Cloud is on a mission to accelerate the adoption of artificial intelligence agents across enterprise computing environments, paving the way for a new era where AI can automate many of the most

M^{2}GRPO: Mamba-based Multi-Agent Group Relative Policy Optimization for Biomimetic Underwater Robots Pursuit

SafetyDGX agent

arXiv:2604.19404v1 Announce Type: cross Abstract: Traditional policy learning methods in cooperative pursuit face fundamental challenges in biomimetic underwater robots, where long-horizon decision ma

Mind the (DH) Gap! A Contrast in Risky Choices Between Reasoning and Conversational LLMs

AgentsDGX agent

arXiv:2602.15173v2 Announce Type: replace Abstract: The use of large language models either as decision support systems, or in agentic workflows, is rapidly transforming the digital ecosystem. However

Sentipolis: Emotion-Aware Agents for Social Simulations

ResearchDGX agent

arXiv:2601.18027v2 Announce Type: replace Abstract: LLM agents are increasingly used for social simulation, yet emotion is often treated as a transient cue, causing emotional amnesia and weak long-hor

Workspace agents can work across tools—pulling context from docs, email, chats, code, and systems, and taking approved actions like updating…

Model ReleasesDGX agent

Workspace agents can work across tools—pulling context from docs, email, chats, code, and systems, and taking approved actions like updating @Linear issues, creating docs, or sending messages. In @Sla

21 Apr 2026

CAPO: Counterfactual Credit Assignment in Sequential Cooperative Teams

SafetyDGX agent

arXiv:2604.17693v1 Announce Type: new Abstract: In cooperative teams where agents act in a fixed order and share a single team reward, it is hard to know how much each agent contributed, and harder st

Dual-Anchoring: Addressing State Drift in Vision-Language Navigation

AgentsDGX agent

arXiv:2604.17473v1 Announce Type: new Abstract: Vision-Language Navigation(VLN) requires an agent to navigate through 3D environments by following natural language instructions. While recent Video Lar

Harness as an Asset: Enforcing Determinism via the Convergent AI Agent Framework (CAAF)

Model ReleasesDGX agent

arXiv:2604.17025v1 Announce Type: cross Abstract: Large Language Models (LLMs) produce a controllability gap in safety-critical engineering: even low rates of undetected constraint violations render a

Nothing better than sprinting on a day-0 release alongside partners that run just as fast. Thank you @sarahmsachs and the @NotionHQ crew. Ex…

AgentsDGX agent

Nothing better than sprinting on a day-0 release alongside partners that run just as fast. Thank you @sarahmsachs and the @NotionHQ crew. Excited to see the day 1 reactions continue to roll in today.

Penny Wise, Pixel Foolish: Bypassing Price Constraints in Multimodal Agents via Visual Adversarial Perturbations

Model ReleasesDGX agent

arXiv:2604.16515v1 Announce Type: new Abstract: The rapid proliferation of Multimodal Large Language Models (MLLMs) has enabled mobile agents to execute high-stakes financial transactions, but their a

Temporal Representations for Exploration: Learning Complex Exploratory Behavior without Extrinsic Rewards

AgentsDGX agent

arXiv:2603.02008v2 Announce Type: replace Abstract: Effective exploration in reinforcement learning requires not only tracking where an agent has been, but also understanding how the agent perceives a

20 Apr 2026

🚀 Introducing Qwen3.6-Max-Preview, an early preview of our next flagship model Highlights: ⚡️ Improved agentic coding capability over Qwen3…

Model ReleasesDGX agent

🚀 Introducing Qwen3.6-Max-Preview, an early preview of our next flagship model Highlights: ⚡️ Improved agentic coding capability over Qwen3.6-Plus 📖 Stronger world knowledge and instruction following

Kimi K2.6 just entered the chat. > beats gpt 5.4 and opus 4.6 on coding. > open source. open weight. > 3x-5x cheaper. > really good for peop…

AgentsDGX agent

Kimi K2.6 just entered the chat. > beats gpt 5.4 and opus 4.6 on coding. > open source. open weight. > 3x-5x cheaper. > really good for people running agents. > insane at running long tasks. > we're t

MemEvoBench: Benchmarking Memory MisEvolution in LLM Agents

Model ReleasesDGX agent

arXiv:2604.15774v1 Announce Type: new Abstract: Equipping Large Language Models (LLMs) with persistent memory enhances interaction continuity and personalization but introduces new safety risks. Speci

Preregistered Belief Revision Contracts

SafetyDGX agent

arXiv:2604.15558v1 Announce Type: new Abstract: Deliberative multi-agent systems allow agents to exchange messages and revise beliefs over time. While this interaction is meant to improve performance,

19 Apr 2026

LLM Artifacts Connected to @karpathy's LLM Knowledge base idea, I've been building out a fun way to generate dynamic artifacts from these kn…

AgentsDGX agent

LLM Artifacts Connected to @karpathy's LLM Knowledge base idea, I've been building out a fun way to generate dynamic artifacts from these knowledge bases with the goal of discovering and revealing mea

releasing qwen3.6 35b a3b carnice edition (inspired by @kaiostephens) to enjoy an hermes agent tailored finetuned on unified memory hardware…

Model ReleasesDGX agent

releasing qwen3.6 35b a3b carnice edition (inspired by @kaiostephens) to enjoy an hermes agent tailored finetuned on unified memory hardware! safetensors and gguf versions (q8, 6,5 and 4) available! f

18 Apr 2026

Salesforce announces Headless 360, an initiative that will give AI agents access to Salesforce's platform capabilities through APIs, MCP tools or CLI commands (Michael Nuñez/VentureBeat)

Model ReleasesDGX agent

Michael Nuñez / VentureBeat: Salesforce announces Headless 360, an initiative that will give AI agents access to Salesforce's platform capabilities through APIs, MCP tools or CLI commands — Salesforce

17 Apr 2026

CCCE: A Continuous Code Calibration Engine for Autonomous Enterprise Codebase Maintenance via Knowledge Graph Traversal and Adaptive Decision Gating

AgentsDGX agent

arXiv:2604.13102v1 Announce Type: cross Abstract: Enterprise software organizations face an escalating challenge in maintaining the integrity, security, and freshness of codebases that span hundreds o

connect, ingest, extract, enrich

AgentsDGX agent

This post likely discusses a four-step framework for data processing and AI workflows: connecting to data sources, ingesting raw data, extracting relevant information, and enriching it with additional

Don't Retrieve, Navigate: Distilling Enterprise Knowledge into Navigable Agent Skills for QA and RAG

Model ReleasesDGX agent

arXiv:2604.14572v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) grounds LLM responses in external evidence but treats the model as a passive consumer of search results: it never

EchoAgent: Towards Reliable Echocardiography Interpretation with 'Eyes','Hands' and 'Minds'

AgentsDGX agent

arXiv:2604.05541v2 Announce Type: replace Abstract: Reliable interpretation of echocardiography (Echo) is crucial for assessing cardiac function, which demands clinicians to synchronously orchestrate

From Risk to Rescue: An Agentic Survival Analysis Framework for Liquidation Prevention

SafetyDGX agent

arXiv:2604.14583v1 Announce Type: new Abstract: Decentralized Finance (DeFi) lending protocols like Aave v3 rely on over-collateralization to secure loans, yet users frequently face liquidation due to

🆕 Harness Engineering: How to Build Software When Humans Steer, Agents Execute https://www.youtube.com/watch?v=am_oeAoUhew @_lopopolo is on…

TutorialsDGX agent

🆕 Harness Engineering: How to Build Software When Humans Steer, Agents Execute https://www.youtube.com/watch?v=am_oeAoUhew @_lopopolo is one of the emerging class of token billionaires at @OpenAI, and

Let's talk content faithfulness. Four days ago, we launched ParseBench, the first document OCR benchmark for AI agents. Its most fundamental…

Model ReleasesDGX agent

Let's talk content faithfulness. Four days ago, we launched ParseBench, the first document OCR benchmark for AI agents. Its most fundamental metric asks: did the parser capture all the text, in order,

NanoClaw partners with Vercel to deliver one-click approvals for AI agents working on sensitive tasks

IndustryDGX agent

NanoCo, the startup behind NanoClaw, a fast-growing alternative to OpenAI Group PBC’s OpenClaw project, said today it’s teaming up with Vercel Inc. and OneCLI to try to fix the “trust problem” holding

OpenAI ratchets up Codex’s agentic capabilities to rival Claude Code

Model ReleasesDGX agent

OpenAI Group PBC today announced a major revamp of its artificial intelligence coding tool Codex, giving it a number of new “agentic” capabilities that enable more complex task automation. The ChatGPT

16 Apr 2026

Anthropic says Opus 4.7 hits 80.6% on Document Reasoning — up from 57.1%. But 'reasoning about documents' ≠ 'parsing documents for agents.' …

Model ReleasesDGX agent

Anthropic says Opus 4.7 hits 80.6% on Document Reasoning — up from 57.1%. But 'reasoning about documents' ≠ 'parsing documents for agents.' We ran it on ParseBench. → Charts: 13.5% → 55.8% (+42.3) — h

Hierarchical Reinforcement Learning with Augmented Step-Level Transitions for LLM Agents

ResearchDGX agent

arXiv:2604.05808v2 Announce Type: replace-cross Abstract: Large language model (LLM) agents have demonstrated strong capabilities in complex interactive decision-making tasks. However, existing LLM ag

Qwen 3.6 is here, and open-source! Run it locally with improved agentic coding capabilities. Try it with Claude Code: ollama launch claude -…

Model ReleasesDGX agent

Qwen 3.6 is here, and open-source! Run it locally with improved agentic coding capabilities. Try it with Claude Code: ollama launch claude --model qwen3.6 Try it with OpenClaw: ollama launch openclaw

SemiFA: An Agentic Multi-Modal Framework for Autonomous Semiconductor Failure Analysis Report Generation

HardwareDGX agent

arXiv:2604.13236v1 Announce Type: new Abstract: Semiconductor failure analysis (FA) requires engineers to examine inspection images, correlate equipment telemetry, consult historical defect records, a

The cognitive companion: a lightweight parallel monitoring architecture for detecting and recovering from reasoning degradation in LLM agents

Model ReleasesDGX agent

arXiv:2604.13759v1 Announce Type: cross Abstract: Large language model (LLM) agents on multi-step tasks suffer reasoning degradation, looping, drift, stuck states, at rates up to 30% on hard tasks. Cu

this is fun. been trying eeg headsets for 10 yrs, the hat makes so much sense :)

AgentsDGX agent

this is fun. been trying eeg headsets for 10 yrs, the hat makes so much sense :) you can now control things with your brain. literally. we're building the most wearable BCI on the planet, with @sabica

ToolOmni: Enabling Open-World Tool Use via Agentic learning with Proactive Retrieval and Grounded Execution

Model ReleasesDGX agent

arXiv:2604.13787v1 Announce Type: new Abstract: Large Language Models (LLMs) enhance their problem-solving capability by utilizing external tools. However, in open-world scenarios with massive and evo

15 Apr 2026

A hierarchical spatial-aware algorithm with efficient reinforcement learning for human-robot task planning and allocation in production

AgentsDGX agent

arXiv:2604.12669v1 Announce Type: new Abstract: In advanced manufacturing systems, humans and robots collaborate to conduct the production process. Effective task planning and allocation (TPA) is cruc

Emergent launches Wingman: a personal AI agent for everyone

Model ReleasesDGX agent

Emergent Labs Inc., a vibe coding platform for building production-ready software, today announced the launch of Wingman, a personal, autonomous artificial intelligence agent that helps people manage

Local-Splitter: A Measurement Study of Seven Tactics for Reducing Cloud LLM Token Usage on Coding-Agent Workloads

Local AiDGX agent

arXiv:2604.12301v1 Announce Type: cross Abstract: We present a systematic measurement study of seven tactics for reducing cloud LLM token usage when a small local model can act as a triage layer in fr

Open-source is the solution to cyber-security because with new AI capabilities, all of the open-source repos will be inspected and patched 1…

AgentsDGX agent

Open-source is the solution to cyber-security because with new AI capabilities, all of the open-source repos will be inspected and patched 100x faster/better than any closed-source system! Let's take

QuarkMedSearch: A Long-Horizon Deep Search Agent for Exploring Medical Intelligence

Model ReleasesDGX agent

arXiv:2604.12867v1 Announce Type: new Abstract: As agentic foundation models continue to evolve, how to further improve their performance in vertical domains has become an important challenge. To this

Rippling AI was the most successful launch we've ever done. On the heels of this launch, Rippling's revenue is now growing 78% YoY (at ARR o…

AgentsDGX agent

Rippling AI was the most successful launch we've ever done. On the heels of this launch, Rippling's revenue is now growing 78% YoY (at ARR over $1 Billion). And this growth rate has now increased, eve

Toward Efficient and Robust Behavior Models for Multi-Agent Driving Simulation

Local AiDGX agent

arXiv:2512.05812v5 Announce Type: replace-cross Abstract: Scalable multi-agent driving simulation requires behavior models that are both realistic and computationally efficient. We address this by opt

When Reasoning Models Hurt Behavioral Simulation: A Solver-Sampler Mismatch in Multi-Agent LLM Negotiation

Model ReleasesDGX agent

arXiv:2604.11840v1 Announce Type: new Abstract: Large language models are increasingly used as agents in social, economic, and policy simulations. A common assumption is that stronger reasoning should

which way, modern man?

AgentsDGX agent

Harrison Chase, the creator of LangChain, shares reflections or opinions on directions for modern AI development, tooling, or the broader technology landscape. The post likely explores strategic or ph

14 Apr 2026

Agentic Exploration of PDE Spaces using Latent Foundation Models for Parameterized Simulations

Model ReleasesDGX agent

arXiv:2604.09584v1 Announce Type: new Abstract: Flow physics and more broadly physical phenomena governed by partial differential equations (PDEs), are inherently continuous, high-dimensional and ofte

ClawBench: Can AI Agents Complete Everyday Online Tasks? 153 tasks, 144 live websites, best model at 33.3% [R]

ResearchDGX agent

ClawBench is a benchmark of 153 everyday web tasks spanning 144 live platforms across 15 categories — from completing purchases and booking appointments to submitting job applications. Unlike existing

Cursor Automations now support event-based triggers for Sentry. Set up agents that automatically respond to new issues, investigate root cau…

ToolsDGX agent

Cursor Automations now support event-based triggers for Sentry. Set up agents that automatically respond to new issues, investigate root causes, open PRs with fixes, and post summaries to Slack. Media

FossID launches Agentic SCA to bring real-time compliance to AI-driven code development

Model ReleasesDGX agent

Software supply chain solutions company FossID AB today announced the launch of Agentic SCA, a new technology layer for software composition analysts that allows for real-time compliance and intellige

From Perception to Autonomous Computational Modeling: A Multi-Agent Approach

Model ReleasesDGX agent

arXiv:2604.06788v2 Announce Type: replace-cross Abstract: We present a solver-agnostic framework in which coordinated large language model (LLM) agents autonomously execute the complete computational

From Pixels to Digital Agents: An Empirical Study on the Taxonomy and Technological Trends of Reinforcement Learning Environments

ResearchDGX agent

arXiv:2603.23964v2 Announce Type: replace Abstract: The remarkable progress of reinforcement learning (RL) is intrinsically tied to the environments used to train and evaluate artificial agents. Movin

HearthNet: Edge Multi-Agent Orchestration for Smart Homes

Model ReleasesDGX agent

arXiv:2604.09618v1 Announce Type: cross Abstract: Smart-home users increasingly want to control their homes in natural language rather than assemble rules, dashboards, and API integrations by hand. At

Hunt Globally: Wide Search AI Agents for Drug Asset Scouting in Investing, Business Development, and Competitive Intelligence

Model ReleasesDGX agent

arXiv:2602.15019v3 Announce Type: replace Abstract: Bio-pharmaceutical innovation has shifted: many new drug assets now originate outside the United States and are disclosed primarily via regional, no

Infusing Theory of Mind into Socially Intelligent LLM Agents

Model ReleasesDGX agent

arXiv:2509.22887v2 Announce Type: replace Abstract: Theory of Mind (ToM)-an understanding of the mental states of others-is a key aspect of human social intelligence, yet, chatbots and LLM-based socia

MERMAID: Memory-Enhanced Retrieval and Reasoning with Multi-Agent Iterative Knowledge Grounding for Veracity Assessment

Model ReleasesDGX agent

arXiv:2601.22361v2 Announce Type: replace-cross Abstract: Assessing the veracity of online content has become increasingly critical. Large language models (LLMs) have recently enabled substantial prog

Mobile GUI Agent Privacy Personalization with Trajectory Induced Preference Optimization

Model ReleasesDGX agent

arXiv:2604.11259v1 Announce Type: new Abstract: Mobile GUI agents powered by Multimodal Large Language Models (MLLMs) can execute complex tasks on mobile devices. Despite this progress, most existing

← Previous
1…122123124125126…300
Next →