AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,958 results
5 May 2026

DigitalOcean raises 2026 and 2027 revenue outlook after AI-driven earnings beat

AgentsDGX agent

Shares of DigitalOcean Holdings Inc. rocketed more than 40% today after the developer-oriented cloud infrastructure provider topped Wall Street targets in its fiscal 2026 first quarter. It also lifted

Forager: a lightweight testbed for continual learning with partial observability in RL

TutorialsDGX agent

arXiv:2605.01131v1 Announce Type: new Abstract: In continual reinforcement learning (CRL), good performance requires never-ending learning, acting, and exploration in a big, partially observable world

Hallucinations Undermine Trust; Metacognition is a Way Forward

AgentsDGX agent

arXiv:2605.01428v1 Announce Type: new Abstract: Despite significant strides in factual reliability, errors -- often termed hallucinations -- remain a major concern for generative AI, especially as LLM

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

IBM charts AI operating model to move enterprises beyond experimentation

AgentsDGX agent

IBM Corp. will use its Think 2026 conference today to outline a broad expansion of its enterprise artificial intelligence portfolio, positioning a new “AI operating model” as the next stage in its cus

Large Language Models for Multi-Robot Systems: A Survey

AgentsDGX agent

arXiv:2502.03814v5 Announce Type: replace Abstract: The rapid advancement of Large Language Models (LLMs) has opened new possibilities in Multi-Robot Systems (MRS), enabling enhanced communication, ta

LLM Ghostbusters: Surgical Hallucination Suppression via Adaptive Unlearning

AgentsDGX agent

arXiv:2605.01047v1 Announce Type: cross Abstract: Hallucinations, outputs that sound plausible but are factually incorrect, remain an open challenge for deployed LLMs. In code generation, models frequ

Lost in the Tower of Babel: The Adverse Effects of Incidental Multilingualism in LLMs

AgentsDGX agent

arXiv:2605.01224v1 Announce Type: new Abstract: This paper argues that contemporary multilingual NLP has converged on a fragile and misleading paradigm of incidental multilingualism. Today's LLMs appe

Outbidding and Outbluffing Elite Humans: Mastering Liar's Poker via Self-Play and Reinforcement Learning

AgentsDGX agent

arXiv:2511.03724v3 Announce Type: replace Abstract: AI researchers have long focused on poker-like games as a testbed for environments characterized by multi-player dynamics, imperfect information, an

Reinforcement Learning from Compiler and Language Server Feedback

SafetyDGX agent

arXiv:2510.22907v2 Announce Type: replace Abstract: Coding agents fail when text-level guesses outrun program facts: they hallucinate APIs, drift to the wrong symbol, and apply edits without evidence

Safe Planning in Interactive Environments via Iterative Policy Updates and Adversarially Robust Conformal Prediction

SafetyDGX agent

arXiv:2511.10586v2 Announce Type: replace-cross Abstract: Safe planning of an autonomous agent in interactive environments -- such as the control of a self-driving vehicle among pedestrians -- poses a

ServiceNow bids to become the control tower for enterprise AI

AgentsDGX agent

ServiceNow Inc. today unveiled a broad expansion of its artificial intelligence platform, stressing governance, security and autonomous execution as foundational requirements for enterprise AI adoptio

Verbal-R3: Verbal Reranker as the Missing Bridge between Retrieval and Reasoning

AgentsDGX agent

arXiv:2605.01399v1 Announce Type: new Abstract: The conventional Retrieval-Augmented Generation (RAG) paradigm of injecting raw retrieved texts into the Large Language Model (LLM)'s context often resu

You can now build your own customized financial apps on @Replit, using your real financial data, securely connected through @Plaid. Whatever…

AgentsDGX agent

You can now build your own customized financial apps on @Replit, using your real financial data, securely connected through @Plaid. Whatever finance app you've always wished existed, you can build it.

4 May 2026

BOLT: Online Lightweight Adaptation for Preparation-Free Heterogeneous Cooperative Perception

SafetyDGX agent

arXiv:2605.00405v1 Announce Type: new Abstract: Most existing heterogeneous cooperative perception methods depend on prior preparation like offline joint training or tailored collaborator-model adapta

Causality-enhanced Decision-Making for Autonomous Mobile Robots in Dynamic Environments

AgentsDGX agent

arXiv:2504.11901v5 Announce Type: replace Abstract: The growing integration of robots in shared environments-such as warehouses, shopping centres, and hospitals-demands a deep understanding of the und

High-Probability Convergence in Decentralized Stochastic Optimization with Gradient Tracking

Model ReleasesDGX agent

arXiv:2605.00281v1 Announce Type: new Abstract: We study high-probability (HP) convergence guarantees in decentralized stochastic optimization, where multiple agents collaborate to jointly train a mod

I have a ~200mb sqlite DB of the entire history of SF criminal court cases for the last 4 years that I want to make publicly accessible for …

AgentsDGX agent

I have a ~200mb sqlite DB of the entire history of SF criminal court cases for the last 4 years that I want to make publicly accessible for anyone to query with AI. It includes roughly 77k cases, 319k

Long-Horizon Model-Based Offline Reinforcement Learning Without Explicit Conservatism

AgentsDGX agent

arXiv:2512.04341v3 Announce Type: replace Abstract: Popular offline reinforcement learning (RL) methods rely on explicit conservatism, penalizing out-of-dataset actions or restricting rollout horizons

NEW paper from Sakana AI (ICLR 2026). A 7B Conductor model just hit SOTA on GPQA-Diamond and LiveCodeBench by orchestrating other LLMs inste…

SafetyDGX agent

NEW paper from Sakana AI (ICLR 2026). A 7B Conductor model just hit SOTA on GPQA-Diamond and LiveCodeBench by orchestrating other LLMs instead of solving problems itself. (great paper! bookmark it!) T

One of the great parts of nous portal is stuff like this

AgentsDGX agent

One of the great parts of nous portal is stuff like this Trinity-Large-Thinking, @arcee_ai's latest model, is now free on Nous Portal for the next week Sign up for Nous Portal to use it in your Hermes

PORTool: Importance-Aware Policy Optimization with Rewarded Tree for Multi-Tool-Integrated Reasoning

SafetyDGX agent

arXiv:2510.26020v2 Announce Type: replace Abstract: Multi-tool-integrated reasoning enables LLM-empowered tool-use agents to solve complex tasks by interleaving natural-language reasoning with calls t

This is probably better messaging than just “own your harness.” Yes, open models will need custom harnesses, but that is a means to an end. …

AgentsDGX agent

This is probably better messaging than just “own your harness.” Yes, open models will need custom harnesses, but that is a means to an end. The end is utilizing models without being handcuffed to Anth

ToolGrad: Efficient Tool-use Dataset Generation with Textual 'Gradients'

AgentsDGX agent

arXiv:2508.04086v2 Announce Type: replace Abstract: Prior work synthesizes tool-use LLM datasets by first generating a user query, followed by complex tool-use annotations like depth-first search (DFS

3 May 2026

i actually don't want this 'but you don't review compiler output either' meme to die. it's the perfect signal for being immediately able to …

AgentsDGX agent

i actually don't want this 'but you don't review compiler output either' meme to die. it's the perfect signal for being immediately able to ignore someone in this space. Interesting article on treatin

2 May 2026

Claude Opus 4.7 just implemented an AlphaZero-style self-play pipeline from scratch. It did this on consumer hardware in three hours, then b…

Model ReleasesDGX agent

Claude Opus 4.7 just implemented an AlphaZero-style self-play pipeline from scratch. It did this on consumer hardware in three hours, then beat the Pascal Pons solver 7 of 8 as first-mover on Connect

The year of the harness 🤝 Harnesses are our most direct layer to turn a model into a great product experience for users by shaping model be…

AgentsDGX agent

The year of the harness 🤝 Harnesses are our most direct layer to turn a model into a great product experience for users by shaping model behavior towards a goal This matters today because models have

1 May 2026

A Collective Variational Principle Unifying Bayesian Inference, Game Theory, and Thermodynamics

Local AiDGX agent

arXiv:2604.27942v1 Announce Type: new Abstract: Collective intelligence emerges across biological, physical, and artificial systems without central coordination, yet a unifying principle governing suc

Can AI Be a Good Peer Reviewer? A Survey of Peer Review Process, Evaluation, and the Future

AgentsDGX agent

arXiv:2604.27924v1 Announce Type: cross Abstract: Peer review is a multi-stage process involving reviews, rebuttals, meta-reviews, final decisions, and subsequent manuscript revisions. Recent advances

Can AI be a moral victim? The role of moral patiency and ownership perceptions in ethical judgments of using AI-generated content

AgentsDGX agent

arXiv:2604.26956v1 Announce Type: cross Abstract: The growing use of generative AI raises ethical concerns about authorship and plagiarism. This study examines how people judge the reuse of AI-generat

Dreaming Across Towns: Semantic Rollout and Town-Adversarial Regularization for Zero-Shot Held-Out-Town Fixed-Route Driving in CARLA

SafetyDGX agent

arXiv:2604.27994v1 Announce Type: new Abstract: Learned driving agents often degrade when deployed in unseen environments. This paper studies a deliberately bounded instance of that problem in the CAR

I have to go out of town for a funeral thru the weekend but I am leaving everyone with one new cool feature inspired by ralph loops and Code…

AgentsDGX agent

I have to go out of town for a funeral thru the weekend but I am leaving everyone with one new cool feature inspired by ralph loops and Codex's upcoming /goal feature. If you use /goal <prompt>, it wi

Interval Orders, Biorders and Credibility-limited Belief Revision

AgentsDGX agent

arXiv:2604.27156v1 Announce Type: new Abstract: Rational belief revision is commonly viewed as being based on a preference order between possible worlds, with the resulting new belief set being those

Let's Measure Information Step-by-Step: AI-Based Evaluation Beyond Vibes

AgentsDGX agent

arXiv:2508.05469v3 Announce Type: replace Abstract: We evaluate artificial intelligence (AI) systems without ground truth by exploiting a link between strategic gaming and information loss. Building o

OptimusKG: Unifying biomedical knowledge in a modern multimodal graph

AgentsDGX agent

arXiv:2604.27269v1 Announce Type: new Abstract: Biomedical knowledge graphs (KGs) are widely used in the life sciences, yet many are derived from unstructured documents and therefore lack schema-level

Randomized trial of an AI therapy chatbot on Mexican women found “improved mental health by 0.3 SD over 6 months with no evidence of an incr…

AgentsDGX agent

Randomized trial of an AI therapy chatbot on Mexican women found “improved mental health by 0.3 SD over 6 months with no evidence of an increase of severe cases; improved sleep, healthful behaviors, d

Receptionist robotics app built under 2 hours thanks to ml intern and reachy mini. Was fun again!

AgentsDGX agent

Receptionist robotics app built under 2 hours thanks to ml intern and reachy mini. Was fun again! Media I'm trying to build an office receptionist app for my reachy mini today with ml intern + @OpenAI

TRUST: A Framework for Decentralized AI Service v.0.1

SafetyDGX agent

arXiv:2604.27132v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) and Multi-Agent Systems (MAS) in high-stakes domains demand reliable verification, yet centralized approaches suffer four

30 Apr 2026

3D Generation for Embodied AI and Robotic Simulation: A Survey

AgentsDGX agent

arXiv:2604.26509v1 Announce Type: cross Abstract: Embodied AI and robotic systems increasingly depend on scalable, diverse, and physically grounded 3D content for simulation-based training and real-wo

Autonomous Knowledge Graph Exploration with Adaptive Breadth-Depth Retrieval

AgentsDGX agent

arXiv:2601.13969v2 Announce Type: replace Abstract: Retrieving evidence for language model queries from knowledge graphs requires balancing broad search across the graph with multi-hop traversal to fo

Benchmarks for Trajectory Safety Evaluation and Diagnosis in OpenClaw and Codex: ATBench-Claw and ATBench-Codex

Model ReleasesDGX agent

arXiv:2604.14858v2 Announce Type: replace Abstract: As agent systems move into increasingly diverse execution settings, trajectory-level safety evaluation and diagnosis require benchmarks that evolve

Lyapunov-Guided Self-Alignment: Test-Time Adaptation for Offline Safe Reinforcement Learning

Model ReleasesDGX agent

arXiv:2604.26516v1 Announce Type: cross Abstract: Offline reinforcement learning (RL) agents often fail when deployed, as the gap between training datasets and real environments leads to unsafe behavi

MappingEvolve: LLM-Driven Code Evolution for Technology Mapping

AgentsDGX agent

arXiv:2604.26591v1 Announce Type: cross Abstract: Technology mapping is a critical yet challenging stage in logic synthesis. While Large Language Models (LLMs) have been applied to generate optimizati

Quoting Andrew Kelley

AgentsDGX agent

It's a common misconception that we can't tell who is using LLM and who is not. I'm sure we didn't catch 100% of LLM-assisted PRs over the past few months, but the kind of mistakes humans make are fun

29 Apr 2026

big theme of 2026 - cost of closed models is too high! really excited to make deepagents work exceptionally well with OSS models

AgentsDGX agent

big theme of 2026 - cost of closed models is too high! really excited to make deepagents work exceptionally well with OSS models Switched out Sonnet 4.6 for GLM 5.1 through @FireworksAI_HQ while doing

great deep dive into how we get deepagents to work well with different families of models

AgentsDGX agent

This post likely discusses best practices and techniques for optimizing DeepAgent performance across diverse model architectures and families, covering strategies for compatibility and effective integ

Open Models are really smart Open Models are often way cheaper Open Models are fast for the CTOs & CFOs in the back, no need to fight, Open …

AgentsDGX agent

Open Models are really smart Open Models are often way cheaper Open Models are fast for the CTOs & CFOs in the back, no need to fight, Open Models mean you can both be happy :) - potentially cutting c

28 Apr 2026

amazing! it’s like talking to an ai from the past

AgentsDGX agent

amazing! it’s like talking to an ai from the past Announcing Talkie: a new, open-weight historical LLM! We trained and finetuned a 13B model on a newly-curated dataset of only pre-1930 data. Try it be

And if you have your own VT-100, check out Devin for Terminal to get started: https://x.com/cognition/status/2048821234281181302

AgentsDGX agent

And if you have your own VT-100, check out Devin for Terminal to get started: https://x.com/cognition/status/2048821234281181302 The terminal hasn’t changed much since the 1970s. What you do with it h

BitRL: Reinforcement Learning with 1-bit Quantized Language Models for Resource-Constrained Edge Deployment

Model ReleasesDGX agent

arXiv:2604.24273v1 Announce Type: new Abstract: The deployment of intelligent reinforcement learning (RL) agents on resource-constrained edge devices remains a fundamental challenge due to the substan

Don't Make the LLM Read the Graph: Make the Graph Think

Model ReleasesDGX agent

arXiv:2604.23057v1 Announce Type: new Abstract: We investigate whether explicit belief graphs improve LLM performance in cooperative multi-agent reasoning. Through 3,000+ controlled trials across four

ESIA: An Energy-Based Spatiotemporal Interaction-Aware Framework for Pedestrian Intention Prediction

AgentsDGX agent

arXiv:2604.23728v1 Announce Type: cross Abstract: Recent advances in autonomous driving have motivated research on pedestrian intention prediction, which aims to infer future crossing decisions and ac

// From Skill Text to Skill Structure // One of the more practical skill papers I've seen this month. SKILL.md files entangle invocation int…

TutorialsDGX agent

// From Skill Text to Skill Structure // One of the more practical skill papers I've seen this month. SKILL.md files entangle invocation interface, execution flow, and tool/resource side effects in on

Google AI 인프라의 미래: 에이전틱 시대를 위한 확장

Model ReleasesDGX agent

* 본 아티클의 원문은 2026년 4월 23일 Google Cloud 블로그(영문)에 게재되었습니다. AI는 질문에 답하는 수준을 넘어 추론하고 행동하는 단계로 진화하고 있습니다. 오늘날의 에이전틱 시대(agentic era)를 선도하고자 하는 기업에는 이러한 새로운 요구사항에 맞춰 설계되고 최적화된 컴퓨팅 인프라가 필요합니다. 오늘 Google Cloud

I’m going to build a bunch of deepagents examples (using deepagent deploy) over the next few days What examples would people want to see?

ApplicationsDGX agent

Harrison Chase announced plans to create multiple deepagents examples using the deepagent deploy tool and solicited community input on which examples would be most valuable. This post likely seeks fee

LEGO: An LLM Skill-Based Front-End Design Generation Platform

Model ReleasesDGX agent

arXiv:2604.23355v1 Announce Type: new Abstract: Existing LLM-based EDA agents are often isolated task-specific systems. This leads to repeated engineering effort and limited reuse of successful design

Leveraging Human Feedback for Semantically-Relevant Skill Discovery

ResearchDGX agent

arXiv:2604.24127v1 Announce Type: cross Abstract: Unsupervised skill discovery in reinforcement learning aims to intrinsically motivate agents to discover diverse and useful behaviours. However, uncon

NeuroClaw Technical Report

Model ReleasesDGX agent

arXiv:2604.24696v1 Announce Type: new Abstract: Agentic artificial intelligence systems promise to accelerate scientific workflows, but neuroimaging poses unique challenges: heterogeneous modalities (

new in ml-intern: you can now actually see what's going on inside added native metric logging + trackio integration. every training run the …

Model ReleasesDGX agent

new in ml-intern: you can now actually see what's going on inside added native metric logging + trackio integration. every training run the agent kicks off now has live curves you can watch in real ti

No Test Cases, No Problem: Distillation-Driven Code Generation for Scientific Workflows

Model ReleasesDGX agent

arXiv:2604.23106v1 Announce Type: cross Abstract: Existing multi-agent Large Language Model (LLM) frameworks for code generation typically use execution feedback and improve iteratively using Input/Ou

On the Convergence of Jacobian-Free Backpropagation for Optimal Control Problems with Implicit Hamiltonians

AgentsDGX agent

arXiv:2602.00921v2 Announce Type: replace-cross Abstract: Optimal feedback control with implicit Hamiltonians poses a fundamental challenge for learning-based value function methods due to the absence

← Previous
1…192193194195196…300
Next →