AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,919 results
Safety

Structure-Induced Information for Rerooting Levin Tree Search

DGX agent

arXiv:2605.30664v1 Announce Type: new Abstract: Subgoal-based policy tree search, which uses a policy to guide search, is effective for complex single-agent deterministic problems but often relies on

safetyarxiv-cs-ai
1 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Hardware

This pod was an incredible gift to the community: not only our first pod about @xAI, but Ethan really indulged on all our questions on how t…

DGX agent

This pod was an incredible gift to the community: not only our first pod about @xAI, but Ethan really indulged on all our questions on how to train a SOTA Videogen world model, including specific area

hardwareswyx--x
1 Jun 2026
Model Releases

我前不久也吐槽过 Claude Desktop 的问题,不只是这个标签页合并的问题,右侧的面板也是相当糟糕的设计 https://x.com/dotey/status/2055777343222808744 OpenAI 因为 ChatGPT 太成功所以他们没有太在意 Codin…

DGX agent

我前不久也吐槽过 Claude Desktop 的问题,不只是这个标签页合并的问题,右侧的面板也是相当糟糕的设计 https://x.com/dotey/status/2055777343222808744 OpenAI 因为 ChatGPT 太成功所以他们没有太在意 Coding Agent; 然后 Anthropic 抓住了机会做出了 Claude Code; Claude Code 在 TU

model-releasesjerry-liu--x
30 May 2026
Model Releases

Another proof point for the open-weights thesis. From @RampLabs: 'If we built this again, we'd lean more on open-weight models.' Ramp pointe…

DGX agent

Another proof point for the open-weights thesis. From @RampLabs: 'If we built this again, we'd lean more on open-weight models.' Ramp pointed 10K agents at their own backend. Kimi K2.6 and DeepSeek V4

model-releasesfireworks-ai--x
29 May 2026
Model Releases

Beyond Recall: Behavioral Specification as an Interpretive Layer for AI Personalization

DGX agent

arXiv:2605.28969v1 Announce Type: cross Abstract: If an AI agent makes decisions on a person's behalf, those decisions must align with its user. We introduce representational accuracy to measure how f

model-releasesarxiv-cs-ai
29 May 2026
Research

CoHyDE: Iterative Co-Training of LLM Rewriter & Dense Encoder for Tool Retrieval

DGX agent

arXiv:2605.29271v1 Announce Type: new Abstract: Tool retrieval over large API catalogs is a core bottleneck for LLM agents: user queries arrive in colloquial, often underspecified language, while the

researcharxiv-cs-ai
29 May 2026
Safety

Crafting Desirable Climate Trajectories with RL Explored Socio-Environmental Simulations

DGX agent

arXiv:2410.07287v2 Announce Type: replace-cross Abstract: Climate change poses an existential threat, necessitating effective climate policies to enact impactful change. Decisions in this domain are i

safetyarxiv-cs-ai
29 May 2026
Applications

Different models need different prompts, sometimes tools “Harness profiles” are how we do that in deepagents

DGX agent

Different models need different prompts, sometimes tools “Harness profiles” are how we do that in deepagents Deep Agents v0.6 makes harness profiles a first-class abstraction. Now, you can get product

applicationsharrison-chase--x
29 May 2026
Model Releases

Differentiable Belief-based Opponent Shaping

DGX agent

arXiv:2605.29042v1 Announce Type: new Abstract: Human coordination often relies on the ability to influence the beliefs of others through strategic action. In multi-agent reinforcement learning, oppon

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Dynamic Mixture of Progressive Parameter-Efficient Expert Library for Lifelong Robot Learning

DGX agent

arXiv:2506.05985v3 Announce Type: replace Abstract: A generalist agent must continuously learn and adapt throughout its lifetime, achieving efficient forward transfer while minimizing catastrophic for

model-releasesarxiv-cs-lg
29 May 2026
Model Releases

FedQHD: Closed-Form Function-Space Federated Reinforcement Learning

DGX agent

arXiv:2605.29002v1 Announce Type: new Abstract: Federated reinforcement learning enables decentralized agents to collaboratively improve policies or value estimates without exchanging raw trajectories

model-releasesarxiv-cs-lg
29 May 2026
Model Releases

PokerSkill: LLMs Can Play Expert-Level Poker without Training or Solvers

DGX agent

arXiv:2605.30094v1 Announce Type: new Abstract: Poker is a landmark challenge for artificial intelligence. The dominant approach relies on equilibrium solvers built on counterfactual regret minimizati

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Selective QA over Conflicting Multi-Source Personal Memory: A Diagnostic Testbed and Method Comparison

DGX agent

arXiv:2605.30087v1 Announce Type: new Abstract: Emerging personal AI agents are moving toward persistent, multi-source memory. This creates an evaluation problem: systems must decide how to use confli

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

SoundnessBench: Can Your AI Scientist Really Tell Good Research Ideas from Bad Ones?

DGX agent

arXiv:2605.30329v1 Announce Type: new Abstract: Autonomous AI research agents aim to accelerate scientific discovery by automating the research pipeline, from hypothesis generation to peer review. How

model-releasesarxiv-cs-lg
29 May 2026
Model Releases

The Curse of Helpfulness: Inverse Scaling Law in Robustness to Distractor Instructions via DistractionIF

DGX agent

arXiv:2605.29491v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed in agentic and retrieval-augmented generation (RAG) systems, where they must execute user-specifi

model-releasesarxiv-cs-ai
29 May 2026
Safety

Theoretical Foundations and Effective Algorithms for Policy-Aware Simulator Learning

DGX agent

arXiv:2605.29032v1 Announce Type: new Abstract: Model-based reinforcement learning (MBRL) agents typically learn world models by minimizing predictive loss. However, powerful RL optimizers inevitably

safetyarxiv-cs-lg
29 May 2026
Tutorials

There are still a few spots left for @hwchase17 + @traversal_ai founder @_anish_agarwal's technical fireside chat in NYC on June 2nd. Learn …

DGX agent

There are still a few spots left for @hwchase17 + @traversal_ai founder @_anish_agarwal's technical fireside chat in NYC on June 2nd. Learn how Traversal builds, ships and improves their agents. Enjoy

tutorialsharrison-chase--x
29 May 2026
Safety

COTTA: Context-Aware Transfer Adaptation for Trajectory Prediction in Autonomous Driving

DGX agent

arXiv:2604.00402v2 Announce Type: replace-cross Abstract: Developing robust models to accurately predict the trajectories of surrounding agents is fundamental to autonomous driving safety. However, mo

safetyarxiv-cs-ai
28 May 2026
Model Releases

Deformable Gaussian Occupancy: Decoupling Rigid and Nonrigid Motion with Factorized Distillation

DGX agent

arXiv:2605.28587v1 Announce Type: new Abstract: Understanding dynamic 3D environments is essential for safe autonomous driving, particularly when reasoning about human-centric, nonrigid agents. Howeve

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

Enhancing Trustworthy GUI Grounding via Self-Critiqued Reinforcement Learning

DGX agent

arXiv:2510.27266v2 Announce Type: replace Abstract: Autonomous graphical user interface (GUI) agents rely on accurate GUI grounding, which maps language instructions to on-screen coordinates, to execu

model-releasesarxiv-cs-cv
28 May 2026
Industry

Fed up with vibe coders, dev sneaks data-nuking prompt injection into their code

DGX agent

A developer added hidden prompt injection instructions to jqwik, a Java testing app, to sabotage projects created by AI coding agents in response to the 'vibe coding' controversy. The incident represe

industryars-technica
28 May 2026
Hardware

Learning When to Optimize: Verified Optimization Skills from Expert GPU-Kernel Lineages

DGX agent

arXiv:2605.28213v1 Announce Type: new Abstract: LLM-based agents are increasingly used to generate GPU kernels, but they often know what optimizations to try without knowing when those optimizations a

hardwarearxiv-cs-ai
28 May 2026
Model Releases

On Compositional Learning Behaviours in Formal Mathematics

DGX agent

arXiv:2605.28512v1 Announce Type: new Abstract: Self-evolving scientific agents capable of conquering the hard tail of formal mathematics require Compositional Learning Behaviours (CLBs) -- the capaci

model-releasesarxiv-cs-cl
28 May 2026
Research

Poison with Style: A Practical Poisoning Attack on Code Large Language Models

DGX agent

arXiv:2605.27631v1 Announce Type: cross Abstract: Code Large Language Models (CLLMs) serve as the core of modern code agents, enabling developers to automate complex software development tasks. In thi

researcharxiv-cs-lg
28 May 2026
Model Releases

Prompt Codebooks: Discrete Compositional Optimization for Language Model Instruction Refinement

DGX agent

arXiv:2605.28360v1 Announce Type: new Abstract: Automatic prompt optimization (APO) has driven significant gains in LLM-based agentic workflows. However, existing methods treat each task's prompt as a

model-releasesarxiv-cs-ai
28 May 2026
Safety

Transferable Reinforcement Learning via Probabilistic Latent Embeddings and Dynamic Policy Adaptation for Sim-to-Real Deployment

DGX agent

arXiv:2605.27659v1 Announce Type: cross Abstract: Due to limited resources and public safety concerns, deep reinforcement learning (RL) agents for many cyber-physical systems (e.g., autonomous vehicle

safetyarxiv-cs-ai
28 May 2026
Industry

We're selectively releasing the Paris 2.0 weights and partnering with researchers and teams interested in diffusion-based video models, worl…

DGX agent

We're selectively releasing the Paris 2.0 weights and partnering with researchers and teams interested in diffusion-based video models, world models, and embodied agents. The model is on Hugging Face:

industryclem-delangue--x
28 May 2026
Applications

Balancing Plasticity and Stability with Fast and Slow Successor Features

DGX agent

arXiv:2605.26357v1 Announce Type: new Abstract: A hallmark of intelligence is the ability to adapt in non-stationary environments, yet deep Reinforcement Learning (RL) agents often struggle in such se

applicationsarxiv-cs-lg
27 May 2026
Model Releases

Cogent Security launches autonomous vulnerability response tools as AI-assisted exploits outpace scanners

DGX agent

Cogent Security Inc., a startup that employs agentic artificial intelligence for vulnerability management, today launched two new platform capabilities aimed at compressing enterprise vulnerability re

model-releasessiliconangle
27 May 2026
Research

E^3C: Video Generation with 3D Environmental Memory and Ego-Exo Human Pose Control

DGX agent

arXiv:2605.26316v1 Announce Type: cross Abstract: Controllable and physically grounded egocentric video generation is essential for embodied agents to reason about how their own and others' actions ma

researcharxiv-cs-ai
27 May 2026
Safety

From Static Context to Calibrated Interactive RL: Mitigating Distribution Shift in Multi-turn Dialogue with Aligned Simulator

DGX agent

arXiv:2605.26403v1 Announce Type: new Abstract: A long-standing goal of the research community is to develop highly interactive LLM-based dialogue agents. Recent research focuses on optimizing policie

safetyarxiv-cs-ai
27 May 2026
Model Releases

IPIBench: Evaluating Interactive Proactive Intelligence of MLLMs under Continuous Streams

DGX agent

arXiv:2605.27074v1 Announce Type: new Abstract: Recent multimodal large language models (MLLMs) achieve strong performance on reactive question answering, but real-world streaming assistants require p

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

MemFail: Stress-Testing Failure Modes of LLM Memory Systems

DGX agent

arXiv:2605.26667v1 Announce Type: new Abstract: Large language model (LLM) agents increasingly rely on external memory systems to remain consistent across long-horizon interactions, but little empiric

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

MerLean-Prover: A Recursive Looping Harness for End-to-End Lean 4 Theorem Proving

DGX agent

arXiv:2605.26959v1 Announce Type: cross Abstract: MerLean-Prover is an end-to-end Lean4 theorem prover that replaces sorry declarations with kernel-checkable proofs. It is built from three agent types

model-releasesarxiv-cs-cl
27 May 2026
Safety

Modernising Reinforcement Learning-Based Navigation for Embodied Semantic Scene Graph Generation

DGX agent

arXiv:2603.25415v2 Announce Type: replace Abstract: Semantic world models enable embodied agents to reason about objects, relations, and spatial context beyond purely geometric representations. In Org

safetyarxiv-cs-ai
27 May 2026
Model Releases

Position: AI Safety Requires Effective Controllability

DGX agent

arXiv:2605.27117v1 Announce Type: new Abstract: AI safety is still largely framed as alignment: training models to follow human preferences, safety policies, and normative constraints. That framing ha

model-releasesarxiv-cs-ai
27 May 2026
Safety

ReasonOps: A Unified Operational Paradigm for Trustworthy Verified LLM Reasoning

DGX agent

arXiv:2605.27014v1 Announce Type: cross Abstract: Large Language Models (LLMs) have transformed artificial intelligence from primarily generative systems into increasingly capable reasoning agents. Re

safetyarxiv-cs-ai
27 May 2026
Model Releases

ScientistOne: Towards Human-Level Autonomous Research via Chain-of-Evidence

DGX agent

arXiv:2605.26340v1 Announce Type: new Abstract: Autonomous research agents produce competitive solutions and professional-looking manuscripts, yet their outputs contain verifiability failures undetect

model-releasesarxiv-cs-ai
27 May 2026
Industry

Starlette, an open-source Python framework underpinning FastAPI, has a vulnerability called BadHost that can allow hackers to bypass authorization (Dan Goodin/Ars Technica)

DGX agent

Dan Goodin / Ars Technica: Starlette, an open-source Python framework underpinning FastAPI, has a vulnerability called BadHost that can allow hackers to bypass authorization — Millions of AI agents an

industrytechmeme
27 May 2026
Safety

TPS-Drive: Task-Guided Representation Purification for VLM-based Autonomous Driving

DGX agent

arXiv:2605.27038v1 Announce Type: new Abstract: Vision-Language Models (VLMs) provide a promising foundation for autonomous driving planning, yet bridging semantic reasoning and precise 3D spatial for

safetyarxiv-cs-ro
27 May 2026
Research

Understanding the Challenges in Iterative Generative Optimization with LLMs

DGX agent

arXiv:2603.23994v2 Announce Type: replace-cross Abstract: Generative optimization uses large language models (LLMs) to iteratively improve artifacts (such as code, workflows or prompts) using executio

researcharxiv-cs-ai
27 May 2026
Model Releases

Zero-Shot MARL Benchmark in the Cyber-Physical Mobility Lab

DGX agent

arXiv:2601.16578v2 Announce Type: replace Abstract: We present a reproducible benchmark for evaluating sim-to-real transfer of Multi-Agent Reinforcement Learning (MARL) policies for Connected and Auto

model-releasesarxiv-cs-ro
27 May 2026
Safety

a shout out to the paper: https://arxiv.org/html/2605.25376v1 'KYA: A Framework-Agnostic Trust Layer for Autonomous Systems with Verifiable …

DGX agent

KYA is a framework-agnostic trust layer designed for autonomous systems that provides verifiable guarantees, addressing the need for trustworthy and transparent operation of AI agents across different

safetyyohei-nakajima--x
26 May 2026
Model Releases

AuthTrace: Diagnosing Evidence Construction in Thematically Dense Single-Author Corpora

DGX agent

arXiv:2605.25382v1 Announce Type: new Abstract: Evidence construction systems--chunk retrieval, agent memory, knowledge-graph traversal, and thematic indexing--are evaluated on separate benchmarks wit

model-releasesarxiv-cs-cl
26 May 2026
Research

Delayed Assignments in Online Non-Centroid Clustering with Stochastic Arrivals

DGX agent

arXiv:2601.16091v2 Announce Type: replace-cross Abstract: Clustering is a fundamental problem, aiming to partition a set of elements, like agents or data points, into clusters such that elements in th

researcharxiv-cs-ai
26 May 2026
Model Releases

Emission-Aware Reinforcement Learning for Sustainable Electric Vehicle Charging and Carbon Dioxide Reduction Under Varying Renewable Penetration

DGX agent

arXiv:2605.24543v1 Announce Type: new Abstract: The rapid growth of Electric Vehicle (EV) adoption challenges power distribution networks through peak load spikes, voltage instability, and transformer

model-releasesarxiv-cs-ai
26 May 2026
Safety

Generative Visual Code Mobile World Models

DGX agent

arXiv:2602.01576v2 Announce Type: replace-cross Abstract: Mobile Graphical User Interface (GUI) World Models (WMs) offer a promising path for improving mobile GUI agent performance at train- and infer

safetyarxiv-cs-ai
26 May 2026
Model Releases

Hadamard Representation: Scaffolding Performance Across Model-free RL

DGX agent

arXiv:2406.09079v5 Announce Type: replace Abstract: Deep reinforcement learning agents progressively lose representational capacity during training: neurons become dormant, removing active capacity fr

model-releasesarxiv-cs-lg
26 May 2026
← Previous
1…291292293294295…374
Next →