AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,919 results
Local Ai

Playing games with knowledge: AI-Induced delusions need game theoretic interventions

DGX agent

arXiv:2605.08409v1 Announce Type: new Abstract: Conversational AI has a fundamental flaw as a knowledge interface: sycophantic chatbots induce epistemic entrenchment and delusional belief spirals even

local-aiarxiv-cs-ai
12 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Position: AI Security Policy Should Target Systems, Not Models

DGX agent

arXiv:2605.09504v1 Announce Type: cross Abstract: We present swarm-attack, an open-source adversarial testing framework in which multiple lightweight LLM agents coordinate through shared memory, paral

model-releasesarxiv-cs-ai
12 May 2026
Safety

Reflective Prompted Policy Optimization: Trajectory-Grounded Revision and Salience Bias

DGX agent

arXiv:2605.08315v1 Announce Type: new Abstract: Existing LLM-based policy optimizers see only scalar rewards: that a policy scored 0.45, but not whether the agent got stuck in a loop, fell into a hole

safetyarxiv-cs-lg
12 May 2026
Safety

SalesSim: Benchmarking and Aligning Multimodal Language Models as Retail User Simulators

DGX agent

arXiv:2605.08334v1 Announce Type: new Abstract: We present SalesSim, a framework and testbed for evaluating the ability of Multimodal Large Language Models (MLLMs) to simulate realistic, persona-drive

safetyarxiv-cs-cl
12 May 2026
Model Releases

Selective Deficits in LLM Mental Self-Modeling in a Behavior-Based Test of Theory of Mind

DGX agent

arXiv:2603.26089v2 Announce Type: replace-cross Abstract: The ability to represent oneself and others as agents with knowledge, intentions, and belief states that guide their behavior - Theory of Mind

model-releasesarxiv-cs-ai
12 May 2026
Safety

Shields to Guarantee Probabilistic Safety in MDPs

DGX agent

arXiv:2605.10888v1 Announce Type: cross Abstract: Shielding is a prominent model-based technique to ensure safety of autonomous agents. Classical shielding aims to ensure that nothing bad ever happens

safetyarxiv-cs-ai
12 May 2026
Tutorials

Sign up if you're in LA!

DGX agent

Sign up if you're in LA! The last AI Agents Happy Hour was so fun, @jvedi and I are going to do it again. This time, co-hosted by @pinecone! We want to see what you're building and learn from your exp

tutorialspinecone--x
12 May 2026
Model Releases

Statistical Model Checking of the Keynes+Schumpeter Model: A Transient Sensitivity Analysis of a Macroeconomic ABM

DGX agent

arXiv:2605.10447v1 Announce Type: cross Abstract: Agent-based models (ABMs) are increasingly used in macroeconomics, but their analysis still often relies on ad hoc Monte Carlo campaigns with heteroge

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Step Rejection Fine-Tuning: A Practical Distillation Recipe

DGX agent

arXiv:2605.10674v1 Announce Type: cross Abstract: Rejection Fine-Tuning (RFT) is a standard method for training LLM agents, where unsuccessful trajectories are discarded from the training set. In the

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Talk to Your Slides: High-Efficiency Slide Editing via Language-Driven Structured Data Manipulation

DGX agent

arXiv:2505.11604v5 Announce Type: replace Abstract: Editing presentation slides is a frequent yet tedious task, ranging from creative layout design to repetitive text maintenance. While recent GUI-bas

model-releasesarxiv-cs-cl
12 May 2026
Safety

TIE: Time Interval Encoding for Video Generation over Events

DGX agent

arXiv:2605.10543v1 Announce Type: new Abstract: Director-style prompting, robotic action prediction, and interactive video agents demand temporal grounding over concurrent events -- a regime in which

safetyarxiv-cs-cv
12 May 2026
Model Releases

Towards Conversational Medical AI with Eyes, Ears and a Voice

DGX agent

arXiv:2605.09272v1 Announce Type: new Abstract: The practice of medicine relies not only upon skillful dialogue but also on the nuanced exchange and interpretation of rich auditory and visual cues bet

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Try Grok Voice

DGX agent

Try Grok Voice Grok Voice Think Fast 1.0 ranks #1 on the Artificial Analysis τ-Voice benchmark for real-world agentic customer service resolution Absolutely outperforming GPT-Realtime-2 (High) and Gem

model-releaseselon-musk--x
12 May 2026
Model Releases

What Parameter Golf taught us about AI-assisted research

DGX agent

Parameter Golf brought together 1,000+ participants and 2,000+ submissions to explore AI-assisted machine learning research, coding agents, quantization, and novel model design under strict constraint

model-releasesopenai
12 May 2026
Applications

What to expect during KB4-CON: Join theCUBE May 14

DGX agent

Human risk management is becoming a practical measure of enterprise security. The old playbook treated employees as the weak link; the new one has to account for people, AI agents and automated decisi

applicationssiliconangle
12 May 2026
Model Releases

When Reviews Disagree: Fine-Grained Contradiction Analysis in Scientific Peer Reviews

DGX agent

arXiv:2605.10171v1 Announce Type: cross Abstract: Scientific peer reviews frequently contain conflicting expert judgments, and the increasing scale of conference submissions makes it challenging for A

model-releasesarxiv-cs-ai
12 May 2026
Applications

AT-VLA: Adaptive Tactile Injection for Enhanced Feedback Reaction in Vision-Language-Action Models

DGX agent

arXiv:2605.07308v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have significantly advanced the capabilities of robotic agents in executing diverse tasks; however, they still face

applicationsarxiv-cs-ro
11 May 2026
Model Releases

Beyond Linear Attention: Softmax Transformers Implement In-Context Reinforcement Learning

DGX agent

arXiv:2605.07333v1 Announce Type: new Abstract: In-context reinforcement learning (ICRL) studies agents that, after pretraining, adapt to new tasks by conditioning on additional context without parame

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

Contrast-X: A Multi-Modal Contrast Image Synthesis Benchmark and Universal Modality Flow Matching

DGX agent

arXiv:2601.15884v2 Announce Type: replace Abstract: Contrast-enhanced imaging is central to oncologic diagnosis, but contrast agents can be contraindicated for many of the patients who need them most.

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

Convergence and Emergence of In-Context Reinforcement Learning with Chain of Thought

DGX agent

arXiv:2605.07123v1 Announce Type: new Abstract: In-context reinforcement learning (ICRL) refers to the ability of RL agents to adapt to new tasks at inference time without parameter updates by conditi

model-releasesarxiv-cs-lg
11 May 2026
Safety

Decentralized Time-Varying Optimization for Streaming Data via Temporal Weighting

DGX agent

arXiv:2605.06971v1 Announce Type: cross Abstract: Classical optimization theory largely focuses on fixed objective functions, whereas many modern learning systems operate in dynamic environments where

safetyarxiv-cs-ai
11 May 2026
Model Releases

Discovering Ordinary Differential Equations with LLM-Based Qualitative and Quantitative Evaluation

DGX agent

arXiv:2605.07323v1 Announce Type: new Abstract: Discovering governing differential equations from observational data is a fundamental challenge in scientific machine learning. Existing symbolic regres

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Echo: KV-Cache-Free Associative Recall with Spectral Koopman Operators

DGX agent

arXiv:2605.06997v1 Announce Type: new Abstract: Long chain-of-thought reasoning and agentic tool-calling produce traces spanning tens of thousands of tokens, yet Transformer KV caches grow linearly wi

model-releasesarxiv-cs-lg
11 May 2026
Safety

Entropy-Regularized Adjoint Matching for Offline Reinforcement Learning

DGX agent

arXiv:2605.06156v2 Announce Type: replace-cross Abstract: Integrating expressive generative policies, such as flow-matching models, into offline reinforcement learning (RL) allows agents to capture co

safetyarxiv-cs-ai
11 May 2026
Tools

From observability to context: What’s next for Arize Phoenix

DGX agent

As agents start changing software, they need a way to verify their work that includes traces, evals, feedback, and APIs. This is where Phoenix goes next — not the next release, but what this product b

toolsarize-ai
11 May 2026
Safety

Multi-Environment POMDPs with Finite-Horizon Objectives

DGX agent

arXiv:2605.07537v1 Announce Type: new Abstract: Partially Observable Markov Decision Processes (POMDPs) are systems in which one agent interacts with a stochastic environment, and receives only partia

safetyarxiv-cs-ai
11 May 2026
Model Releases

Multi-Objective Constraint Inference using Inverse reinforcement learning

DGX agent

arXiv:2605.06951v1 Announce Type: new Abstract: Constraint inference is widely considered essential to align reinforcement learning agents with safety boundaries and operational guidelines by observin

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

OpenAI just released its answer to Claude Mythos

DGX agent

OpenAI is launching Daybreak, an AI initiative focused on detecting and patching vulnerabilities before attackers find them. Daybreak uses the Codex Security AI agent that launched in March to create

model-releasesthe-verge-ai
11 May 2026
Safety

SB-TRPO: Towards Safe Reinforcement Learning with Hard Constraints

DGX agent

arXiv:2512.23770v3 Announce Type: replace-cross Abstract: In safety-critical domains, reinforcement learning (RL) agents must often satisfy strict, zero-cost safety constraints while accomplishing tas

safetyarxiv-cs-ai
11 May 2026
Safety

Sparse Autoencoders as Plug-and-Play Firewalls for Adversarial Attack Detection in VLMs

DGX agent

arXiv:2605.07447v1 Announce Type: cross Abstract: Vision-language models (VLMs) have advanced rapidly and are increasingly deployed in real-world applications, especially with the rise of agent-based

safetyarxiv-cs-ai
11 May 2026
Model Releases

Local AI is having its moment! Below is the number of new GGUF models created each month over the past 8 months & insights from our HF inter…

DGX agent

Local AI is having its moment! Below is the number of new GGUF models created each month over the past 8 months & insights from our HF internal agent (May is partial): - 176,000 total public GGUF mode

model-releasesclem-delangue--x
10 May 2026
Industry

AI Ascent 2026

DGX agent

AI Ascent 2026 is Sequoia Capital's event that convened leading AI researchers and founders including Greg Brockman, Andrej Karpathy, and Demis Hassabis. The event featured discussions on AI agents, s

industrysequoia-capital
8 May 2026
Model Releases

Anthropic gave these out yesterday at code with claude. Added personalized memory and Claude to it. You can just build things. @bcherny @trq…

DGX agent

Anthropic gave these out yesterday at code with claude. Added personalized memory and Claude to it. You can just build things. @bcherny @trq212 Time to add managed agents to this. That’s going to be s

model-releasesboris-cherny--x
8 May 2026
Tutorials

9 themes defining the future of AI-driven customer engagement: Insights from Twilio’s Signal event

DGX agent

Customer engagement is shifting from disconnected interactions to continuous, AI-driven experiences that span channels and adapt in real time — and unified data, orchestration and AI agents are rapidl

tutorialssiliconangle
7 May 2026
Safety

Distilling Bayesian Belief States into Language Models for Auditable Negotiation

DGX agent

arXiv:2605.04507v1 Announce Type: new Abstract: Negotiation agents must infer what their counterpart values, update those beliefs over dialogue turns, and choose actions under uncertainty. End-to-end

safetyarxiv-cs-cl
7 May 2026
Model Releases

🚀 Introducing the genmedia CLI, generative media directly from the command line. Generate images, video, 3D and audio from your terminal, a…

DGX agent

🚀 Introducing the genmedia CLI, generative media directly from the command line. Generate images, video, 3D and audio from your terminal, alongside Claude and other AI agents. • Native terminal workfl

model-releasessonya-huang--x
7 May 2026
Industry

Mozilla says 271 vulnerabilities found by Mythos have 'almost no false positives'

DGX agent

Mozilla ran an agentic harness powered by Claude Mythos Preview across Firefox's source code, identifying 271 security bugs fixed in Firefox 150 . The breakthrough was achieved through improvements in

industryars-technica
7 May 2026
Applications

New #YAAP episode out now 🎙️ @yuvalinthedeep sits down with @mikegchambers from @awsdevelopers to unpack harness engineering and why it's t…

DGX agent

New #YAAP episode out now 🎙️ @yuvalinthedeep sits down with @mikegchambers from @awsdevelopers to unpack harness engineering and why it's the reason most agents never make it to production. 🎧 Listen/W

applicationsai21-labs--x
7 May 2026
Model Releases

OpenClaw and Claude can put your AI-generated podcasts in Spotify

DGX agent

Save to Spotify is a new command-line tool designed specifically for AI agents like OpenClaw, Claude Code, or OpenAI Codex. If you're the kind of person who collects research on a topic, then feeds it

model-releasesthe-verge-ai
7 May 2026
Model Releases

A Benchmark for Interactive World Models with a Unified Action Generation Framework

DGX agent

arXiv:2605.03941v1 Announce Type: new Abstract: Achieving Artificial General Intelligence (AGI) requires agents that learn and interact adaptively, with interactive world models providing scalable env

model-releasesarxiv-cs-cv
6 May 2026
Research

A Sentence Relation-Based Approach to Sanitizing Malicious Instructions

DGX agent

arXiv:2605.01078v1 Announce Type: cross Abstract: Retrieval-augmented generation and tool-integrated LLM agents increasingly depend on external textual sources. This reliance broadens the available at

researcharxiv-cs-ai
6 May 2026
Model Releases

Code with Claude is happening now! ▪︎ 9:00AM - Keynote ▪︎ 10:30AM - What's new in Claude Code ▪︎ 11:15AM - Building on Claude at GitHub scal…

DGX agent

Code with Claude is happening now! ▪︎ 9:00AM - Keynote ▪︎ 10:30AM - What's new in Claude Code ▪︎ 11:15AM - Building on Claude at GitHub scale ▪︎ 12:00PM - Get to production faster with Managed Agents

model-releasesboris-cherny--x
6 May 2026
Industry

Devs are spending only 16% of their time coding. Atlassian is engineering AI to reclaim the rest

DGX agent

The rise of AI coding agents has commoditized code generation, exposing a deeper challenge for software teams: the non-coding friction that consumes the vast majority of a developer’s day. Fixing that

industrysiliconangle
6 May 2026
Model Releases

From Where Things Are to What They’re For: Benchmarking Spatial–Functional Intelligence for Multimodal LLMs

DGX agent

True spatial intelligence for multimodal agents transcends low-level geometric perception, evolving from knowing where things are to understanding what they are for. While existing benchmarks, such as

model-releasesapple-ml-research
6 May 2026
Tools

https://x.com/walden_yan/status/2052070983083942322?s=20

DGX agent

https://x.com/walden_yan/status/2052070983083942322?s=20 Cool to see failure modes of different coding agents in new report from @greptile - seems like Devin is better than humans in almost all catego

toolswindsurf--x
6 May 2026
Safety

Intervention Complexity as a Canonical Reward and a Measure of Intelligence

DGX agent

arXiv:2605.02175v1 Announce Type: new Abstract: The Legg--Hutter universal intelligence measure provides a rigorous scalar assessment of general intelligence as expected reward across all computable e

safetyarxiv-cs-ai
6 May 2026
Research

MEMAUDIT: An Exact Package-Oracle Evaluation Protocol for Budgeted Long-Term LLM Memory Writing

DGX agent

arXiv:2605.02199v1 Announce Type: new Abstract: Long-term LLM agents must compress streams of past interactions into persistent memory before future queries are known. Existing evaluations usually mea

researcharxiv-cs-ai
6 May 2026
Local Ai

On-Device Fine-Tuning via Backprop-Free Zeroth-Order Optimization

DGX agent

arXiv:2511.11362v2 Announce Type: replace-cross Abstract: On-device fine-tuning is a critical capability for edge AI systems, which must support adaptation to different agentic tasks under stringent m

local-aiarxiv-cs-cl
6 May 2026
← Previous
1…295296297298299…374
Next →