AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlog
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,920 results
Safety

Deterministic Decisions for High-Stakes AI. A Zero-Egress Pipeline with the Deployability of RAG and the Accuracy of Machine Learning

DGX agent

arXiv:2606.29280v1 Announce Type: cross Abstract: We identify intervention bias as a previously unquantified failure mode of zero-shot large-language-model (LLM) educational advisory agents: without t

safetyarxiv-cs-ai
30 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

EVAF: A Test-Retest Protocol for Selective Parametric Consolidation

DGX agent

arXiv:2606.29916v1 Announce Type: cross Abstract: Long-running language agents need mechanisms for deciding which experiences should persist after the working context is gone. Retrieval systems can re

model-releasesarxiv-cs-ai
30 Jun 2026
Research

Exploring the Cryptographic Limits of Transformer Networks

DGX agent

arXiv:2606.29389v1 Announce Type: cross Abstract: In recent work it has been shown that colluding AI agents can use steganographic methods to exchange malicious information. Whether a transformer can

researcharxiv-cs-lg
30 Jun 2026
Safety

FutureNav: Unified World-Action Modeling for Vision-and-Language Navigation

DGX agent

arXiv:2606.30367v1 Announce Type: new Abstract: Vision-and-language navigation (VLN) in continuous environments requires an agent to ground instructions in egocentric observations while maintaining sp

safetyarxiv-cs-ro
30 Jun 2026
Model Releases

Google’s Gemini Omni Flash and Nano Banana 2 Lite support slick media content creation at lower costs

DGX agent

Google LLC is enhancing its generative artificial intelligence capabilities for creators with the debut of a pair of new media-focused models in the Gemini Enterprise Agent Platform. The new additions

model-releasessiliconangle
30 Jun 2026
Safety

Hierarchical Decision Making with Structured Policies: A Principled Design via Inverse Optimization

DGX agent

arXiv:2606.28764v1 Announce Type: new Abstract: Hierarchical decision-making frameworks are pivotal for addressing complex control tasks, enabling agents to decompose intricate problems into manageabl

safetyarxiv-cs-lg
30 Jun 2026
Model Releases

Internal-State Probes Read the Situation, Not the Action: Three Negative Results for Pre-Action Misalignment Monitoring

DGX agent

arXiv:2606.30449v1 Announce Type: new Abstract: Probes on model internals could help monitor agentic systems if they identify harmful text or tool actions before those actions are generated. We ask wh

model-releasesarxiv-cs-lg
30 Jun 2026
Safety

Learned Coordination Conventions in Cooperative MARL: Measuring the Translation Gap Between Theory-Informed Roles and Learned Routing

DGX agent

arXiv:2606.29541v1 Announce Type: new Abstract: Role-semantic assignments provide priors over how heterogeneous agents may coordinate, but cooperative MARL systems instead settle on conventions throug

safetyarxiv-cs-ai
30 Jun 2026
Model Releases

Lost in Execution: On the Multilingual Robustness of Tool Calling in Large Language Models

DGX agent

arXiv:2601.05366v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly deployed as agents that invoke external tools through structured function calls. While recent wo

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

MirrorCode: AI can rebuild entire programs from behavior alone

DGX agent

arXiv:2606.30182v1 Announce Type: new Abstract: AI models are rapidly improving at autonomous coding, as shown by benchmark progress and one-off demonstrations such as AI implementing a C compiler. Ho

model-releasesarxiv-cs-ai
30 Jun 2026
Safety

MIThinker: A Plug-and-Play Policy-Optimized Thinker For Motivational Interviewing Counseling

DGX agent

arXiv:2606.29265v1 Announce Type: new Abstract: Reasoning large language models (LLMs) have recently made much progress in complex problem-solving, leveraging internal reasoning (or thought) to guide

safetyarxiv-cs-cl
30 Jun 2026
Local Ai

Modernizing financial services with deployment freedom and transformational AI with AlloyDB Omni

DGX agent

The financial services industry (FSI) operates under a unique set of non-negotiable requirements: the need for strict regulatory compliance, sub-millisecond transactional speeds, and security that ver

local-aigoogle-cloud-ai
30 Jun 2026
Safety

Modification-Considering Value Learning for Reward Hacking Mitigation in RL

DGX agent

arXiv:2606.28955v1 Announce Type: cross Abstract: Reinforcement learning agents can exploit misspecified reward signals to achieve high apparent returns while failing on the intended objective, a fail

safetyarxiv-cs-ai
30 Jun 2026
Model Releases

Optimizing Expert-Designed Crystal Graph Networks for Band-Gap Prediction with an Autonomous LLM Research Loop

DGX agent

arXiv:2606.29717v1 Announce Type: cross Abstract: Predicting a material's properties from its structure is a central, fast-advancing problem in computational materials science. A decade of work has pr

model-releasesarxiv-cs-ai
30 Jun 2026
Safety

RoamFlow: Reinforcement-Aligned One-Step Action MeanFlow Policy for Image-Goal Navigation

DGX agent

arXiv:2606.29934v1 Announce Type: new Abstract: Image-goal navigation is a key challenge in embodied robotics, where an agent must reach a target specified solely by a goal image. While existing reinf

safetyarxiv-cs-ro
30 Jun 2026
Model Releases

SABER-Math: Automated Benchmark for Information Retrieval Evaluation in Mathematics

DGX agent

arXiv:2606.29894v1 Announce Type: cross Abstract: As agentic AI systems tackle more complex mathematical tasks, they increasingly rely on information retrieval (IR) to search problem databases, theore

model-releasesarxiv-cs-ai
30 Jun 2026
Safety

Safety from Honesty in a Disinterested AI Predictor

DGX agent

arXiv:2606.29657v1 Announce Type: new Abstract: As AI systems become more capable, training procedures that optimize for downstream outcomes risk introducing implicit agency: goal-directed behavior th

safetyarxiv-cs-ai
30 Jun 2026
Model Releases

Self-Supervised Theorem Discovery in a Formal Axiomatic System

DGX agent

arXiv:2606.28747v1 Announce Type: new Abstract: Recent artificial intelligence (AI) systems have shown remarkable progress in mathematical reasoning. Many existing approaches, including large language

model-releasesarxiv-cs-ai
30 Jun 2026
Tutorials

The Download: AI “coworkers” and stratospheric internet

DGX agent

This is today’s edition of The Download, our weekday newsletter that provides a daily dose of what’s going on in the world of technology. AI agents are not your “coworkers” Imagine coming in to work t

tutorialsmit-tech-review
30 Jun 2026
Research

TRACE: Temporal Relationship-Aware Conversational Entrainment Detection in Dyadic Speech

DGX agent

arXiv:2606.30543v1 Announce Type: cross Abstract: With the proliferation of speech AI agents, understanding emotional entrainment in conversational interaction has become increasingly important. Emoti

researcharxiv-cs-ai
30 Jun 2026
Model Releases

Translating Natural Language to Strategic Temporal Specifications via LLMs

DGX agent

arXiv:2606.30441v1 Announce Type: cross Abstract: A rigorous formalization of system requirements is a fundamental prerequisite for the verification of Multi-Agent Systems (MAS). However, writing corr

model-releasesarxiv-cs-ai
30 Jun 2026
Safety

X-Mind: Efficient Visual Chain-of-Thought via Predictive World Model for End-to-End Driving

DGX agent

arXiv:2606.28758v1 Announce Type: cross Abstract: Predicting future states is essential for autonomous agents, yet current Vision-Language-Action (VLA) models fundamentally lack this capability, relyi

safetyarxiv-cs-ai
30 Jun 2026
Model Releases

EXPLORE-Bench: Egocentric Scene Prediction with Long-Horizon Reasoning

DGX agent

arXiv:2603.09731v3 Announce Type: replace-cross Abstract: Multimodal large language models (MLLMs) are increasingly considered as a foundation for embodied agents, yet it remains unclear whether they

model-releasesarxiv-cs-ai
29 Jun 2026
Safety

hia-gat: A Heterogeneous Interaction-Aware Graph Attention Network For Frame-Level Traffic Conflict Risk Prediction On Freeways

DGX agent

arXiv:2606.27577v1 Announce Type: cross Abstract: This paper formulates frame-level freeway risk assessment as a multi-agent scene graph-level binary classification problem, where each video or trajec

safetyarxiv-cs-ai
29 Jun 2026
Model Releases

In the next version of Claude Code: subagents run in the background by default, so you can keep talking to Claude while your subagents work …

DGX agent

In the next version of Claude Code: subagents run in the background by default, so you can keep talking to Claude while your subagents work If you want your agent to run in the foreground, just tell C

model-releasesboris-cherny--x
29 Jun 2026
Research

Memora: A Harmonic Memory Representation Balancing Abstraction and Specificity

DGX agent

AI agents can't remember past conversations. They must constantly reload or retrieve context, which grows less efficient as tasks get longer and more complex. Memora solves this with a scalable memory

researchmicrosoft-research
29 Jun 2026
Tutorials

RAE-NWM: Navigation World Model in Dense Visual Representation Space

DGX agent

arXiv:2603.09241v2 Announce Type: replace Abstract: Visual navigation requires agents to reach goals in complex environments through perception and planning. World models address this task by simulati

tutorialsarxiv-cs-cv
29 Jun 2026
Applications

Towards Evaluation of Implicit Software World Models in Coding LLMs

DGX agent

arXiv:2606.27406v1 Announce Type: cross Abstract: Software engineering, whether performed by humans or by AI agents, requires reasoning about how software behaves. We call the internal model that supp

applicationsarxiv-cs-ai
29 Jun 2026
Safety

Towards Value-Constrained Credit Assignment in Fully Delegated AI Cooperatives

DGX agent

arXiv:2606.28217v1 Announce Type: cross Abstract: We propose a framework for reward allocation in fully delegated AI cooperatives where humans are represented by agents that contribute data and partic

safetyarxiv-cs-ai
29 Jun 2026
Local Ai

Understanding Rollout Error in Graph World Models

DGX agent

arXiv:2606.27780v1 Announce Type: new Abstract: World models are often used for planning by rolling learned dynamics forward. Many planning environments, however, are not vectors or images; they are g

local-aiarxiv-cs-ai
29 Jun 2026
Tutorials

> Better Caching – Cache misses are the easiest way to drive your cost up. All of our requests are cache aware, so we’re reusing a warm cach…

DGX agent

> Better Caching – Cache misses are the easiest way to drive your cost up. All of our requests are cache aware, so we’re reusing a warm cache wherever possible. We do this for you in Deep Agents - see

tutorialsharrison-chase--x
27 Jun 2026
Research

Link to the full article: https://magazine.sebastianraschka.com/p/using-local-coding-agents

DGX agent

This article explores the implementation and use of local coding agents—AI systems designed to autonomously write, test, and debug code on local machines without relying on external APIs. The piece li

researchsebastian-raschka--x
27 Jun 2026
Model Releases

Sakana Fugu Technical Report Instead of training one larger model, Sakana AI trains an orchestrator that reads each query and dynamically ro…

DGX agent

Sakana Fugu Technical Report Instead of training one larger model, Sakana AI trains an orchestrator that reads each query and dynamically routes or composes GPT-5.5, Gemini-3.1-Pro, Claude Opus 4.8 an

model-releasesdavid-ha--x
27 Jun 2026
Safety

Automating Potential-based Reward Shaping with Vision Language Model Guidance

DGX agent

arXiv:2606.27180v1 Announce Type: cross Abstract: Sparse rewards are inherently challenging for reinforcement learning agents as they lack intermediate feedback to guide exploration and to correctly a

safetyarxiv-cs-ai
26 Jun 2026
Model Releases

Dynamic workflows (generating harnesses on the fly) are a new form of test-time compute. But LLMs aren't great at building them. I often hav…

DGX agent

Dynamic workflows (generating harnesses on the fly) are a new form of test-time compute. But LLMs aren't great at building them. I often have to steer agents to generate complex patterns. Curious how

model-releasesdair-ai--x
26 Jun 2026
Model Releases

Parametric Open Source Games

DGX agent

arXiv:2606.27068v1 Announce Type: cross Abstract: Open-source game theory studies agents whose behavior may depend on one another's decision procedures, but most existing models use discrete or symbol

model-releasesarxiv-cs-ai
26 Jun 2026
Safety

Proposal-Conditioned Latent Diffusion for Closed-Loop Traffic Scenario Generation

DGX agent

arXiv:2606.27123v1 Announce Type: cross Abstract: Closed-loop traffic simulation remains challenging because it must generate interactive multi-agent behaviors that are scene-consistent and controllab

safetyarxiv-cs-cv
26 Jun 2026
Safety

Radical AI Interpretability

DGX agent

arXiv:2606.26523v1 Announce Type: new Abstract: We develop a framework for interpreting AI systems as agents, drawing on the philosophical tradition of radical interpretation and the tools of mechanis

safetyarxiv-cs-ai
26 Jun 2026
Model Releases

SciFig: Towards Automating Editable Figure Generation for Scientific Papers

DGX agent

arXiv:2601.04390v2 Announce Type: replace Abstract: High-quality methodology figures are central to scientific communication, yet they remain difficult and time-consuming to create. Such figures must

model-releasesarxiv-cs-ai
26 Jun 2026
Applications

The big lesson from training @cursor_ai Composer 2: models exploit flaws in their training environment before learning what you actually wan…

DGX agent

The big lesson from training @cursor_ai Composer 2: models exploit flaws in their training environment before learning what you actually want. Real RL for coding agents means production-faithful envir

applicationsfireworks-ai--x
26 Jun 2026
Model Releases

Today on TITV: -Trump asks OpenAI to stagger release of new model | @leomschwartz & @amir, The Information -Google pressures publishers on A…

DGX agent

Today on TITV: -Trump asks OpenAI to stagger release of new model | @leomschwartz & @amir, The Information -Google pressures publishers on AI licensing | @anngehan -Inside an AI power user’s agent wor

model-releasesallie-k--miller--x
26 Jun 2026
Model Releases

Beyond One-Size-Fits-All: Diagnosis-Driven Online Reinforcement Learning with Offline Priors

DGX agent

arXiv:2606.25527v1 Announce Type: new Abstract: Online reinforcement learning (RL) agents increasingly depend on knowledge acquired offline to achieve practical efficiency. Originally studied in offli

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

BreachRx launches Rex Platform to coordinate AI-era incident response

DGX agent

Incident response company BreachRx Inc. today launched the Rex Platform, an agentic artificial intelligence incident command center built for a future in which AI-accelerated attacks set off several b

model-releasessiliconangle
25 Jun 2026
Model Releases

Improving Zero-Shot Offline RL via Behavioral Task Sampling

DGX agent

arXiv:2604.25496v2 Announce Type: replace Abstract: Offline zero-shot reinforcement learning (RL) aims to learn agents that optimize unseen reward functions without additional environment interaction.

model-releasesarxiv-cs-ai
25 Jun 2026
Safety

SAGE-Nav: Leveraging LLM Planning and Alignment Fusion for Hierarchical Scene Graph-Guided Navigation

DGX agent

arXiv:2606.25497v1 Announce Type: new Abstract: Object-Goal Navigation (ObjNav) requires embodied agents to autonomously locate specified targets using only egocentric visual observations. Existing mo

safetyarxiv-cs-ro
25 Jun 2026
Model Releases

Spatio-Temporal Mixture-of-Modality-Experts Diffusion for Quantitative DCE-MRI Synthesis from Incomplete MR Sequences

DGX agent

arXiv:2606.25535v1 Announce Type: new Abstract: Quantitative maps from dynamic contrast-enhanced MRI (DCE-MRI) are essential for tumor assessment but are often unavailable due to contrast-agent risks

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

USS: Unified Spatial-Semantic Prompts for Embodied Visual Tracking with Latent Dynamics Learning

DGX agent

arXiv:2606.25880v1 Announce Type: new Abstract: Embodied Visual Tracking (EVT) requires an agent to continuously follow a specified target while actively moving through dynamic environments. However,

model-releasesarxiv-cs-cv
25 Jun 2026
Safety

Why Multi-Step Tool-Use Reinforcement Learning Collapses and How Supervisory Signals Fix It

DGX agent

arXiv:2606.26027v1 Announce Type: new Abstract: Tool use enables large language models (LLMs) to perform complex tasks, and recent agentic reinforcement learning (RL) methods show promise for enhancin

safetyarxiv-cs-cl
25 Jun 2026
← Previous
1…287288289290291…374
Next →