AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,771 results
11 Jun 2026

Learning to Inject: Automated Prompt Injection via Reinforcement Learning

SafetyDGX agent

arXiv:2602.05746v2 Announce Type: replace-cross Abstract: Prompt injection is a critical vulnerability in LLM agents, yet the strongest methods still rely on human red-teamers and hand-crafted prompts

Mind the Perspective: Let's Reason Recursively for Theory of Mind

Model ReleasesDGX agent

arXiv:2606.11724v1 Announce Type: new Abstract: Theory of Mind (ToM) reasoning requires inferring agents' beliefs from partial and asymmetric observations, which remains an open challenge for LLMs. Ex

PIGEON: VLM-Driven Object Navigation via Points of Interest Selection

Local AiDGX agent

arXiv:2511.13207v2 Announce Type: replace-cross Abstract: Object navigation in unseen indoor environments requires agents to perform semantic search under partial observability. Vision-language models

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

TouchThinker: Scaling Tactile Commonsense Reasoning to the Open World with Large-scale Data and Action-aware Representation

Model ReleasesDGX agent

arXiv:2606.11637v1 Announce Type: new Abstract: Touch is a key modality for embodied agents to understand the physical world. Although recent work has incorporated tactile signals into language system

10 Jun 2026

A Survey of Robotic Navigation and Manipulation with Physics Simulators in the Era of Embodied AI

Model ReleasesDGX agent

arXiv:2505.01458v2 Announce Type: replace-cross Abstract: Navigation and manipulation are core capabilities in Embodied AI, but training agents to perform them directly in the real world is costly, ti

CoCoSI: Collaborative Cognitive Map Construction for Spatial Intelligence

Model ReleasesDGX agent

arXiv:2606.10401v1 Announce Type: new Abstract: Spatial intelligence is a key frontier for multimodal large language models (MLLMs), enabling them to reason about the physical world from visual experi

Learn how @cursor_ai partnered with Together AI to deliver real-time inference for AI-powered coding in this article from @ce_zhang and @rea…

TutorialsDGX agent

Learn how @cursor_ai partnered with Together AI to deliver real-time inference for AI-powered coding in this article from @ce_zhang and @realDanFu. Cursor's in-editor agents generate code while develo

Nanocoder hit 2,000 GitHub stars 🌟

Local AiDGX agent

Nanocoder is an open coding agent for the terminal built by a community collective rather than a company. It allows users to bring their own model, keep code on their machine, and owe nothing to anyon

One thing I mentioned only in passing in my Fable post is that, for long running tasks, Fable starts to develop its own dialect as its many …

ApplicationsDGX agent

One thing I mentioned only in passing in my Fable post is that, for long running tasks, Fable starts to develop its own dialect as its many agents and tasks reinforce themselves and make Claudish lang

PreAct-Bench: Benchmarking Predictive Monitoring in LLMs

Model ReleasesDGX agent

arXiv:2606.09890v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed as autonomous agents capable of executing multi-step action trajectories toward a given objecti

ReflectiChain: Epistemic Grounding in LLM-Driven World Models for Supply Chain Resilience

Model ReleasesDGX agent

arXiv:2606.10359v1 Announce Type: new Abstract: AI agents in supply chains face a fundamental epistemic gap: large language models (LLMs) interpret policies but lack physical grounding, while reinforc

Rethinking Embodied Navigation via Relational Inductive Bias

SafetyDGX agent

arXiv:2606.10348v1 Announce Type: new Abstract: Object navigation requires an agent to locate a target in an unknown environment through visual observations. Existing methods typically rely on open-vo

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning

Model ReleasesDGX agent

arXiv:2510.14828v3 Announce Type: replace Abstract: Improving the reasoning capabilities of embodied agents is crucial for robots to complete complex human instructions in long-view manipulation tasks

SHAPO: Sharpness-Aware Policy Optimization for Safe Exploration

Model ReleasesDGX agent

arXiv:2606.10228v1 Announce Type: cross Abstract: Safe exploration is a prerequisite for deploying reinforcement learning (RL) agents in safety-critical domains. In this paper, we approach safe explor

9 Jun 2026

A Geometric Theory of Cognition for Machine Intelligence

ResearchDGX agent

arXiv:2512.12225v3 Announce Type: replace Abstract: Developing artificial agents that unify representation, memory, adaptation, and prediction remains a fundamental challenge in artificial intelligenc

Claude Fable 5: Available on Google Cloud

Model ReleasesDGX agent

Claude Fable 5, Anthropic’s latest frontier model, is now generally available on Google Cloud. This launch is the latest proof point of our ongoing commitment to bring the industry's latest models str

Datadog launches more than 100 features at DASH to push autonomous AI ops

Model ReleasesDGX agent

Observability and security platform company Datadog Inc. today unveiled more than 100 new capabilities at its annual DASH 2026 conference, headlined by a major expansion of its Bits AI agents that the

DIJIT: A Robotic Head for an Active Observer

ResearchDGX agent

arXiv:2512.07998v2 Announce Type: replace-cross Abstract: We present DIJIT, a novel binocular robotic head expressly designed for mobile agents that behave as active observers. DIJIT's unique breadth

If you thought AI progress was slowing down, well here's the immediate answer to that. Huge jump in capability across the board. This is goi…

Model ReleasesDGX agent

If you thought AI progress was slowing down, well here's the immediate answer to that. Huge jump in capability across the board. This is going to deliver major improvement in agents across almost all

Language-based Trial and Error Falls Behind in the Era of Experience

Model ReleasesDGX agent

arXiv:2601.21754v3 Announce Type: replace Abstract: While Large Language Models (LLMs) excel in language-based agentic tasks, their applicability to unseen, nonlinguistic environments (e.g., symbolic

LargeMonitor: Monitoring Online Task-Free Continual Learning via Large Pretrained Models

Model ReleasesDGX agent

arXiv:2606.09430v1 Announce Type: cross Abstract: Online task-free continual learning (TFCL) requires intelligent agents to sequentially accumulate knowledge from an unbounded, non-stationary data str

LLM Inference at the Edge: Mobile, NPU, and GPU Performance Efficiency Trade-offs Under Sustained Load

Model ReleasesDGX agent

arXiv:2603.23640v2 Announce Type: replace-cross Abstract: Deploying large language models on-device for always-on personal agents demands sustained inference from hardware tightly constrained in power

MuJoCo-Drones-Gym: A GPU-Accelerated Multi-Drone Simulator for Control and Reinforcement Learning

HardwareDGX agent

arXiv:2606.08039v1 Announce Type: new Abstract: Robotic simulators are a cornerstone of modern research in aerial robotics, serving both as a vehicle for the development of new control algorithms and

Setting a custom price for a model in AgentsView

Model ReleasesDGX agent

TIL: Setting a custom price for a model in AgentsView I've been really enjoying AgentsView by Wes McKinney as a tool for exploring my token usage across different coding agents running on my laptop. C

Storage Insights datasets: Enabling org-wide operational discovery with activity insights

Model ReleasesDGX agent

As enterprise storage footprints scale to billions of objects, AI applications and agentic workloads are fundamentally shifting the role of storage from a passive repository to the foundation of the d

Towards Automated Kernel Generation in the Era of LLMs

HardwareDGX agent

arXiv:2601.15727v3 Announce Type: replace Abstract: The performance of modern AI systems is fundamentally constrained by the quality of their underlying GPU kernels, which translate high-level algorit

Where Instruction Hierarchy Breaks: Diagnosing and Repairing Failures in Reasoning Language Models

Model ReleasesDGX agent

arXiv:2606.07808v1 Announce Type: new Abstract: Reasoning language models deployed in agentic workflows must follow an instruction hierarchy: when instructions from different sources conflict, the mod

WhiFlash: Accelerating Speculative Decoding with Token-Level Cross-Paradigm Routing

SafetyDGX agent

arXiv:2606.07710v1 Announce Type: cross Abstract: The autoregressive nature of large language models (LLMs) remains a significant bottleneck for inference, particularly in complex agentic workloads. W

8 Jun 2026

Accelerated Decentralized Stochastic Gradient Descent for Strongly Convex Optimization

ResearchDGX agent

arXiv:2606.07496v1 Announce Type: new Abstract: Decentralized stochastic optimization is a fundamental paradigm for large-scale learning over networks, where agents communicate only with their neighbo

Beyond Waypoints: A Trajectory-Centric Waypointing Paradigm for Vision-Language Navigation

Model ReleasesDGX agent

arXiv:2606.07244v1 Announce Type: cross Abstract: Vision-Language Navigation in Continuous Environments (VLN-CE) requires agents to follow natural-language instructions while navigating in real-world-

CAPE: Contrastive Action-conditioned Parallel Encoding for Embodied Planning

TutorialsDGX agent

arXiv:2606.07304v1 Announce Type: new Abstract: Embodied agents need to predict the future consequences of candidate actions in order to plan effectively before execution. Existing visual dynamics mod

Learning Explicit Behavioral Models with Adaptive Questions and World-Model Probes

Local AiDGX agent

arXiv:2606.07127v1 Announce Type: new Abstract: Interactive agents trained only against task return can achieve high scores while failing to represent the mechanisms that make their actions succeed. T

NTILC: Neural Tool Invocation via Learned Compression

Model ReleasesDGX agent

arXiv:2606.06566v1 Announce Type: cross Abstract: Agentic tool-calling language models depend on large registries of callable APIs, functions, and local actions. Placing full tool specifications direc

RECAP: Regression Evaluation for Continual Adaptation of Prompts

Model ReleasesDGX agent

arXiv:2606.06698v1 Announce Type: cross Abstract: Production agentic systems routinely face evolving constraints and must comply from the very next interaction. Scenarios like a tool-call notification

Robust Driving Control for Autonomous Vehicles: An Intelligent General-sum Constrained Adversarial Reinforcement Learning Approach

SafetyDGX agent

arXiv:2510.09041v3 Announce Type: replace-cross Abstract: Deep reinforcement learning (DRL) has demonstrated remarkable success in developing autonomous driving policies. However, its vulnerability to

Rubrics are even more flexible than /goal You can define a custom subagent for the grading, including custom tools, prompt, and iteration li…

Model ReleasesDGX agent

Rubrics are even more flexible than /goal You can define a custom subagent for the grading, including custom tools, prompt, and iteration limits. Try it out and let us know what you think! we just shi

7 Jun 2026

NVIDIA, KRAFTON, NC and Reigning ‘League of Legends’ Champions T1 Celebrate RTX Spark at Korea’s PC Bangs

HardwareDGX agent

At GTC Taipei at COMPUTEX last week, NVIDIA unveiled RTX Spark, the superchip that reinvents Windows PCs for the era of personal AI agents. On the heels of this announcement, NVIDIA founder and CEO Je

OpenAI plots biggest ChatGPT overhaul since launch

IndustryDGX agent

OpenAI is planning its biggest ChatGPT overhaul yet, aiming to turn it into a 'superapp' with coding tools and AI agents to boost revenue ahead of a potential stock market listing. The redesigned Chat

6 Jun 2026

Brick-Composer: Using MLLMs for Assembly with Diverse Bricks

Model ReleasesDGX agent

arXiv:2606.05445v1 Announce Type: new Abstract: We dream of AI agents that can read arbitrary designs and construct real-world objects from reusable building blocks. As a first step toward this vision

// Continual Learning Bench // One of the research areas with lots of investments is continual learning. While there are many efforts, there…

TutorialsDGX agent

// Continual Learning Bench // One of the research areas with lots of investments is continual learning. While there are many efforts, there is very little progress in measuring it. So the big questio

DragOn: A Benchmark and Dataset for Drag-Based GUI Interactions

Model ReleasesDGX agent

arXiv:2606.06322v1 Announce Type: new Abstract: GUI agents - vision-based models that control desktops, web browsers, and mobile devices through graphical user interfaces - promise to automate a wide

Five labs, five minds: building a multi-model finance drama on small models

ApplicationsDGX agent

This article describes a collaborative hackathon project involving five research labs that developed a financial simulation drama using small language models, focusing on building complex multi-agent

GIPO: Gaussian Importance Sampling Policy Optimization

SafetyDGX agent

arXiv:2603.03955v2 Announce Type: replace-cross Abstract: Post-training with reinforcement learning (RL) has recently shown strong promise for advancing multimodal agents beyond supervised imitation.

Goedel-Architect: Streamlining Formal Theorem Proving with Blueprint Generation and Refinement

Model ReleasesDGX agent

arXiv:2606.06468v1 Announce Type: new Abstract: We introduce Goedel-Architect, an agentic framework for formal theorem proving in Lean 4 centered on blueprint generation and refinement. A blueprint is

5 Jun 2026

Flow-based Policy Adaptation without Policy Updates

SafetyDGX agent

arXiv:2606.06461v1 Announce Type: new Abstract: Leveraging prior knowledge from pretrained policies, foundation models, or human operators offers an efficient alternative to learning robot skills from

Grok supports worktrees

IndustryDGX agent

Grok supports worktrees Grok Build tip of the day: worktrees! If you're unfamiliar with worktrees, they're essentially lightweight copies of your repo, allowing you to run parallel agents within their

KV-Control: Parameter-Efficient K/V Injection for Trajectory-Controlled Text-to-Motion

Model ReleasesDGX agent

arXiv:2606.05624v1 Announce Type: new Abstract: Text-conditioned 3D human motion models now synthesize plausible motions from prompts, but practical animation and embodied-agent workflows rarely stop

LeanMarathon: Toward Reliable AI Co-Mathematicians through Long-Horizon Lean Autoformalization

Local AiDGX agent

arXiv:2606.05400v1 Announce Type: cross Abstract: Long-horizon autoformalization of research mathematics fails not only at hard lemmas, but at scale: statements drift, dependencies tangle, context dec

LLM-Guided ANN Index Optimization for Human-Object Interaction Retrieval

Model ReleasesDGX agent

arXiv:2606.05489v1 Announce Type: new Abstract: Retrieval systems underpin modern AI applications -- spanning visual search, recommendation engines, and multi-modal question answering. Modern multi-st

PHUMA: Physically Reliable Humanoid Locomotion Dataset

ResearchDGX agent

arXiv:2510.26236v2 Announce Type: replace Abstract: Motion imitation is a promising approach for humanoid locomotion, enabling agents to acquire humanlike behaviors. Existing methods typically rely on

QueryAgent-R1: Bridging Query Generation and Product Retrieval for E-Commerce Query Recommendation

SafetyDGX agent

arXiv:2606.05671v1 Announce Type: new Abstract: Query recommendation in e-commerce search aims to proactively suggest queries that match users' potential interests. However, existing methods mainly op

RiskFlow: Fast and Faithful Safety-Critical Traffic Scenario Generation

SafetyDGX agent

arXiv:2606.06423v1 Announce Type: new Abstract: Safety-critical traffic scenario generation is essential for evaluating autonomous driving systems under rare but high-risk interactions. Existing diffu

The latest AI news we announced in May 2026

Model ReleasesDGX agent

Google's May 2026 AI updates center on the new 'agentic' era, featuring the Gemini 3.5 model and Gemini Omni for advanced reasoning and creation. Gemini Omni is a new model that can create anything fr

4 Jun 2026

Adaptive Information Control for Search-Augmented LLM Reasoning

SafetyDGX agent

arXiv:2602.01672v2 Announce Type: replace Abstract: Search-augmented reasoning agents interleave multi-step reasoning with external retrieval, but uncontrolled retrieval can introduce redundant eviden

BioBlue: Systematic runaway-optimiser-like LLM failure modes on biologically and economically aligned AI safety benchmarks for LLMs with simplified observation format

Model ReleasesDGX agent

arXiv:2509.02655v3 Announce Type: replace-cross Abstract: Many AI alignment discussions of 'runaway optimisation' focus on RL agents: unbounded utility maximisers that over-optimise a proxy objective

CYGNET: Cypher Gate for Neural Execution Triage and Cost Containment

ApplicationsDGX agent

arXiv:2606.04645v1 Announce Type: new Abstract: Language models acting as agents over knowledge graphs generate Cypher queries that fail structurally (crashing at the database) or semantically (execut

Nemotron 3 Ultra (550B-A55B) is here - our strongest open-weight model and full training recipe to date. Heavy emphasis on real-world infere…

Model ReleasesDGX agent

Nemotron 3 Ultra (550B-A55B) is here - our strongest open-weight model and full training recipe to date. Heavy emphasis on real-world inference efficiency for long-context agentic workloads. Everythin

ollama run gemma4:12b Gemma 4 12B is updated on Ollama, and available across all platforms! Try it on: Claude Code ollama launch claude --mo…

Model ReleasesDGX agent

ollama run gemma4:12b Gemma 4 12B is updated on Ollama, and available across all platforms! Try it on: Claude Code ollama launch claude --model gemma4:12b Hermes Agent ollama launch hermes --model gem

Policy Gradient for Continuous-Time Robust Markov Decision Processes

SafetyDGX agent

arXiv:2606.04335v1 Announce Type: new Abstract: The framework of robust Markov decision processes (RMDPs) allows the design of reinforcement learning agents that satisfy performance guarantees under w

R-APS: Compositional Reasoning and In-Context Meta-Learning for Constrained Design via Reflective Adversarial Pareto Search

Local AiDGX agent

arXiv:2606.04823v1 Announce Type: new Abstract: Large language models (LLMs) are fluent on open-ended tasks, yet in agentic settings, where a system must plan, use tools, and act over extended horizon

← Previous
1…229230231232233…297
Next →