AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,771 results
7 Jul 2026

NEST: Nascent Encoded Steganographic Thoughts

Model ReleasesDGX agent

arXiv:2602.14095v2 Announce Type: replace Abstract: Monitoring chain-of-thought (CoT) reasoning is a foundational safety technique for large language model agents; however, this oversight is compromis

Saving GPU Hours in LLM Inference System Development and Online Workloads with Simulation and DBMS-Inspired Cache Replacement Policies

SafetyDGX agent

arXiv:2411.07447v5 Announce Type: replace-cross Abstract: LLMs are increasingly used world-wide from daily tasks to agentic systems and data analytics, requiring significant GPU resources. While LLM i

VERITAS: Towards a General-Purpose Replication Tool for Scientific Research

Model ReleasesDGX agent

arXiv:2607.02931v1 Announce Type: new Abstract: AI tools are accelerating scientific publication while the systems that review it struggle to keep up, and independent verification of published researc

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
6 Jul 2026

OpenClaw landed on @huggingface local apps 🦞🤝🤗 1. Pick any GGUF/MLX model on hf 2. Copy the openclaw onboard setup 3. Volla you've got a …

Local AiDGX agent

OpenClaw landed on @huggingface local apps 🦞🤝🤗 1. Pick any GGUF/MLX model on hf 2. Copy the openclaw onboard setup 3. Volla you've got a tool-calling agent running fully local. no cloud, no keys, no o

'The frontier labs will keep owning discovery. Open source will increasingly own production.' Insightful take on a tired topic 'When you're …

ApplicationsDGX agent

'The frontier labs will keep owning discovery. Open source will increasingly own production.' Insightful take on a tired topic 'When you're running AI agents in production for customer service, latenc

4 Jul 2026

spending the last week at @aidotengineer was awesome. too many great convos to cover them all, but jotted down some things that stood out: -…

Model ReleasesDGX agent

spending the last week at @aidotengineer was awesome. too many great convos to cover them all, but jotted down some things that stood out: - Lots of discussion around open source models. I spoke with

3 Jul 2026

Behind the Refusal: Determining Guardrail Activation via Behavioral Monitoring

SafetyDGX agent

arXiv:2607.02121v1 Announce Type: cross Abstract: As Large Language Models (LLMs) and agentic systems become integrated into real-world applications, ensuring their safety and security is critical. Gu

Collaborative Disagreement Resolution for Scalable Oversight

ResearchDGX agent

arXiv:2607.01251v1 Announce Type: cross Abstract: Debate, where AI agents argue opposing positions, has emerged as a key approach to scalable oversight. However, debate faces a fundamental tension: mo

Hardware-Enforced Semantic Coordination for Safety-Critical Real-Time Autonomous Systems

SafetyDGX agent

arXiv:2607.02376v1 Announce Type: new Abstract: Recent advances in agentic AI are producing increasingly complex autonomous systems that integrate large language models, world models, optimization eng

MAGIK: Mapping to Analogous Goals via Imagination-enabled Knowledge Transfer

SafetyDGX agent

arXiv:2506.01623v4 Announce Type: replace Abstract: Humans excel at analogical reasoning - applying knowledge from one task to a related one with minimal relearning. In contrast, reinforcement learnin

Object Aligner: A Configurable JSON Schema Similarity Score for Graphs, Applied to LLM Prompt Optimization

SafetyDGX agent

arXiv:2607.01972v1 Announce Type: cross Abstract: Large language models (LLMs) are often asked to produce JSON conforming to a fixed schema, powering information extraction, tool calling, agentic plan

PairCoder++: Pair Programming as a Universal Paradigm for Verified Code-Driven Multimodal and Structured-Artifact Generation

Model ReleasesDGX agent

arXiv:2607.01883v1 Announce Type: new Abstract: Code is the medium through which large language models generate structured artifacts: charts, scientific figures, vector graphics, CAD models, 3D scenes

Some notes from @aiDotEngineer world fair: > the energy was incredible. it's magical to have a large group of smart, hungry, technical, and …

Model ReleasesDGX agent

Some notes from @aiDotEngineer world fair: > the energy was incredible. it's magical to have a large group of smart, hungry, technical, and driven people learning from each other under one roof. > the

We're releasing the full slides for our 2 hr deepdive session from the AI Engineer World's Fair. We covered how we build inference engines t…

HardwareDGX agent

We're releasing the full slides for our 2 hr deepdive session from the AI Engineer World's Fair. We covered how we build inference engines to serve agentic workloads at trillion token production scale

2 Jul 2026

Affogato: Open-Vocabulary Affordance Grounding with Automated Data Generation at Scale

ResearchDGX agent

arXiv:2506.12009v2 Announce Type: replace Abstract: Affordance grounding aims to localize where to interact with an object, a fundamental capability for embodied agents. Yet progress is bottlenecked b

AIEWF Daily Dispatch: Autoresearch and the tension between AI and human agency

ToolsDGX agent

This dispatch examines the emerging capabilities and challenges of AI autoresearch systems, analyzing the tension between autonomous AI agents conducting research and the preservation of human agency

Been reading all sorts of posts about the best ways to develop workflows for Fable and it reminds me of how little we actually know about th…

ApplicationsDGX agent

Been reading all sorts of posts about the best ways to develop workflows for Fable and it reminds me of how little we actually know about the best ways to organize work for long-running agents. Nobody

DART-VLN: Test-Time Memory Decay and Anti-Loop Regularization for Discrete Vision-Language Navigation

ResearchDGX agent

arXiv:2607.01043v1 Announce Type: cross Abstract: Memory-based discrete vision-language navigation (VLN) agents must act under partial observability, yet even strong frozen backbones remain vulnerable

Decentralized Geometric Control for Cable-Suspended Payload Transport with Adaptive Mass Estimation

SafetyDGX agent

arXiv:2607.00024v1 Announce Type: new Abstract: Cooperative aerial transport requires controllers that respect nonlinear manifold geometry, operate without centralized coordination, and respect operat

Distributed Online Bandit Submodular Maximization with Bounded Sampling Violations

ResearchDGX agent

arXiv:2607.00680v1 Announce Type: new Abstract: We study distributed online submodular maximization under partition matroid constraints, in which multiple agents select a limited number of actions fro

Dual-Informed Vertical Expansion for Multi-Objective Node Selection in Anytime Conflict-Based Search

SafetyDGX agent

arXiv:2607.00156v1 Announce Type: new Abstract: Conflict-Based Search (CBS) is a leading exact algorithm for Multi-Agent Path Finding (MAPF), but its high-level node-selection rule is usually treated

EgoSafetyBench: A Diagnostic Egocentric Video Benchmark for Evaluating Embodied VLMs as Runtime Safety Guards

Model ReleasesDGX agent

arXiv:2607.00218v1 Announce Type: cross Abstract: Vision-language models (VLMs) are now proposed as runtime safety guards for embodied agents in homes and factories. A deployable guard must catch genu

From Holistic Evaluation to Structured Criteria: Rubrics Across the Evolving LLM Landscape

SafetyDGX agent

arXiv:2606.08625v2 Announce Type: replace Abstract: As Large Language Models (LLMs) advance toward open-ended autonomous agents, the mechanisms used to evaluate and guide their behavior must evolve ac

huge week of really exciting launches at langchain!

Model ReleasesDGX agent

huge week of really exciting launches at langchain! big week at langchain, with a lot of launches: 1/ OpenWiki - auto generate a wiki of a github repo 2/ two different voice agent tutorials 3/ Harbor

OpenWiki is designed to run in the background, without you needing to think about it It'll generate docs, update your AGENTS.md so your agen…

Local AiDGX agent

OpenWiki is designed to run in the background, without you needing to think about it It'll generate docs, update your AGENTS.md so your agent automatically knows how to read the docs, and update itsel

RetailSMV: Exocentric vs. Egocentric Adaptation of Foundation Video World Models in Retail

Model ReleasesDGX agent

arXiv:2607.00310v1 Announce Type: cross Abstract: Foundation video diffusion models are increasingly viewed as world simulators for embodied agents, yet their pretraining on internet-scale generic vid

TRACE: State-Aware Query Processing over Temporal Evidence Graphs for Conversational Data

ResearchDGX agent

arXiv:2607.00339v1 Announce Type: new Abstract: Conversational data is increasingly used as a persistent source of user state for long-running assistants and AI agents. However, querying this data rem

When Classic Cache Policies Fail: Learning-Augmented Replacement for Semantic Retrieval Buffers

ResearchDGX agent

arXiv:2607.00394v1 Announce Type: cross Abstract: LLM agents increasingly rely on retrieval buffers to store and reuse past experience, yet the cache management policies governing these buffers remain

1 Jul 2026

AION: Aerial Indoor Object-Goal Navigation Using Dual-Policy Reinforcement Learning

Model ReleasesDGX agent

arXiv:2601.15614v3 Announce Type: replace Abstract: Object-Goal Navigation (ObjectNav) requires an agent to autonomously explore an unknown environment and navigate toward target objects specified by

Expected Gain-based Escalation in Vertical Federated Learning

ResearchDGX agent

arXiv:2606.31331v1 Announce Type: new Abstract: Collaborative inference can improve predictive performance by integrating complementary information across agents, but applying collaborative fusion to

GUI-AIMA: Aligning Intrinsic Multimodal Attention with a Context Anchor for GUI Grounding

ResearchDGX agent

arXiv:2511.00810v4 Announce Type: replace-cross Abstract: Graphical user interface (GUI) grounding is a key capability for computer-use agents, mapping natural-language instructions to actionable regi

MVP-Nav: Multi-layer Value Map Planner Navigator

TutorialsDGX agent

arXiv:2606.31919v1 Announce Type: cross Abstract: Zero-shot Object Goal Navigation (ZSON) with RGB-only perception poses a fundamental challenge for embodied agents, as the absence of explicit depth i

One Reflection Is Not Enough: Self-Correcting Autonomous Research via Multi-Hypothesis Failure Attribution

Model ReleasesDGX agent

arXiv:2606.31478v1 Announce Type: new Abstract: Autonomous research agents can now draft hypotheses, write code, run experiments, and produce papers, but they remain brittle when experiments fail. Und

Reference-Based Prosody and Rhythm Evaluation for Spoken Dialogue Systems

ResearchDGX agent

arXiv:2606.31055v1 Announce Type: new Abstract: Speech-to-speech (S2S) AI agents are advancing rapidly, yet evaluation lacks interpretable speech-native measures for conversational prosody and rhythm.

Sampling-Based Coordination-Informed Multi-Objective Multi-Robot Reinforcement Learning

SafetyDGX agent

arXiv:2606.30893v1 Announce Type: new Abstract: Multi-robot systems must simultaneously optimize competing objectives while maintaining coordinated behavior. Existing multi-agent reinforcement learnin

The HydroGym Reinforcement Learning Platform for Fluid Dynamics

ResearchDGX agent

arXiv:2512.17534v2 Announce Type: replace-cross Abstract: Modeling and controlling fluids is critical across science and engineering. Effective flow control can increase lift, reduce drag, enhance mix

The Past Is Prologue: A Plug-in Controller for Selective Updates in Sequentially Evolving LLM Memory

SafetyDGX agent

arXiv:2606.31121v1 Announce Type: new Abstract: Sequentially evolving LLM memory enables agents to reuse past experience, but existing systems usually deploy each locally generated memory update witho

Who Determines the Meaning of an Emotion? Affective Sovereignty as an Epistemic Consequence of Measurement Limits

ResearchDGX agent

arXiv:2606.31442v1 Announce Type: new Abstract: Emotion-sensing AI is rapidly becoming embedded in vehicles, home appliances, dialogue agents, and social infrastructure, giving rise to a sphere in whi

30 Jun 2026

Accelerating scientific discovery with Co-Scientist

Model ReleasesDGX agent

arXiv:2502.18864v2 Announce Type: replace Abstract: Scientific discovery is driven by scientists generating novel hypotheses for complex problems that undergo rigorous experimental validation. To augm

CLOSER-VLN: Closed-Loop Self-Verified Retrieval-Augmented Reasoning for Aerial Vision-Language Navigation

Model ReleasesDGX agent

arXiv:2606.28397v1 Announce Type: cross Abstract: Vision-language navigation (VLN) has recently advanced with large language and multimodal models, enabling agents to follow natural-language instructi

Deterministic Decisions for High-Stakes AI. A Zero-Egress Pipeline with the Deployability of RAG and the Accuracy of Machine Learning

SafetyDGX agent

arXiv:2606.29280v1 Announce Type: cross Abstract: We identify intervention bias as a previously unquantified failure mode of zero-shot large-language-model (LLM) educational advisory agents: without t

EVAF: A Test-Retest Protocol for Selective Parametric Consolidation

Model ReleasesDGX agent

arXiv:2606.29916v1 Announce Type: cross Abstract: Long-running language agents need mechanisms for deciding which experiences should persist after the working context is gone. Retrieval systems can re

Exploring the Cryptographic Limits of Transformer Networks

ResearchDGX agent

arXiv:2606.29389v1 Announce Type: cross Abstract: In recent work it has been shown that colluding AI agents can use steganographic methods to exchange malicious information. Whether a transformer can

FutureNav: Unified World-Action Modeling for Vision-and-Language Navigation

SafetyDGX agent

arXiv:2606.30367v1 Announce Type: new Abstract: Vision-and-language navigation (VLN) in continuous environments requires an agent to ground instructions in egocentric observations while maintaining sp

Google’s Gemini Omni Flash and Nano Banana 2 Lite support slick media content creation at lower costs

Model ReleasesDGX agent

Google LLC is enhancing its generative artificial intelligence capabilities for creators with the debut of a pair of new media-focused models in the Gemini Enterprise Agent Platform. The new additions

Hierarchical Decision Making with Structured Policies: A Principled Design via Inverse Optimization

SafetyDGX agent

arXiv:2606.28764v1 Announce Type: new Abstract: Hierarchical decision-making frameworks are pivotal for addressing complex control tasks, enabling agents to decompose intricate problems into manageabl

Internal-State Probes Read the Situation, Not the Action: Three Negative Results for Pre-Action Misalignment Monitoring

Model ReleasesDGX agent

arXiv:2606.30449v1 Announce Type: new Abstract: Probes on model internals could help monitor agentic systems if they identify harmful text or tool actions before those actions are generated. We ask wh

Learned Coordination Conventions in Cooperative MARL: Measuring the Translation Gap Between Theory-Informed Roles and Learned Routing

SafetyDGX agent

arXiv:2606.29541v1 Announce Type: new Abstract: Role-semantic assignments provide priors over how heterogeneous agents may coordinate, but cooperative MARL systems instead settle on conventions throug

Lost in Execution: On the Multilingual Robustness of Tool Calling in Large Language Models

Model ReleasesDGX agent

arXiv:2601.05366v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly deployed as agents that invoke external tools through structured function calls. While recent wo

MirrorCode: AI can rebuild entire programs from behavior alone

Model ReleasesDGX agent

arXiv:2606.30182v1 Announce Type: new Abstract: AI models are rapidly improving at autonomous coding, as shown by benchmark progress and one-off demonstrations such as AI implementing a C compiler. Ho

MIThinker: A Plug-and-Play Policy-Optimized Thinker For Motivational Interviewing Counseling

SafetyDGX agent

arXiv:2606.29265v1 Announce Type: new Abstract: Reasoning large language models (LLMs) have recently made much progress in complex problem-solving, leveraging internal reasoning (or thought) to guide

Modernizing financial services with deployment freedom and transformational AI with AlloyDB Omni

Local AiDGX agent

The financial services industry (FSI) operates under a unique set of non-negotiable requirements: the need for strict regulatory compliance, sub-millisecond transactional speeds, and security that ver

Modification-Considering Value Learning for Reward Hacking Mitigation in RL

SafetyDGX agent

arXiv:2606.28955v1 Announce Type: cross Abstract: Reinforcement learning agents can exploit misspecified reward signals to achieve high apparent returns while failing on the intended objective, a fail

Optimizing Expert-Designed Crystal Graph Networks for Band-Gap Prediction with an Autonomous LLM Research Loop

Model ReleasesDGX agent

arXiv:2606.29717v1 Announce Type: cross Abstract: Predicting a material's properties from its structure is a central, fast-advancing problem in computational materials science. A decade of work has pr

RoamFlow: Reinforcement-Aligned One-Step Action MeanFlow Policy for Image-Goal Navigation

SafetyDGX agent

arXiv:2606.29934v1 Announce Type: new Abstract: Image-goal navigation is a key challenge in embodied robotics, where an agent must reach a target specified solely by a goal image. While existing reinf

SABER-Math: Automated Benchmark for Information Retrieval Evaluation in Mathematics

Model ReleasesDGX agent

arXiv:2606.29894v1 Announce Type: cross Abstract: As agentic AI systems tackle more complex mathematical tasks, they increasingly rely on information retrieval (IR) to search problem databases, theore

Safety from Honesty in a Disinterested AI Predictor

SafetyDGX agent

arXiv:2606.29657v1 Announce Type: new Abstract: As AI systems become more capable, training procedures that optimize for downstream outcomes risk introducing implicit agency: goal-directed behavior th

Self-Supervised Theorem Discovery in a Formal Axiomatic System

Model ReleasesDGX agent

arXiv:2606.28747v1 Announce Type: new Abstract: Recent artificial intelligence (AI) systems have shown remarkable progress in mathematical reasoning. Many existing approaches, including large language

The Download: AI “coworkers” and stratospheric internet

TutorialsDGX agent

This is today’s edition of The Download, our weekday newsletter that provides a daily dose of what’s going on in the world of technology. AI agents are not your “coworkers” Imagine coming in to work t

TRACE: Temporal Relationship-Aware Conversational Entrainment Detection in Dyadic Speech

ResearchDGX agent

arXiv:2606.30543v1 Announce Type: cross Abstract: With the proliferation of speech AI agents, understanding emotional entrainment in conversational interaction has become increasingly important. Emoti

← Previous
1…227228229230231…297
Next →