AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,959 results
11 Jun 2026

LLMs+Graphs: Toward Graph-Native, Synergistic AI Systems

AgentsDGX agent

arXiv:2606.11560v1 Announce Type: cross Abstract: Large Language Models (LLMs) have advanced rapidly, but their limitations in structured and multi-hop reasoning underscore the need for graph-native,

10 Jun 2026

Beyond Memorization: Distinguishing Between Pattern-Based and Epistemic Reasoning in LLMs Using Epistemic Puzzles

Model ReleasesDGX agent

arXiv:2603.21350v2 Announce Type: replace Abstract: Epistemic reasoning requires agents to infer the state of the world from partial observations and information about other agents' knowledge. Prior w

Deep dive: How Lightning Engine delivers 4.9x faster Apache Spark performance

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents
DGX agent

From foundational ETL and analytics to the frontier of generative AI, Apache Spark serves as the architectural backbone for global data processing. However, as data volumes scale, the trade-off betwee

Mobility Anomaly Generation using LLM-Driven Behavior with Kinematic Constraints

AgentsDGX agent

arXiv:2606.10314v1 Announce Type: new Abstract: Although the study of human trajectory anomalies is critical for advancing spatial data mining, empirical research remains severely hindered by a pervas

Trace Only What You Need: Structure-Aware On-Demand Hypergraph Memory for Long-Document Question Answering

AgentsDGX agent

arXiv:2606.10921v1 Announce Type: new Abstract: Long-document question answering (QA) requires large language models (LLMs) to reason over evidence scattered across lengthy documents, where answers of

9 Jun 2026

An Information-Theoretic Definition for Open-Ended Learning

AgentsDGX agent

arXiv:2606.08369v1 Announce Type: cross Abstract: A growing body of work points to the great promise of AI systems that can continually expand their capabilities as they operate in an open-ended envir

CodeTaste: Can LLMs Generate Human-Level Code Refactorings?

Model ReleasesDGX agent

arXiv:2603.04177v2 Announce Type: replace-cross Abstract: LLM coding agents can generate working code, but their solutions often accumulate complexity, duplication, and architectural debt. Human devel

One if by Land, Two if by Sea, Three if by Four Seas, and More to Come -- Values of Perception, Prediction, Communication, and Common Sense in Decision Making

AgentsDGX agent

arXiv:2601.06077v2 Announce Type: replace-cross Abstract: This work aims to rigorously define the values of perception, prediction, communication, and common sense in decision making. The defined quan

SPIN: Decentralized Swarm Control via Tensorized Policy Coordination

SafetyDGX agent

arXiv:2606.07557v1 Announce Type: new Abstract: Decentralized multi-agent swarm coordination on resource-constrained edge platforms remains fundamentally bottlenecked by the exponential scaling of joi

TianJi-Environ: An Autonomous AI Scientist for Atmospheric Environmental Research

AgentsDGX agent

arXiv:2606.07697v1 Announce Type: cross Abstract: As atmospheric environmental prediction continues to improve, interpretable validation of pollution mechanisms and feedback processes has become a mai

When Video Misreads: Closed-Loop Distillation of Reading Heuristics for Exploratory Manipulation Trace QA

AgentsDGX agent

arXiv:2606.08542v1 Announce Type: cross Abstract: Exploratory manipulation often turns an apparent failed attempt into the key evidence for what to do next. For example, a robot pulls a locked cabinet

8 Jun 2026

AnchorWorld: Embodied Egocentric World Simulation with View-based Evolution Customization

AgentsDGX agent

arXiv:2606.07326v1 Announce Type: new Abstract: Despite being a pivotal frontier, interactive world modeling remains underexplored in terms of the versatile controllability required by practical scena

Efficient Coordination and Synchronization of Multi-Robot Systems Under Recurring Linear Temporal Logic

AgentsDGX agent

arXiv:2502.16531v2 Announce Type: replace Abstract: We consider multi-robot systems under recurring tasks formalized as linear temporal logic (LTL) specifications. To solve the planning problem effici

It's finally out!!! @METR_Evals found that more than half of SWEBench results is unmergeable slop. FrontierCode represents over 1000+ hours …

AgentsDGX agent

It's finally out!!! @METR_Evals found that more than half of SWEBench results is unmergeable slop. FrontierCode represents over 1000+ hours of maintainer validated software engineering work most front

two things ready to share from this weekend: 📖 http://learn.activegraph.ai interactive site teaching activegraph concepts blog: https://act…

AgentsDGX agent

two things ready to share from this weekend: 📖 http://learn.activegraph.ai interactive site teaching activegraph concepts blog: https://activegraph.ai/blog/introducing-learn 💻 AG coder (open source) r

6 Jun 2026

Beyond Rewards in Reinforcement Learning for Cyber Defence

SafetyDGX agent

arXiv:2602.04809v3 Announce Type: replace-cross Abstract: Recent years have seen an explosion of interest in autonomous cyber defence agents trained to defend computer networks using deep reinforcemen

New research from Renmin University. Treat skill selection as a harness in its own right. If you design skill routing for personal or edge a…

Local AiDGX agent

New research from Renmin University. Treat skill selection as a harness in its own right. If you design skill routing for personal or edge agents, this work argues that the selection layer is a first-

No frontier lab will own every single point on the pareto frontier around cost/latency and accuracy. Even as the pareto frontier itself adva…

AgentsDGX agent

No frontier lab will own every single point on the pareto frontier around cost/latency and accuracy. Even as the pareto frontier itself advances, there will always be points owned by open-weight model

5 Jun 2026

Ask-to-Clarify: Resolving Instruction Ambiguity through Multi-turn Dialogue

ApplicationsDGX agent

arXiv:2509.15061v2 Announce Type: cross Abstract: The ultimate goal of embodied agents is to create collaborators that can interact with humans, not mere executors that passively follow instructions.

CollabBench: Benchmarking and Unleashing Collaborative Ability of LLMs with Diverse Players via Proactive Engagement

Model ReleasesDGX agent

arXiv:2606.05793v1 Announce Type: new Abstract: While LLM-based agents excel at individual tasks, effective collaboration with realistic human partners remains challenging. Most of the existing conver

Resonant Minds: Closed-Loop Social Avatars with Theory of Mind

AgentsDGX agent

arXiv:2606.05896v1 Announce Type: new Abstract: Creating lifelike digital humans with genuine social intelligence requires unifying cognitive reasoning and multimodal generation within a coherent fram

Staying with the Uncertainty: Uncertainty-Scaffolding Strategies for Artificial Moral Advisors in LLM-to-LLM Simulated Conversations

AgentsDGX agent

arXiv:2606.05890v1 Announce Type: new Abstract: LLMs are increasingly deployed as Artificial Moral Advisors (AMA) in a variety of contexts: what kind of conversational patterns should they display? In

4 Jun 2026

Fireworks was named to @Redpoint's InfraRed 100 which recognizes the companies building the foundation for the next wave of AI. We're just g…

AgentsDGX agent

Fireworks was named to @Redpoint's InfraRed 100 which recognizes the companies building the foundation for the next wave of AI. We're just getting started. Come build with us: https://fireworks.ai/car

NVIDIA Nemotron 3 Ultra is on Fireworks, day zero. Nemotron Ultra is an open model for frontier reasoning and orchestration in long-running …

Model ReleasesDGX agent

NVIDIA Nemotron 3 Ultra is on Fireworks, day zero. Nemotron Ultra is an open model for frontier reasoning and orchestration in long-running autonomous agents. Think use cases like coding agents, deep

we're taking an early bet on open models specifically, because they're SO much cheaper one point of reference: an app outputting 10M tokens/…

AgentsDGX agent

we're taking an early bet on open models specifically, because they're SO much cheaper one point of reference: an app outputting 10M tokens/day costs roughly 250/day on Opus 4.8 versus ~24/day for Min

3 Jun 2026

A 3D Isovist World Model -- Revealing a City's Unseen Geometry and Its Emergent Cross-City Signature

Model ReleasesDGX agent

arXiv:2606.03609v1 Announce Type: cross Abstract: Embodied agents that navigate cities rely on world models that predict how their surroundings will change as they move. But for navigation, what matte

Any2Poster: Any-Source Poster Generation Across Modalities and Domains

Model ReleasesDGX agent

arXiv:2606.02915v1 Announce Type: new Abstract: Visual posters are a compact medium for communicating dense information, yet progress on automatic poster generation remains difficult to measure becaus

Ask When It Pays: Cost-Aware Open-Ended Interaction for Instance Goal Navigation

Model ReleasesDGX agent

arXiv:2606.03175v1 Announce Type: new Abstract: Instance Goal Navigation (IGN) requires an embodied agent to find a specific object instance among distractors from an underspecified natural-language d

CodeHacker: Automated Test Case Generation for Detecting Vulnerabilities in Competitive Programming Solutions

AgentsDGX agent

arXiv:2602.20213v2 Announce Type: replace-cross Abstract: The evaluation of Large Language Models (LLMs) for code generation relies heavily on the quality and robustness of test cases. However, existi

SkillDAG: Self-Evolving Typed Skill Graphs for LLM Skill Selection at Scale

Model ReleasesDGX agent

arXiv:2606.03056v1 Announce Type: new Abstract: As LLM agents adopt large skill libraries, selecting the right subset becomes a structural problem rather than a similarity-matching one: skills depend

2 Jun 2026

Beyond Scalar Rewards: Dense Feedback for LLM Policy Synthesis in Sequential Social Dilemmas

Model ReleasesDGX agent

arXiv:2603.19453v2 Announce Type: replace Abstract: We study LLM policy synthesis: using a language model to iteratively generate programmatic agent policies for multi-agent environments. Rather than

Claudini: Autoresearch Discovers State-of-the-Art Adversarial Attack Algorithms for LLMs

Model ReleasesDGX agent

arXiv:2603.24511v2 Announce Type: replace-cross Abstract: We show that AI agents are capable of discovering novel algorithms for adversarial attacks against LLMs, advancing the state of the art on whi

Don't Ask the LLM to Track Freshness: A Deterministic Recipe for Memory Conflict Resolution

AgentsDGX agent

arXiv:2606.01435v1 Announce Type: new Abstract: LLM-based memory systems increasingly maintain facts that evolve over time, where a recurring failure is conflict resolution: when a fact has multiple c

EvoPool: Evolutionary Programmatic Annotation for Label-Efficient Specialized Supervision

AgentsDGX agent

arXiv:2606.01617v1 Announce Type: cross Abstract: Large language models excel at general tasks but underperform smaller supervised models in specialized, high-stakes domains where training labels are

I-WebGenBench : Evaluating Interactivity in LLM-Generated Scientific Web Applications

Model ReleasesDGX agent

arXiv:2606.00750v1 Announce Type: new Abstract: Recent advances in visual language models have enabled autonomous agents for complex reasoning, tool use, and document understanding. However, existing

Learning Query-Specific Rubrics from Human Preferences for DeepResearch Report Generation

AgentsDGX agent

arXiv:2602.03619v2 Announce Type: replace Abstract: Nowadays, developing reliable DeepResearch-style long-form report generation remains challenging, as training and evaluation lack verifiable reward

Market-Based Replanning for Safety-Critical UAV Swarms in Search and Rescue Missions

SafetyDGX agent

arXiv:2606.01970v1 Announce Type: new Abstract: Reliable autonomous UAV swarms in Search and Rescue (SAR) missions require fault-tolerant coordination capable of sustaining operations despite agent de

PaperVoyager : Building Interactive Web with Visual Language Models

Model ReleasesDGX agent

arXiv:2603.22999v3 Announce Type: replace Abstract: Recent advances in visual language models have enabled autonomous agents for complex reasoning, tool use, and document understanding. However, exist

Quick explainer of our managed deepagents offering

AgentsDGX agent

This post likely provides a brief explanation of a managed deepagents offering, describing what deepagents are, how the managed service works, and its key benefits or use cases. As a post from Harriso

Situation-Aware Interactive MPC Switching for Autonomous Driving

AgentsDGX agent

arXiv:2512.06182v2 Announce Type: replace Abstract: Autonomous driving in interactive traffic scenarios remains challenging because of the mutual influence among vehicles and the inherent uncertainty

🌞This is big Local AI news! A new open-source Computer-Use LLM has just launched. Holo 3.1 is H Company’s (🇫🇷) new local computer-use age…

Model ReleasesDGX agent

🌞This is big Local AI news! A new open-source Computer-Use LLM has just launched. Holo 3.1 is H Company’s (🇫🇷) new local computer-use agent model that beats Qwen3.5-397B, Kimi-K2.5, and Sonnet 4.6! Si

Towards a General Intelligence and Interface for Wearable Health Data

AgentsDGX agent

arXiv:2605.22759v2 Announce Type: replace Abstract: While ubiquitous wearable sensors capture a wealth of behavioral and physiological information, effectively transforming these signals into personal

Visual Persuasion: What Influences Decisions of Vision-Language Models?

SafetyDGX agent

arXiv:2602.15278v2 Announce Type: replace-cross Abstract: The web is littered with images, once created for human consumption and now increasingly interpreted by agents using vision-language models (V

When Knowledge Is Not Free: Cost-Aware Evidence Selection in Retrieval-Augmented Generation

AgentsDGX agent

arXiv:2606.02245v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) typically assumes that external knowledge is free, but many high-quality sources are paywalled, licensed, restricte

1 Jun 2026

Batched Stochastic Linear Bandits with 1-Bit Communication Constraints

AgentsDGX agent

arXiv:2605.30976v1 Announce Type: cross Abstract: We study stochastic linear bandits under a natural combination of batching and communication constraints: the time horizon is partitioned into batches

Comparing LLM-Based Conversational and Graphical Interfaces for Industrial Decision Tasks: An Exploratory Mixed-Methods Study

AgentsDGX agent

arXiv:2605.31224v1 Announce Type: cross Abstract: The use of Generative AI Conversational User Interfaces (CUI) as a new way to access and analyze data is growing in all sectors, and the industrial on

How Trustpilot built a real-time architecture for data enrichment using Gemma

Model ReleasesDGX agent

Processing millions of user reviews in real-time, under strict latency and cost constraints, is no easy task. Trustpilot has been doing exactly that with custom machine learning since long before larg

Social welfare optimisation under institutional reward and punishment

Model ReleasesDGX agent

arXiv:2605.31330v1 Announce Type: cross Abstract: Institutional incentives are widely used to promote cooperation among autonomous, self-regarding agents, from human societies to multi-agent and AI sy

SVI-Bench: A Dynamic Microworld for Strategic Video Intelligence

Model ReleasesDGX agent

arXiv:2605.31529v1 Announce Type: new Abstract: True video intelligence demands more than recognizing what is visible: it requires reasoning about why events unfold, predicting what would change under

Uncertainty-Aware and Temporally Regulated Expert Advice in Reinforcement Learning for Autonomous Driving

SafetyDGX agent

arXiv:2605.30576v1 Announce Type: new Abstract: Exploration in reinforcement learning for autonomous driving is inherently unsafe: agents must experience novel behaviors to learn, yet exploration can

29 May 2026

Code-QA-Bench: Separating Code Reasoning from Documentation Memorization in Repository-Level QA

AgentsDGX agent

arXiv:2605.29277v1 Announce Type: cross Abstract: We present Code-QA-Bench, a fully automated framework for synthesizing repository-level code understanding benchmarks that separates genuine code comp

CompilerDream: Learning a Compiler World Model for General Code Optimization

AgentsDGX agent

arXiv:2404.16077v4 Announce Type: replace-cross Abstract: Effective code optimization in compilers is crucial for computer and software engineering. The success of these optimizations primarily depend

Croissant Tasks: A Metadata Format for Reproducible Machine Learning Evaluations

AgentsDGX agent

arXiv:2605.29786v1 Announce Type: new Abstract: Reproducibility is fundamental to the scientific method, yet remains a critical challenge in machine learning. Contributing factors include underspecifi

Discovering Cooperative Pipelines: Autoresearch for Sequential Social Dilemmas

SafetyDGX agent

arXiv:2605.30003v1 Announce Type: cross Abstract: We study two-level autoresearch for cooperation: an outer-loop AI agent autonomously redesigns the inner-loop pipeline of an LLM policy-synthesis syst

Error as a Lens: Probing LLM Reasoning through Synthetic Misconception Generation

AgentsDGX agent

arXiv:2605.29007v1 Announce Type: new Abstract: Personalized tutoring, teacher training, and education research need access to targeted synthetic misconceptions, but privacy and IRB constraints make l

PhyGenHOI: Physically-Aware 4D Generation of Dynamic Human-Object Interactions

AgentsDGX agent

arXiv:2605.30268v1 Announce Type: cross Abstract: We address the task of generating physically accurate and visually faithful 4D Human-Object Interaction (HOI). Given a static 3D human and target obje

Training Deliberative Monitors for Black-Box Scheming Detection

Model ReleasesDGX agent

arXiv:2605.29601v1 Announce Type: cross Abstract: As autonomous agents become more capable of performing real-world tasks, distinguishing scheming behavior from benign task pursuit may become a centra

unix-ctf: Procedural Environments for Unix-Competence Reinforcement Learning

AgentsDGX agent

arXiv:2605.29115v1 Announce Type: cross Abstract: Unix competence is the ability to use shell and operating-system primitives as first-class tools, not merely to write programs through a terminal. Cur

28 May 2026

An LLM-Based Assistance System for Intuitive and Flexible Capability-Based Planning

AgentsDGX agent

arXiv:2605.28666v1 Announce Type: new Abstract: In modern industry, dynamic environments and the complexity of modular and reconfigurable resources require automated planning of process sequences. Cap

Falsification-driven reinforcement learning for maritime motion planning

AgentsDGX agent

arXiv:2510.06970v2 Announce Type: replace-cross Abstract: Compliance with maritime traffic rules is essential for the safe operation of autonomous vessels, yet training reinforcement learning (RL) age

← Previous
1…176177178179180…300
Next →