AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,973 results
Model Releases

Any2Poster: Any-Source Poster Generation Across Modalities and Domains

DGX agent

arXiv:2606.02915v1 Announce Type: new Abstract: Visual posters are a compact medium for communicating dense information, yet progress on automatic poster generation remains difficult to measure becaus

model-releasesarxiv-cs-cv
3 Jun 2026
Model Releases
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Ask When It Pays: Cost-Aware Open-Ended Interaction for Instance Goal Navigation

DGX agent

arXiv:2606.03175v1 Announce Type: new Abstract: Instance Goal Navigation (IGN) requires an embodied agent to find a specific object instance among distractors from an underspecified natural-language d

model-releasesarxiv-cs-cv
3 Jun 2026
Agents

CodeHacker: Automated Test Case Generation for Detecting Vulnerabilities in Competitive Programming Solutions

DGX agent

arXiv:2602.20213v2 Announce Type: replace-cross Abstract: The evaluation of Large Language Models (LLMs) for code generation relies heavily on the quality and robustness of test cases. However, existi

agentsarxiv-cs-ai
3 Jun 2026
Model Releases

SkillDAG: Self-Evolving Typed Skill Graphs for LLM Skill Selection at Scale

DGX agent

arXiv:2606.03056v1 Announce Type: new Abstract: As LLM agents adopt large skill libraries, selecting the right subset becomes a structural problem rather than a similarity-matching one: skills depend

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

Beyond Scalar Rewards: Dense Feedback for LLM Policy Synthesis in Sequential Social Dilemmas

DGX agent

arXiv:2603.19453v2 Announce Type: replace Abstract: We study LLM policy synthesis: using a language model to iteratively generate programmatic agent policies for multi-agent environments. Rather than

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Claudini: Autoresearch Discovers State-of-the-Art Adversarial Attack Algorithms for LLMs

DGX agent

arXiv:2603.24511v2 Announce Type: replace-cross Abstract: We show that AI agents are capable of discovering novel algorithms for adversarial attacks against LLMs, advancing the state of the art on whi

model-releasesarxiv-cs-ai
2 Jun 2026
Agents

Don't Ask the LLM to Track Freshness: A Deterministic Recipe for Memory Conflict Resolution

DGX agent

arXiv:2606.01435v1 Announce Type: new Abstract: LLM-based memory systems increasingly maintain facts that evolve over time, where a recurring failure is conflict resolution: when a fact has multiple c

agentsarxiv-cs-ai
2 Jun 2026
Agents

EvoPool: Evolutionary Programmatic Annotation for Label-Efficient Specialized Supervision

DGX agent

arXiv:2606.01617v1 Announce Type: cross Abstract: Large language models excel at general tasks but underperform smaller supervised models in specialized, high-stakes domains where training labels are

agentsarxiv-cs-ai
2 Jun 2026
Model Releases

I-WebGenBench : Evaluating Interactivity in LLM-Generated Scientific Web Applications

DGX agent

arXiv:2606.00750v1 Announce Type: new Abstract: Recent advances in visual language models have enabled autonomous agents for complex reasoning, tool use, and document understanding. However, existing

model-releasesarxiv-cs-cl
2 Jun 2026
Agents

Learning Query-Specific Rubrics from Human Preferences for DeepResearch Report Generation

DGX agent

arXiv:2602.03619v2 Announce Type: replace Abstract: Nowadays, developing reliable DeepResearch-style long-form report generation remains challenging, as training and evaluation lack verifiable reward

agentsarxiv-cs-cl
2 Jun 2026
Safety

Market-Based Replanning for Safety-Critical UAV Swarms in Search and Rescue Missions

DGX agent

arXiv:2606.01970v1 Announce Type: new Abstract: Reliable autonomous UAV swarms in Search and Rescue (SAR) missions require fault-tolerant coordination capable of sustaining operations despite agent de

safetyarxiv-cs-ro
2 Jun 2026
Model Releases

PaperVoyager : Building Interactive Web with Visual Language Models

DGX agent

arXiv:2603.22999v3 Announce Type: replace Abstract: Recent advances in visual language models have enabled autonomous agents for complex reasoning, tool use, and document understanding. However, exist

model-releasesarxiv-cs-cl
2 Jun 2026
Agents

Quick explainer of our managed deepagents offering

DGX agent

This post likely provides a brief explanation of a managed deepagents offering, describing what deepagents are, how the managed service works, and its key benefits or use cases. As a post from Harriso

agentsharrison-chase--x
2 Jun 2026
Agents

Situation-Aware Interactive MPC Switching for Autonomous Driving

DGX agent

arXiv:2512.06182v2 Announce Type: replace Abstract: Autonomous driving in interactive traffic scenarios remains challenging because of the mutual influence among vehicles and the inherent uncertainty

agentsarxiv-cs-ro
2 Jun 2026
Model Releases

🌞This is big Local AI news! A new open-source Computer-Use LLM has just launched. Holo 3.1 is H Company’s (🇫🇷) new local computer-use age…

DGX agent

🌞This is big Local AI news! A new open-source Computer-Use LLM has just launched. Holo 3.1 is H Company’s (🇫🇷) new local computer-use agent model that beats Qwen3.5-397B, Kimi-K2.5, and Sonnet 4.6! Si

model-releasesclem-delangue--x
2 Jun 2026
Agents

Towards a General Intelligence and Interface for Wearable Health Data

DGX agent

arXiv:2605.22759v2 Announce Type: replace Abstract: While ubiquitous wearable sensors capture a wealth of behavioral and physiological information, effectively transforming these signals into personal

agentsarxiv-cs-ai
2 Jun 2026
Safety

Visual Persuasion: What Influences Decisions of Vision-Language Models?

DGX agent

arXiv:2602.15278v2 Announce Type: replace-cross Abstract: The web is littered with images, once created for human consumption and now increasingly interpreted by agents using vision-language models (V

safetyarxiv-cs-ai
2 Jun 2026
Agents

When Knowledge Is Not Free: Cost-Aware Evidence Selection in Retrieval-Augmented Generation

DGX agent

arXiv:2606.02245v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) typically assumes that external knowledge is free, but many high-quality sources are paywalled, licensed, restricte

agentsarxiv-cs-cl
2 Jun 2026
Agents

Batched Stochastic Linear Bandits with 1-Bit Communication Constraints

DGX agent

arXiv:2605.30976v1 Announce Type: cross Abstract: We study stochastic linear bandits under a natural combination of batching and communication constraints: the time horizon is partitioned into batches

agentsarxiv-cs-lg
1 Jun 2026
Agents

Comparing LLM-Based Conversational and Graphical Interfaces for Industrial Decision Tasks: An Exploratory Mixed-Methods Study

DGX agent

arXiv:2605.31224v1 Announce Type: cross Abstract: The use of Generative AI Conversational User Interfaces (CUI) as a new way to access and analyze data is growing in all sectors, and the industrial on

agentsarxiv-cs-ai
1 Jun 2026
Model Releases

How Trustpilot built a real-time architecture for data enrichment using Gemma

DGX agent

Processing millions of user reviews in real-time, under strict latency and cost constraints, is no easy task. Trustpilot has been doing exactly that with custom machine learning since long before larg

model-releasesgoogle-cloud-ai
1 Jun 2026
Model Releases

Social welfare optimisation under institutional reward and punishment

DGX agent

arXiv:2605.31330v1 Announce Type: cross Abstract: Institutional incentives are widely used to promote cooperation among autonomous, self-regarding agents, from human societies to multi-agent and AI sy

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

SVI-Bench: A Dynamic Microworld for Strategic Video Intelligence

DGX agent

arXiv:2605.31529v1 Announce Type: new Abstract: True video intelligence demands more than recognizing what is visible: it requires reasoning about why events unfold, predicting what would change under

model-releasesarxiv-cs-cv
1 Jun 2026
Safety

Uncertainty-Aware and Temporally Regulated Expert Advice in Reinforcement Learning for Autonomous Driving

DGX agent

arXiv:2605.30576v1 Announce Type: new Abstract: Exploration in reinforcement learning for autonomous driving is inherently unsafe: agents must experience novel behaviors to learn, yet exploration can

safetyarxiv-cs-ai
1 Jun 2026
Agents

Code-QA-Bench: Separating Code Reasoning from Documentation Memorization in Repository-Level QA

DGX agent

arXiv:2605.29277v1 Announce Type: cross Abstract: We present Code-QA-Bench, a fully automated framework for synthesizing repository-level code understanding benchmarks that separates genuine code comp

agentsarxiv-cs-ai
29 May 2026
Agents

CompilerDream: Learning a Compiler World Model for General Code Optimization

DGX agent

arXiv:2404.16077v4 Announce Type: replace-cross Abstract: Effective code optimization in compilers is crucial for computer and software engineering. The success of these optimizations primarily depend

agentsarxiv-cs-lg
29 May 2026
Agents

Croissant Tasks: A Metadata Format for Reproducible Machine Learning Evaluations

DGX agent

arXiv:2605.29786v1 Announce Type: new Abstract: Reproducibility is fundamental to the scientific method, yet remains a critical challenge in machine learning. Contributing factors include underspecifi

agentsarxiv-cs-ai
29 May 2026
Safety

Discovering Cooperative Pipelines: Autoresearch for Sequential Social Dilemmas

DGX agent

arXiv:2605.30003v1 Announce Type: cross Abstract: We study two-level autoresearch for cooperation: an outer-loop AI agent autonomously redesigns the inner-loop pipeline of an LLM policy-synthesis syst

safetyarxiv-cs-ai
29 May 2026
Agents

Error as a Lens: Probing LLM Reasoning through Synthetic Misconception Generation

DGX agent

arXiv:2605.29007v1 Announce Type: new Abstract: Personalized tutoring, teacher training, and education research need access to targeted synthetic misconceptions, but privacy and IRB constraints make l

agentsarxiv-cs-cl
29 May 2026
Agents

PhyGenHOI: Physically-Aware 4D Generation of Dynamic Human-Object Interactions

DGX agent

arXiv:2605.30268v1 Announce Type: cross Abstract: We address the task of generating physically accurate and visually faithful 4D Human-Object Interaction (HOI). Given a static 3D human and target obje

agentsarxiv-cs-ai
29 May 2026
Model Releases

Training Deliberative Monitors for Black-Box Scheming Detection

DGX agent

arXiv:2605.29601v1 Announce Type: cross Abstract: As autonomous agents become more capable of performing real-world tasks, distinguishing scheming behavior from benign task pursuit may become a centra

model-releasesarxiv-cs-ai
29 May 2026
Agents

unix-ctf: Procedural Environments for Unix-Competence Reinforcement Learning

DGX agent

arXiv:2605.29115v1 Announce Type: cross Abstract: Unix competence is the ability to use shell and operating-system primitives as first-class tools, not merely to write programs through a terminal. Cur

agentsarxiv-cs-ai
29 May 2026
Agents

An LLM-Based Assistance System for Intuitive and Flexible Capability-Based Planning

DGX agent

arXiv:2605.28666v1 Announce Type: new Abstract: In modern industry, dynamic environments and the complexity of modular and reconfigurable resources require automated planning of process sequences. Cap

agentsarxiv-cs-ai
28 May 2026
Agents

Falsification-driven reinforcement learning for maritime motion planning

DGX agent

arXiv:2510.06970v2 Announce Type: replace-cross Abstract: Compliance with maritime traffic rules is essential for the safe operation of autonomous vessels, yet training reinforcement learning (RL) age

agentsarxiv-cs-lg
28 May 2026
Model Releases

Personalized Observation Normalization for Federated Reinforcement Learning in Simulation Environments with Heterogeneity

DGX agent

arXiv:2605.27385v1 Announce Type: cross Abstract: Federated reinforcement learning (FedRL) enables multiple agents to collaboratively train a global policy without sharing raw data, making it ideal fo

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

You Live More Than Once: Towards Hierarchical Skill Meta-Evolving

DGX agent

arXiv:2605.28390v1 Announce Type: new Abstract: Test-time skill evolving is regarded as a new paradigm for enhancing deployed agentic systems. Existing works mainly focus on hard-coded skill evolving

model-releasesarxiv-cs-ai
28 May 2026
Safety

Constrained Meta Reinforcement Learning with Provable Test-Time Safety

DGX agent

arXiv:2601.21845v2 Announce Type: replace Abstract: Meta reinforcement learning (RL) allows agents to leverage experience across a distribution of tasks on which the agent can train at will, enabling

safetyarxiv-cs-lg
27 May 2026
Agents

Cordon-MAS: Defending RAG against Knowledge Poisoning via Information-Flow Control

DGX agent

arXiv:2605.26754v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) increasingly underpins high-stakes applications, yet remains vulnerable to Confundo-style poisoning where adversa

agentsarxiv-cs-ai
27 May 2026
Agents

FoundObj: Self-supervised Foundation Models as Rewards for Label-free 3D Object Segmentation

DGX agent

arXiv:2605.27178v1 Announce Type: cross Abstract: We address the challenging task of 3D object segmentation in complex scene point clouds without relying on any scene-level human annotations during tr

agentsarxiv-cs-ai
27 May 2026
Model Releases

Knowledge Graphs as the Missing Data Layer for LLM-Based Industrial Asset Operations

DGX agent

arXiv:2605.26874v1 Announce Type: cross Abstract: LLM-based agents for industrial asset operations show limited accuracy when reasoning over flat document stores. AssetOpsBench (KDD 2026) establishes

model-releasesarxiv-cs-ai
27 May 2026
Local Ai

Microsoft and Dell believe the answer to rising cloud token costs sits on every employee’s desk

DGX agent

Enterprises are adopting an agentic AI PC strategy as Copilot+ PCs shift AI workloads from expensive cloud inference to secure, high-performance on-device processing. As agentic AI moves from experime

local-aisiliconangle
27 May 2026
Model Releases

The MiniMax M2 series was one of the most widely used open-weight LLM series earlier this year. Now, we got a technical report with some int…

DGX agent

The MiniMax M2 series was one of the most widely used open-weight LLM series earlier this year. Now, we got a technical report with some interesting tidbits. I summarized some of them below: 1. Full a

model-releasessebastian-raschka--x
27 May 2026
Agents

AutoSOTA: An End-to-End Automated Research System for State-of-the-Art AI Model Discovery

DGX agent

arXiv:2604.05550v2 Announce Type: replace Abstract: Artificial intelligence research increasingly depends on prolonged cycles of reproduction, debugging, and iterative refinement to achieve State-Of-T

agentsarxiv-cs-cl
26 May 2026
Agents

EfficientGraph-RAG: Structured Retrieval-State Management for Cross-Task Retrieval-Augmented Generation

DGX agent

arXiv:2605.25379v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) has become the standard way to ground large language models in external knowledge, but many systems still organize

agentsarxiv-cs-cl
26 May 2026
Applications

Hypothesis Generation and Inductive Inference in Children and Language Models

DGX agent

arXiv:2605.24528v1 Announce Type: new Abstract: Real world decision-making requires constructing mental models under uncertainty over evidence, over the underlying causal rules, and over the state of

applicationsarxiv-cs-ai
26 May 2026
Agents

Meta-Engineering Harnesses for AI-Native Software Production: A Contract-Driven Adversarial Verification Architecture with Early Deployment Report

DGX agent

arXiv:2605.25665v1 Announce Type: cross Abstract: AI-native software development is often evaluated at the level of individual models, prompts, or generated artifacts. This framing is insufficient for

agentsarxiv-cs-ai
26 May 2026
Agents

Multi-Persona Debate System for Automated Scientific Hypothesis Generation

DGX agent

arXiv:2605.23917v1 Announce Type: new Abstract: Modern scientific discovery is bottlenecked not by data scarcity, but by the inability to synthesize fragmented knowledge into actionable hypotheses. Th

agentsarxiv-cs-cl
26 May 2026
Agents

PCGRLLM: Large Language Model-Driven Reward Design for Procedural Content Generation Reinforcement Learning

DGX agent

arXiv:2502.10906v2 Announce Type: replace Abstract: Reward design plays a pivotal role in the training of game AIs, requiring substantial domain-specific knowledge and human effort. In recent years, s

agentsarxiv-cs-ai
26 May 2026
← Previous
1…221222223224225…375
Next →