AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,973 results
Tutorials

How Endava is redesigning software delivery around AI agents

DGX agent

Endava, a software services company, is leveraging AI agents to fundamentally transform its software delivery processes and workflows. The case study likely demonstrates how the company is implementin

tutorialsopenai
4 Jun 2026
Model Releases
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Introducing NVIDIA Nemotron 3 Ultra. A frontier smart open model built for long-running agents that need to plan, reason, use tools and keep…

DGX agent

Introducing NVIDIA Nemotron 3 Ultra. A frontier smart open model built for long-running agents that need to plan, reason, use tools and keep working across complex coding, research and enterprise work

model-releasesclem-delangue--x
4 Jun 2026
Safety

Multi-Agent Next-Best-View Optimization for Risk-Averse Planning

DGX agent

arXiv:2606.04158v1 Announce Type: new Abstract: Multi-agent Next-Best-View (NBV) selection for safe path planning in uncertain and unknown environments requires informative, safety-aware, and efficien

safetyarxiv-cs-ro
4 Jun 2026
Model Releases

NVIDIA Nemotron 3 Ultra Powers Faster, More Efficient Reasoning for Long-Running Agents

DGX agent

NVIDIA's Nemotron 3 Ultra is a 550B-parameter Mixture-of-Experts model with 55B active parameters, optimized for orchestrating complex, long-running agent workflows by combining frontier reasoning and

model-releasesnvidia-developer
4 Jun 2026
Safety

PersonaTree: Structured Lifecycle Memory for Person Understanding in LLM Agents

DGX agent

arXiv:2606.04780v1 Announce Type: new Abstract: Persistent LLM agents require memory representations that make the formation of person understanding explicit across long term interaction. Existing age

safetyarxiv-cs-cl
4 Jun 2026
Model Releases

RAMPART: Registry-based Agentic Memory with Priority-Aware Runtime Transformation

DGX agent

arXiv:2606.04628v1 Announce Type: new Abstract: RAMPART is a compile-time memory model and pure in-RAM block registry for LLM-based agents. Context assembly is a programmable runtime operation where c

model-releasesarxiv-cs-cl
4 Jun 2026
Research

SaliMory: Orchestrating Cognitive Memory for Conversational Agents

DGX agent

arXiv:2606.04120v1 Announce Type: cross Abstract: Conversational agents that serve as lifelong companions must maintain persistent memory across all interactions. However, simply expanding context win

researcharxiv-cs-ai
4 Jun 2026
Safety

The Accountability Horizon: An Impossibility Theorem for Governing Human-Agent Collectives

DGX agent

arXiv:2604.07778v2 Announce Type: replace Abstract: Existing accountability frameworks for AI systems, legal, ethical, and regulatory, rest on a shared assumption: for any consequential outcome, at le

safetyarxiv-cs-ai
4 Jun 2026
Model Releases

We're presenting ParseBench at CVPR 2026 today. 🦙 Come learn why document understanding is an AGI-complete problem (an agent can't act on a…

DGX agent

We're presenting ParseBench at CVPR 2026 today. 🦙 Come learn why document understanding is an AGI-complete problem (an agent can't act on a doc it can't correctly read, and reading a real enterprise t

model-releasesjerry-liu--x
4 Jun 2026
Hardware

CoreWeave’s Vera Rubin milestone sets stage for theCUBE’s agentic AI coverage

DGX agent

The agentic AI era is putting new pressure on the infrastructure stack, and CoreWeave Inc.’s latest milestone gives the conversation a sharper edge. This week, the company announced that it has comple

hardwaresiliconangle
3 Jun 2026
Model Releases

Cross-Lingual Token Arbitrage: Optimizing Code Agent Context Windows via Local LLM Preprocessing

DGX agent

arXiv:2606.03618v1 Announce Type: new Abstract: AI-assisted coding agents are bottlenecked by input-token cost. Two pathologies of raw human input drive much of this overhead: tokenization inefficienc

model-releasesarxiv-cs-ai
3 Jun 2026
Safety

Enhancing Operational Safety via Agentic Dialogue Hazard Identification Analysis

DGX agent

arXiv:2606.03812v1 Announce Type: new Abstract: Operational safety in high-stakes domains such as industrial process control, autonomous, and safety-critical systems, demand reliable hazard identifica

safetyarxiv-cs-ai
3 Jun 2026
Model Releases

Hedge-Bench: Benchmarking Agents on Hard, Realistic Tasks Pertaining to Financial Reasoning

DGX agent

arXiv:2606.03918v1 Announce Type: new Abstract: AI agents can increasingly handle the mechanical tasks of financial analysis: retrieving documents, calculating formulas, updating spreadsheets. The har

model-releasesarxiv-cs-ai
3 Jun 2026
Agents

Improve your agent’s tool-calling accuracy with SFT and DPO on Amazon SageMaker AI

DGX agent

In this post, you learn how to use Supervised Fine-Tuning (SFT) and Direct Preference Optimization (DPO) together to improve the tool-calling accuracy of a small language model (SLM). The example uses

agentsaws-ml-blog
3 Jun 2026
Model Releases

PieArena: Ranking and Profiling Language Agents in Realistic Negotiation Scenarios

DGX agent

arXiv:2602.05302v3 Announce Type: replace Abstract: We present an in-depth evaluation of LLMs' ability to negotiate, a central business task requiring strategic reasoning, theory of mind, and economic

model-releasesarxiv-cs-ai
3 Jun 2026
Safety

Self-Refining Agentic Reinforcement Learning for Vision-Conditioned UAV Navigation

DGX agent

arXiv:2606.03963v1 Announce Type: cross Abstract: Deep reinforcement learning has shown strong potential for enabling autonomous robots to learn complex navigational tasks. However, its practical use

safetyarxiv-cs-ai
3 Jun 2026
Model Releases

StepFinder: A Temporal Semantic Framework for Failure Attribution in Multi-Agent Systems

DGX agent

arXiv:2606.03467v1 Announce Type: new Abstract: LLM-based multi-agent systems exhibit remarkable collaborative capabilities in complex multi-step tasks. However, these systems are highly sensitive to

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

TalkPlayData 2: An Agentic Synthetic Data Pipeline for Multimodal Conversational Music Recommendation

DGX agent

arXiv:2509.09685v5 Announce Type: replace-cross Abstract: We present TalkPlayData 2, a synthetic dataset for multimodal conversational music recommendation generated by an agentic data pipeline. In th

model-releasesarxiv-cs-ai
3 Jun 2026
Local Ai

VulnAgent-R2: Evidence-Calibrated Multi-Agent Auditing for Repository-Level Vulnerability Detection

DGX agent

arXiv:2603.13384v2 Announce Type: replace-cross Abstract: Software vulnerabilities often depend on cross-file data flow, build options, framework conventions, and runtime guards, so isolated function

local-aiarxiv-cs-ai
3 Jun 2026
Model Releases

AgentProcessBench: Diagnosing Step-Level Process Quality in Tool-Using Agents

DGX agent

arXiv:2603.14465v2 Announce Type: replace Abstract: While Large Language Models (LLMs) have evolved into tool-using agents, they remain brittle in long-horizon interactions. Unlike mathematical reason

model-releasesarxiv-cs-ai
2 Jun 2026
Safety

Agents on a Tree: Pathwise Coordination for Multi-Objective Molecular Optimization

DGX agent

arXiv:2606.00008v1 Announce Type: new Abstract: Multi-objective molecular optimization requires searching vast chemical spaces under conflicting objectives, where early design decisions strongly const

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

APEX-SQL: Talking to the data via Agentic Exploration for Text-to-SQL

DGX agent

arXiv:2602.16720v2 Announce Type: replace-cross Abstract: Text-to-SQL systems powered by Large Language Models have excelled on academic benchmarks but struggle in complex enterprise environments. The

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Bridging Requirements and Architecture: Multi-Agent Orchestration with External Knowledge and Hierarchical Memory

DGX agent

arXiv:2606.01385v1 Announce Type: cross Abstract: Software architecture design is a critical yet inherently complex and knowledge-intensive phase that requires balancing competing quality attributes a

model-releasesarxiv-cs-ai
2 Jun 2026
Tutorials

Build agents you can trust across any framework with open evals and a control standard

DGX agent

Learn how Microsoft helps developers build trustworthy AI agents with open evaluations, portable runtime controls, production observability, and security workflows that work across frameworks. The pos

tutorialsmicrosoft-foundry
2 Jun 2026
Model Releases

Data agents don't fail at writing SQL. They fail at knowing your business. Schemas show you the columns, but they don't tell you which view …

DGX agent

Data agents don't fail at writing SQL. They fail at knowing your business. Schemas show you the columns, but they don't tell you which view is canonical for ARR, how often each metric updates, or whic

model-releasespinecone--x
2 Jun 2026
Model Releases

ExpWeaver: LLM Agents Learn from Experience via Latent RAG

DGX agent

arXiv:2606.01041v1 Announce Type: new Abstract: Experience learning has achieved promising results in enhancing LLM agent planning and reasoning by integrating past interactions as reusable knowledge.

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

ForeSci: Evaluating LLM Agents for Forward-Looking AI Research Judgment

DGX agent

arXiv:2606.00644v1 Announce Type: new Abstract: AI research often requires decisions before future evidence exists: which bottleneck to attack, which direction to pursue, or where a project should be

model-releasesarxiv-cs-ai
2 Jun 2026
Safety

Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses

DGX agent

arXiv:2606.02373v1 Announce Type: new Abstract: Search agents are often trained as policies over growing transcripts: the model must decide how to search while also remembering what it has seen, which

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

Identifying High-Confidence Social Biases in LLMs for Trustworthy Conversational Tutoring Agents

DGX agent

arXiv:2606.01584v1 Announce Type: cross Abstract: Conversational tutoring agents have been shown to improve learning engagement and student outcomes, and large language models (LLMs) are increasingly

model-releasesarxiv-cs-ai
2 Jun 2026
Safety

Large Language Model Guided Incentive Aware Reward Design for Cooperative Multi-Agent Reinforcement Learning

DGX agent

arXiv:2603.24324v4 Announce Type: replace-cross Abstract: Designing effective auxiliary rewards for cooperative multi-agent systems remains challenging, as misaligned incentives can induce suboptimal

safetyarxiv-cs-ai
2 Jun 2026
Hardware

LayerRoute: Input-Conditioned Adaptive Layer Skipping via LoRA Fine-Tuning for Agentic Language Models

DGX agent

arXiv:2606.01838v1 Announce Type: cross Abstract: Agentic language model systems alternate between two structurally distinct step types: structured tool calls (short, deterministic, low perplexity) an

hardwarearxiv-cs-ai
2 Jun 2026
Model Releases

MemoNoveltyAgent: A Historical Research Memory-Aware Agent Workflow for Paper Novelty Assessment

DGX agent

arXiv:2603.20884v2 Announce Type: replace Abstract: To alleviate the heavy burden of paper screening, researchers increasingly rely on existing AI agents, such as AI reviewers or DeepResearch, for pap

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Microsoft launches Rayfin to let developers and agents build app backends on Fabric

DGX agent

Microsoft Corp. today introduced Rayfin, an open-source software development kit and command-line interface that lets developers and coding agents define an entire application backend in code and depl

model-releasessiliconangle
2 Jun 2026
Model Releases

Microsoft unveils Microsoft Execution Containers for Windows, an OS-level sandbox for AI agents, with OpenAI, Nvidia, Manus, and Nous Research as partners (Michael Nuñez/VentureBeat)

DGX agent

Michael Nuñez / VentureBeat: Microsoft unveils Microsoft Execution Containers for Windows, an OS-level sandbox for AI agents, with OpenAI, Nvidia, Manus, and Nous Research as partners — For the past t

model-releasestechmeme
2 Jun 2026
Industry

Microsoft's Project Solara is an Android OS designed for agents instead of apps

DGX agent

Microsoft has developed Project Solara, a platform for devices that run AI agents instead of apps, based on Android instead of Windows. The platform is Microsoft's bet that AI will open up entirely ne

industryars-technica
2 Jun 2026
Safety

MOSAIC: Modular Orchestration for Structured Agentic Intelligence and Composition

DGX agent

arXiv:2606.00708v1 Announce Type: new Abstract: Automated data science is a structured model-selection problem. A solution must choose data transformations, feature representations, architecture, trai

safetyarxiv-cs-ai
2 Jun 2026
Agents

NestRL: A Nested Training Regime for Mutual Adaptation in Human-AI Teaming

DGX agent

arXiv:2602.17737v2 Announce Type: replace-cross Abstract: Mutual adaptation is a central challenge in human-AI teaming, as humans naturally adjust their strategies in response to an AI agent's behavio

agentsarxiv-cs-lg
2 Jun 2026
Safety

RedDebate: Safer Responses Through Multi-Agent Red Teaming Debates

DGX agent

arXiv:2506.11083v3 Announce Type: replace Abstract: We introduce RedDebate, a novel multi-agent debate framework that provides the foundation for Large Language Models (LLMs) to identify and mitigate

safetyarxiv-cs-cl
2 Jun 2026
Agents

Rehumanizing global health care with agentic AI

DGX agent

The global health care sector is under increasing strain. Decades of chronic underinvestment and constraints in recruitment have coincided with a surge in demand for services for aging populations. Ga

agentsmit-tech-review
2 Jun 2026
Model Releases

RoleCDE:Benchmarking and Mitigating Role-Alignment Trade-offs in Role-Playing Agents

DGX agent

arXiv:2606.01552v1 Announce Type: new Abstract: Role-playing agents(RPAs) are widely used to steer large language models(LLMs) toward role-consistent behavior, yet existing benchmarks mainly evaluate

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

RubricMiddleware helps your agent verify task completion with a grader subagent This is similar to /goal in claude code or codex, but for de…

DGX agent

RubricMiddleware is a LangChain feature that enables agents to verify task completion by delegating grading to a specialized subagent using predefined rubrics. This approach parallels the goal-verific

model-releasesharrison-chase--x
2 Jun 2026
Model Releases

Self-Healing Agentic Orchestrators for Reliable Tool-Augmented Large Language Model Systems

DGX agent

arXiv:2606.01416v1 Announce Type: new Abstract: Tool-augmented large language model (LLM) agents rely on orchestration layers that coordinate planning, retrieval, tool invocation, validation, memory,

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

SkillAdaptor: Self-Adapting Skills for LLM Agents from Trajectories

DGX agent

arXiv:2606.01311v1 Announce Type: cross Abstract: Large language model (LLM) agents increasingly rely on reusable external skills to solve long-horizon interactive tasks. Existing training-free skill

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Strategizing at Speed: A Learned Model Predictive Game for Multi-Agent Drone Racing

DGX agent

arXiv:2602.06925v2 Announce Type: replace Abstract: Autonomous drone racing pushes the boundaries of high-speed motion planning and multi-agent strategic decision-making. Success in this domain requir

model-releasesarxiv-cs-ro
2 Jun 2026
Tutorials

Verification is the hidden bottleneck for knowledge work agents, especially in legal AI — complex, long-horizon work is graded by rubrics wi…

DGX agent

Verification is the hidden bottleneck for knowledge work agents, especially in legal AI — complex, long-horizon work is graded by rubrics with dozens of strict criteria. In new research with @LangChai

tutorialsharrison-chase--x
2 Jun 2026
Safety

COMPASS: Cognitive MCTS-Guided Process Alignment for Safe Search Agents

DGX agent

arXiv:2605.30838v1 Announce Type: new Abstract: LLM-powered search agents enable multi-step reasoning and tool use. However, these capabilities introduce retrieval-induced safety degradation, as harmf

safetyarxiv-cs-ai
1 Jun 2026
Model Releases

Design and Evaluation of Multi-Agent AI Oracle Systems for Prediction Market Resolution

DGX agent

arXiv:2605.30802v1 Announce Type: cross Abstract: Prediction markets aggregate collective intelligence to forecast uncertain events, but their utility depends on reliable outcome resolution. Existing

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

ElasticMem: Latent Memory as a Learnable Resource for LLM Agents

DGX agent

arXiv:2605.30690v1 Announce Type: new Abstract: Long-term memory is essential for LLM agents to reason coherently across extended interactions, personalize responses, and reuse past experience. Howeve

model-releasesarxiv-cs-cl
1 Jun 2026
← Previous
1…146147148149150…375
Next →