AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,963 results
Model Releases

Beyond Description: Cognitively Benchmarking Fine-Grained Action for Embodied Agents

DGX agent

arXiv:2511.18685v4 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) show promising results as decision-making engines for embodied agents operating in complex, physical enviro

model-releasesarxiv-cs-cv
16 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Quantum Circuit Vision: Cost-Aware Evaluation of Visual AI Agents for Quantum Code Generation

DGX agent

arXiv:2607.10057v1 Announce Type: cross Abstract: Can AI agents visually comprehend quantum circuit diagrams and generate verified executable code--and at what cost? We present Quantum Circuit Vision,

model-releasesarxiv-cs-cv
16 Jul 2026
Local Ai

A Learning-Rate-Gated Failure of GRPO in a Small Language and Vision-Language Model Web Agent: A Controlled Null and Its Mechanism

DGX agent

arXiv:2607.12640v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards, and Group Relative Policy Optimization (GRPO) in particular, is now run routinely on a supervised checkp

local-aiarxiv-cs-ai
15 Jul 2026
Model Releases

BREAKING: Inkling by @thinkymachines is 9th overall on Agentic Web App Arena by Design Arena with an Elo of 1257 It's an open-weight model i…

DGX agent

BREAKING: Inkling by @thinkymachines is 9th overall on Agentic Web App Arena by Design Arena with an Elo of 1257 It's an open-weight model in the same performance band as Claude Opus 4.6 by @Anthropic

model-releasessoumith-chintala--x
15 Jul 2026
Hardware

Develop Lightweight USD Runtimes Faster with AI Agents

DGX agent

nanousd‑labs, part of NVIDIA Omniverse Labs, uses AI agents to generate lightweight, spec‑compliant USD runtimes directly from the USD Core Specification, sidestepping the need to adapt large legacy c

hardwarenvidia-developer
15 Jul 2026
Model Releases

Multi-Perspective Agentic Program Repair via Code Property Graphs and Temporal Execution Graphs

DGX agent

arXiv:2607.12605v1 Announce Type: cross Abstract: Large language models (LLMs) have improved automated program repair (APR), but two limitations remain. First, raw execution traces are often too large

model-releasesarxiv-cs-ai
15 Jul 2026
Safety

Operationalising Multi-Dimensional Evaluation for Conversational Agents: A Scalable, Governed Pipeline with Selective Re-evaluation and Model Benchmarking

DGX agent

arXiv:2607.12085v1 Announce Type: new Abstract: Evaluating retail conversational agents requires methods beyond lexical-overlap metrics to assess intent alignment, factuality, helpfulness, clarity, to

safetyarxiv-cs-ai
15 Jul 2026
Safety

New work led by @FlemmingKondrup and @tomjiralerspong highlights an important vulnerability in chain-of-thought monitoring for agents, I hig…

DGX agent

New work led by @FlemmingKondrup and @tomjiralerspong highlights an important vulnerability in chain-of-thought monitoring for agents, I highly recommend giving it a read. Link to the paper: https://a

safetyyoshua-bengio--x
14 Jul 2026
Model Releases

love to see two port cos collaborating😁 get private market data for your agents from @akta_pro with your @monid_ai account!

DGX agent

love to see two port cos collaborating😁 get private market data for your agents from @akta_pro with your @monid_ai account! We just killed PitchBook. Introducing Claude for private market data. Your a

model-releasesyohei-nakajima--x
13 Jul 2026
Local Ai

3 production patterns for AI agents and how to evaluate each one

DGX agent

A local coding agent, an in-app customer assistant, and an AI SRE triaging production logs may all use the same model class—but not the same harness, eval plan, or rollout risk. Mastra CEO Sam Bhagwat

local-aiarize-ai
10 Jul 2026
Agents

Agentic Neural Architecture Search

DGX agent

arXiv:2607.07984v1 Announce Type: new Abstract: Neural architecture search (NAS) methods have grown increasingly efficient, yet they remain bounded by manually engineered search spaces that require su

agentsarxiv-cs-ai
10 Jul 2026
Model Releases

CausalDS: Benchmarking Causal Reasoning in Data-Science Agents

DGX agent

arXiv:2607.08093v1 Announce Type: new Abstract: Large language models (LLMs) increasingly act as integrated data-science agents, combining abstract reasoning with advanced tool use. Yet the relevant b

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Hugging Face Gemma Challenge results are in! 📈 Over 6 days, more than 100 AI agents and humans collaborated to make Gemma 4 inference 5x fa…

DGX agent

Hugging Face Gemma Challenge results are in! 📈 Over 6 days, more than 100 AI agents and humans collaborated to make Gemma 4 inference 5x faster on a single NVIDIA A10G GPU. - Fastest result: 491.8 TPS

model-releasesclem-delangue--x
10 Jul 2026
Research

I have spent years teaching people how AI agents work, but this one does not need an explanation. It just works. Meet the AI employee that 4…

DGX agent

This post highlights a practical AI agent that operates intuitively without requiring explanation of its underlying mechanics, suggesting it has achieved a user-friendly interface or autonomous functi

researchdair-ai--x
10 Jul 2026
Model Releases

The same way, we're probably one of the few AI startups with user network effects, we might become the first one with agent network effects!

DGX agent

The same way, we're probably one of the few AI startups with user network effects, we might become the first one with agent network effects! Hugging Face Gemma Challenge results are in! 📈 Over 6 days,

model-releasesclem-delangue--x
10 Jul 2026
Tools

We've also simplified the project and repo pickers and made them more powerful. Use the pickers to launch agents in fewer clicks.

DGX agent

Cursor has streamlined its project and repository selection interface to improve user efficiency, allowing developers to launch AI agents with fewer clicks. The update combines simplified navigation w

toolscursor--x
10 Jul 2026
Hardware

AI agent startup Lyzr reportedly raising 100M at 500M valuation

DGX agent

Lyzr Inc., a startup that helps enterprises build artificial intelligence agents, is reportedly raising a funding round worth about 100 million. Bloomberg today cited sources as saying that the deal h

hardwaresiliconangle
9 Jul 2026
Model Releases

Evaluating SageMath-Augmented LLM Agents for Computational and Experimental Mathematics

DGX agent

arXiv:2607.06820v1 Announce Type: new Abstract: Recent advances in AI for Mathematics have focused largely on autoformalization and theorem proving, leaving the role of Computer Algebra Systems (CAS)

model-releasesarxiv-cs-ai
9 Jul 2026
Model Releases

GPT-5.6 is now supported in Hermes Agent and available via Nous Portal

DGX agent

Nous Research announced support for GPT-5.6 integration within Hermes Agent, with access available through the Nous Portal platform. This update enables users to leverage the capabilities of GPT-5.6 t

model-releasesnous-research--x
9 Jul 2026
Safety

HiDVFS: Hierarchical Multi-Agent DVFS for Real-Time OpenMP DAG Workloads

DGX agent

arXiv:2601.06425v2 Announce Type: replace-cross Abstract: Leakage power in multicore embedded systems now rivals dynamic power, so DVFS schedulers must respect deadlines and thermal limits, not just a

safetyarxiv-cs-ai
9 Jul 2026
Agents

http://hermes-agent.nousresearch.com

DGX agent

Hermes Agent is a project from Nous Research focused on developing agentic AI systems capable of autonomous reasoning and task execution. The initiative likely explores methods for creating AI agents

agentsnous-research--x
9 Jul 2026
Safety

LLM-powered reasoning in agent-based modeling

DGX agent

arXiv:2607.06757v1 Announce Type: new Abstract: Agent-based modeling (ABM) has the capability to model millions of individuals and their interactions, which is useful for policy making. However, ABMs

safetyarxiv-cs-ai
9 Jul 2026
Industry

Mercor buys Deeptune to build training environments for AI agents

DGX agent

Artificial intelligence training data company Mercor.io Corp. announced today that it has acquired Deeptune Inc., a startup that builds simulated software environments used to train AI agents. Financi

industrysiliconangle
9 Jul 2026
Model Releases

Meta launches flagship Muse Spark 1.1 model with multi-agent upgrades

DGX agent

Meta Platforms Inc. today launched a new flagship large language model optimized to power multi-agent automation workflows. Muse Spark 1.1 is available in the company’s Meta AI chatbot service and via

model-releasessiliconangle
9 Jul 2026
Model Releases

OpenAI debuts ChatGPT Work, an agentic tool for automating business workflows

DGX agent

OpenAI Group PBC today launched a new “agentic” tool called ChatGPT Work as it announced the global rollout of its most advanced model family so far in GPT-5.6. ChatGPT Work is a new mode within ChatG

model-releasessiliconangle
9 Jul 2026
Agents

RIMRULE: Improving Tool-Using Language Agents via MDL-Guided Rule Learning

DGX agent

arXiv:2601.00086v3 Announce Type: replace Abstract: Large language models (LLMs) often struggle to use tools reliably in domain-specific settings, where APIs may be idiosyncratic, under-documented, or

agentsarxiv-cs-cl
9 Jul 2026
Safety

Single-Rollout Asynchronous Optimization for Agentic Reinforcement Learning

DGX agent

arXiv:2607.07508v1 Announce Type: cross Abstract: Reinforcement learning (RL) is becoming increasingly important for post-training large language models (LLMs). Previous RL pipelines for LLMs were mos

safetyarxiv-cs-ai
9 Jul 2026
Safety

Agentic AI for Commercial Insurance Underwriting with Adversarial Self-Critique

DGX agent

arXiv:2602.13213v2 Announce Type: replace Abstract: Commercial insurance underwriting is a labor-intensive process that requires manual review of extensive documentation to assess risk and determine p

safetyarxiv-cs-ai
8 Jul 2026
Agents

Agentic AI for IPoDWDM Network Lifecycle Automation: An MCP-Enabled Architecture

DGX agent

arXiv:2607.05958v1 Announce Type: cross Abstract: We present a distributed, vendor-agnostic multi-MCP architecture for SDN-based automation and autonomous control of multi-vendor, multi-layer IPoDWDM

agentsarxiv-cs-ai
8 Jul 2026
Hardware

Congrats to @SpaceXAI on Grok 4.5 — trained on NVIDIA GB300 NVL72 systems and purpose-built for coding, agentic tasks, and knowledge work. T…

DGX agent

Congrats to @SpaceXAI on Grok 4.5 — trained on NVIDIA GB300 NVL72 systems and purpose-built for coding, agentic tasks, and knowledge work. This is what happens when world-class AI infrastructure meets

hardwareelon-musk--x
8 Jul 2026
Model Releases

Former GitHub CEO Thomas Dohmke's Entire launches a decentralized Git network to handle high coding agent traffic, with servers in the US, the EU, and Australia (Radhika Rajkumar/ZDNET)

DGX agent

Radhika Rajkumar / ZDNET: Former GitHub CEO Thomas Dohmke's Entire launches a decentralized Git network to handle high coding agent traffic, with servers in the US, the EU, and Australia — ZDNET's key

model-releasestechmeme
8 Jul 2026
Safety

From Blueprint to Reality: Modeling and Applying Putnam's Social Capital Theory with LLM-based Multi-agent Simulations

DGX agent

arXiv:2607.06080v1 Announce Type: cross Abstract: Putnam's Social Capital Theory is a foundational framework for collective action and community prosperity. However, traditional empirical methods face

safetyarxiv-cs-ai
8 Jul 2026
Model Releases

IMR: Iterative Mode-World Weighted Regression for Multi-Agent Trajectory Prediction

DGX agent

arXiv:2607.05705v1 Announce Type: cross Abstract: Multi-agent motion prediction is essential for automated vehicles to understand the intentions of surrounding vehicles. However, previous prediction-b

model-releasesarxiv-cs-ai
8 Jul 2026
Safety

Intercepting an Agile Target with Net-Carrying Drones using Competitive Multi-Agent Reinforcement Learning

DGX agent

arXiv:2607.05939v1 Announce Type: new Abstract: This article presents a solution to intercept an agile drone by a team of agile drone carrying catching nets. We formulate the problem as a competitive

safetyarxiv-cs-ro
8 Jul 2026
Model Releases

Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning

DGX agent

arXiv:2607.05458v1 Announce Type: cross Abstract: Large language model (LLM) agents are usually improved by changing prompts, models, or hand-written workflows, while the execution harness around the

model-releasesarxiv-cs-ai
8 Jul 2026
Safety

Lingering Authority: Revocable Resource-and-Effect Capabilities for Coding Agents

DGX agent

arXiv:2606.22504v1 Announce Type: cross Abstract: Coding agents often receive broad tool access for an entire task, even when a resource is needed only for one subgoal. We call this gap lingering auth

safetyarxiv-cs-ai
8 Jul 2026
Industry

Prime Intellect, which helps companies build their own AI agents by offering computing power and specialized tools, raised a 130M Series A at a 1B valuation (Marina Temkin/TechCrunch)

DGX agent

Marina Temkin / TechCrunch: Prime Intellect, which helps companies build their own AI agents by offering computing power and specialized tools, raised a 130M Series A at a 1B valuation — Prime Intelle

industrytechmeme
8 Jul 2026
Safety

TurnOPD: Making On-Policy Distillation Turn-Aware for Efficient Long-Horizon Agent Training

DGX agent

arXiv:2607.05804v1 Announce Type: new Abstract: On-policy distillation (OPD) trains a student policy by matching a stronger teacher on the student's own trajectories, offering a promising framework fo

safetyarxiv-cs-ai
8 Jul 2026
Model Releases

Agent Step Value: State-Transition Measurement with State-Grounded LLM Evaluators

DGX agent

arXiv:2607.04419v1 Announce Type: new Abstract: Most agent evaluations collapse a multi-step trace into a final answer, a success flag, or a trajectory-level score. These aggregates obscure the diagno

model-releasesarxiv-cs-ai
7 Jul 2026
Safety

Agentic Artificial Intelligence for Multistage Physics Experiments at a Large-Scale User Facility Particle Accelerator

DGX agent

arXiv:2509.17255v2 Announce Type: replace-cross Abstract: We present the first language-model-driven agentic artificial intelligence (AI) system to autonomously execute multi-stage physics experiments

safetyarxiv-cs-ai
7 Jul 2026
Model Releases

AgentLTL: A Trace-Verification Framework for Measuring, Enforcing, and Training Procedural Compliance in Tool-Using LLM Agents

DGX agent

arXiv:2607.02599v1 Announce Type: cross Abstract: Tool-using LLM agents are usually evaluated by final-answer correctness or LLM judges. Neither captures how an answer was produced. In safety-critical

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Beyond Task Completion: A Verification-vs.-Conformance Gap in Tool-Evolving Agents

DGX agent

arXiv:2604.00392v2 Announce Type: replace-cross Abstract: Agents that synthesize their own tools ship a second artifact alongside each answer: a software library that future tasks reuse, compose, and

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

CLEANER: Self-Purified Trajectories Boost Agentic Reinforcement Learning

DGX agent

arXiv:2601.15141v2 Announce Type: replace Abstract: Agentic Reinforcement Learning (RL) has empowered Large Language Models (LLMs) to utilize tools like Python interpreters for complex problem-solving

model-releasesarxiv-cs-lg
7 Jul 2026
Safety

GaP: A Graph-as-Policy Multi-Agent Self-Learning Harness For Variational Automation Tasks

DGX agent

arXiv:2607.05369v1 Announce Type: cross Abstract: For robots to work reliably in commercial and industrial applications, can recent advances in agentic coding systems combine interpretable robot progr

safetyarxiv-cs-ai
7 Jul 2026
Model Releases

Governed MCP: Kernel-Level Tool Governance for AI Agents via Logit-Based Safety Primitives

DGX agent

arXiv:2604.16870v2 Announce Type: replace-cross Abstract: AI agents increasingly call external tools (file system, network, APIs) through the Model Context Protocol (MCP). These tool calls are the age

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

IndustryNav: Exploring Spatial Reasoning of Embodied Agents in Dynamic Industrial Navigation

DGX agent

arXiv:2511.17384v2 Announce Type: replace-cross Abstract: While Visual Large Language Models (VLLMs) show great promise as embodied agents, they continue to face substantial challenges in spatial reas

model-releasesarxiv-cs-cv
7 Jul 2026
Applications

Norm Ai nabs 120M at 1.2B valuation to bring AI agents to the law

DGX agent

Nomos Ai Inc., a company building a platform to embed legal operations into artificial intelligence agents, today announced it has raised 120 million in new funding, bringing the company’s valuation t

applicationssiliconangle
7 Jul 2026
Model Releases

Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents

DGX agent

arXiv:2607.03968v1 Announce Type: cross Abstract: Large language models are increasingly deployed as IDE-integrated coding agents that decompose tasks, generate and edit files, run code, and refine ou

model-releasesarxiv-cs-ai
7 Jul 2026
← Previous
1…142143144145146…375
Next →