AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
10,088 results
28 May 2026

Building Community-Centred NLP Resources for Puno Quechua

Model ReleasesDGX agent

arXiv:2605.28253v1 Announce Type: new Abstract: The preservation of under-resourced languages requires digital tools and resources shaped by and for their speakers. We present the first dedicated ASR

CNN sues Perplexity over ‘verbatim’ copycat articles

IndustryDGX agent

CNN has filed a lawsuit against Perplexity, claiming that the startup's AI tools generate 'verbatim' copies of its work, as reported earlier by CNN. The lawsuit, filed in a New York court on Thursday,

Cyclical Entropy Eruption: Entropy Dynamics in Agent Reinforcement Learning

AgentsDGX agent

arXiv:2605.27954v1 Announce Type: new Abstract: Agentic large language models are increasingly used to solve real-world tasks by reasoning over goals, invoking tools, and interacting with external env

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Finding Miscompiles for Fun, Not Profit

Model ReleasesDGX agent

This article likely discusses the discovery and analysis of compiler bugs or miscompilation errors in software development tools, exploring how developers identify these issues and their implications

Grimlock: Guarding High-Agency Systems with eBPF and Attested Channels

SafetyDGX agent

arXiv:2605.27488v1 Announce Type: cross Abstract: Agentic systems increasingly run user-authored orchestration code that invokes tools, spawns subtasks, and delegates work across machines and clouds.

IBM expands digital sovereignty push with new cloud compliance and visibility platform

IndustryDGX agent

IBM Corp. today introduced a new cloud sovereignty tool aimed at helping enterprises demonstrate compliance, operational control and governance as artificial intelligence deployments expand across hyb

Identifiable Bayesian Deep Generative Copulas with Unknown Layer Widths for Data with Arbitrary Marginal Distributions

TutorialsDGX agent

arXiv:2605.27523v1 Announce Type: cross Abstract: Deep generative models offer powerful tools for multivariate data analysis, but their black-box architectures are often unidentified and difficult to

Mag-VLA: Vision-Language-Action Model for Bimanual Magnetically Actuated Microrobot Manipulation

SafetyDGX agent

arXiv:2605.28486v1 Announce Type: new Abstract: Magnetically actuated microrobots have been used as wireless, non-contact manipulation tools at microscales, making them promising for minimally invasiv

MRMMIA: Membership Inference Attacks on Memory in Chat Agents

AgentsDGX agent

arXiv:2605.27825v1 Announce Type: cross Abstract: Membership inference attacks (MIAs) test whether a target data record belongs to a system's private data, and have become a standard tool to measure p

Off-Policy Learning to Reason Works Because It Is More Pessimistic Than You Think

SafetyDGX agent

arXiv:2605.28150v1 Announce Type: new Abstract: Large scale reinforcement learning has become a central tool for improving reasoning in large language models. At this scale, generation is often lagged

Performance and Explainability Requirements of Evolutionary Algorithms in Real-World Physics-Informed Optimization

ApplicationsDGX agent

arXiv:2605.28164v1 Announce Type: cross Abstract: Evolutionary computation offers a variety of tools to solve complex real-world optimization problems. However, research often focuses on smaller, simp

ProvMind: Provenance-grounded reasoning for materials synthesis

Model ReleasesDGX agent

arXiv:2605.28487v1 Announce Type: new Abstract: Materials process optimization requires reasoning over routes, conditions, tools and causal dependencies, yet most computational formulations flatten sy

Riverbed accelerates enterprise adoption of autonomous IT operations

AgentsDGX agent

With artificial intelligence tools proliferating across the enterprise world, observability is a more critical sphere than ever before. Riverbed Technology LLC, which specializes in unified observabil

Semantic Optimal Transport for Sparse Autoencoder Feature Matching and Circuit Compression

ResearchDGX agent

arXiv:2605.28567v1 Announce Type: cross Abstract: Sparse autoencoders (SAEs) have become a central tool for interpreting language models. However, two key SAE analyses that remain difficult to scale a

STR Robot: Design of an Autonomous Mobile Robot from Simulation to Reality

Model ReleasesDGX agent

arXiv:2605.28110v1 Announce Type: new Abstract: With the rapid development of simulation tools, the development and validation of autonomous robotic systems have become more efficient before real-worl

The best design work doesn't happen in a chat box. You need space to explore ideas, create variants, and iterate Meet the new Replit Canvas …

AgentsDGX agent

The best design work doesn't happen in a chat box. You need space to explore ideas, create variants, and iterate Meet the new Replit Canvas Your agentic design tool to build beautiful websites, apps,

The Energy Blind Spot: NVIDIA's Flagship Edge AI Hardware Cannot Support Process-Level Energy Attribution

Local AiDGX agent

arXiv:2605.27599v1 Announce Type: cross Abstract: Agentic AI workloads - where a single user goal triggers multi-step orchestration, tool calls, retries, and failure recovery - are being targeted for

tokenmaxxing is officially over

SafetyDGX agent

tokenmaxxing is officially over Sources: Amazon has shut down an internal leaderboard that tracked employees' use of AI tools after workers tried to boost their scores with needless tasks (@rafeuddin_

TRACES: Proactive Safety Auditing for Multi-Turn LLM Agents via Trajectory-State Modeling

SafetyDGX agent

arXiv:2605.27690v1 Announce Type: new Abstract: LLM agents increasingly operate through multi-turn tool use and environment interaction, where safety risks often emerge from intermediate steps long be

Understanding Automated Program Repair Agents Through the Lens of Traceability: An Empirical Study

AgentsDGX agent

arXiv:2506.08311v2 Announce Type: replace-cross Abstract: Automated Program Repair (APR) agents leverage Large Language Models (LLMs) to autonomously diagnose and fix software bugs through reasoning,

27 May 2026

Cisco and OpenAI redefine enterprise engineering with Codex

ApplicationsDGX agent

Cisco and OpenAI partnered to integrate OpenAI's Codex AI model into Cisco's enterprise engineering tools to enhance software development and automation capabilities. The collaboration aims to improve

FinHarness: An Inline Lifecycle Safety Harness for Finance LLM Agents

SafetyDGX agent

arXiv:2605.27333v1 Announce Type: new Abstract: Finance LLM agents must simultaneously block prompt-induced unauthorized actions and approve legitimate multi-step business workflows. However, boundary

For future-proof, build AI that's composable. Regardless of what you use, all these should be composable, iterative, and customizable: - LLM…

AgentsDGX agent

For future-proof, build AI that's composable. Regardless of what you use, all these should be composable, iterative, and customizable: - LLMs - Evals - Automations - MCP/CLI tools - Skills/Memory/Cont

Generating Robust Portfolios of Optimization Models using Large Language Models

ResearchDGX agent

arXiv:2605.27013v1 Announce Type: new Abstract: Mathematical optimization is a powerful tool for structured decision-making across domains such as resource allocation and planning. Formulating optimiz

'Give Me BF16 or Give Me Death'? Accuracy-Performance Trade-Offs in LLM Quantization

Model ReleasesDGX agent

arXiv:2411.02355v4 Announce Type: replace-cross Abstract: Quantization is a powerful tool for accelerating large language model (LLM) inference, but the accuracy-performance trade-offs across differen

Hermes Agent now has a built-in MCP Catalog

AgentsDGX agent

Nous Research announced that their Hermes Agent now includes an integrated MCP (Model Context Protocol) Catalog, enabling users to discover and access available Model Context Protocol tools and integr

Improving your agent has been a manual process of: ✅ Reading traces ✅ Looking for patterns ✅ Writing evals ✅ Creating fixes Now, LangSmith E…

AgentsDGX agent

LangSmith has introduced automated tools to streamline agent improvement, eliminating the manual workflow of reading execution traces, identifying patterns, writing evaluations, and implementing fixes

Knowledge Graphs as the Missing Data Layer for LLM-Based Industrial Asset Operations

Model ReleasesDGX agent

arXiv:2605.26874v1 Announce Type: cross Abstract: LLM-based agents for industrial asset operations show limited accuracy when reasoning over flat document stores. AssetOpsBench (KDD 2026) establishes

LLMs versus the Halting Problem: Characterizing Program Termination Reasoning

Model ReleasesDGX agent

arXiv:2601.18987v5 Announce Type: replace-cross Abstract: Determining whether a program terminates is a central problem in computer science. Turing's Halting Problem established termination as undecid

Modeling Agentic Technical Debt and Stochastic Tax: A Standalone Framework for Measurement, Simulation, and Dashboarding

AgentsDGX agent

arXiv:2605.27320v1 Announce Type: new Abstract: Agentic AI systems combine probabilistic reasoning with delegated action through tools, context, memory, orchestration, and external workflow integratio

PLAID: A Unified Data Model for Machine Learning on Heterogeneous Physics Simulations

ResearchDGX agent

arXiv:2505.02974v3 Announce Type: replace Abstract: Machine learning-based surrogate models have emerged as a powerful tool to accelerate simulation-driven scientific workflows, but their adoption is

Probabilistic Smoothing with Ratio-Monotone Transforms for Global Optimization

ResearchDGX agent

arXiv:2605.27316v1 Announce Type: new Abstract: Probabilistic smoothing is a standard tool for global optimization, but existing methods rely on Gaussian kernels and specific transforms, often resulti

Self-Verified Distillation: Your Language Model Is Secretly Its Own Synthetic Data Pipeline

Model ReleasesDGX agent

arXiv:2605.26132v1 Announce Type: new Abstract: Can post-trained large language models (LLMs) further improve themselves using only unlabeled prompts, without external teachers or feedback from tools?

The Strongest Teacher Is Not Always the Best Teacher: Student-Centric Answer Selection

AgentsDGX agent

arXiv:2605.26872v1 Announce Type: cross Abstract: LLM training increasingly relies on teacher-generated supervision, from synthetic responses to reasoning traces and tool-use demonstrations. Current p

Towards Controllable Image Generation through Representation-Conditioned Diffusion Models

ResearchDGX agent

arXiv:2605.27343v1 Announce Type: new Abstract: Diffusion models have emerged as powerful tools for high-quality image generation and editing, but guiding these models to produce specific outputs rema

Towards Just-in-Time Adaptive Feedback: Enhancing Student Learning via Knowledge-Grounded LLM

TutorialsDGX agent

arXiv:2605.26405v1 Announce Type: new Abstract: Educational interventions are effective tools for enhancing student learning. While Large Language Models (LLMs) allow for generating adaptive feedback

Vibe coding startup Cognition more than doubles valuation in new $1B+ round

IndustryDGX agent

Cognition Inc., a provider of artificial intelligence programming tools, today announced that it has raised more than 1 billion in funding. Lux Capital, General Catalyst and 8VC led the Series D round

XGrammar-2: Efficient Dynamic Structured Generation Engine for Agentic LLMs

AgentsDGX agent

arXiv:2601.04426v3 Announce Type: replace Abstract: Modern LLM agents increasingly rely on dynamic structured generation, such as tool calling and response protocols. Unlike traditional structured gen

26 May 2026

Agent-X: Evaluating Deep Multimodal Reasoning in Vision-Centric Agentic Tasks

Model ReleasesDGX agent

arXiv:2505.24876v2 Announce Type: replace-cross Abstract: Deep reasoning is fundamental for solving complex tasks, especially in vision-centric scenarios that demand sequential, multimodal understandi

AgentFugue: Agent Scaling for Long-Horizon Tasks through Collective Reasoning

AgentsDGX agent

arXiv:2605.24486v1 Announce Type: new Abstract: Recent progress on long-horizon agentic tasks has been driven largely by scaling up individual agents through stronger models, better tools, and more ef

AION: Next-Generation Tasks and Practical Harness for Time Series

AgentsDGX agent

arXiv:2605.25045v1 Announce Type: new Abstract: Time series research is moving beyond fixed forecasting benchmarks toward realistic tasks that combine prediction, contextual reasoning, tool use, and s

Anticipate and Learn: Unleashing Idle-Time Compute in Proactive Agents

Model ReleasesDGX agent

arXiv:2605.25971v1 Announce Type: new Abstract: While AI agents demonstrate remarkable capabilities in reasoning and tool use, they remain fundamentally reactive: they compute responses only after exp

Beyond Final Answers: Auditing Trajectory-Level Hallucinations in Multi-Agent Industrial Workflows

Model ReleasesDGX agent

arXiv:2605.24219v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed as autonomous agents that reason, use tools, and act over multiple steps. Yet most hallucination

CausalFlow: Causal Attribution and Counterfactual Repair for LLM Agent Failures

AgentsDGX agent

arXiv:2605.25338v1 Announce Type: cross Abstract: Large language model (LLM) agents frequently fail on multi-step tasks involving reasoning, tool use, and environment interaction. While such failures

Code2UML: Agentic LLMs with context engineering for scalable software visualization

Model ReleasesDGX agent

arXiv:2605.24453v1 Announce Type: cross Abstract: Large Language Model (LLM)-based code analysis tools are adopted to automate software documentation tasks. However, the scalability of these approache

Contractual Skills: A GovernSpec Design Framework for Enterprise AI Agents

SafetyDGX agent

arXiv:2605.22634v2 Announce Type: replace-cross Abstract: Skills have become a practical packaging mechanism for agent instructions, workflows, scripts, and reference materials. In enterprise settings

CUA-Gym: Scaling Verifiable Training Environments and Tasks for Computer-Use Agents

Model ReleasesDGX agent

arXiv:2605.25624v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has driven breakthroughs in domains such as math, tool-use, and software engineering, yet its exte

first citation of ActiveGraph on arXiv 'as a complementary runtime', plus they included a working bridge example in the repo 😁

AgentsDGX agent

Yohei Nakajima announced the first arXiv citation of ActiveGraph, positioning it as a complementary runtime tool. The announcement included a working bridge example in the repository, indicating pract

Grok Imagine image and video generation are truly incredible The realism is insanely good to the point where it starts blurring the line bet…

IndustryDGX agent

Elon Musk posted about Grok Imagine, a generative AI tool capable of creating highly realistic images and video, noting that the output quality is so advanced it becomes difficult to distinguish gener

Introducing CHI-Bench on @huggingface: the world’s first long-horizon healthcare benchmark for AI agents. 75 real healthcare workflows + 20 …

Model ReleasesDGX agent

Introducing CHI-Bench on @huggingface: the world’s first long-horizon healthcare benchmark for AI agents. 75 real healthcare workflows + 20 apps + 200+ MCP tools + 1,290 skills + process / outcome rew

IterInject: Indirect Prompt Injection Against LLM Agents via Feedback-Guided Iterative Optimization

Model ReleasesDGX agent

arXiv:2605.24659v1 Announce Type: new Abstract: LLM-based agents are increasingly deployed for complex tasks requiring planning, tool use, and interaction with external services. Their reliance on unt

Message-Passing GNNs Fail to Approximate Sparse Triangular Factorizations

ApplicationsDGX agent

arXiv:2502.01397v3 Announce Type: replace-cross Abstract: Graph Neural Networks (GNNs) have been proposed as a tool for learning sparse matrix preconditioners, which are key components in accelerating

Practical Quantum CIM Empowerment via All-Domestic-Core Agentic Large Model

AgentsDGX agent

arXiv:2605.23934v1 Announce Type: new Abstract: Quantum computing devices are recognized as powerful tools for solving NP-complete problems. However, the intricacy of their modeling presents notable b

SAM: State-Adaptive Memory for Long-Horizon Reasoning Agent

AgentsDGX agent

arXiv:2605.24468v1 Announce Type: new Abstract: Long-horizon agentic reasoning requires large language models to act over long interaction histories containing thoughts, tool calls, observations, and

SEAL: Synergistic Co-Evolution of Agents and Learning Environments

SafetyDGX agent

arXiv:2605.24426v1 Announce Type: new Abstract: Large Language Model (LLM) agents are increasingly improved through interaction, yet most self-evolution methods adapt either the policy or the learning

STORM: Internalized Modeling for Spatial-Temporal Reasoning in Video-Language Models

AgentsDGX agent

arXiv:2605.26014v1 Announce Type: cross Abstract: Many video reasoning tasks require tracking motion, temporal order, and evolving visual states across frames. Existing methods built on large vision-l

SURGE: On the Potential of Large Language Models as General-Purpose Surrogate Code Executors

Model ReleasesDGX agent

arXiv:2502.11167v5 Announce Type: replace-cross Abstract: Neural surrogate models are powerful and efficient tools in data mining. Meanwhile, large language models (LLMs) have demonstrated remarkable

this basically destroys the extrapolations that had Anthropic making two trillion dollars a year.

Model ReleasesDGX agent

this basically destroys the extrapolations that had Anthropic making two trillion dollars a year. It's clear that growth for coding tools such as Claude Code has decelerated from the pace it was since

Towards trustworthy agentic AI: a comprehensive survey of safety, robustness, privacy, and system security

SafetyDGX agent

arXiv:2605.23989v1 Announce Type: new Abstract: Agentic AI systems -- Large Language Models (LLMs) augmented with planning, tool use, memory, and long-horizon interactions -- can execute complex tasks

VineLM: Trie-Based Fine-Grained Control for Agentic Workflows

AgentsDGX agent

arXiv:2605.23914v1 Announce Type: cross Abstract: Agentic workflows interleave configurable LLM stages with tool stages and often include retries or refinement loops. Existing workflow managers profil

← Previous
1…7374757677…169
Next →