AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlog
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
10,103 results
Agents

From Control Boundary to Insurance Claim: Reconstructing AI-Mediated Losses Through the CER Framework

DGX agent

arXiv:2606.03777v1 Announce Type: new Abstract: AI losses that arise through an insured organization's generative or agentic AI system require state reconstruction, not merely event reconstruction, be

agentsarxiv-cs-ai
3 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Research

How Much of a Model Do We Need? Redundancy and Slimmability in Remote Sensing Foundation Models

DGX agent

arXiv:2601.22841v2 Announce Type: replace Abstract: Large-scale foundation models (FMs) in remote sensing (RS) (denoted as RS FMs) are developed following paradigms established in computer vision (CV)

researcharxiv-cs-cv
3 Jun 2026
Agents

Inducing Reasoning Primitives from Agent Traces

DGX agent

arXiv:2606.02994v1 Announce Type: new Abstract: ReAct-style LLM agents often rediscover the same reasoning routines across problems, yet leave those routines trapped in transient scratchpads. We intro

agentsarxiv-cs-ai
3 Jun 2026
Safety

Libra: Efficient Resource Management for Agentic RL Post-Training

DGX agent

arXiv:2606.03077v1 Announce Type: cross Abstract: Reinforcement learning (RL) has become a standard post-training paradigm for large language models (LLMs), extending beyond preference alignment to co

safetyarxiv-cs-ai
3 Jun 2026
Model Releases

SCOPE: Real-Time Natural Language Camera Agent at the Edge

DGX agent

arXiv:2606.02951v1 Announce Type: cross Abstract: Deploying language-driven agents in robotics requires evaluations that reflect real-world task demands: natural-language instructions with reproducibl

model-releasesarxiv-cs-ai
3 Jun 2026
Agents

Uncertainty-Aware Clarification in LLM Agents with Information Gain

DGX agent

arXiv:2606.03135v1 Announce Type: new Abstract: Large Language Model (LLM) agents often operate under underspecified user instructions, where latent uncertainty over user intent leads to erroneous too

agentsarxiv-cs-ai
3 Jun 2026
Safety

Addressing Longstanding Challenges in Cognitive Science with Language Models

DGX agent

arXiv:2511.00206v3 Announce Type: replace Abstract: Cognitive science faces ongoing challenges in research integration, formalization, conceptual clarity, and other areas, in part due to its multiface

safetyarxiv-cs-ai
2 Jun 2026
Safety

Agent Operating Systems (AOS): Integrating Agentic Control Planes into, and Beyond, Traditional Operating Systems

DGX agent

arXiv:2606.01508v1 Announce Type: cross Abstract: Traditional operating systems were designed around deterministic programs, explicit control flow, and human initiated workflows. Their core abstractio

safetyarxiv-cs-ai
2 Jun 2026
Agents

Agent-R1: A Unified and Modular Framework for Agentic Reinforcement Learning

DGX agent

arXiv:2511.14460v2 Announce Type: replace Abstract: Large language models (LLMs) have rapidly evolved from single-turn text generators into the foundation of increasingly capable agents. As these agen

agentsarxiv-cs-cl
2 Jun 2026
Agents

Agentic Authoring of Interactive Multiview Visualizations in Genomics

DGX agent

arXiv:2606.00370v1 Announce Type: cross Abstract: Diverse genomics data, scientific questions, and analysis tasks typically demand highly specialized visualizations. Therefore, users often must custom

agentsarxiv-cs-ai
2 Jun 2026
Agents

Automated Conjecture Resolution with Formal Verification

DGX agent

arXiv:2604.03789v2 Announce Type: replace-cross Abstract: Recent advances in large language models have significantly improved their ability to perform mathematical reasoning, extending from elementar

agentsarxiv-cs-ai
2 Jun 2026
Research

DeepLatent: Think with Images via Parallel Latent Visual Reasoning

DGX agent

arXiv:2606.00562v1 Announce Type: new Abstract: The emerging paradigm of 'thinking with images' embeds visual states into intermediate reasoning steps, defining a new frontier for Vision-Language Mode

researcharxiv-cs-cv
2 Jun 2026
Agents

Grok Build is genuinely amazing right now. It is not just another coding assistant. It is a full agentic system that can plan, write, refact…

DGX agent

Grok Build is genuinely amazing right now. It is not just another coding assistant. It is a full agentic system that can plan, write, refactor, debug, and build complete projects autonomously from a s

agentselon-musk--x
2 Jun 2026
Safety

IstGPT: LLM-based Anomaly Detection for Spatial-Temporal Graph in Industrial Systems

DGX agent

arXiv:2606.01691v1 Announce Type: cross Abstract: Industrial Internet systems face increasing threats from sophisticated industrial control system (ICS) attacks, resulting in critical safety incidents

safetyarxiv-cs-lg
2 Jun 2026
Agents

Learning to Construct Practical Agentic Systems

DGX agent

arXiv:2606.00189v1 Announce Type: cross Abstract: Automated design and optimization of agentic LLM-based systems leads to sophisticated systems that substantially improve result quality over off-the-s

agentsarxiv-cs-ai
2 Jun 2026
Agents

OctoT2I: A Self-Evolving Agentic Text-to-Image Router

DGX agent

arXiv:2606.01803v1 Announce Type: new Abstract: The explosive growth of Text-to-Image (T2I) models, from large-scale versions to lightweight, real-time ones, now faces diminishing marginal returns fro

agentsarxiv-cs-ai
2 Jun 2026
Applications

Predicting the risk of colorectal anastomotic leak based on preoperative mapping of the blood supply of the bowel

DGX agent

arXiv:2606.02156v1 Announce Type: cross Abstract: Anastomotic leak remains one of the most serious complications following colorectal cancer surgery, substantially affecting patient outcomes, recovery

applicationsarxiv-cs-ai
2 Jun 2026
Model Releases

ProtStructQA: A Denotation Threshold in Protein Structural Reasoning

DGX agent

arXiv:2606.00451v1 Announce Type: new Abstract: Protein-language systems are often evaluated by whether they generate plausible biological text, but a structural question has a sharper semantics: it d

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Quality Audio Prototyping: a prototype system for unified sound retrieval and procedural generation

DGX agent

arXiv:2606.00629v1 Announce Type: cross Abstract: Sound design workflows frequently oscillate between time-consuming library searches and the complexity of procedural synthesis, with practitioners typ

model-releasesarxiv-cs-lg
2 Jun 2026
Safety

SafeMCP: Proactive Power Regulation for LLM Agent Defense via Environment-Grounded Look-Ahead Reasoning

DGX agent

arXiv:2606.01991v1 Announce Type: new Abstract: As Large Language Model (LLM) agents increasingly leverage the Model Context Protocol (MCP) to operate in complex environments, the expansion of their a

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

Sandboxed Coding Agents are Competitive Omni-modal Task Solvers

DGX agent

arXiv:2606.00579v1 Announce Type: new Abstract: As multimodal LLMs increasingly target video and audio, it is often assumed that such tasks require native omnimodal models. We show that this is not al

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

SPADE-Bench: Evaluating Spontaneous Strategic Deception in Agents via Plan-Action Divergence

DGX agent

arXiv:2606.02380v1 Announce Type: cross Abstract: As LLM-based agents expand their operational scope, reliability becomes a prerequisite for real-world deployment. However, in practical applications,

model-releasesarxiv-cs-ai
2 Jun 2026
Research

Step-Level Sparse Autoencoder for Reasoning Process Interpretation

DGX agent

arXiv:2603.03031v2 Announce Type: replace Abstract: Large Language Models (LLMs) have achieved strong complex reasoning capabilities through Chain-of-Thought (CoT) reasoning. However, their reasoning

researcharxiv-cs-lg
2 Jun 2026
Local Ai

v0.30.1

DGX agent

Ollama 0.30 provides improved compatibility and performance using llama.cpp, augments the MLX engine on Apple Silicon for broader hardware support, and brings support for a wider range of models inclu

local-aiollama-releases
2 Jun 2026
Model Releases

When Safe Skills Collide: Measuring Compositional Risk in Agent Skill Ecosystems

DGX agent

arXiv:2606.00448v1 Announce Type: cross Abstract: LLM agents increasingly rely on community-contributed skills that expand an agent's operational capability set. We study a core safety problem in agen

model-releasesarxiv-cs-ai
2 Jun 2026
Safety

Counterfactual Evaluation Reveals Hidden Capability Profiles in Clinical LLMs and Agents

DGX agent

arXiv:2605.30590v1 Announce Type: cross Abstract: Two clinical AI systems can score nearly identically on coverage-based rubrics yet behave radically differently when their patient inputs change: one

safetyarxiv-cs-ai
1 Jun 2026
Research

EBuddy: a workflow orchestrator for industrial human-machine collaboration

DGX agent

arXiv:2603.28579v2 Announce Type: replace Abstract: This paper presents EBuddy, a voice-guided workflow orchestrator for natural human-machine collaboration in industrial environments. EBuddy targets

researcharxiv-cs-ro
1 Jun 2026
Model Releases

LegSegNet: A Public Deep Learning System for Lower Extremity CT Tissue Segmentation and Quantification

DGX agent

arXiv:2605.30829v1 Announce Type: new Abstract: Lower extremity computed tomography (CT) contains clinically relevant information for body composition analysis, sarcopenia assessment, and musculoskele

model-releasesarxiv-cs-cv
1 Jun 2026
Model Releases

Mellum2 Technical Report

DGX agent

arXiv:2605.31268v1 Announce Type: new Abstract: We present Mellum 2, an open-weight 12B-parameter Mixture-of-Experts (MoE) language model with 2.5B active parameters per token. Mellum 2 is a general-p

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

The fully-managed Remote MCP Server for AlloyDB is now Generally Available

DGX agent

AI agents possess incredible reasoning capabilities and can perform increasingly complex actions. But the reliability of agentic outcomes depends entirely on the quality of the context they can access

model-releasesgoogle-cloud-ai
1 Jun 2026
Model Releases

Brain-IT-VQA: From Brain Signals to Answers

DGX agent

arXiv:2605.29588v1 Announce Type: cross Abstract: Decoding visual content from fMRI signals recorded while a person views images, and specifically answering questions about the seen images, is a long-

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Cloud CISO Perspectives: How to build an AI-ready security program for the public sector

DGX agent

Welcome to the second Cloud CISO Perspectives for May 2026. Today, Usman Chaudhary, Field CISO, Google Public Sector, offers a guide for CISOs protecting government agencies and critical infrastructur

model-releasesgoogle-cloud-ai
29 May 2026
Model Releases

Developer's guide to Gemini Enterprise and A2UI integration

DGX agent

If you've built a chatbot, you know this conversation: User: 'Book a table for two tomorrow at 7pm.' Agent: 'Okay, for what day?' User: 'Tomorrow.' Agent: 'What time?' A date picker would have ended t

model-releasesgoogle-cloud-ai
29 May 2026
Research

Eigen-Spike Emergence and Quadratic Equivalents for Conjugate Kernels on Nonlinearly Separable Data

DGX agent

arXiv:2605.29669v1 Announce Type: cross Abstract: Recent work in random matrix theory (RMT) has developed the notion of deterministic equivalents: typically linear surrogate models that approximate th

researcharxiv-cs-lg
29 May 2026
Model Releases

GroundAct: Can LLM Agents Ground Actions in Environmental States?

DGX agent

arXiv:2508.05614v2 Announce Type: replace-cross Abstract: LLM agents achieve 85-96% success on tasks where instructions fully specify the action, but drop to 29-53% when action feasibility depends on

model-releasesarxiv-cs-ai
29 May 2026
Agents

Most people training agentic LLMs with RL right now have a silently broken training loop and have no idea. Here's the trap: single-turn RL w…

DGX agent

Most people training agentic LLMs with RL right now have a silently broken training loop and have no idea. Here's the trap: single-turn RL works beautifully. Clean curves, sane rewards, everything con

agentsclem-delangue--x
29 May 2026
Model Releases

TaxDistill: Improving Metagenomic Taxonomic Annotation via Distilled Genomic Foundation Models

DGX agent

arXiv:2605.28868v1 Announce Type: cross Abstract: Metagenomic taxonomic annotation aims to identify the microbial origins of DNA fragments in environmental samples. Traditional methods that rely on se

model-releasesarxiv-cs-ai
29 May 2026
Local Ai

Towards Understanding the Shape of Representations in Protein Language Models

DGX agent

arXiv:2509.24895v2 Announce Type: replace Abstract: While protein language models (PLMs) are one of the most promising avenues of research for future de novo protein design, the way in which they tran

local-aiarxiv-cs-lg
29 May 2026
Model Releases

Announcing the newest cohort of the Google for Startups Accelerator: Middle East, North Africa & Turkey

DGX agent

Google’s mission is to organize the world’s information and make it universally accessible. In high-growth, technically ambitious markets like the Middle East, North Africa, and Türkiye (MENA-T), we f

model-releasesgoogle-cloud-ai
28 May 2026
Local Ai

Decoupling Skeleton and Flesh: Efficient Multimodal Table Reasoning with Disentangled Alignment and Structure-aware Guidance

DGX agent

arXiv:2602.03491v2 Announce Type: replace-cross Abstract: Reasoning over table images remains challenging for Large Vision-Language Models (LVLMs) due to complex layouts and tightly coupled structure-

local-aiarxiv-cs-cl
28 May 2026
Agents

DeltaMCP: Incremental Regeneration via Spec-Aware Transformation for MCP servers

DGX agent

arXiv:2605.28148v1 Announce Type: cross Abstract: The rapid development of LLMs coupled with the introduction of Model Context Protocol (MCP) has revolutionized how intelligent agents interact with AP

agentsarxiv-cs-ai
28 May 2026
Safety

Diagnosing Live Within-Policy Instruction Conflicts in LLM Agents with Witnessed Resolution Profiles

DGX agent

arXiv:2605.27784v1 Announce Type: new Abstract: LLM agents are governed by long-lived natural-language prompt policies, but individually reasonable standing rules can interact in uninspected ways. We

safetyarxiv-cs-ai
28 May 2026
Model Releases

How the University of Central Oklahoma is using AI to streamline analysis of complex criminal cases

DGX agent

In the high-stakes world of forensic science, time is the enemy of justice. The University of Central Oklahoma (UCO) Forensic Science Institute (FSI) was looking for an innovative AI solution that cou

model-releasesgoogle-cloud-ai
28 May 2026
Safety

LACUNA: Safe Agents as Recursive Program Holes

DGX agent

arXiv:2605.28617v1 Announce Type: new Abstract: LLM agents increasingly act by writing code, yet a split persists between the runtime that drives the agent and the code the model writes. The runtime o

safetyarxiv-cs-ai
28 May 2026
Model Releases

LiveBrowseComp: Are Search Agents Searching, or Just Verifying What They Already Know?

DGX agent

arXiv:2605.28721v1 Announce Type: new Abstract: Are LLM-based search agents genuinely searching, or using the web to verify what they already know? We study this question on BrowseComp with three diag

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

MolLingo: Molecule-Native Representations for LLM-Powered Scientific Agents

DGX agent

arXiv:2605.27853v1 Announce Type: new Abstract: We present MolLingo, a multi-agent system that emulates the reasoning process of a chemist to automate molecular design. Existing LLM-based approaches e

model-releasesarxiv-cs-ai
28 May 2026
Research

RCM Constraint-Consistent Dynamic Control in Surgical Robots

DGX agent

arXiv:2509.14075v2 Announce Type: replace Abstract: Robotic-assisted minimally invasive surgery (RAMIS) requires accurate enforcement of the remote center of motion (RCM) constraint to ensure safe too

researcharxiv-cs-ro
28 May 2026
Model Releases

Thermodynamic properties of chemically disordered compounds via AI-driven estimation of partition function with the PULSE method

DGX agent

arXiv:2605.28594v1 Announce Type: cross Abstract: In this article, we present an improved version of the PULSE method (Partition function Unsupervised Learning Sampling and Evaluation) for estimating

model-releasesarxiv-cs-ai
28 May 2026
← Previous
1…116117118119120…211
Next →