AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlog
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
10,103 results
Model Releases

Took some inspiration from @vboykis and converted my first ever talk into a blog post. I talk about the role of agentic search in context en…

DGX agent

Took some inspiration from @vboykis and converted my first ever talk into a blog post. I talk about the role of agentic search in context engineering. Together we build an intuition on the strengths a

model-releasesswyx--x
28 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Advancing Creative Physical Intelligence in Large Multimodal Models

DGX agent

arXiv:2605.26396v1 Announce Type: new Abstract: Large multimodal models (LMMs) have rapidly advanced in perception and reasoning; however, it remains unclear whether these capabilities generalize to d

model-releasesarxiv-cs-ai
27 May 2026
Safety

AI-Driven Contribution Evaluation and Conflict Resolution: A Framework & Design for Group Workload Investigation

DGX agent

arXiv:2511.07667v2 Announce Type: replace Abstract: The equitable assessment of individual contribution in teams remains a persistent challenge, where conflict and disparity in workload can result in

safetyarxiv-cs-ai
27 May 2026
Research

Guiding LLM Post-training Data Engineering with Model Internals from Sparse Autoencoders

DGX agent

arXiv:2605.27354v1 Announce Type: cross Abstract: Model internals encode rich information about how a large language model (LLM) processes its training data; however, post-training data engineering la

researcharxiv-cs-ai
27 May 2026
Local Ai

Persistent AI Agents in Academic Research: A Single-Investigator Implementation Case Study

DGX agent

arXiv:2605.26870v1 Announce Type: cross Abstract: Background: Large language models are typically evaluated as models, benchmarks, or short conversational episodes. Less is known about what happens wh

local-aiarxiv-cs-ai
27 May 2026
Model Releases

Position: AI Safety Requires Effective Controllability

DGX agent

arXiv:2605.27117v1 Announce Type: new Abstract: AI safety is still largely framed as alignment: training models to follow human preferences, safety policies, and normative constraints. That framing ha

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Shopping Companion: A Memory-Augmented LLM Agent for Real-World E-Commerce Tasks

DGX agent

arXiv:2603.14864v2 Announce Type: replace Abstract: In e-commerce, LLM agents show promise for shopping tasks such as recommendations, budget management, and bundle deals, where accurately capturing u

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

SWE-Adept: An LLM-Based Agentic Framework for Deep Codebase Analysis and Structured Issue Resolution

DGX agent

arXiv:2603.01327v2 Announce Type: replace-cross Abstract: Large language models (LLMs) exhibit strong performance on self-contained programming tasks. However, they still struggle with repository-leve

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

Towards Error-Free EHRs: Reasoning-Intensive Consistency Verification Between Clinical Notes and Structured Tables in Electronic Health Records

DGX agent

arXiv:2605.26463v1 Announce Type: cross Abstract: Data consistency between unstructured clinical notes and structured tables in Electronic Health Records (EHRs) is essential for patient safety and cli

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

AgentGrounder: Zero-Shot 3D Visual Pointcloud Grounding using Multimodal Language Models

DGX agent

arXiv:2605.25901v1 Announce Type: cross Abstract: 3D Visual Grounding (3DVG) is an essential capability for embodied AI, requiring agents to localize objects in 3D scenes based on natural language des

model-releasesarxiv-cs-ro
26 May 2026
Research

Deep Learning-Enabled Prediction of Geoeffective CMEs Using SOHO and SDO Observations

DGX agent

arXiv:2605.24748v1 Announce Type: cross Abstract: Understanding and forecasting the geoeffectiveness of a coronal mass ejection (CME) is crucial for protecting infrastructure in the near-Earth space e

researcharxiv-cs-lg
26 May 2026
Model Releases

From Model Scaling to System Scaling: Scaling the Harness in Agentic AI

DGX agent

arXiv:2605.26112v1 Announce Type: new Abstract: This paper studies the next major bottleneck in agentic AI as system scaling, not only model scaling: the design of auditable, persistent, modular, and

model-releasesarxiv-cs-ai
26 May 2026
Research

Fuzzy PyTorch: Rapid Numerical Variability Evaluation for Deep Learning Models

DGX agent

arXiv:2605.25991v1 Announce Type: new Abstract: We introduce Fuzzy PyTorch, a framework for rapid evaluation of numerical variability in deep learning (DL) models. As DL is increasingly applied to div

researcharxiv-cs-lg
26 May 2026
Model Releases

Insuring Every Action: An Authority Frontier Framework for Runtime Actuarial Control of Autonomous AI Agents

DGX agent

arXiv:2605.25632v1 Announce Type: new Abstract: Autonomous AI agents increasingly issue side-effect-bearing actions: database mutations, refunds, payments, external commitments. We propose the Actuari

model-releasesarxiv-cs-ai
26 May 2026
Agents

Methods for Formal Verification of Agent Skills: Three Layers Toward a Mechanically Checkable Capability-Containment Proof

DGX agent

arXiv:2605.23951v1 Announce Type: new Abstract: The companion paper introduced a four-level verification lattice on agent-skill manifests (unverified, declared, tested, formal) and left the top level

agentsarxiv-cs-ai
26 May 2026
Applications

Positivity in classical enumerative geometry: a case study in synchronized AI-assisted mathematics

DGX agent

arXiv:2605.25271v1 Announce Type: cross Abstract: We study the symmetric polynomial prod_{alphain A_{n,d}}igl(1+alpha_1 x_1+dots+alpha_n x_nigr) where A_{n,d}:={alphainZ_{ge 0}^n:|alpha|=d}, which is

applicationsarxiv-cs-ai
26 May 2026
Agents

PRIMA: Operational Patterns for Resilient Multi-Agent Research with Verifiable Identity and Convergent Feedback

DGX agent

arXiv:2605.24775v1 Announce Type: new Abstract: Operating LLMs as coordinated multi-agent research systems over multi-hour runs surfaces failure modes that single-shot evaluation cannot: upstream prov

agentsarxiv-cs-ai
26 May 2026
Model Releases

Querying structural and functional niches on spatial transcriptomics data

DGX agent

arXiv:2410.10652v4 Announce Type: replace-cross Abstract: Cells in multicellular organisms coordinate to form structural and functional niches. With spatial transcriptomics (ST) enabling gene expressi

model-releasesarxiv-cs-lg
26 May 2026
Agents

Retrieval as Reasoning: Self-Evolving Agent-Native Retrieval via LLM-Wiki

DGX agent

arXiv:2605.25480v1 Announce Type: new Abstract: LLM agents require retrieval to behave less like one-shot context fetching and more like reasoning: searching, reading, traversing, and deciding when ev

agentsarxiv-cs-cl
26 May 2026
Research

TaBIIC2: Interactive Building of Ontological Taxonomies using Weighted Self-Organizing Maps

DGX agent

arXiv:2605.24899v1 Announce Type: new Abstract: Ontologies represent the conceptual knowledge of a domain. At the core of an ontology is the taxonomy of concepts and subconcepts that represent specifi

researcharxiv-cs-ai
26 May 2026
Model Releases

Ax-Prover: A Deep Reasoning Agentic Framework for Theorem Proving in Mathematics and Quantum Physics

DGX agent

arXiv:2510.12787v4 Announce Type: replace Abstract: We present Ax-Prover, a multi-agent system for automated theorem proving in Lean that can solve problems across diverse scientific domains and opera

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

Design and Report Benchmarks for Knowledge Work

DGX agent

arXiv:2605.23262v1 Announce Type: new Abstract: The development of LLM agents has led to a growing body of work on knowledge-work AI, including coding, research, and healthcare. However, current knowl

model-releasesarxiv-cs-ai
25 May 2026
Agents

ExpOS: Explainable Open-Surgery Skills Assessment Using 3D Hand Reconstruction

DGX agent

arXiv:2605.23653v1 Announce Type: new Abstract: Timely and transparent feedback is essential for effective surgical training, yet current assessment remains dependent on expert observation, limiting s

agentsarxiv-cs-cv
25 May 2026
Safety

NeuroNL2LTL: A Neurosymbolic Framework for Natural Language Translation of Linear Temporal Logic

DGX agent

arXiv:2605.22874v1 Announce Type: new Abstract: Effectively translating between natural language (NL) and formal logics like Linear Temporal Logic (LTL) requires expertise that limits formal verificat

safetyarxiv-cs-ai
25 May 2026
Research

Gradient Descent as a Perceptron Algorithm: Understanding Dynamics and Implicit Acceleration

DGX agent

arXiv:2512.11587v2 Announce Type: replace Abstract: Even for the gradient descent (GD) method applied to neural network training, understanding its optimization dynamics, including convergence rate, i

researcharxiv-cs-lg
23 May 2026
Research

Turning Trust to Transactions: Tracking Affiliate Marketing and FTC Compliance in YouTube's Influencer Economy

DGX agent

arXiv:2603.04383v2 Announce Type: replace-cross Abstract: YouTube has evolved into a powerful platform where creators monetize their influence through affiliate marketing, raising concerns about trans

researcharxiv-cs-lg
23 May 2026
Model Releases

HealthCraft: A Reinforcement Learning Safety Environment for Emergency Medicine

DGX agent

arXiv:2605.21496v1 Announce Type: cross Abstract: Frontier language models are being deployed into clinical workflows faster than the infrastructure to evaluate them safely. Static medical-QA benchmar

model-releasesarxiv-cs-cl
22 May 2026
Local Ai

Heartbeat-Bound Hierarchical Credentials: Cryptographic Revocation for AI Agent Swarms

DGX agent

arXiv:2605.20704v1 Announce Type: cross Abstract: Autonomous AI agents that spawn sub-agent swarms create a safety gap: existing credential revocation mechanisms, OAuth~2.0 introspection, OCSP, and W3

local-aiarxiv-cs-ai
22 May 2026
Agents

Quality and Security Signals in AI-Generated Python Refactoring Pull Requests

DGX agent

arXiv:2605.21453v1 Announce Type: cross Abstract: As AI agents increasingly contribute to code development and maintenance, there is still limited empirical evidence on the quality and risk characteri

agentsarxiv-cs-ai
22 May 2026
Model Releases

The Blueprint: How Movix fills a gap in dental skills with specialized agentic AI

DGX agent

Welcome to The Blueprint, a regular feature where we highlight how Google Cloud customers are tackling unique and common challenges across industries using the latest AI and cloud technologies. We hop

model-releasesgoogle-cloud-ai
22 May 2026
Research

Component Influence-Driven Fastener Reduction for Robotic Disassemblability-Aware Design Simplification

DGX agent

arXiv:2605.21026v1 Announce Type: new Abstract: To accelerate automated remanufacturing, robotic disassembly must be considered during the product design phase. However, designers currently lack quant

researcharxiv-cs-ro
21 May 2026
Research

Generative AI Practices, Literacy, and Divides: An Empirical Analysis in the Italian Context

DGX agent

arXiv:2512.03671v2 Announce Type: replace Abstract: The rise of generative AI (GenAI) chatbots accessible via conversational interfaces is transforming digital interactions and holds economic promise.

researcharxiv-cs-cl
21 May 2026
Model Releases

MemGym: a Long-Horizon Memory Environment for LLM Agents

DGX agent

arXiv:2605.20833v1 Announce Type: new Abstract: Memory is a central capability for LLM agents operating across long-horizon tasks. Existing memory benchmarks predominantly evaluate retention of person

model-releasesarxiv-cs-cl
21 May 2026
Tutorials

WiXus: A Wheeled-Legged Robot with Wire-Driven Environmental Utilizing to Integrate Mobility and Manipulation

DGX agent

arXiv:2605.20932v1 Announce Type: new Abstract: Wheeled-legged robots, which have wheels at their feet and achieve high mobility by coordinating wheel drive and leg drive, have been developed. These r

tutorialsarxiv-cs-ro
21 May 2026
Agents

Agentic GraphRAG: Navigating Unstructured Financial Data with Collaborative AI

DGX agent

arXiv:2605.18770v1 Announce Type: cross Abstract: We present a collaborative agentic GraphRAG framework for expert analysis of commercial registry data. Public registries are often formally accessible

agentsarxiv-cs-ai
20 May 2026
Model Releases

Benchmark and optimize LLMs on-device with AI Edge Portal

DGX agent

LLMs have become more powerful at smaller sizes, but deploying them to edge devices like smartphones remains a massive challenge. Today, developers have to optimize across a sprawling combination of a

model-releasesgoogle-cloud-ai
20 May 2026
Research

Can LLMs Estimate Cognitive Complexity of Reading Comprehension Items?

DGX agent

arXiv:2510.25064v2 Announce Type: replace Abstract: Estimating the cognitive complexity of reading comprehension (RC) items is crucial for assessing item difficulty before it is administered to learne

researcharxiv-cs-cl
20 May 2026
Model Releases

Google I/O, Gemini Spark, Antigravity

DGX agent

It's hard to find much to write about Google I/O this year because I have a policy of not writing about anything that I can't try out myself, and a lot of the big announcements are 'coming soon'. I ac

model-releasessimon-willison
20 May 2026
Model Releases

Introducing Agent Executor, Google’s distributed Agent Runtime

DGX agent

As models and harnesses improve, agents are taking on increasingly complex tasks that can run for hours or even days. But as we push agents to do more, this has surfaced a new operational problem: lon

model-releasesgoogle-cloud-ai
20 May 2026
Model Releases

JAXenstein: Accelerated Benchmarking for First-Person Environments

DGX agent

arXiv:2605.19926v1 Announce Type: new Abstract: The progression of reinforcement learning algorithms have been driven by challenging benchmarks. The rate in which a researcher can iterate on a problem

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

Physics-in-the-Loop: A Hybrid Agentic Architecture for Validated CAD Engineering Design

DGX agent

arXiv:2605.19717v1 Announce Type: new Abstract: Large Language Models (LLMs) can generate Computer-Aided Design (CAD), yet lack physical comprehension required for reliable engineering design. Instead

model-releasesarxiv-cs-cv
20 May 2026
Applications

Prognostic Value of Lung Ultrasound Biomarkers for Readmission Risk in Congestive Heart Failure: A Pilot Data-Driven Analysis

DGX agent

arXiv:2605.18878v1 Announce Type: cross Abstract: Hospital readmission within 30 days of discharge is a leading driver of morbidity, mortality, and avoidable healthcare expenditure in congestive heart

applicationsarxiv-cs-cv
20 May 2026
Model Releases

STAR-PolyaMath: Multi-Agent Reasoning under Persistent Meta-Strategic Supervision

DGX agent

arXiv:2605.19338v1 Announce Type: cross Abstract: Frontier AI models and multi-agent systems have led to significant improvements in mathematical reasoning. However, for problems requiring extended, l

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Vision Harnessing Agent for Open Ad-hoc Segmentation

DGX agent

arXiv:2605.19410v1 Announce Type: new Abstract: Segmentation has become easy when the concept is known, requiring retrieval of a learned visual grounding from text. It remains hard for open ad-hoc con

model-releasesarxiv-cs-cv
20 May 2026
Agents

YAC: Bridging Natural Language and Interactive Visual Exploration with Generative AI for Biomedical Data Discovery

DGX agent

arXiv:2509.19182v2 Announce Type: replace-cross Abstract: Incorporating natural language input has the potential to improve the capabilities of biomedical data discovery interfaces. However, user inte

agentsarxiv-cs-ai
20 May 2026
Model Releases

AI for Auto-Research: Roadmap & User Guide

DGX agent

arXiv:2605.18661v1 Announce Type: new Abstract: AI-assisted research is crossing a threshold: fully automated systems can now generate research papers for as little as $15, while long-horizon agents c

model-releasesarxiv-cs-ai
19 May 2026
Safety

AMATA: Adaptive Multi-Agent Trajectory Alignment for Knowledge-Intensive Question Answering

DGX agent

arXiv:2605.17352v1 Announce Type: new Abstract: Despite substantial advances in large language models (LLMs), generating factually consistent responses for knowledge-intensive question answering remai

safetyarxiv-cs-cl
19 May 2026
Model Releases

CHI-Bench: Can AI Agents Automate End-to-End, Long-Horizon, Policy-Rich Healthcare Workflows?

DGX agent

arXiv:2605.16679v1 Announce Type: cross Abstract: End-to-end automation of realistic healthcare operations stresses three capabilities underrepresented in current benchmarks: policy density, decisions

model-releasesarxiv-cs-ai
19 May 2026
← Previous
1…117118119120121…211
Next →