AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
10,088 results
15 Jul 2026

Evidence-Grounded AI for Musculoskeletal Care

Model ReleasesDGX agent

arXiv:2607.12527v1 Announce Type: new Abstract: Musculoskeletal diseases are among the leading causes of disability worldwide and create the greatest global need for rehabilitation. Because recovery,

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models

ResearchDGX agent

arXiv:2509.22415v3 Announce Type: replace-cross Abstract: Multimodal large language models (MLLMs) have achieved strong vision-language performance, yet their token-level visual evidence remains diffi

How Many Tasks Are Enough for Agent Benchmark Decisions? A Replay Analysis of Public LLM Agent Benchmarks

Model ReleasesDGX agent

arXiv:2607.12338v1 Announce Type: new Abstract: Agent benchmarks often compare two agents after all tasks have run, but costly evaluations make partial runs tempting. A task fraction alone does not sh

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

just ran this on my phone (17 pro)

Model ReleasesDGX agent

just ran this on my phone (17 pro) Every on-device AI benchmark you've seen was measured on someone else's phone. So we stopped posting numbers and shipped the benchmark instead. Built into the RunAny

LapSurgie: Humanoid Robots Performing Surgery via Teleoperated Handheld Laparoscopy

ApplicationsDGX agent

arXiv:2510.03529v3 Announce Type: replace Abstract: Robotic laparoscopic surgery has gained increasing attention in recent years for its potential to deliver more efficient and precise minimally invas

Note the current expectations are still around test time compute/more tokens for a given task This is not the case Tokens per task will now …

Model ReleasesDGX agent

Note the current expectations are still around test time compute/more tokens for a given task This is not the case Tokens per task will now drop even as quality improves Cost per intelligence equivale

PM-Bench: Evaluating Prospective Memory in LLM Agents

Model ReleasesDGX agent

arXiv:2607.12385v1 Announce Type: new Abstract: A significant challenge in agentic AI is prospective memory: the ability to execute an intention at a specific future cue or state while other activitie

Suno snatched millions of songs from YouTube, Genius, and Deezer

IndustryDGX agent

Suno data obtained in a hacking incident has exposed that the AI music generator was trained by scraping millions of songs and lyrics from online audio platforms, including YouTube Music, Deezer, and

The GEST-Engine: From Event Graphs to Synthetic Video. A Full Technical Report

AgentsDGX agent

arXiv:2607.12231v1 Announce Type: new Abstract: We present the GEST-Engine, a complete system that goes from natural-language text to fully-annotated multi-actor video. At its core is an explicit worl

14 Jul 2026

Claude at scale on Google Cloud: Frontier AI, built for enterprise production

Model ReleasesDGX agent

Running frontier AI in production is demanding — accelerators to manage, latency to hold steady across continents, regulated data to keep in-region, and long-context requests to serve reliably. Claude

Fable, turn my tweet into a thinkpiece (this was pretty funny): There has never been a better time to have opinions about artificial intelli…

Model ReleasesDGX agent

Fable, turn my tweet into a thinkpiece (this was pretty funny): There has never been a better time to have opinions about artificial intelligence. I say this with some authority, because I am currentl

Together AI positions open-weight AI models as the enterprise moat for cost, control and IP

AgentsDGX agent

Enterprises racing to deploy AI at scale are discovering that the biggest constraint isn’t model capability anymore — it’s control. As agentic AI moves from experimentation into core business processe

13 Jul 2026

Building the AI-defined vehicle with Android, Google Cloud, and Nexus SDV

Model ReleasesDGX agent

The automotive industry is moving from building hardware-centric platforms toward building their own sophisticated Software-Defined Vehicle (SDV) architectures. For OEMs, a vehicle is no longer just a

Computer use in Codex got very good on PC. Asking it to do something on your computer and having the cursor move under the control of a ghos…

Model ReleasesDGX agent

Computer use in Codex got very good on PC. Asking it to do something on your computer and having the cursor move under the control of a ghost is one of the things that makes you viscerally realize how

New model for AMD Strix Halo users: My 198B Step 3.7 Flash release was a big hit, but this one may be even better: 298B-parameter Hy3, now r…

Model ReleasesDGX agent

New model for AMD Strix Halo users: My 198B Step 3.7 Flash release was a big hit, but this one may be even better: 298B-parameter Hy3, now running on a new 2-bit FPX codebook designed to map efficient

10 Jul 2026

Cognitive-structured Multimodal Agent for Multimodal Understanding, Generation, and Editing

Model ReleasesDGX agent

arXiv:2607.08497v1 Announce Type: cross Abstract: Recent unified multimodal models show a single architecture can jointly perform vision/language understanding and image generation/editing. However, t

Design optimization and robustness analysis of rigid-link flapping mechanisms

ApplicationsDGX agent

arXiv:2503.21204v3 Announce Type: replace Abstract: Rigid link flapping mechanisms remain the most practical choice for flapping wing micro-aerial vehicles (MAVs) to carry useful payloads and onboard

Self-EvolveRec: Self-Evolving Recommender Systems with LLM-based Directional Feedback

ResearchDGX agent

arXiv:2602.12612v2 Announce Type: replace-cross Abstract: Traditional methods for automating recommender system design, such as Neural Architecture Search (NAS), are often constrained by a fixed searc

Token-Flow Firewall: Semantic Runtime Auditing for Persistent AI Agents

Local AiDGX agent

arXiv:2607.08395v1 Announce Type: cross Abstract: Persistent AI agents extend large language models (LLMs) beyond single-turn interaction into long-lived software systems. Unlike traditional chat assi

TrackStudio: An Integrated Toolkit for Markerless Tracking

TutorialsDGX agent

arXiv:2511.07624v3 Announce Type: replace Abstract: Markerless motion tracking has advanced rapidly in the past 10 years and currently offers powerful opportunities for behavioural, clinical, and biom

9 Jul 2026

Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts

ResearchDGX agent

arXiv:2607.06611v1 Announce Type: cross Abstract: Automatically recognizing the sentiment, positive or negative, from speech is a challenging task, requiring both the analysis of vocal inflections and

From Beats to Breaches:How Offensive AI Infers Sensitive User Information from Playlists

Model ReleasesDGX agent

arXiv:2605.04724v2 Announce Type: replace-cross Abstract: The pervasive integration of AI has enabled Offensive AI: the exploitation of AI for malicious ends across the cyber-kill chain. A critical ma

Predicting LLM Safety Before Release by Simulating Deployment

Model ReleasesDGX agent

arXiv:2607.07184v1 Announce Type: cross Abstract: Pre-deployment safety evaluations aim to inform the downstream risks of releasing a new AI model. Yet most evaluations provide limited evidence about

Safely run AI-generated code in Cloud Run sandboxes

Model ReleasesDGX agent

Here’s a question we hear often at Google Cloud: How do you safely run AI-generated code or untrusted binaries without putting your host application, data, and cloud credentials at risk? In other word

SpaCellAgent: A Self-Evolving LLM-Based Multi-Agent Framework for Trajectory Analysis

AgentsDGX agent

arXiv:2607.07467v1 Announce Type: new Abstract: Spatial and Single-cell transcriptomics are transformative in deciphering cellular dynamics. As the fundamental paradigm for reconstructing cell develop

8 Jul 2026

KAT-Coder-V2.5 Technical Report

SafetyDGX agent

arXiv:2607.05471v1 Announce Type: cross Abstract: We present KAT-Coder-V2.5, a coding-focused agentic model trained to act autonomously inside real, executable repositories rather than as a single-tur

Prompt: ANNALS — a living kingdom in a single file Working title behavior: the app titles itself per seed — 'The Annals of Vaelmere', 'The A…

Model ReleasesDGX agent

Prompt: ANNALS — a living kingdom in a single file Working title behavior: the app titles itself per seed — 'The Annals of Vaelmere', 'The Annals of Osterholt' — because the central conceit is that yo

VASP Agent: An Agentic Framework for Autonomous First-principles Calculations

AgentsDGX agent

arXiv:2512.19458v2 Announce Type: replace Abstract: Large Language Models (LLMs) are increasingly embedded in agentic frameworks for scientific discovery. First-principles materials computation impose

7 Jul 2026

Agentic AI-RAN: Enabling Intent-Driven, Explainable and Self-Evolving Open RAN Intelligence

AgentsDGX agent

arXiv:2602.24115v2 Announce Type: replace Abstract: Open RAN (O-RAN) exposes rich control and telemetry interfaces across the Non-RT RIC, Near-RT RIC, and distributed units, but also makes it harder t

AGL-1: The Enterprise AI Governance Layer as a Control Plane for Trusted Enterprise Intelligence

SafetyDGX agent

arXiv:2607.03516v1 Announce Type: cross Abstract: Enterprise artificial intelligence is moving from isolated experimentation toward operational dependency across copilots, retrieval-augmented generati

Deriving Benchmarking Datasets from Long-Form Recordings: Challenges and Opportunities

ApplicationsDGX agent

arXiv:2607.03201v1 Announce Type: cross Abstract: Long-form recordings (LFRs) of child-centered audio are ecologically valid sources for studying early language development, but three problems limit t

Evaluating Agentic Harness Systems for Autonomous Computational Pathology

Model ReleasesDGX agent

arXiv:2607.02598v1 Announce Type: new Abstract: Autonomous computational pathology (ACP) converts high-level pathology analysis goals into executable, traceable and clinically bounded workflows. Reali

Loop engineering is great until something breaks. Here is how I improve the reliability of my agentic loops. I use human-in-the-loop (HITL).…

Model ReleasesDGX agent

Loop engineering is great until something breaks. Here is how I improve the reliability of my agentic loops. I use human-in-the-loop (HITL). It's easy and extremely effective. Anyone can build this. M

Multi-Turn On-Policy Distillation with Prefix Replay

SafetyDGX agent

arXiv:2607.04763v1 Announce Type: cross Abstract: We study on-policy distillation (OPD) for agentic tasks, where an LLM agent interacts with an environment over multiple turns and a student imitates a

OmniLayout: A Schematic-Coupled Multimodal Benchmark for Constraint-Aware Geometric Reasoning in PCB Layout

Model ReleasesDGX agent

arXiv:2607.03261v1 Announce Type: new Abstract: Recent large language models (LLMs) have demonstrated remarkable progress in 3D spatial reasoning, spatial grounding, and fine-grained geometric underst

Report: 83% of organizations need to upgrade their infrastructure to support agentic AI

Model ReleasesDGX agent

For years, enterprise AI has been synonymous with conversational AI — the customer service bots and digital assistants we interact with every day. But today, the market has shifted. We’ve officially m

SABLE: An NDA-Safe Closed-Loop LLM Framework for Analog Circuit Optimization in Industrial EDA Flows

SafetyDGX agent

arXiv:2607.03701v1 Announce Type: cross Abstract: Large language models (LLMs) can propose circuit-optimization decisions, but industrial analog flows cannot expose foundry PDK content, proprietary sc

Saving GPU Hours in LLM Inference System Development and Online Workloads with Simulation and DBMS-Inspired Cache Replacement Policies

SafetyDGX agent

arXiv:2411.07447v5 Announce Type: replace-cross Abstract: LLMs are increasingly used world-wide from daily tasks to agentic systems and data analytics, requiring significant GPU resources. While LLM i

Search Beyond What Can Be Taught: Evolving the Knowledge Boundary in Agentic Visual Generation

AgentsDGX agent

arXiv:2607.05382v1 Announce Type: cross Abstract: Visual generators excel at rendering, but they confidently fabricate what they do not know. User requests are unbounded, evolving, and deeply long-tai

Self-Specializing Vision-Language Transmon Chip Calibration in a Physics-Grounded Environment

AgentsDGX agent

arXiv:2607.03193v1 Announce Type: cross Abstract: Calibrating a superconducting transmon chip is a sequential decision problem under noise, drift, and a finite budget: an expert must choose experiment

6 Jul 2026

my keynote at AI Engineer World Fair: “A Field Guide to Fable” is live on YouTube! https://youtu.be/9fubhllmsBU?is=ejZTRy8t85FIbSGQ

TutorialsDGX agent

A keynote presentation titled 'A Field Guide to Fable' was delivered at AI Engineer World Fair and is available on YouTube. The talk, given by Thariq, likely covers practical guidance or best practice

4 Jul 2026

We will look back on their work with wonder - how could humans have built this with keyboards alone?

ApplicationsDGX agent

Ethan Mollick reflects on how future generations will be amazed at the complexity and scale of work humans accomplished using only traditional keyboards and interfaces, before presumably more advanced

3 Jul 2026

Another major Grok Build update just landed, packed with new features, extensive bug fixes, and meaningful performance improvements Release …

SafetyDGX agent

Another major Grok Build update just landed, packed with new features, extensive bug fixes, and meaningful performance improvements Release Notes: v0.2.84 — 2026-07-03 Features: • Announcements now up

Don't train the model, evolve the harness. I read a brilliant blog post from Hugging Face where they took a frozen open model scoring 0% on …

Model ReleasesDGX agent

Don't train the model, evolve the harness. I read a brilliant blog post from Hugging Face where they took a frozen open model scoring 0% on a hard legal agent benchmark, left its weights alone, and le

How Indian Dermatologists are Utilizing Artificial Intelligence for Clinical Practice and Workflow Management: A Nationwide Survey with a Special Focus on atopic dermatitis

ResearchDGX agent

arXiv:2607.01252v1 Announce Type: cross Abstract: Background: Dermatology AI has mainly focused on image-based diagnosis, while chronic disease workflows have received less attention. We surveyed Indi

Rethinking Complexity Metrics for LLM-Integrated Applications: Beyond Source Code

ResearchDGX agent

arXiv:2607.01903v1 Announce Type: new Abstract: LLM-integrated applications blend natural language prompts with program code, and much of their runtime behavior originates in the prompt layer rather t

TokenScope: Token-Level Explainability and Interpretability for Code-Oriented Tasks in Large Language Models

ResearchDGX agent

arXiv:2607.01235v1 Announce Type: cross Abstract: Understanding how Large Language Models (LLMs) make token-level decisions during code generation remains a major challenge for both researchers and pr

2 Jul 2026

Cheap Code, Costly Judgment: A Case Study on Governable Agentic Software Engineering

AgentsDGX agent

arXiv:2607.01087v1 Announce Type: cross Abstract: Generative AI is shifting software engineering from a practice organized around scarce implementation effort toward one organized around abundant, low

LRAT-Catcher: Importing SAT Solver Certificates into Lean4 by Reflection

ResearchDGX agent

arXiv:2607.00815v1 Announce Type: cross Abstract: SAT solvers settle combinatorial problems beyond the reach of interactive theorem provers and produce LRAT certificates for independent verification.

Personalization as Inverse Planning: Learning Latent Design Intents for Agentic Slide Generation via Structural Denoising

SafetyDGX agent

arXiv:2607.00407v1 Announce Type: new Abstract: Slide design requires personalizing both deck themes and page layouts. Yet, current AI agent-based methods struggle with fine-grained, page-level design

Understanding Guest Preferences and Optimizing Two-sided Marketplaces: Airbnb as an Example

ResearchDGX agent

arXiv:2607.00280v1 Announce Type: new Abstract: Airbnb is a community based on connection and belonging -- many hosts on Airbnb are everyday people who share their worlds to provide guests with the fe

1 Jul 2026

AlloyDB AI Functions - now with revolutionary performance boosts and cost savings

Model ReleasesDGX agent

AlloyDB is an AI-native database—it isn’t just a passive data store, it intelligently understands and processes your data. With AlloyDB, you get industry-leading vector and hybrid search, near 100% ac

Beyond Compilation: Evaluating Faithful Natural-Language-to-Lean Statement Formalization

Model ReleasesDGX agent

arXiv:2606.31002v1 Announce Type: new Abstract: Theorem-proving benchmarks evaluate proof search against fixed formal statements, but natural-language-to-Lean formalization must generate the formal st

Exploring the relationship between team institutional composition and novelty in academic papers based on fine-grained knowledge entities

ResearchDGX agent

arXiv:2606.31058v1 Announce Type: new Abstract: The composition of author teams is an important factor influencing the novelty of academic papers. However, existing studies have paid limited attention

Visual Prompt Discovery via Semantic Exploration

AgentsDGX agent

arXiv:2603.16250v2 Announce Type: replace-cross Abstract: LVLMs encounter significant challenges in image understanding and visual reasoning, leading to critical perception failures. Visual prompts, w

What If We Allocate Test-Time Compute Adaptively?

Model ReleasesDGX agent

arXiv:2602.01070v5 Announce Type: replace Abstract: Test-time compute scaling allocates inference computation uniformly, uses fixed sampling strategies, and applies verification only for reranking. In

30 Jun 2026

CaveAgent: Transforming LLMs into Stateful Runtime Operators

AgentsDGX agent

arXiv:2601.01569v4 Announce Type: replace Abstract: LLM-based agents are increasingly capable of complex task execution, yet current agentic systems remain constrained by text-centric paradigms that s

Characterizing Large Language Model Agentic Workflows: A Study on N8n Ecosystem

SafetyDGX agent

arXiv:2606.29116v1 Announce Type: new Abstract: Large Language Models (LLMs) are rapidly being adopted in low-code and no-code automation platforms, where non-expert users design workflows that combin

Evidence-Driven LLM Agent for C-to-Synthesizable-C Conversion and Verification

AgentsDGX agent

arXiv:2606.28409v1 Announce Type: cross Abstract: Software-compilable C programs routinely fail to complete the four-stage pipeline of a high-level synthesis (HLS) toolchain -- compilation, C simulati

Experience Graphs: The Data Foundation for Self-Improving Agents

AgentsDGX agent

arXiv:2606.29823v1 Announce Type: cross Abstract: The database community has repeatedly advanced the state of the art by recognizing that new workloads demand new system architectures. We argue that l

← Previous
1…9091929394…169
Next →