AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
10,102 results
Research

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models

DGX agent

arXiv:2509.22415v3 Announce Type: replace-cross Abstract: Multimodal large language models (MLLMs) have achieved strong vision-language performance, yet their token-level visual evidence remains diffi

researcharxiv-cs-ai
15 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

How Many Tasks Are Enough for Agent Benchmark Decisions? A Replay Analysis of Public LLM Agent Benchmarks

DGX agent

arXiv:2607.12338v1 Announce Type: new Abstract: Agent benchmarks often compare two agents after all tasks have run, but costly evaluations make partial runs tempting. A task fraction alone does not sh

model-releasesarxiv-cs-ai
15 Jul 2026
Model Releases

just ran this on my phone (17 pro)

DGX agent

just ran this on my phone (17 pro) Every on-device AI benchmark you've seen was measured on someone else's phone. So we stopped posting numbers and shipped the benchmark instead. Built into the RunAny

model-releasesyohei-nakajima--x
15 Jul 2026
Applications

LapSurgie: Humanoid Robots Performing Surgery via Teleoperated Handheld Laparoscopy

DGX agent

arXiv:2510.03529v3 Announce Type: replace Abstract: Robotic laparoscopic surgery has gained increasing attention in recent years for its potential to deliver more efficient and precise minimally invas

applicationsarxiv-cs-ro
15 Jul 2026
Model Releases

Note the current expectations are still around test time compute/more tokens for a given task This is not the case Tokens per task will now …

DGX agent

Note the current expectations are still around test time compute/more tokens for a given task This is not the case Tokens per task will now drop even as quality improves Cost per intelligence equivale

model-releasesemad-mostaque--x
15 Jul 2026
Model Releases

PM-Bench: Evaluating Prospective Memory in LLM Agents

DGX agent

arXiv:2607.12385v1 Announce Type: new Abstract: A significant challenge in agentic AI is prospective memory: the ability to execute an intention at a specific future cue or state while other activitie

model-releasesarxiv-cs-ai
15 Jul 2026
Industry

Suno snatched millions of songs from YouTube, Genius, and Deezer

DGX agent

Suno data obtained in a hacking incident has exposed that the AI music generator was trained by scraping millions of songs and lyrics from online audio platforms, including YouTube Music, Deezer, and

industrythe-verge-ai
15 Jul 2026
Agents

The GEST-Engine: From Event Graphs to Synthetic Video. A Full Technical Report

DGX agent

arXiv:2607.12231v1 Announce Type: new Abstract: We present the GEST-Engine, a complete system that goes from natural-language text to fully-annotated multi-actor video. At its core is an explicit worl

agentsarxiv-cs-cv
15 Jul 2026
Model Releases

Claude at scale on Google Cloud: Frontier AI, built for enterprise production

DGX agent

Running frontier AI in production is demanding — accelerators to manage, latency to hold steady across continents, regulated data to keep in-region, and long-context requests to serve reliably. Claude

model-releasesgoogle-cloud-ai
14 Jul 2026
Model Releases

Fable, turn my tweet into a thinkpiece (this was pretty funny): There has never been a better time to have opinions about artificial intelli…

DGX agent

Fable, turn my tweet into a thinkpiece (this was pretty funny): There has never been a better time to have opinions about artificial intelligence. I say this with some authority, because I am currentl

model-releasesethan-mollick--x
14 Jul 2026
Agents

Together AI positions open-weight AI models as the enterprise moat for cost, control and IP

DGX agent

Enterprises racing to deploy AI at scale are discovering that the biggest constraint isn’t model capability anymore — it’s control. As agentic AI moves from experimentation into core business processe

agentssiliconangle
14 Jul 2026
Model Releases

Building the AI-defined vehicle with Android, Google Cloud, and Nexus SDV

DGX agent

The automotive industry is moving from building hardware-centric platforms toward building their own sophisticated Software-Defined Vehicle (SDV) architectures. For OEMs, a vehicle is no longer just a

model-releasesgoogle-cloud-ai
13 Jul 2026
Model Releases

Computer use in Codex got very good on PC. Asking it to do something on your computer and having the cursor move under the control of a ghos…

DGX agent

Computer use in Codex got very good on PC. Asking it to do something on your computer and having the cursor move under the control of a ghost is one of the things that makes you viscerally realize how

model-releasesethan-mollick--x
13 Jul 2026
Model Releases

New model for AMD Strix Halo users: My 198B Step 3.7 Flash release was a big hit, but this one may be even better: 298B-parameter Hy3, now r…

DGX agent

New model for AMD Strix Halo users: My 198B Step 3.7 Flash release was a big hit, but this one may be even better: 298B-parameter Hy3, now running on a new 2-bit FPX codebook designed to map efficient

model-releasesclem-delangue--x
13 Jul 2026
Model Releases

Cognitive-structured Multimodal Agent for Multimodal Understanding, Generation, and Editing

DGX agent

arXiv:2607.08497v1 Announce Type: cross Abstract: Recent unified multimodal models show a single architecture can jointly perform vision/language understanding and image generation/editing. However, t

model-releasesarxiv-cs-ai
10 Jul 2026
Applications

Design optimization and robustness analysis of rigid-link flapping mechanisms

DGX agent

arXiv:2503.21204v3 Announce Type: replace Abstract: Rigid link flapping mechanisms remain the most practical choice for flapping wing micro-aerial vehicles (MAVs) to carry useful payloads and onboard

applicationsarxiv-cs-ro
10 Jul 2026
Research

Self-EvolveRec: Self-Evolving Recommender Systems with LLM-based Directional Feedback

DGX agent

arXiv:2602.12612v2 Announce Type: replace-cross Abstract: Traditional methods for automating recommender system design, such as Neural Architecture Search (NAS), are often constrained by a fixed searc

researcharxiv-cs-ai
10 Jul 2026
Local Ai

Token-Flow Firewall: Semantic Runtime Auditing for Persistent AI Agents

DGX agent

arXiv:2607.08395v1 Announce Type: cross Abstract: Persistent AI agents extend large language models (LLMs) beyond single-turn interaction into long-lived software systems. Unlike traditional chat assi

local-aiarxiv-cs-cl
10 Jul 2026
Tutorials

TrackStudio: An Integrated Toolkit for Markerless Tracking

DGX agent

arXiv:2511.07624v3 Announce Type: replace Abstract: Markerless motion tracking has advanced rapidly in the past 10 years and currently offers powerful opportunities for behavioural, clinical, and biom

tutorialsarxiv-cs-cv
10 Jul 2026
Research

Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts

DGX agent

arXiv:2607.06611v1 Announce Type: cross Abstract: Automatically recognizing the sentiment, positive or negative, from speech is a challenging task, requiring both the analysis of vocal inflections and

researcharxiv-cs-ai
9 Jul 2026
Model Releases

From Beats to Breaches:How Offensive AI Infers Sensitive User Information from Playlists

DGX agent

arXiv:2605.04724v2 Announce Type: replace-cross Abstract: The pervasive integration of AI has enabled Offensive AI: the exploitation of AI for malicious ends across the cyber-kill chain. A critical ma

model-releasesarxiv-cs-ai
9 Jul 2026
Model Releases

Predicting LLM Safety Before Release by Simulating Deployment

DGX agent

arXiv:2607.07184v1 Announce Type: cross Abstract: Pre-deployment safety evaluations aim to inform the downstream risks of releasing a new AI model. Yet most evaluations provide limited evidence about

model-releasesarxiv-cs-ai
9 Jul 2026
Model Releases

Safely run AI-generated code in Cloud Run sandboxes

DGX agent

Here’s a question we hear often at Google Cloud: How do you safely run AI-generated code or untrusted binaries without putting your host application, data, and cloud credentials at risk? In other word

model-releasesgoogle-cloud-ai
9 Jul 2026
Agents

SpaCellAgent: A Self-Evolving LLM-Based Multi-Agent Framework for Trajectory Analysis

DGX agent

arXiv:2607.07467v1 Announce Type: new Abstract: Spatial and Single-cell transcriptomics are transformative in deciphering cellular dynamics. As the fundamental paradigm for reconstructing cell develop

agentsarxiv-cs-ai
9 Jul 2026
Safety

KAT-Coder-V2.5 Technical Report

DGX agent

arXiv:2607.05471v1 Announce Type: cross Abstract: We present KAT-Coder-V2.5, a coding-focused agentic model trained to act autonomously inside real, executable repositories rather than as a single-tur

safetyarxiv-cs-ai
8 Jul 2026
Model Releases

Prompt: ANNALS — a living kingdom in a single file Working title behavior: the app titles itself per seed — 'The Annals of Vaelmere', 'The A…

DGX agent

Prompt: ANNALS — a living kingdom in a single file Working title behavior: the app titles itself per seed — 'The Annals of Vaelmere', 'The Annals of Osterholt' — because the central conceit is that yo

model-releasesethan-mollick--x
8 Jul 2026
Agents

VASP Agent: An Agentic Framework for Autonomous First-principles Calculations

DGX agent

arXiv:2512.19458v2 Announce Type: replace Abstract: Large Language Models (LLMs) are increasingly embedded in agentic frameworks for scientific discovery. First-principles materials computation impose

agentsarxiv-cs-ai
8 Jul 2026
Agents

Agentic AI-RAN: Enabling Intent-Driven, Explainable and Self-Evolving Open RAN Intelligence

DGX agent

arXiv:2602.24115v2 Announce Type: replace Abstract: Open RAN (O-RAN) exposes rich control and telemetry interfaces across the Non-RT RIC, Near-RT RIC, and distributed units, but also makes it harder t

agentsarxiv-cs-lg
7 Jul 2026
Safety

AGL-1: The Enterprise AI Governance Layer as a Control Plane for Trusted Enterprise Intelligence

DGX agent

arXiv:2607.03516v1 Announce Type: cross Abstract: Enterprise artificial intelligence is moving from isolated experimentation toward operational dependency across copilots, retrieval-augmented generati

safetyarxiv-cs-ai
7 Jul 2026
Applications

Deriving Benchmarking Datasets from Long-Form Recordings: Challenges and Opportunities

DGX agent

arXiv:2607.03201v1 Announce Type: cross Abstract: Long-form recordings (LFRs) of child-centered audio are ecologically valid sources for studying early language development, but three problems limit t

applicationsarxiv-cs-lg
7 Jul 2026
Model Releases

Evaluating Agentic Harness Systems for Autonomous Computational Pathology

DGX agent

arXiv:2607.02598v1 Announce Type: new Abstract: Autonomous computational pathology (ACP) converts high-level pathology analysis goals into executable, traceable and clinically bounded workflows. Reali

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

Loop engineering is great until something breaks. Here is how I improve the reliability of my agentic loops. I use human-in-the-loop (HITL).…

DGX agent

Loop engineering is great until something breaks. Here is how I improve the reliability of my agentic loops. I use human-in-the-loop (HITL). It's easy and extremely effective. Anyone can build this. M

model-releasesdair-ai--x
7 Jul 2026
Safety

Multi-Turn On-Policy Distillation with Prefix Replay

DGX agent

arXiv:2607.04763v1 Announce Type: cross Abstract: We study on-policy distillation (OPD) for agentic tasks, where an LLM agent interacts with an environment over multiple turns and a student imitates a

safetyarxiv-cs-ai
7 Jul 2026
Model Releases

OmniLayout: A Schematic-Coupled Multimodal Benchmark for Constraint-Aware Geometric Reasoning in PCB Layout

DGX agent

arXiv:2607.03261v1 Announce Type: new Abstract: Recent large language models (LLMs) have demonstrated remarkable progress in 3D spatial reasoning, spatial grounding, and fine-grained geometric underst

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

Report: 83% of organizations need to upgrade their infrastructure to support agentic AI

DGX agent

For years, enterprise AI has been synonymous with conversational AI — the customer service bots and digital assistants we interact with every day. But today, the market has shifted. We’ve officially m

model-releasesgoogle-cloud-ai
7 Jul 2026
Safety

SABLE: An NDA-Safe Closed-Loop LLM Framework for Analog Circuit Optimization in Industrial EDA Flows

DGX agent

arXiv:2607.03701v1 Announce Type: cross Abstract: Large language models (LLMs) can propose circuit-optimization decisions, but industrial analog flows cannot expose foundry PDK content, proprietary sc

safetyarxiv-cs-lg
7 Jul 2026
Safety

Saving GPU Hours in LLM Inference System Development and Online Workloads with Simulation and DBMS-Inspired Cache Replacement Policies

DGX agent

arXiv:2411.07447v5 Announce Type: replace-cross Abstract: LLMs are increasingly used world-wide from daily tasks to agentic systems and data analytics, requiring significant GPU resources. While LLM i

safetyarxiv-cs-ai
7 Jul 2026
Agents

Search Beyond What Can Be Taught: Evolving the Knowledge Boundary in Agentic Visual Generation

DGX agent

arXiv:2607.05382v1 Announce Type: cross Abstract: Visual generators excel at rendering, but they confidently fabricate what they do not know. User requests are unbounded, evolving, and deeply long-tai

agentsarxiv-cs-ai
7 Jul 2026
Agents

Self-Specializing Vision-Language Transmon Chip Calibration in a Physics-Grounded Environment

DGX agent

arXiv:2607.03193v1 Announce Type: cross Abstract: Calibrating a superconducting transmon chip is a sequential decision problem under noise, drift, and a finite budget: an expert must choose experiment

agentsarxiv-cs-ai
7 Jul 2026
Tutorials

my keynote at AI Engineer World Fair: “A Field Guide to Fable” is live on YouTube! https://youtu.be/9fubhllmsBU?is=ejZTRy8t85FIbSGQ

DGX agent

A keynote presentation titled 'A Field Guide to Fable' was delivered at AI Engineer World Fair and is available on YouTube. The talk, given by Thariq, likely covers practical guidance or best practice

tutorialsthariq--x
6 Jul 2026
Applications

We will look back on their work with wonder - how could humans have built this with keyboards alone?

DGX agent

Ethan Mollick reflects on how future generations will be amazed at the complexity and scale of work humans accomplished using only traditional keyboards and interfaces, before presumably more advanced

applicationsethan-mollick--x
4 Jul 2026
Safety

Another major Grok Build update just landed, packed with new features, extensive bug fixes, and meaningful performance improvements Release …

DGX agent

Another major Grok Build update just landed, packed with new features, extensive bug fixes, and meaningful performance improvements Release Notes: v0.2.84 — 2026-07-03 Features: • Announcements now up

safetyelon-musk--x
3 Jul 2026
Model Releases

Don't train the model, evolve the harness. I read a brilliant blog post from Hugging Face where they took a frozen open model scoring 0% on …

DGX agent

Don't train the model, evolve the harness. I read a brilliant blog post from Hugging Face where they took a frozen open model scoring 0% on a hard legal agent benchmark, left its weights alone, and le

model-releasesclem-delangue--x
3 Jul 2026
Research

How Indian Dermatologists are Utilizing Artificial Intelligence for Clinical Practice and Workflow Management: A Nationwide Survey with a Special Focus on atopic dermatitis

DGX agent

arXiv:2607.01252v1 Announce Type: cross Abstract: Background: Dermatology AI has mainly focused on image-based diagnosis, while chronic disease workflows have received less attention. We surveyed Indi

researcharxiv-cs-ai
3 Jul 2026
Research

Rethinking Complexity Metrics for LLM-Integrated Applications: Beyond Source Code

DGX agent

arXiv:2607.01903v1 Announce Type: new Abstract: LLM-integrated applications blend natural language prompts with program code, and much of their runtime behavior originates in the prompt layer rather t

researcharxiv-cs-ai
3 Jul 2026
Research

TokenScope: Token-Level Explainability and Interpretability for Code-Oriented Tasks in Large Language Models

DGX agent

arXiv:2607.01235v1 Announce Type: cross Abstract: Understanding how Large Language Models (LLMs) make token-level decisions during code generation remains a major challenge for both researchers and pr

researcharxiv-cs-ai
3 Jul 2026
Agents

Cheap Code, Costly Judgment: A Case Study on Governable Agentic Software Engineering

DGX agent

arXiv:2607.01087v1 Announce Type: cross Abstract: Generative AI is shifting software engineering from a practice organized around scarce implementation effort toward one organized around abundant, low

agentsarxiv-cs-ai
2 Jul 2026
Research

LRAT-Catcher: Importing SAT Solver Certificates into Lean4 by Reflection

DGX agent

arXiv:2607.00815v1 Announce Type: cross Abstract: SAT solvers settle combinatorial problems beyond the reach of interactive theorem provers and produce LRAT certificates for independent verification.

researcharxiv-cs-ai
2 Jul 2026
← Previous
1…113114115116117…211
Next →