AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
10,102 results
Safety

Personalization as Inverse Planning: Learning Latent Design Intents for Agentic Slide Generation via Structural Denoising

DGX agent

arXiv:2607.00407v1 Announce Type: new Abstract: Slide design requires personalizing both deck themes and page layouts. Yet, current AI agent-based methods struggle with fine-grained, page-level design

safetyarxiv-cs-ai
2 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Understanding Guest Preferences and Optimizing Two-sided Marketplaces: Airbnb as an Example

DGX agent

arXiv:2607.00280v1 Announce Type: new Abstract: Airbnb is a community based on connection and belonging -- many hosts on Airbnb are everyday people who share their worlds to provide guests with the fe

researcharxiv-cs-lg
2 Jul 2026
Model Releases

AlloyDB AI Functions - now with revolutionary performance boosts and cost savings

DGX agent

AlloyDB is an AI-native database—it isn’t just a passive data store, it intelligently understands and processes your data. With AlloyDB, you get industry-leading vector and hybrid search, near 100% ac

model-releasesgoogle-cloud-ai
1 Jul 2026
Model Releases

Beyond Compilation: Evaluating Faithful Natural-Language-to-Lean Statement Formalization

DGX agent

arXiv:2606.31002v1 Announce Type: new Abstract: Theorem-proving benchmarks evaluate proof search against fixed formal statements, but natural-language-to-Lean formalization must generate the formal st

model-releasesarxiv-cs-ai
1 Jul 2026
Research

Exploring the relationship between team institutional composition and novelty in academic papers based on fine-grained knowledge entities

DGX agent

arXiv:2606.31058v1 Announce Type: new Abstract: The composition of author teams is an important factor influencing the novelty of academic papers. However, existing studies have paid limited attention

researcharxiv-cs-cl
1 Jul 2026
Agents

Visual Prompt Discovery via Semantic Exploration

DGX agent

arXiv:2603.16250v2 Announce Type: replace-cross Abstract: LVLMs encounter significant challenges in image understanding and visual reasoning, leading to critical perception failures. Visual prompts, w

agentsarxiv-cs-ai
1 Jul 2026
Model Releases

What If We Allocate Test-Time Compute Adaptively?

DGX agent

arXiv:2602.01070v5 Announce Type: replace Abstract: Test-time compute scaling allocates inference computation uniformly, uses fixed sampling strategies, and applies verification only for reranking. In

model-releasesarxiv-cs-cl
1 Jul 2026
Agents

CaveAgent: Transforming LLMs into Stateful Runtime Operators

DGX agent

arXiv:2601.01569v4 Announce Type: replace Abstract: LLM-based agents are increasingly capable of complex task execution, yet current agentic systems remain constrained by text-centric paradigms that s

agentsarxiv-cs-ai
30 Jun 2026
Safety

Characterizing Large Language Model Agentic Workflows: A Study on N8n Ecosystem

DGX agent

arXiv:2606.29116v1 Announce Type: new Abstract: Large Language Models (LLMs) are rapidly being adopted in low-code and no-code automation platforms, where non-expert users design workflows that combin

safetyarxiv-cs-ai
30 Jun 2026
Agents

Evidence-Driven LLM Agent for C-to-Synthesizable-C Conversion and Verification

DGX agent

arXiv:2606.28409v1 Announce Type: cross Abstract: Software-compilable C programs routinely fail to complete the four-stage pipeline of a high-level synthesis (HLS) toolchain -- compilation, C simulati

agentsarxiv-cs-ai
30 Jun 2026
Agents

Experience Graphs: The Data Foundation for Self-Improving Agents

DGX agent

arXiv:2606.29823v1 Announce Type: cross Abstract: The database community has repeatedly advanced the state of the art by recognizing that new workloads demand new system architectures. We argue that l

agentsarxiv-cs-ai
30 Jun 2026
Model Releases

Forensic Trajectory Signatures for Agent Memory Poisoning Detection

DGX agent

arXiv:2606.30566v1 Announce Type: cross Abstract: We discover a behavioral invariant in LLM agents under persistent memory poisoning: in architectures where routing information is retrieved through ob

model-releasesarxiv-cs-lg
30 Jun 2026
Safety

Generalization error of min-norm interpolators in transfer learning

DGX agent

arXiv:2406.13944v2 Announce Type: replace-cross Abstract: This paper establishes the generalization error of pooled min-ell_2-norm interpolation in transfer learning, where data from diverse distribut

safetyarxiv-cs-lg
30 Jun 2026
Model Releases

Governance Decay: How Context Compaction Silently Erases Safety Constraints in Long-Horizon LLM Agents

DGX agent

arXiv:2606.22528v2 Announce Type: replace Abstract: Modern LLM agents increasingly rely on context compaction, summarization, or eviction to keep long-running sessions within a token budget. We show t

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

OptiMUS-0.3: Using Large Language Models to Model and Solve Optimization Problems at Scale

DGX agent

arXiv:2407.19633v4 Announce Type: replace Abstract: Optimization problems are pervasive in sectors from manufacturing and distribution to healthcare. However, most such problems are still solved heuri

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

OSWorld2.0: Benchmarking Computer Use Agents on Long-Horizon Real-World Tasks

DGX agent

arXiv:2606.29537v1 Announce Type: new Abstract: Existing computer-use benchmarks fail to capture the realism, complexity, and long-horizon demands of real-world computer use, limiting their ability to

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

SADL: What to Ignore? A Benchmark for Subject-Aware Distractor Localization

DGX agent

arXiv:2606.30393v1 Announce Type: new Abstract: Photographs frequently contain visual distractors besides foregrounds and backgrounds of the intended subject, competing for attention and weakening com

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

What's new in Claude Sonnet 5

DGX agent

What's new in Claude Sonnet 5 Claude Sonnet 5 came out this morning. I always head straight for the 'what's new' developer docs because they tend to have more actionable information than the official

model-releasessimon-willison
30 Jun 2026
Model Releases

When Does Overlap Help? OSU-Mem and a Cell-Conditional Analysis of Trajectory Memory for LLM Agents

DGX agent

arXiv:2606.28376v1 Announce Type: cross Abstract: Long-horizon large language model (LLM) agents accumulate interaction trajectories that quickly exceed any practical prompt budget, and existing memor

model-releasesarxiv-cs-ai
30 Jun 2026
Applications

Turner Industries’ Blueprint for a Secure, Cloud-First Infrastructure

DGX agent

Editor’s note: Today’s post is by Scott Gatreau, Director of Information Security, at Turner Industries, a privately owned industrial contractor that provides construction, maintenance, and fabricatio

applicationsgoogle-cloud-ai
29 Jun 2026
Model Releases

Unified Zero-Shot Time Series Forecasting: A Darts Foundation

DGX agent

arXiv:2606.27438v1 Announce Type: new Abstract: Since its initial release in 2020, Darts has become a widely used open-source Python library for time series analysis. A series of foundation models hav

model-releasesarxiv-cs-lg
29 Jun 2026
Model Releases

Yuvion LLM: An Adversarially-Aware Large Language Model for Content And AI Safety

DGX agent

arXiv:2606.27632v1 Announce Type: new Abstract: As large language models are increasingly deployed in real-world systems, safety failures can still lead to harmful outputs and dangerous misuse. We arg

model-releasesarxiv-cs-cl
29 Jun 2026
Agents

Agreed with this. There's a lot of value in both agents and software (i don't think SaaS is dead), but the product interfaces are different:…

DGX agent

Agreed with this. There's a lot of value in both agents and software (i don't think SaaS is dead), but the product interfaces are different: 1. An agent is like a human. There are a core set of interf

agentsjerry-liu--x
28 Jun 2026
Model Releases

I put together a new article on setting up local coding agents with open-weight models. Everything runs 100% locally. I thought it might be …

DGX agent

I put together a new article on setting up local coding agents with open-weight models. Everything runs 100% locally. I thought it might be useful putting this together because many people asked me ab

model-releasessebastian-raschka--x
27 Jun 2026
Model Releases

The take that frontier models aren't ready for use in medicine is dead wrong. I've been using Opus 4.x in my clinical workflow every day for…

DGX agent

The take that frontier models aren't ready for use in medicine is dead wrong. I've been using Opus 4.x in my clinical workflow every day for 6 months. I'm a deep sub-sub-specialist in dermatology - th

model-releasesjeremy-howard--x
27 Jun 2026
Applications

A Pipeline for Generating Longitudinal Synthetic Clinical Notes Using Large Language Models

DGX agent

arXiv:2606.26879v1 Announce Type: new Abstract: Synthetic data is increasingly used to enable the development and evaluation of AI systems in domains where access to real-world data is restricted. In

applicationsarxiv-cs-ai
26 Jun 2026
Applications

Adversarial Robustness of AI-Generated Image Detectors in the Real World

DGX agent

arXiv:2410.01574v4 Announce Type: replace Abstract: The rapid advancement of Generative Artificial Intelligence (GenAI) capabilities is accompanied by a concerning rise in its misuse. In particular th

applicationsarxiv-cs-cv
26 Jun 2026
Agents

AXLE: A Cloud Infrastructure for Lean 4 Theorem Proving Utilities

DGX agent

arXiv:2606.26442v1 Announce Type: cross Abstract: We present AXLE (Axiom Lean Engine), a cloud service for Lean 4 proof manipulation, extraction, and verification. Recent progress in AI for mathematic

agentsarxiv-cs-ai
26 Jun 2026
Agents

The @n8n_io node for the LlamaParse Platform is now an officially verified community node🦙 Out of the box, you get access to parsing, split…

DGX agent

The @n8n_io node for the LlamaParse Platform is now an officially verified community node🦙 Out of the box, you get access to parsing, splitting, classification, structured data extraction and retrieva

agentsjerry-liu--x
26 Jun 2026
Agents

AI SDK 7 is here. This release sets the foundation for agents and AI platforms in production: approvals, durability, telemetry, and more.

DGX agent

AI SDK 7 is here. This release sets the foundation for agents and AI platforms in production: approvals, durability, telemetry, and more. AI SDK 7 is now available. Introducing: reasoning control, age

agentsvercel--x
25 Jun 2026
Research

FeVOS: Foresight Expression Video Object Segmentation

DGX agent

arXiv:2606.25585v1 Announce Type: new Abstract: Existing Referring Video Object Segmentation tasks focus on referring expressions describing events, actions or appearances of relevant objects within t

researcharxiv-cs-cv
25 Jun 2026
Model Releases

FlowID : Enhancing Forensic Identification with Latent Flow-Matching Models

DGX agent

arXiv:2603.29591v2 Announce Type: replace Abstract: Every day, many people die under violent circumstances, whether from crimes, war, migration, or climate disasters. Medico-legal and law enforcement

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

AdversaBench: Automated LLM Red-Teaming with Multi-Judge Confirmation and Cross-Model Transferability

DGX agent

arXiv:2606.24589v1 Announce Type: new Abstract: Scaling adversarial evaluation of large language models requires both a method for generating hard inputs and a reliable way to confirm that resulting f

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

At Hugging Face we've been building our own agent that we use via Slack (Moon Bot). Honestly, building your own is quite simple and you'll b…

DGX agent

At Hugging Face we've been building our own agent that we use via Slack (Moon Bot). Honestly, building your own is quite simple and you'll be happy you did: any model you want (self-hosted if needed),

model-releasesclem-delangue--x
24 Jun 2026
Research

Automatic Part-of-Speech Tagging of Arabic-English Dictionary Senses through WordNet

DGX agent

arXiv:2606.24359v1 Announce Type: new Abstract: This paper proposed an algorithm for part-of-speech (POS) tagging senses of a bilingual dictionary. The algorithm is applied on the Al-Mawrid Arabic-Eng

researcharxiv-cs-cl
24 Jun 2026
Model Releases

DeepBD: A Grounded Agentic Workflow for Variant Prioritization and Diagnosis of Genetic Birth Defects

DGX agent

arXiv:2606.24779v1 Announce Type: cross Abstract: Birth defects are a major cause of fetal loss, neonatal morbidity and long-term disability. In the subset with suspected genetic etiologies, exome and

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

EComAgentBench: Benchmarking Shopping Agents on Long-Horizon Tasks with Distributed Hidden Intent

DGX agent

arXiv:2606.17698v2 Announce Type: replace Abstract: As LLM-based shopping agents enter production, existing benchmarks fail to capture how a shopper's requirements arrive: stated implicitly in the que

model-releasesarxiv-cs-ai
24 Jun 2026
Safety

Red-Teaming the Agentic Red-Team

DGX agent

arXiv:2606.24496v1 Announce Type: cross Abstract: The use of agentic systems to perform offensive security operations has moved from a theoretical possibility to a commoditized capability. However, wh

safetyarxiv-cs-ai
24 Jun 2026
Model Releases

SP-Mind: An Autonomous Reasoning Agent for Spatial Proteomics Analysis

DGX agent

arXiv:2606.24235v1 Announce Type: new Abstract: Spatial proteomics enables single-cell-resolution characterization of protein expression within tissue architecture, playing a critical role in understa

model-releasesarxiv-cs-ai
24 Jun 2026
Safety

Verifiable Foundation Models for Robot Safety

DGX agent

arXiv:2606.23754v1 Announce Type: cross Abstract: Deploying foundation models for robot control raises a central challenge: the expressive power that enables rich, multimodal perception also makes the

safetyarxiv-cs-lg
24 Jun 2026
Research

Accurate identification and measurement of the precipitate area by two-stage deep neural networks in novel chromium-based alloys

DGX agent

arXiv:2606.22112v1 Announce Type: new Abstract: The performance of advanced materials for extreme environments is underpinned by their microstructure, including the size and distribution of reinforcin

researcharxiv-cs-cv
23 Jun 2026
Tutorials

Decision-Focused Learning: When and Why Traditional Prediction Models Fail

DGX agent

arXiv:2606.21773v1 Announce Type: new Abstract: Plugging predictions of unknown parameters into downstream optimization problems, often referred to as the ``predict-then-optimize'' paradigm, has long

tutorialsarxiv-cs-lg
23 Jun 2026
Research

Multi-Depth Concept Extraction for Post-Hoc Vision Encoder Explanation

DGX agent

arXiv:2411.19700v5 Announce Type: replace Abstract: Explainable AI methods for vision models aim to identify the parts of the input that are important for the final prediction and subsequently relate

researcharxiv-cs-cv
23 Jun 2026
Model Releases

RLM-Cascade: Response-Level Speculative Decoding for Cost-Efficient LLM API Serving

DGX agent

arXiv:2606.22840v1 Announce Type: new Abstract: We present RLM-Cascade, a proxy-layer system that applies speculative decoding at the response level to reduce LLM API costs without requiring model arc

model-releasesarxiv-cs-lg
23 Jun 2026
Safety

have been thinking a bunch about model routing and related things current thoughts here, would love feedback: 1/ there is a difference betwe…

DGX agent

have been thinking a bunch about model routing and related things current thoughts here, would love feedback: 1/ there is a difference between 'model routing' and 'model council' 'model routing' = rou

safetyharrison-chase--x
22 Jun 2026
Tutorials

I’m excited to partner with @swyx to host a workshop at the @aiDotEngineer World’s Fair with Carole Robin (Distinguished @StanfordGSB Teache…

DGX agent

I’m excited to partner with @swyx to host a workshop at the @aiDotEngineer World’s Fair with Carole Robin (Distinguished @StanfordGSB Teacher of “Touchy Feely” & Co-Founder of @LITfellows) on developi

tutorialsswyx--x
22 Jun 2026
Agents

A fundamental problem with extending Codex/Cowork/Code to all knowledge work is that they remain very 'software-brained' where the end resul…

DGX agent

A fundamental problem with extending Codex/Cowork/Code to all knowledge work is that they remain very 'software-brained' where the end result (the software) is what is important & that code serves as

agentsethan-mollick--x
21 Jun 2026
Research

An XAI View on Explainable ASP: Methods, Systems, and Perspectives

DGX agent

arXiv:2601.14764v2 Announce Type: replace Abstract: Answer Set Programming (ASP) is a popular declarative reasoning and problem solving approach in symbolic AI. Its rule-based formalism makes it inher

researcharxiv-cs-ai
11 Jun 2026
← Previous
1…114115116117118…211
Next →