AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
5,189 results
Safety

KAT-Coder-V2.5 Technical Report

DGX agent

arXiv:2607.05471v1 Announce Type: cross Abstract: We present KAT-Coder-V2.5, a coding-focused agentic model trained to act autonomously inside real, executable repositories rather than as a single-tur

safetyarxiv-cs-ai
8 Jul 2026
Agents

VASP Agent: An Agentic Framework for Autonomous First-principles Calculations

DGX agent

arXiv:2512.19458v2 Announce Type: replace Abstract: Large Language Models (LLMs) are increasingly embedded in agentic frameworks for scientific discovery. First-principles materials computation impose

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
agentsarxiv-cs-ai
8 Jul 2026
Agents

Agentic AI-RAN: Enabling Intent-Driven, Explainable and Self-Evolving Open RAN Intelligence

DGX agent

arXiv:2602.24115v2 Announce Type: replace Abstract: Open RAN (O-RAN) exposes rich control and telemetry interfaces across the Non-RT RIC, Near-RT RIC, and distributed units, but also makes it harder t

agentsarxiv-cs-lg
7 Jul 2026
Safety

AGL-1: The Enterprise AI Governance Layer as a Control Plane for Trusted Enterprise Intelligence

DGX agent

arXiv:2607.03516v1 Announce Type: cross Abstract: Enterprise artificial intelligence is moving from isolated experimentation toward operational dependency across copilots, retrieval-augmented generati

safetyarxiv-cs-ai
7 Jul 2026
Applications

Deriving Benchmarking Datasets from Long-Form Recordings: Challenges and Opportunities

DGX agent

arXiv:2607.03201v1 Announce Type: cross Abstract: Long-form recordings (LFRs) of child-centered audio are ecologically valid sources for studying early language development, but three problems limit t

applicationsarxiv-cs-lg
7 Jul 2026
Model Releases

Evaluating Agentic Harness Systems for Autonomous Computational Pathology

DGX agent

arXiv:2607.02598v1 Announce Type: new Abstract: Autonomous computational pathology (ACP) converts high-level pathology analysis goals into executable, traceable and clinically bounded workflows. Reali

model-releasesarxiv-cs-cv
7 Jul 2026
Safety

Multi-Turn On-Policy Distillation with Prefix Replay

DGX agent

arXiv:2607.04763v1 Announce Type: cross Abstract: We study on-policy distillation (OPD) for agentic tasks, where an LLM agent interacts with an environment over multiple turns and a student imitates a

safetyarxiv-cs-ai
7 Jul 2026
Model Releases

OmniLayout: A Schematic-Coupled Multimodal Benchmark for Constraint-Aware Geometric Reasoning in PCB Layout

DGX agent

arXiv:2607.03261v1 Announce Type: new Abstract: Recent large language models (LLMs) have demonstrated remarkable progress in 3D spatial reasoning, spatial grounding, and fine-grained geometric underst

model-releasesarxiv-cs-cv
7 Jul 2026
Safety

SABLE: An NDA-Safe Closed-Loop LLM Framework for Analog Circuit Optimization in Industrial EDA Flows

DGX agent

arXiv:2607.03701v1 Announce Type: cross Abstract: Large language models (LLMs) can propose circuit-optimization decisions, but industrial analog flows cannot expose foundry PDK content, proprietary sc

safetyarxiv-cs-lg
7 Jul 2026
Safety

Saving GPU Hours in LLM Inference System Development and Online Workloads with Simulation and DBMS-Inspired Cache Replacement Policies

DGX agent

arXiv:2411.07447v5 Announce Type: replace-cross Abstract: LLMs are increasingly used world-wide from daily tasks to agentic systems and data analytics, requiring significant GPU resources. While LLM i

safetyarxiv-cs-ai
7 Jul 2026
Agents

Search Beyond What Can Be Taught: Evolving the Knowledge Boundary in Agentic Visual Generation

DGX agent

arXiv:2607.05382v1 Announce Type: cross Abstract: Visual generators excel at rendering, but they confidently fabricate what they do not know. User requests are unbounded, evolving, and deeply long-tai

agentsarxiv-cs-ai
7 Jul 2026
Agents

Self-Specializing Vision-Language Transmon Chip Calibration in a Physics-Grounded Environment

DGX agent

arXiv:2607.03193v1 Announce Type: cross Abstract: Calibrating a superconducting transmon chip is a sequential decision problem under noise, drift, and a finite budget: an expert must choose experiment

agentsarxiv-cs-ai
7 Jul 2026
Research

How Indian Dermatologists are Utilizing Artificial Intelligence for Clinical Practice and Workflow Management: A Nationwide Survey with a Special Focus on atopic dermatitis

DGX agent

arXiv:2607.01252v1 Announce Type: cross Abstract: Background: Dermatology AI has mainly focused on image-based diagnosis, while chronic disease workflows have received less attention. We surveyed Indi

researcharxiv-cs-ai
3 Jul 2026
Research

Rethinking Complexity Metrics for LLM-Integrated Applications: Beyond Source Code

DGX agent

arXiv:2607.01903v1 Announce Type: new Abstract: LLM-integrated applications blend natural language prompts with program code, and much of their runtime behavior originates in the prompt layer rather t

researcharxiv-cs-ai
3 Jul 2026
Research

TokenScope: Token-Level Explainability and Interpretability for Code-Oriented Tasks in Large Language Models

DGX agent

arXiv:2607.01235v1 Announce Type: cross Abstract: Understanding how Large Language Models (LLMs) make token-level decisions during code generation remains a major challenge for both researchers and pr

researcharxiv-cs-ai
3 Jul 2026
Agents

Cheap Code, Costly Judgment: A Case Study on Governable Agentic Software Engineering

DGX agent

arXiv:2607.01087v1 Announce Type: cross Abstract: Generative AI is shifting software engineering from a practice organized around scarce implementation effort toward one organized around abundant, low

agentsarxiv-cs-ai
2 Jul 2026
Research

LRAT-Catcher: Importing SAT Solver Certificates into Lean4 by Reflection

DGX agent

arXiv:2607.00815v1 Announce Type: cross Abstract: SAT solvers settle combinatorial problems beyond the reach of interactive theorem provers and produce LRAT certificates for independent verification.

researcharxiv-cs-ai
2 Jul 2026
Safety

Personalization as Inverse Planning: Learning Latent Design Intents for Agentic Slide Generation via Structural Denoising

DGX agent

arXiv:2607.00407v1 Announce Type: new Abstract: Slide design requires personalizing both deck themes and page layouts. Yet, current AI agent-based methods struggle with fine-grained, page-level design

safetyarxiv-cs-ai
2 Jul 2026
Research

Understanding Guest Preferences and Optimizing Two-sided Marketplaces: Airbnb as an Example

DGX agent

arXiv:2607.00280v1 Announce Type: new Abstract: Airbnb is a community based on connection and belonging -- many hosts on Airbnb are everyday people who share their worlds to provide guests with the fe

researcharxiv-cs-lg
2 Jul 2026
Model Releases

Beyond Compilation: Evaluating Faithful Natural-Language-to-Lean Statement Formalization

DGX agent

arXiv:2606.31002v1 Announce Type: new Abstract: Theorem-proving benchmarks evaluate proof search against fixed formal statements, but natural-language-to-Lean formalization must generate the formal st

model-releasesarxiv-cs-ai
1 Jul 2026
Research

Exploring the relationship between team institutional composition and novelty in academic papers based on fine-grained knowledge entities

DGX agent

arXiv:2606.31058v1 Announce Type: new Abstract: The composition of author teams is an important factor influencing the novelty of academic papers. However, existing studies have paid limited attention

researcharxiv-cs-cl
1 Jul 2026
Agents

Visual Prompt Discovery via Semantic Exploration

DGX agent

arXiv:2603.16250v2 Announce Type: replace-cross Abstract: LVLMs encounter significant challenges in image understanding and visual reasoning, leading to critical perception failures. Visual prompts, w

agentsarxiv-cs-ai
1 Jul 2026
Model Releases

What If We Allocate Test-Time Compute Adaptively?

DGX agent

arXiv:2602.01070v5 Announce Type: replace Abstract: Test-time compute scaling allocates inference computation uniformly, uses fixed sampling strategies, and applies verification only for reranking. In

model-releasesarxiv-cs-cl
1 Jul 2026
Agents

CaveAgent: Transforming LLMs into Stateful Runtime Operators

DGX agent

arXiv:2601.01569v4 Announce Type: replace Abstract: LLM-based agents are increasingly capable of complex task execution, yet current agentic systems remain constrained by text-centric paradigms that s

agentsarxiv-cs-ai
30 Jun 2026
Safety

Characterizing Large Language Model Agentic Workflows: A Study on N8n Ecosystem

DGX agent

arXiv:2606.29116v1 Announce Type: new Abstract: Large Language Models (LLMs) are rapidly being adopted in low-code and no-code automation platforms, where non-expert users design workflows that combin

safetyarxiv-cs-ai
30 Jun 2026
Agents

Evidence-Driven LLM Agent for C-to-Synthesizable-C Conversion and Verification

DGX agent

arXiv:2606.28409v1 Announce Type: cross Abstract: Software-compilable C programs routinely fail to complete the four-stage pipeline of a high-level synthesis (HLS) toolchain -- compilation, C simulati

agentsarxiv-cs-ai
30 Jun 2026
Agents

Experience Graphs: The Data Foundation for Self-Improving Agents

DGX agent

arXiv:2606.29823v1 Announce Type: cross Abstract: The database community has repeatedly advanced the state of the art by recognizing that new workloads demand new system architectures. We argue that l

agentsarxiv-cs-ai
30 Jun 2026
Model Releases

Forensic Trajectory Signatures for Agent Memory Poisoning Detection

DGX agent

arXiv:2606.30566v1 Announce Type: cross Abstract: We discover a behavioral invariant in LLM agents under persistent memory poisoning: in architectures where routing information is retrieved through ob

model-releasesarxiv-cs-lg
30 Jun 2026
Safety

Generalization error of min-norm interpolators in transfer learning

DGX agent

arXiv:2406.13944v2 Announce Type: replace-cross Abstract: This paper establishes the generalization error of pooled min-ell_2-norm interpolation in transfer learning, where data from diverse distribut

safetyarxiv-cs-lg
30 Jun 2026
Model Releases

Governance Decay: How Context Compaction Silently Erases Safety Constraints in Long-Horizon LLM Agents

DGX agent

arXiv:2606.22528v2 Announce Type: replace Abstract: Modern LLM agents increasingly rely on context compaction, summarization, or eviction to keep long-running sessions within a token budget. We show t

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

OptiMUS-0.3: Using Large Language Models to Model and Solve Optimization Problems at Scale

DGX agent

arXiv:2407.19633v4 Announce Type: replace Abstract: Optimization problems are pervasive in sectors from manufacturing and distribution to healthcare. However, most such problems are still solved heuri

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

OSWorld2.0: Benchmarking Computer Use Agents on Long-Horizon Real-World Tasks

DGX agent

arXiv:2606.29537v1 Announce Type: new Abstract: Existing computer-use benchmarks fail to capture the realism, complexity, and long-horizon demands of real-world computer use, limiting their ability to

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

SADL: What to Ignore? A Benchmark for Subject-Aware Distractor Localization

DGX agent

arXiv:2606.30393v1 Announce Type: new Abstract: Photographs frequently contain visual distractors besides foregrounds and backgrounds of the intended subject, competing for attention and weakening com

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

When Does Overlap Help? OSU-Mem and a Cell-Conditional Analysis of Trajectory Memory for LLM Agents

DGX agent

arXiv:2606.28376v1 Announce Type: cross Abstract: Long-horizon large language model (LLM) agents accumulate interaction trajectories that quickly exceed any practical prompt budget, and existing memor

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Unified Zero-Shot Time Series Forecasting: A Darts Foundation

DGX agent

arXiv:2606.27438v1 Announce Type: new Abstract: Since its initial release in 2020, Darts has become a widely used open-source Python library for time series analysis. A series of foundation models hav

model-releasesarxiv-cs-lg
29 Jun 2026
Model Releases

Yuvion LLM: An Adversarially-Aware Large Language Model for Content And AI Safety

DGX agent

arXiv:2606.27632v1 Announce Type: new Abstract: As large language models are increasingly deployed in real-world systems, safety failures can still lead to harmful outputs and dangerous misuse. We arg

model-releasesarxiv-cs-cl
29 Jun 2026
Applications

A Pipeline for Generating Longitudinal Synthetic Clinical Notes Using Large Language Models

DGX agent

arXiv:2606.26879v1 Announce Type: new Abstract: Synthetic data is increasingly used to enable the development and evaluation of AI systems in domains where access to real-world data is restricted. In

applicationsarxiv-cs-ai
26 Jun 2026
Applications

Adversarial Robustness of AI-Generated Image Detectors in the Real World

DGX agent

arXiv:2410.01574v4 Announce Type: replace Abstract: The rapid advancement of Generative Artificial Intelligence (GenAI) capabilities is accompanied by a concerning rise in its misuse. In particular th

applicationsarxiv-cs-cv
26 Jun 2026
Agents

AXLE: A Cloud Infrastructure for Lean 4 Theorem Proving Utilities

DGX agent

arXiv:2606.26442v1 Announce Type: cross Abstract: We present AXLE (Axiom Lean Engine), a cloud service for Lean 4 proof manipulation, extraction, and verification. Recent progress in AI for mathematic

agentsarxiv-cs-ai
26 Jun 2026
Research

FeVOS: Foresight Expression Video Object Segmentation

DGX agent

arXiv:2606.25585v1 Announce Type: new Abstract: Existing Referring Video Object Segmentation tasks focus on referring expressions describing events, actions or appearances of relevant objects within t

researcharxiv-cs-cv
25 Jun 2026
Model Releases

FlowID : Enhancing Forensic Identification with Latent Flow-Matching Models

DGX agent

arXiv:2603.29591v2 Announce Type: replace Abstract: Every day, many people die under violent circumstances, whether from crimes, war, migration, or climate disasters. Medico-legal and law enforcement

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

AdversaBench: Automated LLM Red-Teaming with Multi-Judge Confirmation and Cross-Model Transferability

DGX agent

arXiv:2606.24589v1 Announce Type: new Abstract: Scaling adversarial evaluation of large language models requires both a method for generating hard inputs and a reliable way to confirm that resulting f

model-releasesarxiv-cs-ai
24 Jun 2026
Research

Automatic Part-of-Speech Tagging of Arabic-English Dictionary Senses through WordNet

DGX agent

arXiv:2606.24359v1 Announce Type: new Abstract: This paper proposed an algorithm for part-of-speech (POS) tagging senses of a bilingual dictionary. The algorithm is applied on the Al-Mawrid Arabic-Eng

researcharxiv-cs-cl
24 Jun 2026
Model Releases

DeepBD: A Grounded Agentic Workflow for Variant Prioritization and Diagnosis of Genetic Birth Defects

DGX agent

arXiv:2606.24779v1 Announce Type: cross Abstract: Birth defects are a major cause of fetal loss, neonatal morbidity and long-term disability. In the subset with suspected genetic etiologies, exome and

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

EComAgentBench: Benchmarking Shopping Agents on Long-Horizon Tasks with Distributed Hidden Intent

DGX agent

arXiv:2606.17698v2 Announce Type: replace Abstract: As LLM-based shopping agents enter production, existing benchmarks fail to capture how a shopper's requirements arrive: stated implicitly in the que

model-releasesarxiv-cs-ai
24 Jun 2026
Safety

Red-Teaming the Agentic Red-Team

DGX agent

arXiv:2606.24496v1 Announce Type: cross Abstract: The use of agentic systems to perform offensive security operations has moved from a theoretical possibility to a commoditized capability. However, wh

safetyarxiv-cs-ai
24 Jun 2026
Model Releases

SP-Mind: An Autonomous Reasoning Agent for Spatial Proteomics Analysis

DGX agent

arXiv:2606.24235v1 Announce Type: new Abstract: Spatial proteomics enables single-cell-resolution characterization of protein expression within tissue architecture, playing a critical role in understa

model-releasesarxiv-cs-ai
24 Jun 2026
Safety

Verifiable Foundation Models for Robot Safety

DGX agent

arXiv:2606.23754v1 Announce Type: cross Abstract: Deploying foundation models for robot control raises a central challenge: the expressive power that enables rich, multimodal perception also makes the

safetyarxiv-cs-lg
24 Jun 2026
← Previous
1…4445464748…109
Next →