AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
5,189 results
Model Releases

MultiUAV-Plat: An LLM-Oriented Platform, Benchmark and Framework for Multi-UAV Collaborative Task Planning

DGX agent

arXiv:2606.31073v1 Announce Type: new Abstract: Large language models (LLMs) provide a promising interface for high-level robotic task planning, but their use in multi-UAV collaboration remains diffic

model-releasesarxiv-cs-ai
1 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

Nazrin: An Atomic Neural Proof Automation Tactic in Lean 4

DGX agent

arXiv:2602.18767v3 Announce Type: replace-cross Abstract: In Machine-Assisted Theorem Proving, a theorem proving agent searches for a sequence of expressions and tactics that can prove a statement in

agentsarxiv-cs-lg
1 Jul 2026
Agents

OpenLife: Toward Open-World Artificial Life with Autonomous LLM Agents

DGX agent

arXiv:2606.31046v1 Announce Type: new Abstract: Artificial life has explored life-like behavior on many computational substrates, but mostly in researcher-designed closed worlds. We argue that large l

agentsarxiv-cs-ai
1 Jul 2026
Safety

Rethinking On-policy Optimization for Query Augmentation

DGX agent

arXiv:2510.17139v3 Announce Type: replace Abstract: Recent advances in large language models (LLMs) have led to a surge of interest in query augmentation for information retrieval (IR). Two main appro

safetyarxiv-cs-cl
1 Jul 2026
Model Releases

RigorBench: Benchmarking Engineering Process Discipline in Autonomous AI Coding Agents

DGX agent

arXiv:2606.22678v2 Announce Type: replace-cross Abstract: Agentic coding harnesses - such as Agent-Skills, Superpowers, and Agent-Rigor - are increasingly deployed to augment underlying LLMs for real-

model-releasesarxiv-cs-ai
1 Jul 2026
Agents

ShopX: A Foundation Model for Intent-to-Item Fulfillment in Agentic Shopping

DGX agent

arXiv:2606.31693v1 Announce Type: cross Abstract: The wave of AI-native applications is moving shopping beyond page- and feed-based browsing toward intent-driven experiences orchestrated by LLM agents

agentsarxiv-cs-ai
1 Jul 2026
Model Releases

Signed-Permutation Coordinate Transport for RMSNorm Transformers

DGX agent

arXiv:2606.31963v1 Announce Type: cross Abstract: Modern LLM workflows move coordinate-indexed objects across checkpoints: steering vectors, sparse autoencoders, top-k neuron sets, attribution lists,

model-releasesarxiv-cs-cl
1 Jul 2026
Agents

Using AI Agents to Automate Black-Box Audits of Personalization Algorithms at Scale

DGX agent

arXiv:2606.30801v1 Announce Type: new Abstract: Personalization algorithms determine what content users encounter on online platforms. Auditing these systems is difficult because independent auditors

agentsarxiv-cs-cl
1 Jul 2026
Model Releases

When Does Learning to Stop Help? A Cost-Aware Study of Early Exits in Reasoning Models

DGX agent

arXiv:2606.30852v1 Announce Type: new Abstract: Reasoning models spend different amounts of useful computation across instances, but it remains unclear when a learned stopping rule improves over simpl

model-releasesarxiv-cs-ai
1 Jul 2026
Safety

A causal modeling perspective on decision theory

DGX agent

arXiv:2606.29911v1 Announce Type: new Abstract: Decision theory provides a formal framework for how agents should make choices under uncertainty, drawing on ideas from philosophy, probability, and cau

safetyarxiv-cs-ai
30 Jun 2026
Local Ai

A Hybrid Framework For Crypto-Ransomware Detection In Enterprise Shared Storage

DGX agent

arXiv:2606.30586v1 Announce Type: cross Abstract: Most corporate workplace environments enforce policies and technical controls that limit the storage of sensitive data on client endpoints. Consequent

local-aiarxiv-cs-lg
30 Jun 2026
Model Releases

A Machine-Verified Proof of a Quantum-Optimization Conjecture

DGX agent

arXiv:2606.29687v1 Announce Type: cross Abstract: We report a machine-verified resolution of a problem open for over a decade in quantum optimization: the Farhi, Goldstone and Gutmann (FGG) conjecture

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

A Physics-Grounded Benchmark for Multi-Agent Dynamics in World Models

DGX agent

arXiv:2606.28757v1 Announce Type: new Abstract: Generative world models hold immense promise as scalable simulators for autonomous systems, particularly for synthesizing rare but safety-critical multi

model-releasesarxiv-cs-cv
30 Jun 2026
Agents

Adam's Law: Textual Frequency Law on Large Language Models

DGX agent

arXiv:2604.02176v3 Announce Type: replace Abstract: While textual frequency has been validated as relevant to human cognition in reading speed, its relatedness to Large Language Models (LLMs) is seldo

agentsarxiv-cs-cl
30 Jun 2026
Model Releases

Anisotropy Decides Cosine vs. Rank Metrics for Text Embeddings

DGX agent

arXiv:2606.29571v1 Announce Type: new Abstract: The standard way to compare two text embeddings is cosine similarity. Scattered studies report that a different metric does better, but never pin down t

model-releasesarxiv-cs-cl
30 Jun 2026
Local Ai

Anomaly Factory 3D: A Modular Framework for Diverse Pseudo-Anomaly Synthesis in Unsupervised 3D Anomaly Detection

DGX agent

arXiv:2606.29181v1 Announce Type: cross Abstract: Detecting and localizing defects in 3D point clouds is challenging because abnormal samples are scarce and diverse, while training is often limited to

local-aiarxiv-cs-ai
30 Jun 2026
Agents

AutoB2G: Agentic Simulation and Reinforcement Learning for Spatio-Temporal Grid-Interactive Building Control

DGX agent

arXiv:2603.26005v2 Announce Type: replace Abstract: Grid-interactive building control has emerged as a promising approach for improving demand-side flexibility in modern power systems. Realistic studi

agentsarxiv-cs-ai
30 Jun 2026
Research

Bricker to BRACE: A Bracket Exposure RAW Dataset and Restoration Model for Flicker-Banding

DGX agent

arXiv:2606.29845v1 Announce Type: new Abstract: Flicker-banding (FB), arises from temporal aliasing between a camera's rolling shutter and a display's brightness modulation, degrading screen-captured

researcharxiv-cs-cv
30 Jun 2026
Model Releases

CLQT: A Closed-Loop, Cost-Aware, Strategy-Consistent Benchmark for Diagnostic Evaluation of LLM Portfolio-Management Agents

DGX agent

arXiv:2606.29771v1 Announce Type: new Abstract: LLM agents are increasingly cast as autonomous portfolio managers, and benchmarks have moved from financial question-answering to sequential trading. Ye

model-releasesarxiv-cs-ai
30 Jun 2026
Agents

Code Reasoning for Software Engineering Tasks: A Survey and A Call to Action

DGX agent

arXiv:2506.13932v3 Announce Type: replace-cross Abstract: The rise of large language models (LLMs) has led to dramatic improvements across a wide range of natural language tasks. Their performance on

agentsarxiv-cs-ai
30 Jun 2026
Model Releases

COHORT: Collaborative Orchestration for Hardening via Offensive Replay on Emulated Topologies

DGX agent

arXiv:2606.30479v1 Announce Type: cross Abstract: Mitigating an observed adversary in an enterprise network typically takes weeks of expert work: an analyst derives a mitigation tailored to that adver

model-releasesarxiv-cs-ai
30 Jun 2026
Research

Concentration bounds on response-based vector embeddings of black-box generative models

DGX agent

arXiv:2511.08307v2 Announce Type: replace-cross Abstract: Generative models, such as large language models or text-to-image diffusion models, can generate relevant responses to user-given queries. Res

researcharxiv-cs-lg
30 Jun 2026
Model Releases

Conversational Query Engine for Mixed-Modality Heterogeneous Enterprise Data Sources

DGX agent

arXiv:2606.28370v1 Announce Type: cross Abstract: Enterprise business intelligence queries span structured warehouses and unstructured document repositories -- modalities with fundamentally different

model-releasesarxiv-cs-ai
30 Jun 2026
Safety

Critical Interval MSE: Toward Reliable Offline Validation for Robot Manipulation Policies

DGX agent

arXiv:2606.29898v1 Announce Type: cross Abstract: Real-world evaluation is the gold standard for robot policies because it tests them against the physical conditions and deployment challenges they are

safetyarxiv-cs-ai
30 Jun 2026
Research

Cybersecurity is the True Frontier for Generative AI Success or Failure

DGX agent

arXiv:2606.28929v1 Announce Type: cross Abstract: Cybersecurity is a real-life test-bed for many machine learning problems at once, especially when considering modern strides in using Large Language M

researcharxiv-cs-lg
30 Jun 2026
Agents

DEEPMED Search: An Open-Source Agentic Platform for Medical Deep Research with Introspective Verification

DGX agent

arXiv:2606.29746v1 Announce Type: new Abstract: Navigating the deluge of heterogeneous medical data, from academic literature (PubMed) to clinical guidelines (Web) and private knowledge bases, remains

agentsarxiv-cs-ai
30 Jun 2026
Agents

DeepTrans Studio: Turning Expert Interventions into Shared Team Knowledge in Agentic Translation Workflows

DGX agent

arXiv:2606.29727v1 Announce Type: new Abstract: Professional translation is often a team-based process: translators, reviewers, and project managers must coordinate terminology, legal force, and accou

agentsarxiv-cs-ai
30 Jun 2026
Model Releases

DIA-HARM: Dialectal Disparities in Harmful Content Detection Across 50 English Dialects

DGX agent

arXiv:2604.05318v2 Announce Type: replace Abstract: Harmful content detectors, particularly disinformation classifiers, are predominantly developed and evaluated on Standard American English (SAE), le

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

Diagnosing and Mitigating Retrieval Bottlenecks in LLM-Based Cold-Start Recommendation

DGX agent

arXiv:2606.29947v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as rerankers in recommender systems, with the expectation that semantic understanding will help in

model-releasesarxiv-cs-lg
30 Jun 2026
Model Releases

Early Estimation of Language to Latent Alignment in Diffusion Models

DGX agent

arXiv:2512.08505v2 Announce Type: replace Abstract: Conditional diffusion models frequently suffer from language-image misalignments. Due to the ambiguity of intermediate noise corrupted latents, asse

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

EvalSafetyGap: A Hybrid Survey and Conceptual Framework for LLM Evaluation-Safety Failures

DGX agent

arXiv:2606.30219v1 Announce Type: new Abstract: LLM evaluation and AI safety face a shared measurement problem: benchmark scores, reward-model signals, and reported safety metrics can improve while th

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Evaluating Memory in LLM Agents via Incremental Multi-Turn Interactions

DGX agent

arXiv:2507.05257v4 Announce Type: replace-cross Abstract: Recent benchmarks for Large Language Model (LLM) agents primarily focus on evaluating reasoning, planning, and execution capabilities, while a

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Factorizable Normalizing Flows for parameter-dependent density morphing

DGX agent

arXiv:2606.30489v1 Announce Type: cross Abstract: Normalizing Flows excel at modeling a single fixed density, yet many problems across the sciences, such as high energy physics, instead require modeli

model-releasesarxiv-cs-lg
30 Jun 2026
Research

Fast Numbers, Slow Language: Bridging Quantitative and Qualitative Earnings Signals

DGX agent

arXiv:2606.29734v1 Announce Type: new Abstract: Earnings announcements release two types of information sequentially: quantitative surprise (numeric earnings-per-share (EPS)/revenue versus analyst est

researcharxiv-cs-cl
30 Jun 2026
Model Releases

Hierarchical Experimentalist Agents

DGX agent

arXiv:2606.29315v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used to take actions in the real world and support human decision-making, yet most agents rely on parametr

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

How much of an LLM-generated clinical corpus is actually new? A production-scale measurement of content redundancy for provenance classification

DGX agent

arXiv:2606.29605v1 Announce Type: new Abstract: Clinical machine learning increasingly relies on training corpora generated by large language models (LLMs) rather than annotated by clinicians, and suc

model-releasesarxiv-cs-cl
30 Jun 2026
Research

IG-Lens: Exact Additive Probability Attribution Across Transformer Layers via Telescoping Integrated Gradients

DGX agent

arXiv:2606.29693v1 Announce Type: new Abstract: We ask a simple question about decoder-only transformers: between which two layers is the probability of a predicted token actually produced? Existing l

researcharxiv-cs-lg
30 Jun 2026
Tutorials

Informational Frustration in Neural Manifolds: Shannon Bottlenecks and the Limits of Learnability

DGX agent

arXiv:2606.30512v1 Announce Type: cross Abstract: Why overparameterised deep networks generalise so remarkably well remains one of the most stubborn open questions in machine learning theory. Classica

tutorialsarxiv-cs-ai
30 Jun 2026
Safety

InsertAnywhere: Geometrically Grounded and Optics-Aware Video Object Insertion

DGX agent

arXiv:2512.17504v2 Announce Type: replace-cross Abstract: Recent advances in diffusion models have enabled impressive video editing capabilities, yet production-grade Video Object Insertion (VOI) rema

safetyarxiv-cs-ai
30 Jun 2026
Local Ai

KernelSight-LM: A Kernel-Level LLM Inference Simulator

DGX agent

arXiv:2606.28565v1 Announce Type: cross Abstract: As large language models (LLMs) move into production serving, practitioners must rapidly evaluate inference performance across diverse hardware, model

local-aiarxiv-cs-ai
30 Jun 2026
Agents

LAMP: Lean-based Agentic framework with MCP and Proof Repair

DGX agent

arXiv:2606.28841v1 Announce Type: cross Abstract: Large language models are increasingly capable of mathematical reasoning, but the proofs they generate are often unreliable and hard to verify. Intera

agentsarxiv-cs-ai
30 Jun 2026
Model Releases

LLM-Guided Planning for Multi-hop Reasoning over Multimodal Nuclear Regulatory Documents

DGX agent

arXiv:2606.29399v1 Announce Type: new Abstract: Reviewing nuclear regulatory documents requires multi-hop reasoning across tens of thousands of pages, where judgments depend on evidence assembled acro

model-releasesarxiv-cs-ai
30 Jun 2026
Safety

LLM Semantic Signaling Game and Mechanism Design: Systematic Blindness, Awareness Shaping, and Mindset Dynamics

DGX agent

arXiv:2606.29113v1 Announce Type: cross Abstract: Large language models (LLMs) increasingly mediate strategic interactions through natural language, making semantic control a critical element of commu

safetyarxiv-cs-ai
30 Jun 2026
Model Releases

Making Multimodal LLMs Reliable Chart Data Extractors: A Benchmark and Training Framework

DGX agent

arXiv:2606.29808v1 Announce Type: cross Abstract: Chart data extraction, which reverse-engineers data tables from chart images, is essential for reproducibility, analysis, retrieval, and redesign. Exi

model-releasesarxiv-cs-ai
30 Jun 2026
Safety

Mechanistically Eliciting Latent Behaviors in Language Models

DGX agent

arXiv:2606.29604v1 Announce Type: cross Abstract: We aim to discover diverse, generalizable perturbations of LLM internals that can surface hidden behavioral modes. Such perturbations could help resha

safetyarxiv-cs-ai
30 Jun 2026
Applications

Meshtryoshka: Differentiable Rendering of Real-World Scenes via Mesh Rasterization

DGX agent

arXiv:2606.28622v1 Announce Type: new Abstract: Differentiable rendering has emerged as a powerful approach for 3D reconstruction and novel view synthesis. State-of-the-art differentiable rendering me

applicationsarxiv-cs-cv
30 Jun 2026
Model Releases

MirrorCode: AI can rebuild entire programs from behavior alone

DGX agent

arXiv:2606.30182v1 Announce Type: new Abstract: AI models are rapidly improving at autonomous coding, as shown by benchmark progress and one-off demonstrations such as AI implementing a C compiler. Ho

model-releasesarxiv-cs-ai
30 Jun 2026
Safety

Mitigating the Safety-utility Trade-off in LLM Alignment via Adaptive Safe Context Learning

DGX agent

arXiv:2602.13562v2 Announce Type: replace-cross Abstract: While reasoning models have achieved remarkable success in complex reasoning tasks, their increasing power necessitates stringent safety measu

safetyarxiv-cs-ai
30 Jun 2026
← Previous
1…6768697071…109
Next →