AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
20,981 results
Model Releases

When the Judge Should Not Decide: Evidence-Locked, Non-Compensatory Selection Bounds LLM-Judge Failure in Reasoning Pipelines

DGX agent

arXiv:2608.07813v1 Announce Type: new Abstract: An LLM judge deployed inside a reasoning pipeline does not merely measure quality, it decides which answer ships. We show that the cost of that decision

model-releasesarxiv-cs-ai
11 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Local Ai

Who Bridges Safety? Identifying and Targeting Cross-Lingual Shared Safety Pathways

DGX agent

arXiv:2608.09095v1 Announce Type: new Abstract: Uncovering the internal mechanisms underlying the safety capabilities of large language models (LLMs) is crucial for developing trustworthy artificial i

local-aiarxiv-cs-ai
11 Aug 2026
Safety

Who Built This Model? Tracing LLM Lineage via Spectral Fingerprints in Weight Space

DGX agent

arXiv:2608.07786v1 Announce Type: new Abstract: Open-weight large language models (LLMs) are increasingly developed through complex, multi-stage pipelines, leading to intricate lineage relationships t

safetyarxiv-cs-ai
11 Aug 2026
Model Releases

Who Verifies the Benchmark? Decentralizing Trust in Large Language Model Evaluation

DGX agent

arXiv:2608.07762v1 Announce Type: new Abstract: LLM benchmarks can build an organization's reputation and attract customers, but only when results are transparent and verifiable. Unverified claims tha

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Why Does the Future Branch? Identifiable Closure Tests for Stochastic Physical World Models

DGX agent

arXiv:2608.00591v2 Announce Type: replace Abstract: A calibrated stochastic world model can reveal how uncertain a future is without revealing why it branches. The same conditional future law can aris

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

WorldSimProbe: Diagnosing Simulator Faithfulness in Action-Conditioned World Models for Embodied Manipulation

DGX agent

arXiv:2608.09298v1 Announce Type: cross Abstract: Action-conditioned world models (ACWMs) promise to provide embodied AI with scalable predictive simulators for planning, policy evaluation, and data g

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

WuYuEval: A Multi-Level Benchmark for Large Language Models in Solid Waste Management

DGX agent

arXiv:2608.07529v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as technical assistants, but their competence in solid waste management (SWM) remains difficult to

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

X2C: A Dataset Featuring Nuanced Facial Expressions for Realistic Humanoid Imitation

DGX agent

arXiv:2505.11146v3 Announce Type: replace-cross Abstract: Fine-grained facial expression transfer from humans to humanoid agents presents a unique pattern recognition challenge due to the significant

model-releasesarxiv-cs-ai
11 Aug 2026
Safety

Yesterday's Shield, Today's Spear: A Self-Evolving Safety Guardrail in Production

DGX agent

arXiv:2608.08471v1 Announce Type: new Abstract: Deployed LLM safety guardrails are predominantly static: trained once and frozen at release, while new jailbreak techniques and previously un-addressed

safetyarxiv-cs-ai
11 Aug 2026
Model Releases

Your Prompt Is Not the Only Prompt: How Much Do LLMs Weight Structured-Output Schema Descriptions?

DGX agent

arXiv:2608.08254v1 Announce Type: new Abstract: Structured output, where an LLM populates a predefined JSON schema, has become a default mechanism for data labeling and information extraction, but it

model-releasesarxiv-cs-ai
11 Aug 2026
Research

ZetaGPT: A Reference Implementation of Positional--Encoding--Free State--Space--Attention Language Models

DGX agent

arXiv:2608.09432v1 Announce Type: cross Abstract: Transformer-based language models rely on self-attention, whose computation is permutation-equivariant and therefore lacks an intrinsic mechanism for

researcharxiv-cs-ai
11 Aug 2026
Model Releases

ZhuLong: Execution-Grounded LLM Agent for EDA Scripting with Offline API Self-Exploration

DGX agent

arXiv:2608.07925v1 Announce Type: new Abstract: EDA scripting with tool-specific, often undocumented APIs remains a long-tail bottleneck that existing LLMs fail to address. This paper presents ZhuLong

model-releasesarxiv-cs-ai
11 Aug 2026
Research

A Finite E-Group of Nilpotency Class Three

DGX agent

arXiv:2608.07275v1 Announce Type: cross Abstract: A group is an E-group if every element commutes with each of its endomorphic images. Caranti asked whether a finite E-group can have nilpotency class

researcharxiv-cs-ai
10 Aug 2026
Local Ai

A MARL Centered Reference Architecture for Large Language Model Augmentation in Smart Manufacturing

DGX agent

arXiv:2608.07148v1 Announce Type: new Abstract: Modern manufacturing imposes six coupled demands on adaptive control: local decisions with global consequences, partial observability, nonstationarity,

local-aiarxiv-cs-ai
10 Aug 2026
Model Releases

A Multi-Agent Framework for Automated Coarse-Grained Molecular Dynamics of Polymers

DGX agent

arXiv:2608.06694v1 Announce Type: new Abstract: Coarse-grained (CG) molecular dynamics extends polymer simulation beyond the scales accessible to all-atom (AA) methods, but bottom-up CG modeling is la

model-releasesarxiv-cs-ai
10 Aug 2026
Research

A Physics-Inspired Classical Digital Twin of Cortical Dynamics: A Band-Stratified Metriplectic Port-Hamiltonian Neural Network Learned from Brain-Computer-Interface EEG

DGX agent

arXiv:2607.10439v3 Announce Type: replace-cross Abstract: We present a physics-inspired classical digital twin of brain-computer- interface (BCI) data: a graph neural network constrained to a band-str

researcharxiv-cs-ai
10 Aug 2026
Model Releases

A Picture is Worth a Thousand Tokens: How Vision Language Models Cut AI Energy Costs While Improving Accuracy

DGX agent

arXiv:2608.07427v1 Announce Type: new Abstract: LLM inference accounts for over 90% of AI operational energy, scaling directly with input token count---a critical inefficiency for telecom network anal

model-releasesarxiv-cs-ai
10 Aug 2026
Research

A primer on optimal transport for causal inference with observational data

DGX agent

arXiv:2503.07811v3 Announce Type: replace-cross Abstract: The theory of optimal transportation has developed into a powerful and elegant framework for comparing probability distributions, with wide-ra

researcharxiv-cs-ai
10 Aug 2026
Model Releases

Accounting Graph Transformer for Short-History Multi-KPI Forecasting in Small Businesses

DGX agent

arXiv:2608.07037v1 Announce Type: cross Abstract: Small businesses often have only 12-24 months of accounting history, yet planning and risk workflows require coordinated forecasts across financial st

model-releasesarxiv-cs-ai
10 Aug 2026
Agents

ADIAS: Automated Design of Interactive Agentic Systems

DGX agent

arXiv:2608.06410v1 Announce Type: new Abstract: Automated agent design improves agent harnesses through iterative revision, evaluation, and feedback summarization. Existing methods are largely candida

agentsarxiv-cs-ai
10 Aug 2026
Model Releases

Aftab: A Comprehensive Benchmark of CNN Encoders and Advanced Value Functions in Parallelized Q-Networks

DGX agent

arXiv:2608.07335v1 Announce Type: cross Abstract: Recent advancements in deep reinforcement learning have increasingly favored simplified, highly parallelized paradigms. Notably, the Parallelized Q-Ne

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Agent Memory Distillation: Empowering Small LLM Agents with Hierarchical Teacher Memory

DGX agent

arXiv:2608.07169v1 Announce Type: new Abstract: Memory systems have shown promise for improving agent performance, but their potential remains largely unexplored for small language models, which strug

model-releasesarxiv-cs-ai
10 Aug 2026
Agents

Agentic AI: User Empowerment or Enclosure?

DGX agent

arXiv:2608.06510v1 Announce Type: cross Abstract: Agentic AI promises a more flexible form of digital agency: systems that can act on users' behalf, from filtering content to negotiating prices to sel

agentsarxiv-cs-ai
10 Aug 2026
Agents

Agentic Planning for Symbolic Execution

DGX agent

arXiv:2608.06397v1 Announce Type: cross Abstract: Symbolic execution seeks to explore feasible program paths, yet a practical run may exhaust its resources while much program behaviour remains unreach

agentsarxiv-cs-ai
10 Aug 2026
Agents

AgentPatch: Coarse-to-Fine Weak-Task Repair for Merging Agentic Multimodal Large Language Models

DGX agent

arXiv:2608.06699v1 Announce Type: new Abstract: Agentic multimodal large language models (MLLMs) extend multimodal perception and reasoning with planning, tool use, and interaction in dynamic environm

agentsarxiv-cs-ai
10 Aug 2026
Agents

An Agentic Hybrid Top-Down and Bottom-Up Approach to Knowledge Graph Generation

DGX agent

arXiv:2608.07023v1 Announce Type: cross Abstract: Organizing thousands of unstandardized, multilingual expertise declarations is a persistent challenge for Human Resources (HR) platforms, directly imp

agentsarxiv-cs-ai
10 Aug 2026
Agents

An End-to-End Agent Auditing Engine

DGX agent

arXiv:2608.07346v1 Announce Type: new Abstract: With the rapid advancement of large language models (LLMs), harnesses have become essential infrastructure for deploying agents across a wide range of d

agentsarxiv-cs-ai
10 Aug 2026
Model Releases

Artificial Intelligence Can Match Domain Experts in Evidence Extraction and Critical Appraisal of Microbial Oncogenesis Research Publications

DGX agent

arXiv:2608.07250v1 Announce Type: cross Abstract: Confirmed oncogenic microbes contribute significantly to cancer burden. Identifying novel microbial oncogenicity could yield strategies that will redu

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Ask-E: An Environment for Calibrated Question Generation

DGX agent

arXiv:2608.06933v1 Announce Type: cross Abstract: Today, we improve models by training and evaluating them on problems at the frontier of their abilities. Creating such problems is itself a demanding

model-releasesarxiv-cs-ai
10 Aug 2026
Applications

Assessing AI-generated music detection in real-world broadcast monitoring

DGX agent

arXiv:2608.07359v1 Announce Type: cross Abstract: The proliferation of AI-generated music in broadcast media raises concerns about transparency and fair compensation, but reliable detection under real

applicationsarxiv-cs-ai
10 Aug 2026
Research

Authoring and Management of Transparent Research Integrity Assessments of Randomised Clinical Trial Publications Using LLM-assisted Tools and Provenance Knowledge Graphs

DGX agent

arXiv:2608.07202v1 Announce Type: new Abstract: Systematic reviews of Randomised Controlled Trials (RCTs) are routinely used as evidence for clinical care guidelines. Such evidence has to meet high re

researcharxiv-cs-ai
10 Aug 2026
Safety

AutoIntervene: Calibrated Intervention for Action-Chunking Imitation Learning Policies

DGX agent

arXiv:2608.07065v1 Announce Type: cross Abstract: Action-chunking visuomotor policies learn from demonstrations and improve temporal consistency by predicting short action sequences rather than single

safetyarxiv-cs-ai
10 Aug 2026
Safety

Automated item evaluation: Predicting item acceptance and rejection using LLM-generated critiques

DGX agent

arXiv:2608.06609v1 Announce Type: new Abstract: Automated item evaluation (AIE) refers to the use of computational methods to assess item quality without requiring manual expert review or field testin

safetyarxiv-cs-ai
10 Aug 2026
Model Releases

AutoMOOSE: An Agentic AI for Autonomous Phase-Field Simulation

DGX agent

arXiv:2603.20986v2 Announce Type: replace Abstract: Phase-field modeling links thermodynamics and kinetics to microstructural evolution, but multiphysics frameworks such as MOOSE require expertise to

model-releasesarxiv-cs-ai
10 Aug 2026
Agents

Autonomous discovery of accelerator commissioning algorithms

DGX agent

arXiv:2608.07138v1 Announce Type: cross Abstract: Simulated commissioning has become essential for de-risking modern light-source design and commissioning, but the procedures being simulated are still

agentsarxiv-cs-ai
10 Aug 2026
Applications

Autonomy-of-Heads: Data-Free Sparse Attention from Frozen Query-Key Geometry

DGX agent

arXiv:2608.06849v1 Announce Type: cross Abstract: Long-context LLM inference is bottlenecked by quadratic attention computation and growing KV-cache costs. Existing sparse attention and KV-compression

applicationsarxiv-cs-ai
10 Aug 2026
Research

Beyond 'AI Language': The case for the idiolectal nature of LLM output

DGX agent

arXiv:2608.06589v1 Announce Type: cross Abstract: While large language model outputs are frequently analysed as a collective super variety termed 'AI language,' this chapter argues that this perspecti

researcharxiv-cs-ai
10 Aug 2026
Model Releases

Beyond Foundation Models: Dimension-Aware Neural Architecture Search with Small-Data Representation Models for Cryocooler Lifetime Prediction

DGX agent

arXiv:2608.06993v1 Announce Type: cross Abstract: Large-scale pretrained time-series models achieve strong results through large-scale pretraining and task-agnostic representation learning, but they r

model-releasesarxiv-cs-ai
10 Aug 2026
Research

Beyond Isolation: Unlocking Reinforcement Learning Component Synergy for Sample-Efficient Continuous Control

DGX agent

arXiv:2608.07086v1 Announce Type: cross Abstract: Reinforcement learning systems are significantly more complex than other machine learning paradigms due to inherent properties, causing RL system desi

researcharxiv-cs-ai
10 Aug 2026
Research

Beyond Routing Weights: Faithful Response-Level Interpretation of Mixture-of-Experts Reward Models via Contribution Contrast

DGX agent

arXiv:2608.06400v1 Announce Type: new Abstract: Reward models are central to learning from human preferences, yet identifying what drives their predictions remains challenging. Recent sparse Mixture-o

researcharxiv-cs-ai
10 Aug 2026
Model Releases

Beyond Starry Night: Shortcut-Aware Control-State Planning for Artist-Grounded Text to Image Generation

DGX agent

arXiv:2608.06751v1 Announce Type: cross Abstract: Artist-grounded image generation requires more than appending an artist name to a prompt. Image models often respond to artist names through canonical

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Beyond Text Matching: Towards Reference-Free Evaluation for Human-Oriented Binary Reverse Engineering

DGX agent

arXiv:2608.07038v1 Announce Type: cross Abstract: Human-Oriented Binary Reverse Engineering (HOBRE) aims to transform decompiled pseudocode into a more human-friendly representation, thereby reducing

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Beyond the Black Box: Interpretable Models of Human Randomisation Failures

DGX agent

arXiv:2608.07220v1 Announce Type: new Abstract: Mixed strategy equilibrium predicts i.i.d play: past actions should not help predict future decisions. Human players, however, systematically depart fro

model-releasesarxiv-cs-ai
10 Aug 2026
Safety

bioMoR: Biology-Guided Mixture-of-Recursions for Effective Genomic Learning

DGX agent

arXiv:2608.06727v1 Announce Type: new Abstract: Transformer models for high-dimensional omics analysis process thousands of genes or pathways, although only a subset requires deep computation. Mixture

safetyarxiv-cs-ai
10 Aug 2026
Agents

Blast Radius

DGX agent

arXiv:2608.07440v1 Announce Type: new Abstract: Agentic coding faces growing problems of affordability and wasted tokens. We introduce Blast Radius, a predictive memory management layer that estimates

agentsarxiv-cs-ai
10 Aug 2026
Safety

Blind to the Pivotal Vote: Aggregate Independence Metrics Miss Where Verification Actually Helps

DGX agent

arXiv:2608.06940v1 Announce Type: new Abstract: LLM judge panels are a standard evaluation tool, but prior work reports highly correlated panel errors: nine judges provide roughly the effective inform

safetyarxiv-cs-ai
10 Aug 2026
Agents

BONSAI: Evolvability-Guided Tree Search over Skills

DGX agent

arXiv:2608.07056v1 Announce Type: new Abstract: A skill is a naturallanguage document that steers a frozen agent whose weights cannot be updated so any capability the agent lacks must be supplied in p

agentsarxiv-cs-ai
10 Aug 2026
Local Ai

Boundary Density Likelihood for Direct Event-Time Supervision

DGX agent

arXiv:2408.12792v2 Announce Type: replace Abstract: Event detection turns long recordings into a sparse set of ranked timestamps. Yet many sequence models are trained for samplewise segmentation and o

local-aiarxiv-cs-ai
10 Aug 2026
← Previous
1…1718192021…438
Next →