AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Tutorials

Selective Test-Time Compute Scaling for Click-Through Rate Prediction via Uncertainty-Triggered Feature Path Exploration

DGX agent

arXiv:2605.24989v1 Announce Type: cross Abstract: Scaling test-time compute has proven highly effective for language models, yet this opportunity remains largely unexplored for industrial Click-Throug

tutorialsarxiv-cs-ai
26 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

SemanticZip: A Pilot Framework for Lossy Text Compression with LLMs as Semantic Decompressors

DGX agent

arXiv:2605.24541v1 Announce Type: cross Abstract: Text compression for large language model (LLM) systems is usually framed as token deletion, retrieval, summarization, or exact reconstruction. We stu

model-releasesarxiv-cs-ai
26 May 2026
Research

Sensing Intelligence as a Trainable Metamaterial Property

DGX agent

arXiv:2605.23967v1 Announce Type: cross Abstract: In biological systems, sensing is not performed by the brain alone: the body deforms, vibrates, and filters external stimuli before they are transduce

researcharxiv-cs-ai
26 May 2026
Research

SentGraph: Hierarchical Sentence Graph for Multi-hop Retrieval-Augmented Question Answering

DGX agent

arXiv:2601.03014v3 Announce Type: replace-cross Abstract: Traditional Retrieval-Augmented Generation (RAG) effectively supports single-hop question answering with large language models but faces signi

researcharxiv-cs-ai
26 May 2026
Applications

SEP-Attack: A Simple and Effective Paradigm for Transfer-Based Textual Adversarial Attack

DGX agent

arXiv:2605.24958v1 Announce Type: cross Abstract: Despite the strong performance of deep neural networks in modern Web and language applications, they remain vulnerable to adversarial attacks, especia

applicationsarxiv-cs-ai
26 May 2026
Applications

SeqRoute: Global Budget-Aware Sequential LLM Routing via Offline Reinforcement Learning

DGX agent

arXiv:2605.25424v1 Announce Type: cross Abstract: Existing LLM routing frameworks treat queries as independent events, neglecting the sequential nature of real-world user sessions constrained by globa

applicationsarxiv-cs-ai
26 May 2026
Safety

Side-by-side Comparison Amplifies Dialect Bias in Language Models

DGX agent

arXiv:2605.24384v1 Announce Type: cross Abstract: Language models (LMs) can exhibit systematic biases against speakers based on variations in their dialects, even in the absence of a dialect label, a

safetyarxiv-cs-ai
26 May 2026
Local Ai

Signs Beat Floats: Low-Rank Double-Binary Adaptation for On-Device Fine-Tuning

DGX agent

arXiv:2605.24058v1 Announce Type: cross Abstract: On-device adaptation of large language models commonly keeps a quantized base model frozen while training and deploying a small, task-specific LoRA ad

local-aiarxiv-cs-ai
26 May 2026
Applications

Simulating Human Memory with Language Models

DGX agent

arXiv:2605.25680v1 Announce Type: cross Abstract: Language models are increasingly being deployed as user simulators, but their memory is far more reliable than that of real users. To measure this gap

applicationsarxiv-cs-ai
26 May 2026
Model Releases

'Si'multaneous 'S'patial-'T'emporal Message Passing for Dynamic Graph Representation Learning

DGX agent

arXiv:2605.25548v1 Announce Type: cross Abstract: Dynamic graph neural networks (DGNNs) that operate on snapshot sequences typically fall into one of two categories. Temporal-first approaches build pe

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

SimuWoB: Simulating Real-World Mobile Apps for Fast and Faithful GUI Agent Benchmarking

DGX agent

arXiv:2605.25160v1 Announce Type: new Abstract: Mobile GUI agents powered by large language models have progressed rapidly, creating urgent needs for realistic and comprehensive evaluation. Existing b

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

SkillEvolBench: Benchmarking the Evolution from Episodic Experience to Procedural Skills

DGX agent

arXiv:2605.24117v1 Announce Type: new Abstract: Large language model (LLM) agents accumulate rich episodic trajectories while solving real-world tasks, but it remains unclear whether such experience c

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Small Models, Strong Priors: Architectural Inductive Bias for Parameter-Efficient Neural PDE Solvers

DGX agent

arXiv:2605.25949v1 Announce Type: cross Abstract: Neural PDE solvers have followed the scaling trajectory of vision and language, with recent foundation models reaching billions of parameters. We argu

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Smart Timing for Mining: A Deep Learning Framework for Bitcoin Hardware ROI Prediction

DGX agent

arXiv:2512.05402v2 Announce Type: replace-cross Abstract: Bitcoin mining hardware acquisition requires strategic timing due to volatile markets, rapid technological obsolescence, and protocol-driven r

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

SMDD-Bench: Can LLMs Solve Real-World Small Molecule Drug Design Tasks?

DGX agent

arXiv:2605.21740v2 Announce Type: replace Abstract: LLM agents have incredible potential for scientific discovery applications. However, the performance of LLM agents on real-world, small molecule dru

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

SODE: Analyzing Social Dynamics in LLM Agents

DGX agent

arXiv:2605.23949v1 Announce Type: cross Abstract: As Large Language Models (LLMs) evolve into interactive agents, understanding their behavioral alignment within human social dynamics becomes essentia

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models

DGX agent

arXiv:2506.18543v2 Announce Type: replace-cross Abstract: The rapid proliferation of Large Language Models (LLMs) has heightened concerns regarding their exposure to jailbreak attacks, which craft adv

model-releasesarxiv-cs-ai
26 May 2026
Agents

SoK: DARPA's AI Cyber Challenge (AIxCC): Competition Design, Architectures, and Lessons Learned

DGX agent

arXiv:2602.07666v3 Announce Type: replace-cross Abstract: DARPA's AI Cyber Challenge (AIxCC, 2023--2025) is the largest competition to date for building fully autonomous cyber reasoning systems (CRSs)

agentsarxiv-cs-ai
26 May 2026
Research

Solving Combinatorial Counting Problems with Weighted First-Order Model Counting

DGX agent

arXiv:2605.24845v1 Announce Type: new Abstract: Combinatorial counting problems pervade artificial intelligence, statistics, and discrete mathematics. Whether the task is enumerating subsets, multiset

researcharxiv-cs-ai
26 May 2026
Model Releases

SomaliBench Eval: Measuring English-to-Somali Refusal Gaps in Open-Weight Language Models

DGX agent

arXiv:2605.25420v1 Announce Type: cross Abstract: Large language model safety evaluation remains heavily English-centered, leaving low-resource languages under-measured even when models are deployed g

model-releasesarxiv-cs-ai
26 May 2026
Research

SPA-Cache: Singular Proxies for Adaptive Caching in Diffusion Language Models

DGX agent

arXiv:2602.02544v2 Announce Type: replace-cross Abstract: While Diffusion Language Models (DLMs) offer a flexible, arbitrary-order alternative to the autoregressive paradigm, their non-causal nature p

researcharxiv-cs-ai
26 May 2026
Applications

SPACE: Unifying Symmetric and Asymmetric Routing Problems for Generalist Neural Solver

DGX agent

arXiv:2605.24484v1 Announce Type: new Abstract: Generalist neural routing solvers have shown great potential in solving diverse vehicle routing problems (VRPs) with a unified model. However, existing

applicationsarxiv-cs-ai
26 May 2026
Research

Spacetime Formation under Requirements: Contextual Realization and Form-Dependent Probability

DGX agent

arXiv:2605.23943v1 Announce Type: new Abstract: Quantum cognition often explains order effects, contextuality, and violations of the law of total probability by replacing classical probability with qu

researcharxiv-cs-ai
26 May 2026
Agents

SPARK: Search Personalization via Agent-Driven Retrieval and Knowledge-sharing

DGX agent

arXiv:2512.24008v3 Announce Type: replace Abstract: Personalized search demands the ability to model users' evolving, multi-dimensional information needs; a challenge for systems constrained by static

agentsarxiv-cs-ai
26 May 2026
Safety

SpecAlign: A Semantic Alignment Framework for SystemVerilog Assertion Generation

DGX agent

arXiv:2605.25181v1 Announce Type: new Abstract: Existing Large Language Model (LLM) approaches to SystemVerilog Assertion (SVA) generation primarily focus on syntactic validity and formal verification

safetyarxiv-cs-ai
26 May 2026
Research

Specification-Based Code-Text-Code Reengineering for LLM-Mediated Software Evolution

DGX agent

arXiv:2605.25232v1 Announce Type: cross Abstract: Direct Code2Code transformation remains challenging to control because it can preserve surface-level syntax while introducing semantic drift, hidden b

researcharxiv-cs-ai
26 May 2026
Local Ai

SpecPrune-VLA: Accelerating Vision-Language-Action Models via Action-Aware Self-Speculative Pruning

DGX agent

arXiv:2509.05614v3 Announce Type: replace-cross Abstract: Pruning is a typical acceleration technique for compute-bound models by removing computation on unimportant values. Recently, it has been appl

local-aiarxiv-cs-ai
26 May 2026
Model Releases

Spectral Probe-Circuits: A Three-Step Recipe for Identifying Attention-Head Circuits in Pretrained Transformers

DGX agent

arXiv:2605.24059v1 Announce Type: cross Abstract: We present a three-step recipe for identifying attention-head circuits in pretrained transformers. A per-head spectral signal -- the time-integrated p

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Spectral Retrieval: Multi-Scale Sinc Convolution over Token Embeddings for Localized Retrieval in LLM Multi-Agent Systems

DGX agent

arXiv:2605.24764v1 Announce Type: cross Abstract: [Abridged] - Spectral Retrieval is a plug-in re-ranking stage that interpolates between per-token MaxSim and mean-pool retrieval through a multi-scale

model-releasesarxiv-cs-ai
26 May 2026
Research

Squeezing Capacity from Multimodal Large Language Models for Subject-driven Generation

DGX agent

arXiv:2605.26111v1 Announce Type: cross Abstract: Subject-driven image generation aims to synthesize new images that preserve the identity of the given subject while following textual instructions. Ex

researcharxiv-cs-ai
26 May 2026
Safety

StakeBench: Evaluating Language Understanding Grounded in Market Commitment

DGX agent

arXiv:2605.26074v1 Announce Type: cross Abstract: Existing financial NLP benchmarks often rely on labels supplied by outside observers, measuring how language is perceived rather than what speakers ha

safetyarxiv-cs-ai
26 May 2026
Research

Step-TP: A Grounded, Step-Level Dataset with Chain-of-Thought Reasoning for LLM-Guided Tensor Program Optimization

DGX agent

arXiv:2605.25954v1 Announce Type: cross Abstract: Despite the strong reasoning capabilities of large language models (LLMs), optimizing the execution efficiency of tensor programs remains challenging

researcharxiv-cs-ai
26 May 2026
Safety

Stop Comparing LLM Agents Without Disclosing the Harness

DGX agent

arXiv:2605.23950v1 Announce Type: new Abstract: This position paper argues that, for long-horizon tasks evaluated across models with comparable frontier capability, the agent execution harness, namely

safetyarxiv-cs-ai
26 May 2026
Safety

Strat-Reasoner: Reinforcing Strategic Reasoning of LLMs in Multi-Agent Games

DGX agent

arXiv:2605.04906v2 Announce Type: replace Abstract: While Large Language Models (LLMs) excel in certain reasoning tasks, they struggle in multi-agent games where the final outcome depends on the joint

safetyarxiv-cs-ai
26 May 2026
Model Releases

STREAM: A Data-Centric Framework for Mining High-Value Task-Oriented Dialogues from Streaming Media

DGX agent

arXiv:2605.25162v1 Announce Type: cross Abstract: Large language models for vertical domains are bottlenecked by the scarcity of complex, domain-specific task-oriented dialogues. Existing data acquisi

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

StructBreak: Structural Cognitive Overload-Induced Safety Failures in MLLMs

DGX agent

arXiv:2605.25534v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) excel at structural reasoning yet suffer from a sharp logical brittleness in structural consistency. We term th

model-releasesarxiv-cs-ai
26 May 2026
Research

Subspace Aggregation Query and Index Generation for Multidimensional Resource Space Model

DGX agent

arXiv:2505.02129v3 Announce Type: replace-cross Abstract: Organizing large-scale resources in a multidimensional semantic space is an approach to efficiently managing and querying resources from diffe

researcharxiv-cs-ai
26 May 2026
Safety

Subspace-Guided Semantic and Topological Invariant Registration for Annotation-Free Ultrasound Plane Quality Control

DGX agent

arXiv:2605.25396v1 Announce Type: cross Abstract: Reliable quality control (QC) of ultrasound images is essential for both real-time acquisition guidance and retrospective clinical audit, yet existing

safetyarxiv-cs-ai
26 May 2026
Safety

Summoning the Oracle to Slay It: Mitigating Look-Ahead Bias in Financial Backtesting with Large Language Models

DGX agent

arXiv:2605.24564v1 Announce Type: new Abstract: Backtesting large language models (LLMs) on historical financial data is unreliable because pre-training cuts off after the events happened. An LLM trai

safetyarxiv-cs-ai
26 May 2026
Research

TaBIIC2: Interactive Building of Ontological Taxonomies using Weighted Self-Organizing Maps

DGX agent

arXiv:2605.24899v1 Announce Type: new Abstract: Ontologies represent the conceptual knowledge of a domain. At the core of an ontology is the taxonomy of concepts and subconcepts that represent specifi

researcharxiv-cs-ai
26 May 2026
Safety

Task-Aligned Self-Supervised Learning for Medical Image Analysis: A Systematic Review and Practical Design Guidelines

DGX agent

arXiv:2605.23995v1 Announce Type: cross Abstract: Self-supervised learning (SSL) has emerged as a promising paradigm for addressing the annotation bottleneck in medical imaging by learning representat

safetyarxiv-cs-ai
26 May 2026
Model Releases

Teaching large language models to reason like expert diagnosticians

DGX agent

arXiv:2509.12194v2 Announce Type: replace Abstract: Differential diagnosis is an iterative process that integrates patient information with broader medical knowledge. Clinical case series such as the

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Teaching Through Analogies: A Modular Pipeline for Educational Analogy Generation

DGX agent

arXiv:2605.24211v1 Announce Type: cross Abstract: Analogies help learners understand unfamiliar concepts by relating them to known concepts. Despite recent advances, large language models (LLMs) conti

model-releasesarxiv-cs-ai
26 May 2026
Applications

Temporal Concept Drift in Legal Judgment Prediction: Neural Baselines Across Three Epochs of Ukrainian Court Decisions

DGX agent

arXiv:2605.24452v1 Announce Type: cross Abstract: Legal NLP benchmarks evaluate models on randomly split data, implicitly assuming that legal language is stationary. We test this assumption by fine-tu

applicationsarxiv-cs-ai
26 May 2026
Agents

Test-Time Deep Thinking to Explore Implicit Rules

DGX agent

arXiv:2605.24828v1 Announce Type: new Abstract: With the continuous advancement of Large Language Models (LLMs), intelligent agents are becoming increasingly vital. However, these agents often fail in

agentsarxiv-cs-ai
26 May 2026
Model Releases

Test-Time Self-Adaptive Conditioning for Stable Audio-Driven Talking-Head Generation

DGX agent

arXiv:2605.25488v1 Announce Type: cross Abstract: Audio-driven talking-head generation has achieved remarkable progress with recent models such as AniTalker, FLOAT, and Sonic. Despite their success, m

model-releasesarxiv-cs-ai
26 May 2026
Research

TGFormer: Towards Temporal Graph Transformer with Auto-Correlation Mechanism

DGX agent

arXiv:2605.24971v1 Announce Type: cross Abstract: The growing interest in Temporal Graph Neural Networks (TGNNs) stems from their ability to model complex dynamics and deliver superior performance. Ho

researcharxiv-cs-ai
26 May 2026
Safety

The Concept Allocation Zone: Tracking How Concepts Form Across Transformer Depth

DGX agent

arXiv:2605.24856v1 Announce Type: cross Abstract: Concept formation in transformer language models is depth-extended, not a single-layer event: concepts emerge gradually across a contiguous region of

safetyarxiv-cs-ai
26 May 2026
← Previous
1…273274275276277…448
Next →