AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
12 May 2026

SLAM: Structural Linguistic Activation Marking for Language Models

Model ReleasesDGX agent

arXiv:2605.05443v2 Announce Type: replace-cross Abstract: LLM watermarks must be detectable without compromising text quality, yet most existing schemes bias the next-token distribution and pay for de

SLASH the Sink: Sharpening Structural Attention Inside LLMs

SafetyDGX agent

arXiv:2605.10503v1 Announce Type: new Abstract: Large Language Models (LLMs) show remarkable semantic understanding but often struggle with structural understanding when processing graph topologies in

SLayerGen: a Crystal Generative Model for all Space and Layer Groups

ResearchDGX agent

arXiv:2605.08262v1 Announce Type: cross Abstract: Crystal generative models have shown rapid progress for accelerating the discovery of bulk, periodic materials. However, many material systems such as


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

SLIM: Sparse Latent Steering for Interpretable and Property-Directed LLM-Based Molecular Editing

Model ReleasesDGX agent

arXiv:2605.10831v1 Announce Type: cross Abstract: Large language models possess strong chemical reasoning capabilities, making them effective molecular editors. However, property-relevant information

SlimQwen: Exploring the Pruning and Distillation in Large MoE Model Pre-training

ResearchDGX agent

arXiv:2605.08738v1 Announce Type: cross Abstract: Structured pruning and knowledge distillation (KD) are typical techniques for compressing large language models, but it remains unclear how they shoul

Slipstream: Trajectory-Grounded Compaction Validation for Long-Horizon Agents

AgentsDGX agent

arXiv:2605.08580v1 Announce Type: cross Abstract: To cope with the large contexts that long-horizon LLM agents produce, modern frameworks increasingly rely on compaction -- invoking an LLM to rewrite

SmartEval: A Benchmark for Evaluating LLM-Generated Smart Contracts from Natural Language Specifications

Model ReleasesDGX agent

arXiv:2605.09610v1 Announce Type: cross Abstract: We introduce SmartEval, a benchmark for systematically evaluating the quality of Solidity smart contracts generated by large language models (LLMs) fr

SnareNet: Flexible Repair Layers for Neural Networks with Hard Constraints

SafetyDGX agent

arXiv:2602.09317v2 Announce Type: replace-cross Abstract: Neural networks are increasingly used as fast surrogate models across various domains, but unconstrained predictions can violate physical, ope

SoK: A Systematic Bidirectional Literature Review of AI & DLT Convergence

AgentsDGX agent

arXiv:2605.10515v1 Announce Type: cross Abstract: The integration of Artificial Intelligence (AI) with Distributed Ledger Technology (DLT) has become a growing research area, yet contributions tend to

Sparsity Moves Computation: How FFN Architecture Reshapes Attention in Small Transformers

Model ReleasesDGX agent

arXiv:2605.09403v1 Announce Type: cross Abstract: Architectural choices inside the Transformer feedforward network (FFN) block do not merely affect the block itself; they reshape the computations lear

Spatial Priming Outperforms Semantic Prompting: A Grid-Based Approach to Improving LLM Accuracy on Chart Data Extraction

ResearchDGX agent

arXiv:2605.08220v1 Announce Type: new Abstract: The automated extraction of data from scientific charts is a critical task for large-scale literature analysis. While multimodal Large Language Models (

Spectral Characterization and Mitigation of Sequential Knowledge Editing Collapse

Model ReleasesDGX agent

arXiv:2601.11042v2 Announce Type: replace-cross Abstract: Sequential knowledge editing in large language models often causes catastrophic collapse of the model's general abilities, especially for para

Spectral Transformer Neural Processes

SafetyDGX agent

arXiv:2605.09498v1 Announce Type: cross Abstract: Time series, spatial data, and images are natural applications of Neural Processes. However, when such data exhibit strong periodicity and quasi-perio

SPECTRE: Hybrid Ordinary-Parallel Speculative Serving for Resource-Efficient LLM Inference

ApplicationsDGX agent

arXiv:2605.08151v1 Announce Type: cross Abstract: LLM serving platforms are increasingly deployed as multi-model cloud systems, where user demand is often long-tailed: a few popular large models recei

Speech-based Psychological Crisis Assessment using LLMs

ResearchDGX agent

arXiv:2605.10027v1 Announce Type: cross Abstract: Psychological support hotlines provide critical support for individuals experiencing mental health emergencies, yet current assessments largely rely o

Statistical Model Checking of the Keynes+Schumpeter Model: A Transient Sensitivity Analysis of a Macroeconomic ABM

Model ReleasesDGX agent

arXiv:2605.10447v1 Announce Type: cross Abstract: Agent-based models (ABMs) are increasingly used in macroeconomics, but their analysis still often relies on ad hoc Monte Carlo campaigns with heteroge

Step Rejection Fine-Tuning: A Practical Distillation Recipe

Model ReleasesDGX agent

arXiv:2605.10674v1 Announce Type: cross Abstract: Rejection Fine-Tuning (RFT) is a standard method for training LLM agents, where unsuccessful trajectories are discarded from the training set. In the

StereoTales: A Multilingual Framework for Open-Ended Stereotype Discovery in LLMs

Local AiDGX agent

arXiv:2605.10442v1 Announce Type: cross Abstract: Multilingual studies of social bias in open-ended LLM generation remain limited: most existing benchmarks are English-centric, template-based, or rest

Strategic commitments shape collective cybersecurity under AI inequality

Model ReleasesDGX agent

arXiv:2605.09415v1 Announce Type: new Abstract: The growing integration of AI into cybersecurity is reshaping the balance between attackers and defenders. When access to advanced AI-enabled defence to

Strategic Exploitation in LLM Agent Markets: A Simulation Framework for E-Commerce Trust

Model ReleasesDGX agent

arXiv:2605.10059v1 Announce Type: new Abstract: Agent-based modeling (ABM) has long been used in economics to study human behavior, and large language model (LLM) agents now enable new forms of social

Structure-Centric Graph Foundation Model via Geometric Bases

ResearchDGX agent

arXiv:2605.08689v1 Announce Type: cross Abstract: Graph foundation models (GFMs) seek transferable representations across graph domains but are limited by structural heterogeneity and incompatible nod

Sub-JEPA: Subspace Gaussian Regularization for Stable End-to-End World Models

SafetyDGX agent

arXiv:2605.09241v1 Announce Type: cross Abstract: Joint-Embedding Predictive Architectures (JEPAs) provide a simpleframework for learning world models by predicting future latent representations.Howev

Sufficient conditions for a Heuristic Rating Estimation Method application

ResearchDGX agent

arXiv:2605.08991v1 Announce Type: new Abstract: A series of papers has introduced the Heuristic Rating Estimation method, which evaluates a set of alternatives based on pairwise comparisons and the we

Supervised Dimensionality Reduction Revisited: Why LDA on Frozen CNN Features Deserves a Second Look

ResearchDGX agent

arXiv:2604.03928v2 Announce Type: replace-cross Abstract: Frozen pretrained image representations are widely used for transfer learning: a backbone is kept fixed, feature vectors are extracted, and a

Supervised Mixture-of-Experts for Surgical Grasping and Retraction

SafetyDGX agent

arXiv:2601.21971v2 Announce Type: replace-cross Abstract: Imitation learning has achieved remarkable success in robotic manipulation, yet its application to surgical robotics remains challenging due t

Swarm Skills: A Portable, Self-Evolving Multi-Agent System Specification for Coordination Engineering

AgentsDGX agent

arXiv:2605.10052v1 Announce Type: cross Abstract: As artificial intelligence engineering paradigms shift from single-agent Prompt and Context Engineering toward multi-agent extbf{Coordination Engineer

SWIFT: Prompt-Adaptive Memory for Efficient Interactive Long Video Generation

Local AiDGX agent

arXiv:2605.09442v1 Announce Type: cross Abstract: Streaming long-video generation faces a central challenge in continuous semantic switching, requiring adaptive memory to preserve coherent visual evol

Switching-Geometry Analysis of Deflated Q-Value Iteration

SafetyDGX agent

arXiv:2605.10811v1 Announce Type: cross Abstract: This paper develops a joint spectral radius (JSR) framework for analyzing rank-one deflated Q-value iteration (Q-VI) in discounted Markov decision pro

SynerDiff: Synergetic Continuous Batching for Fast and Parallel Diffusion Model Inference

ResearchDGX agent

arXiv:2605.08835v1 Announce Type: new Abstract: The expansion of Artificial Intelligence-generated content service requires diffusion model serving to simultaneously achieve high throughput and low ta

TAD: Temporal-Aware Trajectory Self-Distillation for Fast and Accurate Diffusion LLM

ResearchDGX agent

arXiv:2605.09536v1 Announce Type: cross Abstract: Diffusion large language models (dLLMs) offer a promising paradigm for parallel text generation, but in practice they face an accuracy-parallelism tra

Task-Agnostic Noisy Label Detection via Standardized Loss Aggregation

ResearchDGX agent

arXiv:2605.10165v1 Announce Type: cross Abstract: Noisy labels are common in large-scale medical imaging datasets due to inter-observer variability and ambiguous cases. We propose a statistically grou

Task complexity shapes internal representations and robustness in neural networks

ResearchDGX agent

arXiv:2508.05463v2 Announce Type: replace-cross Abstract: Neural networks excel across a wide range of tasks, yet remain black boxes. In particular, how their internal representations are shaped by th

Teacher-Aware Evolution of Heuristic Programs from Learned Optimization Policies

Local AiDGX agent

arXiv:2605.10634v1 Announce Type: new Abstract: LLM-based automatic heuristic design has shown promise for generating executable heuristics for combinatorial optimization, but existing methods mainly

Teaching Molecular Dynamics to a Non-Autoregressive Ionic Transport Predictor

TutorialsDGX agent

arXiv:2605.09311v1 Announce Type: cross Abstract: Unlike most static material properties widely studied in the machine learning literature, ionic transport properties are inherently dynamic, making th

Team-Based Self-Play With Dual Adaptive Weighting for Fine-Tuning LLMs

SafetyDGX agent

arXiv:2605.09922v1 Announce Type: cross Abstract: While recent self-training approaches have reduced reliance on human-labeled data for aligning LLMs, they still face critical limitations: (i) sensiti

Temperature and Persona Shape LLM Agent Consensus With Minimal Accuracy Gains in Qualitative Coding

AgentsDGX agent

arXiv:2507.11198v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) enable new possibilities for qualitative research at scale, including annotation and qualitative coding of educat

Text-Guided Multi-Scale Frequency Representation Adaptation

Model ReleasesDGX agent

arXiv:2605.08181v1 Announce Type: cross Abstract: Parameter-efficient fine-tuning methods introduce a small number of training parameters, enabling pre-trained models to adapt rapidly to new data dist

TextBridgeGNN: Pre-training Graph Neural Network for Cross-Domain Recommendation via Text-Guided Transfer

TutorialsDGX agent

arXiv:2601.02366v2 Announce Type: replace-cross Abstract: Graph-based recommendation has achieved great success in recent years. The classical graph recommendation model utilizes ID embedding to store

TFM-Retouche: A Lightweight Input-Space Adapter for Tabular Foundation Models

Model ReleasesDGX agent

arXiv:2605.06047v2 Announce Type: replace-cross Abstract: Tabular foundation models (TFMs), such as TabPFN-2.6, TabICLv2, ConTextTab, Mitra, LimiX, and TabDPT, achieve strong zero-shot performance thr

The Accountability Paradox: How Platform API Restrictions Undermine AI Transparency Mandates

SafetyDGX agent

arXiv:2505.11577v4 Announce Type: replace-cross Abstract: Recent application programming interface (API) restrictions on major social media platforms challenge compliance with the EU Digital Services

The Agent Use of Agent Beings: Agent Cybernetics Is the Missing Science of Foundation Agents

AgentsDGX agent

arXiv:2605.10754v1 Announce Type: new Abstract: LLM-based foundation agents that perceive, reason, and act across thousands of reasoning steps are rapidly becoming the dominant paradigm for deploying

The Art of the Jailbreak: Formulating Jailbreak Attacks for LLM Security Beyond Binary Scoring

SafetyDGX agent

arXiv:2605.09225v1 Announce Type: cross Abstract: Jailbreak attacks -- adversarial prompts that bypass LLM alignment through purely linguistic manipulation -- pose a growing operational security threa

The Attacker in the Mirror: Breaking Self-Consistency in Safety via Anchored Bipolicy Self-Play

Model ReleasesDGX agent

arXiv:2605.08427v1 Announce Type: new Abstract: Self-play red team is an established approach to improving AI safety in which different instances of the same model play attacker and defender roles in

The autoPET3 Challenge: Automated Lesion Segmentation in Whole-Body PET/CT nicode{x2013} Multitracer Multicenter Generalization

Model ReleasesDGX agent

arXiv:2605.05775v2 Announce Type: replace-cross Abstract: We report the design and results of the third autoPET challenge (MICCAI 2024), which benchmarked automated lesion segmentation in whole-body P

The Bystander Effect in Multi-Agent Reasoning: Quantifying Cognitive Loafing in Collaborative Interactions

SafetyDGX agent

arXiv:2605.10698v1 Announce Type: cross Abstract: Multi-agent systems (MAS) assume that collaborating inherently improves Large Language Model (LLM) reasoning. We challenge this by demonstrating that

The Cartesian Shortcut: Re-evaluate Vision Reasoning in Polar Coordinate Space

ResearchDGX agent

arXiv:2605.09883v1 Announce Type: cross Abstract: As current Multimodal Large Language Models rapidly saturate canonical visual reasoning benchmarks, a key question emerges: do these strong scores gen

The DSA's Blind Spot: Algorithmic Audit of Advertising and Minor Profiling on TikTok

ResearchDGX agent

arXiv:2603.05653v2 Announce Type: replace-cross Abstract: Adolescents spend an increasing amount of their time in digital environments where their still-developing cognitive capacities leave them unab

The Echo Amplifies the Knowledge: Somatic Marker Analogues in Language Models via Emotion Vector Re-Injection

Model ReleasesDGX agent

arXiv:2605.08611v1 Announce Type: new Abstract: Current language model memory systems store what happened but not how it felt. This distinction -- between semantic memory (knowing about a past event)

The First Drop of Ink: Nonlinear Impact of Misleading Information in Long-Context Reasoning

AgentsDGX agent

arXiv:2605.10828v1 Announce Type: new Abstract: As large language models are increasingly deployed in retrieval-augmented generation and agentic systems that accumulate extensive context, understandin

The Generalized Turing Test: A Foundation for Comparing Intelligence

ResearchDGX agent

arXiv:2605.10851v1 Announce Type: new Abstract: We introduce the Generalized Turing Test (GTT), a formal framework for comparing the capabilities of arbitrary agents via indistinguishability. For agen

The Geometric Reasoner: Manifold-Informed Latent Foresight Search for Long-Context Reasoning

ResearchDGX agent

arXiv:2601.18832v3 Announce Type: replace-cross Abstract: Scaling test-time compute enhances long chain-of-thought (CoT) reasoning, yet existing approaches face a fundamental trade-off between computa

The Geometric Wall: Manifold Structure Predicts Layerwise Sparse Autoencoder Scaling Laws

Model ReleasesDGX agent

arXiv:2605.09887v1 Announce Type: cross Abstract: Sparse autoencoders (SAEs) operationalise the linear representation hypothesis: they reconstruct model activations as sparse linear combinations of in

The Geometry of Forgetting: Temporal Knowledge Drift as an Independent Axis in LLM Representations

Model ReleasesDGX agent

arXiv:2605.09195v1 Announce Type: new Abstract: Large language models confidently produce outdated answers, and no existing method can detect them. We show this is not an engineering failure but a str

The Gordian Knot for VLMs: Diagrammatic Knot Reasoning as a Hard Benchmark

Model ReleasesDGX agent

arXiv:2605.09900v1 Announce Type: new Abstract: A vision-language model can look at a knot diagram and report what it sees, yet fail to act on that structure. KnotBench pairs an 858,318-image corpus f

The Grounding Gap: How LLMs Anchor the Meaning of Abstract Concepts Differently from Humans

SafetyDGX agent

arXiv:2605.08837v1 Announce Type: cross Abstract: Abstract concepts - justice, theory, availability - have no single perceivable referent; in the human brain, their meaning emerges from a web of exper

The Last Word Often Wins: A Format Confound in Chain-of-Thought Corruption Studies

Model ReleasesDGX agent

arXiv:2605.10799v1 Announce Type: cross Abstract: Corruption studies, the primary tool for evaluating chain-of-thought (CoT) faithfulness, identify which chain positions are 'computationally important

The Metacognitive Probe: Five Behavioural Calibration Diagnostics for LLMs

Model ReleasesDGX agent

arXiv:2605.09844v1 Announce Type: new Abstract: The Metacognitive Probe is an exploratory five-task, 15-slot diagnostic that decomposes an LLM's confidence behaviour into five behaviourally-distinct d

The Open-Box Fallacy: Why AI Deployment Needs a Calibrated Verification Regime

ResearchDGX agent

arXiv:2605.10601v1 Announce Type: new Abstract: AI deployment in sensitive domains such as health care, credit, employment, and criminal justice is often treated as unsafe to authorize until model int

The Pokemon Theorem and other Fairness Impossibility Results

SafetyDGX agent

arXiv:2605.09221v1 Announce Type: cross Abstract: Fairness impossibility results often look like distinct scalar incompatibility statements. We show that several share one RKHS geometry: fairness crit

The Reciprocity Gradient

AgentsDGX agent

arXiv:2605.08323v1 Announce Type: cross Abstract: Communication is fundamental to sustaining reciprocity and cooperation in strategic interactions. We identify and formulate the influence attribution

← Previous
1…274275276277278…358
Next →