AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
11 Jun 2026

Sovereign Assurance Boundary: Certificate-Bound Admission for Agentic Infrastructure

SafetyDGX agent

arXiv:2606.11632v1 Announce Type: cross Abstract: Agentic infrastructure introduces a critical control-plane authorization problem: non-deterministic reasoning systems can propose high-stakes mutation

Sparse probes and murky physics: a case study of interpretability challenges in a foundation model for continuum dynamics

Model ReleasesDGX agent

arXiv:2606.11657v1 Announce Type: cross Abstract: Generative AI emulators are increasingly used in scientific domains where we already have strong theory, benchmarks, and physical intuition. This rais

Sparsified Kolmogorov-Arnold Networks for Interpretable Quantum State Tomography

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.11814v1 Announce Type: cross Abstract: Machine-learning approaches to quantum state tomography can achieve high reconstruction fidelity, but the physical structure used by the trained model

SPEA2^+: Improved Density Estimation in SPEA2 with Provable Runtime Guarantees

Model ReleasesDGX agent

arXiv:2606.12382v1 Announce Type: cross Abstract: The Strength Pareto Evolutionary Algorithm 2 (SPEA2) is a popular and prominent evolutionary algorithm for solving multi-objective optimisation proble

SPEAR: A System for Post-Quantization Error-Adaptive Recovery Enabling Efficient Low-Bit LLM Serving

Model ReleasesDGX agent

arXiv:2606.11244v1 Announce Type: cross Abstract: Efficient large language model (LLM) serving is increasingly constrained by deployment cost. Quantization is a key technique for reducing serving cost

SpikeDecoder: Realizing the GPT Architecture with Spiking Neural Networks

ResearchDGX agent

arXiv:2606.12287v1 Announce Type: cross Abstract: The Transformer architecture is widely regarded as the most powerful tool for natural language processing, but due to a high number of complex operati

StatefulDiscovery: Evidence-Calibrated Claim Formation in Open-Ended Scientific Discovery

AgentsDGX agent

arXiv:2606.11851v1 Announce Type: new Abstract: Open-ended scientific discovery asks agents to move beyond executing analyses for predefined questions. Across multiple rounds of exploration, a discove

Steering Where to Listen: Instruction-Based Activation Steering Redirects Temporal Attention in Large Audio-Language Models

ResearchDGX agent

arXiv:2606.11400v1 Announce Type: cross Abstract: Large Audio-Language Models (LALMs) excel at audio understanding but expose little about where in an audio signal they attend. We introduce instructio

Substrate Asymmetry in User-Side Memory: A Diagnostic Framework

Model ReleasesDGX agent

arXiv:2606.11712v1 Announce Type: cross Abstract: User-side memory in LLMs is typically scored as a single 'personalization' capability: given a user's history, is the output more user-aware? We show

Sustainability assessment using multimodal AI agents

AgentsDGX agent

arXiv:2507.17012v2 Announce Type: replace Abstract: Reducing the rapidly growing environmental impact of the computing industry requires assessing the emissions of electronics at scale. However, a tra

SVoT: State-aware Visualization-of-Thought for Spatial Reasoning via Reinforcement Learning

SafetyDGX agent

arXiv:2606.11770v1 Announce Type: new Abstract: Spatial reasoning remains a challenge for Multimodal Large Language Models (MLLMs), as it requires reliable multi-hop inference over both intermediate s

System Report for CCL25-Eval Task 5: New Dataset and LoRA-Fine-Tuned Qwen2.5

Model ReleasesDGX agent

arXiv:2606.12392v1 Announce Type: cross Abstract: Recently, large language models (LLMs) have achieved promising progress in the fields of classical Chinese translation and the generation of classical

T2MM: An LLM Supported Architecture For Inquiry-Based Modeling

ApplicationsDGX agent

arXiv:2606.11210v1 Announce Type: cross Abstract: Model Construction is a foundational practice in science learning that relies on visualization and interactivity. Large Language Models, increasingly

T2S: A Rehearsal-Based Approach for Extraction-Resistant Model Watermarking

ResearchDGX agent

arXiv:2606.11698v1 Announce Type: cross Abstract: Model watermarking safeguards AI model intellectual property by embedding distinctive knowledge that induces unique behavioral signatures. The primary

Tabular Foundation Models for Clinical Survival Analysis via Survival-Aware Adaptation

ResearchDGX agent

arXiv:2606.12006v1 Announce Type: cross Abstract: Predicting time-to-event outcomes such as mortality is a fundamental task in clinical decision-making, commonly addressed through survival analysis. W

TAHOE: Text-to-SQL with Automated Hint Optimization from Experience

Model ReleasesDGX agent

arXiv:2606.12387v1 Announce Type: cross Abstract: Large Language Models (LLMs) have democratized database access through Text-to-SQL, but moving from prototypes to production remains difficult. Real d

TAROT: Task-Adaptive Refinement of LLM-prior Graphs for Few-shot Tabular Learning

ApplicationsDGX agent

arXiv:2606.11640v1 Announce Type: cross Abstract: Few-shot tabular learning provides a cost-effective approach for real-world applications where annotation is costly and collecting sufficient samples

Task-Aligned Stability Analysis of Vision-Language Models for Autonomous Driving Hazard Detection

Model ReleasesDGX agent

arXiv:2606.11889v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly used for scene understanding in autonomous driving, but robustness analysis often relies on task-agnost

Task-Aware Structured Memory for Dynamic Multi-modal In-Context Learning

SafetyDGX agent

arXiv:2606.11853v1 Announce Type: cross Abstract: Multi-modal large language models (MLLMs) depend on in-context learning (ICL) for rapid task adaptation, but their scalability is severely limited by

TextHOI-3D: Text-to-3D Hand-Object Interaction via Discrete Multi-View Generation and Joint Mesh Optimization

ResearchDGX agent

arXiv:2606.11805v1 Announce Type: cross Abstract: Text-conditioned 3D generation has progressed rapidly for images and isolated objects, but producing a hand-object mesh remains challenging: the outpu

'That's AI Slop, You Bot!' Studying Accusations, Evidence, and Credibility in Online Discourse Towards LLM-Generated Comments

ApplicationsDGX agent

arXiv:2606.12073v1 Announce Type: cross Abstract: Generative AI has made fluent prose cheap to produce, breaking the old promise to readers that good writing meant real thinking. How have readers resp

The Algorithm Is Not the Behavior: Learned Priors Override Look-Ahead in a Chess-Playing Neural Network

SafetyDGX agent

arXiv:2508.21380v3 Announce Type: replace-cross Abstract: Recent mechanistic work has uncovered learned algorithms within neural networks, from modular arithmetic to search and planning in game-playin

The Art of Interrogation: Consistency Amplifies Factuality in Spatial Reasoning

SafetyDGX agent

arXiv:2606.11918v1 Announce Type: new Abstract: Current Large Reasoning Models (LRMs) exhibit remarkable general capabilities but significantly underperform in spatial reasoning tasks. Existing approa

The Dynamics of Human and AI-Generated Language: How Semantics Fluctuates across Different Timescales

ResearchDGX agent

arXiv:2606.11371v1 Announce Type: cross Abstract: Spoken language, whether produced by humans or large language models (LLM), unfolds over time with varying semantic content. However, we still lack si

The Environmental Cost of LLMs in AIED: Reporting and Practices

ApplicationsDGX agent

arXiv:2606.11215v1 Announce Type: cross Abstract: Large Language Model (LLM) usage in recent years has become increasingly widespread in the Artificial Intelligence in Education (AIED) community. Whil

The Impossibility of Eliciting Latent Knowledge

AgentsDGX agent

arXiv:2606.12268v1 Announce Type: new Abstract: Advanced AI systems have extensive knowledge of their environments; in fact, their knowledge may (far) exceed that of their developers or users. Consequ

The Latent Color Subspace: Emergent Order in High-Dimensional Chaos

ResearchDGX agent

arXiv:2603.12261v2 Announce Type: replace-cross Abstract: Text-to-image generation models have advanced rapidly, yet achieving fine-grained control over generated images remains difficult, largely due

The Power of Test-Time Training for Approximate Sampling

ResearchDGX agent

arXiv:2606.11437v1 Announce Type: cross Abstract: Efficiently sampling from a complex probability distribution is a fundamental problem which has become increasingly pertinent in recent years with the

The Standard Interpretable Model: A general theory of interpretable machine learning to deductively design interpretable methods using Lagrangian mechanics

Model ReleasesDGX agent

arXiv:2606.12289v1 Announce Type: cross Abstract: As Artificial Intelligence models grow in complexity, interpretability has become an indispensable tool for understanding, debugging, and controlling

The Structural Attention Tax: How Retrieval Format Hijacks In-Context Learning Independent of Content

Model ReleasesDGX agent

arXiv:2606.11198v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) systems inject external knowledge to improve LLM outputs, yet the format of injected content -- distinct from its

The Unreasonable Effectiveness of Discrete-Time Gaussian Process Mixtures for Robot Policy Learning

SafetyDGX agent

arXiv:2505.03296v2 Announce Type: replace-cross Abstract: We present Mixture of Discrete-time Gaussian Processes (MiDiGap), a novel approach for flexible policy representation and imitation learning i

TileFuse: A Fused Mixed-Precision Kernel Library for Efficient Quantized LLM Inference on AMD NPUs

Local AiDGX agent

arXiv:2606.11357v1 Announce Type: cross Abstract: With the growing demand for on-device LLM inference, edge SoCs increasingly integrate NPUs to improve performance and energy efficiency under tight po

Time-Series Foundation Model Embeddings for Remaining Useful Life Estimation

Model ReleasesDGX agent

arXiv:2606.11990v1 Announce Type: cross Abstract: Remaining Useful Life (RUL) prediction is essential for industrial predictive maintenance, yet many learning-based approaches rely on extensive featur

To Intervene or Not: Guiding Inference-time Alignment with Probabilistic Model Blending

SafetyDGX agent

arXiv:2606.11201v1 Announce Type: cross Abstract: The wide deployment of LLMs has made model alignment necessary to make newly trained models safely and effectively respond to user instructions. Among

TouchThinker: Scaling Tactile Commonsense Reasoning to the Open World with Large-scale Data and Action-aware Representation

Model ReleasesDGX agent

arXiv:2606.11637v1 Announce Type: new Abstract: Touch is a key modality for embodied agents to understand the physical world. Although recent work has incorporated tactile signals into language system

Toward Generalist Autonomous Research via Hypothesis-Tree Refinement

Model ReleasesDGX agent

arXiv:2606.11926v1 Announce Type: cross Abstract: Scientific progress depends on a repeated loop of exploration, experimentation, and abstraction. Researchers test candidate directions, interpret the

Toward Preference-aligned Large Language Models via Residual-based Model Steering

SafetyDGX agent

arXiv:2509.23982v2 Announce Type: replace-cross Abstract: Preference alignment is a critical step in making Large Language Models (LLMs) useful and aligned with (human) preferences. Existing approache

Toward Trustworthy AI: Multi-Target Adversarial Attacks and Robust Defenses for Continuous Data Summarization

Model ReleasesDGX agent

arXiv:2606.11804v1 Announce Type: new Abstract: Trustworthy AI requires reliable data-processing pipelines, not only robust downstream predictive models. As an upstream component, data summarization d

Towards a Bridge Layer Between Bibliographic and Formalized Mathematical Knowledge

SafetyDGX agent

arXiv:2606.11430v1 Announce Type: cross Abstract: Mathematical knowledge is split between bibliographic databases (e.g., MathSciNet, zbMATH Open) and formal proof libraries (e.g., Lean mathlib), preve

Towards Data-free and Training-free Compression for Speech Foundation Models Using Parameter Clustering

Model ReleasesDGX agent

arXiv:2606.11836v1 Announce Type: cross Abstract: This paper presents a novel data-free and training-free compression approach for speech foundation models using channelwise clustering via k-means. Mo

Towards Deep Learning Surrogate for the Forward Problem in Electrocardiology: A Scalable Alternative to Physics-Based Models

ResearchDGX agent

arXiv:2512.13765v2 Announce Type: replace-cross Abstract: The forward problem in electrocardiology, computing body surface potentials from cardiac electrical activity, is traditionally solved using ph

Towards Fully Automated Exam Grading: Fairness-Aware Recognition of Handwritten Answers with Foundation Models

Model ReleasesDGX agent

arXiv:2606.11477v1 Announce Type: cross Abstract: Correcting handwritten exams by hand is time-consuming and error-prone, particularly for large cohorts, while fully digital exams tend to force a dida

Towards Responsibly Non-Compliant Machines

AgentsDGX agent

arXiv:2606.12147v1 Announce Type: new Abstract: We consider the problem of engineering autonomous intelligent agents that are capable to responsibly not comply with user requests. We argue that machin

TreeSeeker: Tree-Structured Trial, Error, and Return in Deep Search

AgentsDGX agent

arXiv:2606.11662v1 Announce Type: new Abstract: Deep search requires agents to answer complex questions through multi-step web search, browsing, evidence comparison, and synthesis. A central challenge

Unifying Learning Dynamics and Generalization in Transformers Scaling Law

ApplicationsDGX agent

arXiv:2512.22088v3 Announce Type: replace-cross Abstract: The scaling law, a cornerstone of Large Language Model (LLM) development, predicts improvements in model performance with increasing computati

Unstable Features, Reproducible Subspaces: Understanding Seed Dependence in Sparse Autoencoders

ResearchDGX agent

arXiv:2606.12138v1 Announce Type: cross Abstract: Sparse autoencoders (SAEs) are widely used to interpret neural network representations, but their utility depends on whether the learned features are

Using Explainability as a Training-Time Reliability Signal for Efficient ECG Classification

Model ReleasesDGX agent

arXiv:2606.12252v1 Announce Type: cross Abstract: Training deep neural networks for clinical time-series analysis is computationally demanding, yet many healthcare settings lack the resources required

VIA-SD: Verification via Intra-Model Routing for Speculative Decoding

ResearchDGX agent

arXiv:2606.12243v1 Announce Type: cross Abstract: Speculative decoding (SD) addresses the high inference costs of LLMs by having lightweight drafters generate candidates for large verifiers to validat

What Limits Does Quantization Place on Dense Top-k Retrieval? A Theoretical Study

ResearchDGX agent

arXiv:2606.11780v1 Announce Type: cross Abstract: We establish conditions for embedding a corpus of N documents as d-dimensional vectors such that every k-subset S subseteq [N] is realizable as a resu

When Context Returns: Toward Robust Internalization in On-Policy Distillation

SafetyDGX agent

arXiv:2606.11627v1 Announce Type: cross Abstract: Recent work has shown that on-policy distillation can internalize privileged context, such as system prompts or task hints, into a student model so th

When Do Data-Driven Systems Exhibit the Capability to Infer?

ResearchDGX agent

arXiv:2606.11769v1 Announce Type: new Abstract: The European AI Act is the first comprehensive regulation of artificial intelligence (AI), setting out extensive obligations, particularly for so-called

When Generic Prompt Improvements Hurt: Evaluation-Driven Iteration for LLM Applications

Model ReleasesDGX agent

arXiv:2601.22025v2 Announce Type: replace-cross Abstract: Evaluating Large Language Model (LLM) applications differs from conventional software testing because outputs are probabilistic, semantically

When Poison Fails After Retrieval: Revisiting Corpus Poisoning under Chunking and Reranking Pipelines

ResearchDGX agent

arXiv:2606.11265v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) systems are vulnerable to corpus poisoning attacks that manipulate downstream model outputs through malicious kno

When Probing Accuracy Saturates, Fragility Resolves: A Complementary Metric for LLM Pre-Training Analysis

ResearchDGX agent

arXiv:2606.11375v1 Announce Type: cross Abstract: Standard linear probing declares a property 'encoded' when a classifier on hidden states achieves high accuracy. The protocol works well on a snapshot

When Researchers Say Mental Model/Theory of Mind of AI, What Are They Really Talking About?

SafetyDGX agent

arXiv:2510.02660v2 Announce Type: replace-cross Abstract: When researchers claim AI systems possess ToM or mental models, they are fundamentally discussing behavioral predictions and bias corrections

When Roleplaying, Do Models Believe What They Say?

Model ReleasesDGX agent

arXiv:2606.11502v1 Announce Type: cross Abstract: Language models can state that 'the Earth orbits the Sun' and, when role-playing Aristotle, assert the opposite. Recent work argues that persona adopt

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning

Model ReleasesDGX agent

arXiv:2606.11816v1 Announce Type: cross Abstract: Forecasting real-world events requires language-model agents to reason under uncertainty from incomplete, time-bounded information. Yet evaluating whe

10 Jun 2026

3SPO: State-Score-Supervised Policy Optimization for LLM Agents

SafetyDGX agent

arXiv:2606.09961v1 Announce Type: cross Abstract: Training large language models (LLMs) as autonomous agents via reinforcement learning (RL) has enabled frontier models to achieve superhuman performan

A Bayesian Network Approach for Enhancing Security-Focused Decision Support Systems

TutorialsDGX agent

arXiv:2606.10782v1 Announce Type: cross Abstract: The adoption and integration of heterogeneous stacks in most of today's open-source based networks brings clear benefits like interoperability and ava

A complementary study on PlanGPT: Evaluation with defined Performance Metrics and comparison with a planner

Model ReleasesDGX agent

arXiv:2606.10489v1 Announce Type: new Abstract: Automated Planning is a subfield of Artificial Intelligence (AI) where the main objective is generating a sequence of actions, known as a plan, that hel

← Previous
1…133134135136137…358
Next →