AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
23 Jun 2026

Can Reasoning Models Detect Changes to their Chains of Thought?

ResearchDGX agent

arXiv:2606.22085v1 Announce Type: cross Abstract: There are many reasons one may want to edit a model's chain of thought (CoT) -- e.g., to prefill it with reasoning from a stronger model or to remove

CAT-Translate: Building Compact Open-Source Models for Japanese-English Translation

ApplicationsDGX agent

arXiv:2606.21413v1 Announce Type: cross Abstract: Nowadays, large multilingual translation models demonstrate impressive translation capabilities in the machine translation benchmarks. This raises a p

CATCH: Channel-Aware multivariate Time Series Anomaly Detection via Frequency Patching

ApplicationsDGX agent

arXiv:2410.12261v5 Announce Type: replace Abstract: Anomaly detection in multivariate time series is challenging as heterogeneous subsequence anomalies may occur. Reconstruction-based methods, which f


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Causal Discovery in the Era of Agents

AgentsDGX agent

arXiv:2606.23608v1 Announce Type: cross Abstract: Recent attempts to combine large language models (LLMs) with causal discovery ask models to infer pairwise directions, propose graph structures, or in

Causal Gaussian Processes for Robust Treatment Effect Evaluation with Unobserved Confounding

SafetyDGX agent

arXiv:2606.21809v1 Announce Type: new Abstract: The presence of confounding bias poses a key challenge in policy evaluation, as the target causal effects of actions are not identifiable (i.e., underde

Causal Variational Deep Embedding: A Family of Interventional Generators for Confounded Images

ResearchDGX agent

arXiv:2606.21806v1 Announce Type: new Abstract: Deep generative models reproduce the observational distribution of their training data, inheriting any spurious associations it contains. A common sourc

Causally Fair Node Classification on Non-IID Graph Data

SafetyDGX agent

arXiv:2505.01652v2 Announce Type: replace Abstract: Fair machine learning seeks to identify and mitigate biases in predictions against unfavorable populations characterized by demographic attributes,

CELEUS: Certifiable and Efficient LLM Evaluation via E-Processes

ApplicationsDGX agent

arXiv:2606.20820v1 Announce Type: new Abstract: Can we trust evaluation scores to capture an LLM's true real-world performance? Certifiable evaluation answers this question by providing guarantee for

Central limit theorem for the averaged Adam optimizer

ResearchDGX agent

arXiv:2606.21433v1 Announce Type: cross Abstract: In this article, we analyse convergence of the averaged Adam optimizer to an attracting zero of the Adam vector field. We provide a central limit theo

Certified World Models: Predictability Across Configuration, Horizon, and Resolution

ResearchDGX agent

arXiv:2606.13092v2 Announce Type: replace Abstract: Scale buys interpolation; structure buys certifiable transfer. A world model's average error does not say whether a particular rollout can be truste

CFAgentBench: A Reproducible Environment and Benchmark for Autonomous Construction-Finance Agents

Model ReleasesDGX agent

arXiv:2606.22000v1 Announce Type: cross Abstract: We introduce CFAgentBench, a reproducible, self-hostable environment and benchmark for autonomous construction-finance agents: a CFO/controller-class

Chain-of-Goals Hierarchical Policy for Long-Horizon Offline Goal-Conditioned RL

SafetyDGX agent

arXiv:2602.03389v2 Announce Type: replace Abstract: Offline goal-conditioned reinforcement learning remains challenging for long-horizon tasks. While hierarchical approaches mitigate this issue by dec

Channel Location Constrains the Auditability of Subliminal Learning

Local AiDGX agent

arXiv:2606.22019v1 Announce Type: new Abstract: Subliminal learning lets a student inherit a teacher's hidden trait from distillation data that never names it. We ask when such transfer can be audited

Chem2Gen-Bench: Benchmarking Chemical-to-Genetic Translation in Perturbation Response Space

Model ReleasesDGX agent

arXiv:2606.21109v1 Announce Type: new Abstract: Virtual-cell and perturbation models are increasingly used to predict cellular responses for biomedical discovery, but chemical and genetic perturbation

CIExplainer++: Generating Causal and Interpretable Explanations for Graph Neural Networks

ResearchDGX agent

arXiv:2606.20747v1 Announce Type: new Abstract: Explainable Artificial Intelligence aims to make black-box models more trustworthy by presenting, in a human-understandable manner, the elements that le

Circuit realization and hardware linearization of monotone operator equilibrium networks

ResearchDGX agent

arXiv:2509.13793v2 Announce Type: replace-cross Abstract: It is shown that the port behavior of a resistor-diode network corresponds to the solution of a ReLU monotone operator equilibrium network (a

Circuit Synchronization Precedes Generalization: A Causal Precursor to Grokking

ResearchDGX agent

arXiv:2606.12966v2 Announce Type: replace Abstract: Grokking is the delayed generalisation phenomenon where a transformer trained on modular arithmetic abruptly transitions from near-chance to near-pe

CITADEL: CSI-Based Jamming Detection and Open-Set Classification for IIoT Networks

HardwareDGX agent

arXiv:2606.22939v1 Announce Type: cross Abstract: Radio frequency jamming poses a critical threat to the availability of wireless Industrial Internet of Things (IIoT) networks. Existing detection and

ClayBuddy: A Framework, Evaluation, & Mitigation of Coding Agent Failures

SafetyDGX agent

arXiv:2606.19380v2 Announce Type: replace-cross Abstract: Software engineering and deployment are increasingly delegated to AI coding agents. The scale of their adoption is surfacing rare, but highly

CLIP-guided Diffusion Model for Backdoor Generation in Sensor-based Human Activity Recognition

ResearchDGX agent

arXiv:2606.22837v1 Announce Type: new Abstract: Sensors are critical components of modern intelligent devices. The proliferation of the Internet of Things (IoT) and wearable mobile devices has enabled

Clipping the Price of Adaptivity at the Tail

Model ReleasesDGX agent

arXiv:2606.22669v1 Announce Type: new Abstract: Adaptive stochastic convex optimization (SCO) methods face a fundamental ``price of adaptivity'' barrier: under the standard set of assumptions, they ca

Closing the Calibration Gap in Semantic Caching

ResearchDGX agent

arXiv:2606.19719v2 Announce Type: replace-cross Abstract: Semantic caching cuts LLM inference costs by serving a cached response to semantically similar queries. Standard practice evaluates these syst

Cluster-Specific Localized Drift Detection for Efficient Batch Model Adaptation under Controlled Distribution Shift

Model ReleasesDGX agent

arXiv:2606.22026v1 Announce Type: new Abstract: Machine learning systems deployed in dynamic environments frequently operate under nonstationary data distributions, where controlled distribution shift

CogFormer: Learn All Your Models Once

TutorialsDGX agent

arXiv:2603.20520v2 Announce Type: replace-cross Abstract: Simulation-based inference (SBI) with neural networks has accelerated and transformed cognitive modeling workflows. SBI enables modelers to fi

Coherence Under Commitment: Probing Generalization and Vacuous Memorization in LLM Logical Reasoning

ResearchDGX agent

arXiv:2606.21083v1 Announce Type: cross Abstract: Large language models (LLMs) deployed for logical reasoning in knowledge-intensive domains exhibit a subtle but critical failure: coherence can be vac

Cohort-Anchored Foundation Models for Electronic Health Records: From Risk Scores to Auditable Peer Cohorts

SafetyDGX agent

arXiv:2606.21885v1 Announce Type: new Abstract: Foundation models have achieved remarkable performance across medical question answering, imaging, and electronic health record (EHR) tasks, yet reliabl

Collapsed Effective Operators for Higher-order Structures

ResearchDGX agent

arXiv:2606.23517v1 Announce Type: new Abstract: Higher-order structures are powerful relational modeling tools, yet existing spectral operators decompose the topology into separate ranks, leaving prac

Combinatorial Allocation Bandits with Nonlinear Arm Utility

ResearchDGX agent

arXiv:2603.07005v2 Announce Type: replace Abstract: A matching platform is a system that matches participants of different types, such as companies and job-seekers. In such a platform, maximizing matc

Combinatorial Sparse PCA Beyond the Spiked Identity Model

ApplicationsDGX agent

arXiv:2603.02607v2 Announce Type: replace-cross Abstract: Sparse PCA is one of the most well-studied problems in high-dimensional statistics. In this problem, we are given samples from a distribution

Comparative Evaluation of Machine Learning and Deep Learning Models for Wound-Rotor Synchronous Motor Performance Prediction

Model ReleasesDGX agent

arXiv:2606.21230v1 Announce Type: new Abstract: Wound rotor synchronous motors have emerged as a strong alternative that eliminates dependence on REEs. However, WRSM design requires the simultaneous o

Compress the Easy, Explore the Hard: Difficulty-Aware Entropy Regularization for Efficient LLM Reasoning

ApplicationsDGX agent

arXiv:2602.22642v2 Announce Type: replace Abstract: Chain-of-Thought (CoT) has substantially empowered Large Language Models (LLMs) to tackle complex reasoning tasks, yet the verbose nature of explici

Concept-Constrained Prompt Learning for Few-Shot CLIP Adaptation

Model ReleasesDGX agent

arXiv:2606.22567v1 Announce Type: new Abstract: Few-shot prompt learning is an effective strategy for adapting CLIP to downstream tasks, but class-only prompt optimization can overfit base-class super

Concordia: JIT-Compiled Persistent-Kernel Checkpointing for Fault-Tolerant LLM Inference

HardwareDGX agent

arXiv:2606.23521v1 Announce Type: cross Abstract: Long-running LLM agents keep valuable state resident on GPUs: KV caches, request schedulers, communication state, and sometimes online adapters. Losin

Conditional Flow Matching for Visually-Guided Acoustic Highlighting

SafetyDGX agent

arXiv:2602.03762v3 Announce Type: replace-cross Abstract: Visually-guided acoustic highlighting seeks to rebalance audio in alignment with the accompanying video, creating a coherent audio-visual expe

Conditional neural control variates for variance reduction in Bayesian inverse problems

ResearchDGX agent

arXiv:2602.21357v2 Announce Type: replace-cross Abstract: Bayesian inference for inverse problems involves computing expectations under posterior distributions--e.g., posterior means, variances, or pr

Confidently Wrong: Severity-Aware Calibration of Prompt-Injection Detectors under Attack Shift

Model ReleasesDGX agent

arXiv:2606.22659v1 Announce Type: cross Abstract: Prompt-injection detectors are deployed as guards: a model scores an input and a downstream system trusts or blocks it on that score. I study the conf

Conformal and kNN Predictive Uncertainty Quantification Algorithms in Metric Spaces

ResearchDGX agent

arXiv:2507.15741v2 Announce Type: replace-cross Abstract: This paper introduces a framework for uncertainty quantification in regression models defined on metric spaces. Using a proposed notion of hom

Continuous Behavioral Authentication via Multi-Expert BERT Log Analysis for Secure Data Sharing

SafetyDGX agent

arXiv:2606.21900v1 Announce Type: cross Abstract: Continuous authentication for mobile and zero-trust systems requires nonintrusive evidence confirming the enrolled user-device context remains valid a

Continuous-Time Probabilistic Correctors for Uncertainty-Aware Physics-Based Spacecraft Trajectory Forecasting

SafetyDGX agent

arXiv:2606.21021v1 Announce Type: new Abstract: Long-horizon spacecraft trajectory forecasting suffers from error accumulation due to the absence of corrective observations in the forecast regime, mak

Convergence Analysis of Nystrom Subsampling in Covariate Shift Adaptation for Misspecified case

ResearchDGX agent

arXiv:2606.22259v1 Announce Type: cross Abstract: This paper investigates convergence properties of regularized Nystrom subsampling applied to the unsupervised domain adaptation problem under covariat

Convergence of Gradient Descent for General Neural Network Architectures Beyond the NTK Regime

ResearchDGX agent

arXiv:2606.23364v1 Announce Type: new Abstract: Training dynamics is central to understanding neural networks, yet its theoretical analysis remains difficult even for simple architectures and becomes

Convergence Rate Analysis of LION

ResearchDGX agent

arXiv:2411.07724v2 Announce Type: replace Abstract: The LION (evoLved sIgn mOmeNtum) optimizer for deep neural network training was found by Google via program search, with the simple sign update yet

CoorDex: Coordinating Body and Hand Priors for Continuous Dexterous Humanoid Loco-Manipulation

SafetyDGX agent

arXiv:2606.23680v1 Announce Type: cross Abstract: Humanoid loco-manipulation is often simplified into a stop-and-go process: walking to an object, stopping to manipulate it, and then resuming locomoti

Counsel: A Meta-Evaluation Dataset for Agentic Tasks

SafetyDGX agent

arXiv:2606.21627v1 Announce Type: cross Abstract: As agentic systems tackle increasingly complex multi-step tasks, evaluating their trajectories presents a major bottleneck - human annotation of a sin

Counterfactual learning of new adaptive instructional policies using logged data

SafetyDGX agent

arXiv:2606.23015v1 Announce Type: new Abstract: Optimizing instructional policies in Intelligent Tutoring Systems (ITS) typically requires costly online experimentation or student simulators that may

CQD-SHAP: Explainable Complex Query Answering via Shapley Values

ResearchDGX agent

arXiv:2510.15623v2 Announce Type: replace Abstract: Complex query answering (CQA) goes beyond the widely studied link prediction task by addressing more sophisticated queries that require multi-hop re

CRAX: Fast Safe Reinforcement Learning Benchmarking

Model ReleasesDGX agent

arXiv:2606.20376v2 Announce Type: replace Abstract: Safety is a core concern for deploying reinforcement learning (RL) agents in real-world domains such as robotics and autonomous driving. While bench

CrediBench: Building Web-Scale Network Datasets for Information Integrity

ResearchDGX agent

arXiv:2509.23340v4 Announce Type: replace-cross Abstract: Automatically assessing the credibility of online sources presents an invaluable tool for navigating today's information ecosystem. However, e

CuratorKIT : Data Curation and Synthetic Data Generation for LLM Post-Training

ResearchDGX agent

arXiv:2606.21631v1 Announce Type: cross Abstract: Data curation is a critical part of post-training pipelines for large language models, yet existing tools often treat ingestion, deduplication, synthe

Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model

Model ReleasesDGX agent

arXiv:2606.22317v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) is widely viewed as a promising path toward continuously improving large language models. Recent w

Darwin Mobile Agent: A Roadmap for Self-Evolution

SafetyDGX agent

arXiv:2606.20622v1 Announce Type: cross Abstract: The goal of artificial intelligence is to create agents capable of general, adaptive behaviour in open-ended environments. Guided by the 'Bitter Lesso

Data Evolution by Wittgenstein's Rule Following

ResearchDGX agent

arXiv:2606.22674v1 Announce Type: cross Abstract: This paper introduces Wittgenstein's Rule Following (WRF) data evolution, a framework in philomatics for evolving or generating a new dataset from a s

Data Pruning: Redundant, Problematic, and Interdependent Samples

Model ReleasesDGX agent

arXiv:2606.21916v1 Announce Type: new Abstract: The performance of deep learning models is affected by not only data quantity but also data quality. Data pruning is a process by which practitioners ca

DataClaw0: Agentic Tailoring Multimodal Data from Raw Streams

Model ReleasesDGX agent

arXiv:2606.21337v1 Announce Type: new Abstract: Massive unstructured multimodal streams suffer from high 'data entropy,' impeding both efficient human knowledge acquisition and high-quality AI post-tr

DataMIL: Selecting Data for Robot Imitation Learning with Datamodels

SafetyDGX agent

arXiv:2505.09603v2 Announce Type: replace-cross Abstract: Recently, the robotics community has amassed ever larger and more diverse datasets to train generalist policies. However, while these policies

Dataset-Aware Cold-Start Active Learning for Annotation-Efficient 3D Medical Image Segmentation

Local AiDGX agent

arXiv:2606.20765v1 Announce Type: cross Abstract: Deep learning for 3D medical image segmentation requires extensive manual annotations, a major bottleneck in volumetric medical imaging. Active learni

DCD-PFN: A Decoupling-Aware Foundation Model for Causal Discovery

Local AiDGX agent

arXiv:2606.21212v1 Announce Type: new Abstract: Causal discovery is critical for understanding complex data-generating mechanisms, yet traditional algorithms often struggle with highly non-linear and

Dead-Direction Signatures: A Cheap Spectral Reading of Singular Complexity

Local AiDGX agent

arXiv:2606.21158v1 Announce Type: new Abstract: Singular learning theory characterises the complexity of a deep network through the geometry of its loss singularities. The local learning coefficient (

Decision-Focused Learning: When and Why Traditional Prediction Models Fail

TutorialsDGX agent

arXiv:2606.21773v1 Announce Type: new Abstract: Plugging predictions of unknown parameters into downstream optimization problems, often referred to as the ``predict-then-optimize'' paradigm, has long

Decodable but Not Faithful: Coupling Natural-Language Rationales to Programmatic Verifiers

ResearchDGX agent

arXiv:2606.21678v1 Announce Type: new Abstract: Language models can generate plausible rationales for their predictions, but these explanations may not faithfully represent the model's internal reason

← Previous
1…7576777879…243
Next →