AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
13 May 2026

Causal Bias Detection in Generative Artifical Intelligence

SafetyDGX agent

arXiv:2605.11365v1 Announce Type: cross Abstract: Automated systems built on artificial intelligence (AI) are increasingly deployed across high-stakes domains, raising critical concerns about fairness

Causal Fairness for Survival Analysis

SafetyDGX agent

arXiv:2605.11362v1 Announce Type: new Abstract: In the data-driven era, large-scale datasets are routinely collected and analyzed using machine learning (ML) and artificial intelligence (AI) to inform

ChunkFlow: Communication-Aware Chunked Prefetching for Layerwise Offloading in Distributed Diffusion Transformer Inference

HardwareDGX agent

arXiv:2605.11335v1 Announce Type: cross Abstract: Layerwise offloading reduces the GPU memory footprint of large diffusion transformer (DiT) inference by prefetching upcoming layers from host memory,


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Clarity: The Flexibility-Interpretability Trade-Off in Sparsity-aware Concept Bottleneck Models

SafetyDGX agent

arXiv:2601.21944v2 Announce Type: replace Abstract: The widespread adoption of deep learning models in computer vision has intensified concerns about interpretability. Despite strong performance, thes

Compositional Neural Operators for Multi-Dimensional Fluid Dynamics

ResearchDGX agent

arXiv:2605.11691v1 Announce Type: new Abstract: Partial differential equations (PDEs) govern diverse physical phenomena, yet high-fidelity numerical solutions are computationally expensive and Machine

Constrained Stochastic Spectral Preconditioning Converges for Nonconvex Objectives

ResearchDGX agent

arXiv:2605.11850v1 Announce Type: cross Abstract: In this work, we develop proximal preconditioned gradient methods with a focus on spectral gradient methods providing a proximal extension to the Muon

Context Steering: A New Paradigm for Compression-based Embeddings by Synthesizing Relevant Information Features

ApplicationsDGX agent

arXiv:2508.14780v2 Announce Type: replace Abstract: Compression-based dissimilarities (CD) offer a flexible and domain-agnostic means of measuring similarity by identifying implicit information throug

Continuous Discovery of Vulnerabilities in LLM Serving Systems with Fuzzing

SafetyDGX agent

arXiv:2605.11202v1 Announce Type: cross Abstract: LLM inference and serving systems have become security-critical infrastructure; however, many of their most concerning failures arise from the serving

Control Charts for Multi-agent Systems

AgentsDGX agent

arXiv:2605.11135v1 Announce Type: cross Abstract: Generative agents have proven to be powerful assistants in a wide variety of contexts. Given this success, users are now deploying agents with minimal

CORE: Cyclic Orthotope Relation Embedding for Knowledge Graph Completion

Model ReleasesDGX agent

arXiv:2605.11159v1 Announce Type: new Abstract: Knowledge graph completion (KGC) aims to automatically infer missing facts in multi-relational data by mapping entities and relations into continuous re

COSMOS: Model-Agnostic Personalized Federated Learning with Clustered Server Models and Pseudo-Label-Only Communication

Local AiDGX agent

arXiv:2605.11165v1 Announce Type: new Abstract: Federated learning (FL) in heterogeneous environments remains challenging because client models often differ in both architecture and data distribution.

Crash Assessment via Mesh-Based Graph Neural Networks and Physics-Aware Attention

Model ReleasesDGX agent

arXiv:2605.11784v1 Announce Type: cross Abstract: Full-vehicle crash simulations are computationally expensive, limiting their use in iterative design exploration. This work investigates learned hybri

CTFusion: A CTF-based Benchmark for LLM Agent Evaluation

Model ReleasesDGX agent

arXiv:2605.11504v1 Announce Type: new Abstract: Recent advances in Large Language Models (LLMs) have enabled agentic systems for complex, multi-step tasks; cybersecurity is emerging as a prominent app

Curriculum Learning-Guided Progressive Distillation in Large Language Models

ResearchDGX agent

arXiv:2605.11260v1 Announce Type: new Abstract: Knowledge distillation is a key technique for transferring the capabilities of large language models (LLMs) into smaller, more efficient student models.

DAG: A Dual Correlation Network for Time Series Forecasting with Exogenous Variables

ResearchDGX agent

arXiv:2509.14933v3 Announce Type: replace Abstract: Time series forecasting is essential in various domains. Compared to relying solely on endogenous variables (i.e., target variables), considering ex

Debiased Model-based Representations for Sample-efficient Continuous Control

SafetyDGX agent

arXiv:2605.11711v1 Announce Type: new Abstract: Model-based representations recently stand out as a promising framework that embeds latent dynamics information into the representations for downstream

Debiasing Message Passing to Mitigate Popularity Bias in GNN-based Collaborative Filtering

SafetyDGX agent

arXiv:2605.11145v1 Announce Type: cross Abstract: Collaborative filtering (CF) models based on graph neural networks (GNNs) achieve strong performance in recommender systems by propagating user-item s

Decomposing the Generalization Gap in PROTAC Activity Prediction: Variance Attribution and the Inter-Laboratory Ceiling

SafetyDGX agent

arXiv:2605.11764v1 Announce Type: new Abstract: Machine-learning predictors of biochemical activity often exhibit large random-split-to-leave-one-target-out generalisation gaps that have been document

DeconDTN-Toolkit: A Library for Evaluation and Enhancement of Robustness to Provenance Shift

ResearchDGX agent

arXiv:2605.11237v1 Announce Type: new Abstract: Despite the burgeoning body of work on distribution shifts, provenance shift-where the relationship between data source and label changes at deployment-

Deep Learning for Protein Complex Prediction and Design

ResearchDGX agent

arXiv:2605.11189v1 Announce Type: new Abstract: Accurately modeling and designing protein complex structures is a central problem in computational structural biology, with broad implications for under

Deep Minds and Shallow Probes

ApplicationsDGX agent

arXiv:2605.11448v1 Announce Type: new Abstract: Neural representations are not unique objects. Even when two systems realize the same downstream computation, their hidden coordinates may differ by rep

Delay-Empowered Causal Hierarchical Reinforcement Learning

ApplicationsDGX agent

arXiv:2605.12261v1 Announce Type: new Abstract: Many real-world tasks involve delayed effects, where the outcomes of actions emerge after varying time lags. Existing delay-aware reinforcement learning

Delightful Gradients Accelerate Corner Escape

Model ReleasesDGX agent

arXiv:2605.11908v1 Announce Type: new Abstract: Softmax policy gradient converges at O(1/t), but its transient behavior near sub-optimal corners of the simplex can be exponentially slow. The bottlenec

Detecting In-Person Conversations in Noisy Real-World Environments with Smartwatch Audio and Motion Sensing

ApplicationsDGX agent

arXiv:2507.12002v2 Announce Type: replace Abstract: Social interactions play a crucial role in shaping human behavior, relationships, and societies. It encompasses various forms of communication, such

Detecting overfitting in Neural Networks during long-horizon grokking using Random Matrix Theory

ResearchDGX agent

arXiv:2605.12394v1 Announce Type: new Abstract: Training Neural Networks (NNs) without overfitting is difficult; detecting that overfitting is difficult as well. We present a novel Random Matrix Theor

Dirichlet process mixtures of block g priors for model selection and prediction in linear models

ResearchDGX agent

arXiv:2411.00471v3 Announce Type: replace-cross Abstract: This paper introduces Dirichlet process mixtures of block g priors for model selection and prediction in linear models. These priors are exten

DisagMoE: Computation-Communication overlapped MoE Training via Disaggregated AF-Pipe Parallelism

Model ReleasesDGX agent

arXiv:2605.11005v1 Announce Type: new Abstract: Mixture-of-experts (MoE) architectures enable trillion-parameter LLMs with sparsely activated experts. Expert parallelism (EP) is a widely adopted MoE t

Discrete Flow Matching for Offline-to-Online Reinforcement Learning

SafetyDGX agent

arXiv:2605.12379v1 Announce Type: new Abstract: Many reinforcement learning (RL) tasks have discrete action spaces, but most generative policy methods based on diffusion and flow matching are designed

Dissecting Discrete Soft Actor-Critic: Limitations and Principled Alternatives

SafetyDGX agent

arXiv:2509.09838v2 Announce Type: replace Abstract: While Soft Actor-Critic (SAC) is highly effective in continuous control, its discrete counterpart (DSAC) performs poorly on challenging discrete-act

Distributed Quantum Gaussian Processes for Multi-Agent Systems

Local AiDGX agent

arXiv:2602.15006v2 Announce Type: replace-cross Abstract: Gaussian Processes (GPs) are a powerful tool for probabilistic modeling, but their performance is often constrained in complex, large-scale re

DiVeQ: Differentiable Vector Quantization Using the Reparameterization Trick

ResearchDGX agent

arXiv:2509.26469v3 Announce Type: replace Abstract: Vector quantization is common in deep models, yet its hard assignments block gradients and hinder end-to-end training. We propose DiVeQ, which treat

DP-{lambda}CGD: Efficient Noise Correlation for Differentially Private Model Training

ResearchDGX agent

arXiv:2601.22334v2 Announce Type: replace Abstract: Differentially private stochastic gradient descent (DP-SGD) is the gold standard for training machine learning models with formal differential priva

DreamPolicy: A Unified World-model Policy for Scalable Humanoid Locomotion

SafetyDGX agent

arXiv:2505.18780v3 Announce Type: replace-cross Abstract: Achieving versatile humanoid locomotion with a single policy presents a critical scalability challenge. Prevailing methods often rely on disti

DriftXpress: Faster Drifting Models via Projected RKHS Fields

ResearchDGX agent

arXiv:2605.12183v1 Announce Type: new Abstract: Drifting Models have emerged as a new paradigm for one-step generative modeling, achieving strong image quality without iterative inference. The premise

Drop the Act: Probe-Filtered RL for Faithful Chain-of-Thought Reasoning

Model ReleasesDGX agent

arXiv:2605.11467v1 Announce Type: new Abstract: Reasoning models post-hoc rationalize answers they have already committed to internally, producing chains of *reasoning theater*: deliberative-looking s

Dual-Temporal LSTM with Hybrid Attention for Airline Passenger Load Factor Forecasting: Integrating Intra-Flight and Inter-Flight Booking Dynamics

ApplicationsDGX agent

arXiv:2605.11569v1 Announce Type: cross Abstract: Accurate short-term demand forecasting is crucial to airline revenue management, yet most existing systems fail to meet this need because current mode

ECTO: Exogenous-Conditioned Temporal Operator for Ultra-Short-Term Wind Power Forecasting

SafetyDGX agent

arXiv:2605.12196v1 Announce Type: new Abstract: Accurate ultra-short-term wind power forecasting is critical for grid dispatch and reserve management, yet remains challenging due to the non-stationary

Efficient Adjoint Matching for Fine-tuning Diffusion Models

ResearchDGX agent

arXiv:2605.11480v1 Announce Type: new Abstract: Reward fine-tuning has become a common approach for aligning pretrained diffusion and flow models with human preferences in text-to-image generation. Am

Efficient Algorithms for Logistic Contextual Slate Bandits with Bandit Feedback

Model ReleasesDGX agent

arXiv:2506.13163v3 Announce Type: replace Abstract: We study the Logistic Contextual Slate Bandit problem, where, at each round, an agent selects a slate of N items from an exponentially large set (of

Efficient and Adaptive Human Activity Recognition via LLM Backbones

Model ReleasesDGX agent

arXiv:2605.12019v1 Announce Type: new Abstract: Human Activity Recognition (HAR) is a core task in pervasive computing systems, where models must operate under strict computational constraints while r

Efficient and provably convergent end-to-end training of deep neural networks with linear constraints

ApplicationsDGX agent

arXiv:2605.11526v1 Announce Type: cross Abstract: Training a deep neural network with the outputs of selected layers satisfying linear constraints is required in many contemporary data-driven applicat

Efficient LLM Reasoning via Variational Posterior Guidance with Efficiency Awareness

Model ReleasesDGX agent

arXiv:2605.11019v1 Announce Type: new Abstract: Although large language models rely on chain-of-thought for complex reasoning, the overthinking phenomenon severely degrades inference efficiency. Exist

Efficient Remote KV Cache Reuse with GPU-native Video Codec

HardwareDGX agent

arXiv:2602.09725v3 Announce Type: replace-cross Abstract: Remote KV cache reuse fetches KV cache for identical contexts from remote storage, avoiding recomputation, accelerating LLM inference. While i

EHR-RAGp: Retrieval-Augmented Prototype-Guided Foundation Model for Electronic Health Records

SafetyDGX agent

arXiv:2605.12335v1 Announce Type: cross Abstract: Electronic Health Records (EHR) contain rich longitudinal patient information and are widely used in predictive modeling applications. However, effect

Elicitation-Augmented Bayesian Optimization

ResearchDGX agent

arXiv:2605.12079v1 Announce Type: new Abstract: Human-in-the-loop Bayesian optimization (HITL BO) methods utilize human expertise to improve the sample-efficiency of BO. Most HITL BO methods assume th

Enabling AI-Native Mobility in 6G: A Real-World Dataset for Handover, Beam Management, and Timing Advance

ApplicationsDGX agent

arXiv:2605.12453v1 Announce Type: cross Abstract: To address the issues of high interruption time and measurement report overhead under user equipment (UE) mobility especially in high speed 5G use cas

Enabling Performant and Flexible Model-Internal Observability for LLM Inference

SafetyDGX agent

arXiv:2605.11093v1 Announce Type: new Abstract: Today's inference-time workloads increasingly depend on timely access to a model's internal states. We present DMI-Lib, a high-speed deep model inspecto

Enforcing Constraints in Generative Sampling via Adaptive Correction Scheduling

SafetyDGX agent

arXiv:2605.11214v1 Announce Type: new Abstract: Hard constraints in generative sampling are typically enforced by projection, applied either once at the end of sampling or after every update. This bin

Environment-Adaptive Preference Optimization for Wildfire Prediction

ApplicationsDGX agent

arXiv:2605.12435v1 Announce Type: new Abstract: Predicting rare extreme events such as wildfires from meteorological data requires models that remain reliable under evolving environmental conditions.

EpiCastBench: Datasets and Benchmarks for Multivariate Epidemic Forecasting

ResearchDGX agent

arXiv:2605.11598v1 Announce Type: new Abstract: The increasing adoption of data-driven decision-making in public health has established epidemic forecasting as a critical area of research. Recent adva

Epistemic Uncertainty for Test-Time Discovery

SafetyDGX agent

arXiv:2605.11328v1 Announce Type: new Abstract: Automated scientific discovery using large language models relies on identifying genuinely novel solutions. Standard reinforcement learning penalizes hi

EqOD: Symmetry-Informed Stability Selection for PDE Identification

ResearchDGX agent

arXiv:2605.11524v1 Announce Type: new Abstract: Data-driven identification of partial differential equations (PDEs) relies on sparse regression over a candidate library of differential operators, wher

Error whitening: Why Gauss-Newton outperforms Newton

ResearchDGX agent

arXiv:2605.11316v1 Announce Type: new Abstract: The Gauss-Newton matrix is widely viewed as a positive semidefinite approximation of the Hessian, yet mounting empirical evidence shows that Gauss-Newto

EsoLang-Bench: Evaluating Genuine Reasoning in Large Language Models via Esoteric Programming Languages

Model ReleasesDGX agent

arXiv:2603.09678v2 Announce Type: replace-cross Abstract: Large language models achieve near-ceiling performance on code generation benchmarks, yet most of the programming languages used by popular be

Estimating Subgraph Importance with Structural Prior Domain Knowledge

ApplicationsDGX agent

arXiv:2605.12009v1 Announce Type: new Abstract: We propose a subgraph importance estimation method for pretrained Graph Neural Networks (GNNs) on graph-level tasks, formulated as a linear Group Lasso

Events as Triggers for Behavioral Diversity in Multi-Agent Reinforcement Learning

SafetyDGX agent

arXiv:2605.12388v1 Announce Type: cross Abstract: Effective multi-agent cooperation requires agents to adopt diverse behaviors as task conditions evolve-and to do so at the right moment. Yet, current

Evolutionary Task Discovery: Advancing Reasoning Frontiers via Skill Composition and Complexity Scaling

ResearchDGX agent

arXiv:2605.11666v1 Announce Type: new Abstract: The reasoning frontier of Large Language Models (LLMs) has advanced significantly through modern post-training paradigms (e.g., Reinforcement Learning f

Exact Stiefel Optimization for Probabilistic PLS: Closed-Form Updates, Error Bounds, and Calibrated Uncertainty

Model ReleasesDGX agent

arXiv:2605.11607v1 Announce Type: cross Abstract: Probabilistic partial least squares (PPLS) is a central likelihood-based model for two-view learning when one needs both interpretable latent factors

Expected Batch Optimal Transport Plans and Consequences for Flow Matching

SafetyDGX agent

arXiv:2605.12174v1 Announce Type: new Abstract: Solving optimal transport (OT) on random minibatches is a common surrogate for exact OT in large-scale learning. In flow matching (FM), this surrogate i

ExploitGym: Can AI Agents Turn Security Vulnerabilities into Real Attacks?

Model ReleasesDGX agent

arXiv:2605.11086v1 Announce Type: cross Abstract: AI agents are rapidly gaining capabilities that could significantly reshape cybersecurity, making rigorous evaluation urgent. A critical capability is

← Previous
1…162163164165166…243
Next →