AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Model Releases

PPU-Bench:Real World Benchmark for Personalized Partial Unlearning in Vision Language Models

DGX agent

arXiv:2605.08800v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) may memorize sensitive cross-modal information during pretraining. However, existing MLLM unlearning benchmar

model-releasesarxiv-cs-ai
12 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Research

Practical Wi-Fi-based Motion Recognition Under Variable Traffic Patterns

DGX agent

arXiv:2605.08308v1 Announce Type: cross Abstract: Wi-Fi sensing detects human motions and activities by analysing the channel state information (CSI) derived from Wi-Fi transmissions. However, the imp

researcharxiv-cs-ai
12 May 2026
Model Releases

PrAg-PO: Prompt Augmented Policy Optimization for Robust and Diverse Mathematical Reasoning

DGX agent

arXiv:2602.03190v3 Announce Type: replace-cross Abstract: Reinforcement learning algorithms such as group-relative policy optimization (GRPO) have shown strong potential for improving the mathematical

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Prediction Bottlenecks Don't Discover Causal Structure (But Here's What They Actually Do)

DGX agent

arXiv:2605.09169v1 Announce Type: cross Abstract: A Mamba state-space model trained only for next-step prediction appears to recover Granger-causal structure through a simple readout S = |W_{out} W_{i

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

PrepBench: How Far Are We from Natural-Language-Driven Data Preparation?

DGX agent

arXiv:2605.08687v1 Announce Type: cross Abstract: Data preparation is a central and time-consuming stage in data analysis workflows. Traditionally, commercial tools have relied on graphical user inter

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Pretraining large language models with MXFP4

DGX agent

arXiv:2605.09825v1 Announce Type: cross Abstract: Why does full-pipeline FP4 training of large language models often diverge, even when forward activations and activation gradients remain stable? We a

model-releasesarxiv-cs-ai
12 May 2026
Local Ai

Preventing Rank Collapse in Federated Low-Rank Adaptation with Client Heterogeneity

DGX agent

arXiv:2602.13486v2 Announce Type: replace-cross Abstract: Federated low-rank adaptation (FedLoRA) has facilitated communication-efficient and privacy-preserving fine-tuning of foundation models for do

local-aiarxiv-cs-ai
12 May 2026
Safety

Primal-Dual Guided Decoding for Constrained Discrete Diffusion

DGX agent

arXiv:2605.09749v1 Announce Type: new Abstract: Discrete diffusion models generate structured sequences by progressively unmasking tokens, but enforcing global property constraints during generation r

safetyarxiv-cs-ai
12 May 2026
Model Releases

PrimeKG-CL: A Continual Graph Learning Benchmark on Evolving Biomedical Knowledge Graphs

DGX agent

arXiv:2605.10529v1 Announce Type: new Abstract: Biomedical knowledge graphs underwrite drug repurposing and clinical decision support, yet the upstream ontologies they depend on update on independent

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Priming: Hybrid State Space Models From Pre-trained Transformers

DGX agent

arXiv:2605.08301v1 Announce Type: cross Abstract: Hybrid State-Space models combine Attention with recurrent State-Space Model (SSM) layers, balancing eidetic memory from Attention with compressed fad

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

PRISM: Generation-Time Detection and Mitigation of Secret Leakage in Multi-Agent LLM Pipelines

DGX agent

arXiv:2605.10614v1 Announce Type: new Abstract: Multi-agent LLM systems introduce a security risk in which sensitive information accessed by one agent can propagate through shared context and reappear

model-releasesarxiv-cs-ai
12 May 2026
Safety

Privacy-Aware Video Anomaly Detection through Orthogonal Subspace Projection

DGX agent

arXiv:2605.08651v1 Announce Type: cross Abstract: Video anomaly detection (VAD) systems often prioritize accuracy while overlooking privacy concerns, limiting their suitability for real-world deployme

safetyarxiv-cs-ai
12 May 2026
Local Ai

Privacy-Preserving Federated Learning: Integrating Zero-Knowledge Proofs in Scalable Distributed Architectures

DGX agent

arXiv:2605.08152v1 Announce Type: cross Abstract: The intersection of Artificial Intelligence (AI) and distributed systems has given rise to Federated Learning (FL), a paradigm that enables decentrali

local-aiarxiv-cs-ai
12 May 2026
Model Releases

ProactBench: Beyond What The User Asked For

DGX agent

arXiv:2605.09228v1 Announce Type: cross Abstract: Most LLM benchmarks score how well a model responds to explicit requests. They leave unmeasured a different conversational ability: noticing and actin

model-releasesarxiv-cs-ai
12 May 2026
Research

Probing Cross-modal Information Hubs in Audio-Visual LLMs

DGX agent

arXiv:2605.10815v1 Announce Type: new Abstract: Audio-visual large language models (AVLLMs) have recently emerged as a powerful architecture capable of jointly reasoning over audio, visual, and textua

researcharxiv-cs-ai
12 May 2026
Research

Probing Routing-Conditional Calibration in Attention-Residual Transformers

DGX agent

arXiv:2605.09850v1 Announce Type: cross Abstract: Post-hoc calibration is usually evaluated as a function of logits or softmax confidence alone, even as routing-augmented architectures increasingly ac

researcharxiv-cs-ai
12 May 2026
Model Releases

Probing the Critical Point (CritPt) of AI Reasoning: a Frontier Physics Research Benchmark

DGX agent

arXiv:2509.26574v4 Announce Type: replace Abstract: While large language models (LLMs) with reasoning capabilities are progressing rapidly on high-school math competitions and coding, can they reason

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Probing the Impact of Scale on Data-Efficient, Generalist Transformer World Models for Atari

DGX agent

arXiv:2605.08578v1 Announce Type: cross Abstract: Developing generalist systems that retain human-like data efficiency is a central challenge. While world models (WMs) offer a promising path, existing

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Process Matters more than Output for Distinguishing Humans from Machines

DGX agent

arXiv:2605.06524v2 Announce Type: replace Abstract: Reliable human-machine discrimination is becoming increasingly important as large language models and autonomous agents are deployed in online setti

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Prompt-Activation Duality: Improving Activation Steering via Attention-Level Interventions

DGX agent

arXiv:2605.10664v1 Announce Type: cross Abstract: Activation steering controls language model behavior by adding directions to internal representations at inference time, but standard residual-stream

model-releasesarxiv-cs-ai
12 May 2026
Research

PromptDx: Differentiable Prompt Tuning for Multimodal In-Context Alzheimer's Diagnosis

DGX agent

arXiv:2605.08585v1 Announce Type: cross Abstract: Deep learning models in medical imaging typically operate as parametric memory, diagnosing patients by recalling fixed knowledge learned during traini

researcharxiv-cs-ai
12 May 2026
Safety

PromptGuard: Soft Prompt-Guided Unsafe Content Moderation for Text-to-Image Models

DGX agent

arXiv:2501.03544v5 Announce Type: replace-cross Abstract: Recent text-to-image (T2I) models have exhibited remarkable performance in generating high-quality images from text descriptions. However, the

safetyarxiv-cs-ai
12 May 2026
Applications

Prospective Compression in Human Abstraction Learning

DGX agent

arXiv:2605.09985v1 Announce Type: new Abstract: A core challenge in program synthesis is online library learning: the incremental acquisition of reusable abstractions under uncertainty about future ta

applicationsarxiv-cs-ai
12 May 2026
Safety

ProteinOPD: Towards Effective and Efficient Preference Alignment for Protein Design

DGX agent

arXiv:2605.10189v1 Announce Type: cross Abstract: Designing proteins with desired functions or properties represents a core goal in synthetic biology and drug discovery. Recent advances in protein lan

safetyarxiv-cs-ai
12 May 2026
Research

Provable Anytime Ensemble Sampling Algorithms in Nonlinear Contextual Bandits

DGX agent

arXiv:2510.10730v2 Announce Type: replace-cross Abstract: We provide a unified algorithmic framework for ensemble sampling in nonlinear contextual bandits and develop corresponding regret bounds for t

researcharxiv-cs-ai
12 May 2026
Research

Provable Sparse Inversion and Token Relabel Enhanced One-shot Federated Learning with ViTs

DGX agent

arXiv:2605.10748v1 Announce Type: cross Abstract: One-Shot Federated Learning, where a central server learns a global model in a single communication round, has emerged as a promising paradigm. Howeve

researcharxiv-cs-ai
12 May 2026
Tutorials

PruneTIR: Inference-Time Tool Call Pruning for Effective yet Efficient Tool-Integrated Reasoning

DGX agent

arXiv:2605.09931v1 Announce Type: cross Abstract: Tool-integrated reasoning (TIR) enables large language models (LLMs) to enhance their capabilities by interacting with external tools, such as code in

tutorialsarxiv-cs-ai
12 May 2026
Safety

Pseudo-Deliberation in Language Models: When Reasoning Fails to Align Values and Actions

DGX agent

arXiv:2605.09893v1 Announce Type: cross Abstract: Large language models (LLMs) are often evaluated based on their stated values, yet these do not reliably translate into their actions, a discrepancy t

safetyarxiv-cs-ai
12 May 2026
Local Ai

PYTHALAB-MERA: Validation-Grounded Memory, Retrieval, and Acceptance Control for Frozen-LLM Coding Agents

DGX agent

arXiv:2605.08468v1 Announce Type: cross Abstract: Local LLM-based coding agents increasingly work in settings where correctness is earned through execution feedback, persistent state, and bounded repa

local-aiarxiv-cs-ai
12 May 2026
Safety

Q-learning with Adjoint Matching

DGX agent

arXiv:2601.14234v3 Announce Type: replace-cross Abstract: We propose Q-learning with Adjoint Matching (QAM), a novel TD-based reinforcement learning (RL) algorithm that tackles a long-standing challen

safetyarxiv-cs-ai
12 May 2026
Tutorials

Quantile Geometry Regularization for Distributional Reinforcement Learning

DGX agent

arXiv:2605.08182v1 Announce Type: cross Abstract: Quantile-based distributional reinforcement learning methods learn return distributions through sampled quantile regression, but their bootstrapped ta

tutorialsarxiv-cs-ai
12 May 2026
Model Releases

Qwen Goes Brrr: Off-the-Shelf RAG for Ukrainian Multi-Domain Document Understanding

DGX agent

arXiv:2605.10296v1 Announce Type: cross Abstract: We participated in the Fifth UNLP shared task on multi-domain document understanding, where systems must answer Ukrainian multiple-choice questions fr

model-releasesarxiv-cs-ai
12 May 2026
Agents

RADAR: Redundancy-Aware Diffusion for Multi-Agent Communication Structure Generation

DGX agent

arXiv:2605.09907v1 Announce Type: new Abstract: Compared with individual agents, large language model based multi-agent systems have shown great capabilities consistently across diverse tasks, includi

agentsarxiv-cs-ai
12 May 2026
Applications

RAG-HAR: Retrieval Augmented Generation-based Human Activity Recognition

DGX agent

arXiv:2512.08984v2 Announce Type: replace-cross Abstract: Human Activity Recognition (HAR) underpins applications in healthcare, rehabilitation, fitness tracking, and smart environments, yet existing

applicationsarxiv-cs-ai
12 May 2026
Research

Randomized PCA Forest for Unsupervised Outlier Detection

DGX agent

arXiv:2508.12776v3 Announce Type: replace-cross Abstract: We propose a novel unsupervised outlier detection method based on Randomized Principal Component Analysis (PCA). Motivated by the performance

researcharxiv-cs-ai
12 May 2026
Local Ai

RAwR: Role-Aware Rewiring via Approximate Equitable Partition

DGX agent

arXiv:2605.09457v1 Announce Type: cross Abstract: While Graph Neural Networks (GNNs) have demonstrated significant efficacy in node classification tasks, where predictions rely on local neighborhood i

local-aiarxiv-cs-ai
12 May 2026
Model Releases

RDEx-CASK: Cauchy Mutation, Archive, and Stagnation Kick for RDEx-CSOP

DGX agent

arXiv:2605.09652v1 Announce Type: cross Abstract: We extend RDEx-CSOP with 3 changes that target stagnation & late-stage variance, plus minor parameter tuning. The second scale factor in the standard

model-releasesarxiv-cs-ai
12 May 2026
Research

RDKV: Rate-Distortion Bit Allocation for Joint Eviction and Quantization of the KV Cache

DGX agent

arXiv:2605.08317v1 Announce Type: cross Abstract: Large language models (LLMs) have shown strong performance across diverse tasks, but their inference with long input contexts is bottlenecked by memor

researcharxiv-cs-ai
12 May 2026
Safety

Re-Triggering Safeguards within LLMs for Jailbreak Detection

DGX agent

arXiv:2605.10611v1 Announce Type: cross Abstract: This paper proposes a jailbreaking prompt detection method for large language models (LLMs) to defend against jailbreak attacks. Although recent LLMs

safetyarxiv-cs-ai
12 May 2026
Model Releases

Re^2Math: Benchmarking Theorem Retrieval in Research-Level Mathematics

DGX agent

arXiv:2605.09012v1 Announce Type: new Abstract: Large language models are increasingly capable at closed-world mathematical reasoning, but research assistance also requires source-grounded use of the

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Real vs. Semi-Simulated: Rethinking Evaluation for Treatment Effect Estimation

DGX agent

arXiv:2605.10430v1 Announce Type: cross Abstract: Estimating heterogeneous treatment effects with machine learning has attracted substantial attention in both academic research and industrial practice

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

REAP: Automatic Curation of Coding Agent Benchmarks from Interactive Production Usage

DGX agent

arXiv:2604.01527v3 Announce Type: replace-cross Abstract: Production deployment of AI coding agents requires fast, reproducible evaluation signals. Existing industrial practices trade off speed and fi

model-releasesarxiv-cs-ai
12 May 2026
Agents

REAP: Reinforcement-Learning End-to-End Autonomous Parking with Gaussian Splatting Simulator for Real2Sim2Real Transfer

DGX agent

arXiv:2605.08713v1 Announce Type: cross Abstract: In recent years, autonomous parking has made significant advances, yet parking tasks still face challenges in extreme scenarios such as mechanical and

agentsarxiv-cs-ai
12 May 2026
Research

Reasoning-Aware Training for Time Series Forecasting

DGX agent

arXiv:2605.08625v1 Announce Type: cross Abstract: Time Series Foundation Models (TSFMs) excel at numerical forecasting but operate as black boxes lacking qualitative reasoning. Conversely, applying LL

researcharxiv-cs-ai
12 May 2026
Safety

Reasoning Compression with Mixed-Policy Distillation

DGX agent

arXiv:2605.08776v1 Announce Type: new Abstract: Reasoning-centric large language models (LLMs) achieve strong performance by generating intermediate reasoning trajectories, but often incur excessive t

safetyarxiv-cs-ai
12 May 2026
Safety

Reasoning Is Not Free: Robust Adaptive Cost-Efficient Routing for LLM-as-a-Judge

DGX agent

arXiv:2605.10805v1 Announce Type: new Abstract: Reasoning-capable large language models (LLMs) have recently been adopted as automated judges, but their benefits and costs in LLM-as-a-Judge settings r

safetyarxiv-cs-ai
12 May 2026
Research

Reconciling Consistency-Based Diagnosis with Actual-Causality-Based Explanations

DGX agent

arXiv:2605.08688v1 Announce Type: new Abstract: We establish, from the point of view of Explainable AI (XAI), connections between Consistency-Based Diagnosis (CBD), on one side, and Actual Causality a

researcharxiv-cs-ai
12 May 2026
Model Releases

Recovering Physical Dynamics from Discrete Observations via Intrinsic Differential Consistency

DGX agent

arXiv:2605.08454v1 Announce Type: cross Abstract: Recovering continuous-time dynamics from discrete observations is difficult because local supervision (e.g., pointwise regression targets, derivative

model-releasesarxiv-cs-ai
12 May 2026
← Previous
1…340341342343344…448
Next →