AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
84,661Total entries
1Added by human
84,660Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Safety

Reward Bias Substitution: Single-Axis Bias Mitigations Redirect Optimization Pressure

DGX agent

arXiv:2605.27996v1 Announce Type: new Abstract: Single-axis mitigations of reward-model biases (e.g., reducing proxy reliance on length, sycophancy, or style) can rotate optimization pressure onto cor

safetyarxiv-cs-ai
28 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Risk-Controlled Lean-as-Judge for Natural-Language Mathematical Reasoning

DGX agent

arXiv:2605.28365v1 Announce Type: new Abstract: Lean is increasingly used to judge natural-language mathematical answers, but its signal is partial: many answers never formalize, and a failed proof ma

researcharxiv-cs-ai
28 May 2026
Research

RL Squeezes, SFT Expands: A Comparative Study of Reasoning LLMs

DGX agent

arXiv:2509.21128v2 Announce Type: replace Abstract: Large language models (LLMs) are typically trained by reinforcement learning (RL) with verifiable rewards (RLVR) and supervised fine-tuning (SFT) on

researcharxiv-cs-ai
28 May 2026
Safety

Routing-Aligned Fine-Tuning for Multilingual Downstream Tasks in Mixture-of-Experts Models

DGX agent

arXiv:2605.28306v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) models have emerged as a dominant paradigm for efficient LLM scaling, yet adapting them to non-English downstream tasks remai

safetyarxiv-cs-ai
28 May 2026
Local Ai

ROVER: Routing Object-Centric Visual Evidence for Grounded Multi-Image Reasoning

DGX agent

arXiv:2605.27959v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have increasingly localized and interleaved visual evidence for deliberative reasoning. Grounding-based appro

local-aiarxiv-cs-ai
28 May 2026
Research

RULER: Representation-Level Verification of Machine Unlearning

DGX agent

arXiv:2605.27569v1 Announce Type: new Abstract: Machine unlearning aims to remove the influence of specific training records from a deployed model without retraining from scratch. Current protocols ve

researcharxiv-cs-ai
28 May 2026
Safety

SafeMed-R1: Clinician-Audited Safety and Ethics Alignment for Medical Large Language Models

DGX agent

arXiv:2605.28338v1 Announce Type: new Abstract: Large language models(LLMs) increasingly match expert performance on licensing examinations, yet routine clinical use remains limited because governance

safetyarxiv-cs-ai
28 May 2026
Model Releases

SAME: Stabilized Mixture-of-Experts for Multimodal Continual Instruction Tuning

DGX agent

arXiv:2602.01990v2 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) achieve strong performance through instruction tuning, but real-world deployment requires them to con

model-releasesarxiv-cs-ai
28 May 2026
Safety

SARAD: LLM-Based Safety-Aware Hybrid Reinforcement Learning with Collision Prediction for Autonomous Driving

DGX agent

arXiv:2605.28583v1 Announce Type: cross Abstract: Ensuring both safety and efficiency in decision-making for autonomous driving systems remains a fundamental challenge. Traditional Deep Reinforcement

safetyarxiv-cs-ai
28 May 2026
Research

Satisfiability Solving with LLMs: A Matched-Pair Evaluation of Reasoning Capability

DGX agent

arXiv:2605.28602v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for tasks that implicitly reduce to Boolean satisfiability (SAT), yet their reasoning ability on SAT

researcharxiv-cs-ai
28 May 2026
Research

Score Based Error Correcting Code Decoder

DGX agent

arXiv:2605.28358v1 Announce Type: cross Abstract: Error-correcting codes enable reliable communication, yet practical soft decoding remains challenging across code families and block lengths. We propo

researcharxiv-cs-ai
28 May 2026
Agents

Securing Retrieval-Augmented Generation: A Taxonomy of Attacks, Defenses, and Future Directions

DGX agent

arXiv:2604.08304v2 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) extends large language models (LLMs) with external knowledge, but this access path also introduces securi

agentsarxiv-cs-ai
28 May 2026
Research

SelfJudge: Faster Speculative Decoding via Self-Supervised Judge Verification

DGX agent

arXiv:2510.02329v2 Announce Type: replace-cross Abstract: Speculative decoding accelerates LLM inference by verifying candidate tokens from a draft model against a larger target model. Recent judge de

researcharxiv-cs-ai
28 May 2026
Research

Semantic Flow Regularization: Teaching LLMs to Generate Diverse Yet Coherent Responses

DGX agent

arXiv:2605.27971v1 Announce Type: cross Abstract: When large language models are fine-tuned to generate persona- or tone-conditioned responses, their output diversity is severely limited--a failure we

researcharxiv-cs-ai
28 May 2026
Research

Semantic-level Backdoor Attack against Text-to-Image Diffusion Models

DGX agent

arXiv:2602.04898v3 Announce Type: replace-cross Abstract: Text-to-image (T2I) diffusion models are widely adopted for their strong generative capabilities, yet remain vulnerable to backdoor attacks. E

researcharxiv-cs-ai
28 May 2026
Research

Semantic Optimal Transport for Sparse Autoencoder Feature Matching and Circuit Compression

DGX agent

arXiv:2605.28567v1 Announce Type: cross Abstract: Sparse autoencoders (SAEs) have become a central tool for interpreting language models. However, two key SAE analyses that remain difficult to scale a

researcharxiv-cs-ai
28 May 2026
Safety

Sense Representations Are Inducible Interfaces

DGX agent

arXiv:2605.28669v1 Announce Type: cross Abstract: Sense representations (explicit, per-token meaning decompositions) are useful for disambiguation, steering, and cross-lingual alignment, but existing

safetyarxiv-cs-ai
28 May 2026
Research

Short-Term Gain, Long-Term Fragility: AI Labor Substitution and the Erosion of Sustainable Capability

DGX agent

arXiv:2605.27399v1 Announce Type: cross Abstract: What looks like acceleration can be a quiet transfer of burden from the present to the future. Attempts to replace human labor with AI systems are oft

researcharxiv-cs-ai
28 May 2026
Applications

Show, Don't TELL: Explainable AI-Generated Text Detection

DGX agent

arXiv:2605.27921v1 Announce Type: new Abstract: Research on AI-generated text detection has presented a number of approaches to discern human from AI prose, some of which achieving high in-distributio

applicationsarxiv-cs-ai
28 May 2026
Safety

Simulation-Informed Diffusion for Decentralized Multi-robot Motion Planning

DGX agent

arXiv:2605.27697v1 Announce Type: cross Abstract: Decentralized multi-robot motion planning requires each robot to generate collision-free trajectories from local observations, without global sensing

safetyarxiv-cs-ai
28 May 2026
Research

Sinc Kolmogorov-Arnold network and its application for solving PDEs with singularities

DGX agent

arXiv:2410.04096v2 Announce Type: replace-cross Abstract: In this paper, we propose to use Sinc interpolation in the context of Kolmogorov-Arnold Networks, neural networks with learnable activation fu

researcharxiv-cs-ai
28 May 2026
Safety

Singular Vectors of Attention Heads Align with Features

DGX agent

arXiv:2602.13524v2 Announce Type: replace-cross Abstract: Identifying feature representations in language models is a central task in mechanistic interpretability. Several recent studies have made the

safetyarxiv-cs-ai
28 May 2026
Safety

Skill-Conditioned Gated Self-Distillation for LLM Reasoning

DGX agent

arXiv:2605.28791v1 Announce Type: cross Abstract: On-policy self-distillation (SD) improves LLM reasoning by using teacher-side privileged information (PI) to turn sparse verifier outcomes into dense

safetyarxiv-cs-ai
28 May 2026
Safety

SKILLC: Learning Autonomous Skill Internalization in LLM Agents via Contrastive Credit Assignment

DGX agent

arXiv:2605.27899v1 Announce Type: new Abstract: Structured skill prompts improve exploration in long-horizon agentic reinforcement learning (RL). Skill-augmented RL methods retain external skills at i

safetyarxiv-cs-ai
28 May 2026
Model Releases

SkillGrad: Optimizing Agent Skills Like Gradient Descent

DGX agent

arXiv:2605.27760v1 Announce Type: new Abstract: Agent skills provide a lightweight way to adapt LLM agents to specialized domains by storing reusable procedural knowledge in structured files. However,

model-releasesarxiv-cs-ai
28 May 2026
Safety

Smaller, Younger, and More Impactful: How AI-Assisted Writing Transforms Research Teams

DGX agent

arXiv:2605.27404v1 Announce Type: cross Abstract: The era of Big Science has long been defined by increasingly large and specialized research teams pushing the frontiers of knowledge. However, recent

safetyarxiv-cs-ai
28 May 2026
Research

SmartDirector: Keyframe-Conditioned Cinematic Video Generation with Narrative Pacing Control

DGX agent

arXiv:2605.27891v1 Announce Type: cross Abstract: The narrative quality of a video fundamentally determines its perceptual value. Although existing video generation methods can produce visually appeal

researcharxiv-cs-ai
28 May 2026
Model Releases

SmartIterator: Visual Analytics Workflows for Supervising Unsupervised Data Grouping

DGX agent

arXiv:2605.28219v1 Announce Type: cross Abstract: Unsupervised learning methods -- topic modeling, partition-based and density-based clustering -- produce data groupings without human guidance, yet ch

model-releasesarxiv-cs-ai
28 May 2026
Applications

SMILE-Next: Teaching Large Language Models to Detect, Classify, and Reason about Laughter

DGX agent

arXiv:2605.28084v1 Announce Type: cross Abstract: Laughter is a complex social signal that conveys communicative intent beyond amusement. While prior work has focused on isolated laughter analysis tas

applicationsarxiv-cs-ai
28 May 2026
Model Releases

SNARE: Adaptive Scenario Synthesis for Eliciting Overeager Behavior in Coding Agents

DGX agent

arXiv:2605.28122v1 Announce Type: cross Abstract: A coding agent executes a benign task as a sequence of shell, file, and network actions, any of which can quietly exceed the authorized scope while th

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Snippet-Driven Supply Chain Discovery with LLMs: Scaling Visibility in China

DGX agent

arXiv:2605.27845v1 Announce Type: cross Abstract: Financial and economic research often relies on structured supply-chain disclosures and commercial databases. In China, supplier--customer disclosure

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Snowveil: A Framework for Decentralised Preference Discovery

DGX agent

arXiv:2512.18444v2 Announce Type: replace-cross Abstract: Aggregating subjective preferences in social choice traditionally assumes a trusted central authority. In contrast, this paper formalises Dece

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

SONIC-O1: A Real-World Benchmark for Evaluating Multimodal Large Language Models on Audio-Video Understanding

DGX agent

arXiv:2601.21666v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) are a major focus of recent AI research. However, most prior work focuses on static image understanding, wh

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Soro: A Lightweight Foundation Model and Chatbot for Tajik

DGX agent

arXiv:2605.27379v1 Announce Type: new Abstract: We present Soro, a family of Tajik-specialized conversational large language models (LLMs) designed for real-world deployment under tight compute and co

model-releasesarxiv-cs-ai
28 May 2026
Safety

SPAR: Support-Preserving Action Rectification

DGX agent

arXiv:2605.27877v1 Announce Type: cross Abstract: Offline policy improvement faces an inherent conflict between maximizing value and fitting the data distribution. While in-sample weighted regression

safetyarxiv-cs-ai
28 May 2026
Safety

SPARD: Defending Harmful Fine-Tuning Attack via Safety Projection with Relevance-Diversity Data Selection

DGX agent

arXiv:2605.28030v1 Announce Type: cross Abstract: Fine-tuning large language models often undermines their safety alignment, a problem further amplified by harmful fine-tuning attacks in which adversa

safetyarxiv-cs-ai
28 May 2026
Research

Speaking of Language: Reflections on Metalanguage Research in NLP

DGX agent

arXiv:2604.02645v2 Announce Type: replace-cross Abstract: This work aims to shine a spotlight on the topic of metalanguage. We first define metalanguage, link it to NLP and LLMs, and then discuss our

researcharxiv-cs-ai
28 May 2026
Model Releases

SSR3D-LLM: Structured Spatial Reasoning via Latent Steps for Fine-Grained Grounding in Unified 3D-LLMs

DGX agent

arXiv:2605.28490v1 Announce Type: cross Abstract: 3D object grounding localizes referred objects in a 3D scene from natural language. Unified instance-centric 3D-LLMs aim to solve grounding together w

model-releasesarxiv-cs-ai
28 May 2026
Research

STAB: Specification-driven Testing for Algorithmic Bottlenecks

DGX agent

arXiv:2605.27981v1 Announce Type: new Abstract: Evaluating the efficiency of algorithmic code requires test cases that expose runtime bottlenecks. Previous methods generate efficiency test cases eithe

researcharxiv-cs-ai
28 May 2026
Safety

STARS: Spike Tail-Aware Relational Synthesis for ANN-to-SNN Data-Free Knowledge Distillation

DGX agent

arXiv:2605.27409v1 Announce Type: cross Abstract: SNNs promise energy-efficient and low-latency inference, but their performance still trails that of ANNs. ANN-to-SNN knowledge distillation helps narr

safetyarxiv-cs-ai
28 May 2026
Research

STFlow: Data-Coupled Flow Matching for Geometric Trajectory Simulation

DGX agent

arXiv:2505.18647v3 Announce Type: replace-cross Abstract: Simulating trajectories of dynamical systems is a fundamental problem in a wide range of fields such as molecular dynamics, biochemistry, and

researcharxiv-cs-ai
28 May 2026
Model Releases

Stochastic Gradient Descent with Momentum is Algorithmically Stable

DGX agent

arXiv:2605.28517v1 Announce Type: cross Abstract: Stochastic gradient descent with momentum (SGDM) is one of the most widely used optimization algorithms in machine learning. While optimization proper

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

StoryLens: Preference-Aligned Story Rewriting via Context-Aware Narrative Enrichment

DGX agent

arXiv:2605.28073v1 Announce Type: cross Abstract: Story rewriting aims to adapt existing narratives to diverse reader preferences while preserving plot consistency and narrative coherence. Unlike conv

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

StoryMI: Steerable Multi-Agent Therapeutic Dialogue Generation

DGX agent

arXiv:2605.27393v1 Announce Type: cross Abstract: Large language models (LLMs) can generate fluent dialogue, but prior works lack situational grounding, dynamic strategy control, and evaluation aligne

model-releasesarxiv-cs-ai
28 May 2026
Safety

Structured Agent Distillation for Large Language Model

DGX agent

arXiv:2505.13820v5 Announce Type: replace-cross Abstract: Large language models (LLMs) exhibit strong capabilities as decision-making agents by interleaving reasoning and actions, as seen in ReAct-sty

safetyarxiv-cs-ai
28 May 2026
Model Releases

Structured Belief State and the First Precision-Aware Benchmark for LLM Memory Retrieval

DGX agent

arXiv:2605.11325v2 Announce Type: replace-cross Abstract: Every major benchmark for LLM memory systems, LoCoMo foremost, measures whether a model answered correctly, not whether the memory system retr

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

SuiChat-CN: Benchmarking Contextual Suicide Risk Assessment in Chinese Group Chats

DGX agent

arXiv:2605.27911v1 Announce Type: new Abstract: Suicide is a critical global public health challenge, causing approximately 720,000 deaths each year and calling for timely, effective prevention strate

model-releasesarxiv-cs-ai
28 May 2026
Safety

Supervised Distributional Reduction via Optimal Transport and Dependence Maximization

DGX agent

arXiv:2605.27619v1 Announce Type: cross Abstract: Learning representations that capture both intrinsic data geometry and target-relevant structure remains a fundamental challenge, particularly in sett

safetyarxiv-cs-ai
28 May 2026
← Previous
1…252253254255256…448
Next →