AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
28 May 2026

RAGe: A Retrieval-Augmented Generation Evaluation Framework

ResearchDGX agent

arXiv:2605.27445v1 Announce Type: cross Abstract: Deploying Large Language Model (LLM) applications, particularly those relying on Retrieval-Augmented Generation (RAG), remains challenging due to high

RE-TRIANGLE: Does TRIANGLE Enable Multimodal Alignment Beyond Cosine Similarity in Retrieval?

SafetyDGX agent

arXiv:2605.27436v1 Announce Type: cross Abstract: Multimodal alignment is critical for bridging the semantic gap in information retrieval. However, traditional pairwise strategies introduce a geometri

Reading or Guessing? Visual Grounding Failures of Vision-Language Models for OCR in Ancient Greek Editions

Local AiDGX agent

arXiv:2605.27750v1 Announce Type: cross Abstract: Recent work has shown that Vision-Language Models (VLMs) used for optical character recognition (OCR) can generate plausible but visually unsupported


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Reasoning and Planning with Dynamically Changing Norms

AgentsDGX agent

arXiv:2605.27622v1 Announce Type: new Abstract: To safely interact with humans, AI agents must both know our norms and consider them during planning. However, such norm-guided planning has been less e

Reasoning Matters: Mitigate Hallucination in Multimodal Large Reasoning Models via Reasoning-Conditioned Preference Optimization

SafetyDGX agent

arXiv:2605.27906v1 Announce Type: new Abstract: Multimodal Large Reasoning Models introduce the reasoning paradigm, demonstrating strong capabilities on complex vision-language tasks. However, they st

REC-CBM: Rubric-Aware Error-Correction Concept Bottleneck Models for Trustworthy Open-Ended Grading

ApplicationsDGX agent

arXiv:2605.27402v1 Announce Type: cross Abstract: Open-ended grading is central to equitable and personalized education, yet manual grading remains time-consuming and costly, underscoring the need for

REED: Post-Training Representation Editing for Cross-Domain Linguistic Steganalysis

Model ReleasesDGX agent

arXiv:2605.28298v1 Announce Type: new Abstract: In real-world scenarios of linguistic steganalysis, tested texts usually come from unseen domains with different vocabularies, topics, writing styles, a

ReflexGrad: Within-Episode Failure Recovery in LLM Agents via Progress-Gated Dual-Process Routing

Model ReleasesDGX agent

arXiv:2511.14584v3 Announce Type: replace-cross Abstract: We present ReflexGrad, a dual-process architecture for within-episode failure recovery in LLM agents without demonstrations. When agents commi

Refusal Before Decoding: Detecting and Exploiting Refusal Signals in Intermediate LLM Activations

SafetyDGX agent

arXiv:2605.28553v1 Announce Type: new Abstract: In this paper, we investigate whether refusal behavior can be predicted from LLM intermediate activations before decoding using linear probes trained on

Regression Language Models for Code

Model ReleasesDGX agent

arXiv:2509.26476v2 Announce Type: replace-cross Abstract: We study code-to-metric regression: predicting numeric outcomes of code executions, a challenging task due to the open-ended nature of program

Relational Semantic Reasoning on 3D Scene Graphs for Open World Interactive Object Search

Model ReleasesDGX agent

arXiv:2603.05642v2 Announce Type: replace-cross Abstract: Open-world interactive object search in household environments requires understanding semantic relationships between objects and their surroun

RelaxFlow: Text-Driven Amodal 3D Generation

ResearchDGX agent

arXiv:2603.05425v2 Announce Type: replace-cross Abstract: Image-to-3D generation faces inherent semantic ambiguity under occlusion, where partial observation alone is often insufficient to determine o

Relevant Is Not Warranted: Evidence-Force Calibration for Cited RAG

Model ReleasesDGX agent

arXiv:2605.28044v1 Announce Type: new Abstract: Cited RAG evaluation often treats visible sources as a grounding signal, but a real, topically relevant citation can still under-warrant the attached wo

ReSAE: Residualized Sparse Autoencoders for Multi-Layer Transformer Interventions

Model ReleasesDGX agent

arXiv:2605.27819v1 Announce Type: cross Abstract: Sparse autoencoders are usually trained one layer at a time, even though transformer residual stream activations are strongly coupled across depth. Th

ResearchLoop: An Evidence-Gated Control Plane for AI-Assisted Research

ApplicationsDGX agent

arXiv:2605.28282v1 Announce Type: new Abstract: AI-assisted research compresses ideation, implementation, evaluation, and manuscript writing into a single interactive loop. This compression is useful,

Residualized Temporal Sparse Autoencoders for Interpreting Diffusion Models

ResearchDGX agent

arXiv:2605.27813v1 Announce Type: cross Abstract: Text-to-image diffusion models generate images through an iterative denoising process, so internal neural layers produce trajectories of activations r

Resource-Constrained Affect Modelling via Variance Regularisation Pruning

Model ReleasesDGX agent

arXiv:2605.27479v1 Announce Type: cross Abstract: Affective computing systems are increasingly embedded in pervasive and interactive environments, such as adaptive games, assistive technologies, and r

Restoring the Sweet Spot: Pass-Rate Weighted Self-Distillation for LLM Reasoning

SafetyDGX agent

arXiv:2605.27765v1 Announce Type: cross Abstract: Self-Distillation Policy Optimization (SDPO) provides dense token-level credit assignment for reinforcement learning with large language models by lev

Rethinking Memory as Continuously Evolving Connectivity

AgentsDGX agent

arXiv:2605.28773v1 Announce Type: cross Abstract: Existing memory-augmented LLM agents often treat memory as a static repository with pre-defined representations and fixed retrieval pipelines, which i

Revealing Algorithmic Deductive Circuits for Logical Reasoning

TutorialsDGX agent

arXiv:2605.27824v1 Announce Type: new Abstract: Recent studies have shown that Large Language Models (LLMs) can achieve strong reasoning performance by incorporating functional symbolic representation

Reverse Probing: Supervised Token-level Uncertainty Quantification for Large Language Models in Clinical Text

Local AiDGX agent

arXiv:2605.28740v1 Announce Type: cross Abstract: As large language models are increasingly deployed for clinical text, ensuring they can reliably signal their own uncertainty becomes critical. Most e

Revisiting Anthropomorphic Reflection Markers in Large Language Model Reasoning

ResearchDGX agent

arXiv:2605.28305v1 Announce Type: cross Abstract: Large Language Models (LLMs) often produce explicit reflective traces during complex reasoning, accompanied by anthropomorphic markers such as wait, h

Revisiting Change Detection Methods for their Application to Serac Fall Time-Lapse Monitoring

ApplicationsDGX agent

arXiv:2605.28100v1 Announce Type: cross Abstract: In an era where climate change aggravates environmental uncertainties, the identification and detection of event precursors are becoming crucial to mi

Revisiting Graph Autoencoders as Implicit Contrastive Learners

ResearchDGX agent

arXiv:2410.10241v2 Announce Type: replace-cross Abstract: Graph autoencoders (GAEs) and graph contrastive learning (GCL) are two major paradigms for self-supervised representation learning on graphs,

Reward Bias Substitution: Single-Axis Bias Mitigations Redirect Optimization Pressure

SafetyDGX agent

arXiv:2605.27996v1 Announce Type: new Abstract: Single-axis mitigations of reward-model biases (e.g., reducing proxy reliance on length, sycophancy, or style) can rotate optimization pressure onto cor

Risk-Controlled Lean-as-Judge for Natural-Language Mathematical Reasoning

ResearchDGX agent

arXiv:2605.28365v1 Announce Type: new Abstract: Lean is increasingly used to judge natural-language mathematical answers, but its signal is partial: many answers never formalize, and a failed proof ma

RL Squeezes, SFT Expands: A Comparative Study of Reasoning LLMs

ResearchDGX agent

arXiv:2509.21128v2 Announce Type: replace Abstract: Large language models (LLMs) are typically trained by reinforcement learning (RL) with verifiable rewards (RLVR) and supervised fine-tuning (SFT) on

Routing-Aligned Fine-Tuning for Multilingual Downstream Tasks in Mixture-of-Experts Models

SafetyDGX agent

arXiv:2605.28306v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) models have emerged as a dominant paradigm for efficient LLM scaling, yet adapting them to non-English downstream tasks remai

ROVER: Routing Object-Centric Visual Evidence for Grounded Multi-Image Reasoning

Local AiDGX agent

arXiv:2605.27959v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have increasingly localized and interleaved visual evidence for deliberative reasoning. Grounding-based appro

RULER: Representation-Level Verification of Machine Unlearning

ResearchDGX agent

arXiv:2605.27569v1 Announce Type: new Abstract: Machine unlearning aims to remove the influence of specific training records from a deployed model without retraining from scratch. Current protocols ve

SafeMed-R1: Clinician-Audited Safety and Ethics Alignment for Medical Large Language Models

SafetyDGX agent

arXiv:2605.28338v1 Announce Type: new Abstract: Large language models(LLMs) increasingly match expert performance on licensing examinations, yet routine clinical use remains limited because governance

SAME: Stabilized Mixture-of-Experts for Multimodal Continual Instruction Tuning

Model ReleasesDGX agent

arXiv:2602.01990v2 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) achieve strong performance through instruction tuning, but real-world deployment requires them to con

SARAD: LLM-Based Safety-Aware Hybrid Reinforcement Learning with Collision Prediction for Autonomous Driving

SafetyDGX agent

arXiv:2605.28583v1 Announce Type: cross Abstract: Ensuring both safety and efficiency in decision-making for autonomous driving systems remains a fundamental challenge. Traditional Deep Reinforcement

Satisfiability Solving with LLMs: A Matched-Pair Evaluation of Reasoning Capability

ResearchDGX agent

arXiv:2605.28602v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for tasks that implicitly reduce to Boolean satisfiability (SAT), yet their reasoning ability on SAT

Score Based Error Correcting Code Decoder

ResearchDGX agent

arXiv:2605.28358v1 Announce Type: cross Abstract: Error-correcting codes enable reliable communication, yet practical soft decoding remains challenging across code families and block lengths. We propo

Securing Retrieval-Augmented Generation: A Taxonomy of Attacks, Defenses, and Future Directions

AgentsDGX agent

arXiv:2604.08304v2 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) extends large language models (LLMs) with external knowledge, but this access path also introduces securi

SelfJudge: Faster Speculative Decoding via Self-Supervised Judge Verification

ResearchDGX agent

arXiv:2510.02329v2 Announce Type: replace-cross Abstract: Speculative decoding accelerates LLM inference by verifying candidate tokens from a draft model against a larger target model. Recent judge de

Semantic Flow Regularization: Teaching LLMs to Generate Diverse Yet Coherent Responses

ResearchDGX agent

arXiv:2605.27971v1 Announce Type: cross Abstract: When large language models are fine-tuned to generate persona- or tone-conditioned responses, their output diversity is severely limited--a failure we

Semantic-level Backdoor Attack against Text-to-Image Diffusion Models

ResearchDGX agent

arXiv:2602.04898v3 Announce Type: replace-cross Abstract: Text-to-image (T2I) diffusion models are widely adopted for their strong generative capabilities, yet remain vulnerable to backdoor attacks. E

Semantic Optimal Transport for Sparse Autoencoder Feature Matching and Circuit Compression

ResearchDGX agent

arXiv:2605.28567v1 Announce Type: cross Abstract: Sparse autoencoders (SAEs) have become a central tool for interpreting language models. However, two key SAE analyses that remain difficult to scale a

Sense Representations Are Inducible Interfaces

SafetyDGX agent

arXiv:2605.28669v1 Announce Type: cross Abstract: Sense representations (explicit, per-token meaning decompositions) are useful for disambiguation, steering, and cross-lingual alignment, but existing

Short-Term Gain, Long-Term Fragility: AI Labor Substitution and the Erosion of Sustainable Capability

ResearchDGX agent

arXiv:2605.27399v1 Announce Type: cross Abstract: What looks like acceleration can be a quiet transfer of burden from the present to the future. Attempts to replace human labor with AI systems are oft

Show, Don't TELL: Explainable AI-Generated Text Detection

ApplicationsDGX agent

arXiv:2605.27921v1 Announce Type: new Abstract: Research on AI-generated text detection has presented a number of approaches to discern human from AI prose, some of which achieving high in-distributio

Simulation-Informed Diffusion for Decentralized Multi-robot Motion Planning

SafetyDGX agent

arXiv:2605.27697v1 Announce Type: cross Abstract: Decentralized multi-robot motion planning requires each robot to generate collision-free trajectories from local observations, without global sensing

Sinc Kolmogorov-Arnold network and its application for solving PDEs with singularities

ResearchDGX agent

arXiv:2410.04096v2 Announce Type: replace-cross Abstract: In this paper, we propose to use Sinc interpolation in the context of Kolmogorov-Arnold Networks, neural networks with learnable activation fu

Singular Vectors of Attention Heads Align with Features

SafetyDGX agent

arXiv:2602.13524v2 Announce Type: replace-cross Abstract: Identifying feature representations in language models is a central task in mechanistic interpretability. Several recent studies have made the

Skill-Conditioned Gated Self-Distillation for LLM Reasoning

SafetyDGX agent

arXiv:2605.28791v1 Announce Type: cross Abstract: On-policy self-distillation (SD) improves LLM reasoning by using teacher-side privileged information (PI) to turn sparse verifier outcomes into dense

SKILLC: Learning Autonomous Skill Internalization in LLM Agents via Contrastive Credit Assignment

SafetyDGX agent

arXiv:2605.27899v1 Announce Type: new Abstract: Structured skill prompts improve exploration in long-horizon agentic reinforcement learning (RL). Skill-augmented RL methods retain external skills at i

SkillGrad: Optimizing Agent Skills Like Gradient Descent

Model ReleasesDGX agent

arXiv:2605.27760v1 Announce Type: new Abstract: Agent skills provide a lightweight way to adapt LLM agents to specialized domains by storing reusable procedural knowledge in structured files. However,

Smaller, Younger, and More Impactful: How AI-Assisted Writing Transforms Research Teams

SafetyDGX agent

arXiv:2605.27404v1 Announce Type: cross Abstract: The era of Big Science has long been defined by increasingly large and specialized research teams pushing the frontiers of knowledge. However, recent

SmartDirector: Keyframe-Conditioned Cinematic Video Generation with Narrative Pacing Control

ResearchDGX agent

arXiv:2605.27891v1 Announce Type: cross Abstract: The narrative quality of a video fundamentally determines its perceptual value. Although existing video generation methods can produce visually appeal

SmartIterator: Visual Analytics Workflows for Supervising Unsupervised Data Grouping

Model ReleasesDGX agent

arXiv:2605.28219v1 Announce Type: cross Abstract: Unsupervised learning methods -- topic modeling, partition-based and density-based clustering -- produce data groupings without human guidance, yet ch

SMILE-Next: Teaching Large Language Models to Detect, Classify, and Reason about Laughter

ApplicationsDGX agent

arXiv:2605.28084v1 Announce Type: cross Abstract: Laughter is a complex social signal that conveys communicative intent beyond amusement. While prior work has focused on isolated laughter analysis tas

SNARE: Adaptive Scenario Synthesis for Eliciting Overeager Behavior in Coding Agents

Model ReleasesDGX agent

arXiv:2605.28122v1 Announce Type: cross Abstract: A coding agent executes a benign task as a sequence of shell, file, and network actions, any of which can quietly exceed the authorized scope while th

Snippet-Driven Supply Chain Discovery with LLMs: Scaling Visibility in China

Model ReleasesDGX agent

arXiv:2605.27845v1 Announce Type: cross Abstract: Financial and economic research often relies on structured supply-chain disclosures and commercial databases. In China, supplier--customer disclosure

Snowveil: A Framework for Decentralised Preference Discovery

Model ReleasesDGX agent

arXiv:2512.18444v2 Announce Type: replace-cross Abstract: Aggregating subjective preferences in social choice traditionally assumes a trusted central authority. In contrast, this paper formalises Dece

SONIC-O1: A Real-World Benchmark for Evaluating Multimodal Large Language Models on Audio-Video Understanding

Model ReleasesDGX agent

arXiv:2601.21666v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) are a major focus of recent AI research. However, most prior work focuses on static image understanding, wh

Soro: A Lightweight Foundation Model and Chatbot for Tajik

Model ReleasesDGX agent

arXiv:2605.27379v1 Announce Type: new Abstract: We present Soro, a family of Tajik-specialized conversational large language models (LLMs) designed for real-world deployment under tight compute and co

SPAR: Support-Preserving Action Rectification

SafetyDGX agent

arXiv:2605.27877v1 Announce Type: cross Abstract: Offline policy improvement faces an inherent conflict between maximizing value and fitting the data distribution. While in-sample weighted regression

SPARD: Defending Harmful Fine-Tuning Attack via Safety Projection with Relevance-Diversity Data Selection

SafetyDGX agent

arXiv:2605.28030v1 Announce Type: cross Abstract: Fine-tuning large language models often undermines their safety alignment, a problem further amplified by harmful fine-tuning attacks in which adversa

← Previous
1…201202203204205…358
Next →