AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
Model Releases

STAPO: Stabilizing Reinforcement Learning for LLMs by Silencing Rare Spurious Tokens

DGX agent

arXiv:2602.15620v4 Announce Type: replace Abstract: Reinforcement Learning (RL) has significantly improved large language model reasoning, but existing RL fine-tuning methods rely heavily on heuristic

model-releasesarxiv-cs-cl
13 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Steering Without Breaking: Mechanistically Informed Interventions for Discrete Diffusion Language Models

DGX agent

arXiv:2605.10971v1 Announce Type: cross Abstract: Discrete diffusion language models (DLMs) generate text by iteratively denoising all positions in parallel, offering an alternative to autoregressive

researcharxiv-cs-cl
13 May 2026
Model Releases

StepCodeReasoner: Aligning Code Reasoning with Stepwise Execution Traces via Reinforcement Learning

DGX agent

arXiv:2605.11922v1 Announce Type: cross Abstract: Existing code reasoning methods primarily supervise final code outputs, ignoring intermediate states, often leading to reward hacking where correct an

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

StoicLLM: Preference Optimization for Philosophical Alignment in Small Language Models

DGX agent

arXiv:2605.11483v1 Announce Type: new Abstract: While large language models excel at factual adaptation, their ability to internalize nuanced philosophical frameworks under severe data constraints rem

model-releasesarxiv-cs-cl
13 May 2026
Research

Stopping Computation for Converged Tokens in Masked Diffusion-LM Decoding

DGX agent

arXiv:2602.06412v3 Announce Type: replace Abstract: Masked Diffusion Language Models generate sequences via iterative sampling that progressively unmasks tokens. However, they still recompute the atte

researcharxiv-cs-cl
13 May 2026
Research

Stories in Space: In-Context Learning Trajectories in Conceptual Belief Space

DGX agent

arXiv:2605.12412v1 Announce Type: new Abstract: Large Language Models (LLMs) update their behavior in context, which can be viewed as a form of Bayesian inference. However, the structure of the latent

researcharxiv-cs-cl
13 May 2026
Tutorials

Synthetic Function Demonstrations Improve Generation in Low-Resource Programming Languages

DGX agent

arXiv:2503.18760v2 Announce Type: replace Abstract: A key consideration when training an LLM is whether the target language is more or less resourced, for example English compared to Welsh, or Python

tutorialsarxiv-cs-cl
13 May 2026
Applications

TabDLM: Free-Form Tabular Data Generation via Joint Numerical-Language Diffusion

DGX agent

arXiv:2602.22586v2 Announce Type: replace-cross Abstract: Synthetic tabular data generation has attracted growing attention due to its importance for data augmentation, foundation models, and privacy.

applicationsarxiv-cs-cl
13 May 2026
Safety

Taming Extreme Tokens: Covariance-Aware GRPO with Gaussian-Kernel Advantage Reweighting

DGX agent

arXiv:2605.11538v1 Announce Type: new Abstract: Group Relative Policy Optimization (GRPO) has emerged as a promising approach for improving the reasoning capabilities of large language models. However

safetyarxiv-cs-cl
13 May 2026
Research

Task-Adaptive Embedding Refinement via Test-time LLM Guidance

DGX agent

arXiv:2605.12487v1 Announce Type: new Abstract: We explore the effectiveness of an LLM-guided query refinement paradigm for extending the usability of embedding models to challenging zero-shot search

researcharxiv-cs-cl
13 May 2026
Model Releases

Test-Time Compute for Dense Retrieval: Agentic Program Generation with Frozen Embedding Models

DGX agent

arXiv:2605.11374v1 Announce Type: cross Abstract: Test-time compute is widely believed to benefit only large reasoning models. We show it also helps small embedding models. Most modern embedding check

model-releasesarxiv-cs-cl
13 May 2026
Local Ai

TextSeal: A Localized LLM Watermark for Provenance & Distillation Protection

DGX agent

arXiv:2605.12456v1 Announce Type: cross Abstract: We introduce TextSeal, a state-of-the-art watermark for large language models. Building on Gumbel-max sampling, TextSeal introduces dual-key generatio

local-aiarxiv-cs-cl
13 May 2026
Research

The Algorithmic Caricature: Auditing LLM-Generated Political Discourse Across Crisis Events

DGX agent

arXiv:2605.12452v1 Announce Type: new Abstract: Large Language Models (LLMs) can generate fluent political text at scale, raising concerns about synthetic discourse during crises and social conflict.

researcharxiv-cs-cl
13 May 2026
Research

The Bicameral Model: Bidirectional Hidden-State Coupling Between Parallel Language Models

DGX agent

arXiv:2605.11167v1 Announce Type: new Abstract: Existing multi-model and tool-augmented systems communicate by generating text, serializing every exchange through the output vocabulary. Can two pretra

researcharxiv-cs-cl
13 May 2026
Research

The Challenge and Reward of Fair Play in Narrative: A Computational Approach

DGX agent

arXiv:2507.13841v2 Announce Type: replace Abstract: Good storytelling involves surprise -- unpredictability in how the story unfolds -- and sense-making, the requirement that the story forms a coheren

researcharxiv-cs-cl
13 May 2026
Model Releases

Three Regimes of Context-Parametric Conflict: A Predictive Framework and Empirical Validation

DGX agent

arXiv:2605.11574v1 Announce Type: new Abstract: The literature on how large language models handle conflict between their training knowledge and a contradicting document presents a persistent empirica

model-releasesarxiv-cs-cl
13 May 2026
Hardware

To Err Is Human; To Annotate, SILICON? Toward Robust Reproducibility in LLM Annotation

DGX agent

arXiv:2412.14461v4 Announce Type: replace Abstract: Unstructured text data annotation is foundational to management research. LLMs offer a cost-effective and scalable alternative to human annotation,

hardwarearxiv-cs-cl
13 May 2026
Safety

TokenRatio: Principled Token-Level Preference Optimization via Ratio Matching

DGX agent

arXiv:2605.12288v1 Announce Type: new Abstract: Direct Preference Optimization (DPO) is a widely used RL-free method for aligning language models from pairwise preferences, but it models preferences o

safetyarxiv-cs-cl
13 May 2026
Safety

Towards Fine-Grained Code-Switch Speech Translation with Semantic Space Alignment

DGX agent

arXiv:2511.10670v2 Announce Type: replace Abstract: Code-switching (CS) speech translation (ST) aims to translate speech that alternates between multiple languages into a target language text, posing

safetyarxiv-cs-cl
13 May 2026
Applications

Towards Visually-Guided Movie Subtitle Translation for Indic Languages

DGX agent

arXiv:2605.11993v1 Announce Type: new Abstract: Movie subtitle translation is inherently multimodal, yet text-only systems often miss visual cues needed to convey emotion, action, and social nuance, e

applicationsarxiv-cs-cl
13 May 2026
Safety

Toxicity Detection Should Measure Contextual Harm, Not Text-Intrinsic Badness

DGX agent

arXiv:2503.16072v4 Announce Type: replace-cross Abstract: Toxicity detection has become core safety infrastructure for online moderation, dataset filtering, and deployed language-model systems. Yet mo

safetyarxiv-cs-cl
13 May 2026
Research

Training-Inference Consistent Segmented Execution for Long-Context LLMs

DGX agent

arXiv:2605.11744v1 Announce Type: new Abstract: Transformer-based large language models face severe scalability challenges in long-context generation due to the computational and memory costs of full-

researcharxiv-cs-cl
13 May 2026
Safety

Understanding the Performance Gap in Preference Learning: A Dichotomy of RLHF and DPO

DGX agent

arXiv:2505.19770v5 Announce Type: replace-cross Abstract: We present a fine-grained theoretical analysis of the performance gap between two-stage reinforcement learning from human feedback~(RLHF) and

safetyarxiv-cs-cl
13 May 2026
Applications

UniVLR: Unifying Text and Vision in Visual Latent Reasoning for Multimodal LLMs

DGX agent

arXiv:2605.11856v1 Announce Type: cross Abstract: Multimodal large language models are increasingly expected to perform thinking with images, yet existing visual latent reasoning methods still rely on

applicationsarxiv-cs-cl
13 May 2026
Agents

Unlocking LLM Creativity in Science through Analogical Reasoning

DGX agent

arXiv:2605.11258v1 Announce Type: cross Abstract: Autonomous science promises to augment scientific discovery, particularly in complex fields like biomedicine. However, this requires AI systems that c

agentsarxiv-cs-cl
13 May 2026
Model Releases

VERDI: Single-Call Confidence Estimation for Verification-Based LLM Judges via Decomposed Inference

DGX agent

arXiv:2605.11334v1 Announce Type: cross Abstract: LLM-as-Judge systems are widely deployed for automated evaluation, yet practitioners lack reliable methods to know when a judge's verdict should be tr

model-releasesarxiv-cs-cl
13 May 2026
Research

Video-Language Understanding: A Survey from Model Architecture, Model Training, and Data Perspectives

DGX agent

arXiv:2406.05615v4 Announce Type: replace Abstract: Humans use multiple senses to comprehend the environment. Vision and language are two of the most vital senses since they allow us to easily communi

researcharxiv-cs-cl
13 May 2026
Tutorials

What makes a word hard to learn? Modeling L1 influence on English vocabulary difficulty

DGX agent

arXiv:2605.12281v1 Announce Type: new Abstract: What makes a word difficult to learn, and how does the difficulty depend on the learner's native language? We computationally model vocabulary difficult

tutorialsarxiv-cs-cl
13 May 2026
Research

When Emotion Becomes Trigger: Emotion-style dynamic Backdoor Attack Parasitising Large Language Models

DGX agent

arXiv:2605.11612v1 Announce Type: new Abstract: Backdoor vulnerabilities widely exist in the fine-tuning of large language models(LLMs). Most backdoor poisoning methods operate mainly at the token lev

researcharxiv-cs-cl
13 May 2026
Research

When the Gold Standard Isn't Necessarily Standard: Challenges of Evaluating the Translation of User-Generated Content

DGX agent

arXiv:2512.17738v2 Announce Type: replace Abstract: User-generated content (UGC) is characterised by frequent use of non-standard language, from spelling errors to expressive choices such as slang, ch

researcharxiv-cs-cl
13 May 2026
Safety

World Action Models: The Next Frontier in Embodied AI

DGX agent

arXiv:2605.12090v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have achieved strong semantic generalization for embodied policy learning, yet they learn reactive observation-to-

safetyarxiv-cs-cl
13 May 2026
Model Releases

YFPO: A Preliminary Study of Yoked Feature Preference Optimization with Neuron-Guided Rewards for Mathematical Reasoning

DGX agent

arXiv:2605.11906v1 Announce Type: new Abstract: Preference optimization has become an important post-training paradigm for improving the reasoning abilities of large language models. Existing methods

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

100,000+ Movie Reviews from Kazakhstan: Russian, Kazakh, and Code-Switched Texts

DGX agent

arXiv:2605.08600v1 Announce Type: new Abstract: We present a new publicly available corpus of 100,502 movie reviews from Kazakhstan collected from kino.kz, spanning 2001-2025 and covering 4,943 unique

model-releasesarxiv-cs-cl
12 May 2026
Research

A Computational Operationalisation of Competing Maturational Theories of Syntactic Development via Statistical Grammar Induction

DGX agent

arXiv:2605.08476v1 Announce Type: new Abstract: This paper is concerned with what intermediate syntactic categories children acquire during first language development, and in what order. Maturational

researcharxiv-cs-cl
12 May 2026
Research

A Single-Layer Model Can Do Language Modeling

DGX agent

arXiv:2605.10643v1 Announce Type: new Abstract: Modern language models scale depth by stacking layers, each holding its own state - a per-layer KV cache in transformers, a per-layer matrix in Mamba, G

researcharxiv-cs-cl
12 May 2026
Research

A Single Layer to Explain Them All:Understanding Massive Activations in Large Language Models

DGX agent

arXiv:2605.08504v1 Announce Type: new Abstract: We investigate the origins of massive activations in large language models (LLMs) and identify a specific layer named the extbf{Massive Emergence Layer

researcharxiv-cs-cl
12 May 2026
Hardware

AAAC: Activation-Aware Adaptive Codebooks for 4-bit LLM Weight Quantization

DGX agent

arXiv:2605.08692v1 Announce Type: cross Abstract: Post-training weight-only quantization to 4 bits is widely used to reduce the memory and compute costs of large language model inference. Existing PTQ

hardwarearxiv-cs-cl
12 May 2026
Local Ai

AdaSwitch: Adaptive Switching between Small and Large Agents for Effective Cloud-Local Collaborative Learning

DGX agent

arXiv:2410.13181v2 Announce Type: replace Abstract: Recent advancements in large language models (LLMs) have been remarkable. Users face a choice between using cloud-based LLMs for generation quality

local-aiarxiv-cs-cl
12 May 2026
Safety

AgentReview: Exploring Peer Review Dynamics with LLM Agents

DGX agent

arXiv:2406.12708v3 Announce Type: replace Abstract: Peer review is fundamental to the integrity and advancement of scientific publication. Traditional methods of peer review analyses often rely on exp

safetyarxiv-cs-cl
12 May 2026
Safety

Aligning LLM Uncertainty with Human Disagreement in Subjectivity Analysis

DGX agent

arXiv:2605.10415v1 Announce Type: new Abstract: Large language models for subjectivity analysis are typically trained with aggregated labels, which compress variations in human judgment into a single

safetyarxiv-cs-cl
12 May 2026
Model Releases

An Annotation Scheme and Classifier for Personal Facts in Dialogue

DGX agent

arXiv:2605.10339v1 Announce Type: new Abstract: The advancement of Large Language Models (LLMs) has enabled their application in personalized dialogue systems. We present an extended annotation scheme

model-releasesarxiv-cs-cl
12 May 2026
Research

ANCHOR: Abductive Network Construction with Hierarchical Orchestration for Reliable Probability Inference in Large Language Models

DGX agent

arXiv:2605.10328v1 Announce Type: new Abstract: A central challenge in large-scale decision-making under incomplete information is estimating reliable probabilities. Recent approaches leverage Large L

researcharxiv-cs-cl
12 May 2026
Tutorials

Annotations Mitigate Post-Training Mode Collapse

DGX agent

arXiv:2605.09995v1 Announce Type: new Abstract: Post-training (via supervised fine-tuning) improves instruction-following, but often induces semantic mode collapse by biasing models toward low-entropy

tutorialsarxiv-cs-cl
12 May 2026
Model Releases

Architecture, Not Scale: Circuit Localization in Large Language Models

DGX agent

arXiv:2605.08853v1 Announce Type: new Abstract: Mechanistic interpretability assumes that circuit analysis becomes harder as models scale. We challenge this assumption by showing that the attention ar

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

ASTRA-QA: A Benchmark for Abstract Question Answering over Documents

DGX agent

arXiv:2605.10168v1 Announce Type: new Abstract: Document-based question answering (QA) increasingly includes abstract questions that require synthesizing scattered information from long documents or a

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

Attention Grounded Enhancement for Visual Document Retrieval

DGX agent

arXiv:2511.13415v2 Announce Type: replace-cross Abstract: Visual document retrieval requires understanding heterogeneous and multi-modal content to satisfy implicit information needs. Recent advances

model-releasesarxiv-cs-cl
12 May 2026
Agents

Autonomous Continual Learning for Environment Adaptation of Computer-Use Agents

DGX agent

arXiv:2602.10356v2 Announce Type: replace Abstract: Real-world digital environments are highly diverse and dynamic. These characteristics cause agents to frequently encounter unseen environments and d

agentsarxiv-cs-cl
12 May 2026
Model Releases

BabelDOC: Better Layout-Preserving PDF Translation via Intermediate Representation

DGX agent

arXiv:2605.10845v1 Announce Type: cross Abstract: As global cross-lingual communication intensifies, language barriers in visually rich documents such as PDFs remain a practical bottleneck. Existing d

model-releasesarxiv-cs-cl
12 May 2026
← Previous
1…96979899100…161
Next →