AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
Safety

Your Students Don't Use LLMs Like You Wish They Did

DGX agent

arXiv:2604.23486v1 Announce Type: new Abstract: Educational NLP systems are typically evaluated using engagement metrics and satisfaction surveys, which are at best a proxy for meeting pedagogical goa

safetyarxiv-cs-cl
28 Apr 2026
Research

Zero-shot Large Language Models for Automatic Readability Assessment

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2604.24470v1 Announce Type: new Abstract: Unsupervised automatic readability assessment (ARA) methods have important practical and research applications (e.g., ensuring medical or educational ma

researcharxiv-cs-cl
28 Apr 2026
Research

Aggregate vs. Personalized Judges in Business Idea Evaluation: Evidence from Expert Disagreement

DGX agent

arXiv:2604.22517v1 Announce Type: new Abstract: Evaluating LLM-generated business ideas is often harder to scale than generating them. Unlike standard NLP benchmarks, business idea evaluation relies o

researcharxiv-cs-cl
27 Apr 2026
Research

An End-to-End Ukrainian RAG for Local Deployment. Optimized Hybrid Search and Lightweight Generation

DGX agent

arXiv:2604.22095v1 Announce Type: new Abstract: This paper presents a highly efficient Retrieval-Augmented Generation (RAG) system built specifically for Ukrainian document question answering, which a

researcharxiv-cs-cl
27 Apr 2026
Agents

Behavioral Canaries: Auditing Private Retrieved Context Usage in RL Fine-Tuning

DGX agent

arXiv:2604.22191v1 Announce Type: cross Abstract: In agentic workflows, LLMs frequently process retrieved contexts that are legally protected from further training. However, auditors currently lack a

agentsarxiv-cs-cl
27 Apr 2026
Research

BERAG: Bayesian Ensemble Retrieval-Augmented Generation for Knowledge-based Visual Question Answering

DGX agent

arXiv:2604.22678v1 Announce Type: new Abstract: A common approach to question answering with retrieval-augmented generation (RAG) is to concatenate documents into a single context and pass it to a lan

researcharxiv-cs-cl
27 Apr 2026
Model Releases

Bridging the Long-Tail Gap: Robust Retrieval-Augmented Relation Completion via Multi-Stage Paraphrase Infusion

DGX agent

arXiv:2604.22261v1 Announce Type: new Abstract: Large language models (LLMs) struggle with relation completion (RC), both with and without retrieval-augmented generation (RAG), particularly when the r

model-releasesarxiv-cs-cl
27 Apr 2026
Research

Can QPP Choose the Right Query Variant? Evaluating Query Variant Selection for RAG Pipelines

DGX agent

arXiv:2604.22661v1 Announce Type: cross Abstract: Large Language Models (LLMs) have made query reformulation ubiquitous in modern retrieval and Retrieval-Augmented Generation (RAG) pipelines, enabling

researcharxiv-cs-cl
27 Apr 2026
Model Releases

CLARITY: A Framework and Benchmark for Conversational Language Ambiguity and Unanswerability in Interactive NL2SQL Systems

DGX agent

arXiv:2604.22313v1 Announce Type: new Abstract: NL2SQL systems deployed in industry settings often encounter ambiguous or unanswerable queries, particularly in interactive scenarios with incomplete us

model-releasesarxiv-cs-cl
27 Apr 2026
Safety

Context-Fidelity Boosting: Enhancing Faithful Generation through Watermark-Inspired Decoding

DGX agent

arXiv:2604.22335v1 Announce Type: new Abstract: Large language models (LLMs) often produce content that contradicts or overlooks information provided in the input context, a phenomenon known as faithf

safetyarxiv-cs-cl
27 Apr 2026
Model Releases

Dharma, Data and Deception: An LLM-Powered Rhetorical Analysis of Cow-Urine Health Claims on YouTube

DGX agent

arXiv:2604.22606v1 Announce Type: new Abstract: Health misinformation remains one of the most pressing challenges on social media, particularly when cultural traditions intersect with scientific-sound

model-releasesarxiv-cs-cl
27 Apr 2026
Model Releases

DimABSA: Building Multilingual and Multidomain Datasets for Dimensional Aspect-Based Sentiment Analysis

DGX agent

arXiv:2601.23022v3 Announce Type: replace Abstract: Aspect-Based Sentiment Analysis (ABSA) focuses on extracting sentiment at a fine-grained aspect level and has been widely applied across real-world

model-releasesarxiv-cs-cl
27 Apr 2026
Research

Dissociating Decodability and Causal Use in Bracket-Sequence Transformers

DGX agent

arXiv:2604.22128v1 Announce Type: new Abstract: When trained on tasks requiring an understanding of hierarchical structure, transformers have been found to represent this hierarchy in distinct ways: i

researcharxiv-cs-cl
27 Apr 2026
Applications

Dynamically Acquiring Text Content to Enable the Classification of Lesser-known Entities for Real-world Tasks

DGX agent

arXiv:2604.22325v1 Announce Type: new Abstract: Existing Natural Language Processing (NLP) resources often lack the task-specific information required for real-world problems and provide limited cover

applicationsarxiv-cs-cl
27 Apr 2026
Agents

Efficient Multi-Agent System Training with Data Influence-Oriented Tree Search

DGX agent

arXiv:2502.00955v2 Announce Type: replace Abstract: Monte Carlo Tree Search (MCTS) based methods provide promising approaches for generating synthetic data to enhance the self-training of Large Langua

agentsarxiv-cs-cl
27 Apr 2026
Local Ai

Fine-Grained Analysis of Shared Syntactic Mechanisms in Language Models

DGX agent

arXiv:2604.22166v1 Announce Type: new Abstract: While language models demonstrate sophisticated syntactic capabilities, the extent to which their internal mechanisms align with cross-constructional pr

local-aiarxiv-cs-cl
27 Apr 2026
Research

From graphemic dependence to lexical structure: a Markovian perspective on Dante's Commedia

DGX agent

arXiv:2604.22626v1 Announce Type: new Abstract: This study investigates the structural organisation of Dante's Divina Commedia through a symbolic representation based on vowel-consonant (V/C) encoding

researcharxiv-cs-cl
27 Apr 2026
Model Releases

From Interpretability to Performance: Optimizing Retrieval Heads for Long-Context Language Models

DGX agent

arXiv:2601.11020v3 Announce Type: replace Abstract: Advances in mechanistic interpretability have identified special attention heads, known as retrieval heads, that are responsible for retrieving info

model-releasesarxiv-cs-cl
27 Apr 2026
Agents

HACHIMI: Scalable and Controllable Student Persona Generation via Orchestrated Agents

DGX agent

arXiv:2603.04855v3 Announce Type: replace Abstract: Student Personas (SPs) are emerging as infrastructure for educational LLMs, yet prior work often relies on ad-hoc prompting or hand-crafted profiles

agentsarxiv-cs-cl
27 Apr 2026
Model Releases

How Do AI Agents Spend Your Money? Analyzing and Predicting Token Consumption in Agentic Coding Tasks

DGX agent

arXiv:2604.22750v1 Announce Type: new Abstract: The wide adoption of AI agents in complex human workflows is driving rapid growth in LLM token consumption. When agents are deployed on tasks that requi

model-releasesarxiv-cs-cl
27 Apr 2026
Safety

How Large Language Models Balance Internal Knowledge with User and Document Assertions

DGX agent

arXiv:2604.22193v1 Announce Type: new Abstract: Large language models (LLMs) often need to balance their internal parametric knowledge with external information, such as user beliefs and content from

safetyarxiv-cs-cl
27 Apr 2026
Safety

Identifying and typifying demographic unfairness in phoneme-level embeddings of self-supervised speech recognition models

DGX agent

arXiv:2604.22631v1 Announce Type: new Abstract: Modern automatic speech recognition (ASR) systems have been observed to function better for certain speaker groups (SGs) than others, despite recent gai

safetyarxiv-cs-cl
27 Apr 2026
Research

Identifying the Periodicity of Information in Natural Language

DGX agent

arXiv:2510.27241v2 Announce Type: replace Abstract: Recent theoretical advancement of information density in natural language has brought the following question on desk: To what degree does natural la

researcharxiv-cs-cl
27 Apr 2026
Hardware

Incentivizing Neuro-symbolic Language-based Reasoning in VLMs via Reinforcement Learning

DGX agent

arXiv:2604.22062v1 Announce Type: new Abstract: There are 7,407 languages in the world. But, what about the languages that are not there in the world? Are humans so narrow minded that we don't care ab

hardwarearxiv-cs-cl
27 Apr 2026
Model Releases

Intrinsic Fingerprint of LLMs: Continue Training is NOT All You Need to Steal A Model!

DGX agent

arXiv:2507.03014v2 Announce Type: replace-cross Abstract: Large language models (LLMs) face significant copyright and intellectual property challenges as the cost of training increases and model reuse

model-releasesarxiv-cs-cl
27 Apr 2026
Applications

Knowledge-driven Augmentation and Retrieval for Integrative Temporal Adaptation

DGX agent

arXiv:2604.22098v1 Announce Type: new Abstract: Time introduces fundamental challenges in model development and deployment: models are usually trained on historical data while deployed on future data

applicationsarxiv-cs-cl
27 Apr 2026
Model Releases

Language Specific Knowledge: Do Models Know Better in X than in English?

DGX agent

arXiv:2505.14990v3 Announce Type: replace Abstract: Often, multilingual language models are trained with the objective to map semantically similar content (in different languages) in the same latent s

model-releasesarxiv-cs-cl
27 Apr 2026
Research

Large Language Models Decide Early and Explain Later

DGX agent

arXiv:2604.22266v1 Announce Type: new Abstract: Large Language Models often achieve strong performance by generating long intermediate chain-of-thought reasoning. However, it remains unclear when a mo

researcharxiv-cs-cl
27 Apr 2026
Research

LATMiX: Learnable Affine Transformations for Microscaling Quantization of LLMs

DGX agent

arXiv:2602.17681v2 Announce Type: replace-cross Abstract: Post-training quantization (PTQ) is a widely used approach for reducing the memory and compute costs of large language models (LLMs). Recent s

researcharxiv-cs-cl
27 Apr 2026
Research

LayerBoost: Layer-Aware Attention Reduction for Efficient LLMs

DGX agent

arXiv:2604.22050v1 Announce Type: cross Abstract: Transformers are mostly relying on softmax attention, which introduces quadratic complexity with respect to sequence length and remains a major bottle

researcharxiv-cs-cl
27 Apr 2026
Model Releases

LLMs as Assessors: Right for the Right Reason?

DGX agent

arXiv:2601.08919v2 Announce Type: replace-cross Abstract: A good deal of recent research has focused on how Large Language Models (LLMs) may be used as judges in place of humans to evaluate the qualit

model-releasesarxiv-cs-cl
27 Apr 2026
Research

Measuring and Mitigating Persona Distortions from AI Writing Assistance

DGX agent

arXiv:2604.22503v1 Announce Type: new Abstract: Hundreds of millions of people use artificial intelligence (AI) for writing assistance. Here, we evaluated how AI writing assistance distorts writer per

researcharxiv-cs-cl
27 Apr 2026
Research

Multi-Token Prediction via Self-Distillation

DGX agent

arXiv:2602.06019v2 Announce Type: replace Abstract: Existing techniques for accelerating language model inference, such as speculative decoding, require training auxiliary speculator models and buildi

researcharxiv-cs-cl
27 Apr 2026
Research

MultiTok: Variable-Length Tokenization for Efficient LLMs Adapted from LZW Compression

DGX agent

arXiv:2410.21548v3 Announce Type: replace Abstract: Large language models have drastically changed the prospects of AI by introducing technologies for more complex natural language processing. However

researcharxiv-cs-cl
27 Apr 2026
Research

Neural Recovery of Historical Lexical Structure in Bantu Languages from Modern Data

DGX agent

arXiv:2604.22730v1 Announce Type: cross Abstract: We investigate whether neural models trained exclusively on modern morphological data can recover cross-lingual lexical structure consistent with hist

researcharxiv-cs-cl
27 Apr 2026
Research

NeuronMLP: Efficient LLM Inference via Singular Value Decomposition Compression and Tiling on AWS Trainium

DGX agent

arXiv:2510.25977v4 Announce Type: replace Abstract: Emerging AI accelerators have started to gain attention and offer new opportunities for efficient inference of large language models (LLMs). Trainiu

researcharxiv-cs-cl
27 Apr 2026
Research

NiuTrans.LMT: Toward Inclusive and Scalable Multilingual Machine Translation with LLMs

DGX agent

arXiv:2511.07003v2 Announce Type: replace Abstract: Large language models have significantly advanced Multilingual Machine Translation (MMT), yet scaling to many languages while keeping quality robust

researcharxiv-cs-cl
27 Apr 2026
Research

Outcome Rewards Do Not Guarantee Verifiable or Causally Important Reasoning

DGX agent

arXiv:2604.22074v1 Announce Type: new Abstract: Reinforcement Learning from Verifiable Rewards (RLVR) on chain-of-thought reasoning has become a standard part of language model post-training recipes.

researcharxiv-cs-cl
27 Apr 2026
Model Releases

PL-MTEB: Polish Massive Text Embedding Benchmark

DGX agent

arXiv:2405.10138v2 Announce Type: replace Abstract: In this paper, we introduce the Polish Massive Text Embedding Benchmark (PL-MTEB), a comprehensive benchmark for text embeddings in the Polish langu

model-releasesarxiv-cs-cl
27 Apr 2026
Applications

Predicting Liquidity-Aware Bond Yields using Causal GANs and Deep Reinforcement Learning with LLM Evaluation

DGX agent

arXiv:2502.17011v2 Announce Type: replace-cross Abstract: Financial bond yield forecasting is challenging due to data scarcity, nonlinear macroeconomic dependencies, and evolving market conditions. In

applicationsarxiv-cs-cl
27 Apr 2026
Research

Preference Heads in Large Language Models: A Mechanistic Framework for Interpretable Personalization

DGX agent

arXiv:2604.22345v1 Announce Type: new Abstract: Large Language Models (LLMs) exhibit strong implicit personalization ability, yet most existing approaches treat this behavior as a black box, relying o

researcharxiv-cs-cl
27 Apr 2026
Research

PrivUn: Unveiling Latent Ripple Effects and Shallow Forgetting in Privacy Unlearning

DGX agent

arXiv:2604.22076v1 Announce Type: cross Abstract: Large language models (LLMs) often memorize private information during training, raising serious privacy concerns. While machine unlearning has emerge

researcharxiv-cs-cl
27 Apr 2026
Safety

Recognition Without Authorization: LLMs and the Moral Order of Online Advice

DGX agent

arXiv:2604.22143v1 Announce Type: cross Abstract: Large language models are increasingly used to mediate everyday interpersonal dilemmas, yet how their advisory defaults interact with the concentrated

safetyarxiv-cs-cl
27 Apr 2026
Applications

Representational Harms in LLM-Generated Narratives Against Global Majority Nationalities

DGX agent

arXiv:2604.22749v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for text generation tasks from everyday use to high-stakes enterprise and government applications, in

applicationsarxiv-cs-cl
27 Apr 2026
Research

RouteLMT: Learned Sample Routing for Hybrid LLM Translation Deployment

DGX agent

arXiv:2604.22520v1 Announce Type: new Abstract: Large Language Models (LLMs) have achieved remarkable performance in Machine Translation (MT), but deploying them at scale remains prohibitively expensi

researcharxiv-cs-cl
27 Apr 2026
Safety

Selective Contrastive Learning For Gloss Free Sign Language Translation

DGX agent

arXiv:2604.22374v1 Announce Type: new Abstract: Sign language translation (SLT) converts continuous sign videos into spoken-language text, yet it remains challenging due to the intrinsic modality mism

safetyarxiv-cs-cl
27 Apr 2026
Research

Selective Rotary Position Embedding

DGX agent

arXiv:2511.17388v2 Announce Type: replace Abstract: Position information is essential for language modeling. In softmax transformers, Rotary Position Embeddings (extit{RoPE}) encode positions through

researcharxiv-cs-cl
27 Apr 2026
Model Releases

SHAPE: Unifying Safety, Helpfulness and Pedagogy for Educational LLMs

DGX agent

arXiv:2604.22134v1 Announce Type: new Abstract: Large Language Models (LLMs) have been widely explored in educational scenarios. We identify a critical vulnerability in current educational LLMs, pedag

model-releasesarxiv-cs-cl
27 Apr 2026
← Previous
1…123124125126127…161
Next →