AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
Research

When Efficient Communication Explains Convexity

DGX agent

arXiv:2602.02821v2 Announce Type: replace Abstract: Much recent work has argued that the variation in the languages of the world can be explained from the perspective of efficient communication; in pa

researcharxiv-cs-cl
12 May 2026
Research

Where do aspectual variants of light verb constructions belong?

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2605.10605v1 Announce Type: new Abstract: Expressions with an aspectual variant of a light verb, e.g. 'take on debt' vs. 'have debt', are frequent in texts but often difficult to classify betwee

researcharxiv-cs-cl
12 May 2026
Model Releases

Where Does Long-Context Supervision Actually Go? Effective-Context Exposure Balancing

DGX agent

arXiv:2605.10544v1 Announce Type: new Abstract: Long-context adaptation is often viewed as window scaling, but this misses a token-level supervision mismatch: in packed training with document masking,

model-releasesarxiv-cs-cl
12 May 2026
Research

Why is prompting hard? Understanding prompts on binary sequence predictors

DGX agent

arXiv:2502.10760v2 Announce Type: replace Abstract: Frontier models can be prompted or conditioned to do many tasks, but finding good prompts is not always easy, nor is understanding some performant p

researcharxiv-cs-cl
12 May 2026
Model Releases

WildClawBench: A Benchmark for Real-World, Long-Horizon Agent Evaluation

DGX agent

arXiv:2605.10912v1 Announce Type: new Abstract: Large language and vision-language models increasingly power agents that act on a user's behalf through command-line interface (CLI) harnesses. However,

model-releasesarxiv-cs-cl
12 May 2026
Research

XPERT: Expert Knowledge Transfer for Effective Training of Language Models

DGX agent

arXiv:2605.08842v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) language models organize knowledge into explicitly routed expert modules, making expert-level representations traceable and ana

researcharxiv-cs-cl
12 May 2026
Research

YEZE at SemEval-2026 Task 9: Detecting Multilingual, Multicultural and Multievent Online Polarization via Heterogeneous Ensembling

DGX agent

arXiv:2605.06231v2 Announce Type: replace Abstract: This paper presents our system for SemEval-2026 Task 9: Detecting Multilingual, Multicultural and Multievent Online Polarization, which identifies p

researcharxiv-cs-cl
12 May 2026
Research

A Comparative Analysis of Classical Machine Learning and Deep Learning Approaches for Sentiment Classification on IMDb Movie Reviews

DGX agent

arXiv:2605.07811v1 Announce Type: new Abstract: This paper presents a comparative study of classical machine learning and deep learning methods for sentiment classification on the IMDb movie reviews d

researcharxiv-cs-cl
11 May 2026
Model Releases

A Reproducible Multi-Architecture Baseline for Token-Level Chinese Metaphor Identification under the MIPVU Framework

DGX agent

arXiv:2605.07170v1 Announce Type: new Abstract: Metaphor is pervasive in everyday language, yet token-level computational identification of metaphor-related words in Chinese under the MIPVU framework

model-releasesarxiv-cs-cl
11 May 2026
Safety

Accurate and Efficient Statistical Testing for Word Semantic Breadth

DGX agent

arXiv:2605.08048v1 Announce Type: new Abstract: Measuring the breadth of a word's meaning, or its spread across contexts, has become feasible with contextualized token embeddings. A word type can be r

safetyarxiv-cs-cl
11 May 2026
Model Releases

Anatomy of Unlearning: The Dual Impact of Fact Salience and Model Fine-Tuning

DGX agent

arXiv:2602.19612v3 Announce Type: replace Abstract: Machine Unlearning (MU) enables Large Language Models (LLMs) to remove unsafe or outdated information. However, existing work assumes that all facts

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

Ask Early, Ask Late, Ask Right: When Does Clarification Timing Matter for Long-Horizon Agents?

DGX agent

arXiv:2605.07937v1 Announce Type: new Abstract: Long-horizon AI agents execute complex workflows spanning hundreds of sequential actions, yet a single wrong assumption early on can cascade into irreve

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

Ask Patients with Patience: Enabling LLMs for Human-Centric Medical Dialogue with Grounded Reasoning

DGX agent

arXiv:2502.07143v3 Announce Type: replace Abstract: The severe shortage of medical doctors limits access to timely and reliable healthcare, leaving millions underserved. Large language models (LLMs) o

model-releasesarxiv-cs-cl
11 May 2026
Research

Bayesian Attention Mechanism: A Probabilistic Framework for Positional Encoding and Context Length Extrapolation

DGX agent

arXiv:2505.22842v4 Announce Type: replace Abstract: Transformer-based language models rely on positional encoding (PE) to handle token order and support context length extrapolation. However, existing

researcharxiv-cs-cl
11 May 2026
Model Releases

Benchmarked Yet Not Measured -- Generative AI Should be Evaluated Against Real-World Utility

DGX agent

arXiv:2605.06856v1 Announce Type: cross Abstract: Generative AI systems achieve impressive performance on standard benchmarks yet fail to deliver real-world utility, a disconnect we identify across 28

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

Beyond Factual Accuracy: Evaluating Global Reasoning Integrity in RAG Systems with LogicScore

DGX agent

arXiv:2601.15050v4 Announce Type: replace Abstract: Current evaluation methods for Retrieval Augmented Generation (RAG) suffer from extit{factual myopia}: they relentlessly emphasize factual accuracy

model-releasesarxiv-cs-cl
11 May 2026
Safety

Beyond 'I cannot fulfill this request': Alleviating Rigid Rejection in LLMs via Label Enhancement

DGX agent

arXiv:2605.07883v1 Announce Type: new Abstract: Large Language Models (LLMs) rely on safety alignment to obey safe requests while refusing harmful ones. However, traditional refusal mechanisms often l

safetyarxiv-cs-cl
11 May 2026
Research

Beyond Reasoning: Reinforcement Learning Unlocks Parametric Knowledge in LLMs

DGX agent

arXiv:2605.07153v1 Announce Type: new Abstract: Reinforcement learning (RL) has achieved remarkable success in LLM reasoning, but whether it can also improve direct recall of parametric knowledge rema

researcharxiv-cs-cl
11 May 2026
Applications

Beyond Single Ground Truth: Reference Monism as Epistemic Injustice in ASR Evaluation

DGX agent

arXiv:2605.07084v1 Announce Type: new Abstract: Automatic speech recognition (ASR) evaluation compares system output to ground truth transcripts, with Word Error Rate (WER) quantifying the distance be

applicationsarxiv-cs-cl
11 May 2026
Research

Bridging Textual Profiles and Latent User Embeddings for Personalization

DGX agent

arXiv:2605.06981v1 Announce Type: cross Abstract: Personalized systems rely on user representations to connect behavioral history with downstream recommendation applications. Existing methods typicall

researcharxiv-cs-cl
11 May 2026
Safety

Can David Beat Goliath? On Multi-Hop Reasoning with Resource-Constrained Agents

DGX agent

arXiv:2601.21699v2 Announce Type: replace Abstract: Multi-turn reasoning agents solve complex questions by decomposing them into intermediate retrieval or tool-use steps, for accumulating supporting e

safetyarxiv-cs-cl
11 May 2026
Applications

Can LLMs Take Retrieved Information with a Grain of Salt?

DGX agent

arXiv:2605.06919v1 Announce Type: new Abstract: Large language models have demonstrated impressive retrieval-augmented capabilities. However, a crucial area remains underexplored: their ability to app

applicationsarxiv-cs-cl
11 May 2026
Model Releases

Chain-based Distillation for Effective Initialization of Variable-Sized Small Language Models

DGX agent

arXiv:2605.07783v1 Announce Type: new Abstract: Large language models (LLMs) achieve strong performance but remain costly to deploy in resource-constrained settings. Training small language models (SL

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

ChartREG++: Towards Benchmarking and Improving Chart Referring Expression Grounding under Diverse referring clues and Multi-Target Referring

DGX agent

arXiv:2605.07415v1 Announce Type: cross Abstract: Referring expression grounding is a core problem in visual grounding and is widely used as a diagnostic of spatial grounding and reasoning in vision a

model-releasesarxiv-cs-cl
11 May 2026
Hardware

CktFormalizer: Autoformalization of Natural Language into Circuit Representations

DGX agent

arXiv:2605.07782v1 Announce Type: new Abstract: LLMs can generate hardware descriptions from natural language specifications, but the resulting Verilog often contains width mismatches, combinational l

hardwarearxiv-cs-cl
11 May 2026
Research

CLIPer: Tailoring Diverse User Preference via Classifier-Guided Inference-Time Personalization

DGX agent

arXiv:2605.07162v1 Announce Type: new Abstract: Personalized LLMs can significantly enhance user experiences by tailoring responses to preferences such as helpfulness, conciseness, and humor. However,

researcharxiv-cs-cl
11 May 2026
Research

Comprehensiveness Metrics for Automatic Evaluation of Factual Recall in Text Generation

DGX agent

arXiv:2510.07926v2 Announce Type: replace Abstract: Despite demonstrating remarkable performance across a wide range of tasks, large language models (LLMs) have also been found to frequently produce o

researcharxiv-cs-cl
11 May 2026
Tutorials

Conformal Path Reasoning: Trustworthy Knowledge Graph Question Answering via Path-Level Calibration

DGX agent

arXiv:2605.08077v1 Announce Type: new Abstract: Knowledge Graph Question Answering (KGQA) has shown promise for grounded and interpretable reasoning, yet existing approaches often fail to provide reli

tutorialsarxiv-cs-cl
11 May 2026
Model Releases

Data Contamination in Neural Hieroglyphic Translation: A Reproducibility Study

DGX agent

arXiv:2605.07453v1 Announce Type: new Abstract: Ancient and endangered languages pose a unique challenge for NLP: their datasets are inherently scarce, difficult to expand, and built from formulaic co

model-releasesarxiv-cs-cl
11 May 2026
Research

DiffRetriever: Parallel Representative Tokens for Retrieval with Diffusion Language Models

DGX agent

arXiv:2605.07210v1 Announce Type: cross Abstract: PromptReps showed that an autoregressive language model can be used directly as a retriever by prompting it to generate dense and sparse representatio

researcharxiv-cs-cl
11 May 2026
Research

Don't Ignore the Tail: Decoupling top-K Probabilities for Efficient Language Model Distillation

DGX agent

arXiv:2602.20816v3 Announce Type: replace Abstract: The core learning signal used in language model distillation is the standard Kullback-Leibler (KL) divergence between the student and teacher distri

researcharxiv-cs-cl
11 May 2026
Research

ExpThink: Experience-Guided Reinforcement Learning for Adaptive Chain-of-Thought Compression

DGX agent

arXiv:2605.07501v1 Announce Type: cross Abstract: Large reasoning models (LRMs) achieve strong performance via extended chain-of-thought (CoT) reasoning, yet suffer from excessive token consumption an

researcharxiv-cs-cl
11 May 2026
Model Releases

FinReasoning: A Hierarchical Benchmark for Reliable Financial Research Reporting

DGX agent

arXiv:2603.19254v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly deployed in financial research workflows, where their role is evolving from single-model assistance fo

model-releasesarxiv-cs-cl
11 May 2026
Safety

From 0-Order Selection to 2-Order Judgment: Combinatorial Hardening Exposes Compositional Failures in Frontier LLMs

DGX agent

arXiv:2605.07268v1 Announce Type: new Abstract: Multiple-choice reasoning benchmarks face dual challenges: rapid saturation from advancing models and data contamination that undermines static evaluati

safetyarxiv-cs-cl
11 May 2026
Research

From Standalone LLMs to Integrated Intelligence: A Survey of Compound Al Systems

DGX agent

arXiv:2506.04565v2 Announce Type: replace-cross Abstract: Compound AI Systems (CAIS) are an emerging paradigm that integrates large language models (LLMs) with external components, including retriever

researcharxiv-cs-cl
11 May 2026
Local Ai

Generating training datasets for legal chatbots in Korean

DGX agent

arXiv:2605.07432v1 Announce Type: new Abstract: Chatbots are robots that can communicate with humans using text or voice signals. Legal chatbots improve access to justice, since legal representation a

local-aiarxiv-cs-cl
11 May 2026
Model Releases

GLiGuard: Schema-Conditioned Classification for LLM Safeguard

DGX agent

arXiv:2605.07982v1 Announce Type: new Abstract: Ensuring safe, policy-compliant outputs from large language models requires real-time content moderation that can scale across multiple safety dimension

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

Gradient-Based LoRA Rank Allocation Under GRPO: An Empirical Study

DGX agent

arXiv:2605.07366v1 Announce Type: new Abstract: Adaptive rank allocation for LoRA, allocating more parameters to important layers and fewer to unimportant ones, consistently improves efficiency under

model-releasesarxiv-cs-cl
11 May 2026
Research

GRaSp: Automatic Example Optimization for In-Context Learning in Low-Data Tasks

DGX agent

arXiv:2605.07454v1 Announce Type: new Abstract: In-context learning enables large language models to adapt to new tasks, but their performance is highly sensitive to the selected examples. Finding eff

researcharxiv-cs-cl
11 May 2026
Safety

Guidance Is Not a Hyperparameter: Learning Dynamic Control in Diffusion Language Models

DGX agent

arXiv:2605.07701v1 Announce Type: new Abstract: Classifier-Free Guidance (CFG) is a widely used mechanism for controlling diffusion-based generative models, yet its guidance scale is typically treated

safetyarxiv-cs-cl
11 May 2026
Tutorials

How to Train Your Latent Diffusion Language Model Jointly With the Latent Space

DGX agent

arXiv:2605.07933v1 Announce Type: new Abstract: Latent diffusion models offer an attractive alternative to discrete diffusion for non-autoregressive text generation by operating on continuous text rep

tutorialsarxiv-cs-cl
11 May 2026
Safety

How Value Induction Reshapes LLM Behaviour

DGX agent

arXiv:2605.07925v1 Announce Type: new Abstract: Conversational Large Language Models are post-trained on language that expresses specific behavioural traits, such as curiosity, open-mindedness, and em

safetyarxiv-cs-cl
11 May 2026
Tutorials

Human-like fleeting memory improves language learning but impairs reading time prediction in transformer language models

DGX agent

arXiv:2508.05803v2 Announce Type: replace Abstract: Human memory is fleeting. As words are processed, the exact wordforms that make up incoming sentences are rapidly lost. Cognitive scientists have lo

tutorialsarxiv-cs-cl
11 May 2026
Applications

Hybrid TF--IDF Logistic Regression and MLP Neural Baseline for Indonesian Three-Class Sentiment Analysis on Social Media Text

DGX agent

arXiv:2605.07793v1 Announce Type: new Abstract: This paper presents a compact three-class sentiment analysis study for Indonesian social media text. The task is formulated with positive, negative, and

applicationsarxiv-cs-cl
11 May 2026
Model Releases

Intent-Driven Semantic ID Generation for Grounded Conversational News Recommendation

DGX agent

arXiv:2605.07613v1 Announce Type: new Abstract: Conversational news recommendation requires grounding each suggestion in a rapidly evolving article corpus while addressing implicit user intents that l

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

InterLV-Search: Benchmarking Interleaved Multimodal Agentic Search

DGX agent

arXiv:2605.07510v1 Announce Type: cross Abstract: Existing benchmarks for multimodal agentic search evaluate multimodal search and visual browsing, but visual evidence is either confined to the input

model-releasesarxiv-cs-cl
11 May 2026
Research

Interpreting Speaker Characteristics in the Dimensions of Self-Supervised Speech Features

DGX agent

arXiv:2603.03096v2 Announce Type: replace-cross Abstract: How do speech models trained through self-supervised learning structure their representations? Previous studies have looked at how information

researcharxiv-cs-cl
11 May 2026
Local Ai

Is She Even Relevant? When BERT Ignores Explicit Gender Cues

DGX agent

arXiv:2605.07622v1 Announce Type: new Abstract: Gender bias in large language models has primarily been investigated for English, while languages with grammatical or morphological gender remain compar

local-aiarxiv-cs-cl
11 May 2026
← Previous
1…101102103104105…161
Next →