AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
23 Apr 2026

Which Reasoning Trajectories Teach Students to Reason Better? A Simple Metric of Informative Alignment

SafetyDGX agent

arXiv:2601.14249v4 Announce Type: replace Abstract: Long chain-of-thought (CoT) trajectories provide rich supervision signals for distilling reasoning from teacher to student LLMs. However, both prior

Whose Story Gets Told? Positionality and Bias in LLM Summaries of Life Narratives

SafetyDGX agent

arXiv:2604.20131v1 Announce Type: new Abstract: Increasingly, studies are exploring using Large Language Models (LLMs) for accelerated or scaled qualitative analysis of text data. While we can compare

WISCA: A Lightweight Model Transition Method to Improve LLM Training via Weight Scaling

ResearchDGX agent

arXiv:2508.16676v2 Announce Type: replace-cross Abstract: Transformer architecture gradually dominates the LLM field. Recent advances in training optimization for Transformer-based large language mode


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
22 Apr 2026

A Bolu: A Structured Dataset for the Computational Analysis of Sardinian Improvisational Poetry

ApplicationsDGX agent

arXiv:2604.19584v1 Announce Type: new Abstract: The growing interest of Natural Language Processing (NLP) in minority languages has not yet bridged the gap in the preservation of oral linguistic herit

A Mechanism and Optimization Study on the Impact of Information Density on User-Generated Content Named Entity Recognition

ResearchDGX agent

arXiv:2604.18944v1 Announce Type: new Abstract: Named Entity Recognition (NER) models trained on clean, high-resource corpora exhibit catastrophic performance collapse when deployed on noisy, sparse U

A Self-Evolving Framework for Efficient Terminal Agents via Observational Context Compression

AgentsDGX agent

arXiv:2604.19572v1 Announce Type: new Abstract: As model capabilities advance, research has increasingly shifted toward long-horizon, multi-turn terminal-centric agentic tasks, where raw environment f

AlignCultura: Towards Culturally Aligned Large Language Models?

Model ReleasesDGX agent

arXiv:2604.19016v1 Announce Type: new Abstract: Cultural alignment in Large Language Models (LLMs) is essential for producing contextually aware, respectful, and trustworthy outputs. Without it, model

An Answer is just the Start: Related Insight Generation for Open-Ended Document-Grounded QA

ResearchDGX agent

arXiv:2604.19685v1 Announce Type: new Abstract: Answering open-ended questions remains challenging for AI systems because it requires synthesis, judgment, and exploration beyond factual retrieval, and

An Empirical Study of Multi-Generation Sampling for Jailbreak Detection in Large Language Models

Model ReleasesDGX agent

arXiv:2604.18775v1 Announce Type: new Abstract: Detecting jailbreak behaviour in large language models remains challenging, particularly when strongly aligned models produce harmful outputs only rarel

Are Large Language Models Economically Viable for Industry Deployment?

Model ReleasesDGX agent

arXiv:2604.19342v1 Announce Type: new Abstract: Generative AI-powered by Large Language Models (LLMs)-is increasingly deployed in industry across healthcare decision support, financial analytics, ente

Article and Comment Frames Shape the Quality of Online Comments

ResearchDGX agent

arXiv:2603.27889v2 Announce Type: replace Abstract: Framing theory posits that how information is presented shapes audience responses, but computational work has largely ignored audience reactions. Wh

Bangla Key2Text: Text Generation from Keywords for a Low Resource Language

Model ReleasesDGX agent

arXiv:2604.19508v1 Announce Type: new Abstract: This paper introduces extit{Bangla Key2Text}, a large-scale dataset of 2.6 million Bangla keyword--text pairs designed for keyword-driven text generatio

Beyond Indistinguishability: Measuring Extraction Risk in LLM APIs

ResearchDGX agent

arXiv:2604.18697v1 Announce Type: cross Abstract: Indistinguishability properties such as differential privacy bounds or low empirically measured membership inference are widely treated as proxies to

Beyond Marginal Distributions: A Framework to Evaluate the Representativeness of Demographic-Aligned LLMs

SafetyDGX agent

arXiv:2601.15755v3 Announce Type: replace Abstract: Large language models are increasingly used to represent human opinions, values, or beliefs, and their steerability towards these ideals is an activ

Beyond Rating: A Comprehensive Evaluation and Benchmark for AI Reviews

Model ReleasesDGX agent

arXiv:2604.19502v1 Announce Type: new Abstract: The rapid adoption of Large Language Models (LLMs) has spurred interest in automated peer review; however, progress is currently stifled by benchmarks t

Can AI-Generated Persuasion Be Detected? Persuaficial Benchmark and AI vs. Human Linguistic Differences

Model ReleasesDGX agent

arXiv:2601.04925v2 Announce Type: replace Abstract: Large Language Models (LLMs) can generate highly persuasive text, raising concerns about their misuse for propaganda, manipulation, and other harmfu

Can Continual Pre-training Bridge the Performance Gap between General-purpose and Specialized Language Models in the Medical Domain?

Model ReleasesDGX agent

arXiv:2604.19394v1 Announce Type: new Abstract: This paper narrows the performance gap between small, specialized models and significantly larger general-purpose models through domain adaptation via c

Capturing Classic Authorial Style in Long-Form Story Generation with GRPO Fine-Tuning

SafetyDGX agent

arXiv:2512.05747v3 Announce Type: replace Abstract: Evaluating and optimising authorial style in long-form story generation remains challenging because style is often assessed with ad hoc prompting an

Cell-Based Representation of Relational Binding in Language Models

ResearchDGX agent

arXiv:2604.19052v1 Announce Type: new Abstract: Understanding a discourse requires tracking entities and the relations that hold between them. While Large Language Models (LLMs) perform well on relati

Comparing energy consumption and accuracy in text classification inference

ResearchDGX agent

arXiv:2508.14170v2 Announce Type: replace Abstract: The increasing deployment of large language models (LLMs) in natural language processing (NLP) tasks raises concerns about energy efficiency and sus

Comparison of sEMG Encoding Accuracy Across Speech Modes Using Articulatory and Phoneme Features

ResearchDGX agent

arXiv:2604.18920v1 Announce Type: cross Abstract: We test whether Speech Articulatory Coding (SPARC) features can linearly predict surface electromyography (sEMG) envelopes across aloud, mimed, and su

Computational Narrative Understanding for Expressive Text-to-Speech

ResearchDGX agent

arXiv:2509.04072v2 Announce Type: replace-cross Abstract: Recent advances in text-to-speech (TTS) have been driven by large, multi-domain speech corpora, yet the expressive potential of audiobook data

Construction of Knowledge Graph based on Language Model

ResearchDGX agent

arXiv:2604.19137v1 Announce Type: new Abstract: Knowledge Graph (KG) can effectively integrate valuable information from massive data, and thus has been rapidly developed and widely used in many field

ContextLeak: Auditing Leakage in Private In-Context Learning Methods

TutorialsDGX agent

arXiv:2512.16059v2 Announce Type: replace-cross Abstract: In-Context Learning (ICL) has become a standard technique for adapting Large Language Models (LLMs) to specialized tasks by supplying task-spe

Cross-lingual Matryoshka Representation Learning across Speech and Text

ResearchDGX agent

arXiv:2602.19991v2 Announce Type: replace Abstract: Speakers of under-represented languages face both a language barrier, as most online knowledge is in a few dominant languages, and a modality barrie

DASH-KV: Accelerating Long-Context LLM Inference via Asymmetric KV Cache Hashing

ResearchDGX agent

arXiv:2604.19351v1 Announce Type: new Abstract: The quadratic computational complexity of the standard attention mechanism constitutes a fundamental bottleneck for large language models in long-contex

Debating the Unspoken: Role-Anchored Multi-Agent Reasoning for Half-Truth Detection

AgentsDGX agent

arXiv:2604.19005v1 Announce Type: new Abstract: Half-truths, claims that are factually correct yet misleading due to omitted context, remain a blind spot for fact verification systems focused on expli

Deep Supervised Contrastive Learning of Pitch Contours for Robust Pitch Accent Classification in Seoul Korean

Model ReleasesDGX agent

arXiv:2604.19477v1 Announce Type: cross Abstract: The intonational structure of Seoul Korean has been defined with discrete tonal categories within the Autosegmental-Metrical model of intonational pho

Detoxification for LLM: From Dataset Itself

Local AiDGX agent

arXiv:2604.19124v1 Announce Type: new Abstract: Existing detoxification methods for large language models mainly focus on post-training stage or inference time, while few tackle the source of toxicity

Diagnosable ColBERT: Debugging Late-Interaction Retrieval Models Using a Learned Latent Space as Reference

SafetyDGX agent

arXiv:2604.19566v1 Announce Type: cross Abstract: Reliable biomedical and clinical retrieval requires more than strong ranking performance: it requires a practical way to find systematic model failure

Discovering a Shared Logical Subspace: Steering LLM Logical Reasoning via Alignment of Natural-Language and Symbolic Views

SafetyDGX agent

arXiv:2604.19716v1 Announce Type: new Abstract: Large Language Models (LLMs) still struggle with multi-step logical reasoning. Existing approaches either purely refine the reasoning chain in natural l

Disparities In Negation Understanding Across Languages In Vision-Language Models

Model ReleasesDGX agent

arXiv:2604.18942v1 Announce Type: new Abstract: Vision-language models (VLMs) exhibit affirmation bias: a systematic tendency to select positive captions ('X is present') even when the correct descrip

Do Emotions Influence Moral Judgment in Large Language Models?

SafetyDGX agent

arXiv:2604.19125v1 Announce Type: new Abstract: Large language models have been extensively studied for emotion recognition and moral reasoning as distinct capabilities, yet the extent to which emotio

Does Self-Consistency Improve the Recall of Encyclopedic Knowledge?

Model ReleasesDGX agent

arXiv:2604.19395v1 Announce Type: new Abstract: While self-consistency is known to improve performance on symbolic reasoning, its effect on the recall of encyclopedic knowledge is unclear due to a lac

Dynamic Model Routing and Cascading for Efficient LLM Inference: A Survey

ResearchDGX agent

arXiv:2603.04445v2 Announce Type: replace-cross Abstract: The rapid growth of large language models (LLMs) with diverse capabilities, costs, and domains has created a critical need for intelligent mod

Emotion-Cause Pair Extraction in Conversations via Semantic Decoupling and Graph Alignment

Model ReleasesDGX agent

arXiv:2604.19547v1 Announce Type: new Abstract: Emotion-Cause Pair Extraction in Conversations (ECPEC) aims to identify the set of causal relations between emotion utterances and their triggering caus

Enhancing Unsupervised Keyword Extraction in Academic Papers through Integrating Highlights with Abstract

ResearchDGX agent

arXiv:2604.19505v1 Announce Type: cross Abstract: Automatic keyword extraction from academic papers is a key area of interest in natural language processing and information retrieval. Although previou

Epistemic orientation in parliamentary discourse is associated with deliberative democracy

ResearchDGX agent

arXiv:2604.19699v1 Announce Type: new Abstract: The pursuit of truth is central to democratic deliberation and governance, yet political discourse reflects varying epistemic orientations, ranging from

EssayCBM: Rubric-Aligned Concept Bottleneck Models for Transparent Essay Grading

ResearchDGX agent

arXiv:2512.20817v2 Announce Type: replace Abstract: Automated essay scoring (AES) has advanced significantly with neural language models, yet most systems remain opaque, offering little visibility int

Evaluating LLM-Driven Summarisation of Parliamentary Debates with Computational Argumentation

SafetyDGX agent

arXiv:2604.19331v1 Announce Type: new Abstract: Understanding how policy is debated and justified in parliament is a fundamental aspect of the democratic process. However, the volume and complexity of

Exploring Language-Agnosticity in Function Vectors: A Case Study in Machine Translation

ApplicationsDGX agent

arXiv:2604.19678v1 Announce Type: new Abstract: Function vectors (FVs) are vector representations of tasks extracted from model activations during in-context learning. While prior work has shown that

FoNE: Precise Single-Token Number Embeddings via Fourier Features

TutorialsDGX agent

arXiv:2502.09741v2 Announce Type: replace Abstract: Large Language Models (LLMs) typically represent numbers using multiple tokens, which requires the model to aggregate these tokens to interpret nume

From Proof to Program: Characterizing Tool-Induced Reasoning Hallucinations in Large Language Models

Model ReleasesDGX agent

arXiv:2511.10899v2 Announce Type: replace Abstract: Tool-augmented Language Models (TaLMs) can invoke external tools to solve problems beyond their parametric capacity. However, it remains unclear whe

HarDBench: A Benchmark for Draft-Based Co-Authoring Jailbreak Attacks for Safe Human-LLM Collaborative Writing

Model ReleasesDGX agent

arXiv:2604.19274v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used as co-authors in collaborative writing, where users begin with rough drafts and rely on LLMs to compl

Headlines You Won't Forget: Can Pronoun Insertion Increase Memorability?

ResearchDGX agent

arXiv:2604.19189v1 Announce Type: new Abstract: For news headlines to influence beliefs and drive action, relevant information needs to be retained and retrievable from memory. In this probing study w

HoWToBench: Holistic Evaluation for LLM's Capability in Human-level Writing using Tree of Writing

Model ReleasesDGX agent

arXiv:2604.19071v1 Announce Type: new Abstract: Evaluating the writing capabilities of large language models (LLMs) remains a significant challenge due to the multidimensional nature of writing skills

Hybrid Architectures for Language Models: Systematic Analysis and Design Insights

ResearchDGX agent

arXiv:2510.04800v3 Announce Type: replace Abstract: Recent progress in large language models demonstrates that hybrid architectures--combining self-attention mechanisms with structured state space mod

Improving the Distributional Alignment of LLMs using Supervision

Model ReleasesDGX agent

arXiv:2507.00439v4 Announce Type: replace Abstract: The ability to accurately align LLMs with diverse population groups on subjective questions would have great value. In this work, we show that addin

Investigating Counterfactual Unfairness in LLMs towards Identities through Humor

SafetyDGX agent

arXiv:2604.18729v1 Announce Type: new Abstract: Humor holds up a mirror to social perception: what we find funny often reflects who we are and how we judge others. When language models engage with hum

LegalBench-BR: A Benchmark for Evaluating Large Language Models on Brazilian Legal Decision Classification

Model ReleasesDGX agent

arXiv:2604.18878v1 Announce Type: new Abstract: We introduce LegalBench-BR, the first public benchmark for evaluating language models on Brazilian legal text classification. The dataset comprises 3,10

Less Is More: Cognitive Load and the Single-Prompt Ceiling in LLM Mathematical Reasoning

Model ReleasesDGX agent

arXiv:2604.18897v1 Announce Type: new Abstract: We present a systematic empirical study of prompt engineering for formal mathematical reasoning in the context of the SAIR Equational Theories Stage 1 c

LogosKG: Hardware-Optimized Scalable and Interpretable Knowledge Graph Retrieval

SafetyDGX agent

arXiv:2604.18913v1 Announce Type: new Abstract: Knowledge graphs (KGs) are increasingly integrated with large language models (LLMs) to provide structured, verifiable reasoning. A core operation in th

Lost in Translation: Do LVLM Judges Generalize Across Languages?

Model ReleasesDGX agent

arXiv:2604.19405v1 Announce Type: new Abstract: Automatic evaluators such as reward models play a central role in the alignment and evaluation of large vision-language models (LVLMs). Despite their gr

Mango: Multi-Agent Web Navigation via Global-View Optimization

Model ReleasesDGX agent

arXiv:2604.18779v1 Announce Type: new Abstract: Existing web agents typically initiate exploration from the root URL, which is inefficient for complex websites with deep hierarchical structures. Witho

Micro Language Models Enable Instant Responses

Model ReleasesDGX agent

arXiv:2604.19642v1 Announce Type: new Abstract: Edge devices such as smartwatches and smart glasses cannot continuously run even the smallest 100M-1B parameter language models due to power and compute

Mind the Unseen Mass: Unmasking LLM Hallucinations via Soft-Hybrid Alphabet Estimation

ResearchDGX agent

arXiv:2604.19162v1 Announce Type: new Abstract: This paper studies uncertainty quantification for large language models (LLMs) under black-box access, where only a small number of responses can be sam

MiroThinker: Pushing the Performance Boundaries of Open-Source Research Agents via Model, Context, and Interactive Scaling

Model ReleasesDGX agent

arXiv:2511.11793v3 Announce Type: replace Abstract: We present MiroThinker v1.0, an open-source research agent designed to advance tool-augmented reasoning and information-seeking capabilities. Unlike

Mitigating Judgment Preference Bias in Large Language Models through Group-Based Polling

SafetyDGX agent

arXiv:2510.08145v2 Announce Type: replace Abstract: Large Language Models (LLMs) as automatic evaluators, commonly referred to as LLM-as-a-Judge, have also attracted growing attention. This approach p

Model-Agnostic Meta Learning for Class Imbalance Adaptation

ResearchDGX agent

arXiv:2604.18759v1 Announce Type: new Abstract: Class imbalance is a widespread challenge in NLP tasks, significantly hindering robust performance across diverse domains and applications. We introduce

Multi-Task Reinforcement Learning for Enhanced Multimodal LLM-as-a-Judge

SafetyDGX agent

arXiv:2603.11665v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have been widely adopted as MLLM-as-a-Judges due to their strong alignment with human judgment across vario

← Previous
1…102103104105106…129
Next →