AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
Research

Beyond Indistinguishability: Measuring Extraction Risk in LLM APIs

DGX agent

arXiv:2604.18697v1 Announce Type: cross Abstract: Indistinguishability properties such as differential privacy bounds or low empirically measured membership inference are widely treated as proxies to

researcharxiv-cs-cl
22 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Beyond Marginal Distributions: A Framework to Evaluate the Representativeness of Demographic-Aligned LLMs

DGX agent

arXiv:2601.15755v3 Announce Type: replace Abstract: Large language models are increasingly used to represent human opinions, values, or beliefs, and their steerability towards these ideals is an activ

safetyarxiv-cs-cl
22 Apr 2026
Model Releases

Beyond Rating: A Comprehensive Evaluation and Benchmark for AI Reviews

DGX agent

arXiv:2604.19502v1 Announce Type: new Abstract: The rapid adoption of Large Language Models (LLMs) has spurred interest in automated peer review; however, progress is currently stifled by benchmarks t

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

Can AI-Generated Persuasion Be Detected? Persuaficial Benchmark and AI vs. Human Linguistic Differences

DGX agent

arXiv:2601.04925v2 Announce Type: replace Abstract: Large Language Models (LLMs) can generate highly persuasive text, raising concerns about their misuse for propaganda, manipulation, and other harmfu

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

Can Continual Pre-training Bridge the Performance Gap between General-purpose and Specialized Language Models in the Medical Domain?

DGX agent

arXiv:2604.19394v1 Announce Type: new Abstract: This paper narrows the performance gap between small, specialized models and significantly larger general-purpose models through domain adaptation via c

model-releasesarxiv-cs-cl
22 Apr 2026
Safety

Capturing Classic Authorial Style in Long-Form Story Generation with GRPO Fine-Tuning

DGX agent

arXiv:2512.05747v3 Announce Type: replace Abstract: Evaluating and optimising authorial style in long-form story generation remains challenging because style is often assessed with ad hoc prompting an

safetyarxiv-cs-cl
22 Apr 2026
Research

Cell-Based Representation of Relational Binding in Language Models

DGX agent

arXiv:2604.19052v1 Announce Type: new Abstract: Understanding a discourse requires tracking entities and the relations that hold between them. While Large Language Models (LLMs) perform well on relati

researcharxiv-cs-cl
22 Apr 2026
Research

Comparing energy consumption and accuracy in text classification inference

DGX agent

arXiv:2508.14170v2 Announce Type: replace Abstract: The increasing deployment of large language models (LLMs) in natural language processing (NLP) tasks raises concerns about energy efficiency and sus

researcharxiv-cs-cl
22 Apr 2026
Research

Comparison of sEMG Encoding Accuracy Across Speech Modes Using Articulatory and Phoneme Features

DGX agent

arXiv:2604.18920v1 Announce Type: cross Abstract: We test whether Speech Articulatory Coding (SPARC) features can linearly predict surface electromyography (sEMG) envelopes across aloud, mimed, and su

researcharxiv-cs-cl
22 Apr 2026
Research

Computational Narrative Understanding for Expressive Text-to-Speech

DGX agent

arXiv:2509.04072v2 Announce Type: replace-cross Abstract: Recent advances in text-to-speech (TTS) have been driven by large, multi-domain speech corpora, yet the expressive potential of audiobook data

researcharxiv-cs-cl
22 Apr 2026
Research

Construction of Knowledge Graph based on Language Model

DGX agent

arXiv:2604.19137v1 Announce Type: new Abstract: Knowledge Graph (KG) can effectively integrate valuable information from massive data, and thus has been rapidly developed and widely used in many field

researcharxiv-cs-cl
22 Apr 2026
Tutorials

ContextLeak: Auditing Leakage in Private In-Context Learning Methods

DGX agent

arXiv:2512.16059v2 Announce Type: replace-cross Abstract: In-Context Learning (ICL) has become a standard technique for adapting Large Language Models (LLMs) to specialized tasks by supplying task-spe

tutorialsarxiv-cs-cl
22 Apr 2026
Research

Cross-lingual Matryoshka Representation Learning across Speech and Text

DGX agent

arXiv:2602.19991v2 Announce Type: replace Abstract: Speakers of under-represented languages face both a language barrier, as most online knowledge is in a few dominant languages, and a modality barrie

researcharxiv-cs-cl
22 Apr 2026
Research

DASH-KV: Accelerating Long-Context LLM Inference via Asymmetric KV Cache Hashing

DGX agent

arXiv:2604.19351v1 Announce Type: new Abstract: The quadratic computational complexity of the standard attention mechanism constitutes a fundamental bottleneck for large language models in long-contex

researcharxiv-cs-cl
22 Apr 2026
Agents

Debating the Unspoken: Role-Anchored Multi-Agent Reasoning for Half-Truth Detection

DGX agent

arXiv:2604.19005v1 Announce Type: new Abstract: Half-truths, claims that are factually correct yet misleading due to omitted context, remain a blind spot for fact verification systems focused on expli

agentsarxiv-cs-cl
22 Apr 2026
Model Releases

Deep Supervised Contrastive Learning of Pitch Contours for Robust Pitch Accent Classification in Seoul Korean

DGX agent

arXiv:2604.19477v1 Announce Type: cross Abstract: The intonational structure of Seoul Korean has been defined with discrete tonal categories within the Autosegmental-Metrical model of intonational pho

model-releasesarxiv-cs-cl
22 Apr 2026
Local Ai

Detoxification for LLM: From Dataset Itself

DGX agent

arXiv:2604.19124v1 Announce Type: new Abstract: Existing detoxification methods for large language models mainly focus on post-training stage or inference time, while few tackle the source of toxicity

local-aiarxiv-cs-cl
22 Apr 2026
Safety

Diagnosable ColBERT: Debugging Late-Interaction Retrieval Models Using a Learned Latent Space as Reference

DGX agent

arXiv:2604.19566v1 Announce Type: cross Abstract: Reliable biomedical and clinical retrieval requires more than strong ranking performance: it requires a practical way to find systematic model failure

safetyarxiv-cs-cl
22 Apr 2026
Safety

Discovering a Shared Logical Subspace: Steering LLM Logical Reasoning via Alignment of Natural-Language and Symbolic Views

DGX agent

arXiv:2604.19716v1 Announce Type: new Abstract: Large Language Models (LLMs) still struggle with multi-step logical reasoning. Existing approaches either purely refine the reasoning chain in natural l

safetyarxiv-cs-cl
22 Apr 2026
Model Releases

Disparities In Negation Understanding Across Languages In Vision-Language Models

DGX agent

arXiv:2604.18942v1 Announce Type: new Abstract: Vision-language models (VLMs) exhibit affirmation bias: a systematic tendency to select positive captions ('X is present') even when the correct descrip

model-releasesarxiv-cs-cl
22 Apr 2026
Safety

Do Emotions Influence Moral Judgment in Large Language Models?

DGX agent

arXiv:2604.19125v1 Announce Type: new Abstract: Large language models have been extensively studied for emotion recognition and moral reasoning as distinct capabilities, yet the extent to which emotio

safetyarxiv-cs-cl
22 Apr 2026
Model Releases

Does Self-Consistency Improve the Recall of Encyclopedic Knowledge?

DGX agent

arXiv:2604.19395v1 Announce Type: new Abstract: While self-consistency is known to improve performance on symbolic reasoning, its effect on the recall of encyclopedic knowledge is unclear due to a lac

model-releasesarxiv-cs-cl
22 Apr 2026
Research

Dynamic Model Routing and Cascading for Efficient LLM Inference: A Survey

DGX agent

arXiv:2603.04445v2 Announce Type: replace-cross Abstract: The rapid growth of large language models (LLMs) with diverse capabilities, costs, and domains has created a critical need for intelligent mod

researcharxiv-cs-cl
22 Apr 2026
Model Releases

Emotion-Cause Pair Extraction in Conversations via Semantic Decoupling and Graph Alignment

DGX agent

arXiv:2604.19547v1 Announce Type: new Abstract: Emotion-Cause Pair Extraction in Conversations (ECPEC) aims to identify the set of causal relations between emotion utterances and their triggering caus

model-releasesarxiv-cs-cl
22 Apr 2026
Research

Enhancing Unsupervised Keyword Extraction in Academic Papers through Integrating Highlights with Abstract

DGX agent

arXiv:2604.19505v1 Announce Type: cross Abstract: Automatic keyword extraction from academic papers is a key area of interest in natural language processing and information retrieval. Although previou

researcharxiv-cs-cl
22 Apr 2026
Research

Epistemic orientation in parliamentary discourse is associated with deliberative democracy

DGX agent

arXiv:2604.19699v1 Announce Type: new Abstract: The pursuit of truth is central to democratic deliberation and governance, yet political discourse reflects varying epistemic orientations, ranging from

researcharxiv-cs-cl
22 Apr 2026
Research

EssayCBM: Rubric-Aligned Concept Bottleneck Models for Transparent Essay Grading

DGX agent

arXiv:2512.20817v2 Announce Type: replace Abstract: Automated essay scoring (AES) has advanced significantly with neural language models, yet most systems remain opaque, offering little visibility int

researcharxiv-cs-cl
22 Apr 2026
Safety

Evaluating LLM-Driven Summarisation of Parliamentary Debates with Computational Argumentation

DGX agent

arXiv:2604.19331v1 Announce Type: new Abstract: Understanding how policy is debated and justified in parliament is a fundamental aspect of the democratic process. However, the volume and complexity of

safetyarxiv-cs-cl
22 Apr 2026
Applications

Exploring Language-Agnosticity in Function Vectors: A Case Study in Machine Translation

DGX agent

arXiv:2604.19678v1 Announce Type: new Abstract: Function vectors (FVs) are vector representations of tasks extracted from model activations during in-context learning. While prior work has shown that

applicationsarxiv-cs-cl
22 Apr 2026
Tutorials

FoNE: Precise Single-Token Number Embeddings via Fourier Features

DGX agent

arXiv:2502.09741v2 Announce Type: replace Abstract: Large Language Models (LLMs) typically represent numbers using multiple tokens, which requires the model to aggregate these tokens to interpret nume

tutorialsarxiv-cs-cl
22 Apr 2026
Model Releases

From Proof to Program: Characterizing Tool-Induced Reasoning Hallucinations in Large Language Models

DGX agent

arXiv:2511.10899v2 Announce Type: replace Abstract: Tool-augmented Language Models (TaLMs) can invoke external tools to solve problems beyond their parametric capacity. However, it remains unclear whe

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

HarDBench: A Benchmark for Draft-Based Co-Authoring Jailbreak Attacks for Safe Human-LLM Collaborative Writing

DGX agent

arXiv:2604.19274v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used as co-authors in collaborative writing, where users begin with rough drafts and rely on LLMs to compl

model-releasesarxiv-cs-cl
22 Apr 2026
Research

Headlines You Won't Forget: Can Pronoun Insertion Increase Memorability?

DGX agent

arXiv:2604.19189v1 Announce Type: new Abstract: For news headlines to influence beliefs and drive action, relevant information needs to be retained and retrievable from memory. In this probing study w

researcharxiv-cs-cl
22 Apr 2026
Model Releases

HoWToBench: Holistic Evaluation for LLM's Capability in Human-level Writing using Tree of Writing

DGX agent

arXiv:2604.19071v1 Announce Type: new Abstract: Evaluating the writing capabilities of large language models (LLMs) remains a significant challenge due to the multidimensional nature of writing skills

model-releasesarxiv-cs-cl
22 Apr 2026
Research

Hybrid Architectures for Language Models: Systematic Analysis and Design Insights

DGX agent

arXiv:2510.04800v3 Announce Type: replace Abstract: Recent progress in large language models demonstrates that hybrid architectures--combining self-attention mechanisms with structured state space mod

researcharxiv-cs-cl
22 Apr 2026
Model Releases

Improving the Distributional Alignment of LLMs using Supervision

DGX agent

arXiv:2507.00439v4 Announce Type: replace Abstract: The ability to accurately align LLMs with diverse population groups on subjective questions would have great value. In this work, we show that addin

model-releasesarxiv-cs-cl
22 Apr 2026
Safety

Investigating Counterfactual Unfairness in LLMs towards Identities through Humor

DGX agent

arXiv:2604.18729v1 Announce Type: new Abstract: Humor holds up a mirror to social perception: what we find funny often reflects who we are and how we judge others. When language models engage with hum

safetyarxiv-cs-cl
22 Apr 2026
Model Releases

LegalBench-BR: A Benchmark for Evaluating Large Language Models on Brazilian Legal Decision Classification

DGX agent

arXiv:2604.18878v1 Announce Type: new Abstract: We introduce LegalBench-BR, the first public benchmark for evaluating language models on Brazilian legal text classification. The dataset comprises 3,10

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

Less Is More: Cognitive Load and the Single-Prompt Ceiling in LLM Mathematical Reasoning

DGX agent

arXiv:2604.18897v1 Announce Type: new Abstract: We present a systematic empirical study of prompt engineering for formal mathematical reasoning in the context of the SAIR Equational Theories Stage 1 c

model-releasesarxiv-cs-cl
22 Apr 2026
Safety

LogosKG: Hardware-Optimized Scalable and Interpretable Knowledge Graph Retrieval

DGX agent

arXiv:2604.18913v1 Announce Type: new Abstract: Knowledge graphs (KGs) are increasingly integrated with large language models (LLMs) to provide structured, verifiable reasoning. A core operation in th

safetyarxiv-cs-cl
22 Apr 2026
Model Releases

Lost in Translation: Do LVLM Judges Generalize Across Languages?

DGX agent

arXiv:2604.19405v1 Announce Type: new Abstract: Automatic evaluators such as reward models play a central role in the alignment and evaluation of large vision-language models (LVLMs). Despite their gr

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

Mango: Multi-Agent Web Navigation via Global-View Optimization

DGX agent

arXiv:2604.18779v1 Announce Type: new Abstract: Existing web agents typically initiate exploration from the root URL, which is inefficient for complex websites with deep hierarchical structures. Witho

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

Micro Language Models Enable Instant Responses

DGX agent

arXiv:2604.19642v1 Announce Type: new Abstract: Edge devices such as smartwatches and smart glasses cannot continuously run even the smallest 100M-1B parameter language models due to power and compute

model-releasesarxiv-cs-cl
22 Apr 2026
Research

Mind the Unseen Mass: Unmasking LLM Hallucinations via Soft-Hybrid Alphabet Estimation

DGX agent

arXiv:2604.19162v1 Announce Type: new Abstract: This paper studies uncertainty quantification for large language models (LLMs) under black-box access, where only a small number of responses can be sam

researcharxiv-cs-cl
22 Apr 2026
Model Releases

MiroThinker: Pushing the Performance Boundaries of Open-Source Research Agents via Model, Context, and Interactive Scaling

DGX agent

arXiv:2511.11793v3 Announce Type: replace Abstract: We present MiroThinker v1.0, an open-source research agent designed to advance tool-augmented reasoning and information-seeking capabilities. Unlike

model-releasesarxiv-cs-cl
22 Apr 2026
Safety

Mitigating Judgment Preference Bias in Large Language Models through Group-Based Polling

DGX agent

arXiv:2510.08145v2 Announce Type: replace Abstract: Large Language Models (LLMs) as automatic evaluators, commonly referred to as LLM-as-a-Judge, have also attracted growing attention. This approach p

safetyarxiv-cs-cl
22 Apr 2026
Research

Model-Agnostic Meta Learning for Class Imbalance Adaptation

DGX agent

arXiv:2604.18759v1 Announce Type: new Abstract: Class imbalance is a widespread challenge in NLP tasks, significantly hindering robust performance across diverse domains and applications. We introduce

researcharxiv-cs-cl
22 Apr 2026
Safety

Multi-Task Reinforcement Learning for Enhanced Multimodal LLM-as-a-Judge

DGX agent

arXiv:2603.11665v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have been widely adopted as MLLM-as-a-Judges due to their strong alignment with human judgment across vario

safetyarxiv-cs-cl
22 Apr 2026
← Previous
1…128129130131132…161
Next →