AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
26 May 2026

WhisTLE: Deeply Supervised, Text-Only Domain Adaptation for Pretrained Speech Recognition Transformers

ApplicationsDGX agent

arXiv:2509.10452v2 Announce Type: replace Abstract: Pretrained automatic speech recognition (ASR) models such as Whisper perform well but still need domain adaptation to handle unseen parlance. In man

WhoSaidIt: Human-LLM Collaborative Annotation for Text-Based Multilingual Speaker-Attribute Classification

Model ReleasesDGX agent

arXiv:2605.26070v1 Announce Type: new Abstract: Annotating speaker attributes from text is inherently ambiguous, particularly in multilingual settings where demographic and social cues are implicit an

WISE: Web Information Satire and Fakeness Evaluation

ApplicationsDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2512.24000v3 Announce Type: replace Abstract: Distinguishing fake or untrue news from satire or humor poses a unique challenge due to their overlapping linguistic features and divergent intent.

Word Class Representations Spontaneously Emerge from Successor Representations Trained on Natural Language

ResearchDGX agent

arXiv:2605.24585v1 Announce Type: new Abstract: Language models are typically trained to predict the next token in a sequence. Here, we explore an alternative predictive principle from reinforcement l

25 May 2026

A Comparative Evaluation of Structural Topic Models and BERTopic for Short, Open-Ended Survey Responses

Model ReleasesDGX agent

arXiv:2605.23093v1 Announce Type: new Abstract: Topic modeling in applied psychology increasingly spans two methodological traditions: probabilistic bag-of-words models and newer embedding-based appro

A graph-based analysis of semantic types and coercion in contextualized word embeddings

ResearchDGX agent

arXiv:2605.23710v1 Announce Type: new Abstract: Semantic type mismatch between a noun and its context is central to coercion phenomena. This paper introduces a graph-based method to examine how lexica

A Reproducible Universal Dependencies-Style Pipeline for Katharevousa Greek Parliamentary Text

Model ReleasesDGX agent

arXiv:2605.22978v1 Announce Type: new Abstract: Katharevousa Greek remains poorly served by contemporary NLP pipelines despite its importance for legal, administrative, and parliamentary archives. We

A Survey of Text and Speech Resources for Hausa and Fongbe: Availability, Quality, and Gaps for NLP Development

ResearchDGX agent

arXiv:2605.22828v1 Announce Type: new Abstract: This survey provides a comprehensive catalog of publicly available text and speech resources for two West African languages: Hausa, an Afroasiatic langu

AI-Friendly LaTeX: Using LaTeX Code as a Knowledge Source for Retrieval-Augmented Generation

ResearchDGX agent

arXiv:2605.22923v1 Announce Type: cross Abstract: Large language models can answer questions about textbooks, lecture notes, and programming exercises more reliably when their answers are grounded in

AraHopeCorpus: Annotation Guidelines and Dataset for Hope Speech in Arabic Social Media Crisis Discourse

Model ReleasesDGX agent

arXiv:2605.23325v1 Announce Type: new Abstract: Social media has become a crucial arena for shaping public narratives during armed conflicts, providing space for both harmful and constructive communic

ARES: Automated Rubric Synthesis for Scalable LLM Reinforcement Learning

ApplicationsDGX agent

arXiv:2605.23454v1 Announce Type: new Abstract: Rubric-based rewards offer a promising way to extend reinforcement learning (RL) for large language models beyond tasks with automatically verifiable an

Articulatory strategy as a source of variation in acoustic vowel dynamics

ApplicationsDGX agent

arXiv:2605.23416v1 Announce Type: new Abstract: Acoustic vowel dynamics have some speaker-identifying characteristics, which have been ascribed to individual properties of articulatory strategies: for

Asking For An Old Friend: Diagnosing and Mitigating Temporal Failure Modes in LLM-based Statutory Question Answering

Model ReleasesDGX agent

arXiv:2605.23497v1 Announce Type: new Abstract: Large language models are increasingly used for legal research, yet their fixed training cutoffs and reliance on static parametric knowledge are at odds

Benchmarking Gaslighting Attacks Against Speech Large Language Models

ResearchDGX agent

arXiv:2509.19858v2 Announce Type: replace Abstract: As Speech Large Language Models (Speech LLMs) become increasingly integrated into voice-based applications, ensuring their robustness against manipu

Benchmarking Google Embeddings 2 against Open-Source Models for Multilingual Dense Retrieval and RAG Systems

Model ReleasesDGX agent

arXiv:2605.23618v1 Announce Type: new Abstract: We benchmark Google Embeddings (GE2), a Vertex-AI-hosted bi-encoder with 2,048-token context and explicit task-type conditioning, against five open-sour

Beyond Log Likelihood: Probability-Based Objectives for Supervised Fine-Tuning across the Model Capability Continuum

ResearchDGX agent

arXiv:2510.00526v3 Announce Type: replace Abstract: Supervised fine-tuning (SFT) is the standard approach for post-training large language models (LLMs), yet it often shows limited generalization. We

BURMESE-SAN: Burmese NLP Benchmark for Evaluating Large Language Models

Model ReleasesDGX agent

arXiv:2602.18788v3 Announce Type: replace Abstract: We introduce BURMESE-SAN, the first holistic benchmark that systematically evaluates large language models (LLMs) for Burmese across three core NLP

Can AI Guess What You Know? Performance Comparison of Large Language Models for Human Domain Knowledge Estimation From Communication Logs

Model ReleasesDGX agent

arXiv:2605.22971v1 Announce Type: new Abstract: Employees often struggle to identify ``who knows what,'' leading to organizational productivity losses. We investigate whether Large Language Models (LL

ChartFI: Benchmarking Faithfulness and Insightfulness of Chart Descriptions from Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2605.23694v1 Announce Type: new Abstract: Chart descriptions are essential for accessibility, cross-modal retrieval, and assisting readers in extracting insights from complex visualizations. As

ClimateChat-300K: A Multi-Modal Facebook Dataset for Understanding Diverse Perspectives in Climate Communication

SafetyDGX agent

arXiv:2605.23326v1 Announce Type: new Abstract: We present ClimateChat-300K, a large-scale dataset of 299,329 public Facebook posts about climate change collected between May 2020 and May 2024 through

CultivAgents: Cultivating Relationship-Centered Multi-Agent Systems for Personalized Gardening

Local AiDGX agent

arXiv:2605.23193v1 Announce Type: cross Abstract: Gardening is critical to support well-being, cultural continuity, and food autonomy, yet existing digital tools often provide generic advice that over

Cultural Adaptation in Large Language Models for Political Discourse

Model ReleasesDGX agent

arXiv:2605.23332v1 Announce Type: new Abstract: The integration of large language models into political discourse analysis creates new opportunities for comparative research, policy analysis, and civi

Decomposing Queries into Tool Calls for Long-Video Keyframe Retrieval

Model ReleasesDGX agent

arXiv:2605.23826v1 Announce Type: cross Abstract: Keyframe selection is a direct way to provide verifiable visual evidence for long-video question answering (QA). Queries differ in what they require,

DELICATE: Diachronic Entity LInking using Classes And Temporal Evidence

ResearchDGX agent

arXiv:2511.10404v2 Announce Type: replace Abstract: In spite of the remarkable advancements in the field of Natural Language Processing, the task of Entity Linking (EL) remains challenging in the fiel

DFKI-MLT at SemEval-2026 TASK 7: Steering Multilingual Models Towards Cultural Knowledge

Model ReleasesDGX agent

arXiv:2605.23069v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used across diverse linguistic and cultural contexts, yet their cultural knowledge remains uneven across r

Differences in Typological Alignment in Language Models' Treatment of Differential Argument Marking

SafetyDGX agent

arXiv:2602.17653v2 Announce Type: replace Abstract: Recent work has shown that language models (LMs) trained on synthetic corpora can exhibit typological preferences that resemble cross-linguistic reg

Emotion Recognition in Sign Language Conversation

ApplicationsDGX agent

arXiv:2605.23328v1 Announce Type: new Abstract: Emotion Recognition in Conversation is a core component of affective computing, while current resources of sign language emotion datasets primarily focu

Entropy-Aware On-Policy Distillation of Language Models

SafetyDGX agent

arXiv:2603.07079v2 Announce Type: replace-cross Abstract: On-policy distillation is a promising approach for transferring knowledge between language models, where a student learns from dense token-lev

EquiSumm : A Gender Bias-Aware Framework for Inclusive Tweet Summarization

SafetyDGX agent

arXiv:2605.23412v1 Announce Type: new Abstract: While social media platforms, such as Twitter, provide a medium for large-scale opinion sharing during news events, it is manually impossible for indivi

Evaluating Counterfactual Strategic Reasoning in Large Language Models

ResearchDGX agent

arXiv:2603.19167v2 Announce Type: replace Abstract: We evaluate Large Language Models (LLMs) in repeated game-theoretic settings to assess whether strategic performance reflects genuine reasoning or r

Evaluating Customized vs. Generalist Transformer-based Models for Legal Contract Classification

ApplicationsDGX agent

arXiv:2508.07849v2 Announce Type: replace Abstract: Despite advances in legal NLP, no comprehensive evaluation of Transformer-based models customized for legal tasks (referred to as `legal-specific' m

Evaluating Memory Structure in LLM Agents

Model ReleasesDGX agent

arXiv:2602.11243v2 Announce Type: replace-cross Abstract: Modern LLM-based agents and chat assistants rely on long-term memory frameworks to store reusable knowledge, recall user preferences, and augm

Fast-dDrive: Efficient Block-Diffusion VLM for Autonomous Driving

SafetyDGX agent

arXiv:2605.23163v1 Announce Type: new Abstract: End-to-end autonomous driving via Vision-Language-Action (VLA) models demands a precarious balance between high-fidelity trajectory planning and efficie

From Correctness to Preference: A Framework for Personalized Agentic Reinforcement Learning

SafetyDGX agent

arXiv:2605.23382v1 Announce Type: new Abstract: Agentic reinforcement learning (Agentic RL) has achieved strong progress in tasks with clear success signals. However, many real-world agent application

GEMQ: Global Expert-Level Mixed-Precision Quantization for MoE LLMs

ResearchDGX agent

arXiv:2605.23078v1 Announce Type: cross Abstract: Mixture-of-Experts Large Language Models (MoE-LLMs) achieve strong performance but incur substantial memory overhead due to massive expert parameters.

HawkesLLM: Semantic Uncertainty Propagation in Agentic Text Simulation

Local AiDGX agent

arXiv:2605.23043v1 Announce Type: new Abstract: Agentic text-simulation systems write in sequence, with each item becoming possible context for later steps. That makes uncertainty path-dependent: an e

Hidden Human-Like Nature of Machine-Generated Texts: Theory and Detection Enhancement

ResearchDGX agent

arXiv:2605.23190v1 Announce Type: new Abstract: Machine-generated texts (MGTs) produced by large language models (LLMs) are increasingly prevalent across various applications, while their potential mi

Hierarchical Concept Geometry in Language Models Emerges from Word Co-occurrence

Model ReleasesDGX agent

arXiv:2605.23821v1 Announce Type: new Abstract: We propose a distributional theory of how hypernymy -- the ``is-a'' relation between general and specific concepts -- is encoded geometrically in langua

How Far Are We from Generating Missing Modalities with Foundation Models?

Model ReleasesDGX agent

arXiv:2506.03530v3 Announce Type: replace-cross Abstract: Multimodal foundation models have demonstrated impressive capabilities across diverse tasks. However, their potential as plug-and-play solutio

How Human-Like Are Large Language Models? A Register-Aware Linguistic Evaluation Framework

ApplicationsDGX agent

arXiv:2605.23651v1 Announce Type: new Abstract: While factual correctness and task-performance have been in focus of Large Language Model (LLM) research for a long time, the fundamental question of ho

Improving Sampling for Masked Diffusion Models via Information Gain

ResearchDGX agent

arXiv:2602.18176v3 Announce Type: replace Abstract: Masked Diffusion Models (MDMs) enable flexible decoding orders, yet existing samplers remain largely greedy, selecting locally certain tokens withou

InfiGFusion: Graph-on-Logits Distillation via Efficient Gromov-Wasserstein for Model Fusion

ResearchDGX agent

arXiv:2505.13893v2 Announce Type: replace Abstract: Recent advances in large language models (LLMs) have intensified efforts to fuse heterogeneous open-source models into a unified system that inherit

Is a Document Educational or Just Wikipedia-Style? -- Pitfalls of Classifier-Based Quality Filtering

ApplicationsDGX agent

arXiv:2605.23721v1 Announce Type: new Abstract: Classifier-based Quality Filtering has recently emerged as a fundamental technique in constructing pre-training corpora. The ability to deploy a single

Knowledge Distillation for Low-Resource Open-source Text-to-SQL Model

ApplicationsDGX agent

arXiv:2605.22843v1 Announce Type: new Abstract: Text-to-SQL converts natural language questions into executable SQL queries, enabling non-technical users to access relational databases for analytics a

Learnability-Informed Fine-Tuning of Diffusion Language Models

TutorialsDGX agent

arXiv:2605.22939v1 Announce Type: new Abstract: We aim to improve the reasoning capabilities of diffusion language models (DLMs). While SFT is a popular post-training recipe for autoregressive models,

Metadata Predictability Is Not Evidence Dependence: An Intervention-Based Audit for Weak-Label Benchmarks

Model ReleasesDGX agent

arXiv:2605.23701v1 Announce Type: new Abstract: We study a protocol-level test for weak-label benchmarks: whether benchmark outputs change when the provided evidence is intervened on. Metadata-only sh

ModeSwitch-LLM: A Lightweight Phase-Aware Controller for Cross-Mode LLM Inference on a Single GPU

Model ReleasesDGX agent

arXiv:2605.23057v1 Announce Type: cross Abstract: ModeSwitch-LLM is a lightweight request-boundary controller for improving single-GPU large language model inference efficiency by routing each request

Multi-SpatialMLLM: Multi-Frame Spatial Understanding with Multi-Modal Large Language Models

Model ReleasesDGX agent

arXiv:2505.17015v2 Announce Type: replace-cross Abstract: Multi-modal large language models (MLLMs) have rapidly advanced in visual tasks, yet their spatial understanding remains limited to single ima

Multilingual Knowledge Transfer under Data Constraints via Lexical Interventions

ResearchDGX agent

arXiv:2605.23885v1 Announce Type: new Abstract: Cross-lingual knowledge transfer is critical for building high-performing multilingual language models for languages with insufficient training data. Wh

Multilingual Steering by Design: Multilingual Sparse Autoencoders and Principled Layer Selection

Model ReleasesDGX agent

arXiv:2605.23036v1 Announce Type: new Abstract: Sparse autoencoders (SAEs) enable feature-level mechanistic interpretability and activation steering in large language models (LLMs), but SAE-based lang

Naturalistic measure of social norms alignment

SafetyDGX agent

arXiv:2605.23420v1 Announce Type: new Abstract: Social norms reflect shared expectations on acceptable behavior. Measuring social norms alignment remains challenging, with existing approaches typicall

NLG Evaluation: Past, Present, Future

SafetyDGX agent

arXiv:2605.23715v1 Announce Type: new Abstract: Natural Language Generation (NLG) evaluation has changed dramatically since 1990, and will continue to evolve in the future. In 1990, when NLG had close

OpenSkillEval: Automatically Auditing the Open Skill Ecosystem for LLM Agents

Model ReleasesDGX agent

arXiv:2605.23657v1 Announce Type: new Abstract: Skills, i.e., structured workflow instructions distilled for large language models (LLMs), are becoming an increasingly important mechanism for improvin

Pooling and Semantic Shift: The Fundamental Challenges in Long Text Embedding and Retrieval

ResearchDGX agent

arXiv:2603.21437v2 Announce Type: replace Abstract: Transformer-based embedding models frequently exhibit geometric pathologies, such as anisotropy and length-induced representation collapse, which ca

PROGRESSLM: Towards Progress Reasoning in Vision-Language Models

Model ReleasesDGX agent

arXiv:2601.15224v2 Announce Type: replace-cross Abstract: Estimating task progress requires reasoning over long-horizon dynamics rather than recognizing static visual content. While modern Vision-Lang

Query-Adaptive Semantic Chunking for Retrieval-Augmented Generation: A Dynamic Strategy with Contextual Window Expansion

AgentsDGX agent

arXiv:2605.22834v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) systems depend critically on document chunking quality for retrieving relevant context. Fixed chunking segments doc

RADAR: Relative Angular Divergence Across Representations

ResearchDGX agent

arXiv:2605.23028v1 Announce Type: cross Abstract: Machine learning methods rely on data. However, gathering suitable data can be challenging due to availability constraints, cost, or the need for doma

RAS: Reflection-Augmented Scaling with In-Context Learning for Executable Cypher Query Generation

ResearchDGX agent

arXiv:2605.22937v1 Announce Type: new Abstract: Inference-time scaling can reduce errors in structured query generation, but methods to allocate the compute for query code generation remains underexpl

Robust LLM Watermarking with Minimal Semantic Distortion for IP Protection

ResearchDGX agent

arXiv:2605.23175v1 Announce Type: cross Abstract: Proprietary large language models (LLMs) face risks of intellectual property (IP) violation, as adversaries can replicate an LLM by collecting input-o

Same Model, Different Weakness: How Language and Modality Reshape the Jailbreak Attack Surface in Frontier MLLMs

Model ReleasesDGX agent

arXiv:2605.23157v1 Announce Type: new Abstract: The attack surface of a multimodal large language model (MLLM) is language-dependent in ways that reveal the mechanistic structure of alignment failures

← Previous
1…6263646566…129
Next →