AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
28 Apr 2026

In-Sync: Adaptation of Speech Aware Large Language Models for ASR with Word Level Timestamp Predictions

SafetyDGX agent

arXiv:2604.22817v1 Announce Type: cross Abstract: Recent advances in speech-aware language models have coupled strong acoustic encoders with large language models, enabling systems that move beyond tr

Indirect Question Answering in English, German and Bavarian: A Challenging Task for High- and Low-Resource Languages Alike

ResearchDGX agent

arXiv:2603.15130v2 Announce Type: replace Abstract: Indirectness is a common feature of daily communication, yet is underexplored in NLP research for both low-resource as well as high-resource languag

Investigating the Representation of Backchannels and Fillers in Fine-tuned Language Models

TutorialsDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2509.20237v2 Announce Type: replace Abstract: Backchannels and fillers are important linguistic expressions in dialogue, but often treated as 'noise' to be bypassed in modern transformer-based l

IRIS: Interleaved Reinforcement with Incremental Staged Curriculum for Cross-Lingual Mathematical Reasoning

SafetyDGX agent

arXiv:2604.24114v1 Announce Type: new Abstract: Curriculum learning helps language models tackle complex reasoning by gradually increasing task difficulty. However, it often fails to generate consiste

JudgeSense: A Benchmark for Prompt Sensitivity in LLM-as-a-Judge Systems

Model ReleasesDGX agent

arXiv:2604.23478v1 Announce Type: new Abstract: Large language models are increasingly deployed as automated judges for evaluating other models, yet the stability of their verdicts under semantically

Knowledge Vector of Logical Reasoning in Large Language Models

ResearchDGX agent

arXiv:2604.23877v1 Announce Type: new Abstract: Logical reasoning serve as a central capability in LLMs and includes three main forms: deductive, inductive, and abductive reasoning. In this work, we s

Large language model-enabled automated data extraction for concrete materials informatics

ResearchDGX agent

arXiv:2604.22938v1 Announce Type: cross Abstract: The promise of data-driven materials discovery remains constrained by the scarcity of large, high-quality, and accessible experimental datasets. Here,

Learning Evidence of Depression Symptoms via Prompt Induction

ResearchDGX agent

arXiv:2604.24376v1 Announce Type: new Abstract: Depression places substantial pressure on mental health services, and many people describe their experiences outside clinical settings in high-volume us

Learning Selective LLM Autonomy from Copilot Feedback in Enterprise Customer Support Workflows

SafetyDGX agent

arXiv:2604.23855v1 Announce Type: new Abstract: We present a deployed system that automates end-to-end customer support workflows inside an enterprise Business Process Management (BPM) platform. The a

LegalDrill: Diagnosis-Driven Synthesis for Legal Reasoning in Small Language Models

ApplicationsDGX agent

arXiv:2604.23809v1 Announce Type: new Abstract: Small language models (SLMs) are promising for real-world deployment due to their efficiency and low operational cost. However, their limited capacity s

LinguDistill: Recovering Linguistic Ability in Vision-Language Models via Selective Cross-Modal Distillation

SafetyDGX agent

arXiv:2604.00829v3 Announce Type: replace-cross Abstract: Adapting pretrained language models (LMs) into vision-language models (VLMs) can degrade their native linguistic capability due to representat

Long-Context Aware Upcycling: A New Frontier for Hybrid LLM Scaling

Model ReleasesDGX agent

arXiv:2604.24715v1 Announce Type: new Abstract: Hybrid sequence models that combine efficient Transformer components with linear sequence modeling blocks are a promising alternative to pure Transforme

LongFlow: Efficient KV Cache Compression for Reasoning Models

Model ReleasesDGX agent

arXiv:2603.11504v2 Announce Type: replace-cross Abstract: Recent reasoning models such as OpenAI-o1 and DeepSeek-R1 have shown strong performance on complex tasks including mathematical reasoning and

Looking for the Bottleneck in Fine-grained Temporal Relation Classification

Model ReleasesDGX agent

arXiv:2604.24620v1 Announce Type: new Abstract: Temporal relation classification is the task of determining the temporal relation between pairs of temporal entities in a text. Despite recent advanceme

Making Dialogue Grounding Data Rich: A Three-Tier Data Synthesis Framework for Generalized Referring Expression Comprehension

ResearchDGX agent

arXiv:2512.02791v2 Announce Type: replace Abstract: Dialogue-Based Generalized Referring Expression Comprehension (GREC) requires models to ground the expression and unlimited targets in complex visua

Measuring Temporal Linguistic Emergence in Diffusion Language Models

ResearchDGX agent

arXiv:2604.23235v1 Announce Type: new Abstract: Diffusion language models expose an explicit denoising trajectory, making it possible to ask when different kinds of information become measurable durin

MEG-RAG: Quantifying Multi-modal Evidence Grounding for Evidence Selection in RAG

Model ReleasesDGX agent

arXiv:2604.24564v1 Announce Type: new Abstract: Multimodal Retrieval-Augmented Generation (MRAG) addresses key limitations of Multimodal Large Language Models (MLLMs), such as hallucination and outdat

Mind the Gap: Evaluating Model- and Agentic-Level Vulnerabilities in LLMs with Action Graphs

Model ReleasesDGX agent

arXiv:2509.04802v3 Announce Type: replace Abstract: As large language models increasingly deployed into agentic systems, existing methods face critical gaps in observing, assessing, and mitigating dep

MIPIC: Matryoshka Representation Learning via Self-Distilled Intra-Relational and Progressive Information Chaining

SafetyDGX agent

arXiv:2604.24374v1 Announce Type: new Abstract: Representation learning is fundamental to NLP, but building embeddings that work well at different computational budgets is challenging. Matryoshka Repr

Multimodal QUD: Inquisitive Questions from Scientific Figures

ResearchDGX agent

arXiv:2604.23733v1 Announce Type: new Abstract: Asking inquisitive questions while reading, and looking for their answers, is an important part in human discourse comprehension, curiosity, and creativ

Neural Grammatical Error Correction for Romanian

ResearchDGX agent

arXiv:2604.23627v1 Announce Type: new Abstract: Resources for Grammatical Error Correction (GEC) in non-English languages are scarce, while available spellcheckers in these languages are mostly limite

OLaPh: Optimal Language Phonemizer

Model ReleasesDGX agent

arXiv:2509.20086v2 Announce Type: replace Abstract: Phonemization is a critical component in text-to-speech synthesis. Traditional approaches rely on deterministic transformations and lexica, while ne

On Emergent Social World Models -- Evidence for Functional Integration of Theory of Mind and Pragmatic Reasoning in Language Models

ResearchDGX agent

arXiv:2602.10298v2 Announce Type: replace Abstract: This paper investigates whether LMs recruit shared computational mechanisms for general Theory of Mind (ToM) and language-specific pragmatic reasoni

One Size Fits None: Heuristic Collapse in LLM Investment Advice

ApplicationsDGX agent

arXiv:2604.23837v1 Announce Type: new Abstract: Large language models are increasingly deployed as advisors in high-stakes domains -- answering medical questions, interpreting legal documents, recomme

OS-SPEAR: A Toolkit for the Safety, Performance,Efficiency, and Robustness Analysis of OS Agents

SafetyDGX agent

arXiv:2604.24348v1 Announce Type: new Abstract: The evolution of Multimodal Large Language Models (MLLMs) has shifted the focus from text generation to active behavioral execution, particularly via OS

Overcoming Copyright Barriers in Corpus Distribution Through Non-Reversible Hashing

SafetyDGX agent

arXiv:2604.23412v1 Announce Type: new Abstract: While annotated corpora are crucial in the field of natural language processing (NLP), those containing copyrighted material are difficult to exchange a

PeeriScope: A Multi-Faceted Framework for Evaluating Peer Review Quality

ResearchDGX agent

arXiv:2604.24071v1 Announce Type: new Abstract: The increasing scale and variability of peer review in scholarly venues has created an urgent need for systematic, interpretable, and extensible tools t

Personality Shapes Gender Bias in Persona-Conditioned LLM Narratives Across English and Hindi: An Empirical Investigation

SafetyDGX agent

arXiv:2604.23600v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed in persona-driven applications such as education, customer service, and social platforms, where m

Position: Logical Soundness is not a Reliable Criterion for Neurosymbolic Fact-Checking with LLMs

SafetyDGX agent

arXiv:2604.04177v2 Announce Type: replace Abstract: As large language models (LLMs) are increasing integrated into fact-checking pipelines, formal logic is often proposed as a rigorous means by which

Preserving Long-Tailed Expert Information in Mixture-of-Experts Tuning

SafetyDGX agent

arXiv:2604.23036v1 Announce Type: cross Abstract: Despite MoE models leading many benchmarks, supervised fine-tuning (SFT) for the MoE architectures remains difficult because its router layers are fra

Process Supervision of Confidence Margin for Calibrated LLM Reasoning

ResearchDGX agent

arXiv:2604.23333v1 Announce Type: cross Abstract: Scaling test-time computation with reinforcement learning (RL) has emerged as a reliable path to improve large language models (LLM) reasoning ability

Propagation Structure-Semantic Transfer Learning for Robust Fake News Detection

TutorialsDGX agent

arXiv:2604.23974v1 Announce Type: new Abstract: Fake news generally refers to false information that is spread deliberately to deceive people, which has detrimental social effects. Existing fake news

Psychologically-Grounded Graph Modeling for Interpretable Depression Detection

Model ReleasesDGX agent

arXiv:2604.24126v1 Announce Type: new Abstract: Automatic depression detection from conversational interactions holds significant promise for scalable screening but remains hindered by severe data sca

Rank, Head-Channel Non-Identifiability, and Symmetry Breaking: A Precise Analysis of Representational Collapse in Transformers

Model ReleasesDGX agent

arXiv:2604.23681v1 Announce Type: cross Abstract: A widely cited result by Dong et al. (2021) showed that Transformers built from self-attention alone, without skip connections or feed-forward layers,

Reducing Redundancy in Retrieval-Augmented Generation through Chunk Filtering

ResearchDGX agent

arXiv:2604.24334v1 Announce Type: new Abstract: Standard Retrieval-Augmented Generation (RAG) chunking methods often create excessive redundancy, increasing storage costs and slowing retrieval. This s

Resource-Lean Lexicon Induction for German Dialects

Model ReleasesDGX agent

arXiv:2604.23824v1 Announce Type: new Abstract: Automatic induction of high-quality dictionaries is essential for building lexical resources, yet low-resource languages and dialects pose several chall

Revisiting Greedy Decoding for Visual Question Answering: A Calibration Perspective

ResearchDGX agent

arXiv:2604.23443v1 Announce Type: new Abstract: Stochastic sampling strategies are widely adopted in large language models (LLMs) to balance output coherence and diversity. These heuristics are often

Robust Audio-Text Retrieval via Cross-Modal Attention and Hybrid Loss

Model ReleasesDGX agent

arXiv:2604.23323v1 Announce Type: new Abstract: Audio-text retrieval enables semantic alignment between audio content and natural language queries, supporting applications in multimedia search, access

RouteNLP: Closed-Loop LLM Routing with Conformal Cascading and Distillation Co-Optimization

Model ReleasesDGX agent

arXiv:2604.23577v1 Announce Type: new Abstract: Serving diverse NLP workloads with large language models is costly: at one enterprise partner, inference costs exceeded $200K/month despite over 70% of

SEARCH-R: Structured Entity-Aware Retrieval with Chain-of-Reasoning Navigator for Multi-hop Question Answering

ResearchDGX agent

arXiv:2604.24515v1 Announce Type: new Abstract: Multi-hop Question Answering (MHQA) aims to answer questions that require multi-step reasoning. It presents two key challenges: generating correct reaso

Secure On-Premise Deployment of Open-Weights Large Language Models in Radiology: An Isolation-First Architecture with Prospective Pilot Evaluation

Model ReleasesDGX agent

arXiv:2604.22768v1 Announce Type: cross Abstract: Purpose: To design, implement, evaluate, and report on the regulatory requirements of a self-hosted LLM infrastructure for radiology adhering to the p

Sentiment and Emotion Classification of Indonesian E-Commerce Reviews via Multi-Task BiLSTM and AutoML Benchmarking

ResearchDGX agent

arXiv:2604.24720v1 Announce Type: new Abstract: Indonesian marketplace reviews mix standard vocabulary with slang, regional loanwords, numeric shorthands, and emoji, making lexicon-based sentiment too

ShredBench: Evaluating the Semantic Reasoning Capabilities of Multimodal LLMs in Document Reconstruction

Model ReleasesDGX agent

arXiv:2604.23813v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable performance in Visually Rich Document Understanding (VRDU) tasks, but their capabili

Spectro-Temporal Modulation Representation Framework for Human-Imitated Speech Detection

ResearchDGX agent

arXiv:2604.23241v1 Announce Type: cross Abstract: Human-imitated speech poses a greater challenge than AI-generated speech for both human listeners and automatic detection systems. Unlike AI-generated

Stabilizing Efficient Reasoning with Step-Level Advantage Selection

Model ReleasesDGX agent

arXiv:2604.24003v1 Announce Type: new Abstract: Large language models (LLMs) achieve strong reasoning performance by allocating substantial computation at inference time, often generating long and ver

Stress-Testing Emotional Support Models: Moving from Homogeneous to Diverse Help Seekers

Model ReleasesDGX agent

arXiv:2601.07698v2 Announce Type: replace Abstract: As emotional support chatbots have recently gained significant traction across both research and industry, a common evaluation strategy has emerged:

Structural Pruning of Large Vision Language Models: A Comprehensive Study on Pruning Dynamics, Recovery, and Data Efficiency

Model ReleasesDGX agent

arXiv:2604.24380v1 Announce Type: new Abstract: While Large Vision Language Models (LVLMs) demonstrate impressive capabilities, their substantial computational and memory requirements pose deployment

Supernodes and Halos: Loss-Critical Hubs in LLM Feed-Forward Layers

Model ReleasesDGX agent

arXiv:2604.23475v1 Announce Type: cross Abstract: We study the organization of channel-level importance in transformer feed-forward networks (FFNs). Using a Fisher-style loss proxy (LP) based on activ

Swa-bhasha Resource Hub: Romanized Sinhala to Sinhala Transliteration Systems and Data Resources

ResearchDGX agent

arXiv:2507.09245v2 Announce Type: replace Abstract: The Swa-bhasha Resource Hub provides a comprehensive collection of data resources and algorithms developed for Romanized Sinhala to Sinhala translit

SWE-Pruner: Self-Adaptive Context Pruning for Coding Agents

AgentsDGX agent

arXiv:2601.16746v3 Announce Type: replace-cross Abstract: LLM agents have demonstrated remarkable capabilities in software development, but their performance is hampered by long interaction contexts,

SWE-QA: Can Language Models Answer Repository-level Code Questions?

Model ReleasesDGX agent

arXiv:2509.14635v2 Announce Type: replace Abstract: Understanding and reasoning about entire software repositories is an essential capability for intelligent software engineering tools. While existing

SwissGov-RSD: A Human-annotated, Cross-lingual Benchmark for Token-level Recognition of Semantic Differences Between Related Documents

Model ReleasesDGX agent

arXiv:2512.07538v3 Announce Type: replace Abstract: Recognizing semantic differences across documents is crucial for text generation evaluation and content alignment, especially in cross-lingual setti

Switch Attention: Towards Dynamic and Fine-grained Hybrid Transformers

Model ReleasesDGX agent

arXiv:2603.26380v2 Announce Type: replace Abstract: The attention mechanism has been the core component in modern transformer architectures. However, the computation of standard full attention scales

Synthetic Eggs in Many Baskets: The Impact of Synthetic Data Diversity on LLM Fine-Tuning

SafetyDGX agent

arXiv:2511.01490v2 Announce Type: replace Abstract: As synthetic data becomes widely used in language model development, understanding its impact on model behavior is crucial. This paper investigates

Talker-T2AV: Joint Talking Audio-Video Generation with Autoregressive Diffusion Modeling

ResearchDGX agent

arXiv:2604.23586v1 Announce Type: cross Abstract: Joint audio-video generation models have shown that unified generation yields stronger cross-modal coherence than cascaded approaches. However, existi

TexOCR: Advancing Document OCR Models for Compilable Page-to-LaTeX Reconstruction

Model ReleasesDGX agent

arXiv:2604.22880v1 Announce Type: new Abstract: Existing document OCR largely targets plain text or Markdown, discarding the structural and executable properties that make LaTeX essential for scientif

The Chameleon's Limit: Investigating Persona Collapse and Homogenization in Large Language Models

AgentsDGX agent

arXiv:2604.24698v1 Announce Type: new Abstract: Applications based on large language models (LLMs), such as multi-agent simulations, require population diversity among agents. We identify a pervasive

The Collapse of Heterogeneity in Silicon Philosophers

SafetyDGX agent

arXiv:2604.23575v1 Announce Type: cross Abstract: Silicon samples are increasingly used as a low-cost substitute for human panels and have been shown to reproduce aggregate human opinion with high fid

The Limits of Artificial Companionship

ApplicationsDGX agent

arXiv:2604.23601v1 Announce Type: cross Abstract: This Article argues that conversations with companion chatbot should be subject to a clear structural distinction between commercial and non-commercia

The Surprising Effectiveness of Membership Inference with Simple N-Gram Coverage

Model ReleasesDGX agent

arXiv:2508.09603v2 Announce Type: replace Abstract: Membership inference attacks serves as useful tool for fair use of language models, such as detecting potential copyright infringement and auditing

← Previous
1…979899100101…129
Next →