AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
21 Apr 2026

Lil: Less is Less When Applying Post-Training Sparse-Attention Algorithms in Long-Decode Stage

ResearchDGX agent

arXiv:2601.03043v3 Announce Type: replace Abstract: Large language models (LLMs) demonstrate strong capabilities across a wide range of complex tasks and are increasingly deployed at scale, placing si

Linear-Time and Constant-Memory Text Embeddings Based on Recurrent Language Models

ResearchDGX agent

arXiv:2604.18199v1 Announce Type: new Abstract: Transformer-based embedding models suffer from quadratic computational and linear memory complexity, limiting their utility for long sequences. We propo

LiveFact: A Dynamic, Time-Aware Benchmark for LLM-Driven Fake News Detection

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.04815v2 Announce Type: replace Abstract: The rapid development of Large Language Models (LLMs) has transformed fake news detection and fact-checking tasks from simple classification to comp

Lizard: An Efficient Linearization Framework for Large Language Models

Model ReleasesDGX agent

arXiv:2507.09025v4 Announce Type: replace Abstract: We propose Lizard, a linearization framework that transforms pretrained Transformer-based Large Language Models (LLMs) into subquadratic architectur

LLM as Graph Kernel: Rethinking Message Passing on Text-Rich Graphs

ResearchDGX agent

arXiv:2603.14937v2 Announce Type: replace-cross Abstract: Text-rich graphs, which integrate complex structural dependencies with abundant textual information, are ubiquitous yet remain challenging for

LLM Hypnosis: Exploiting User Feedback for Unauthorized Knowledge Injection to All Users

ResearchDGX agent

arXiv:2507.02850v3 Announce Type: replace Abstract: We describe a vulnerability in language models (LMs) trained with user feedback, whereby a single user can persistently alter LM knowledge and behav

LLMAR: A Tuning-Free Recommendation Framework for Sparse and Text-Rich Industrial Domains

ResearchDGX agent

arXiv:2604.16379v1 Announce Type: cross Abstract: Industrial B2B applications (e.g., construction site risk prediction, material procurement) face extreme data sparsity yet feature rich textual intera

LOGICAL-COMMONSENSEQA: A Benchmark for Logical Commonsense Reasoning

Model ReleasesDGX agent

arXiv:2601.16504v3 Announce Type: replace Abstract: Commonsense reasoning often involves evaluating multiple plausible interpretations rather than selecting a single atomic answer, yet most benchmarks

Logical Computational Linguistics

ResearchDGX agent

arXiv:2604.17346v1 Announce Type: new Abstract: In this book we promote logical computational linguistics as opposed to statistical computational linguistics. In particular, we provide a logical seman

LogicDiff: Logic-Guided Denoising Improves Zero-Shot Reasoning in Masked Diffusion Language Models

ResearchDGX agent

arXiv:2603.26771v2 Announce Type: replace Abstract: Masked diffusion language models (MDLMs) generate text by iteratively unmasking tokens from a fully masked sequence. Their standard confidence-based

Logit Arithmetic Elicits Long Reasoning Capabilities Without Training

Model ReleasesDGX agent

arXiv:2510.09354v2 Announce Type: replace Abstract: Large reasoning models exhibit long chain-of-thought reasoning with complex strategies such as backtracking and self-verification. Yet, these capabi

LoRA on the Go: Instance-level Dynamic LoRA Selection and Merging

Model ReleasesDGX agent

arXiv:2511.07129v3 Announce Type: replace Abstract: Low-Rank Adaptation (LoRA) has emerged as a parameter-efficient approach for fine-tuning large language models. However, conventional LoRA adapters

LQM: Linguistically Motivated Multidimensional Quality Metrics for Machine Translation

ResearchDGX agent

arXiv:2604.18490v1 Announce Type: new Abstract: Existing MT evaluation frameworks, including automatic metrics and human evaluation schemes such as Multidimensional Quality Metrics (MQM), are largely

LTRR: Learning To Rank Retrievers for LLMs

ResearchDGX agent

arXiv:2506.13743v2 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) systems typically rely on a single fixed retriever, despite growing evidence that no single retriever performs

ltzGLUE: Luxembourgish General Language Understanding Evaluation

Model ReleasesDGX agent

arXiv:2604.17976v1 Announce Type: new Abstract: This paper presents ltzGLUE, the first Natural Language Understanding (NLU) benchmark for Luxembourgish (LTZ) based on the popular GLUE benchmark for En

LVLMs and Humans Ground Differently in Referential Communication

ResearchDGX agent

arXiv:2601.19792v3 Announce Type: replace Abstract: For generative AI agents to partner effectively with human users, the ability to accurately predict human intent is critical. But this ability to co

Macaron: Controlled, Human-Written Benchmark for Multilingual and Multicultural Reasoning via Template-Filling

Model ReleasesDGX agent

arXiv:2602.10732v2 Announce Type: replace Abstract: Multilingual benchmarks rarely test reasoning over culturally grounded premises: translated datasets keep English-centric scenarios, while culture-f

MAPLE: A Meta-learning Framework for Cross-Prompt Essay Scoring

TutorialsDGX agent

arXiv:2604.17569v1 Announce Type: new Abstract: Automated Essay Scoring (AES) faces significant challenges in cross-prompt settings, where models must generalize to unseen writing prompts. To address

Mapping Election Toxicity on Social Media across Issue, Ideology, and Psychosocial Dimensions

ResearchDGX agent

arXiv:2604.16765v1 Announce Type: cross Abstract: Online political hostility is pervasive, yet it remains unclear how toxicity varies across campaign issues and political ideology, and what psychosoci

MARA: A Multimodal Adaptive Retrieval-Augmented Framework for Document Question Answering

ResearchDGX agent

arXiv:2604.16313v1 Announce Type: cross Abstract: Retrieval-based multimodal document QA aims to identify and integrate relevant information from visually rich documents with complex multimodal struct

MASS-RAG: Multi-Agent Synthesis Retrieval-Augmented Generation

AgentsDGX agent

arXiv:2604.18509v1 Announce Type: new Abstract: Large language models (LLMs) are widely used in retrieval-augmented generation (RAG) to incorporate external knowledge at inference time. However, when

Matrix: Peer-to-Peer Multi-Agent Synthetic Data Generation Framework

AgentsDGX agent

arXiv:2511.21686v2 Announce Type: replace Abstract: Synthetic data has become increasingly important for training large language models, especially when real data is scarce, expensive, or privacy-sens

Maximizing Local Entropy Where It Matters: Prefix-Aware Localized LLM Unlearning

Model ReleasesDGX agent

arXiv:2601.03190v3 Announce Type: replace Abstract: Machine unlearning aims to forget sensitive knowledge from Large Language Models (LLMs) while maintaining general utility. However, existing approac

MeasHalu: Mitigation of Scientific Measurement Hallucinations for Large Language Models with Enhanced Reasoning

Model ReleasesDGX agent

arXiv:2604.16929v1 Announce Type: new Abstract: The accurate extraction of scientific measurements from literature is a critical yet challenging task in AI4Science, enabling large-scale analysis and i

Measuring Distribution Shift in User Prompts and Its Effects on LLM Performance

ApplicationsDGX agent

arXiv:2604.17650v1 Announce Type: new Abstract: LLMs are increasingly deployed in dynamic, real-world settings, where the distribution of user prompts can shift substantially over time as new tasks, p

Measuring Representation Robustness in Large Language Models for Geometry

Model ReleasesDGX agent

arXiv:2604.16421v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly evaluated on mathematical reasoning, yet their robustness to equivalent problem representations remains po

Measuring Social Bias in Vision-Language Models with Face-Only Counterfactuals from Real Photos

Model ReleasesDGX agent

arXiv:2601.06931v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) are increasingly deployed in socially consequential settings, raising concerns about social bias driven by demog

Measuring the Gap Between Media Coverage and Public Information Demand: Evidence from the 2026 Lebanon Conflict

ResearchDGX agent

arXiv:2604.16417v1 Announce Type: cross Abstract: This study examines the relationship between media coverage and public information demand during the Lebanon conflict in March 2026. Using a dataset o

Medical thinking with multiple images

Model ReleasesDGX agent

arXiv:2604.16506v1 Announce Type: cross Abstract: Large language models perform well on many medical QA benchmarks, but real clinical reasoning often requires integrating evidence across multiple imag

MedPRMBench: A Fine-grained Benchmark for Process Reward Models in Medical Reasoning

Model ReleasesDGX agent

arXiv:2604.17282v1 Announce Type: new Abstract: Process-Level Reward Models (PRMs) are essential for guiding complex reasoning in large language models, yet existing PRM benchmarks cover only general

MedRedFlag: Investigating how LLMs Redirect Misconceptions in Real-World Health Communication

Model ReleasesDGX agent

arXiv:2601.09853v2 Announce Type: replace Abstract: Real-world health questions from patients often unintentionally embed false assumptions or premises. In such cases, safe medical communication typic

MegaRAG: Multimodal Knowledge Graph-Based Retrieval Augmented Generation

ResearchDGX agent

arXiv:2512.20626v2 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) enables large language models (LLMs) to dynamically access external information, which is powerful for an

MemBuilder: Reinforcing LLMs for Long-Term Memory Construction via Attributed Dense Rewards

Model ReleasesDGX agent

arXiv:2601.05488v3 Announce Type: replace Abstract: Maintaining consistency in long-term dialogues remains a fundamental challenge for LLMs, as standard retrieval mechanisms often fail to capture the

MetaLint: Easy-to-Hard Generalization for Code Linting

Model ReleasesDGX agent

arXiv:2507.11687v4 Announce Type: replace-cross Abstract: Large language models excel at code generation but struggle with code linting, particularly in generalizing to unseen or evolving best practic

MetaMem: Evolving Meta-Memory for Knowledge Utilization through Self-Reflective Symbolic Optimization

TutorialsDGX agent

arXiv:2602.11182v2 Announce Type: replace Abstract: Existing memory systems enable Large Language Models (LLMs) to support long-horizon human-LLM interactions by persisting historical interactions bey

MHSafeEval: Role-Aware Interaction-Level Evaluation of Mental Health Safety in Large Language Models

SafetyDGX agent

arXiv:2604.17730v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly explored as scalable tools for mental health counseling, yet evaluating their safety remains challenging d

Migrant Voices, Local News: Insights on Bridging Community Needs with Media Content

ResearchDGX agent

arXiv:2604.16651v1 Announce Type: new Abstract: Research shows news consumption differs across demographics, yet little is known about non-mainstream audiences, especially in relation to local media.

Mira-Embeddings-V1: Domain-Adapted Semantic Reranking for Recruitment via LLM-Synthesized Data

Local AiDGX agent

arXiv:2604.17738v1 Announce Type: new Abstract: Candidate sourcing for recruiters is best viewed as a two-stage retrieval and reranking pipeline with recall as the primary objective under a limited re

Missing-by-Design: Certifiable Modality Deletion for Revocable Multimodal Sentiment Analysis

Model ReleasesDGX agent

arXiv:2602.16144v3 Announce Type: replace Abstract: As multimodal systems increasingly process sensitive personal data, the ability to selectively revoke specific data modalities has become a critical

Mitigating Multimodal Hallucination via Phase-wise Self-reward

ResearchDGX agent

arXiv:2604.17982v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) still struggle with vision hallucination, where generated responses are inconsistent with the visual input. Exist

Mix and Match: Context Pairing for Scalable Topic-Controlled Educational Summarisation

SafetyDGX agent

arXiv:2604.18087v1 Announce Type: new Abstract: Topic-controlled summarisation enables users to generate summaries focused on specific aspects of source documents. This paper investigates a data augme

MM-JudgeBias: A Benchmark for Evaluating Compositional Biases in MLLM-as-a-Judge

Model ReleasesDGX agent

arXiv:2604.18164v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have been increasingly used as automatic evaluators-a paradigm known as MLLM-as-a-Judge. However, their reliabi

MNAFT: modality neuron-aware fine-tuning of multimodal large language models for image translation

Model ReleasesDGX agent

arXiv:2604.16943v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have shown impressive capabilities, yet they often struggle to effectively capture the fine-grained textual inf

MoCo: A One-Stop Shop for Model Collaboration Research

SafetyDGX agent

arXiv:2601.21257v2 Announce Type: replace Abstract: Advancing beyond single monolithic language models (LMs), recent research increasingly recognizes the importance of model collaboration, where multi

Model in Distress: Sentiment Analysis on French Synthetic Social Media

Model ReleasesDGX agent

arXiv:2604.18226v1 Announce Type: new Abstract: Automated analysis of customer feedback on social media is hindered by three challenges: the high cost of annotated training data, the scarcity of evalu

Modeling Human Perspectives with Socio-Demographic Representations

ApplicationsDGX agent

arXiv:2604.18069v1 Announce Type: new Abstract: Humans often hold different perspectives on the same issues. In many NLP tasks, annotation disagreement can reflect valid subjective perspectives. Model

Modeling Multi-Dimensional Cognitive States in Large Language Models under Cognitive Crowding

Model ReleasesDGX agent

arXiv:2604.17174v1 Announce Type: new Abstract: Modeling human cognitive states is essential for advanced artificial intelligence. Existing Large Language Models (LLMs) mainly address isolated tasks s

Modeling Multiple Support Strategies within a Single Turn for Emotional Support Conversations

Model ReleasesDGX agent

arXiv:2604.17972v1 Announce Type: new Abstract: Emotional Support Conversation (ESC) aims to assist individuals experiencing distress by generating empathetic and supportive dialogue. While prior work

Modular Representation Compression: Adapting LLMs for Efficient and Effective Recommendations

TutorialsDGX agent

arXiv:2604.18146v1 Announce Type: cross Abstract: Recently, large language models (LLMs) have advanced recommendation systems (RSs), and recent works have begun to explore how to integrate LLMs into i

MoE-nD: Per-Layer Mixture-of-Experts Routing for Multi-Axis KV Cache Compression

ResearchDGX agent

arXiv:2604.17695v1 Announce Type: cross Abstract: KV cache memory is the dominant bottleneck for long-context LLM inference. Existing compression methods each act on a single axis of the four-dimensio

More Than Meets the Eye: Measuring the Semiotic Gap in Vision-Language Models via Semantic Anchorage

Model ReleasesDGX agent

arXiv:2604.17354v1 Announce Type: new Abstract: Vision-Language Models (VLMs) excel at photorealistic generation, yet often struggle to represent abstract meaning such as idiomatic interpretations of

More Than Sum of Its Parts: Deciphering Intent Shifts in Multimodal Hate Speech Detection

Model ReleasesDGX agent

arXiv:2603.21298v2 Announce Type: replace Abstract: Combating hate speech on social media is critical for securing cyberspace, yet relies heavily on the efficacy of automated detection systems. As con

MoRI: Learning Motivation-Grounded Reasoning for Scientific Ideation in Large Language Models

AgentsDGX agent

arXiv:2603.19044v2 Announce Type: replace Abstract: Scientific ideation aims to propose novel solutions within a given scientific context. Existing LLM-based agentic approaches emulate human research

MoVE: Translating Laughter and Tears via Mixture of Vocalization Experts in Speech-to-Speech Translation

ApplicationsDGX agent

arXiv:2604.17435v1 Announce Type: new Abstract: Recent Speech-to-Speech Translation (S2ST) systems achieve strong semantic accuracy yet consistently strip away non-verbal vocalizations (NVs), such as

MTSQL-R1: Towards Long-Horizon Multi-Turn Text-to-SQL via Agentic Training

Model ReleasesDGX agent

arXiv:2510.12831v3 Announce Type: replace Abstract: Multi-turn Text-to-SQL aims to translate a user's conversational utterances into executable SQL while preserving dialogue coherence and grounding to

Multilingual Training and Evaluation Resources for Vision-Language Models

ResearchDGX agent

arXiv:2604.18347v1 Announce Type: new Abstract: Vision Language Models (VLMs) achieved rapid progress in the recent years. However, despite their growth, VLMs development is heavily grounded on Englis

Multimodal Claim Extraction for Fact-Checking

Model ReleasesDGX agent

arXiv:2604.16311v1 Announce Type: new Abstract: Automated Fact-Checking (AFC) relies on claim extraction as a first step, yet existing methods largely overlook the multimodal nature of today's misinfo

Multimodal In-context Learning for ASR of Low-resource Languages

Model ReleasesDGX agent

arXiv:2601.05707v2 Announce Type: replace Abstract: Automatic speech recognition (ASR) still covers only a small fraction of the world's languages, mainly due to supervised data scarcity. In-context l

Multimodal Policy Internalization for Conversational Agents

SafetyDGX agent

arXiv:2510.09474v2 Announce Type: replace Abstract: Modern conversational agents like ChatGPT and Alexa+ rely on predefined policies specifying metadata, response styles, and tool-usage rules. As thes

Multimodal Sentiment Analysis with Missing Modality: A Knowledge-Transfer Approach

ResearchDGX agent

arXiv:2401.10747v5 Announce Type: replace-cross Abstract: Multimodal sentiment analysis aims to identify the emotions expressed by individuals through visual, language, and acoustic cues. However, mos

← Previous
1…108109110111112…129
Next →