AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
Research

Latent Abstraction for Retrieval-Augmented Generation

DGX agent

arXiv:2604.17866v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) has become a standard approach for enhancing large language models (LLMs) with external knowledge, mitigating hallu

researcharxiv-cs-cl
21 Apr 2026
Research
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Latent Phase-Shift Rollback: Inference-Time Error Correction via Residual Stream Monitoring and KV-Cache Steering

DGX agent

arXiv:2604.18567v1 Announce Type: cross Abstract: Large language models frequently commit unrecoverable reasoning errors mid-generation: once a wrong step is taken, subsequent tokens compound the mist

researcharxiv-cs-cl
21 Apr 2026
Model Releases

Latent Preference Modeling for Cross-Session Personalized Tool Calling

DGX agent

arXiv:2604.17886v1 Announce Type: new Abstract: Users often omit essential details in their requests to LLM-based agents, resulting in under-specified inputs for tool use. This poses a fundamental cha

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

LEAF: Knowledge Distillation of Text Embedding Models with Teacher-Aligned Representations

DGX agent

arXiv:2509.12539v2 Announce Type: replace-cross Abstract: We present LEAF ('Lightweight Embedding Alignment Framework'), a knowledge distillation framework for text embedding models. A key distinguish

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Learning to Control Summaries with Score Ranking

DGX agent

arXiv:2604.17197v1 Announce Type: new Abstract: Recent advances in summarization research focus on improving summary quality across multiple criteria, such as completeness, conciseness, and faithfulne

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Learning to Retrieve User History and Generate User Profiles for Personalized Persuasiveness Prediction

DGX agent

arXiv:2601.05654v3 Announce Type: replace Abstract: Estimating the persuasiveness of messages is critical in various applications, from recommender systems to safety assessment of LLMs. While it is im

model-releasesarxiv-cs-cl
21 Apr 2026
Local Ai

Learning to Seek Help: Dynamic Collaboration Between Small and Large Language Models

DGX agent

arXiv:2604.17827v1 Announce Type: new Abstract: Large language models (LLMs) offer strong capabilities but raise cost and privacy concerns, whereas small language models (SLMs) facilitate efficient an

local-aiarxiv-cs-cl
21 Apr 2026
Safety

Less Noise, More Voice: Reinforcement Learning for Reasoning via Instruction Purification

DGX agent

arXiv:2601.21244v3 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has advanced LLM reasoning, but remains constrained by inefficient exploration under lim

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

Leveraging Large Language Models for Sarcastic Speech Annotation in Sarcasm Detection

DGX agent

arXiv:2506.00955v2 Announce Type: replace Abstract: Sarcasm fundamentally alters meaning through tone and context, yet detecting it in speech remains a challenge due to data scarcity. In addition, exi

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

LexRel: Benchmarking Legal Relation Extraction for Chinese Civil Cases

DGX agent

arXiv:2512.12643v2 Announce Type: replace Abstract: Legal relations serve as an important analytical framework for dispute resolution in civil cases. However, legal relations in Chinese civil cases re

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

LiFT: Does Instruction Fine-Tuning Improve In-Context Learning for Longitudinal Modelling by Large Language Models?

DGX agent

arXiv:2604.16382v1 Announce Type: new Abstract: Longitudinal NLP tasks require reasoning over temporally ordered text to detect persistence and change in human behavior and opinions. However, in-conte

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

LIFT the Veil for the Truth: Principal Weights Emerge after Rank Reduction for Reasoning-Focused Supervised Fine-Tuning

DGX agent

arXiv:2506.00772v2 Announce Type: replace-cross Abstract: Recent studies have shown that supervised fine-tuning of LLMs on a small number of high-quality datasets can yield strong reasoning capabiliti

model-releasesarxiv-cs-cl
21 Apr 2026
Research

Lil: Less is Less When Applying Post-Training Sparse-Attention Algorithms in Long-Decode Stage

DGX agent

arXiv:2601.03043v3 Announce Type: replace Abstract: Large language models (LLMs) demonstrate strong capabilities across a wide range of complex tasks and are increasingly deployed at scale, placing si

researcharxiv-cs-cl
21 Apr 2026
Research

Linear-Time and Constant-Memory Text Embeddings Based on Recurrent Language Models

DGX agent

arXiv:2604.18199v1 Announce Type: new Abstract: Transformer-based embedding models suffer from quadratic computational and linear memory complexity, limiting their utility for long sequences. We propo

researcharxiv-cs-cl
21 Apr 2026
Model Releases

LiveFact: A Dynamic, Time-Aware Benchmark for LLM-Driven Fake News Detection

DGX agent

arXiv:2604.04815v2 Announce Type: replace Abstract: The rapid development of Large Language Models (LLMs) has transformed fake news detection and fact-checking tasks from simple classification to comp

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Lizard: An Efficient Linearization Framework for Large Language Models

DGX agent

arXiv:2507.09025v4 Announce Type: replace Abstract: We propose Lizard, a linearization framework that transforms pretrained Transformer-based Large Language Models (LLMs) into subquadratic architectur

model-releasesarxiv-cs-cl
21 Apr 2026
Research

LLM as Graph Kernel: Rethinking Message Passing on Text-Rich Graphs

DGX agent

arXiv:2603.14937v2 Announce Type: replace-cross Abstract: Text-rich graphs, which integrate complex structural dependencies with abundant textual information, are ubiquitous yet remain challenging for

researcharxiv-cs-cl
21 Apr 2026
Research

LLM Hypnosis: Exploiting User Feedback for Unauthorized Knowledge Injection to All Users

DGX agent

arXiv:2507.02850v3 Announce Type: replace Abstract: We describe a vulnerability in language models (LMs) trained with user feedback, whereby a single user can persistently alter LM knowledge and behav

researcharxiv-cs-cl
21 Apr 2026
Research

LLMAR: A Tuning-Free Recommendation Framework for Sparse and Text-Rich Industrial Domains

DGX agent

arXiv:2604.16379v1 Announce Type: cross Abstract: Industrial B2B applications (e.g., construction site risk prediction, material procurement) face extreme data sparsity yet feature rich textual intera

researcharxiv-cs-cl
21 Apr 2026
Model Releases

LOGICAL-COMMONSENSEQA: A Benchmark for Logical Commonsense Reasoning

DGX agent

arXiv:2601.16504v3 Announce Type: replace Abstract: Commonsense reasoning often involves evaluating multiple plausible interpretations rather than selecting a single atomic answer, yet most benchmarks

model-releasesarxiv-cs-cl
21 Apr 2026
Research

Logical Computational Linguistics

DGX agent

arXiv:2604.17346v1 Announce Type: new Abstract: In this book we promote logical computational linguistics as opposed to statistical computational linguistics. In particular, we provide a logical seman

researcharxiv-cs-cl
21 Apr 2026
Research

LogicDiff: Logic-Guided Denoising Improves Zero-Shot Reasoning in Masked Diffusion Language Models

DGX agent

arXiv:2603.26771v2 Announce Type: replace Abstract: Masked diffusion language models (MDLMs) generate text by iteratively unmasking tokens from a fully masked sequence. Their standard confidence-based

researcharxiv-cs-cl
21 Apr 2026
Model Releases

Logit Arithmetic Elicits Long Reasoning Capabilities Without Training

DGX agent

arXiv:2510.09354v2 Announce Type: replace Abstract: Large reasoning models exhibit long chain-of-thought reasoning with complex strategies such as backtracking and self-verification. Yet, these capabi

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

LoRA on the Go: Instance-level Dynamic LoRA Selection and Merging

DGX agent

arXiv:2511.07129v3 Announce Type: replace Abstract: Low-Rank Adaptation (LoRA) has emerged as a parameter-efficient approach for fine-tuning large language models. However, conventional LoRA adapters

model-releasesarxiv-cs-cl
21 Apr 2026
Research

LQM: Linguistically Motivated Multidimensional Quality Metrics for Machine Translation

DGX agent

arXiv:2604.18490v1 Announce Type: new Abstract: Existing MT evaluation frameworks, including automatic metrics and human evaluation schemes such as Multidimensional Quality Metrics (MQM), are largely

researcharxiv-cs-cl
21 Apr 2026
Research

LTRR: Learning To Rank Retrievers for LLMs

DGX agent

arXiv:2506.13743v2 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) systems typically rely on a single fixed retriever, despite growing evidence that no single retriever performs

researcharxiv-cs-cl
21 Apr 2026
Model Releases

ltzGLUE: Luxembourgish General Language Understanding Evaluation

DGX agent

arXiv:2604.17976v1 Announce Type: new Abstract: This paper presents ltzGLUE, the first Natural Language Understanding (NLU) benchmark for Luxembourgish (LTZ) based on the popular GLUE benchmark for En

model-releasesarxiv-cs-cl
21 Apr 2026
Research

LVLMs and Humans Ground Differently in Referential Communication

DGX agent

arXiv:2601.19792v3 Announce Type: replace Abstract: For generative AI agents to partner effectively with human users, the ability to accurately predict human intent is critical. But this ability to co

researcharxiv-cs-cl
21 Apr 2026
Model Releases

Macaron: Controlled, Human-Written Benchmark for Multilingual and Multicultural Reasoning via Template-Filling

DGX agent

arXiv:2602.10732v2 Announce Type: replace Abstract: Multilingual benchmarks rarely test reasoning over culturally grounded premises: translated datasets keep English-centric scenarios, while culture-f

model-releasesarxiv-cs-cl
21 Apr 2026
Tutorials

MAPLE: A Meta-learning Framework for Cross-Prompt Essay Scoring

DGX agent

arXiv:2604.17569v1 Announce Type: new Abstract: Automated Essay Scoring (AES) faces significant challenges in cross-prompt settings, where models must generalize to unseen writing prompts. To address

tutorialsarxiv-cs-cl
21 Apr 2026
Research

Mapping Election Toxicity on Social Media across Issue, Ideology, and Psychosocial Dimensions

DGX agent

arXiv:2604.16765v1 Announce Type: cross Abstract: Online political hostility is pervasive, yet it remains unclear how toxicity varies across campaign issues and political ideology, and what psychosoci

researcharxiv-cs-cl
21 Apr 2026
Research

MARA: A Multimodal Adaptive Retrieval-Augmented Framework for Document Question Answering

DGX agent

arXiv:2604.16313v1 Announce Type: cross Abstract: Retrieval-based multimodal document QA aims to identify and integrate relevant information from visually rich documents with complex multimodal struct

researcharxiv-cs-cl
21 Apr 2026
Agents

MASS-RAG: Multi-Agent Synthesis Retrieval-Augmented Generation

DGX agent

arXiv:2604.18509v1 Announce Type: new Abstract: Large language models (LLMs) are widely used in retrieval-augmented generation (RAG) to incorporate external knowledge at inference time. However, when

agentsarxiv-cs-cl
21 Apr 2026
Agents

Matrix: Peer-to-Peer Multi-Agent Synthetic Data Generation Framework

DGX agent

arXiv:2511.21686v2 Announce Type: replace Abstract: Synthetic data has become increasingly important for training large language models, especially when real data is scarce, expensive, or privacy-sens

agentsarxiv-cs-cl
21 Apr 2026
Model Releases

Maximizing Local Entropy Where It Matters: Prefix-Aware Localized LLM Unlearning

DGX agent

arXiv:2601.03190v3 Announce Type: replace Abstract: Machine unlearning aims to forget sensitive knowledge from Large Language Models (LLMs) while maintaining general utility. However, existing approac

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

MeasHalu: Mitigation of Scientific Measurement Hallucinations for Large Language Models with Enhanced Reasoning

DGX agent

arXiv:2604.16929v1 Announce Type: new Abstract: The accurate extraction of scientific measurements from literature is a critical yet challenging task in AI4Science, enabling large-scale analysis and i

model-releasesarxiv-cs-cl
21 Apr 2026
Applications

Measuring Distribution Shift in User Prompts and Its Effects on LLM Performance

DGX agent

arXiv:2604.17650v1 Announce Type: new Abstract: LLMs are increasingly deployed in dynamic, real-world settings, where the distribution of user prompts can shift substantially over time as new tasks, p

applicationsarxiv-cs-cl
21 Apr 2026
Model Releases

Measuring Representation Robustness in Large Language Models for Geometry

DGX agent

arXiv:2604.16421v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly evaluated on mathematical reasoning, yet their robustness to equivalent problem representations remains po

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Measuring Social Bias in Vision-Language Models with Face-Only Counterfactuals from Real Photos

DGX agent

arXiv:2601.06931v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) are increasingly deployed in socially consequential settings, raising concerns about social bias driven by demog

model-releasesarxiv-cs-cl
21 Apr 2026
Research

Measuring the Gap Between Media Coverage and Public Information Demand: Evidence from the 2026 Lebanon Conflict

DGX agent

arXiv:2604.16417v1 Announce Type: cross Abstract: This study examines the relationship between media coverage and public information demand during the Lebanon conflict in March 2026. Using a dataset o

researcharxiv-cs-cl
21 Apr 2026
Model Releases

Medical thinking with multiple images

DGX agent

arXiv:2604.16506v1 Announce Type: cross Abstract: Large language models perform well on many medical QA benchmarks, but real clinical reasoning often requires integrating evidence across multiple imag

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

MedPRMBench: A Fine-grained Benchmark for Process Reward Models in Medical Reasoning

DGX agent

arXiv:2604.17282v1 Announce Type: new Abstract: Process-Level Reward Models (PRMs) are essential for guiding complex reasoning in large language models, yet existing PRM benchmarks cover only general

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

MedRedFlag: Investigating how LLMs Redirect Misconceptions in Real-World Health Communication

DGX agent

arXiv:2601.09853v2 Announce Type: replace Abstract: Real-world health questions from patients often unintentionally embed false assumptions or premises. In such cases, safe medical communication typic

model-releasesarxiv-cs-cl
21 Apr 2026
Research

MegaRAG: Multimodal Knowledge Graph-Based Retrieval Augmented Generation

DGX agent

arXiv:2512.20626v2 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) enables large language models (LLMs) to dynamically access external information, which is powerful for an

researcharxiv-cs-cl
21 Apr 2026
Model Releases

MemBuilder: Reinforcing LLMs for Long-Term Memory Construction via Attributed Dense Rewards

DGX agent

arXiv:2601.05488v3 Announce Type: replace Abstract: Maintaining consistency in long-term dialogues remains a fundamental challenge for LLMs, as standard retrieval mechanisms often fail to capture the

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

MetaLint: Easy-to-Hard Generalization for Code Linting

DGX agent

arXiv:2507.11687v4 Announce Type: replace-cross Abstract: Large language models excel at code generation but struggle with code linting, particularly in generalizing to unseen or evolving best practic

model-releasesarxiv-cs-cl
21 Apr 2026
Tutorials

MetaMem: Evolving Meta-Memory for Knowledge Utilization through Self-Reflective Symbolic Optimization

DGX agent

arXiv:2602.11182v2 Announce Type: replace Abstract: Existing memory systems enable Large Language Models (LLMs) to support long-horizon human-LLM interactions by persisting historical interactions bey

tutorialsarxiv-cs-cl
21 Apr 2026
Safety

MHSafeEval: Role-Aware Interaction-Level Evaluation of Mental Health Safety in Large Language Models

DGX agent

arXiv:2604.17730v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly explored as scalable tools for mental health counseling, yet evaluating their safety remains challenging d

safetyarxiv-cs-cl
21 Apr 2026
← Previous
1…135136137138139…161
Next →