AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
20 May 2026

SETUP: Sentence-level English-To-Uniform Meaning Representation Parser

ApplicationsDGX agent

arXiv:2512.07068v3 Announce Type: replace Abstract: Uniform Meaning Representation (UMR) is a novel graph-based semantic representation which captures the core meaning of a text, with flexibility inco

SLoW: Select Low-frequency Words! Automatic Dictionary Selection for Translation on Large Language Models

Model ReleasesDGX agent

arXiv:2507.18902v2 Announce Type: replace Abstract: There are more than 7,000 languages around the world, and current Large Language Models (LLMs) only support hundreds of languages. Dictionary-based

SPHERICAL KV: Angle-Domain Attention and Rate-Distortion Retention for Efficient Long-Context Inference

ResearchDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.18856v1 Announce Type: cross Abstract: Long-context inference is increasingly constrained by the KV cache: resident memory grows with context length, and decoding becomes limited by repeate

Structured Style-Rewrite with Chain-of-Thought Planning for Low-Resource Character Dialogue

ResearchDGX agent

arXiv:2603.05933v2 Announce Type: replace Abstract: Applying Small Language Models (SLMs) to Chinese character-driven generation remains challenging due to data scarcity and the difficulty of disentan

Taming the Thinker: Conditional Entropy Shaping for Adaptive LLM Reasoning

Model ReleasesDGX agent

arXiv:2605.19358v1 Announce Type: new Abstract: Entropy-based deep reasoning has emerged as a promising direction for improving the reasoning capabilities of Large Language Models (LLMs), but existing

Text-to-SPARQL Generation with Reinforcement Learning: A GRPO-based Approach on DBLP

SafetyDGX agent

arXiv:2605.20066v1 Announce Type: new Abstract: Knowledge graph question answering seeks to translate natural language questions into executable queries over knowledge graphs, but existing approaches

The Annotation Scarcity Paradox in Low-Resource NLP Evaluation: A Decade of Acceleration and Emerging Constraints

Model ReleasesDGX agent

arXiv:2605.19066v1 Announce Type: new Abstract: Over the past decade, low-resource natural language processing (NLP) has experienced explosive growth, propelled by cross-lingual transfer, massively mu

The Wikidata Query Logs Dataset

AgentsDGX agent

arXiv:2602.14594v2 Announce Type: replace Abstract: We present the Wikidata Query Logs (WDQL) dataset, a dataset consisting of 335k question-query pairs over the Wikidata knowledge graph. It is over 1

TIDE: Efficient and Lossless MoE Diffusion LLM Inference with I/O-aware Expert Offload

HardwareDGX agent

arXiv:2605.20179v1 Announce Type: new Abstract: Diffusion Large Language Models (dLLMs) have emerged as a competitive alternative to autoregressive (AR) models, offering better hardware utilization an

Time to REFLECT: Can We Trust LLM Judges for Evidence-based Research Agents?

Model ReleasesDGX agent

arXiv:2605.19196v1 Announce Type: new Abstract: Deep research agents increasingly automate complex information-seeking tasks, producing evidence-grounded reports via multi-step reasoning, tool use, an

Towards Consistent Detection of Cognitive Distortions: LLM-Based Annotation and Dataset-Agnostic Evaluation

Model ReleasesDGX agent

arXiv:2511.01482v2 Announce Type: replace Abstract: Text-based automated Cognitive Distortion detection is a challenging task due to its subjective nature, with low agreement scores observed even amon

Towards Trust Calibration in Socially Interactive Agents: Investigating Gendered Multimodal Behaviors Generation with LLMs

Model ReleasesDGX agent

arXiv:2605.19798v1 Announce Type: new Abstract: As Socially Interactive Agents (SIAs) become increasingly integrated into daily life, the ability to calibrate user trust to an agent's actual capabilit

Trust or Abstain? A Self-Aware RAG Approach

Model ReleasesDGX agent

arXiv:2605.18792v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) improves large language models (LLMs) by incorporating external evidence, but it also introduces knowledge confli

UCCI: Calibrated Uncertainty for Cost-Optimal LLM Cascade Routing

HardwareDGX agent

arXiv:2605.18796v1 Announce Type: cross Abstract: LLM cascades and model routing promise lower inference cost by sending easy queries to a small model and escalating hard ones to a large model, but mo

Unified Deployment-Aware Evaluation of Open Reasoning Language Models

Model ReleasesDGX agent

arXiv:2604.07035v2 Announce Type: replace Abstract: Open reasoning language models are often compared under mixed sample sizes, partially standardized prompts, and accuracy-centered summaries, which m

What Are LLMs Doing to Scientific Communication? Measuring Changes in Writing Practices and Reading Experience

ResearchDGX agent

arXiv:2605.19936v1 Announce Type: new Abstract: Has the style of scientific communication changed due to the growing use of large language models in the writing process? We address this question in th

Where Does Authorship Signal Emerge in Encoder-Based Language Models?

ResearchDGX agent

arXiv:2605.19908v1 Announce Type: new Abstract: Authorship attribution models fine-tuned with the same pretrained encoder, data, and loss can differ four-fold in performance depending only on their sc

XNote: Benchmarking Automated Community Notes Generation for Image-based Contextual Deception

Model ReleasesDGX agent

arXiv:2603.22453v2 Announce Type: replace Abstract: Community Notes have emerged as an effective crowd-sourced mechanism for combating online deception on social media platforms. However, its reliance

ZeroSearch: Incentivize the Search Capability of LLMs without Searching

Model ReleasesDGX agent

arXiv:2505.04588v3 Announce Type: replace Abstract: Effective information searching is essential for enhancing the reasoning and generation capabilities of large language models (LLMs). Recent researc

19 May 2026

A Data-Efficient Path to Multilingual LLMs: Language Expansion via Post-training PARAMDelta Integration into Upcycled MoE

Model ReleasesDGX agent

arXiv:2605.18083v1 Announce Type: new Abstract: Expanding Large Language Models~(LLMs) to new languages is a costly endeavor, demanding extensive Continued Pre-Training~(CPT) and data-intensive alignm

A Pilot Benchmark for NL-to-FOL Translation in Planetary Exploration

Model ReleasesDGX agent

arXiv:2605.17911v1 Announce Type: new Abstract: Future planetary exploration envisions autonomous robotic agents operating under severe communication constraints, without global positioning, and with

ACIL: Auto Chain of Thoughts for In-Context Learning

ResearchDGX agent

arXiv:2605.17088v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) have shown that Chain-of-Thought (CoT) reasoning can substantially improve performance on complex reason

AI Agents May Always Fall for Prompt Injections

SafetyDGX agent

arXiv:2605.17634v1 Announce Type: cross Abstract: Prompt injection is the most critical vulnerability in deployed AI agents. Despite recent progress, we show that the prevailing defense paradigm (data

AI Alignment Breaks at the Edge

SafetyDGX agent

arXiv:2602.20042v2 Announce Type: replace Abstract: General Alignment has improved average-case helpfulness and safety, but current alignment practice still rewards confident, single-turn responses. T

AMATA: Adaptive Multi-Agent Trajectory Alignment for Knowledge-Intensive Question Answering

SafetyDGX agent

arXiv:2605.17352v1 Announce Type: new Abstract: Despite substantial advances in large language models (LLMs), generating factually consistent responses for knowledge-intensive question answering remai

Analyzing Error Propagation in Korean Spoken QA with ASR-LLM Cascades

ResearchDGX agent

arXiv:2605.17443v1 Announce Type: new Abstract: We analyze how automatic speech recognition (ASR) errors propagate through ASR-LLM cascades in Korean spoken question answering (SQA), focusing on downs

Ancient Greek to Modern Greek Machine Translation: A Novel Benchmark and Fine-Tuning Experiments on LLMs and NMT Models

Model ReleasesDGX agent

arXiv:2605.18504v1 Announce Type: new Abstract: Machine Translation (MT) for Ancient Greek (AG) to Modern Greek (MG) is a low-resource task, constrained by the lack of large-scale, high-quality parall

Artificial Intolerance: Stigmatizing Language in Clinical Documentation Skews Large Language Model Decision-Making

SafetyDGX agent

arXiv:2605.17228v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed in high-stakes domains such as clinical decision support and medical documentation. However, the

AutoVecCoder: Teaching LLMs to Generate Explicitly Vectorized Code

ResearchDGX agent

arXiv:2605.17978v1 Announce Type: new Abstract: Vectorization via Single Instruction, Multiple Data (SIMD) architectures is a cornerstone of high-performance computing. To fully exploit hardware poten

BELIEF: Structured Evidence Modeling and Uncertainty-Aware Fusion for Biomedical Question Answering

ResearchDGX agent

arXiv:2605.17435v1 Announce Type: new Abstract: Biomedical question answering often requires decisions from retrieved literature whose relevance, quality, and support for candidate answers are uneven.

Beyond Neural Incompatibility: Cross-Scale Knowledge Transfer in Language Models through Latent Semantic Alignment

Model ReleasesDGX agent

arXiv:2510.24208v2 Announce Type: replace Abstract: Language Models (LMs) encode substantial knowledge in their parameters, yet it remains unclear how to transfer such knowledge in a fine-grained mann

Beyond Sentiment Classification: A Generative Framework for Emotion Intensity Evaluation in Text

ApplicationsDGX agent

arXiv:2605.16613v1 Announce Type: new Abstract: We introduce a novel approach to emotion modeling that shifts the focus from identification to evaluation, addressing the limitations of discrete classi

Beyond the Final Actor: Modeling the Dual Roles of Creator and Editor for Fine-Grained LLM-Generated Text Detection

SafetyDGX agent

arXiv:2604.04932v3 Announce Type: replace Abstract: The misuse of large language models (LLMs) requires precise detection of synthetic text. Existing works mainly follow binary or ternary classificati

Beyond Transcripts: Iterative Peer-Editing with Audio Unlocks High-Quality Human Summaries of Conversational Speech

SafetyDGX agent

arXiv:2605.17652v1 Announce Type: new Abstract: There are not enough established benchmarks for the task fo speech summarization. Creating new benchmarks demands human annotation, as LLMs could embed

Bridging the Gap: Converting Read Text to Conversational Dialogue

ResearchDGX agent

arXiv:2605.18001v1 Announce Type: new Abstract: In recent advancements within speech processing, converting read speech to conversational speech has gained significant attention. The primary challenge

Can LLMs Generate and Solve Linguistic Olympiad Puzzles?

Model ReleasesDGX agent

arXiv:2509.21820v2 Announce Type: replace Abstract: In this paper, we introduce a combination of novel and exciting tasks: the solution and generation of linguistic puzzles. We focus on puzzles used i

Closing the Gap at CRAC 2026: Two-Stage Adaptation for LLM-Based Multilingual Coreference Resolution

Model ReleasesDGX agent

arXiv:2605.16984v1 Announce Type: new Abstract: We present our submission to the LLM track of the 2026 Computational Models of Reference, Anaphora and Coreference (CRAC 2026) shared task. With an aver

CompactAttention: Accelerating Chunked Prefill with Block-Union KV Selection

Model ReleasesDGX agent

arXiv:2605.16839v1 Announce Type: new Abstract: Chunked prefill has become a widely adopted serving strategy for long-context large language models, but efficient attention computation in this regime

Compounding Disadvantage: Auditing Intersectional Bias in LLM-Generated Explanations Across Indian and American STEM Education

Model ReleasesDGX agent

arXiv:2601.14506v3 Announce Type: replace-cross Abstract: Large language models are increasingly deployed in STEM education for personalized instruction and feedback across institutions in high- and l

Compress the Context, Keep the Commitments: A Formal Framework for Verifiable LLM Context Compression

SafetyDGX agent

arXiv:2605.17304v1 Announce Type: cross Abstract: LLM context is not just tokens; it is a set of commitments. Long-running conversations accumulate goals, constraints, decisions, preferences, tool res

Confidence Geometry Reveals Trace-Level Correctness in Large Language Model Reasoning

ResearchDGX agent

arXiv:2605.16824v1 Announce Type: cross Abstract: Large language models (LLMs) generate not only reasoning text, but also token-level confidence trajectories that record how uncertainty evolves during

Constrained Code Generation with Discrete Diffusion

Local AiDGX agent

arXiv:2605.16829v1 Announce Type: new Abstract: Discrete diffusion models are a powerful, emerging paradigm for code generation. They construct programs through iterative refinement of partially corru

DISA: Offline Importance Sampling for Distribution-Matching LLM-RL

SafetyDGX agent

arXiv:2605.17295v1 Announce Type: cross Abstract: Modern reasoning agents are increasingly evaluated on their ability to generate multiple valid solution paths, plans, or tool-use traces for a given i

Disentangling Ambiguity from Instability in Large Language Models: A Clinical Text-to-SQL Case Study

Model ReleasesDGX agent

arXiv:2602.12015v2 Announce Type: replace Abstract: Deploying large language models for clinical Text-to-SQL requires distinguishing two qualitatively different causes of output diversity: (i) input a

Do LLM Agents Mirror Socio-Cognitive Effects in Power-Asymmetric Conversations?

SafetyDGX agent

arXiv:2605.17694v1 Announce Type: new Abstract: Power differences shape human communication through well documented socio cognitive effects, including language coordination, pronoun usage, authority b

Dual-Space Knowledge Distillation with Key-Query Matching for Large Language Models with Vocabulary Mismatch

SafetyDGX agent

arXiv:2603.22056v2 Announce Type: replace Abstract: Large language models (LLMs) achieve state-of-the-art (SOTA) performance across language tasks, but are costly to deploy due to their size and resou

E-PMQ: Expert-Guided Post-Merge Quantization with Merged-Weight Anchoring

ResearchDGX agent

arXiv:2605.16882v1 Announce Type: new Abstract: Low-resource deployment constraints have made model quantization essential for deploying neural networks while preserving performance. Meanwhile, model

Early Stopping Chain-of-thoughts in Large Language Models

ResearchDGX agent

arXiv:2509.14004v2 Announce Type: replace Abstract: Reasoning large language models (LLMs) have demonstrated superior capacities in solving complicated problems by generating long chain-of-thoughts (C

Easier to Judge than to Find: Predicting In-Context Learning Success for Demonstration Selection

Model ReleasesDGX agent

arXiv:2605.18512v1 Announce Type: new Abstract: In-context learning (ICL) is highly sensitive to which demonstrations appear in the prompt, but selecting them is expensive because the space of possibl

Embodied Task Planning via Graph-Informed Action Generation with Large Language Models

Model ReleasesDGX agent

arXiv:2601.21841v3 Announce Type: replace Abstract: While Large Language Models (LLMs) have demonstrated strong zero-shot reasoning capabilities, their deployment as embodied agents still faces fundam

Embracing Anisotropy: Turning Massive Activations into Interpretable Control Knobs for Large Language Models

ResearchDGX agent

arXiv:2603.00029v2 Announce Type: replace Abstract: Large Language Models (LLMs) exhibit highly anisotropic internal representations, often characterized by massive activations, a phenomenon where a s

EnvFactory: Scaling Tool-Use Agents via Executable Environments Synthesis and Robust RL

AgentsDGX agent

arXiv:2605.18703v1 Announce Type: new Abstract: Equipping LLMs with tool-use capabilities via Agentic Reinforcement Learning (Agentic RL) is bottlenecked by two challenges: the lack of scalable, robus

Evaluation Drift in LLM Personality Induction: Are We Moving the Goalpost?

ResearchDGX agent

arXiv:2605.16996v1 Announce Type: new Abstract: Can large language models reliably express a human-like personality, or are they merely mimicking surface cues without a stable underlying profile? To i

Evolve the Method, Not the Prompts: Evolutionary Synthesis of Jailbreak Attacks on LLMs

Model ReleasesDGX agent

arXiv:2511.12710v2 Announce Type: replace Abstract: Automated red teaming frameworks for Large Language Models (LLMs) have become increasingly sophisticated, yet many still formulate attack optimizati

Factual Inconsistencies in Multilingual Wikipedia Tables

SafetyDGX agent

arXiv:2507.18406v2 Announce Type: replace Abstract: Wikipedia serves as a globally accessible knowledge source with content in over 300 languages. Despite covering the same topics, the different versi

FastOCR: Dynamic Visual Fixation via KV Cache Pruning for Efficient Document Parsing

Local AiDGX agent

arXiv:2605.17447v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have shown strong promise on Optical Character Recognition (OCR), yet the sheer number of visual tokens required to enco

FIM-LoRA: Task-Informative Rank Allocation for LoRA via Calibration-Time Gradient-Variance Estimation

Model ReleasesDGX agent

arXiv:2605.16800v1 Announce Type: cross Abstract: Low-rank adaptation (LoRA) assigns a uniform rank to every adapted weight matrix - a practical convenience that ignores a fundamental reality: differe

FinAuditing: A Financial Taxonomy-Structured Multi-Document Benchmark for Evaluating LLMs

Model ReleasesDGX agent

arXiv:2510.08886v3 Announce Type: replace Abstract: Going beyond simple text processing, financial auditing requires detecting semantic, structural, and numerical inconsistencies across large-scale di

Finding Sense in Nonsense with Generated Contexts: Perspectives from Humans and Language Models

ResearchDGX agent

arXiv:2602.11699v3 Announce Type: replace Abstract: Nonsensical and anomalous sentences have been instrumental in the development of computational models of semantic interpretation. A core challenge i

Firefly: Illuminating Large-Scale Verified Tool-Call Data Generation from Real APIs

Model ReleasesDGX agent

arXiv:2605.17558v1 Announce Type: cross Abstract: Training tool-calling agents requires large-scale trajectory data with verifiable labels, yet existing approaches either synthesize environments that

← Previous
1…6970717273…129
Next →