AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlog
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,759 results
Research

FlexDraft: Flexible Speculative Decoding via Attention Tuning and Bonus-Guided Calibration

DGX agent

arXiv:2605.20022v1 Announce Type: new Abstract: Speculative decoding accelerates memory-bound LLM inference without quality degradation by using a fast drafter to propose multiple candidate tokens and

researcharxiv-cs-cl
20 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Research

From Seeing to Thinking: Decoupling Perception and Reasoning Improves Post-Training of Vision-Language Models

DGX agent

arXiv:2605.20177v1 Announce Type: new Abstract: Recent advances in vision-language models (VLMs) emphasize long chain-of-thought reasoning; yet, we find that their performance on visual tasks is prima

researcharxiv-cs-cl
20 May 2026
Model Releases

GoLongRL: Capability-Oriented Long Context Reinforcement Learning with Multitask Alignment

DGX agent

arXiv:2605.19577v1 Announce Type: new Abstract: We present GoLongRL, a fully open-source, capability-oriented post-training recipe for long-context reinforcement learning with verifiable rewards (RLVR

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

GRAB: A Risk Taxonomy--Grounded Benchmark for Unsupervised Topic Discovery in Financial Disclosures

DGX agent

arXiv:2509.21698v2 Announce Type: replace Abstract: Risk categorization in 10-K risk disclosures matters for oversight and investment, yet no public benchmark evaluates unsupervised topic models for t

model-releasesarxiv-cs-cl
20 May 2026
Research

HALvest-Contrastive: Retrieval-Like Authorship Attribution with Patch-Level Late Interaction

DGX agent

arXiv:2407.20595v4 Announce Type: replace-cross Abstract: Deciding whether two pieces of text share an author is made difficult by topical confound: two writers covering the same topic often look more

researcharxiv-cs-cl
20 May 2026
Safety

How Do Document Parsers Break? Auditing Structural Vulnerability in Document Intelligence

DGX agent

arXiv:2605.19309v1 Announce Type: new Abstract: Document Layout Analysis (DLA) pipelines provide structured page representations for retrieval-augmented generation, long-document question answering, a

safetyarxiv-cs-cl
20 May 2026
Model Releases

K-Quantization and its Impact on Output Performance

DGX agent

arXiv:2605.19645v1 Announce Type: new Abstract: Recent advancements in large language models (LLMs) have shown their remarkable capacities in many NLP tasks. However, their substantial size often pres

model-releasesarxiv-cs-cl
20 May 2026
Research

KoRe: Compact Knowledge Representations for Large Language Models

DGX agent

arXiv:2605.20170v1 Announce Type: new Abstract: Modern Large Language Models (LLMs) have shown impressive performances in user-facing tasks such as question answering, as well as consistent improvemen

researcharxiv-cs-cl
20 May 2026
Safety

LambdaPO: A Lambda Style Policy Optimization for Reasoning Language Models

DGX agent

arXiv:2605.19416v1 Announce Type: new Abstract: Group Relative Policy Optimization(GRPO) has become a cornerstone of modern reinforcement learning alignment, prized for its efficacy in foregoing an ex

safetyarxiv-cs-cl
20 May 2026
Tutorials

Language models struggle with compartmentalization

DGX agent

arXiv:2605.19284v1 Announce Type: new Abstract: In the training data used by large language models (LLMs), the same latent concept is often presented in multiple distinct ways: the same facts appear i

tutorialsarxiv-cs-cl
20 May 2026
Research

Language Mutations Sustain the Persistences of Conspiracy Theories on Social Media

DGX agent

arXiv:2605.20050v1 Announce Type: new Abstract: This study investigates how language mutations affect the persistent diffusion of conspiracy theories on social media. Drawing on a three-year dataset o

researcharxiv-cs-cl
20 May 2026
Model Releases

Library Hallucinations in LLM-Generated Code: A Risk Analysis Grounded in Developer Queries

DGX agent

arXiv:2509.22202v3 Announce Type: replace-cross Abstract: Large language models (LLMs) now play a central role in code generation, yet they continue to hallucinate, frequently inventing non-existent l

model-releasesarxiv-cs-cl
20 May 2026
Research

LLM-Based Financial Sentiment Analysis in Arabic: Evidence from Saudi Markets

DGX agent

arXiv:2605.19714v1 Announce Type: new Abstract: Investor sentiment shapes financial markets, yet modeling sentiment in Arabic financial contexts remains challenging due to linguistic complexity and li

researcharxiv-cs-cl
20 May 2026
Applications

LLM-MC-Affect: LLM-Based Monte Carlo Modeling of Affective Trajectories and Latent Ambiguity for Interpersonal Dynamic Insight

DGX agent

arXiv:2601.03645v2 Announce Type: replace Abstract: Emotional coordination is a core property of human interaction that shapes how relational meaning is constructed in real time. While text-based affe

applicationsarxiv-cs-cl
20 May 2026
Model Releases

LLMEval-Logic: A Solver-Verified Chinese Benchmark for Logical Reasoning of LLMs with Adversarial Hardening

DGX agent

arXiv:2605.19597v1 Announce Type: new Abstract: Evaluating large language models (LLMs) on natural-language logical reasoning is essential because rule-governed tasks require conclusions to follow str

model-releasesarxiv-cs-cl
20 May 2026
Research

Lost in Interpretation: The Plausibility-Faithfulness Trade-off in Cross-Lingual Explanations

DGX agent

arXiv:2605.19274v1 Announce Type: new Abstract: LLMs deployed multilingually are often audited via English explanations for non-English inputs. We evaluate extractive explanations ''where the model id

researcharxiv-cs-cl
20 May 2026
Model Releases

m3BERT: A Modern, Multi-lingual, Matryoshka Bidirectional Encoder

DGX agent

arXiv:2605.19568v1 Announce Type: new Abstract: Embedding models are pivotal in industrial information retrieval systems like search and advertising. However, existing pretrained models often exhibit

model-releasesarxiv-cs-cl
20 May 2026
Safety

Measuring Stereotype and Deviation Biases in Large Language Models

DGX agent

arXiv:2508.06649v3 Announce Type: replace Abstract: Large language models (LLMs) are widely applied across diverse domains, raising concerns about their limitations and potential risks. In this study,

safetyarxiv-cs-cl
20 May 2026
Research

Mind Your Moras: Orthography-Aware Error Analysis of Neural Japanese Morphological Generation

DGX agent

arXiv:2605.20043v1 Announce Type: new Abstract: We present an orthography-aware error analysis of Japanese past-tense morphological inflection, treating hiragana not merely as a transcriptional medium

researcharxiv-cs-cl
20 May 2026
Model Releases

MixRea: Benchmarking Explicit-Implicit Reasoning in Large Language Models

DGX agent

arXiv:2605.20128v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly integrated into high-stakes decision-making. Inspired by the theory of inattentional blindness in human co

model-releasesarxiv-cs-cl
20 May 2026
Agents

MMoA: An AI-Agent framework with recurrence for Memoried Mixure-of-Agent

DGX agent

arXiv:2605.19194v1 Announce Type: new Abstract: The Mixture-of-Agents (MoA) framework has shown promise in improving large language model (LLM) performance by aggregating outputs from multiple agents.

agentsarxiv-cs-cl
20 May 2026
Model Releases

MTraining: Distributed Dynamic Sparse Attention for Efficient Ultra-Long Context Training

DGX agent

arXiv:2510.18830v2 Announce Type: replace Abstract: The adoption of long context windows has become a standard feature in Large Language Models (LLMs), as extended contexts significantly enhance their

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

OpenCompass: A Universal Evaluation Platform for Large Language Models

DGX agent

arXiv:2605.19276v1 Announce Type: new Abstract: In recent years, the field of artificial intelligence has undergone a paradigm shift from task-specific small-scale models to general-purpose large lang

model-releasesarxiv-cs-cl
20 May 2026
Hardware

OScaR: The Occam's Razor for Extreme KV Cache Quantization in LLMs and Beyond

DGX agent

arXiv:2605.19660v1 Announce Type: cross Abstract: The rapid advancement toward long-context reasoning and multi-modal intelligence has made the memory footprint of the Key-Value (KV) cache a dominant

hardwarearxiv-cs-cl
20 May 2026
Agents

PASC: Pipeline-Aware Conformal Prediction with Joint Coverage Guarantees for Multi-Stage NLP and LLM Pipelines

DGX agent

arXiv:2605.18812v1 Announce Type: cross Abstract: Modern NLP and LLM systems are pipelines: named entity recognition (NER) -> entity disambiguation (NED) -> entity typing, retrieval-augmented generati

agentsarxiv-cs-cl
20 May 2026
Model Releases

Prompting language influences diagnostic reasoning and accuracy of large language models

DGX agent

arXiv:2605.19173v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly explored for clinical decision support, yet most evaluations are conducted in English, leaving their relia

model-releasesarxiv-cs-cl
20 May 2026
Research

Qayyem: A Real-time Platform for Scoring Proficiency of Arabic Essays

DGX agent

arXiv:2603.01009v2 Announce Type: replace Abstract: Over the past years, Automated Essay Scoring (AES) systems have gained increasing attention as scalable and consistent solutions for assessing the p

researcharxiv-cs-cl
20 May 2026
Model Releases

Rethinking How to Remember: Beyond Atomic Facts in Lifelong LLM Agent Memory

DGX agent

arXiv:2605.19952v1 Announce Type: new Abstract: To enable reliable long-term interaction, LLM agents require a memory system that can faithfully store, efficiently retrieve, and deeply reason over acc

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

Retrieval-Augmented Generation for Natural Language Processing: A Survey

DGX agent

arXiv:2407.13193v4 Announce Type: replace Abstract: Large language models (LLMs) have achieved strong empirical performance in various fields, benefiting from their huge amount of parameters that stor

model-releasesarxiv-cs-cl
20 May 2026
Research

Retrieval-Augmented Linguistic Calibration

DGX agent

arXiv:2605.19344v1 Announce Type: new Abstract: Linguistic cues such as 'I believe' and 'probably' offer an intuitive interface for communicating confidence, yet a generalisable, principled calibratio

researcharxiv-cs-cl
20 May 2026
Safety

Rewarding Beliefs, Not Actions: Consistency-Guided Credit Assignment for Long-Horizon Agents

DGX agent

arXiv:2605.20061v1 Announce Type: new Abstract: Reinforcement learning from verifiable rewards (RLVR) is a promising paradigm for improving large language model (LLM) agents on long-horizon interactiv

safetyarxiv-cs-cl
20 May 2026
Model Releases

Rewriting History: A Recipe for Interventional Analyses to Study Data Effects on Model Behavior

DGX agent

arXiv:2510.14261v2 Announce Type: replace Abstract: We present an experimental recipe for studying the relationship between training data and language model (LM) behavior. We outline steps for interve

model-releasesarxiv-cs-cl
20 May 2026
Research

Scaling Evaluation-time Compute with Reasoning Models as Evaluators

DGX agent

arXiv:2503.19877v2 Announce Type: replace Abstract: As language model (LM) outputs get more and more natural, it is becoming more difficult than ever to evaluate their quality. Simultaneously, increas

researcharxiv-cs-cl
20 May 2026
Model Releases

SciCustom: A Framework for Custom Evaluation of Scientific Capabilities in Large Language Models

DGX agent

arXiv:2605.19357v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly applied to scientific research, yet existing evaluations often fail to reflect the fine-grained capabiliti

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

Self-Filtered Distillation with LLMs-generated Trust Indicators for Reliable Patent Classification

DGX agent

arXiv:2510.05431v4 Announce Type: replace Abstract: Organizing large-scale patent corpora according to classification schemes is a core information management task that determines the accuracy and eff

model-releasesarxiv-cs-cl
20 May 2026
Applications

SETUP: Sentence-level English-To-Uniform Meaning Representation Parser

DGX agent

arXiv:2512.07068v3 Announce Type: replace Abstract: Uniform Meaning Representation (UMR) is a novel graph-based semantic representation which captures the core meaning of a text, with flexibility inco

applicationsarxiv-cs-cl
20 May 2026
Model Releases

SLoW: Select Low-frequency Words! Automatic Dictionary Selection for Translation on Large Language Models

DGX agent

arXiv:2507.18902v2 Announce Type: replace Abstract: There are more than 7,000 languages around the world, and current Large Language Models (LLMs) only support hundreds of languages. Dictionary-based

model-releasesarxiv-cs-cl
20 May 2026
Research

SPHERICAL KV: Angle-Domain Attention and Rate-Distortion Retention for Efficient Long-Context Inference

DGX agent

arXiv:2605.18856v1 Announce Type: cross Abstract: Long-context inference is increasingly constrained by the KV cache: resident memory grows with context length, and decoding becomes limited by repeate

researcharxiv-cs-cl
20 May 2026
Research

Structured Style-Rewrite with Chain-of-Thought Planning for Low-Resource Character Dialogue

DGX agent

arXiv:2603.05933v2 Announce Type: replace Abstract: Applying Small Language Models (SLMs) to Chinese character-driven generation remains challenging due to data scarcity and the difficulty of disentan

researcharxiv-cs-cl
20 May 2026
Model Releases

Taming the Thinker: Conditional Entropy Shaping for Adaptive LLM Reasoning

DGX agent

arXiv:2605.19358v1 Announce Type: new Abstract: Entropy-based deep reasoning has emerged as a promising direction for improving the reasoning capabilities of Large Language Models (LLMs), but existing

model-releasesarxiv-cs-cl
20 May 2026
Safety

Text-to-SPARQL Generation with Reinforcement Learning: A GRPO-based Approach on DBLP

DGX agent

arXiv:2605.20066v1 Announce Type: new Abstract: Knowledge graph question answering seeks to translate natural language questions into executable queries over knowledge graphs, but existing approaches

safetyarxiv-cs-cl
20 May 2026
Model Releases

The Annotation Scarcity Paradox in Low-Resource NLP Evaluation: A Decade of Acceleration and Emerging Constraints

DGX agent

arXiv:2605.19066v1 Announce Type: new Abstract: Over the past decade, low-resource natural language processing (NLP) has experienced explosive growth, propelled by cross-lingual transfer, massively mu

model-releasesarxiv-cs-cl
20 May 2026
Agents

The Wikidata Query Logs Dataset

DGX agent

arXiv:2602.14594v2 Announce Type: replace Abstract: We present the Wikidata Query Logs (WDQL) dataset, a dataset consisting of 335k question-query pairs over the Wikidata knowledge graph. It is over 1

agentsarxiv-cs-cl
20 May 2026
Hardware

TIDE: Efficient and Lossless MoE Diffusion LLM Inference with I/O-aware Expert Offload

DGX agent

arXiv:2605.20179v1 Announce Type: new Abstract: Diffusion Large Language Models (dLLMs) have emerged as a competitive alternative to autoregressive (AR) models, offering better hardware utilization an

hardwarearxiv-cs-cl
20 May 2026
Model Releases

Time to REFLECT: Can We Trust LLM Judges for Evidence-based Research Agents?

DGX agent

arXiv:2605.19196v1 Announce Type: new Abstract: Deep research agents increasingly automate complex information-seeking tasks, producing evidence-grounded reports via multi-step reasoning, tool use, an

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

Towards Consistent Detection of Cognitive Distortions: LLM-Based Annotation and Dataset-Agnostic Evaluation

DGX agent

arXiv:2511.01482v2 Announce Type: replace Abstract: Text-based automated Cognitive Distortion detection is a challenging task due to its subjective nature, with low agreement scores observed even amon

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

Towards Trust Calibration in Socially Interactive Agents: Investigating Gendered Multimodal Behaviors Generation with LLMs

DGX agent

arXiv:2605.19798v1 Announce Type: new Abstract: As Socially Interactive Agents (SIAs) become increasingly integrated into daily life, the ability to calibrate user trust to an agent's actual capabilit

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

Trust or Abstain? A Self-Aware RAG Approach

DGX agent

arXiv:2605.18792v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) improves large language models (LLMs) by incorporating external evidence, but it also introduces knowledge confli

model-releasesarxiv-cs-cl
20 May 2026
← Previous
1…8788899091…162
Next →