AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,646 results
10 Apr 2026

Contextual Earnings-22: A Speech Recognition Benchmark with Custom Vocabulary in the Wild

Model ReleasesDGX agent

arXiv:2604.07354v1 Announce Type: new Abstract: The accuracy frontier of speech-to-text systems has plateaued on academic benchmarks.1 In contrast, industrial benchmarks and adoption in high-stakes do

Contextualising (Im)plausible Events Triggers Figurative Language

SafetyDGX agent

arXiv:2604.07885v1 Announce Type: new Abstract: This work explores the connection between (non-)literalness and plausibility at the example of subject-verb-object events in English. We design a system

Cram Less to Fit More: Training Data Pruning Improves Memorization of Facts

ResearchDGX agent

arXiv:2604.08519v1 Announce Type: new Abstract: Large language models (LLMs) can struggle to memorize factual knowledge in their parameters, often leading to hallucinations and poor performance on kno


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Cross-Tokenizer LLM Distillation through a Byte-Level Interface

ResearchDGX agent

arXiv:2604.07466v1 Announce Type: new Abstract: Cross-tokenizer distillation (CTD), the transfer of knowledge from a teacher to a student language model when the two use different tokenizers, remains

Current LLMs still cannot 'talk much' about grammar modules: Evidence from syntax

ResearchDGX agent

arXiv:2603.20114v4 Announce Type: replace Abstract: We aim to examine the extent to which Large Language Models (LLMs) can 'talk much' about grammar modules, providing evidence from syntax core proper

CycleChart: A Unified Consistency-Based Learning Framework for Bidirectional Chart Understanding and Generation

Model ReleasesDGX agent

arXiv:2512.19173v2 Announce Type: replace Abstract: Current chart-related tasks, such as chart generation (NL2Chart), chart schema parsing, chart data parsing, and chart question answering (ChartQA),

Data Selection for Multi-turn Dialogue Instruction Tuning

Local AiDGX agent

arXiv:2604.07892v1 Announce Type: new Abstract: Instruction-tuned language models increasingly rely on large multi-turn dialogue corpora, but these datasets are often noisy and structurally inconsiste

Decompose, Look, and Reason: Reinforced Latent Reasoning for VLMs

SafetyDGX agent

arXiv:2604.07518v1 Announce Type: new Abstract: Vision-Language Models often struggle with complex visual reasoning due to the visual information loss in textual CoT. Existing methods either add the c

Demystifying OPD: Length Inflation and Stabilization Strategies for Large Language Models

SafetyDGX agent

arXiv:2604.08527v1 Announce Type: new Abstract: On-policy distillation (OPD) trains student models under their own induced distribution while leveraging supervision from stronger teachers. We identify

Detecting HIV-Related Stigma in Clinical Narratives Using Large Language Models

Model ReleasesDGX agent

arXiv:2604.07717v1 Announce Type: new Abstract: Human immunodeficiency virus (HIV)-related stigma is a critical psychosocial determinant of health for people living with HIV (PLWH), influencing mental

Differentially Private Language Generation and Identification in the Limit

ResearchDGX agent

arXiv:2604.08504v1 Announce Type: cross Abstract: We initiate the study of language generation in the limit, a model recently introduced by Kleinberg and Mullainathan [KM24], under the constraint of d

Diffusion Language Models Know the Answer Before Decoding

ResearchDGX agent

arXiv:2508.19982v5 Announce Type: replace Abstract: Diffusion language models (DLMs) have recently emerged as an alternative to autoregressive approaches, offering parallel sequence generation and fle

Distributed Multi-Layer Editing for Rule-Level Knowledge in Large Language Models

Model ReleasesDGX agent

arXiv:2604.08284v1 Announce Type: new Abstract: Large language models store not only isolated facts but also rules that support reasoning across symbolic expressions, natural language explanations, an

DIVERSED: Relaxed Speculative Decoding via Dynamic Ensemble Verification

ResearchDGX agent

arXiv:2604.07622v1 Announce Type: new Abstract: Speculative decoding is an effective technique for accelerating large language model inference by drafting multiple tokens in parallel. In practice, its

Don't Overthink It: Inter-Rollout Action Agreement as a Free Adaptive-Compute Signal for LLM Agents

Model ReleasesDGX agent

arXiv:2604.08369v1 Announce Type: cross Abstract: Inference-time compute scaling has emerged as a powerful technique for improving the reliability of large language model (LLM) agents, but existing me

Double: Breaking the Acceleration Limit via Double Retrieval Speculative Parallelism

ResearchDGX agent

arXiv:2601.05524v2 Announce Type: replace Abstract: Parallel Speculative Decoding (PSD) accelerates traditional Speculative Decoding (SD) by overlapping draft generation with verification. However, it

DQA: Diagnostic Question Answering for IT Support

ApplicationsDGX agent

arXiv:2604.05350v2 Announce Type: replace Abstract: Enterprise IT support interactions are fundamentally diagnostic: effective resolution requires iterative evidence gathering from ambiguous user repo

Dual-Pool Token-Budget Routing for Cost-Efficient and Reliable LLM Serving

Model ReleasesDGX agent

arXiv:2604.08075v1 Announce Type: new Abstract: Production vLLM fleets typically provision each instance for the worst-case context length, leading to substantial KV-cache over-allocation and under-ut

DYCP: Dynamic Context Pruning for Long-Form Dialogue with LLMs

ResearchDGX agent

arXiv:2601.07994v4 Announce Type: replace Abstract: Large Language Models (LLMs) increasingly operate over long-form dialogues with frequent topic shifts. While recent LLMs support extended context wi

E2Edev: Benchmarking Large Language Models in End-to-End Software Development Task

Model ReleasesDGX agent

arXiv:2510.14509v3 Announce Type: replace-cross Abstract: The rapid advancement in large language models (LLMs) has demonstrated significant potential in End-to-End Software Development (E2ESD). Howev

Efficient and Effective Internal Memory Retrieval for LLM-Based Healthcare Prediction

Model ReleasesDGX agent

arXiv:2604.07659v1 Announce Type: new Abstract: Large language models (LLMs) hold significant promise for healthcare, yet their reliability in high-stakes clinical settings is often compromised by hal

Efficient PRM Training Data Synthesis via Formal Verification

ResearchDGX agent

arXiv:2505.15960v3 Announce Type: replace Abstract: Process Reward Models (PRMs) have emerged as a promising approach for improving LLM reasoning capabilities by providing process supervision over rea

Efficient Provably Secure Linguistic Steganography via Range Coding

ResearchDGX agent

arXiv:2604.08052v1 Announce Type: new Abstract: Linguistic steganography involves embedding secret messages within seemingly innocuous texts to enable covert communication. Provable security, which is

Emotion Concepts and their Function in a Large Language Model

Model ReleasesDGX agent

arXiv:2604.07729v1 Announce Type: cross Abstract: Large language models (LLMs) sometimes appear to exhibit emotional reactions. We investigate why this is the case in Claude Sonnet 4.5 and explore imp

EMSDialog: Synthetic Multi-person Emergency Medical Service Dialogue Generation from Electronic Patient Care Reports via Multi-LLM Agents

AgentsDGX agent

arXiv:2604.07549v1 Announce Type: new Abstract: Conversational diagnosis prediction requires models to track evolving evidence in streaming clinical conversations and decide when to commit to a diagno

Enabling Intrinsic Reasoning over Dense Geospatial Embeddings with DFR-Gemma

Model ReleasesDGX agent

arXiv:2604.07490v1 Announce Type: new Abstract: Representation learning for geospatial and spatio-temporal data plays a critical role in enabling general-purpose geospatial intelligence. Recent geospa

Entropy-Gradient Grounding: Training-Free Evidence Retrieval in Vision-Language Models

ResearchDGX agent

arXiv:2604.08456v1 Announce Type: cross Abstract: Despite rapid progress, pretrained vision-language models still struggle when answers depend on tiny visual details or on combining clues spread acros

Evaluating LLMs for Demographic-Targeted Social Bias Detection: A Comprehensive Benchmark Study

Model ReleasesDGX agent

arXiv:2510.04641v3 Announce Type: replace Abstract: Large-scale web-scraped text corpora used to train general-purpose AI models often contain harmful demographic-targeted social biases, creating a re

EventWeave: A Dynamic Framework for Capturing Core and Supporting Events in Dialogue Systems

TutorialsDGX agent

arXiv:2503.23078v3 Announce Type: replace Abstract: Large language models have improved dialogue systems, but often process conversational turns in isolation, overlooking the event structures that gui

Exploring Plan Space through Conversation: An Agentic Framework for LLM-Mediated Explanations in Planning

AgentsDGX agent

arXiv:2603.02070v2 Announce Type: replace-cross Abstract: When automating plan generation for a real-world sequential decision problem, the goal is often not to replace the human planner, but to facil

FinTruthQA: A Benchmark for AI-Driven Financial Disclosure Quality Assessment in Investor -- Firm Interactions

Model ReleasesDGX agent

arXiv:2406.12009v5 Announce Type: replace Abstract: Accurate and transparent financial information disclosure is essential for market efficiency, investor decision-making, and corporate governance. Ch

Floating or Suggesting Ideas? A Large-Scale Contrastive Analysis of Metaphorical and Literal Verb-Object Constructions

ResearchDGX agent

arXiv:2604.08275v1 Announce Type: new Abstract: Metaphor pervades everyday language, allowing speakers to express abstract concepts via concrete domains. While prior work has studied metaphors cogniti

Flux Attention: Context-Aware Hybrid Attention for Efficient LLMs Inference

Model ReleasesDGX agent

arXiv:2604.07394v1 Announce Type: cross Abstract: The quadratic computational complexity of standard attention mechanisms presents a severe scalability bottleneck for LLMs in long-context scenarios. W

Formalizing building-up constructions of self-dual codes through isotropic lines in Lean

ResearchDGX agent

arXiv:2604.08485v1 Announce Type: cross Abstract: The purpose of this paper is two-fold. First we show that Kim's building-up construction of binary self-dual codes is equivalent to Chinburg-Zhang's H

From Fragments to Facts: A Curriculum-Driven DPO Approach for Generating Hindi News Veracity Explanations

Model ReleasesDGX agent

arXiv:2507.05179v4 Announce Type: replace Abstract: In an era of rampant misinformation, generating reliable news explanations is vital, especially for under-represented languages like Hindi. Lacking

From Ground Truth to Measurement: A Statistical Framework for Human Labeling

SafetyDGX agent

arXiv:2604.07591v1 Announce Type: cross Abstract: Supervised machine learning assumes that labeled data provide accurate measurements of the concepts models are meant to learn. Yet in practice, human

Graph Neural Networks for Misinformation Detection: Performance-Efficiency Trade-offs

Model ReleasesDGX agent

arXiv:2604.08131v1 Announce Type: new Abstract: The rapid spread of online misinformation has led to increasingly complex detection models, including large language models and hybrid architectures. Ho

GRASS: Gradient-based Adaptive Layer-wise Importance Sampling for Memory-efficient Large Language Model Fine-tuning

Model ReleasesDGX agent

arXiv:2604.07808v1 Announce Type: new Abstract: Full-parameter fine-tuning of large language models is constrained by substantial GPU memory requirements. Low-rank adaptation methods mitigate this cha

GroupGPT: A Token-efficient and Privacy-preserving Agentic Framework for Multi-User Chat Assistant

Model ReleasesDGX agent

arXiv:2603.01059v3 Announce Type: replace Abstract: Recent advances in large language models (LLMs) have enabled increasingly capable chatbots. However, most existing systems focus on single-user sett

Guaranteeing Knowledge Integration with Joint Decoding for Retrieval-Augmented Generation

ResearchDGX agent

arXiv:2604.08046v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) significantly enhances Large Language Models (LLMs) by providing access to external knowledge. However, current res

Guardian-as-an-Advisor: Advancing Next-Generation Guardian Models for Trustworthy LLMs

SafetyDGX agent

arXiv:2604.07655v1 Announce Type: cross Abstract: Hard-gated safety checkers often over-refuse and misalign with a vendor's model spec; prevailing taxonomies also neglect robustness and honesty, yield

Hallucination Detection and Evaluation of Large Language Model

Local AiDGX agent

arXiv:2512.22416v2 Announce Type: replace Abstract: Hallucinations in Large Language Models (LLMs) pose a significant challenge, generating misleading or unverifiable content that undermines trust and

HCRE: LLM-based Hierarchical Classification for Cross-Document Relation Extraction with a Prediction-then-Verification Strategy

ResearchDGX agent

arXiv:2604.07937v1 Announce Type: new Abstract: Cross-document relation extraction (RE) aims to identify relations between the head and tail entities located in different documents. Existing approache

HiCI: Hierarchical Construction-Integration for Long-Context Attention

Model ReleasesDGX agent

arXiv:2603.20843v2 Announce Type: replace Abstract: Long-context language modeling is commonly framed as a scalability challenge of token-level attention, yet local-to-global information structuring r

How Independent are Large Language Models? A Statistical Framework for Auditing Behavioral Entanglement and Reweighting Verifier Ensembles

SafetyDGX agent

arXiv:2604.07650v1 Announce Type: cross Abstract: The rapid growth of the large language model (LLM) ecosystem raises a critical question: are seemingly diverse models truly independent? Shared pretra

How Psychological Learning Paradigms Shaped and Constrained Artificial Intelligence

SafetyDGX agent

arXiv:2603.18203v3 Announce Type: replace Abstract: Current artificial intelligence systems struggle with systematic compositional reasoning: the capacity to recombine known components in novel config

HumanLLM: Benchmarking and Improving LLM Anthropomorphism via Human Cognitive Patterns

SafetyDGX agent

arXiv:2601.10198v3 Announce Type: replace Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities in reasoning and generation, serving as the foundation for advanced persona s

Hybrid CNN-Transformer Architecture for Arabic Speech Emotion Recognition

ResearchDGX agent

arXiv:2604.07357v1 Announce Type: new Abstract: Recognizing emotions from speech using machine learning has become an active research area due to its importance in building human-centered applications

HyperMem: Hypergraph Memory for Long-Term Conversations

Model ReleasesDGX agent

arXiv:2604.08256v1 Announce Type: new Abstract: Long-term memory is essential for conversational agents to maintain coherence, track persistent tasks, and provide personalized interactions across exte

IatroBench: Pre-Registered Evidence of Iatrogenic Harm from AI Safety Measures

Model ReleasesDGX agent

arXiv:2604.07709v1 Announce Type: cross Abstract: Ask a frontier model how to taper six milligrams of alprazolam (psychiatrist retired, ten days of pills left, abrupt cessation causes seizures) and it

Initialisation Determines the Basin: Efficient Codebook Optimisation for Extreme LLM Quantization

Model ReleasesDGX agent

arXiv:2604.08118v1 Announce Type: new Abstract: Additive quantization enables extreme LLM compression with O(1) lookup-table dequantization, making it attractive for edge deployment. Yet at 2-bit prec

Iterative Formalization and Planning in Partially Observable Environments

ResearchDGX agent

arXiv:2505.13126v3 Announce Type: replace-cross Abstract: Using LLMs not to predict plans but to formalize an environment into the Planning Domain Definition Language (PDDL) has been shown to improve

Kathleen: Oscillator-Based Byte-Level Text Classification Without Tokenization or Attention

Model ReleasesDGX agent

arXiv:2604.07969v1 Announce Type: new Abstract: We present Kathleen, a text classification architecture that operates directly on raw UTF-8 bytes using frequency-domain processing -- requiring no toke

KEO: Knowledge Extraction on OMIn via Knowledge Graphs and RAG for Safety-Critical Aviation Maintenance

Model ReleasesDGX agent

arXiv:2510.05524v2 Announce Type: replace Abstract: We present Knowledge Extraction on OMIn (KEO), a domain-specific knowledge extraction and reasoning framework with large language models (LLMs) in s

KV Cache Offloading for Context-Intensive Tasks

Model ReleasesDGX agent

arXiv:2604.08426v1 Announce Type: cross Abstract: With the growing demand for long-context LLMs across a wide range of applications, the key-value (KV) cache has become a critical bottleneck for both

Large Language Model Post-Training: A Unified View of Off-Policy and On-Policy Learning

SafetyDGX agent

arXiv:2604.07941v1 Announce Type: new Abstract: Post-training has become central to turning pretrained large language models (LLMs) into aligned and deployable systems. Recent progress spans supervise

Learning is Forgetting: LLM Training As Lossy Compression

TutorialsDGX agent

arXiv:2604.07569v1 Announce Type: cross Abstract: Despite the increasing prevalence of large language models (LLMs), we still have a limited understanding of how their representational spaces are stru

Learning to Negotiate: Multi-Agent Deliberation for Collective Value Alignment in LLMs

SafetyDGX agent

arXiv:2603.10476v2 Announce Type: replace Abstract: LLM alignment has progressed in single-agent settings through paradigms such as RL with human feedback (RLHF), while recent work explores scalable a

Learning Who Disagrees: Demographic Importance Weighting for Modeling Annotator Distributions with DiADEM

SafetyDGX agent

arXiv:2604.08425v1 Announce Type: cross Abstract: When humans label subjective content, they disagree, and that disagreement is not noise. It reflects genuine differences in perspective shaped by anno

Lexical Tone is Hard to Quantize: Probing Discrete Speech Units in Mandarin and Yoruba

ResearchDGX agent

arXiv:2604.07467v1 Announce Type: new Abstract: Discrete speech units (DSUs) are derived by quantising representations from models trained using self-supervised learning (SSL). They are a popular repr

← Previous
1…124125126127128
Next →