AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
Research

Layer-wise Probing of wav2vec 2.0 and Whisper for Consonant Cluster Reduction in African American English

DGX agent

arXiv:2606.23948v1 Announce Type: new Abstract: Self-supervised and supervised speech models are increasingly used to investigate which linguistic information their internal representations encode, an

researcharxiv-cs-cl
24 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

Less is More: Quality-Aware Training Data Selection for Scientific Summarization

DGX agent

arXiv:2606.24828v1 Announce Type: new Abstract: Scientific long-document summarization datasets commonly treat author-written abstracts as gold reference summaries, although their quality and alignmen

safetyarxiv-cs-cl
24 Jun 2026
Safety

Light-weight Pronunciation Assessment via Discrete Speech Token Surprisal

DGX agent

arXiv:2606.19910v2 Announce Type: replace Abstract: Training automated pronunciation assessment often relies on labeled learner errors or non-native corpora that are costly to collect. We propose a li

safetyarxiv-cs-cl
24 Jun 2026
Research

Measuring User's Mental Models of Speech Translation in Human-AI Collaboration

DGX agent

arXiv:2606.24644v1 Announce Type: new Abstract: Millions of people use machine translation (MT) tools daily, yet little is known about their perception of what systems can and cannot do. This paper st

researcharxiv-cs-cl
24 Jun 2026
Model Releases

MedBench v5: A Dynamic, Process-Oriented, and Hallucination-Aware Benchmark for Clinical Multimodal Models

DGX agent

arXiv:2606.24155v1 Announce Type: new Abstract: Existing medical AI benchmarks lack process visibility, atomic skill evaluation, and integrated hallucination detection. We introduce MedBench v5, a red

model-releasesarxiv-cs-cl
24 Jun 2026
Research

Meet UD_Czech-PDTC: A Large and Genre-Rich Treebank in Universal Dependencies

DGX agent

arXiv:2606.24337v1 Announce Type: new Abstract: Czech has been part of Universal Dependencies since its first release in 2015. It has also been one of the best represented languages, with the Prague D

researcharxiv-cs-cl
24 Jun 2026
Model Releases

MEMPROBE: Probing Long-Term Agent Memory via Hidden User-State Recovery

DGX agent

arXiv:2606.24595v1 Announce Type: new Abstract: Long-term memory promises LLM agents that grow more capable across sessions, maintaining an accurate, evolving understanding of the user that interactio

model-releasesarxiv-cs-cl
24 Jun 2026
Research

MERGE: Minimal Expression-Replacement GEneralization Test for Natural Language Inference

DGX agent

arXiv:2510.24295v2 Announce Type: replace Abstract: As many benchmarks have become saturated, it has become increasingly important to create new datasets that evaluate the generalization capacity of c

researcharxiv-cs-cl
24 Jun 2026
Research

ModTGCN: Modularity-aware Graph Neural Networks for Text Classification

DGX agent

arXiv:2606.23694v1 Announce Type: new Abstract: Graph-based text classification models typically rely on local neighborhood aggregation and overlook global community structure, despite semantic docume

researcharxiv-cs-cl
24 Jun 2026
Research

MorfFlex: Handling Rich Morphology

DGX agent

arXiv:2606.24366v1 Announce Type: new Abstract: We present MorfFlex, a morphological dictionary architecture suitable for languages with extensive regularity in both inflection and derivation. As the

researcharxiv-cs-cl
24 Jun 2026
Model Releases

NatureBench: Can Coding Agents Match the Published SOTA of Nature-Family Papers?

DGX agent

arXiv:2606.24530v1 Announce Type: new Abstract: We introduce NatureBench, a cross-discipline benchmark of 90 tasks distilled from peer-reviewed Nature-family publications, designed to evaluate whether

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

ParaPairAudioBench: Paralinguistic Pairwise Audio Benchmark for LALM-as-a-Judge

DGX agent

arXiv:2606.24648v1 Announce Type: cross Abstract: Large Audio-Language Models (LALMs) have been widely used as judge models for the automatic evaluation of generated speech. However, prior approaches

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

PEARL: Self-Evolving Assistant for Time Management with Reinforcement Learning

DGX agent

arXiv:2601.11957v4 Announce Type: replace Abstract: Overlapping calendar invitations force busy professionals to repeatedly decide which meetings to attend, reschedule, or decline. We refer to this pr

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

PETRA: Transforming Web Text for Petroleum-Engineering Domain Adaptation

DGX agent

arXiv:2606.24346v1 Announce Type: cross Abstract: Petroleum-engineering search exposes a supervision gap for strong general retrievers: relevant evidence exists in public web text, but domain relevanc

model-releasesarxiv-cs-cl
24 Jun 2026
Research

PORTER: Language-Grounded Event Representations for Portable Structured EHR Foundation Models

DGX agent

arXiv:2606.24102v1 Announce Type: new Abstract: Most electronic health record (EHR) foundation models encode clinical events as discrete event tokens from a fixed vocabulary and therefore cannot direc

researcharxiv-cs-cl
24 Jun 2026
Research

Posterior Refinement: Fast Language Generation via Any-Order Flow Maps

DGX agent

arXiv:2606.24773v1 Announce Type: new Abstract: Non-autoregressive generation offers a powerful paradigm for iterative refinement, allowing models to recursively critique, erase and regenerate arbitra

researcharxiv-cs-cl
24 Jun 2026
Research

Prague Dependency Treebank -- Consolidated 2.0: Enriching a Complex Annotation Scheme

DGX agent

arXiv:2606.24324v1 Announce Type: new Abstract: The Prague Dependency Treebank framework is unique in its attempt to systematically include and link different layers of language, including a meaning r

researcharxiv-cs-cl
24 Jun 2026
Safety

Progressive Alignment Objectives for Aligner-Encoder based ASR

DGX agent

arXiv:2606.24147v1 Announce Type: cross Abstract: Aligner-Encoders are recently proposed seq2seq end-to-end ASR models that replace decoder attention by predicting the uth token directly from the u-th

safetyarxiv-cs-cl
24 Jun 2026
Model Releases

QuechuaTok: Morphological Boundary Accuracy as a Necessary Metric for Tokenizer Evaluation in Agglutinative Low-Resource Languages

DGX agent

arXiv:2606.23943v1 Announce Type: new Abstract: Tokenization is a foundational step in NLP pipelines, yet standard evaluation metrics such as fertility rate fail to capture morphological correctness f

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

Qwen-AgentWorld: Language World Models for General Agents

DGX agent

arXiv:2606.24597v1 Announce Type: new Abstract: A world model predicts environment dynamics based on current observations and actions, serving as a core cognitive mechanism for reasoning and planning.

model-releasesarxiv-cs-cl
24 Jun 2026
Research

Removing Noise, not Finding Gold: Quality Filtering for Large-Scale Pretraining

DGX agent

arXiv:2510.00866v3 Announce Type: replace-cross Abstract: Large-scale models are pretrained on massive web-crawled datasets containing documents of mixed quality, making data filtering essential. A po

researcharxiv-cs-cl
24 Jun 2026
Model Releases

RoPE-Aware Bit Allocation for KV-Cache Quantization

DGX agent

arXiv:2606.24033v1 Announce Type: cross Abstract: Existing low-bit KV-cache quantizers often treat each cached key as a flat vector. Under RoPE, however, a key's contribution to a future attention log

model-releasesarxiv-cs-cl
24 Jun 2026
Research

Same Lesson, Different Story: Cross-Lingual Reconstruction of Cultural Narratives in Large Language Models

DGX agent

arXiv:2606.24610v1 Announce Type: new Abstract: The evaluation of cultural grounding context becomes complex when multiple cultures convey the same moral lesson. This challenge is particularly relevan

researcharxiv-cs-cl
24 Jun 2026
Model Releases

SciZoom: A Large-scale Benchmark for Hierarchical Scientific Summarization across the LLM Era

DGX agent

arXiv:2603.16131v2 Announce Type: replace Abstract: The explosive growth of AI research has created unprecedented information overload, increasing the demand for scientific summarization at multiple l

model-releasesarxiv-cs-cl
24 Jun 2026
Research

Sentence-Level Contextual Entrainment in Large Language Models

DGX agent

arXiv:2606.24077v1 Announce Type: new Abstract: Contextual entrainment, which is a newly discovered phenomenon in large language models (LLMs), refers to the tendency of a model to assign higher proba

researcharxiv-cs-cl
24 Jun 2026
Local Ai

SHERLOC: Structured Diagnostic Localization for Code Repair Agents

DGX agent

arXiv:2606.24820v1 Announce Type: new Abstract: LLM agents solve repository-level coding tasks through multi-turn tool use, but utilize half their budget on locating faults before editing. Dedicated l

local-aiarxiv-cs-cl
24 Jun 2026
Model Releases

SICI: A Semantic-Pragmatic Complexity Index Reveals Regime Shifts in LLM Stance Detection

DGX agent

arXiv:2606.13189v2 Announce Type: replace Abstract: Prompt-based LLMs are increasingly used for stance detection, but harder examples are not always repaired by clearer instructions, reasoning prompts

model-releasesarxiv-cs-cl
24 Jun 2026
Applications

The Sparse Frontier: Sparse Attention Trade-offs in Transformer LLMs

DGX agent

arXiv:2504.17768v3 Announce Type: replace Abstract: Sparse attention offers a promising strategy to extend long-context capabilities in Transformer LLMs, yet its efficiency-accuracy trade-offs remain

applicationsarxiv-cs-cl
24 Jun 2026
Research

The Warrant Gap: Claim-Conditioned Re-scoring for Fact-Checking

DGX agent

arXiv:2606.24627v1 Announce Type: new Abstract: Fact-checking systems built on LLMs achieve high verdict accuracy on standard benchmarks, yet routinely output Supports labels whose cited evidence does

researcharxiv-cs-cl
24 Jun 2026
Agents

Thinking While Speaking: Inference-Time Knowledge Transfer for Responsive and Intelligent Conversational Voice Agents

DGX agent

arXiv:2511.07397v2 Announce Type: replace Abstract: Voice agents face a fundamental tension: the reasoning, retrieval, and tool use that make foundation models capable are iterative and slow, while co

agentsarxiv-cs-cl
24 Jun 2026
Model Releases

To Compare, or Not to Compare: On Methodological Practices in Evaluating Social Bias

DGX agent

arXiv:2606.24596v1 Announce Type: new Abstract: As Large Language Models are increasingly deployed in critical applications, robustly evaluating their social biases is paramount. However, the current

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

Transformer-Based Language Models Across Domain Verticals: Architectures, Applications and Critical Assessment

DGX agent

arXiv:2606.24331v1 Announce Type: new Abstract: Transformer-based language models have become the default substrate for natural language processing and the pace of new releases has made it hard for pr

model-releasesarxiv-cs-cl
24 Jun 2026
Safety

TruncProof: A Guardrail for LLM-based JSON Generation under Token-Length Constraints

DGX agent

arXiv:2605.13076v2 Announce Type: replace Abstract: The LLM-based generation of machine-readable outputs such as JSON has attracted significant attention for integration with external systems. However

safetyarxiv-cs-cl
24 Jun 2026
Safety

UOL@IDEM at BEA 2026 Shared Task 1: Neural Fusion and Feature-Rich Modeling for L1-Aware Vocabulary Difficulty Prediction

DGX agent

arXiv:2606.24501v1 Announce Type: new Abstract: This paper describes UOL@IDEM's closed-track submission to the BEA 2026 shared task on L1-aware vocabulary difficulty prediction. We model the task as r

safetyarxiv-cs-cl
24 Jun 2026
Research

VieSpeaker: A Large-Scale Vietnamese Speaker Recognition Dataset Beyond Visual Dependency

DGX agent

arXiv:2606.24066v1 Announce Type: cross Abstract: Speaker recognition has advanced rapidly with large-scale training datasets, yet Vietnamese remains under-resourced, with existing corpora limited in

researcharxiv-cs-cl
24 Jun 2026
Research

What's Missing in Vision-Language Models? Probing Their Struggles with Causal Order Reasoning

DGX agent

arXiv:2506.00869v3 Announce Type: replace Abstract: Despite the impressive performance of vision-language models (VLMs) on downstream tasks, their ability to understand and reason about causal relatio

researcharxiv-cs-cl
24 Jun 2026
Model Releases

When Top-1 Fails: Calibrating LoRA Monitors for Masked Diffusion LMs

DGX agent

arXiv:2606.24119v1 Announce Type: cross Abstract: Discrete diffusion language model (DLM) fine-tuning inherits inexpensive diagnostics from denoising-time confidence monitors, but their PEFT-training

model-releasesarxiv-cs-cl
24 Jun 2026
Research

3-Key-Input: Exploring the Theoretical Minimum Keys for Text Entry

DGX agent

arXiv:2606.11642v1 Announce Type: cross Abstract: How far can we reduce the number of physical keys if we endow an ambiguous keyboard with modern language models? Fewer keys increase hardware design f

researcharxiv-cs-cl
11 Jun 2026
Research

A Controlled Study of Decoding-Time Truthfulness Methods on Instruction-Tuned LLMs

DGX agent

arXiv:2606.12160v1 Announce Type: new Abstract: In this work, we introduce CHAIR (Classifier of Hallucination As ImproveR), a supervised framework for detecting hallucinations by analyzing internal lo

researcharxiv-cs-cl
11 Jun 2026
Research

A Geometric Profile of Semantic Information in Text: Frame-Conditional Uniqueness and a Trade-Off Triangle for Scalar Summaries

DGX agent

arXiv:2606.11222v1 Announce Type: new Abstract: How much meaning does a text carry? Shannon's theory measures uncertainty over symbols and is intentionally indifferent to meaning, while pairwise metri

researcharxiv-cs-cl
11 Jun 2026
Model Releases

A PubMed-Scale Dataset of Structured Biomedical Abstracts

DGX agent

arXiv:2606.11361v1 Announce Type: cross Abstract: Structured abstracts are important for biomedical literature processing, by facilitating information retrieval, text mining, and knowledge synthesis.

model-releasesarxiv-cs-cl
11 Jun 2026
Research

A Resource for Enthymeme Detection in Controversial Political Discourse

DGX agent

arXiv:2606.12186v1 Announce Type: new Abstract: Enthymemes, arguments with unstated premises or conclusions, are pervasive in persuasive discourse, yet their annotation remains notoriously subjective.

researcharxiv-cs-cl
11 Jun 2026
Model Releases

Adaptive Multi-Resolution Procedural Knowledge Compression for Large Language Models

DGX agent

arXiv:2606.12203v1 Announce Type: new Abstract: Large language models (LLMs) are widely used to tackle complex tasks with autonomous workflows. Recently, reusable natural language skills have emerged

model-releasesarxiv-cs-cl
11 Jun 2026
Model Releases

Agent Skill Evaluation and Evolution: Frameworks and Benchmarks

DGX agent

arXiv:2606.11435v1 Announce Type: new Abstract: The growth of agent skills has transformed how agentic systems are built, evaluated, and deployed. As skill libraries continue to scale, rigorous evalua

model-releasesarxiv-cs-cl
11 Jun 2026
Research

Agreement in Representation Space for Open-Ended Self-Consistency

DGX agent

arXiv:2606.12003v1 Announce Type: new Abstract: Self-consistency improves LLM reasoning by sampling multiple outputs and selecting the most consistent answer, but existing formulations largely rely on

researcharxiv-cs-cl
11 Jun 2026
Model Releases

AI Coding Agents Can Reproduce Social Science Findings

DGX agent

arXiv:2606.11447v1 Announce Type: new Abstract: Recent anecdotal evidence suggests that AI coding agents can reproduce published findings when provided with original data and code; yet systematic eval

model-releasesarxiv-cs-cl
11 Jun 2026
Research

AI4SLT: Empirical Processes in Lean 4 for Formal Statistical Learning Theory

DGX agent

arXiv:2602.02285v2 Announce Type: replace-cross Abstract: We present the first comprehensive Lean 4 formalization of statistical learning theory (SLT) grounded in empirical process theory. Our en-to-e

researcharxiv-cs-cl
11 Jun 2026
Model Releases

An Ontology-Guided Multi-Anchor Graph Retrieval Framework for Traffic Legal Liability Determination

DGX agent

arXiv:2606.11910v1 Announce Type: new Abstract: Traffic law liability determination is critical for assigning legal penalties, requiring the simultaneous identification of interdependent statutory pro

model-releasesarxiv-cs-cl
11 Jun 2026
← Previous
1…4546474849…161
Next →