AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
24 Jun 2026

PEARL: Self-Evolving Assistant for Time Management with Reinforcement Learning

Model ReleasesDGX agent

arXiv:2601.11957v4 Announce Type: replace Abstract: Overlapping calendar invitations force busy professionals to repeatedly decide which meetings to attend, reschedule, or decline. We refer to this pr

PETRA: Transforming Web Text for Petroleum-Engineering Domain Adaptation

Model ReleasesDGX agent

arXiv:2606.24346v1 Announce Type: cross Abstract: Petroleum-engineering search exposes a supervision gap for strong general retrievers: relevant evidence exists in public web text, but domain relevanc

PORTER: Language-Grounded Event Representations for Portable Structured EHR Foundation Models

ResearchDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.24102v1 Announce Type: new Abstract: Most electronic health record (EHR) foundation models encode clinical events as discrete event tokens from a fixed vocabulary and therefore cannot direc

Posterior Refinement: Fast Language Generation via Any-Order Flow Maps

ResearchDGX agent

arXiv:2606.24773v1 Announce Type: new Abstract: Non-autoregressive generation offers a powerful paradigm for iterative refinement, allowing models to recursively critique, erase and regenerate arbitra

Prague Dependency Treebank -- Consolidated 2.0: Enriching a Complex Annotation Scheme

ResearchDGX agent

arXiv:2606.24324v1 Announce Type: new Abstract: The Prague Dependency Treebank framework is unique in its attempt to systematically include and link different layers of language, including a meaning r

Progressive Alignment Objectives for Aligner-Encoder based ASR

SafetyDGX agent

arXiv:2606.24147v1 Announce Type: cross Abstract: Aligner-Encoders are recently proposed seq2seq end-to-end ASR models that replace decoder attention by predicting the uth token directly from the u-th

QuechuaTok: Morphological Boundary Accuracy as a Necessary Metric for Tokenizer Evaluation in Agglutinative Low-Resource Languages

Model ReleasesDGX agent

arXiv:2606.23943v1 Announce Type: new Abstract: Tokenization is a foundational step in NLP pipelines, yet standard evaluation metrics such as fertility rate fail to capture morphological correctness f

Qwen-AgentWorld: Language World Models for General Agents

Model ReleasesDGX agent

arXiv:2606.24597v1 Announce Type: new Abstract: A world model predicts environment dynamics based on current observations and actions, serving as a core cognitive mechanism for reasoning and planning.

Removing Noise, not Finding Gold: Quality Filtering for Large-Scale Pretraining

ResearchDGX agent

arXiv:2510.00866v3 Announce Type: replace-cross Abstract: Large-scale models are pretrained on massive web-crawled datasets containing documents of mixed quality, making data filtering essential. A po

RoPE-Aware Bit Allocation for KV-Cache Quantization

Model ReleasesDGX agent

arXiv:2606.24033v1 Announce Type: cross Abstract: Existing low-bit KV-cache quantizers often treat each cached key as a flat vector. Under RoPE, however, a key's contribution to a future attention log

Same Lesson, Different Story: Cross-Lingual Reconstruction of Cultural Narratives in Large Language Models

ResearchDGX agent

arXiv:2606.24610v1 Announce Type: new Abstract: The evaluation of cultural grounding context becomes complex when multiple cultures convey the same moral lesson. This challenge is particularly relevan

SciZoom: A Large-scale Benchmark for Hierarchical Scientific Summarization across the LLM Era

Model ReleasesDGX agent

arXiv:2603.16131v2 Announce Type: replace Abstract: The explosive growth of AI research has created unprecedented information overload, increasing the demand for scientific summarization at multiple l

Sentence-Level Contextual Entrainment in Large Language Models

ResearchDGX agent

arXiv:2606.24077v1 Announce Type: new Abstract: Contextual entrainment, which is a newly discovered phenomenon in large language models (LLMs), refers to the tendency of a model to assign higher proba

SHERLOC: Structured Diagnostic Localization for Code Repair Agents

Local AiDGX agent

arXiv:2606.24820v1 Announce Type: new Abstract: LLM agents solve repository-level coding tasks through multi-turn tool use, but utilize half their budget on locating faults before editing. Dedicated l

SICI: A Semantic-Pragmatic Complexity Index Reveals Regime Shifts in LLM Stance Detection

Model ReleasesDGX agent

arXiv:2606.13189v2 Announce Type: replace Abstract: Prompt-based LLMs are increasingly used for stance detection, but harder examples are not always repaired by clearer instructions, reasoning prompts

The Sparse Frontier: Sparse Attention Trade-offs in Transformer LLMs

ApplicationsDGX agent

arXiv:2504.17768v3 Announce Type: replace Abstract: Sparse attention offers a promising strategy to extend long-context capabilities in Transformer LLMs, yet its efficiency-accuracy trade-offs remain

The Warrant Gap: Claim-Conditioned Re-scoring for Fact-Checking

ResearchDGX agent

arXiv:2606.24627v1 Announce Type: new Abstract: Fact-checking systems built on LLMs achieve high verdict accuracy on standard benchmarks, yet routinely output Supports labels whose cited evidence does

Thinking While Speaking: Inference-Time Knowledge Transfer for Responsive and Intelligent Conversational Voice Agents

AgentsDGX agent

arXiv:2511.07397v2 Announce Type: replace Abstract: Voice agents face a fundamental tension: the reasoning, retrieval, and tool use that make foundation models capable are iterative and slow, while co

To Compare, or Not to Compare: On Methodological Practices in Evaluating Social Bias

Model ReleasesDGX agent

arXiv:2606.24596v1 Announce Type: new Abstract: As Large Language Models are increasingly deployed in critical applications, robustly evaluating their social biases is paramount. However, the current

Transformer-Based Language Models Across Domain Verticals: Architectures, Applications and Critical Assessment

Model ReleasesDGX agent

arXiv:2606.24331v1 Announce Type: new Abstract: Transformer-based language models have become the default substrate for natural language processing and the pace of new releases has made it hard for pr

TruncProof: A Guardrail for LLM-based JSON Generation under Token-Length Constraints

SafetyDGX agent

arXiv:2605.13076v2 Announce Type: replace Abstract: The LLM-based generation of machine-readable outputs such as JSON has attracted significant attention for integration with external systems. However

UOL@IDEM at BEA 2026 Shared Task 1: Neural Fusion and Feature-Rich Modeling for L1-Aware Vocabulary Difficulty Prediction

SafetyDGX agent

arXiv:2606.24501v1 Announce Type: new Abstract: This paper describes UOL@IDEM's closed-track submission to the BEA 2026 shared task on L1-aware vocabulary difficulty prediction. We model the task as r

VieSpeaker: A Large-Scale Vietnamese Speaker Recognition Dataset Beyond Visual Dependency

ResearchDGX agent

arXiv:2606.24066v1 Announce Type: cross Abstract: Speaker recognition has advanced rapidly with large-scale training datasets, yet Vietnamese remains under-resourced, with existing corpora limited in

What's Missing in Vision-Language Models? Probing Their Struggles with Causal Order Reasoning

ResearchDGX agent

arXiv:2506.00869v3 Announce Type: replace Abstract: Despite the impressive performance of vision-language models (VLMs) on downstream tasks, their ability to understand and reason about causal relatio

When Top-1 Fails: Calibrating LoRA Monitors for Masked Diffusion LMs

Model ReleasesDGX agent

arXiv:2606.24119v1 Announce Type: cross Abstract: Discrete diffusion language model (DLM) fine-tuning inherits inexpensive diagnostics from denoising-time confidence monitors, but their PEFT-training

11 Jun 2026

3-Key-Input: Exploring the Theoretical Minimum Keys for Text Entry

ResearchDGX agent

arXiv:2606.11642v1 Announce Type: cross Abstract: How far can we reduce the number of physical keys if we endow an ambiguous keyboard with modern language models? Fewer keys increase hardware design f

A Controlled Study of Decoding-Time Truthfulness Methods on Instruction-Tuned LLMs

ResearchDGX agent

arXiv:2606.12160v1 Announce Type: new Abstract: In this work, we introduce CHAIR (Classifier of Hallucination As ImproveR), a supervised framework for detecting hallucinations by analyzing internal lo

A Geometric Profile of Semantic Information in Text: Frame-Conditional Uniqueness and a Trade-Off Triangle for Scalar Summaries

ResearchDGX agent

arXiv:2606.11222v1 Announce Type: new Abstract: How much meaning does a text carry? Shannon's theory measures uncertainty over symbols and is intentionally indifferent to meaning, while pairwise metri

A PubMed-Scale Dataset of Structured Biomedical Abstracts

Model ReleasesDGX agent

arXiv:2606.11361v1 Announce Type: cross Abstract: Structured abstracts are important for biomedical literature processing, by facilitating information retrieval, text mining, and knowledge synthesis.

A Resource for Enthymeme Detection in Controversial Political Discourse

ResearchDGX agent

arXiv:2606.12186v1 Announce Type: new Abstract: Enthymemes, arguments with unstated premises or conclusions, are pervasive in persuasive discourse, yet their annotation remains notoriously subjective.

Adaptive Multi-Resolution Procedural Knowledge Compression for Large Language Models

Model ReleasesDGX agent

arXiv:2606.12203v1 Announce Type: new Abstract: Large language models (LLMs) are widely used to tackle complex tasks with autonomous workflows. Recently, reusable natural language skills have emerged

Agent Skill Evaluation and Evolution: Frameworks and Benchmarks

Model ReleasesDGX agent

arXiv:2606.11435v1 Announce Type: new Abstract: The growth of agent skills has transformed how agentic systems are built, evaluated, and deployed. As skill libraries continue to scale, rigorous evalua

Agreement in Representation Space for Open-Ended Self-Consistency

ResearchDGX agent

arXiv:2606.12003v1 Announce Type: new Abstract: Self-consistency improves LLM reasoning by sampling multiple outputs and selecting the most consistent answer, but existing formulations largely rely on

AI Coding Agents Can Reproduce Social Science Findings

Model ReleasesDGX agent

arXiv:2606.11447v1 Announce Type: new Abstract: Recent anecdotal evidence suggests that AI coding agents can reproduce published findings when provided with original data and code; yet systematic eval

AI4SLT: Empirical Processes in Lean 4 for Formal Statistical Learning Theory

ResearchDGX agent

arXiv:2602.02285v2 Announce Type: replace-cross Abstract: We present the first comprehensive Lean 4 formalization of statistical learning theory (SLT) grounded in empirical process theory. Our en-to-e

An Ontology-Guided Multi-Anchor Graph Retrieval Framework for Traffic Legal Liability Determination

Model ReleasesDGX agent

arXiv:2606.11910v1 Announce Type: new Abstract: Traffic law liability determination is critical for assigning legal penalties, requiring the simultaneous identification of interdependent statutory pro

Benchmarking Large Language Models for Safety Data Extraction

Model ReleasesDGX agent

arXiv:2606.11204v1 Announce Type: new Abstract: Accurate extraction of structured information from Safety Data Sheets (SDS) remains challenging in industrial safety due to heterogeneous document forma

Beyond Compaction: Structured Context Eviction for Long-Horizon Agents

Model ReleasesDGX agent

arXiv:2606.11213v1 Announce Type: new Abstract: We present Context Window Lifecycle (CWL), a context-management scheme that gives long-horizon LLM agents an effectively unbounded working horizon. As a

Beyond Fully Random Masking: Attention-Guided Denoising and Optimization for Diffusion Language Models

ResearchDGX agent

arXiv:2606.12273v1 Announce Type: new Abstract: Diffusion large language models (dLLMs) offer an efficient alternative to autoregressive models through parallel decoding, yet existing post-training me

Beyond Third-Person Audits: Situated Interaction Auditing for User-Centered LLM Bias Research

SafetyDGX agent

arXiv:2606.12247v1 Announce Type: cross Abstract: Research on bias in large language models (LLMs) has predominantly focused on third-person audits, which study how models represent or evaluate demogr

BioMamba: Domain-Adaptive Biomedical Language Models

Model ReleasesDGX agent

arXiv:2408.02600v3 Announce Type: replace Abstract: Background. Biomedical language models should improve performance on biomedical text while retaining general-language-modeling fluency. For Mamba-ba

Breaking Entropy Bounds: Accelerating RL Training via MTP with Rejection Sampling

AgentsDGX agent

arXiv:2606.12370v1 Announce Type: cross Abstract: Reinforcement learning (RL) has become a key component in modern large language models, yet the rollout stage remains the key bottleneck in RL trainin

Building Social World Models with Large Language Models

Model ReleasesDGX agent

arXiv:2606.11482v1 Announce Type: cross Abstract: Understanding and predicting how social beliefs evolve in response to events -- from policy changes to scientific breakthroughs -- remains a fundament

Can AI Reason Like an Urban Planner? Benchmarking Large Language Models Against Professional Judgment

SafetyDGX agent

arXiv:2606.11678v1 Announce Type: new Abstract: Problem, Research Strategy, and Findings: The rise of large language models (LLMs) raises a key question for urban planning: which forms of professional

Can News Predict the Market? Limits of Zero-Shot Financial NLP and the Role of Explainable AI

ResearchDGX agent

arXiv:2606.12210v1 Announce Type: new Abstract: Can financial news reliably predict short-term stock movements? Despite advances in large language models, this question remains unresolved. We revisit

Claw-SWE-Bench: A Benchmark for Evaluating OpenClaw-style Agent Harnesses on Coding Tasks

Model ReleasesDGX agent

arXiv:2606.12344v1 Announce Type: cross Abstract: General-purpose agents such as OpenClaw are increasingly used as autonomous tool users, but their coding ability is difficult to measure under SWE-ben

Compatibility-Aware Dynamic Fine-Tuning for Large Language Models

SafetyDGX agent

arXiv:2606.11206v1 Announce Type: new Abstract: Supervised Fine-Tuning (SFT) is the predominant paradigm for aligning large language models (LLMs), yet it suffers from optimization instability and lim

Context-Aware Multimodal Claim Verification in Spoken Dialogues

Model ReleasesDGX agent

arXiv:2606.11420v1 Announce Type: new Abstract: Every day, millions absorb claims from podcasts and streams that no fact-checker ever sees. Spoken misinformation is built through conversation, where c

Context-Driven Incremental Compression for Multi-Turn Dialogue Generation

ResearchDGX agent

arXiv:2606.12411v1 Announce Type: new Abstract: Modern conversational agents condition on an ever-growing dialogue history at each turn, incurring redundant attention and encoding costs that grow with

Debiasing Without Protected Attributes: Latent Concept Erasure from Textual Profiles

Model ReleasesDGX agent

arXiv:2606.12088v1 Announce Type: new Abstract: Most fairness research in NLP assumes direct access to protected attributes such as gender, race, or nationality. In practice, however, such information

Decoding Multimodal Cues: Unveiling the Implicit Meaning Behind Hateful Videos

ResearchDGX agent

arXiv:2606.11953v1 Announce Type: new Abstract: Hateful videos have become prevalent on online platforms, highlighting an urgent need for effective detection. However, existing studies primarily focus

Detecting AI-Generated Content on Social Media with Multi-modal Language Models

ApplicationsDGX agent

arXiv:2606.11200v1 Announce Type: new Abstract: Generative AI has enabled the creation of photorealistic images and videos that are increasingly disseminated on social media, often used for spam, misi

Detecting Sensitive Personal Information in Japanese Pre-Training Corpora for Large Language Models

ResearchDGX agent

arXiv:2606.12114v1 Announce Type: new Abstract: Sensitive personal information can appear in large-scale pre-training corpora for large language models (LLMs). Detecting and filtering such information

Doc-to-Atom: Learning to Compile and Compose Memory Atoms

ResearchDGX agent

arXiv:2606.12400v1 Announce Type: new Abstract: Long input sequences are central to document understanding and multi-step reasoning in Large Language Models, yet the quadratic cost of attention makes

Dummy Backdoor as a Defense: Removing Unknown Backdoors via Shared Internal Mechanisms for Generative LLMs

SafetyDGX agent

arXiv:2606.11648v1 Announce Type: cross Abstract: Backdoor attacks pose a serious threat to the safety and reliability of Large Language Models (LLMs), as they cause models to behave normally on clean

Energy-Efficient On-Device RAG on a Mobile NPU: System Design and Benchmark on Snapdragon X Elite

Model ReleasesDGX agent

arXiv:2606.11257v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) pipelines are compute-intensive, combining embedding, retrieval, reranking, and large language model (LLM) generati

Evaluating Bias in Phoneme-Based Automatic Speech Recognition Systems: An Analysis of IPA Transcription Models

SafetyDGX agent

arXiv:2606.11639v1 Announce Type: new Abstract: The popularization of automatic speech recognition (ASR) systems has increased exploration of the demographic biases related to race, age, gender, and a

EverydayGPT: Confidence-Gated Routing for Efficient and Safe Hybrid GPT-RAG Conversational QA

Model ReleasesDGX agent

arXiv:2606.11212v1 Announce Type: new Abstract: Standard Retrieval-Augmented Generation (RAG) pipelines route every query through retrieval and generation unconditionally, incurring unnecessary comput

External Experience Serving in Production LLM Systems: A Deployment-Oriented Study of Quality-Cost Trade-offs

SafetyDGX agent

arXiv:2606.11806v1 Announce Type: new Abstract: Production LLM systems accumulate reusable operational experience, but the practical deployment issue is not merely whether such experience can help. It

Factions Within, Uncertain Across: Within-Document Reader Sub-Groups in Social Highlighting

ResearchDGX agent

arXiv:2606.11613v1 Announce Type: cross Abstract: When many people highlight the same document, is the crowd a single consensus, or is it internally structured into reader sub-groups that mark differe

← Previous
1…3637383940…129
Next →