AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
30 Jun 2026

Translationese as a Rational Response to Translation Task Difficulty

ApplicationsDGX agent

arXiv:2603.12050v2 Announce Type: replace Abstract: Translations systematically diverge from texts originally produced in the target language, a phenomenon widely referred to as translationese. Transl

Travel-Oriented Reasoning Large Language Model via Domain-Specific Knowledge Graphs

Model ReleasesDGX agent

arXiv:2606.29254v1 Announce Type: new Abstract: Large language models (LLMs) demonstrate broad reasoning abilities but struggle with accuracy and reliability in specialized domains such as travel, whe

TriageRA-CCF: Source-Side Clinical Confidence and Coverage Signals for Adaptive Rank Budgeting in Medical LLMs

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.29375v1 Announce Type: new Abstract: Medical large language models are commonly adapted with a fixed low-rank budget, even though medical questions differ substantially in confidence, clini

Turn-Averaged SAEs for Feature Discovery and Long-Context Attribution

ResearchDGX agent

arXiv:2606.28548v1 Announce Type: new Abstract: Sparse autoencoders (SAEs) have become a useful tool for extracting interpretable features in language models. However, standard SAE architectures opera

Uncertainty-Aware Generation and Decision-Making Under Ambiguity

ApplicationsDGX agent

arXiv:2606.30578v1 Announce Type: new Abstract: With rapidly improving capabilities, Large Language Models (LLMs) are increasingly used in many complex real-world tasks. Beyond requiring in-depth know

Uncovering Salience-Driven Dynamics in Consumer Confidence with Generative Social Simulation

SafetyDGX agent

arXiv:2606.30395v1 Announce Type: cross Abstract: Consumer confidence is typically modeled as a persistent macroeconomic index, yet its movements arise from households that interpret economic informat

Understanding Evaluation Illusion in Diffusion Large Language Models

ResearchDGX agent

arXiv:2606.29228v1 Announce Type: new Abstract: Despite the capability of parallel decoding, diffusion large language models (dLLMs) require many denoising steps to maintain generation quality, motiva

Unified Enhancement of the Generalization and Robustness of Language Models via Bi-Stage Optimization

Model ReleasesDGX agent

arXiv:2503.16550v2 Announce Type: replace Abstract: Neural network language models (LMs) are confronted with significant challenges in generalization and robustness. Currently, many studies focus on i

Unveiling Novelty Evolution in the field of Library and Information Science in China

ResearchDGX agent

arXiv:2606.29872v1 Announce Type: cross Abstract: This study analyzes the novelty distribution of scholarly papers in the field of Library and Information Science (LIS) in China, with a focus on diffe

wav2VOT: Automatic estimation of voice onset time, closure duration, and burst realisation with wav2vec2

ResearchDGX agent

arXiv:2606.28857v1 Announce Type: cross Abstract: While automatic tools for speech annotation are now commonplace within phonetic research pipelines, many tasks require substantial manual correction o

When Does Sparsity Mitigate the Curse of Depth in LLMs

ResearchDGX agent

arXiv:2603.15389v2 Announce Type: replace Abstract: Recent work has demonstrated the curse of depth in large language models (LLMs), where later layers contribute less to learning and representation t

When Is a Draft Accepted? A Theory of Acceptance in Speculative Decoding

Local AiDGX agent

arXiv:2606.30265v1 Announce Type: cross Abstract: Speculative decoding accelerates language model inference by using a fast drafter to propose candidate tokens that are then verified by a larger targe

Which Tokens Need Context? A Reference-Based Analysis of Translation Responsibility Using Fertility and Entropy

ResearchDGX agent

arXiv:2606.29489v1 Announce Type: new Abstract: When humans translate, not every word depends equally on the surrounding context. Some tokens, particularly function words like pronouns and auxiliaries

Who Plays Which Role When? Communication Role Dynamics for Peer Recognition and Team Performance Prediction

TutorialsDGX agent

arXiv:2606.28544v1 Announce Type: cross Abstract: Team roles offer an interpretable lens on collaboration, yet computational studies of roles often rely on domain-specific personas or data-driven clus

Why Struggle with Continuous Latents? Interpretable Discrete Latent Reasoning via Rendered Compression

Model ReleasesDGX agent

arXiv:2606.29712v1 Announce Type: new Abstract: Large language models achieve high reasoning performance via explicit chain-of-thought and reinforcement learning, but require long output sequences and

29 Jun 2026

A Survey of Automated Presentation Coaching: Systems, Methods, and Open Challenges

ResearchDGX agent

arXiv:2606.27380v1 Announce Type: new Abstract: Automated coaching for oral presentations sits at the intersection of computer-assisted pronunciation training (CAPT), prosody modeling, and speech synt

A Tree-of-Thoughts Inspired Hybrid Approach for Legal Case Judgement Summarization using LLMs

Model ReleasesDGX agent

arXiv:2606.28044v1 Announce Type: new Abstract: In recent times, Large Language Models (LLMs) are increasingly being used for legal case judgement summarization. Most prior works have tried traditiona

AI Persuasive Framing in Collective Dilemmas

ResearchDGX agent

arXiv:2606.27951v1 Announce Type: cross Abstract: AI agents are promising tools that can act as flexible behavioral nudges to enhance human cooperation in addressing large-scale societal problems. How

Aloe-Vision: Robust Vision-Language Models for Healthcare

Model ReleasesDGX agent

arXiv:2606.27500v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) specialized in healthcare are emerging as a promising research direction due to their potential impact in clinica

An Empirical Analysis of Factual Errors in Human-Written Text and its Application

Model ReleasesDGX agent

arXiv:2606.27959v1 Announce Type: new Abstract: Factual Error Detection (FED), which is the task of identifying factually incorrect spans in a given text, has long been recognized as an important rese

Causal Connections: Leveraging Multilingual Fine-Tuning for Financial QA@FinCausal 2026

Model ReleasesDGX agent

arXiv:2606.27446v1 Announce Type: new Abstract: This paper describes team HSA_CORAL's submission to the FinCausal 2026 shared task on extracting cause-effect relations from financial narratives via ex

Check Yourself Before You Wreck Yourself: Selectively Quitting Improves LLM Agent Safety

SafetyDGX agent

arXiv:2510.16492v4 Announce Type: replace Abstract: As Large Language Model (LLM) agents increasingly operate in complex environments with real-world consequences, their safety becomes critical. While

Cluster, Route, Escalate: Cascaded Framework for Cost-Aware LLM Serving

ApplicationsDGX agent

arXiv:2606.27457v1 Announce Type: cross Abstract: Efficient deployment of large language models (LLMs) in production forces a trade-off between accuracy and cost. Operators often default to a single m

Continual Memorization of Factoids in Language Models

ResearchDGX agent

arXiv:2411.07175v3 Announce Type: replace Abstract: As new knowledge rapidly accumulates, language models (LMs) with pretrained knowledge quickly become obsolete. A common approach to updating LMs is

Delayed Verification Destabilizes Multi-Agent LLM Belief: Instability Thresholds and Optimal Corrector Placement

AgentsDGX agent

arXiv:2606.27409v1 Announce Type: cross Abstract: Multi-agent large language model (LLM) systems often rely on verifier and critic agents to suppress hallucinations, but verification is delayed. Durin

Developmental approach reveals the statistical learning of Neural Language Models: Transformers generalize from the most abstract statistical patterns

TutorialsDGX agent

arXiv:2606.27460v1 Announce Type: new Abstract: In this study, we use a developmental approach to investigate the statistical learning and mental representation of neural language models (NLM). A seri

EntMTP: Accelerating LLM Inference with Entropy Guided Multi Token Prediction

ResearchDGX agent

arXiv:2606.27550v1 Announce Type: new Abstract: Multi-token prediction has been shown to increase data density during training, improve downstream text-generation quality, and serves as the defacto ap

Formalizing Latent Thoughts: Four Axioms of Thought Representation in LLMs

Model ReleasesDGX agent

arXiv:2606.27378v1 Announce Type: new Abstract: We introduce an axiomatic evaluation framework for latent thought representations in LLMs, comprising metrics that are independent of downstream benchma

HPRO: Hierarchical Progressive Reward Optimization via Preference Extraction for Emotional Text-to-Speech

TutorialsDGX agent

arXiv:2606.28249v1 Announce Type: cross Abstract: Recently, Large Language Model (LLM)-based Text-to-Speech (TTS) models have achieved remarkable naturalness. However, the standard Supervised Fine-Tun

Joint Transcription and Decryption of Images of Encrypted Handwritten Documents: A Comparison with the Traditional Pipeline

ApplicationsDGX agent

arXiv:2606.27700v1 Announce Type: cross Abstract: Historical encrypted manuscripts present a challenging problem at the intersection of cryptology, linguistics, paleography, and computer vision. Curre

Ko-WideSearch: A Korean Breadth-Search Benchmark for Exhaustive Set Enumeration by Web Agents

Model ReleasesDGX agent

arXiv:2606.27595v1 Announce Type: new Abstract: Web-agent benchmarks overwhelmingly measure depth -- pinning one obscure answer behind a chain of constraints -- while breadth, exhaustively enumerating

Learning Complementary Action Modeling from Automotive Maintenance Instructions

ResearchDGX agent

arXiv:2606.27808v1 Announce Type: new Abstract: A minute lexical variation can reverse the procedural meaning of an instruction even when the rest of the sentence remains unchanged. In automotive main

Learning to Evict from Key-Value Cache

Model ReleasesDGX agent

arXiv:2602.10238v2 Announce Type: replace Abstract: The growing size of Large Language Models (LLMs) makes efficient inference challenging, primarily due to the memory demands of the autoregressive Ke

Masked Language Flow Models

ResearchDGX agent

arXiv:2606.27617v1 Announce Type: new Abstract: Masked Diffusion Models (MDMs) promise fast, parallel language generation, but their reverse transition factorises across token positions -- an approxim

Mechanism-Driven Monitors for Preemptive Detection of LLM Training Instability

ResearchDGX agent

arXiv:2606.28116v1 Announce Type: new Abstract: Frontier large language model training consumes massive accelerator fleets and long wall-clock computation, making stability failures costly when they o

Mitigating Position Bias in Transformers via Layer-Specific Positional Embedding Scaling

SafetyDGX agent

arXiv:2606.27705v1 Announce Type: new Abstract: Large Language Models (LLMs) still struggle with the ``lost-in-the-middle'' problem, where critical information located in the middle of long-context in

Multimodal Evaluator Preference Collapse: Cross-Modal Coupling in Self-Evolving Agents

Model ReleasesDGX agent

arXiv:2606.16682v3 Announce Type: replace-cross Abstract: When AI agents use language models to evaluate their own outputs in a feedback loop, systematic biases emerge. We show that Evaluator Preferen

On the Effect of Uncertainty on Layer-wise Inference Dynamics

TutorialsDGX agent

arXiv:2507.06722v2 Announce Type: replace Abstract: Understanding how large language models (LLMs) internally represent and process their predictions is central to detecting uncertainty and preventing

Recall Before Rerank: Benchmarking Deep Learning Models for Large-Scale Code-to-Code Retrieval

ResearchDGX agent

arXiv:2606.27401v1 Announce Type: cross Abstract: Semantic code search and clone detection are essential for software development, maintenance, and reuse. This paper evaluates the effectiveness, effic

Retaining by Doing: The Role of On-Policy Data in Mitigating Forgetting

Model ReleasesDGX agent

arXiv:2510.18874v3 Announce Type: replace-cross Abstract: Adapting language models (LMs) to new tasks via post-training carries the risk of degrading existing capabilities -- a phenomenon classically

Safe Language Generation in the Limit

ApplicationsDGX agent

arXiv:2601.08648v2 Announce Type: replace Abstract: Recent results in learning a language in the limit have shown that, although language identification is impossible, language generation is tractable

Scaling limit of the Random Language Model

ResearchDGX agent

arXiv:2606.28105v1 Announce Type: cross Abstract: We develop a quantitative theory of the Random Language Model (RLM), an ensemble of stochastic context-free grammars, in a scaling limit where the num

Self-Stigma Is Not a Monolith, but Generic Empathy Is: Persona-Conditioned LLM Support for People Who Use Drugs

ResearchDGX agent

arXiv:2606.23387v2 Announce Type: replace Abstract: Self-stigma predicts treatment avoidance and disengagement among people who use drugs (PWUD), yet conversational systems aiming to provide support t

SIGNER: Temporally Grounded Sign Language Generation via Time-Resolved Conditioning

ResearchDGX agent

arXiv:2506.07460v2 Announce Type: replace-cross Abstract: Sign language generation (SLG), also known as text-to-sign generation, aims to bridge the communication gap between signers and non-signers. U

Textual Belief States for World Models: Identifiable Representation Learning Under Strict Mediation

TutorialsDGX agent

arXiv:2606.27681v1 Announce Type: cross Abstract: World models in partially observed environments rely on latent representations that summarize interaction history, but in many modern LLM-based archit

The Curse of Multiple Mediators: Hidden Interaction Effects in Activation Patching

ResearchDGX agent

arXiv:2606.27510v1 Announce Type: cross Abstract: Activation patching is the primary tool in mechanistic interpretability. It attributes causal responsibility for a model behavior to each of its indiv

The Signal-Coverage Matrix: Stratifying Type and Semantic Errors in Statement Autoformalization

Model ReleasesDGX agent

arXiv:2606.28013v1 Announce Type: new Abstract: Headline type-correctness (TC%) of LLM autoformalization has climbed from sim53% to sim76% in two years, yet this scalar conceals which errors each meth

ToxiREX: A Dataset on Toxic REasoning in ConteXt

ResearchDGX agent

arXiv:2606.27981v1 Announce Type: new Abstract: We introduce a new, contextual, multilingual dataset called ToxiREX: Toxic REasoning in ConteXt. The dataset consists of threads of Reddit comments and

Training-free Truthfulness Detection via Sparse MLP Value Vectors

Model ReleasesDGX agent

arXiv:2509.17932v2 Announce Type: replace Abstract: Large language models (LLMs) are prone to generating factually incorrect content, motivating methods for assessing truthfulness from internal model

Vision-Default, Prior-Override: Causal Mechanisms of Perception-Knowledge Conflict in Vision-Language Models

ResearchDGX agent

arXiv:2606.28273v1 Announce Type: new Abstract: Vision-language models must reconcile visual evidence with memorized world knowledge when the two conflict. How they resolve this conflict shapes the re

When Search Agents Should Ask: DiscoBench for Clarification-Aware Deep Search

Model ReleasesDGX agent

arXiv:2606.27669v1 Announce Type: new Abstract: Search agents powered by large language models (LLMs) are increasingly used to solve complex information-seeking tasks, requiring multi-step retrieval a

Yuvion LLM: An Adversarially-Aware Large Language Model for Content And AI Safety

Model ReleasesDGX agent

arXiv:2606.27632v1 Announce Type: new Abstract: As large language models are increasingly deployed in real-world systems, safety failures can still lead to harmful outputs and dangerous misuse. We arg

26 Jun 2026

A Systematic Survey of Semantic Role Labeling in the Era of Pretrained Language Models

ResearchDGX agent

arXiv:2502.08660v4 Announce Type: replace Abstract: Semantic role labeling (SRL) is a central natural language processing task for understanding predicate-argument structures within texts and enabling

Adversarial Diffusion Across Modalities: A Fusion Survey of Attacks, Defenses, and Evaluation for Text, Vision, and Vision-Language Models

Model ReleasesDGX agent

arXiv:2606.26566v1 Announce Type: cross Abstract: Adversarial evaluation of AI systems has matured along four largely disconnected tracks: diffusion-based attacks on text and large language models (LL

Analyzing and Encoding the Al-Mawrid Arabic-English Dictionary with the ISO Language Markup Framework and TEI Lex-0

ResearchDGX agent

arXiv:2606.18205v2 Announce Type: replace Abstract: This paper presents a robust methodology for the systematic digitization and encoding of the Al-Mawrid Arabic-English dictionary, transforming it fr

AnySimLite: A Lightweight Few-Shot Similarity Encoder for On-Device Speech-Adjacent Classification

Local AiDGX agent

arXiv:2606.26452v1 Announce Type: new Abstract: To minimize privacy concerns and inference latency on edge devices like smartphones, lightweight on-device models remain important for end-user applicat

Assessing Post-Reform Changes in Risk Disclosure Quality with a Multidimensional Text Analysis Approach

SafetyDGX agent

arXiv:2606.26522v1 Announce Type: new Abstract: While corporate narrative disclosures provide crucial information to capital markets, comprehensively evaluating their qualitative changes over time rem

Axon: A Synthesizing Superoptimizer for Tensor Programs

ResearchDGX agent

arXiv:2606.26344v1 Announce Type: cross Abstract: Writing high performance kernels for AI accelerators requires deep expertise in tiling, instruction selection, data layout, and operator fusion placin

Beyond Perplexity: UTF-8 Validity in Byte-aware Language Models

Model ReleasesDGX agent

arXiv:2606.14122v2 Announce Type: replace Abstract: Byte-level tokenization enables language models to handle any Unicode input, but models can generate invalid UTF-8 sequences when encountering rare

Beyond Surface Forms: A Comprehensive, Mechanism-Oriented Taxonomy of Indirect Linguistic Encoding for LLM-Based Coded Language Detection

Model ReleasesDGX agent

arXiv:2606.27314v1 Announce Type: new Abstract: To avoid moderation and surveillance on social media, some users routinely invent indirect linguistic expressions (ILE) that camouflage sensitive meanin

← Previous
1…3132333435…129
Next →