AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
Applications

Translationese as a Rational Response to Translation Task Difficulty

DGX agent

arXiv:2603.12050v2 Announce Type: replace Abstract: Translations systematically diverge from texts originally produced in the target language, a phenomenon widely referred to as translationese. Transl

applicationsarxiv-cs-cl
30 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Travel-Oriented Reasoning Large Language Model via Domain-Specific Knowledge Graphs

DGX agent

arXiv:2606.29254v1 Announce Type: new Abstract: Large language models (LLMs) demonstrate broad reasoning abilities but struggle with accuracy and reliability in specialized domains such as travel, whe

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

TriageRA-CCF: Source-Side Clinical Confidence and Coverage Signals for Adaptive Rank Budgeting in Medical LLMs

DGX agent

arXiv:2606.29375v1 Announce Type: new Abstract: Medical large language models are commonly adapted with a fixed low-rank budget, even though medical questions differ substantially in confidence, clini

model-releasesarxiv-cs-cl
30 Jun 2026
Research

Turn-Averaged SAEs for Feature Discovery and Long-Context Attribution

DGX agent

arXiv:2606.28548v1 Announce Type: new Abstract: Sparse autoencoders (SAEs) have become a useful tool for extracting interpretable features in language models. However, standard SAE architectures opera

researcharxiv-cs-cl
30 Jun 2026
Applications

Uncertainty-Aware Generation and Decision-Making Under Ambiguity

DGX agent

arXiv:2606.30578v1 Announce Type: new Abstract: With rapidly improving capabilities, Large Language Models (LLMs) are increasingly used in many complex real-world tasks. Beyond requiring in-depth know

applicationsarxiv-cs-cl
30 Jun 2026
Safety

Uncovering Salience-Driven Dynamics in Consumer Confidence with Generative Social Simulation

DGX agent

arXiv:2606.30395v1 Announce Type: cross Abstract: Consumer confidence is typically modeled as a persistent macroeconomic index, yet its movements arise from households that interpret economic informat

safetyarxiv-cs-cl
30 Jun 2026
Research

Understanding Evaluation Illusion in Diffusion Large Language Models

DGX agent

arXiv:2606.29228v1 Announce Type: new Abstract: Despite the capability of parallel decoding, diffusion large language models (dLLMs) require many denoising steps to maintain generation quality, motiva

researcharxiv-cs-cl
30 Jun 2026
Model Releases

Unified Enhancement of the Generalization and Robustness of Language Models via Bi-Stage Optimization

DGX agent

arXiv:2503.16550v2 Announce Type: replace Abstract: Neural network language models (LMs) are confronted with significant challenges in generalization and robustness. Currently, many studies focus on i

model-releasesarxiv-cs-cl
30 Jun 2026
Research

Unveiling Novelty Evolution in the field of Library and Information Science in China

DGX agent

arXiv:2606.29872v1 Announce Type: cross Abstract: This study analyzes the novelty distribution of scholarly papers in the field of Library and Information Science (LIS) in China, with a focus on diffe

researcharxiv-cs-cl
30 Jun 2026
Research

wav2VOT: Automatic estimation of voice onset time, closure duration, and burst realisation with wav2vec2

DGX agent

arXiv:2606.28857v1 Announce Type: cross Abstract: While automatic tools for speech annotation are now commonplace within phonetic research pipelines, many tasks require substantial manual correction o

researcharxiv-cs-cl
30 Jun 2026
Research

When Does Sparsity Mitigate the Curse of Depth in LLMs

DGX agent

arXiv:2603.15389v2 Announce Type: replace Abstract: Recent work has demonstrated the curse of depth in large language models (LLMs), where later layers contribute less to learning and representation t

researcharxiv-cs-cl
30 Jun 2026
Local Ai

When Is a Draft Accepted? A Theory of Acceptance in Speculative Decoding

DGX agent

arXiv:2606.30265v1 Announce Type: cross Abstract: Speculative decoding accelerates language model inference by using a fast drafter to propose candidate tokens that are then verified by a larger targe

local-aiarxiv-cs-cl
30 Jun 2026
Research

Which Tokens Need Context? A Reference-Based Analysis of Translation Responsibility Using Fertility and Entropy

DGX agent

arXiv:2606.29489v1 Announce Type: new Abstract: When humans translate, not every word depends equally on the surrounding context. Some tokens, particularly function words like pronouns and auxiliaries

researcharxiv-cs-cl
30 Jun 2026
Tutorials

Who Plays Which Role When? Communication Role Dynamics for Peer Recognition and Team Performance Prediction

DGX agent

arXiv:2606.28544v1 Announce Type: cross Abstract: Team roles offer an interpretable lens on collaboration, yet computational studies of roles often rely on domain-specific personas or data-driven clus

tutorialsarxiv-cs-cl
30 Jun 2026
Model Releases

Why Struggle with Continuous Latents? Interpretable Discrete Latent Reasoning via Rendered Compression

DGX agent

arXiv:2606.29712v1 Announce Type: new Abstract: Large language models achieve high reasoning performance via explicit chain-of-thought and reinforcement learning, but require long output sequences and

model-releasesarxiv-cs-cl
30 Jun 2026
Research

A Survey of Automated Presentation Coaching: Systems, Methods, and Open Challenges

DGX agent

arXiv:2606.27380v1 Announce Type: new Abstract: Automated coaching for oral presentations sits at the intersection of computer-assisted pronunciation training (CAPT), prosody modeling, and speech synt

researcharxiv-cs-cl
29 Jun 2026
Model Releases

A Tree-of-Thoughts Inspired Hybrid Approach for Legal Case Judgement Summarization using LLMs

DGX agent

arXiv:2606.28044v1 Announce Type: new Abstract: In recent times, Large Language Models (LLMs) are increasingly being used for legal case judgement summarization. Most prior works have tried traditiona

model-releasesarxiv-cs-cl
29 Jun 2026
Research

AI Persuasive Framing in Collective Dilemmas

DGX agent

arXiv:2606.27951v1 Announce Type: cross Abstract: AI agents are promising tools that can act as flexible behavioral nudges to enhance human cooperation in addressing large-scale societal problems. How

researcharxiv-cs-cl
29 Jun 2026
Model Releases

Aloe-Vision: Robust Vision-Language Models for Healthcare

DGX agent

arXiv:2606.27500v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) specialized in healthcare are emerging as a promising research direction due to their potential impact in clinica

model-releasesarxiv-cs-cl
29 Jun 2026
Model Releases

An Empirical Analysis of Factual Errors in Human-Written Text and its Application

DGX agent

arXiv:2606.27959v1 Announce Type: new Abstract: Factual Error Detection (FED), which is the task of identifying factually incorrect spans in a given text, has long been recognized as an important rese

model-releasesarxiv-cs-cl
29 Jun 2026
Model Releases

Causal Connections: Leveraging Multilingual Fine-Tuning for Financial QA@FinCausal 2026

DGX agent

arXiv:2606.27446v1 Announce Type: new Abstract: This paper describes team HSA_CORAL's submission to the FinCausal 2026 shared task on extracting cause-effect relations from financial narratives via ex

model-releasesarxiv-cs-cl
29 Jun 2026
Safety

Check Yourself Before You Wreck Yourself: Selectively Quitting Improves LLM Agent Safety

DGX agent

arXiv:2510.16492v4 Announce Type: replace Abstract: As Large Language Model (LLM) agents increasingly operate in complex environments with real-world consequences, their safety becomes critical. While

safetyarxiv-cs-cl
29 Jun 2026
Applications

Cluster, Route, Escalate: Cascaded Framework for Cost-Aware LLM Serving

DGX agent

arXiv:2606.27457v1 Announce Type: cross Abstract: Efficient deployment of large language models (LLMs) in production forces a trade-off between accuracy and cost. Operators often default to a single m

applicationsarxiv-cs-cl
29 Jun 2026
Research

Continual Memorization of Factoids in Language Models

DGX agent

arXiv:2411.07175v3 Announce Type: replace Abstract: As new knowledge rapidly accumulates, language models (LMs) with pretrained knowledge quickly become obsolete. A common approach to updating LMs is

researcharxiv-cs-cl
29 Jun 2026
Agents

Delayed Verification Destabilizes Multi-Agent LLM Belief: Instability Thresholds and Optimal Corrector Placement

DGX agent

arXiv:2606.27409v1 Announce Type: cross Abstract: Multi-agent large language model (LLM) systems often rely on verifier and critic agents to suppress hallucinations, but verification is delayed. Durin

agentsarxiv-cs-cl
29 Jun 2026
Tutorials

Developmental approach reveals the statistical learning of Neural Language Models: Transformers generalize from the most abstract statistical patterns

DGX agent

arXiv:2606.27460v1 Announce Type: new Abstract: In this study, we use a developmental approach to investigate the statistical learning and mental representation of neural language models (NLM). A seri

tutorialsarxiv-cs-cl
29 Jun 2026
Research

EntMTP: Accelerating LLM Inference with Entropy Guided Multi Token Prediction

DGX agent

arXiv:2606.27550v1 Announce Type: new Abstract: Multi-token prediction has been shown to increase data density during training, improve downstream text-generation quality, and serves as the defacto ap

researcharxiv-cs-cl
29 Jun 2026
Model Releases

Formalizing Latent Thoughts: Four Axioms of Thought Representation in LLMs

DGX agent

arXiv:2606.27378v1 Announce Type: new Abstract: We introduce an axiomatic evaluation framework for latent thought representations in LLMs, comprising metrics that are independent of downstream benchma

model-releasesarxiv-cs-cl
29 Jun 2026
Tutorials

HPRO: Hierarchical Progressive Reward Optimization via Preference Extraction for Emotional Text-to-Speech

DGX agent

arXiv:2606.28249v1 Announce Type: cross Abstract: Recently, Large Language Model (LLM)-based Text-to-Speech (TTS) models have achieved remarkable naturalness. However, the standard Supervised Fine-Tun

tutorialsarxiv-cs-cl
29 Jun 2026
Applications

Joint Transcription and Decryption of Images of Encrypted Handwritten Documents: A Comparison with the Traditional Pipeline

DGX agent

arXiv:2606.27700v1 Announce Type: cross Abstract: Historical encrypted manuscripts present a challenging problem at the intersection of cryptology, linguistics, paleography, and computer vision. Curre

applicationsarxiv-cs-cl
29 Jun 2026
Model Releases

Ko-WideSearch: A Korean Breadth-Search Benchmark for Exhaustive Set Enumeration by Web Agents

DGX agent

arXiv:2606.27595v1 Announce Type: new Abstract: Web-agent benchmarks overwhelmingly measure depth -- pinning one obscure answer behind a chain of constraints -- while breadth, exhaustively enumerating

model-releasesarxiv-cs-cl
29 Jun 2026
Research

Learning Complementary Action Modeling from Automotive Maintenance Instructions

DGX agent

arXiv:2606.27808v1 Announce Type: new Abstract: A minute lexical variation can reverse the procedural meaning of an instruction even when the rest of the sentence remains unchanged. In automotive main

researcharxiv-cs-cl
29 Jun 2026
Model Releases

Learning to Evict from Key-Value Cache

DGX agent

arXiv:2602.10238v2 Announce Type: replace Abstract: The growing size of Large Language Models (LLMs) makes efficient inference challenging, primarily due to the memory demands of the autoregressive Ke

model-releasesarxiv-cs-cl
29 Jun 2026
Research

Masked Language Flow Models

DGX agent

arXiv:2606.27617v1 Announce Type: new Abstract: Masked Diffusion Models (MDMs) promise fast, parallel language generation, but their reverse transition factorises across token positions -- an approxim

researcharxiv-cs-cl
29 Jun 2026
Research

Mechanism-Driven Monitors for Preemptive Detection of LLM Training Instability

DGX agent

arXiv:2606.28116v1 Announce Type: new Abstract: Frontier large language model training consumes massive accelerator fleets and long wall-clock computation, making stability failures costly when they o

researcharxiv-cs-cl
29 Jun 2026
Safety

Mitigating Position Bias in Transformers via Layer-Specific Positional Embedding Scaling

DGX agent

arXiv:2606.27705v1 Announce Type: new Abstract: Large Language Models (LLMs) still struggle with the ``lost-in-the-middle'' problem, where critical information located in the middle of long-context in

safetyarxiv-cs-cl
29 Jun 2026
Model Releases

Multimodal Evaluator Preference Collapse: Cross-Modal Coupling in Self-Evolving Agents

DGX agent

arXiv:2606.16682v3 Announce Type: replace-cross Abstract: When AI agents use language models to evaluate their own outputs in a feedback loop, systematic biases emerge. We show that Evaluator Preferen

model-releasesarxiv-cs-cl
29 Jun 2026
Tutorials

On the Effect of Uncertainty on Layer-wise Inference Dynamics

DGX agent

arXiv:2507.06722v2 Announce Type: replace Abstract: Understanding how large language models (LLMs) internally represent and process their predictions is central to detecting uncertainty and preventing

tutorialsarxiv-cs-cl
29 Jun 2026
Research

Recall Before Rerank: Benchmarking Deep Learning Models for Large-Scale Code-to-Code Retrieval

DGX agent

arXiv:2606.27401v1 Announce Type: cross Abstract: Semantic code search and clone detection are essential for software development, maintenance, and reuse. This paper evaluates the effectiveness, effic

researcharxiv-cs-cl
29 Jun 2026
Model Releases

Retaining by Doing: The Role of On-Policy Data in Mitigating Forgetting

DGX agent

arXiv:2510.18874v3 Announce Type: replace-cross Abstract: Adapting language models (LMs) to new tasks via post-training carries the risk of degrading existing capabilities -- a phenomenon classically

model-releasesarxiv-cs-cl
29 Jun 2026
Applications

Safe Language Generation in the Limit

DGX agent

arXiv:2601.08648v2 Announce Type: replace Abstract: Recent results in learning a language in the limit have shown that, although language identification is impossible, language generation is tractable

applicationsarxiv-cs-cl
29 Jun 2026
Research

Scaling limit of the Random Language Model

DGX agent

arXiv:2606.28105v1 Announce Type: cross Abstract: We develop a quantitative theory of the Random Language Model (RLM), an ensemble of stochastic context-free grammars, in a scaling limit where the num

researcharxiv-cs-cl
29 Jun 2026
Research

Self-Stigma Is Not a Monolith, but Generic Empathy Is: Persona-Conditioned LLM Support for People Who Use Drugs

DGX agent

arXiv:2606.23387v2 Announce Type: replace Abstract: Self-stigma predicts treatment avoidance and disengagement among people who use drugs (PWUD), yet conversational systems aiming to provide support t

researcharxiv-cs-cl
29 Jun 2026
Research

SIGNER: Temporally Grounded Sign Language Generation via Time-Resolved Conditioning

DGX agent

arXiv:2506.07460v2 Announce Type: replace-cross Abstract: Sign language generation (SLG), also known as text-to-sign generation, aims to bridge the communication gap between signers and non-signers. U

researcharxiv-cs-cl
29 Jun 2026
Tutorials

Textual Belief States for World Models: Identifiable Representation Learning Under Strict Mediation

DGX agent

arXiv:2606.27681v1 Announce Type: cross Abstract: World models in partially observed environments rely on latent representations that summarize interaction history, but in many modern LLM-based archit

tutorialsarxiv-cs-cl
29 Jun 2026
Research

The Curse of Multiple Mediators: Hidden Interaction Effects in Activation Patching

DGX agent

arXiv:2606.27510v1 Announce Type: cross Abstract: Activation patching is the primary tool in mechanistic interpretability. It attributes causal responsibility for a model behavior to each of its indiv

researcharxiv-cs-cl
29 Jun 2026
Model Releases

The Signal-Coverage Matrix: Stratifying Type and Semantic Errors in Statement Autoformalization

DGX agent

arXiv:2606.28013v1 Announce Type: new Abstract: Headline type-correctness (TC%) of LLM autoformalization has climbed from sim53% to sim76% in two years, yet this scalar conceals which errors each meth

model-releasesarxiv-cs-cl
29 Jun 2026
Research

ToxiREX: A Dataset on Toxic REasoning in ConteXt

DGX agent

arXiv:2606.27981v1 Announce Type: new Abstract: We introduce a new, contextual, multilingual dataset called ToxiREX: Toxic REasoning in ConteXt. The dataset consists of threads of Reddit comments and

researcharxiv-cs-cl
29 Jun 2026
← Previous
1…3940414243…161
Next →