AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,759 results
Research

The Latin Substrate: How Language Models Represent and Mediate Script Choice

DGX agent

arXiv:2605.31363v1 Announce Type: new Abstract: Many languages are written in multiple scripts, requiring large language models (LLMs) to generate equivalent linguistic content in distinct orthographi

researcharxiv-cs-cl
1 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Tutorials

The relative strength of hierarchical structure and statistics differs across the measures in naturalistic reading

DGX agent

arXiv:2509.23195v2 Announce Type: replace Abstract: The hierarchical syntactic structure and non-hierarchical, statistical, or sequential factors have long been framed as rival theories in accounting

tutorialsarxiv-cs-cl
1 Jun 2026
Applications

Towards Effective Long-Video Event Prediction via Multi-Level Event Semantics Mining

DGX agent

arXiv:2605.31069v1 Announce Type: cross Abstract: Accurately predicting future events is fundamental to content understanding and decision-making across various domains. While prior research has prima

applicationsarxiv-cs-cl
1 Jun 2026
Research

Towards Efficient LLMs Annealing with Principled Sample Selection

DGX agent

arXiv:2605.31175v1 Announce Type: new Abstract: The annealing phase is a pivotal convergence stage in LLM pre-training that ultimately determines final model quality. However, effectively selecting tr

researcharxiv-cs-cl
1 Jun 2026
Model Releases

TRACE: Discovering Task-Specific Parameter via Adaptation-Aware Probing for Continual Fine-Tuning

DGX agent

arXiv:2605.31025v1 Announce Type: new Abstract: In real-world deployment, LLMs are often adapted continually across tasks to keep LLMs up-to-date in production, where new fine-tuning should preserve p

model-releasesarxiv-cs-cl
1 Jun 2026
Safety

Traceable by Design: An LLM Pipeline and Dashboard for EU Regulatory Consultation Analysis

DGX agent

arXiv:2605.30995v1 Announce Type: cross Abstract: Public consultations generate large volumes of data in the form of stakeholder submissions that are practically unfeasible to analyse manually. We pre

safetyarxiv-cs-cl
1 Jun 2026
Tutorials

Trading Complexity for Expressivity Through Structured Generalized Linear Token Mixing

DGX agent

arXiv:2605.31367v1 Announce Type: cross Abstract: Token mixing layers play a key role in how language models can learn and generate long-range dependencies. Their efficiency relies on the necessary tr

tutorialsarxiv-cs-cl
1 Jun 2026
Model Releases

Translation Analytics for Freelancers II: Benchmarking Local LLMs for Confidential Translation Workflows

DGX agent

arXiv:2605.31452v1 Announce Type: new Abstract: Building on our previous work, this paper develops practical, low-barrier methods for freelance translators and smaller language service providers to ev

model-releasesarxiv-cs-cl
1 Jun 2026
Research

TransLPRNet: Lite Vision-Language Network for Single/Dual-line Chinese License Plate Recognition

DGX agent

arXiv:2507.17335v2 Announce Type: replace-cross Abstract: License plate recognition in open environments is widely applicable across various domains; however, the diversity of license plate types and

researcharxiv-cs-cl
1 Jun 2026
Model Releases

Triaging Threats to Specialized Guardrails

DGX agent

arXiv:2605.30693v1 Announce Type: cross Abstract: Building robust safety guardrails is essential for deploying Large Language Models across diverse real-world applications. However, this goal remains

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

TSM-Bench: Detecting LLM-Generated Text in Real-World Wikipedia Editing Practices

DGX agent

arXiv:2605.31113v1 Announce Type: new Abstract: Automatically detecting machine-generated text (MGT) is critical to maintaining the knowledge integrity of user-generated content (UGC) platforms such a

model-releasesarxiv-cs-cl
1 Jun 2026
Safety

UniAudio-Token: Empowering Semantic Speech Tokenizers with General Audio Perception

DGX agent

arXiv:2605.31521v1 Announce Type: new Abstract: Semantic speech tokenizers have become a widely used interface for Audio-LLMs, owing to their compact single-codebook design and strong linguistic align

safetyarxiv-cs-cl
1 Jun 2026
Model Releases

UniDial-EvalKit: A Unified Toolkit for Evaluating Multi-Faceted Conversational Abilities

DGX agent

arXiv:2603.23160v2 Announce Type: replace Abstract: Benchmarking large language models (LLMs) and agents in multi-turn interactive scenarios is essential for understanding their practical capabilities

model-releasesarxiv-cs-cl
1 Jun 2026
Research

Unlocking Fine-Grained Translation Quality Estimation in LRMs through Synergistically Evolving Implicit and Explicit Reasoning

DGX agent

arXiv:2605.31378v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) still struggle with fine-grained translation quality estimation (QE), even with long reasoning chains. We argue that LRMs

researcharxiv-cs-cl
1 Jun 2026
Research

Weights to Code: Extracting Interpretable Algorithms from the Discrete Transformer

DGX agent

arXiv:2601.05770v3 Announce Type: replace-cross Abstract: Algorithm extraction aims to synthesize executable programs directly from models trained on algorithmic tasks, enabling de novo recovery of ex

researcharxiv-cs-cl
1 Jun 2026
Safety

What Am I Missing? Question-Answering as Hidden State Probing

DGX agent

arXiv:2605.31561v1 Announce Type: new Abstract: Test-time reasoning has become a significant field of study since the introduction of chain-of-thought reasoning in large language models (LLMs). Howeve

safetyarxiv-cs-cl
1 Jun 2026
Local Ai

When English Rewrites Local Knowledge: Global Narrative Dominance in Large Language Models

DGX agent

arXiv:2605.30481v1 Announce Type: new Abstract: Large language models (LLMs) are widely used as cross-lingual knowledge interfaces. However, culturally grounded questions often reflect globally domina

local-aiarxiv-cs-cl
1 Jun 2026
Research

Wind Turbine Maintenance Log Labelling Framework: LLM-Driven Data Correction and Enrichment via Semantic Extraction of Reliability Intelligence

DGX agent

arXiv:2605.31281v1 Announce Type: new Abstract: As wind turbine fleets age, data-driven reliability engineering is essential to optimise their operation and maintenance for service life extension and

researcharxiv-cs-cl
1 Jun 2026
Model Releases

Your Multimodal Speech Model Says I Have a Face for Radio

DGX agent

arXiv:2605.30472v1 Announce Type: new Abstract: As large neural models have become better at language tasks, researchers are increasingly building multi- and omnimodal models that handle more modaliti

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

A Dual-Path Architecture for Scaling Compute and Capacity in LLMs

DGX agent

arXiv:2605.30202v1 Announce Type: new Abstract: Looped transformers apply a shared block multiple times and have emerged as a parameter-efficient route to scaling compute in language models. However,

model-releasesarxiv-cs-cl
29 May 2026
Safety

A Modular Architecture for Typologically Controlled Lexicon Generation

DGX agent

arXiv:2605.28824v1 Announce Type: new Abstract: Constructing artificial lexicons that are pronounceable, typologically plausible, and semantically structured remains an open challenge in computational

safetyarxiv-cs-cl
29 May 2026
Safety

A Study on Question-Answer Dataset for LLM Safety Evaluation with a Focus on Illegal Activities

DGX agent

arXiv:2605.29340v1 Announce Type: new Abstract: In this paper, we discuss question-answer dataset for LLM safety evaluation, with a focus on illegal activities. Specifically, on the basis of manual an

safetyarxiv-cs-cl
29 May 2026
Applications

Accommodation Goes Both Ways: Studying Linguistic Convergence Between Humans and Language Models

DGX agent

arXiv:2605.29278v1 Announce Type: new Abstract: As LLMs become increasingly integrated into daily life, understanding how their presence will shape human linguistic behavior is an open question. We pr

applicationsarxiv-cs-cl
29 May 2026
Safety

ActTraitBench: Quantifying the Knowledge-Decision Gap in Large Language Models via Human-Grounded Behavioral Validation

DGX agent

arXiv:2605.29791v1 Announce Type: new Abstract: While Large Language Models (LLMs) can convincingly simulate personas in explicit self-reports, they often deviate in implicit behavioral decisions, rev

safetyarxiv-cs-cl
29 May 2026
Model Releases

Adapting Multilingual Embedding Models to Turkish via Cross-Lingual Tokenizer Surgery and Offline Distillation

DGX agent

arXiv:2605.29992v1 Announce Type: new Abstract: Sentence embeddings are a foundational component for semantic search, clustering, classification, and retrieval-augmented generation. This paper present

model-releasesarxiv-cs-cl
29 May 2026
Research

Adaptive Targeted Dynamic Chunking for Tokenization-Free Hierarchical Model

DGX agent

arXiv:2605.30080v1 Announce Type: new Abstract: Tokenization-free hierarchical models are emerging as a promising alternative to traditional Large Language Models (LLMs), addressing inherent preproces

researcharxiv-cs-cl
29 May 2026
Model Releases

AfriScience-MT: Towards Decolonizing Science in Africa through Text Translation

DGX agent

arXiv:2605.29741v1 Announce Type: new Abstract: The dominance of colonial languages in African education and scientific communication limits how hundreds of millions of speakers of African languages a

model-releasesarxiv-cs-cl
29 May 2026
Research

Analyzing Persona Effects in Generated Explanations from Multimodal LLM Agents in Urban Perception

DGX agent

arXiv:2605.29064v1 Announce Type: new Abstract: We study how persona prompting shapes language generated by multimodal large language models in an urban perception setting. Using 59,808 annotations fr

researcharxiv-cs-cl
29 May 2026
Research

Attention Asymmetry in AI Layoff Discourse on X: A Computational Analysis of Capital vs Labour Amplification

DGX agent

arXiv:2605.29367v1 Announce Type: new Abstract: When workers lose jobs to AI-driven restructuring, two very different conversations happen on X (formerly Twitter) at the same time. Tech executives and

researcharxiv-cs-cl
29 May 2026
Model Releases

'Be My Cheese?': Cultural Nuance Benchmarking for Machine Translation in Multilingual LLMs

DGX agent

arXiv:2602.04729v2 Announce Type: replace Abstract: We present a large-scale human evaluation benchmark for assessing cultural localisation in machine translation produced by state-of-the-art multilin

model-releasesarxiv-cs-cl
29 May 2026
Model Releases

Beyond English and Evasion: A Human-Annotated Multi-Domain Benchmark for High-Stakes LLM Safety Evaluation in Chinese

DGX agent

arXiv:2605.29667v1 Announce Type: new Abstract: When Large Language Models (LLMs) are deployed in Chinese-language settings, a troubling pattern emerges: safety systems that work well in English break

model-releasesarxiv-cs-cl
29 May 2026
Research

Beyond Transcripts: A Renewed Perspective on Audio Chaptering

DGX agent

arXiv:2602.08979v2 Announce Type: replace-cross Abstract: Audio chaptering, the task of segmenting long-form audio into coherent sections, is increasingly important for navigating podcasts, lectures,

researcharxiv-cs-cl
29 May 2026
Agents

Bosses, Kings, and the Commons: Cooperation Under Power Asymmetry in LLM Societies

DGX agent

arXiv:2605.29062v1 Announce Type: new Abstract: Communities can sustainably manage shared resources (commons) through self-governance and cooperative norms, a central finding of Ostrom's theory of sel

agentsarxiv-cs-cl
29 May 2026
Model Releases

BrahmicTokenizer-131K: An Indic-Capable Drop-In Replacement for o200k_base

DGX agent

arXiv:2605.29379v1 Announce Type: new Abstract: We present BrahmicTokenizer-131K, a 131,072-vocabulary byte-level BPE tokenizer that closes the Brahmic compression gap at the 131K-vocabulary class whi

model-releasesarxiv-cs-cl
29 May 2026
Safety

Calibration Is Not Enough: Evaluating Confidence Estimation Under Language Variations

DGX agent

arXiv:2601.08064v2 Announce Type: replace Abstract: Confidence estimation (CE) indicates how reliable the answers of large language models are and impacts user trust and decision-making. Existing eval

safetyarxiv-cs-cl
29 May 2026
Model Releases

Casual as an Anchor: Resolving Supervision Misalignment in Formality Transfer Dataset

DGX agent

arXiv:2605.29365v1 Announce Type: new Abstract: Formality transfer is commonly framed as a symmetric bidirectional task between informal and formal registers. We argue that this framing conceals a sup

model-releasesarxiv-cs-cl
29 May 2026
Agents

Catalyst-Agent: Autonomous heterogeneous catalyst screening with an LLM Agent

DGX agent

arXiv:2603.01311v2 Announce Type: replace Abstract: The discovery of novel catalysts tailored for particular applications is a major challenge for the twenty-first century. Traditional methods for thi

agentsarxiv-cs-cl
29 May 2026
Safety

Causal Interventions on Continuous Variables: A Case Study on Verb Bias in Steering Vectors for In-Context Learning

DGX agent

arXiv:2605.29971v1 Announce Type: new Abstract: Causal interventions in language model representations have largely targeted discrete features, like grammatical number. However, language models must a

safetyarxiv-cs-cl
29 May 2026
Research

CCS: Clinical Consensus Selection for Radiology Report Generation

DGX agent

arXiv:2605.30131v1 Announce Type: new Abstract: Radiology report generation (RRG) is commonly formulated as a single-path generation task, where a multimodal large language model (MLLM) produces one d

researcharxiv-cs-cl
29 May 2026
Local Ai

Classification of non-analyzable word types in web documents to implement an effective Korean e-learning system

DGX agent

arXiv:2605.29638v1 Announce Type: new Abstract: E-learning systems should deliver contents that reflect various phenomena of the language as it is used. In addition to formal Korean, e-learning system

local-aiarxiv-cs-cl
29 May 2026
Research

Cognitive Loop of Thought: Reversible Hierarchical Markov Chain for Efficient Mathematical Reasoning

DGX agent

arXiv:2604.06805v2 Announce Type: replace Abstract: Multi-step Chain-of-Thought (CoT) has significantly advanced the mathematical reasoning capabilities of LLMs by leveraging explicit reasoning steps.

researcharxiv-cs-cl
29 May 2026
Model Releases

CommunityFact: A Dynamic, Multilingual, Multi-domain Benchmark for Misinformation Detection in the Wild

DGX agent

arXiv:2605.30241v1 Announce Type: new Abstract: Misinformation verification increasingly occurs in public, fast-moving, and multilingual online settings, where static benchmarks provide an incomplete

model-releasesarxiv-cs-cl
29 May 2026
Model Releases

Comparative Evaluation of Machine Translation Systems on Images with Text

DGX agent

arXiv:2605.29476v1 Announce Type: new Abstract: This work presents a comparative evaluation of machine translation systems applied to images containing textual information, a task that lies at the int

model-releasesarxiv-cs-cl
29 May 2026
Model Releases

COMPOSE: Composing Future Theorems from Citations and Formal Structure

DGX agent

arXiv:2605.30333v1 Announce Type: new Abstract: A plausible future mathematical claim must satisfy two constraints: it should follow the direction of prior work and respect the formal dependencies tha

model-releasesarxiv-cs-cl
29 May 2026
Agents

CONCAT: Consensus- and Confidence-Driven Ad Hoc Teaming for Efficient LLM-Based Multi-Agent Systems

DGX agent

arXiv:2605.29612v1 Announce Type: cross Abstract: Although large language model (LLM) based multi-agent systems (MAS) show their capability to solve complex tasks and achieve higher performance over s

agentsarxiv-cs-cl
29 May 2026
Model Releases

Converted, Not Equivalent: Benchmarking Codebase Conversion via Observational Equivalence

DGX agent

arXiv:2605.29054v1 Announce Type: cross Abstract: Coding agents increasingly act as codebase-scale collaborators that can assist with codebase conversion, but this progress has exposed a critical weak

model-releasesarxiv-cs-cl
29 May 2026
Research

CorPipe at CRAC 2026: Empty Nodes and Cross-Lingual Transfer in Multilingual Coreference Resolution

DGX agent

arXiv:2605.30133v1 Announce Type: new Abstract: We introduce CorPipe 26, our winning submission to the CRAC 2026 Shared Task on Multilingual Coreference Resolution. The fifth edition of this shared ta

researcharxiv-cs-cl
29 May 2026
Model Releases

CriticalKV: Optimizing KV Cache Eviction from an Output Perturbation Perspective

DGX agent

arXiv:2502.03805v2 Announce Type: replace Abstract: Large language models have revolutionized natural language processing but face significant challenges of high storage and runtime costs, due to the

model-releasesarxiv-cs-cl
29 May 2026
← Previous
1…6768697071…162
Next →