AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
Research

BitNet Text Embeddings

DGX agent

arXiv:2606.25674v1 Announce Type: new Abstract: LLM-based text embedders have substantially improved retrieval and semantic representation quality, but their deployment remains costly: large backbone

researcharxiv-cs-cl
25 Jun 2026
Research

CLEF HIPE-2026: Evaluating Accurate and Efficient Person-Place Relation Extraction from Multilingual Historical Texts

DGX agent
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

arXiv:2602.17663v2 Announce Type: replace-cross Abstract: HIPE-2026 is a CLEF evaluation lab dedicated to person-place relation extraction from noisy, multilingual historical texts. Building on the HI

researcharxiv-cs-cl
25 Jun 2026
Local Ai

Cliff Tokens: Identifying Single-Token Failure Triggers in LLM Mathematical Reasoning

DGX agent

arXiv:2606.25524v1 Announce Type: cross Abstract: Large language models (LLMs) reach high accuracy in mathematical reasoning, but individual traces on the same problem diverge; some arrive at the corr

local-aiarxiv-cs-cl
25 Jun 2026
Model Releases

CoLA: Cross-Modal Low-rank Adaptation for Multimodal Downstream Tasks

DGX agent

arXiv:2604.03314v2 Announce Type: replace-cross Abstract: Foundation models have revolutionized AI, but adapting them efficiently for multimodal tasks, particularly in dual-stream architectures compos

model-releasesarxiv-cs-cl
25 Jun 2026
Research

ConPress: Learning Efficient Reasoning from Multi-Question Contextual Pressure

DGX agent

arXiv:2602.01472v2 Announce Type: replace Abstract: Large reasoning models (LRMs) typically solve reasoning-intensive tasks by generating long chain-of-thought (CoT) traces, leading to substantial inf

researcharxiv-cs-cl
25 Jun 2026
Safety

Constituency Structure over Eojeol in Korean Treebanks

DGX agent

arXiv:2512.22487v2 Announce Type: replace Abstract: The design of Korean constituency treebanks raises a central representational question concerning the choice of terminal units. Although Korean word

safetyarxiv-cs-cl
25 Jun 2026
Model Releases

Constraint Tax in Open-Weight LLMs: An Empirical Study of Tool Calling Suppression Under Structured Output Constraints

DGX agent

arXiv:2606.25605v1 Announce Type: new Abstract: Tool Calling and Structured Output are two core capabilities of modern Agent systems, yet their interaction under joint deployment conditions remains in

model-releasesarxiv-cs-cl
25 Jun 2026
Applications

Cross-Modal Robustness Transfer (CMRT): Training Robust Speech Translation Models Using Adversarial Text

DGX agent

arXiv:2602.11933v2 Announce Type: replace Abstract: End-to-End Speech Translation (E2E-ST) has seen significant advancements, yet current models are primarily benchmarked on curated, 'clean' datasets.

applicationsarxiv-cs-cl
25 Jun 2026
Research

Data-Driven Evolution of Library and Information Science Research Methods (1990-2022): A Perspective Based on Fine-grained Method Entities

DGX agent

arXiv:2606.25320v1 Announce Type: cross Abstract: Since the 1990s, advancements in big data and information technology have increasingly driven data-centric research in the field of Library and Inform

researcharxiv-cs-cl
25 Jun 2026
Safety

daVinci-kernel: Co-Evolving Skill Selection, Summarization, and Utilization via RL for GPU Kernel Optimization

DGX agent

arXiv:2606.16497v2 Announce Type: replace-cross Abstract: GPU kernel optimization represents a paradigm where functional correctness is assumed and execution efficiency is the objective. We present da

safetyarxiv-cs-cl
25 Jun 2026
Model Releases

Detect, Unlearn, Restore: Defending Text Summarization Models Against Data Poisoning

DGX agent

arXiv:2606.26036v1 Announce Type: new Abstract: Training-time data poisoning during fine-tuning poses a significant threat to large language models (LLMs) deployed for abstractive text summarization,

model-releasesarxiv-cs-cl
25 Jun 2026
Agents

Diagnosing and Mitigating Compounding Failures in Agentic Persuasion via Taxonomic Strategy Retrieval

DGX agent

arXiv:2606.24976v1 Announce Type: cross Abstract: Foundation-model agents in multi-step, open-ended environments frequently suffer from compounding errors, where early mistakes contaminate long-horizo

agentsarxiv-cs-cl
25 Jun 2026
Safety

Digital Twin-Driven Adaptive Sim-to-Real Alignment via Reinforcement Learning for Vibration-Based Bearing Health Monitoring Under Data Scarcity

DGX agent

arXiv:2606.24954v1 Announce Type: cross Abstract: Vibration-based health monitoring of rotating machinery requires reliable fault diagnosis under operational data constraints, yet condition assessment

safetyarxiv-cs-cl
25 Jun 2026
Model Releases

Do Encoders Suffice? A Systematic Comparison of Encoder and Decoder Safety Judges for LLM Adversarial Evaluation

DGX agent

arXiv:2606.25782v1 Announce Type: new Abstract: With the widespread adoption of large language models (LLMs) in chatbots and everyday applications, companies increasingly need guardrails that are effe

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

Do Thinking Tokens Help with Safety?

DGX agent

arXiv:2606.25013v1 Announce Type: cross Abstract: Today's reasoning models use thinking tokens to attain stronger performance on benchmarks than their instruction-tuned counterparts. It is also genera

model-releasesarxiv-cs-cl
25 Jun 2026
Tutorials

Does Translation-Enhanced Speech Encoder Pre-training Affect Speech LLMs?

DGX agent

arXiv:2606.25444v1 Announce Type: cross Abstract: Connecting a pre-trained speech encoder to a Large Language Model (LLM) is the standard architecture for building Speech LLMs. However, a structural m

tutorialsarxiv-cs-cl
25 Jun 2026
Model Releases

Dream at SemEval-2026 Task 13: SALSA for Single-Pass Machine-Generated Code Detection

DGX agent

arXiv:2606.25102v1 Announce Type: new Abstract: Large language models have transformed code generation, raising concerns around authorship, assessment integrity, and software trust. SemEval-2026 Task

model-releasesarxiv-cs-cl
25 Jun 2026
Research

Dustin: Draft-Augmented Sparse Verification for Efficient Long-Context Generation with Speculative Decoding

DGX agent

arXiv:2606.24957v1 Announce Type: new Abstract: While speculative decoding improves inference throughput for multi-batch long-context Large Language Models (LLMs), its efficiency is often limited by a

researcharxiv-cs-cl
25 Jun 2026
Research

Dziri Voicebot: An End-to-End Low-Resource Speech-to-Speech Conversational System for Algerian Dialect

DGX agent

arXiv:2606.26003v1 Announce Type: new Abstract: Automatic speech and language technologies are still heavily biased toward high-resource languages, limiting their applicability to dialectal and low-re

researcharxiv-cs-cl
25 Jun 2026
Local Ai

Efficient and Trainable Language Model Test-Time Scaling via Local Branch Routing

DGX agent

arXiv:2606.25354v1 Announce Type: new Abstract: Test-time scaling improves language-model reasoning, but existing approaches often face a difficult trade-off: long chain-of-thought sampling remains si

local-aiarxiv-cs-cl
25 Jun 2026
Research

Emergent Capabilities Arise Randomly from Learning Sparse Attention Patterns

DGX agent

arXiv:2606.25010v1 Announce Type: cross Abstract: Neural scaling laws for transformer language models predict smooth improvements in pretraining loss with increasing parameters, but downstream capabil

researcharxiv-cs-cl
25 Jun 2026
Research

Error-Aware TF-IDF Retrieval-Augmented Generation for ASR Error Correction

DGX agent

arXiv:2606.24915v1 Announce Type: new Abstract: End-to-end automatic speech recognition systems frequently hallucinate rare entities and domain-specific terms, especially in low-resource languages. Wh

researcharxiv-cs-cl
25 Jun 2026
Research

Evaluating Japanese Dialect Robustness Across Speech and Text-based Large Language Models

DGX agent

arXiv:2606.25436v1 Announce Type: cross Abstract: Dialogue systems based on large language models (LLMs) have advanced significantly in recent years. However, dialectal variation remains a major chall

researcharxiv-cs-cl
25 Jun 2026
Model Releases

Evaluating LLMs on Real-World Software Performance Optimization

DGX agent

arXiv:2606.25530v1 Announce Type: cross Abstract: Software performance optimization is a notoriously complex and manual task. Despite the growing use of Large Language Models (LLMs) for code refinemen

model-releasesarxiv-cs-cl
25 Jun 2026
Research

Fault of Our Stars: Behavioral Drivers of Rating-Sentiment Incongruence

DGX agent

arXiv:2606.25518v1 Announce Type: new Abstract: When people share experiences online, they often express thoughts in two ways: a star rating and a written review. In sentiment analysis, ratings are wi

researcharxiv-cs-cl
25 Jun 2026
Safety

Fully Differentiable Neural Forced Alignment via Soft Dynamic Programming

DGX agent

arXiv:2606.25460v1 Announce Type: cross Abstract: Recent advances in sequence modeling have significantly improved ASR systems, bringing them close to human-level recognition accuracy and enhancing ro

safetyarxiv-cs-cl
25 Jun 2026
Safety

Generalised Medical Phrase Grounding

DGX agent

arXiv:2512.01085v3 Announce Type: replace-cross Abstract: Medical phrase grounding (MPG) maps textual descriptions of radiological findings to corresponding image regions. These grounded reports are e

safetyarxiv-cs-cl
25 Jun 2026
Local Ai

Graph-Based Phonetic Error Correction of Noisy ASR

DGX agent

arXiv:2606.24889v1 Announce Type: new Abstract: Automatic speech recognition (ASR) systems, despite low overall word error rates, produce residual lexical errors that disproportionately affect semanti

local-aiarxiv-cs-cl
25 Jun 2026
Model Releases

Hitting a Moving Target: Test-Time Adaptation for AI Text Detection under Continual Distribution Shift

DGX agent

arXiv:2606.25152v1 Announce Type: new Abstract: Deployed approaches for AI text detection often rely on training-time access to labeled datasets of both human-written and AI-generated text. This appro

model-releasesarxiv-cs-cl
25 Jun 2026
Safety

Homogeneity Bias in Open-Weight LLMs Is Robust to Decoding Hyperparameters

DGX agent

arXiv:2501.02211v2 Announce Type: replace-cross Abstract: Large language models (LLMs) reproduce homogeneity bias -- the tendency to portray marginalized groups as more internally similar than dominan

safetyarxiv-cs-cl
25 Jun 2026
Research

How Large Language Models Source Brand Reputation Across Languages and Markets

DGX agent

arXiv:2606.25787v1 Announce Type: cross Abstract: When a large language model (LLM) answers a question about a company, it grounds the answer in retrieved web sources, and those sources decide what th

researcharxiv-cs-cl
25 Jun 2026
Applications

How Pragmatics Shape Articulation: A Computational Case Study in STEM ASL Discourse

DGX agent

arXiv:2510.23842v2 Announce Type: replace Abstract: Most state-of-the-art sign language models are trained on interpreter or isolated vocabulary data, which overlooks the variability that characterize

applicationsarxiv-cs-cl
25 Jun 2026
Model Releases

How Reliable Is Your Jailbreak Judge? Calibration and Adversarial Robustness of Automated ASR Scoring

DGX agent

arXiv:2606.25487v1 Announce Type: new Abstract: Almost every paper on LLM jailbreaks and prompt injection reports an attack-success rate (ASR), and that number is assigned not by people but by an auto

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

How Robust is OCR-Reasoning? Evaluating OCR-Reasoning Robustness of Vision-Language Models under Visual Perturbations

DGX agent

arXiv:2606.26041v1 Announce Type: cross Abstract: Vision-language models (VLMs) have achieved strong performance on OCR-based benchmarks and increasingly focused on text-rich understanding, but their

model-releasesarxiv-cs-cl
25 Jun 2026
Research

Hybrid-IR: Dual-Path Hybrid Retrieval with Iterative Reasoning for Complex Medical Question Answering

DGX agent

arXiv:2606.25338v1 Announce Type: new Abstract: Large language models (LLMs) have shown promising performance across a wide range of biomedical applications, including medical question answering (QA),

researcharxiv-cs-cl
25 Jun 2026
Research

Improved Large Language Diffusion Models

DGX agent

arXiv:2606.25331v1 Announce Type: new Abstract: Modern large language models are predominantly trained with autoregressive factorization and causal attention. We present iLLaDA, an 8B masked diffusion

researcharxiv-cs-cl
25 Jun 2026
Model Releases

IndicContextEval: A Benchmark for Evaluating Context Utilisation in Audio Large Language Models Across 8 Indic Languages

DGX agent

arXiv:2606.19157v2 Announce Type: replace-cross Abstract: AudioLLMs enable speech recognition conditioned on textual prompts such as domain descriptions or entity lists. However, it remains unclear wh

model-releasesarxiv-cs-cl
25 Jun 2026
Research

Introducing corpora Hlava Cor and Hlava AD: Human Label Variation in Coreference and Discourse Relations

DGX agent

arXiv:2606.25383v1 Announce Type: new Abstract: As previous research on annotator disagreement in discourse phenomena has shown, understanding text coherence varies considerably from one individual to

researcharxiv-cs-cl
25 Jun 2026
Research

Invisible to humans, visible to machines: a preregistered audit of Unicode fidelity across four biomedical bibliographic APIs

DGX agent

arXiv:2606.24897v1 Announce Type: cross Abstract: Biomedical text mining, scientometrics, and the construction of training corpora for biomedical large language models (LLMs) all assume that the abstr

researcharxiv-cs-cl
25 Jun 2026
Agents

Is GraphRAG Needed? From Basic RAG to Graph-/Agentic Solutions with Context Optimization

DGX agent

arXiv:2606.25656v1 Announce Type: new Abstract: As advanced RAG variants like GraphRAG and Agentic RAG emerge, one leading question is when and how to use them. Here, we introduce a framework for diff

agentsarxiv-cs-cl
25 Jun 2026
Hardware

JetSpec: Breaking the Scaling Ceiling of Speculative Decoding with Parallel Tree Drafting

DGX agent

arXiv:2606.18394v2 Announce Type: replace Abstract: Speculative decoding (SD) accelerates autoregressive Large Language Models (LLMs) by drafting multiple tokens and verifying them in parallel, but it

hardwarearxiv-cs-cl
25 Jun 2026
Applications

Learning Diachronic Representations of Ancient Greek Letterforms

DGX agent

arXiv:2606.24984v1 Announce Type: cross Abstract: Learning representations that remain robust across centuries of variation in handwriting is a key challenge in diachronic representation learning. Tak

applicationsarxiv-cs-cl
25 Jun 2026
Tutorials

Learning task-specific subspaces via interventional post-training of speech foundation models

DGX agent

arXiv:2606.17967v2 Announce Type: replace Abstract: Speech foundation models, pre-trained on large corpora of unlabelled speech data, produce general-purpose representations which are useful across ta

tutorialsarxiv-cs-cl
25 Jun 2026
Research

Learning to Erase Private Knowledge from Multi-Documents for Retrieval-Augmented Large Language Models

DGX agent

arXiv:2504.09910v2 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) is a promising technique for applying LLMs to proprietary domains. However, retrieved documents may contain sen

researcharxiv-cs-cl
25 Jun 2026
Research

LLM-ACES: Closed-Loop Discovery of Dynamical Systems with LLM-Guided Adaptive Search

DGX agent

arXiv:2606.25039v1 Announce Type: cross Abstract: Recovering governing Ordinary Differential Equations (ODEs) from data is a central challenge in modeling dynamical systems across scientific domains.

researcharxiv-cs-cl
25 Jun 2026
Safety

LLM-Based Scientific Peer Review: Methods, Benchmarks, and Reliability Challenges

DGX agent

arXiv:2606.25057v1 Announce Type: new Abstract: The rapid growth of scientific submissions has pushed traditional peer review toward its scalability limits, motivating the exploration of large languag

safetyarxiv-cs-cl
25 Jun 2026
Model Releases

LLM Performance on a Real, Double-Marked GCSE Benchmark

DGX agent

arXiv:2606.24973v1 Announce Type: new Abstract: We introduce a dataset of 32,534 double-marked real student responses to GCSE mock exams (GCSEs are the UK's national exams, taken at age ~16), spanning

model-releasesarxiv-cs-cl
25 Jun 2026
Applications

Measuring Research Difficulty of Academic Papers: A Case Study in Natural Language Processing

DGX agent

arXiv:2606.25307v1 Announce Type: cross Abstract: With the rapid growth of the number of academic papers, systematically evaluating the difficulty of research and its relationship to academic impact o

applicationsarxiv-cs-cl
25 Jun 2026
← Previous
1…4243444546…161
Next →