AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
Safety

Hierarchical Policy Optimization for Simultaneous Translation of Unbounded Speech

DGX agent

arXiv:2604.21045v1 Announce Type: new Abstract: Simultaneous speech translation (SST) generates translations while receiving partial speech input. Recent advances show that large language models (LLMs

safetyarxiv-cs-cl
24 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

How Much Is One Recurrence Worth? Iso-Depth Scaling Laws for Looped Language Models

DGX agent

arXiv:2604.21106v1 Announce Type: cross Abstract: We measure how much one extra recurrence is worth to a looped (depth-recurrent) language model, in equivalent unique parameters. From an iso-depth swe

researcharxiv-cs-cl
24 Apr 2026
Model Releases

Hyperloop Transformers

DGX agent

arXiv:2604.21254v1 Announce Type: cross Abstract: LLM architecture research generally aims to maximize model quality subject to fixed compute/latency budgets. However, many applications of interest su

model-releasesarxiv-cs-cl
24 Apr 2026
Agents

Improving Clinical Diagnosis with Counterfactual Multi-Agent Reasoning

DGX agent

arXiv:2603.27820v2 Announce Type: replace Abstract: Clinical diagnosis is a complex reasoning process in which clinicians gather evidence, form hypotheses, and test them against alternative explanatio

agentsarxiv-cs-cl
24 Apr 2026
Model Releases

It's High Time: A Survey of Temporal Question Answering

DGX agent

arXiv:2505.20243v4 Announce Type: replace Abstract: Time plays a critical role in how information is generated, retrieved, and interpreted. In this survey, we provide a comprehensive overview of Tempo

model-releasesarxiv-cs-cl
24 Apr 2026
Research

Job Skill Extraction via LLM-Centric Multi-Module Framework

DGX agent

arXiv:2604.21525v1 Announce Type: new Abstract: Span-level skill extraction from job advertisements underpins candidate-job matching and labor-market analytics, yet generative large language models (L

researcharxiv-cs-cl
24 Apr 2026
Model Releases

Kernel-Smith: A Unified Recipe for Evolutionary Kernel Optimization

DGX agent

arXiv:2603.28342v2 Announce Type: replace Abstract: We present Kernel-Smith, a framework for high-performance GPU kernel and operator generation that combines a stable evaluation-driven evolutionary a

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

Language as a Latent Variable for Reasoning Optimization

DGX agent

arXiv:2604.21593v1 Announce Type: new Abstract: As LLMs reduce English-centric bias, a surprising trend emerges: non-English responses sometimes outperform English on reasoning tasks. We hypothesize t

model-releasesarxiv-cs-cl
24 Apr 2026
Safety

Learning Dynamic Representations and Policies from Multimodal Clinical Time-Series with Informative Missingness

DGX agent

arXiv:2604.21235v1 Announce Type: cross Abstract: Multimodal clinical records contain structured measurements and clinical notes recorded over time, offering rich temporal information about the evolut

safetyarxiv-cs-cl
24 Apr 2026
Research

Learning State-Tracking from Code Using Linear RNNs

DGX agent

arXiv:2602.14814v2 Announce Type: replace-cross Abstract: Over the last years, state-tracking tasks, particularly permutation composition, have become a testbed to understand the limits of sequence mo

researcharxiv-cs-cl
24 Apr 2026
Research

Listen and Chant Before You Read: The Ladder of Beauty in LM Pre-Training

DGX agent

arXiv:2604.21265v1 Announce Type: new Abstract: We show that pre-training a Transformer on music before language significantly accelerates language acquisition. Using piano performances (MAESTRO datas

researcharxiv-cs-cl
24 Apr 2026
Research

Losing our Tail, Again: (Un)Natural Selection & Multilingual LLMs

DGX agent

arXiv:2507.03933v3 Announce Type: replace Abstract: Multilingual Large Language Models considerably changed how technologies influence language. While previous technologies could mediate or assist hum

researcharxiv-cs-cl
24 Apr 2026
Safety

Machine Behavior in Relational Moral Dilemmas: Moral Rightness, Predicted Human Behavior, and Model Decisions

DGX agent

arXiv:2604.21871v1 Announce Type: new Abstract: Human moral judgment is context-dependent and modulated by interpersonal relationships. As large language models (LLMs) increasingly function as decisio

safetyarxiv-cs-cl
24 Apr 2026
Research

Machine learning and digital pragmatics: Which word category influences emoji use most?

DGX agent

arXiv:2604.21108v1 Announce Type: new Abstract: This study investigates Machine Learning (ML) in the prediction of emojis in Arabic tweets employing the (state-of-the-art) MARBERT model. A corpus of 1

researcharxiv-cs-cl
24 Apr 2026
Applications

Mapping the Political Discourse in the Brazilian Chamber of Deputies: A Multi-Faceted Computational Approach

DGX agent

arXiv:2604.21897v1 Announce Type: new Abstract: Analyses of legislative behavior often rely on voting records, overlooking the rich semantic and rhetorical content of political speech. In this paper,

applicationsarxiv-cs-cl
24 Apr 2026
Model Releases

MathDuels: Evaluating LLMs as Problem Posers and Solvers

DGX agent

arXiv:2604.21916v1 Announce Type: new Abstract: As frontier language models attain near-ceiling performance on static mathematical benchmarks, existing evaluations are increasingly unable to different

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

Measuring Opinion Bias and Sycophancy via LLM-based Coercion

DGX agent

arXiv:2604.21564v1 Announce Type: new Abstract: Large language models increasingly shape the information people consume: they are embedded in search, consulted for professional advice, deployed as age

model-releasesarxiv-cs-cl
24 Apr 2026
Research

Misinformation Span Detection in Videos via Audio Transcripts

DGX agent

arXiv:2604.21767v1 Announce Type: new Abstract: Online misinformation is one of the most challenging issues lately, yielding severe consequences, including political polarization, attacks on democracy

researcharxiv-cs-cl
24 Apr 2026
Research

MKJ at SemEval-2026 Task 9: A Comparative Study of Generalist, Specialist, and Ensemble Strategies for Multilingual Polarization

DGX agent

arXiv:2604.21370v1 Announce Type: new Abstract: We present a systematic study of multilingual polarization detection across 22 languages for SemEval-2026 Task 9 (Subtask 1), contrasting multilingual g

researcharxiv-cs-cl
24 Apr 2026
Model Releases

Multilingual and Domain-Agnostic Tip-of-the-Tongue Query Generation for Simulated Evaluation

DGX agent

arXiv:2604.21096v1 Announce Type: cross Abstract: Tip-of-the-Tongue (ToT) retrieval benchmarks have largely focused on English, limiting their applicability to multilingual information access. In this

model-releasesarxiv-cs-cl
24 Apr 2026
Research

Multilinguality at the Edge: Developing Language Models for the Global South

DGX agent

arXiv:2604.21637v1 Announce Type: new Abstract: Where and how language models (LMs) are deployed determines who can benefit from them. However, there are several challenges that prevent effective depl

researcharxiv-cs-cl
24 Apr 2026
Research

Optimal Aggregation of LLM and PRM Signals for Efficient Test-Time Scaling

DGX agent

arXiv:2510.13918v2 Announce Type: replace Abstract: Process reward models (PRMs) are a cornerstone of test-time scaling (TTS), designed to verify and select the best responses from large language mode

researcharxiv-cs-cl
24 Apr 2026
Model Releases

OptiVerse: A Comprehensive Benchmark towards Optimization Problem Solving

DGX agent

arXiv:2604.21510v1 Announce Type: new Abstract: While Large Language Models (LLMs) demonstrate remarkable reasoning, complex optimization tasks remain challenging, requiring domain knowledge and robus

model-releasesarxiv-cs-cl
24 Apr 2026
Safety

Participation and Representation in Local Government Speech

DGX agent

arXiv:2604.21202v1 Announce Type: cross Abstract: Local government meetings are the most common formal channel through which residents speak directly with elected officials, contest policies, and shap

safetyarxiv-cs-cl
24 Apr 2026
Research

Phonological Subspace Collapse Is Aetiology-Specific and Cross-Lingually Stable: Evidence from 3,374 Speakers

DGX agent

arXiv:2604.21706v1 Announce Type: new Abstract: We previously introduced a training-free method for dysarthria severity assessment based on d-prime separability of phonological feature subspaces in fr

researcharxiv-cs-cl
24 Apr 2026
Research

Preferences of a Voice-First Nation: Large-Scale Pairwise Evaluation and Preference Analysis for TTS in Indian Languages

DGX agent

arXiv:2604.21481v1 Announce Type: new Abstract: Crowdsourced pairwise evaluation has emerged as a scalable approach for assessing foundation models. However, applying it to Text to Speech(TTS) introdu

researcharxiv-cs-cl
24 Apr 2026
Research

Prefix Parsing is Just Parsing

DGX agent

arXiv:2604.21191v1 Announce Type: new Abstract: Prefix parsing asks whether an input prefix can be extended to a complete string generated by a given grammar. In the weighted setting, it also provides

researcharxiv-cs-cl
24 Apr 2026
Model Releases

ReFACT: A Benchmark for Scientific Confabulation Detection with Positional Error Annotations

DGX agent

arXiv:2509.25868v3 Announce Type: replace Abstract: The mechanisms underlying scientific confabulation in Large Language Models (LLMs) remain poorly understood. We introduce ReFACT (Reddit False And C

model-releasesarxiv-cs-cl
24 Apr 2026
Research

Revisiting Non-Verbatim Memorization in Large Language Models: The Role of Entity Surface Forms

DGX agent

arXiv:2604.21882v1 Announce Type: new Abstract: Understanding what kinds of factual knowledge large language models (LLMs) memorize is essential for evaluating their reliability and limitations. Entit

researcharxiv-cs-cl
24 Apr 2026
Model Releases

RewardBench 2: Advancing Reward Model Evaluation

DGX agent

arXiv:2506.01937v2 Announce Type: replace Abstract: Reward models are used throughout the post-training of language models to capture nuanced signals from preference data and provide a training target

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

Seeing Isn't Believing: Uncovering Blind Spots in Evaluator Vision-Language Models

DGX agent

arXiv:2604.21523v1 Announce Type: cross Abstract: Large Vision-Language Models (VLMs) are increasingly used to evaluate outputs of other models, for image-to-text (I2T) tasks such as visual question a

model-releasesarxiv-cs-cl
24 Apr 2026
Research

SemEval-2026 Task 4: Narrative Story Similarity and Narrative Representation Learning

DGX agent

arXiv:2604.21782v1 Announce Type: new Abstract: We present the shared task on narrative similarity and narrative representation learning - NSNRL (pronounced 'nass-na-rel'). The task operationalizes na

researcharxiv-cs-cl
24 Apr 2026
Model Releases

SlideAgent: Hierarchical Agentic Framework for Multi-Page Visual Document Understanding

DGX agent

arXiv:2510.26615v3 Announce Type: replace Abstract: Retrieval-augmented generation (RAG) extends large language models (LLMs) with external knowledge, but it must balance limited effective context, re

model-releasesarxiv-cs-cl
24 Apr 2026
Research

Slot Machines: How LLMs Keep Track of Multiple Entities

DGX agent

arXiv:2604.21139v1 Announce Type: new Abstract: Language models must bind entities to the attributes they possess and maintain several such binding relationships within a context. We study how multipl

researcharxiv-cs-cl
24 Apr 2026
Model Releases

SocraticKG: Knowledge Graph Construction via QA-Driven Fact Extraction

DGX agent

arXiv:2601.10003v2 Announce Type: replace Abstract: Constructing Knowledge Graphs (KGs) from unstructured text provides a structured framework for knowledge representation and reasoning, yet current L

model-releasesarxiv-cs-cl
24 Apr 2026
Research

StegoStylo: Squelching Stylometric Scrutiny through Steganographic Stitching

DGX agent

arXiv:2601.09056v3 Announce Type: replace-cross Abstract: Stylometry--the identification of an author through analysis of a text's style (i.e., authorship attribution)--serves many constructive purpos

researcharxiv-cs-cl
24 Apr 2026
Research

Sub-Token Routing in LoRA for Adaptation and Query-Aware KV Compression

DGX agent

arXiv:2604.21335v1 Announce Type: cross Abstract: Sub-token routing offers a finer control axis for transformer efficiency than the coarse units used in most prior work, such as tokens, pages, heads,

researcharxiv-cs-cl
24 Apr 2026
Model Releases

Subject-level Inference for Realistic Text Anonymization Evaluation

DGX agent

arXiv:2604.21211v1 Announce Type: new Abstract: Current text anonymization evaluation relies on span-based metrics that fail to capture what an adversary could actually infer, and assumes a single dat

model-releasesarxiv-cs-cl
24 Apr 2026
Local Ai

TabSHAP

DGX agent

arXiv:2604.21120v1 Announce Type: cross Abstract: Large Language Models (LLMs) fine-tuned on serialized tabular data are emerging as powerful alternatives to traditional tree-based models, particularl

local-aiarxiv-cs-cl
24 Apr 2026
Model Releases

The Root Theorem of Context Engineering

DGX agent

arXiv:2604.20874v1 Announce Type: cross Abstract: Every system that maintains a large language model conversation beyond a single session faces two inescapable constraints: the context window is finit

model-releasesarxiv-cs-cl
24 Apr 2026
Safety

'This Wasn't Made for Me': Recentering User Experience and Emotional Impact in the Evaluation of ASR Bias

DGX agent

arXiv:2604.21148v1 Announce Type: new Abstract: Studies on bias in Automatic Speech Recognition (ASR) tend to focus on reporting error rates for speakers of underrepresented dialects, yet less researc

safetyarxiv-cs-cl
24 Apr 2026
Research

TRACES: Tagging Reasoning Steps for Adaptive Cost-Efficient Early-Stopping

DGX agent

arXiv:2604.21057v1 Announce Type: new Abstract: The field of Language Reasoning Models (LRMs) has been very active over the past few years with advances in training and inference techniques enabling L

researcharxiv-cs-cl
24 Apr 2026
Research

Tuning for TraceTarnish: Techniques, Trends, and Testing Tangible Traits

DGX agent

arXiv:2512.03465v3 Announce Type: replace-cross Abstract: In this study, we more rigorously evaluated our attack script extit{TraceTarnish}, which leverages adversarial stylometry principles to anonym

researcharxiv-cs-cl
24 Apr 2026
Research

UKP_Psycontrol at SemEval-2026 Task 2: Modeling Valence and Arousal Dynamics from Text

DGX agent

arXiv:2604.21534v1 Announce Type: new Abstract: This paper presents our system developed for SemEval-2026 Task 2. The task requires modeling both current affect and short-term affective change in chro

researcharxiv-cs-cl
24 Apr 2026
Research

Unlocking the Power of Large Language Models for Multi-table Entity Matching

DGX agent

arXiv:2604.21238v1 Announce Type: new Abstract: Multi-table entity matching (MEM) addresses the limitations of dual-table approaches by enabling simultaneous identification of equivalent entities acro

researcharxiv-cs-cl
24 Apr 2026
Agents

Unveiling Unicode's Unseen Underpinnings in Undermining Authorship Attribution

DGX agent

arXiv:2508.15840v5 Announce Type: replace-cross Abstract: When using a public communication channel--whether formal or informal, such as commenting or posting on social media--end users have no expect

agentsarxiv-cs-cl
24 Apr 2026
Research

Weighting What Matters: Boosting Sample Efficiency in Medical Report Generation via Token Reweighting

DGX agent

arXiv:2604.21082v1 Announce Type: new Abstract: Training vision-language models (VLMs) for medical report generation is often hindered by the scarcity of high-quality annotated data. This work evaluat

researcharxiv-cs-cl
24 Apr 2026
Model Releases

When Agents Look the Same: Quantifying Distillation-Induced Similarity in Tool-Use Behaviors

DGX agent

arXiv:2604.21255v1 Announce Type: new Abstract: Model distillation is a primary driver behind the rapid progress of LLM agents, yet it often leads to behavioral homogenization. Many emerging agents sh

model-releasesarxiv-cs-cl
24 Apr 2026
← Previous
1…125126127128129…161
Next →