AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
6 May 2026

When Prompts Interact: Assessing Prompt Arithmetic for Deconfounding under Distribution Shift

Model ReleasesDGX agent

arXiv:2605.03096v1 Announce Type: cross Abstract: In classification tasks, models may rely on confounding variables to achieve strong in-distribution performance, capturing spurious features that fail

When Should a Language Model Trust Itself? Same-Model Self-Verification as a Conditional Confidence Signal

Model ReleasesDGX agent

arXiv:2605.02915v1 Announce Type: new Abstract: Same-model self-verification, prompting a model to audit its own predicted answer, is a plausible confidence signal for selective prediction, but its pr

When to Think, When to Speak: Learning Disclosure Policies for LLM Reasoning

SafetyDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.03314v1 Announce Type: new Abstract: In single-stream autoregressive interfaces, the same tokens both update the model state and constitute an irreversible public commitment. This coupling

Workspace-Bench 1.0: Benchmarking AI Agents on Workspace Tasks with Large-Scale File Dependencies

Model ReleasesDGX agent

arXiv:2605.03596v1 Announce Type: cross Abstract: Workspace learning requires AI agents to identify, reason over, exploit, and update explicit and implicit dependencies among heterogeneous files in a

5 May 2026

A framework for analyzing concept representations in neural models

ResearchDGX agent

arXiv:2605.01381v1 Announce Type: new Abstract: Understanding how neural models represent human-interpretable concepts is challenging. Prior work has explored linear concept subspaces from diverse per

A Language for Describing Agentic LLM Contexts

AgentsDGX agent

arXiv:2605.01920v1 Announce Type: cross Abstract: Large language models are increasingly used within larger systems ('LLM agents'). These make a sequence of LLM calls, each call providing the LLM with

A Multi-View Media Profiling Suite: Resources, Evaluation, and Analysis

SafetyDGX agent

arXiv:2605.01336v1 Announce Type: new Abstract: News outlets shape public opinion at a scale that makes automated detection of political bias and factuality essential. However, the field still lacks u

A multilingual hallucination benchmark: MultiWikiQHalluA

Model ReleasesDGX agent

arXiv:2605.02504v1 Announce Type: new Abstract: Most hallucination evaluations focus on English, leaving it unclear whether findings transfer to lower-resource languages. We investigate faithfulness h

A Multimodal Dataset for Visually Grounded Ambiguity in Machine Translation

ResearchDGX agent

arXiv:2605.02035v1 Announce Type: new Abstract: Ambiguity resolution is a key challenge in multimodal machine translation (MMT), where models must genuinely leverage visual input to map an ambiguous e

A Systematic Benchmark of Machine Transliteration Models for the Tajik-Farsi Language Pair: A Comparative Study from Rule-Based to Transformer Architectures

Model ReleasesDGX agent

arXiv:2605.02270v1 Announce Type: new Abstract: This paper presents the first comprehensive comparative analysis of modern machine learning architectures for transliteration between Tajik (Cyrillic sc

A Systematic Exploration of Text Decomposition and Budget Distribution in Differentially Private Text Obfuscation

ResearchDGX agent

arXiv:2605.01065v1 Announce Type: new Abstract: The goal of differentially private text obfuscation is to obfuscate, or 'perturb', input texts with Differential Privacy (DP) guarantees, such that the

A Theoretical Game of Attacks via Compositional Skills

SafetyDGX agent

arXiv:2605.01034v1 Announce Type: new Abstract: As large language models grow increasingly capable, concerns about their safe deployment have intensified. While numerous alignment strategies aim to re

Accurate Legal Reasoning at Scale: Neuro-Symbolic Offloading and Structural Auditability for Robust Legal Adjudication

Model ReleasesDGX agent

arXiv:2605.02472v1 Announce Type: new Abstract: Legal texts often contain computational legal clauses--provisions whose understanding requires complex logic. While frontier Large Reasoning Models (LRM

Adaptive GoGI-Skip: Coupling Goal-Gradient Importance with Dynamic Uncertainty for Efficient Reasoning

SafetyDGX agent

arXiv:2505.08392v3 Announce Type: replace Abstract: Chain-of-Thought (CoT) prompting trades inference speed for reasoning accuracy. Existing compressors force a compromise as static gradient technique

Addressing Data Scarcity in Bangla Fake News Detection: An LLM-Based Dataset Augmentation Approach

Model ReleasesDGX agent

arXiv:2605.01292v1 Announce Type: new Abstract: The growing spread of misinformation in digital media highlights the need for reliable fake news detection systems, yet progress in under-resourced lang

AgentXRay: White-Boxing Agentic Systems via Workflow Reconstruction

AgentsDGX agent

arXiv:2602.05353v3 Announce Type: replace-cross Abstract: Large Language Models have shown strong capabilities in complex problem solving, yet many agentic systems remain difficult to interpret and co

ALIGNS: Unlocking nomological networks in psychological measurement through a large language model

SafetyDGX agent

arXiv:2509.09723v3 Announce Type: replace Abstract: Psychological measurement is critical to many disciplines. Despite advances in measurement, building nomological networks, theoretical maps of how c

An Information-theoretic Propagation Denoising and Fusion Framework for Fake News Detection

TutorialsDGX agent

arXiv:2605.02259v1 Announce Type: new Abstract: Incomplete propagation data significantly hinders robust fake news detection. Recent approaches leverage large language models to simulate missing user

Argumentation for Explainable and Globally Contestable Decision Support with LLMs

ResearchDGX agent

arXiv:2603.14643v2 Announce Type: replace-cross Abstract: Large language models (LLMs) exhibit strong general capabilities, but their deployment in high-stakes domains is hindered by their opacity and

ARGUS: Policy-Adaptive Ad Governance via Evolving Reinforcement with Adversarial Umpiring

SafetyDGX agent

arXiv:2605.02200v1 Announce Type: new Abstract: Online advertising governance faces significant challenges due to the non-stationary nature of regulatory policies, where emerging mandates (e.g., restr

Arithmetic in the Wild: Llama uses Base-10 Addition to Reason About Cyclic Concepts

Model ReleasesDGX agent

arXiv:2605.01148v1 Announce Type: cross Abstract: Does structure in representations imply structure in computation? We study how Llama-3.1-8B reasons over cyclic concepts (e.g., 'what month is six mon

Artificial intelligence language technologies in multilingual healthcare: Grand challenges ahead

SafetyDGX agent

arXiv:2605.01441v1 Announce Type: new Abstract: AI language technologies (AILTs), increasingly enabled by large language models (LLMs), are becoming embedded in multilingual healthcare workflows for t

ATLAS: Article Tracking, Linking, and Analysis of Swedish Encyclopedias

ResearchDGX agent

arXiv:2605.02466v1 Announce Type: new Abstract: The digitization of old encyclopedias represents an important step to improve access to historically structured knowledge. Often, however, this process

ATR-Bench: A Federated Learning Benchmark for Adaptation, Trust, and Reasoning

Model ReleasesDGX agent

arXiv:2505.16850v2 Announce Type: replace-cross Abstract: Federated Learning (FL) has emerged as a promising paradigm for collaborative model training while preserving data privacy across decentralize

Attention Sink Forges Native MoE in Attention Layers: Sink-Aware Training to Address Head Collapse

ResearchDGX agent

arXiv:2602.01203v2 Announce Type: replace Abstract: Large Language Models (LLMs) often assign disproportionate attention to the first token, a phenomenon known as the attention sink. Several recent ap

Attention Sinks in Massively Multilingual Neural Machine Translation:Discovery, Analysis, and Mitigation

SafetyDGX agent

arXiv:2605.01229v1 Announce Type: cross Abstract: Cross-attention patterns in neural machine translation (NMT) are widely used to study how multilingual models align linguistic structure. We report a

Attribution-Guided Pruning for Insight and Control: Circuit Discovery and Targeted Correction in Small-scale LLMs

ApplicationsDGX agent

arXiv:2506.13727v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are widely deployed in real-world applications, yet their internal mechanisms remain difficult to interpret and c

Auditing demographic bias in AI-based emergency police dispatch: a cross-lingual evaluation of eleven large language models

SafetyDGX agent

arXiv:2605.01451v1 Announce Type: new Abstract: Large language models (LLMs) are rapidly being integrated into high-stakes public safety systems, including emergency call triage and dispatch decision

Automated Interpretability and Feature Discovery in Language Models with Agents

Model ReleasesDGX agent

arXiv:2605.01555v1 Announce Type: new Abstract: We introduce an autonomous multiagent framework for mechanistic interpretability that automates both explaining and finding internal features in large l

Automatic Correction of Writing Anomalies in Hausa Texts

ResearchDGX agent

arXiv:2506.03820v2 Announce Type: replace Abstract: Hausa texts are often characterized by writing anomalies, such as incorrect character substitutions and spacing errors, which sometimes hinder natur

Automatic Reflection Level Classification in Hungarian Student Essays

ApplicationsDGX agent

arXiv:2605.02402v1 Announce Type: new Abstract: Reflective thinking is a key competency in education, but assessing reflective writing remains a time-consuming and subjective task for education expert

Balalaika: Data-Centric, Prosody-Aware Annotation Pipeline for Russian Speech

ResearchDGX agent

arXiv:2507.13563v2 Announce Type: replace Abstract: We introduce Balalaika, an open-source, data-centric pipeline for processing audio and producing prosody-aware annotations. It combines semantic VAD

BaldWhisper: Faster Whisper with Head Shearing and Layer Merging

ResearchDGX agent

arXiv:2510.08599v2 Announce Type: replace-cross Abstract: Pruning large pre-trained transformers in a data-scarce scenario is challenging, as it often requires massive retraining data to recover perfo

Beating the Style Detector: Three Hours of Agentic Research on the AI-Text Arms Race

Model ReleasesDGX agent

arXiv:2605.02620v1 Announce Type: new Abstract: Reproducing an empirical NLP study used to take weeks. Given the released data and a modern agentic-research harness, we redo every experiment of a rece

Benchmarking LightGBM and BiLSTM for Sentiment Analysis on Indonesian E-Commerce Reviews

ResearchDGX agent

arXiv:2605.01322v1 Announce Type: new Abstract: This study presents a comparative analysis between two primary approaches in Natural Language Processing (NLP): Machine Learning (ML) utilizing the PyCa

Benchmarking Retrieval Strategies for Biomedical Retrieval-Augmented Generation: A Controlled Empirical Study

Model ReleasesDGX agent

arXiv:2605.02520v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) offers a well-established path to grounding large language model (LLM) outputs in external knowledge, yet the quest

Beyond Perplexity: Character Distribution Signatures and the MDTA Benchmark for AI Text Detection

Model ReleasesDGX agent

arXiv:2605.01647v1 Announce Type: new Abstract: Training-free AI text detection methods primarily rely on model log-probabilities, achieving strong performance through approaches like Binoculars and D

Beyond Semantic Relevance: Counterfactual Risk Minimization for Robust Retrieval-Augmented Generation

SafetyDGX agent

arXiv:2605.01302v1 Announce Type: new Abstract: Standard Retrieval-Augmented Generation (RAG) systems predominantly rely on semantic relevance as a proxy for utility. However, this assumption collapse

Beyond Sentiment: A Multi-Agent Pipeline for Actionable Business Advice from Reviews

AgentsDGX agent

arXiv:2601.12024v2 Announce Type: replace-cross Abstract: Customer reviews contain valuable signals about service quality, but converting large-scale review corpora into actionable business recommenda

BIM Information Extraction Through LLM-based Adaptive Exploration

Model ReleasesDGX agent

arXiv:2605.01698v1 Announce Type: new Abstract: BIM models provide structured representations of building geometry, semantics, and topology, yet extracting specific information from them remains remar

Breaking the Silence: A Dataset and Benchmark for Bangla Text-to-Gloss Translation

Model ReleasesDGX agent

arXiv:2504.02293v3 Announce Type: replace Abstract: Gloss is a written approximation that bridges Sign Language (SL) and its corresponding spoken language. Despite a deaf and hard-of-hearing populatio

Bucketing the Good Apples: A Method for Diagnosing and Improving Causal Abstraction

ResearchDGX agent

arXiv:2605.02234v1 Announce Type: cross Abstract: We present a method for diagnosing interpretation in neural networks by identifying an input subspace where a proposed interpretation is highly faithf

Can AI Debias the News? LLM Interventions Improve Cross-Partisan Receptivity but LLMs Overestimate Their Own Effectiveness

HardwareDGX agent

arXiv:2605.01006v1 Announce Type: new Abstract: Partisan news media erode cross-partisan trust, but large language models (LLMs) offer a potential means of debiasing such content at scale. Across two

Can professional translators identify machine-generated text?

ResearchDGX agent

arXiv:2601.15828v3 Announce Type: replace Abstract: This study investigates whether professional translators without prior specialized training can reliably identify short stories generated in Italian

Causal2Vec: Improving Decoder-only LLMs as Embedding Models through a Contextual Token

Model ReleasesDGX agent

arXiv:2507.23386v3 Announce Type: replace Abstract: Decoder-only large language models (LLMs) have been increasingly adopted to build embedding models for diverse tasks. To overcome the inherent limit

Chain of Evidence: Pixel-Level Visual Attribution for Iterative Retrieval-Augmented Generation

ResearchDGX agent

arXiv:2605.01284v1 Announce Type: cross Abstract: Iterative Retrieval-Augmented Generation (iRAG) has emerged as a powerful paradigm for answering complex multi-hop questions by progressively retrievi

CLaC at SemEval-2026 Task 6: Response Clarity Detection in Political Discourse

Model ReleasesDGX agent

arXiv:2605.02170v1 Announce Type: new Abstract: In this paper, we present our system for SemEval-2026 Task 6 (CLARITY) on response clarity and evasion detection in question-answer pairs from U.S. pres

CLEAR: Revealing How Noise and Ambiguity Degrade Reliability in LLMs for Medicine

ApplicationsDGX agent

arXiv:2605.01011v1 Announce Type: new Abstract: Medical large language model (LLM) evaluations rely on simplified, exam-style benchmarks that rarely reflect the ambiguity of real-world medical inquiri

COCORELI: Enforcing Execution Preconditions for Reliable Collaborative Instruction Following

Model ReleasesDGX agent

arXiv:2509.04470v2 Announce Type: replace Abstract: Autonomous agents executing human instructions must operate reliably even when instructions are incomplete. While recent approaches improve detectio

Code-switching in text and speech challenges information-theoretic speaker design

ResearchDGX agent

arXiv:2408.04596v2 Announce Type: replace Abstract: In this work, we use language modeling to investigate the factors that influence insertional code-switching. Code-switching occurs when a speaker al

CoFrGeNet: Continued Fraction Architectures for Language Generation

ResearchDGX agent

arXiv:2601.21766v3 Announce Type: replace Abstract: Transformers are arguably the preferred architecture for language generation. In this paper, inspired by continued fractions, we introduce a new fun

Compared to What? Baselines and Metrics for Counterfactual Prompting

SafetyDGX agent

arXiv:2605.01048v1 Announce Type: new Abstract: Counterfactual prompting (i.e., perturbing a single factor and measuring output change) is widely used to evaluate things like LLM bias and CoT faithful

Component-Aware Self-Speculative Decoding in Hybrid Language Models

Model ReleasesDGX agent

arXiv:2605.01106v1 Announce Type: new Abstract: Speculative decoding accelerates autoregressive inference by drafting candidate tokens with a fast model and verifying them in parallel with the target.

Compositional Multi-hop Factual Error Correction via Decomposition-and-Injection

ApplicationsDGX agent

arXiv:2605.02277v1 Announce Type: new Abstract: Factual Error Correction (FEC) aims to revise inaccurate text into statements that are factually consistent with external evidence. Although recent meth

Compute Optimal Tokenization

Model ReleasesDGX agent

arXiv:2605.01188v1 Announce Type: new Abstract: Scaling laws enable the optimal selection of data amount and language model size, yet the impact of the data unit, the token, on this relationship remai

Confident, Calibrated, or Complicit: Safety Alignment and Ideological Bias in LLM Hate Speech Detection

SafetyDGX agent

arXiv:2509.00673v2 Announce Type: replace Abstract: We investigate the efficacy of Large Language Models (LLMs) in detecting implicit and explicit hate speech, examining how models with minimal safety

Constructing Interpretable Features from Compositional Neuron Groups

Model ReleasesDGX agent

arXiv:2506.10920v2 Announce Type: replace Abstract: A central goal for mechanistic interpretability has been to identify the right units of analysis in large language models (LLMs) that causally expla

ContextualJailbreak: Evolutionary Red-Teaming via Simulated Conversational Priming

Model ReleasesDGX agent

arXiv:2605.02647v1 Announce Type: new Abstract: Large language models (LLMs) remain vulnerable to jailbreak attacks that bypass safety alignment and elicit harmful responses. A growing body of work sh

Control Reinforcement Learning: Interpretable Token-Level Steering of LLMs via Sparse Autoencoder Features

Model ReleasesDGX agent

arXiv:2602.10437v3 Announce Type: replace-cross Abstract: Sparse autoencoders (SAEs) decompose language model activations into interpretable features, but existing methods reveal only which features a

Controlled Paraphrase Geometry in Sentence Embedding Space: Local Manifold Modeling and Latent Probing

ResearchDGX agent

arXiv:2605.01073v1 Announce Type: new Abstract: The paper studies the local geometry of embedding clouds induced by controlled local classes of semantically close sentences. The central question is ho

← Previous
1…8687888990…129
Next →