AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
Model Releases

Accurate Legal Reasoning at Scale: Neuro-Symbolic Offloading and Structural Auditability for Robust Legal Adjudication

DGX agent

arXiv:2605.02472v1 Announce Type: new Abstract: Legal texts often contain computational legal clauses--provisions whose understanding requires complex logic. While frontier Large Reasoning Models (LRM

model-releasesarxiv-cs-cl
5 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

Adaptive GoGI-Skip: Coupling Goal-Gradient Importance with Dynamic Uncertainty for Efficient Reasoning

DGX agent

arXiv:2505.08392v3 Announce Type: replace Abstract: Chain-of-Thought (CoT) prompting trades inference speed for reasoning accuracy. Existing compressors force a compromise as static gradient technique

safetyarxiv-cs-cl
5 May 2026
Model Releases

Addressing Data Scarcity in Bangla Fake News Detection: An LLM-Based Dataset Augmentation Approach

DGX agent

arXiv:2605.01292v1 Announce Type: new Abstract: The growing spread of misinformation in digital media highlights the need for reliable fake news detection systems, yet progress in under-resourced lang

model-releasesarxiv-cs-cl
5 May 2026
Agents

AgentXRay: White-Boxing Agentic Systems via Workflow Reconstruction

DGX agent

arXiv:2602.05353v3 Announce Type: replace-cross Abstract: Large Language Models have shown strong capabilities in complex problem solving, yet many agentic systems remain difficult to interpret and co

agentsarxiv-cs-cl
5 May 2026
Safety

ALIGNS: Unlocking nomological networks in psychological measurement through a large language model

DGX agent

arXiv:2509.09723v3 Announce Type: replace Abstract: Psychological measurement is critical to many disciplines. Despite advances in measurement, building nomological networks, theoretical maps of how c

safetyarxiv-cs-cl
5 May 2026
Tutorials

An Information-theoretic Propagation Denoising and Fusion Framework for Fake News Detection

DGX agent

arXiv:2605.02259v1 Announce Type: new Abstract: Incomplete propagation data significantly hinders robust fake news detection. Recent approaches leverage large language models to simulate missing user

tutorialsarxiv-cs-cl
5 May 2026
Research

Argumentation for Explainable and Globally Contestable Decision Support with LLMs

DGX agent

arXiv:2603.14643v2 Announce Type: replace-cross Abstract: Large language models (LLMs) exhibit strong general capabilities, but their deployment in high-stakes domains is hindered by their opacity and

researcharxiv-cs-cl
5 May 2026
Safety

ARGUS: Policy-Adaptive Ad Governance via Evolving Reinforcement with Adversarial Umpiring

DGX agent

arXiv:2605.02200v1 Announce Type: new Abstract: Online advertising governance faces significant challenges due to the non-stationary nature of regulatory policies, where emerging mandates (e.g., restr

safetyarxiv-cs-cl
5 May 2026
Model Releases

Arithmetic in the Wild: Llama uses Base-10 Addition to Reason About Cyclic Concepts

DGX agent

arXiv:2605.01148v1 Announce Type: cross Abstract: Does structure in representations imply structure in computation? We study how Llama-3.1-8B reasons over cyclic concepts (e.g., 'what month is six mon

model-releasesarxiv-cs-cl
5 May 2026
Safety

Artificial intelligence language technologies in multilingual healthcare: Grand challenges ahead

DGX agent

arXiv:2605.01441v1 Announce Type: new Abstract: AI language technologies (AILTs), increasingly enabled by large language models (LLMs), are becoming embedded in multilingual healthcare workflows for t

safetyarxiv-cs-cl
5 May 2026
Research

ATLAS: Article Tracking, Linking, and Analysis of Swedish Encyclopedias

DGX agent

arXiv:2605.02466v1 Announce Type: new Abstract: The digitization of old encyclopedias represents an important step to improve access to historically structured knowledge. Often, however, this process

researcharxiv-cs-cl
5 May 2026
Model Releases

ATR-Bench: A Federated Learning Benchmark for Adaptation, Trust, and Reasoning

DGX agent

arXiv:2505.16850v2 Announce Type: replace-cross Abstract: Federated Learning (FL) has emerged as a promising paradigm for collaborative model training while preserving data privacy across decentralize

model-releasesarxiv-cs-cl
5 May 2026
Research

Attention Sink Forges Native MoE in Attention Layers: Sink-Aware Training to Address Head Collapse

DGX agent

arXiv:2602.01203v2 Announce Type: replace Abstract: Large Language Models (LLMs) often assign disproportionate attention to the first token, a phenomenon known as the attention sink. Several recent ap

researcharxiv-cs-cl
5 May 2026
Safety

Attention Sinks in Massively Multilingual Neural Machine Translation:Discovery, Analysis, and Mitigation

DGX agent

arXiv:2605.01229v1 Announce Type: cross Abstract: Cross-attention patterns in neural machine translation (NMT) are widely used to study how multilingual models align linguistic structure. We report a

safetyarxiv-cs-cl
5 May 2026
Applications

Attribution-Guided Pruning for Insight and Control: Circuit Discovery and Targeted Correction in Small-scale LLMs

DGX agent

arXiv:2506.13727v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are widely deployed in real-world applications, yet their internal mechanisms remain difficult to interpret and c

applicationsarxiv-cs-cl
5 May 2026
Safety

Auditing demographic bias in AI-based emergency police dispatch: a cross-lingual evaluation of eleven large language models

DGX agent

arXiv:2605.01451v1 Announce Type: new Abstract: Large language models (LLMs) are rapidly being integrated into high-stakes public safety systems, including emergency call triage and dispatch decision

safetyarxiv-cs-cl
5 May 2026
Model Releases

Automated Interpretability and Feature Discovery in Language Models with Agents

DGX agent

arXiv:2605.01555v1 Announce Type: new Abstract: We introduce an autonomous multiagent framework for mechanistic interpretability that automates both explaining and finding internal features in large l

model-releasesarxiv-cs-cl
5 May 2026
Research

Automatic Correction of Writing Anomalies in Hausa Texts

DGX agent

arXiv:2506.03820v2 Announce Type: replace Abstract: Hausa texts are often characterized by writing anomalies, such as incorrect character substitutions and spacing errors, which sometimes hinder natur

researcharxiv-cs-cl
5 May 2026
Applications

Automatic Reflection Level Classification in Hungarian Student Essays

DGX agent

arXiv:2605.02402v1 Announce Type: new Abstract: Reflective thinking is a key competency in education, but assessing reflective writing remains a time-consuming and subjective task for education expert

applicationsarxiv-cs-cl
5 May 2026
Research

Balalaika: Data-Centric, Prosody-Aware Annotation Pipeline for Russian Speech

DGX agent

arXiv:2507.13563v2 Announce Type: replace Abstract: We introduce Balalaika, an open-source, data-centric pipeline for processing audio and producing prosody-aware annotations. It combines semantic VAD

researcharxiv-cs-cl
5 May 2026
Research

BaldWhisper: Faster Whisper with Head Shearing and Layer Merging

DGX agent

arXiv:2510.08599v2 Announce Type: replace-cross Abstract: Pruning large pre-trained transformers in a data-scarce scenario is challenging, as it often requires massive retraining data to recover perfo

researcharxiv-cs-cl
5 May 2026
Model Releases

Beating the Style Detector: Three Hours of Agentic Research on the AI-Text Arms Race

DGX agent

arXiv:2605.02620v1 Announce Type: new Abstract: Reproducing an empirical NLP study used to take weeks. Given the released data and a modern agentic-research harness, we redo every experiment of a rece

model-releasesarxiv-cs-cl
5 May 2026
Research

Benchmarking LightGBM and BiLSTM for Sentiment Analysis on Indonesian E-Commerce Reviews

DGX agent

arXiv:2605.01322v1 Announce Type: new Abstract: This study presents a comparative analysis between two primary approaches in Natural Language Processing (NLP): Machine Learning (ML) utilizing the PyCa

researcharxiv-cs-cl
5 May 2026
Model Releases

Benchmarking Retrieval Strategies for Biomedical Retrieval-Augmented Generation: A Controlled Empirical Study

DGX agent

arXiv:2605.02520v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) offers a well-established path to grounding large language model (LLM) outputs in external knowledge, yet the quest

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Beyond Perplexity: Character Distribution Signatures and the MDTA Benchmark for AI Text Detection

DGX agent

arXiv:2605.01647v1 Announce Type: new Abstract: Training-free AI text detection methods primarily rely on model log-probabilities, achieving strong performance through approaches like Binoculars and D

model-releasesarxiv-cs-cl
5 May 2026
Safety

Beyond Semantic Relevance: Counterfactual Risk Minimization for Robust Retrieval-Augmented Generation

DGX agent

arXiv:2605.01302v1 Announce Type: new Abstract: Standard Retrieval-Augmented Generation (RAG) systems predominantly rely on semantic relevance as a proxy for utility. However, this assumption collapse

safetyarxiv-cs-cl
5 May 2026
Agents

Beyond Sentiment: A Multi-Agent Pipeline for Actionable Business Advice from Reviews

DGX agent

arXiv:2601.12024v2 Announce Type: replace-cross Abstract: Customer reviews contain valuable signals about service quality, but converting large-scale review corpora into actionable business recommenda

agentsarxiv-cs-cl
5 May 2026
Model Releases

BIM Information Extraction Through LLM-based Adaptive Exploration

DGX agent

arXiv:2605.01698v1 Announce Type: new Abstract: BIM models provide structured representations of building geometry, semantics, and topology, yet extracting specific information from them remains remar

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Breaking the Silence: A Dataset and Benchmark for Bangla Text-to-Gloss Translation

DGX agent

arXiv:2504.02293v3 Announce Type: replace Abstract: Gloss is a written approximation that bridges Sign Language (SL) and its corresponding spoken language. Despite a deaf and hard-of-hearing populatio

model-releasesarxiv-cs-cl
5 May 2026
Research

Bucketing the Good Apples: A Method for Diagnosing and Improving Causal Abstraction

DGX agent

arXiv:2605.02234v1 Announce Type: cross Abstract: We present a method for diagnosing interpretation in neural networks by identifying an input subspace where a proposed interpretation is highly faithf

researcharxiv-cs-cl
5 May 2026
Hardware

Can AI Debias the News? LLM Interventions Improve Cross-Partisan Receptivity but LLMs Overestimate Their Own Effectiveness

DGX agent

arXiv:2605.01006v1 Announce Type: new Abstract: Partisan news media erode cross-partisan trust, but large language models (LLMs) offer a potential means of debiasing such content at scale. Across two

hardwarearxiv-cs-cl
5 May 2026
Research

Can professional translators identify machine-generated text?

DGX agent

arXiv:2601.15828v3 Announce Type: replace Abstract: This study investigates whether professional translators without prior specialized training can reliably identify short stories generated in Italian

researcharxiv-cs-cl
5 May 2026
Model Releases

Causal2Vec: Improving Decoder-only LLMs as Embedding Models through a Contextual Token

DGX agent

arXiv:2507.23386v3 Announce Type: replace Abstract: Decoder-only large language models (LLMs) have been increasingly adopted to build embedding models for diverse tasks. To overcome the inherent limit

model-releasesarxiv-cs-cl
5 May 2026
Research

Chain of Evidence: Pixel-Level Visual Attribution for Iterative Retrieval-Augmented Generation

DGX agent

arXiv:2605.01284v1 Announce Type: cross Abstract: Iterative Retrieval-Augmented Generation (iRAG) has emerged as a powerful paradigm for answering complex multi-hop questions by progressively retrievi

researcharxiv-cs-cl
5 May 2026
Model Releases

CLaC at SemEval-2026 Task 6: Response Clarity Detection in Political Discourse

DGX agent

arXiv:2605.02170v1 Announce Type: new Abstract: In this paper, we present our system for SemEval-2026 Task 6 (CLARITY) on response clarity and evasion detection in question-answer pairs from U.S. pres

model-releasesarxiv-cs-cl
5 May 2026
Applications

CLEAR: Revealing How Noise and Ambiguity Degrade Reliability in LLMs for Medicine

DGX agent

arXiv:2605.01011v1 Announce Type: new Abstract: Medical large language model (LLM) evaluations rely on simplified, exam-style benchmarks that rarely reflect the ambiguity of real-world medical inquiri

applicationsarxiv-cs-cl
5 May 2026
Model Releases

COCORELI: Enforcing Execution Preconditions for Reliable Collaborative Instruction Following

DGX agent

arXiv:2509.04470v2 Announce Type: replace Abstract: Autonomous agents executing human instructions must operate reliably even when instructions are incomplete. While recent approaches improve detectio

model-releasesarxiv-cs-cl
5 May 2026
Research

Code-switching in text and speech challenges information-theoretic speaker design

DGX agent

arXiv:2408.04596v2 Announce Type: replace Abstract: In this work, we use language modeling to investigate the factors that influence insertional code-switching. Code-switching occurs when a speaker al

researcharxiv-cs-cl
5 May 2026
Research

CoFrGeNet: Continued Fraction Architectures for Language Generation

DGX agent

arXiv:2601.21766v3 Announce Type: replace Abstract: Transformers are arguably the preferred architecture for language generation. In this paper, inspired by continued fractions, we introduce a new fun

researcharxiv-cs-cl
5 May 2026
Safety

Compared to What? Baselines and Metrics for Counterfactual Prompting

DGX agent

arXiv:2605.01048v1 Announce Type: new Abstract: Counterfactual prompting (i.e., perturbing a single factor and measuring output change) is widely used to evaluate things like LLM bias and CoT faithful

safetyarxiv-cs-cl
5 May 2026
Model Releases

Component-Aware Self-Speculative Decoding in Hybrid Language Models

DGX agent

arXiv:2605.01106v1 Announce Type: new Abstract: Speculative decoding accelerates autoregressive inference by drafting candidate tokens with a fast model and verifying them in parallel with the target.

model-releasesarxiv-cs-cl
5 May 2026
Applications

Compositional Multi-hop Factual Error Correction via Decomposition-and-Injection

DGX agent

arXiv:2605.02277v1 Announce Type: new Abstract: Factual Error Correction (FEC) aims to revise inaccurate text into statements that are factually consistent with external evidence. Although recent meth

applicationsarxiv-cs-cl
5 May 2026
Model Releases

Compute Optimal Tokenization

DGX agent

arXiv:2605.01188v1 Announce Type: new Abstract: Scaling laws enable the optimal selection of data amount and language model size, yet the impact of the data unit, the token, on this relationship remai

model-releasesarxiv-cs-cl
5 May 2026
Safety

Confident, Calibrated, or Complicit: Safety Alignment and Ideological Bias in LLM Hate Speech Detection

DGX agent

arXiv:2509.00673v2 Announce Type: replace Abstract: We investigate the efficacy of Large Language Models (LLMs) in detecting implicit and explicit hate speech, examining how models with minimal safety

safetyarxiv-cs-cl
5 May 2026
Model Releases

Constructing Interpretable Features from Compositional Neuron Groups

DGX agent

arXiv:2506.10920v2 Announce Type: replace Abstract: A central goal for mechanistic interpretability has been to identify the right units of analysis in large language models (LLMs) that causally expla

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

ContextualJailbreak: Evolutionary Red-Teaming via Simulated Conversational Priming

DGX agent

arXiv:2605.02647v1 Announce Type: new Abstract: Large language models (LLMs) remain vulnerable to jailbreak attacks that bypass safety alignment and elicit harmful responses. A growing body of work sh

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Control Reinforcement Learning: Interpretable Token-Level Steering of LLMs via Sparse Autoencoder Features

DGX agent

arXiv:2602.10437v3 Announce Type: replace-cross Abstract: Sparse autoencoders (SAEs) decompose language model activations into interpretable features, but existing methods reveal only which features a

model-releasesarxiv-cs-cl
5 May 2026
Research

Controlled Paraphrase Geometry in Sentence Embedding Space: Local Manifold Modeling and Latent Probing

DGX agent

arXiv:2605.01073v1 Announce Type: new Abstract: The paper studies the local geometry of embedding clouds induced by controlled local classes of semantically close sentences. The central question is ho

researcharxiv-cs-cl
5 May 2026
← Previous
1…108109110111112…161
Next →