AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
Agents

Self-Induced Outcome Potential: Turn-Level Credit Assignment for Agents without Verifiers

DGX agent

arXiv:2605.04984v1 Announce Type: cross Abstract: Long-horizon LLM agents depend on intermediate information-gathering turns, yet training feedback is usually observed only at the final answer, becaus

agentsarxiv-cs-cl
7 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Self-Prompting Small Language Models for Privacy-Sensitive Clinical Information Extraction

DGX agent

arXiv:2605.04221v1 Announce Type: new Abstract: Clinical named entity recognition from dental progress notes is challenging because documentation is highly unstructured, domain-specific, and often pri

model-releasesarxiv-cs-cl
7 May 2026
Research

Sentiment Analysis and Customer Satisfaction Prediction on E-Commerce Platforms Based on YouTube Comments Using the XGBoost Algorithm

DGX agent

arXiv:2605.04887v1 Announce Type: new Abstract: The exponential expansion of digital commerce in Indonesia has significantly shifted consumer interactions toward video-centric social networks, particu

researcharxiv-cs-cl
7 May 2026
Model Releases

Single-Position Intervention Fails: Distributed Output Templates Drive In-Context Learning

DGX agent

arXiv:2605.04061v1 Announce Type: cross Abstract: Understanding how large language models encode task identity from few-shot demonstrations is a central open problem in mechanistic interpretability. P

model-releasesarxiv-cs-cl
7 May 2026
Model Releases

Sparse Autoencoder Decomposition of Clinical Sequence Model Representations: Feature Complexity, Task Specialisation, and Mortality Prediction

DGX agent

arXiv:2605.04072v1 Announce Type: cross Abstract: Sparse autoencoders (SAEs) have been applied to large language models and protein language models, but not systematically to electronic health record

model-releasesarxiv-cs-cl
7 May 2026
Safety

Sparse Tokens Suffice: Jailbreaking Audio Language Models via Token-Aware Gradient Optimization

DGX agent

arXiv:2605.04700v1 Announce Type: cross Abstract: Jailbreak attacks on audio language models (ALMs) optimize audio perturbations to elicit unsafe generations, and they typically update the entire wave

safetyarxiv-cs-cl
7 May 2026
Model Releases

SpecPL: Disentangling Spectral Granularity for Prompt Learning

DGX agent

arXiv:2605.04504v1 Announce Type: cross Abstract: Existing prompt learning for VLMs exhibits a modality asymmetry, predominantly optimizing text tokens while still relying on frozen visual encoder as

model-releasesarxiv-cs-cl
7 May 2026
Local Ai

Stabilizing LLM Supervised Fine-Tuning via Explicit Distributional Control

DGX agent

arXiv:2605.04468v1 Announce Type: cross Abstract: Post-training large language models (LLMs) often suffers from catastrophic forgetting, where improvements on a target objective degrade previously acq

local-aiarxiv-cs-cl
7 May 2026
Model Releases

Storage Is Not Memory: A Retrieval-Centered Architecture for Agent Recall

DGX agent

arXiv:2605.04897v1 Announce Type: new Abstract: Extraction at ingestion is the wrong primitive for agent memory: content discarded before the query is known cannot be recovered at retrieval time. We p

model-releasesarxiv-cs-cl
7 May 2026
Model Releases

StoryAlign: Evaluating and Training Reward Models for Story Generation

DGX agent

arXiv:2605.04831v1 Announce Type: new Abstract: Story generation aims to automatically produce coherent, structured, and engaging narratives. Although large language models (LLMs) have significantly a

model-releasesarxiv-cs-cl
7 May 2026
Model Releases

SWAN: Semantic Watermarking with Abstract Meaning Representation

DGX agent

arXiv:2605.04305v1 Announce Type: new Abstract: We introduce SWAN (Semantic Watermarking with Abstract Meaning Representation), a novel framework that embeds watermark signatures into the semantic str

model-releasesarxiv-cs-cl
7 May 2026
Model Releases

TabEmbed: Benchmarking and Learning Generalist Embeddings for Tabular Understanding

DGX agent

arXiv:2605.04962v1 Announce Type: new Abstract: Foundation models have established unified representations for natural language processing, yet this paradigm remains largely unexplored for tabular dat

model-releasesarxiv-cs-cl
7 May 2026
Research

TajikNLP: An Open-Source Toolkit for Comprehensive Text Processing of Tajik (Cyrillic Script)

DGX agent

arXiv:2605.04583v1 Announce Type: new Abstract: The Tajik language, written in Cyrillic script, remains severely under-resourced in terms of publicly available natural language processing (NLP) toolki

researcharxiv-cs-cl
7 May 2026
Model Releases

Telegraph English: Semantic Prompt Compression via Structured Symbolic Rewriting

DGX agent

arXiv:2605.04426v1 Announce Type: new Abstract: We introduce Telegraph English (TE), a prompt-compression protocol that rewrites natural language into a symbol-rich, formally-structured dialect. Where

model-releasesarxiv-cs-cl
7 May 2026
Local Ai

Text Corpora as Concept Fields: Black-Box Hallucination and Novelty Measurement

DGX agent

arXiv:2605.05103v1 Announce Type: new Abstract: We introduce the **Concept Field** of a text corpus: a local drift field with pointwise uncertainty, estimated in sentence-embedding space from the delt

local-aiarxiv-cs-cl
7 May 2026
Research

The First Token Knows: Single-Decode Confidence for Hallucination Detection

DGX agent

arXiv:2605.05166v1 Announce Type: new Abstract: Self-consistency detects hallucinations by generating multiple sampled answers to a question and measuring agreement, but this requires repeated decodin

researcharxiv-cs-cl
7 May 2026
Research

The Impact of Vocabulary Overlaps on Knowledge Transfer in Multilingual Machine Translation

DGX agent

arXiv:2605.04196v1 Announce Type: new Abstract: Knowledge transfer, especially across related languages, has been found beneficial for multilingual neural machine translation (MNMT), but some aspects

researcharxiv-cs-cl
7 May 2026
Research

The Impossibility Triangle of Long-Context Modeling

DGX agent

arXiv:2605.05066v1 Announce Type: new Abstract: We identify and prove a fundamental trade-off governing long-sequence models: no model can simultaneously achieve (i) per-step computation independent o

researcharxiv-cs-cl
7 May 2026
Research

The Newsworthiness of Brazilian Distress: A Peak Analysis on Time Series of International Media Attention to Disasters in Brazil

DGX agent

arXiv:2605.04552v1 Announce Type: new Abstract: Media coverage influences disaster response, yet the drivers of international media attention to local events remain unevenly understood. Brazil offers

researcharxiv-cs-cl
7 May 2026
Research

The Pinocchio Dimension: Phenomenality of Experience as the Primary Axis of LLM Psychometric Differences

DGX agent

arXiv:2605.05080v1 Announce Type: new Abstract: We administer 45 validated psychometric questionnaires to 50 large language models (LLMs) to identify the dimensions along which LLMs differ psychometri

researcharxiv-cs-cl
7 May 2026
Model Releases

The Shape of Beliefs: Geometry, Dynamics, and Interventions along Representation Manifolds of Language Models' Posteriors

DGX agent

arXiv:2602.02315v2 Announce Type: replace Abstract: Large language models (LLMs) form implicit beliefs (posteriors over latent variables) from prompts, but we lack a mechanistic account of how these b

model-releasesarxiv-cs-cl
7 May 2026
Research

Towards Distillation-Resistant Large Language Models: An Information-Theoretic Perspective

DGX agent

arXiv:2602.03396v3 Announce Type: replace Abstract: Proprietary large language models (LLMs) embody substantial economic value and are generally exposed only as black-box APIs, yet adversaries can sti

researcharxiv-cs-cl
7 May 2026
Research

Towards Self-Referential Analytic Assessment: A Profile-Based Approach to L2 Writing Evaluation with LLMs

DGX agent

arXiv:2605.04298v1 Announce Type: new Abstract: Automated essay scoring (AES) research often relies on rank-based correlation metrics to validate analytic assessment. However, such metrics obscure bot

researcharxiv-cs-cl
7 May 2026
Model Releases

TSCG: Deterministic Tool-Schema Compilation for Agentic LLM Deployments

DGX agent

arXiv:2605.04107v1 Announce Type: cross Abstract: Production agent frameworks (OpenAI Function Calling, Anthropic Tool Use, MCP) transmit tool schemas as JSON, a format designed for machine parsing, n

model-releasesarxiv-cs-cl
7 May 2026
Model Releases

UFAL-CUNI at SemEval-2026 Task 11: An Efficient Modular Neuro-symbolic Method for Syllogistic Reasoning

DGX agent

arXiv:2605.04941v1 Announce Type: new Abstract: This paper describes our system submitted to SemEval-2026 Task 11: Disentangling Content and Formal Reasoning in Large Language Models. We present an ef

model-releasesarxiv-cs-cl
7 May 2026
Safety

Uncertainty-Aware Exploratory Direct Preference Optimization for Multimodal Large Language Models

DGX agent

arXiv:2605.04874v1 Announce Type: cross Abstract: Direct Preference Optimization (DPO) has proven to be an effective solution for mitigating hallucination in Multimodal Large Language Models (MLLMs) b

safetyarxiv-cs-cl
7 May 2026
Local Ai

Uncovering Cross-Objective Interference in Multi-Objective Alignment

DGX agent

arXiv:2602.06869v2 Announce Type: replace Abstract: We study a persistent failure mode in multi-objective alignment for large language models (LLMs): training improves performance on only a subset of

local-aiarxiv-cs-cl
7 May 2026
Research

Unintended Negative Impacts of Promotional Language in Patent Evaluation

DGX agent

arXiv:2605.04926v1 Announce Type: new Abstract: Promotional language has been increasingly used to aid the communication of innovative ideas in science. Yet, less is known about its role in the contex

researcharxiv-cs-cl
7 May 2026
Local Ai

UniVer: A Unified Perspective for Multi-step and Multi-draft Speculative Decoding

DGX agent

arXiv:2605.04543v1 Announce Type: new Abstract: Speculative decoding accelerates Large Language Models via draft-then-verify, where verification can be framed as an Optimal Transport (OT) problem. Exi

local-aiarxiv-cs-cl
7 May 2026
Research

When Relations Break: Analyzing Relation Hallucination in Vision-Language Model Under Rotation and Noise

DGX agent

arXiv:2605.05045v1 Announce Type: cross Abstract: Vision-language models (VLMs) achieve strong multimodal performance but remain prone to relation hallucination, which requires accurate reasoning over

researcharxiv-cs-cl
7 May 2026
Safety

Why Expert Alignment Is Hard: Evidence from Subjective Evaluation

DGX agent

arXiv:2605.04972v1 Announce Type: new Abstract: Aligning large language models with expert judgment is especially difficult in subjective evaluation tasks, where experts may disagree, rely on tacit cr

safetyarxiv-cs-cl
7 May 2026
Research

Why Geometric Continuity Emerges in Deep Neural Networks: Residual Connections and Rotational Symmetry Breaking

DGX agent

arXiv:2605.04971v1 Announce Type: cross Abstract: Weight matrices in deep networks exhibit geometric continuity -- principal singular vectors of adjacent layers point in similar directions. While this

researcharxiv-cs-cl
7 May 2026
Research

A Comparison of Traditional Machine Learning Algorithms and LSTM-Based Deep Learning Models for Email Sentiment Analysis

DGX agent

arXiv:2605.03440v1 Announce Type: new Abstract: The rapid growth of electronic communication has necessitated more robust systems for email classification and sentiment detection. This study presents

researcharxiv-cs-cl
6 May 2026
Research

A Comprehensive Analysis of Tokenization and Self-Supervised Learning in End-to-End Automatic Speech Recognition applied on French Language

DGX agent

arXiv:2605.03696v1 Announce Type: new Abstract: The performance of end-to-end automatic speech recognition (ASR) systems enables their increasing integration into numerous applications. While there ar

researcharxiv-cs-cl
6 May 2026
Research

A Paradigm for Interpreting Metrics and Identifying Critical Errors in Automatic Speech Recognition

DGX agent

arXiv:2605.03671v1 Announce Type: new Abstract: The most commonly used metrics for evaluating automatic speech transcriptions, namely Word Error Rate (WER) and Character Error Rate (CER), have been he

researcharxiv-cs-cl
6 May 2026
Safety

ADAPTS: Agentic Decomposition for Automated Protocol-agnostic Tracking of Symptoms

DGX agent

arXiv:2605.03212v1 Announce Type: cross Abstract: Modeling latent clinical constructs from unconstrained clinical interactions is a unique challenge in affective computing. We present ADAPTS (Agentic

safetyarxiv-cs-cl
6 May 2026
Model Releases

AfriqueLLM: How Data Mixing and Model Architecture Impact Continued Pre-training for African Languages

DGX agent

arXiv:2601.06395v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly multilingual, yet open models continue to underperform relative to proprietary systems, with the gap m

model-releasesarxiv-cs-cl
6 May 2026
Model Releases

AfriVox-v2: A Domain-Verticalized Benchmark for In-the-Wild African Speech Recognition

DGX agent

arXiv:2605.03590v1 Announce Type: new Abstract: Recent large language models (LLMs) show strong speech recognition and translation capabilities for high-resource languages. However, African languages

model-releasesarxiv-cs-cl
6 May 2026
Model Releases

Agentic-imodels: Evolving agentic interpretability tools via autoresearch

DGX agent

arXiv:2605.03808v1 Announce Type: cross Abstract: Agentic data science (ADS) systems are rapidly improving their capability to autonomously analyze, fit, and interpret data, potentially moving towards

model-releasesarxiv-cs-cl
6 May 2026
Research

An ERP Study of Recursive Possessive Parsing in ASD Children and Its Cognitive Neuro Mechanisms

DGX agent

arXiv:2605.03447v1 Announce Type: new Abstract: Recursive structures are a core property of human language, yet little is known about how children with autism spectrum disorder (ASD) process complex r

researcharxiv-cs-cl
6 May 2026
Model Releases

Annotation Quality in Aspect-Based Sentiment Analysis: A Case Study Comparing Experts, Students, Crowdworkers, and Large Language Model

DGX agent

arXiv:2605.03624v1 Announce Type: new Abstract: Aspect-Based Sentiment Analysis (ABSA) enables fine-grained opinion analysis by identifying sentiments toward specific aspects or targets within a text.

model-releasesarxiv-cs-cl
6 May 2026
Research

Atomic Fact-Checking Increases Clinician Trust in Large Language Model Recommendations for Oncology Decision Support: A Randomized Controlled Trial

DGX agent

arXiv:2605.03916v1 Announce Type: new Abstract: Question: Does atomic fact-checking, which decomposes AI treatment recommendations into individually verifiable claims linked to source guideline docume

researcharxiv-cs-cl
6 May 2026
Model Releases

AutoRAGTuner: A Declarative Framework for Automatic Optimization of RAG Pipelines

DGX agent

arXiv:2605.02967v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) enhances LLMs, but performance is highly sensitive to complex architecture designs and hyper-parameter configurat

model-releasesarxiv-cs-cl
6 May 2026
Local Ai

Benchmarking Local Language Models for Social Robots using Edge Devices

DGX agent

arXiv:2605.03111v1 Announce Type: cross Abstract: Social-educational robots designed for socially interactive pedagogical support, such as the Robot Study Companion (RSC), rely on responsive, privacy-

local-aiarxiv-cs-cl
6 May 2026
Research

Benchmarking Logistic Regression, SVM, Naive Bayes, and IndoBERT Fine-Tuning for Sentiment Analysis on Indonesian Product Reviews

DGX agent

arXiv:2605.03439v1 Announce Type: new Abstract: The exponential growth of e-commerce platforms in Indonesia has generated a massive volume of user-generated product reviews. Analyzing the sentiment of

researcharxiv-cs-cl
6 May 2026
Model Releases

Benchmarking Parameter-Efficient Fine-Tuning of Large Language Models for Low-Resource Tajik Text Generation with the Tajik Web Corpus

DGX agent

arXiv:2605.03742v1 Announce Type: new Abstract: This paper is devoted to the adaptation of generative large language models for the Tajik language, a low-resource language with Cyrillic script. To ove

model-releasesarxiv-cs-cl
6 May 2026
Research

BiMind: A Dual-Head Reasoning Model with Attention-Geometry Adapter for Incorrect Information Detection

DGX agent

arXiv:2604.06022v2 Announce Type: replace Abstract: Incorrect information poses significant challenges by disrupting content veracity and integrity, yet most detection approaches struggle to jointly b

researcharxiv-cs-cl
6 May 2026
Local Ai

BIT.UA-AAUBS at ArchEHR-QA 2026: Evaluating Open-Source and Proprietary LLMs via Prompting in Low-Resource QA

DGX agent

arXiv:2605.03618v1 Announce Type: new Abstract: This paper presents the joint participation of the BIT.UA and AAUBS groups in the ArchEHR-QA 2026 shared task, which focuses on clinical question answer

local-aiarxiv-cs-cl
6 May 2026
← Previous
1…105106107108109…161
Next →