AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
7 May 2026

TajikNLP: An Open-Source Toolkit for Comprehensive Text Processing of Tajik (Cyrillic Script)

ResearchDGX agent

arXiv:2605.04583v1 Announce Type: new Abstract: The Tajik language, written in Cyrillic script, remains severely under-resourced in terms of publicly available natural language processing (NLP) toolki

Telegraph English: Semantic Prompt Compression via Structured Symbolic Rewriting

Model ReleasesDGX agent

arXiv:2605.04426v1 Announce Type: new Abstract: We introduce Telegraph English (TE), a prompt-compression protocol that rewrites natural language into a symbol-rich, formally-structured dialect. Where

Text Corpora as Concept Fields: Black-Box Hallucination and Novelty Measurement

Local AiDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.05103v1 Announce Type: new Abstract: We introduce the **Concept Field** of a text corpus: a local drift field with pointwise uncertainty, estimated in sentence-embedding space from the delt

The First Token Knows: Single-Decode Confidence for Hallucination Detection

ResearchDGX agent

arXiv:2605.05166v1 Announce Type: new Abstract: Self-consistency detects hallucinations by generating multiple sampled answers to a question and measuring agreement, but this requires repeated decodin

The Impact of Vocabulary Overlaps on Knowledge Transfer in Multilingual Machine Translation

ResearchDGX agent

arXiv:2605.04196v1 Announce Type: new Abstract: Knowledge transfer, especially across related languages, has been found beneficial for multilingual neural machine translation (MNMT), but some aspects

The Impossibility Triangle of Long-Context Modeling

ResearchDGX agent

arXiv:2605.05066v1 Announce Type: new Abstract: We identify and prove a fundamental trade-off governing long-sequence models: no model can simultaneously achieve (i) per-step computation independent o

The Newsworthiness of Brazilian Distress: A Peak Analysis on Time Series of International Media Attention to Disasters in Brazil

ResearchDGX agent

arXiv:2605.04552v1 Announce Type: new Abstract: Media coverage influences disaster response, yet the drivers of international media attention to local events remain unevenly understood. Brazil offers

The Pinocchio Dimension: Phenomenality of Experience as the Primary Axis of LLM Psychometric Differences

ResearchDGX agent

arXiv:2605.05080v1 Announce Type: new Abstract: We administer 45 validated psychometric questionnaires to 50 large language models (LLMs) to identify the dimensions along which LLMs differ psychometri

The Shape of Beliefs: Geometry, Dynamics, and Interventions along Representation Manifolds of Language Models' Posteriors

Model ReleasesDGX agent

arXiv:2602.02315v2 Announce Type: replace Abstract: Large language models (LLMs) form implicit beliefs (posteriors over latent variables) from prompts, but we lack a mechanistic account of how these b

Towards Distillation-Resistant Large Language Models: An Information-Theoretic Perspective

ResearchDGX agent

arXiv:2602.03396v3 Announce Type: replace Abstract: Proprietary large language models (LLMs) embody substantial economic value and are generally exposed only as black-box APIs, yet adversaries can sti

Towards Self-Referential Analytic Assessment: A Profile-Based Approach to L2 Writing Evaluation with LLMs

ResearchDGX agent

arXiv:2605.04298v1 Announce Type: new Abstract: Automated essay scoring (AES) research often relies on rank-based correlation metrics to validate analytic assessment. However, such metrics obscure bot

TSCG: Deterministic Tool-Schema Compilation for Agentic LLM Deployments

Model ReleasesDGX agent

arXiv:2605.04107v1 Announce Type: cross Abstract: Production agent frameworks (OpenAI Function Calling, Anthropic Tool Use, MCP) transmit tool schemas as JSON, a format designed for machine parsing, n

UFAL-CUNI at SemEval-2026 Task 11: An Efficient Modular Neuro-symbolic Method for Syllogistic Reasoning

Model ReleasesDGX agent

arXiv:2605.04941v1 Announce Type: new Abstract: This paper describes our system submitted to SemEval-2026 Task 11: Disentangling Content and Formal Reasoning in Large Language Models. We present an ef

Uncertainty-Aware Exploratory Direct Preference Optimization for Multimodal Large Language Models

SafetyDGX agent

arXiv:2605.04874v1 Announce Type: cross Abstract: Direct Preference Optimization (DPO) has proven to be an effective solution for mitigating hallucination in Multimodal Large Language Models (MLLMs) b

Uncovering Cross-Objective Interference in Multi-Objective Alignment

Local AiDGX agent

arXiv:2602.06869v2 Announce Type: replace Abstract: We study a persistent failure mode in multi-objective alignment for large language models (LLMs): training improves performance on only a subset of

Unintended Negative Impacts of Promotional Language in Patent Evaluation

ResearchDGX agent

arXiv:2605.04926v1 Announce Type: new Abstract: Promotional language has been increasingly used to aid the communication of innovative ideas in science. Yet, less is known about its role in the contex

UniVer: A Unified Perspective for Multi-step and Multi-draft Speculative Decoding

Local AiDGX agent

arXiv:2605.04543v1 Announce Type: new Abstract: Speculative decoding accelerates Large Language Models via draft-then-verify, where verification can be framed as an Optimal Transport (OT) problem. Exi

When Relations Break: Analyzing Relation Hallucination in Vision-Language Model Under Rotation and Noise

ResearchDGX agent

arXiv:2605.05045v1 Announce Type: cross Abstract: Vision-language models (VLMs) achieve strong multimodal performance but remain prone to relation hallucination, which requires accurate reasoning over

Why Expert Alignment Is Hard: Evidence from Subjective Evaluation

SafetyDGX agent

arXiv:2605.04972v1 Announce Type: new Abstract: Aligning large language models with expert judgment is especially difficult in subjective evaluation tasks, where experts may disagree, rely on tacit cr

Why Geometric Continuity Emerges in Deep Neural Networks: Residual Connections and Rotational Symmetry Breaking

ResearchDGX agent

arXiv:2605.04971v1 Announce Type: cross Abstract: Weight matrices in deep networks exhibit geometric continuity -- principal singular vectors of adjacent layers point in similar directions. While this

6 May 2026

A Comparison of Traditional Machine Learning Algorithms and LSTM-Based Deep Learning Models for Email Sentiment Analysis

ResearchDGX agent

arXiv:2605.03440v1 Announce Type: new Abstract: The rapid growth of electronic communication has necessitated more robust systems for email classification and sentiment detection. This study presents

A Comprehensive Analysis of Tokenization and Self-Supervised Learning in End-to-End Automatic Speech Recognition applied on French Language

ResearchDGX agent

arXiv:2605.03696v1 Announce Type: new Abstract: The performance of end-to-end automatic speech recognition (ASR) systems enables their increasing integration into numerous applications. While there ar

A Paradigm for Interpreting Metrics and Identifying Critical Errors in Automatic Speech Recognition

ResearchDGX agent

arXiv:2605.03671v1 Announce Type: new Abstract: The most commonly used metrics for evaluating automatic speech transcriptions, namely Word Error Rate (WER) and Character Error Rate (CER), have been he

ADAPTS: Agentic Decomposition for Automated Protocol-agnostic Tracking of Symptoms

SafetyDGX agent

arXiv:2605.03212v1 Announce Type: cross Abstract: Modeling latent clinical constructs from unconstrained clinical interactions is a unique challenge in affective computing. We present ADAPTS (Agentic

AfriqueLLM: How Data Mixing and Model Architecture Impact Continued Pre-training for African Languages

Model ReleasesDGX agent

arXiv:2601.06395v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly multilingual, yet open models continue to underperform relative to proprietary systems, with the gap m

AfriVox-v2: A Domain-Verticalized Benchmark for In-the-Wild African Speech Recognition

Model ReleasesDGX agent

arXiv:2605.03590v1 Announce Type: new Abstract: Recent large language models (LLMs) show strong speech recognition and translation capabilities for high-resource languages. However, African languages

Agentic-imodels: Evolving agentic interpretability tools via autoresearch

Model ReleasesDGX agent

arXiv:2605.03808v1 Announce Type: cross Abstract: Agentic data science (ADS) systems are rapidly improving their capability to autonomously analyze, fit, and interpret data, potentially moving towards

An ERP Study of Recursive Possessive Parsing in ASD Children and Its Cognitive Neuro Mechanisms

ResearchDGX agent

arXiv:2605.03447v1 Announce Type: new Abstract: Recursive structures are a core property of human language, yet little is known about how children with autism spectrum disorder (ASD) process complex r

Annotation Quality in Aspect-Based Sentiment Analysis: A Case Study Comparing Experts, Students, Crowdworkers, and Large Language Model

Model ReleasesDGX agent

arXiv:2605.03624v1 Announce Type: new Abstract: Aspect-Based Sentiment Analysis (ABSA) enables fine-grained opinion analysis by identifying sentiments toward specific aspects or targets within a text.

Atomic Fact-Checking Increases Clinician Trust in Large Language Model Recommendations for Oncology Decision Support: A Randomized Controlled Trial

ResearchDGX agent

arXiv:2605.03916v1 Announce Type: new Abstract: Question: Does atomic fact-checking, which decomposes AI treatment recommendations into individually verifiable claims linked to source guideline docume

AutoRAGTuner: A Declarative Framework for Automatic Optimization of RAG Pipelines

Model ReleasesDGX agent

arXiv:2605.02967v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) enhances LLMs, but performance is highly sensitive to complex architecture designs and hyper-parameter configurat

Benchmarking Local Language Models for Social Robots using Edge Devices

Local AiDGX agent

arXiv:2605.03111v1 Announce Type: cross Abstract: Social-educational robots designed for socially interactive pedagogical support, such as the Robot Study Companion (RSC), rely on responsive, privacy-

Benchmarking Logistic Regression, SVM, Naive Bayes, and IndoBERT Fine-Tuning for Sentiment Analysis on Indonesian Product Reviews

ResearchDGX agent

arXiv:2605.03439v1 Announce Type: new Abstract: The exponential growth of e-commerce platforms in Indonesia has generated a massive volume of user-generated product reviews. Analyzing the sentiment of

Benchmarking Parameter-Efficient Fine-Tuning of Large Language Models for Low-Resource Tajik Text Generation with the Tajik Web Corpus

Model ReleasesDGX agent

arXiv:2605.03742v1 Announce Type: new Abstract: This paper is devoted to the adaptation of generative large language models for the Tajik language, a low-resource language with Cyrillic script. To ove

BiMind: A Dual-Head Reasoning Model with Attention-Geometry Adapter for Incorrect Information Detection

ResearchDGX agent

arXiv:2604.06022v2 Announce Type: replace Abstract: Incorrect information poses significant challenges by disrupting content veracity and integrity, yet most detection approaches struggle to jointly b

BIT.UA-AAUBS at ArchEHR-QA 2026: Evaluating Open-Source and Proprietary LLMs via Prompting in Low-Resource QA

Local AiDGX agent

arXiv:2605.03618v1 Announce Type: new Abstract: This paper presents the joint participation of the BIT.UA and AAUBS groups in the ArchEHR-QA 2026 shared task, which focuses on clinical question answer

CC-OCR V2: Benchmarking Large Multimodal Models for Literacy in Real-world Document Processing

Model ReleasesDGX agent

arXiv:2605.03903v1 Announce Type: new Abstract: Large Multimodal Models (LMMs) have recently shown strong performance on Optical Character Recognition (OCR) tasks, demonstrating their promising capabi

Correct Is Not Enough: Training Reasoning Planners with Executor-Grounded Rewards

Local AiDGX agent

arXiv:2605.03862v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards has become a common way to improve explicit reasoning in large language models, but final-answer correc

CreativityBench: Evaluating Agent Creative Reasoning via Affordance-Based Tool Repurposing

Model ReleasesDGX agent

arXiv:2605.02910v1 Announce Type: cross Abstract: Recent advances in large language models have led to strong performance on reasoning and environment-interaction tasks, yet their ability for creative

CuraView: A Multi-Agent Framework for Medical Hallucination Detection with GraphRAG-Enhanced Knowledge Verification

Model ReleasesDGX agent

arXiv:2605.03476v1 Announce Type: new Abstract: Discharge summaries require extracting critical information from lengthy electronic health records (EHRs), a process that is labor-intensive when perfor

Detecting Stealth Sycophancy in Mental-Health Dialogue with Dynamic Emotional Signature Graphs

Model ReleasesDGX agent

arXiv:2605.03472v1 Announce Type: new Abstract: As conversational AI therapists are increasingly used in psychological support settings, reliable offline evaluation of therapeutic response quality rem

Direct Simultaneous Translation Activation for Large Audio-Language Models

TutorialsDGX agent

arXiv:2509.15692v2 Announce Type: replace-cross Abstract: Simultaneous speech-to-text translation (Simul-S2TT) aims to translate speech into target text in real time, outputting translations while rec

Do Not Waste Your Rollouts: Recycling Search Experience for Efficient Test-Time Scaling

ResearchDGX agent

arXiv:2601.21684v2 Announce Type: replace Abstract: Test-Time Scaling enhances the reasoning capabilities of Large Language Models by allocating additional inference compute to broaden the exploration

Effective Performance Measurement: Challenges and Opportunities in KPI Extraction from Earnings Calls

Model ReleasesDGX agent

arXiv:2605.03147v1 Announce Type: new Abstract: Earnings calls are a key source of financial information about public companies. However, extracting information from these calls is difficult. Unlike t

EQUITRIAGE: A Fairness Audit of Gender Bias in LLM-Based Emergency Department Triage

Model ReleasesDGX agent

arXiv:2605.03998v1 Announce Type: new Abstract: Emergency department triage assigns patients an acuity score that determines treatment priority, and clinical evidence documents persistent gender dispa

Evaluating Reasoning Models for Queries with Presuppositions

ResearchDGX agent

arXiv:2605.03050v1 Announce Type: new Abstract: Millions of users turn to AI models for their information needs. It is conceivable that a large number of user queries contain assumptions that may be f

ExaGPT: Example-Based Machine-Generated Text Detection for Human Interpretability

ResearchDGX agent

arXiv:2502.11336v2 Announce Type: replace Abstract: Detecting texts generated by Large Language Models (LLMs) could cause grave mistakes due to incorrect decisions, such as undermining students' acade

Exposing LLM Safety Gaps Through Mathematical Encoding:New Attacks and Systematic Analysis

Model ReleasesDGX agent

arXiv:2605.03441v1 Announce Type: cross Abstract: Large language models (LLMs) employ safety mechanisms to prevent harmful outputs, yet these defenses primarily rely on semantic pattern matching. We s

Feature-Augmented Transformers for Robust AI-Text Detection Across Domains and Generators

Model ReleasesDGX agent

arXiv:2605.03969v1 Announce Type: new Abstract: AI-generated text is nowadays produced at scale across domains and heterogeneous generation pipelines, making robustness to distribution shift a central

FINER-SQL: Boosting Small Language Models for Text-to-SQL

SafetyDGX agent

arXiv:2605.03465v1 Announce Type: cross Abstract: Large language models have driven major advances in Text-to-SQL generation. However, they suffer from high computational cost, long latency, and data

From prompting to evidence-based translation: A RAG+prompt system for Japanese-Chinese translation and its pedagogical potential

ResearchDGX agent

arXiv:2605.03387v1 Announce Type: new Abstract: Large language models perform well on high-resource pairs but are less reliable for Japanese-Chinese sentences containing noun-modifying clause construc

Geolocating News about Extreme Climate Events: A Comparative Analysis of Off-the-Shelf Tools for Toponym Identification in German

ResearchDGX agent

arXiv:2605.03414v1 Announce Type: new Abstract: Determining the geolocation of extreme climate events and disasters in texts is a common problem in climate impact and adaptation research. Named-entity

Geometric Deviation as an Unsupervised Pre-Generation Reliability Signal: Probing LLM Representations for Answerability

Model ReleasesDGX agent

arXiv:2605.03196v1 Announce Type: new Abstract: A reliable language model should be able to signal, prior to generation, when a query falls outside its knowledge. We investigate whether representation

GLEAN: Active Generalized Category Discovery with Diverse LLM Feedback

ApplicationsDGX agent

arXiv:2502.18414v2 Announce Type: replace Abstract: Generalized Category Discovery (GCD) is a practical and challenging open-world task that aims to recognize both known and novel categories in unlabe

Hierarchical Memorization in Large Language Models: Evidence from Citation Generation

Model ReleasesDGX agent

arXiv:2511.08877v2 Announce Type: replace Abstract: Large language models (LLMs) generate fluent text across a wide range of tasks, but the fabrication of non-existent academic citations remains a cri

How Language Models Process Negation

Model ReleasesDGX agent

arXiv:2605.03052v1 Announce Type: new Abstract: We study how Large Language Models (LLMs) process negation mechanistically. First, we establish that even though open-weight models often provide wrong

Hybrid Models for Natural Language Reasoning: The Case of Syllogistic Logic

Model ReleasesDGX agent

arXiv:2510.09472v2 Announce Type: replace Abstract: Despite the remarkable progress in neural models, their ability to generalize, a cornerstone for applications such as logical reasoning, remains a c

InvisibleInk: High-Utility and Low-Cost Text Generation with Differential Privacy

ResearchDGX agent

arXiv:2507.02974v3 Announce Type: replace-cross Abstract: As major progress in LLM-based long-form text generation enables paradigms such as retrieval-augmented generation (RAG) and inference-time sca

Kanade: A Simple Disentangled Tokenizer for Spoken Language Modeling

ResearchDGX agent

arXiv:2602.00594v2 Announce Type: replace Abstract: A good language model starts with a good tokenizer. Tokenization is especially important for speech modeling, which must handle continuous signals t

LitVISTA: A Benchmark for Narrative Orchestration in Literary Text

Model ReleasesDGX agent

arXiv:2601.06445v2 Announce Type: replace Abstract: Computational narrative analysis aims to capture rhythm, tension, and emotional dynamics in literary texts. Existing large language models can gener

← Previous
1…8485868788…129
Next →