AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
22 Apr 2026

Multilingual Language Models Encode Script Over Linguistic Structure

Model ReleasesDGX agent

arXiv:2604.05090v2 Announce Type: replace Abstract: Multilingual language models (LMs) organize representations for typologically and orthographically diverse languages into a shared parameter space,

OmniParser V2: Structured-Points-of-Thought for Unified Visual Text Parsing and Its Generality to Multimodal Large Language Models

ResearchDGX agent

arXiv:2502.16161v2 Announce Type: replace-cross Abstract: Visually-situated text parsing (VsTP) has recently seen notable advancements, driven by the growing demand for automated document understandin

OmniVoice: Towards Omnilingual Zero-Shot Text-to-Speech with Diffusion Language Models

ResearchDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.00688v3 Announce Type: replace Abstract: We present OmniVoice, a massively multilingual zero-shot text-to-speech (TTS) model that scales to over 600 languages. At its core is a novel diffus

On Temperature-Constrained Non-Deterministic Machine Translation: Potential and Evaluation

ApplicationsDGX agent

arXiv:2601.13729v2 Announce Type: replace Abstract: In recent years, the non-deterministic properties of language models have garnered considerable attention and have shown a significant influence on

Once Correct, Still Wrong: Counterfactual Hallucination in Multilingual Vision-Language Models

Model ReleasesDGX agent

arXiv:2602.05437v2 Announce Type: replace Abstract: Vision-language models (VLMs) can achieve high accuracy while still accepting culturally plausible but visually incorrect interpretations. Existing

One Persona, Many Cues, Different Results: How Sociodemographic Cues Impact LLM Personalization

SafetyDGX agent

arXiv:2601.18572v2 Announce Type: replace Abstract: Personalization of LLMs by sociodemographic subgroup often improves user experience, but can also introduce or amplify biases and unfair outcomes ac

Pause or Fabricate? Training Language Models for Grounded Reasoning

ResearchDGX agent

arXiv:2604.19656v1 Announce Type: new Abstract: Large language models have achieved remarkable progress on complex reasoning tasks. However, they often implicitly fabricate information when inputs are

Persuasion with Large Language Models: A Survey of Empirical Evidence, Study Methodologies, and Ethical Implications

SafetyDGX agent

arXiv:2411.06837v2 Announce Type: replace Abstract: The rapid rise of Large Language Models (LLMs) has created new disruptive possibilities for persuasive communication, enabling fully-automated, pers

PolarQuant: Optimal Gaussian Weight Quantization via Hadamard Rotation for LLM Compression

Local AiDGX agent

arXiv:2603.29078v2 Announce Type: replace Abstract: We present PolarQuant, a post-training weight quantization method for large language models (LLMs) that exploits the distributional structure of neu

Position: LLM Watermarking Should Align Stakeholders' Incentives for Practical Adoption

ApplicationsDGX agent

arXiv:2510.18333v2 Announce Type: replace-cross Abstract: Despite progress in watermarking algorithms for large language models (LLMs), real-world deployment remains limited. We argue that this gap st

Prioritizing the Best: Incentivizing Reliable Multimodal Reasoning by Rewarding Beyond Answer Correctness

SafetyDGX agent

arXiv:2604.18892v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) improves multimodal reasoning by rewarding verifiable final answers. Yet answer-correct trajectori

Probing for Reading Times

SafetyDGX agent

arXiv:2604.18712v1 Announce Type: new Abstract: Probing has shown that language model representations encode rich linguistic information, but it remains unclear whether they also capture cognitive sig

Proposing Topic Models and Evaluation Frameworks for Analyzing Associations with External Outcomes: An Application to Leadership Analysis Using Large-Scale Corporate Review Data

SafetyDGX agent

arXiv:2604.18919v1 Announce Type: new Abstract: Analyzing topics extracted from text data in relation to external outcomes is important across fields such as computational social science and organizat

Rank-Turbulence Delta and Interpretable Approaches to Stylometric Delta Metrics

Model ReleasesDGX agent

arXiv:2604.19499v1 Announce Type: new Abstract: This article introduces two new measures for authorship attribution - Rank-Turbulence Delta and Jensen-Shannon Delta - which generalise Burrows's classi

ReflectMT: Internalizing Reflection for Efficient and High-Quality Machine Translation

Model ReleasesDGX agent

arXiv:2604.19144v1 Announce Type: new Abstract: Recent years have witnessed growing interest in applying Large Reasoning Models (LRMs) to Machine Translation (MT). Existing approaches predominantly ad

Remask, Don't Replace: Token-to-Mask Refinement in Masked Diffusion Language Models

ResearchDGX agent

arXiv:2604.18738v1 Announce Type: new Abstract: Masked diffusion language models such as LLaDA2.1 rely on Token-to-Token (T2T) editing to correct their own generation errors: whenever a different toke

Rethinking Information Synthesis in Multimodal Question Answering A Multi-Agent Perspective

AgentsDGX agent

arXiv:2505.20816v2 Announce Type: replace Abstract: Recent advances in multimodal question answering have primarily focused on combining heterogeneous modalities or fine-tuning multimodal large langua

Scripts Through Time: A Survey of the Evolving Role of Transliteration in NLP

ResearchDGX agent

arXiv:2604.18722v1 Announce Type: new Abstract: Cross-lingual transfer in NLP is often hindered by the ``script barrier'' where differences in writing systems inhibit transfer learning between languag

SitEmb-v1.5: Improved Context-Aware Dense Retrieval for Semantic Association and Long Story Comprehension

Model ReleasesDGX agent

arXiv:2508.01959v2 Announce Type: replace Abstract: Retrieval-augmented generation (RAG) over long documents typically involves splitting the text into smaller chunks, which serve as the basic units f

Stable-RAG: Mitigating Retrieval-Permutation-Induced Hallucinations in Retrieval-Augmented Generation

SafetyDGX agent

arXiv:2601.02993v4 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) has become a key paradigm for reducing factual hallucinations in Large Language Models (LLMs), yet little is kn

STAR-Teaming: A Strategy-Response Multiplex Network Approach to Automated LLM Red Teaming

SafetyDGX agent

arXiv:2604.18976v1 Announce Type: new Abstract: While Large Language Models (LLMs) are widely used, they remain susceptible to jailbreak prompts that can elicit harmful or inappropriate responses. Thi

StochasTok: Improving Fine-Grained Subword Understanding in LLMs

ResearchDGX agent

arXiv:2506.01687v3 Announce Type: replace Abstract: Subword-level understanding is integral to numerous tasks, including understanding multi-digit numbers, spelling mistakes, abbreviations, rhyming, a

STReasoner: Empowering LLMs for Spatio-Temporal Reasoning in Time Series via Spatial-Aware Reinforcement Learning

Model ReleasesDGX agent

arXiv:2601.03248v2 Announce Type: replace Abstract: Spatio-temporal reasoning in time series involves the explicit synthesis of temporal dynamics, spatial dependencies, and textual context. This capab

Superficial Success vs. Internal Breakdown: An Empirical Study of Generalization in Adaptive Multi-Agent Systems

AgentsDGX agent

arXiv:2604.18951v1 Announce Type: cross Abstract: Adaptive multi-agent systems (MAS) are increasingly adopted to tackle complex problems.However, the narrow task coverage of their optimization raises

Syntax as a Rosetta Stone: Universal Dependencies for In-Context Coptic Translation

ResearchDGX agent

arXiv:2604.18758v1 Announce Type: new Abstract: Low-resource machine translation requires methods that differ from those used for high-resource languages. This paper proposes a novel in-context learni

TabReX : Tabular Referenceless eXplainable Evaluation

Model ReleasesDGX agent

arXiv:2512.15907v2 Announce Type: replace Abstract: Evaluating the quality of tables generated by large language models (LLMs) remains an open challenge: existing metrics either flatten tables into te

TabXEval: Why this is a Bad Table? An eXhaustive Rubric for Table Evaluation

Model ReleasesDGX agent

arXiv:2505.22176v3 Announce Type: replace Abstract: Evaluating tables qualitatively and quantitatively poses a significant challenge, as standard metrics often overlook subtle structural and content-l

Take Out Your Calculators: Estimating the Real Difficulty of Question Items with LLM Student Simulations

Model ReleasesDGX agent

arXiv:2601.09953v2 Announce Type: replace Abstract: Standardized math assessments require expensive human pilot studies to establish the difficulty of test items. We investigate the predictive value o

Temporal Leakage in Search-Engine Date-Filtered Web Retrieval: A Retrospective Forecasting Case Study

ApplicationsDGX agent

arXiv:2602.00758v2 Announce Type: replace Abstract: Search-engine date filters are widely used to enforce pre-cutoff retrieval in retrospective evaluations of search-augmented forecasters. We show thi

The Alignment Waltz: Jointly Training Agents to Collaborate for Safety

SafetyDGX agent

arXiv:2510.08240v2 Announce Type: replace Abstract: Harnessing the power of LLMs requires a delicate dance between being helpful and harmless. This creates a fundamental tension between two competing

'The Order in the Horse's Heart': A Case Study in LLM-Assisted Stylometry for the Discovery of Biblical Allusion in Modern Literary Fiction

Local AiDGX agent

arXiv:2604.19447v1 Announce Type: new Abstract: We present a dual-track pipeline for detecting biblical allusions in literary fiction and apply it to the novels of Cormac McCarthy. A bottom-up embeddi

The signal is the ceiling: Measurement limits of LLM-predicted experience ratings from open-ended survey text

SafetyDGX agent

arXiv:2604.19645v1 Announce Type: new Abstract: An earlier paper (Hong, Potteiger, and Zapata 2026) established that an unoptimized GPT 4.1 prompt predicts fan-reported experience ratings within one p

The 'Small World of Words' German Free-Association Norms

ResearchDGX agent

arXiv:2604.19620v1 Announce Type: new Abstract: Free-association norms provide essential empirical data for investigating linguistic, semantic, and cultural phenomena in the cognitive sciences. Althou

Towards a Linguistic Evaluation of Narratives: A Quantitative Stylistic Framework

ResearchDGX agent

arXiv:2604.19261v1 Announce Type: new Abstract: The evaluation of narrative quality remains a complex challenge, as it involves subjective factors such as plot, character development, and emotional im

TRN-R1-Zero: Text-rich Network Reasoning via LLMs with Reinforcement Learning Only

SafetyDGX agent

arXiv:2604.19070v1 Announce Type: new Abstract: Zero-shot reasoning on text-rich networks (TRNs) remains a challenging frontier, as models must integrate textual semantics with relational structure wi

VCE: A zero-cost hallucination mitigation method of LVLMs via visual contrastive editing

Model ReleasesDGX agent

arXiv:2604.19412v1 Announce Type: cross Abstract: Large vision-language models (LVLMs) frequently suffer from Object Hallucination (OH), wherein they generate descriptions containing objects that are

VIGIL: An Extensible System for Real-Time Detection and Mitigation of Cognitive Bias Triggers

SafetyDGX agent

arXiv:2604.03261v2 Announce Type: replace Abstract: The rise of generative AI is posing increasing risks to online information integrity and civic discourse. Most concretely, such risks can materialis

VimRAG: Navigating Massive Visual Context in Retrieval-Augmented Generation via Multimodal Memory Graph

SafetyDGX agent

arXiv:2602.12735v2 Announce Type: replace-cross Abstract: Effectively retrieving, reasoning, and understanding multimodal information remains a critical challenge for agentic systems. Traditional Retr

VISTA: Verification In Sequential Turn-based Assessment

ResearchDGX agent

arXiv:2510.27052v5 Announce Type: replace Abstract: Hallucination--defined here as generating statements unsupported or contradicted by available evidence or conversational context--remains a major ob

Visual-TableQA: Open-Domain Benchmark for Reasoning over Table Images

Model ReleasesDGX agent

arXiv:2509.07966v2 Announce Type: replace-cross Abstract: Visual reasoning over structured data such as tables is a critical capability for modern vision-language models (VLMs), yet current benchmarks

Voice of India: A Large-Scale Benchmark for Real-World Speech Recognition in India

Model ReleasesDGX agent

arXiv:2604.19151v1 Announce Type: new Abstract: Existing Indic ASR benchmarks often use scripted, clean speech and leaderboard driven evaluation that encourages dataset specific overfitting. In additi

Watch the Weights: Unsupervised monitoring and control of fine-tuned LLMs

Model ReleasesDGX agent

arXiv:2508.00161v3 Announce Type: replace-cross Abstract: The releases of powerful open-weight large language models (LLMs) are often not accompanied by access to their full training data. Existing in

What Makes an LLM a Good Optimizer? A Trajectory Analysis of LLM-Guided Evolutionary Search

Local AiDGX agent

arXiv:2604.19440v1 Announce Type: new Abstract: Recent work has demonstrated the promise of orchestrating large language models (LLMs) within evolutionary and agentic optimization systems. However, th

When and What to Ask: AskBench and Rubric-Guided RLVR for LLM Clarification

Model ReleasesDGX agent

arXiv:2602.11199v2 Announce Type: replace Abstract: Large language models (LLMs) often respond even when prompts omit critical details or include misleading information, leading to hallucinations or r

When Does Verification Pay Off? A Closer Look at LLMs as Solution Verifiers

ResearchDGX agent

arXiv:2512.02304v2 Announce Type: replace Abstract: Large language models (LLMs) can act as both problem solvers and solution verifiers, where the latter select high-quality answers from a pool of sol

When Safety Fails Before the Answer: Benchmarking Harmful Behavior Detection in Reasoning Chains

Model ReleasesDGX agent

arXiv:2604.19001v1 Announce Type: new Abstract: Large reasoning models (LRMs) produce complex, multi-step reasoning traces, yet safety evaluation remains focused on final outputs, overlooking how harm

Why Does Self-Distillation (Sometimes) Degrade the Reasoning Capability of LLMs?

Model ReleasesDGX agent

arXiv:2603.24472v2 Announce Type: replace Abstract: Self-distillation has emerged as an effective post-training paradigm for LLMs, often improving performance while shortening reasoning traces. Howeve

21 Apr 2026

A Community-Based Approach for Stance Distribution and Argument Organization

ResearchDGX agent

arXiv:2604.16852v1 Announce Type: new Abstract: The proliferation of online debate platforms and social media has led to an unprecedented volume of argumentative content on controversial topics from m

A Computational Method for Measuring 'Open Codes' in Qualitative Analysis

ResearchDGX agent

arXiv:2411.12142v4 Announce Type: replace Abstract: Qualitative analysis is critical to understanding human datasets in many social science disciplines. A central method in this process is inductive c

A Goal Without a Plan Is Just a Wish: Efficient and Effective Global Planner Training for Long-Horizon Agent Tasks

AgentsDGX agent

arXiv:2510.05608v2 Announce Type: replace Abstract: Agents based on large language models (LLMs) struggle with brainless trial-and-error and generating hallucinatory actions due to a lack of global pl

A Multi-Agent Approach for Claim Verification from Tabular Data Documents

AgentsDGX agent

arXiv:2604.17225v1 Announce Type: new Abstract: We present a novel approach for claim verification from tabular data documents. Recent LLM-based approaches either employ complex pretraining/fine-tunin

A multimodal and temporal foundation model for virtual patient representations at healthcare system scale

ApplicationsDGX agent

arXiv:2604.18570v1 Announce Type: cross Abstract: Modern medicine generates vast multimodal data across siloed systems, yet no existing model integrates the full breadth and temporal depth of the clin

A novel LSTM music generator based on the fractional time-frequency feature extraction

ResearchDGX agent

arXiv:2604.17823v1 Announce Type: cross Abstract: In this paper, we propose a novel approach for generating music based on an artificial intelligence (AI) system. We analyze the features of music and

A Survey on the Security of Long-Term Memory in LLM Agents: Toward Mnemonic Sovereignty

AgentsDGX agent

arXiv:2604.16548v1 Announce Type: cross Abstract: Research on large language model (LLM) security is shifting from 'will the model leak training data' to a more consequential question: can an agent wi

A Text-To-Text Alignment Algorithm for Better Evaluation of Modern Speech Recognition Systems

SafetyDGX agent

arXiv:2509.24478v2 Announce Type: replace Abstract: Modern neural networks have greatly improved performance across speech recognition benchmarks. However, gains are often driven by frequent words wit

A Transformer and Prototype-based Interpretable Model for Contextual Sarcasm Detection

Model ReleasesDGX agent

arXiv:2503.11838v2 Announce Type: replace Abstract: Sarcasm detection, with its figurative nature, poses unique challenges for affective systems designed to perform sentiment analysis. While these sys

A Universal Avoidance Method for Diverse Multi-branch Generation

ResearchDGX agent

arXiv:2604.17323v1 Announce Type: new Abstract: Modern generative models still lack human-level creativity, particularly in multi-branch diversity. Prior approaches to address this problem often incur

Abstain-R1: Calibrated Abstention and Post-Refusal Clarification via Verifiable RL

Model ReleasesDGX agent

arXiv:2604.17073v1 Announce Type: new Abstract: Reinforcement fine-tuning improves the reasoning ability of large language models, but it can also encourage them to answer unanswerable queries by gues

AdaExplore: Failure-Driven Adaptation and Diversity-Preserving Search for Efficient Kernel Generation

AgentsDGX agent

arXiv:2604.16625v1 Announce Type: new Abstract: Recent large language model (LLM) agents have shown promise in using execution feedback for test-time adaptation. However, robust self-improvement remai

Adaptive Text Anonymization: Learning Privacy-Utility Trade-offs via Prompt Optimization

Model ReleasesDGX agent

arXiv:2602.20743v2 Announce Type: replace Abstract: Anonymizing textual documents is a highly context-sensitive problem: the appropriate balance between privacy protection and utility preservation var

← Previous
1…103104105106107…129
Next →