AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
Model Releases

Multilingual Language Models Encode Script Over Linguistic Structure

DGX agent

arXiv:2604.05090v2 Announce Type: replace Abstract: Multilingual language models (LMs) organize representations for typologically and orthographically diverse languages into a shared parameter space,

model-releasesarxiv-cs-cl
22 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Research

OmniParser V2: Structured-Points-of-Thought for Unified Visual Text Parsing and Its Generality to Multimodal Large Language Models

DGX agent

arXiv:2502.16161v2 Announce Type: replace-cross Abstract: Visually-situated text parsing (VsTP) has recently seen notable advancements, driven by the growing demand for automated document understandin

researcharxiv-cs-cl
22 Apr 2026
Research

OmniVoice: Towards Omnilingual Zero-Shot Text-to-Speech with Diffusion Language Models

DGX agent

arXiv:2604.00688v3 Announce Type: replace Abstract: We present OmniVoice, a massively multilingual zero-shot text-to-speech (TTS) model that scales to over 600 languages. At its core is a novel diffus

researcharxiv-cs-cl
22 Apr 2026
Applications

On Temperature-Constrained Non-Deterministic Machine Translation: Potential and Evaluation

DGX agent

arXiv:2601.13729v2 Announce Type: replace Abstract: In recent years, the non-deterministic properties of language models have garnered considerable attention and have shown a significant influence on

applicationsarxiv-cs-cl
22 Apr 2026
Model Releases

Once Correct, Still Wrong: Counterfactual Hallucination in Multilingual Vision-Language Models

DGX agent

arXiv:2602.05437v2 Announce Type: replace Abstract: Vision-language models (VLMs) can achieve high accuracy while still accepting culturally plausible but visually incorrect interpretations. Existing

model-releasesarxiv-cs-cl
22 Apr 2026
Safety

One Persona, Many Cues, Different Results: How Sociodemographic Cues Impact LLM Personalization

DGX agent

arXiv:2601.18572v2 Announce Type: replace Abstract: Personalization of LLMs by sociodemographic subgroup often improves user experience, but can also introduce or amplify biases and unfair outcomes ac

safetyarxiv-cs-cl
22 Apr 2026
Research

Pause or Fabricate? Training Language Models for Grounded Reasoning

DGX agent

arXiv:2604.19656v1 Announce Type: new Abstract: Large language models have achieved remarkable progress on complex reasoning tasks. However, they often implicitly fabricate information when inputs are

researcharxiv-cs-cl
22 Apr 2026
Safety

Persuasion with Large Language Models: A Survey of Empirical Evidence, Study Methodologies, and Ethical Implications

DGX agent

arXiv:2411.06837v2 Announce Type: replace Abstract: The rapid rise of Large Language Models (LLMs) has created new disruptive possibilities for persuasive communication, enabling fully-automated, pers

safetyarxiv-cs-cl
22 Apr 2026
Local Ai

PolarQuant: Optimal Gaussian Weight Quantization via Hadamard Rotation for LLM Compression

DGX agent

arXiv:2603.29078v2 Announce Type: replace Abstract: We present PolarQuant, a post-training weight quantization method for large language models (LLMs) that exploits the distributional structure of neu

local-aiarxiv-cs-cl
22 Apr 2026
Applications

Position: LLM Watermarking Should Align Stakeholders' Incentives for Practical Adoption

DGX agent

arXiv:2510.18333v2 Announce Type: replace-cross Abstract: Despite progress in watermarking algorithms for large language models (LLMs), real-world deployment remains limited. We argue that this gap st

applicationsarxiv-cs-cl
22 Apr 2026
Safety

Prioritizing the Best: Incentivizing Reliable Multimodal Reasoning by Rewarding Beyond Answer Correctness

DGX agent

arXiv:2604.18892v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) improves multimodal reasoning by rewarding verifiable final answers. Yet answer-correct trajectori

safetyarxiv-cs-cl
22 Apr 2026
Safety

Probing for Reading Times

DGX agent

arXiv:2604.18712v1 Announce Type: new Abstract: Probing has shown that language model representations encode rich linguistic information, but it remains unclear whether they also capture cognitive sig

safetyarxiv-cs-cl
22 Apr 2026
Safety

Proposing Topic Models and Evaluation Frameworks for Analyzing Associations with External Outcomes: An Application to Leadership Analysis Using Large-Scale Corporate Review Data

DGX agent

arXiv:2604.18919v1 Announce Type: new Abstract: Analyzing topics extracted from text data in relation to external outcomes is important across fields such as computational social science and organizat

safetyarxiv-cs-cl
22 Apr 2026
Model Releases

Rank-Turbulence Delta and Interpretable Approaches to Stylometric Delta Metrics

DGX agent

arXiv:2604.19499v1 Announce Type: new Abstract: This article introduces two new measures for authorship attribution - Rank-Turbulence Delta and Jensen-Shannon Delta - which generalise Burrows's classi

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

ReflectMT: Internalizing Reflection for Efficient and High-Quality Machine Translation

DGX agent

arXiv:2604.19144v1 Announce Type: new Abstract: Recent years have witnessed growing interest in applying Large Reasoning Models (LRMs) to Machine Translation (MT). Existing approaches predominantly ad

model-releasesarxiv-cs-cl
22 Apr 2026
Research

Remask, Don't Replace: Token-to-Mask Refinement in Masked Diffusion Language Models

DGX agent

arXiv:2604.18738v1 Announce Type: new Abstract: Masked diffusion language models such as LLaDA2.1 rely on Token-to-Token (T2T) editing to correct their own generation errors: whenever a different toke

researcharxiv-cs-cl
22 Apr 2026
Agents

Rethinking Information Synthesis in Multimodal Question Answering A Multi-Agent Perspective

DGX agent

arXiv:2505.20816v2 Announce Type: replace Abstract: Recent advances in multimodal question answering have primarily focused on combining heterogeneous modalities or fine-tuning multimodal large langua

agentsarxiv-cs-cl
22 Apr 2026
Research

Scripts Through Time: A Survey of the Evolving Role of Transliteration in NLP

DGX agent

arXiv:2604.18722v1 Announce Type: new Abstract: Cross-lingual transfer in NLP is often hindered by the ``script barrier'' where differences in writing systems inhibit transfer learning between languag

researcharxiv-cs-cl
22 Apr 2026
Model Releases

SitEmb-v1.5: Improved Context-Aware Dense Retrieval for Semantic Association and Long Story Comprehension

DGX agent

arXiv:2508.01959v2 Announce Type: replace Abstract: Retrieval-augmented generation (RAG) over long documents typically involves splitting the text into smaller chunks, which serve as the basic units f

model-releasesarxiv-cs-cl
22 Apr 2026
Safety

Stable-RAG: Mitigating Retrieval-Permutation-Induced Hallucinations in Retrieval-Augmented Generation

DGX agent

arXiv:2601.02993v4 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) has become a key paradigm for reducing factual hallucinations in Large Language Models (LLMs), yet little is kn

safetyarxiv-cs-cl
22 Apr 2026
Safety

STAR-Teaming: A Strategy-Response Multiplex Network Approach to Automated LLM Red Teaming

DGX agent

arXiv:2604.18976v1 Announce Type: new Abstract: While Large Language Models (LLMs) are widely used, they remain susceptible to jailbreak prompts that can elicit harmful or inappropriate responses. Thi

safetyarxiv-cs-cl
22 Apr 2026
Research

StochasTok: Improving Fine-Grained Subword Understanding in LLMs

DGX agent

arXiv:2506.01687v3 Announce Type: replace Abstract: Subword-level understanding is integral to numerous tasks, including understanding multi-digit numbers, spelling mistakes, abbreviations, rhyming, a

researcharxiv-cs-cl
22 Apr 2026
Model Releases

STReasoner: Empowering LLMs for Spatio-Temporal Reasoning in Time Series via Spatial-Aware Reinforcement Learning

DGX agent

arXiv:2601.03248v2 Announce Type: replace Abstract: Spatio-temporal reasoning in time series involves the explicit synthesis of temporal dynamics, spatial dependencies, and textual context. This capab

model-releasesarxiv-cs-cl
22 Apr 2026
Agents

Superficial Success vs. Internal Breakdown: An Empirical Study of Generalization in Adaptive Multi-Agent Systems

DGX agent

arXiv:2604.18951v1 Announce Type: cross Abstract: Adaptive multi-agent systems (MAS) are increasingly adopted to tackle complex problems.However, the narrow task coverage of their optimization raises

agentsarxiv-cs-cl
22 Apr 2026
Research

Syntax as a Rosetta Stone: Universal Dependencies for In-Context Coptic Translation

DGX agent

arXiv:2604.18758v1 Announce Type: new Abstract: Low-resource machine translation requires methods that differ from those used for high-resource languages. This paper proposes a novel in-context learni

researcharxiv-cs-cl
22 Apr 2026
Model Releases

TabReX : Tabular Referenceless eXplainable Evaluation

DGX agent

arXiv:2512.15907v2 Announce Type: replace Abstract: Evaluating the quality of tables generated by large language models (LLMs) remains an open challenge: existing metrics either flatten tables into te

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

TabXEval: Why this is a Bad Table? An eXhaustive Rubric for Table Evaluation

DGX agent

arXiv:2505.22176v3 Announce Type: replace Abstract: Evaluating tables qualitatively and quantitatively poses a significant challenge, as standard metrics often overlook subtle structural and content-l

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

Take Out Your Calculators: Estimating the Real Difficulty of Question Items with LLM Student Simulations

DGX agent

arXiv:2601.09953v2 Announce Type: replace Abstract: Standardized math assessments require expensive human pilot studies to establish the difficulty of test items. We investigate the predictive value o

model-releasesarxiv-cs-cl
22 Apr 2026
Applications

Temporal Leakage in Search-Engine Date-Filtered Web Retrieval: A Retrospective Forecasting Case Study

DGX agent

arXiv:2602.00758v2 Announce Type: replace Abstract: Search-engine date filters are widely used to enforce pre-cutoff retrieval in retrospective evaluations of search-augmented forecasters. We show thi

applicationsarxiv-cs-cl
22 Apr 2026
Safety

The Alignment Waltz: Jointly Training Agents to Collaborate for Safety

DGX agent

arXiv:2510.08240v2 Announce Type: replace Abstract: Harnessing the power of LLMs requires a delicate dance between being helpful and harmless. This creates a fundamental tension between two competing

safetyarxiv-cs-cl
22 Apr 2026
Local Ai

'The Order in the Horse's Heart': A Case Study in LLM-Assisted Stylometry for the Discovery of Biblical Allusion in Modern Literary Fiction

DGX agent

arXiv:2604.19447v1 Announce Type: new Abstract: We present a dual-track pipeline for detecting biblical allusions in literary fiction and apply it to the novels of Cormac McCarthy. A bottom-up embeddi

local-aiarxiv-cs-cl
22 Apr 2026
Safety

The signal is the ceiling: Measurement limits of LLM-predicted experience ratings from open-ended survey text

DGX agent

arXiv:2604.19645v1 Announce Type: new Abstract: An earlier paper (Hong, Potteiger, and Zapata 2026) established that an unoptimized GPT 4.1 prompt predicts fan-reported experience ratings within one p

safetyarxiv-cs-cl
22 Apr 2026
Research

The 'Small World of Words' German Free-Association Norms

DGX agent

arXiv:2604.19620v1 Announce Type: new Abstract: Free-association norms provide essential empirical data for investigating linguistic, semantic, and cultural phenomena in the cognitive sciences. Althou

researcharxiv-cs-cl
22 Apr 2026
Research

Towards a Linguistic Evaluation of Narratives: A Quantitative Stylistic Framework

DGX agent

arXiv:2604.19261v1 Announce Type: new Abstract: The evaluation of narrative quality remains a complex challenge, as it involves subjective factors such as plot, character development, and emotional im

researcharxiv-cs-cl
22 Apr 2026
Safety

TRN-R1-Zero: Text-rich Network Reasoning via LLMs with Reinforcement Learning Only

DGX agent

arXiv:2604.19070v1 Announce Type: new Abstract: Zero-shot reasoning on text-rich networks (TRNs) remains a challenging frontier, as models must integrate textual semantics with relational structure wi

safetyarxiv-cs-cl
22 Apr 2026
Model Releases

VCE: A zero-cost hallucination mitigation method of LVLMs via visual contrastive editing

DGX agent

arXiv:2604.19412v1 Announce Type: cross Abstract: Large vision-language models (LVLMs) frequently suffer from Object Hallucination (OH), wherein they generate descriptions containing objects that are

model-releasesarxiv-cs-cl
22 Apr 2026
Safety

VIGIL: An Extensible System for Real-Time Detection and Mitigation of Cognitive Bias Triggers

DGX agent

arXiv:2604.03261v2 Announce Type: replace Abstract: The rise of generative AI is posing increasing risks to online information integrity and civic discourse. Most concretely, such risks can materialis

safetyarxiv-cs-cl
22 Apr 2026
Safety

VimRAG: Navigating Massive Visual Context in Retrieval-Augmented Generation via Multimodal Memory Graph

DGX agent

arXiv:2602.12735v2 Announce Type: replace-cross Abstract: Effectively retrieving, reasoning, and understanding multimodal information remains a critical challenge for agentic systems. Traditional Retr

safetyarxiv-cs-cl
22 Apr 2026
Research

VISTA: Verification In Sequential Turn-based Assessment

DGX agent

arXiv:2510.27052v5 Announce Type: replace Abstract: Hallucination--defined here as generating statements unsupported or contradicted by available evidence or conversational context--remains a major ob

researcharxiv-cs-cl
22 Apr 2026
Model Releases

Visual-TableQA: Open-Domain Benchmark for Reasoning over Table Images

DGX agent

arXiv:2509.07966v2 Announce Type: replace-cross Abstract: Visual reasoning over structured data such as tables is a critical capability for modern vision-language models (VLMs), yet current benchmarks

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

Voice of India: A Large-Scale Benchmark for Real-World Speech Recognition in India

DGX agent

arXiv:2604.19151v1 Announce Type: new Abstract: Existing Indic ASR benchmarks often use scripted, clean speech and leaderboard driven evaluation that encourages dataset specific overfitting. In additi

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

Watch the Weights: Unsupervised monitoring and control of fine-tuned LLMs

DGX agent

arXiv:2508.00161v3 Announce Type: replace-cross Abstract: The releases of powerful open-weight large language models (LLMs) are often not accompanied by access to their full training data. Existing in

model-releasesarxiv-cs-cl
22 Apr 2026
Local Ai

What Makes an LLM a Good Optimizer? A Trajectory Analysis of LLM-Guided Evolutionary Search

DGX agent

arXiv:2604.19440v1 Announce Type: new Abstract: Recent work has demonstrated the promise of orchestrating large language models (LLMs) within evolutionary and agentic optimization systems. However, th

local-aiarxiv-cs-cl
22 Apr 2026
Model Releases

When and What to Ask: AskBench and Rubric-Guided RLVR for LLM Clarification

DGX agent

arXiv:2602.11199v2 Announce Type: replace Abstract: Large language models (LLMs) often respond even when prompts omit critical details or include misleading information, leading to hallucinations or r

model-releasesarxiv-cs-cl
22 Apr 2026
Research

When Does Verification Pay Off? A Closer Look at LLMs as Solution Verifiers

DGX agent

arXiv:2512.02304v2 Announce Type: replace Abstract: Large language models (LLMs) can act as both problem solvers and solution verifiers, where the latter select high-quality answers from a pool of sol

researcharxiv-cs-cl
22 Apr 2026
Model Releases

When Safety Fails Before the Answer: Benchmarking Harmful Behavior Detection in Reasoning Chains

DGX agent

arXiv:2604.19001v1 Announce Type: new Abstract: Large reasoning models (LRMs) produce complex, multi-step reasoning traces, yet safety evaluation remains focused on final outputs, overlooking how harm

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

Why Does Self-Distillation (Sometimes) Degrade the Reasoning Capability of LLMs?

DGX agent

arXiv:2603.24472v2 Announce Type: replace Abstract: Self-distillation has emerged as an effective post-training paradigm for LLMs, often improving performance while shortening reasoning traces. Howeve

model-releasesarxiv-cs-cl
22 Apr 2026
Research

A Community-Based Approach for Stance Distribution and Argument Organization

DGX agent

arXiv:2604.16852v1 Announce Type: new Abstract: The proliferation of online debate platforms and social media has led to an unprecedented volume of argumentative content on controversial topics from m

researcharxiv-cs-cl
21 Apr 2026
← Previous
1…129130131132133…161
Next →