AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
Model Releases

Omni-RRM: Advancing Omni Reward Modeling via Automatic Rubric-Grounded Preference Synthesis

DGX agent

arXiv:2602.00846v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) struggle with alignment due to the limitations of existing reward models (RMs), which are predominantly vis

model-releasesarxiv-cs-cl
8 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

On the feasibility of dependency parsing of non-human sequences without a gold standard. Is evaluation possible in other species?

DGX agent

arXiv:2607.06542v1 Announce Type: new Abstract: Dependency parsing consists of finding a tree representation for a sequence. Unsupervised dependency parsing aims to develop parsing methods without a g

researcharxiv-cs-cl
8 Jul 2026
Model Releases

Pluralis v0.1: Towards a Multicultural, Multimodal, Multilingual Benchmark for AI Risk and Reliability

DGX agent

arXiv:2607.06196v1 Announce Type: new Abstract: Current AI safety evaluation and benchmarking frameworks predominantly rely on Western-centric culture-agnostic defaults that mask critical regional law

model-releasesarxiv-cs-cl
8 Jul 2026
Agents

PolyJarvis: An LLM-Orchestrated Agent for Automated All-Atom Molecular Dynamics of Amorphous Homopolymers

DGX agent

arXiv:2604.02537v2 Announce Type: replace Abstract: All-atom molecular dynamics (MD) simulations can predict polymer properties from molecular structure, yet their execution requires specialized exper

agentsarxiv-cs-cl
8 Jul 2026
Model Releases

Population-Level Profiling of DSM-5 Depressive Symptoms Among Self-Reported ADHD and ASD Users on Twitter: An Exploratory Study Using Advanced NLP and Statistical Analysis

DGX agent

arXiv:2607.05626v1 Announce Type: new Abstract: Background: Depression frequently co-occurs with ADHD and autism spectrum disorder (ASD), but population-level differences in symptom expression between

model-releasesarxiv-cs-cl
8 Jul 2026
Research

Prompting Complexity: Shortest Prompts for Texts and Behaviors in LLMs

DGX agent

arXiv:2607.06145v1 Announce Type: new Abstract: In this paper, we define the quantity of prompting complexity: for a fixed instruction-tuned language model, what is the shortest plausible prompt that

researcharxiv-cs-cl
8 Jul 2026
Safety

Quantifying Retriever-Generator Alignment in RAG with Local Explanations

DGX agent

arXiv:2601.21803v2 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) systems combine dense retrievers and language models to ground their outputs in external documents. However, th

safetyarxiv-cs-cl
8 Jul 2026
Research

Revisiting the Relation Between Language Model Perplexity and ASR Word Error Rate for Modern End-to-End Speech Recognition

DGX agent

arXiv:2607.05612v1 Announce Type: new Abstract: Language model (LM) perplexity (PPL) has historically been used as a proxy for automatic speech recognition (ASR) word error rate (WER), with prior work

researcharxiv-cs-cl
8 Jul 2026
Model Releases

SpanUQ: Span-Level Uncertainty Quantification for Large Language Model Generation

DGX agent

arXiv:2607.05721v1 Announce Type: new Abstract: Uncertainty estimation is essential not only for the trustworthy deployment of large language models (LLMs) but also as a foundation for self-refinement

model-releasesarxiv-cs-cl
8 Jul 2026
Research

Text Distance from Nested and Hierarchical Repetitions: A Compression-Based Perspective

DGX agent

arXiv:2607.05416v1 Announce Type: new Abstract: We present a new method for structural sequence analysis grounded in Algorithmic Information Theory (AIT). At its core is the Ladderpath approach, which

researcharxiv-cs-cl
8 Jul 2026
Research

Transferring Natural Language Datasets Between Languages Using Large Language Models for Modern Decision Support and Sci-Tech Analytical Systems

DGX agent

arXiv:2410.14074v2 Announce Type: replace Abstract: The decision-making process to rule R&D relies on information related to current trends in particular research areas. In this work, we investigated

researcharxiv-cs-cl
8 Jul 2026
Safety

Truthful or Fabricated? Using Causal Attribution to Mitigate Reward Hacking in Explanations

DGX agent

arXiv:2504.05294v3 Announce Type: replace Abstract: Chain-of-thought explanations are widely used to inspect the decision process of large language models (LLMs) and to evaluate the trustworthiness of

safetyarxiv-cs-cl
8 Jul 2026
Research

UCSC NLP at SemEval-2026 Task 10: Boundary-Aware Span Extraction and RoBERTa Classification for Conspiracy Detection

DGX agent

arXiv:2607.05689v1 Announce Type: new Abstract: We present our systems for SemEval-2026 Task 10 (PsyCoMark), addressing conspiracy marker extraction (Subtask 1) and document-level conspiracy detection

researcharxiv-cs-cl
8 Jul 2026
Research

Umm... With Transformers? Insights from Filled Pause Use across Four Slavic Parliaments

DGX agent

arXiv:2607.05964v1 Announce Type: new Abstract: Filled pauses (FPs) are a universal feature of spontaneous speech, yet most studies rely on small, single-language corpora, limiting the generalisabilit

researcharxiv-cs-cl
8 Jul 2026
Agents

When Does Tool Use Increase the Expressive Power of Finite-Precision Recurrent Models?

DGX agent

arXiv:2607.06155v1 Announce Type: cross Abstract: Modern sequence models are increasingly deployed as agents that interleave token generation with calls to external tools. We give an exact, architectu

agentsarxiv-cs-cl
8 Jul 2026
Research

Where to cut, how deep: BPE and Unigram-LM on chemistry SMILES

DGX agent

arXiv:2607.05691v1 Announce Type: new Abstract: Every chemical language model reading SMILES begins with a tokenizer, yet the field has inherited byte-pair encoding (BPE) from natural language with li

researcharxiv-cs-cl
8 Jul 2026
Safety

WordVoice: Explicit and Decoupled Multi-Dimensional Word-Level Control for LLM-Based TTS

DGX agent

arXiv:2607.06461v1 Announce Type: cross Abstract: While recent Large Language Model (LLM)-based Text-to-Speech (TTS) systems have achieved remarkable naturalness, they predominantly rely on implicit e

safetyarxiv-cs-cl
8 Jul 2026
Model Releases

AI Wizards at EXIST 2026: Hierarchical Soft-Label Learning for Multimodal Sexism Identification in Memes

DGX agent

arXiv:2607.04410v1 Announce Type: new Abstract: We present the AI Wizards submission to EXIST 2026 for multimodal sexism identification in memes. The task is composed of three, increasingly harder sub

model-releasesarxiv-cs-cl
7 Jul 2026
Safety

Alignment-Guided Largest Table Overlap Size Estimation

DGX agent

arXiv:2607.03049v1 Announce Type: new Abstract: Fast estimation of the size of the largest overlap between tables enables blocking and query-by-table retrieval in large table repositories. The first a

safetyarxiv-cs-cl
7 Jul 2026
Model Releases

Anchored Self-Play for Code Repair

DGX agent

arXiv:2607.03523v1 Announce Type: cross Abstract: Code repair is an important capability for language models (LMs): given a buggy program and unit tests, an LM must produce a fixed program that passes

model-releasesarxiv-cs-cl
7 Jul 2026
Research

Angry but Accurate: Detecting and Profiling the Counter-Misinformation Ecosystem on Twitter

DGX agent

arXiv:2607.02900v1 Announce Type: cross Abstract: On social media, many users actively push back against false claims. Understanding who pushes back and how they do so matters, as this corrective acti

researcharxiv-cs-cl
7 Jul 2026
Research

Annotating Korean adnominal ending constructions in corpus data: Beyond relative-clause identification

DGX agent

arXiv:2607.03681v1 Announce Type: new Abstract: The Korean adnominal ending exttt{ETM} occurs in diverse noun-modifying constructions, including relative-clause-like modifiers, adjectival and copular

researcharxiv-cs-cl
7 Jul 2026
Safety

Autonomous Information Seeking: A Roadmap for Agentic Recommender Systems

DGX agent

arXiv:2607.04433v1 Announce Type: cross Abstract: The rapid integration of large language model-based agents into recommender systems has driven a shift from static, ranking-based pipelines toward aut

safetyarxiv-cs-cl
7 Jul 2026
Research

Beyond Satisfaction: Learning Associations Between Content, Reviews, and Well-Being

DGX agent

arXiv:2607.02539v1 Announce Type: cross Abstract: Digital platforms commonly optimize for satisfaction using signals such as ratings, likes, and sentiment, implicitly treating satisfaction as a proxy

researcharxiv-cs-cl
7 Jul 2026
Research

Can Dialects Be Steered Like Languages? Sparse Neurons and Distributed Directions in Arabic LLMs

DGX agent

arXiv:2607.03936v1 Announce Type: new Abstract: A key challenge in Arabic NLP is the scarcity of dialectal data relative to Modern Standard Arabic (MSA), causing LLMs to overproduce MSA and struggle w

researcharxiv-cs-cl
7 Jul 2026
Safety

Can temporal article-level credibility signals improve domain-level credibility prediction?

DGX agent

arXiv:2607.04560v1 Announce Type: new Abstract: Web domain credibility evaluation is vital for combating misinformation. It is conducted by examining factors such as domain type, transparency, and ove

safetyarxiv-cs-cl
7 Jul 2026
Research

Candidate-Constrained Retrieval-Augmented Generation for LongEval-RAG: System Design and Empirical Analysis

DGX agent

arXiv:2607.04008v1 Announce Type: new Abstract: We present a candidate-constrained retrieval-augmented generation system for LongEval-RAG, where each query is associated with an organizer-provided can

researcharxiv-cs-cl
7 Jul 2026
Research

CARD: Cross-component Audio Representation Distillation for Encoder-Free Audio Captioning

DGX agent

arXiv:2607.04619v1 Announce Type: cross Abstract: Modern automated audio captioning systems pair a frozen audio encoder with a large language model (LLM) via a trainable projector, incurring the encod

researcharxiv-cs-cl
7 Jul 2026
Research

Characterizing the Temporal, Emotional, and Social Patterns of Adolescent Substance Use Discussions on Reddit

DGX agent

arXiv:2607.04566v1 Announce Type: new Abstract: Adolescence is a critical developmental period marked by heightened emotional sensitivity, social stress, and vulnerability to substance use. However, t

researcharxiv-cs-cl
7 Jul 2026
Model Releases

Chasing Moving Targets with Online Self-Play Reinforcement Learning for Safer Language Models

DGX agent

arXiv:2506.07468v4 Announce Type: replace-cross Abstract: Conventional large language model (LLM) safety alignment relies on a reactive, disjoint loop: attackers exploit a static model, then defenders

model-releasesarxiv-cs-cl
7 Jul 2026
Applications

CompLLM: Compression for Long Context Q&A

DGX agent

arXiv:2509.19228v2 Announce Type: replace Abstract: Large Language Models (LLMs) face significant computational challenges when processing long contexts due to the quadratic complexity of self-attenti

applicationsarxiv-cs-cl
7 Jul 2026
Safety

Coverage-Controlled Preference Mining from Noisy Claim Verification for Evidence-Grounded Generation

DGX agent

arXiv:2603.10494v2 Announce Type: replace Abstract: Evidence-grounded generation produces summaries whose claims should be supported by supplied evidence, but claim-level verifiers provide noisy feedb

safetyarxiv-cs-cl
7 Jul 2026
Safety

CrossHallu: Do Hallucination Signals Generalize Across Languages and Domains in Large Language Model's Internals?

DGX agent

arXiv:2607.04029v1 Announce Type: new Abstract: Recent hallucination detection techniques in large language models (LLMs) focus on directly extracting features from a model's internal representations

safetyarxiv-cs-cl
7 Jul 2026
Local Ai

Curated retrieval versus open web search in public AI information services: a coverage-trust trade-off

DGX agent

arXiv:2607.05217v1 Announce Type: cross Abstract: Public institutions increasingly use large language models (LLMs) to answer citizens' questions, often pairing a curated knowledge base with live web

local-aiarxiv-cs-cl
7 Jul 2026
Model Releases

Curriculum-Guided Layer Scaling for Language Model Pretraining

DGX agent

arXiv:2506.11389v4 Announce Type: replace Abstract: As the cost of pretraining large language models grows, there is continued interest in strategies to improve learning efficiency during this core tr

model-releasesarxiv-cs-cl
7 Jul 2026
Safety

Decomposed Prompting Does Not Fix Knowledge Gaps, But Helps Models Say 'I Don't Know'

DGX agent

arXiv:2602.04853v2 Announce Type: replace Abstract: Large language models often struggle to recognize their knowledge limits in closed-book question answering, leading to confident hallucinations. Whi

safetyarxiv-cs-cl
7 Jul 2026
Local Ai

DELTA-TTS: Adapting Autoregressive Model into Diffusion Language Model for Text-to-Speech

DGX agent

arXiv:2607.04140v1 Announce Type: cross Abstract: Autoregressive (AR) text-to-speech (TTS) models generate discrete speech tokens sequentially, which makes inference slow and can degrade robustness by

local-aiarxiv-cs-cl
7 Jul 2026
Safety

Distill Where the Student Goes: Teacher-Regularized RL for English-Evidence Cross-Lingual RAG

DGX agent

arXiv:2607.02966v1 Announce Type: new Abstract: Cross-lingual retrieval-augmented generation (RAG) is often deployed in an English-evidence regime, where users query in diverse languages but retrieved

safetyarxiv-cs-cl
7 Jul 2026
Research

Does It Fail to See or Fail to Know? Attributing Errors in Vision-Language Models

DGX agent

arXiv:2607.04683v1 Announce Type: cross Abstract: Vision-language models (VLMs) perform well on visual question answering with high-quality images but struggle when questions require knowledge beyond

researcharxiv-cs-cl
7 Jul 2026
Research

Don't Commit Alone: Joint Token Commitment in Diffusion Large Language Models

DGX agent

arXiv:2607.04469v1 Announce Type: new Abstract: Diffusion large language models (dLLMs) commit multiple tokens per denoising step by decoding each selected position independently from the shared conte

researcharxiv-cs-cl
7 Jul 2026
Research

DuplexChat: Constructing Speaker-Separated Full-Duplex Dialogue Speech at Scale for Spoken Dialogue Language Modeling

DGX agent

arXiv:2607.04941v1 Announce Type: new Abstract: Full-duplex spoken dialogue models are trained on conversational speech in which each speaker is represented as a separate stream, but existing large-sc

researcharxiv-cs-cl
7 Jul 2026
Agents

EdgeBench: Unveiling Scaling Laws of Learning from Real-World Environments

DGX agent

arXiv:2607.05155v1 Announce Type: new Abstract: Pretraining scaling laws reveal that model capability improves predictably with data and compute. But learning from real world environments after deploy

agentsarxiv-cs-cl
7 Jul 2026
Model Releases

Evaluating Large Language Models for Antisemitic Incident Classification

DGX agent

arXiv:2607.04890v1 Announce Type: new Abstract: Addressing hate and violence in society requires timely detection of hateful events from public reporting, but automated identification of hateful event

model-releasesarxiv-cs-cl
7 Jul 2026
Safety

Fair-GPTQ: Bias-Aware Quantization for Large Language Models

DGX agent

arXiv:2509.15206v3 Announce Type: replace Abstract: The high memory demands of generative language models have drawn attention to quantization, which reduces memory usage by mapping model weights to l

safetyarxiv-cs-cl
7 Jul 2026
Safety

Faithfulness to Refusal: A Causal Audit of Neuron Selectors

DGX agent

arXiv:2607.05355v1 Announce Type: new Abstract: Attribution scores increasingly identify which neuron rows of a language model matter for applications such as pruning, interpretability, and editing fo

safetyarxiv-cs-cl
7 Jul 2026
Research

Fidelity-Diversity Metrics for Text

DGX agent

arXiv:2607.04563v1 Announce Type: new Abstract: As language modeling technology matures, there is an increasing research focus on the composition and curation of datasets used to train these models. F

researcharxiv-cs-cl
7 Jul 2026
Agents

FOI-O: An NZ-first ontology and verification methods package for Freedom of Information process modelling

DGX agent

arXiv:2607.02947v1 Announce Type: cross Abstract: Public official-information request records contain process signals. They can support research, workflow review, and human-supervised agent help. Yet

agentsarxiv-cs-cl
7 Jul 2026
Model Releases

FormalRx: Rectify and eXamine Semantic Failures in Autoformalization

DGX agent

arXiv:2607.04655v1 Announce Type: new Abstract: The veracious semantic alignment in autoformalization is significant for formal mathematical reasoning. However, existing evaluations provide only opaqu

model-releasesarxiv-cs-cl
7 Jul 2026
← Previous
1…3031323334…161
Next →