AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
8 Jul 2026

UCSC NLP at SemEval-2026 Task 10: Boundary-Aware Span Extraction and RoBERTa Classification for Conspiracy Detection

ResearchDGX agent

arXiv:2607.05689v1 Announce Type: new Abstract: We present our systems for SemEval-2026 Task 10 (PsyCoMark), addressing conspiracy marker extraction (Subtask 1) and document-level conspiracy detection

Umm... With Transformers? Insights from Filled Pause Use across Four Slavic Parliaments

ResearchDGX agent

arXiv:2607.05964v1 Announce Type: new Abstract: Filled pauses (FPs) are a universal feature of spontaneous speech, yet most studies rely on small, single-language corpora, limiting the generalisabilit

When Does Tool Use Increase the Expressive Power of Finite-Precision Recurrent Models?

AgentsDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2607.06155v1 Announce Type: cross Abstract: Modern sequence models are increasingly deployed as agents that interleave token generation with calls to external tools. We give an exact, architectu

Where to cut, how deep: BPE and Unigram-LM on chemistry SMILES

ResearchDGX agent

arXiv:2607.05691v1 Announce Type: new Abstract: Every chemical language model reading SMILES begins with a tokenizer, yet the field has inherited byte-pair encoding (BPE) from natural language with li

WordVoice: Explicit and Decoupled Multi-Dimensional Word-Level Control for LLM-Based TTS

SafetyDGX agent

arXiv:2607.06461v1 Announce Type: cross Abstract: While recent Large Language Model (LLM)-based Text-to-Speech (TTS) systems have achieved remarkable naturalness, they predominantly rely on implicit e

7 Jul 2026

AI Wizards at EXIST 2026: Hierarchical Soft-Label Learning for Multimodal Sexism Identification in Memes

Model ReleasesDGX agent

arXiv:2607.04410v1 Announce Type: new Abstract: We present the AI Wizards submission to EXIST 2026 for multimodal sexism identification in memes. The task is composed of three, increasingly harder sub

Alignment-Guided Largest Table Overlap Size Estimation

SafetyDGX agent

arXiv:2607.03049v1 Announce Type: new Abstract: Fast estimation of the size of the largest overlap between tables enables blocking and query-by-table retrieval in large table repositories. The first a

Anchored Self-Play for Code Repair

Model ReleasesDGX agent

arXiv:2607.03523v1 Announce Type: cross Abstract: Code repair is an important capability for language models (LMs): given a buggy program and unit tests, an LM must produce a fixed program that passes

Angry but Accurate: Detecting and Profiling the Counter-Misinformation Ecosystem on Twitter

ResearchDGX agent

arXiv:2607.02900v1 Announce Type: cross Abstract: On social media, many users actively push back against false claims. Understanding who pushes back and how they do so matters, as this corrective acti

Annotating Korean adnominal ending constructions in corpus data: Beyond relative-clause identification

ResearchDGX agent

arXiv:2607.03681v1 Announce Type: new Abstract: The Korean adnominal ending exttt{ETM} occurs in diverse noun-modifying constructions, including relative-clause-like modifiers, adjectival and copular

Autonomous Information Seeking: A Roadmap for Agentic Recommender Systems

SafetyDGX agent

arXiv:2607.04433v1 Announce Type: cross Abstract: The rapid integration of large language model-based agents into recommender systems has driven a shift from static, ranking-based pipelines toward aut

Beyond Satisfaction: Learning Associations Between Content, Reviews, and Well-Being

ResearchDGX agent

arXiv:2607.02539v1 Announce Type: cross Abstract: Digital platforms commonly optimize for satisfaction using signals such as ratings, likes, and sentiment, implicitly treating satisfaction as a proxy

Can Dialects Be Steered Like Languages? Sparse Neurons and Distributed Directions in Arabic LLMs

ResearchDGX agent

arXiv:2607.03936v1 Announce Type: new Abstract: A key challenge in Arabic NLP is the scarcity of dialectal data relative to Modern Standard Arabic (MSA), causing LLMs to overproduce MSA and struggle w

Can temporal article-level credibility signals improve domain-level credibility prediction?

SafetyDGX agent

arXiv:2607.04560v1 Announce Type: new Abstract: Web domain credibility evaluation is vital for combating misinformation. It is conducted by examining factors such as domain type, transparency, and ove

Candidate-Constrained Retrieval-Augmented Generation for LongEval-RAG: System Design and Empirical Analysis

ResearchDGX agent

arXiv:2607.04008v1 Announce Type: new Abstract: We present a candidate-constrained retrieval-augmented generation system for LongEval-RAG, where each query is associated with an organizer-provided can

CARD: Cross-component Audio Representation Distillation for Encoder-Free Audio Captioning

ResearchDGX agent

arXiv:2607.04619v1 Announce Type: cross Abstract: Modern automated audio captioning systems pair a frozen audio encoder with a large language model (LLM) via a trainable projector, incurring the encod

Characterizing the Temporal, Emotional, and Social Patterns of Adolescent Substance Use Discussions on Reddit

ResearchDGX agent

arXiv:2607.04566v1 Announce Type: new Abstract: Adolescence is a critical developmental period marked by heightened emotional sensitivity, social stress, and vulnerability to substance use. However, t

Chasing Moving Targets with Online Self-Play Reinforcement Learning for Safer Language Models

Model ReleasesDGX agent

arXiv:2506.07468v4 Announce Type: replace-cross Abstract: Conventional large language model (LLM) safety alignment relies on a reactive, disjoint loop: attackers exploit a static model, then defenders

CompLLM: Compression for Long Context Q&A

ApplicationsDGX agent

arXiv:2509.19228v2 Announce Type: replace Abstract: Large Language Models (LLMs) face significant computational challenges when processing long contexts due to the quadratic complexity of self-attenti

Coverage-Controlled Preference Mining from Noisy Claim Verification for Evidence-Grounded Generation

SafetyDGX agent

arXiv:2603.10494v2 Announce Type: replace Abstract: Evidence-grounded generation produces summaries whose claims should be supported by supplied evidence, but claim-level verifiers provide noisy feedb

CrossHallu: Do Hallucination Signals Generalize Across Languages and Domains in Large Language Model's Internals?

SafetyDGX agent

arXiv:2607.04029v1 Announce Type: new Abstract: Recent hallucination detection techniques in large language models (LLMs) focus on directly extracting features from a model's internal representations

Curated retrieval versus open web search in public AI information services: a coverage-trust trade-off

Local AiDGX agent

arXiv:2607.05217v1 Announce Type: cross Abstract: Public institutions increasingly use large language models (LLMs) to answer citizens' questions, often pairing a curated knowledge base with live web

Curriculum-Guided Layer Scaling for Language Model Pretraining

Model ReleasesDGX agent

arXiv:2506.11389v4 Announce Type: replace Abstract: As the cost of pretraining large language models grows, there is continued interest in strategies to improve learning efficiency during this core tr

Decomposed Prompting Does Not Fix Knowledge Gaps, But Helps Models Say 'I Don't Know'

SafetyDGX agent

arXiv:2602.04853v2 Announce Type: replace Abstract: Large language models often struggle to recognize their knowledge limits in closed-book question answering, leading to confident hallucinations. Whi

DELTA-TTS: Adapting Autoregressive Model into Diffusion Language Model for Text-to-Speech

Local AiDGX agent

arXiv:2607.04140v1 Announce Type: cross Abstract: Autoregressive (AR) text-to-speech (TTS) models generate discrete speech tokens sequentially, which makes inference slow and can degrade robustness by

Distill Where the Student Goes: Teacher-Regularized RL for English-Evidence Cross-Lingual RAG

SafetyDGX agent

arXiv:2607.02966v1 Announce Type: new Abstract: Cross-lingual retrieval-augmented generation (RAG) is often deployed in an English-evidence regime, where users query in diverse languages but retrieved

Does It Fail to See or Fail to Know? Attributing Errors in Vision-Language Models

ResearchDGX agent

arXiv:2607.04683v1 Announce Type: cross Abstract: Vision-language models (VLMs) perform well on visual question answering with high-quality images but struggle when questions require knowledge beyond

Don't Commit Alone: Joint Token Commitment in Diffusion Large Language Models

ResearchDGX agent

arXiv:2607.04469v1 Announce Type: new Abstract: Diffusion large language models (dLLMs) commit multiple tokens per denoising step by decoding each selected position independently from the shared conte

DuplexChat: Constructing Speaker-Separated Full-Duplex Dialogue Speech at Scale for Spoken Dialogue Language Modeling

ResearchDGX agent

arXiv:2607.04941v1 Announce Type: new Abstract: Full-duplex spoken dialogue models are trained on conversational speech in which each speaker is represented as a separate stream, but existing large-sc

EdgeBench: Unveiling Scaling Laws of Learning from Real-World Environments

AgentsDGX agent

arXiv:2607.05155v1 Announce Type: new Abstract: Pretraining scaling laws reveal that model capability improves predictably with data and compute. But learning from real world environments after deploy

Evaluating Large Language Models for Antisemitic Incident Classification

Model ReleasesDGX agent

arXiv:2607.04890v1 Announce Type: new Abstract: Addressing hate and violence in society requires timely detection of hateful events from public reporting, but automated identification of hateful event

Fair-GPTQ: Bias-Aware Quantization for Large Language Models

SafetyDGX agent

arXiv:2509.15206v3 Announce Type: replace Abstract: The high memory demands of generative language models have drawn attention to quantization, which reduces memory usage by mapping model weights to l

Faithfulness to Refusal: A Causal Audit of Neuron Selectors

SafetyDGX agent

arXiv:2607.05355v1 Announce Type: new Abstract: Attribution scores increasingly identify which neuron rows of a language model matter for applications such as pruning, interpretability, and editing fo

Fidelity-Diversity Metrics for Text

ResearchDGX agent

arXiv:2607.04563v1 Announce Type: new Abstract: As language modeling technology matures, there is an increasing research focus on the composition and curation of datasets used to train these models. F

FOI-O: An NZ-first ontology and verification methods package for Freedom of Information process modelling

AgentsDGX agent

arXiv:2607.02947v1 Announce Type: cross Abstract: Public official-information request records contain process signals. They can support research, workflow review, and human-supervised agent help. Yet

FormalRx: Rectify and eXamine Semantic Failures in Autoformalization

Model ReleasesDGX agent

arXiv:2607.04655v1 Announce Type: new Abstract: The veracious semantic alignment in autoformalization is significant for formal mathematical reasoning. However, existing evaluations provide only opaqu

From Gentlemen to Frontiermen: Masculine Formations in English-Language Fiction (1771--1930)

ResearchDGX agent

arXiv:2607.03323v1 Announce Type: new Abstract: Masculinity in nineteenth-century fiction is not a single ideal but a field of competing scripts. Drawing on 150 British and American canonical novels f

GameEngineBench: Evaluating Coding Agents on Real C++ Runtime Environments

Model ReleasesDGX agent

arXiv:2607.03525v1 Announce Type: cross Abstract: Game engines provide real-time simulation, rendering, physics, interaction, networking, and asset pipelines, making them valuable not only for games b

Generative Pseudo-Labeling for Pre-Ranking with LLMs

SafetyDGX agent

arXiv:2602.20995v2 Announce Type: replace-cross Abstract: Pre-ranking is a critical stage in industrial recommendation systems, tasked with efficiently scoring thousands of recalled items for downstre

GRAFT: Grafted Reference Audio for Fine-grained Pronunciation in Zero-shot Text-to-Speech

Model ReleasesDGX agent

arXiv:2607.02633v1 Announce Type: cross Abstract: We present GRAFT, a per-word pronunciation conditioning mechanism for text-to-speech neural codec language modeling. Existing systems reach high intel

GRASP: Graph-Reasoning Aided Survey Planning for High-Fidelity Related Work Generation

ResearchDGX agent

arXiv:2607.03709v1 Announce Type: new Abstract: Writing a literature review requires a deep understanding of the relationships among cited papers: how they build on, challenge, or offer alternative pe

HiSAC: Hierarchical Sparse Activation Compression for Ultra-long Sequence Modeling in Recommenders

ApplicationsDGX agent

arXiv:2602.21009v2 Announce Type: replace-cross Abstract: Modern recommender systems leverage ultra-long user behavior sequences to capture dynamic preferences, but end-to-end modeling is infeasible i

How Much is Left? LLMs Linearly Encode Their Remaining Output Length

TutorialsDGX agent

arXiv:2607.05316v1 Announce Type: new Abstract: Large language models generate one token at a time, yet their responses show remarkably consistent length structure: step-by-step solutions converge in

How to Build Digital Humans? From Priors to Photorealistic Avatars

TutorialsDGX agent

arXiv:2607.04341v1 Announce Type: cross Abstract: This state-of-the-art report provides an overview of controllable 3D human avatar creation. We describe current 3D avatar systems, which typically con

How Utilitarian Are OpenAI's Models Really? Replicating and Reinterpreting Pfeffer, Krugel, and Uhl (2025)

SafetyDGX agent

arXiv:2603.22730v2 Announce Type: replace Abstract: Pfeffer, Krugel, and Uhl (2025) report that OpenAI's reasoning model o1-mini produces more utilitarian responses to the trolley problem and footbrid

Hyper-KGGen: A Skill-Driven Knowledge Extractor for High-Quality Knowledge Hypergraph Generation

Model ReleasesDGX agent

arXiv:2602.19543v2 Announce Type: replace Abstract: Knowledge hypergraphs surpass traditional binary knowledge graphs by encapsulating complex n-ary atomic facts, providing a more comprehensive paradi

Identifiability Without Gaussianity: Symbolic World Models and Near-Infinite Temporal Consistency

SafetyDGX agent

arXiv:2606.12471v2 Announce Type: replace-cross Abstract: Klindt, LeCun, and Balestriero (arXiv:2605.26379) proved that Joint-Embedding Predictive Architectures (JEPAs) achieve linear identifiability,

Improving LLMs via Validator-to-Generator Alignment

SafetyDGX agent

arXiv:2607.02668v1 Announce Type: new Abstract: Large language models are inconsistent: varying prompts or including unrelated information can lead to unexpected changes in model outputs. The generato

Is Your Benchmark Still Useful? Dynamic Benchmarking for Code Language Models

Model ReleasesDGX agent

arXiv:2503.06643v2 Announce Type: replace-cross Abstract: In this paper, we tackle a critical challenge in model evaluation: how to keep code benchmarks useful when models might have already seen them

Jointly Improving Dialect Identification and ASR in Indian Languages using Multimodal Feature Fusion

ResearchDGX agent

arXiv:2607.02862v1 Announce Type: new Abstract: Automatic Speech Recognition (ASR) and Dialect Identification (DID) are crucial for Indian languages, many of which are low-resource and exhibit signifi

Knowing When Not to Answer: Lightweight KB-Aligned OOD Detection for Safe RAG

SafetyDGX agent

arXiv:2508.02296v3 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) systems are increasingly deployed in high-stakes domains, where safety depends not only on how a system answers

Knowing When to Stop: Predicting Execution-Consistency Convergence in Text-to-SQL

Model ReleasesDGX agent

arXiv:2607.03991v1 Announce Type: cross Abstract: Repeated LLM calls are the standard way to estimate how trustworthy a Text-to-SQL result is: run the pipeline multiple times, judge each SQL execution

Knowledge Knows, Verbalization Tells: Disentangling Latent Directions for Mathematical Solvability in LLMs

ResearchDGX agent

arXiv:2607.05013v1 Announce Type: new Abstract: Although LLMs have made significant progress in mathematical reasoning, determining whether a mathematical problem is solvable remains a fundamental yet

Lacuna Inc. at SemEval-2026 Task 4: Structurally Gated State-Space Models for Disentangling Narrative Similarity

SafetyDGX agent

arXiv:2607.03482v1 Announce Type: new Abstract: In this paper, we present the Invariant-Variant Disentangled State-Space Model (IVD-SSM), our submission to SemEval-2026 Task 4 on Narrative Story Simil

Language Models as Higher-Order Planning Formalizers

ResearchDGX agent

arXiv:2603.23844v2 Announce Type: replace Abstract: Recent work provides overwhelming evidence that LLMs, even those trained to scale their reasoning trace, quickly deteriorate at planning as problems

Large-scale dataset of automatically classified rhetorical sections in scientific papers

ResearchDGX agent

arXiv:2607.03381v1 Announce Type: cross Abstract: Scientific papers follow rhetorical structures that organize content into sections such as Introduction, Methods, Results, and Discussion. Automatical

Latent Visual Cache for Video Reasoning

SafetyDGX agent

arXiv:2607.02607v1 Announce Type: cross Abstract: Video reasoning requires Large Multimodal Models (LMMs) to remain grounded in dense evidence, yet existing systems largely adopt 'read-once, generate-

Learning from Lost Provenance: Multiple Instance Learning for Cancer Registry Tumor Group Classification

ResearchDGX agent

arXiv:2607.03481v1 Announce Type: new Abstract: Modernizing cancer registries with deep learning is opening new opportunities to automate labor-intensive tasks such as the coding of pathology reports.

Learning When to Attend: Conditional Memory Access for Long-Context LLMs

Model ReleasesDGX agent

arXiv:2603.17484v2 Announce Type: replace Abstract: Language models struggle to generalize beyond pretraining context lengths, limiting long-horizon reasoning and retrieval. Continued pretraining on l

Legible-by-Construction: Attention and End-to-End Transformers

Model ReleasesDGX agent

arXiv:2607.04319v1 Announce Type: new Abstract: A companion paper showed that a transformer's feed-forward layer can be rebuilt from explicit fuzzy set operations - intersection, set-difference, and a

← Previous
1…2425262728…129
Next →