AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,646 results
28 Jul 2026

Mwando: Leveraging AI to Preserve and Teach shiKomori

AgentsDGX agent

arXiv:2607.23481v1 Announce Type: new Abstract: This paper presents Mwando, a virtual educational assistant designed to support the teaching and preservation of shiKomori, the language of the Comoros

Neural Dynamics of AI-attributed Irony Reveal a Partial Intentional Stance

ResearchDGX agent

arXiv:2510.17168v2 Announce Type: replace Abstract: As Large Language Models (LLMs) are increasingly deployed as social agents and trained to produce humor and irony, a question emerges: when encounte

No Optimal Language Set Exists for Multilingual Instruction Tuning: Insights from a Linguistically-Informed Study

Model ReleasesDGX agent

arXiv:2410.07809v2 Announce Type: replace Abstract: Multilingual instruction tuning (MIT) is challenged by the curse of multilinguality, data scarcity, and high computational cost. A natural hypothesi


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Occluded Oculus: Operationalizing Stylistic Obscurement

ResearchDGX agent

arXiv:2607.24411v1 Announce Type: cross Abstract: What did it take for Hermes, the devout messenger of the Olympian gods, to slay Argus Panoptes, the multi-eyed giant of Greek myth? As the perfect gua

Omni-Prune: Query-Aware Unified Token Pruning for Efficient Omnimodal Large Language Models

HardwareDGX agent

arXiv:2607.23445v1 Announce Type: cross Abstract: Omnimodal large language models (OmniLLMs) are rapidly extending multimodal reasoning to cover synchronized audio and video. However, the resulting au

Open Your Model's Eyes: Video and Context-Aware Multimodal Backchannel Prediction

SafetyDGX agent

arXiv:2607.22729v1 Announce Type: cross Abstract: Backchannels, which signal listener states like empathy and understanding, are fundamental to natural human interaction. However, current approaches r

Parallel Tokenizers: Rethinking Encoder Models' Vocabulary Design in Cross-Lingual Transfer of Low-Resource Languages

SafetyDGX agent

arXiv:2510.06128v2 Announce Type: replace Abstract: Tokenization forms the basis of multilingual language models, yet existing methods often limit cross-lingual transfer by mapping semantically equiva

PatiGonit22K: A Comprehensive Dataset for Solving Complex Bengali MWPs

Model ReleasesDGX agent

arXiv:2607.22859v1 Announce Type: new Abstract: Mathematical Word Problems (MWPs) are an important benchmark for evaluating natural language understanding and quantitative reasoning. Despite recent pr

PIVOT: Efficient Query-Group Indexing for Token-Level Sparse Attention

Model ReleasesDGX agent

arXiv:2607.24593v1 Announce Type: new Abstract: Token-level sparse attention, as implemented by DeepSeek Sparse Attention (DSA) in production systems, makes the downstream attention efficient but shif

PlanCraft: Sketch, Refine, and Furnish for Architect-Inspired Progressive 3D Residential Scene Generation

AgentsDGX agent

arXiv:2607.23491v1 Announce Type: cross Abstract: Two structural insights have been overlooked in automated residential floor plan generation. First, design is inherently progressive. Architects begin

Pointer-Augmented Autoregressive Generation of Patent Claims with Joint Topology and Content Decoding

Model ReleasesDGX agent

arXiv:2607.24040v1 Announce Type: new Abstract: Autoregressive decoders emit flat token sequences and cannot enforce hierarchical constraints across output segments, a limitation that becomes acute in

PReSS: An Automated Black-Box Framework for Evaluating Political Stance Stability in LLMs

SafetyDGX agent

arXiv:2504.17052v4 Announce Type: replace Abstract: Existing evaluations of political bias in large language models (LLMs) typically classify outputs as left- or right-leaning. We extend this perspect

Reading Between the Signs: Predicting Future Suicidal Ideation from Adolescent Social Media Texts

ResearchDGX agent

arXiv:2509.03530v2 Announce Type: replace Abstract: Suicide is a leading cause of death, yet predicting it remains a significant challenge. Risk factors such as depression or substance use are commonl

Rethinking the Generation Order of Block Diffusion Language Models

ResearchDGX agent

arXiv:2607.24306v1 Announce Type: new Abstract: Diffusion language models enable flexible arbitrary-order generation, but existing sampling methods are mostly designed for early masked diffusion model

Retrieval-Augmented Large Language Models as Components of Cognitive Computing architecture for Regulatory Knowledge Management

Local AiDGX agent

arXiv:2607.24352v1 Announce Type: new Abstract: The aim of this article is to verify whether integrating large language models (LLMs) with the Retrieval-Augmented Generation (RAG) architecture enables

RM-Distiller: Exploiting Generative LLM for Reward Model Distillation

SafetyDGX agent

arXiv:2601.14032v2 Announce Type: replace Abstract: Reward models (RMs) play a pivotal role in aligning large language models (LLMs) with human preferences. Due to the difficulty of obtaining high-qua

SAFE-MEME: Structured Reasoning Framework for Robust Hate Speech Detection in Memes

Model ReleasesDGX agent

arXiv:2412.20541v2 Announce Type: replace Abstract: Memes act as cryptic tools for sharing sensitive ideas, often requiring contextual knowledge to interpret them correctly. It makes multimodal meme m

Self-Authored Verification Is Unreliable in Heuristic Self-Improving Agents

SafetyDGX agent

arXiv:2607.24300v1 Announce Type: new Abstract: Self-improving agents accumulate capability by repeatedly rewriting procedural policies, controllers, or heuristic rules. They typically rely on self-au

Similarity Is Not Logic: Factored Inference for Dual-Encoder Vision-Language Models

ResearchDGX agent

arXiv:2607.23052v1 Announce Type: cross Abstract: Dual-encoder vision-language models (VLMs) expose a similarity interface that enables zero-shot retrieval but fails compositional constraints: queries

Simple Language Normalization Wins: Cross-Lingual Speaker Verification for the TidyVoice 2026 Challenge

ResearchDGX agent

arXiv:2607.22923v1 Announce Type: new Abstract: Cross-lingual mismatch remains a key source of overall degradation in modern speaker verification. The TidyVoice2026 Challenge targets this setting with

SINT-Flow: Schema Integration using Large Language Model Workflows

Model ReleasesDGX agent

arXiv:2607.24492v1 Announce Type: new Abstract: The goal of schema integration is, given a set of input schemata or tables, to derive a global, unified schema that is able to represent the concepts, a

SMART: LLM-Augmented Hybrid Retrieval for Dynamic Product Ads

ResearchDGX agent

arXiv:2607.23121v1 Announce Type: cross Abstract: Dynamic Product Ads (DPA) require retrieving relevant items from multi-million product catalogs, balancing two competing objectives: retargeting (re-s

Speech Signals Complement LLMs for Predicting Interpersonal Attraction in Speed Dating

ResearchDGX agent

arXiv:2607.23037v1 Announce Type: new Abstract: Large language models (LLMs) can predict interpersonal attraction from conversation transcripts, but it remains unclear what a speech predictor can add

Systematic Analysis of Large Language Models and Transformer-Based Machine Translation for English-Tamil and Tamil-English Across Diverse Datasets

SafetyDGX agent

arXiv:2607.24515v1 Announce Type: new Abstract: The challenge of Machine Translation for low resource languages such as Tamil is primarily caused by the restricted amount of parallel data for these la

Tailored untruths: How personalisation challenges LLM safeguards

SafetyDGX agent

arXiv:2510.12993v3 Announce Type: replace Abstract: Large Language Models (LLMs) can generate highly persuasive disinformation, yet little is known about how effectively they personalise it across lan

The Best Programming Language for Tokenmaxxing: An Investigation of Coding Agent Behavior Across Programming Languages

AgentsDGX agent

arXiv:2607.22807v1 Announce Type: cross Abstract: Although coding agents are now very effective in a variety of programming languages, this paper first shows that the cost (in tokens) can very signifi

The Cross-Domain Generalization Cost of Offensive Language Detection

ResearchDGX agent

arXiv:2607.23512v1 Announce Type: new Abstract: Offensive language detection models generally suffer performance degradation when deployed across datasets and across languages, yet most existing studi

The Few-shot Dilemma: Over-prompting Large Language Models

Model ReleasesDGX agent

arXiv:2509.13196v2 Announce Type: replace Abstract: Over-prompting, a phenomenon where excessive examples in prompts lead to diminished performance in Large Language Models (LLMs), challenges the conv

The JEPA Paradox in Language: The Geometry of Linguistic Alternatives

ResearchDGX agent

arXiv:2607.23531v1 Announce Type: new Abstract: Joint-Embedding Predictive Architectures (JEPAs) are effective for images, video, and audio, yet deterministic JEPA-style latent prediction has not beco

Toward Automated Detection of Documentation Inconsistencies in Electronic Health Records

Model ReleasesDGX agent

arXiv:2607.22954v1 Announce Type: new Abstract: Objective: To characterize the kinds of internal documentation inconsistencies a general-domain large language model (LLM) can surface from real-world d

Towards Cultural Bridge by Bahnaric-Vietnamese Translation Using Transfer Learning of Sequence-To-Sequence Pre-training Language Model

ResearchDGX agent

arXiv:2505.11421v2 Announce Type: replace Abstract: This work explores the journey towards achieving Bahnaric-Vietnamese translation for the sake of culturally bridging the two ethnic groups in Vietna

TRIDENT: Benchmarking LLM Safety in Finance, Medicine, and Law

Model ReleasesDGX agent

arXiv:2507.21134v2 Announce Type: replace Abstract: As large language models (LLMs) are increasingly deployed in high-risk domains such as law, finance, and medicine, systematically evaluating their d

Two Regimes of Chain-of-Thought Unfaithfulness: Behavioral Detection Fails Where Models Are Wrong

Model ReleasesDGX agent

arXiv:2607.23458v1 Announce Type: new Abstract: Chain-of-thought (CoT) explanations support oversight only if they are faithful: the stated reasoning must actually produce the answer. Auditing black-b

Verbalized Particle Posterior: Bayesian Inference over Natural Language Hypotheses

Model ReleasesDGX agent

arXiv:2607.22961v1 Announce Type: cross Abstract: Verbalized Machine Learning (VML) parameterizes a model as a natural-language prompt that an LLM evaluates as f(x; theta). The framework is interpreta

What do Reward Models Memorize?

ResearchDGX agent

arXiv:2607.24484v1 Announce Type: cross Abstract: This paper studies what discriminatively trained reward models (RMs) memorize by measuring counterfactual memorization on two human preference dataset

Where Quality Breaks in Compressed Short-Text Generation: Staged Bottleneck Localization

ApplicationsDGX agent

arXiv:2607.24176v1 Announce Type: new Abstract: Compressed short-text generators can fail in two different places: the codec may discard information before generation starts, or the latent generator m

Which Models Perform Better in Inheritance Reasoning?

Model ReleasesDGX agent

arXiv:2606.13751v4 Announce Type: replace Abstract: This paper presents the participation of team PSL in the QIAS 2026 Shared Task on Arabic Islamic inheritance reasoning. The task evaluates the abili

Who Gets Named: Citation Type Predicts Individual Naming by Grounded Language Models, and a Roster Instrument Captures 0.5% of It

Model ReleasesDGX agent

arXiv:2607.23893v1 Announce Type: cross Abstract: Prior work on AI brand visibility measures the firm: does a model recommend a company, and does that track its reputation. This study asks the questio

Wired for Overconfidence: A Mechanistic Perspective on Inflated Verbalized Confidence in LLMs

ResearchDGX agent

arXiv:2604.01457v3 Announce Type: replace Abstract: Large language models are often not just wrong, but confidently wrong: when they produce factually incorrect answers, they tend to verbalize overly

You Talkin to Me?: A Network Analysis of Gendered Speaker-Addressee Patterns in Film Screenplays

SafetyDGX agent

arXiv:2607.22656v1 Announce Type: cross Abstract: Objective: This paper investigates the gendered structure of speaker addressee relationships in film dialogue, asking not merely who speaks, but who i

Zing: Social Mind for LLMs

Model ReleasesDGX agent

arXiv:2607.23740v1 Announce Type: new Abstract: As large language models move from isolated task solving toward long-term service in human environments, they require social intelligence: the ability t

27 Jul 2026

A Consensus-Based Framework for Relative Preference Evaluation of Large Language Models

SafetyDGX agent

arXiv:2607.21632v1 Announce Type: new Abstract: Traditional benchmarks for LLMs primarily rely on static datasets and objective scoring metrics, which often fail to capture differences in response qua

A Factorial Study of Synthetic Data Generation for Low-Resource Machine Translation using Grammar Books

ResearchDGX agent

arXiv:2607.22376v1 Announce Type: new Abstract: Most endangered languages lack the parallel data required for machine translation, despite the existence of descriptive grammar books. We introduce a pi

Adversarial Prompts for Acceptance Collapse in Speculative Decoding

SafetyDGX agent

arXiv:2607.21804v1 Announce Type: cross Abstract: Lossless acceleration schemes, such as speculative decoding, promise significant inference speedups by relying on dynamic token-level alignment betwee

Adversarial Style Optimization: Enhancing VLM Jailbreaks by GRPO-based Stylistic Triggers Optimization

SafetyDGX agent

arXiv:2607.21619v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have achieved impressive performance, but their safety alignment remains vulnerable to jailbreak attacks. Exist

Agentic Evaluation of Copyright Law Compliance

Model ReleasesDGX agent

arXiv:2607.21799v1 Announce Type: new Abstract: Large language model (LLM) agents increasingly perform commercial tasks that involve retrieving external content such as images and, where appropriate,

Analysing Self-Harm Representations in Language Models: a Cross-Architecture Study

Model ReleasesDGX agent

arXiv:2607.21988v1 Announce Type: new Abstract: Self-harm content is particularly challenging to detect using NLP techniques, and is also a high-stakes task which requires the highest accuracy to enab

Analyzing Toxic Behavior and Its Impact on the Mastodon Community

ResearchDGX agent

arXiv:2607.21980v1 Announce Type: new Abstract: Mastodon as a decentralized federation of independently moderated social servers poses unique challenges for the detection and mitigation of toxic conte

Benchmarking Fine-tuning and Retrieval Strategies for a Multimodal Language Model on the NRC Reactor Operator Licensing Examination

Model ReleasesDGX agent

arXiv:2607.22067v1 Announce Type: new Abstract: The integration of large language models (LLMs) into the nuclear power industry requires outputs grounded in domain-specific knowledge. This study evalu

Biomedical Machine Translation for Low-Resource Arabic-Script Languages via Cross-Lingual Transfer and LoRA Adapter Merging

ApplicationsDGX agent

arXiv:2607.22300v1 Announce Type: new Abstract: We present a systematic study of healthcare-domain cross-lingual transfer to address the scarcity of biomedical NMT resources for Arabic-script language

Cross-Tokenizer On-Policy Distillation via Byte-Prefix Marginalization

Model ReleasesDGX agent

arXiv:2607.22334v1 Announce Type: cross Abstract: Open-weight language models from different families exhibit complementary capabilities, motivating their consolidation into a compact student through

Data Quality over Capacity: Internalizing Documents into LoRA Adapters for Closed-Book QA

Model ReleasesDGX agent

arXiv:2607.21861v1 Announce Type: new Abstract: We study baking documents directly into the weights of a 4-bit Gemma-4-e4b model via LoRA, so a system can answer questions about a corpus closed-book:

DBA-Bench: A Production-Fidelity Benchmark for LLM-Based Database Operations Agents

Model ReleasesDGX agent

arXiv:2607.22165v1 Announce Type: cross Abstract: LLM-based database agents show promise, but differing task scopes, testbeds, and metrics hinder comparison. We identify four gaps between evaluation a

Developing and Validating the Spanish Version of the Large Language Models Dependency Scale (LLM-D12-SP)

ResearchDGX agent

arXiv:2607.22041v1 Announce Type: new Abstract: There is a growing need for reliable and culturally validated instruments to assess psychological dependency on large language models (LLMs), particular

Diffusion Models in Medical Image Inpainting: Challenges, Solution Taxonomy, and Future Directions

ResearchDGX agent

arXiv:2607.21904v1 Announce Type: cross Abstract: Image inpainting aims to reconstruct missing or corrupted regions of an image while preserving as much as possible, visual and semantic consistency. I

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models

Model ReleasesDGX agent

arXiv:2607.21617v1 Announce Type: cross Abstract: Vision Language Models (VLMs) are increasingly used in place of traditional OCR pipelines for document understanding. In this paper, we show they do n

DWT-Fusion: A Signal-Based Framework for Training-Free LLM-Generated Text Detection

Model ReleasesDGX agent

arXiv:2607.22026v1 Announce Type: new Abstract: Detecting LLM-generated text remains challenging under zero-shot and training-free conditions, especially when detectors must generalize across datasets

Dynamic Commonsense Coordination for Empathetic Response Generation

Model ReleasesDGX agent

arXiv:2607.22136v1 Announce Type: new Abstract: Empathetic Response Generation (ERG) requires models to recognize users' emotions and generate empathetic responses. Commonsense knowledge has been show

Enjoy Your Talk: A Human-Centered Benchmark for Multi-Turn Dialogue with Decoupled User Simulation, Target Modeling, and Judging

Model ReleasesDGX agent

arXiv:2607.10428v2 Announce Type: replace Abstract: Evaluating large language models (LLMs) as multi-turn conversational partners requires probing capabilities that single-turn benchmarks miss: person

Enough is as good as a feast: A Comprehensive Analysis of How Reinforcement Learning Mitigates Task Conflicts in LLMs

Model ReleasesDGX agent

arXiv:2607.22039v1 Announce Type: new Abstract: Model merging plays a crucial role in consolidating multiple specialized models into a single, unified model, especially in the era of large language mo

← Previous
1…1718192021…128
Next →