AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
Research

Similarity Is Not Logic: Factored Inference for Dual-Encoder Vision-Language Models

DGX agent

arXiv:2607.23052v1 Announce Type: cross Abstract: Dual-encoder vision-language models (VLMs) expose a similarity interface that enables zero-shot retrieval but fails compositional constraints: queries

researcharxiv-cs-cl
28 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Simple Language Normalization Wins: Cross-Lingual Speaker Verification for the TidyVoice 2026 Challenge

DGX agent

arXiv:2607.22923v1 Announce Type: new Abstract: Cross-lingual mismatch remains a key source of overall degradation in modern speaker verification. The TidyVoice2026 Challenge targets this setting with

researcharxiv-cs-cl
28 Jul 2026
Model Releases

SINT-Flow: Schema Integration using Large Language Model Workflows

DGX agent

arXiv:2607.24492v1 Announce Type: new Abstract: The goal of schema integration is, given a set of input schemata or tables, to derive a global, unified schema that is able to represent the concepts, a

model-releasesarxiv-cs-cl
28 Jul 2026
Research

SMART: LLM-Augmented Hybrid Retrieval for Dynamic Product Ads

DGX agent

arXiv:2607.23121v1 Announce Type: cross Abstract: Dynamic Product Ads (DPA) require retrieving relevant items from multi-million product catalogs, balancing two competing objectives: retargeting (re-s

researcharxiv-cs-cl
28 Jul 2026
Research

Speech Signals Complement LLMs for Predicting Interpersonal Attraction in Speed Dating

DGX agent

arXiv:2607.23037v1 Announce Type: new Abstract: Large language models (LLMs) can predict interpersonal attraction from conversation transcripts, but it remains unclear what a speech predictor can add

researcharxiv-cs-cl
28 Jul 2026
Safety

Systematic Analysis of Large Language Models and Transformer-Based Machine Translation for English-Tamil and Tamil-English Across Diverse Datasets

DGX agent

arXiv:2607.24515v1 Announce Type: new Abstract: The challenge of Machine Translation for low resource languages such as Tamil is primarily caused by the restricted amount of parallel data for these la

safetyarxiv-cs-cl
28 Jul 2026
Safety

Tailored untruths: How personalisation challenges LLM safeguards

DGX agent

arXiv:2510.12993v3 Announce Type: replace Abstract: Large Language Models (LLMs) can generate highly persuasive disinformation, yet little is known about how effectively they personalise it across lan

safetyarxiv-cs-cl
28 Jul 2026
Agents

The Best Programming Language for Tokenmaxxing: An Investigation of Coding Agent Behavior Across Programming Languages

DGX agent

arXiv:2607.22807v1 Announce Type: cross Abstract: Although coding agents are now very effective in a variety of programming languages, this paper first shows that the cost (in tokens) can very signifi

agentsarxiv-cs-cl
28 Jul 2026
Research

The Cross-Domain Generalization Cost of Offensive Language Detection

DGX agent

arXiv:2607.23512v1 Announce Type: new Abstract: Offensive language detection models generally suffer performance degradation when deployed across datasets and across languages, yet most existing studi

researcharxiv-cs-cl
28 Jul 2026
Model Releases

The Few-shot Dilemma: Over-prompting Large Language Models

DGX agent

arXiv:2509.13196v2 Announce Type: replace Abstract: Over-prompting, a phenomenon where excessive examples in prompts lead to diminished performance in Large Language Models (LLMs), challenges the conv

model-releasesarxiv-cs-cl
28 Jul 2026
Research

The JEPA Paradox in Language: The Geometry of Linguistic Alternatives

DGX agent

arXiv:2607.23531v1 Announce Type: new Abstract: Joint-Embedding Predictive Architectures (JEPAs) are effective for images, video, and audio, yet deterministic JEPA-style latent prediction has not beco

researcharxiv-cs-cl
28 Jul 2026
Model Releases

Toward Automated Detection of Documentation Inconsistencies in Electronic Health Records

DGX agent

arXiv:2607.22954v1 Announce Type: new Abstract: Objective: To characterize the kinds of internal documentation inconsistencies a general-domain large language model (LLM) can surface from real-world d

model-releasesarxiv-cs-cl
28 Jul 2026
Research

Towards Cultural Bridge by Bahnaric-Vietnamese Translation Using Transfer Learning of Sequence-To-Sequence Pre-training Language Model

DGX agent

arXiv:2505.11421v2 Announce Type: replace Abstract: This work explores the journey towards achieving Bahnaric-Vietnamese translation for the sake of culturally bridging the two ethnic groups in Vietna

researcharxiv-cs-cl
28 Jul 2026
Model Releases

TRIDENT: Benchmarking LLM Safety in Finance, Medicine, and Law

DGX agent

arXiv:2507.21134v2 Announce Type: replace Abstract: As large language models (LLMs) are increasingly deployed in high-risk domains such as law, finance, and medicine, systematically evaluating their d

model-releasesarxiv-cs-cl
28 Jul 2026
Model Releases

Two Regimes of Chain-of-Thought Unfaithfulness: Behavioral Detection Fails Where Models Are Wrong

DGX agent

arXiv:2607.23458v1 Announce Type: new Abstract: Chain-of-thought (CoT) explanations support oversight only if they are faithful: the stated reasoning must actually produce the answer. Auditing black-b

model-releasesarxiv-cs-cl
28 Jul 2026
Model Releases

Verbalized Particle Posterior: Bayesian Inference over Natural Language Hypotheses

DGX agent

arXiv:2607.22961v1 Announce Type: cross Abstract: Verbalized Machine Learning (VML) parameterizes a model as a natural-language prompt that an LLM evaluates as f(x; theta). The framework is interpreta

model-releasesarxiv-cs-cl
28 Jul 2026
Research

What do Reward Models Memorize?

DGX agent

arXiv:2607.24484v1 Announce Type: cross Abstract: This paper studies what discriminatively trained reward models (RMs) memorize by measuring counterfactual memorization on two human preference dataset

researcharxiv-cs-cl
28 Jul 2026
Applications

Where Quality Breaks in Compressed Short-Text Generation: Staged Bottleneck Localization

DGX agent

arXiv:2607.24176v1 Announce Type: new Abstract: Compressed short-text generators can fail in two different places: the codec may discard information before generation starts, or the latent generator m

applicationsarxiv-cs-cl
28 Jul 2026
Model Releases

Which Models Perform Better in Inheritance Reasoning?

DGX agent

arXiv:2606.13751v4 Announce Type: replace Abstract: This paper presents the participation of team PSL in the QIAS 2026 Shared Task on Arabic Islamic inheritance reasoning. The task evaluates the abili

model-releasesarxiv-cs-cl
28 Jul 2026
Model Releases

Who Gets Named: Citation Type Predicts Individual Naming by Grounded Language Models, and a Roster Instrument Captures 0.5% of It

DGX agent

arXiv:2607.23893v1 Announce Type: cross Abstract: Prior work on AI brand visibility measures the firm: does a model recommend a company, and does that track its reputation. This study asks the questio

model-releasesarxiv-cs-cl
28 Jul 2026
Research

Wired for Overconfidence: A Mechanistic Perspective on Inflated Verbalized Confidence in LLMs

DGX agent

arXiv:2604.01457v3 Announce Type: replace Abstract: Large language models are often not just wrong, but confidently wrong: when they produce factually incorrect answers, they tend to verbalize overly

researcharxiv-cs-cl
28 Jul 2026
Safety

You Talkin to Me?: A Network Analysis of Gendered Speaker-Addressee Patterns in Film Screenplays

DGX agent

arXiv:2607.22656v1 Announce Type: cross Abstract: Objective: This paper investigates the gendered structure of speaker addressee relationships in film dialogue, asking not merely who speaks, but who i

safetyarxiv-cs-cl
28 Jul 2026
Model Releases

Zing: Social Mind for LLMs

DGX agent

arXiv:2607.23740v1 Announce Type: new Abstract: As large language models move from isolated task solving toward long-term service in human environments, they require social intelligence: the ability t

model-releasesarxiv-cs-cl
28 Jul 2026
Safety

A Consensus-Based Framework for Relative Preference Evaluation of Large Language Models

DGX agent

arXiv:2607.21632v1 Announce Type: new Abstract: Traditional benchmarks for LLMs primarily rely on static datasets and objective scoring metrics, which often fail to capture differences in response qua

safetyarxiv-cs-cl
27 Jul 2026
Research

A Factorial Study of Synthetic Data Generation for Low-Resource Machine Translation using Grammar Books

DGX agent

arXiv:2607.22376v1 Announce Type: new Abstract: Most endangered languages lack the parallel data required for machine translation, despite the existence of descriptive grammar books. We introduce a pi

researcharxiv-cs-cl
27 Jul 2026
Safety

Adversarial Prompts for Acceptance Collapse in Speculative Decoding

DGX agent

arXiv:2607.21804v1 Announce Type: cross Abstract: Lossless acceleration schemes, such as speculative decoding, promise significant inference speedups by relying on dynamic token-level alignment betwee

safetyarxiv-cs-cl
27 Jul 2026
Safety

Adversarial Style Optimization: Enhancing VLM Jailbreaks by GRPO-based Stylistic Triggers Optimization

DGX agent

arXiv:2607.21619v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have achieved impressive performance, but their safety alignment remains vulnerable to jailbreak attacks. Exist

safetyarxiv-cs-cl
27 Jul 2026
Model Releases

Agentic Evaluation of Copyright Law Compliance

DGX agent

arXiv:2607.21799v1 Announce Type: new Abstract: Large language model (LLM) agents increasingly perform commercial tasks that involve retrieving external content such as images and, where appropriate,

model-releasesarxiv-cs-cl
27 Jul 2026
Model Releases

Analysing Self-Harm Representations in Language Models: a Cross-Architecture Study

DGX agent

arXiv:2607.21988v1 Announce Type: new Abstract: Self-harm content is particularly challenging to detect using NLP techniques, and is also a high-stakes task which requires the highest accuracy to enab

model-releasesarxiv-cs-cl
27 Jul 2026
Research

Analyzing Toxic Behavior and Its Impact on the Mastodon Community

DGX agent

arXiv:2607.21980v1 Announce Type: new Abstract: Mastodon as a decentralized federation of independently moderated social servers poses unique challenges for the detection and mitigation of toxic conte

researcharxiv-cs-cl
27 Jul 2026
Model Releases

Benchmarking Fine-tuning and Retrieval Strategies for a Multimodal Language Model on the NRC Reactor Operator Licensing Examination

DGX agent

arXiv:2607.22067v1 Announce Type: new Abstract: The integration of large language models (LLMs) into the nuclear power industry requires outputs grounded in domain-specific knowledge. This study evalu

model-releasesarxiv-cs-cl
27 Jul 2026
Applications

Biomedical Machine Translation for Low-Resource Arabic-Script Languages via Cross-Lingual Transfer and LoRA Adapter Merging

DGX agent

arXiv:2607.22300v1 Announce Type: new Abstract: We present a systematic study of healthcare-domain cross-lingual transfer to address the scarcity of biomedical NMT resources for Arabic-script language

applicationsarxiv-cs-cl
27 Jul 2026
Model Releases

Cross-Tokenizer On-Policy Distillation via Byte-Prefix Marginalization

DGX agent

arXiv:2607.22334v1 Announce Type: cross Abstract: Open-weight language models from different families exhibit complementary capabilities, motivating their consolidation into a compact student through

model-releasesarxiv-cs-cl
27 Jul 2026
Model Releases

Data Quality over Capacity: Internalizing Documents into LoRA Adapters for Closed-Book QA

DGX agent

arXiv:2607.21861v1 Announce Type: new Abstract: We study baking documents directly into the weights of a 4-bit Gemma-4-e4b model via LoRA, so a system can answer questions about a corpus closed-book:

model-releasesarxiv-cs-cl
27 Jul 2026
Model Releases

DBA-Bench: A Production-Fidelity Benchmark for LLM-Based Database Operations Agents

DGX agent

arXiv:2607.22165v1 Announce Type: cross Abstract: LLM-based database agents show promise, but differing task scopes, testbeds, and metrics hinder comparison. We identify four gaps between evaluation a

model-releasesarxiv-cs-cl
27 Jul 2026
Research

Developing and Validating the Spanish Version of the Large Language Models Dependency Scale (LLM-D12-SP)

DGX agent

arXiv:2607.22041v1 Announce Type: new Abstract: There is a growing need for reliable and culturally validated instruments to assess psychological dependency on large language models (LLMs), particular

researcharxiv-cs-cl
27 Jul 2026
Research

Diffusion Models in Medical Image Inpainting: Challenges, Solution Taxonomy, and Future Directions

DGX agent

arXiv:2607.21904v1 Announce Type: cross Abstract: Image inpainting aims to reconstruct missing or corrupted regions of an image while preserving as much as possible, visual and semantic consistency. I

researcharxiv-cs-cl
27 Jul 2026
Model Releases

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models

DGX agent

arXiv:2607.21617v1 Announce Type: cross Abstract: Vision Language Models (VLMs) are increasingly used in place of traditional OCR pipelines for document understanding. In this paper, we show they do n

model-releasesarxiv-cs-cl
27 Jul 2026
Model Releases

DWT-Fusion: A Signal-Based Framework for Training-Free LLM-Generated Text Detection

DGX agent

arXiv:2607.22026v1 Announce Type: new Abstract: Detecting LLM-generated text remains challenging under zero-shot and training-free conditions, especially when detectors must generalize across datasets

model-releasesarxiv-cs-cl
27 Jul 2026
Model Releases

Dynamic Commonsense Coordination for Empathetic Response Generation

DGX agent

arXiv:2607.22136v1 Announce Type: new Abstract: Empathetic Response Generation (ERG) requires models to recognize users' emotions and generate empathetic responses. Commonsense knowledge has been show

model-releasesarxiv-cs-cl
27 Jul 2026
Model Releases

Enjoy Your Talk: A Human-Centered Benchmark for Multi-Turn Dialogue with Decoupled User Simulation, Target Modeling, and Judging

DGX agent

arXiv:2607.10428v2 Announce Type: replace Abstract: Evaluating large language models (LLMs) as multi-turn conversational partners requires probing capabilities that single-turn benchmarks miss: person

model-releasesarxiv-cs-cl
27 Jul 2026
Model Releases

Enough is as good as a feast: A Comprehensive Analysis of How Reinforcement Learning Mitigates Task Conflicts in LLMs

DGX agent

arXiv:2607.22039v1 Announce Type: new Abstract: Model merging plays a crucial role in consolidating multiple specialized models into a single, unified model, especially in the era of large language mo

model-releasesarxiv-cs-cl
27 Jul 2026
Model Releases

Evaluation design conditions the expert-vs-auto MeSH gap: a controlled comparison of bag-of-words and BiomedBERT on the Cohen benchmark

DGX agent

arXiv:2607.21685v1 Announce Type: new Abstract: A systematic review begins with someone reading thousands of abstracts to identify the few that are relevant, and classifiers are used to prioritise tha

model-releasesarxiv-cs-cl
27 Jul 2026
Research

From Isolated Tasks to Structured Capabilities: A Multilayer Taxonomy for Large Language Models

DGX agent

arXiv:2607.22182v1 Announce Type: new Abstract: Large language model (LLM) evaluation spans diverse tasks and benchmarks, yet evidence remains organized around tasks rather than the capabilities they

researcharxiv-cs-cl
27 Jul 2026
Safety

From Obligation to Specification: A Survey on Validating EU AI Act Requirements in RE

DGX agent

arXiv:2607.21608v1 Announce Type: cross Abstract: With the EU AI Act entering into force, organizations developing or operating AI systems face new obligations on transparency, risk management, and tr

safetyarxiv-cs-cl
27 Jul 2026
Local Ai

From Seasonality to Semantics: Benchmarking a Hybrid Probabilistic Forecasting System for Roadblocks in Bolivia

DGX agent

arXiv:2607.21785v1 Announce Type: cross Abstract: Roadblocks in Bolivia are a social conflict phenomenon with devastating economic impacts, estimated at losses equivalent to 4% of the national Gross D

local-aiarxiv-cs-cl
27 Jul 2026
Tutorials

FSE: Continual Learning for Named Entity Recognition by Fast-Slow Experts

DGX agent

arXiv:2607.22075v1 Announce Type: new Abstract: Continual Learning for Named Entity Recognition (CLNER) enable models to incrementally learn new entity types without forgetting previously acquired one

tutorialsarxiv-cs-cl
27 Jul 2026
Applications

grapheme-kit: Grapheme-Level Metrics and Text Processing for Multilingual NLP

DGX agent

arXiv:2607.22456v1 Announce Type: new Abstract: Existing lexical distance, similarity, and evaluation metrics operate on Unicode code points, which can misrepresent errors in writing systems where a s

applicationsarxiv-cs-cl
27 Jul 2026
← Previous
1…2324252627…161
Next →