AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,759 results
Model Releases

DASH: Fast Differentiable Architecture Search for Hybrid Attention in Minutes on a Single GPU

DGX agent

arXiv:2605.20936v1 Announce Type: cross Abstract: Hybrid attention architectures are becoming an increasingly important paradigm for improving LLM inference efficiency while preserving model quality,

model-releasesarxiv-cs-cl
21 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Data Scaling as Progressive Coverage of a Predictive Contribution Spectrum

DGX agent

arXiv:2605.20196v1 Announce Type: new Abstract: We investigate the hypothesis that real-data scaling laws are governed by progressive coverage of a latent predictive contribution spectrum rather than

researcharxiv-cs-cl
21 May 2026
Model Releases

DEL: Digit Entropy Loss for Numerical Learning of Large Language Models

DGX agent

arXiv:2605.20369v1 Announce Type: new Abstract: Number prediction stands as a fundamental capability of large language models (LLMs) in mathematical problem-solving and code generation. The widely ado

model-releasesarxiv-cs-cl
21 May 2026
Safety

DelTA: Discriminative Token Credit Assignment for Reinforcement Learning from Verifiable Rewards

DGX agent

arXiv:2605.21467v1 Announce Type: cross Abstract: Reinforcement learning from verifiable rewards (RLVR) has emerged as a central technique for improving the reasoning capabilities of large language mo

safetyarxiv-cs-cl
21 May 2026
Research

Dictionary Insertion Prompting for Multilingual Reasoning on Multilingual Large Language Models

DGX agent

arXiv:2411.01141v2 Announce Type: replace Abstract: There are two shortages in the current Large Language Models (LLMs) era. The first is short of multilingual models, where most LLMs are English-cent

researcharxiv-cs-cl
21 May 2026
Model Releases

DiMextsuperscript{3}: Bridging Multilingual and Multimodal Models via Direction- and Magnitude-Aware Merging

DGX agent

arXiv:2605.12960v2 Announce Type: replace Abstract: Towards more general and human-like intelligence, large language models should seamlessly integrate both multilingual and multimodal capabilities; h

model-releasesarxiv-cs-cl
21 May 2026
Research

Direct Translation between Sign Languages

DGX agent

arXiv:2605.20588v1 Announce Type: new Abstract: The field of sign language translation has witnessed significant progress in the translation between sign and spoken languages, but the translation betw

researcharxiv-cs-cl
21 May 2026
Safety

Distribution-Aware Reward: Reinforcement Learning over Predictive Distributions for LLM Regression

DGX agent

arXiv:2605.20740v1 Announce Type: cross Abstract: Large language models can predict real-valued quantities from heterogeneous inputs such as text, code, and molecular strings, but most training object

safetyarxiv-cs-cl
21 May 2026
Safety

Distributional Alignment as a Criterion for Designing Task Vectors in In-Context Learning

DGX agent

arXiv:2605.20730v1 Announce Type: new Abstract: In-context learning (ICL) allows large language models (LLMs) to adapt to new tasks through demonstrations, yet it suffers from escalating inference cos

safetyarxiv-cs-cl
21 May 2026
Model Releases

DIVE: Embedding Compression via Self-Limiting Gradient Updates

DGX agent

arXiv:2605.20689v1 Announce Type: new Abstract: High-dimensional embeddings from large language models impose significant storage and computational costs on vector search systems. Recent embedding com

model-releasesarxiv-cs-cl
21 May 2026
Safety

Divide-Prompt-Refine: a Training-Free, Structure-Aware Framework for Biomedical Abstract Generation

DGX agent

arXiv:2605.20628v1 Announce Type: new Abstract: Biomedical abstracts play a critical role in downstream NLP applications, such as information retrieval, biocuration, and biomedical knowledge discovery

safetyarxiv-cs-cl
21 May 2026
Local Ai

DNACHUNKER: Learnable Tokenization for DNA Language Models

DGX agent

arXiv:2601.03019v4 Announce Type: replace-cross Abstract: DNA language models are increasingly used to represent genomic sequence, yet their effectiveness depends critically on how raw nucleotides are

local-aiarxiv-cs-cl
21 May 2026
Research

Do as I Say, Not as I Do: Instruction-Induction Conflict in LLMs

DGX agent

arXiv:2605.20382v1 Announce Type: new Abstract: Language models are trained to follow instructions, but they are also powerful pattern completers. What happens when these two objectives conflict? We c

researcharxiv-cs-cl
21 May 2026
Model Releases

Do LLMs Know What Luxembourgish Borrows? Probing Lexical Neology in Low-Resource Multilingual Models

DGX agent

arXiv:2605.21227v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for writing assistance in small contact languages, yet it is unclear whether they respect community n

model-releasesarxiv-cs-cl
21 May 2026
Safety

Do No Harm? Hallucination and Actor-Level Abuse in Web-Deployed Medical Large Language Models

DGX agent

arXiv:2605.20591v1 Announce Type: new Abstract: Medical large language models (LLMs), including custom medical GPTs (MedGPTs) and open-source models, are increasingly deployed on web platforms to prov

safetyarxiv-cs-cl
21 May 2026
Agents

Draw2Think: Harnessing Geometry Reasoning through Constraint Engine Interaction

DGX agent

arXiv:2605.20743v1 Announce Type: cross Abstract: Vision-language models solve geometry problems with rising accuracy, yet their intermediate states remain latent and unverifiable: a relation expresse

agentsarxiv-cs-cl
21 May 2026
Model Releases

DrugRAG: Enhancing Pharmacy LLM Performance Through A Novel Retrieval-Augmented Generation Pipeline

DGX agent

arXiv:2512.14896v2 Announce Type: replace Abstract: In our study, we evaluated large language model (LLM) performance on pharmacy licensure-style question-answering tasks and developed an external kno

model-releasesarxiv-cs-cl
21 May 2026
Research

END: Early Noise Dropping for Efficient and Effective Context Denoising

DGX agent

arXiv:2502.18915v3 Announce Type: replace Abstract: Large Language Models (LLMs) have demonstrated remarkable performance across a wide range of natural language processing tasks. However, they are of

researcharxiv-cs-cl
21 May 2026
Research

Enhancing Scientific Discourse: Machine Translation for the Scientific Domain

DGX agent

arXiv:2605.20912v1 Announce Type: new Abstract: The increasing volume of scientific research necessitates effective communication across language barriers. Machine translation (MT) offers a promising

researcharxiv-cs-cl
21 May 2026
Safety

Enhancing Speech Large Language Models through Reinforced Behavior Alignment

DGX agent

arXiv:2509.03526v2 Announce Type: replace Abstract: The recent advancements of Large Language Models (LLMs) have spurred considerable research interest in extending their linguistic capabilities beyon

safetyarxiv-cs-cl
21 May 2026
Research

EpiCache: Episodic KV Cache Management for Long-Term Conversation on Resource-Constrained Environments

DGX agent

arXiv:2509.17396v4 Announce Type: replace Abstract: Modern large language models (LLMs) extend context lengths to millions of tokens, enabling coherent, personalized responses grounded in long convers

researcharxiv-cs-cl
21 May 2026
Safety

EvalMORAAL: Interpretable Chain-of-Thought and LLM-as-Judge Evaluation for Moral Alignment in Large Language Models

DGX agent

arXiv:2510.05942v3 Announce Type: replace Abstract: We present EvalMORAAL, a transparent chain-of-thought (CoT) framework that uses two scoring methods (log-probabilities and direct ratings) plus a mo

safetyarxiv-cs-cl
21 May 2026
Applications

Evaluating Speech Articulation Synthesis with Articulatory Phoneme Recognition

DGX agent

arXiv:2605.20920v1 Announce Type: new Abstract: Recent advances in machine learning and the availability of articulatory datasets allow vocal tract synthesis to be conditioned on phonetic sequences, a

applicationsarxiv-cs-cl
21 May 2026
Local Ai

Explainable AI: Context-Aware Layer-Wise Integrated Gradients for Explaining Transformer Models

DGX agent

arXiv:2602.16608v2 Announce Type: replace Abstract: Transformer models achieve state-of-the-art performance across domains and tasks, yet their deeply layered representations make their predictions di

local-aiarxiv-cs-cl
21 May 2026
Model Releases

Findings of the Counter Turing Test: AI-Generated Text Detection

DGX agent

arXiv:2605.20761v1 Announce Type: new Abstract: The rapid proliferation of AI-generated text has introduced significant challenges in maintaining the integrity of digital content. Advanced generative

model-releasesarxiv-cs-cl
21 May 2026
Research

Findings of the Fifth Shared Task on Multilingual Coreference Resolution: Expanding Datasets for Long-Range Entities

DGX agent

arXiv:2605.21369v1 Announce Type: new Abstract: This paper describes the fifth edition of the Shared Task on Multilingual Coreference Resolution, held in conjunction with the CODI-CRAC 2026 workshop.

researcharxiv-cs-cl
21 May 2026
Model Releases

Fine-grained Claim-level RAG Benchmark for Law

DGX agent

arXiv:2605.21071v1 Announce Type: new Abstract: The rapid progress of large language models (LLMs) is shifting semantic search toward a question-answering paradigm, where users ask questions and LLMs

model-releasesarxiv-cs-cl
21 May 2026
Research

Flow Map Language Models: One-step Language Modeling via Continuous Denoising

DGX agent

arXiv:2602.16813v3 Announce Type: replace Abstract: Language models based on discrete diffusion have attracted widespread interest for their potential to provide faster generation than autoregressive

researcharxiv-cs-cl
21 May 2026
Tutorials

FlowLM: Few-Step Language Modeling via Diffusion-to-Flow Adaptation

DGX agent

arXiv:2605.20199v1 Announce Type: new Abstract: We present FlowLM, a flow matching language model transformed from pre-trained diffusion language models via efficient fine-tuning. By re-aligning the c

tutorialsarxiv-cs-cl
21 May 2026
Model Releases

Gated Normalization Removal and Scale Anchoring in Pre-Norm Transformers

DGX agent

arXiv:2602.10408v2 Announce Type: replace-cross Abstract: Normalization layers are standard in transformers, but it is not clear whether their sample-dependent computations are necessary throughout bo

model-releasesarxiv-cs-cl
21 May 2026
Research

Generative AI Practices, Literacy, and Divides: An Empirical Analysis in the Italian Context

DGX agent

arXiv:2512.03671v2 Announce Type: replace Abstract: The rise of generative AI (GenAI) chatbots accessible via conversational interfaces is transforming digital interactions and holds economic promise.

researcharxiv-cs-cl
21 May 2026
Model Releases

Geometry-Lite: Interpretable Safety Probing via Layer-Wise Margin Geometry

DGX agent

arXiv:2605.20241v1 Announce Type: cross Abstract: Prompt-level safety probes for large language models use hidden-state representations to separate safe from unsafe prompts, but strong average detecti

model-releasesarxiv-cs-cl
21 May 2026
Applications

GradeLegal: Automated Grading for German Legal Cases

DGX agent

arXiv:2605.21076v1 Announce Type: new Abstract: Grading German legal exam solutions faces growing volumes and a shortage of qualified graders, delaying feedback and creating a bottleneck. At the same

applicationsarxiv-cs-cl
21 May 2026
Model Releases

GraphRAG on Consumer Hardware: Benchmarking Local LLMs for Healthcare EHR Schema Retrieval

DGX agent

arXiv:2605.20815v1 Announce Type: new Abstract: Graph-based Retrieval Augmented Generation (GraphRAG) extends retrieval-augmented generation to support structured reasoning over complex corpora, but i

model-releasesarxiv-cs-cl
21 May 2026
Research

Hiding in Plain Sight: Finding MAHA on Reddit

DGX agent

arXiv:2605.20435v1 Announce Type: cross Abstract: Make America Healthy Again (MAHA) is a national health movement that encompasses a striking mix of beliefs, from broadly accepted concerns about good

researcharxiv-cs-cl
21 May 2026
Research

How Open Must Language Models be to Enable Reliable Scientific Inference?

DGX agent

arXiv:2603.26539v2 Announce Type: replace Abstract: How does the extent to which a model is open or closed impact the scientific inferences that can be drawn from research that involves it? In this pa

researcharxiv-cs-cl
21 May 2026
Model Releases

HRM-Text: Efficient Pretraining Beyond Scaling

DGX agent

arXiv:2605.20613v1 Announce Type: new Abstract: The current pretraining paradigm for large language models relies on massive compute and internet-scale raw text, creating a significant barrier to foun

model-releasesarxiv-cs-cl
21 May 2026
Applications

'I didn't Make the Micro Decisions': Measuring, Inducing, and Exposing Goal-Level AI Contributions in Collaboration

DGX agent

arXiv:2605.21363v1 Announce Type: new Abstract: As large language models (LLMs) increasingly shape how users form, refine, and extend their goals, attributing contributions in human-AI collaboration b

applicationsarxiv-cs-cl
21 May 2026
Model Releases

Improving Quantized Model Performance in Qualitative Analysis with Multi-Pass Prompt Verification

DGX agent

arXiv:2605.20193v1 Announce Type: new Abstract: Quantized Large Language Models (LLMs) are used more often in qualitative analysis because they run fast and need fewer computing resources. This study

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

InternBootcamp Technical Report: Boosting LLM Reasoning with Verifiable Task Scaling

DGX agent

arXiv:2508.08636v2 Announce Type: replace Abstract: Large language models (LLMs) have revolutionized artificial intelligence by enabling complex reasoning capabilities. While recent advancements in re

model-releasesarxiv-cs-cl
21 May 2026
Research

Interpretable Discriminative Text Representations via Agreement and Label Disentanglement

DGX agent

arXiv:2605.20693v1 Announce Type: new Abstract: Interpretable text representations should expose coordinates that are not only predictive, but also meaningful enough for independent auditors to apply.

researcharxiv-cs-cl
21 May 2026
Research

iReasoner: Trajectory-Aware Intrinsic Reasoning Supervision for Self-Evolving Large Multimodal Models

DGX agent

arXiv:2601.05877v3 Announce Type: replace Abstract: Recent work shows that large multimodal models (LMMs) can self-improve from unlabeled data via self-play and intrinsic feedback. Yet existing self-e

researcharxiv-cs-cl
21 May 2026
Research

Iterative LLM-based improvement for French Clinical Interview Transcription and Speaker Diarization

DGX agent

arXiv:2603.00086v2 Announce Type: replace Abstract: Automatic speech recognition for French medical conversations remains challenging, with word error rates often exceeding 30% in spontaneous clinical

researcharxiv-cs-cl
21 May 2026
Model Releases

JobArabi: An Arabic Corpus and Analysis of Job Announcements from Social Media

DGX agent

arXiv:2605.20960v1 Announce Type: new Abstract: This paper introduces JobArabi, a large-scale corpus of Arabic job announcements collected from social media between January 2024 and October 2025. The

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

LamPO: A Lambda Style Policy Optimization for Reasoning Language Models

DGX agent

arXiv:2605.21235v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has become an effective paradigm for improving reasoning language models on tasks such as mathemat

model-releasesarxiv-cs-cl
21 May 2026
Research

Large Language Models Unpack Complex Political Opinions through Target-Stance Extraction

DGX agent

arXiv:2603.23531v2 Announce Type: replace Abstract: Political polarization emerges from a complex interplay of beliefs about policies, figures, and issues. However, most computational analyses reduce

researcharxiv-cs-cl
21 May 2026
Safety

LASH: Adaptive Semantic Hybridization for Black-Box Jailbreaking of Large Language Models

DGX agent

arXiv:2605.21362v1 Announce Type: new Abstract: Jailbreak attacks expose a persistent gap between the intended safety behavior of aligned large language models and their behavior under adversarial pro

safetyarxiv-cs-cl
21 May 2026
Model Releases

Lean Refactor: Multi-Objective Controllable Proof Optimization via Agentic Strategy Search

DGX agent

arXiv:2605.20244v1 Announce Type: cross Abstract: We present Lean Refactor, a plug-and-play retrieval-augmented agentic framework for multi-objective, controllable, and version-robust refactoring of L

model-releasesarxiv-cs-cl
21 May 2026
← Previous
1…8485868788…162
Next →