AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,646 results
Research

Robust and Secure Code Watermarking for Large Language Models via ML/Crypto Codesign

DGX agent

arXiv:2502.02068v3 Announce Type: replace-cross Abstract: This paper introduces RoSeMary, the first-of-its-kind ML/Crypto codesign watermarking framework that regulates LLM-generated code to avoid int

researcharxiv-cs-cl
12 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Seeds Before Objectives: Rethinking Evaluation for Low-Resource Garhwali ASR

DGX agent

arXiv:2608.10670v1 Announce Type: new Abstract: At corpus sizes typical of low-resource dialects, single-run comparisons can yield gains that do not replicate. We show this for Garhwali, an under-reso

model-releasesarxiv-cs-cl
12 Aug 2026
Applications

Self-Knowledge Retrieval Augmented Generation Framework for Patent Matching

DGX agent

arXiv:2608.11030v1 Announce Type: cross Abstract: Patent retrieval and matching based on large language models (LLMs) play a vital role in intellectual property protection. However, due to the complex

applicationsarxiv-cs-cl
12 Aug 2026
Research

Share First, Route What Remains: A Unified Framework for Token-Adaptive MoE Computation

DGX agent

arXiv:2608.10392v1 Announce Type: cross Abstract: Mixture-of-experts (MoE) models have recently moved beyond routing a fixed number of complete experts. Shared-expert designs preserve reusable knowled

researcharxiv-cs-cl
12 Aug 2026
Research

Simplex Relaxation for Discrete Diffusion

DGX agent

arXiv:2608.10615v1 Announce Type: new Abstract: Discrete diffusion models for categorical generation are defined by a corruption kernel, which determines the intermediate state space and the associate

researcharxiv-cs-cl
12 Aug 2026
Research

StreamFlow: Dynamic Memory Flows for Streaming Video Understanding

DGX agent

arXiv:2608.10949v1 Announce Type: cross Abstract: Streaming video understanding requires multimodal large language models (MLLMs) to preserve relevant evidence from continuously evolving streams under

researcharxiv-cs-cl
12 Aug 2026
Research

TEAMMix: Taxonomy Enrichment Augmentation and Minority-augmented Mixing Strategy for LLM-enhanced Weak-Supervised Hierarchical Text Classification

DGX agent

arXiv:2608.11044v1 Announce Type: new Abstract: Hierarchical Text Classification (HTC), as a critical text mining task, faces challenges such as complex label hierarchies and class imbalance. Existing

researcharxiv-cs-cl
12 Aug 2026
Model Releases

TemMed-Bench: Evaluating Temporal Medical Image Reasoning in Vision-Language Models

DGX agent

arXiv:2509.25143v2 Announce Type: replace-cross Abstract: Existing medical reasoning benchmarks for vision-language models primarily focus on analyzing a patient's condition based on an image from a s

model-releasesarxiv-cs-cl
12 Aug 2026
Safety

Templated or fully Synthetic? Prompt construction as a confound in measuring LLM political stance beyond writing assistance

DGX agent

arXiv:2608.11008v1 Announce Type: new Abstract: Political stance detection in LLMs has long been dominated by closed-ended, multiple-choice political survey questions---originally designed for humans,

safetyarxiv-cs-cl
12 Aug 2026
Safety

The Hidden Puppet Master: Predicting Human Belief Change in Manipulative LLM Dialogues

DGX agent

arXiv:2603.20907v5 Announce Type: replace Abstract: As users increasingly turn to LLMs for practical and personal advice, they become vulnerable to subtle steering toward hidden incentives misaligned

safetyarxiv-cs-cl
12 Aug 2026
Local Ai

The Illusion of Cross-Lingual Safety in Low-Resource Languages

DGX agent

arXiv:2608.11146v1 Announce Type: new Abstract: Safety alignment in large language models (LLMs) is largely developed in English, assuming these safeguards generalize across multilingual settings. How

local-aiarxiv-cs-cl
12 Aug 2026
Model Releases

The Multilingual Quantization Tax: Structural Collapse and Typological Fragility in Edge SLMs

DGX agent

arXiv:2608.09941v1 Announce Type: new Abstract: While 4-bit weight quantization is critical for deploying Small Language Models (SLMs) on edge devices, evaluations of the resulting performance degrada

model-releasesarxiv-cs-cl
12 Aug 2026
Safety

The Parser Already Knows: Lightweight Bias Correction in Constrained Decoding

DGX agent

arXiv:2608.10137v1 Announce Type: new Abstract: Grammar Constrained Decoding (GCD) forces Language Models (LMs) to produce syntactically valid outputs by masking out non-conforming tokens at each step

safetyarxiv-cs-cl
12 Aug 2026
Agents

The Signal Rail: A Deterministic Motion Grammar for Communicating Conversational Agent State in Terminal Interfaces

DGX agent

arXiv:2608.10689v1 Announce Type: cross Abstract: Terminal interfaces to conversational agents report rich internal state (listening, thinking, executing tools, awaiting input, failing) almost entirel

agentsarxiv-cs-cl
12 Aug 2026
Model Releases

UT-ACA: Uncertainty-Triggered Adaptive Context Allocation for Long-Context Inference

DGX agent

arXiv:2603.18446v2 Announce Type: replace Abstract: Long-context inference remains challenging for large language models due to attention dilution and out-of-distribution degradation. Context selectio

model-releasesarxiv-cs-cl
12 Aug 2026
Model Releases

VisEditBench: Can Vision-Language Models Edit Visualization Code from Multimodal Feedback?

DGX agent

arXiv:2608.10408v1 Announce Type: new Abstract: Vision-language models (VLMs) have shown strong capabilities in generating visualization code from textual or visual specifications. However, real-world

model-releasesarxiv-cs-cl
12 Aug 2026
Model Releases

VoxSumm: A Multilingual Corpus of Long-Form Spoken News for Joint Summarization and Translation

DGX agent

arXiv:2608.10359v1 Announce Type: cross Abstract: As information increasingly traverses linguistic boundaries, users require concise cross-lingual representations of long-form content. Nevertheless, l

model-releasesarxiv-cs-cl
12 Aug 2026
Agents

What Iterated Self-Feeding Probes of Language Models Measure, and a test that separates the construction from the model

DGX agent

arXiv:2608.10986v1 Announce Type: new Abstract: A growing class of methods probes a language model by feeding it its own output: self-consistency, iterated refinement, agentic loops. We ask what such

agentsarxiv-cs-cl
12 Aug 2026
Research

When Is a General Factor Distinguishable? Non-Proportionality, Stable Structure, and the Bifactor Decision

DGX agent

arXiv:2608.10731v1 Announce Type: cross Abstract: Whether an additional general dimension is necessary beyond correlated first-order factors is a property of the population covariance matrix, not of a

researcharxiv-cs-cl
12 Aug 2026
Applications

Who Gets Heeded? An Obligation-Level Audit of Responsiveness in EPA Rulemaking

DGX agent

arXiv:2608.10329v1 Announce Type: cross Abstract: Notice-and-comment rulemaking gives any affected party the same formal right to influence federal regulation, but formal access is not substantive cap

applicationsarxiv-cs-cl
12 Aug 2026
Research

X2-Turn: Frame-Synchronous Dual-Head Modeling for Joint Streaming ASR and Turn State Prediction

DGX agent

arXiv:2608.10878v1 Announce Type: new Abstract: Accurate and responsive turn-taking is essential for spoken dialogue systems, which must distinguish in real time between user interruptions, backchanne

researcharxiv-cs-cl
12 Aug 2026
Applications

Accurate but Natural? Diagnosing Grammatical and Idiomatic Gaps in Japanese EFL Writing

DGX agent

arXiv:2608.09289v1 Announce Type: new Abstract: Second language writing research distinguishes grammatical accuracy from native-like idiomaticity, yet automated writing evaluation often conflates thes

applicationsarxiv-cs-cl
11 Aug 2026
Research

Accurate Ensembles, Fragile Narratives: Multi-Scale Stacking and a Fidelity Audit of LLM-Generated Explanations for Credit Risk

DGX agent

arXiv:2608.08126v1 Announce Type: cross Abstract: Credit scoring increasingly relies on models whose decision logic cannot be read off their parameters, in tension with supervisory expectations that a

researcharxiv-cs-cl
11 Aug 2026
Safety

An Agentic Generative Large Language Model for Treatment Planning of Colorectal Cancer

DGX agent

arXiv:2608.09142v1 Announce Type: new Abstract: Treatment planning in precision oncology requires synthesizing heterogeneous patient information with rapidly evolving clinical guidelines to ensure gui

safetyarxiv-cs-cl
11 Aug 2026
Research

AnchorFold: A Focus-Then-Fold Framework via Recursive Attention Propagation for Efficient Multi-Vector Visual Document Retrieval

DGX agent

arXiv:2608.08732v1 Announce Type: cross Abstract: Multi-vector vision-language retrievers enable fine-grained Visual Document Retrieval (VDR) through late interaction, but storing and scoring hundreds

researcharxiv-cs-cl
11 Aug 2026
Model Releases

APEX-VW: A Document-Level English-Spanish Post-Editing Dataset in the Healthcare Domain

DGX agent

arXiv:2608.08059v1 Announce Type: new Abstract: Post-Editing (PE) of Machine Translation (MT) output often involves repeating the same lexical and terminological corrections across many segments, espe

model-releasesarxiv-cs-cl
11 Aug 2026
Hardware

AraSSM: A bidirectional state-space encoder for Arabic masked language modeling

DGX agent

arXiv:2608.08256v1 Announce Type: new Abstract: Pretrained Transformer encoders such as AraBERT, MARBERT, and CAMeLBERT have become the standard backbone for Arabic natural language understanding, but

hardwarearxiv-cs-cl
11 Aug 2026
Research

Archer: Adaptive Reuse of Cached Hidden States for Efficient Rollback in Diffusion Language Models

DGX agent

arXiv:2608.08086v1 Announce Type: new Abstract: Diffusion language models (DLMs) iteratively refine a sequence, allowing earlier predictions to be revised as context evolves. This rollback capability

researcharxiv-cs-cl
11 Aug 2026
Model Releases

ATLAS: Agentic Taxonomy of Large-Scale Software Ecosystems

DGX agent

arXiv:2606.21597v2 Announce Type: replace-cross Abstract: The open-source ecosystem on GitHub lacks a systematic hierarchical taxonomy of software repositories. GitHub Topics, the dominant organizatio

model-releasesarxiv-cs-cl
11 Aug 2026
Safety

Beyond cognacy

DGX agent

arXiv:2507.03005v3 Announce Type: replace Abstract: Computational phylogenetics has become an established tool in historical linguistics, with many language families now analyzed using likelihood-base

safetyarxiv-cs-cl
11 Aug 2026
Model Releases

Beyond Direct Identifiers: Probabilistic Privacy Risk Estimation for Privacy-Conscious LLM Query Delegation

DGX agent

arXiv:2608.09140v1 Announce Type: cross Abstract: Recent work on protecting privacy during user-LLM interactions often focuses on direct, explicit identifiers: the personally-identifiable information

model-releasesarxiv-cs-cl
11 Aug 2026
Model Releases

Beyond the Capability Boundary: Zeroth-Order Optimization for Self-Evolving LLM Agents

DGX agent

arXiv:2608.09292v1 Announce Type: cross Abstract: Self-evolving methods improve the capabilities of LLM agents by sampling trajectories from the underlying LLMs and learning from these trajectories. H

model-releasesarxiv-cs-cl
11 Aug 2026
Model Releases

BibTeX Citation Errors in Scientific Publishing Agents: Evaluation and Mitigation

DGX agent

arXiv:2604.03159v2 Announce Type: replace-cross Abstract: Large language models with web search are increasingly used in scientific publishing agents, yet they produce BibTeX entries with pervasive fi

model-releasesarxiv-cs-cl
11 Aug 2026
Research

Commitment Before Realization: When Classifier-Free Guidance Becomes Unnecessary in Masked Diffusion Language Models

DGX agent

arXiv:2608.08082v1 Announce Type: new Abstract: Classifier-free guidance (CFG) is usually kept on throughout masked diffusion language model decoding, although its benefit varies across prompts and ov

researcharxiv-cs-cl
11 Aug 2026
Research

Comparing British and American Audio Description of Movies

DGX agent

arXiv:2608.09792v1 Announce Type: new Abstract: Narrating the visual component of movies is known as audio description. It is a narrative technique designed to enable blind and visually impaired indiv

researcharxiv-cs-cl
11 Aug 2026
Applications

Consilience for Verifier-Free Test-Time Scaling

DGX agent

arXiv:2608.09898v1 Announce Type: new Abstract: Test-time scaling often uses an external verifier, such as compilers and test cases in coding or trained value functions in robotics applications, to ob

applicationsarxiv-cs-cl
11 Aug 2026
Model Releases

Conversation as Measurement in Clinical Encounters: Observable Phase Structure, Partially Observable Patient State

DGX agent

arXiv:2608.08868v1 Announce Type: new Abstract: Many modern AI systems analyze conversational traces to infer aspects of human interaction and state, implicitly assuming that such information is recov

model-releasesarxiv-cs-cl
11 Aug 2026
Research

Data Repetition Beats Data Scaling in Long-CoT Supervised Fine-Tuning

DGX agent

arXiv:2602.11149v2 Announce Type: replace Abstract: Supervised fine-tuning (SFT) on chain-of-thought data is an essential post-training step for reasoning language models. Standard machine learning in

researcharxiv-cs-cl
11 Aug 2026
Model Releases

Decoding-Level Taboo: A Diagnostic Stress Test for LLM Robustness

DGX agent

arXiv:2608.09900v1 Announce Type: new Abstract: Large language model evaluations typically focus on performance under nominal conditions, creating an illusion of capability where models comfortably wa

model-releasesarxiv-cs-cl
11 Aug 2026
Hardware

Depth-adaptive Inference of Looped Language Models via Continuous Depth Batching

DGX agent

arXiv:2608.09444v1 Announce Type: cross Abstract: A main promise of looped language models (LMs) is depth-adaptive inference. By iterating a block of shared layers a variable number of times, the mode

hardwarearxiv-cs-cl
11 Aug 2026
Research

Detection of Self-Introductions in Legislative Testimony

DGX agent

arXiv:2608.07891v1 Announce Type: new Abstract: Self-introductions are common in legislative committee testimonies. Successfully detecting them and extracting the speaker's name is enormously helpful

researcharxiv-cs-cl
11 Aug 2026
Model Releases

DevIntent: How Much Does LLM-Generated Code Violate Developer Intent?

DGX agent

arXiv:2608.07614v1 Announce Type: cross Abstract: Code generated by LLMs can violate a developer's implicit intentions when given an ambiguous prompt, yet standard benchmarks measure only whether code

model-releasesarxiv-cs-cl
11 Aug 2026
Safety

Discovering and Causally Validating Emotion-Sensitive Neurons in Large Audio-Language Models

DGX agent

arXiv:2601.03115v2 Announce Type: replace Abstract: Emotion is a central dimension of spoken communication, yet, we still lack a mechanistic account of how modern large audio-language models (LALMs) e

safetyarxiv-cs-cl
11 Aug 2026
Research

DS@GT ARC at Touche: Large Language Models for Retrieval-Augmented Debate

DGX agent

arXiv:2608.08143v1 Announce Type: cross Abstract: We extend the DS@GT ARC working-note submission to the Touche 2025 Retrieval-Augmented Debate task. The task has two subtasks: generating the next utt

researcharxiv-cs-cl
11 Aug 2026
Model Releases

ELICITED: EHR-grounded Longitudinal Interactive Conversations for Information-seeking Triage Evaluation and Decision-making

DGX agent

arXiv:2608.09024v1 Announce Type: new Abstract: Emergency-department (ED) triage requires clinicians to rapidly identify patients who need immediate attention, determine who can safely wait, and prior

model-releasesarxiv-cs-cl
11 Aug 2026
Applications

Embedding Initialization for Unseen Low-resource Languages in Multilingual NMT: A Case Study on Limbum-English Translation

DGX agent

arXiv:2608.07629v1 Announce Type: new Abstract: Multilingual neural machine translation models such as NLLB-200 cover 200 languages but leave thousands unsupported, including most Grassfields Bantu la

applicationsarxiv-cs-cl
11 Aug 2026
Model Releases

EmoS: A Theory-Grounded Framework for Evaluating and Aligning Emotional Intelligence in Spoken Language Models

DGX agent

arXiv:2608.09189v1 Announce Type: new Abstract: Despite significant advances in instruction-following and auditory comprehension, the evaluation of Emotional Intelligence (EI) in Spoken Language Model

model-releasesarxiv-cs-cl
11 Aug 2026
Research

EvalConvoLearn: An Open-Source Framework for Evaluating Grounded Learner Simulations in Tutoring Conversations

DGX agent

arXiv:2608.07497v1 Announce Type: cross Abstract: Conversational learner simulations are valuable tools for testing learning theories, evaluating instructional materials and automated tutors, or power

researcharxiv-cs-cl
11 Aug 2026
← Previous
1234…160
Next →