AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlog
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,759 results
Research

When Helpful Context Leaks: Privacy Risks in Domain-Adapted ASR

DGX agent

arXiv:2605.28211v1 Announce Type: new Abstract: SpeechLLMs are increasingly deployed in professional settings where domain customisation is standard practice: users supply context in prompts with sens

researcharxiv-cs-cl
28 May 2026
Research
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

When Seekers Are Hard to Help: Evaluating Emotional Support Dialogue Systems in Worst-Case Interactions

DGX agent

arXiv:2605.28228v1 Announce Type: new Abstract: Emotional Support Dialogue Systems (ESDSes) are increasingly evaluated and trained with LLM-simulated seekers. However, such simulated seekers often beh

researcharxiv-cs-cl
28 May 2026
Research

Which Heads Matter for Reasoning? RL-Guided KV Cache Compression

DGX agent

arXiv:2510.08525v3 Announce Type: replace Abstract: Reasoning large language models exhibit complex reasoning behaviors via extended chain-of-thought generation that are highly fragile to information

researcharxiv-cs-cl
28 May 2026
Tutorials

Why Gaussian Diffusion Models Fail on Discrete Data and How to Prevent It?

DGX agent

arXiv:2604.02028v2 Announce Type: replace Abstract: Diffusion models have become a standard approach for generative modeling in continuous domains, yet their application to discrete data remains chall

tutorialsarxiv-cs-cl
28 May 2026
Research

Why We Need Speech to Evaluate Speech Translation

DGX agent

arXiv:2605.28227v1 Announce Type: new Abstract: Speech translation models are increasingly capable of preserving speech-specific information (e.g., speaker gender, prosody, and emphasis), yet evaluati

researcharxiv-cs-cl
28 May 2026
Safety

xKV: Cross-Layer KV-Cache Compression via Aligned Singular Vector Extraction

DGX agent

arXiv:2503.18893v2 Announce Type: replace Abstract: Long-context Large Language Models (LLMs) enable powerful applications but incur high memory costs due to the key-value states (KV-Cache). Recent st

safetyarxiv-cs-cl
28 May 2026
Model Releases

You Only Align Once: Propagating Cooperative Behaviors in Multi-Agent Systems through Seed Agents

DGX agent

arXiv:2605.27586v1 Announce Type: cross Abstract: Ensuring agent behaviors in distributed open multi-agent systems remains challenging, especially as populations grow and unaligned agents may exist. W

model-releasesarxiv-cs-cl
28 May 2026
Research

A Method for Learning Large-Scale Computational Construction Grammars from Semantically Annotated Corpora

DGX agent

arXiv:2603.12754v2 Announce Type: replace Abstract: We present a method for learning large-scale, broad-coverage construction grammars from corpora of language use. Starting from utterances annotated

researcharxiv-cs-cl
27 May 2026
Research

Accountable Human-AI Deliberation with LLMs: Scaling Collective Intelligence through Symbiotic Scaffolding

DGX agent

arXiv:2605.26940v1 Announce Type: new Abstract: Large language models (LLMs) can support democratic deliberation at scales previously constrained by turn-taking and facilitation bandwidth. Recent work

researcharxiv-cs-cl
27 May 2026
Model Releases

AdaSD: Adaptive Speculative Decoding for Efficient Language Model Inference

DGX agent

arXiv:2512.11280v2 Announce Type: replace Abstract: Large language models (LLMs) have achieved remarkable performance across a wide range of tasks, but their increasing parameter sizes significantly s

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

ADRD-Bench: A Preliminary LLM Benchmark for Alzheimer's Disease and Related Dementias

DGX agent

arXiv:2602.11460v2 Announce Type: replace Abstract: Large language models (LLMs) have shown great potential for healthcare applications. However, existing evaluation benchmarks provide minimal coverag

model-releasesarxiv-cs-cl
27 May 2026
Research

Agreement Between Large Language Models and Human Raters in Essay Scoring: A Research Synthesis

DGX agent

arXiv:2512.14561v2 Announce Type: replace Abstract: Despite the growing promise of large language models (LLMs) in automated essay scoring (AES), empirical findings regarding their reliability compare

researcharxiv-cs-cl
27 May 2026
Local Ai

AIDG: A Formal Decomposition of Information Extraction and Containment Asymmetries in Multi-Turn LLM Dialogue

DGX agent

arXiv:2602.17443v2 Announce Type: replace Abstract: Multi-turn LLM evaluation is typically reported as a single win-rate scalar, conflating distinct capabilities. We introduce AIDG (Adversarial Inform

local-aiarxiv-cs-cl
27 May 2026
Model Releases

AlbanianLLMSafety: A Safety Evaluation Dataset for Large Language Models in Albanian

DGX agent

arXiv:2605.26954v1 Announce Type: new Abstract: Safety evaluation of Large Language Models (LLMs) has largely focused on high-resource languages, leaving low-resource languages critically underserved.

model-releasesarxiv-cs-cl
27 May 2026
Safety

AlignEvoSkill: Towards Knowledge-Aware and Task-Aligned Agent Skill Evolution

DGX agent

arXiv:2506.23149v2 Announce Type: replace Abstract: Reusable skills play a key role in improving LLM-based agents, but existing skill-evolution methods often fail to ensure that evolved skills both co

safetyarxiv-cs-cl
27 May 2026
Research

Anchored Decoding: Provably Reducing Copyright Risk for Any Language Model

DGX agent

arXiv:2602.07120v2 Announce Type: replace Abstract: Language models (LMs) tend to memorize portions of their training data and emit verbatim spans. When the underlying sources are sensitive or copyrig

researcharxiv-cs-cl
27 May 2026
Model Releases

Are Video Models Zero-Shot Learners and Reasoners in Education? EduVideoBench, A Knowledge-Skills-Attitude Benchmark for Educational Video Generation

DGX agent

arXiv:2605.26918v1 Announce Type: new Abstract: Video generation models (VGMs) are rapidly entering classrooms, yet existing benchmarks evaluate only perceptual quality, intrinsic faithfulness, generi

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

Attribute-Based Diagnosis of LLM Alignment with Hate Speech Annotations

DGX agent

arXiv:2605.27025v1 Announce Type: new Abstract: Hate speech annotation is costly, subjective, and prone to annotator disagreement, making large-scale dataset construction challenging. We systematicall

model-releasesarxiv-cs-cl
27 May 2026
Safety

BAIT: Boundary-Guided Disclosure Escalation via Self-Conditioned Reasoning

DGX agent

arXiv:2605.27110v1 Announce Type: cross Abstract: In this work, we propose BAIT (Boundary-Aware Iterative Trap), a three-step jailbreak framework that approaches malicious goals through internal discl

safetyarxiv-cs-cl
27 May 2026
Model Releases

BESPOKE: Benchmark for Search-Augmented Large Language Model Personalization via Diagnostic Feedback

DGX agent

arXiv:2509.21106v2 Announce Type: replace Abstract: Search-augmented large language models (LLMs) have advanced information-seeking tasks by integrating retrieval into generation, reducing users' cogn

model-releasesarxiv-cs-cl
27 May 2026
Research

Beyond Binary: Speech Representations Across the Cognitive Score Hierarchy

DGX agent

arXiv:2605.27189v1 Announce Type: new Abstract: This study examines the relationship between speech representations and the hierarchical structure of cognitive assessment in mild cognitive impairment.

researcharxiv-cs-cl
27 May 2026
Agents

Beyond Self-Talk: A Communication-Centric Survey of LLM-Based Multi-Agent Systems

DGX agent

arXiv:2502.14321v3 Announce Type: replace-cross Abstract: Large language model-based multi-agent systems have recently gained significant attention due to their potential for complex, collaborative, a

agentsarxiv-cs-cl
27 May 2026
Model Releases

BeyondSWE: Can Current Code Agent Survive Beyond Single-Repo Bug Fixing?

DGX agent

arXiv:2603.03194v2 Announce Type: replace Abstract: Current code-agent benchmarks primarily evaluate localized issue resolution within a single target repository, leaving under-tested many software en

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

BhashaSetu: A Data-Centric Approach to Low-Resource Machine Translation

DGX agent

arXiv:2605.27050v1 Announce Type: new Abstract: We present BhashaSetu, a linguistically enriched English--Marathi parallel dataset addressing persistent data limitations in low-resource neural machine

model-releasesarxiv-cs-cl
27 May 2026
Local Ai

Bounded Path Context: A Controlled Study of Visible Path History in LLM-Based Knowledge Graph Question Answering

DGX agent

arXiv:2605.26645v1 Announce Type: new Abstract: LLM-based knowledge-graph question answering (KGQA) delegates graph traversal to language models, turning each question into a sequence of local relatio

local-aiarxiv-cs-cl
27 May 2026
Research

Chartographer: Counterfactual Chart Generation for Evaluating Vision-Language Models

DGX agent

arXiv:2605.27311v1 Announce Type: new Abstract: Chart question-answering (QA) benchmarks aim to pose questions that require visual reasoning to correctly answer, but models can often reach solutions t

researcharxiv-cs-cl
27 May 2026
Research

Clozing the Gap: Exploring Why Language Model Surprisal Outperforms Cloze Surprisal

DGX agent

arXiv:2601.09886v2 Announce Type: replace Abstract: How predictable a word is can be quantified in two ways: using human responses to the cloze task or using probabilities from language models (LMs).W

researcharxiv-cs-cl
27 May 2026
Research

Conceptual Steganography

DGX agent

arXiv:2605.26537v1 Announce Type: new Abstract: Language Models (LMs) emit Chains-of-Thought (CoTs) that drive much of their capability. However, the same sequence that carries useful reasoning can al

researcharxiv-cs-cl
27 May 2026
Safety

Conv-to-Bench: Evaluating Language Models Via User-Assistant Dialogues In Code Tasks

DGX agent

arXiv:2605.26440v1 Announce Type: new Abstract: The rapid advancement of Large Language Models (LLMs) has outpaced the scalability of traditional evaluation benchmarks, which remain heavily dependent

safetyarxiv-cs-cl
27 May 2026
Safety

Cultural Value Alignment Via Latent Activation Steering in Large Language Models

DGX agent

arXiv:2605.26365v1 Announce Type: new Abstract: Large Language Models (LLMs) often exhibit homogenized cultural perspectives. While the World Values Survey (WVS) provides a gold standard for mapping h

safetyarxiv-cs-cl
27 May 2026
Model Releases

Curation and Extraction of Drug-Related Entities from Reddit Platform

DGX agent

arXiv:2605.26445v1 Announce Type: new Abstract: Physicians learn primarily about illicit drugs from clinical overdose cases, limiting their understanding of real-world usage. Meanwhile, drug users sha

model-releasesarxiv-cs-cl
27 May 2026
Tutorials

Dissecting Multimodal In-Context Learning: Modality Asymmetries and Circuit Dynamics in modern Transformers

DGX agent

arXiv:2601.20796v2 Announce Type: replace Abstract: Transformer-based multimodal large language models often exhibit in-context learning (ICL) abilities. Motivated by this phenomenon, we ask: how do t

tutorialsarxiv-cs-cl
27 May 2026
Model Releases

DunbaaBERT: From Sacrifice to Semantics

DGX agent

arXiv:2605.26935v1 Announce Type: new Abstract: Large language models have achieved strong performance across many NLP tasks, yet Urdu remains comparatively underexplored due to limited resources and

model-releasesarxiv-cs-cl
27 May 2026
Safety

Efficient Agentic Reinforcement Learning with On-Policy Intrinsic Knowledge Boundary Enhancement

DGX agent

arXiv:2605.26952v1 Announce Type: new Abstract: Agentic reinforcement learning (RL) has proven effective for training LLM-based agents with external tool-use capabilities. However, we identify that ag

safetyarxiv-cs-cl
27 May 2026
Research

Energy-Gated Attention and Wavelet Positional Encoding: Complementary Inductive Biases for Transformer Attention

DGX agent

arXiv:2605.26355v1 Announce Type: cross Abstract: Standard transformer attention computes pairwise token similarity but treats all tokens as equally salient and all positions as equally local, regardl

researcharxiv-cs-cl
27 May 2026
Model Releases

ENPMR-Bench: Benchmarking Proactive Memory Retrieval for Emotional Support Agents

DGX agent

arXiv:2605.27240v1 Announce Type: new Abstract: Memory-augmented language agents are increasingly deployed in affective applications such as emotional support, where understanding and responding to us

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

Entropy Sentinel: Continuous LLM Accuracy Monitoring from Decoding Entropy Traces in STEM

DGX agent

arXiv:2601.09001v4 Announce Type: replace Abstract: Deploying LLMs raises two coupled challenges: (1) monitoring--estimating where a model underperforms as traffic and domains drift--and (2) improveme

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

EpiCurveBench: Evaluating VLMs on Epidemic Curve Digitization

DGX agent

arXiv:2605.27195v1 Announce Type: new Abstract: Chart-to-data extraction with vision-language models (VLMs) is increasingly evaluated on benchmarks that show diminishing headroom (frontier VLMs exceed

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

Evi-Steer: Learning to Steer Biomedical Vision-Language Models through Efficient and Generalizable Evidential Tuning

DGX agent

arXiv:2605.26292v1 Announce Type: cross Abstract: Parameter-efficient adaptation of vision-language foundation models is crucial for precise multimodal understanding of biomedical images, yet existing

model-releasesarxiv-cs-cl
27 May 2026
Research

Evidence Absence Is Not Evidence Insufficiency: Diagnosing NEI Construction Artifacts in Fact Verification

DGX agent

arXiv:2605.26663v1 Announce Type: new Abstract: Evidence absence is not evidence insufficiency, but fact verification benchmarks can make them observationally similar. The Not Enough Information (NEI)

researcharxiv-cs-cl
27 May 2026
Research

ExTax: Explainable Disinformation Detection via Persuasion, Emotion, and Narrative Role Taxonomies

DGX agent

arXiv:2605.27045v1 Announce Type: new Abstract: The democratization of LLMs has accelerated the generation and circulation of highly fluent disinformation, making traditional syntax-semantic verificat

researcharxiv-cs-cl
27 May 2026
Model Releases

FAB-Bench: A Framework for Adaptive RAG Benchmarking in Semiconductor Manufacturing

DGX agent

arXiv:2605.26476v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) has become critical for knowledge-intensive applications, yet evaluating its performance in vertical domains remain

model-releasesarxiv-cs-cl
27 May 2026
Safety

FalAR: A Large-scale Speaker-Annotated European Portuguese Speech Corpus of Parliamentary Sessions

DGX agent

arXiv:2605.27062v1 Announce Type: new Abstract: State-of-the-art performance for Automatic Speech Recognition (ASR) largely depends on the availability of large-scale labeled corpora. This creates a d

safetyarxiv-cs-cl
27 May 2026
Safety

FinHarness: An Inline Lifecycle Safety Harness for Finance LLM Agents

DGX agent

arXiv:2605.27333v1 Announce Type: new Abstract: Finance LLM agents must simultaneously block prompt-induced unauthorized actions and approve legitimate multi-step business workflows. However, boundary

safetyarxiv-cs-cl
27 May 2026
Research

Formalization of Malagasy conjugation

DGX agent

arXiv:2605.27161v1 Announce Type: new Abstract: This paper reports the core linguistic work performed to construct a dictionary-based morphological analyser for Malagasy simple verbs. It uses the Unit

researcharxiv-cs-cl
27 May 2026
Research

From Snippets to Semantics: Rethinking Evidence Granularity for Multilingual Fact Verification

DGX agent

arXiv:2605.26755v1 Announce Type: new Abstract: Multilingual fact verification requires evidence that is both relevant and sufficiently complete for reliable factuality prediction. However, existing s

researcharxiv-cs-cl
27 May 2026
Tutorials

Generating Logically Consistent Synthetic Supply Chain Data with LLM-Driven Knowledge Graph Reasoning

DGX agent

arXiv:2605.26823v1 Announce Type: new Abstract: Synthetic data offers a promising solution to two persistent barriers in supply chain analytics: data scarcity and data privacy. However, for synthetic

tutorialsarxiv-cs-cl
27 May 2026
Research

Granuscore: A Reference-Free Measure of Granularity for Text Analysis and Question Answering

DGX agent

arXiv:2605.26620v1 Announce Type: new Abstract: Natural language conveys information at varying levels of granularity, from fine-grained references to broad descriptions. While granularity is fundamen

researcharxiv-cs-cl
27 May 2026
← Previous
1…7374757677…162
Next →