AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
26 Jun 2026

Cascaded Multi-Granularity Pruning for On-Device LLM Inference in Industrial IoT

Model ReleasesDGX agent

arXiv:2606.26861v1 Announce Type: new Abstract: Deploying large language models (LLMs) on Industrial Internet of Things (IIoT) edge devices demands extreme compression, yet existing structured pruning

Closing the Quality Gap in Low-Resource Text-to-Speech: LoRA Fine-Tuning of VoxCPM2 for Khmer and Korean

Model ReleasesDGX agent

arXiv:2606.26618v1 Announce Type: new Abstract: Large pretrained text-to-speech (TTS) models sound almost human for well-resourced languages, but much worse for languages that are rare in their traini

Comparing BERT Sentence-Pair Classification and Few-Shot LLM Prompting for Detecting Threat and Solution Framing in German Climate News


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases
DGX agent

arXiv:2606.26489v1 Announce Type: new Abstract: News media play a central role in shaping public perceptions of climate change, and whether coverage emphasizes threats or solutions has measurable effe

Compiler-Driven Approximation Tuning for Hyperdimensional Computing

ResearchDGX agent

arXiv:2606.26547v1 Announce Type: cross Abstract: As Moore's law reaches its physical and economic limits, domain-specific approaches are increasingly employed to accelerate machine learning workloads

Compositionality and the lexicon in evolutionary semantics

ResearchDGX agent

arXiv:2606.27228v1 Announce Type: new Abstract: Formal semantics has shown that sentence meanings arise by recursively composing lexical meanings, yet much of the literature on semantic universals mod

ConvMemory v3: A Validity Context Layer for Conversational Memory via Target-Conditioned Relation Verification

Model ReleasesDGX agent

arXiv:2606.26753v1 Announce Type: new Abstract: Conversational memory retrieval optimizes relevance, yet a retrieved memory can be relevant and simultaneously outdated: a later turn updates, corrects,

DanceOPD: On-Policy Generative Field Distillation

Local AiDGX agent

arXiv:2606.27377v1 Announce Type: cross Abstract: Modern image generation demands a single model that unifies diverse capabilities, including text-to-image (T2I), local editing, and global editing. Ho

Does AI Reviewer See the Full Picture? Attacking and Defending Multimodal Peer Review

Model ReleasesDGX agent

arXiv:2606.12716v2 Announce Type: replace Abstract: The integration of Large Language Models (LLMs) and Multimodal LLMs (MLLMs) into scientific peer-review workflows introduces novel and significant r

DualEval: Joint Model-Item Calibration for Unified LLM Evaluation

Model ReleasesDGX agent

arXiv:2606.26429v1 Announce Type: cross Abstract: Current LLM evaluation relies on two complementary but often disconnected signals: static benchmarks with objective correctness labels and arena-style

Dynamic-dLLM: Dynamic Cache-Budget and Adaptive Parallel Decoding for Training-Free Acceleration of Diffusion LLM

ResearchDGX agent

arXiv:2606.26120v1 Announce Type: new Abstract: Diffusion Large Language Models (dLLMs) offer a promising alternative to autoregressive models, excelling in text generation tasks due to their bidirect

Embarrassingly Simple Self-Distillation Improves Code Generation

Model ReleasesDGX agent

arXiv:2604.01193v2 Announce Type: replace Abstract: Can a large language model (LLM) improve at code generation using only its own raw outputs, without a verifier, a teacher model, or reinforcement le

Epiphany-Aware KV Cache Eviction Without the Attention Matrix

ApplicationsDGX agent

arXiv:2606.26472v1 Announce Type: cross Abstract: As reasoning models emit chains of thought tens of thousands of tokens long, KV cache increasingly becomes a deployment bottleneck. Existing cache evi

Erase-then-Delta Attention: Decoupling Erase and Write Addresses in Delta-Rule Linear Attention

ResearchDGX agent

arXiv:2606.26560v1 Announce Type: new Abstract: Delta-rule linear attention improves recurrent memory updates by correcting what is already stored at the current write address before writing new conte

Evaluation Pitfalls and Challenges in Multimedia Event Extraction

ApplicationsDGX agent

arXiv:2606.26775v1 Announce Type: new Abstract: Multimedia event extraction aims to jointly identify events and their arguments across multiple modalities, such as text and images, to support more com

EvoEmbedding: Evolvable Representations for Long-Context Retrieval and Agentic Memory

AgentsDGX agent

arXiv:2606.21649v2 Announce Type: replace Abstract: Existing embedding models are inherently static: they encode text segments in isolation, ignoring their surrounding context and temporal order. This

Extracting Problem and Method Sentence from Scientific Papers: A Context-enhanced Transformer Using Formulaic Expression Desensitization

TutorialsDGX agent

arXiv:2606.26481v1 Announce Type: new Abstract: Billions of scientific papers lead to the need to identify essential parts from the massive text. Scientific research is an activity from putting forwar

Eyes-on-Me: Scalable RAG Poisoning through Transferable Attention-Steering Attractors

ResearchDGX agent

arXiv:2510.00586v3 Announce Type: replace-cross Abstract: Existing data poisoning attacks on retrieval-augmented generation (RAG) systems scale poorly because they require costly optimization of poiso

FBK's Long-form SpeechLLMs for IWSLT 2026 Instruction Following

ResearchDGX agent

arXiv:2606.26819v1 Announce Type: new Abstract: This paper describes our submission to the IWSLT 2026 Instruction Following shared task. SpeechLLMs are developed for both short-form and long-form spee

Forecasting With LLMs: Improved Generalization Through Feature Steering

SafetyDGX agent

arXiv:2606.27199v1 Announce Type: new Abstract: Successful forecasting involves identifying patterns between historical and future states of the world which generalize to future observations. We apply

From Guessing to Placeholding: A Cost-Theoretic Framework for Uncertainty-Aware Code Completion

Model ReleasesDGX agent

arXiv:2604.01849v2 Announce Type: replace Abstract: While Large Language Models (LLMs) have demonstrated exceptional proficiency in code completion, they typically adhere to a Hard Completion (HC) par

From Vajrayana Tara to Bengali Baul: A Computational Study of Lexical Transmission Across Buddhist, Shakta, and Vaishnava Traditions in Bengal

ResearchDGX agent

arXiv:2606.26803v1 Announce Type: new Abstract: We present a computational corpus study of vocabulary relationships across eight tradition layers of Bengali and Sanskrit devotional literature spanning

From Weights to Features: SAE-Guided Activation Regularization for LLM Continual Learning

Model ReleasesDGX agent

arXiv:2606.26629v1 Announce Type: cross Abstract: Weight-space regularization methods such as Elastic Weight Consolidation (EWC) are the standard approach to catastrophic forgetting in continual learn

GAVEL: Grounded Caption Error Verification and Localization

Model ReleasesDGX agent

arXiv:2606.26923v1 Announce Type: new Abstract: Vision-language models (VLMs) often produce hallucinated or inconsistent outputs, where text and images are not properly aligned. Addressing this issue

GenRecal: Generation after Recalibration from Large to Small Vision-Language Models

ApplicationsDGX agent

arXiv:2506.15681v4 Announce Type: replace Abstract: Recent advancements in vision-language models (VLMs) have leveraged large language models (LLMs) to achieve performance on par with closed-source sy

HarmVideoBench: Benchmarking Harmful Video Understanding in Large Multimodal Models

Model ReleasesDGX agent

arXiv:2606.27187v1 Announce Type: cross Abstract: Large vision-language models (LVLMs) have recently shown immense potential in automated content moderation, sparking growing interest in developing ha

Heterogeneous Neural Predictivity from Language Models During Naturalistic Comprehension

Local AiDGX agent

arXiv:2606.26880v1 Announce Type: new Abstract: Language-model representations provide structured, high-dimensional annotations of naturalistic language stimuli and can serve as informative neural pre

HierBias: Context-Conditioned Hierarchical Media Bias Detection with Multi-Task Type Classification

SafetyDGX agent

arXiv:2606.26100v1 Announce Type: new Abstract: Media bias detection is a critical task for ensuring fair and balanced information dissemination, yet existing sentence-level approaches classify each s

How Surprising Is Historical Italian to Language Models? Tokenization Tax, Comprehension Tax, and a Simple Mitigation

ApplicationsDGX agent

arXiv:2606.27275v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly critical to digital library workflows, yet their ability to process historical language remains poorly und

HyperDFlash: MHC-Aligned Block Speculative Decoding with Gated Residual Reduction

Model ReleasesDGX agent

arXiv:2606.26744v1 Announce Type: cross Abstract: We present HyperDFlash, a block-parallel speculative decoding framework tailored to the novel multi-hyper-connection (MHC) architecture proposed by De

Improving General Role-Playing Agents via Psychology-Grounded Reasoning and Role-Aware Policy Optimization

SafetyDGX agent

arXiv:2606.27025v1 Announce Type: new Abstract: Building general-purpose role-playing agents that faithfully portray any character from a natural-language profile remains challenging. The dominant par

Jailbreaking for the Average Jane: Choosing Optimal Jailbreaks via Bandit Algorithms for Automatically Enhanced Queries

Model ReleasesDGX agent

arXiv:2606.26936v1 Announce Type: cross Abstract: With a profusion of jailbreaks for LLMs now widely known, a growing concern is that non-expert malicious actors ('the average Jane') could elicit acti

Just how sure are you? Improving Verbalized Uncertainty Calibration in Medical VQA

SafetyDGX agent

arXiv:2606.27023v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) applied to Medical Visual Question Answering (VQA) tend to produce overconfident outputs regardless of actual

Learning from the Self-future: On-policy Self-distillation for dLLMs

SafetyDGX agent

arXiv:2606.18195v2 Announce Type: replace Abstract: On-policy self-distillation (OPSD) has proven effective for post-training large language models (LLMs), yet its application to diffusion LLMs (dLLMs

Learning User Simulators with Turing Rewards

AgentsDGX agent

arXiv:2606.19336v2 Announce Type: replace Abstract: Learning to simulate human users in interactive settings could advance the training of agent assistants, evaluation of personalization systems, rese

Linguistics and Human Brain: A Perspective of Computational Neuroscience

SafetyDGX agent

arXiv:2602.08275v3 Announce Type: replace-cross Abstract: Elucidating the language-brain relationship requires bridging the methodological gap between the abstract theoretical frameworks of linguistic

LLM-Based Examination of Eligibility Criteria from Securities Prospectuses at the German Central Bank

ApplicationsDGX agent

arXiv:2606.27316v1 Announce Type: new Abstract: Verifying the eligibility of securities as collateral is a key responsibility of the German Central Bank. However, manually verifying these assets again

LMs as Task-Specific Knowledge Bases: An Interpretability Analysis

Model ReleasesDGX agent

arXiv:2606.27237v1 Announce Type: new Abstract: Language models (LMs) capture large amounts of factual knowledge applicable to a wide range of tasks, motivating the view of their parameters as a knowl

Mapping Political-Elite Networks in Europe with a Multilingual Joint Entity-Relation Extraction Pipeline

ApplicationsDGX agent

arXiv:2606.27347v1 Announce Type: new Abstract: Whether political elites organise into rent-seeking coalitions that capture public resources or civic networks that sustain governance is a central ques

MinGram: A Minimalist Unigram Tokenizer with High Compression and Competitive Morphological Alignment

SafetyDGX agent

arXiv:2606.27019v1 Announce Type: new Abstract: The Unigram tokenizer uses an elegant representation which makes it straightforward to edit vocabularies, but its training is comparatively heavy and co

Multilingual Reasoning Cascades Need More Context

ResearchDGX agent

arXiv:2606.27306v1 Announce Type: new Abstract: Translation cascades for reasoning translate the query from another language to English, reason in English, and translate the answer back to the origina

Nemotron-TwoTower: Diffusion Language Modeling with Pretrained Autoregressive Context

Model ReleasesDGX agent

arXiv:2606.26493v1 Announce Type: new Abstract: Diffusion language models offer a promising alternative to autoregressive models due to their potential for parallel and iterative generation. However,

Neural Speaker Diarization via Multilingual Training: Evaluation on Low-Resource Nepali-Hindi Speech

SafetyDGX agent

arXiv:2606.26144v1 Announce Type: cross Abstract: Speaker diarization, the task of determining 'who spoke when' in a multi-speaker recording, is a critical component in applications such as meeting tr

OI-Bench: An Option Injection Benchmark for Evaluating LLM Susceptibility to Directive Interference

Model ReleasesDGX agent

arXiv:2601.13300v2 Announce Type: replace Abstract: Benchmarking large language models (LLMs) is critical for understanding their capabilities, limitations, and robustness. In addition to interface ar

OPID: On-Policy Skill Distillation for Agentic Reinforcement Learning

SafetyDGX agent

arXiv:2606.26790v1 Announce Type: new Abstract: Outcome-based reinforcement learning provides a stable optimization backbone for language agents, but its sparse trajectory-level rewards provide little

Orthogonal Hierarchical Decomposition for Structure-Aware Table Understanding with Large Language Models

ResearchDGX agent

arXiv:2602.01969v2 Announce Type: replace Abstract: Complex tables with multi-level headers, merged cells and heterogeneous layouts pose persistent challenges for LLMs in both understanding and reason

Overcoming State Inertia: Minimally Invasive Temporal Alignment for Evolving Contexts

SafetyDGX agent

arXiv:2512.03704v3 Announce Type: replace Abstract: Long-context dialogue systems suffer from state inertia, where models over-attend to history and fail to adapt to evolving intents. We demonstrate t

Paved with True Intents: Intent-Aware Training Improves LLM Safety Classification Across Training Regimes

SafetyDGX agent

arXiv:2606.27210v1 Announce Type: new Abstract: We argue that safety classifiers should model user intent as an explicit signal between the prompt and the final label. To study this, we introduce AIMS

Phonetic and semantic analyses of spoken corpora of Beijing and Taiwan Mandarin indicate that the neutral tone is a lexical tone

ResearchDGX agent

arXiv:2606.26360v1 Announce Type: new Abstract: The neutral, or floating, tone of Mandarin Chinese is a tone with an enigmatic set of properties. It has been described as a reduced tone, or as a tone

ProfileFoundry: A Synthetic Person-Object Substrate for Privacy, Memory, and Tool-Use Evaluation in LLM Agent

Local AiDGX agent

arXiv:2606.26403v1 Announce Type: new Abstract: Foundation-model research increasingly needs data about people: user state, personal histories, relationships, contact-like fields, documents, and longi

Prompt, Plan, Extract: Zero-Shot Agentic LLMs Workflows for Lung Pathology Extraction from Clinical Narratives

AgentsDGX agent

arXiv:2606.19852v2 Announce Type: replace Abstract: Information extraction from pathology reports is essential for cancer staging, tumor registry population. Yet key data remains embedded in narrative

RedVox: Safety and Fairness Gaps in Speech Models Across Languages

Model ReleasesDGX agent

arXiv:2606.26968v1 Announce Type: new Abstract: Speech-capable models are increasingly deployed in real-world applications across languages. Yet their safety and fairness beyond English settings and u

Reproducibility Study of 'AlphaEdit: Null-Space Constrained Knowledge Editing for Language Models'

Model ReleasesDGX agent

arXiv:2606.26783v1 Announce Type: cross Abstract: Fang et al. (2025) introduced a null-space constrained projection, named AlphaEdit, for locate-then-edit knowledge editing methods, theoretically guar

SocialPersona: Benchmarking Personalized Profiling and Response with Multimodal Social-Media Context

Model ReleasesDGX agent

arXiv:2606.26654v1 Announce Type: new Abstract: Personalized language-model assistants are often evaluated through a memory lens: can a model recall preferences users have explicitly stated in dialogu

Soft Token Alignment for Cross-Lingual Reasoning

SafetyDGX agent

arXiv:2606.26466v1 Announce Type: new Abstract: Multilingual large language models often produce inconsistent reasoning and answers for semantically equivalent prompts in different languages. Prior wo

Somatic in the East, Psychological in the West?: Investigating Clinically-Grounded Cross-Cultural Depression Symptom Expression in LLMs

SafetyDGX agent

arXiv:2508.03247v2 Announce Type: replace Abstract: Prior clinical psychology research shows that Western individuals with depression tend to report psychological symptoms, while Eastern individuals r

Spurious Rewards Paradox: Mechanistically Understanding How RLVR Activates Memorization Shortcuts in LLMs

Model ReleasesDGX agent

arXiv:2601.11061v2 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) is highly effective for enhancing LLM reasoning, yet recent evidence shows models like Q

Staying VIGILant: Mitigating Visual Laziness via Counterfactual Visual Alignment in MLLMs

SafetyDGX agent

arXiv:2606.26387v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) extend large language models (LLMs) with visual perception, enabling joint reasoning over images and text. De

Structure Before Collapse: Transient semantic geometry in next-token prediction

TutorialsDGX agent

arXiv:2606.26749v1 Announce Type: cross Abstract: Neural Collapse predicts that balanced one-hot classification pushes model representations to be equally far from each other; a symmetric configuratio

Syntactic Belief Update as the Driver of Garden Path Processing Difficulty

ResearchDGX agent

arXiv:2606.27206v1 Announce Type: new Abstract: Garden path sentences present a processing difficulty for humans -- the sentence prefix leads the listener towards one interpretation, until the listene

Term-Centric Hierarchy Induction from Heterogeneous Corpora

Model ReleasesDGX agent

arXiv:2606.26963v1 Announce Type: new Abstract: Organizing knowledge from diverse text sources into interpretable hierarchies is crucial for tasks such as policy analysis, innovation monitoring, and e

← Previous
1…3233343536…129
Next →