AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
Model Releases

Continual LLM Upcycling: A Predictor-Gated Bank-Wise Sparsity Training Recipe for Dense-to-Sparse LLMs

DGX agent

arXiv:2606.10722v1 Announce Type: new Abstract: We study dense-to-sparse continual training as a way to construct channel-sparse large language models from dense checkpoints. Starting from a Qwen2.5-8

model-releasesarxiv-cs-cl
10 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

ConvMemory v2: A Recall-Preserving Top-10 Evidence Reranker for Conversational Memory Retrieval

DGX agent

arXiv:2606.10842v1 Announce Type: new Abstract: We describe ConvMemory v2, an opt-in token-evidence reranker that sits after the lightweight ConvMemory v1 reranker and reorders only v1's protected top

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

CoTAL: Human-in-the-Loop Prompt Engineering for Generalizable Formative Assessment Scoring and Feedback

DGX agent

arXiv:2504.02323v4 Announce Type: replace Abstract: Large language models (LLMs) have created new opportunities to assist teachers and support student learning. While researchers have explored various

model-releasesarxiv-cs-cl
10 Jun 2026
Agents

Data Journalist Agent: Transforming Data into Verifiable Multimodal Stories

DGX agent

arXiv:2606.11176v1 Announce Type: cross Abstract: Data tells stories that shape society; the data journalist's job is to turn raw information into stories non-experts can trust. A high-quality news fe

agentsarxiv-cs-cl
10 Jun 2026
Research

DECSELFMASK: Leveraging Unlabeled Text via Self-Relevance-Guided Masking for Decoder-Only Classification

DGX agent

arXiv:2606.09466v2 Announce Type: replace Abstract: Classification tasks require annotated data, which can often be expensive, time-consuming, or even unfeasible to collect. This is the case of the me

researcharxiv-cs-cl
10 Jun 2026
Local Ai

Density Field State Space Models: 1-Bit Distillation, Efficient Inference, and Knowledge Organization in Mamba-2

DGX agent

arXiv:2606.10932v1 Announce Type: new Abstract: We present Density Field State Space Models (DF-SSM), a framework for compressing SSMs to a 1-bit scaffold with int8 low-rank correction. Applied to Mam

local-aiarxiv-cs-cl
10 Jun 2026
Model Releases

Do Vision-Language Models See or Guess? Measuring and Reducing Textual-Prior Reliance with a Phrasing-Controlled Benchmark

DGX agent

arXiv:2606.10400v1 Announce Type: new Abstract: Vision-language models (VLMs) are increasingly deployed where answers must follow from what is in the image, yet they often answer from textual priors,

model-releasesarxiv-cs-cl
10 Jun 2026
Safety

Does Reasoning Preserve Alignment? On the Trustworthiness of Large Reasoning Models

DGX agent

arXiv:2606.11046v1 Announce Type: new Abstract: Instruction-tuned LLMs are increasingly converted into reasoning models through post-training to improve multi-step task performance. This conversion is

safetyarxiv-cs-cl
10 Jun 2026
Safety

Early-Token Confidence Predicts Reasoning Quality in Multi-Agent LLM Debate

DGX agent

arXiv:2606.10307v1 Announce Type: new Abstract: Evaluating reasoning quality in multi-agent LLM systems is challenging, especially for open-ended tasks without reference answers. We investigate whethe

safetyarxiv-cs-cl
10 Jun 2026
Safety

Enhancing Multilingual LLM-based ASR with Mixture of Experts and Dynamic Downsampling

DGX agent

arXiv:2606.10439v1 Announce Type: cross Abstract: The rapid progress of large language models (LLMs) has opened up a new frontier for automatic speech recognition (ASR), making their effective integra

safetyarxiv-cs-cl
10 Jun 2026
Tutorials

Entropy, Disagreement, and the Limits of Foundation Models in Genomics

DGX agent

arXiv:2604.04287v2 Announce Type: replace-cross Abstract: Foundation models in genomics have shown mixed success compared to their counterparts in natural language processing. Yet, the reasons for the

tutorialsarxiv-cs-cl
10 Jun 2026
Research

From Genes to Tokens: a GWAS-inspired Approach for Interpretable Stylometric Analysis

DGX agent

arXiv:2606.09543v2 Announce Type: replace Abstract: This short paper introduces a stylometric interpretation method inspired by genome-wide association studies (GWAS). Each 'gene' token's association

researcharxiv-cs-cl
10 Jun 2026
Model Releases

From Observation to Intervention: A Causal Audit of Expert Importance in Mixture-of-Experts Models

DGX agent

arXiv:2606.10703v1 Announce Type: cross Abstract: Interpretability methods routinely use population-level summary statistics over observed model behaviour to license claims about the effects of target

model-releasesarxiv-cs-cl
10 Jun 2026
Applications

Generative Archetype-Grounded Item Representations for Sequential Recommendation

DGX agent

arXiv:2606.11023v1 Announce Type: cross Abstract: Sequential recommendation aims to predict users' next interaction with items by analyzing their historical behavior. However, the limited quality of i

applicationsarxiv-cs-cl
10 Jun 2026
Model Releases

GhazalBench: Evaluating LLM Understanding and Canonical Surface-Form Access in Persian Ghazals

DGX agent

arXiv:2603.09979v2 Announce Type: replace Abstract: Persian poetry plays an active role in Iranian cultural practice, where verses by canonical poets such as Hafez are frequently quoted, paraphrased,

model-releasesarxiv-cs-cl
10 Jun 2026
Research

How Does Reasoning Flow? Tracing Attention-Induced Information Flow for Targeted RL in LLMs

DGX agent

arXiv:2606.10646v1 Announce Type: cross Abstract: Token-level credit assignment remains a key obstacle for reinforcement learning (RL) in large language models (LLMs), where RL recipes typically treat

researcharxiv-cs-cl
10 Jun 2026
Research

inversedMixup: Data Augmentation via Inverting Mixed Embeddings

DGX agent

arXiv:2601.21543v3 Announce Type: replace Abstract: Mixup generates augmented samples by linearly interpolating inputs and labels with a controllable ratio. However, since it operates at the latent em

researcharxiv-cs-cl
10 Jun 2026
Safety

It Takes One to Bias Them All: Breaking Bad with One-Shot GRPO

DGX agent

arXiv:2606.10931v1 Announce Type: new Abstract: Warning: This paper contains several toxic and offensive statements. Modern large language models (LLMs) are typically aligned through large-scale post-

safetyarxiv-cs-cl
10 Jun 2026
Model Releases

KCSAT-ML: Probing Reasoning Models with Nationwide-Cohort Human Difficulty

DGX agent

arXiv:2606.10403v1 Announce Type: new Abstract: Math reasoning benchmarks have proliferated, yet most lack a per-item difficulty signal grounded in actual human performance. We introduce KCSAT-ML, a d

model-releasesarxiv-cs-cl
10 Jun 2026
Research

Large Language Models as Modal Models in Linguistics

DGX agent

arXiv:2606.10467v1 Announce Type: new Abstract: The rapid advancement of large language models (LLMs) has intensified debates about their significance for linguistic theory. These debates are commonly

researcharxiv-cs-cl
10 Jun 2026
Research

Leveraging Social Media Data for COVID-19 Studies

DGX agent

arXiv:2606.10459v1 Announce Type: cross Abstract: Nowadays, social media networks have become widely preferred sources of information. Especially during the time of the Coronavirus disease 2019 COVID

researcharxiv-cs-cl
10 Jun 2026
Safety

Lightweight Latent Reasoning for Narrative Tasks

DGX agent

arXiv:2512.02240v2 Announce Type: replace Abstract: Large language models (LLMs) tackle complex tasks by generating long chains of thought or 'reasoning traces' that act as latent variables in the gen

safetyarxiv-cs-cl
10 Jun 2026
Safety

Measuring Human Value Expression in Social Media Texts: Calibrated LLM Annotation and Encoder Transfer

DGX agent

arXiv:2606.11018v1 Announce Type: new Abstract: Measuring subjective constructs in naturally occurring social media text requires annotation procedures that are theoretically grounded, empirically val

safetyarxiv-cs-cl
10 Jun 2026
Safety

Mechanistic Analysis of Alignment Algorithms in Language Models

DGX agent

arXiv:2606.09850v1 Announce Type: cross Abstract: Post-training alignment algorithms are predominantly evaluated as black boxes, obscuring how they reshape language models' internal computations. We p

safetyarxiv-cs-cl
10 Jun 2026
Model Releases

MIRAGE: A Polarity-Flipping Encoding Subspace in LLM Agents

DGX agent

arXiv:2606.10304v1 Announce Type: new Abstract: When LLM agents are coerced into covertly encoding sensitive data (Base64, ROT13, acrostic, synonym chains, and beyond), the resulting outputs evade out

model-releasesarxiv-cs-cl
10 Jun 2026
Safety

Multi-Faceted Interactivity Alignment in Full-Duplex Speech Models

DGX agent

arXiv:2606.11167v1 Announce Type: new Abstract: Full-duplex spoken dialogue models can listen and speak simultaneously, making them a promising architecture for natural conversation. However, current

safetyarxiv-cs-cl
10 Jun 2026
Safety

Multilingual Word-Level Forced Alignment with Self-Supervised Representations and Learned Dynamic Programming

DGX agent

arXiv:2606.10675v1 Announce Type: new Abstract: We present a method for accurate multilingual word-level forced alignment, consisting of an alignment encoder and a learned alignment decoder. The encod

safetyarxiv-cs-cl
10 Jun 2026
Model Releases

N-GRPO: Embedding-Level Neighbor Mixing for Enhanced Policy Optimization

DGX agent

arXiv:2606.10768v1 Announce Type: cross Abstract: The success of Large Language Models in mathematical reasoning relies heavily on the generation of diverse and valid solution paths during the rollout

model-releasesarxiv-cs-cl
10 Jun 2026
Research

Open Korean Corpora: A Practical Report

DGX agent

arXiv:2012.15621v3 Announce Type: replace Abstract: Korean is often referred to as a low-resource language in the research community. While this claim is partially true, it is also because the availab

researcharxiv-cs-cl
10 Jun 2026
Model Releases

OpenRTLSet: A Fully Open-Source Dataset for Large Language Model-based Verilog Module Design

DGX agent

arXiv:2606.10285v1 Announce Type: new Abstract: OpenRTLSet introduces the largest fully open-source dataset for hardware design, offering over 131,000 diverse Verilog code samples to the research comm

model-releasesarxiv-cs-cl
10 Jun 2026
Safety

PADD: Path-Aligned Decompression Distillation for Non-Router Teacher to Guide MoE Student Learning

DGX agent

arXiv:2606.10369v1 Announce Type: new Abstract: As large language models (LLMs) continue to scale, it becomes increasingly challenging to grow model capacity under fixed computation budgets. We propos

safetyarxiv-cs-cl
10 Jun 2026
Safety

ParaBridge: Bridging Paralinguistic Perception and Dialogue Behavior in Speech Language Models

DGX agent

arXiv:2606.10581v1 Announce Type: new Abstract: Speech carries more information than just words: a child's voice, a fearful tone, or a noisy background should all lead a sufficiently competent spoken-

safetyarxiv-cs-cl
10 Jun 2026
Model Releases

Parallel Causal Associative Fields: Gated Sparse Memory for Long-Context Language Modeling

DGX agent

arXiv:2606.10435v1 Announce Type: cross Abstract: Transformers achieve strong language modeling performance by providing direct token-to-token communication paths, but causal self-attention scales qua

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

Parametric Knowledge is Not All You Need: Toward Honest Large Language Models via Retrieval of Pretraining Data

DGX agent

arXiv:2601.21218v2 Announce Type: replace Abstract: Large language models (LLMs) are highly capable of answering questions, but they are often unaware of their own knowledge boundary, i.e., knowing wh

model-releasesarxiv-cs-cl
10 Jun 2026
Research

Pre-AF 13: An Interpretable Atrial Fibrillation Risk Score Mined from Discharge Reports

DGX agent

arXiv:2606.10725v1 Announce Type: cross Abstract: Background. Atrial fibrillation (AF) is the most prevalent cardiac arrhythmia and a major determinant of prognosis. Established AF risk scores rely on

researcharxiv-cs-cl
10 Jun 2026
Research

Prefilling-dLLM: Predictive Prefilling for Long-Context Inference in Diffusion Language Models

DGX agent

arXiv:2606.10537v1 Announce Type: new Abstract: Diffusion large language models (dLLMs) re-encode the entire prefix at every denoising step, causing recomputation that scales quadratically with contex

researcharxiv-cs-cl
10 Jun 2026
Model Releases

ProbeLLM: Automating Principled Diagnosis of LLM Failures

DGX agent

arXiv:2602.12966v2 Announce Type: replace Abstract: Understanding how and why large language models (LLMs) fail is becoming a central challenge as models rapidly evolve and static evaluations fall beh

model-releasesarxiv-cs-cl
10 Jun 2026
Agents

Pushing the Limits of LLM Tool Calling via Experiential Knowledge Integration and Activation

DGX agent

arXiv:2606.10875v1 Announce Type: new Abstract: Large language models (LLMs) rely on tool use to act as autonomous agents, yet often fail in multi-step execution due to insufficient tool-related knowl

agentsarxiv-cs-cl
10 Jun 2026
Model Releases

REAL: A Reasoning-Enhanced Graph Framework for Long-Term Memory Management of LLMs

DGX agent

arXiv:2606.10694v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly expected to interact with users over long time horizons. However, due to their finite context window, LLMs

model-releasesarxiv-cs-cl
10 Jun 2026
Safety

Recovering the Zipfian Distribution in Unsupervised Term Discovery

DGX agent

arXiv:2606.10781v1 Announce Type: cross Abstract: Unsupervised term discovery involves segmenting unlabelled speech into word- or syllable-like units and clustering these into a lexicon of candidate t

safetyarxiv-cs-cl
10 Jun 2026
Model Releases

RedAct: Redacting Agent Capability Traces for Procedural Skill Protection

DGX agent

arXiv:2606.10813v1 Announce Type: cross Abstract: Users rely on execution traces to observe agent behavior, diagnose failures, and ensure accountability. These traces contain rich procedural detail, i

model-releasesarxiv-cs-cl
10 Jun 2026
Research

Representation-Aware Advantage Estimation: Your Reward Model Provides More Than A Scalar Output

DGX agent

arXiv:2606.10528v1 Announce Type: cross Abstract: Current reinforcement learning from human feedback (RLHF) methods primarily rely on scalar rewards from a trained reward model (RM). While effective,

researcharxiv-cs-cl
10 Jun 2026
Model Releases

Revisiting Metric Reliability for Fine-grained Evaluation of Machine Translation and Summarization in Indian Languages

DGX agent

arXiv:2510.07061v2 Announce Type: replace Abstract: While automatic metrics drive progress in Machine Translation (MT) and Text Summarization (TS), existing metrics have been developed and validated a

model-releasesarxiv-cs-cl
10 Jun 2026
Research

Scaling Self-Supervised Speech Models Uncovers Deep Linguistic Relationships: Evidence from the Pacific Cluster

DGX agent

arXiv:2603.07238v2 Announce Type: replace Abstract: Similarities between language representations derived from Self-Supervised Speech Models (S3Ms) have been observed to primarily reflect geographic p

researcharxiv-cs-cl
10 Jun 2026
Safety

Selection, Not Salience: The Shape and Limits of Personalization in Social Highlighting

DGX agent

arXiv:2606.10398v1 Announce Type: cross Abstract: Does personalizing what a reader sees pay off, and where does it stop? Using a social web highlighter and a co-readership identity control (the same d

safetyarxiv-cs-cl
10 Jun 2026
Model Releases

Small Data, Big Noise: Adversarial Training for Robust Parameter-Efficient Fine-Tuning

DGX agent

arXiv:2606.10610v1 Announce Type: new Abstract: Parameter-Efficient Fine-Tuning (PEFT) has become essential for adapting foundation models to downstream NLP tasks. However, current PEFT methods often

model-releasesarxiv-cs-cl
10 Jun 2026
Safety

Speaker Group Encoding in Self-supervised Speech Recognition Models

DGX agent

arXiv:2606.10654v1 Announce Type: new Abstract: We investigate what self-supervised speech recognition models (S3Ms) learn about speaker groups (SGs). We examine several states of S3Ms: pretrained, fi

safetyarxiv-cs-cl
10 Jun 2026
Safety

SpeechJBB: Probing Safety Alignment and Comprehension in Large Audio Language Models under Code-Switched Speech

DGX agent

arXiv:2606.06037v2 Announce Type: cross Abstract: Large audio language models (LALMs) are increasingly deployed in real-world applications, yet their safety alignment is still primarily evaluated on m

safetyarxiv-cs-cl
10 Jun 2026
← Previous
1…4849505152…161
Next →