AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
Safety

MIThinker: A Plug-and-Play Policy-Optimized Thinker For Motivational Interviewing Counseling

DGX agent

arXiv:2606.29265v1 Announce Type: new Abstract: Reasoning large language models (LLMs) have recently made much progress in complex problem-solving, leveraging internal reasoning (or thought) to guide

safetyarxiv-cs-cl
30 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Research

Mitigating Batch Effects in Histopathology via Language-Mediated Robust Embedding Generation

DGX agent

arXiv:2606.28697v1 Announce Type: cross Abstract: Pathology foundation models (PFMs) have demonstrated strong potential across clinical and scientific applications, yet their performance is often hind

researcharxiv-cs-cl
30 Jun 2026
Model Releases

MixSarc: A Bangla-English Code-Mixed Corpus for Implicit Meaning Identification

DGX agent

arXiv:2602.21608v2 Announce Type: replace Abstract: Bangla-English code-mixing is widespread across South Asian social media, yet resources for implicit meaning identification in this setting remain s

model-releasesarxiv-cs-cl
30 Jun 2026
Research

Model Directions, Not Words: Mechanistic Topic Models Using Sparse Autoencoders

DGX agent

arXiv:2507.23220v2 Announce Type: replace Abstract: Traditional topic models are effective at uncovering latent themes in large text collections. However, due to their reliance on bag-of-words represe

researcharxiv-cs-cl
30 Jun 2026
Safety

MOPD: Multi-Teacher On-Policy Distillation for Capability Integration in LLM Post-Training

DGX agent

arXiv:2606.30406v1 Announce Type: new Abstract: Modern large language models (LLMs) rely on reinforcement learning during post-training to push specific capabilities, yet integrating multiple capabili

safetyarxiv-cs-cl
30 Jun 2026
Model Releases

Morphing into Hybrid Attention Models

DGX agent

arXiv:2606.30562v1 Announce Type: new Abstract: Hybrid attention models improve long-context efficiency by retaining only a subset of full-attention layers and replacing the remaining layers with line

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

Multi-Agentic System Leveraging Open-Source LLMs to Mitigate Disinformation Threats

DGX agent

arXiv:2606.30259v1 Announce Type: new Abstract: In contemporary societies, the threat of disinformation has reached alarming levels, exacerbated by the proliferation of electronic communication, socia

model-releasesarxiv-cs-cl
30 Jun 2026
Research

Multi-Block Diffusion Language Models

DGX agent

arXiv:2606.29215v1 Announce Type: cross Abstract: Block Diffusion Language Models (BD-LMs) improve diffusion-based text generation with KV caching and flexible-length generation. A natural next step i

researcharxiv-cs-cl
30 Jun 2026
Model Releases

Multimodal Mathematical Reasoning with Diverse Solving Perspective

DGX agent

arXiv:2507.02804v2 Announce Type: replace Abstract: Recent progress in large-scale reinforcement learning (RL) has notably enhanced the reasoning capabilities of large language models (LLMs), especial

model-releasesarxiv-cs-cl
30 Jun 2026
Safety

Node-to-Neighborhood Semantic Consistency: Text-Topology Alignment for TAGs Anomaly Detection

DGX agent

arXiv:2606.30009v1 Announce Type: new Abstract: Graph anomaly detection (GAD) on text-attributed graphs (TAGs) is vital for applications such as fraud detection and academic integrity verification. Ex

safetyarxiv-cs-cl
30 Jun 2026
Model Releases

Not-quite-human tastes: the stylized omnivorousness of LLM survey surrogates

DGX agent

arXiv:2606.30085v1 Announce Type: new Abstract: Large-language models have proven to be remarkable if inconsistent parrots of public attitudes and opinions. The extent to which LLMs are able to produc

model-releasesarxiv-cs-cl
30 Jun 2026
Research

OLIVE: View-Augmented Latent Prediction with Waveform Reconstruction for Speech SSL

DGX agent

arXiv:2606.30356v1 Announce Type: new Abstract: We propose Online Latent prediction with Invariant Views and rEconstruction (OLIVE), a self-supervised speech representation learning framework that joi

researcharxiv-cs-cl
30 Jun 2026
Safety

Online Experiential Learning for Language Models

DGX agent

arXiv:2603.16856v2 Announce Type: replace Abstract: The prevailing paradigm for improving large language models relies on offline training with human annotations or simulated environments, leaving the

safetyarxiv-cs-cl
30 Jun 2026
Applications

Open but Incompatible: A License Compatibility Analysis of Corpora for Low-Resource African Languages

DGX agent

arXiv:2606.28867v1 Announce Type: new Abstract: Creative Commons licenses dominate African NLP corpus releases, but their compatibility rules are rarely applied. CC-BY-SA and CC-BY-NC cannot be combin

applicationsarxiv-cs-cl
30 Jun 2026
Model Releases

Parametric Skills

DGX agent

arXiv:2606.30015v1 Announce Type: new Abstract: Since intelligence fundamentally relies on efficient skill acquisition (Chollet, 2019), the ability to leverage skills is critical. For LLMs, skills, ma

model-releasesarxiv-cs-cl
30 Jun 2026
Research

PASTA: A Paraphrasing And Self-Training Approach for Knowledge Updating in LLMs

DGX agent

arXiv:2606.28898v1 Announce Type: new Abstract: Knowledge updating in pre-trained Large Language Models (LLMs) remains an important challenge. While continual training provides a potential avenue for

researcharxiv-cs-cl
30 Jun 2026
Safety

Phonological Perception of Sign Language Models

DGX agent

arXiv:2606.28667v1 Announce Type: new Abstract: Sign languages are compositional systems where meaning arises by combining sublexical phonological parameters, such as handshape, location, and movement

safetyarxiv-cs-cl
30 Jun 2026
Research

Poller: Are LLMs Suitable for Evaluating the Poetry Understanding Task?

DGX agent

arXiv:2606.30556v1 Announce Type: new Abstract: Traditional automatic evaluation methods have been shown to be unsuitable for modern Chinese poetry because of the distinct nature of this literary genr

researcharxiv-cs-cl
30 Jun 2026
Research

Preference-ASR: A Preference-Aware Test Set for Benchmarking ASR in the Era of Speech LLMs

DGX agent

arXiv:2606.29534v1 Announce Type: new Abstract: Popular ASR test sets adopt inconsistent conventions for numbers, disfluencies, entities, and casing, while standard normalizers erase the format distin

researcharxiv-cs-cl
30 Jun 2026
Safety

Preserving Fairness and Safety in Quantized LLMs Through Critical Weight Protection

DGX agent

arXiv:2601.12033v2 Announce Type: replace Abstract: Quantization is widely adopted to reduce the computational cost of large language models (LLMs); however, its implications for fairness and safety,

safetyarxiv-cs-cl
30 Jun 2026
Safety

REAR: Test-time Preference Realignment through Reward Decomposition

DGX agent

arXiv:2606.30339v1 Announce Type: new Abstract: Aligning large language models (LLMs) with diverse user preferences is a critical yet challenging task. While post-training methods can adapt models to

safetyarxiv-cs-cl
30 Jun 2026
Safety

Regime-Aware Peer Specialization for Robust RAG under Heterogeneous Knowledge Conflicts

DGX agent

arXiv:2606.30518v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) improves language models by grounding generation in external context. However, it can be fragile when the retrieved

safetyarxiv-cs-cl
30 Jun 2026
Model Releases

Representational Depth of Evaluation Awareness Shifts With Scale in Open-Weight Language Models

DGX agent

arXiv:2606.29196v1 Announce Type: cross Abstract: Do language models know when they are being tested? This question matters for AI safety: a model that recognises an evaluation context could alter its

model-releasesarxiv-cs-cl
30 Jun 2026
Safety

Resolution Thresholds in VLM Detection of Harmful ASCII Art Across Construction Modes and Languages

DGX agent

arXiv:2606.29649v1 Announce Type: new Abstract: Large Vision-Language Models (VLMs) are increasingly deployed as content moderation tools, yet they remain vulnerable to jailbreak attacks in which harm

safetyarxiv-cs-cl
30 Jun 2026
Research

Revealing the Technology Development of Natural Language Processing: A Scientific Entity-Centric Perspective

DGX agent

arXiv:2606.29836v1 Announce Type: new Abstract: Most studies on technology development have been conducted from a thematic perspective, but the topics are coarse-grained and insufficient to accurately

researcharxiv-cs-cl
30 Jun 2026
Model Releases

Scaling the Horizon, Not the Parameters: Reaching Trillion-Parameter Performance with a 35B Agent

DGX agent

arXiv:2606.30616v1 Announce Type: new Abstract: We introduce Agents-A1, a 35B Mixture-of-Experts Agentic Model that reaches trillion-parameter-level performance by scaling the agent horizon. We invest

model-releasesarxiv-cs-cl
30 Jun 2026
Safety

SEAD: Competence-Aware On-Policy Distillation via Entropy-Guided Supervision

DGX agent

arXiv:2606.28562v1 Announce Type: new Abstract: On-policy distillation (OPD) has a property absent in offline distillation and RL: teacher supervision quality depends on student competence. Incoherent

safetyarxiv-cs-cl
30 Jun 2026
Tutorials

See, Think, Learn: A Self-Taught Multimodal Reasoner

DGX agent

arXiv:2512.02456v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) have achieved remarkable progress in integrating visual perception with language understanding. However, effecti

tutorialsarxiv-cs-cl
30 Jun 2026
Model Releases

SHOVIR: A Benchmark for Evaluating Vision Shortcut Learning in Radiology Report Generation

DGX agent

arXiv:2606.30201v1 Announce Type: cross Abstract: Current evaluation protocols for Vision-Language Models (VLMs) in Radiology Report Generation (RRG) rely on report-level metrics that measure lexical

model-releasesarxiv-cs-cl
30 Jun 2026
Research

Sifei at SemEval-2026 Task 8: Hybrid Retrieval and Query Rewriting for Multi-Turn RAG

DGX agent

arXiv:2606.28352v1 Announce Type: cross Abstract: Multi-turn retrieval-augmented generation (RAG) is challenging due to evolving user intent, conversational noise, and strict context limits. We propos

researcharxiv-cs-cl
30 Jun 2026
Local Ai

Smooth Scaling Laws Hide Stepwise Token Learning

DGX agent

arXiv:2606.29858v1 Announce Type: new Abstract: Language model loss follows remarkably regular scaling laws over model and data size, yet it remains unclear why the aggregate loss should exhibit a pow

local-aiarxiv-cs-cl
30 Jun 2026
Research

SPARKLING: Balancing Signal Preservation and Symmetry Breaking for Width-Progressive Learning

DGX agent

arXiv:2602.02472v2 Announce Type: replace-cross Abstract: Progressive Learning (PL) reduces pre-training computational overhead by gradually increasing model scale. While prior work has extensively ex

researcharxiv-cs-cl
30 Jun 2026
Safety

Sparse Autoencoders are Capable LLM Jailbreak Mitigators

DGX agent

arXiv:2602.12418v2 Announce Type: replace-cross Abstract: Jailbreak attacks remain a persistent threat to large language model safety. We propose Context-Conditioned Delta Steering (CC-Delta), an SAE-

safetyarxiv-cs-cl
30 Jun 2026
Safety

SpecMind: Cognitively Inspired, Interactive Multi-Turn Framework for Postcondition Inference

DGX agent

arXiv:2602.20610v3 Announce Type: replace-cross Abstract: Specifications are vital for ensuring program correctness, yet writing them manually remains challenging and time-intensive. Recent large lang

safetyarxiv-cs-cl
30 Jun 2026
Model Releases

SrDetection: A Self-Referential Framework for Data Leakage Detection in Code Large Language Models

DGX agent

arXiv:2606.29815v1 Announce Type: new Abstract: Evaluating code large language models (Code LLMs) requires reliable detection of data leakage, where benchmark performance is artificially inflated by e

model-releasesarxiv-cs-cl
30 Jun 2026
Safety

SRUM: Fine-Grained Self-Rewarding for Unified Multimodal Models

DGX agent

arXiv:2510.12784v2 Announce Type: replace-cross Abstract: Recently, remarkable progress has been made in Unified Multimodal Models (UMMs), which integrate vision-language generation and understanding

safetyarxiv-cs-cl
30 Jun 2026
Safety

Structure-Preserving Document Translation via Multi-Stage LLM Pipeline: A Case Study in Marathi

DGX agent

arXiv:2606.28796v1 Announce Type: new Abstract: Government documents in India are predominantly issued in regional languages such as Marathi, creating substantial accessibility barriers for non-native

safetyarxiv-cs-cl
30 Jun 2026
Research

Supporting Workflow Reproducibility by Linking Bioinformatics Tools across Papers and Executable Code

DGX agent

arXiv:2603.08195v2 Announce Type: replace Abstract: Motivation: The rapid growth of biological data has intensified the need for transparent, reproducible, and well-documented computational workflows.

researcharxiv-cs-cl
30 Jun 2026
Local Ai

SurrogateShield: Beyond Redaction for High-Utility, Privacy-Preserving LLM Interactions

DGX agent

arXiv:2606.29567v1 Announce Type: cross Abstract: LLM-based assistants transmit user queries verbatim to third-party API endpoints that lie outside the user's audit or control. When those queries cont

local-aiarxiv-cs-cl
30 Jun 2026
Safety

The Alignment Auditor: A Bayesian Framework for Verifying and Refining LLM Objectives

DGX agent

arXiv:2510.06096v3 Announce Type: replace-cross Abstract: The objectives that Large Language Models (LLMs) implicitly optimize remain dangerously opaque, making trustworthy alignment and auditing a gr

safetyarxiv-cs-cl
30 Jun 2026
Model Releases

The Digital Afterlife of Empires: Four Language Models Converge on the Same Imperial Cartography of Writing

DGX agent

arXiv:2606.28325v1 Announce Type: cross Abstract: Large language models process the world's writing systems with radical inequality. We constructed the Digital Script Representation Index (DSRI), a se

model-releasesarxiv-cs-cl
30 Jun 2026
Research

The Effect of Scripts and Formats on LLM Numeracy

DGX agent

arXiv:2601.15251v2 Announce Type: replace Abstract: Large language models (LLMs) have achieved impressive proficiency in basic arithmetic, rivaling human-level performance on standard numerical tasks.

researcharxiv-cs-cl
30 Jun 2026
Tutorials

The Hidden Cost of Resampling: How Imbalance Correction Degrades Probability Calibration in Tree Ensembles

DGX agent

arXiv:2606.29720v1 Announce Type: cross Abstract: Resampling methods such as SMOTE and random under/over-sampling are standard tools for class-imbalanced classification, almost always evaluated by min

tutorialsarxiv-cs-cl
30 Jun 2026
Model Releases

The NTNU System at the S&I Challenge 2025 SLA Open Track

DGX agent

arXiv:2506.05121v3 Announce Type: replace Abstract: A recent line of research on spoken language assessment (SLA) employs neural models such as BERT and wav2vec 2.0 (W2V) to evaluate speaking proficie

model-releasesarxiv-cs-cl
30 Jun 2026
Research

ThinkProbe: Beyond Accuracy -- Structural Profiling of Open-Ended LLM Reasoning Traces via Non-Generative Thought Graphs

DGX agent

arXiv:2606.29067v1 Announce Type: new Abstract: We present ThinkProbe, a framework for structural analysis of LLM reasoning traces. ThinkProbe converts each trace into a Thought Graph a directed graph

researcharxiv-cs-cl
30 Jun 2026
Model Releases

Thunder-KoNUBench: A Corpus-Aligned Benchmark for Korean Negation Understanding

DGX agent

arXiv:2601.04693v2 Announce Type: replace Abstract: Although negation is known to challenge large language models (LLMs), benchmarks for evaluating negation understanding-especially in Korean-are scar

model-releasesarxiv-cs-cl
30 Jun 2026
Safety

Timesteps of Mamba Align with Human Reading Times

DGX agent

arXiv:2606.29904v1 Announce Type: new Abstract: This study demonstrates an alignment of per-word processing time in a popular state-space language model Mamba and human readers. In Mamba, the recurren

safetyarxiv-cs-cl
30 Jun 2026
Safety

Towards Physical Intuitions for Alignment Dynamics: A Case Study With Randomness Crystallization

DGX agent

arXiv:2606.29933v1 Announce Type: new Abstract: The alignment of language models is typically studied through the lens of capability benchmarks, but the dynamics of how models change during post-train

safetyarxiv-cs-cl
30 Jun 2026
← Previous
1…3839404142…161
Next →