AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlog
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,759 results
Tutorials

Task-Routed Mixture-of-Experts with Cognitive Appraisal for Implicit Sentiment Analysis

DGX agent

arXiv:2605.20916v1 Announce Type: new Abstract: Implicit sentiment analysis is challenging because sentiment toward an aspect is often inferred from events rather than expressed through explicit opini

tutorialsarxiv-cs-cl
21 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Terminal-World: Scaling Terminal-Agent Environments via Agent Skills

DGX agent

arXiv:2605.20876v1 Announce Type: new Abstract: Terminal agents extend Large Language Models with the ability to execute tasks directly in command-line environments, but their progress is bottlenecked

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Text Analytics Evaluation Framework: A Case Study on LLMs and Social Media

DGX agent

arXiv:2605.21338v1 Announce Type: new Abstract: LLMs have demonstrated exceptional proficiency in a wide range of NLP tasks. However, a notable gap remains in practical data analysis scenarios, partic

model-releasesarxiv-cs-cl
21 May 2026
Research

TextReg: Mitigating Prompt Distributional Overfitting via Regularized Text-Space Optimization

DGX agent

arXiv:2605.21318v1 Announce Type: new Abstract: Large language models (LLMs) are highly sensitive to the prompts used to specify task objectives and behavioral constraints. Many recent prompt optimiza

researcharxiv-cs-cl
21 May 2026
Applications

The Generation-Recognition Asymmetry: Six Dimensions of a Fundamental Divide in Formal Language Theory

DGX agent

arXiv:2603.10139v2 Announce Type: replace Abstract: Every formal grammar defines a language and can in principle be used in three ways: to generate strings (production), to recognize them (parsing), o

applicationsarxiv-cs-cl
21 May 2026
Research

The Hidden Signal of Verifier Strictness: Controlling and Improving Step-Wise Verification via Selective Latent Steering

DGX agent

arXiv:2605.20745v1 Announce Type: cross Abstract: Generative verifiers have emerged as a promising paradigm for step-wise verification, but their verification behavior is often poorly calibrated: they

researcharxiv-cs-cl
21 May 2026
Safety

The Illusion of Intervention: Your LLM-Simulated Experiment is an Observational Study

DGX agent

arXiv:2605.20767v1 Announce Type: new Abstract: Large language models (LLMs) show potential as simulators of human behavior, offering a scalable way to study responses to interventions. However, becau

safetyarxiv-cs-cl
21 May 2026
Model Releases

The Visual Iconicity Challenge: Evaluating Vision-Language Models on Sign Language Form-Meaning Mapping

DGX agent

arXiv:2510.08482v3 Announce Type: replace-cross Abstract: Iconicity, the resemblance between linguistic form and meaning, is pervasive in signed languages, offering a natural testbed for visual ground

model-releasesarxiv-cs-cl
21 May 2026
Research

Thinking-while-speaking: A Controlled, Interleaved Reasoning Method for Real-Time Speech Generation

DGX agent

arXiv:2605.20946v1 Announce Type: new Abstract: The thinking-while-speaking paradigm aims to make AI communication more human. A key challenge is maintaining fluent speech while performing deep reason

researcharxiv-cs-cl
21 May 2026
Safety

Towards Context-Invariant Safety Alignment for Large Language Models

DGX agent

arXiv:2605.20994v1 Announce Type: new Abstract: Preference-based post-training aligns LLMs with human intent, yet safety behavior often remains brittle. A model may refuse a harmful request in a stand

safetyarxiv-cs-cl
21 May 2026
Applications

Towards the Anonymization of the Language Modeling

DGX agent

arXiv:2501.02407v3 Announce Type: replace Abstract: Rapid advances in Natural Language Processing (NLP) have revolutionized many fields, including healthcare. However, these advances raise significant

applicationsarxiv-cs-cl
21 May 2026
Model Releases

Toxic Subword Pruning for Dialogue Response Generation on Large Language Models

DGX agent

arXiv:2410.04155v2 Announce Type: replace Abstract: How to defend large language models (LLMs) from generating toxic content is an important research area. Yet, most research focused on various model

model-releasesarxiv-cs-cl
21 May 2026
Research

Tracing the ongoing emergence of human-like reasoning in Large Language Models

DGX agent

arXiv:2605.21299v1 Announce Type: new Abstract: Humans effortlessly go beyond literal meanings: If you mow the lawn, I will give you fifty dollars, is typically understood as implying that the speaker

researcharxiv-cs-cl
21 May 2026
Model Releases

Training Language Agents to Learn from Experience

DGX agent

arXiv:2605.20477v1 Announce Type: cross Abstract: Language agents can adapt from experience in interactive environments, but current reflection-based methods can only self-correct within a single task

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Under Pressure: Emotional Framing Induces Measurable Behavioral Shifts and Structured Internal Geometry in Small Language Models

DGX agent

arXiv:2605.20202v1 Announce Type: new Abstract: I study whether emotionally framed evaluation follow-ups change both the behavior and the calm-relative internal representations of small, locally deplo

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs

DGX agent

arXiv:2505.19075v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have demonstrated remarkable general capabilities, but enhancing skills such as reasoning often demands substanti

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

WCXB: A Multi-Type Web Content Extraction Benchmark

DGX agent

arXiv:2605.21097v1 Announce Type: new Abstract: Web content extraction - isolating a page's main content from surrounding boilerplate - is a prerequisite for search indexing, retrieval-augmented gener

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

What Do Biomedical NER and Entity Linking Benchmarks Measure? A Corpus-Centric Diagnostic Framework

DGX agent

arXiv:2605.20537v1 Announce Type: new Abstract: Biomedical named entity recognition (NER) and entity linking (EL) strongly depend on annotated corpora, but the utility of these resources for benchmark

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

When Irregularity Helps: A Subclass Analysis of Inductive Bias in Neural Morphology

DGX agent

arXiv:2605.20558v1 Announce Type: new Abstract: Neural morphological generation systems often achieve high aggregate accuracy on benchmark datasets, yet such performance can conceal systematic errors

model-releasesarxiv-cs-cl
21 May 2026
Research

When Reasoning Supervision Hurts: TTCW-Based Long-Form Literary Review Generation

DGX agent

arXiv:2605.20364v1 Announce Type: new Abstract: Automatic evaluation of long-form literary writing remains challenging, as generic LLM-as-Judge approaches may not fully capture creativity-related dime

researcharxiv-cs-cl
21 May 2026
Research

You Are What You Say: Exploiting Linguistic Content for VoicePrivacy Attacks

DGX agent

arXiv:2506.09521v2 Announce Type: replace-cross Abstract: Speaker anonymization systems hide the identity of speakers while preserving other information such as linguistic content and emotions. To eva

researcharxiv-cs-cl
21 May 2026
Model Releases

You Only Need Minimal RLVR Training: Extrapolating LLMs via Rank-1 Trajectories

DGX agent

arXiv:2605.21468v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has become a dominant paradigm for improving reasoning in large language models (LLMs), yet the

model-releasesarxiv-cs-cl
21 May 2026
Research

A Data-Driven Approach to Idiomaticity Based on Experts' Criteria in Theoretical Linguistics

DGX agent

arXiv:2605.19575v1 Announce Type: new Abstract: The article observes data analysis of 286 multi-word expressions (MWEs) based on 16 lexical, grammatical and other criteria described in theoretical boo

researcharxiv-cs-cl
20 May 2026
Agents

A Multi-Agent Framework for Feature-Constrained Difficulty Control in Reading Comprehension Item Generation

DGX agent

arXiv:2605.19316v1 Announce Type: new Abstract: Recent studies in difficulty-controlled reading comprehension item generation have leveraged large language models (LLMs) to produce items by adjusting

agentsarxiv-cs-cl
20 May 2026
Model Releases

Acoustic scattering AI for non-invasive object classifications: A case study on hair assessment

DGX agent

arXiv:2506.14148v2 Announce Type: replace-cross Abstract: This paper presents a novel non-invasive object classification approach using acoustic scattering, demonstrated through a case study on hair a

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

Agent Meltdowns: The Road to Hell Is Paved with Helpful Agents

DGX agent

arXiv:2605.19149v1 Announce Type: new Abstract: Agents operating with computer and Web use inevitably encounter errors: inaccessible webpages, missing files, local and remote misconfigurations, etc. T

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

An LLM-Based System for Argument Mining

DGX agent

arXiv:2605.13793v2 Announce Type: replace Abstract: Arguments are a fundamental aspect of human reasoning, in which claims are supported, challenged, and weighed against one another. We present an end

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

Are Tools Always Beneficial? Learning to Invoke Tools Adaptively for Dual-Mode Multimodal LLM Reasoning

DGX agent

arXiv:2605.19852v1 Announce Type: new Abstract: Tool-augmented reasoning has emerged as a promising direction for enhancing the reasoning capabilities of multimodal large language models (MLLMs). Howe

model-releasesarxiv-cs-cl
20 May 2026
Applications

CAIT: A Syntactic Parsing Toolkit for Child-Adult InTeractions

DGX agent

arXiv:2605.19718v1 Announce Type: new Abstract: CHILDES is a paramount resource for language acquisition studies -- yet computational tools for analyzing its syntactic structure remain limited. Levera

applicationsarxiv-cs-cl
20 May 2026
Model Releases

Can Large Language Models Reliably Correct Errors in Low-Resource ASR? A Contamination-Aware Case Study on West Frisian

DGX agent

arXiv:2605.19711v1 Announce Type: new Abstract: Automatic speech recognition (ASR) has improved substantially in recent years, yet performance remains limited for low-resource languages. Large languag

model-releasesarxiv-cs-cl
20 May 2026
Research

Can LLMs Estimate Cognitive Complexity of Reading Comprehension Items?

DGX agent

arXiv:2510.25064v2 Announce Type: replace Abstract: Estimating the cognitive complexity of reading comprehension (RC) items is crucial for assessing item difficulty before it is administered to learne

researcharxiv-cs-cl
20 May 2026
Safety

CEPO: RLVR Self-Distillation using Contrastive Evidence Policy Optimization

DGX agent

arXiv:2605.19436v1 Announce Type: cross Abstract: When a model produces a correct solution under reinforcement learning with verifiable rewards (RLVR), every token receives the same reward signal rega

safetyarxiv-cs-cl
20 May 2026
Applications

CLIF: Concept-Level Influence Functions for Transparent Bottleneck Models

DGX agent

arXiv:2605.19848v1 Announce Type: new Abstract: In recent years, the black-box nature of deep learning models has limited their application in high-stakes domains such as medical diagnosis and finance

applicationsarxiv-cs-cl
20 May 2026
Model Releases

ClinSeekAgent: Automating Multimodal Evidence Seeking for Agentic Clinical Reasoning

DGX agent

arXiv:2605.20176v1 Announce Type: new Abstract: Large language models (LLMs) and agentic systems have shown promise for clinical decision support, but existing works largely assume that evidence has a

model-releasesarxiv-cs-cl
20 May 2026
Research

Context-Aware Detection and Victim-Centered Response Generation for Online Harassment in Private Messaging

DGX agent

arXiv:2512.14700v2 Announce Type: replace-cross Abstract: Online harassment is a widespread social and public health concern, yet most computational approaches for detecting and addressing harassment

researcharxiv-cs-cl
20 May 2026
Tutorials

Critique-Guided Distillation for Robust Reasoning via Refinement

DGX agent

arXiv:2505.11628v4 Announce Type: replace Abstract: Supervised fine-tuning with expert demonstrations often produces models that imitate outputs without internalizing the reasoning processes needed fo

tutorialsarxiv-cs-cl
20 May 2026
Safety

Cross-modal Consistency Guidance for Robust Emotion Control in Auto-Regressive TTS Models

DGX agent

arXiv:2510.13293v3 Announce Type: replace Abstract: While Text-to-Speech (TTS) systems enable emotional control via natural-language instructions, expressiveness, naturalness, and speech quality degra

safetyarxiv-cs-cl
20 May 2026
Research

Cubit: Token Mixer with Kernel Ridge Regression

DGX agent

arXiv:2605.06501v2 Announce Type: replace-cross Abstract: Since its introduction in 2017, the Transformer has become one of the most widely adopted architectures in modern deep learning. Despite exten

researcharxiv-cs-cl
20 May 2026
Agents

DECOR: Auditing LLM Deception via Information Manipulation Theory

DGX agent

arXiv:2605.19270v1 Announce Type: new Abstract: Large language models can deceive by subtly manipulating truthful information -- omitting key facts, shifting focus, or obscuring meaning -- making such

agentsarxiv-cs-cl
20 May 2026
Research

Difficulty-Controllable Cloze Question Distractor Generation

DGX agent

arXiv:2511.01526v2 Announce Type: replace Abstract: Multiple-choice cloze questions are commonly used to assess linguistic proficiency and comprehension. However, generating high-quality distractors r

researcharxiv-cs-cl
20 May 2026
Tutorials

Drifting Objectives for Refining Discrete Diffusion Language Models

DGX agent

arXiv:2605.19470v1 Announce Type: new Abstract: Discrete diffusion language models (DDLMs) generate text by iteratively denoising categorical token sequences, while recent drifting methods for continu

tutorialsarxiv-cs-cl
20 May 2026
Research

ECG-R1: Protocol-Guided and Modality-Agnostic MLLM for Reliable ECG Interpretation

DGX agent

arXiv:2602.04279v2 Announce Type: replace Abstract: Electrocardiography (ECG) serves as an indispensable diagnostic tool in clinical practice, yet existing multimodal large language models (MLLMs) rem

researcharxiv-cs-cl
20 May 2026
Research

Efficient Pre-Training with Token Superposition

DGX agent

arXiv:2605.06546v2 Announce Type: replace Abstract: Pre-training of Large Language Models is often prohibitively expensive and inefficient at scale, requiring complex and invasive modifications in ord

researcharxiv-cs-cl
20 May 2026
Research

EnsemHalDet: Robust VLM Hallucination Detection via Ensemble of Internal State Detectors

DGX agent

arXiv:2604.02784v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) excel at multimodal tasks, but they remain vulnerable to hallucinations that are factually incorrect or unground

researcharxiv-cs-cl
20 May 2026
Model Releases

Federated Learning for ICD Classification with Lightweight Models and Pretrained Embeddings

DGX agent

arXiv:2507.03122v2 Announce Type: replace-cross Abstract: This study investigates the feasibility and performance of federated learning (FL) for multi-label ICD code classification using clinical note

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

FedMental: Evaluating Federated Learning for Mental Health Detection from Social Media Data

DGX agent

arXiv:2605.18936v1 Announce Type: cross Abstract: Social media text data are often used to train Machine Learning (ML) models to identify users exhibiting high-risk mental health behaviors. However, s

model-releasesarxiv-cs-cl
20 May 2026
Research

Fine-tuning language encoding models on slow fMRI improves prediction for fast ECoG

DGX agent

arXiv:2605.19224v1 Announce Type: new Abstract: Neuroscientists have recently turned to intracranial brain recording methods, like electrocorticography (ECoG), for human experiments because of the fin

researcharxiv-cs-cl
20 May 2026
Research

Fingerprinting LLMs via Prompt Injection

DGX agent

arXiv:2509.25448v3 Announce Type: replace-cross Abstract: Large language models (LLMs) are often modified after release through post-processing such as post-training or quantization, which makes it ch

researcharxiv-cs-cl
20 May 2026
← Previous
1…8687888990…162
Next →