AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
Model Releases

Entropy of Ukrainian

DGX agent

arXiv:2604.27534v1 Announce Type: new Abstract: In natural language processing, the entropy of a language is a measure of its unpredictability and complexity. The first study on this subject was condu

model-releasesarxiv-cs-cl
1 May 2026
Research

EviMem: Evidence-Gap-Driven Iterative Retrieval for Long-Term Conversational Memory

DGX agent
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.27695v1 Announce Type: cross Abstract: Long-term conversational memory requires retrieving evidence scattered across multiple sessions, yet single-pass retrieval fails on temporal and multi

researcharxiv-cs-cl
1 May 2026
Safety

Exploration Hacking: Can LLMs Learn to Resist RL Training?

DGX agent

arXiv:2604.28182v1 Announce Type: cross Abstract: Reinforcement learning (RL) has become essential to the post-training of large language models (LLMs) for reasoning, agentic capabilities and alignmen

safetyarxiv-cs-cl
1 May 2026
Safety

Exploring Applications of Transfer-State Large Language Models: Cognitive Profiling and Socratic AI Tutoring

DGX agent

arXiv:2604.27454v1 Announce Type: new Abstract: Large language models (LLMs) sometimes exhibit qualitative shifts in response style under sustained self-referential dialogue conditions (Berg et al., 2

safetyarxiv-cs-cl
1 May 2026
Model Releases

Exploring the Limits of Pruning: Task-Specific Neurons, Model Collapse, and Recovery in Task-Specific Large Language Models

DGX agent

arXiv:2604.27115v1 Announce Type: new Abstract: Neuron pruning is widely used to reduce the computational cost and parameter footprint of large language models, yet it remains unclear whether neurons

model-releasesarxiv-cs-cl
1 May 2026
Research

From Coarse to Fine: Benchmarking and Reward Modeling for Writing-Centric Generation Tasks

DGX agent

arXiv:2604.27453v1 Announce Type: new Abstract: Large language models have achieved remarkable progress in text generation but still struggle with generative writing tasks. In terms of evaluation, exi

researcharxiv-cs-cl
1 May 2026
Applications

From Unstructured to Structured: LLM-Guided Attribute Graphs for Entity Search and Ranking

DGX agent

arXiv:2604.27410v1 Announce Type: cross Abstract: Entity search, i.e., finding the most similar entities to a query entity, faces unique challenges in e-commerce, where product similarity varies acros

applicationsarxiv-cs-cl
1 May 2026
Research

Geometry-Calibrated Conformal Abstention for Language Models

DGX agent

arXiv:2604.27914v1 Announce Type: new Abstract: When language models lack relevant knowledge for a given query, they frequently generate plausible responses that can be hallucinations, rather than adm

researcharxiv-cs-cl
1 May 2026
Research

HATS: An Open data set Integrating Human Perception Applied to the Evaluation of Automatic Speech Recognition Metrics

DGX agent

arXiv:2604.27542v1 Announce Type: new Abstract: Conventionally, Automatic Speech Recognition (ASR) systems are evaluated on their ability to correctly recognize each word contained in a speech signal.

researcharxiv-cs-cl
1 May 2026
Model Releases

HealthBench Professional: Evaluating Large Language Models on Real Clinician Chats

DGX agent

arXiv:2604.27470v1 Announce Type: new Abstract: Millions of clinicians use ChatGPT to support clinical care, but evaluations of the most common use cases in model-clinician conversations are limited.

model-releasesarxiv-cs-cl
1 May 2026
Research

Hypencoder Revisited: Reproducibility and Analysis of Non-Linear Scoring for First-Stage Retrieval

DGX agent

arXiv:2604.27037v1 Announce Type: cross Abstract: The Hypencoder, proposed by Killingback et al., is a retrieval framework that replaces the fixed inner-product scoring function used in standard bi-en

researcharxiv-cs-cl
1 May 2026
Safety

In-context Learning vs. Instruction Tuning: The Case of Small and Multilingual Language Models

DGX agent

arXiv:2503.01611v3 Announce Type: replace Abstract: Instruction following is a critical ability for Large Language Models to perform downstream tasks. The standard approach to instruction tuning has r

safetyarxiv-cs-cl
1 May 2026
Applications

JaiTTS: A Thai Voice Cloning Model

DGX agent

arXiv:2604.27607v1 Announce Type: new Abstract: We present JaiTTS-v1.0, a state-of-the-art Thai voice cloning text-to-speech model built through continual training on a large Thai-centric speech corpu

applicationsarxiv-cs-cl
1 May 2026
Research

Language Ideologies in a Multilingual Society: An LLM-based Analysis of Luxembourgish News Comments

DGX agent

arXiv:2604.27661v1 Announce Type: new Abstract: Detecting language ideologies is a valuable yet complex task for understanding how identities are constructed through discourse. In Luxembourg's multicu

researcharxiv-cs-cl
1 May 2026
Safety

Latent-GRPO: Group Relative Policy Optimization for Latent Reasoning

DGX agent

arXiv:2604.27998v1 Announce Type: cross Abstract: Latent reasoning offers a more efficient alternative to explicit reasoning by compressing intermediate reasoning into continuous representations and s

safetyarxiv-cs-cl
1 May 2026
Research

Length Value Model: Scalable Value Pretraining for Token-Level Length Modeling

DGX agent

arXiv:2604.27039v1 Announce Type: new Abstract: Token serves as the fundamental unit of computation in modern autoregressive models, and generation length directly influences both inference cost and r

researcharxiv-cs-cl
1 May 2026
Safety

Linguistically Informed Multimodal Fusion for Vietnamese Scene-Text Image Captioning: Dataset, Graph Framework, and Phonological Attention

DGX agent

arXiv:2604.27712v1 Announce Type: cross Abstract: Scene-text image captioning requires fusing three information streams -- visual features, OCR-detected text, and linguistic knowledge -- to generate d

safetyarxiv-cs-cl
1 May 2026
Applications

LLM-based User Profile Management for Recommender System

DGX agent

arXiv:2502.14541v3 Announce Type: replace Abstract: The rapid advancement of Large Language Models (LLMs) has opened new opportunities in recommender systems by enabling zero-shot recommendation witho

applicationsarxiv-cs-cl
1 May 2026
Research

LLMs Capture Emotion Labels, Not Emotion Uncertainty: Distributional Analysis and Calibration of Human--LLM Judgment Gaps

DGX agent

arXiv:2604.27345v1 Announce Type: new Abstract: Human annotators frequently disagree on emotion labels, yet most evaluations of Large Language Model (LLM) emotion annotation collapse these judgments i

researcharxiv-cs-cl
1 May 2026
Model Releases

M-DaQ: Retrieving Samples with Multilingual Diversity and Quality for Instruction Fine-Tuning Datasets

DGX agent

arXiv:2509.15549v2 Announce Type: replace Abstract: Multilingual instruction fine-tuning (IFT) empowers large language models to generalize across diverse linguistic and cultural contexts; however, hi

model-releasesarxiv-cs-cl
1 May 2026
Research

Measuring research data reuse in scholarly publications using generative artificial intelligence: Open Science Indicator development and preliminary results

DGX agent

arXiv:2604.28061v1 Announce Type: cross Abstract: Numerous metascience studies and other initiatives have begun to monitor the prevalence of open science practices when it is more important to underst

researcharxiv-cs-cl
1 May 2026
Model Releases

MiniCPM-o 4.5: Towards Real-Time Full-Duplex Omni-Modal Interaction

DGX agent

arXiv:2604.27393v1 Announce Type: new Abstract: Recent progress in multimodal large language models (MLLMs) has brought AI capabilities from static offline data processing to real-time streaming inter

model-releasesarxiv-cs-cl
1 May 2026
Model Releases

Models Recall What They Violate: Constraint Adherence in Multi-Turn LLM Ideation

DGX agent

arXiv:2604.28031v1 Announce Type: new Abstract: When researchers iteratively refine ideas with large language models, do the models preserve fidelity to the original objective? We introduce DriftBench

model-releasesarxiv-cs-cl
1 May 2026
Local Ai

Multi-Level Narrative Evaluation Outperforms Lexical Features for Mental Health

DGX agent

arXiv:2604.27846v1 Announce Type: new Abstract: How people narrate their experiences offers a window into how the mind organizes them. Computational approaches to therapeutic writing have evolved from

local-aiarxiv-cs-cl
1 May 2026
Model Releases

MultiBLiMP 1.0: A Massively Multilingual Benchmark of Linguistic Minimal Pairs

DGX agent

arXiv:2504.02768v4 Announce Type: replace Abstract: We introduce MultiBLiMP 1.0, a massively multilingual benchmark of linguistic minimal pairs, covering 101 languages and 2 types of subject-verb agre

model-releasesarxiv-cs-cl
1 May 2026
Research

On the Proper Treatment of Units in Surprisal Theory

DGX agent

arXiv:2604.28147v1 Announce Type: new Abstract: Surprisal theory links human processing effort to the predictability of an upcoming linguistic unit, but empirical work often leaves the notion of a uni

researcharxiv-cs-cl
1 May 2026
Model Releases

Perturbation Probing: A Two-Pass-per-Prompt Diagnostic for FFN Behavioral Circuits in Aligned LLMs

DGX agent

arXiv:2604.27401v1 Announce Type: new Abstract: Perturbation probing generates task-specific causal hypotheses for FFN neurons in large language models using two forward passes per prompt and no backp

model-releasesarxiv-cs-cl
1 May 2026
Research

PPA-Plan: Proactive Pitfall Avoidance for Reliable Planning in Long-Context LLM Reasoning

DGX agent

arXiv:2601.11908v2 Announce Type: replace Abstract: Large language models (LLMs) struggle with reasoning over long contexts where relevant information is sparsely distributed. Although plan-and-execut

researcharxiv-cs-cl
1 May 2026
Research

Proactive Dialogue Model with Intent Prediction

DGX agent

arXiv:2604.27379v1 Announce Type: new Abstract: Dialogue models are inherently reactive, responding to the current user turn without anticipating upcoming intents, which leads to redundant interaction

researcharxiv-cs-cl
1 May 2026
Research

Qualitative Evaluation of Language Model Rescoring in Automatic Speech Recognition

DGX agent

arXiv:2604.27533v1 Announce Type: new Abstract: Evaluating automatic speech recognition (ASR) systems is a classical but difficult and still open problem, which often boils down to focusing only on th

researcharxiv-cs-cl
1 May 2026
Agents

Real-World Doctor Agent with Proactive Consultation through Multi-Agent Reinforcement Learning

DGX agent

arXiv:2505.19630v4 Announce Type: replace Abstract: Large language models (LLMs) struggle in real-world clinical consultations. Single-turn consultation systems require patients to describe all sympto

agentsarxiv-cs-cl
1 May 2026
Tutorials

Reasoning over Object Descriptions Improves Coreference Resolution in Task-Based Dialogue Systems

DGX agent

arXiv:2604.27850v1 Announce Type: new Abstract: Task-based dialogue systems assist users in achieving specific goals, such as executing actions or retrieving information, through natural language inte

tutorialsarxiv-cs-cl
1 May 2026
Model Releases

RoadMapper: A Multi-Agent System for Roadmap Generation of Solving Complex Research Problems

DGX agent

arXiv:2604.27616v1 Announce Type: new Abstract: People commonly leverage structured content to accelerate knowledge acquisition and research problem solving. Among these, roadmaps guide researchers th

model-releasesarxiv-cs-cl
1 May 2026
Research

ScaleBox: Enabling High-Fidelity and Scalable Code Verification for Large Language Models

DGX agent

arXiv:2604.27467v1 Announce Type: cross Abstract: Code sandboxes have emerged as a critical infrastructure for advancing the coding capabilities of large language models, providing verifiable feedback

researcharxiv-cs-cl
1 May 2026
Research

Selective Augmentation: Improving Universal Automatic Phonetic Transcription via G2P Bootstrapping

DGX agent

arXiv:2604.27204v1 Announce Type: new Abstract: In the field of universal automatic phonetic transcription (APT), clean and diverse training transcriptions are required. However, such high-quality dat

researcharxiv-cs-cl
1 May 2026
Research

Semantic Structure of Feature Space in Large Language Models

DGX agent

arXiv:2604.27169v1 Announce Type: new Abstract: We show that the geometric relations between semantic features in large language models' hidden states closely mirror human psychological associations.

researcharxiv-cs-cl
1 May 2026
Applications

Sentiment Analysis of AI Adoption in Indonesian Higher Education Using Machine Learning and Transformer-Based Models

DGX agent

arXiv:2604.27439v1 Announce Type: new Abstract: This study analyzes Indonesian student opinions on the adoption of artificial intelligence in higher education using two approaches: TF-IDF-based machin

applicationsarxiv-cs-cl
1 May 2026
Model Releases

Skills-Coach: A Self-Evolving Skill Optimizer via Training-Free GRPO

DGX agent

arXiv:2604.27488v1 Announce Type: new Abstract: We introduce Skills-Coach, a novel automated framework designed to significantly enhance the self-evolution of skills within Large Language Model (LLM)-

model-releasesarxiv-cs-cl
1 May 2026
Safety

Stable Behavior, Limited Variation: Persona Validity in LLM Agents for Urban Sentiment Perception

DGX agent

arXiv:2604.28048v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly used as proxies for human perception in urban analysis, yet it remains unclear whether persona prompting p

safetyarxiv-cs-cl
1 May 2026
Safety

Supercharging Agenda Setting Research: The ParlaCAP Dataset of 28 European Parliaments and a Scalable Multilingual LLM-Based Classification

DGX agent

arXiv:2602.16516v2 Announce Type: replace Abstract: This paper introduces ParlaCAP, a large-scale dataset for analyzing parliamentary agenda setting across Europe, and proposes a cost-effective method

safetyarxiv-cs-cl
1 May 2026
Research

Syntactically-guided Information Maintenance in Sentence Comprehension

DGX agent

arXiv:2604.27468v1 Announce Type: new Abstract: Maintaining information in context is essential in successful real-time language comprehension, but maintenance is cognitively costly and can slow proce

researcharxiv-cs-cl
1 May 2026
Model Releases

Targeted Linguistic Analysis of Sign Language Models with Minimal Translation Pairs

DGX agent

arXiv:2604.27232v1 Announce Type: new Abstract: Models of sign language have historically lagged behind those for spoken language (text and speech). Recent work has greatly improved their performance

model-releasesarxiv-cs-cl
1 May 2026
Research

To Diff or Not to Diff? Structure-Aware and Adaptive Output Formats for Efficient LLM-based Code Editing

DGX agent

arXiv:2604.27296v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used for code editing, yet the prevalent full-code generation paradigm suffers from severe efficiency bo

researcharxiv-cs-cl
1 May 2026
Applications

TwinGate: Stateful Defense against Decompositional Jailbreaks in Untraceable Traffic via Asymmetric Contrastive Learning

DGX agent

arXiv:2604.27861v1 Announce Type: cross Abstract: Decompositional jailbreaks pose a critical threat to large language models (LLMs) by allowing adversaries to fragment a malicious objective into a seq

applicationsarxiv-cs-cl
1 May 2026
Research

Universal statistical laws governing culinary design

DGX agent

arXiv:2604.28021v1 Announce Type: cross Abstract: Cooking is a cultural expression of human creativity that transcends geography and time through the orchestration of ingredients and techniques, much

researcharxiv-cs-cl
1 May 2026
Safety

Vision-Language Models Mistake Head Orientation for Gaze Direction: Nonverbal Conversation Cues

DGX agent

arXiv:2506.05412v3 Announce Type: replace-cross Abstract: Where someone looks is a nonverbal communication cue that children and adults readily use. How well can Vision-Language Models (VLMs) infer ga

safetyarxiv-cs-cl
1 May 2026
Model Releases

WebMall -- A Multi-Shop Benchmark for Evaluating Web Agents

DGX agent

arXiv:2508.13024v3 Announce Type: replace Abstract: LLM-based web agents have the potential to automate long-running web tasks, such as searching for products in multiple e-shops and subsequently orde

model-releasesarxiv-cs-cl
1 May 2026
Research

Why Mean Pooling Works: Quantifying Second-Order Collapse in Text Embeddings

DGX agent

arXiv:2604.27398v1 Announce Type: new Abstract: For constructing text embeddings, mean pooling, which averages token embeddings, is the standard approach. This paper examines whether mean pooling actu

researcharxiv-cs-cl
1 May 2026
← Previous
1…115116117118119…161
Next →