AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,646 results
14 Apr 2026

MCGA: A Multi-task Classical Chinese Literary Genre Audio Corpus

ResearchDGX agent

arXiv:2601.09270v3 Announce Type: replace Abstract: With the rapid advancement of Multimodal Large Language Models (MLLMs), their potential has gained significant attention in Chinese Classical Studie

Measuring What Matters!! Assessing Therapeutic Principles in Mental-Health Conversation

Model ReleasesDGX agent

arXiv:2604.05795v2 Announce Type: replace Abstract: The increasing use of large language models in mental health applications calls for principled evaluation frameworks that assess alignment with psyc

MEDSYN: Benchmarking Multi-EviDence SYNthesis in Complex Clinical Cases for Multimodal Large Language Models

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2602.21950v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) have shown great potential in medical applications, yet existing benchmarks inadequately capture real-world

MemDLM: Memory-Enhanced DLM Training

Model ReleasesDGX agent

arXiv:2603.22241v2 Announce Type: replace Abstract: Diffusion Language Models (DLMs) offer attractive advantages over Auto-Regressive (AR) models, such as full-attention parallel decoding and flexible

MerNav: A Highly Generalizable Memory-Execute-Review Framework for Zero-Shot Object Goal Navigation

ApplicationsDGX agent

arXiv:2602.05467v2 Announce Type: replace-cross Abstract: Visual Language Navigation (VLN) is one of the fundamental capabilities for embodied intelligence and a critical challenge that urgently needs

MimicLM: Zero-Shot Voice Imitation through Autoregressive Modeling of Pseudo-Parallel Speech Corpora

SafetyDGX agent

arXiv:2604.11552v1 Announce Type: cross Abstract: Voice imitation aims to transform source speech to match a reference speaker's timbre and speaking style while preserving linguistic content. A straig

MIXAR: Scaling Autoregressive Pixel-based Language Models to Multiple Languages and Scripts

ResearchDGX agent

arXiv:2604.11575v1 Announce Type: new Abstract: Pixel-based language models are gaining momentum as alternatives to traditional token-based approaches, promising to circumvent tokenization challenges.

NameBERT: Scaling Name-Based Nationality Classification with LLM-Augmented Open Academic Data

SafetyDGX agent

arXiv:2604.10401v1 Announce Type: new Abstract: Inferring nationality from personal names is a critical capability for equity and bias monitoring, personalization, and a valuable tool in biomedical an

Nationality encoding in language model hidden states: Probing culturally differentiated representations in persona-conditioned academic text

Model ReleasesDGX agent

arXiv:2604.10151v1 Announce Type: new Abstract: Large language models are increasingly used as writing tools and pedagogical resources in English for Academic Purposes, but it remains unclear whether

NOSE: Neural Olfactory-Semantic Embedding with Tri-Modal Orthogonal Contrastive Learning

SafetyDGX agent

arXiv:2604.10452v1 Announce Type: new Abstract: Olfaction lies at the intersection of chemical structure, neural encoding, and linguistic perception, yet existing representation methods fail to fully

Not All Denoising Steps Are Equal: Model Scheduling for Faster Masked Diffusion Language Models

ResearchDGX agent

arXiv:2604.02340v2 Announce Type: replace-cross Abstract: Recent advances in masked diffusion language models (MDLMs) narrow the quality gap to autoregressive LMs, but their sampling remains expensive

OccuBench: Evaluating AI Agents on Real-World Professional Tasks via Language World Models

Model ReleasesDGX agent

arXiv:2604.10866v1 Announce Type: new Abstract: AI agents are expected to perform professional work across hundreds of occupational domains (from emergency department triage to nuclear reactor safety

ODUTQA-MDC: A Task for Open-Domain Underspecified Tabular QA with Multi-turn Dialogue-based Clarification

Model ReleasesDGX agent

arXiv:2604.10159v1 Announce Type: new Abstract: The advancement of large language models (LLMs) has enhanced tabular question answering (Tabular QA), yet they struggle with open-domain queries exhibit

Omnimodal Dataset Distillation via High-order Proxy Alignment

Model ReleasesDGX agent

arXiv:2604.10666v1 Announce Type: cross Abstract: Dataset distillation compresses large-scale datasets into compact synthetic sets while preserving training performance, but existing methods are large

PatchRecall: Patch-Driven Retrieval for Automated Program Repair

ResearchDGX agent

arXiv:2604.10481v1 Announce Type: cross Abstract: Retrieving the correct set of files from a large codebase is a crucial step in Automated Program Repair (APR). High recall is necessary to ensure that

Pay Less Attention to Function Words for Free Robustness of Vision-Language Models

ResearchDGX agent

arXiv:2512.07222v3 Announce Type: replace-cross Abstract: To address the trade-off between robustness and performance for robust VLM, we observe that function words could incur vulnerability of VLMs a

Phonological distances for linguistic typology and the origin of Indo-European languages

ResearchDGX agent

arXiv:2604.11565v1 Announce Type: new Abstract: We show that short-range phoneme dependencies encode large-scale patterns of linguistic relatedness, with direct implications for quantitative typology

Physical Commonsense Reasoning for Lower-Resourced Languages and Dialects: a Study on Basque

ResearchDGX agent

arXiv:2602.14812v3 Announce Type: replace Abstract: Physical commonsense reasoning represents a fundamental capability of human intelligence, enabling individuals to understand their environment, pred

PICon: A Multi-Turn Interrogation Framework for Evaluating Persona Agent Consistency

SafetyDGX agent

arXiv:2603.25620v2 Announce Type: replace Abstract: Large language model (LLM)-based persona agents are rapidly being adopted as scalable proxies for human participants across diverse domains. Yet the

Please Make it Sound like Human: Encoder-Decoder vs. Decoder-Only Transformers for AI-to-Human Text Style Transfer

Model ReleasesDGX agent

arXiv:2604.11687v1 Announce Type: new Abstract: AI-generated text has become common in academic and professional writing, prompting research into detection methods. Less studied is the reverse: system

Polyglot Teachers: Evaluating Language Models for Multilingual Synthetic Data Generation

Model ReleasesDGX agent

arXiv:2604.11290v1 Announce Type: new Abstract: Synthesizing supervised finetuning (SFT) data from language models (LMs) to teach smaller models multilingual tasks has become increasingly common. Howe

Position-Agnostic Pre-Projection for Transformer Attention: Nonlinear Feature Construction and Content Skip Before Q/K/V

ResearchDGX agent

arXiv:2604.10791v1 Announce Type: new Abstract: We propose two complementary modifications to transformer attention blocks. First, a non-linear pre-projection MLP is inserted between layer norm and Q/

Preference Learning Unlocks LLMs' Psycho-Counseling Skills

ResearchDGX agent

arXiv:2502.19731v2 Announce Type: replace Abstract: Applying large language models (LLMs) to assist in psycho-counseling is an emerging and meaningful approach, driven by the significant gap between p

ProGAL-VLA: Grounded Alignment through Prospective Reasoning in Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2604.09824v1 Announce Type: cross Abstract: Vision language action (VLA) models enable generalist robotic agents but often exhibit language ignorance, relying on visual shortcuts and remaining i

ProUIE: A Macro-to-Micro Progressive Learning Method for LLM-based Universal Information Extraction

SafetyDGX agent

arXiv:2604.10633v1 Announce Type: new Abstract: LLM-based universal information extraction (UIE) methods often rely on additional information beyond the original training data, which increases trainin

Psychological Concept Neurons: Can Neural Control Bias Probing and Shift Generation in LLMs?

Local AiDGX agent

arXiv:2604.11802v1 Announce Type: new Abstract: Using psychological constructs such as the Big Five, large language models (LLMs) can imitate specific personality profiles and predict a user's persona

QFS-Composer: Query-focused summarization pipeline for less resourced languages

SafetyDGX agent

arXiv:2604.10687v1 Announce Type: new Abstract: Large language models (LLMs) demonstrate strong performance in text summarization, yet their effectiveness drops significantly across languages with res

Quantifying the Climate Risk of Generative AI: Region-Aware Carbon Accounting with G-TRACE and the AI Sustainability Pyramid

SafetyDGX agent

arXiv:2511.04776v2 Announce Type: replace-cross Abstract: Generative Artificial Intelligence (GenAI) represents a rapidly expanding digital infrastructure whose energy demand and associated CO2 emissi

RCBSF: A Multi-Agent Framework for Automated Contract Revision via Stackelberg Game

Model ReleasesDGX agent

arXiv:2604.10740v1 Announce Type: new Abstract: Despite the widespread adoption of Large Language Models (LLMs) in Legal AI, their utility for automated contract revision remains impeded by hallucinat

Reason Only When Needed: Efficient Generative Reward Modeling via Model-Internal Uncertainty

ResearchDGX agent

arXiv:2604.10072v1 Announce Type: new Abstract: Recent advancements in the Generative Reward Model (GRM) have demonstrated its potential to enhance the reasoning abilities of LLMs through Chain-of-Tho

Reasoning Resides in Layers: Restoring Temporal Reasoning in Video-Language Models with Layer-Selective Merging

SafetyDGX agent

arXiv:2604.11399v1 Announce Type: cross Abstract: Multimodal adaptation equips large language models (LLMs) with perceptual capabilities, but often weakens the reasoning ability inherited from languag

RedNote-Vibe: A Dataset for Capturing Temporal Dynamics of AI-Generated Text in Lifestyle Social Media

ResearchDGX agent

arXiv:2509.22055v2 Announce Type: replace Abstract: We introduce RedNote-Vibe, a dataset spanning five years (pre-LLM to July 2025) sourced from lifestyle platform RedNote (Xiaohongshu), capturing the

Relational Probing: LM-to-Graph Adaptation for Financial Prediction

HardwareDGX agent

arXiv:2604.10212v1 Announce Type: new Abstract: Language models can be used to identify relationships between financial entities in text. However, while structured output mechanisms exist, prompting-b

Relax: An Asynchronous Reinforcement Learning Engine for Omni-Modal Post-Training at Scale

Model ReleasesDGX agent

arXiv:2604.11554v1 Announce Type: new Abstract: Reinforcement learning (RL) post-training has proven effective at unlocking reasoning, self-reflection, and tool-use capabilities in large language mode

Reproduction Beyond Benchmarks: ConstBERT and ColBERT-v2 Across Backends and Query Distributions

ResearchDGX agent

arXiv:2604.09982v1 Announce Type: cross Abstract: Reproducibility must validate architectural robustness, not just numerical accuracy. We evaluate ColBERT-v2 and ConstBERT across five dimensions, find

Rethinking LLM Watermark Detection in Black-Box Settings: A Non-Intrusive Third-Party Framework

SafetyDGX agent

arXiv:2603.14968v2 Announce Type: replace-cross Abstract: While watermarking serves as a critical mechanism for LLM provenance, existing secret-key schemes tightly couple detection with injection, req

Revisiting Compositionality in Dual-Encoder Vision-Language Models: The Role of Inference

SafetyDGX agent

arXiv:2604.11496v1 Announce Type: cross Abstract: Dual-encoder Vision-Language Models (VLMs) such as CLIP are often characterized as bag-of-words systems due to their poor performance on compositional

Revisiting Epistemic Markers in Confidence Estimation: Can Markers Accurately Reflect Large Language Models' Uncertainty?

SafetyDGX agent

arXiv:2505.24778v3 Announce Type: replace Abstract: As large language models (LLMs) are increasingly used in high-stakes domains, accurately assessing their confidence is crucial. Humans typically exp

RiTeK: A Dataset for Large Language Models Complex Reasoning over Textual Knowledge Graphs in Medicine

Model ReleasesDGX agent

arXiv:2410.13987v3 Announce Type: replace Abstract: Answering complex real-world questions in the medical domain often requires accurate retrieval from medical Textual Knowledge Graphs (medical TKGs),

RUMLEM: A Dictionary-Based Lemmatizer for Romansh

ResearchDGX agent

arXiv:2604.11233v1 Announce Type: new Abstract: Lemmatization -- the task of mapping an inflected word form to its dictionary form -- is a crucial component of many NLP applications. In this paper, we

Saar-Voice: A Multi-Speaker Saarbrucken Dialect Speech Corpus

Local AiDGX agent

arXiv:2604.11803v1 Announce Type: new Abstract: Natural language processing (NLP) and speech technologies have made significant progress in recent years; however, they remain largely focused on standa

SafeConstellations: Mitigating Over-Refusals in LLMs Through Task-Aware Representation Steering

SafetyDGX agent

arXiv:2508.11290v3 Announce Type: replace Abstract: LLMs increasingly exhibit over-refusal behavior, where safety mechanisms cause models to reject benign instructions that seemingly resemble harmful

Seeing No Evil: Blinding Large Vision-Language Models to Safety Instructions via Adversarial Attention Hijacking

Model ReleasesDGX agent

arXiv:2604.10299v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) rely on attention-based retrieval of safety instructions to maintain alignment during generation. Existing attack

Seeing Through Deception: Uncovering Misleading Creator Intent in Multimodal News with Vision-Language Models

Model ReleasesDGX agent

arXiv:2505.15489v4 Announce Type: replace-cross Abstract: The impact of multimodal misinformation arises not only from factual inaccuracies but also from the misleading narratives that creators delibe

Self-Calibrating Language Models via Test-Time Discriminative Distillation

ResearchDGX agent

arXiv:2604.09624v1 Announce Type: new Abstract: Large language models (LLMs) are systematically overconfident: they routinely express high certainty on questions they often answer incorrectly. Existin

Self-Correcting RAG: Enhancing Faithfulness via MMKP Context Selection and NLI-Guided MCTS

ResearchDGX agent

arXiv:2604.10734v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) substantially extends the knowledge boundary of large language models. However, it still faces two major challenges

Self-Evolving LLM Memory Extraction Across Heterogeneous Tasks

Model ReleasesDGX agent

arXiv:2604.11610v1 Announce Type: new Abstract: As LLM-based assistants become persistent and personalized, they must extract and retain useful information from past conversations as memory. However,

Semantic-Space Exploration and Exploitation in RLVR for LLM Reasoning

ResearchDGX agent

arXiv:2509.23808v4 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) for LLM reasoning is often framed as balancing exploration and exploitation in action sp

SEPTQ: A Simple and Effective Post-Training Quantization Paradigm for Large Language Models

ResearchDGX agent

arXiv:2604.10091v1 Announce Type: new Abstract: Large language models (LLMs) have shown remarkable performance in various domains, but they are constrained by massive computational and storage costs.

SHARE: Social-Humanities AI for Research and Education

Model ReleasesDGX agent

arXiv:2604.11152v1 Announce Type: new Abstract: This intermediate technical report introduces the SHARE family of base models and the MIRROR user interface. The SHARE models are the first causal langu

Sign Language Recognition in the Age of LLMs

Model ReleasesDGX agent

arXiv:2604.11225v1 Announce Type: cross Abstract: Recent Vision Language Models (VLMs) have demonstrated strong performance across a wide range of multimodal reasoning tasks. This raises the question

Simulating Organized Group Behavior: New Framework, Benchmark, and Analysis

Model ReleasesDGX agent

arXiv:2604.09874v1 Announce Type: new Abstract: Simulating how organized groups (e.g., corporations) make decisions (e.g., responding to a competitor's move) is essential for understanding real-world

Single-Agent LLMs Outperform Multi-Agent Systems on Multi-Hop Reasoning Under Equal Thinking Token Budgets

Model ReleasesDGX agent

arXiv:2604.02460v2 Announce Type: replace Abstract: Recent work reports strong performance from multi-agent LLM systems (MAS), but these gains are often confounded by increased test-time computation.

SODA: Semi On-Policy Black-Box Distillation for Large Language Models

Model ReleasesDGX agent

arXiv:2604.03873v2 Announce Type: replace-cross Abstract: Black-box knowledge distillation for large language models presents a strict trade-off. Simple off-policy methods (e.g., sequence-level knowle

Solver-Independent Automated Problem Formulation via LLMs for High-Cost Simulation-Driven Design

ResearchDGX agent

arXiv:2512.18682v2 Announce Type: replace Abstract: In the high-cost simulation-driven design domain, translating ambiguous design requirements into a mathematical optimization formulation is a bottle

SpectralLoRA: Is Low-Frequency Structure Sufficient for LoRA Adaptation? A Spectral Analysis of Weight Updates

ResearchDGX agent

arXiv:2604.10649v1 Announce Type: cross Abstract: We present a systematic empirical study of the spectral structure of LoRA weight updates. Through 2D Discrete Cosine Transform (DCT) analysis of train

SpeechLess: Micro-utterance with Personalized Spatial Memory-aware Assistant in Everyday Augmented Reality

ResearchDGX agent

arXiv:2602.00793v2 Announce Type: replace-cross Abstract: Speaking aloud to a wearable AR assistant in public can be socially awkward, and re-articulating the same requests every day creates unnecessa

Spoiler Alert: Narrative Forecasting as a Metric for Tension in LLM Storytelling

Model ReleasesDGX agent

arXiv:2604.09854v1 Announce Type: new Abstract: LLMs have so far failed both to generate consistently compelling stories and to recognize this failure--on the leading creative-writing benchmark (EQ-Be

StableToken: A Noise-Robust Semantic Speech Tokenizer for Resilient SpeechLLMs

ResearchDGX agent

arXiv:2509.22220v2 Announce Type: replace Abstract: Prevalent semantic speech tokenizers, designed to capture linguistic content, are surprisingly fragile. We find they are not robust to meaning-irrel

Stop Fixating on Prompts: Reasoning Hijacking and Constraint Tightening for Red-Teaming LLM Agents

SafetyDGX agent

arXiv:2604.05549v2 Announce Type: replace Abstract: With the widespread application of LLM-based agents across various domains, their complexity has introduced new security threats. Existing red-team

← Previous
1…121122123124125…128
Next →