AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,646 results
14 Apr 2026

Structure-Grounded Knowledge Retrieval via Code Dependencies for Multi-Step Data Reasoning

AgentsDGX agent

arXiv:2604.10516v1 Announce Type: new Abstract: Selecting the right knowledge is critical when using large language models (LLMs) to solve domain-specific data analysis tasks. However, most retrieval-

Structured Causal Video Reasoning via Multi-Objective Alignment

SafetyDGX agent

arXiv:2604.04415v2 Announce Type: replace Abstract: Human understanding of video dynamics is typically grounded in a structured mental representation of entities, actions, and temporal relations, rath

STU-PID: Steering Token Usage via PID Controller for Efficient Large Language Model Reasoning

ResearchDGX agent

arXiv:2506.18831v2 Announce Type: replace Abstract: Large Language Models employing extended chain-of-thought (CoT) reasoning often suffer from the overthinking phenomenon, generating excessive and re


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Toward Generalized Cross-Lingual Hateful Language Detection with Web-Scale Data and Ensemble LLM Annotations

Model ReleasesDGX agent

arXiv:2604.09625v1 Announce Type: new Abstract: We study whether large-scale unlabelled web data and LLM-based synthetic annotations can improve multilingual hate speech detection. Starting from texts

Towards Efficient Large Vision-Language Models: A Comprehensive Survey on Inference Strategies

ResearchDGX agent

arXiv:2603.27960v2 Announce Type: replace-cross Abstract: Although Large Vision Language Models (LVLMs) have demonstrated impressive multimodal reasoning capabilities, their scalability and deployment

TRACE: An Experiential Framework for Coherent Multi-hop Knowledge Graph Question Answering

TutorialsDGX agent

arXiv:2604.11193v1 Announce Type: new Abstract: Multi-hop Knowledge Graph Question Answering (KGQA) requires coherent reasoning across relational paths, yet existing methods often treat each reasoning

Training-Free Cross-Lingual Dysarthria Severity Assessment via Phonological Subspace Analysis in Self-Supervised Speech Representations

SafetyDGX agent

arXiv:2604.10123v1 Announce Type: new Abstract: Dysarthric speech severity assessment typically requires trained clinicians or supervised models built from labelled pathological speech, limiting scala

Transactional Attention: Semantic Sponsorship for KV-Cache Retention

ResearchDGX agent

arXiv:2604.11288v1 Announce Type: new Abstract: At K=16 tokens (0.4% of a 4K context), every existing KV-cache compression method achieves 0% on credential retrieval. The failure mode is dormant token

Triviality Corrected Endogenous Reward

SafetyDGX agent

arXiv:2604.11522v1 Announce Type: new Abstract: Reinforcement learning for open-ended text generation is constrained by the lack of verifiable rewards, necessitating reliance on judge models that requ

Turing or Cantor: That is the Question

ResearchDGX agent

arXiv:2604.10418v1 Announce Type: new Abstract: Alan Turing is considered as a founder of current computer science together with Kurt Godel, Alonzo Church and John von Neumann. In this paper multiple

Ultra-Low-Dimensional Prompt Tuning via Random Projection

Model ReleasesDGX agent

arXiv:2502.04501v3 Announce Type: replace Abstract: Large language models achieve state-of-the-art performance but are increasingly costly to fine-tune. Prompt tuning is a parameter-efficient fine-tun

Utilizing and Calibrating Hindsight Process Rewards via Reinforcement with Mutual Information Self-Evaluation

SafetyDGX agent

arXiv:2604.11611v1 Announce Type: new Abstract: To overcome the sparse reward challenge in reinforcement learning (RL) for agents based on large language models (LLMs), we propose Mutual Information S

VeriInteresting: An Empirical Study of Model Prompt Interactions in Verilog Code Generation

Model ReleasesDGX agent

arXiv:2603.08715v2 Announce Type: replace-cross Abstract: Rapid advances in language models (LMs) have created new opportunities for automated code generation while complicating trade-offs between mod

VisText-Mosquito: A Unified Multimodal Dataset for Visual Detection, Segmentation, and Textual Explanation on Mosquito Breeding Sites

ResearchDGX agent

arXiv:2506.14629v3 Announce Type: replace-cross Abstract: Mosquito-borne diseases pose a major global health risk, requiring early detection and proactive control of breeding sites to prevent outbreak

Visual Late Chunking: An Empirical Study of Contextual Chunking for Efficient Visual Document Retrieval

ResearchDGX agent

arXiv:2604.10167v1 Announce Type: cross Abstract: Multi-vector models dominate Visual Document Retrieval (VDR) due to their fine-grained matching capabilities, but their high storage and computational

VLN-NF: Feasibility-Aware Vision-and-Language Navigation with False-Premise Instructions

Model ReleasesDGX agent

arXiv:2604.10533v1 Announce Type: cross Abstract: Conventional Vision-and-Language Navigation (VLN) benchmarks assume instructions are feasible and the referenced target exists, leaving agents ill-equ

Weird Generalization is Weirdly Brittle

SafetyDGX agent

arXiv:2604.10022v1 Announce Type: new Abstract: Weird generalization is a phenomenon in which models fine-tuned on data from a narrow domain (e.g. insecure code) develop surprising traits that manifes

What Do Vision-Language Models Encode for Personalized Image Aesthetics Assessment?

ApplicationsDGX agent

arXiv:2604.11374v1 Announce Type: cross Abstract: Personalized image aesthetics assessment (PIAA) is an important research problem with practical real-world applications. While methods based on vision

What Factors Affect LLMs and RLLMs in Financial Question Answering?

SafetyDGX agent

arXiv:2507.08339v4 Announce Type: replace Abstract: Recently, large language models (LLMs) and reasoning large language models (RLLMs) have gained considerable attention from many researchers. RLLMs e

When Meaning Isn't Literal: Exploring Idiomatic Meaning Across Languages and Modalities

Model ReleasesDGX agent

arXiv:2604.10787v1 Announce Type: new Abstract: Idiomatic reasoning, deeply intertwined with metaphor and culture, remains a blind spot for contemporary language models, whose progress skews toward su

Who Wrote This Line? Evaluating the Detection of LLM-Generated Classical Chinese Poetry

Model ReleasesDGX agent

arXiv:2604.10101v1 Announce Type: new Abstract: The rapid development of large language models (LLMs) has extended text generation tasks into the literary domain. However, AI-generated literary creati

Why Code, Why Now: An Information-Theoretic Perspective on the Limits of Machine Learning

Local AiDGX agent

arXiv:2602.13934v4 Announce Type: replace-cross Abstract: This paper offers a new perspective on the limits of machine learning: the ceiling on progress is set not by model size or algorithm choice bu

Why Don't You Know? Evaluating the Impact of Uncertainty Sources on Uncertainty Quantification in LLMs

ApplicationsDGX agent

arXiv:2604.10495v1 Announce Type: new Abstract: As Large Language Models (LLMs) are increasingly deployed in real-world applications, reliable uncertainty quantification (UQ) becomes critical for safe

Why Supervised Fine-Tuning Fails to Learn: A Systematic Study of Incomplete Learning in Large Language Models

Model ReleasesDGX agent

arXiv:2604.10079v1 Announce Type: new Abstract: Supervised Fine-Tuning (SFT) is the standard approach for adapting large language models (LLMs) to downstream tasks. However, we observe a persistent fa

YIELD: A Large-Scale Dataset and Evaluation Framework for Information Elicitation Agents

SafetyDGX agent

arXiv:2604.10968v1 Announce Type: new Abstract: Most conversational agents (CAs) are designed to satisfy user needs through user-driven interactions. However, many real-world settings, such as academi

ZARA: Training-Free Motion Time-Series Reasoning via Evidence-Grounded LLM Agents

Model ReleasesDGX agent

arXiv:2508.04038v2 Announce Type: replace Abstract: Motion sensor time-series are central to Human Activity Recognition (HAR), yet conventional approaches are constrained to fixed activity sets and ty

13 Apr 2026

A Representation-Level Assessment of Bias Mitigation in Foundation Models

SafetyDGX agent

arXiv:2604.08561v1 Announce Type: new Abstract: We investigate how successful bias mitigation reshapes the embedding space of encoder-only and decoder-only foundation models, offering an internal audi

Across the Levels of Analysis: Explaining Predictive Processing in Humans Requires More Than Machine-Estimated Probabilities

ResearchDGX agent

arXiv:2604.09466v1 Announce Type: new Abstract: Under the lens of Marr's levels of analysis, we critique and extend two claims about language models (LMs) and language processing: first, that predicti

Agentic Jackal: Live Execution and Semantic Value Grounding for Text-to-JQL

Model ReleasesDGX agent

arXiv:2604.09470v1 Announce Type: new Abstract: Translating natural language into Jira Query Language (JQL) requires resolving ambiguous field references, instance-specific categorical values, and com

Anchored Sliding Window: Toward Robust and Imperceptible Linguistic Steganography

Model ReleasesDGX agent

arXiv:2604.09066v1 Announce Type: new Abstract: Linguistic steganography based on language models typically assumes that steganographic texts are transmitted without alteration, making them fragile to

Arbitration Failure, Not Perceptual Blindness: How Vision-Language Models Resolve Visual-Linguistic Conflicts

ResearchDGX agent

arXiv:2604.09364v1 Announce Type: cross Abstract: When a Vision-Language Model (VLM) sees a blue banana and answers 'yellow', is the problem of perception or arbitration? We explore the question in te

Attention-Based Sampler for Diffusion Language Models

ResearchDGX agent

arXiv:2604.08564v1 Announce Type: new Abstract: Auto-regressive models (ARMs) have established a dominant paradigm in language modeling. However, their strictly sequential decoding paradigm imposes fu

Automated Instruction Revision (AIR): A Structured Comparison of Task Adaptation Strategies for LLM

Model ReleasesDGX agent

arXiv:2604.09418v1 Announce Type: new Abstract: This paper studies Automated Instruction Revision (AIR), a rule-induction-based method for adapting large language models (LLMs) to downstream tasks usi

BEDTime: A Unified Benchmark for Automatically Describing Time Series

Model ReleasesDGX agent

arXiv:2509.05215v3 Announce Type: replace Abstract: Recent works propose complex multi-modal models that handle both time series and language, ultimately claiming high performance on complex tasks lik

Breaking Block Boundaries: Anchor-based History-stable Decoding for Diffusion Large Language Models

Model ReleasesDGX agent

arXiv:2604.08964v1 Announce Type: new Abstract: Diffusion Large Language Models (dLLMs) have recently become a promising alternative to autoregressive large language models (ARMs). Semi-autoregressive

CodeScout: Contextual Problem Statement Enhancement for Software Agents

AgentsDGX agent

arXiv:2603.05744v2 Announce Type: replace Abstract: Current AI-powered code assistance tools often struggle with poorly-defined problem statements that lack sufficient task context and requirements sp

Confident in a Confidence Score: Investigating the Sensitivity of Confidence Scores to Supervised Fine-Tuning

ApplicationsDGX agent

arXiv:2604.08974v1 Announce Type: new Abstract: Uncertainty quantification is a set of techniques that measure confidence in language models. They can be used, for example, to detect hallucinations or

Cross-Lingual Attention Distillation with Personality-Informed Generative Augmentation for Multilingual Personality Recognition

Model ReleasesDGX agent

arXiv:2604.08851v1 Announce Type: new Abstract: While significant work has been done on personality recognition, the lack of multilingual datasets remains an unresolved challenge. To address this, we

EthicMind: A Risk-Aware Framework for Ethical-Emotional Alignment in Multi-Turn Dialogue

SafetyDGX agent

arXiv:2604.09265v1 Announce Type: new Abstract: Intelligent dialogue systems are increasingly deployed in emotionally and ethically sensitive settings, where failures in either emotional attunement or

EVOKE: Emotion Vocabulary Of Korean and English

ResearchDGX agent

arXiv:2602.10414v2 Announce Type: replace Abstract: This paper introduces EVOKE (Emotion Vocabulary of Korean and English), a Korean-English parallel dataset of emotion words. The dataset offers compr

EXAONE 4.5 Technical Report

Model ReleasesDGX agent

arXiv:2604.08644v1 Announce Type: new Abstract: This technical report introduces EXAONE 4.5, the first open-weight vision language model released by LG AI Research. EXAONE 4.5 is architected by integr

Exploiting Web Search Tools of AI Agents for Data Exfiltration

ResearchDGX agent

arXiv:2510.09093v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are now routinely used to autonomously execute complex tasks, from natural language processing to dynamic workflo

Exploring Cross-lingual Latent Transplantation: Mutual Opportunities and Open Challenges

ResearchDGX agent

arXiv:2412.12686v3 Announce Type: replace Abstract: Current large language models (LLMs) often exhibit imbalances in multilingual capabilities and cultural adaptability, largely attributed to their En

Facet-Level Tracing of Evidence Uncertainty and Hallucination in RAG

Model ReleasesDGX agent

arXiv:2604.09174v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) aims to reduce hallucination by grounding answers in retrieved evidence, yet hallucinated answers remain common eve

Fast-dVLM: Efficient Block-Diffusion VLM via Direct Conversion from Autoregressive VLM

AgentsDGX agent

arXiv:2604.06832v2 Announce Type: replace Abstract: Vision-language models (VLMs) predominantly rely on autoregressive decoding, which generates tokens one at a time and fundamentally limits inference

Few-Shot Contrastive Adaptation for Audio Abuse Detection in Low-Resource Indic Languages

ResearchDGX agent

arXiv:2604.09094v1 Announce Type: cross Abstract: Abusive speech detection is becoming increasingly important as social media shifts towards voice-based interaction, particularly in multilingual and l

FP8-RL: A Practical and Stable Low-Precision Stack for LLM Reinforcement Learning

SafetyDGX agent

arXiv:2601.18150v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) for large language models (LLMs) is increasingly bottlenecked by rollout (generation), where long output sequence

From Reasoning to Agentic: Credit Assignment in Reinforcement Learning for Large Language Models

Model ReleasesDGX agent

arXiv:2604.09459v1 Announce Type: new Abstract: Reinforcement learning (RL) for large language models (LLMs) increasingly relies on sparse, outcome-level rewards -- yet determining which actions withi

Grammar as a Behavioral Biometric: Using Cognitively Motivated Grammar Models for Authorship Verification

ResearchDGX agent

arXiv:2403.08462v3 Announce Type: replace Abstract: Authorship Verification (AV) is a key area of research in digital text forensics, which addresses the fundamental question of whether two texts were

GRASP: Grounded CoT Reasoning with Dual-Stage Optimization for Multimodal Sarcasm Target Identification

Model ReleasesDGX agent

arXiv:2604.08879v1 Announce Type: new Abstract: Moving beyond the traditional binary classification paradigm of Multimodal Sarcasm Detection, Multimodal Sarcasm Target Identification (MSTI) presents a

Growing a Multi-head Twig via Distillation and Reinforcement Learning to Accelerate Large Vision-Language Models

ResearchDGX agent

arXiv:2503.14075v3 Announce Type: replace-cross Abstract: Large vision-language models (VLMs) have demonstrated remarkable capabilities in open-world multimodal understanding, yet their high computati

Hierarchical Alignment: Enforcing Hierarchical Instruction-Following in LLMs through Logical Consistency

SafetyDGX agent

arXiv:2604.09075v1 Announce Type: new Abstract: Large language models increasingly operate under multiple instructions from heterogeneous sources with different authority levels, including system poli

Is More Data Worth the Cost? Dataset Scaling Laws in a Tiny Attention-Only Decoder

ResearchDGX agent

arXiv:2604.09389v1 Announce Type: cross Abstract: Training Transformer language models is expensive, as performance typically improves with increasing dataset size and computational budget. Although s

LADR: Locality-Aware Dynamic Rescue for Efficient Text-to-Image Generation with Diffusion Large Language Models

Local AiDGX agent

arXiv:2603.13450v2 Announce Type: replace-cross Abstract: Discrete Diffusion Language Models have emerged as a compelling paradigm for unified multimodal generation, yet their deployment is hindered b

Localizing Task Recognition and Task Learning in In-Context Learning via Attention Head Analysis

ResearchDGX agent

arXiv:2509.24164v2 Announce Type: replace Abstract: We investigate the mechanistic underpinnings of in-context learning (ICL) in large language models by reconciling two dominant perspectives: the com

MAB-DQA: Addressing Query Aspect Importance in Document Question Answering with Multi-Armed Bandits

SafetyDGX agent

arXiv:2604.08952v1 Announce Type: new Abstract: Document Question Answering (DQA) involves generating answers from a document based on a user's query, representing a key task in document understanding

Many Ways to Be Fake: Benchmarking Fake News Detection Under Strategy-Driven AI Generation

Model ReleasesDGX agent

arXiv:2604.09514v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) have enabled the large-scale generation of highly fluent and deceptive news-like content. While prior wo

MedConceal: A Benchmark for Clinical Hidden-Concern Reasoning Under Partial Observability

Model ReleasesDGX agent

arXiv:2604.08788v1 Announce Type: new Abstract: Patient-clinician communication is an asymmetric-information problem: patients often do not disclose fears, misconceptions, or practical barriers unless

Mnemis: Dual-Route Retrieval on Hierarchical Graphs for Long-Term LLM Memory

Model ReleasesDGX agent

arXiv:2602.15313v2 Announce Type: replace Abstract: AI Memory, specifically how models organizes and retrieves historical messages, becomes increasingly valuable to Large Language Models (LLMs), yet e

MSMO-ABSA: Multi-Scale and Multi-Objective Optimization for Cross-Lingual Aspect-Based Sentiment Analysis

SafetyDGX agent

arXiv:2502.13718v2 Announce Type: replace Abstract: Aspect-based sentiment analysis (ABSA) garnered growing research interest in multilingual contexts in the past. However, the majority of the studies

← Previous
1…122123124125126…128
Next →