AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,646 results
14 Apr 2026

Defending against Backdoor Attacks via Module Switching

ResearchDGX agent

arXiv:2504.05902v2 Announce Type: replace-cross Abstract: Backdoor attacks pose a serious threat to deep neural networks (DNNs), allowing adversaries to implant triggers for hidden behaviors in infere

Dialectic-Med: Mitigating Diagnostic Hallucinations via Counterfactual Adversarial Multi-Agent Debate

SafetyDGX agent

arXiv:2604.11258v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) in healthcare suffer from severe confirmation bias, often hallucinating visual details to support initial, pote

Different types of syntactic agreement recruit the same units within large language models

Local AiDGX agent

arXiv:2512.03676v2 Announce Type: replace Abstract: Large language models (LLMs) can reliably distinguish grammatical from ungrammatical sentences, but how grammatical knowledge is represented within


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Discourse Coherence and Response-Guided Context Rewriting for Multi-Party Dialogue Generation

ResearchDGX agent

arXiv:2604.06784v2 Announce Type: replace Abstract: Previous research on multi-party dialogue generation has predominantly leveraged structural information inherent in dialogues to directly inform the

Doc-PP: Document Policy Preservation Benchmark for Large Vision-Language Models

Model ReleasesDGX agent

arXiv:2601.03926v2 Announce Type: replace Abstract: The deployment of Large Vision-Language Models (LVLMs) for real-world document question answering is often constrained by dynamic, user-defined poli

DuET: Dual Execution for Test Output Prediction with Generated Code and Pseudocode

ResearchDGX agent

arXiv:2604.11514v1 Announce Type: cross Abstract: This work addresses test output prediction, a key challenge in test case generation. To improve the reliability of predicted outputs by LLMs, prior ap

Dynamic Adaptive Attention and Supervised Contrastive Learning: A Novel Hybrid Framework for Text Sentiment Classification

ResearchDGX agent

arXiv:2604.10459v1 Announce Type: new Abstract: The exponential growth of user-generated movie reviews on digital platforms has made accurate text sentiment classification a cornerstone task in natura

EEPO: Exploration-Enhanced Policy Optimization via Sample-Then-Forget

SafetyDGX agent

arXiv:2510.05837v2 Announce Type: replace Abstract: Balancing exploration and exploitation remains a central challenge in reinforcement learning with verifiable rewards (RLVR) for large language model

Eliciting Medical Reasoning with Knowledge-enhanced Data Synthesis: A Semi-Supervised Reinforcement Learning Approach

Model ReleasesDGX agent

arXiv:2604.11547v1 Announce Type: cross Abstract: While large language models hold promise for complex medical applications, their development is hindered by the scarcity of high-quality reasoning dat

End-to-end Contrastive Language-Speech Pretraining Model For Long-form Spoken Question Answering

SafetyDGX agent

arXiv:2511.09282v3 Announce Type: replace-cross Abstract: Significant progress has been made in spoken question answering (SQA) in recent years. However, many existing methods, including large audio l

Enhancing Multilingual RAG Systems with Debiased Language Preference-Guided Query Fusion

Local AiDGX agent

arXiv:2601.02956v3 Announce Type: replace Abstract: Multilingual Retrieval-Augmented Generation (mRAG) systems often exhibit a perceived preference for high-resource languages, particularly English, r

Evaluating Memory Capability in Continuous Lifelog Scenario

Model ReleasesDGX agent

arXiv:2604.11182v1 Announce Type: new Abstract: Nowadays, wearable devices can continuously lifelog ambient conversations, creating substantial opportunities for memory systems. However, existing benc

Evaluating Small Open LLMs for Medical Question Answering: A Practical Framework

Model ReleasesDGX agent

arXiv:2604.10535v1 Announce Type: cross Abstract: Incorporating large language models (LLMs) in medical question answering demands more than high average accuracy: a model that returns substantively d

EviCare: Enhancing Diagnosis Prediction with Deep Model-Guided Evidence for In-Context Reasoning

TutorialsDGX agent

arXiv:2604.10455v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) have enabled promising progress in diagnosis prediction from electronic health records (EHRs). However,

EvoDiagram: Agentic Editable Diagram Creation via Design Expertise Evolution

Model ReleasesDGX agent

arXiv:2604.09568v1 Announce Type: cross Abstract: High-fidelity diagram creation requires the complex orchestration of semantic topology, visual styling, and spatial layout, posing a significant chall

Expect the Unexpected? Testing the Surprisal of Salient Entities

ResearchDGX agent

arXiv:2604.10724v1 Announce Type: new Abstract: Previous work examining the Uniform Information Density (UID) hypothesis has shown that while information as measured by surprisal metrics is distribute

Exploring Knowledge Purification in Multi-Teacher Knowledge Distillation for LLMs

ResearchDGX agent

arXiv:2602.01064v2 Announce Type: replace Abstract: Knowledge distillation has emerged as a pivotal technique for transferring knowledge from stronger large language models (LLMs) to smaller, more eff

FAITH: Factuality Alignment through Integrating Trustworthiness and Honestness

SafetyDGX agent

arXiv:2604.10189v1 Announce Type: new Abstract: Large Language Models (LLMs) can generate factually inaccurate content even if they have corresponding knowledge, which critically undermines their reli

Find Your Optimal Teacher: Personalized Data Synthesis via Router-Guided Multi-Teacher Distillation

TutorialsDGX agent

arXiv:2510.10925v2 Announce Type: replace-cross Abstract: Training student models on synthetic data generated by strong teacher models is a promising way to distilling the capabilities of teachers. Ho

FlashMem: Distilling Intrinsic Latent Memory via Computation Reuse

Model ReleasesDGX agent

arXiv:2601.05505v2 Announce Type: replace Abstract: The stateless architecture of Large Language Models inherently lacks the mechanism to preserve dynamic context, compelling agents to redundantly rep

From Perception to Autonomous Computational Modeling: A Multi-Agent Approach

Model ReleasesDGX agent

arXiv:2604.06788v2 Announce Type: replace-cross Abstract: We present a solver-agnostic framework in which coordinated large language model (LLM) agents autonomously execute the complete computational

From Speech-to-Spatial: Grounding Utterances on A Live Shared View with Augmented Reality

ResearchDGX agent

arXiv:2602.03059v2 Announce Type: replace-cross Abstract: We introduce Speech-to-Spatial, a referent disambiguation framework that converts verbal remote-assistance instructions into spatially grounde

Generation-Augmented Generation: A Plug-and-Play Framework for Private Knowledge Injection in Large Language Models

Model ReleasesDGX agent

arXiv:2601.08209v3 Announce Type: replace Abstract: In domains such as materials science, biomedicine, and finance, high-stakes deployment of large language models (LLMs) requires injecting private, d

GenProve: Learning to Generate Text with Fine-Grained Provenance

SafetyDGX agent

arXiv:2601.04932v2 Announce Type: replace Abstract: Large language models (LLM) often hallucinate, and while adding citations is a common solution, it is frequently insufficient for accountability as

Geometry-Aware Localized Watermarking for Copyright Protection in Embedding-as-a-Service

Model ReleasesDGX agent

arXiv:2604.11344v1 Announce Type: cross Abstract: Embedding-as-a-Service (EaaS) has become an important semantic infrastructure for natural language and multimedia applications, but it is highly vulne

Head-wise Modality Specialization within MLLMs for Robust Fake News Detection under Missing Modality

ApplicationsDGX agent

arXiv:2604.09711v1 Announce Type: cross Abstract: Multimodal fake news detection (MFND) aims to verify news credibility by jointly exploiting textual and visual evidence. However, real-world news diss

HeceTokenizer: A Syllable-Based Tokenization Approach for Turkish Retrieval

Model ReleasesDGX agent

arXiv:2604.10665v1 Announce Type: new Abstract: HeceTokenizer is a syllable-based tokenizer for Turkish that exploits the deterministic six-pattern phonological structure of the language to construct

Hidden Failures in Robustness: Why Supervised Uncertainty Quantification Needs Better Evaluation

ResearchDGX agent

arXiv:2604.11662v1 Announce Type: new Abstract: Recent work has shown that the hidden states of large language models contain signals useful for uncertainty estimation and hallucination detection, mot

HiEdit: Lifelong Model Editing with Hierarchical Reinforcement Learning

Model ReleasesDGX agent

arXiv:2604.11214v1 Announce Type: new Abstract: Lifelong model editing (LME) aims to sequentially rectify outdated or inaccurate knowledge in deployed LLMs while minimizing side effects on unrelated i

Hierarchical Textual Knowledge for Enhanced Image Clustering

TutorialsDGX agent

arXiv:2604.11144v1 Announce Type: cross Abstract: Image clustering aims to group images in an unsupervised fashion. Traditional methods focus on knowledge from visual space, making it difficult to dis

Hijacking Text Heritage: Hiding the Human Signature through Homoglyphic Substitution

ResearchDGX agent

arXiv:2604.10271v1 Announce Type: cross Abstract: In what way could a data breach involving government-issued IDs such as passports, driver's licenses, etc., rival a random voluntary disclosure on a n

HistLens: Mapping Idea Change across Concepts and Corpora

ResearchDGX agent

arXiv:2604.11749v1 Announce Type: new Abstract: Language change both reflects and shapes social processes, and the semantic evolution of foundational concepts provides a measurable trace of historical

How Robust Are Large Language Models for Clinical Numeracy? An Empirical Study on Numerical Reasoning Abilities in Clinical Contexts

Model ReleasesDGX agent

arXiv:2604.11133v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly being explored for clinical question answering and decision support, yet safe deployment critically requir

How You Ask Matters! Adaptive RAG Robustness to Query Variations

Model ReleasesDGX agent

arXiv:2604.10745v1 Announce Type: new Abstract: Adaptive Retrieval-Augmented Generation (RAG) promises accuracy and efficiency by dynamically triggering retrieval only when needed and is widely used i

HTAA: Enhancing LLM Planning via Hybrid Toolset Agentization & Adaptation

AgentsDGX agent

arXiv:2604.10917v1 Announce Type: new Abstract: Enabling large language models to scale and reliably use hundreds of tools is critical for real-world applications, yet challenging due to the inefficie

Human vs. Machine Deception: Distinguishing AI-Generated and Human-Written Fake News Using Ensemble Learning

ResearchDGX agent

arXiv:2604.09960v1 Announce Type: new Abstract: The rapid adoption of large language models has introduced a new class of AI-generated fake news that coexists with traditional human-written misinforma

HumorGen: Cognitive Synergy for Humor Generation in Large Language Models via Persona-Based Distillation

Model ReleasesDGX agent

arXiv:2604.09629v1 Announce Type: new Abstract: Humor generation poses a significant challenge for Large Language Models (LLMs), because their standard training objective - predicting the most likely

HyperGraphPro: Progress-Aware Reinforcement Learning for Structure-Guided Hypergraph RAG

SafetyDGX agent

arXiv:2601.17755v2 Announce Type: replace Abstract: Graph Retrieval-Augmented Generation (GraphRAG) has emerged as a promising paradigm that organizes external knowledge into structured graphs of enti

Improving LLM Unlearning Robustness via Random Perturbations

ResearchDGX agent

arXiv:2501.19202v5 Announce Type: replace Abstract: Here, we show that current LLM unlearning methods inherently reduce models' robustness, causing them to misbehave even when a single non-adversarial

Infusing Theory of Mind into Socially Intelligent LLM Agents

Model ReleasesDGX agent

arXiv:2509.22887v2 Announce Type: replace Abstract: Theory of Mind (ToM)-an understanding of the mental states of others-is a key aspect of human social intelligence, yet, chatbots and LLM-based socia

Instruction Data Selection via Answer Divergence

ResearchDGX agent

arXiv:2604.10448v1 Announce Type: new Abstract: Instruction tuning relies on large instruction-response corpora whose quality and composition strongly affect downstream performance. We propose Answer

Judge Like Human Examiners: A Weighted Importance Multi-Point Evaluation Framework for Generative Tasks with Long-form Answers

SafetyDGX agent

arXiv:2604.11246v1 Announce Type: new Abstract: Evaluating the quality of model responses remains challenging in generative tasks with long-form answers, as the expected answers usually contain multip

K-Way Energy Probes for Metacognition Reduce to Softmax in Discriminative Predictive Coding Networks

Model ReleasesDGX agent

arXiv:2604.11011v1 Announce Type: cross Abstract: We present this as a negative result with an explanatory mechanism, not as a formal upper bound. Predictive coding networks (PCNs) admit a K-way energ

KCS: Diversify Multi-hop Question Generation with Knowledge Composition Sampling

ResearchDGX agent

arXiv:2508.20567v2 Announce Type: replace Abstract: Multi-hop question answering faces substantial challenges due to data sparsity, which increases the likelihood of language models learning spurious

Knowing What to Stress: A Discourse-Conditioned Text-to-Speech Benchmark

Model ReleasesDGX agent

arXiv:2604.10580v1 Announce Type: new Abstract: Spoken meaning often depends not only on what is said, but also on which word is emphasized. The same sentence can convey correction, contrast, or clari

ks-pret-5m: a 5 million word, 12 million token kashmiri pretraining dataset

Model ReleasesDGX agent

arXiv:2604.11066v1 Announce Type: new Abstract: We present KS-PRET-5M, the largest publicly available pretraining dataset for the Kashmiri language, comprising 5,090,244 (5.09M) words, 27,692,959 (27.

LaMI: Augmenting Large Language Models via Late Multi-Image Fusion

Model ReleasesDGX agent

arXiv:2406.13621v2 Announce Type: replace Abstract: Commonsense reasoning often requires both textual and visual knowledge, yet Large Language Models (LLMs) trained solely on text lack visual groundin

LangFlow: Continuous Diffusion Rivals Discrete in Language Modeling

ResearchDGX agent

arXiv:2604.11748v1 Announce Type: new Abstract: Continuous diffusion models have achieved strong performance across domains such as images. However, in language modeling, prior continuous diffusion la

LASQ: A Low-resource Aspect-based Sentiment Quadruple Extraction Dataset

ResearchDGX agent

arXiv:2604.10417v1 Announce Type: new Abstract: In recent years, aspect-based sentiment analysis (ABSA) has made rapid progress and shown strong practical value. However, existing research and benchma

LayerNorm Induces Recency Bias in Transformer Decoders

SafetyDGX agent

arXiv:2509.21042v3 Announce Type: replace Abstract: Causal self-attention provides positional information to Transformer decoders. Prior work has shown that stacks of causal self-attention layers alon

LIFT: A Novel Framework for Enhancing Long-Context Understanding of LLMs via Long Input Fine-Tuning

Model ReleasesDGX agent

arXiv:2502.14644v5 Announce Type: replace Abstract: Long context understanding remains challenging for large language models due to their limited context windows. This paper introduces Long Input Fine

LingoLoop Attack: Trapping MLLMs via Linguistic Context and State Entrapment into Endless Loops

ResearchDGX agent

arXiv:2506.14493v3 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have shown great promise but require substantial computational resources during inference. Attackers can ex

Linguistic Accommodation Between Neurodivergent Communities on Reddit:A Communication Accommodation Theory Analysis of ADHD and Autism Groups

ResearchDGX agent

arXiv:2604.10063v1 Announce Type: new Abstract: Social media research on mental health has focused predominantly on detecting and diagnosing conditions at the individual level. In this work, we shift

LLMs for Game Theory: Entropy-Guided In-Context Learning and Adaptive CoT Reasoning

ResearchDGX agent

arXiv:2601.10775v2 Announce Type: replace Abstract: We propose a novel LLM-based framework for reasoning in discrete, game-theoretic tasks, illustrated with Tic-Tac-Toe. The method integrates in-conte

Look Twice before You Leap: A Rational Framework for Localized Adversarial Anonymization

Local AiDGX agent

arXiv:2512.06713v3 Announce Type: replace-cross Abstract: Current LLM-based frameworks for text anonymization usually rely on remote API services from powerful LLMs, which creates an inherent privacy

Lost in Diffusion: Uncovering Hallucination Patterns and Failure Modes in Diffusion Large Language Models

ResearchDGX agent

arXiv:2604.10556v1 Announce Type: new Abstract: While Diffusion Large Language Models (dLLMs) have emerged as a promising non-autoregressive paradigm comparable to autoregressive (AR) models, their fa

M2-Verify: A Large-Scale Multidomain Benchmark for Checking Multimodal Claim Consistency

Model ReleasesDGX agent

arXiv:2604.01306v2 Announce Type: replace Abstract: Evaluating scientific arguments requires assessing the strict consistency between a claim and its underlying multimodal evidence. However, existing

MAESTRO: Meta-learning Adaptive Estimation of Scalarization Trade-offs for Reward Optimization

SafetyDGX agent

arXiv:2601.07208v2 Announce Type: replace-cross Abstract: Group-Relative Policy Optimization (GRPO) has emerged as an efficient paradigm for aligning Large Language Models (LLMs), yet its efficacy is

MASH: Modeling Abstention via Selective Help-Seeking

AgentsDGX agent

arXiv:2510.01152v2 Announce Type: replace Abstract: LLMs cannot reliably recognize their parametric knowledge boundaries and often hallucinate answers to outside-of-boundary questions. In this paper,

MCAT: Scaling Many-to-Many Speech-to-Text Translation with MLLMs to 70 Languages

Model ReleasesDGX agent

arXiv:2512.01512v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have achieved great success in Speech-to-Text Translation (S2TT) tasks. However, current research is constr

← Previous
1…120121122123124…128
Next →