AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,646 results
Research

Catalog-Native LLM: Speaking Item-ID Dialect with Less Entanglement for Recommendation

DGX agent

arXiv:2510.05125v2 Announce Type: replace Abstract: While collaborative filtering delivers predictive accuracy and efficiency, and Large Language Models (LLMs) enable expressive and generalizable reas

researcharxiv-cs-cl
14 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Challenging the Boundaries of Reasoning: An Olympiad-Level Math Benchmark for Large Language Models

DGX agent

arXiv:2503.21380v3 Announce Type: replace Abstract: The rapid advancement of large reasoning models has saturated existing math benchmarks, underscoring the urgent need for more challenging evaluation

model-releasesarxiv-cs-cl
14 Apr 2026
Agents

ChatInject: Abusing Chat Templates for Prompt Injection in LLM Agents

DGX agent

arXiv:2509.22830v3 Announce Type: replace Abstract: The growing deployment of large language model (LLM) based agents that interact with external environments has created new attack surfaces for adver

agentsarxiv-cs-cl
14 Apr 2026
Model Releases

ChemPro: A Progressive Chemistry Benchmark for Large Language Models

DGX agent

arXiv:2602.03108v3 Announce Type: replace Abstract: We introduce ChemPro, a progressive benchmark with 4100 natural language question-answer pairs in Chemistry, across 4 coherent sections of difficult

model-releasesarxiv-cs-cl
14 Apr 2026
Safety

Claim2Vec: Embedding Fact-Check Claims for Multilingual Similarity and Clustering

DGX agent

arXiv:2604.09812v1 Announce Type: new Abstract: Recurrent claims present a major challenge for automated fact-checking systems designed to combat misinformation, especially in multilingual settings. W

safetyarxiv-cs-cl
14 Apr 2026
Model Releases

ClaimDB: A Fact Verification Benchmark over Large Structured Data

DGX agent

arXiv:2601.14698v2 Announce Type: replace Abstract: Real-world fact-checking often involves verifying claims grounded in structured data at scale. Despite substantial progress in fact-verification ben

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

CLSGen: A Dual-Head Fine-Tuning Framework for Joint Probabilistic Classification and Verbalized Explanation

DGX agent

arXiv:2604.11801v1 Announce Type: new Abstract: With the recent progress of Large Language Models (LLMs), there is a growing interest in applying these models to solve complex and challenging problems

model-releasesarxiv-cs-cl
14 Apr 2026
Local Ai

CodeComp: Structural KV Cache Compression for Agentic Coding

DGX agent

arXiv:2604.10235v1 Announce Type: new Abstract: Agentic code tasks such as fault localization and patch generation require processing long codebases under tight memory constraints, where the Key-Value

local-aiarxiv-cs-cl
14 Apr 2026
Model Releases

Comparative Analysis of Large Language Models in Healthcare

DGX agent

arXiv:2604.10316v1 Announce Type: new Abstract: Background: Large Language Models (LLMs) are transforming artificial intelligence applications in healthcare due to their ability to understand, generat

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

CounterBench: Evaluating and Improving Counterfactual Reasoning in Large Language Models

DGX agent

arXiv:2502.11008v2 Announce Type: replace Abstract: Counterfactual reasoning is widely recognized as one of the most challenging and intricate aspects of causality in artificial intelligence. In this

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Decomposing and Reducing Hidden Measurement Error in LLM Evaluation Pipelines

DGX agent

arXiv:2604.11581v1 Announce Type: new Abstract: LLM evaluations drive which models get deployed, which safety standards get adopted, and which research conclusions get published. Yet these scores carr

model-releasesarxiv-cs-cl
14 Apr 2026
Research

DeCoVec: Building Decoding Space based Task Vector for Large Language Models via In-Context Learning

DGX agent

arXiv:2604.11129v1 Announce Type: new Abstract: Task vectors, representing directions in model or activation spaces that encode task-specific behaviors, have emerged as a promising tool for steering l

researcharxiv-cs-cl
14 Apr 2026
Research

Defending against Backdoor Attacks via Module Switching

DGX agent

arXiv:2504.05902v2 Announce Type: replace-cross Abstract: Backdoor attacks pose a serious threat to deep neural networks (DNNs), allowing adversaries to implant triggers for hidden behaviors in infere

researcharxiv-cs-cl
14 Apr 2026
Safety

Dialectic-Med: Mitigating Diagnostic Hallucinations via Counterfactual Adversarial Multi-Agent Debate

DGX agent

arXiv:2604.11258v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) in healthcare suffer from severe confirmation bias, often hallucinating visual details to support initial, pote

safetyarxiv-cs-cl
14 Apr 2026
Local Ai

Different types of syntactic agreement recruit the same units within large language models

DGX agent

arXiv:2512.03676v2 Announce Type: replace Abstract: Large language models (LLMs) can reliably distinguish grammatical from ungrammatical sentences, but how grammatical knowledge is represented within

local-aiarxiv-cs-cl
14 Apr 2026
Research

Discourse Coherence and Response-Guided Context Rewriting for Multi-Party Dialogue Generation

DGX agent

arXiv:2604.06784v2 Announce Type: replace Abstract: Previous research on multi-party dialogue generation has predominantly leveraged structural information inherent in dialogues to directly inform the

researcharxiv-cs-cl
14 Apr 2026
Model Releases

Doc-PP: Document Policy Preservation Benchmark for Large Vision-Language Models

DGX agent

arXiv:2601.03926v2 Announce Type: replace Abstract: The deployment of Large Vision-Language Models (LVLMs) for real-world document question answering is often constrained by dynamic, user-defined poli

model-releasesarxiv-cs-cl
14 Apr 2026
Research

DuET: Dual Execution for Test Output Prediction with Generated Code and Pseudocode

DGX agent

arXiv:2604.11514v1 Announce Type: cross Abstract: This work addresses test output prediction, a key challenge in test case generation. To improve the reliability of predicted outputs by LLMs, prior ap

researcharxiv-cs-cl
14 Apr 2026
Research

Dynamic Adaptive Attention and Supervised Contrastive Learning: A Novel Hybrid Framework for Text Sentiment Classification

DGX agent

arXiv:2604.10459v1 Announce Type: new Abstract: The exponential growth of user-generated movie reviews on digital platforms has made accurate text sentiment classification a cornerstone task in natura

researcharxiv-cs-cl
14 Apr 2026
Safety

EEPO: Exploration-Enhanced Policy Optimization via Sample-Then-Forget

DGX agent

arXiv:2510.05837v2 Announce Type: replace Abstract: Balancing exploration and exploitation remains a central challenge in reinforcement learning with verifiable rewards (RLVR) for large language model

safetyarxiv-cs-cl
14 Apr 2026
Model Releases

Eliciting Medical Reasoning with Knowledge-enhanced Data Synthesis: A Semi-Supervised Reinforcement Learning Approach

DGX agent

arXiv:2604.11547v1 Announce Type: cross Abstract: While large language models hold promise for complex medical applications, their development is hindered by the scarcity of high-quality reasoning dat

model-releasesarxiv-cs-cl
14 Apr 2026
Safety

End-to-end Contrastive Language-Speech Pretraining Model For Long-form Spoken Question Answering

DGX agent

arXiv:2511.09282v3 Announce Type: replace-cross Abstract: Significant progress has been made in spoken question answering (SQA) in recent years. However, many existing methods, including large audio l

safetyarxiv-cs-cl
14 Apr 2026
Local Ai

Enhancing Multilingual RAG Systems with Debiased Language Preference-Guided Query Fusion

DGX agent

arXiv:2601.02956v3 Announce Type: replace Abstract: Multilingual Retrieval-Augmented Generation (mRAG) systems often exhibit a perceived preference for high-resource languages, particularly English, r

local-aiarxiv-cs-cl
14 Apr 2026
Model Releases

Evaluating Memory Capability in Continuous Lifelog Scenario

DGX agent

arXiv:2604.11182v1 Announce Type: new Abstract: Nowadays, wearable devices can continuously lifelog ambient conversations, creating substantial opportunities for memory systems. However, existing benc

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Evaluating Small Open LLMs for Medical Question Answering: A Practical Framework

DGX agent

arXiv:2604.10535v1 Announce Type: cross Abstract: Incorporating large language models (LLMs) in medical question answering demands more than high average accuracy: a model that returns substantively d

model-releasesarxiv-cs-cl
14 Apr 2026
Tutorials

EviCare: Enhancing Diagnosis Prediction with Deep Model-Guided Evidence for In-Context Reasoning

DGX agent

arXiv:2604.10455v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) have enabled promising progress in diagnosis prediction from electronic health records (EHRs). However,

tutorialsarxiv-cs-cl
14 Apr 2026
Model Releases

EvoDiagram: Agentic Editable Diagram Creation via Design Expertise Evolution

DGX agent

arXiv:2604.09568v1 Announce Type: cross Abstract: High-fidelity diagram creation requires the complex orchestration of semantic topology, visual styling, and spatial layout, posing a significant chall

model-releasesarxiv-cs-cl
14 Apr 2026
Research

Expect the Unexpected? Testing the Surprisal of Salient Entities

DGX agent

arXiv:2604.10724v1 Announce Type: new Abstract: Previous work examining the Uniform Information Density (UID) hypothesis has shown that while information as measured by surprisal metrics is distribute

researcharxiv-cs-cl
14 Apr 2026
Research

Exploring Knowledge Purification in Multi-Teacher Knowledge Distillation for LLMs

DGX agent

arXiv:2602.01064v2 Announce Type: replace Abstract: Knowledge distillation has emerged as a pivotal technique for transferring knowledge from stronger large language models (LLMs) to smaller, more eff

researcharxiv-cs-cl
14 Apr 2026
Safety

FAITH: Factuality Alignment through Integrating Trustworthiness and Honestness

DGX agent

arXiv:2604.10189v1 Announce Type: new Abstract: Large Language Models (LLMs) can generate factually inaccurate content even if they have corresponding knowledge, which critically undermines their reli

safetyarxiv-cs-cl
14 Apr 2026
Tutorials

Find Your Optimal Teacher: Personalized Data Synthesis via Router-Guided Multi-Teacher Distillation

DGX agent

arXiv:2510.10925v2 Announce Type: replace-cross Abstract: Training student models on synthetic data generated by strong teacher models is a promising way to distilling the capabilities of teachers. Ho

tutorialsarxiv-cs-cl
14 Apr 2026
Model Releases

FlashMem: Distilling Intrinsic Latent Memory via Computation Reuse

DGX agent

arXiv:2601.05505v2 Announce Type: replace Abstract: The stateless architecture of Large Language Models inherently lacks the mechanism to preserve dynamic context, compelling agents to redundantly rep

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

From Perception to Autonomous Computational Modeling: A Multi-Agent Approach

DGX agent

arXiv:2604.06788v2 Announce Type: replace-cross Abstract: We present a solver-agnostic framework in which coordinated large language model (LLM) agents autonomously execute the complete computational

model-releasesarxiv-cs-cl
14 Apr 2026
Research

From Speech-to-Spatial: Grounding Utterances on A Live Shared View with Augmented Reality

DGX agent

arXiv:2602.03059v2 Announce Type: replace-cross Abstract: We introduce Speech-to-Spatial, a referent disambiguation framework that converts verbal remote-assistance instructions into spatially grounde

researcharxiv-cs-cl
14 Apr 2026
Model Releases

Generation-Augmented Generation: A Plug-and-Play Framework for Private Knowledge Injection in Large Language Models

DGX agent

arXiv:2601.08209v3 Announce Type: replace Abstract: In domains such as materials science, biomedicine, and finance, high-stakes deployment of large language models (LLMs) requires injecting private, d

model-releasesarxiv-cs-cl
14 Apr 2026
Safety

GenProve: Learning to Generate Text with Fine-Grained Provenance

DGX agent

arXiv:2601.04932v2 Announce Type: replace Abstract: Large language models (LLM) often hallucinate, and while adding citations is a common solution, it is frequently insufficient for accountability as

safetyarxiv-cs-cl
14 Apr 2026
Model Releases

Geometry-Aware Localized Watermarking for Copyright Protection in Embedding-as-a-Service

DGX agent

arXiv:2604.11344v1 Announce Type: cross Abstract: Embedding-as-a-Service (EaaS) has become an important semantic infrastructure for natural language and multimedia applications, but it is highly vulne

model-releasesarxiv-cs-cl
14 Apr 2026
Applications

Head-wise Modality Specialization within MLLMs for Robust Fake News Detection under Missing Modality

DGX agent

arXiv:2604.09711v1 Announce Type: cross Abstract: Multimodal fake news detection (MFND) aims to verify news credibility by jointly exploiting textual and visual evidence. However, real-world news diss

applicationsarxiv-cs-cl
14 Apr 2026
Model Releases

HeceTokenizer: A Syllable-Based Tokenization Approach for Turkish Retrieval

DGX agent

arXiv:2604.10665v1 Announce Type: new Abstract: HeceTokenizer is a syllable-based tokenizer for Turkish that exploits the deterministic six-pattern phonological structure of the language to construct

model-releasesarxiv-cs-cl
14 Apr 2026
Research

Hidden Failures in Robustness: Why Supervised Uncertainty Quantification Needs Better Evaluation

DGX agent

arXiv:2604.11662v1 Announce Type: new Abstract: Recent work has shown that the hidden states of large language models contain signals useful for uncertainty estimation and hallucination detection, mot

researcharxiv-cs-cl
14 Apr 2026
Model Releases

HiEdit: Lifelong Model Editing with Hierarchical Reinforcement Learning

DGX agent

arXiv:2604.11214v1 Announce Type: new Abstract: Lifelong model editing (LME) aims to sequentially rectify outdated or inaccurate knowledge in deployed LLMs while minimizing side effects on unrelated i

model-releasesarxiv-cs-cl
14 Apr 2026
Tutorials

Hierarchical Textual Knowledge for Enhanced Image Clustering

DGX agent

arXiv:2604.11144v1 Announce Type: cross Abstract: Image clustering aims to group images in an unsupervised fashion. Traditional methods focus on knowledge from visual space, making it difficult to dis

tutorialsarxiv-cs-cl
14 Apr 2026
Research

Hijacking Text Heritage: Hiding the Human Signature through Homoglyphic Substitution

DGX agent

arXiv:2604.10271v1 Announce Type: cross Abstract: In what way could a data breach involving government-issued IDs such as passports, driver's licenses, etc., rival a random voluntary disclosure on a n

researcharxiv-cs-cl
14 Apr 2026
Research

HistLens: Mapping Idea Change across Concepts and Corpora

DGX agent

arXiv:2604.11749v1 Announce Type: new Abstract: Language change both reflects and shapes social processes, and the semantic evolution of foundational concepts provides a measurable trace of historical

researcharxiv-cs-cl
14 Apr 2026
Model Releases

How Robust Are Large Language Models for Clinical Numeracy? An Empirical Study on Numerical Reasoning Abilities in Clinical Contexts

DGX agent

arXiv:2604.11133v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly being explored for clinical question answering and decision support, yet safe deployment critically requir

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

How You Ask Matters! Adaptive RAG Robustness to Query Variations

DGX agent

arXiv:2604.10745v1 Announce Type: new Abstract: Adaptive Retrieval-Augmented Generation (RAG) promises accuracy and efficiency by dynamically triggering retrieval only when needed and is widely used i

model-releasesarxiv-cs-cl
14 Apr 2026
Agents

HTAA: Enhancing LLM Planning via Hybrid Toolset Agentization & Adaptation

DGX agent

arXiv:2604.10917v1 Announce Type: new Abstract: Enabling large language models to scale and reliably use hundreds of tools is critical for real-world applications, yet challenging due to the inefficie

agentsarxiv-cs-cl
14 Apr 2026
Research

Human vs. Machine Deception: Distinguishing AI-Generated and Human-Written Fake News Using Ensemble Learning

DGX agent

arXiv:2604.09960v1 Announce Type: new Abstract: The rapid adoption of large language models has introduced a new class of AI-generated fake news that coexists with traditional human-written misinforma

researcharxiv-cs-cl
14 Apr 2026
← Previous
1…150151152153154…160
Next →