AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
Safety

Whose Facts Win? LLM Source Preferences under Knowledge Conflicts

DGX agent

arXiv:2601.03746v3 Announce Type: replace Abstract: As large language models (LLMs) are more frequently used in retrieval-augmented generation pipelines, it is increasingly relevant to study their beh

safetyarxiv-cs-cl
20 Apr 2026
Safety
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

WildFeedback: Aligning LLMs With In-situ User Interactions And Feedback

DGX agent

arXiv:2408.15549v4 Announce Type: replace Abstract: As large language models (LLMs) continue to advance, aligning these models with human preferences has emerged as a critical challenge. Traditional a

safetyarxiv-cs-cl
20 Apr 2026
Model Releases

Wisdom is Knowing What not to Say: Hallucination-Free LLMs Unlearning via Attention Shifting

DGX agent

arXiv:2510.17210v3 Announce Type: replace Abstract: The increase in computing power and the necessity of AI-assisted decision-making boost the growing application of large language models (LLMs). Alon

model-releasesarxiv-cs-cl
20 Apr 2026
Research

A Linguistics-Aware LLM Watermarking via Syntactic Predictability

DGX agent

arXiv:2510.13829v3 Announce Type: replace Abstract: As large language models (LLMs) continue to advance rapidly, reliable governance tools have become critical. Publicly verifiable watermarking is par

researcharxiv-cs-cl
17 Apr 2026
Model Releases

AccelOpt: A Self-Improving LLM Agentic System for AI Accelerator Kernel Optimization

DGX agent

arXiv:2511.15915v2 Announce Type: replace-cross Abstract: We present AccelOpt, a self-improving large language model (LLM) agentic system that autonomously optimizes kernels for emerging AI acclerator

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

Acceptance Dynamics Across Cognitive Domains in Speculative Decoding

DGX agent

arXiv:2604.14682v1 Announce Type: cross Abstract: Speculative decoding accelerates large language model (LLM) inference. It uses a small draft model to propose a tree of future tokens. A larger target

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

ADAPT: Benchmarking Commonsense Planning under Unspecified Affordance Constraints

DGX agent

arXiv:2604.14902v1 Announce Type: cross Abstract: Intelligent embodied agents should not simply follow instructions, as real-world environments often involve unexpected conditions and exceptions. Howe

model-releasesarxiv-cs-cl
17 Apr 2026
Research

Adaptive Layer Selection for Layer-Wise Token Pruning in LLM Inference

DGX agent

arXiv:2601.07667v2 Announce Type: replace Abstract: Due to the prevalence of large language models (LLMs), key-value (KV) cache reduction for LLM inference has received remarkable attention. Among num

researcharxiv-cs-cl
17 Apr 2026
Hardware

AdaSplash-2: Faster Differentiable Sparse Attention

DGX agent

arXiv:2604.15180v1 Announce Type: cross Abstract: Sparse attention has been proposed as a way to alleviate the quadratic cost of transformers, a central bottleneck in long-context training. A promisin

hardwarearxiv-cs-cl
17 Apr 2026
Research

AIM: Asymmetric Information Masking for Visual Question Answering Continual Learning

DGX agent

arXiv:2604.14779v1 Announce Type: cross Abstract: In continual visual question answering (VQA), existing Continual Learning (CL) methods are mostly built for symmetric, unimodal architectures. However

researcharxiv-cs-cl
17 Apr 2026
Applications

An Underexplored Frontier: Large Language Models for Rare Disease Patient Education and Communication -- A scoping review

DGX agent

arXiv:2604.14179v1 Announce Type: new Abstract: Rare diseases affect over 300 million people worldwide and are characterized by complex care pathways, limited clinical expertise, and substantial unmet

applicationsarxiv-cs-cl
17 Apr 2026
Model Releases

Anonpsy: A Graph-Based Framework for Structure-Preserving De-identification of Psychiatric Narratives

DGX agent

arXiv:2601.13503v2 Announce Type: replace Abstract: Psychiatric narratives encode patient identity not only through explicit identifiers but also through idiosyncratic life events embedded in their cl

model-releasesarxiv-cs-cl
17 Apr 2026
Agents

APEX-MEM: Agentic Semi-Structured Memory with Temporal Reasoning for Long-Term Conversational AI

DGX agent

arXiv:2604.14362v1 Announce Type: new Abstract: Large language models still struggle with reliable long-term conversational memory: simply enlarging context windows or applying naive retrieval often i

agentsarxiv-cs-cl
17 Apr 2026
Tutorials

Attention to Mamba: A Recipe for Cross-Architecture Distillation

DGX agent

arXiv:2604.14191v1 Announce Type: new Abstract: State Space Models (SSMs) such as Mamba have become a popular alternative to Transformer models, due to their reduced memory consumption and higher thro

tutorialsarxiv-cs-cl
17 Apr 2026
Research

Attribution, Citation, and Quotation: A Survey of Evidence-based Text Generation with Large Language Models

DGX agent

arXiv:2508.15396v2 Announce Type: replace Abstract: The increasing adoption of large language models (LLMs) has raised serious concerns about their reliability and trustworthiness. As a result, a grow

researcharxiv-cs-cl
17 Apr 2026
Model Releases

Benchmarking Linguistic Adaptation in Comparable-Sized LLMs: A Study of Llama-3.1-8B, Mistral-7B-v0.1, and Qwen3-8B on Romanized Nepali

DGX agent

arXiv:2604.14171v1 Announce Type: new Abstract: Romanized Nepali, the Nepali language written in the Latin alphabet, is the dominant medium for informal digital communication in Nepal, yet it remains

model-releasesarxiv-cs-cl
17 Apr 2026
Agents

Beyond Literal Mapping: Benchmarking and Improving Non-Literal Translation Evaluation

DGX agent

arXiv:2601.07338v2 Announce Type: replace Abstract: Large Language Models (LLMs) have significantly advanced Machine Translation (MT), applying them to linguistically complex domains-such as Social Ne

agentsarxiv-cs-cl
17 Apr 2026
Research

Beyond Translation: Evaluating Mathematical Reasoning Capabilities of LLMs in Sinhala and Tamil

DGX agent

arXiv:2602.14517v2 Announce Type: replace Abstract: Large language models (LLMs) have achieved strong results in mathematical reasoning, and are increasingly deployed as tutoring and learning support

researcharxiv-cs-cl
17 Apr 2026
Model Releases

BiCon-Gate: Consistency-Gated De-colloquialisation for Dialogue Fact-Checking

DGX agent

arXiv:2604.14389v1 Announce Type: new Abstract: Automated fact-checking in dialogue involves multi-turn conversations where colloquial language is frequent yet understudied. To address this gap, we pr

model-releasesarxiv-cs-cl
17 Apr 2026
Safety

Blinded Multi-Rater Comparative Evaluation of a Large Language Model and Clinician-Authored Responses in CGM-Informed Diabetes Counseling

DGX agent

arXiv:2604.15124v1 Announce Type: new Abstract: Continuous glucose monitoring (CGM) is central to diabetes care, but explaining CGM patterns clearly and empathetically remains time-intensive. Evidence

safetyarxiv-cs-cl
17 Apr 2026
Safety

BoundRL: Efficient Structured Text Segmentation through Reinforced Boundary Generation

DGX agent

arXiv:2510.20151v2 Announce Type: replace Abstract: Structured texts refer to texts containing structured elements beyond plain texts, such as code snippets and placeholders. Such structured texts inc

safetyarxiv-cs-cl
17 Apr 2026
Agents

CAMO: An Agentic Framework for Automated Causal Discovery from Micro Behaviors to Macro Emergence in LLM Agent Simulations

DGX agent

arXiv:2604.14691v1 Announce Type: cross Abstract: LLM-empowered agent simulations are increasingly used to study social emergence, yet the micro-to-macro causal mechanisms behind macro outcomes often

agentsarxiv-cs-cl
17 Apr 2026
Applications

Can Large Language Models Detect Methodological Flaws? Evidence from Gesture Recognition for UAV-Based Rescue Operation Based on Deep Learning

DGX agent

arXiv:2604.14161v1 Announce Type: new Abstract: Reliable evaluation is essential in machine learning research, yet methodological flaws-particularly data leakage-continue to undermine the validity of

applicationsarxiv-cs-cl
17 Apr 2026
Model Releases

CausalDetox: Causal Head Selection and Intervention for Language Model Detoxification

DGX agent

arXiv:2604.14602v1 Announce Type: new Abstract: Large language models (LLMs) frequently generate toxic content, posing significant risks for safe deployment. Current mitigation strategies often degrad

model-releasesarxiv-cs-cl
17 Apr 2026
Tutorials

CausalEmbed: Auto-Regressive Multi-Vector Generation in Latent Space for Visual Document Embedding

DGX agent

arXiv:2601.21262v3 Announce Type: replace Abstract: Although Multimodal Large Language Models (MLLMs) have shown remarkable potential in Visual Document Retrieval (VDR) through generating high-quality

tutorialsarxiv-cs-cl
17 Apr 2026
Research

Challenges in Translating Technical Lectures: Insights from the NPTEL

DGX agent

arXiv:2602.08698v2 Announce Type: replace Abstract: This study examines the practical applications and methodological implications of Machine Translation in Indian Languages, specifically Bangla, Mala

researcharxiv-cs-cl
17 Apr 2026
Applications

Chinese Essay Rhetoric Recognition Using LoRA, In-context Learning and Model Ensemble

DGX agent

arXiv:2604.14167v1 Announce Type: new Abstract: Rhetoric recognition is a critical component in automated essay scoring. By identifying rhetorical elements in student writing, AI systems can better as

applicationsarxiv-cs-cl
17 Apr 2026
Model Releases

Chinese Language Is Not More Efficient Than English in Vibe Coding: A Preliminary Study on Token Cost and Problem-Solving Rate

DGX agent

arXiv:2604.14210v1 Announce Type: new Abstract: A claim has been circulating on social media and practitioner forums that Chinese prompts are more token-efficient than English for LLM coding tasks, po

model-releasesarxiv-cs-cl
17 Apr 2026
Research

Chronological Knowledge Retrieval: A Retrieval-Augmented Generation Approach to Construction Project Documentation

DGX agent

arXiv:2604.14169v1 Announce Type: new Abstract: In large-scale construction projects, the continuous evolution of decisions generates extensive records, most often captured in meeting minutes. Since d

researcharxiv-cs-cl
17 Apr 2026
Safety

ClimateCause: Complex and Implicit Causal Structures in Climate Reports

DGX agent

arXiv:2604.14856v1 Announce Type: new Abstract: Understanding climate change requires reasoning over complex causal networks. Yet, existing causal discovery datasets predominantly capture explicit, di

safetyarxiv-cs-cl
17 Apr 2026
Model Releases

CobwebTM: Probabilistic Concept Formation for Lifelong and Hierarchical Topic Modeling

DGX agent

arXiv:2604.14489v1 Announce Type: new Abstract: Topic modeling seeks to uncover latent semantic structure in text corpora with minimal supervision. Neural approaches achieve strong performance but req

model-releasesarxiv-cs-cl
17 Apr 2026
Research

Cognitive Alpha Mining via LLM-Driven Code-Based Evolution

DGX agent

arXiv:2511.18850v2 Announce Type: replace Abstract: Discovering effective predictive signals, or 'alphas,' from financial data with high dimensionality and extremely low signal-to-noise ratio remains

researcharxiv-cs-cl
17 Apr 2026
Model Releases

Comparison of Modern Multilingual Text Embedding Techniques for Hate Speech Detection Task

DGX agent

arXiv:2604.14907v1 Announce Type: new Abstract: Online hate speech and abusive language pose a growing challenge for content moderation, especially in multilingual settings and for low-resource langua

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

Compressed-Sensing-Guided, Inference-Aware Structured Reduction for Large Language Models

DGX agent

arXiv:2604.14156v1 Announce Type: new Abstract: Large language models deliver strong generative performance but at the cost of massive parameter counts, memory use, and decoding latency. Prior work ha

model-releasesarxiv-cs-cl
17 Apr 2026
Research

Compressing Sequences in the Latent Embedding Space: K-Token Merging for Large Language Models

DGX agent

arXiv:2604.15153v1 Announce Type: new Abstract: Large Language Models (LLMs) incur significant computational and memory costs when processing long prompts, as full self-attention scales quadratically

researcharxiv-cs-cl
17 Apr 2026
Safety

ConfLayers: Adaptive Confidence-based Layer Skipping for Self-Speculative Decoding

DGX agent

arXiv:2604.14612v1 Announce Type: cross Abstract: Self-speculative decoding is an inference technique for large language models designed to speed up generation without sacrificing output quality. It c

safetyarxiv-cs-cl
17 Apr 2026
Safety

Context Over Content: Exposing Evaluation Faking in Automated Judges

DGX agent

arXiv:2604.15224v1 Announce Type: cross Abstract: The extit{LLM-as-a-judge} paradigm has become the operational backbone of automated AI evaluation pipelines, yet rests on an unverified assumption: th

safetyarxiv-cs-cl
17 Apr 2026
Model Releases

Controlling Authority Retrieval: A Missing Retrieval Objective for Authority-Governed Knowledge

DGX agent

arXiv:2604.14488v1 Announce Type: cross Abstract: In any domain where knowledge accumulates under formal authority -- law, drug regulation, software security -- a later document can formally void an e

model-releasesarxiv-cs-cl
17 Apr 2026
Safety

CoopEval: Benchmarking Cooperation-Sustaining Mechanisms and LLM Agents in Social Dilemmas

DGX agent

arXiv:2604.15267v1 Announce Type: cross Abstract: It is increasingly important that LLM agents interact effectively and safely with other goal-pursuing agents, yet, recent works report the opposite tr

safetyarxiv-cs-cl
17 Apr 2026
Model Releases

CoPA: Benchmarking Personalized Question Answering with Data-Informed Cognitive Factors

DGX agent

arXiv:2604.14773v1 Announce Type: new Abstract: While LLMs have demonstrated remarkable potential in Question Answering (QA), evaluating personalization remains a critical bottleneck. Existing paradig

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

Correcting Suppressed Log-Probabilities in Language Models with Post-Transformer Adapters

DGX agent

arXiv:2604.14174v1 Announce Type: new Abstract: Alignment-tuned language models frequently suppress factual log-probabilities on politically sensitive topics despite retaining the knowledge in their h

model-releasesarxiv-cs-cl
17 Apr 2026
Research

Cosine-Similarity Routing with Semantic Anchors for Interpretable Mixture-of-Experts Language Models

DGX agent

arXiv:2509.14255v2 Announce Type: replace Abstract: Mixture-of-Experts (MoE) models improve efficiency through sparse activation, but their learned gating functions provide limited insight into routin

researcharxiv-cs-cl
17 Apr 2026
Research

Counting Without Numbers and Finding Without Words

DGX agent

arXiv:2603.24470v2 Announce Type: replace-cross Abstract: Every year, 10 million pets enter shelters, separated from their families. Despite desperate searches by both guardians and lost animals, 70%

researcharxiv-cs-cl
17 Apr 2026
Agents

CROP: Token-Efficient Reasoning in Large Language Models via Regularized Prompt Optimization

DGX agent

arXiv:2604.14214v1 Announce Type: new Abstract: Large Language Models utilizing reasoning techniques improve task performance but incur significant latency and token costs due to verbose generation. E

agentsarxiv-cs-cl
17 Apr 2026
Safety

CURA: Clinical Uncertainty Risk Alignment for Language Model-Based Risk Prediction

DGX agent

arXiv:2604.14651v1 Announce Type: new Abstract: Clinical language models (LMs) are increasingly applied to support clinical risk prediction from free-text notes, yet their uncertainty estimates often

safetyarxiv-cs-cl
17 Apr 2026
Research

CURaTE: Continual Unlearning in Real Time with Ensured Preservation of LLM Knowledge

DGX agent

arXiv:2604.14644v1 Announce Type: new Abstract: The inability to filter out in advance all potentially problematic data from the pre-training of large language models has given rise to the need for me

researcharxiv-cs-cl
17 Apr 2026
Hardware

DA-Cramming: Enhancing Cost-Effective Language Model Pretraining with Dependency Agreement Integration

DGX agent

arXiv:2311.04799v2 Announce Type: replace Abstract: Pretraining language models is still a challenge for many researchers due to its substantial computational costs. As such, there is growing interest

hardwarearxiv-cs-cl
17 Apr 2026
Research

Dark & Stormy: Modeling Humor in Sentences from the Bulwer-Lytton Fiction Contest

DGX agent

arXiv:2510.24538v2 Announce Type: replace Abstract: Textual humor is enormously diverse and computational studies need to account for this range, including intentionally bad humor. In this paper, we c

researcharxiv-cs-cl
17 Apr 2026
← Previous
1…142143144145146…161
Next →