AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,646 results
Model Releases

ADAPT: Benchmarking Commonsense Planning under Unspecified Affordance Constraints

DGX agent

arXiv:2604.14902v1 Announce Type: cross Abstract: Intelligent embodied agents should not simply follow instructions, as real-world environments often involve unexpected conditions and exceptions. Howe

model-releasesarxiv-cs-cl
17 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Adaptive Layer Selection for Layer-Wise Token Pruning in LLM Inference

DGX agent

arXiv:2601.07667v2 Announce Type: replace Abstract: Due to the prevalence of large language models (LLMs), key-value (KV) cache reduction for LLM inference has received remarkable attention. Among num

researcharxiv-cs-cl
17 Apr 2026
Hardware

AdaSplash-2: Faster Differentiable Sparse Attention

DGX agent

arXiv:2604.15180v1 Announce Type: cross Abstract: Sparse attention has been proposed as a way to alleviate the quadratic cost of transformers, a central bottleneck in long-context training. A promisin

hardwarearxiv-cs-cl
17 Apr 2026
Research

AIM: Asymmetric Information Masking for Visual Question Answering Continual Learning

DGX agent

arXiv:2604.14779v1 Announce Type: cross Abstract: In continual visual question answering (VQA), existing Continual Learning (CL) methods are mostly built for symmetric, unimodal architectures. However

researcharxiv-cs-cl
17 Apr 2026
Applications

An Underexplored Frontier: Large Language Models for Rare Disease Patient Education and Communication -- A scoping review

DGX agent

arXiv:2604.14179v1 Announce Type: new Abstract: Rare diseases affect over 300 million people worldwide and are characterized by complex care pathways, limited clinical expertise, and substantial unmet

applicationsarxiv-cs-cl
17 Apr 2026
Model Releases

Anonpsy: A Graph-Based Framework for Structure-Preserving De-identification of Psychiatric Narratives

DGX agent

arXiv:2601.13503v2 Announce Type: replace Abstract: Psychiatric narratives encode patient identity not only through explicit identifiers but also through idiosyncratic life events embedded in their cl

model-releasesarxiv-cs-cl
17 Apr 2026
Agents

APEX-MEM: Agentic Semi-Structured Memory with Temporal Reasoning for Long-Term Conversational AI

DGX agent

arXiv:2604.14362v1 Announce Type: new Abstract: Large language models still struggle with reliable long-term conversational memory: simply enlarging context windows or applying naive retrieval often i

agentsarxiv-cs-cl
17 Apr 2026
Tutorials

Attention to Mamba: A Recipe for Cross-Architecture Distillation

DGX agent

arXiv:2604.14191v1 Announce Type: new Abstract: State Space Models (SSMs) such as Mamba have become a popular alternative to Transformer models, due to their reduced memory consumption and higher thro

tutorialsarxiv-cs-cl
17 Apr 2026
Research

Attribution, Citation, and Quotation: A Survey of Evidence-based Text Generation with Large Language Models

DGX agent

arXiv:2508.15396v2 Announce Type: replace Abstract: The increasing adoption of large language models (LLMs) has raised serious concerns about their reliability and trustworthiness. As a result, a grow

researcharxiv-cs-cl
17 Apr 2026
Model Releases

Benchmarking Linguistic Adaptation in Comparable-Sized LLMs: A Study of Llama-3.1-8B, Mistral-7B-v0.1, and Qwen3-8B on Romanized Nepali

DGX agent

arXiv:2604.14171v1 Announce Type: new Abstract: Romanized Nepali, the Nepali language written in the Latin alphabet, is the dominant medium for informal digital communication in Nepal, yet it remains

model-releasesarxiv-cs-cl
17 Apr 2026
Agents

Beyond Literal Mapping: Benchmarking and Improving Non-Literal Translation Evaluation

DGX agent

arXiv:2601.07338v2 Announce Type: replace Abstract: Large Language Models (LLMs) have significantly advanced Machine Translation (MT), applying them to linguistically complex domains-such as Social Ne

agentsarxiv-cs-cl
17 Apr 2026
Research

Beyond Translation: Evaluating Mathematical Reasoning Capabilities of LLMs in Sinhala and Tamil

DGX agent

arXiv:2602.14517v2 Announce Type: replace Abstract: Large language models (LLMs) have achieved strong results in mathematical reasoning, and are increasingly deployed as tutoring and learning support

researcharxiv-cs-cl
17 Apr 2026
Model Releases

BiCon-Gate: Consistency-Gated De-colloquialisation for Dialogue Fact-Checking

DGX agent

arXiv:2604.14389v1 Announce Type: new Abstract: Automated fact-checking in dialogue involves multi-turn conversations where colloquial language is frequent yet understudied. To address this gap, we pr

model-releasesarxiv-cs-cl
17 Apr 2026
Safety

Blinded Multi-Rater Comparative Evaluation of a Large Language Model and Clinician-Authored Responses in CGM-Informed Diabetes Counseling

DGX agent

arXiv:2604.15124v1 Announce Type: new Abstract: Continuous glucose monitoring (CGM) is central to diabetes care, but explaining CGM patterns clearly and empathetically remains time-intensive. Evidence

safetyarxiv-cs-cl
17 Apr 2026
Safety

BoundRL: Efficient Structured Text Segmentation through Reinforced Boundary Generation

DGX agent

arXiv:2510.20151v2 Announce Type: replace Abstract: Structured texts refer to texts containing structured elements beyond plain texts, such as code snippets and placeholders. Such structured texts inc

safetyarxiv-cs-cl
17 Apr 2026
Agents

CAMO: An Agentic Framework for Automated Causal Discovery from Micro Behaviors to Macro Emergence in LLM Agent Simulations

DGX agent

arXiv:2604.14691v1 Announce Type: cross Abstract: LLM-empowered agent simulations are increasingly used to study social emergence, yet the micro-to-macro causal mechanisms behind macro outcomes often

agentsarxiv-cs-cl
17 Apr 2026
Applications

Can Large Language Models Detect Methodological Flaws? Evidence from Gesture Recognition for UAV-Based Rescue Operation Based on Deep Learning

DGX agent

arXiv:2604.14161v1 Announce Type: new Abstract: Reliable evaluation is essential in machine learning research, yet methodological flaws-particularly data leakage-continue to undermine the validity of

applicationsarxiv-cs-cl
17 Apr 2026
Model Releases

CausalDetox: Causal Head Selection and Intervention for Language Model Detoxification

DGX agent

arXiv:2604.14602v1 Announce Type: new Abstract: Large language models (LLMs) frequently generate toxic content, posing significant risks for safe deployment. Current mitigation strategies often degrad

model-releasesarxiv-cs-cl
17 Apr 2026
Tutorials

CausalEmbed: Auto-Regressive Multi-Vector Generation in Latent Space for Visual Document Embedding

DGX agent

arXiv:2601.21262v3 Announce Type: replace Abstract: Although Multimodal Large Language Models (MLLMs) have shown remarkable potential in Visual Document Retrieval (VDR) through generating high-quality

tutorialsarxiv-cs-cl
17 Apr 2026
Research

Challenges in Translating Technical Lectures: Insights from the NPTEL

DGX agent

arXiv:2602.08698v2 Announce Type: replace Abstract: This study examines the practical applications and methodological implications of Machine Translation in Indian Languages, specifically Bangla, Mala

researcharxiv-cs-cl
17 Apr 2026
Applications

Chinese Essay Rhetoric Recognition Using LoRA, In-context Learning and Model Ensemble

DGX agent

arXiv:2604.14167v1 Announce Type: new Abstract: Rhetoric recognition is a critical component in automated essay scoring. By identifying rhetorical elements in student writing, AI systems can better as

applicationsarxiv-cs-cl
17 Apr 2026
Model Releases

Chinese Language Is Not More Efficient Than English in Vibe Coding: A Preliminary Study on Token Cost and Problem-Solving Rate

DGX agent

arXiv:2604.14210v1 Announce Type: new Abstract: A claim has been circulating on social media and practitioner forums that Chinese prompts are more token-efficient than English for LLM coding tasks, po

model-releasesarxiv-cs-cl
17 Apr 2026
Research

Chronological Knowledge Retrieval: A Retrieval-Augmented Generation Approach to Construction Project Documentation

DGX agent

arXiv:2604.14169v1 Announce Type: new Abstract: In large-scale construction projects, the continuous evolution of decisions generates extensive records, most often captured in meeting minutes. Since d

researcharxiv-cs-cl
17 Apr 2026
Safety

ClimateCause: Complex and Implicit Causal Structures in Climate Reports

DGX agent

arXiv:2604.14856v1 Announce Type: new Abstract: Understanding climate change requires reasoning over complex causal networks. Yet, existing causal discovery datasets predominantly capture explicit, di

safetyarxiv-cs-cl
17 Apr 2026
Model Releases

CobwebTM: Probabilistic Concept Formation for Lifelong and Hierarchical Topic Modeling

DGX agent

arXiv:2604.14489v1 Announce Type: new Abstract: Topic modeling seeks to uncover latent semantic structure in text corpora with minimal supervision. Neural approaches achieve strong performance but req

model-releasesarxiv-cs-cl
17 Apr 2026
Research

Cognitive Alpha Mining via LLM-Driven Code-Based Evolution

DGX agent

arXiv:2511.18850v2 Announce Type: replace Abstract: Discovering effective predictive signals, or 'alphas,' from financial data with high dimensionality and extremely low signal-to-noise ratio remains

researcharxiv-cs-cl
17 Apr 2026
Model Releases

Comparison of Modern Multilingual Text Embedding Techniques for Hate Speech Detection Task

DGX agent

arXiv:2604.14907v1 Announce Type: new Abstract: Online hate speech and abusive language pose a growing challenge for content moderation, especially in multilingual settings and for low-resource langua

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

Compressed-Sensing-Guided, Inference-Aware Structured Reduction for Large Language Models

DGX agent

arXiv:2604.14156v1 Announce Type: new Abstract: Large language models deliver strong generative performance but at the cost of massive parameter counts, memory use, and decoding latency. Prior work ha

model-releasesarxiv-cs-cl
17 Apr 2026
Research

Compressing Sequences in the Latent Embedding Space: K-Token Merging for Large Language Models

DGX agent

arXiv:2604.15153v1 Announce Type: new Abstract: Large Language Models (LLMs) incur significant computational and memory costs when processing long prompts, as full self-attention scales quadratically

researcharxiv-cs-cl
17 Apr 2026
Safety

ConfLayers: Adaptive Confidence-based Layer Skipping for Self-Speculative Decoding

DGX agent

arXiv:2604.14612v1 Announce Type: cross Abstract: Self-speculative decoding is an inference technique for large language models designed to speed up generation without sacrificing output quality. It c

safetyarxiv-cs-cl
17 Apr 2026
Safety

Context Over Content: Exposing Evaluation Faking in Automated Judges

DGX agent

arXiv:2604.15224v1 Announce Type: cross Abstract: The extit{LLM-as-a-judge} paradigm has become the operational backbone of automated AI evaluation pipelines, yet rests on an unverified assumption: th

safetyarxiv-cs-cl
17 Apr 2026
Model Releases

Controlling Authority Retrieval: A Missing Retrieval Objective for Authority-Governed Knowledge

DGX agent

arXiv:2604.14488v1 Announce Type: cross Abstract: In any domain where knowledge accumulates under formal authority -- law, drug regulation, software security -- a later document can formally void an e

model-releasesarxiv-cs-cl
17 Apr 2026
Safety

CoopEval: Benchmarking Cooperation-Sustaining Mechanisms and LLM Agents in Social Dilemmas

DGX agent

arXiv:2604.15267v1 Announce Type: cross Abstract: It is increasingly important that LLM agents interact effectively and safely with other goal-pursuing agents, yet, recent works report the opposite tr

safetyarxiv-cs-cl
17 Apr 2026
Model Releases

CoPA: Benchmarking Personalized Question Answering with Data-Informed Cognitive Factors

DGX agent

arXiv:2604.14773v1 Announce Type: new Abstract: While LLMs have demonstrated remarkable potential in Question Answering (QA), evaluating personalization remains a critical bottleneck. Existing paradig

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

Correcting Suppressed Log-Probabilities in Language Models with Post-Transformer Adapters

DGX agent

arXiv:2604.14174v1 Announce Type: new Abstract: Alignment-tuned language models frequently suppress factual log-probabilities on politically sensitive topics despite retaining the knowledge in their h

model-releasesarxiv-cs-cl
17 Apr 2026
Research

Cosine-Similarity Routing with Semantic Anchors for Interpretable Mixture-of-Experts Language Models

DGX agent

arXiv:2509.14255v2 Announce Type: replace Abstract: Mixture-of-Experts (MoE) models improve efficiency through sparse activation, but their learned gating functions provide limited insight into routin

researcharxiv-cs-cl
17 Apr 2026
Research

Counting Without Numbers and Finding Without Words

DGX agent

arXiv:2603.24470v2 Announce Type: replace-cross Abstract: Every year, 10 million pets enter shelters, separated from their families. Despite desperate searches by both guardians and lost animals, 70%

researcharxiv-cs-cl
17 Apr 2026
Agents

CROP: Token-Efficient Reasoning in Large Language Models via Regularized Prompt Optimization

DGX agent

arXiv:2604.14214v1 Announce Type: new Abstract: Large Language Models utilizing reasoning techniques improve task performance but incur significant latency and token costs due to verbose generation. E

agentsarxiv-cs-cl
17 Apr 2026
Safety

CURA: Clinical Uncertainty Risk Alignment for Language Model-Based Risk Prediction

DGX agent

arXiv:2604.14651v1 Announce Type: new Abstract: Clinical language models (LMs) are increasingly applied to support clinical risk prediction from free-text notes, yet their uncertainty estimates often

safetyarxiv-cs-cl
17 Apr 2026
Research

CURaTE: Continual Unlearning in Real Time with Ensured Preservation of LLM Knowledge

DGX agent

arXiv:2604.14644v1 Announce Type: new Abstract: The inability to filter out in advance all potentially problematic data from the pre-training of large language models has given rise to the need for me

researcharxiv-cs-cl
17 Apr 2026
Hardware

DA-Cramming: Enhancing Cost-Effective Language Model Pretraining with Dependency Agreement Integration

DGX agent

arXiv:2311.04799v2 Announce Type: replace Abstract: Pretraining language models is still a challenge for many researchers due to its substantial computational costs. As such, there is growing interest

hardwarearxiv-cs-cl
17 Apr 2026
Research

Dark & Stormy: Modeling Humor in Sentences from the Bulwer-Lytton Fiction Contest

DGX agent

arXiv:2510.24538v2 Announce Type: replace Abstract: Textual humor is enormously diverse and computational studies need to account for this range, including intentionally bad humor. In this paper, we c

researcharxiv-cs-cl
17 Apr 2026
Applications

De-Anonymization at Scale via Tournament-Style Attribution

DGX agent

arXiv:2601.12407v2 Announce Type: replace-cross Abstract: As LLMs rapidly advance and enter real-world use, their privacy implications are increasingly important. We study an authorship de-anonymizati

applicationsarxiv-cs-cl
17 Apr 2026
Research

Decoupling Scores and Text: The Politeness Principle in Peer Review

DGX agent

arXiv:2604.14162v1 Announce Type: new Abstract: Authors often struggle to interpret peer review feedback, deriving false hope from polite comments or feeling confused by specific low scores. To invest

researcharxiv-cs-cl
17 Apr 2026
Research

DeepPrune: Parallel Scaling without Inter-trace Redundancy

DGX agent

arXiv:2510.08483v2 Announce Type: replace Abstract: Parallel scaling has emerged as a powerful paradigm to enhance reasoning capabilities in large language models (LLMs) by generating multiple Chain-o

researcharxiv-cs-cl
17 Apr 2026
Model Releases

DharmaOCR: Specialized Small Language Models for Structured OCR that outperform Open-Source and Commercial Baselines

DGX agent

arXiv:2604.14314v1 Announce Type: cross Abstract: This manuscript introduces DharmaOCR Full and Lite, a pair of specialized small language models (SSLMs) for structured OCR that jointly optimize trans

model-releasesarxiv-cs-cl
17 Apr 2026
Research

Diagnosing LLM Judge Reliability: Conformal Prediction Sets and Transitivity Violations

DGX agent

arXiv:2604.15302v1 Announce Type: cross Abstract: LLM-as-judge frameworks are increasingly used for automatic NLG evaluation, yet their per-instance reliability remains poorly understood. We present a

researcharxiv-cs-cl
17 Apr 2026
Tutorials

DiscoTrace: Representing and Comparing Answering Strategies of Humans and LLMs in Information-Seeking Question Answering

DGX agent

arXiv:2604.15140v1 Announce Type: new Abstract: We introduce DiscoTrace, a method to identify the rhetorical strategies that answerers use when responding to information-seeking questions. DiscoTrace

tutorialsarxiv-cs-cl
17 Apr 2026
← Previous
1…141142143144145…160
Next →