AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,646 results
Model Releases

HumorGen: Cognitive Synergy for Humor Generation in Large Language Models via Persona-Based Distillation

DGX agent

arXiv:2604.09629v1 Announce Type: new Abstract: Humor generation poses a significant challenge for Large Language Models (LLMs), because their standard training objective - predicting the most likely

model-releasesarxiv-cs-cl
14 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

HyperGraphPro: Progress-Aware Reinforcement Learning for Structure-Guided Hypergraph RAG

DGX agent

arXiv:2601.17755v2 Announce Type: replace Abstract: Graph Retrieval-Augmented Generation (GraphRAG) has emerged as a promising paradigm that organizes external knowledge into structured graphs of enti

safetyarxiv-cs-cl
14 Apr 2026
Research

Improving LLM Unlearning Robustness via Random Perturbations

DGX agent

arXiv:2501.19202v5 Announce Type: replace Abstract: Here, we show that current LLM unlearning methods inherently reduce models' robustness, causing them to misbehave even when a single non-adversarial

researcharxiv-cs-cl
14 Apr 2026
Model Releases

Infusing Theory of Mind into Socially Intelligent LLM Agents

DGX agent

arXiv:2509.22887v2 Announce Type: replace Abstract: Theory of Mind (ToM)-an understanding of the mental states of others-is a key aspect of human social intelligence, yet, chatbots and LLM-based socia

model-releasesarxiv-cs-cl
14 Apr 2026
Research

Instruction Data Selection via Answer Divergence

DGX agent

arXiv:2604.10448v1 Announce Type: new Abstract: Instruction tuning relies on large instruction-response corpora whose quality and composition strongly affect downstream performance. We propose Answer

researcharxiv-cs-cl
14 Apr 2026
Safety

Judge Like Human Examiners: A Weighted Importance Multi-Point Evaluation Framework for Generative Tasks with Long-form Answers

DGX agent

arXiv:2604.11246v1 Announce Type: new Abstract: Evaluating the quality of model responses remains challenging in generative tasks with long-form answers, as the expected answers usually contain multip

safetyarxiv-cs-cl
14 Apr 2026
Model Releases

K-Way Energy Probes for Metacognition Reduce to Softmax in Discriminative Predictive Coding Networks

DGX agent

arXiv:2604.11011v1 Announce Type: cross Abstract: We present this as a negative result with an explanatory mechanism, not as a formal upper bound. Predictive coding networks (PCNs) admit a K-way energ

model-releasesarxiv-cs-cl
14 Apr 2026
Research

KCS: Diversify Multi-hop Question Generation with Knowledge Composition Sampling

DGX agent

arXiv:2508.20567v2 Announce Type: replace Abstract: Multi-hop question answering faces substantial challenges due to data sparsity, which increases the likelihood of language models learning spurious

researcharxiv-cs-cl
14 Apr 2026
Model Releases

Knowing What to Stress: A Discourse-Conditioned Text-to-Speech Benchmark

DGX agent

arXiv:2604.10580v1 Announce Type: new Abstract: Spoken meaning often depends not only on what is said, but also on which word is emphasized. The same sentence can convey correction, contrast, or clari

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

ks-pret-5m: a 5 million word, 12 million token kashmiri pretraining dataset

DGX agent

arXiv:2604.11066v1 Announce Type: new Abstract: We present KS-PRET-5M, the largest publicly available pretraining dataset for the Kashmiri language, comprising 5,090,244 (5.09M) words, 27,692,959 (27.

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

LaMI: Augmenting Large Language Models via Late Multi-Image Fusion

DGX agent

arXiv:2406.13621v2 Announce Type: replace Abstract: Commonsense reasoning often requires both textual and visual knowledge, yet Large Language Models (LLMs) trained solely on text lack visual groundin

model-releasesarxiv-cs-cl
14 Apr 2026
Research

LangFlow: Continuous Diffusion Rivals Discrete in Language Modeling

DGX agent

arXiv:2604.11748v1 Announce Type: new Abstract: Continuous diffusion models have achieved strong performance across domains such as images. However, in language modeling, prior continuous diffusion la

researcharxiv-cs-cl
14 Apr 2026
Research

LASQ: A Low-resource Aspect-based Sentiment Quadruple Extraction Dataset

DGX agent

arXiv:2604.10417v1 Announce Type: new Abstract: In recent years, aspect-based sentiment analysis (ABSA) has made rapid progress and shown strong practical value. However, existing research and benchma

researcharxiv-cs-cl
14 Apr 2026
Safety

LayerNorm Induces Recency Bias in Transformer Decoders

DGX agent

arXiv:2509.21042v3 Announce Type: replace Abstract: Causal self-attention provides positional information to Transformer decoders. Prior work has shown that stacks of causal self-attention layers alon

safetyarxiv-cs-cl
14 Apr 2026
Model Releases

LIFT: A Novel Framework for Enhancing Long-Context Understanding of LLMs via Long Input Fine-Tuning

DGX agent

arXiv:2502.14644v5 Announce Type: replace Abstract: Long context understanding remains challenging for large language models due to their limited context windows. This paper introduces Long Input Fine

model-releasesarxiv-cs-cl
14 Apr 2026
Research

LingoLoop Attack: Trapping MLLMs via Linguistic Context and State Entrapment into Endless Loops

DGX agent

arXiv:2506.14493v3 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have shown great promise but require substantial computational resources during inference. Attackers can ex

researcharxiv-cs-cl
14 Apr 2026
Research

Linguistic Accommodation Between Neurodivergent Communities on Reddit:A Communication Accommodation Theory Analysis of ADHD and Autism Groups

DGX agent

arXiv:2604.10063v1 Announce Type: new Abstract: Social media research on mental health has focused predominantly on detecting and diagnosing conditions at the individual level. In this work, we shift

researcharxiv-cs-cl
14 Apr 2026
Research

LLMs for Game Theory: Entropy-Guided In-Context Learning and Adaptive CoT Reasoning

DGX agent

arXiv:2601.10775v2 Announce Type: replace Abstract: We propose a novel LLM-based framework for reasoning in discrete, game-theoretic tasks, illustrated with Tic-Tac-Toe. The method integrates in-conte

researcharxiv-cs-cl
14 Apr 2026
Local Ai

Look Twice before You Leap: A Rational Framework for Localized Adversarial Anonymization

DGX agent

arXiv:2512.06713v3 Announce Type: replace-cross Abstract: Current LLM-based frameworks for text anonymization usually rely on remote API services from powerful LLMs, which creates an inherent privacy

local-aiarxiv-cs-cl
14 Apr 2026
Research

Lost in Diffusion: Uncovering Hallucination Patterns and Failure Modes in Diffusion Large Language Models

DGX agent

arXiv:2604.10556v1 Announce Type: new Abstract: While Diffusion Large Language Models (dLLMs) have emerged as a promising non-autoregressive paradigm comparable to autoregressive (AR) models, their fa

researcharxiv-cs-cl
14 Apr 2026
Model Releases

M2-Verify: A Large-Scale Multidomain Benchmark for Checking Multimodal Claim Consistency

DGX agent

arXiv:2604.01306v2 Announce Type: replace Abstract: Evaluating scientific arguments requires assessing the strict consistency between a claim and its underlying multimodal evidence. However, existing

model-releasesarxiv-cs-cl
14 Apr 2026
Safety

MAESTRO: Meta-learning Adaptive Estimation of Scalarization Trade-offs for Reward Optimization

DGX agent

arXiv:2601.07208v2 Announce Type: replace-cross Abstract: Group-Relative Policy Optimization (GRPO) has emerged as an efficient paradigm for aligning Large Language Models (LLMs), yet its efficacy is

safetyarxiv-cs-cl
14 Apr 2026
Agents

MASH: Modeling Abstention via Selective Help-Seeking

DGX agent

arXiv:2510.01152v2 Announce Type: replace Abstract: LLMs cannot reliably recognize their parametric knowledge boundaries and often hallucinate answers to outside-of-boundary questions. In this paper,

agentsarxiv-cs-cl
14 Apr 2026
Model Releases

MCAT: Scaling Many-to-Many Speech-to-Text Translation with MLLMs to 70 Languages

DGX agent

arXiv:2512.01512v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have achieved great success in Speech-to-Text Translation (S2TT) tasks. However, current research is constr

model-releasesarxiv-cs-cl
14 Apr 2026
Research

MCGA: A Multi-task Classical Chinese Literary Genre Audio Corpus

DGX agent

arXiv:2601.09270v3 Announce Type: replace Abstract: With the rapid advancement of Multimodal Large Language Models (MLLMs), their potential has gained significant attention in Chinese Classical Studie

researcharxiv-cs-cl
14 Apr 2026
Model Releases

Measuring What Matters!! Assessing Therapeutic Principles in Mental-Health Conversation

DGX agent

arXiv:2604.05795v2 Announce Type: replace Abstract: The increasing use of large language models in mental health applications calls for principled evaluation frameworks that assess alignment with psyc

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

MEDSYN: Benchmarking Multi-EviDence SYNthesis in Complex Clinical Cases for Multimodal Large Language Models

DGX agent

arXiv:2602.21950v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) have shown great potential in medical applications, yet existing benchmarks inadequately capture real-world

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

MemDLM: Memory-Enhanced DLM Training

DGX agent

arXiv:2603.22241v2 Announce Type: replace Abstract: Diffusion Language Models (DLMs) offer attractive advantages over Auto-Regressive (AR) models, such as full-attention parallel decoding and flexible

model-releasesarxiv-cs-cl
14 Apr 2026
Applications

MerNav: A Highly Generalizable Memory-Execute-Review Framework for Zero-Shot Object Goal Navigation

DGX agent

arXiv:2602.05467v2 Announce Type: replace-cross Abstract: Visual Language Navigation (VLN) is one of the fundamental capabilities for embodied intelligence and a critical challenge that urgently needs

applicationsarxiv-cs-cl
14 Apr 2026
Safety

MimicLM: Zero-Shot Voice Imitation through Autoregressive Modeling of Pseudo-Parallel Speech Corpora

DGX agent

arXiv:2604.11552v1 Announce Type: cross Abstract: Voice imitation aims to transform source speech to match a reference speaker's timbre and speaking style while preserving linguistic content. A straig

safetyarxiv-cs-cl
14 Apr 2026
Research

MIXAR: Scaling Autoregressive Pixel-based Language Models to Multiple Languages and Scripts

DGX agent

arXiv:2604.11575v1 Announce Type: new Abstract: Pixel-based language models are gaining momentum as alternatives to traditional token-based approaches, promising to circumvent tokenization challenges.

researcharxiv-cs-cl
14 Apr 2026
Safety

NameBERT: Scaling Name-Based Nationality Classification with LLM-Augmented Open Academic Data

DGX agent

arXiv:2604.10401v1 Announce Type: new Abstract: Inferring nationality from personal names is a critical capability for equity and bias monitoring, personalization, and a valuable tool in biomedical an

safetyarxiv-cs-cl
14 Apr 2026
Model Releases

Nationality encoding in language model hidden states: Probing culturally differentiated representations in persona-conditioned academic text

DGX agent

arXiv:2604.10151v1 Announce Type: new Abstract: Large language models are increasingly used as writing tools and pedagogical resources in English for Academic Purposes, but it remains unclear whether

model-releasesarxiv-cs-cl
14 Apr 2026
Safety

NOSE: Neural Olfactory-Semantic Embedding with Tri-Modal Orthogonal Contrastive Learning

DGX agent

arXiv:2604.10452v1 Announce Type: new Abstract: Olfaction lies at the intersection of chemical structure, neural encoding, and linguistic perception, yet existing representation methods fail to fully

safetyarxiv-cs-cl
14 Apr 2026
Research

Not All Denoising Steps Are Equal: Model Scheduling for Faster Masked Diffusion Language Models

DGX agent

arXiv:2604.02340v2 Announce Type: replace-cross Abstract: Recent advances in masked diffusion language models (MDLMs) narrow the quality gap to autoregressive LMs, but their sampling remains expensive

researcharxiv-cs-cl
14 Apr 2026
Model Releases

OccuBench: Evaluating AI Agents on Real-World Professional Tasks via Language World Models

DGX agent

arXiv:2604.10866v1 Announce Type: new Abstract: AI agents are expected to perform professional work across hundreds of occupational domains (from emergency department triage to nuclear reactor safety

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

ODUTQA-MDC: A Task for Open-Domain Underspecified Tabular QA with Multi-turn Dialogue-based Clarification

DGX agent

arXiv:2604.10159v1 Announce Type: new Abstract: The advancement of large language models (LLMs) has enhanced tabular question answering (Tabular QA), yet they struggle with open-domain queries exhibit

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Omnimodal Dataset Distillation via High-order Proxy Alignment

DGX agent

arXiv:2604.10666v1 Announce Type: cross Abstract: Dataset distillation compresses large-scale datasets into compact synthetic sets while preserving training performance, but existing methods are large

model-releasesarxiv-cs-cl
14 Apr 2026
Research

PatchRecall: Patch-Driven Retrieval for Automated Program Repair

DGX agent

arXiv:2604.10481v1 Announce Type: cross Abstract: Retrieving the correct set of files from a large codebase is a crucial step in Automated Program Repair (APR). High recall is necessary to ensure that

researcharxiv-cs-cl
14 Apr 2026
Research

Pay Less Attention to Function Words for Free Robustness of Vision-Language Models

DGX agent

arXiv:2512.07222v3 Announce Type: replace-cross Abstract: To address the trade-off between robustness and performance for robust VLM, we observe that function words could incur vulnerability of VLMs a

researcharxiv-cs-cl
14 Apr 2026
Research

Phonological distances for linguistic typology and the origin of Indo-European languages

DGX agent

arXiv:2604.11565v1 Announce Type: new Abstract: We show that short-range phoneme dependencies encode large-scale patterns of linguistic relatedness, with direct implications for quantitative typology

researcharxiv-cs-cl
14 Apr 2026
Research

Physical Commonsense Reasoning for Lower-Resourced Languages and Dialects: a Study on Basque

DGX agent

arXiv:2602.14812v3 Announce Type: replace Abstract: Physical commonsense reasoning represents a fundamental capability of human intelligence, enabling individuals to understand their environment, pred

researcharxiv-cs-cl
14 Apr 2026
Safety

PICon: A Multi-Turn Interrogation Framework for Evaluating Persona Agent Consistency

DGX agent

arXiv:2603.25620v2 Announce Type: replace Abstract: Large language model (LLM)-based persona agents are rapidly being adopted as scalable proxies for human participants across diverse domains. Yet the

safetyarxiv-cs-cl
14 Apr 2026
Model Releases

Please Make it Sound like Human: Encoder-Decoder vs. Decoder-Only Transformers for AI-to-Human Text Style Transfer

DGX agent

arXiv:2604.11687v1 Announce Type: new Abstract: AI-generated text has become common in academic and professional writing, prompting research into detection methods. Less studied is the reverse: system

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Polyglot Teachers: Evaluating Language Models for Multilingual Synthetic Data Generation

DGX agent

arXiv:2604.11290v1 Announce Type: new Abstract: Synthesizing supervised finetuning (SFT) data from language models (LMs) to teach smaller models multilingual tasks has become increasingly common. Howe

model-releasesarxiv-cs-cl
14 Apr 2026
Research

Position-Agnostic Pre-Projection for Transformer Attention: Nonlinear Feature Construction and Content Skip Before Q/K/V

DGX agent

arXiv:2604.10791v1 Announce Type: new Abstract: We propose two complementary modifications to transformer attention blocks. First, a non-linear pre-projection MLP is inserted between layer norm and Q/

researcharxiv-cs-cl
14 Apr 2026
Research

Preference Learning Unlocks LLMs' Psycho-Counseling Skills

DGX agent

arXiv:2502.19731v2 Announce Type: replace Abstract: Applying large language models (LLMs) to assist in psycho-counseling is an emerging and meaningful approach, driven by the significant gap between p

researcharxiv-cs-cl
14 Apr 2026
Model Releases

ProGAL-VLA: Grounded Alignment through Prospective Reasoning in Vision-Language-Action Models

DGX agent

arXiv:2604.09824v1 Announce Type: cross Abstract: Vision language action (VLA) models enable generalist robotic agents but often exhibit language ignorance, relying on visual shortcuts and remaining i

model-releasesarxiv-cs-cl
14 Apr 2026
← Previous
1…151152153154155…160
Next →