AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
Agents

Source-Modality Monitoring in Vision-Language Models

DGX agent

arXiv:2604.22038v1 Announce Type: new Abstract: We define and investigate source-modality monitoring -- the ability of multimodal models to track and communicate the input source from which pieces of

agentsarxiv-cs-cl
27 Apr 2026
Research

STEM: Structure-Tracing Evidence Mining for Knowledge Graphs-Driven Retrieval-Augmented Generation

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2604.22282v1 Announce Type: new Abstract: Knowledge Graph-based Question Answering (KGQA) plays a pivotal role in complex reasoning tasks but remains constrained by two persistent challenges: th

researcharxiv-cs-cl
27 Apr 2026
Safety

Survey Response Generation: Generating Closed-Ended Survey Responses In-Silico with Large Language Models

DGX agent

arXiv:2510.11586v2 Announce Type: replace Abstract: Many in-silico simulations of human survey responses with large language models (LLMs) focus on generating closed-ended survey responses, whereas LL

safetyarxiv-cs-cl
27 Apr 2026
Safety

System-Mediated Attention Imbalances Make Vision-Language Models Say Yes

DGX agent

arXiv:2601.12430v2 Announce Type: replace Abstract: Vision-language model (VLM) hallucination is commonly linked to imbalanced allocation of attention across input modalities: system, image and text.

safetyarxiv-cs-cl
27 Apr 2026
Agents

The Bitter Lesson of Diffusion Language Models for Agentic Workflows: A Comprehensive Reality Check

DGX agent

arXiv:2601.12979v3 Announce Type: replace Abstract: The pursuit of real-time agentic interaction has driven interest in Diffusion-based Large Language Models (dLLMs) as alternatives to auto-regressive

agentsarxiv-cs-cl
27 Apr 2026
Safety

Thinking Without Words: Efficient Latent Reasoning with Abstract Chain-of-Thought

DGX agent

arXiv:2604.22709v1 Announce Type: new Abstract: While long, explicit chains-of-thought (CoT) have proven effective on complex reasoning tasks, they are costly to generate during inference. Non-verbal

safetyarxiv-cs-cl
27 Apr 2026
Model Releases

Toward Automated Robustness Evaluation of Mathematical Reasoning

DGX agent

arXiv:2506.05038v2 Announce Type: replace Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities in various reasoning-intensive tasks. However, these models exhibit unexpecte

model-releasesarxiv-cs-cl
27 Apr 2026
Research

Tracing the complexity profiles of different linguistic phenomena through the intrinsic dimension of LLM representations

DGX agent

arXiv:2601.03779v2 Announce Type: replace Abstract: We explore intrinsic dimension (ID) of LLM representations as a marker of linguistic complexity. Specifically, we test whether ID differences across

researcharxiv-cs-cl
27 Apr 2026
Safety

TTS-PRISM: A Perceptual Reasoning and Interpretable Speech Model for Fine-Grained Diagnosis

DGX agent

arXiv:2604.22225v1 Announce Type: new Abstract: While generative text-to-speech (TTS) models approach human-level quality, monolithic metrics fail to diagnose fine-grained acoustic artifacts or explai

safetyarxiv-cs-cl
27 Apr 2026
Model Releases

UNIKIE-BENCH: Benchmarking Large Multimodal Models for Key Information Extraction in Visual Documents

DGX agent

arXiv:2602.07038v2 Announce Type: replace-cross Abstract: Key Information Extraction (KIE) from real-world documents remains challenging due to substantial variations in layout structures, visual qual

model-releasesarxiv-cs-cl
27 Apr 2026
Research

Using Embedding Models to Improve Probabilistic Race Prediction

DGX agent

arXiv:2604.22555v1 Announce Type: new Abstract: Estimating racial disparity requires individual-level race data, which are often unavailable due to the sensitivity of collecting such information. To a

researcharxiv-cs-cl
27 Apr 2026
Research

Voice Under Revision: Large Language Models and the Normalization of Personal Narrative

DGX agent

arXiv:2604.22142v1 Announce Type: new Abstract: This study examines how large language model rewriting alters the style and narrative texture of personal narratives. It analyzes 300 personal narrative

researcharxiv-cs-cl
27 Apr 2026
Model Releases

When Cow Urine Cures Constipation on YouTube: Limits of LLMs in Detecting Culture-specific Health Misinformation

DGX agent

arXiv:2604.22002v1 Announce Type: new Abstract: Social media platforms have become primary channels for health information in the Global South. Using gomutra (cow urine) discourse on YouTube in India

model-releasesarxiv-cs-cl
27 Apr 2026
Research

Where Should LoRA Go? Component-Type Placement in Hybrid Language Models

DGX agent

arXiv:2604.22127v1 Announce Type: new Abstract: Hybrid language models that interleave attention with recurrent components are increasingly competitive with pure Transformers, yet standard LoRA practi

researcharxiv-cs-cl
27 Apr 2026
Research

Zero-Shot Morphological Discovery in Low-Resource Bantu Languages via Cross-Lingual Transfer and Unsupervised Clustering

DGX agent

arXiv:2604.22723v1 Announce Type: cross Abstract: We present a method for discovering morphological features in low-resource Bantu languages by combining cross-lingual transfer learning with unsupervi

researcharxiv-cs-cl
27 Apr 2026
Model Releases

AFRILANGTUTOR: Advancing Language Tutoring and Culture Education in Low-Resource Languages with Large Language Models

DGX agent

arXiv:2604.20996v1 Announce Type: new Abstract: How can language learning systems be developed for languages that lack sufficient training resources? This challenge is increasingly faced by developers

model-releasesarxiv-cs-cl
24 Apr 2026
Safety

AgentGL: Towards Agentic Graph Learning with LLMs via Reinforcement Learning

DGX agent

arXiv:2604.05846v2 Announce Type: replace Abstract: Large Language Models (LLMs) increasingly rely on agentic capabilities-iterative retrieval, tool use, and decision-making-to overcome the limits of

safetyarxiv-cs-cl
24 Apr 2026
Agents

AgenticQwen: Training Small Agentic Language Models with Dual Data Flywheels for Industrial-Scale Tool Use

DGX agent

arXiv:2604.21590v1 Announce Type: new Abstract: Modern industrial applications increasingly demand language models that act as agents, capable of multi-step reasoning and tool use in real-world settin

agentsarxiv-cs-cl
24 Apr 2026
Model Releases

AITP: Traffic Accident Responsibility Allocation via Multimodal Large Language Models

DGX agent

arXiv:2604.20878v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable progress in Traffic Accident Detection (TAD) and Traffic Accident Understanding (TAU).

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

AUDITA: A New Dataset to Audit Humans vs. AI Skill at Audio QA

DGX agent

arXiv:2604.21766v1 Announce Type: new Abstract: Existing audio question answering benchmarks largely emphasize sound event classification or caption-grounded queries, often enabling models to succeed

model-releasesarxiv-cs-cl
24 Apr 2026
Agents

Automating Computational Reproducibility in Social Science: Comparing Prompt-Based and Agent-Based Approaches

DGX agent

arXiv:2602.08561v3 Announce Type: replace-cross Abstract: Reproducing computational research is often assumed to be as simple as rerunning the original code with provided data. In practice, missing pa

agentsarxiv-cs-cl
24 Apr 2026
Applications

Back to the Future: The Role of Past and Future Context Predictability in Incremental Language Production

DGX agent

arXiv:2511.07752v3 Announce Type: replace Abstract: Contextual predictability shapes how we choose and encode words in production. The effects of a word's predictability given preceding or past contex

applicationsarxiv-cs-cl
24 Apr 2026
Model Releases

Beyond N-gram: Data-Aware X-GRAM Extraction for Efficient Embedding Parameter Scaling

DGX agent

arXiv:2604.21724v1 Announce Type: new Abstract: Large token-indexed lookup tables provide a compute-decoupled scaling path, but their practical gains are often limited by poor parameter efficiency and

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

Beyond Pixels: Introspective and Interactive Grounding for Visualization Agents

DGX agent

arXiv:2604.21134v1 Announce Type: new Abstract: Vision-Language Models (VLMs) frequently misread values, hallucinate details, and confuse overlapping elements in charts. Current approaches rely solely

model-releasesarxiv-cs-cl
24 Apr 2026
Applications

Capabilities and Evaluation Biases of Large Language Models in Classical Chinese Poetry Generation: A Case Study on Tang Poetry

DGX agent

arXiv:2510.15313v2 Announce Type: replace Abstract: Large Language Models (LLMs) are increasingly applied to creative domains, yet their performance in classical Chinese poetry generation and evaluati

applicationsarxiv-cs-cl
24 Apr 2026
Safety

CARE: Counselor-Aligned Response Engine for Online Mental-Health Support

DGX agent

arXiv:2604.21352v1 Announce Type: new Abstract: Mental health challenges are increasing worldwide, straining emotional support services and leading to counselor overload. This can result in delayed re

safetyarxiv-cs-cl
24 Apr 2026
Safety

CE-GPPO: Coordinating Entropy via Gradient-Preserving Clipping Policy Optimization in Reinforcement Learning

DGX agent

arXiv:2509.20712v5 Announce Type: replace-cross Abstract: Reinforcement learning (RL) has become a powerful paradigm for optimizing large language models (LLMs) to handle complex reasoning tasks. A co

safetyarxiv-cs-cl
24 Apr 2026
Model Releases

CI-Work: Benchmarking Contextual Integrity in Enterprise LLM Agents

DGX agent

arXiv:2604.21308v1 Announce Type: cross Abstract: Enterprise LLM agents can dramatically improve workplace productivity, but their core capability, retrieving and using internal context to act on a us

model-releasesarxiv-cs-cl
24 Apr 2026
Applications

Cross-Domain Data Selection and Augmentation for Automatic Compliance Detection

DGX agent

arXiv:2604.21469v1 Announce Type: new Abstract: Automating the detection of regulatory compliance remains a challenging task due to the complexity and variability of legal texts. Models trained on one

applicationsarxiv-cs-cl
24 Apr 2026
Model Releases

Decoupled DiLoCo for Resilient Distributed Pre-training

DGX agent

arXiv:2604.21428v1 Announce Type: new Abstract: Modern large-scale language model pre-training relies heavily on the single program multiple data (SPMD) paradigm, which requires tight coupling across

model-releasesarxiv-cs-cl
24 Apr 2026
Research

DMAP: A Distribution Map for Text

DGX agent

arXiv:2602.11871v2 Announce Type: replace Abstract: Large Language Models (LLMs) are a powerful tool for statistical text analysis, with derived sequences of next-token probability distributions offer

researcharxiv-cs-cl
24 Apr 2026
Model Releases

Do LLMs Overthink Basic Math Reasoning? Benchmarking the Accuracy-Efficiency Tradeoff in Language Models

DGX agent

arXiv:2507.04023v3 Announce Type: replace Abstract: Large language models (LLMs) achieve impressive performance on complex mathematical benchmarks yet sometimes fail on basic math reasoning while gene

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

Dr. Assistant: Enhancing Clinical Diagnostic Inquiry via Structured Diagnostic Reasoning Data and Reinforcement Learning

DGX agent

arXiv:2601.13690v2 Announce Type: replace Abstract: Clinical Decision Support Systems (CDSSs) provide reasoning and inquiry guidance for physicians, yet they face notable challenges, including high ma

model-releasesarxiv-cs-cl
24 Apr 2026
Local Ai

DWTSumm: Discrete Wavelet Transform for Document Summarization

DGX agent

arXiv:2604.21070v1 Announce Type: new Abstract: Summarizing long, domain-specific documents with large language models (LLMs) remains challenging due to context limitations, information loss, and hall

local-aiarxiv-cs-cl
24 Apr 2026
Applications

EduCoder: An Open-Source Annotation System for Education Transcript Data

DGX agent

arXiv:2507.05385v4 Announce Type: replace Abstract: We introduce EduCoder, a domain-specialized tool designed to support utterance-level annotation of educational dialogue. While general-purpose text

applicationsarxiv-cs-cl
24 Apr 2026
Safety

Entropy Ratio Clipping as a Soft Global Constraint for Stable Reinforcement Learning

DGX agent

arXiv:2512.05591v2 Announce Type: replace-cross Abstract: Large language model post-training relies on reinforcement learning to improve model capability and alignment quality. However, the off-policy

safetyarxiv-cs-cl
24 Apr 2026
Research

Evaluation of Automatic Speech Recognition Using Generative Large Language Models

DGX agent

arXiv:2604.21928v1 Announce Type: new Abstract: Automatic Speech Recognition (ASR) is traditionally evaluated using Word Error Rate (WER), a metric that is insensitive to meaning. Embedding-based sema

researcharxiv-cs-cl
24 Apr 2026
Model Releases

EVENT5Ws: A Large Dataset for Open-Domain Event Extraction from Documents

DGX agent

arXiv:2604.21890v1 Announce Type: new Abstract: Event extraction identifies the central aspects of events from text. It supports event understanding and analysis, which is crucial for tasks such as in

model-releasesarxiv-cs-cl
24 Apr 2026
Tutorials

Explainable Disentangled Representation Learning for Generalizable Authorship Attribution in the Era of Generative AI

DGX agent

arXiv:2604.21300v1 Announce Type: new Abstract: Learning robust representations of authorial style is crucial for authorship attribution and AI-generated text detection. However, existing methods ofte

tutorialsarxiv-cs-cl
24 Apr 2026
Tutorials

Exploring Continual Fine-Tuning for Enhancing Language Ability in Large Language Model

DGX agent

arXiv:2410.16006v3 Announce Type: replace Abstract: A common challenge towards the adaptability of Large Language Models (LLMs) is their ability to learn new languages over time without hampering the

tutorialsarxiv-cs-cl
24 Apr 2026
Research

Finding Meaning in Embeddings: Concept Separation Curves

DGX agent

arXiv:2604.21555v1 Announce Type: new Abstract: Sentence embedding techniques aim to encode key concepts of a sentence's meaning in a vector space. However, the majority of evaluation approaches for s

researcharxiv-cs-cl
24 Apr 2026
Research

Fixation Sequences as Time Series: A Topological Approach to Dyslexia Detection

DGX agent

arXiv:2604.21698v1 Announce Type: new Abstract: Persistent homology, a method from topological data analysis, extracts robust, multi-scale features from data. It produces stable representations of tim

researcharxiv-cs-cl
24 Apr 2026
Safety

Flipping Against All Odds: Reducing LLM Coin Flip Bias via Verbalized Rejection Sampling

DGX agent

arXiv:2506.09998v2 Announce Type: replace-cross Abstract: Large language models (LLMs) can often accurately describe probability distributions using natural language, yet they still struggle to genera

safetyarxiv-cs-cl
24 Apr 2026
Safety

From If-Statements to ML Pipelines: Revisiting Bias in Code-Generation

DGX agent

arXiv:2604.21716v1 Announce Type: new Abstract: Prior work evaluates code generation bias primarily through simple conditional statements, which represent only a narrow slice of real-world programming

safetyarxiv-cs-cl
24 Apr 2026
Safety

From Past To Path: Masked History Learning for Next-Item Prediction in Generative Recommendation

DGX agent

arXiv:2509.23649v2 Announce Type: replace-cross Abstract: Generative recommendation, which directly generates item identifiers, has emerged as a promising paradigm for recommendation systems. However,

safetyarxiv-cs-cl
24 Apr 2026
Research

From Tokens to Concepts: Leveraging SAE for SPLADE

DGX agent

arXiv:2604.21511v1 Announce Type: cross Abstract: Learned Sparse IR models, such as SPLADE, offer an excellent efficiency-effectiveness tradeoff. However, they rely on the underlying backbone vocabula

researcharxiv-cs-cl
24 Apr 2026
Model Releases

GerAV: Towards New Heights in German Authorship Verification using Fine-Tuned LLMs on a New Benchmark

DGX agent

arXiv:2601.13711v2 Announce Type: replace Abstract: Authorship verification (AV) is the task of determining whether two texts were written by the same author and has been studied extensively, predomin

model-releasesarxiv-cs-cl
24 Apr 2026
Research

GRISP: Guided Recurrent IRI Selection over SPARQL Skeletons

DGX agent

arXiv:2604.21133v1 Announce Type: new Abstract: We present GRISP (Guided Recurrent IRI Selection over SPARQL Skeletons), a novel SPARQL-based question-answering method over knowledge graphs based on f

researcharxiv-cs-cl
24 Apr 2026
← Previous
1…124125126127128…161
Next →