AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
Safety

OS-SPEAR: A Toolkit for the Safety, Performance,Efficiency, and Robustness Analysis of OS Agents

DGX agent

arXiv:2604.24348v1 Announce Type: new Abstract: The evolution of Multimodal Large Language Models (MLLMs) has shifted the focus from text generation to active behavioral execution, particularly via OS

safetyarxiv-cs-cl
28 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

Overcoming Copyright Barriers in Corpus Distribution Through Non-Reversible Hashing

DGX agent

arXiv:2604.23412v1 Announce Type: new Abstract: While annotated corpora are crucial in the field of natural language processing (NLP), those containing copyrighted material are difficult to exchange a

safetyarxiv-cs-cl
28 Apr 2026
Research

PeeriScope: A Multi-Faceted Framework for Evaluating Peer Review Quality

DGX agent

arXiv:2604.24071v1 Announce Type: new Abstract: The increasing scale and variability of peer review in scholarly venues has created an urgent need for systematic, interpretable, and extensible tools t

researcharxiv-cs-cl
28 Apr 2026
Safety

Personality Shapes Gender Bias in Persona-Conditioned LLM Narratives Across English and Hindi: An Empirical Investigation

DGX agent

arXiv:2604.23600v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed in persona-driven applications such as education, customer service, and social platforms, where m

safetyarxiv-cs-cl
28 Apr 2026
Safety

Position: Logical Soundness is not a Reliable Criterion for Neurosymbolic Fact-Checking with LLMs

DGX agent

arXiv:2604.04177v2 Announce Type: replace Abstract: As large language models (LLMs) are increasing integrated into fact-checking pipelines, formal logic is often proposed as a rigorous means by which

safetyarxiv-cs-cl
28 Apr 2026
Safety

Preserving Long-Tailed Expert Information in Mixture-of-Experts Tuning

DGX agent

arXiv:2604.23036v1 Announce Type: cross Abstract: Despite MoE models leading many benchmarks, supervised fine-tuning (SFT) for the MoE architectures remains difficult because its router layers are fra

safetyarxiv-cs-cl
28 Apr 2026
Research

Process Supervision of Confidence Margin for Calibrated LLM Reasoning

DGX agent

arXiv:2604.23333v1 Announce Type: cross Abstract: Scaling test-time computation with reinforcement learning (RL) has emerged as a reliable path to improve large language models (LLM) reasoning ability

researcharxiv-cs-cl
28 Apr 2026
Tutorials

Propagation Structure-Semantic Transfer Learning for Robust Fake News Detection

DGX agent

arXiv:2604.23974v1 Announce Type: new Abstract: Fake news generally refers to false information that is spread deliberately to deceive people, which has detrimental social effects. Existing fake news

tutorialsarxiv-cs-cl
28 Apr 2026
Model Releases

Psychologically-Grounded Graph Modeling for Interpretable Depression Detection

DGX agent

arXiv:2604.24126v1 Announce Type: new Abstract: Automatic depression detection from conversational interactions holds significant promise for scalable screening but remains hindered by severe data sca

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

Rank, Head-Channel Non-Identifiability, and Symmetry Breaking: A Precise Analysis of Representational Collapse in Transformers

DGX agent

arXiv:2604.23681v1 Announce Type: cross Abstract: A widely cited result by Dong et al. (2021) showed that Transformers built from self-attention alone, without skip connections or feed-forward layers,

model-releasesarxiv-cs-cl
28 Apr 2026
Research

Reducing Redundancy in Retrieval-Augmented Generation through Chunk Filtering

DGX agent

arXiv:2604.24334v1 Announce Type: new Abstract: Standard Retrieval-Augmented Generation (RAG) chunking methods often create excessive redundancy, increasing storage costs and slowing retrieval. This s

researcharxiv-cs-cl
28 Apr 2026
Model Releases

Resource-Lean Lexicon Induction for German Dialects

DGX agent

arXiv:2604.23824v1 Announce Type: new Abstract: Automatic induction of high-quality dictionaries is essential for building lexical resources, yet low-resource languages and dialects pose several chall

model-releasesarxiv-cs-cl
28 Apr 2026
Research

Revisiting Greedy Decoding for Visual Question Answering: A Calibration Perspective

DGX agent

arXiv:2604.23443v1 Announce Type: new Abstract: Stochastic sampling strategies are widely adopted in large language models (LLMs) to balance output coherence and diversity. These heuristics are often

researcharxiv-cs-cl
28 Apr 2026
Model Releases

Robust Audio-Text Retrieval via Cross-Modal Attention and Hybrid Loss

DGX agent

arXiv:2604.23323v1 Announce Type: new Abstract: Audio-text retrieval enables semantic alignment between audio content and natural language queries, supporting applications in multimedia search, access

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

RouteNLP: Closed-Loop LLM Routing with Conformal Cascading and Distillation Co-Optimization

DGX agent

arXiv:2604.23577v1 Announce Type: new Abstract: Serving diverse NLP workloads with large language models is costly: at one enterprise partner, inference costs exceeded $200K/month despite over 70% of

model-releasesarxiv-cs-cl
28 Apr 2026
Research

SEARCH-R: Structured Entity-Aware Retrieval with Chain-of-Reasoning Navigator for Multi-hop Question Answering

DGX agent

arXiv:2604.24515v1 Announce Type: new Abstract: Multi-hop Question Answering (MHQA) aims to answer questions that require multi-step reasoning. It presents two key challenges: generating correct reaso

researcharxiv-cs-cl
28 Apr 2026
Model Releases

Secure On-Premise Deployment of Open-Weights Large Language Models in Radiology: An Isolation-First Architecture with Prospective Pilot Evaluation

DGX agent

arXiv:2604.22768v1 Announce Type: cross Abstract: Purpose: To design, implement, evaluate, and report on the regulatory requirements of a self-hosted LLM infrastructure for radiology adhering to the p

model-releasesarxiv-cs-cl
28 Apr 2026
Research

Sentiment and Emotion Classification of Indonesian E-Commerce Reviews via Multi-Task BiLSTM and AutoML Benchmarking

DGX agent

arXiv:2604.24720v1 Announce Type: new Abstract: Indonesian marketplace reviews mix standard vocabulary with slang, regional loanwords, numeric shorthands, and emoji, making lexicon-based sentiment too

researcharxiv-cs-cl
28 Apr 2026
Model Releases

ShredBench: Evaluating the Semantic Reasoning Capabilities of Multimodal LLMs in Document Reconstruction

DGX agent

arXiv:2604.23813v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable performance in Visually Rich Document Understanding (VRDU) tasks, but their capabili

model-releasesarxiv-cs-cl
28 Apr 2026
Research

Spectro-Temporal Modulation Representation Framework for Human-Imitated Speech Detection

DGX agent

arXiv:2604.23241v1 Announce Type: cross Abstract: Human-imitated speech poses a greater challenge than AI-generated speech for both human listeners and automatic detection systems. Unlike AI-generated

researcharxiv-cs-cl
28 Apr 2026
Model Releases

Stabilizing Efficient Reasoning with Step-Level Advantage Selection

DGX agent

arXiv:2604.24003v1 Announce Type: new Abstract: Large language models (LLMs) achieve strong reasoning performance by allocating substantial computation at inference time, often generating long and ver

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

Stress-Testing Emotional Support Models: Moving from Homogeneous to Diverse Help Seekers

DGX agent

arXiv:2601.07698v2 Announce Type: replace Abstract: As emotional support chatbots have recently gained significant traction across both research and industry, a common evaluation strategy has emerged:

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

Structural Pruning of Large Vision Language Models: A Comprehensive Study on Pruning Dynamics, Recovery, and Data Efficiency

DGX agent

arXiv:2604.24380v1 Announce Type: new Abstract: While Large Vision Language Models (LVLMs) demonstrate impressive capabilities, their substantial computational and memory requirements pose deployment

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

Supernodes and Halos: Loss-Critical Hubs in LLM Feed-Forward Layers

DGX agent

arXiv:2604.23475v1 Announce Type: cross Abstract: We study the organization of channel-level importance in transformer feed-forward networks (FFNs). Using a Fisher-style loss proxy (LP) based on activ

model-releasesarxiv-cs-cl
28 Apr 2026
Research

Swa-bhasha Resource Hub: Romanized Sinhala to Sinhala Transliteration Systems and Data Resources

DGX agent

arXiv:2507.09245v2 Announce Type: replace Abstract: The Swa-bhasha Resource Hub provides a comprehensive collection of data resources and algorithms developed for Romanized Sinhala to Sinhala translit

researcharxiv-cs-cl
28 Apr 2026
Agents

SWE-Pruner: Self-Adaptive Context Pruning for Coding Agents

DGX agent

arXiv:2601.16746v3 Announce Type: replace-cross Abstract: LLM agents have demonstrated remarkable capabilities in software development, but their performance is hampered by long interaction contexts,

agentsarxiv-cs-cl
28 Apr 2026
Model Releases

SWE-QA: Can Language Models Answer Repository-level Code Questions?

DGX agent

arXiv:2509.14635v2 Announce Type: replace Abstract: Understanding and reasoning about entire software repositories is an essential capability for intelligent software engineering tools. While existing

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

SwissGov-RSD: A Human-annotated, Cross-lingual Benchmark for Token-level Recognition of Semantic Differences Between Related Documents

DGX agent

arXiv:2512.07538v3 Announce Type: replace Abstract: Recognizing semantic differences across documents is crucial for text generation evaluation and content alignment, especially in cross-lingual setti

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

Switch Attention: Towards Dynamic and Fine-grained Hybrid Transformers

DGX agent

arXiv:2603.26380v2 Announce Type: replace Abstract: The attention mechanism has been the core component in modern transformer architectures. However, the computation of standard full attention scales

model-releasesarxiv-cs-cl
28 Apr 2026
Safety

Synthetic Eggs in Many Baskets: The Impact of Synthetic Data Diversity on LLM Fine-Tuning

DGX agent

arXiv:2511.01490v2 Announce Type: replace Abstract: As synthetic data becomes widely used in language model development, understanding its impact on model behavior is crucial. This paper investigates

safetyarxiv-cs-cl
28 Apr 2026
Research

Talker-T2AV: Joint Talking Audio-Video Generation with Autoregressive Diffusion Modeling

DGX agent

arXiv:2604.23586v1 Announce Type: cross Abstract: Joint audio-video generation models have shown that unified generation yields stronger cross-modal coherence than cascaded approaches. However, existi

researcharxiv-cs-cl
28 Apr 2026
Model Releases

TexOCR: Advancing Document OCR Models for Compilable Page-to-LaTeX Reconstruction

DGX agent

arXiv:2604.22880v1 Announce Type: new Abstract: Existing document OCR largely targets plain text or Markdown, discarding the structural and executable properties that make LaTeX essential for scientif

model-releasesarxiv-cs-cl
28 Apr 2026
Agents

The Chameleon's Limit: Investigating Persona Collapse and Homogenization in Large Language Models

DGX agent

arXiv:2604.24698v1 Announce Type: new Abstract: Applications based on large language models (LLMs), such as multi-agent simulations, require population diversity among agents. We identify a pervasive

agentsarxiv-cs-cl
28 Apr 2026
Safety

The Collapse of Heterogeneity in Silicon Philosophers

DGX agent

arXiv:2604.23575v1 Announce Type: cross Abstract: Silicon samples are increasingly used as a low-cost substitute for human panels and have been shown to reproduce aggregate human opinion with high fid

safetyarxiv-cs-cl
28 Apr 2026
Applications

The Limits of Artificial Companionship

DGX agent

arXiv:2604.23601v1 Announce Type: cross Abstract: This Article argues that conversations with companion chatbot should be subject to a clear structural distinction between commercial and non-commercia

applicationsarxiv-cs-cl
28 Apr 2026
Model Releases

The Surprising Effectiveness of Membership Inference with Simple N-Gram Coverage

DGX agent

arXiv:2508.09603v2 Announce Type: replace Abstract: Membership inference attacks serves as useful tool for fair use of language models, such as detecting potential copyright infringement and auditing

model-releasesarxiv-cs-cl
28 Apr 2026
Safety

Training a General Purpose Automated Red Teaming Model

DGX agent

arXiv:2604.23067v1 Announce Type: cross Abstract: Automated methods for red teaming LLMs are an important tool to identify LLM vulnerabilities that may not be covered in static benchmarks, allowing fo

safetyarxiv-cs-cl
28 Apr 2026
Research

Translate or Simplify First: An Analysis of Cross-lingual Text Simplification in English and French

DGX agent

arXiv:2604.23844v1 Announce Type: new Abstract: Cross-Lingual Text Simplification (CLTS) aims to make content more accessible across languages by simultaneously addressing both linguistic complexity a

researcharxiv-cs-cl
28 Apr 2026
Safety

TSAssistant: A Human-in-the-Loop Agentic Framework for Automated Target Safety Assessment

DGX agent

arXiv:2604.23938v1 Announce Type: new Abstract: Target Safety Assessment (TSA) requires systematic integration of heterogeneous evidence, including genetic, transcriptomic, target homology, pharmacolo

safetyarxiv-cs-cl
28 Apr 2026
Agents

Uncertainty Quantification for LLM Function-Calling

DGX agent

arXiv:2604.22985v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed to autonomously solve real-world tasks. A key ingredient for this is the LLM Function-Calling par

agentsarxiv-cs-cl
28 Apr 2026
Model Releases

V-SEAM: Visual Semantic Editing and Attention Modulating for Causal Interpretability of Vision-Language Models

DGX agent

arXiv:2509.14837v2 Announce Type: replace Abstract: Recent advances in causal interpretability have extended from language models to vision-language models (VLMs), seeking to reveal their internal mec

model-releasesarxiv-cs-cl
28 Apr 2026
Applications

VeriLLMed: Interactive Visual Debugging of Medical Large Language Models with Knowledge Graphs

DGX agent

arXiv:2604.23356v1 Announce Type: new Abstract: Large language models (LLMs) show promise in medical diagnosis, but real-world deployment remains challenging due to high-stakes clinical decisions and

applicationsarxiv-cs-cl
28 Apr 2026
Research

What Prompts Don't Say: Understanding and Managing Underspecification in LLM Prompts

DGX agent

arXiv:2505.13360v3 Announce Type: replace Abstract: Prompt underspecification is a common challenge when interacting with LLMs. In this paper, we present an in-depth analysis of this problem, showing

researcharxiv-cs-cl
28 Apr 2026
Safety

When Annotators Agree but Labels Disagree: The Projection Problem in Stance Detection

DGX agent

arXiv:2603.24231v2 Announce Type: replace Abstract: Stance detection is nearly always formulated as classifying text into Favor, Against, or Neutral. This convention was inherited from debate analysis

safetyarxiv-cs-cl
28 Apr 2026
Model Releases

When Does Removing LayerNorm Help? Activation Bounding as a Regime-Dependent Implicit Regularizer

DGX agent

arXiv:2604.23434v1 Announce Type: cross Abstract: Dynamic Tanh (DyT) removes LayerNorm by bounding activations with a learned tanh(alpha x). We show that this bounding is a regime-dependent implicit r

model-releasesarxiv-cs-cl
28 Apr 2026
Applications

When Silence Matters: The Impact of Irrelevant Audio on Text Reasoning in Large Audio-Language Models

DGX agent

arXiv:2510.00626v3 Announce Type: replace-cross Abstract: Large audio-language models (LALMs) unify speech and text processing, but their robustness in noisy real-world settings remains underexplored.

applicationsarxiv-cs-cl
28 Apr 2026
Research

When to Commit? Towards Variable-Size Self-Contained Blocks for Discrete Diffusion Language Models

DGX agent

arXiv:2604.23994v1 Announce Type: cross Abstract: Discrete diffusion language models (dLLMs) enable parallel token updates with bidirectional attention, yet practical generation typically adopts block

researcharxiv-cs-cl
28 Apr 2026
Research

XITE: Cross-lingual Interpolation for Transfer using Embeddings

DGX agent

arXiv:2604.23589v1 Announce Type: new Abstract: Facilitating cross-lingual transfer in multilingual language models remains a critical challenge. Towards this goal, we propose an embedding-based data

researcharxiv-cs-cl
28 Apr 2026
← Previous
1…122123124125126…161
Next →