AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,646 results
Safety

ProUIE: A Macro-to-Micro Progressive Learning Method for LLM-based Universal Information Extraction

DGX agent

arXiv:2604.10633v1 Announce Type: new Abstract: LLM-based universal information extraction (UIE) methods often rely on additional information beyond the original training data, which increases trainin

safetyarxiv-cs-cl
14 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Local Ai

Psychological Concept Neurons: Can Neural Control Bias Probing and Shift Generation in LLMs?

DGX agent

arXiv:2604.11802v1 Announce Type: new Abstract: Using psychological constructs such as the Big Five, large language models (LLMs) can imitate specific personality profiles and predict a user's persona

local-aiarxiv-cs-cl
14 Apr 2026
Safety

QFS-Composer: Query-focused summarization pipeline for less resourced languages

DGX agent

arXiv:2604.10687v1 Announce Type: new Abstract: Large language models (LLMs) demonstrate strong performance in text summarization, yet their effectiveness drops significantly across languages with res

safetyarxiv-cs-cl
14 Apr 2026
Safety

Quantifying the Climate Risk of Generative AI: Region-Aware Carbon Accounting with G-TRACE and the AI Sustainability Pyramid

DGX agent

arXiv:2511.04776v2 Announce Type: replace-cross Abstract: Generative Artificial Intelligence (GenAI) represents a rapidly expanding digital infrastructure whose energy demand and associated CO2 emissi

safetyarxiv-cs-cl
14 Apr 2026
Model Releases

RCBSF: A Multi-Agent Framework for Automated Contract Revision via Stackelberg Game

DGX agent

arXiv:2604.10740v1 Announce Type: new Abstract: Despite the widespread adoption of Large Language Models (LLMs) in Legal AI, their utility for automated contract revision remains impeded by hallucinat

model-releasesarxiv-cs-cl
14 Apr 2026
Research

Reason Only When Needed: Efficient Generative Reward Modeling via Model-Internal Uncertainty

DGX agent

arXiv:2604.10072v1 Announce Type: new Abstract: Recent advancements in the Generative Reward Model (GRM) have demonstrated its potential to enhance the reasoning abilities of LLMs through Chain-of-Tho

researcharxiv-cs-cl
14 Apr 2026
Safety

Reasoning Resides in Layers: Restoring Temporal Reasoning in Video-Language Models with Layer-Selective Merging

DGX agent

arXiv:2604.11399v1 Announce Type: cross Abstract: Multimodal adaptation equips large language models (LLMs) with perceptual capabilities, but often weakens the reasoning ability inherited from languag

safetyarxiv-cs-cl
14 Apr 2026
Research

RedNote-Vibe: A Dataset for Capturing Temporal Dynamics of AI-Generated Text in Lifestyle Social Media

DGX agent

arXiv:2509.22055v2 Announce Type: replace Abstract: We introduce RedNote-Vibe, a dataset spanning five years (pre-LLM to July 2025) sourced from lifestyle platform RedNote (Xiaohongshu), capturing the

researcharxiv-cs-cl
14 Apr 2026
Hardware

Relational Probing: LM-to-Graph Adaptation for Financial Prediction

DGX agent

arXiv:2604.10212v1 Announce Type: new Abstract: Language models can be used to identify relationships between financial entities in text. However, while structured output mechanisms exist, prompting-b

hardwarearxiv-cs-cl
14 Apr 2026
Model Releases

Relax: An Asynchronous Reinforcement Learning Engine for Omni-Modal Post-Training at Scale

DGX agent

arXiv:2604.11554v1 Announce Type: new Abstract: Reinforcement learning (RL) post-training has proven effective at unlocking reasoning, self-reflection, and tool-use capabilities in large language mode

model-releasesarxiv-cs-cl
14 Apr 2026
Research

Reproduction Beyond Benchmarks: ConstBERT and ColBERT-v2 Across Backends and Query Distributions

DGX agent

arXiv:2604.09982v1 Announce Type: cross Abstract: Reproducibility must validate architectural robustness, not just numerical accuracy. We evaluate ColBERT-v2 and ConstBERT across five dimensions, find

researcharxiv-cs-cl
14 Apr 2026
Safety

Rethinking LLM Watermark Detection in Black-Box Settings: A Non-Intrusive Third-Party Framework

DGX agent

arXiv:2603.14968v2 Announce Type: replace-cross Abstract: While watermarking serves as a critical mechanism for LLM provenance, existing secret-key schemes tightly couple detection with injection, req

safetyarxiv-cs-cl
14 Apr 2026
Safety

Revisiting Compositionality in Dual-Encoder Vision-Language Models: The Role of Inference

DGX agent

arXiv:2604.11496v1 Announce Type: cross Abstract: Dual-encoder Vision-Language Models (VLMs) such as CLIP are often characterized as bag-of-words systems due to their poor performance on compositional

safetyarxiv-cs-cl
14 Apr 2026
Safety

Revisiting Epistemic Markers in Confidence Estimation: Can Markers Accurately Reflect Large Language Models' Uncertainty?

DGX agent

arXiv:2505.24778v3 Announce Type: replace Abstract: As large language models (LLMs) are increasingly used in high-stakes domains, accurately assessing their confidence is crucial. Humans typically exp

safetyarxiv-cs-cl
14 Apr 2026
Model Releases

RiTeK: A Dataset for Large Language Models Complex Reasoning over Textual Knowledge Graphs in Medicine

DGX agent

arXiv:2410.13987v3 Announce Type: replace Abstract: Answering complex real-world questions in the medical domain often requires accurate retrieval from medical Textual Knowledge Graphs (medical TKGs),

model-releasesarxiv-cs-cl
14 Apr 2026
Research

RUMLEM: A Dictionary-Based Lemmatizer for Romansh

DGX agent

arXiv:2604.11233v1 Announce Type: new Abstract: Lemmatization -- the task of mapping an inflected word form to its dictionary form -- is a crucial component of many NLP applications. In this paper, we

researcharxiv-cs-cl
14 Apr 2026
Local Ai

Saar-Voice: A Multi-Speaker Saarbrucken Dialect Speech Corpus

DGX agent

arXiv:2604.11803v1 Announce Type: new Abstract: Natural language processing (NLP) and speech technologies have made significant progress in recent years; however, they remain largely focused on standa

local-aiarxiv-cs-cl
14 Apr 2026
Safety

SafeConstellations: Mitigating Over-Refusals in LLMs Through Task-Aware Representation Steering

DGX agent

arXiv:2508.11290v3 Announce Type: replace Abstract: LLMs increasingly exhibit over-refusal behavior, where safety mechanisms cause models to reject benign instructions that seemingly resemble harmful

safetyarxiv-cs-cl
14 Apr 2026
Model Releases

Seeing No Evil: Blinding Large Vision-Language Models to Safety Instructions via Adversarial Attention Hijacking

DGX agent

arXiv:2604.10299v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) rely on attention-based retrieval of safety instructions to maintain alignment during generation. Existing attack

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Seeing Through Deception: Uncovering Misleading Creator Intent in Multimodal News with Vision-Language Models

DGX agent

arXiv:2505.15489v4 Announce Type: replace-cross Abstract: The impact of multimodal misinformation arises not only from factual inaccuracies but also from the misleading narratives that creators delibe

model-releasesarxiv-cs-cl
14 Apr 2026
Research

Self-Calibrating Language Models via Test-Time Discriminative Distillation

DGX agent

arXiv:2604.09624v1 Announce Type: new Abstract: Large language models (LLMs) are systematically overconfident: they routinely express high certainty on questions they often answer incorrectly. Existin

researcharxiv-cs-cl
14 Apr 2026
Research

Self-Correcting RAG: Enhancing Faithfulness via MMKP Context Selection and NLI-Guided MCTS

DGX agent

arXiv:2604.10734v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) substantially extends the knowledge boundary of large language models. However, it still faces two major challenges

researcharxiv-cs-cl
14 Apr 2026
Model Releases

Self-Evolving LLM Memory Extraction Across Heterogeneous Tasks

DGX agent

arXiv:2604.11610v1 Announce Type: new Abstract: As LLM-based assistants become persistent and personalized, they must extract and retain useful information from past conversations as memory. However,

model-releasesarxiv-cs-cl
14 Apr 2026
Research

Semantic-Space Exploration and Exploitation in RLVR for LLM Reasoning

DGX agent

arXiv:2509.23808v4 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) for LLM reasoning is often framed as balancing exploration and exploitation in action sp

researcharxiv-cs-cl
14 Apr 2026
Research

SEPTQ: A Simple and Effective Post-Training Quantization Paradigm for Large Language Models

DGX agent

arXiv:2604.10091v1 Announce Type: new Abstract: Large language models (LLMs) have shown remarkable performance in various domains, but they are constrained by massive computational and storage costs.

researcharxiv-cs-cl
14 Apr 2026
Model Releases

SHARE: Social-Humanities AI for Research and Education

DGX agent

arXiv:2604.11152v1 Announce Type: new Abstract: This intermediate technical report introduces the SHARE family of base models and the MIRROR user interface. The SHARE models are the first causal langu

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Sign Language Recognition in the Age of LLMs

DGX agent

arXiv:2604.11225v1 Announce Type: cross Abstract: Recent Vision Language Models (VLMs) have demonstrated strong performance across a wide range of multimodal reasoning tasks. This raises the question

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Simulating Organized Group Behavior: New Framework, Benchmark, and Analysis

DGX agent

arXiv:2604.09874v1 Announce Type: new Abstract: Simulating how organized groups (e.g., corporations) make decisions (e.g., responding to a competitor's move) is essential for understanding real-world

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Single-Agent LLMs Outperform Multi-Agent Systems on Multi-Hop Reasoning Under Equal Thinking Token Budgets

DGX agent

arXiv:2604.02460v2 Announce Type: replace Abstract: Recent work reports strong performance from multi-agent LLM systems (MAS), but these gains are often confounded by increased test-time computation.

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

SODA: Semi On-Policy Black-Box Distillation for Large Language Models

DGX agent

arXiv:2604.03873v2 Announce Type: replace-cross Abstract: Black-box knowledge distillation for large language models presents a strict trade-off. Simple off-policy methods (e.g., sequence-level knowle

model-releasesarxiv-cs-cl
14 Apr 2026
Research

Solver-Independent Automated Problem Formulation via LLMs for High-Cost Simulation-Driven Design

DGX agent

arXiv:2512.18682v2 Announce Type: replace Abstract: In the high-cost simulation-driven design domain, translating ambiguous design requirements into a mathematical optimization formulation is a bottle

researcharxiv-cs-cl
14 Apr 2026
Research

SpectralLoRA: Is Low-Frequency Structure Sufficient for LoRA Adaptation? A Spectral Analysis of Weight Updates

DGX agent

arXiv:2604.10649v1 Announce Type: cross Abstract: We present a systematic empirical study of the spectral structure of LoRA weight updates. Through 2D Discrete Cosine Transform (DCT) analysis of train

researcharxiv-cs-cl
14 Apr 2026
Research

SpeechLess: Micro-utterance with Personalized Spatial Memory-aware Assistant in Everyday Augmented Reality

DGX agent

arXiv:2602.00793v2 Announce Type: replace-cross Abstract: Speaking aloud to a wearable AR assistant in public can be socially awkward, and re-articulating the same requests every day creates unnecessa

researcharxiv-cs-cl
14 Apr 2026
Model Releases

Spoiler Alert: Narrative Forecasting as a Metric for Tension in LLM Storytelling

DGX agent

arXiv:2604.09854v1 Announce Type: new Abstract: LLMs have so far failed both to generate consistently compelling stories and to recognize this failure--on the leading creative-writing benchmark (EQ-Be

model-releasesarxiv-cs-cl
14 Apr 2026
Research

StableToken: A Noise-Robust Semantic Speech Tokenizer for Resilient SpeechLLMs

DGX agent

arXiv:2509.22220v2 Announce Type: replace Abstract: Prevalent semantic speech tokenizers, designed to capture linguistic content, are surprisingly fragile. We find they are not robust to meaning-irrel

researcharxiv-cs-cl
14 Apr 2026
Safety

Stop Fixating on Prompts: Reasoning Hijacking and Constraint Tightening for Red-Teaming LLM Agents

DGX agent

arXiv:2604.05549v2 Announce Type: replace Abstract: With the widespread application of LLM-based agents across various domains, their complexity has introduced new security threats. Existing red-team

safetyarxiv-cs-cl
14 Apr 2026
Agents

Structure-Grounded Knowledge Retrieval via Code Dependencies for Multi-Step Data Reasoning

DGX agent

arXiv:2604.10516v1 Announce Type: new Abstract: Selecting the right knowledge is critical when using large language models (LLMs) to solve domain-specific data analysis tasks. However, most retrieval-

agentsarxiv-cs-cl
14 Apr 2026
Safety

Structured Causal Video Reasoning via Multi-Objective Alignment

DGX agent

arXiv:2604.04415v2 Announce Type: replace Abstract: Human understanding of video dynamics is typically grounded in a structured mental representation of entities, actions, and temporal relations, rath

safetyarxiv-cs-cl
14 Apr 2026
Research

STU-PID: Steering Token Usage via PID Controller for Efficient Large Language Model Reasoning

DGX agent

arXiv:2506.18831v2 Announce Type: replace Abstract: Large Language Models employing extended chain-of-thought (CoT) reasoning often suffer from the overthinking phenomenon, generating excessive and re

researcharxiv-cs-cl
14 Apr 2026
Model Releases

Toward Generalized Cross-Lingual Hateful Language Detection with Web-Scale Data and Ensemble LLM Annotations

DGX agent

arXiv:2604.09625v1 Announce Type: new Abstract: We study whether large-scale unlabelled web data and LLM-based synthetic annotations can improve multilingual hate speech detection. Starting from texts

model-releasesarxiv-cs-cl
14 Apr 2026
Research

Towards Efficient Large Vision-Language Models: A Comprehensive Survey on Inference Strategies

DGX agent

arXiv:2603.27960v2 Announce Type: replace-cross Abstract: Although Large Vision Language Models (LVLMs) have demonstrated impressive multimodal reasoning capabilities, their scalability and deployment

researcharxiv-cs-cl
14 Apr 2026
Tutorials

TRACE: An Experiential Framework for Coherent Multi-hop Knowledge Graph Question Answering

DGX agent

arXiv:2604.11193v1 Announce Type: new Abstract: Multi-hop Knowledge Graph Question Answering (KGQA) requires coherent reasoning across relational paths, yet existing methods often treat each reasoning

tutorialsarxiv-cs-cl
14 Apr 2026
Safety

Training-Free Cross-Lingual Dysarthria Severity Assessment via Phonological Subspace Analysis in Self-Supervised Speech Representations

DGX agent

arXiv:2604.10123v1 Announce Type: new Abstract: Dysarthric speech severity assessment typically requires trained clinicians or supervised models built from labelled pathological speech, limiting scala

safetyarxiv-cs-cl
14 Apr 2026
Research

Transactional Attention: Semantic Sponsorship for KV-Cache Retention

DGX agent

arXiv:2604.11288v1 Announce Type: new Abstract: At K=16 tokens (0.4% of a 4K context), every existing KV-cache compression method achieves 0% on credential retrieval. The failure mode is dormant token

researcharxiv-cs-cl
14 Apr 2026
Safety

Triviality Corrected Endogenous Reward

DGX agent

arXiv:2604.11522v1 Announce Type: new Abstract: Reinforcement learning for open-ended text generation is constrained by the lack of verifiable rewards, necessitating reliance on judge models that requ

safetyarxiv-cs-cl
14 Apr 2026
Research

Turing or Cantor: That is the Question

DGX agent

arXiv:2604.10418v1 Announce Type: new Abstract: Alan Turing is considered as a founder of current computer science together with Kurt Godel, Alonzo Church and John von Neumann. In this paper multiple

researcharxiv-cs-cl
14 Apr 2026
Model Releases

Ultra-Low-Dimensional Prompt Tuning via Random Projection

DGX agent

arXiv:2502.04501v3 Announce Type: replace Abstract: Large language models achieve state-of-the-art performance but are increasingly costly to fine-tune. Prompt tuning is a parameter-efficient fine-tun

model-releasesarxiv-cs-cl
14 Apr 2026
Safety

Utilizing and Calibrating Hindsight Process Rewards via Reinforcement with Mutual Information Self-Evaluation

DGX agent

arXiv:2604.11611v1 Announce Type: new Abstract: To overcome the sparse reward challenge in reinforcement learning (RL) for agents based on large language models (LLMs), we propose Mutual Information S

safetyarxiv-cs-cl
14 Apr 2026
← Previous
1…152153154155156…160
Next →