AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,646 results
Research

Quantum Vision Theory Applied to Audio Classification for Deepfake Speech Detection

DGX agent

arXiv:2604.08104v1 Announce Type: new Abstract: We propose Quantum Vision (QV) theory as a new perspective for deep learning-based audio classification, applied to deepfake speech detection. Inspired

researcharxiv-cs-cl
10 Apr 2026
Research
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Rag Performance Prediction for Question Answering

DGX agent

arXiv:2604.07985v1 Announce Type: new Abstract: We address the task of predicting the gain of using RAG (retrieval augmented generation) for question answering with respect to not using it. We study t

researcharxiv-cs-cl
10 Apr 2026
Applications

Reasoning-Based Refinement of Unsupervised Text Clusters with LLMs

DGX agent

arXiv:2604.07562v1 Announce Type: new Abstract: Unsupervised methods are widely used to induce latent semantic structure from large text collections, yet their outputs often contain incoherent, redund

applicationsarxiv-cs-cl
10 Apr 2026
Agents

Reasoning Graphs: Deterministic Agent Accuracy through Evidence-Centric Chain-of-Thought Feedback

DGX agent

arXiv:2604.07595v1 Announce Type: cross Abstract: Language model agents reason from scratch on every query: each time an agent retrieves evidence and deliberates, the chain of thought is discarded and

agentsarxiv-cs-cl
10 Apr 2026
Safety

Reasoning Within the Mind: Dynamic Multimodal Interleaving in Latent Space

DGX agent

arXiv:2512.12623v3 Announce Type: replace-cross Abstract: Recent advancements in Multimodal Large Language Models (MLLMs) have significantly enhanced cross-modal understanding and reasoning by incorpo

safetyarxiv-cs-cl
10 Apr 2026
Research

ReCellTy: Domain-Specific Knowledge Graph Retrieval-Augmented LLMs Reasoning Workflow for Single-Cell Annotation

DGX agent

arXiv:2505.00017v2 Announce Type: replace Abstract: With the rapid development of large language models (LLMs), their application to cell type annotation has drawn increasing attention. However, gener

researcharxiv-cs-cl
10 Apr 2026
Safety

ReflectRM: Boosting Generative Reward Models via Self-Reflection within a Unified Judgment Framework

DGX agent

arXiv:2604.07506v1 Announce Type: cross Abstract: Reward Models (RMs) are critical components in the Reinforcement Learning from Human Feedback (RLHF) pipeline, directly determining the alignment qual

safetyarxiv-cs-cl
10 Apr 2026
Research

Rethinking Data Mixing from the Perspective of Large Language Models

DGX agent

arXiv:2604.07963v1 Announce Type: new Abstract: Data mixing strategy is essential for large language model (LLM) training. Empirical evidence shows that inappropriate strategies can significantly redu

researcharxiv-cs-cl
10 Apr 2026
Model Releases

Rethinking Entropy Allocation in LLM-based ASR: Understanding the Dynamics between Speech Encoders and LLMs

DGX agent

arXiv:2604.08003v1 Announce Type: cross Abstract: Integrating large language models (LLMs) into automatic speech recognition (ASR) has become a dominant paradigm. Although recent LLM-based ASR models

model-releasesarxiv-cs-cl
10 Apr 2026
Research

SAT: Balancing Reasoning Accuracy and Efficiency with Stepwise Adaptive Thinking

DGX agent

arXiv:2604.07922v1 Announce Type: cross Abstract: Large Reasoning Models (LRMs) have revolutionized complex problem-solving, yet they exhibit a pervasive 'overthinking', generating unnecessarily long

researcharxiv-cs-cl
10 Apr 2026
Model Releases

sciwrite-lint: Verification Infrastructure for the Age of Science Vibe-Writing

DGX agent

arXiv:2604.08501v1 Announce Type: cross Abstract: Science currently offers two options for quality assurance, both inadequate. Journal gatekeeping claims to verify both integrity and contribution, but

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

SealQA: Raising the Bar for Reasoning in Search-Augmented Language Models

DGX agent

arXiv:2506.01062v4 Announce Type: replace Abstract: We introduce SealQA, a new challenge benchmark for evaluating SEarch-Augmented Language models on fact-seeking questions where web search yields con

model-releasesarxiv-cs-cl
10 Apr 2026
Tutorials

Search-R3: Unifying Reasoning and Embedding in Large Language Models

DGX agent

arXiv:2510.07048v2 Announce Type: replace Abstract: Despite their remarkable natural language understanding capabilities, Large Language Models (LLMs) have been underutilized for retrieval tasks. We p

tutorialsarxiv-cs-cl
10 Apr 2026
Research

See the Forest for the Trees: Loosely Speculative Decoding via Visual-Semantic Guidance for Efficient Inference of Video LLMs

DGX agent

arXiv:2604.05650v2 Announce Type: replace Abstract: Video Large Language Models (Video-LLMs) excel in video understanding but suffer from high inference latency during autoregressive generation. Specu

researcharxiv-cs-cl
10 Apr 2026
Safety

Seeing but Not Thinking: Routing Distraction in Multimodal Mixture-of-Experts

DGX agent

arXiv:2604.08541v1 Announce Type: cross Abstract: Multimodal Mixture-of-Experts (MoE) models have achieved remarkable performance on vision-language tasks. However, we identify a puzzling phenomenon t

safetyarxiv-cs-cl
10 Apr 2026
Safety

Seeing Like an AI: How LLMs Apply (and Misapply) Wikipedia Neutrality Norms

DGX agent

arXiv:2407.04183v4 Announce Type: replace Abstract: Large language models (LLMs) are trained on broad corpora and then used in communities with specialized norms. Is providing LLMs with community rule

safetyarxiv-cs-cl
10 Apr 2026
Research

SeLaR: Selective Latent Reasoning in Large Language Models

DGX agent

arXiv:2604.08299v1 Announce Type: new Abstract: Chain-of-Thought (CoT) has become a cornerstone of reasoning in large language models, yet its effectiveness is constrained by the limited expressivenes

researcharxiv-cs-cl
10 Apr 2026
Safety

Self-Debias: Self-correcting for Debiasing Large Language Models

DGX agent

arXiv:2604.08243v1 Announce Type: new Abstract: Although Large Language Models (LLMs) demonstrate remarkable reasoning capabilities, inherent social biases often cascade throughout the Chain-of-Though

safetyarxiv-cs-cl
10 Apr 2026
Model Releases

Sell More, Play Less: Benchmarking LLM Realistic Selling Skill

DGX agent

arXiv:2604.07054v2 Announce Type: replace Abstract: Sales dialogues require multi-turn, goal-directed persuasion under asymmetric incentives, which makes them a challenging setting for large language

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Sensitivity-Positional Co-Localization in GQA Transformers

DGX agent

arXiv:2604.07766v1 Announce Type: new Abstract: We investigate a fundamental structural question in Grouped Query Attention (GQA) transformers: do the layers most sensitive to task correctness coincid

model-releasesarxiv-cs-cl
10 Apr 2026
Local Ai

SepSeq: A Training-Free Framework for Long Numerical Sequence Processing in LLMs

DGX agent

arXiv:2604.07737v1 Announce Type: new Abstract: While transformer-based Large Language Models (LLMs) theoretically support massive context windows, they suffer from severe performance degradation when

local-aiarxiv-cs-cl
10 Apr 2026
Agents

SkillClaw: Let Skills Evolve Collectively with Agentic Evolver

DGX agent

arXiv:2604.08377v1 Announce Type: cross Abstract: Large language model (LLM) agents such as OpenClaw rely on reusable skills to perform complex tasks, yet these skills remain largely static after depl

agentsarxiv-cs-cl
10 Apr 2026
Model Releases

Small Vision-Language Models are Smart Compressors for Long Video Understanding

DGX agent

arXiv:2604.08120v1 Announce Type: cross Abstract: Adapting Multimodal Large Language Models (MLLMs) for hour-long videos is bottlenecked by context limits. Dense visual streams saturate token budgets

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

SOLAR: Communication-Efficient Model Adaptation via Subspace-Oriented Latent Adapter Reparametrization

DGX agent

arXiv:2604.08368v1 Announce Type: cross Abstract: Parameter-efficient fine-tuning (PEFT) methods, such as LoRA, enable scalable adaptation of foundation models by injecting low-rank adapters. However,

model-releasesarxiv-cs-cl
10 Apr 2026
Research

Splits! Flexible Sociocultural Linguistic Investigation at Scale

DGX agent

arXiv:2504.04640v3 Announce Type: replace Abstract: Variation in language use, shaped by speakers' sociocultural background and specific context of use, offers a rich lens into cultural perspectives,

researcharxiv-cs-cl
10 Apr 2026
Research

Squeeze Evolve: Unified Multi-Model Orchestration for Verifier-Free Evolution

DGX agent

arXiv:2604.07725v1 Announce Type: cross Abstract: We show that verifier-free evolution is bottlenecked by both diversity and efficiency: without external correction, repeated evolution accelerates col

researcharxiv-cs-cl
10 Apr 2026
Model Releases

Stacked from One: Multi-Scale Self-Injection for Context Window Extension

DGX agent

arXiv:2603.04759v2 Announce Type: replace Abstract: The limited context window of contemporary large language models (LLMs) remains a primary bottleneck for their broader application across diverse do

model-releasesarxiv-cs-cl
10 Apr 2026
Agents

Stay Focused: Problem Drift in Multi-Agent Debate

DGX agent

arXiv:2502.19559v3 Announce Type: replace Abstract: Multi-agent debate - multiple instances of large language models discussing problems in turn-based interaction - has shown promise for solving knowl

agentsarxiv-cs-cl
10 Apr 2026
Applications

Stop Listening to Me! How Multi-turn Conversations Can Degrade LLM Diagnostic Reasoning

DGX agent

arXiv:2603.11394v2 Announce Type: replace Abstract: Patients and clinicians are increasingly using chatbots powered by large language models (LLMs) for healthcare inquiries. While state-of-the-art LLM

applicationsarxiv-cs-cl
10 Apr 2026
Agents

SubSearch: Intermediate Rewards for Unsupervised Guided Reasoning in Complex Retrieval

DGX agent

arXiv:2604.07415v1 Announce Type: cross Abstract: Large language models (LLMs) are probabilistic in nature and perform more reliably when augmented with external information. As complex queries often

agentsarxiv-cs-cl
10 Apr 2026
Model Releases

Symbiotic-MoE: Unlocking the Synergy between Generation and Understanding

DGX agent

arXiv:2604.07753v1 Announce Type: cross Abstract: Empowering Large Multimodal Models (LMMs) with image generation often leads to catastrophic forgetting in understanding tasks due to severe gradient c

model-releasesarxiv-cs-cl
10 Apr 2026
Safety

SYN-DIGITS: A Synthetic Control Framework for Calibrated Digital Twin Simulation

DGX agent

arXiv:2604.07513v1 Announce Type: cross Abstract: AI-based persona simulation -- often referred to as digital twin simulation -- is increasingly used for market research, recommender systems, and soci

safetyarxiv-cs-cl
10 Apr 2026
Safety

Synthetic Data for any Differentiable Target

DGX agent

arXiv:2604.08423v1 Announce Type: new Abstract: What are the limits of controlling language models via synthetic training data? We develop a reinforcement learning (RL) primitive, the Dataset Policy G

safetyarxiv-cs-cl
10 Apr 2026
Tutorials

TEC: A Collection of Human Trial-and-error Trajectories for Problem Solving

DGX agent

arXiv:2604.06734v2 Announce Type: replace Abstract: Trial-and-error is a fundamental strategy for humans to solve complex problems and a necessary capability for Artificial Intelligence (AI) systems o

tutorialsarxiv-cs-cl
10 Apr 2026
Model Releases

TEMPER: Testing Emotional Perturbation in Quantitative Reasoning

DGX agent

arXiv:2604.07801v1 Announce Type: new Abstract: Large language models are trained and evaluated on quantitative reasoning tasks written in clean, emotionally neutral language. However, real-world quer

model-releasesarxiv-cs-cl
10 Apr 2026
Research

Testimole-Conversational: A 30-Billion-Word Italian Discussion Board Corpus (1996-2024) for Language Modeling and Sociolinguistic Research

DGX agent

arXiv:2602.14819v2 Announce Type: replace Abstract: We present 'Testimole-conversational' a massive collection of discussion boards messages in the Italian language. The large size of the corpus, more

researcharxiv-cs-cl
10 Apr 2026
Applications

exttt{SEM-CTRL}: Semantically Controlled Decoding

DGX agent

arXiv:2503.01804v4 Announce Type: replace Abstract: Ensuring both syntactic and semantic correctness in Large Language Model (LLM) outputs remains a significant challenge, despite being critical for r

applicationsarxiv-cs-cl
10 Apr 2026
Safety

The Art of (Mis)alignment: How Fine-Tuning Methods Effectively Misalign and Realign LLMs in Post-Training

DGX agent

arXiv:2604.07754v1 Announce Type: cross Abstract: The deployment of large language models (LLMs) raises significant ethical and safety concerns. While LLM alignment techniques are adopted to improve m

safetyarxiv-cs-cl
10 Apr 2026
Model Releases

Tool Retrieval Bridge: Aligning Vague Instructions with Retriever Preferences via Bridge Model

DGX agent

arXiv:2604.07816v1 Announce Type: new Abstract: Tool learning has emerged as a promising paradigm for large language models (LLMs) to address real-world challenges. Due to the extensive and irregularl

model-releasesarxiv-cs-cl
10 Apr 2026
Agents

TOOLCAD: Exploring Tool-Using Large Language Models in Text-to-CAD Generation with Reinforcement Learning

DGX agent

arXiv:2604.07960v1 Announce Type: cross Abstract: Computer-Aided Design (CAD) is an expert-level task that relies on long-horizon reasoning and coherent modeling actions. Large Language Models (LLMs)

agentsarxiv-cs-cl
10 Apr 2026
Research

Towards Hierarchical Multi-Step Reward Models for Enhanced Reasoning in Large Language Models

DGX agent

arXiv:2503.13551v5 Announce Type: replace Abstract: Recent studies show that Large Language Models (LLMs) achieve strong reasoning capabilities through supervised fine-tuning or reinforcement learning

researcharxiv-cs-cl
10 Apr 2026
Model Releases

Towards Real-world Human Behavior Simulation: Benchmarking Large Language Models on Long-horizon, Cross-scenario, Heterogeneous Behavior Traces

DGX agent

arXiv:2604.08362v1 Announce Type: new Abstract: The emergence of Large Language Models (LLMs) has illuminated the potential for a general-purpose user simulator. However, existing benchmarks remain co

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

TR-EduVSum: A Turkish-Focused Dataset and Consensus Framework for Educational Video Summarization

DGX agent

arXiv:2604.07553v1 Announce Type: new Abstract: This study presents a framework for generating the gold-standard summary fully automatically and reproducibly based on multiple human summaries of Turki

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Training Data Size Sensitivity in Unsupervised Rhyme Recognition

DGX agent

arXiv:2604.08156v1 Announce Type: new Abstract: Rhyme is deceptively intuitive: what is or is not a rhyme is constructed historically, scholars struggle with rhyme classification, and people disagree

model-releasesarxiv-cs-cl
10 Apr 2026
Tutorials

Transforming the Voice of the Customer: Large Language Models for Identifying Customer Needs

DGX agent

arXiv:2503.01870v2 Announce Type: replace Abstract: Identifying customer needs (CNs) is fundamental to product innovation and marketing strategy. Yet for over thirty years, Voice-of-the-Customer (VOC)

tutorialsarxiv-cs-cl
10 Apr 2026
Model Releases

TSUBASA: Improving Long-Horizon Personalization via Evolving Memory and Self-Learning with Context Distillation

DGX agent

arXiv:2604.07894v1 Announce Type: new Abstract: Personalized large language models (PLLMs) have garnered significant attention for their ability to align outputs with individual's needs and preference

model-releasesarxiv-cs-cl
10 Apr 2026
Applications

Understanding Structured Financial Data with LLMs: A Case Study on Fraud Detection

DGX agent

arXiv:2512.13040v2 Announce Type: replace-cross Abstract: Detecting fraud in financial transactions typically relies on tabular models that demand heavy feature engineering to handle high-dimensional

applicationsarxiv-cs-cl
10 Apr 2026
Model Releases

Verify Before You Commit: Towards Faithful Reasoning in LLM Agents via Self-Auditing

DGX agent

arXiv:2604.08401v1 Announce Type: cross Abstract: In large language model (LLM) agents, reasoning trajectories are treated as reliable internal beliefs for guiding actions and updating memory. However

model-releasesarxiv-cs-cl
10 Apr 2026
← Previous
1…157158159160
Next →