AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlog
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Model Releases

SongBench: A Fine-Grained Multi-Aspect Benchmark for Song Quality Assessment

DGX agent

arXiv:2604.25937v1 Announce Type: cross Abstract: Recent advancements in Text-to-Song generation have enabled realistic musical content production, yet existing evaluation benchmarks lack the professi

model-releasesarxiv-cs-ai
30 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Tutorials

Speech Emotion Recognition Using MFCC Features and LSTM-Based Deep Learning Model

DGX agent

arXiv:2604.25938v1 Announce Type: cross Abstract: Speech Emotion Recognition (SER) is the use of machines to detect the emotional state of humans based on the speech, which is gaining importance in na

tutorialsarxiv-cs-ai
30 Apr 2026
Agents

Star-Fusion: A Multi-modal Transformer Architecture for Discrete Celestial Orientation via Spherical Topology

DGX agent

arXiv:2604.26582v1 Announce Type: cross Abstract: Reliable celestial attitude determination is a critical requirement for autonomous spacecraft navigation, yet traditional 'Lost-in-Space' (LIS) algori

agentsarxiv-cs-ai
30 Apr 2026
Applications

STLGT: A Scalable Trace-Based Linear Graph Transformer for Tail Latency Prediction in Microservices

DGX agent

arXiv:2604.26422v1 Announce Type: cross Abstract: Accurate end-to-end tail-latency forecasting is critical for proactive SLO management in microservice systems. However, modeling long-range dependency

applicationsarxiv-cs-ai
30 Apr 2026
Model Releases

StratMem-Bench: Evaluating Strategic Memory Use in Virtual Character Conversation Beyond Factual Recall

DGX agent

arXiv:2604.26243v1 Announce Type: cross Abstract: Achieving realistic human-like conversation for virtual characters requires not only a simple memorization and recall of past events, but also the str

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Stress Testing Factual Consistency Metrics for Long-Document Summarization

DGX agent

arXiv:2511.07689v2 Announce Type: replace-cross Abstract: Evaluating the factual consistency of abstractive text summarization remains a significant challenge, particularly for long documents, where c

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Structural Generalization on SLOG without Hand-Written Rules

DGX agent

arXiv:2604.26157v1 Announce Type: cross Abstract: Structural generalization in semantic parsing requires systems to apply learned compositional rules to novel structural combinations. Existing approac

model-releasesarxiv-cs-ai
30 Apr 2026
Safety

Student Guides Teacher: Weak-to-Strong Inference via Spectral Orthogonal Exploration

DGX agent

arXiv:2601.06160v2 Announce Type: replace Abstract: Large Language Models (LLMs) often suffer from ''Reasoning Collapse'' on challenging mathematical reasoning tasks, where stochastic sampling produce

safetyarxiv-cs-ai
30 Apr 2026
Research

SynSur: An end-to-end generative pipeline for synthetic industrial surface defect generation and detection

DGX agent

arXiv:2604.26633v1 Announce Type: cross Abstract: The bottleneck in learning-based industrial defect detection is often limited not by model capacity, but by the scarcity of labeled defect data: defec

researcharxiv-cs-ai
30 Apr 2026
Safety

Tatemae: Detecting Alignment Faking via Tool Selection in LLMs

DGX agent

arXiv:2604.26511v1 Announce Type: cross Abstract: Alignment faking (AF) occurs when an LLM strategically complies with training objectives to avoid value modification, reverting to prior preferences o

safetyarxiv-cs-ai
30 Apr 2026
Agents

TDD Governance for Multi-Agent Code Generation via Prompt Engineering

DGX agent

arXiv:2604.26615v1 Announce Type: cross Abstract: Large language models (LLMs) accelerate software development but often exhibit instability, non-determinism, and weak adherence to development discipl

agentsarxiv-cs-ai
30 Apr 2026
Safety

Test-Time Safety Alignment

DGX agent

arXiv:2604.26167v1 Announce Type: cross Abstract: Recent work has shown that a model's input word embeddings can serve as effective control variables for steering its behavior toward outputs that sati

safetyarxiv-cs-ai
30 Apr 2026
Safety

Text Style Transfer with Machine Translation for Graphic Designs

DGX agent

arXiv:2604.26361v1 Announce Type: cross Abstract: Globalization of graphic designs such as those used in marketing materials and magazines is increasingly important for communication to broad audience

safetyarxiv-cs-ai
30 Apr 2026
Research

Text-Utilization for Encoder-dominated Speech Recognition Models

DGX agent

arXiv:2604.26514v1 Announce Type: cross Abstract: This paper investigates efficient methods for utilizing text-only data to improve speech recognition, focusing on encoder-dominated models that facili

researcharxiv-cs-ai
30 Apr 2026
Research

The Dual Role of Abstracting over the Irrelevant in Symbolic Explanations: Cognitive Effort vs. Understanding

DGX agent

arXiv:2602.03467v2 Announce Type: replace Abstract: Explanations are central to human cognition, yet AI systems often produce outputs that are difficult to understand. While symbolic AI offers a trans

researcharxiv-cs-ai
30 Apr 2026
Research

The Fools are Certain; the Wise are Doubtful: Exploring LLM Confidence in Code Completion

DGX agent

arXiv:2508.16131v2 Announce Type: replace-cross Abstract: Code completion entails the task of providing missing tokens given a surrounding context. It can boost developer productivity while providing

researcharxiv-cs-ai
30 Apr 2026
Model Releases

TildeOpen LLM: Leveraging Curriculum Learning to Achieve Equitable Language Representation

DGX agent

arXiv:2603.08182v2 Announce Type: replace-cross Abstract: Large language models often underperform in many European languages due to the dominance of English and a few high-resource languages in train

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Time Blindness: Why Video-Language Models Can't See What Humans Can?

DGX agent

arXiv:2505.24867v2 Announce Type: replace-cross Abstract: Recent advances in vision-language models (VLMs) have made impressive strides in understanding spatio-temporal relationships in videos. Howeve

model-releasesarxiv-cs-ai
30 Apr 2026
Applications

TimeMM: Time-as-Operator Spectral Filtering for Dynamic Multimodal Recommendation

DGX agent

arXiv:2604.26247v1 Announce Type: cross Abstract: Multimodal recommendation improves user modeling by integrating collaborative signals with heterogeneous item content. In real applications, user inte

applicationsarxiv-cs-ai
30 Apr 2026
Model Releases

TinyR1-32B-Preview: Boosting Accuracy with Branch-Merge Distillation

DGX agent

arXiv:2503.04872v3 Announce Type: replace-cross Abstract: The challenge of reducing the size of Large Language Models (LLMs) while maintaining their performance has gained significant attention. Howev

model-releasesarxiv-cs-ai
30 Apr 2026
Local Ai

TLPO: Token-Level Policy Optimization for Mitigating Language Confusion in Large Language Models

DGX agent

arXiv:2604.26553v1 Announce Type: cross Abstract: Large language models (LLMs) demonstrate strong multilingual capabilities, yet often fail to consistently generate responses in the intended language,

local-aiarxiv-cs-ai
30 Apr 2026
Research

ToolPRM: Fine-Grained Inference Scaling of Structured Outputs for Function Calling

DGX agent

arXiv:2510.14703v2 Announce Type: replace Abstract: Large language models (LLMs) excel at function calling, but inference scaling has been explored mainly for unstructured generation. We propose an in

researcharxiv-cs-ai
30 Apr 2026
Agents

Training Computer Use Agents to Assess the Usability of Graphical User Interfaces

DGX agent

arXiv:2604.26020v1 Announce Type: cross Abstract: Usability testing with experts and potential users can assess the effectiveness, efficiency, and user satisfaction of graphical user interfaces (GUIs)

agentsarxiv-cs-ai
30 Apr 2026
Model Releases

Training-Free Adaptation of New-Generation LLMs using Legacy Clinical Models

DGX agent

arXiv:2601.03423v3 Announce Type: replace-cross Abstract: Adapting language models to the clinical domain through continued pretraining and instruction tuning requires costly retraining for each new m

model-releasesarxiv-cs-ai
30 Apr 2026
Safety

Translating Under Pressure: Domain-Aware LLMs for Crisis Communication

DGX agent

arXiv:2604.26597v1 Announce Type: cross Abstract: Timely and reliable multilingual communication is critical during natural and human-induced disasters, but developing effective solutions for crisis c

safetyarxiv-cs-ai
30 Apr 2026
Research

Tree-of-Text: A Tree-based Prompting Framework for Table-to-Text Generation in the Sports Domain

DGX agent

arXiv:2604.26501v1 Announce Type: cross Abstract: Generating sports game reports from structured tables is a complex table-to-text task that demands both precise data interpretation and fluent narrati

researcharxiv-cs-ai
30 Apr 2026
Research

Turning the TIDE: Cross-Architecture Distillation for Diffusion Large Language Models

DGX agent

arXiv:2604.26951v1 Announce Type: cross Abstract: Diffusion large language models (dLLMs) offer parallel decoding and bidirectional context, but state-of-the-art dLLMs require billions of parameters f

researcharxiv-cs-ai
30 Apr 2026
Safety

Uncertainty-Aware Reward Discounting for Mitigating Reward Hacking

DGX agent

arXiv:2604.26360v1 Announce Type: cross Abstract: Reinforcement learning (RL) systems typically optimize scalar reward functions that assume precise and reliable evaluation of outcomes. However, real-

safetyarxiv-cs-ai
30 Apr 2026
Research

Unified 4D World Action Modeling from Video Priors with Asynchronous Denoising

DGX agent

arXiv:2604.26694v1 Announce Type: cross Abstract: We propose X-WAM, a Unified 4D World Model that unifies real-time robotic action execution and high-fidelity 4D world synthesis (video + 3D reconstruc

researcharxiv-cs-ai
30 Apr 2026
Model Releases

Value-Guided Iterative Refinement and the DIQ-H Benchmark for Evaluating VLM Robustness

DGX agent

arXiv:2512.03992v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) are essential for embodied AI and safety-critical applications, such as robotics and autonomous systems. However

model-releasesarxiv-cs-ai
30 Apr 2026
Research

Vertex Features for Neural Global Illumination

DGX agent

arXiv:2508.07852v2 Announce Type: replace-cross Abstract: Recent research on learnable neural representations has been widely adopted in the field of 3D scene reconstruction and neural rendering appli

researcharxiv-cs-ai
30 Apr 2026
Safety

Vibe Check: Understanding the Effects of LLM-Based Conversational Agents' Personality and Alignment on User Perceptions in Goal-Oriented Tasks

DGX agent

arXiv:2509.09870v2 Announce Type: replace-cross Abstract: Large language models (LLMs) enable conversational agents (CAs) to express distinctive personalities, raising new questions about how such des

safetyarxiv-cs-ai
30 Apr 2026
Local Ai

ViCrop-Det: Spatial Attention Entropy Guided Cropping for Training-Free Small-Object Detection

DGX agent

arXiv:2604.26806v1 Announce Type: cross Abstract: Transformer-based architectures have established a dominant paradigm in global semantic perception; however, they remain fundamentally constrained by

local-aiarxiv-cs-ai
30 Apr 2026
Model Releases

When to Retrieve During Reasoning: Adaptive Retrieval for Large Reasoning Models

DGX agent

arXiv:2604.26649v1 Announce Type: cross Abstract: Large reasoning models such as DeepSeek-R1 and OpenAI o1 generate extended chains of thought spanning thousands of tokens, yet their integration with

model-releasesarxiv-cs-ai
30 Apr 2026
Research

When to Vote, When to Rewrite: Disagreement-Guided Strategy Routing for Test-Time Scaling

DGX agent

arXiv:2604.26644v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) achieve strong performance on mathematical reasoning tasks but remain unreliable on challenging instances. Existing test-t

researcharxiv-cs-ai
30 Apr 2026
Research

Why Attend to Everything? Focus is the Key

DGX agent

arXiv:2604.03260v2 Announce Type: replace-cross Abstract: Standard attention scales quadratically with sequence length. Efficient attention methods reduce this O(n^2) cost, but when retrofitted into p

researcharxiv-cs-ai
30 Apr 2026
Safety

A Decoupled Human-in-the-Loop System for Controlled Autonomy in Agentic Workflows

DGX agent

arXiv:2604.23049v1 Announce Type: new Abstract: AI agents are increasingly deployed to execute tasks and make decisions within agentic workflows, introducing new requirements for safe and controlled a

safetyarxiv-cs-ai
28 Apr 2026
Model Releases

A Fano-Style Accuracy Upper Bound for LLM Single-Pass Reasoning in Multi-Hop QA

DGX agent

arXiv:2509.21199v3 Announce Type: replace Abstract: Multi-Hop Question Answering (MHQA) requires integrating dispersed, interdependent evidence through sequential reasoning under noise. This task is c

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

A General Framework for Generative Self-supervised Learning in Non-invasive Estimation of Physiological Parameters Using Photoplethysmography

DGX agent

arXiv:2604.22780v1 Announce Type: cross Abstract: Aligning physiological parameter labels with large-scale photoplethysmographic (PPG) data for deep learning is challenging and resource-intensive. Whi

model-releasesarxiv-cs-ai
28 Apr 2026
Safety

A Lightweight Explainable Guardrail for Prompt Safety

DGX agent

arXiv:2602.15853v2 Announce Type: replace-cross Abstract: We propose a lightweight explainable guardrail (LEG) method to detect unsafe prompts. LEG uses a multi-task learning architecture to jointly l

safetyarxiv-cs-ai
28 Apr 2026
Research

A Lower Bound for the Number of Linear Regions of Ternary ReLU Regression Neural Networks

DGX agent

arXiv:2507.16079v2 Announce Type: replace-cross Abstract: With the advancement of deep learning, reducing computational complexity and memory consumption has become a critical challenge, and ternary n

researcharxiv-cs-ai
28 Apr 2026
Research

A Milestone in Formalization: The Sphere Packing Problem in Dimension 8

DGX agent

arXiv:2604.23468v1 Announce Type: cross Abstract: In 2016, Viazovska famously solved the sphere packing problem in dimension 8, using modular forms to construct a 'magic' function satisfying optimalit

researcharxiv-cs-ai
28 Apr 2026
Model Releases

A Parametric Memory Head for Continual Generative Retrieval

DGX agent

arXiv:2604.23388v1 Announce Type: cross Abstract: Generative information retrieval (GenIR) consolidates retrieval into a single neural model that decodes document identifiers (docids) directly from qu

model-releasesarxiv-cs-ai
28 Apr 2026
Safety

A Self-Supervised Framework for Space Object Behaviour Characterisation

DGX agent

arXiv:2504.06176v3 Announce Type: replace-cross Abstract: Foundation Models, which leverage large neural networks pre-trained on unlabelled data before fine-tuning for specific tasks, are increasingly

safetyarxiv-cs-ai
28 Apr 2026
Agents

A Systematic Approach for Large Language Models Debugging

DGX agent

arXiv:2604.23027v1 Announce Type: new Abstract: Large language models (LLMs) have become central to modern AI workflows, powering applications from open-ended text generation to complex agent-based re

agentsarxiv-cs-ai
28 Apr 2026
Model Releases

A systematic evaluation of vision-language models for observational astronomical reasoning tasks

DGX agent

arXiv:2604.24589v1 Announce Type: new Abstract: Vision-language models (VLMs) are increasingly proposed as general-purpose tools for scientific data interpretation, yet their reliability on real astro

model-releasesarxiv-cs-ai
28 Apr 2026
Safety

A Taxonomy and Resolution Strategy for Client-Level Disagreements in Federated Learning

DGX agent

arXiv:2604.23386v1 Announce Type: cross Abstract: Federated Learning (FL) typically assumes unconditional collaboration, a premise that overlooks the complexities of real-world, multi-stakeholder envi

safetyarxiv-cs-ai
28 Apr 2026
Model Releases

A Theoretical Framework for Auxiliary-Loss-Free Load Balancing of Sparse Mixture-of-Experts in Large-Scale AI Models

DGX agent

arXiv:2512.03915v3 Announce Type: replace-cross Abstract: In large-scale AI training, Sparse Mixture-of-Experts (s-MoE) layers enable scaling by activating only a small subset of experts per token. An

model-releasesarxiv-cs-ai
28 Apr 2026
← Previous
1…373374375376377…448
Next →