AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
Human
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,851 results
28 Apr 2026

Speech Enhancement Based on Drifting Models

Model ReleasesDGX agent

arXiv:2604.24199v1 Announce Type: cross Abstract: We propose Speech Enhancement based on Drifting Models (DriftSE), a novel generative framework that formulates denoising as an equilibrium problem. Ra

Speech-FT: Merging Pre-trained And Fine-Tuned Speech Representation Models For Cross-Task Generalization

Model ReleasesDGX agent

arXiv:2502.12672v4 Announce Type: replace-cross Abstract: Fine-tuning speech representation models can enhance performance on specific tasks but often compromises their cross-task generalization abili

Sphere-Depth: A Benchmark for Depth Estimation Methods with Varying Spherical Camera Orientations

Model ReleasesDGX agent

arXiv:2604.23432v1 Announce Type: cross Abstract: Reliable depth estimation from spherical images is crucial for 360{eg} vision in robotic navigation and immersive scene understanding. However, the on

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

SPLIT: Separating Physical-Contact via Latent Arithmetic in Image-Based Tactile Sensors

ResearchDGX agent

arXiv:2604.24449v1 Announce Type: cross Abstract: Training machine learning models for robotic tactile sensing requires vast amounts of data, yet obtaining realistic interaction data remains a challen

SRL-CLIP: Efficient CLIP Video Adaptation via Structured Semantic Role Labels

TutorialsDGX agent

arXiv:2401.07669v2 Announce Type: replace Abstract: Adapting CLIP for videos has gained popularity due to its semantic and rich representation. While CLIP is a good starting point, it typically underg

SSR-Zero: Simple Self-Rewarding Reinforcement Learning for Machine Translation

Model ReleasesDGX agent

arXiv:2505.16637v4 Announce Type: replace-cross Abstract: Large language models (LLMs) have recently demonstrated remarkable capabilities in machine translation (MT). However, most advanced MT-specifi

Stabilizing Efficient Reasoning with Step-Level Advantage Selection

Model ReleasesDGX agent

arXiv:2604.24003v1 Announce Type: new Abstract: Large language models (LLMs) achieve strong reasoning performance by allocating substantial computation at inference time, often generating long and ver

StackFeat RL: Reinforcement Learning over Iterative Dual Criterion Feature Selection for Stable Biomarker Discovery

SafetyDGX agent

arXiv:2604.22892v1 Announce Type: new Abstract: Feature selection in high-dimensional genomic data (d gg n) demands methods that are simultaneously accurate, sparse, and stable. Existing approaches ei

STAND: Semantic Anchoring Constraint with Dual-Granularity Disambiguation for Remote Sensing Image Change Captioning

ResearchDGX agent

arXiv:2604.23309v1 Announce Type: new Abstract: Remote sensing image change captioning (RSICC) aims to describe the difference between two remote sensing images. While recent methods have explored vid

Statistical Test for Diffusion-Based Anomaly Localization via Selective Inference

Local AiDGX agent

arXiv:2402.11789v5 Announce Type: replace-cross Abstract: Anomaly localization in images -- identifying regions that deviate from normal patterns -- is vital in applications such as medical diagnosis

Statistically-Guided Meta-Learning for Cross-Deployment Activity Recognition in Distributed Fiber-Optic Sensing

ApplicationsDGX agent

arXiv:2511.17902v3 Announce Type: replace-cross Abstract: Distributed Fiber Optic Sensing (DFOS) is promising for long-range perimeter security, yet practical deployment faces three key obstacles: sev

STELLAR-E: a Synthetic, Tailored, End-to-end LLM Application Rigorous Evaluator

ResearchDGX agent

arXiv:2604.24544v1 Announce Type: new Abstract: The increasing reliance on Large Language Models (LLMs) across diverse sectors highlights the need for robust domain-specific and language-specific eval

StereoSpace: Depth-Free Synthesis of Stereo Geometry via End-to-End Diffusion in a Canonical Space

TutorialsDGX agent

arXiv:2512.10959v2 Announce Type: replace Abstract: We introduce StereoSpace, a diffusion-based framework for monocular-to-stereo synthesis that models geometry purely through viewpoint conditioning,

Stochastic KV Routing: Enabling Adaptive Depth-Wise Cache Sharing

ResearchDGX agent

arXiv:2604.22782v1 Announce Type: cross Abstract: Serving transformer language models with high throughput requires caching Key-Values (KVs) to avoid redundant computation during autoregressive genera

Stochastic simultaneous optimistic optimization

ResearchDGX agent

arXiv:2604.24537v1 Announce Type: new Abstract: We study the problem of global maximization of a function f given a finite number of evaluations perturbed by noise. We consider a very weak assumption

StoryTR: Narrative-Centric Video Temporal Retrieval with Theory of Mind Reasoning

Model ReleasesDGX agent

arXiv:2604.23198v1 Announce Type: new Abstract: Current video moment retrieval excels at action-centric tasks but struggles with narrative content. Models can see extit{what is happening} but fail to

Strategic Bidding in 6G Spectrum Auctions with Large Language Models

Model ReleasesDGX agent

arXiv:2604.24156v1 Announce Type: cross Abstract: Efficient and fair spectrum allocation is a central challenge in 6G networks, where massive connectivity and heterogeneous services continuously compe

StratRAG: A Multi-Hop Retrieval Evaluation Dataset for Retrieval-Augmented Generation Systems

Model ReleasesDGX agent

arXiv:2604.22757v1 Announce Type: cross Abstract: We introduce StratRAG, an open-source retrieval evaluation dataset for benchmarking Retrieval-Augmented Generation (RAG) systems on multi-hop reasonin

Stress-Testing Emotional Support Models: Moving from Homogeneous to Diverse Help Seekers

Model ReleasesDGX agent

arXiv:2601.07698v2 Announce Type: replace Abstract: As emotional support chatbots have recently gained significant traction across both research and industry, a common evaluation strategy has emerged:

StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval

Model ReleasesDGX agent

arXiv:2601.20597v2 Announce Type: replace Abstract: Continual Text-to-Video Retrieval (CTVR) is a challenging multimodal continual learning setting, where models must incrementally learn new semantic

Structural Enforcement of Goal Integrity in AI Agents via Separation-of-Powers Architecture

SafetyDGX agent

arXiv:2604.23646v1 Announce Type: new Abstract: Recent evidence suggests that frontier AI systems can exhibit agentic misalignment, generating and executing harmful actions derived from internally con

Structural Pruning of Large Vision Language Models: A Comprehensive Study on Pruning Dynamics, Recovery, and Data Efficiency

Model ReleasesDGX agent

arXiv:2604.24380v1 Announce Type: new Abstract: While Large Vision Language Models (LVLMs) demonstrate impressive capabilities, their substantial computational and memory requirements pose deployment

Structure Guided Retrieval-Augmented Generation for Factual Queries

TutorialsDGX agent

arXiv:2604.22843v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) has been proposed to mitigate hallucinations in large language models (LLMs), where generated outputs may be fact

Supernodes and Halos: Loss-Critical Hubs in LLM Feed-Forward Layers

Model ReleasesDGX agent

arXiv:2604.23475v1 Announce Type: cross Abstract: We study the organization of channel-level importance in transformer feed-forward networks (FFNs). Using a Fisher-style loss proxy (LP) based on activ

Supporting Family-School Partnerships with Robot-Facilitated Home-Based Activities

ResearchDGX agent

arXiv:2604.23978v1 Announce Type: new Abstract: Family-school partnerships (FSP) are critical to children's development, yet families often face barriers such as time constraints, fragmented communica

Surface Sensitivity in Lean 4 Autoformalization

ResearchDGX agent

arXiv:2604.23135v1 Announce Type: new Abstract: Natural-language variation poses a key challenge in Lean autoformalization: semantically equivalent paraphrases of the same theorem statements can induc

Survey in Characterizing Semantic Change

TutorialsDGX agent

arXiv:2402.19088v5 Announce Type: replace-cross Abstract: Live languages continuously evolve to integrate the cultural change of human societies. This evolution manifests through neologisms (new words

Swa-bhasha Resource Hub: Romanized Sinhala to Sinhala Transliteration Systems and Data Resources

ResearchDGX agent

arXiv:2507.09245v2 Announce Type: replace Abstract: The Swa-bhasha Resource Hub provides a comprehensive collection of data resources and algorithms developed for Romanized Sinhala to Sinhala translit

SwarmDrive: Semantic V2V Coordination for Latency-Constrained Cooperative Autonomous Driving

AgentsDGX agent

arXiv:2604.22852v1 Announce Type: cross Abstract: Cloud-hosted LLM inference for autonomous driving adds round-trip delay and depends on stable connectivity, while purely local edge models struggle un

SWE-Pruner: Self-Adaptive Context Pruning for Coding Agents

AgentsDGX agent

arXiv:2601.16746v3 Announce Type: replace-cross Abstract: LLM agents have demonstrated remarkable capabilities in software development, but their performance is hampered by long interaction contexts,

SWE-QA: Can Language Models Answer Repository-level Code Questions?

Model ReleasesDGX agent

arXiv:2509.14635v2 Announce Type: replace Abstract: Understanding and reasoning about entire software repositories is an essential capability for intelligent software engineering tools. While existing

SwissGov-RSD: A Human-annotated, Cross-lingual Benchmark for Token-level Recognition of Semantic Differences Between Related Documents

Model ReleasesDGX agent

arXiv:2512.07538v3 Announce Type: replace Abstract: Recognizing semantic differences across documents is crucial for text generation evaluation and content alignment, especially in cross-lingual setti

Switch Attention: Towards Dynamic and Fine-grained Hybrid Transformers

Model ReleasesDGX agent

arXiv:2603.26380v2 Announce Type: replace Abstract: The attention mechanism has been the core component in modern transformer architectures. However, the computation of standard full attention scales

SycoPhantasy: Quantifying Sycophancy and Hallucination in Small Open Weight VLMs for Vision-Language Scoring of Fantasy Characters

Model ReleasesDGX agent

arXiv:2604.24346v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly deployed as evaluators in tasks requiring nuanced image understanding, yet their reliability in scoring

Symbolic recovery of PDEs from measurement data

ResearchDGX agent

arXiv:2602.15603v2 Announce Type: replace Abstract: Models based on partial differential equations (PDEs) are powerful for describing a wide range of complex phenomena in the natural sciences. Accurat

Symmetric Equilibrium Propagation for Thermodynamic Diffusion Training

Model ReleasesDGX agent

arXiv:2604.23806v1 Announce Type: cross Abstract: The reverse process in score-based diffusion models is formally equivalent to overdamped Langevin dynamics in a time-dependent energy landscape. In ou

Synthetic Eggs in Many Baskets: The Impact of Synthetic Data Diversity on LLM Fine-Tuning

SafetyDGX agent

arXiv:2511.01490v2 Announce Type: replace Abstract: As synthetic data becomes widely used in language model development, understanding its impact on model behavior is crucial. This paper investigates

SynthPert: Enhancing LLM Biological Reasoning via Synthetic Reasoning Traces for Cellular Perturbation Prediction

Model ReleasesDGX agent

arXiv:2509.25346v2 Announce Type: replace Abstract: Predicting cellular responses to genetic perturbations represents a fundamental challenge in systems biology, critical for advancing therapeutic dis

TACO: Efficient Communication Compression of Intermediate Tensors for Scalable Tensor-Parallel LLM Training

Model ReleasesDGX agent

arXiv:2604.24088v1 Announce Type: cross Abstract: Handling communication overhead in large-scale tensor-parallel training remains a critical challenge due to the dense, near-zero distributions of inte

Talker-T2AV: Joint Talking Audio-Video Generation with Autoregressive Diffusion Modeling

ResearchDGX agent

arXiv:2604.23586v1 Announce Type: cross Abstract: Joint audio-video generation models have shown that unified generation yields stronger cross-modal coherence than cascaded approaches. However, existi

Talking Slide Avatars: Open-Source Multimodal Communication Approach for Teaching

ApplicationsDGX agent

arXiv:2604.23703v1 Announce Type: cross Abstract: Slide-based teaching is widely used in higher education, yet in online, hybrid, and asynchronous contexts, slides often lose the instructor presence,

Tandem: Riding Together with Large and Small Language Models for Efficient Reasoning

TutorialsDGX agent

arXiv:2604.23623v1 Announce Type: new Abstract: Recent advancements in large language models (LLMs) have catalyzed the rise of reasoning-intensive inference paradigms, where models perform explicit st

Task-guided Spatiotemporal Network with Diffusion Augmentation for EEG-based Dementia Diagnosis and MMSE Prediction

ResearchDGX agent

arXiv:2604.23964v1 Announce Type: cross Abstract: Patients with dementia typically exhibit cognitive impairment, which is routinely assessed using the Mini-Mental State Examination (MMSE). Concurrentl

TCOD: Exploring Temporal Curriculum in On-Policy Distillation for Multi-turn Autonomous Agents

SafetyDGX agent

arXiv:2604.24005v1 Announce Type: cross Abstract: On-policy distillation (OPD) has shown strong potential for transferring reasoning ability from frontier or domain-specific models to smaller students

TeachMaster: Generative Teaching via Code

AgentsDGX agent

arXiv:2601.04204v2 Announce Type: replace-cross Abstract: The scalability of high-quality online education is hindered by the high costs and slow cycles of manual content creation. Despite advancement

TEMPO: Transformers for Temporal Disease Progression from Cross-Sectional Data

ResearchDGX agent

arXiv:2604.23368v1 Announce Type: new Abstract: Event-Based Models (EBMs) infer biomarker progression from cross-sectional data but typically only as ordinal sequences and rely on rigid model assumpti

Tessera: Secure, Near-Line-Rate Weight Streaming for UMA Edge Accelerators

ResearchDGX agent

arXiv:2604.23205v1 Announce Type: cross Abstract: Deploying proprietary Deep Neural Networks (DNNs) on commodity edge devices demands hardware-backed Digital Rights Management (DRM) capable of withsta

Test of Time: Rethinking Temporal Signal of Benchmark Contamination

Model ReleasesDGX agent

arXiv:2509.00072v3 Announce Type: replace Abstract: Post-cutoff performance decay has been widely interpreted as a temporal signal for benchmark contamination. We critically examine this belief and de

Test-Time Adaptation for Unsupervised Combinatorial Optimization

Local AiDGX agent

arXiv:2601.21048v2 Announce Type: replace Abstract: Unsupervised neural combinatorial optimization (NCO) enables learning powerful solvers without access to ground-truth solutions. Existing approaches

TexOCR: Advancing Document OCR Models for Compilable Page-to-LaTeX Reconstruction

Model ReleasesDGX agent

arXiv:2604.22880v1 Announce Type: new Abstract: Existing document OCR largely targets plain text or Markdown, discarding the structural and executable properties that make LaTeX essential for scientif

Text-Guided Multimodal Unified Industrial Anomaly Detection

SafetyDGX agent

arXiv:2604.22899v1 Announce Type: new Abstract: Industrial anomaly detection based on RGB-3D multimodal data has emerged as a mainstream paradigm for intelligent quality inspection. However, existing

Text Tells the Cost: Predicting and Analyzing Repayment Effort of Self-Admitted Technical Debt

ResearchDGX agent

arXiv:2309.06020v3 Announce Type: replace-cross Abstract: Technical debt refers to the consequences of sub-optimal decisions made during software development that prioritize short-term benefits over l

TextGround4M: A Prompt-Aligned Dataset for Layout-Aware Text Rendering

Model ReleasesDGX agent

arXiv:2604.24459v1 Announce Type: new Abstract: Despite recent advances in text-to-image generation, models still struggle to accurately render prompt-specified text with correct spatial layout -- esp

The Alignment Target Problem: Divergent Moral Judgments of Humans, AI Systems, and Their Designers

Model ReleasesDGX agent

arXiv:2604.24155v1 Announce Type: cross Abstract: The quest to align machine behavior with human values raises fundamental questions about the moral frameworks that should govern AI decision-making. M

The Chameleon's Limit: Investigating Persona Collapse and Homogenization in Large Language Models

AgentsDGX agent

arXiv:2604.24698v1 Announce Type: new Abstract: Applications based on large language models (LLMs), such as multi-agent simulations, require population diversity among agents. We identify a pervasive

The Collapse of Heterogeneity in Silicon Philosophers

SafetyDGX agent

arXiv:2604.23575v1 Announce Type: cross Abstract: Silicon samples are increasingly used as a low-cost substitute for human panels and have been shown to reproduce aggregate human opinion with high fid

The Consensus Trap: Dissecting Subjectivity and the 'Ground Truth' Illusion in Data Annotation

SafetyDGX agent

arXiv:2602.11318v3 Announce Type: replace Abstract: In machine learning, 'ground truth' refers to the assumed correct labels used to train and evaluate models. However, the foundational 'ground truth'

The High Cost of Incivility: Quantifying Interaction Inefficiency via Multi-Agent Monte Carlo Simulations

AgentsDGX agent

arXiv:2512.08345v2 Announce Type: replace Abstract: Workplace toxicity is widely recognized as detrimental to organizational culture, yet quantifying its direct impact on operational efficiency remain

The Imbalanced User-AI Relationships as an Ethical Failure of Front-End Design in Healthcare AI

SafetyDGX agent

arXiv:2604.22767v1 Announce Type: cross Abstract: Ethical discourse on AI in healthcare has focused predominantly on back-end concerns such as bias, fairness and explainability, while the front-end in

The Kerimov-Alekberli Model: An Information-Geometric Framework for Real-Time System Stability

Model ReleasesDGX agent

arXiv:2604.24083v1 Announce Type: new Abstract: This study introduces the Kerimov-Alekberli model, a novel information-geometric framework that redefines AI safety by formally linking non-equilibrium

← Previous
1…838839840841842…998
Next →