AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
25,885 results
30 Apr 2026

Seamless Indoor-Outdoor Mapping for INGENIOUS First Responders

ResearchDGX agent

arXiv:2604.26368v1 Announce Type: new Abstract: In several applications it is desired to have 3D models not only from the outdoor spaces but also from inside the building. In the context of First Resp

Semantic Embeddings of Chemical Elements for Enhanced Materials Inference and Discovery

ResearchDGX agent

arXiv:2502.14912v2 Announce Type: replace Abstract: We present a framework for generating universal semantic embeddings of chemical elements to advance materials inference and discovery. This framewor

Semantic Foam: Unifying Spatial and Semantic Scene Decomposition

ResearchDGX agent

arXiv:2604.26262v1 Announce Type: new Abstract: Modern scene reconstruction methods, such as 3D Gaussian Splatting, enable photo-realistic novel view synthesis at real-time speeds. However, their adop

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Semi-supervised learning with max-margin graph cuts

ResearchDGX agent

arXiv:2604.26818v1 Announce Type: new Abstract: This paper proposes a novel algorithm for semisupervised learning. This algorithm learns graph cuts that maximize the margin with respect to the labels

SG-UniBuc-NLP at SemEval-2026 Task 6: Multi-Head RoBERTa with Chunking for Long-Context Evasion Detection

ResearchDGX agent

arXiv:2604.26375v1 Announce Type: cross Abstract: We describe our system for SemEval-2026 Task 6 (CLARITY: Unmasking Political Question Evasions), which classifies English political interview response

Similarity Choice and Negative Scaling in Supervised Contrastive Learning for Deepfake Audio Detection

ResearchDGX agent

arXiv:2604.26057v1 Announce Type: cross Abstract: Supervised contrastive learning (SupCon) is widely used to shape representations, but has seen limited targeted study for audio deepfake detection. Ex

SkyReels-Text: Fine-Grained Font-Controllable Text Editing for Poster Design

ResearchDGX agent

arXiv:2511.13285v2 Announce Type: replace Abstract: Artistic design, particularly poster design, often demands rapid yet precise modification of textual content while preserving visual harmony and typ

SnapPose3D: Diffusion-Based Single-Frame 2D-to-3D Lifting of Human Poses

ResearchDGX agent

arXiv:2604.26620v1 Announce Type: new Abstract: Depth ambiguity and joint uncertainty are the two main obstacles in obtaining accurate human pose predictions by 2D-to-3D lifting methods proposed in th

Source-Free Bistable Fluidic Gripper for Size-Selective and Stiffness-Adaptive Grasping

ResearchDGX agent

arXiv:2511.03691v2 Announce Type: replace Abstract: Conventional fluid-driven soft grippers typically depend on external sources, which limit portability and long-term autonomy. This work introduces a

SpatialFusion: Endowing Unified Image Generation with Intrinsic 3D Geometric Awareness

ResearchDGX agent

arXiv:2604.26341v1 Announce Type: new Abstract: Recent unified image generation models have achieved remarkable success by employing MLLMs for semantic understanding and diffusion backbones for image

Spatially-constrained clustering of geospatial features for heat vulnerability assessment of favelas in Rio de Janeiro

ResearchDGX agent

arXiv:2604.26133v1 Announce Type: new Abstract: Informal settlements face disproportionate exposure to climate-related health hazards. However, existing methodologies lack systematic approaches to lin

SpecTr-GBV: Multi-Draft Block Verification Accelerating Speculative Decoding

ResearchDGX agent

arXiv:2604.25925v1 Announce Type: new Abstract: Autoregressive language models suffer from high inference latency due to their sequential decoding nature. Speculative decoding (SD) mitigates this by e

Stochastic Scaling Limits and Synchronization by Noise in Deep Transformer Models

ResearchDGX agent

arXiv:2604.26898v1 Announce Type: cross Abstract: We prove pathwise convergence of the layerwise evolution of tokens in a finite-depth, finite-width transformer model with MultiLayer Perceptron (MLP)

SynSur: An end-to-end generative pipeline for synthetic industrial surface defect generation and detection

ResearchDGX agent

arXiv:2604.26633v1 Announce Type: cross Abstract: The bottleneck in learning-based industrial defect detection is often limited not by model capacity, but by the scarcity of labeled defect data: defec

Text-Utilization for Encoder-dominated Speech Recognition Models

ResearchDGX agent

arXiv:2604.26514v1 Announce Type: cross Abstract: This paper investigates efficient methods for utilizing text-only data to improve speech recognition, focusing on encoder-dominated models that facili

The Bandit's Blind Spot: The Critical Role of User State Representation in Recommender Systems

ResearchDGX agent

arXiv:2604.26651v1 Announce Type: cross Abstract: With the increasing availability of online information, recommender systems have become an important tool for many web-based systems. Due to the conti

The devil is in the details: Enhancing Video Virtual Try-On via Keyframe-Driven Details Injection

ResearchDGX agent

arXiv:2512.20340v3 Announce Type: replace Abstract: Although diffusion transformer (DiT)-based video virtual try-on (VVT) has made significant progress in synthesizing realistic videos, existing metho

The Dual Role of Abstracting over the Irrelevant in Symbolic Explanations: Cognitive Effort vs. Understanding

ResearchDGX agent

arXiv:2602.03467v2 Announce Type: replace Abstract: Explanations are central to human cognition, yet AI systems often produce outputs that are difficult to understand. While symbolic AI offers a trans

The Serial Scaling Hypothesis

ResearchDGX agent

arXiv:2507.12549v4 Announce Type: replace Abstract: While machine learning has advanced through massive parallelization, we identify a critical blind spot: some problems are fundamentally sequential.

These folks are trying to ban open source. They're looking to take away your freedom to choose. They're also looking to take away the rights…

ResearchDGX agent

These folks are trying to ban open source. They're looking to take away your freedom to choose. They're also looking to take away the rights of businesses like Cursor to fine tune and make their produ

Through a Compressed Lens: Investigating The Impact of Quantization on Factual Knowledge Recall

ResearchDGX agent

arXiv:2505.13963v3 Announce Type: replace Abstract: Quantization methods are widely used to accelerate inference and streamline the deployment of large language models (LLMs). Although quantization's

ToolPRM: Fine-Grained Inference Scaling of Structured Outputs for Function Calling

ResearchDGX agent

arXiv:2510.14703v2 Announce Type: replace Abstract: Large language models (LLMs) excel at function calling, but inference scaling has been explored mainly for unstructured generation. We propose an in

Tree-of-Text: A Tree-based Prompting Framework for Table-to-Text Generation in the Sports Domain

ResearchDGX agent

arXiv:2604.26501v1 Announce Type: cross Abstract: Generating sports game reports from structured tables is a complex table-to-text task that demands both precise data interpretation and fluent narrati

✨🧠 Tribe v2, our latest model of human brain responses to sound, sight and language can now be (partly) explored on your phone📱: ▶️demo: h…

ResearchDGX agent

✨🧠 Tribe v2, our latest model of human brain responses to sound, sight and language can now be (partly) explored on your phone📱: ▶️demo: https://aidemos.atmeta.com/tribev2/ 📄paper: https://ai.meta.com

Trump's war on science.

ResearchDGX agent

Trump's war on science. The Trump administration has downsized US science by historic margins — but it's not just via grant or workforce cuts. Our new @nature analysis reveals the government has cut m

Turning the TIDE: Cross-Architecture Distillation for Diffusion Large Language Models

ResearchDGX agent

arXiv:2604.26951v1 Announce Type: cross Abstract: Diffusion large language models (dLLMs) offer parallel decoding and bidirectional context, but state-of-the-art dLLMs require billions of parameters f

Two Heads Are Better Than One: Async Knowledge Injection for Speech AI with Tandem Architecture Blog: https://pub.sakana.ai/kame/ 🐢 KAME: T…

ResearchDGX agent

Two Heads Are Better Than One: Async Knowledge Injection for Speech AI with Tandem Architecture Blog: https://pub.sakana.ai/kame/ 🐢 KAME: Tandem Architecture for Enhancing Knowledge in Real-Time Speec

U-FaceBP: Uncertainty-aware Bayesian Ensemble Deep Learning for Face Video-based Blood Pressure Estimation

ResearchDGX agent

arXiv:2412.10679v3 Announce Type: replace Abstract: Blood pressure (BP) measurement is crucial for daily health assessment. Remote photoplethysmography (rPPG), which extracts pulse waves from face vid

Understanding DNNs in Feature Interaction Models: A Dimensional Collapse Perspective

ResearchDGX agent

arXiv:2604.26489v1 Announce Type: new Abstract: DNNs have gained widespread adoption in feature interaction recommendation models. However, there has been a longstanding debate on their roles. On one

Unified 4D World Action Modeling from Video Priors with Asynchronous Denoising

ResearchDGX agent

arXiv:2604.26694v1 Announce Type: cross Abstract: We propose X-WAM, a Unified 4D World Model that unifies real-time robotic action execution and high-fidelity 4D world synthesis (video + 3D reconstruc

Unsupervised Graph Modeling for Anomaly Detection in Accounting Subject Relationships

ResearchDGX agent

arXiv:2604.26216v1 Announce Type: new Abstract: This paper addresses the problem of anomaly detection in accounting subject association structures, proposing a structured modeling and unsupervised dis

When Hidden States Drift: Can KV Caches Rescue Long-Range Speculative Decoding?

ResearchDGX agent

arXiv:2604.26412v1 Announce Type: new Abstract: Speculative decoding accelerates LLM inference, but SOTA hidden-state-based drafters suffer from long-range decay: draft accuracy degrades as the specul

When to Vote, When to Rewrite: Disagreement-Guided Strategy Routing for Test-Time Scaling

ResearchDGX agent

arXiv:2604.26644v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) achieve strong performance on mathematical reasoning tasks but remain unreliable on challenging instances. Existing test-t

Why Attend to Everything? Focus is the Key

ResearchDGX agent

arXiv:2604.03260v2 Announce Type: replace-cross Abstract: Standard attention scales quadratically with sequence length. Efficient attention methods reduce this O(n^2) cost, but when retrofitted into p

World2VLM: Distilling World Model Imagination into VLMs for Dynamic Spatial Reasoning

ResearchDGX agent

arXiv:2604.26934v1 Announce Type: new Abstract: Vision-language models (VLMs) have shown strong performance on static visual understanding, yet they still struggle with dynamic spatial reasoning that

29 Apr 2026

8DNA: 8D Neural Asset Light Transport by Distribution Learning

ResearchDGX agent

arXiv:2604.25129v1 Announce Type: cross Abstract: High-fidelity 3D assets exhibit intriguing global illumination effects like subsurface scattering, glossy interreflections, and fine-scale fiber scatt

'A decade ago, AI was supposed to replace radiologists. Today, radiologists make more than $500,000 per year, and their employment continues…

ResearchDGX agent

'A decade ago, AI was supposed to replace radiologists. Today, radiologists make more than $500,000 per year, and their employment continues to grow, see chart below. Reading scans is a task, not a jo

A graph generation pipeline for critical infrastructures based on heuristics, images and depth data

ResearchDGX agent

arXiv:2512.07269v2 Announce Type: replace Abstract: Virtual representations of physical critical infrastructures, such as water or energy plants, are used for simulations and digital twins to ensure r

A paradox of AI fluency

ResearchDGX agent

arXiv:2604.25905v1 Announce Type: new Abstract: How much does a user's skill with AI shape what AI actually delivers for them? This question is critical for users, AI product builders, and society at

A systematic literature Review for Transformer-based Software Vulnerability detection

ApplicationsDGX agent

arXiv:2604.24822v1 Announce Type: cross Abstract: Context: Software vulnerabilities pose significant security threats to software systems, especially as software is increasingly used across many areas

A Unifying Framework for Unsupervised Concept Extraction

ResearchDGX agent

arXiv:2604.24936v1 Announce Type: new Abstract: Techniques for concept extraction, such as sparse autoencoders and transcoders, aim to extract high-level symbolic concepts from low-level nonsymbolic r

Accuracy Improvement of Cell Image Segmentation Using Feedback Former

ResearchDGX agent

arXiv:2408.12974v4 Announce Type: replace Abstract: Semantic segmentation of microscopy cell images by deep learning is a significant technique. We considered that the Transformers, which have recentl

Adaptive Meta-Learning Stochastic Gradient Hamiltonian Monte Carlo Simulation for Bayesian Updating of Structural Dynamic Models

ResearchDGX agent

arXiv:2604.25710v1 Announce Type: cross Abstract: In the last few decades, Markov chain Monte Carlo (MCMC) methods have been widely applied to Bayesian updating of structural dynamic models in the fie

An analysis of sensor selection for fruit picking with suction-based grippers

ResearchDGX agent

arXiv:2604.24906v1 Announce Type: cross Abstract: Robotic fruit harvesting often fails to reliably detect whether a fruit has been successfully picked, limiting efficiency and increasing crop damage.

ARQ: A Mixed-Precision Quantization Framework for Accurate and Certifiably Robust DNNs

ResearchDGX agent

arXiv:2410.24214v3 Announce Type: replace-cross Abstract: Mixed precision quantization has become an important technique for optimizing the execution of deep neural networks (DNNs). Certified robustne

Automated detection of pediatric congenital heart disease from phonocardiograms using deep and handcrafted feature fusion

ResearchDGX agent

arXiv:2604.24767v1 Announce Type: cross Abstract: Congenital heart disease (CHD) is the most common type of birth defect, impacting about 1% of live births worldwide. Echocardiography, the gold-standa

Backtranslation Augmented Direct Preference Optimization for Neural Machine Translation

ResearchDGX agent

arXiv:2604.25702v1 Announce Type: new Abstract: Contemporary neural machine translation (NMT) systems are almost exclusively built by training on supervised parallel data. Despite the tremendous progr

Benchmarking Logistic Regression, SVM, and LightGBM Against BiLSTM with Attention for Sentiment Analysis on Indonesian Product Reviews

ResearchDGX agent

arXiv:2604.25452v1 Announce Type: new Abstract: Sentiment analysis of product reviews on e-commerce platforms plays a critical role in automatically understanding customer satisfaction and providing a

Benchmarking PyCaret AutoML Against IndoBERT Fine-Tuning for Sentiment Analysis on Indonesian IKN Twitter Data

ResearchDGX agent

arXiv:2604.25392v1 Announce Type: new Abstract: This paper benchmarks a classical machine learning approach based on PyCaret AutoML against a deep learning approach based on IndoBERT fine-tuning for b

Beyond Fidelity: Semantic Similarity Assessment in Low-Level Image Processing

ResearchDGX agent

arXiv:2604.25408v1 Announce Type: new Abstract: Low-level image processing has long been evaluated mainly from the perspective of visual fidelity. However, with the rise of deep learning and generativ

Biased Dreams: Limitations to Epistemic Uncertainty Quantification in Latent Space Models

ResearchDGX agent

arXiv:2604.25416v1 Announce Type: new Abstract: Model-Based Reinforcement Learning distinguishes between physical dynamics models operating on proprioceptive inputs and latent dynamics models operatin

Bridging the Indoor-Outdoor Gap: Cross-Technology Ranging for Seamless Robot Navigation

ResearchDGX agent

arXiv:2604.25541v1 Announce Type: cross Abstract: Mobile robots that move between outdoor and indoor environments still struggle with consistent positioning. Satellite-based and terrestrial ranging ea

Calibrated Fusion for Heterogeneous Graph-Vector Retrieval in Multi-Hop QA

ResearchDGX agent

arXiv:2603.28886v2 Announce Type: replace-cross Abstract: Graph-augmented retrieval combines dense similarity with graph-based relevance signals such as Personalized PageRank (PPR), but these scores h

Can We Change the Stroke Size for Easier Diffusion?

ResearchDGX agent

arXiv:2603.26783v2 Announce Type: replace Abstract: Diffusion models can be challenged in the low signal-to-noise regime, where they have to make pixel-level predictions despite the presence of high n

Categorical Optimization with Bayesian Anchored Latent Trust Regions for Structural Design under High-Dimensional Uncertainty

ResearchDGX agent

arXiv:2604.25241v1 Announce Type: new Abstract: Categorical structural optimization under aleatoric uncertainty is challenging because each design variable must be selected from a finite catalog of ad

CodeOCR: On the Effectiveness of Vision Language Models in Code Understanding

ResearchDGX agent

arXiv:2602.01785v2 Announce Type: replace Abstract: Large Language Models (LLMs) have achieved remarkable success in source code understanding, yet as software systems grow in scale, computational eff

Comparative Study of Bending Analysis using Physics-Informed Neural Networks and Numerical Dynamic Deflection in Perforated nanobeam

ResearchDGX agent

arXiv:2604.24768v1 Announce Type: new Abstract: In this chapter, we investigate the bending behavior of a perforated nanobeam subjected to sinusoidal loading using an efficient and computationally rob

Conditional Flow Matching for Probabilistic Downscaling of Maximum 3-day Snowfall in Alaska

ResearchDGX agent

arXiv:2604.25172v1 Announce Type: cross Abstract: Precipitation in complex terrain is governed by orographic processes operating at scales of a few kilometers, yet climate models typically run at reso

CoreFlow: Low-Rank Matrix Generative Models

ResearchDGX agent

arXiv:2604.24959v1 Announce Type: new Abstract: Learning matrix-valued distributions from high-dimensional and possibly incomplete training data is challenging: ambient-space generative modeling is co

Cornserve: A Distributed Serving System for Any-to-Any Multimodal Models

ResearchDGX agent

arXiv:2603.12118v2 Announce Type: replace Abstract: Any-to-Any models are an emerging class of multimodal models that accept combinations of multimodal data (e.g., text, image, video, audio) as input

← Previous
1…292293294295296…432
Next →