AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

research

GridTimelineEvolution
18,859 results
13 Apr 2026

LLMs Underperform Graph-Based Parsers on Supervised Relation Extraction for Complex Graphs

ResearchDGX agent

arXiv:2604.08752v1 Announce Type: cross Abstract: Relation extraction represents a fundamental component in the process of creating knowledge graphs, among other applications. Large language models (L

LoBE-GS: Load-Balanced and Efficient 3D Gaussian Splatting for Large-Scale Scene Reconstruction

ResearchDGX agent

arXiv:2510.01767v2 Announce Type: replace Abstract: 3D Gaussian Splatting (3DGS) has established itself as an efficient representation for real-time, high-fidelity 3D scene reconstruction. However, sc

Localizing Task Recognition and Task Learning in In-Context Learning via Attention Head Analysis

ResearchDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2509.24164v2 Announce Type: replace Abstract: We investigate the mechanistic underpinnings of in-context learning (ICL) in large language models by reconciling two dominant perspectives: the com

Loom: A Scalable Analytical Neural Computer Architecture

ResearchDGX agent

arXiv:2604.08816v1 Announce Type: new Abstract: We present Loom, a computer architecture that executes programs compiled from C inside a looped transformer whose weights are derived analytically. The

Mandatory In-Person Presentation in CVPR 2026 [D]

ResearchDGX agent

This Reddit thread on r/MachineLearning discusses CVPR 2026's policy requiring accepted papers to be registered under an in-person author registration, with virtual attendance still permitted if circu

MASS: Mesh-inellipse Aligned Deformable Surfel Splatting for Hand Reconstruction and Rendering from Egocentric Monocular Video

ResearchDGX agent

arXiv:2604.08943v1 Announce Type: new Abstract: Reconstructing high-fidelity 3D hands from egocentric monocular videos remains a challenge due to the limitations in capturing high-resolution geometry,

Measurement-Consistent Langevin Corrector for Stabilizing Latent Diffusion Inverse Problem Solvers

ResearchDGX agent

arXiv:2601.04791v3 Announce Type: replace Abstract: While latent diffusion models (LDMs) have emerged as powerful priors for inverse problems, existing LDM-based solvers frequently suffer from instabi

Meta-Learned Basis Adaptation for Parametric Linear PDEs

ResearchDGX agent

arXiv:2604.09289v1 Announce Type: new Abstract: We propose a hybrid physics-informed framework for solving families of parametric linear partial differential equations (PDEs) by combining a meta-learn

Modality-Aware Zero-Shot Pruning and Sparse Attention for Efficient Multimodal Edge Inference

ResearchDGX agent

arXiv:2604.08971v1 Announce Type: new Abstract: Edge devices increasingly run multimodal sensing pipelines that must remain accurate despite fluctuating power budgets and unpredictable sensor dropout.

Multi-task Just Recognizable Difference for Video Coding for Machines: Database, Model, and Coding Application

ResearchDGX agent

arXiv:2604.09421v1 Announce Type: cross Abstract: Just Recognizable Difference (JRD) boosts coding efficiency for machine vision through visibility threshold modeling, but is currently limited to a si

Music Audio-Visual Question Answering Requires Specialized Multimodal Designs

ResearchDGX agent

arXiv:2505.20638v2 Announce Type: replace-cross Abstract: While recent Multimodal Large Language Models exhibit impressive capabilities for general multimodal tasks, specialized domains like music nec

[N] AMA Announcement: Max Welling (VAEs, GNNs, AI4Science & CuspAI)

ResearchDGX agent

This r/MachineLearning AMA announcement features Max Welling, a prominent figure in machine learning renowned for his foundational contributions to probabilistic deep learning, including the co-develo

Nested Radially Monotone Polar Occupancy Estimation: Clinically-Grounded Optic Disc and Cup Segmentation for Glaucoma Screening

ResearchDGX agent

arXiv:2604.09062v1 Announce Type: new Abstract: Valid segmentation of the optic disc (OD) and optic cup (OC) from fundus photographs is essential for glaucoma screening. Unfortunately, existing deep l

一つのニューラルネットに符号と記号は創発しうるか? 「Neural Computers」論文から考える @rmaruy https://rmaruy.hatenablog.com/entry/2026/04/11/223828

ResearchDGX agent

This Japanese blog post explores whether signs and symbols can emerge within a single neural network, drawing on analysis of the 'Neural Computers' paper. The discussion likely examines the intersecti

No Single Best Model for Diversity: Learning a Router for Sample Diversity

ResearchDGX agent

arXiv:2604.02319v2 Announce Type: replace Abstract: When posed with prompts that permit a large number of valid answers, comprehensively generating them is the first step towards satisfying a wide ran

OASIS: Online Activation Subspace Learning for Memory-Efficient Training

ResearchDGX agent

arXiv:2604.09406v1 Announce Type: new Abstract: Training large language models (LLMs) is constrained by memory requirements, with activations accounting for a substantial fraction of the total footpri

Offline Local Search for Online Stochastic Bandits

ResearchDGX agent

arXiv:2604.09423v1 Announce Type: new Abstract: Combinatorial multi-armed bandits provide a fundamental online decision-making environment where a decision-maker interacts with an environment across T

On the Limits of Layer Pruning for Generative Reasoning in Large Language Models

ResearchDGX agent

arXiv:2602.01997v2 Announce Type: replace-cross Abstract: Recent work has shown that layer pruning can effectively compress large language models (LLMs) while retaining strong performance on classific

On the Terminology and Geometric Aspects of Redundant Parallel Manipulators

ResearchDGX agent

arXiv:2604.09156v1 Announce Type: new Abstract: Parallel kinematics machines (PKM) can exhibit kinematic as well as actuation redundancy. While the meaning of kinematic redundancy has been clarified a

One Interface, Many Robots: Unified Real-Time Low-Level Motion Planning for Collaborative Arms

ResearchDGX agent

arXiv:2604.08787v1 Announce Type: new Abstract: This paper proposes a common interface for real-time low-level motion planning of collaborative robotic arms, aimed at enabling broader applicability an

Online Quantile Regression for Nonparametric Additive Models

ResearchDGX agent

arXiv:2604.08969v1 Announce Type: cross Abstract: This paper introduces a projected functional gradient descent algorithm (P-FGD) for training nonparametric additive quantile regression models in onli

Optimal Multi-bit Generative Watermarking Schemes Under Worst-Case False-Alarm Constraints

ResearchDGX agent

arXiv:2604.08759v1 Announce Type: cross Abstract: This paper considers the problem of multi-bit generative watermarking for large language models under a worst-case false-alarm constraint. Prior work

Out-of-the-box: Black-box Causal Attacks on Object Detectors

ResearchDGX agent

arXiv:2512.03730v2 Announce Type: replace-cross Abstract: Adversarial perturbations are a useful way to expose vulnerabilities in object detectors. Existing perturbation methods are frequently white-b

Overhang Tower: Resource-Rational Adaptation in Sequential Physical Planning

ResearchDGX agent

arXiv:2604.09072v1 Announce Type: new Abstract: Humans effortlessly navigate the physical world by predicting how objects behave under gravity and contact forces, yet how such judgments support sequen

Overstating Attitudes, Ignoring Networks: LLM Biases in Simulating Misinformation Susceptibility

ResearchDGX agent

arXiv:2602.04674v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly used as proxies for human judgment in computational social science, yet their ability to reprodu

p1: Better Prompt Optimization with Fewer Prompts

ResearchDGX agent

arXiv:2604.08801v1 Announce Type: cross Abstract: Prompt optimization improves language models without updating their weights by searching for a better system prompt, but its effectiveness varies wide

P3P Made Easy

ResearchDGX agent

arXiv:2508.01312v4 Announce Type: replace Abstract: We revisit the classical Perspective-Three-Point (P3P) problem, which aims to recover the absolute pose of a calibrated camera from three 2D-3D corr

PaceLLM: Brain-Inspired Large Language Models for Long-Context Understanding

ResearchDGX agent

arXiv:2506.17310v3 Announce Type: replace-cross Abstract: While Large Language Models (LLMs) demonstrate strong performance across domains, their long-context capabilities are limited by transient neu

PDE-regularized Dynamics-informed Diffusion with Uncertainty-aware Filtering for Long-Horizon Dynamics

ResearchDGX agent

arXiv:2604.09058v1 Announce Type: cross Abstract: Long-horizon spatiotemporal prediction remains a challenging problem due to cumulative errors, noise amplification, and the lack of physical consisten

Persona-E^2: A Human-Grounded Dataset for Personality-Shaped Emotional Responses to Textual Events

ResearchDGX agent

arXiv:2604.09162v1 Announce Type: cross Abstract: Most affective computing research treats emotion as a static property of text, focusing on the writer's sentiment while overlooking the reader's persp

Physics-guided surrogate learning enables zero-shot control of turbulent wings

ResearchDGX agent

arXiv:2604.09434v1 Announce Type: cross Abstract: Turbulent boundary layers over aerodynamic surfaces are a major source of aircraft drag, yet their control remains challenging due to multiscale dynam

PoseGen: In-Context LoRA Finetuning for Pose-Controllable Long Human Video Generation

ResearchDGX agent

arXiv:2508.05091v2 Announce Type: replace Abstract: Generating temporally coherent, long-duration videos with precise control over subject identity and movement remains a fundamental challenge for con

Practical Bayesian Inference for Speech SNNs: Uncertainty and Loss-Landscape Smoothing

ResearchDGX agent

arXiv:2604.08624v1 Announce Type: cross Abstract: Spiking Neural Networks (SNNs) are naturally suited for speech processing tasks due to their specific dynamics, which allows them to handle temporal d

PRADA: Probability-Ratio-Based Attribution and Detection of Autoregressive-Generated Images

ResearchDGX agent

arXiv:2511.20068v2 Announce Type: replace Abstract: Autoregressive (AR) image generation has recently emerged as a powerful paradigm for image synthesis. Leveraging the generation principle of large l

PRAGMA: Revolut Foundation Model

ResearchDGX agent

arXiv:2604.08649v1 Announce Type: cross Abstract: Modern financial systems generate vast quantities of transactional and event-level data that encode rich economic signals. This paper presents PRAGMA,

Predictive Entropy Links Calibration and Paraphrase Sensitivity in Medical Vision-Language Models

ResearchDGX agent

arXiv:2604.08941v1 Announce Type: new Abstract: Medical Vision Language Models VLMs suffer from two failure modes that threaten safe deployment mis calibrated confidence and sensitivity to question re

Prototype-Regularized Federated Learning for Cross-Domain Aspect Sentiment Triplet Extraction

ResearchDGX agent

arXiv:2604.09123v1 Announce Type: new Abstract: Aspect Sentiment Triplet Extraction (ASTE) aims to extract all sentiment triplets of aspect terms, opinion terms, and sentiment polarities from a senten

PS-TTS: Phonetic Synchronization in Text-to-Speech for Achieving Natural Automated Dubbing

ResearchDGX agent

arXiv:2604.09111v1 Announce Type: cross Abstract: Recently, artificial intelligence-based dubbing technology has advanced, enabling automated dubbing (AD) to convert the source speech of a video into

PSIRNet: Deep Learning-based Free-breathing Rapid Acquisition Late Enhancement Imaging

ResearchDGX agent

arXiv:2604.08781v1 Announce Type: cross Abstract: Purpose: To develop and evaluate a deep learning (DL) method for free-breathing phase-sensitive inversion recovery (PSIR) late gadolinium enhancement

RAM: Recover Any 3D Human Motion in-the-Wild

ResearchDGX agent

arXiv:2603.19929v2 Announce Type: replace-cross Abstract: RAM incorporates a motion-aware semantic tracker with adaptive Kalman filtering to achieve robust identity association under severe occlusions

Ranked Activation Shift for Post-Hoc Out-of-Distribution Detection

ResearchDGX agent

arXiv:2604.08572v1 Announce Type: cross Abstract: State-of-the-art post-hoc out-of-distribution detection methods rely on intermediate layer activation editing. However, they exhibit inconsistent perf

Rays as Pixels: Learning A Joint Distribution of Videos and Camera Trajectories

ResearchDGX agent

arXiv:2604.09429v1 Announce Type: cross Abstract: Recovering camera parameters from images and rendering scenes from novel viewpoints have long been treated as separate tasks in computer vision and gr

REACT3D: Recovering Articulations for Interactive Physical 3D Scenes

ResearchDGX agent

arXiv:2510.11340v4 Announce Type: replace Abstract: Interactive 3D scenes are increasingly vital for embodied intelligence, yet existing datasets remain limited due to the labor-intensive process of a

Reasoning Models Will Sometimes Lie About Their Reasoning

ResearchDGX agent

arXiv:2601.07663v3 Announce Type: replace Abstract: Hint-based faithfulness evaluations have established that Large Reasoning Models (LRMs) may not say what they think: they do not always volunteer in

RecaLLM: Addressing the Lost-in-Thought Phenomenon with Explicit In-Context Retrieval

ResearchDGX agent

arXiv:2604.09494v1 Announce Type: cross Abstract: We propose RecaLLM, a set of reasoning language models post-trained to make effective use of long-context information. In-context retrieval, which ide

Reflection of Episodes: Learning to Play Game from Expert and Self Experiences

ResearchDGX agent

arXiv:2502.13388v4 Announce Type: replace Abstract: StarCraft II is a complex and dynamic real-time strategy (RTS) game environment, which is very suitable for artificial intelligence and reinforcemen

Regime-Conditional Retrieval: Theory and a Transferable Router for Two-Hop QA

ResearchDGX agent

arXiv:2604.09019v1 Announce Type: cross Abstract: Two-hop QA retrieval splits queries into two regimes determined by whether the hop-2 entity is explicitly named in the question (Q-dominant) or only i

Reservoir observer enhanced with residual calibration and attention mechanism

ResearchDGX agent

arXiv:2604.08592v1 Announce Type: new Abstract: Reservoir observers provide a data-driven approach to the inference of unmeasured variables from observed ones for nonlinear dynamical systems. While pr

Revisiting Anisotropy in Language Transformers: The Geometry of Learning Dynamics

ResearchDGX agent

arXiv:2604.08764v1 Announce Type: new Abstract: Since their introduction, Transformer architectures have dominated Natural Language Processing (NLP). However, recent research has highlighted an inhere

Revisiting the Capacity Gap in Chain-of-Thought Distillation from a Practical Perspective

ResearchDGX agent

arXiv:2604.08880v1 Announce Type: cross Abstract: Chain-of-thought (CoT) distillation transfers reasoning behaviors from a strong teacher to a smaller student, but prior work reports a capacity gap: d

Robust Adaptive Backstepping Impedance Control of Robots in Unknown Environments

ResearchDGX agent

arXiv:2604.09323v1 Announce Type: new Abstract: This paper presents a Robust Adaptive Backstepping Impedance Control (RABIC) strategy for robots operating in contact-rich and uncertain environments. T

Sample Complexity of Composite Quantum Hypothesis Testing

ResearchDGX agent

arXiv:2601.08588v4 Announce Type: replace-cross Abstract: This paper investigates symmetric composite binary quantum hypothesis testing (QHT), where the goal is to determine which of two uncertainty s

Scalable High-Recall Constraint-Satisfaction-Based Information Retrieval for Clinical Trials Matching

ResearchDGX agent

arXiv:2604.08849v1 Announce Type: cross Abstract: Clinical trials are central to evidence-based medicine, yet many struggle to meet enrollment targets, despite the availability of over half a million

Scaling flow-based approaches for topology sampling in SU(3) gauge theory

ResearchDGX agent

arXiv:2510.25704v2 Announce Type: replace-cross Abstract: We develop a methodology based on out-of-equilibrium simulations to mitigate topological freezing when approaching the continuum limit of latt

Screen, Cache, and Match: A Training-Free Causality-Consistent Reference Frame Framework for Human Animation

ResearchDGX agent

arXiv:2601.22160v2 Announce Type: replace-cross Abstract: Human animation aims to generate temporally coherent and visually consistent videos over long sequences, yet modeling long-range dependencies

Self-Supervised Slice-to-Volume Reconstruction with Gaussian Representations for Fetal MRI

ResearchDGX agent

arXiv:2601.22990v2 Announce Type: replace-cross Abstract: Reconstructing 3D fetal MR volumes from motion-corrupted stacks of 2D slices is a crucial and challenging task. Conventional slice-to-volume r

🎶 Share Your Thoughts on Music Description using AI! (Short Survey) [R]

ResearchDGX agent

This Reddit post on r/MachineLearning is a short community survey seeking opinions and experiences related to AI-based music description — that is, the use of AI models to automatically generate natur

Sharp description of local minima in the loss landscape of high-dimensional two-layer ReLU neural networks

ResearchDGX agent

arXiv:2604.09412v1 Announce Type: cross Abstract: We study the population loss landscape of two-layer ReLU networks of the form sum_{k=1}^K ReLU(w_k^op x) in a realisable teacher-student setting with

SIC3D: Style Image Conditioned Text-to-3D Gaussian Splatting Generation

ResearchDGX agent

arXiv:2604.08760v1 Announce Type: new Abstract: Recent progress in text-to-3D object generation enables the synthesis of detailed geometry from text input by leveraging 2D diffusion models and differe

Silhouette Loss: Differentiable Global Structure Learning for Deep Representations

ResearchDGX agent

arXiv:2604.08573v1 Announce Type: cross Abstract: Learning discriminative representations is a central goal of supervised deep learning. While cross-entropy (CE) remains the dominant objective for cla

← Previous
1…305306307308309…315
Next →