AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

research

GridTimelineEvolution
19,191 results
2 Jul 2026

GADA: Geometry-Aware Deformable Aggregation for Image-Based Gaussian Splatting

ResearchDGX agent

arXiv:2607.00595v1 Announce Type: new Abstract: Gaussian Splatting has achieved significant improvements by incorporating warping-based techniques. However, such methods suffer from pixel-level inaccu

GaussianEmoTalker: Real-Time Emotional Talking Head Synthesis with Audio-Driven and Blendshape-Based 3D Gaussian Splatting

ResearchDGX agent

arXiv:2607.00959v1 Announce Type: new Abstract: Audio-driven talking head synthesis has achieved impressive progress in lip synchronization and visual quality, yet generating expressive emotional avat

Generated Contents Enrichment

ResearchDGX agent

arXiv:2405.03650v4 Announce Type: replace Abstract: We study Generated Contents Enrichment (GCE), a conditional image-generation task in which a sparse scene description is first enriched through an e


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Generative Refinement for Low-Budget Black-Box Optimization

ResearchDGX agent

arXiv:2607.00691v1 Announce Type: new Abstract: Black-box optimization is a fundamental science and engineering tool that makes it possible to optimize objectives without gradient information. Unfortu

GimbalDiffusion: Gravity-Aware Camera Control for Video Generation

ResearchDGX agent

arXiv:2512.09112v3 Announce Type: replace Abstract: Recent progress in text-to-video generation has achieved remarkable realism, yet fine-grained control over camera motion and orientation remains elu

Goal-oriented learning of stochastic differential equations using error bounds on path-space observables

ResearchDGX agent

arXiv:2603.20467v2 Announce Type: replace-cross Abstract: Stochastic differential equations (SDEs), which serve as the governing equations for dynamical systems in a broad range of applications, can b

Graded strength of comparative illusions is explained by Bayesian inference

ResearchDGX agent

arXiv:2511.14642v2 Announce Type: replace Abstract: Like visual processing, language processing is susceptible to illusions in which people systematically misperceive stimuli. In one such case--the co

Group-invariant Coresets for Data-efficient Active Learning

ResearchDGX agent

arXiv:2607.01089v1 Announce Type: cross Abstract: Active learning reduces labeling cost by querying the most informative unlabeled samples, but standard coreset methods ignore known data symmetries an

Guaranteed Escape for a Bouncing Robot in Pipe Chains

ResearchDGX agent

arXiv:2607.00221v1 Announce Type: cross Abstract: We study the symmetric bouncing of a point robot within orthogonally-joined rectangles with equal width, which we refer to as pipes. We provide an exh

Hate Speech Detection in Turkish and Arabic Languages: A Comprehensive Study

ResearchDGX agent

arXiv:2607.00143v1 Announce Type: cross Abstract: Online hate speech has been linked to a global rise in violence against minorities, including incidents such as mass shootings, lynchings, and ethnic

hermes now supports reference-image editing with your codex/chatgpt login drop in a source image + up to 16 reference images, and it transfo…

ResearchDGX agent

Hermes now supports reference-image editing functionality that allows users to input a source image along with up to 16 reference images for transformation, accessible through Codex/ChatGPT login cred

Hey, That's My Model! Introducing Chain & Hash, An LLM Fingerprinting Technique

ResearchDGX agent

arXiv:2407.10887v4 Announce Type: replace-cross Abstract: Growing concerns over the theft and misuse of Large Language Models (LLMs) underscore the need for effective fingerprinting to link a model to

HieDG: A Hierarchical Discrete Geometry-Guided Framework for Multi-Animal Tracking

ResearchDGX agent

arXiv:2607.00494v1 Announce Type: new Abstract: Multi-animal tracking (MAT) is critical for wildlife monitoring and behavioral analysis, yet remains challenging due to uniform appearance, high density

High-dimensional Embedding Prior for Noisy K-space Domain MRIReconstruction

ResearchDGX agent

arXiv:2607.01176v1 Announce Type: new Abstract: Magnetic resonance imaging (MRI) reconstruction under realistic acquisition conditions can be fundamentally viewed as estimating the underlying k-space

How Do We Engage with Other Disciplines? A Framework to Study Meaningful Interdisciplinary Discourse in Scholarly Publications

ResearchDGX agent

arXiv:2601.17020v2 Announce Type: replace-cross Abstract: With the rising popularity of interdisciplinary work and increasing institutional incentives in this direction, there is a growing need to und

How Early Is Early Enough? Design-Dependent Observation-Window Sufficiency in Subscription Churn Prediction

ResearchDGX agent

arXiv:2607.00473v1 Announce Type: new Abstract: How many days of early behavior suffice for subscription churn prediction? In the public KKBox dataset, the early indicator of churn is typically an ind

How Ethos and Pathos Appeals Resonate in Reader Interpretations of Social Media Messages

ResearchDGX agent

arXiv:2607.00873v1 Announce Type: new Abstract: Rhetorical strategies and their influence on audiences are often studied through social media posts and comments. However, this focus overlooks the univ

How Much Do RF Drone Benchmarks Overstate? A Controlled Study and Theory of Data Leakage in UAV Signal Identification

ResearchDGX agent

arXiv:2607.01025v1 Announce Type: cross Abstract: Radio-frequency (RF) sensing is a central modality for counter-unmanned-aerial-system (counter-UAS) defence because it exploits the control, telemetry

Human-Centric Transferable Tactile Pre-Training for Dexterous Robotic Manipulation

ResearchDGX agent

arXiv:2607.01067v1 Announce Type: cross Abstract: As an essential modality for dexterous and contact-rich tasks, tactile sensing provides precise force feedback that cannot be reliably inferred from v

I’m looking to hire a Program Manager to help manage Sakana AI’s fast growing Recursive Self-Improvement (RSI) Lab 🚀 RSI Lab (English): htt…

ResearchDGX agent

I’m looking to hire a Program Manager to help manage Sakana AI’s fast growing Recursive Self-Improvement (RSI) Lab 🚀 RSI Lab (English): https://sakana.ai/rsi-lab/ RSI Lab (日本語): https://sakana.ai/rsi-

Image-Domain Tilt Constrained Distributed Fusion for Maneuvering UAV Tracking with Multi-Camera Electro-Optical Observations

ResearchDGX agent

arXiv:2607.01008v1 Announce Type: cross Abstract: Short-horizon prediction is essential for electro-optical UAV tracking, especially when the target is small, maneuvering, or intermittently observed.

Ink3D: Sculpting 3D Assets with Extremely Complex Textures via Video Generative Models

ResearchDGX agent

arXiv:2607.01222v1 Announce Type: new Abstract: Recent 3D generative models can synthesize high-quality geometry but often struggle to reproduce intricate textures from reference images, largely due t

Invariance Pair Guidance: Robustness to Spurious Correlations via Corrective Gradients

ResearchDGX agent

arXiv:2502.18975v2 Announce Type: replace Abstract: Machine learning models are inherently bound to the distribution of the training data, often exploiting non-causal shortcuts. As a result, achieving

Iterated Invariant EKF for 3D Landmark-Aided Inertial Navigation

ResearchDGX agent

arXiv:2607.00145v1 Announce Type: new Abstract: Inertial navigation systems aided by three-dimensional landmark measurements constitute a fundamental problem in robotic perception and state estimation

Joint Medical Image Enhancement and Segmentation with Diffusion-based Symbiotic Information Interaction

ResearchDGX agent

arXiv:2607.00058v1 Announce Type: new Abstract: Image quality is critical for accurate medical diagnosis. However, MRI, CT, and ultrasound images are often of low resolution and quality due to cost co

K-Inverse-RFM: A Modified RFM that Bridges the Gap to Neural Networks for Data-Corrupted Mathematical Tasks

ResearchDGX agent

arXiv:2607.00329v1 Announce Type: cross Abstract: Recursive Feature Machines (RFMs) are a class of kernel machines that utilize the Average Gradient Outer Product (AGOP) as a mechanism for feature lea

Know When to Stop: Segment-Level Credit Assignment for Reducing Overthinking

ResearchDGX agent

arXiv:2607.00482v1 Announce Type: new Abstract: Reasoning language models frequently overthink: generating extended chains of behaviors such as hedging, approach abandonment, and self contradiction th

KnowledgeDebugger -- an Exploration Tool for Knowledge Localization and Editing in Transformers

ResearchDGX agent

arXiv:2607.01000v1 Announce Type: new Abstract: Recent research has increasingly focused on understanding how Transformers store and process knowledge, as well as how this knowledge can be edited. Res

Language-Critique Imitation Learning from Suboptimal Demonstrations

ResearchDGX agent

arXiv:2607.01225v1 Announce Type: cross Abstract: Prior work on imitation learning from suboptimal demonstrations typically relies on compressed supervision signals such as confidence estimates, discr

Learning Structured Reasoning via Tractable Trajectory Control

ResearchDGX agent

Large language models can exhibit emergent reasoning behaviors, often manifested as recurring lexical patterns (e.g., “wait,” indicating verification). However, complex reasoning trajectories remain s

Learning to Compose: Revisiting Proxy Task Design for Zero-Shot Composed Image Retrieval

ResearchDGX agent

arXiv:2607.00374v1 Announce Type: cross Abstract: Composed Image Retrieval (CIR) retrieves a target image from a reference image and a textual modification. While supervised CIR relies on costly tripl

Learning Unmasking Policies for Diffusion Language Models

ResearchDGX agent

Diffusion (Large) Language Models (dLLMs) now match the downstream performance of their autoregressive counterparts on many tasks, while holding the promise of being more efficient during inference. O

Leveraging Multimodality for Real-Time Classification of Transients and Variables found by the Zwicky Transient Facility

ResearchDGX agent

arXiv:2607.00228v1 Announce Type: cross Abstract: Modern time-domain surveys such as the Zwicky Transient Facility (ZTF) generate hundreds of thousands of alerts each night, making real-time decisions

LeVLJEPA: End-to-End Vision-Language Pretraining Without Negatives

ResearchDGX agent

arXiv:2607.00784v1 Announce Type: cross Abstract: Vision-language pretraining remains dominated by contrastive objectives, whereas vision-only self-supervised learning has largely adopted non-contrast

Loss Smoothing for Stable Adaptation Under Distribution Shift

ResearchDGX agent

arXiv:2607.00634v1 Announce Type: cross Abstract: In settings such as fine-tuning and reinforcement learning, neural networks are often adapted under distribution shift. Standard adaptation methods ty

Low Perplexity is Repetition: A One-Dimensional Self-Conditioning Attractor in Continuous Diffusion LMs

ResearchDGX agent

arXiv:2607.00588v1 Announce Type: new Abstract: Continuous diffusion language models such as ELF report record-low generative perplexity (Gen-PPL). We find a catch: these models repeat far more than h

LRAT-Catcher: Importing SAT Solver Certificates into Lean4 by Reflection

ResearchDGX agent

arXiv:2607.00815v1 Announce Type: cross Abstract: SAT solvers settle combinatorial problems beyond the reach of interactive theorem provers and produce LRAT certificates for independent verification.

Magic-MM-Embedding: Towards Visual-Token-Efficient Universal Multimodal Embedding with MLLMs

ResearchDGX agent

arXiv:2602.05275v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have shown immense promise in universal multimodal retrieval, which aims to find relevant items of various

Measuring the Gap Between Human and LLM Research Ideas

ResearchDGX agent

arXiv:2607.01233v1 Announce Type: cross Abstract: LLMs are increasingly used to brainstorm research ideas, but existing evaluations mostly judge individual ideas by novelty, feasibility, or expert pre

Memory-Native Non-Terrestrial Networks for Embodied Intelligence

ResearchDGX agent

arXiv:2607.00029v1 Announce Type: cross Abstract: Non-terrestrial networks (NTN) provide ubiquitous connectivity for embodied intelligence (EI), enabling robots in wilderness to leverage cloud resourc

MemoryLLM: Plug-n-Play Interpretable Feed-Forward Memory for Transformers

ResearchDGX agent

Understanding how transformer components operate in LLMs is important, as it is at the core of recent technological advances in artificial intelligence. In this work, we revisit the challenges associa

Message Passing Enables Efficient Reasoning

ResearchDGX agent

arXiv:2607.01077v1 Announce Type: new Abstract: While inference-time scaling has improved the reasoning abilities of large language models (LLMs), the need to generate long chains-of-thought (CoTs) is

MetaHOPE: A Metaphor-Oriented Evaluation Framework for Analysing MT and LLM Translation Errors

ResearchDGX agent

arXiv:2607.00848v1 Announce Type: new Abstract: In this opinion paper, we propose MetaHOPE, an error severity-aware annotation framework for evaluating metaphor translations. Metaphors present challen

MG-RWKV: Multi-Grained Context-Aware RWKV for Temporal Forgery Localization

ResearchDGX agent

arXiv:2607.00902v1 Announce Type: new Abstract: Driven by Artificial Intelligence-Generated Content (AIGC), the authenticity of audio-visual content is facing severe challenges. Temporal Forgery Local

MG-SpaIR: Multi-grade Sparse-guided Implicit Representation for Training-Data-Free Image Restoration

ResearchDGX agent

arXiv:2607.00138v1 Announce Type: new Abstract: MG-SpaIR is a training-data-free framework for restoring a clean image from a single observation corrupted by a mixture of blur, downsampling, noise, an

MineRobot: An Actuator-Centered Kinematic Modeling and Solving Framework for Underground Mining Robots

ResearchDGX agent

arXiv:2603.22055v2 Announce Type: replace-cross Abstract: Underground mining robots are increasingly modeled for planning, operator training, and digital-twin workflows, where reliable actuator-level

Mixture of Distributions Matters: Dynamic Sparse Attention for Efficient Video Diffusion Transformers

ResearchDGX agent

arXiv:2601.11641v3 Announce Type: replace Abstract: While Diffusion Transformers (DiTs) have achieved notable progress in video generation, this long-sequence generation task remains constrained by th

Moire Video Authentication: A Physical Signature Against AI Video Generation

ResearchDGX agent

arXiv:2604.01654v2 Announce Type: replace-cross Abstract: Recent advances in video generation have made AI-synthesized content increasingly difficult to distinguish from real footage. We propose a phy

MonoMSK: Monocular 3D Musculoskeletal Dynamics Estimation

ResearchDGX agent

arXiv:2511.19326v2 Announce Type: replace Abstract: Reconstructing biomechanically realistic 3D human motion - recovering both kinematics (motion) and kinetics (forces) - is a critical challenge. Whil

Multi-Embodiment Robotic Retargeting via Guided Diffusion Model

ResearchDGX agent

arXiv:2505.20857v2 Announce Type: replace Abstract: Motion retargeting for specific robot from existing motion datasets is one critical step in transferring motion patterns from human behaviors to and

MVDGC: Joint 3D and 2D Multi-view Pedestrian Detection via Dual Geometric Constraints

ResearchDGX agent

arXiv:2607.00273v1 Announce Type: new Abstract: The core challenge in multi-view pedestrian detection (MVPD) lies in effective aggregation of visual features from different viewpoints for robust occlu

Neural Certificate Pricing for Combinatorial Optimization Problems

ResearchDGX agent

arXiv:2607.01185v1 Announce Type: new Abstract: Combinatorial optimization (CO) problems are difficult because certifiable discrete structure induces exponential search. One needs to search over the s

NoPA: Non-Parametric Online 3D Scene Graph Generation

ResearchDGX agent

arXiv:2607.00529v1 Announce Type: new Abstract: Classic 3D scene graph generation approaches fail to work in real-time due to the heavy computational cost of environment mapping and the need to genera

On Robustness and Chain-of-Thought Consistency of RL-Finetuned VLMs

ResearchDGX agent

Reinforcement learning (RL) finetuning has become a key technique for enhancing large language models (LLMs) on reasoning-intensive tasks, motivating its extension to vision language models (VLMs). Wh

OnPoint: Offline-to-Online Multi-Level Distillation for Point-Supervised Online Temporal Action Localization

ResearchDGX agent

arXiv:2607.00289v1 Announce Type: new Abstract: Temporal Action Localization (TAL) typically relies on segment annotations or offline access to full videos, limiting scalability and online use. We int

Oops, SIGReg did it again! Large scale (CC12M->Datacomp-L) vision-language JEPA pretraining beats CLIP and SigLIP objectives! Thanks to SIGR…

ResearchDGX agent

Oops, SIGReg did it again! Large scale (CC12M->Datacomp-L) vision-language JEPA pretraining beats CLIP and SigLIP objectives! Thanks to SIGReg, our LeVLJEPA has no collapse, no EMA, no stop-gradient,

OpFML: Pipeline for ML-based Operational Inference

ResearchDGX agent

arXiv:2601.11046v2 Announce Type: replace Abstract: Machine learning models for climate and Earth science are becoming increasingly capable, yet model deployment into operational use remains a largely

Optimal any-angle path planning in static and dynamic environments

ResearchDGX agent

arXiv:2607.00065v1 Announce Type: cross Abstract: Any-angle path planning extends traditional graph-based path planning by allowing movement between any pair of vertices, rather than being restricted

Optimization on the Oblique Manifold for Sparse Simplex Constraints via Multiplicative Updates

ResearchDGX agent

arXiv:2503.24075v4 Announce Type: replace-cross Abstract: Low-rank optimization problems with sparse simplex constraints involve variables that must satisfy nonnegativity, sparsity, and sum-to-1 condi

PanoGrounder: Bridging 2D and 3D with Panoramic Scene Representations for VLM-based 3D Visual Grounding

ResearchDGX agent

arXiv:2512.20907v2 Announce Type: replace Abstract: 3D Visual Grounding (3DVG) is a critical bridge from vision-language perception to robotics, requiring both language understanding and 3D scene reas

← Previous
1…8081828384…320
Next →