AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
25,881 results
23 Jul 2026

RELTA-SGLD: Relative-Growth Localized Taming for Nonconvex Stochastic-Gradient Langevin Learning

ResearchDGX agent

arXiv:2607.19544v1 Announce Type: cross Abstract: We introduce RELTA-SGLD, a taming scheme that stabilizes superlinear stochastic-gradient updates while reducing unnecessary suppression of the origina

Reproducing Recurrent Transformers: The CoTFormer

ResearchDGX agent

arXiv:2607.19405v1 Announce Type: new Abstract: The CoTFormer architecture formalizes Chain-of-Thought as a form of recurrent latent computation, preserving intermediate states as attendable represent

Rethinking Uncertainty Evaluation in Large Language Models

ResearchDGX agent

arXiv:2607.19367v1 Announce Type: new Abstract: Calibration is the primary criterion for evaluating LLM confidence, but it is insufficient: it admits trivially incoherent estimators, depends on the ev

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Reward-Aware Population Scaling of Evolutionary Strategies in LLM Fine-Tuning

ResearchDGX agent

arXiv:2607.19408v1 Announce Type: new Abstract: Using Evolutionary Strategies (ES) for fine-tuning large language models is attractive because it is memory-efficient, parallel, and compatible with bla

Robots Acquire Manipulation Skills in Seconds from a Single Human Video

ResearchDGX agent

arXiv:2607.20033v1 Announce Type: new Abstract: The ability to acquire skills rapidly and effortlessly while retaining those already mastered is essential for robots. However, current methods still re

Robust Multi-View Classification under Noisy Supervision via Global Anchor Consensus

ResearchDGX agent

arXiv:2607.18561v1 Announce Type: cross Abstract: In recent years, multi-view learning has attracted increasing attention, as it integrates the complementary information of heterogeneous views. Most e

RPPNet: Perceptually-Grouped Rhythm-Pitch Primitives for Long-Term Structure Melody Generation via Boundary-Aware Modeling

ResearchDGX agent

arXiv:2607.19776v1 Announce Type: cross Abstract: Existing symbolic music generation models typically use bars as the basic structural unit. However, human perception of musical phrases often does not

SciTrek: Evaluating and Improving Long-Context Numerical Reasoning over Scientific Articles

ResearchDGX agent

arXiv:2509.21028v4 Announce Type: replace Abstract: We introduce SciTrek, a synthetic question-answering dataset for assessing and improving long-context numerical reasoning in large language models (

Semantic Richness or Geometric Reasoning? The Fragility of VLM's Visual Invariance

ResearchDGX agent

arXiv:2604.01848v4 Announce Type: replace Abstract: This work investigates the fundamental fragility of state-of-the-art Vision-Language Models (VLMs) under basic geometric transformations. While mode

Sentence Splitter: Uncovering Latent Factual Structure for Self-Supervised Learning

ResearchDGX agent

arXiv:2607.19845v1 Announce Type: cross Abstract: This paper introduces Sentence Splitter, a self-supervised framework built upon a T5-based encoder--decoder architecture for uncovering the latent fac

Signed Rectified Flow: Negativity-Controlled Generation

ResearchDGX agent

arXiv:2607.18516v1 Announce Type: cross Abstract: We introduce Signed Rectified Flow (Signed RF), a generalization of Rectified Flow that targets the signed measure pi^{sign} = (1+alpha)pi^+ - alphapi

SIINR: Structurally Informed Implicit Neural Representations for super-resolution with uncertainty quantification of clinical quality diffusion MRI datasets

ResearchDGX agent

arXiv:2607.19943v1 Announce Type: new Abstract: Diffusion Magnetic Resonance Imaging (dMRI) is a powerful tool for probing brain microstructure, but clinical acquisitions are often limited by low out-

SoftReason: A Fully Differentiable Neuro-Soft-Symbolic Deductive Reasoning Architecture over High-Dimensional Perceptual Data

ResearchDGX agent

arXiv:2607.20402v1 Announce Type: new Abstract: In many reasoning problems, the premises are not observed as discrete symbols, but must be inferred from high-dimensional inputs. Further, the predicate

Spectral DPPs via NEPv: A Scalable Continuous Relaxation of Determinantal MAP for Diversity-Aware Data Selection

ResearchDGX agent

arXiv:2606.19411v2 Announce Type: replace Abstract: Selecting a small, diverse, high-quality subset from a massive pool of candidates is a recurring primitive in modern machine learning -- data curati

Statistical Early Stopping for Reasoning Models

ResearchDGX agent

arXiv:2602.13935v2 Announce Type: replace Abstract: While LLMs have seen substantial improvement in reasoning capabilities, they also sometimes overthink, generating unnecessary reasoning steps, parti

Strong Gravitational Lensing Posterior Sampling in Pixel-Space Using Diffusion Models and Recurrent Inference Machines

ResearchDGX agent

arXiv:2607.19459v1 Announce Type: cross Abstract: Modeling galaxy-galaxy strong gravitational lenses to infer the brightness of the source galaxy and the mass distribution of the foreground galaxy is

SUPER Module for Detail-Sensitive and Cost-Efficient U-Net Variant Decoders

ResearchDGX agent

arXiv:2511.11015v2 Announce Type: replace Abstract: Skip-connected U-Net variants are widely used for dense inverse problems, yet their decoders commonly recover resolution through spatial upscaling,

Surprise Forcing: What to Remember, When to Skip in Long Video Generation

ResearchDGX agent

arXiv:2607.18436v1 Announce Type: new Abstract: Streaming autoregressive diffusion makes minute-scale video synthesis practical, but its bounded context and fixed denoising schedule allocate resources

SwiftGS: Episodic Priors for Immediate Satellite Surface Recovery

ResearchDGX agent

arXiv:2603.18634v3 Announce Type: replace-cross Abstract: Rapid, large-scale 3D reconstruction from multi-date satellite imagery is vital for environmental monitoring, urban planning, and disaster res

Synthetic and Derived Training Images for Campus Waste Detection: A Multi-Seed Evaluation with YOLOv8n

ResearchDGX agent

arXiv:2607.19535v2 Announce Type: new Abstract: Incorrect disposal can contaminate campus recycling streams, and a bin-mounted camera could provide feedback as an item is discarded. We evaluated wheth

Taming the Security-Energy Paradox: A Green AI Approach to Optimized Android Malware Detection

ResearchDGX agent

arXiv:2607.20003v1 Announce Type: cross Abstract: An increase in advanced Android malware requires the use of deep learning models, which can run on Android devices. But there is a trade-off between s

Team RAS in 11th ABAW Competition: Multimodal Ambivalence Recognition Approach

ResearchDGX agent

arXiv:2607.14702v2 Announce Type: replace Abstract: Automatic recognition of ambivalence and hesitancy is challenging because these states may be expressed through inconsistent linguistic, acoustic, f

Tensor Network Machine Learning for Wildfire Susceptibility Mapping: from Grokking Dynamics to Quantum Mixedness of Class Representations

ResearchDGX agent

arXiv:2607.19503v1 Announce Type: cross Abstract: A quantum-inspired tensor network framework for wildfire susceptibility classification in the Gargano region is introduced, leveraging AlphaEarth embe

Test-Time Training for Modality Order Consistency in Vision-Language Models

ResearchDGX agent

arXiv:2607.20351v1 Announce Type: cross Abstract: We find that vision-language models are sensitive to a specific semantically irrelevant change: the order in which the image and question are presente

Text Template Tokens Are Implicit Semantic Registers in Diffusion Transformers

ResearchDGX agent

arXiv:2607.19139v1 Announce Type: new Abstract: Text-to-image diffusion transformers (DiTs) jointly process text and image tokens, yet their internal computation during denoising remains poorly unders

The C-index illusion: discrimination without calibration in published survival models

ResearchDGX agent

arXiv:2607.19526v1 Announce Type: new Abstract: 'Stop Chasing the C-index when Evaluating Survival Analysis Models' (ICML 2026, Spotlight) argued normatively, on synthetic data, that evaluating surviv

The JEPA Predictor: A Transferable Operator for Occluded Feature Completion

ResearchDGX agent

arXiv:2607.16274v2 Announce Type: replace Abstract: Joint-Embedding Predictive Architectures (JEPAs) train a predictor jointly with their encoder, but downstream deployment discards the predictor and

The Orthogonalized Read Is a Removable Training Scaffold for Recurrent Memory

ResearchDGX agent

arXiv:2607.19390v1 Announce Type: new Abstract: A recent report finds that orthogonalizing the mLSTM memory matrix at read time (five Newton-Schulz iterations, trained through) substantially improves

The PAR dataset: Prostate biopsy whole slide images from an underrepresented Middle Eastern population

ResearchDGX agent

arXiv:2512.03854v2 Announce Type: replace Abstract: Artificial intelligence (AI) is increasingly used in digital pathology. Publicly available histopathology datasets remain scarce, and those that do

The Quadrilateral Loss: Additivity as a Measurable Behavior of Dense Neural Networks

ResearchDGX agent

arXiv:2607.20201v1 Announce Type: cross Abstract: Additive models buy interpretability by forbidding feature interactions, a constraint that neural instantiations enforce architecturally. We introduce

Theory-to-Practice Gap for Neural Networks and Neural Operators

ResearchDGX agent

arXiv:2503.18219v2 Announce Type: replace Abstract: This work studies the sampling complexity of learning with ReLU neural networks and neural operators. For mappings belonging to relevant approximati

Timeripple: Accelerating vDiTs by Understanding the Spatio-Temporal Correlations in Latent Space

ResearchDGX agent

arXiv:2511.12035v2 Announce Type: replace-cross Abstract: The recent surge in video generation has shown the growing demand for high-quality video synthesis using large vision models. Existing video g

Total Variation Distance Estimation in Autoregressive Models

ResearchDGX agent

arXiv:2607.19510v1 Announce Type: new Abstract: Modern LLM deployments use a number of implementation choices and inference optimizations (e.g., batching, custom kernels, and quantization) on top of f

Toward Reliable RGB-D Semantic Segmentation: Handling Missing Modalities via Condition Dropout

ResearchDGX agent

arXiv:2607.20326v1 Announce Type: cross Abstract: RGB-D semantic segmentation has achieved remarkable progress, yet most models assume that RGB and depth are always available. In practice, failures or

Toward Seasonal Guidelines for Robust Deep-Learning Sentinel-2 Building Detection in Different Area Types

ResearchDGX agent

arXiv:2607.19994v2 Announce Type: new Abstract: Sentinel-2 imagery offers open access, global coverage, and frequent revisit times, making it attractive for practical building mapping at scale; howeve

Trace: A Taxonomy-Guided Environment for Multidomain Visual Reasoning

ResearchDGX agent

arXiv:2607.19790v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has substantially improved language-model reasoning, yet its extension to vision-language models r

Trusting What You Cannot See: Auditable Fine-Tuning and Inference for Proprietary AI

ResearchDGX agent

arXiv:2603.07466v2 Announce Type: replace-cross Abstract: Cloud-based infrastructure has become the dominant platform for deploying large models, particularly large language models (LLMs). Fine-tuning

Understanding Developer Pain Points in Federated Learning: Insights from Stack Overflow and GitHub

ResearchDGX agent

arXiv:2607.19621v1 Announce Type: cross Abstract: Federated Learning (FL) enables collaborative model training without centralizing raw data, but building and operating FL systems remains difficult du

User-Centric Modeling of Transactional Sequences with Explainable State Space Models

ResearchDGX agent

arXiv:2607.20228v1 Announce Type: new Abstract: We propose a hybrid approach for user-centric modeling of transactional event sequences that combines contrastive representation learning (CoLES) with S

UVFaceFusion: Fast Multi-view Topologically Consistent Face Reconstruction in the Wild via UV-space Neural Fusion

ResearchDGX agent

arXiv:2607.18798v1 Announce Type: new Abstract: Reconstructing high-fidelity facial geometry with an assigned topology is essential for digital avatar creation and animation, yet existing automated me

V2F: Vision-Informed Grasp Force Prediction for Damage-Aware Robotic Handling of Date Fruits

ResearchDGX agent

arXiv:2607.19804v1 Announce Type: new Abstract: This paper presents a vision-informed grasp force prediction framework for robotic handling of date fruits. Addressing the dual challenge of high detach

Vera: Identity-Faithful Human Subject-to-Video Generation

ResearchDGX agent

arXiv:2607.20247v1 Announce Type: new Abstract: Subject-to-video (S2V) generation has made substantial progress in preserving reference subjects across diverse categories, yet generic subject consiste

VizRAG: Enhancing Retrieval-Augmented Generation with Hypergraph Visualization

ResearchDGX agent

arXiv:2607.19830v1 Announce Type: new Abstract: Hypergraph-based RAG systems surpass traditional graph-based approaches by organizing complex n-ary atomic facts among entities, rather than relying sol

WanSong v1.0 Technical Report

ResearchDGX agent

arXiv:2607.14749v3 Announce Type: replace-cross Abstract: Music generation foundation models have recently attracted significant industry attention. However, achieving efficient generation and high-fi

Wavefront Parallelization for Efficient Learned Image Compression

ResearchDGX agent

arXiv:2607.19082v1 Announce Type: cross Abstract: Autoregressive context models are foundational for learned image compression,but they suffer from slow serial inference. Existing acceleration methods

When Does Consensus Beat Voting? A Critical Analysis of Statistical Label Fusion in Medical Image Segmentation

ResearchDGX agent

arXiv:2607.19402v1 Announce Type: new Abstract: This paper provides a rigorous, self-contained investigation of consensus segmentation. We derive the mathematical foundations from first principles --

World Modeling with JEPA has recently gained traction thanks to a novel anti-collapse mechanism called 'SIGReg' (by @ylecun and @randall_bal…

ResearchDGX agent

World Modeling with JEPA has recently gained traction thanks to a novel anti-collapse mechanism called 'SIGReg' (by @ylecun and @randall_balestr). The math is clean, but it is rarely explained from fi

Zero-Observation User Reactivation with Gap-Driven Dimensional Gating

ResearchDGX agent

arXiv:2607.19802v1 Announce Type: cross Abstract: Sequential recommendation (SR) models capture continuously observed behavior, but a returning user may have no interactions for months or years. We de

22 Jul 2026

Closed source safeguards that infantalize us all and leave American companies defenseless are a menace. Gated access is a menace. Who cares …

ResearchDGX agent

Closed source safeguards that infantalize us all and leave American companies defenseless are a menace. Gated access is a menace. Who cares if 100 companies get to defend themselves because they got o

microsoft/Fara1.5-27B · Hugging Face

Model ReleasesDGX agent

Fara1.5-27B is a multimodal computer use agent (CUA) for web browsers, from Microsoft Research AI Frontiers. It observes the browser through screenshots and acts on the user's behalf by emitting struc

21 Jul 2026

Highly performant open weights frontier models such as Kimi are a competitive threat to OpenAI & Anthropic, but probably for everyone else t…

ResearchDGX agent

Highly performant open weights frontier models such as Kimi are a competitive threat to OpenAI & Anthropic, but probably for everyone else these are a win. Hope more US entities will release top quali

Not because Andrej is saying it but I think voice is goated. And you can mix it with other modalities for even richer prompting. I recorded …

ResearchDGX agent

Not because Andrej is saying it but I think voice is goated. And you can mix it with other modalities for even richer prompting. I recorded a session a few weeks back to demo the power of multimodal p

Open-source is not the cause of the cybersecurity crisis, it's the solution! https://huggingface.co/fdtn-ai

ResearchDGX agent

In July 2026, Clement Delangue tweeted that “Open‑source is not the cause of the cybersecurity crisis, it’s the solution!” linking to the Hugging Face page *fdtn‑ai* (https://huggingface.co/fdtn-ai).

16 Jul 2026

A 3DGS-Driven Dynamic Viewpoint and Vibrotactile Framework for Subsea Teleoperation Validated via fNIRS

ResearchDGX agent

arXiv:2607.13067v1 Announce Type: new Abstract: Teleoperating remotely operated vehicles (ROVs) in flooded, cluttered infrastructure is fundamentally limited by narrow 2D egocentric views and subsea c

A Bayesian framework for the uncanny valley in humanoid robot design

ResearchDGX agent

arXiv:2607.13060v1 Announce Type: new Abstract: The uncanny valley is a long-standing empirical rule in humanoid robot design: making robots more human-like can reduce, rather than increase, affinity.

A Hybrid Mamba for Audio-Visual Navigation

ResearchDGX agent

arXiv:2607.13110v1 Announce Type: cross Abstract: Since the paradigm centered on convolutional neural networks and recurrent architectures was established in 2020, the fundamental backbone networks fo

A novel network for classification of cuneiform tablet metadata

ResearchDGX agent

arXiv:2603.03892v2 Announce Type: replace-cross Abstract: In this paper, we present a network structure for classifying metadata of cuneiform tablets. The problem is of practical importance, as the si

A novel unsupervised machine learning strategy to handle multimodal cardiac PET/MRI data

ResearchDGX agent

arXiv:2607.13936v1 Announce Type: new Abstract: Arrhythmogenic left ventricular cardiomyopathy is a genetic myocardial disease difficult to diagnose due to the lack of gold standard criteria. Simultan

A Space-Time Transformer for Precipitation Nowcasting

ResearchDGX agent

arXiv:2511.11090v3 Announce Type: replace Abstract: Until recently, numerical weather prediction (NWP) models have stood rivalless in operational forecasting despite a few limitations. Namely, physica

Accuracy-Preserving Stability Regularization for Large-Scale Retail Demand Forecasting

ResearchDGX agent

arXiv:2607.13331v1 Announce Type: new Abstract: Retail demand forecasts are reused across replenishment, capacity, labor, and transportation planning cycles. Point-error objectives do not constrain ab

← Previous
1…104105106107108…432
Next →