AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
25,881 results
21 Apr 2026

全网都在吹的LeCun新论文,90%的解读都是错的。 他们说生成式AI是死路,说过去三年花的几百亿全白费了,说15M参数的小模型就能吊打万亿大模型。 这些全是营销号的夸张, 我觉得这篇论文的真正分量比他们吹的还要重。 Yann LeCun团队这次解决了JEPA困扰了好几年的表征坍…

ResearchDGX agent

全网都在吹的LeCun新论文,90%的解读都是错的。 他们说生成式AI是死路,说过去三年花的几百亿全白费了,说15M参数的小模型就能吊打万亿大模型。 这些全是营销号的夸张, 我觉得这篇论文的真正分量比他们吹的还要重。 Yann LeCun团队这次解决了JEPA困扰了好几年的表征坍缩问题。 以前的世界模型,学着学着就会把狗车人都压成一模一样的向量,什么都学不到。 这次他们只加了一个极其优雅的数学正则

LEPO: nderline{L}atent Rnderline{e}asoning nderline{P}olicy nderline{O}ptimization for Large Language~Models

ResearchDGX agent

arXiv:2604.17892v1 Announce Type: new Abstract: Recently, latent reasoning has been introduced into large language models (LLMs) to leverage rich information within a continuous space. However, withou

Leveraging Kernel Symmetry for Joint Compression and Error Mitigation in Edge Model Transfer

ResearchDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.17371v1 Announce Type: cross Abstract: This paper investigates communication-efficient neural network transmission by exploiting structured symmetry constraints in convolutional kernels. In

Leveraging VR Robot Games to Facilitate Data Collection for Embodied Intelligence Tasks

ResearchDGX agent

arXiv:2604.16903v1 Announce Type: new Abstract: Collecting embodied interaction data at scale remains costly and difficult due to the limited accessibility of conventional interfaces. We present a gam

Light-Adapted Electroretinogram and Oscillatory Potentials (LEOPs) Dataset for Autism Spectrum Disorder and Typically Developing Individuals

ResearchDGX agent

arXiv:2604.16981v1 Announce Type: cross Abstract: The LEOPs (Light-ERG-Oscillatory Potentials) dataset provides light-adapted (LA) electroretinogram (ERG) and Oscillatory Potentials (OPs) waveforms fo

Lil: Less is Less When Applying Post-Training Sparse-Attention Algorithms in Long-Decode Stage

ResearchDGX agent

arXiv:2601.03043v3 Announce Type: replace Abstract: Large language models (LLMs) demonstrate strong capabilities across a wide range of complex tasks and are increasingly deployed at scale, placing si

Linear-Time and Constant-Memory Text Embeddings Based on Recurrent Language Models

ResearchDGX agent

arXiv:2604.18199v1 Announce Type: new Abstract: Transformer-based embedding models suffer from quadratic computational and linear memory complexity, limiting their utility for long sequences. We propo

Linking Exteroception and Proprioception through Improved Contact Modeling for Soft Growing Robots

ResearchDGX agent

arXiv:2507.10694v2 Announce Type: replace Abstract: Passive deformation due to compliance is a commonly used benefit of soft robots, providing opportunities to achieve robust actuation with few active

LLaVA-Octopus: Unlocking Instruction-Driven Adaptive Projector Fusion for Video Understanding

ResearchDGX agent

arXiv:2501.05067v3 Announce Type: replace Abstract: In this paper, we introduce LLaVA-Octopus, a novel video multimodal large language model. LLaVA-Octopus adaptively weights features from different v

LLM as Graph Kernel: Rethinking Message Passing on Text-Rich Graphs

ResearchDGX agent

arXiv:2603.14937v2 Announce Type: replace-cross Abstract: Text-rich graphs, which integrate complex structural dependencies with abundant textual information, are ubiquitous yet remain challenging for

LLM-AUG: Robust Wireless Data Augmentation with In-Context Learning in Large Language Models

ResearchDGX agent

arXiv:2604.17770v1 Announce Type: new Abstract: Data scarcity remains a fundamental bottleneck in applying deep learning to wireless communication problems, particularly in scenarios where collecting

LLM Hypnosis: Exploiting User Feedback for Unauthorized Knowledge Injection to All Users

ResearchDGX agent

arXiv:2507.02850v3 Announce Type: replace Abstract: We describe a vulnerability in language models (LMs) trained with user feedback, whereby a single user can persistently alter LM knowledge and behav

LLMAR: A Tuning-Free Recommendation Framework for Sparse and Text-Rich Industrial Domains

ResearchDGX agent

arXiv:2604.16379v1 Announce Type: cross Abstract: Industrial B2B applications (e.g., construction site risk prediction, material procurement) face extreme data sparsity yet feature rich textual intera

LLMs can persuade only psychologically susceptible humans on societal issues, via trust in AI and emotional appeals, amid logical fallacies

ResearchDGX agent

arXiv:2604.16935v1 Announce Type: cross Abstract: Scarce longitudinal evidence examines LLMs' persuasiveness and humanness along time-evolving psychological frameworks. We introduce Talk2AI, a longitu

Local Inconsistency Resolution: The Interplay between Attention and Control in Probabilistic Models

ResearchDGX agent

arXiv:2604.17140v1 Announce Type: cross Abstract: We present a generic algorithm for learning and approximate inference with an intuitive epistemic interpretation: iteratively focus on a subset of the

LogicDiff: Logic-Guided Denoising Improves Zero-Shot Reasoning in Masked Diffusion Language Models

ResearchDGX agent

arXiv:2603.26771v2 Announce Type: replace Abstract: Masked diffusion language models (MDLMs) generate text by iteratively unmasking tokens from a fully masked sequence. Their standard confidence-based

LoRaQ: Optimized Low Rank Approximation for 4-bit Quantization

ResearchDGX agent

arXiv:2604.18117v1 Announce Type: new Abstract: Post-training quantization (PTQ) is essential for deploying large diffusion transformers on resource-constrained hardware, but aggressive 4-bit quantiza

LoReC: Rethinking Large Language Models for Graph Data Analysis

ResearchDGX agent

arXiv:2604.17897v1 Announce Type: new Abstract: The advent of Large Language Models (LLMs) has fundamentally reshaped the way we interact with graphs, giving rise to a new paradigm called GraphLLM. As

Love this work from Aksel and the post-training team at Hugging Face! Turns out the HF ecosystem (papers, datasets, models all accessible th…

Model ReleasesDGX agent

Love this work from Aksel and the post-training team at Hugging Face! Turns out the HF ecosystem (papers, datasets, models all accessible through CLI, skills and md files) is perfect for running SOTA

Low Light Image Enhancement Challenge at NTIRE 2026

ResearchDGX agent

arXiv:2604.17669v1 Announce Type: new Abstract: This paper presents a comprehensive review of the NTIRE 2026 Low Light Image Enhancement Challenge, highlighting the proposed solutions and final result

Lower Bounds and Proximally Anchored SGD for Non-Convex Minimization Under Unbounded Variance

ResearchDGX agent

arXiv:2604.16620v1 Announce Type: new Abstract: Analysis of Stochastic Gradient Descent (SGD) and its variants typically relies on the assumption of uniformly bounded variance, a condition that freque

LQM: Linguistically Motivated Multidimensional Quality Metrics for Machine Translation

ResearchDGX agent

arXiv:2604.18490v1 Announce Type: new Abstract: Existing MT evaluation frameworks, including automatic metrics and human evaluation schemes such as Multidimensional Quality Metrics (MQM), are largely

LTRR: Learning To Rank Retrievers for LLMs

ResearchDGX agent

arXiv:2506.13743v2 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) systems typically rely on a single fixed retriever, despite growing evidence that no single retriever performs

LVLMs and Humans Ground Differently in Referential Communication

ResearchDGX agent

arXiv:2601.19792v3 Announce Type: replace Abstract: For generative AI agents to partner effectively with human users, the ability to accurately predict human intent is critical. But this ability to co

Machine Learning Based Prediction of Proton Conductivity in Metal-Organic Frameworks

ResearchDGX agent

arXiv:2407.09514v3 Announce Type: replace-cross Abstract: Recently, metal-organic frameworks (MOFs) have demonstrated their potential as solid-state electrolytes in proton exchange membrane fuel cells

Mapping Election Toxicity on Social Media across Issue, Ideology, and Psychosocial Dimensions

ResearchDGX agent

arXiv:2604.16765v1 Announce Type: cross Abstract: Online political hostility is pervasive, yet it remains unclear how toxicity varies across campaign issues and political ideology, and what psychosoci

MARA: A Multimodal Adaptive Retrieval-Augmented Framework for Document Question Answering

ResearchDGX agent

arXiv:2604.16313v1 Announce Type: cross Abstract: Retrieval-based multimodal document QA aims to identify and integrate relevant information from visually rich documents with complex multimodal struct

Matched-Learning-Rate Analysis of Attention Drift and Transfer Retention in Fine-Tuned CLIP

ResearchDGX agent

arXiv:2604.16410v1 Announce Type: new Abstract: CLIP adaptation can improve in-domain accuracy while degrading out-of-domain transfer, but comparisons between Full Fine-Tuning (Full FT) and LoRA are o

Medial Axis Aware Learning of Signed Distance Functions

ResearchDGX agent

arXiv:2604.16512v1 Announce Type: new Abstract: We propose a novel variational method to compute a highly accurate global signed distance function (SDF) to a given point cloud. To this end, the jump s

MegaRAG: Multimodal Knowledge Graph-Based Retrieval Augmented Generation

ResearchDGX agent

arXiv:2512.20626v2 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) enables large language models (LLMs) to dynamically access external information, which is powerful for an

MetaCloak-JPEG: JPEG-Robust Adversarial Perturbation for Preventing Unauthorized DreamBooth-Based Deepfake Generation

ResearchDGX agent

arXiv:2604.18537v1 Announce Type: new Abstract: The rapid progress of subject-driven text-to-image synthesis, and in particular DreamBooth, has enabled a consent-free deepfake pipeline: an adversary n

Mitigating Multimodal Hallucination via Phase-wise Self-reward

ResearchDGX agent

arXiv:2604.17982v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) still struggle with vision hallucination, where generated responses are inconsistent with the visual input. Exist

MLE-UVAD: Minimal Latent Entropy Autoencoder for Fully Unsupervised Video Anomaly Detection

ResearchDGX agent

arXiv:2603.23868v2 Announce Type: replace Abstract: In this paper, we address the challenging problem of single-scene, fully unsupervised video anomaly detection (VAD), where raw videos containing bot

Modeling Biomechanical Constraint Violations for Language-Agnostic Lip-Sync Deepfake Detection

ResearchDGX agent

arXiv:2604.16808v1 Announce Type: new Abstract: Current lip-sync deepfake detectors rely on pixel-level artifacts or audio-visual correspondence, failing to generalize across languages because these c

Modeling, Control and Self-sensing of Dielectric Elastomer Soft Actuators: A Review

ResearchDGX agent

arXiv:2604.17199v1 Announce Type: new Abstract: Dielectric elastomer actuators (DEAs) have garnered extensive attention especially in soft robotic applications over the past few decades owing to the a

MoE-nD: Per-Layer Mixture-of-Experts Routing for Multi-Axis KV Cache Compression

ResearchDGX agent

arXiv:2604.17695v1 Announce Type: cross Abstract: KV cache memory is the dominant bottleneck for long-context LLM inference. Existing compression methods each act on a single axis of the four-dimensio

Multi-Beholder: Biomarker Prediction for Low-Grade Glioma with Multiple Instance Learning and One-Class Classification

ResearchDGX agent

arXiv:2310.07464v2 Announce Type: replace-cross Abstract: Biomarker detection is an indispensable part of the diagnosis and treatment of low-grade glioma (LGG). However, current LGG biomarker detectio

Multi-Label Phase Diagram Prediction in Complex Alloys via Physics-Informed Graph Attention Networks

ResearchDGX agent

arXiv:2604.16468v1 Announce Type: new Abstract: Accurate phase equilibria are foundational to alloy design because they encode the underlying thermodynamics governing stability, transformations, and p

Multi-Scale Reversible Chaos Game Representation: A Unified Framework for Sequence Classification

ResearchDGX agent

arXiv:2604.18477v1 Announce Type: new Abstract: Biological classification with interpretability remains a challenging task. For this, we introduce a novel encoding framework, Multi-Scale Reversible Ch

Multi-stage Planning for Multi-target Surveillance using Aircrafts Equipped with Synthetic Aperture Radars Aware of Target Visibility

ResearchDGX agent

arXiv:2604.16962v1 Announce Type: new Abstract: Generating trajectories for synthetic aperture radar (SAR)-equipped aircraft poses significant challenges due to terrain constraints, and the need for s

Multilingual Training and Evaluation Resources for Vision-Language Models

ResearchDGX agent

arXiv:2604.18347v1 Announce Type: new Abstract: Vision Language Models (VLMs) achieved rapid progress in the recent years. However, despite their growth, VLMs development is heavily grounded on Englis

Multimodal Fusion of Histopathology Images and Electronic Health Records for Early Breast Cancer Diagnosis

ResearchDGX agent

arXiv:2604.17122v1 Announce Type: new Abstract: Breast cancer is a leading cause of cancer-related mortality worldwide, and timely accurate diagnosis is critical to improving survival outcomes. While

MUSEG: Reinforcing Video Temporal Understanding via Timestamp-Aware Multi-Segment Grounding

ResearchDGX agent

arXiv:2505.20715v2 Announce Type: replace-cross Abstract: Video temporal understanding is crucial for multimodal large language models (MLLMs) to reason over events in videos. Despite recent advances

MuSteerNet: Human Reaction Generation from Videos via Observation-Reaction Mutual Steering

ResearchDGX agent

arXiv:2603.20187v2 Announce Type: replace Abstract: Video-driven human reaction generation aims to synthesize 3D human motions that directly react to observed video sequences, which is crucial for bui

Negative Momentum for Convex-Concave Optimization

ResearchDGX agent

arXiv:2604.17145v1 Announce Type: cross Abstract: This paper revisits momentum in the context of min-max optimization. Momentum is a celebrated mechanism for accelerating gradient dynamics in settings

Neural Adjoint Method for Meta-optics: Accelerating Volumetric Inverse Design via Fourier Neural Operators

ResearchDGX agent

arXiv:2604.17425v1 Announce Type: new Abstract: Meta-optics promises compact, high-performance imaging and color routing. However, designing high-performance structures is a high-dimensional optimizat

Neural Network-Based Adaptive Event-Triggered Control for Dual-Arm Unmanned Aerial Manipulator Systems

ResearchDGX agent

arXiv:2604.17048v1 Announce Type: new Abstract: This paper investigates the control problem of dual-arm unmanned aerial manipulator systems (DAUAMs). Strong coupling between the dual-arm and the multi

Neural Network-Based Score Estimation in Diffusion Models: Optimization and Generalization

ResearchDGX agent

arXiv:2401.15604v4 Announce Type: replace Abstract: Diffusion models have become a leading paradigm in generative AI, with score estimation via denoising score matching as a central component. While r

Neural Operator: Is data all you need to model the world? An insight into the paradigm of data-driven scientific ML

ResearchDGX agent

arXiv:2301.13331v3 Announce Type: replace-cross Abstract: Numerical approximations of partial differential equations (PDEs) are routinely employed to formulate the solution of physics, engineering, an

Neuroscience Inspired Graph Operators Towards Edge-Deployable Virtual Sensing for Irregular Geometries

ResearchDGX agent

arXiv:2604.16722v1 Announce Type: new Abstract: Predicting full-field physics through the real-time virtual sensing of engineering systems can enhance limited physical sensors but often requires spars

New Fourth-Order Grayscale Indicator-Based Telegraph Diffusion Model for Image Despeckling

ResearchDGX agent

arXiv:2509.26010v2 Announce Type: replace Abstract: Second-order PDE models have been widely used for suppressing multiplicative noise, but they often introduce blocky artifacts in the early stages of

NI Sampling: Accelerating Discrete Diffusion Sampling by Token Order Optimization

ResearchDGX agent

arXiv:2604.18471v1 Announce Type: new Abstract: Discrete diffusion language models (dLLMs) have recently emerged as a promising alternative to traditional autoregressive approaches, offering the flexi

No One Fits All: From Fixed Prompting to Learned Routing in Multilingual LLMs

ResearchDGX agent

arXiv:2604.16937v1 Announce Type: new Abstract: Translation-based prompting is widely used in multilingual LLMs, yet its effectiveness varies across languages and tasks. We evaluate prompting strategi

No-Worse Context-Aware Decoding: Preventing Neutral Regression in Context-Conditioned Generation

ResearchDGX agent

arXiv:2604.16686v1 Announce Type: new Abstract: Large language models (LLMs) can answer questions and summarize documents when conditioned on external contexts (e.g., retrieved evidence), yet context

Non-Stationarity in the Embedding Space of Time Series Foundation Models

ResearchDGX agent

arXiv:2604.16428v1 Announce Type: new Abstract: Time series foundation models (TSFMs) are widely used as generic feature extractors, yet the notion of non-stationarity in their embedding spaces remain

NPCNet: Navigator-Driven Pseudo Text for Deep Clustering of Early Sepsis Phenotyping

ResearchDGX agent

arXiv:2602.03562v2 Announce Type: replace Abstract: Electronic Health Records (EHRs) provide high-dimensional temporal data essential for patient modeling; however, conventional algorithmic approaches

OASIS: On-Demand Hierarchical Event Memory for Streaming Video Reasoning

ResearchDGX agent

arXiv:2604.17052v1 Announce Type: new Abstract: Streaming video reasoning requires models to operate in a setting where history grows without bound while meaningful evidence remains scarce. In such a

OD3: Optimization-free Dataset Distillation for Object Detection

ResearchDGX agent

arXiv:2506.01942v2 Announce Type: replace Abstract: Training large neural networks on large-scale datasets requires substantial computational resources, particularly for dense prediction tasks such as

On Different Notions of Redundancy in Conditional-Independence-Based Discovery of Graphical Models

ResearchDGX agent

arXiv:2502.08531v3 Announce Type: replace Abstract: Conditional-independence-based discovery uses statistical tests to identify a graphical model that represents the independence structure of variable

On the Interpolation Effect of Score Smoothing in Diffusion Models

ResearchDGX agent

arXiv:2502.19499v3 Announce Type: replace Abstract: Diffusion models have achieved remarkable progress in various domains with an intriguing ability to produce new data that do not exist in the traini

← Previous
1…316317318319320…432
Next →