AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “papers”

GridTimelineEvolution
12,153 results
4 May 2026

AIDA-ReID: Adaptive Intermediate Domain Adaptation for Generalizable and Source-Free Person Re-Identification

ResearchDGX agent

arXiv:2605.00111v1 Announce Type: new Abstract: Person re-identification (Re-ID) aims to match images of the same individual across non-overlapping camera views and remains challenging due to domain s

AirFM-DDA: Air-Interface Foundation Model in the Delay-Doppler-Angle Domain for AI-Native 6G

Local AiDGX agent

arXiv:2605.00020v1 Announce Type: new Abstract: The success of large foundation models is catalyzing a new paradigm for AI-native 6G network design: wireless foundation models for physical layer desig

Alethia: A Foundational Encoder for Voice Deepfakes

Model ReleasesDGX agent

arXiv:2605.00251v1 Announce Type: cross Abstract: Existing voice deepfake detection and localization models rely heavily on representations extracted from speech foundation models (SFMs). However, dow

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Batch Normalization for Neural Networks on Complex Domains

ResearchDGX agent

arXiv:2605.00467v1 Announce Type: new Abstract: Riemannian neural networks have proven effective in solving a variety of machine learning tasks. The key to their success lies in the development of pri

Beyond Heuristics: Learnable Density Control for 3D Gaussian Splatting

Model ReleasesDGX agent

arXiv:2605.00408v1 Announce Type: new Abstract: While 3D Gaussian Splatting (3DGS) has demonstrated impressive real-time rendering performance, its efficacy remains constrained by a reliance on heuris

Broadband Wide Field of View Imaging with Computational Mirrors

ResearchDGX agent

arXiv:2605.00029v1 Announce Type: cross Abstract: Traditional glass-based optics are typically optimized for narrow spectral bands, such as the visible (400-700nm) or shortwave infrared (1000-1800nm).

Build, Judge, Optimize: A Blueprint for Continuous Improvement of Multi-Agent Consumer Assistants

AgentsDGX agent

arXiv:2603.03565v2 Announce Type: replace-cross Abstract: Conversational shopping assistants (CSAs) represent a compelling application of agentic AI, but moving from prototype to production reveals tw

Causality-enhanced Decision-Making for Autonomous Mobile Robots in Dynamic Environments

AgentsDGX agent

arXiv:2504.11901v5 Announce Type: replace Abstract: The growing integration of robots in shared environments-such as warehouses, shopping centres, and hospitals-demands a deep understanding of the und

Comparative Analysis of Polygon-Based and Global Machine Learning Models for Bus Occupancy Prediction

Local AiDGX agent

arXiv:2605.00083v1 Announce Type: new Abstract: Accurate forecasting of bus ridership (passengers numbers) is crucial for efficient management and optimization of public transport systems. Traditional

Deepfakes: we need to re-think the concept of 'real' images

Model ReleasesDGX agent

arXiv:2509.21864v2 Announce Type: replace Abstract: The wide availability and low usability barrier of modern image generation models has triggered the reasonable fear of criminal misconduct and negat

Diffusion Models for Solving Inverse Problems via Posterior Sampling with Piecewise Guidance

ResearchDGX agent

arXiv:2507.18654v2 Announce Type: replace-cross Abstract: Diffusion models are powerful tools for sampling from high-dimensional distributions by progressively transforming pure noise into structured

Directed Social Regard: Surfacing Targeted Advocacy, Opposition, Aid, Harms, and Victimization in Online Media

ResearchDGX agent

arXiv:2605.00776v1 Announce Type: new Abstract: The language in online platforms, influence operations, and political rhetoric frequently directs a mix of pro-social sentiment (e.g., advocacy, helpful

Discrete Cosine Transform Based Decorrelated Attention for Vision Transformers

ResearchDGX agent

arXiv:2405.13901v4 Announce Type: replace Abstract: Self-attention is central to the success of Transformer architectures; however, learning the query, key, and value projections from random initializ

Distance metric learning for conditional anomaly detection

ResearchDGX agent

arXiv:2605.00490v1 Announce Type: new Abstract: Anomaly detection methods can be very useful in identifying unusual or interesting patterns in data. A recently proposed conditional anomaly detection f

DMDSC: A Dynamic-Margin Deep Simplex Classifier for Open-Set Recognition on Medical Image Datasets

ResearchDGX agent

arXiv:2605.00675v1 Announce Type: new Abstract: Medical imaging datasets are often characterized by extreme class imbalances, where rare pathologies are significantly underrepresented compared to comm

Do Open-Loop Metrics Predict Closed-Loop Driving? A Cross-Benchmark Correlation Study of NAVSIM and Bench2Drive

Model ReleasesDGX agent

arXiv:2605.00066v1 Announce Type: new Abstract: Open-loop evaluation offers fast, reproducible assessment of autonomous driving planners, but its ability to predict real closed-loop driving performanc

DPU or GPU for Accelerating Neural Networks Inference -- Why not both? Split CNN Inference

HardwareDGX agent

arXiv:2605.00174v1 Announce Type: cross Abstract: Video and image streaming on edge devices requires low latency. To address this, Neural Networks (NNs) are widely used, and prior work mainly focuses

Efficient Spatio-Temporal Vegetation Pixel Classification with Vision Transformers

Model ReleasesDGX agent

arXiv:2605.00296v1 Announce Type: new Abstract: Plant phenology-the study of recurrent life cycle events-is essential for understanding ecosystem dynamics and their responses to climate change impacts

EGREFINE: An Execution-Grounded Optimization Framework for Text-to-SQL Schema Refinement

Local AiDGX agent

arXiv:2605.00628v1 Announce Type: cross Abstract: Text-to-SQL enables non-expert users to query databases in natural language, yet real-world schemas often suffer from ambiguous, abbreviated, or incon

Evaluating the Architectural Reasoning Capabilities of LLM Provers via the Obfuscated Natural Number Game

Model ReleasesDGX agent

arXiv:2605.00677v1 Announce Type: new Abstract: While Large Language Models have achieved notable success on formal mathematics benchmarks such as MiniF2F, it remains unclear whether these results ste

Exploring the Limits of End-to-End Feature-Affinity Propagation for Single-Point Supervised Infrared Small Target Detection

ResearchDGX agent

arXiv:2605.00722v1 Announce Type: new Abstract: Single-point supervised infrared small target detection (IRSTD) drastically reduces dense annotation costs. Current state-of-the-art (SOTA) methods achi

From Backward Spreading to Forward Replay: Revisiting Target Construction in LLM Parameter Editing

Model ReleasesDGX agent

arXiv:2605.00358v1 Announce Type: new Abstract: LLM parameter editing methods commonly rely on computing an ideal target hidden-state at a target layer (referred as anchor point) and distributing the

Generative Modeling under Non-Monotone MAR Missingness via Approximate Wasserstein Gradient Flows

ResearchDGX agent

arXiv:2604.04567v2 Announce Type: replace-cross Abstract: The prevalence of missing values in data science poses a substantial risk to any further analyses. Despite a wealth of research, principled no

Geometric analysis of attractor boundaries and storage capacity limits in kernel Hopfield networks

ApplicationsDGX agent

arXiv:2605.00366v1 Announce Type: cross Abstract: High-capacity associative memories based on Kernel Logistic Regression (KLR) exhibit strong storage capabilities, but the dynamical and geometric mech

GOR-IS: 3D Gaussian Object Removal in the Intrinsic Space

ApplicationsDGX agent

arXiv:2605.00498v1 Announce Type: new Abstract: Recent advances in Neural Radiance Fields (NeRF) and 3D Gaussian Splatting (3DGS) have made it standard practice to reconstruct 3D scenes from multi-vie

Gradient Regularized Newton Boosting Trees with Global Convergence

ResearchDGX agent

arXiv:2605.00581v1 Announce Type: cross Abstract: Gradient Boosting Decision Trees (GBDTs) dominate tabular machine learning, with modern implementations like XGBoost, LightGBM, and CatBoost being bas

High-Probability Convergence in Decentralized Stochastic Optimization with Gradient Tracking

Model ReleasesDGX agent

arXiv:2605.00281v1 Announce Type: new Abstract: We study high-probability (HP) convergence guarantees in decentralized stochastic optimization, where multiple agents collaborate to jointly train a mod

How Well Does GPT-4o Understand Vision? Evaluating Multimodal Foundation Models on Standard Computer Vision Tasks

Model ReleasesDGX agent

arXiv:2507.01955v3 Announce Type: replace Abstract: Multimodal foundation models (MFMs), such as GPT-4o, have recently made remarkable progress. However, their detailed visual understanding beyond que

Human-in-the-Loop Meta Bayesian Optimization for Fusion Energy and Scientific Applications

ResearchDGX agent

arXiv:2605.00068v1 Announce Type: new Abstract: Inertial Confinement Fusion (ICF) holds transformative promise for sustainable, near-limitless clean energy, yet remains constrained by prohibitively hi

I want to tell you the story of a young woman who you have probably never heard of. Her name is Mary Anne. She was born on a remote island i…

IndustryDGX agent

I want to tell you the story of a young woman who you have probably never heard of. Her name is Mary Anne. She was born on a remote island in Scotland, where life was harsh and unforgiving. On May 2,

Information-geometric adaptive sampling for graph diffusion

ResearchDGX agent

arXiv:2605.00250v1 Announce Type: cross Abstract: Standard diffusion models for graph generation typically rely on uniform time-stepping, an approach that overlooks the non-homogeneous dynamics of dis

InpaintSLat: Inpainting Structured 3D Latents via Initial Noise Optimization

SafetyDGX agent

arXiv:2605.00664v1 Announce Type: new Abstract: We present a training-free approach for controllable 3D inpainting based on initial noise optimization. In the structured 3D latent diffusion framework,

Introducing nanowhale 🐳! A tiny DeepSeek model fully pretrained by an agent. Inspired by @karpathy's nanochat, we gave ml-intern the task o…

Model ReleasesDGX agent

Introducing nanowhale 🐳! A tiny DeepSeek model fully pretrained by an agent. Inspired by @karpathy's nanochat, we gave ml-intern the task of training a tiny MoE with all the architectural advancements

Introducing WARM-VR: Benchmark Dataset for Multimodal Wearable Affect Recognition in Virtual Reality

Model ReleasesDGX agent

arXiv:2605.00184v1 Announce Type: new Abstract: With the growing integration of human-computer interaction into everyday life, advances in machine learning have enabled systems to better perceive and

Koopman-Assisted Reinforcement Learning

ResearchDGX agent

arXiv:2403.02290v2 Announce Type: replace-cross Abstract: The Bellman equation and its continuous form, the Hamilton-Jacobi-Bellman equation, are ubiquitous in reinforcement learning and control theor

Last-Iterate Analyses of FTRL with the 1/2-Tsallis Entropy in Stochastic Bandits

ResearchDGX agent

arXiv:2510.22819v2 Announce Type: replace Abstract: The convergence analysis of online learning algorithms is central to machine learning theory, where the last-iterate convergence is particularly imp

Learn where to Click from Yourself: On-Policy Self-Distillation for GUI Grounding

SafetyDGX agent

arXiv:2605.00642v1 Announce Type: cross Abstract: Graphical User Interface (GUI) grounding maps natural language instructions to the visual coordinates of target elements and serves as a core capabili

Lightweight Domain Adaptation of a Large Language Model for Legal Assistance in the Indian Context

Model ReleasesDGX agent

arXiv:2505.22003v2 Announce Type: replace Abstract: In India, access to legal assistance for the general public has been observed to have a critical gap, as many citizens are not able to take full adv

LIMSSR: LLM-Driven Sequence-to-Score Reasoning under Training-Time Incomplete Multimodal Observations

ApplicationsDGX agent

arXiv:2605.00434v1 Announce Type: new Abstract: Real-world multimodal learning is often hindered by missing modalities. While Incomplete Multimodal Learning (IML) has gained traction, existing methods

LLM-Oriented Information Retrieval: A Denoising-First Perspective

Model ReleasesDGX agent

arXiv:2605.00505v1 Announce Type: cross Abstract: Modern information retrieval (IR) is no longer consumed primarily by humans but increasingly by large language models (LLMs) via retrieval-augmented g

Memory in the LLM Era: Modular Architectures and Strategies in a Unified Framework

AgentsDGX agent

arXiv:2604.01707v2 Announce Type: replace Abstract: Memory emerges as the core module in the large language model (LLM)-based agents for long-horizon complex tasks (e.g., multi-turn dialogue, game pla

MMAudioReverbs: Video-Guided Acoustic Modeling for Dereverberation and Room Impulse Response Estimation

ResearchDGX agent

arXiv:2605.00431v1 Announce Type: cross Abstract: Although recent video-to-audio (V2A) models excelled at synthesizing semantically plausible sounds from visual inputs, they do not explicitly model ro

Model-Based Reinforcement Learning with Double Oracle Efficiency in Policy Optimization and Offline Estimation

SafetyDGX agent

arXiv:2605.00393v1 Announce Type: new Abstract: Reinforcement learning (RL) in large environments often suffers from severe computational bottlenecks, as conventional regret minimization algorithms re

Modeling Subjective Urban Perception with Human Gaze

ResearchDGX agent

arXiv:2605.00764v1 Announce Type: new Abstract: Urban perception describes how people subjectively evaluate urban environments, shaping how cities are experienced and understood. Existing computationa

Near-optimal and Efficient First-Order Algorithm for Multi-Task Learning with Shared Linear Representation

ResearchDGX agent

arXiv:2605.00473v1 Announce Type: new Abstract: Multi-task learning (MTL) has emerged as a pivotal paradigm in machine learning by leveraging shared structures across multiple related tasks. Despite i

Physically Native World Models: A Hamiltonian Perspective on Generative World Modeling

AgentsDGX agent

arXiv:2605.00412v1 Announce Type: cross Abstract: World models have recently re-emerged as a central paradigm for embodied intelligence, robotics, autonomous driving, and model-based reinforcement lea

PhysiGen: Integrating Collision-Aware Physical Constraints for High-Fidelity Human-Human Interaction Generation

ResearchDGX agent

arXiv:2605.00517v1 Announce Type: new Abstract: Despite substantial progress in text-driven 3D human motion synthesis, generating realistic multi-person interaction sequences remains challenging. Nota

PORTool: Importance-Aware Policy Optimization with Rewarded Tree for Multi-Tool-Integrated Reasoning

SafetyDGX agent

arXiv:2510.26020v2 Announce Type: replace Abstract: Multi-tool-integrated reasoning enables LLM-empowered tool-use agents to solve complex tasks by interleaving natural-language reasoning with calls t

PPLLaVA: Varied Video Sequence Understanding With Prompt Guidance

SafetyDGX agent

arXiv:2411.02327v4 Announce Type: replace Abstract: In the past year, video-based large language models (Video LLMs) have achieved impressive progress, particularly in their ability to process long vi

Probing Multimodal Large Language Models on Cognitive Biases in Chinese Short-Video Misinformation

Model ReleasesDGX agent

arXiv:2601.06600v2 Announce Type: replace Abstract: Short-video platforms have become major channels for misinformation, where deceptive claims frequently leverage visual experiments and social cues.

Quantum Interval Bound Propagation for Certified Training of Quantum Neural Networks

ResearchDGX agent

arXiv:2605.00747v1 Announce Type: cross Abstract: Quantum machine learning is a promising field for efficiently learning features of a dataset to perform a specified task, such as classification. Inte

Representation in large language models

ResearchDGX agent

arXiv:2501.00885v2 Announce Type: replace Abstract: The extraordinary success of recent Large Language Models (LLMs) on a diverse array of tasks has led to an explosion of scientific and philosophical

ResRL: Boosting LLM Reasoning via Negative Sample Projection Residual Reinforcement Learning

SafetyDGX agent

arXiv:2605.00380v1 Announce Type: cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) enhances reasoning of Large Language Models (LLMs) but usually exhibits limited generation diver

Rethinking LLM Ensembling from the Perspective of Mixture Models

ResearchDGX agent

arXiv:2605.00419v1 Announce Type: cross Abstract: Model ensembling is a well-established technique for improving the performance of machine learning models. Conventionally, this involves averaging the

Revealing graph bandits for maximizing local influence

ResearchDGX agent

arXiv:2605.00489v1 Announce Type: new Abstract: We study a graph bandit setting where the objective of the learner is to detect the most influential node of a graph by requesting as little information

Reward Modeling from Natural Language Human Feedback

ResearchDGX agent

arXiv:2601.07349v3 Announce Type: replace Abstract: Reinforcement Learning with Verifiable reward (RLVR) on preference data has become the mainstream approach for training Generative Reward Models (GR

Scaling Federated Linear Contextual Bandits via Sketching

ApplicationsDGX agent

arXiv:2605.00500v1 Announce Type: new Abstract: In federated contextual linear bandits, high data dimensionality incurs prohibitive computation and communication costs: local agents perform O(d^3)-tim

Selfie-Capture Dynamics as an Auxiliary Signal Against Deepfakes and Injection Attacks for Mobile Identity Verification

Model ReleasesDGX agent

arXiv:2605.00218v1 Announce Type: cross Abstract: Mobile remote identity verification (RIdV) systems are exposed to attacks that manipulate or replace the facial video stream, including presentation a

Smart Profit-Aware Crop Advisory System: Kisan AI

Model ReleasesDGX agent

arXiv:2605.00133v1 Announce Type: new Abstract: Modern crop advisory systems exhibit a critical limitation termed extit{economic blindness}. These systems primarily optimize for biological yield, ofte

Soft Graph Diffusion Transformer for MIMO Detection

SafetyDGX agent

arXiv:2605.00449v1 Announce Type: cross Abstract: Learning-based MIMO detection has shown strong empirical performance, yet existing methods typically rely on fixed-depth architectures without explici

← Previous
1…167168169170171…203
Next →