AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
Human
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,295 results
10 Apr 2026

Non-identifiability of Explanations from Model Behavior in Deep Networks of Image Authenticity Judgments

ResearchDGX agent

arXiv:2604.07254v1 Announce Type: cross Abstract: Deep neural networks can predict human judgments, but this does not imply that they rely on human-like information or reveal the cues underlying those

Nonparametric Instrumental Regression via Kernel Methods is Minimax Optimal

ResearchDGX agent

arXiv:2411.19653v2 Announce Type: replace-cross Abstract: We study the kernel instrumental variable (KIV) algorithm, a kernel-based two-stage least-squares method for nonparametric instrumental variab

Not All Tokens See Equally: Perception-Grounded Policy Optimization for Large Vision-Language Models

Model ReleasesDGX agent
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.01840v2 Announce Type: replace Abstract: While Reinforcement Learning from Verifiable Rewards (RLVR) has advanced reasoning in Large Vision-Language Models (LVLMs), prevailing frameworks su

Novel Interpretable and Robust Web-based AI Platform for Phishing Email Detection

ApplicationsDGX agent

arXiv:2405.11619v2 Announce Type: replace-cross Abstract: Phishing emails continue to pose a significant threat, causing financial losses and security breaches. This study addresses limitations in exi

Novel View Synthesis as Video Completion

ResearchDGX agent

arXiv:2604.08500v1 Announce Type: new Abstract: We tackle the problem of sparse novel view synthesis (NVS) using video diffusion models; given K (approx 5) multi-view images of a scene and their

NTIRE 2026 Challenge on Bitstream-Corrupted Video Restoration: Methods and Results

Model ReleasesDGX agent

arXiv:2604.06945v2 Announce Type: replace Abstract: This paper reports on the NTIRE 2026 Challenge on Bitstream-Corrupted Video Restoration (BSCVR). The challenge aims to advance research on recoverin

Object-Centric Stereo Ranging for Autonomous Driving: From Dense Disparity to Census-Based Template Matching

HardwareDGX agent

arXiv:2604.07980v1 Announce Type: new Abstract: Accurate depth estimation is critical for autonomous driving perception systems, particularly for long range vehicle detection on highways. Traditional

OceanMAE: A Foundation Model for Ocean Remote Sensing

TutorialsDGX agent

arXiv:2604.08171v1 Announce Type: new Abstract: Accurate ocean mapping is essential for applications such as bathymetry estimation, seabed characterization, marine litter detection, and ecosystem moni

ODE-free Neural Flow Matching for One-Step Generative Modeling

TutorialsDGX agent

arXiv:2604.06413v1 Announce Type: new Abstract: Diffusion and flow matching models generate samples by learning time-dependent vector fields whose integration transports noise to data, requiring tens

ODYN: An All-Shifted Non-Interior-Point Method for Quadratic Programming in Robotics and AI

Model ReleasesDGX agent

arXiv:2602.16005v2 Announce Type: replace-cross Abstract: We introduce ODYN, a novel all-shifted primal-dual non-interior-point quadratic programming (QP) solver designed to efficiently handle challen

OmniJigsaw: Enhancing Omni-Modal Reasoning via Modality-Orchestrated Reordering

ResearchDGX agent

arXiv:2604.08209v1 Announce Type: new Abstract: To extend the reinforcement learning post-training paradigm to omni-modal models for concurrently bolstering video-audio understanding and collaborative

OmniTabBench: Mapping the Empirical Frontiers of GBDTs, Neural Networks, and Foundation Models for Tabular Data at Scale

Model ReleasesDGX agent

arXiv:2604.06814v1 Announce Type: cross Abstract: While traditional tree-based ensemble methods have long dominated tabular tasks, deep neural networks and emerging foundation models have challenged t

On Emotion-Sensitive Decision Making of Small Language Model Agents

Model ReleasesDGX agent

arXiv:2604.06562v1 Announce Type: new Abstract: Small language models (SLM) are increasingly used as interactive decision-making agents, yet most decision-oriented evaluations ignore emotion as a caus

On Integrating Resilience and Human Oversight into LLM-Assisted Modeling Workflows for Digital Twins

Model ReleasesDGX agent

arXiv:2603.25898v2 Announce Type: replace-cross Abstract: LLM-assisted modeling holds the potential to rapidly build executable Digital Twins of complex systems from only coarse descriptions and senso

On-Policy Distillation of Language Models for Autonomous Vehicle Motion Planning

Model ReleasesDGX agent

arXiv:2604.07944v1 Announce Type: new Abstract: Large language models (LLMs) have recently demonstrated strong potential for autonomous vehicle motion planning by reformulating trajectory prediction a

On the Global Photometric Alignment for Low-Level Vision

SafetyDGX agent

arXiv:2604.08172v1 Announce Type: new Abstract: Supervised low-level vision models rely on pixel-wise losses against paired references, yet paired training sets exhibit per-pair photometric inconsiste

On the Price of Privacy for Language Identification and Generation

ResearchDGX agent

arXiv:2604.07238v1 Announce Type: new Abstract: As large language models (LLMs) are increasingly trained on sensitive user data, understanding the fundamental cost of privacy in language learning beco

On the Robustness of Diffusion-Based Image Compression to Bit-Flip Errors

ResearchDGX agent

arXiv:2604.05743v2 Announce Type: replace-cross Abstract: Modern image compression methods are typically optimized for the rate--distortion--perception trade-off, whereas their robustness to bit-level

On the Step Length Confounding in LLM Reasoning Data Selection

ResearchDGX agent

arXiv:2604.06834v1 Announce Type: cross Abstract: Large reasoning models have recently demonstrated strong performance on complex tasks that require long chain-of-thought reasoning, through supervised

On the Uphill Battle of Image frequency Analysis

ResearchDGX agent

arXiv:2604.07563v1 Announce Type: new Abstract: This work is a follow up on the newly proposed clustering algorithm called The Inverse Square Mean Shift Algorithm. In this paper a special case of algo

Once4All: Skeleton-Guided SMT Solver Fuzzing with LLM-Synthesized Generators

ResearchDGX agent

arXiv:2508.20340v4 Announce Type: replace-cross Abstract: Satisfiability Modulo Theory (SMT) solvers are foundational to modern systems and programming languages research, providing the foundation for

One Life to Learn: Inferring Symbolic World Models for Stochastic Environments from Unguided Exploration

AgentsDGX agent

arXiv:2510.12088v2 Announce Type: replace Abstract: Symbolic world modeling requires inferring and representing an environment's transitional dynamics as an executable program. Prior work has focused

Ontology-based knowledge graph infrastructure for interoperable atomistic simulation data

ResearchDGX agent

arXiv:2604.06230v1 Announce Type: cross Abstract: The reuse of atomistic simulation data is often limited by heterogeneous formats, incomplete metadata, and a lack of standardized representations of w

Open-Ended Instruction Realization with LLM-Enabled Multi-Planner Scheduling in Autonomous Vehicles

Model ReleasesDGX agent

arXiv:2604.08031v1 Announce Type: cross Abstract: Most Human-Machine Interaction (HMI) research overlooks the maneuvering needs of passengers in autonomous driving (AD). Natural language offers an int

OpenPRC: A Unified Open-Source Framework for Physics-to-Task Evaluation in Physical Reservoir Computing

HardwareDGX agent

arXiv:2604.07423v1 Announce Type: new Abstract: Physical Reservoir Computing (PRC) leverages the intrinsic nonlinear dynamics of physical substrates, mechanical, optical, spintronic, and beyond, as fi

OpenSpatial: A Principled Data Engine for Empowering Spatial Intelligence

ApplicationsDGX agent

arXiv:2604.07296v2 Announce Type: replace Abstract: Spatial understanding is a fundamental cornerstone of human-level intelligence. Nonetheless, current research predominantly focuses on domain-specif

OpenTrack3D: Towards Accurate and Generalizable Open-Vocabulary 3D Instance Segmentation

ResearchDGX agent

arXiv:2512.03532v2 Announce Type: replace Abstract: Generalizing open-vocabulary 3D instance segmentation (OV-3DIS) to diverse, unstructured, and mesh-free environments is crucial for robotics and AR/

OpenVLThinkerV2: A Generalist Multimodal Reasoning Model for Multi-domain Visual Tasks

SafetyDGX agent

arXiv:2604.08539v1 Announce Type: cross Abstract: Group Relative Policy Optimization (GRPO) has emerged as the de facto Reinforcement Learning (RL) objective driving recent advancements in Multimodal

Operator Learning for Surrogate Modeling of Wave-Induced Forces from Sea Surface Waves

ResearchDGX agent

arXiv:2604.06433v1 Announce Type: cross Abstract: Wave setup plays a significant role in transferring wave-induced energy to currents and causing an increase in water elevation. This excess momentum f

Optimal Decay Spectra for Linear Recurrences

ResearchDGX agent

arXiv:2604.07658v1 Announce Type: cross Abstract: Linear recurrent models offer linear-time sequence processing but often suffer from suboptimal long-range memory. We trace this to the decay spectrum:

Optimal Rates for Pure {arepsilon}-Differentially Private Stochastic Convex Optimization with Heavy Tails

Model ReleasesDGX agent

arXiv:2604.06492v1 Announce Type: new Abstract: We study stochastic convex optimization (SCO) with heavy-tailed gradients under pure epsilon-differential privacy (DP). Instead of assuming a bound on t

ORACLE-SWE: Quantifying the Contribution of Oracle Information Signals on SWE Agents

AgentsDGX agent

arXiv:2604.07789v1 Announce Type: cross Abstract: Recent advances in language model (LM) agents have significantly improved automated software engineering (SWE). Prior work has proposed various agenti

OrgForge: A Multi-Agent Simulation Framework for Verifiable Synthetic Corporate Corpora

AgentsDGX agent

arXiv:2603.14997v2 Announce Type: replace Abstract: Building and evaluating enterprise AI systems requires synthetic organizational corpora that are internally consistent, temporally structured, and c

Orion-Lite: Distilling LLM Reasoning into Efficient Vision-Only Driving Models

Model ReleasesDGX agent

arXiv:2604.08266v1 Announce Type: new Abstract: Leveraging the general world knowledge of Large Language Models (LLMs) holds significant promise for improving the ability of autonomous driving systems

oslash Source Models Leak What They Shouldn't nrightarrow: Unlearning Zero-Shot Transfer in Domain Adaptation Through Adversarial Optimization

Model ReleasesDGX agent

arXiv:2604.08238v1 Announce Type: new Abstract: The increasing adaptation of vision models across domains, such as satellite imagery and medical scans, has raised an emerging privacy risk: models may

OV-Stitcher: A Global Context-Aware Framework for Training-Free Open-Vocabulary Semantic Segmentation

ResearchDGX agent

arXiv:2604.08110v1 Announce Type: new Abstract: Training-free open-vocabulary semantic segmentation(TF-OVSS) has recently attracted attention for its ability to perform dense prediction by leveraging

OVS-DINO: Open-Vocabulary Segmentation via Structure-Aligned SAM-DINO with Language Guidance

SafetyDGX agent

arXiv:2604.08461v1 Announce Type: new Abstract: Open-Vocabulary Segmentation (OVS) aims to segment image regions beyond predefined category sets by leveraging semantic descriptions. While CLIP based a

OxEnsemble: Fair Ensembles for Low-Data Classification

SafetyDGX agent

arXiv:2512.09665v2 Announce Type: replace Abstract: We address the problem of fair classification in settings where data is scarce and unbalanced across demographic groups. Such low-data regimes are c

PAC-Bayesian Bounds on Constrained f-Entropic Risk Measures

ResearchDGX agent

arXiv:2510.11169v2 Announce Type: replace-cross Abstract: PAC generalization bounds on the risk, when expressed in terms of the expected loss, are often insufficient to capture imbalances between subg

PANC: Prior-Aware Normalized Cut via Anchor-Augmented Token Graphs

ResearchDGX agent

arXiv:2602.06912v2 Announce Type: replace Abstract: Unsupervised segmentation from self-supervised ViT patches holds promise but lacks robustness: multi-object scenes confound saliency cues, and low-s

PanoSAM2: Lightweight Distortion- and Memory-aware Adaptions of SAM2 for 360 Video Object Segmentation

ResearchDGX agent

arXiv:2604.07901v1 Announce Type: new Abstract: 360 video object segmentation (360VOS) aims to predict temporally-consistent masks in 360 videos, offering full-scene coverage, benefiting applications,

Paragraph Segmentation Revisited: Towards a Standard Task for Structuring Speech

ResearchDGX agent

arXiv:2512.24517v2 Announce Type: replace Abstract: Automatic speech transcripts are often delivered as unstructured word streams that impede readability and repurposing. We recast paragraph segmentat

ParkSense: Where Should a Delivery Driver Park? Leveraging Idle AV Compute and Vision-Language Models

AgentsDGX agent

arXiv:2604.07912v1 Announce Type: new Abstract: Finding parking consumes a disproportionate share of food delivery time, yet no system addresses precise parking-spot selection relative to merchant ent

ParseBench: A Document Parsing Benchmark for AI Agents

Model ReleasesDGX agent

arXiv:2604.08538v1 Announce Type: new Abstract: AI agents are changing the requirements for document parsing. What matters is semantic correctness: parsed output must preserve the structure and

Part^{2}GS: Part-aware Modeling of Articulated Objects using 3D Gaussian Splatting

SafetyDGX agent

arXiv:2506.17212v2 Announce Type: replace Abstract: Articulated objects are common in the real world, yet modeling their structure and motion remains a challenging task for 3D reconstruction methods.

PASK: Toward Intent-Aware Proactive Agents with Long-Term Memory

Model ReleasesDGX agent

arXiv:2604.08000v1 Announce Type: cross Abstract: Proactivity is a core expectation for AGI. Prior work remains largely confined to laboratory settings, leaving a clear gap in real-world proactive age

Path Regularization: A Near-Complete and Optimal Nonasymptotic Generalization Theory for Multilayer Neural Networks and Double Descent Phenomenon

SafetyDGX agent

arXiv:2503.02129v2 Announce Type: replace-cross Abstract: Path regularization has shown to be a very effective regularization to train neural networks, leading to a better generalization property than

PD-SOVNet: A Physics-Driven Second-Order Vibration Operator Network for Estimating Wheel Polygonal Roughness from Axle-Box Vibrations

ApplicationsDGX agent

arXiv:2604.06620v1 Announce Type: new Abstract: Quantitative estimation of wheel polygonal roughness from axle-box vibration signals is a challenging yet practically relevant problem for rail-vehicle

PEER: Unified Process-Outcome Reinforcement Learning for Structured Empathetic Reasoning

SafetyDGX agent

arXiv:2508.09521v2 Announce Type: replace Abstract: Emotional support conversations require more than fluent responses. Supporters need to understand the seeker's situation and emotions, adopt an appr

PeReGrINE: Evaluating Personalized Review Fidelity with User Item Graph Context

Model ReleasesDGX agent

arXiv:2604.07788v1 Announce Type: cross Abstract: We introduce PeReGrINE, a benchmark and evaluation framework for personalized review generation grounded in graph-structured user--item evidence. PeRe

Personalized RewardBench: Evaluating Reward Models with Human Aligned Personalization

Model ReleasesDGX agent

arXiv:2604.07343v1 Announce Type: cross Abstract: Pluralistic alignment has emerged as a critical frontier in the development of Large Language Models (LLMs), with reward models (RMs) serving as a cen

Personalizing Text-to-Image Generation to Individual Taste

SafetyDGX agent

arXiv:2604.07427v1 Announce Type: new Abstract: Modern text-to-image (T2I) models generate high-fidelity visuals but remain indifferent to individual user preferences. While existing reward models opt

Phantasia: Context-Adaptive Backdoors in Vision Language Models

ResearchDGX agent

arXiv:2604.08395v1 Announce Type: new Abstract: Recent advances in Vision-Language Models (VLMs) have greatly enhanced the integration of visual perception and linguistic reasoning, driving rapid prog

Phantom: Physics-Infused Video Generation via Joint Modeling of Visual and Latent Physical Dynamics

ApplicationsDGX agent

arXiv:2604.08503v1 Announce Type: new Abstract: Recent advances in generative video modeling, driven by large-scale datasets and powerful architectures, have yielded remarkable visual realism. However

PhyAVBench: A Challenging Audio Physics-Sensitivity Benchmark for Physically Grounded Text-to-Audio-Video Generation

Model ReleasesDGX agent

arXiv:2512.23994v3 Announce Type: replace-cross Abstract: Text-to-audio-video (T2AV) generation is central to applications such as filmmaking and world modeling. However, current models often fail to

PhyEdit: Towards Real-World Object Manipulation via Physically-Grounded Image Editing

Model ReleasesDGX agent

arXiv:2604.07230v2 Announce Type: replace Abstract: Achieving physically accurate object manipulation in image editing is essential for its potential applications in interactive world models. However,

Physical Adversarial Attacks on AI Surveillance Systems:Detection, Tracking, and Visible--Infrared Evasion

ResearchDGX agent

arXiv:2604.06865v1 Announce Type: cross Abstract: Physical adversarial attacks are increasingly studied in settings that resemble deployed surveillance systems rather than isolated image benchmarks. I

Physical Knot Classification Beyond Accuracy: A Benchmark and Diagnostic Study

Model ReleasesDGX agent

arXiv:2603.23286v3 Announce Type: replace Abstract: Physical knot classification is a fine-grained task in which the intended cue is rope crossing structure, but high accuracy may still come from appe

Physically Plausible Human-Object Rendering from Sparse Views via 3D Gaussian Splatting

TutorialsDGX agent

arXiv:2503.09640v2 Announce Type: replace-cross Abstract: Rendering realistic human-object interactions (HOIs) from sparse-view inputs is a challenging yet crucial task for various real-world applicat

Physics-Informed Functional Link Constrained Framework with Domain Mapping for Solving Bending Analysis of an Exponentially Loaded Perforated Beam

Model ReleasesDGX agent

arXiv:2604.07025v1 Announce Type: cross Abstract: This article presents a novel and comprehensive approach for analyzing bending behavior of the tapered perforated beam under an exponential load. The

← Previous
1…981982983984985…989
Next →