AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
Human
84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,851 results
28 Apr 2026

Nonlinear Non-Gaussian Density Steering with Input and Noise Channel Mismatch: Sinkhorn with Memory for Solving the Control-affine Schrodinger Bridge Problem

ResearchDGX agent

arXiv:2604.23370v1 Announce Type: cross Abstract: Solutions to the Schrodinger bridge problem and its generalizations yield feedback control policies for optimal density steering over a controlled dif

Not All Directions Matter: Towards Structured and Task-Aware Low-Rank Model Adaptation

Model ReleasesDGX agent

arXiv:2603.14228v2 Announce Type: replace Abstract: Low-Rank Adaptation (LoRA) has become a cornerstone of parameter-efficient fine-tuning (PEFT). Yet, its efficacy is hampered by two fundamental limi

NVILA: Efficient Frontier Visual Language Models

ResearchDGX agent
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2412.04468v3 Announce Type: replace Abstract: Visual language models (VLMs) have made significant advances in accuracy in recent years. However, their efficiency has received much less attention

OAMVOS:2nd Report for 5th PVUW MOSE Track

SafetyDGX agent

arXiv:2604.22837v1 Announce Type: cross Abstract: SAM-based dense trackers provide strong short-term mask propagation but remain fragile under long occlusion, fast motion, viewpoint change, and distra

ODE-GS: Latent ODEs for Dynamic Scene Extrapolation with 3D Gaussian Splatting

Model ReleasesDGX agent

arXiv:2506.05480v4 Announce Type: replace-cross Abstract: We introduce ODE-GS, a novel approach that integrates 3D Gaussian Splatting with latent neural ordinary differential equations (ODEs) to enabl

OLaPh: Optimal Language Phonemizer

Model ReleasesDGX agent

arXiv:2509.20086v2 Announce Type: replace Abstract: Phonemization is a critical component in text-to-speech synthesis. Traditional approaches rely on deterministic transformations and lexica, while ne

Omni-o3: Deep Nested Omnimodal Deduction for Deliberative Audio-Visual Reasoning

SafetyDGX agent

arXiv:2604.24191v1 Announce Type: new Abstract: Omnimodal understanding entails a massive, highly redundant search space of cross-modal interactions, demanding focused and deliberative reasoning. Curr

OmniSch: A Multimodal PCB Schematic Benchmark For Structured Diagram Visual Reasoning

Model ReleasesDGX agent

arXiv:2604.00270v2 Announce Type: replace Abstract: Recent large multimodal models (LMMs) have made rapid progress in visual grounding, document understanding, and diagram reasoning tasks. However, th

OmniShotCut: Holistic Relational Shot Boundary Detection with Shot-Query Transformer

Model ReleasesDGX agent

arXiv:2604.24762v1 Announce Type: new Abstract: Shot Boundary Detection (SBD) aims to automatically identify shot changes and divide a video into coherent shots. While SBD was widely studied in the li

On Common Misconceptions in Classical Vehicle Dynamics

ResearchDGX agent

arXiv:2604.22815v1 Announce Type: cross Abstract: Classical vehicle dynamics contains several widely adopted misconceptions that, while intuitively appealing, may lead to inconsistencies when examined

On-Device Vision Training, Deployment, and Inference on a Thumb-Sized Microcontroller

Model ReleasesDGX agent

arXiv:2604.23012v1 Announce Type: cross Abstract: This paper presents a complete, end-to-end on-device vision machine learning pipeline, comprising data acquisition, two-layer CNN training with Adam o

On Emergent Social World Models -- Evidence for Functional Integration of Theory of Mind and Pragmatic Reasoning in Language Models

ResearchDGX agent

arXiv:2602.10298v2 Announce Type: replace Abstract: This paper investigates whether LMs recruit shared computational mechanisms for general Theory of Mind (ToM) and language-specific pragmatic reasoni

On (not) learning the Mobius function

ResearchDGX agent

arXiv:2604.23427v1 Announce Type: cross Abstract: We prove lower bounds on learning the Mobius or Liouville function with a variety of standard learning techniques, including kernel methods, noisy gra

On the Complementarity of Quantum and Classical Features: Adaptive Hybrid Quantum-Classical Feature Fusion for Breast Cancer Classification

ResearchDGX agent

arXiv:2604.22903v1 Announce Type: cross Abstract: The integration of quantum machine learning with classical deep learning offers promising avenues for medical image analysis by mapping data into high

On the Convergence of Jacobian-Free Backpropagation for Optimal Control Problems with Implicit Hamiltonians

AgentsDGX agent

arXiv:2602.00921v2 Announce Type: replace-cross Abstract: Optimal feedback control with implicit Hamiltonians poses a fundamental challenge for learning-based value function methods due to the absence

On the Convergence Theory of Pipeline Gradient-based Analog In-memory Training

ResearchDGX agent

arXiv:2410.15155v3 Announce Type: replace Abstract: Aiming to accelerate the training of large deep neural networks (DNN) in an energy-efficient way, analog in-memory computing (AIMC) emerges as a sol

On the Existence of an Inverse Solution for Preference-Based Reductions in Argumentation

ResearchDGX agent

arXiv:2604.22958v1 Announce Type: new Abstract: Preference-based argumentation frameworks (PAFs) extend Dung's approach to abstract argumentation (AAFs) by encoding preferences over arguments. Such pr

On the Memorization of Consistency Distillation for Diffusion Models

ResearchDGX agent

arXiv:2604.23552v1 Announce Type: cross Abstract: Diffusion models are central to modern generative modeling, and understanding how they balance memorization and generalization is critical for reliabl

On the Peril of (Even a Little) Nonstationarity in Satisficing Regret Minimization

ResearchDGX agent

arXiv:2603.18514v2 Announce Type: replace-cross Abstract: Motivated by the principle of satisficing in decision-making, we study satisficing regret guarantees for nonstationary K-armed bandits. We sho

On the Reasoning Abilities of Masked Diffusion Language Models

ResearchDGX agent

arXiv:2510.13117v3 Announce Type: replace-cross Abstract: Masked diffusion models (MDMs) for text offer a compelling alternative to traditional autoregressive language models. Parallel generation make

On the Surprising Effectiveness of a Single Global Merging in Decentralized Learning

Model ReleasesDGX agent

arXiv:2507.06542v4 Announce Type: replace Abstract: Decentralized learning provides a scalable alternative to parameter-server-based training, yet its performance is often hindered by limited peer-to-

One Identity, Many Roles: Multimodal Entity Coreference for Enhanced Video Situation Recognition

ResearchDGX agent

arXiv:2604.23173v1 Announce Type: new Abstract: Video Situation Recognition (VidSitu) addresses the challenging problem of 'who did what to whom, with what, how, and where' in a video. It tests thorou

One-Shot Real-World Demonstration Synthesis for Scalable Bimanual Manipulation

SafetyDGX agent

arXiv:2512.09297v3 Announce Type: replace Abstract: Learning dexterous bimanual manipulation policies critically depends on large-scale, high-quality demonstrations, yet current paradigms face inheren

One Size Fits None: Heuristic Collapse in LLM Investment Advice

ApplicationsDGX agent

arXiv:2604.23837v1 Announce Type: new Abstract: Large language models are increasingly deployed as advisors in high-stakes domains -- answering medical questions, interpreting legal documents, recomme

OntoLogX: Ontology-Guided Knowledge Graph Extraction from Cybersecurity Logs with Large Language Models

Model ReleasesDGX agent

arXiv:2510.01409v2 Announce Type: replace Abstract: System logs represent a valuable source of Cyber Threat Intelligence (CTI), capturing attacker behaviors, exploited vulnerabilities, and traces of m

Open-Vocabulary Semantic Segmentation Network Integrating Object-Level Label and Scene-Level Semantic Features for Multimodal Remote Sensing Images

ApplicationsDGX agent

arXiv:2604.24125v1 Announce Type: new Abstract: Semantic segmentation of multi-modal remote sensing imagery plays a pivotal role in land use/land cover (LULC) mapping, environmental monitoring, and pr

OpenPodcar2: a robust, ROS2 vehicle for self-driving research

SafetyDGX agent

arXiv:2604.24242v1 Announce Type: new Abstract: OpenPodcar2 is a robust, ROS2-interfaced, low-cost, open source hardware and software, autonomous vehicle platform based on an off-the-shelf, hard-canop

OpenVO: Open-World Visual Odometry with Temporal Dynamics Awareness

AgentsDGX agent

arXiv:2602.19035v2 Announce Type: replace Abstract: We introduce OpenVO, a novel framework for Open-world Visual Odometry (VO) with temporal awareness under limited input conditions. OpenVO effectivel

Optimal Experimental Design for Reliable Learning of History-Dependent Constitutive Laws

Model ReleasesDGX agent

arXiv:2603.12365v2 Announce Type: replace-cross Abstract: History-dependent constitutive models serve as macroscopic closures for the aggregated effects of micromechanics. Their parameters are typical

OptProver: Bridging Olympiad and Optimization through Continual Training in Formal Theorem Proving

Model ReleasesDGX agent

arXiv:2604.23712v1 Announce Type: cross Abstract: Recent advances in formal theorem proving have focused on Olympiad-level mathematics, leaving undergraduate domains largely unexplored. Optimization,

Oracle Noise: Faster Semantic Spherical Alignment for Interpretable Latent Optimization

SafetyDGX agent

arXiv:2604.23540v1 Announce Type: new Abstract: Text-to-image diffusion models have achieved remarkable generative capabilities, yet accurately aligning complex textual prompts with synthesized layout

Orthogonal Representation Learning for Estimating Causal Quantities

SafetyDGX agent

arXiv:2502.04274v4 Announce Type: replace Abstract: End-to-end representation learning has become a powerful tool for estimating causal quantities from high-dimensional observational data, but its eff

OS-SPEAR: A Toolkit for the Safety, Performance,Efficiency, and Robustness Analysis of OS Agents

SafetyDGX agent

arXiv:2604.24348v1 Announce Type: new Abstract: The evolution of Multimodal Large Language Models (MLLMs) has shifted the focus from text generation to active behavioral execution, particularly via OS

Otherness as a Quality in Designing Expressive Robotic Touch

ResearchDGX agent

arXiv:2604.23402v1 Announce Type: cross Abstract: Haptic technologies have advanced rapidly, yet exploration of robotic touch remains dominated by replicating realistic environmental cues or hand gest

Out of Spuriousity: Improving Robustness to Spurious Correlations without Group Annotations

TutorialsDGX agent

arXiv:2407.14974v2 Announce Type: replace-cross Abstract: Machine learning models are known to learn spurious correlations, i.e., features having strong relations with class labels but no causal relat

Overcoming Copyright Barriers in Corpus Distribution Through Non-Reversible Hashing

SafetyDGX agent

arXiv:2604.23412v1 Announce Type: new Abstract: While annotated corpora are crucial in the field of natural language processing (NLP), those containing copyrighted material are difficult to exchange a

Parameter Efficiency Is Not Memory Efficiency: Rethinking Fine-Tuning for On-Device LLM Adaptation

Model ReleasesDGX agent

arXiv:2604.22783v1 Announce Type: cross Abstract: Parameter-Efficient Fine-Tuning (PEFT) has become the standard for adapting large language models (LLMs). In this work we challenge the wide-spread as

Parameter-Efficient Multi-Task Learning via Progressive Task-Specific Adaptation

Model ReleasesDGX agent

arXiv:2509.19602v2 Announce Type: replace Abstract: Parameter-efficient fine-tuning methods have emerged as a promising solution for adapting pre-trained models to various downstream tasks. While thes

PARASITE: Conditional System Prompt Poisoning to Hijack LLMs

AgentsDGX agent

arXiv:2505.16888v4 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly deployed via third-party system prompts downloaded from public marketplaces. We identify a criti

ParkingScenes: A Structured Dataset for End-to-End Autonomous Parking in Simulation Scenes

Model ReleasesDGX agent

arXiv:2604.22835v1 Announce Type: cross Abstract: Autonomous parking remains a critical yet challenging task in intelligent driving systems, particularly within constrained urban environments where ma

Partition-of-Unity Gaussian Kolmogorov-Arnold Networks

ResearchDGX agent

arXiv:2604.23599v1 Announce Type: cross Abstract: Gaussian basis functions provide an efficient and flexible alternative to spline activations in KANs. In this work, we introduce the partition-of-unit

Passage-Aware Structural Mapping for RGB-D Visual SLAM

ResearchDGX agent

arXiv:2604.24707v1 Announce Type: new Abstract: Doorways and passages are critical structural elements for indoor robot navigation, yet they remain underexplored in modern Visual SLAM (VSLAM) framewor

Patching LLM Like Software: A Lightweight Method for Improving Safety Policy in Large Language Models

Model ReleasesDGX agent

arXiv:2511.08484v2 Announce Type: replace Abstract: We propose patching for large language models (LLMs) like software versions, a lightweight and modular approach for addressing safety vulnerabilitie

PathMoG: A Pathway-Centric Modular Graph Neural Network for Multi-Omics Survival Prediction

ResearchDGX agent

arXiv:2604.24371v1 Announce Type: cross Abstract: Cancer survival prediction from multi-omics data remains challenging because prognostic signals are high-dimensional, heterogeneous, and distributed a

Patterns vs. Patients: Evaluating LLMs against Mental Health Professionals on Personality Disorder Diagnosis through First-Person Narratives

Model ReleasesDGX agent

arXiv:2512.20298v2 Announce Type: replace-cross Abstract: Growing reliance on LLMs for psychiatric self-assessment raises questions about their ability to interpret qualitative patient narratives. Thi

PDF-WuKong: A Large Multimodal Model for Efficient Long PDF Reading with End-to-End Sparse Sampling

Model ReleasesDGX agent

arXiv:2410.05970v3 Announce Type: replace-cross Abstract: Multimodal document understanding is a challenging task to process and comprehend large amounts of textual and visual information. Recent adva

Pedestrians play chicken with an autonomous vehicle

SafetyDGX agent

arXiv:2604.24384v1 Announce Type: new Abstract: Automated vehicles (AVs) are commonly programmed to yield unconditionally to pedestrians in the interest of safety. However, this design choice can give

Peer Identity Bias in Multi-Agent LLM Evaluation: An Empirical Study Using the TRUST Democratic Discourse Analysis Pipeline

SafetyDGX agent

arXiv:2604.22971v1 Announce Type: cross Abstract: The TRUST democratic discourse analysis pipeline exposes its large language model (LLM) components to peer model identity through multiple structural

PeeriScope: A Multi-Faceted Framework for Evaluating Peer Review Quality

ResearchDGX agent

arXiv:2604.24071v1 Announce Type: new Abstract: The increasing scale and variability of peer review in scholarly venues has created an urgent need for systematic, interpretable, and extensible tools t

PEPS: Positional Encoding Projected Sampling -- Extended

TutorialsDGX agent

arXiv:2604.24167v1 Announce Type: new Abstract: Implicit neural representations (INRs) are increasingly being used as tools to map coordinates to signals, encompassing applications from neural fields

Perfecting Aircraft Maneuvers with Reinforcement Learning

ResearchDGX agent

arXiv:2604.24338v1 Announce Type: new Abstract: This paper evaluates an advanced jet trainer's utilization of artificial intelligence (AI)-based aircraft aerobatic maneuvers with the intention of deve

Personality Shapes Gender Bias in Persona-Conditioned LLM Narratives Across English and Hindi: An Empirical Investigation

SafetyDGX agent

arXiv:2604.23600v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed in persona-driven applications such as education, customer service, and social platforms, where m

Personalized Worked Example Generation from Student Code Submissions using Pattern-based Knowledge Components

ResearchDGX agent

arXiv:2604.24758v1 Announce Type: cross Abstract: Adaptive programming practice often relies on fixed libraries of worked examples and practice problems, which require substantial authoring effort and

Personalizing Causal Audio-Driven Facial Motion via Dynamic Multi-modal Retrieval

ResearchDGX agent

arXiv:2604.23692v1 Announce Type: cross Abstract: Audio-driven facial animation is essential for immersive digital interaction, yet existing frameworks fail to reconcile real-time streaming with high-

PExA: Parallel Exploration Agent for Complex Text-to-SQL

Model ReleasesDGX agent

arXiv:2604.22934v1 Announce Type: new Abstract: LLM-based agents for text-to-SQL often struggle with latency-performance trade-off, where performance improvements come at the cost of latency or vice v

Phase-Separated Complex Hilbert PCA on Markerless 3D Pose Estimation Data: A Global Phase Network and Its Extension to a Continuous Field on the Body Surface

ResearchDGX agent

arXiv:2604.24415v1 Announce Type: cross Abstract: Quantitative analysis of the kinematic chain in sports motion is essential for performance evaluation and injury prevention. Conventional methods such

PhysCodeBench: Benchmarking Physics-Aware Symbolic Simulation of 3D Scenes via Self-Corrective Multi-Agent Refinement

Model ReleasesDGX agent

arXiv:2604.23580v1 Announce Type: cross Abstract: Physics-aware symbolic simulation of 3D scenes is critical for robotics, embodied AI, and scientific computing, requiring models to understand natural

PhySE: A Psychological Framework for Real-Time AR-LLM Social Engineering Attacks

AgentsDGX agent

arXiv:2604.23148v1 Announce Type: new Abstract: The emerging threat of AR-LLM-based Social Engineering (AR-LLM-SE) attacks (e.g. SEAR) poses a significant risk to real-world social interactions. In su

Physics-Informed Temporal U-Net for High-Fidelity Fluid Interpolation

ResearchDGX agent

arXiv:2604.23372v1 Announce Type: cross Abstract: Reconstructing high-fidelity fluid dynamics from sparse temporal observations is quite challenging, mainly due to the chaotic and non-linear nature of

PhysLayer: Language-Guided Layered Animation with Depth-Aware Physics

SafetyDGX agent

arXiv:2604.23574v1 Announce Type: new Abstract: Existing image-to-video generation methods often produce physically implausible motions and lack precise control over object dynamics. While prior appro

← Previous
1…834835836837838…998
Next →