AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
Human
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,295 results
15 Apr 2026

MODIX: A Training-Free Multimodal Information-Driven Positional Index Scaling for Vision-Language Models

SafetyDGX agent

arXiv:2604.12537v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have achieved remarkable progress in multimodal understanding, yet their positional encoding mechanisms remain suboptima

MoDora: Tree-Based Semi-Structured Document Analysis System

Local AiDGX agent

arXiv:2602.23061v3 Announce Type: replace-cross Abstract: Semi-structured documents integrate diverse interleaved data elements (e.g., tables, charts, hierarchical paragraphs) arranged in various and

MolMem: Memory-Augmented Agentic Reinforcement Learning for Sample-Efficient Molecular Optimization

SafetyDGX agent

arXiv:2604.12237v1 Announce Type: cross Abstract: In drug discovery, molecular optimization aims to iteratively refine a lead compound to improve molecular properties while preserving structural simil

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Monte Carlo Stochastic Depth for Uncertainty Estimation in Deep Learning

Model ReleasesDGX agent

arXiv:2604.12719v1 Announce Type: new Abstract: The deployment of deep neural networks in safety-critical systems necessitates reliable and efficient uncertainty quantification (UQ). A practical and w

MoshiRAG: Asynchronous Knowledge Retrieval for Full-Duplex Speech Language Models

Model ReleasesDGX agent

arXiv:2604.12928v1 Announce Type: new Abstract: Speech-to-speech language models have recently emerged to enhance the naturalness of conversational AI. In particular, full-duplex models are distinguis

M^star: Every Task Deserves Its Own Memory Harness

AgentsDGX agent

arXiv:2604.11811v1 Announce Type: cross Abstract: Large language model agents rely on specialized memory systems to accumulate and reuse knowledge during extended interactions. Recent architectures ty

Multi-Head Residual-Gated DeepONet for Coherent Nonlinear Wave Dynamics

Model ReleasesDGX agent

arXiv:2604.11972v1 Announce Type: new Abstract: Coherent nonlinear wave dynamics are often strongly shaped by a compact set of physically meaningful descriptors of the initial state. Traditional neura

MultiDocFusion: Hierarchical and Multimodal Chunking Pipeline for Enhanced RAG on Long Industrial Documents

ResearchDGX agent

arXiv:2604.12352v1 Announce Type: new Abstract: RAG-based QA has emerged as a powerful method for processing long industrial documents. However, conventional text chunking approaches often neglect com

Multilingual Multi-Label Emotion Classification at Scale with Synthetic Data

ResearchDGX agent

arXiv:2604.12633v1 Announce Type: new Abstract: Emotion classification in multilingual settings remains constrained by the scarcity of annotated data: existing corpora are predominantly English, singl

Mutual Information Surprise: Rethinking Unexpectedness in Autonomous Systems

SafetyDGX agent

arXiv:2508.17403v3 Announce Type: replace Abstract: A community of researchers appears to think that a machine can be surprised and have introduced various surprise measures, principally the Shannon S

MVAdapt: Zero-Shot Multi-Vehicle Adaptation for End-to-End Autonomous Driving

Model ReleasesDGX agent

arXiv:2604.11854v1 Announce Type: cross Abstract: End-to-End (E2E) autonomous driving models are usually trained and evaluated with a fixed ego-vehicle, even though their driving policy is implicitly

Narrative-Driven Paper-to-Slide Generation via ArcDeck

Model ReleasesDGX agent

arXiv:2604.11969v1 Announce Type: new Abstract: We introduce ArcDeck, a multi-agent framework that formulates paper-to-slide generation as a structured narrative reconstruction task. Unlike existing m

Narrative over Numbers: The Identifiable Victim Effect and its Amplification Under Alignment and Reasoning in Large Language Models

Model ReleasesDGX agent

arXiv:2604.12076v1 Announce Type: cross Abstract: The Identifiable Victim Effect (IVE) - the tendency to allocate greater resources to a specific, narratively described victim than to a statistically

Navigating the Accuracy-Size Trade-Off with Flexible Model Merging

ResearchDGX agent

arXiv:2505.23209v3 Announce Type: replace Abstract: Model merging has emerged as an efficient method to combine multiple single-task fine-tuned models. The merged model can enjoy multi-task capabiliti

NaviRAG: Towards Active Knowledge Navigation for Retrieval-Augmented Generation

AgentsDGX agent

arXiv:2604.12766v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) typically relies on a flat retrieval paradigm that maps queries directly to static, isolated text segments. This ap

Nemotron 3 Super: Open, Efficient Mixture-of-Experts Hybrid Mamba-Transformer Model for Agentic Reasoning

Model ReleasesDGX agent

arXiv:2604.12374v1 Announce Type: cross Abstract: We describe the pre-training, post-training, and quantization of Nemotron 3 Super, a 120 billion (active 12 billion) parameter hybrid Mamba-Attention

Neural Dynamic GI: Random-Access Neural Compression for Temporal Lightmaps in Dynamic Lighting Environments

ResearchDGX agent

arXiv:2604.12625v1 Announce Type: cross Abstract: High-quality global illumination (GI) in real-time rendering is commonly achieved using precomputed lighting techniques, with lightmap as the standard

NeuroPareto: Calibrated Acquisition for Costly Many-Goal Search in Vast Parameter Spaces

Model ReleasesDGX agent

arXiv:2602.03901v4 Announce Type: replace Abstract: The pursuit of optimal trade-offs in high-dimensional search spaces under stringent computational constraints poses a fundamental challenge for cont

No More Stale Feedback: Co-Evolving Critics for Open-World Agent Learning

SafetyDGX agent

arXiv:2601.06794v2 Announce Type: replace Abstract: Critique-guided reinforcement learning (RL) has emerged as a powerful paradigm for training LLM agents by augmenting sparse outcome rewards with nat

NoisePrints: Distortion-Free Watermarks for Authorship in Private Diffusion Models

TutorialsDGX agent

arXiv:2510.13793v2 Announce Type: replace Abstract: With the rapid adoption of diffusion models for visual content generation, proving authorship and protecting copyright have become critical. This ch

Not All Turns Are Equally Hard: Adaptive Thinking Budgets For Efficient Multi-Turn Reasoning

SafetyDGX agent

arXiv:2604.05164v2 Announce Type: replace-cross Abstract: As LLM reasoning performance plateau, improving inference-time compute efficiency is crucial to mitigate overthinking and long thinking traces

NTIRE 2026 The 3rd Restore Any Image Model (RAIM) Challenge: Professional Image Quality Assessment (Track 1)

Model ReleasesDGX agent

arXiv:2604.12512v1 Announce Type: cross Abstract: In this paper, we present an overview of the NTIRE 2026 challenge on the 3rd Restore Any Image Model in the Wild, specifically focusing on Track 1: Pr

Nucleus-Image: Sparse MoE for Image Generation

Model ReleasesDGX agent

arXiv:2604.12163v1 Announce Type: new Abstract: We present Nucleus-Image, a text-to-image generation model that establishes a new Pareto frontier in quality-versus-efficiency by matching or exceeding

Observing the unobserved confounding through its effects: toward randomized trial-like estimates from real-world survival data

ApplicationsDGX agent

arXiv:2604.12137v1 Announce Type: cross Abstract: Background: Randomized controlled trials (RCTs) are costly, time-consuming, and often infeasible, while treatment-effect estimation from observational

Obtaining Partition Crossover masks using Statistical Linkage Learning for solving noised optimization problems with hidden variable dependency structure

ApplicationsDGX agent

arXiv:2604.11862v1 Announce Type: cross Abstract: In optimization problems, some variable subsets may have a joint non-linear or non-monotonical influence on the function value. Therefore, knowledge o

OctoTools: An Agentic Framework with Extensible Tools for Complex Reasoning

AgentsDGX agent

arXiv:2502.11271v2 Announce Type: replace-cross Abstract: Solving complex reasoning tasks may involve visual understanding, domain knowledge retrieval, numerical calculation, and multi-step reasoning.

OFA-Diffusion Compression: Compressing Diffusion Model in One-Shot Manner

Model ReleasesDGX agent

arXiv:2604.12668v1 Announce Type: new Abstract: The Diffusion Probabilistic Model (DPM) achieves remarkable performance in image generation, while its increasing parameter size and computational overh

Offline-Online Reinforcement Learning for Linear Mixture MDPs

SafetyDGX agent

arXiv:2604.11994v1 Announce Type: new Abstract: We study offline-online reinforcement learning in linear mixture Markov decision processes (MDPs) under environment shift. In the offline phase, data ar

Olmo 3

Model ReleasesDGX agent

arXiv:2512.13961v2 Announce Type: replace Abstract: We introduce Olmo 3, a family of state-of-the-art, fully-open language models at the 7B and 32B parameter scales. Olmo 3 model construction targets

OmniFood8K: Single-Image Nutrition Estimation via Hierarchical Frequency-Aligned Fusion

ResearchDGX agent

arXiv:2604.12356v1 Announce Type: new Abstract: Accurate estimation of food nutrition plays a vital role in promoting healthy dietary habits and personalized diet management. Most existing food datase

OmniHands: Towards Robust 4D Hand Mesh Recovery via A Versatile Transformer

Model ReleasesDGX agent

arXiv:2405.20330v4 Announce Type: replace-cross Abstract: In this paper, we introduce OmniHands, a universal approach to recovering interactive hand meshes and their relative movement from monocular o

On Efficient Variants of Segment Anything Model: A Survey

ApplicationsDGX agent

arXiv:2410.04960v5 Announce Type: replace Abstract: The Segment Anything Model (SAM) is a foundational model for image segmentation tasks, known for its strong generalization across diverse applicatio

On Higher-Order Geometric Refinements of Classical Covariance Asymptotics: An Approach via Intrinsic and Extrinsic Information Geometry

Model ReleasesDGX agent

arXiv:2604.12725v1 Announce Type: cross Abstract: Classical Fisher-information asymptotics describe the covariance of regular efficient estimators through the local quadratic approximation of the log-

On the continuum limit of t-SNE for data visualization

Model ReleasesDGX agent

arXiv:2604.12041v1 Announce Type: cross Abstract: This work is concerned with the continuum limit of a graph-based data visualization technique called the t-Distributed Stochastic Neighbor Embedding (

On the Convergence Analysis of Muon

ResearchDGX agent

arXiv:2505.23737v2 Announce Type: replace-cross Abstract: The majority of parameters in neural networks are naturally represented as matrices. However, most commonly used optimizers treat these matrix

On the Geometry of Receiver Operating Characteristic and Precision-Recall Curves

ApplicationsDGX agent

arXiv:2504.02169v3 Announce Type: replace-cross Abstract: We study the geometry of Receiver Operating Characteristic (ROC) and Precision-Recall (PR) curves in binary classification problems. The key f

On the Mathematical Relationship Between Layer Normalization and Dynamic Activation Functions

ResearchDGX agent

arXiv:2503.21708v4 Announce Type: replace-cross Abstract: Layer normalization (LN) is an essential component of modern neural networks. While many alternative techniques have been proposed, none of th

One Model for All: Unified Try-On and Try-Off in Any Pose via LLM-Inspired Bidirectional Tweedie Diffusion

ApplicationsDGX agent

arXiv:2508.04559v3 Announce Type: replace Abstract: Recent diffusion-based approaches have made significant advances in image-based virtual try-on, enabling more realistic and end-to-end garment synth

One Token Away from Collapse: The Fragility of Instruction-Tuned Helpfulness

ResearchDGX agent

arXiv:2604.13006v1 Announce Type: cross Abstract: Instruction-tuned large language models produce helpful, structured responses, but how robust is this helpfulness when trivially constrained? We show

One View Is Enough! Monocular Training for In-the-Wild Novel View Generation

ResearchDGX agent

arXiv:2603.23488v2 Announce Type: replace Abstract: Monocular novel-view synthesis has long required multi-view image pairs for supervision, limiting training data scale and diversity. We argue it is

OpenTME: An Open Dataset of AI-powered H&E Tumor Microenvironment Profiles from TCGA

ResearchDGX agent

arXiv:2604.12075v1 Announce Type: cross Abstract: The tumor microenvironment (TME) plays a central role in cancer progression, treatment response, and patient outcomes, yet large-scale, consistent, an

Operationalising the Right to be Forgotten in LLMs: A Lightweight Sequential Unlearning Framework for Privacy-Aligned Deployment in Politically Sensitive Environments

Model ReleasesDGX agent

arXiv:2604.12459v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed in politically sensitive environments, where memorisation of personal data or confidential conten

Orthogonal Subspace Projection for Continual Machine Unlearning via SVD-Based LoRA

Model ReleasesDGX agent

arXiv:2604.12526v1 Announce Type: cross Abstract: Continual machine unlearning aims to remove the influence of data that should no longer be retained, while preserving the usefulness of the model on e

OSC: Hardware Efficient W4A4 Quantization via Outlier Separation in Channel Dimension

ResearchDGX agent

arXiv:2604.12782v1 Announce Type: cross Abstract: While 4-bit quantization is essential for high-throughput deployment of Large Language Models, activation outliers often lead to significant accuracy

OVAL: Open-Vocabulary Augmented Memory Model for Lifelong Object Goal Navigation

AgentsDGX agent

arXiv:2604.12872v1 Announce Type: new Abstract: Object Goal Navigation (ObjectNav) refers to an agent navigating to an object in an unseen environment, which is an ability often required in the accomp

PAINT: Partner-Agnostic Intent-Aware Cooperative Transport with Legged Robots

SafetyDGX agent

arXiv:2604.12852v1 Announce Type: new Abstract: Collaborative transport requires robots to infer partner intent through physical interaction while maintaining stable loco-manipulation. This becomes pa

PAL: Personal Adaptive Learner

ApplicationsDGX agent

arXiv:2604.13017v1 Announce Type: new Abstract: AI-driven education platforms have made some progress in personalisation, yet most remain constrained to static adaptation--predefined quizzes, uniform

Parallax: Why AI Agents That Think Must Never Act

SafetyDGX agent

arXiv:2604.12986v1 Announce Type: cross Abstract: Autonomous AI agents are rapidly transitioning from experimental tools to operational infrastructure, with projections that 80% of enterprise applicat

Parametric Interpolation of Dynamic Mode Decomposition for Predicting Nonlinear Systems

Model ReleasesDGX agent

arXiv:2604.12103v1 Announce Type: cross Abstract: We present parameter-interpolated dynamic mode decomposition (piDMD), a parametric reduced-order modeling framework that embeds known parameter-affine

Parcae: Scaling Laws For Stable Looped Language Models

Model ReleasesDGX agent

arXiv:2604.12946v1 Announce Type: new Abstract: Traditional fixed-depth architectures scale quality by increasing training FLOPs, typically through increased parameterization, at the expense of a high

ParetoBandit: Budget-Paced Adaptive Routing for Non-Stationary LLM Serving

Model ReleasesDGX agent

arXiv:2604.00136v2 Announce Type: replace-cross Abstract: Multi-model LLM serving operates in a non-stationary, noisy environment: providers revise pricing, model quality can shift or regress without

PC-MIL: Decoupling Feature Resolution from Supervision Scale in Whole-Slide Learning

Local AiDGX agent

arXiv:2604.12100v1 Announce Type: new Abstract: Whole-slide image (WSI) classification in computational pathology is commonly formulated as slide-level Multiple Instance Learning (MIL) with a single g

PDF-GS: Progressive Distractor Filtering for Robust 3D Gaussian Splatting

ApplicationsDGX agent

arXiv:2604.12580v1 Announce Type: new Abstract: Recent advances in 3D Gaussian Splatting (3DGS) have enabled impressive real-time photorealistic rendering. However, conventional training pipelines inh

Perception-Aware Policy Optimization for Multimodal Reasoning

SafetyDGX agent

arXiv:2507.06448v5 Announce Type: replace Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has proven to be a highly effective strategy for endowing Large Language Models (LLMs) with ro

Physically Accurate Rigid-Body Dynamics in Particle-Based Simulation

Model ReleasesDGX agent

arXiv:2603.14634v3 Announce Type: replace Abstract: Robotics demands simulation that can reason about the diversity of real-world physical interactions, from rigid to deformable objects and fluids. Cu

Physics-Grounded Monocular Vehicle Distance Estimation Using Standardized License Plate Typography

SafetyDGX agent

arXiv:2604.12239v1 Announce Type: new Abstract: Accurate inter-vehicle distance estimation is a cornerstone of Advanced Driver Assistance Systems (ADAS) and autonomous driving. While LiDAR and radar p

Pi-HOC: Pairwise 3D Human-Object Contact Estimation

ApplicationsDGX agent

arXiv:2604.12923v1 Announce Type: new Abstract: Resolving real-world human-object interactions in images is a many-to-many challenge, in which disentangling fine-grained concurrent physical contact is

PianoFlow: Music-Aware Streaming Piano Motion Generation with Bimanual Coordination

ResearchDGX agent

arXiv:2604.12856v1 Announce Type: new Abstract: Audio-driven bimanual piano motion generation requires precise modeling of complex musical structures and dynamic cross-hand coordination. However, exis

Pictorial and apictorial polygonal jigsaw puzzles from arbitrary number of crossing cuts

ResearchDGX agent

arXiv:2008.07644v3 Announce Type: replace-cross Abstract: Jigsaw puzzle solving, the problem of constructing a coherent whole from a set of non-overlapping unordered visual fragments, is fundamental t

PILOT: Planning via Internalized Latent Optimization Trajectories for Large Language Models

ResearchDGX agent

arXiv:2601.19917v2 Announce Type: replace Abstract: Strategic planning is critical for multi-step reasoning, yet compact Large Language Models (LLMs) often lack the capacity to formulate global strate

← Previous
1…930931932933934…989
Next →