AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “papers”

GridTimelineEvolution
11,255 results
Model Releases

Grid Games: The Power of Multiple Grids for Quantizing Large Language Models

DGX agent

arXiv:2605.12327v1 Announce Type: new Abstract: A major recent advance in quantization is given by microscaled 4-bit formats such as NVFP4 and MXFP4, quantizing values into small groups sharing a scal

model-releasesarxiv-cs-lg
13 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Grokking or Glitching? How Low-Precision Drives Slingshot Loss Spikes

DGX agent

arXiv:2605.06152v2 Announce Type: replace-cross Abstract: Deep neural networks exhibit periodic loss spikes during unregularized long-term training, a phenomenon known as the 'Slingshot Mechanism.' Ex

model-releasesarxiv-cs-cl
13 May 2026
Tutorials

GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization

DGX agent

arXiv:2605.12369v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models aim for general robot learning by aligning action as a modality within powerful Vision-Language Models (VLMs). Exist

tutorialsarxiv-cs-ro
13 May 2026
Model Releases

HE-SNR: Uncovering Latent Logic via Entropy for Guiding Mid-Training on SWE-bench

DGX agent

arXiv:2601.20255v2 Announce Type: replace-cross Abstract: SWE-bench has emerged as the premier benchmark for evaluating Large Language Models on complex software engineering tasks. While these capabil

model-releasesarxiv-cs-cl
13 May 2026
Local Ai

Hierarchical LLM-Driven Control for HAPS-Assisted UAV Networks: Joint Optimization of Flight and Connectivity

DGX agent

arXiv:2605.11509v1 Announce Type: cross Abstract: Uncrewed aerial vehicles (UAVs) are increasingly deployed in complex networked environments, yet the joint optimization of multi-UAV motion control an

local-aiarxiv-cs-lg
13 May 2026
Research

Hypernetworks for Dynamic Feature Selection

DGX agent

arXiv:2605.12278v1 Announce Type: new Abstract: Dynamic feature selection (DFS) is a machine learning framework in which features are acquired sequentially for individual samples under budget constrai

researcharxiv-cs-lg
13 May 2026
Research

Improving the Performance and Learning Stability of Parallelizable RNNs Designed for Ultra-Low Power Applications

DGX agent

arXiv:2605.11855v1 Announce Type: new Abstract: Sequence learning is dominated by Transformers and parallelizable recurrent neural networks (RNNs) such as state-space models, yet learning long-term de

researcharxiv-cs-lg
13 May 2026
Safety

Incentivizing Truthfulness and Collaborative Fairness in Bayesian Learning

DGX agent

arXiv:2605.11889v1 Announce Type: new Abstract: Collaborative machine learning involves training high-quality models using datasets from a number of sources. To incentivize sources to share data, exis

safetyarxiv-cs-lg
13 May 2026
Research

Information theoretic underpinning of self-supervised learning by clustering

DGX agent

arXiv:2605.11870v1 Announce Type: new Abstract: Self-supervised learning (SSL) is recognized as an essential tool for building foundation models for Artificial Intelligence applications. The advances

researcharxiv-cs-lg
13 May 2026
Applications

Instruct-ICL: Instruction-Guided In-Context Learning for Post-Disaster Damage Assessment

DGX agent

arXiv:2605.11439v1 Announce Type: new Abstract: Rapid and accurate situational awareness is essential for effective response during natural disasters, where delays in analysis can significantly hinder

applicationsarxiv-cs-cv
13 May 2026
Local Ai

Instruction Lens Score: Your Instruction Contributes a Powerful Object Hallucination Detector for Multimodal Large Language Models

DGX agent

arXiv:2605.12258v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have achieved remarkable progress, yet the object hallucination remains a critical challenge for reliable deplo

local-aiarxiv-cs-lg
13 May 2026
Model Releases

Intention-Conditioned Flow Occupancy Models

DGX agent

arXiv:2506.08902v4 Announce Type: replace Abstract: Large-scale pre-training has fundamentally changed how machine learning research is done today: large foundation models are trained once, and then c

model-releasesarxiv-cs-lg
13 May 2026
Research

Interactive State Space Model with Cross-Modal Local Scanning for Depth Super-Resolution

DGX agent

arXiv:2605.11934v1 Announce Type: new Abstract: Guided depth super-resolution (GDSR) reconstructs HR depth maps from LR inputs with HR RGB guidance. Existing methods either model each modality indepen

researcharxiv-cs-cv
13 May 2026
Safety

Internalizing Curriculum Judgment for LLM Reinforcement Fine-Tuning

DGX agent

arXiv:2605.11235v1 Announce Type: new Abstract: In LLM Reinforcement Fine-Tuning (RFT), curriculum learning drives both efficiency and performance. Yet, current methods externalize curriculum judgment

safetyarxiv-cs-lg
13 May 2026
Applications

Interpretability Can Be Actionable

DGX agent

arXiv:2605.11161v1 Announce Type: new Abstract: Interpretability aims to explain the behavior of deep neural networks. Despite rapid growth, there is mounting concern that much of this work has not tr

applicationsarxiv-cs-lg
13 May 2026
Model Releases

Investigating simple target-covariate relationships for Chronos-2 and TabPFN-TS

DGX agent

arXiv:2605.12200v1 Announce Type: new Abstract: Time Series Foundation Models (TSFMs) have recently achieved state-of-the-art performance, often outperforming supervised models in zero-shot settings.

model-releasesarxiv-cs-lg
13 May 2026
Research

Latent Chain-of-Thought Improves Structured-Data Transformers

DGX agent

arXiv:2605.11262v1 Announce Type: new Abstract: Chain-of-thought and more broadly test-time compute are known to augment the expressive capabilities of language models and have led to major innovation

researcharxiv-cs-lg
13 May 2026
Model Releases

Learnable Multi-level Discrete Wavelet Transforms for 3D Gaussian Splatting Frequency Modulation

DGX agent

arXiv:2602.14199v2 Announce Type: replace-cross Abstract: 3D Gaussian Splatting (3DGS) has emerged as a powerful approach for novel view synthesis. However, the number of Gaussian primitives often gro

model-releasesarxiv-cs-cv
13 May 2026
Tutorials

Learning density ratios in causal inference using Bregman-Riesz regression

DGX agent

arXiv:2510.16127v2 Announce Type: replace-cross Abstract: The ratio of two probability density functions is a fundamental quantity that appears in many areas of statistics and machine learning, includ

tutorialsarxiv-cs-lg
13 May 2026
Tutorials

Learning Feature Encoder with Synthetic Anomalies for Weakly Supervised Graph Anomaly Detection

DGX agent

arXiv:2605.11749v1 Announce Type: new Abstract: Weakly supervised graph anomaly detection aims to unveil unusual graph instances, e.g., nodes, whose behaviors significantly differ from normal ones, gi

tutorialsarxiv-cs-lg
13 May 2026
Model Releases

Learning Subspace-Preserving Sparse Attention Graphs from Heterogeneous Multiview Data

DGX agent

arXiv:2605.11881v1 Announce Type: new Abstract: The high-dimensional features extracted from large-scale unlabeled data via various pretrained models with diverse architectures are referred to as hete

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

LiBrA-Net: Lie-Algebraic Bilateral Affine Fields for Real-Time 4K Video Dehazing

DGX agent

arXiv:2605.11508v1 Announce Type: new Abstract: Currently, there is a gap in the field of ultra-high-definition (UHD) video dehazing due to the lack of a benchmark for evaluation. Furthermore, existin

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

LOFT: Low-Rank Orthogonal Fine-Tuning via Task-Aware Support Selection

DGX agent

arXiv:2605.11872v1 Announce Type: new Abstract: Orthogonal parameter-efficient fine-tuning (PEFT) adapts pretrained weights through structure-preserving multiplicative transformations, but existing me

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Long Story Short: Disentangling Compositionality and Long-Caption Understanding in Contrastive VLMs

DGX agent

arXiv:2509.19207v2 Announce Type: replace Abstract: Contrastive vision-language models (VLMs) have made significant progress in binding visual and textual information, yet understanding long, composit

model-releasesarxiv-cs-cv
13 May 2026
Research

Manifold Sampling via Entropy Maximization

DGX agent

arXiv:2605.12338v1 Announce Type: new Abstract: Sampling from constrained distributions has a wide range of applications, including in Bayesian optimization and robotics. Prior work establishes conver

researcharxiv-cs-lg
13 May 2026
Model Releases

Measuring Five-Nines Reliability: Sample-Efficient LLM Evaluation in Saturated Benchmarks

DGX agent

arXiv:2605.11209v1 Announce Type: new Abstract: While existing benchmarks demonstrate the near-perfect performance of large language models (LLMs) on various tasks, this apparent saturation often obsc

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Mitigating Context-Memory Conflicts in LLMs through Dynamic Cognitive Reconciliation Decoding

DGX agent

arXiv:2605.12185v1 Announce Type: new Abstract: Large language models accumulate extensive parametric knowledge through pre-training. However, knowledge conflicts occur when outdated or incorrect para

model-releasesarxiv-cs-cl
13 May 2026
Research

Mobile Traffic Camera Calibration from Road Geometry for UAV-Based Traffic Surveillance

DGX agent

arXiv:2605.11900v1 Announce Type: new Abstract: Unmanned aerial vehicles (UAVs) can provide flexible traffic surveillance where fixed roadside cameras are unavailable, costly, or impractical. However,

researcharxiv-cs-cv
13 May 2026
Safety

Morphologically Equivariant Flow Matching for Bimanual Mobile Manipulation

DGX agent

arXiv:2605.12228v1 Announce Type: new Abstract: Mobile manipulation requires coordinated control of high-dimensional, bimanual robots. Imitation learning methods have been broadly used to solve these

safetyarxiv-cs-ro
13 May 2026
Local Ai

Muon is Not That Special: Random or Inverted Spectra Work Just as Well

DGX agent

arXiv:2605.11181v1 Announce Type: new Abstract: The recent empirical success of the Muon optimizer has renewed interest in non-Euclidean optimization, typically justified by similarities with second-o

local-aiarxiv-cs-lg
13 May 2026
Model Releases

NavOL: Navigation Policy with Online Imitation Learning

DGX agent

arXiv:2605.11762v1 Announce Type: new Abstract: Learning robust navigation policies remains a core challenge in robotics. Offline imitation learning suffers from distribution shift and compounding err

model-releasesarxiv-cs-ro
13 May 2026
Research

Not Worth Mentioning? A Pilot Study on Salient Proposition Annotation

DGX agent

arXiv:2603.27358v2 Announce Type: replace Abstract: Despite a long tradition of work on extractive summarization, which by nature aims to recover the most important propositions in a text, little work

researcharxiv-cs-cl
13 May 2026
Applications

OmniThoughtVis: A Scalable Distillation Pipeline for Deployable Multimodal Reasoning Models

DGX agent

arXiv:2605.11629v1 Announce Type: new Abstract: Recent multimodal large language models (MLLMs) have shown strong chain-of-thought (CoT) reasoning ability on vision-language tasks, but their direct de

applicationsarxiv-cs-cl
13 May 2026
Agents

On Problems of Implicit Context Compression for Software Engineering Agents

DGX agent

arXiv:2605.11051v1 Announce Type: cross Abstract: LLM-based Software Engineering agents face a critical bottleneck: context length limitations cause failures on complex, long-horizon tasks. One promis

agentsarxiv-cs-cl
13 May 2026
Safety

On the Importance of Multistability for Horizon Generalization in Reinforcement Learning

DGX agent

arXiv:2605.12206v1 Announce Type: new Abstract: In reinforcement learning (RL), agents acting in partially observable Markov decision processes (POMDPs) must rely on memory, typically encoded in a rec

safetyarxiv-cs-lg
13 May 2026
Research

OUI as a Structural Observable: Towards an Activation-Centric View of Neural Network Training

DGX agent

arXiv:2605.11570v1 Announce Type: new Abstract: Activation functions are what make deep networks expressive: without them, the model collapses to a linear map. Yet we still evaluate training mostly fr

researcharxiv-cs-lg
13 May 2026
Model Releases

Output Composability of QLoRA PEFT Modules for Plug-and-Play Attribute-Controlled Text Generation

DGX agent

arXiv:2605.12345v1 Announce Type: new Abstract: Parameter-efficient fine-tuning (PEFT) techniques offer task-specific fine-tuning at a fraction of the cost of full fine-tuning, but require separate fi

model-releasesarxiv-cs-cl
13 May 2026
Safety

OverNaN: NaN-Aware Oversampling for Imbalanced Learning with Meaningful Missingness

DGX agent

arXiv:2605.11525v1 Announce Type: new Abstract: Missing values are routinely treated as defects to be eliminated through deletion or imputation prior to machine learning. In many applied domains, howe

safetyarxiv-cs-lg
13 May 2026
Research

PairDropGS: Paired Dropout-Induced Consistency Regularization for Sparse-View Gaussian Splatting

DGX agent

arXiv:2605.12072v1 Announce Type: new Abstract: Dropout-based sparse-view 3D Gaussian Splatting (3DGS) methods alleviate overfitting by randomly suppressing Gaussian primitives during training. Existi

researcharxiv-cs-cv
13 May 2026
Tutorials

PG-3DGS: Optimizing 3D Gaussian Splatting to Satisfy Physics Objectives

DGX agent

arXiv:2605.11266v1 Announce Type: new Abstract: Recent advances in Gaussian Splatting have enabled fast, high-fidelity 3D scene generation, yet these methods remain purely visual and lack an understan

tutorialsarxiv-cs-cv
13 May 2026
Model Releases

Picasso: Holistic Scene Reconstruction with Physics-Constrained Sampling

DGX agent

arXiv:2602.08058v2 Announce Type: replace Abstract: In the presence of occlusions and measurement noise, geometrically accurate scene reconstructions -- which fit the sensor data -- can still be physi

model-releasesarxiv-cs-cv
13 May 2026
Research

PointCaM: Cut-and-Mix for Open-Set Point Cloud Learning

DGX agent

arXiv:2212.02011v3 Announce Type: replace Abstract: Point cloud learning is receiving increasing attention. However, most existing point cloud models lack the practical ability to deal with the unavoi

researcharxiv-cs-cv
13 May 2026
Safety

PointGS: Semantic-Consistent Unsupervised 3D Point Cloud Segmentation with 3D Gaussian Splatting

DGX agent

arXiv:2605.11520v1 Announce Type: new Abstract: Unsupervised point cloud segmentation is critical for embodied artificial intelligence and autonomous driving, as it mitigates the prohibitive cost of d

safetyarxiv-cs-cv
13 May 2026
Safety

Position: Universal Aesthetic Alignment Narrows Artistic Expression

DGX agent

arXiv:2512.11883v3 Announce Type: replace-cross Abstract: Over-aligning image generation models to a generalized aesthetic preference conflicts with user intent, particularly when 'anti-aesthetic' out

safetyarxiv-cs-cv
13 May 2026
Applications

PrivacySIM: Evaluating LLM Simulation of User Privacy Behavior

DGX agent

arXiv:2605.12147v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used to simulate human behavior, but their ability to simulate individual privacy decisions is not well

applicationsarxiv-cs-lg
13 May 2026
Model Releases

Procedural-skill SFT across capacity tiers: A W-Shaped pre-SFT Trajectory and Regime-Asymmetric Mechanism on 0.8B-4B Qwen3.5 Models

DGX agent

arXiv:2605.11907v1 Announce Type: new Abstract: We measure procedural-skill SFT contribution across three Qwen3.5 dense scales (0.8B, 2B, 4B) on a 200-task / 40-skill holdout, with Claude Haiku 4.5 as

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

PVLM: Parsing-Aware Vision Language Model with Dynamic Contrastive Learning for Zero-Shot Deepfake Attribution

DGX agent

arXiv:2504.14129v4 Announce Type: replace Abstract: The challenge of tracing the source attribution of forged faces has gained significant attention due to the rapid advancement of generative models.

model-releasesarxiv-cs-cv
13 May 2026
Safety

Question Difficulty Estimation for Large Language Models via Answer Plausibility Scoring

DGX agent

arXiv:2605.12398v1 Announce Type: new Abstract: Estimating question difficulty is a critical component in evaluating and improving large language models (LLMs) for question answering (QA). Existing ap

safetyarxiv-cs-cl
13 May 2026
← Previous
1…173174175176177…235
Next →