AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,419
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,934
  • Model Releases23,909
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,419
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,934
  • Model Releases23,909
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

88,419Total entries
1Added by human
88,418Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,638 results
26 May 2026

XRPO: Pushing the limits of GRPO with Targeted Exploration and Exploitation

SafetyDGX agent

arXiv:2510.06672v3 Announce Type: replace Abstract: Reinforcement learning algorithms such as GRPO have driven recent advances in large language model (LLM) reasoning. While scaling the number of roll

25 May 2026

A comprehensive evaluation of pretraining strategies for channel-agnostic contrastive self-supervision of biosignals

ResearchDGX agent

arXiv:2410.19842v2 Announce Type: replace-cross Abstract: Contrastive learning yields impressive results for self-supervision in computer vision. The approach relies on the creation of positive pairs,

A drone-based framework for coral habitat mapping via weakly supervised segmentation

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research
DGX agent

arXiv:2508.18958v2 Announce Type: replace-cross Abstract: Obtaining pixel-level annotations over large spatial extents remains a major bottleneck for deploying machine learning in ecological applicati

A European Multi-Center Breast Cancer MRI Dataset

Model ReleasesDGX agent

arXiv:2506.00474v3 Announce Type: replace-cross Abstract: Early detection of breast cancer is critical for improving patient outcomes. While mammography remains the primary screening modality, magneti

Anytime Training with Schedule-Free Spectral Optimization

Model ReleasesDGX agent

arXiv:2605.23061v1 Announce Type: cross Abstract: Standard neural network training relies on learning-rate schedules tied to a fixed horizon, leading to strong path dependence and costly re-tuning as

AraHopeCorpus: Annotation Guidelines and Dataset for Hope Speech in Arabic Social Media Crisis Discourse

Model ReleasesDGX agent

arXiv:2605.23325v1 Announce Type: new Abstract: Social media has become a crucial arena for shaping public narratives during armed conflicts, providing space for both harmful and constructive communic

Archimedean Copula Inference via Taylor-Mode AD

Model ReleasesDGX agent

arXiv:2605.23134v1 Announce Type: new Abstract: No existing nested Archimedean copula tool handles all three of (a) arbitrary per-variable (right-)censoring in survival analysis, (b) arbitrary nesting

Brain-LLM Alignment Tracks Training Data, Not Typology

Model ReleasesDGX agent

arXiv:2605.23032v1 Announce Type: cross Abstract: Brain-LLM alignment is well established in English, yet the brain's language network is neuroanatomically universal across languages. Does alignment a

CARE: Class-Adaptive Expert Consensus for Reliable Learning with Long-Tailed Noisy Labels

Model ReleasesDGX agent

arXiv:2605.23254v1 Announce Type: new Abstract: Learning from real-world data is frequently hindered by the compound challenge of long-tailed class distributions and noisy annotations. Existing method

Decomposing Queries into Tool Calls for Long-Video Keyframe Retrieval

Model ReleasesDGX agent

arXiv:2605.23826v1 Announce Type: cross Abstract: Keyframe selection is a direct way to provide verifiable visual evidence for long-video question answering (QA). Queries differ in what they require,

Do Synthetic Brain MRIs Reliably Improve Tumour Classification? A StyleGAN2-ADA Class-Plane Augmentation Study on BRISC 2025

Model ReleasesDGX agent

arXiv:2605.23094v1 Announce Type: cross Abstract: Generative augmentation is often proposed as a remedy for small medical-image datasets, but synthetic images are only useful when they improve downstr

Enhancing Deep Neural Network Reliability with Refinement and Calibration

ResearchDGX agent

arXiv:2605.23249v1 Announce Type: cross Abstract: Although deep neural networks (DNNs) achieve high predictive accuracy, their confidence estimates are often unreliable, potentially compromising user

Evaluating PhaseNet on Teleseismic Data with MsPASS

Local AiDGX agent

arXiv:2605.22837v1 Announce Type: cross Abstract: Numerous studies have shown that the machine-learning picker PhaseNet produces accurate P and S picks on local earthquake signals, but its performance

EvalVerse: Pipeline-Aware and Expert-Calibrated Benchmarking for Professional Cinematic Video Generation

AgentsDGX agent

arXiv:2605.23271v1 Announce Type: cross Abstract: The rapid evolution of generative video foundation models has propelled the field toward professional-grade cinematic synthesis. To achieve such deman

From Simulation to Discovery: AI Enabled Probabilistic Emulation of Mechanistic Crop Systems

ResearchDGX agent

arXiv:2605.22848v1 Announce Type: cross Abstract: Global food security depends on predicting crop responses to climate variability, yet process based crop models remain too computationally expensive f

Given how much of the original 'bottle of water per generated email' water estimate came from guesses at the architecture of GPT-4, it would…

Model ReleasesDGX agent

Given how much of the original 'bottle of water per generated email' water estimate came from guesses at the architecture of GPT-4, it would be very much in @OpenAI's interest to publish the architect

GT-HarmBench: Benchmarking AI Safety Risks Through the Lens of Game Theory

Model ReleasesDGX agent

arXiv:2602.12316v2 Announce Type: replace Abstract: Frontier AI systems are increasingly capable and deployed in high-stakes multi-agent environments. However, existing AI safety benchmarks largely ev

Hidden Human-Like Nature of Machine-Generated Texts: Theory and Detection Enhancement

ResearchDGX agent

arXiv:2605.23190v1 Announce Type: new Abstract: Machine-generated texts (MGTs) produced by large language models (LLMs) are increasingly prevalent across various applications, while their potential mi

Human-Centered Learning Mechanics: A Dynamical Framework for Entropy-Regulated Representation Learning

Model ReleasesDGX agent

arXiv:2605.22940v1 Announce Type: cross Abstract: Deep learning is increasingly viewed as a dynamical process in parameter space, yet many existing theories still treat training as a closed optimizati

IntentionNav: A Benchmark for Intent-Driven Object Navigation from Implicit Human Instruction

Model ReleasesDGX agent

arXiv:2605.23187v1 Announce Type: new Abstract: Existing object navigation benchmarks usually tell an embodied agent which object category to find, such as microwave or chair. Human-facing embodied AI

Investigating Robot Control Policy Learning for Autonomous X-ray-guided Spine Procedures

Model ReleasesDGX agent

arXiv:2511.03882v2 Announce Type: replace-cross Abstract: Imitation learning-based robot control policies are enjoying renewed interest in video-based robotics. However, it remains unclear whether thi

L-FAME: Longitudinal Focused Attention Meditation EEG Dataset and Benchmark

Model ReleasesDGX agent

arXiv:2605.22893v1 Announce Type: cross Abstract: We introduce a novel Longitudinal Focused Attention Meditation Electroencephalography (L-FAME) dataset and an accompanying benchmark, designed to fost

LLAMA LIMA: A Living Meta-Analysis on the Effects of Generative AI on Learning Mathematics

Model ReleasesDGX agent

arXiv:2601.18685v3 Announce Type: replace-cross Abstract: The capabilities of generative AI in mathematics education are rapidly evolving, posing significant challenges for research to keep pace. Rese

LLM Sparsity Prior for Robust Feature Selection

ResearchDGX agent

arXiv:2605.23102v1 Announce Type: cross Abstract: Large language models (LLMs) offer a scalable mechanism to elicit domain-informed prior information for high-dimensional variable selection. However,

Low-Cost Hard-Label Adversarial Attack with Theoretical Foundations

ResearchDGX agent

arXiv:2601.14300v3 Announce Type: replace Abstract: Hard-label black-box attacks, relying solely on top-1 predictions, represent one of the most challenging yet practically threat models. Despite rece

MARS: Magnitude-Aware Rank Statistics

ResearchDGX agent

arXiv:2605.23563v1 Announce Type: new Abstract: Comprehensive evaluation of machine learning models is the key to make sure that they perform as robustly and consistently as desired. In order to summa

MedSAE: Dissecting MedCLIP Representations with Sparse Autoencoders

ApplicationsDGX agent

arXiv:2510.26411v2 Announce Type: replace Abstract: Artificial intelligence in healthcare requires models that are accurate and interpretable. We advance mechanistic interpretability in medical vision

MELT: A Behavioral Trace Dataset for High-Risk Memecoin Launch Detection

Model ReleasesDGX agent

arXiv:2602.13480v2 Announce Type: cross Abstract: Launchpads have become the dominant mechanism for issuing memecoins, exposing investors to a new class of high-risk launches that existing rug-pull de

Move on Muon : A Hamiltonian probability gradient flow perspective of Muon optimizer

Model ReleasesDGX agent

arXiv:2605.23871v1 Announce Type: cross Abstract: We develop a gradient flow on the space of probability measures defined on matrix-valued parameters induced by regularized Muon, an analytically smoot

Nonlinear Transformations Against Unlearnable Datasets

TutorialsDGX agent

arXiv:2406.02883v2 Announce Type: replace Abstract: Automated scraping stands out as a common method for collecting data in deep learning models without the authorization of data owners. Recent studie

NP-LoRA: Null Space Projection for Subject-Style LoRA Fusion

Model ReleasesDGX agent

arXiv:2511.11051v3 Announce Type: replace Abstract: Low-Rank Adaptation (LoRA) fusion enables the composition of subject and style representations for controllable generation without retraining. Howev

PhotoFlow: Agentic 3D Virtual Photography Missions

Model ReleasesDGX agent

arXiv:2605.23771v1 Announce Type: cross Abstract: Virtual photography asks an agent to enter a prepared 3D scene with no preselected camera pose or reference image, infer a suitable shot from scene in

PiD: Fast and High-Resolution Latent Decoding with Pixel Diffusion

HardwareDGX agent

arXiv:2605.23902v1 Announce Type: new Abstract: Most practical high-resolution text-to-image systems, including latent diffusion and autoregressive models, perform generation in a compact latent space

Pooling and Semantic Shift: The Fundamental Challenges in Long Text Embedding and Retrieval

ResearchDGX agent

arXiv:2603.21437v2 Announce Type: replace Abstract: Transformer-based embedding models frequently exhibit geometric pathologies, such as anisotropy and length-induced representation collapse, which ca

ProtDBench: A Unified Benchmark of Protein Binder Design and Evaluation

Model ReleasesDGX agent

arXiv:2605.04118v2 Announce Type: replace-cross Abstract: Recent advances in de novo protein binder design have enabled increasing experimental validation, yet reported in silico metrics remain diffic

Resilience Characterization of AI-Native Wireless Receivers via Persistent Homology

Model ReleasesDGX agent

arXiv:2605.22886v1 Announce Type: cross Abstract: AI-native wireless receivers based on deep learning exhibit remarkable performance under stationary channel conditions, yet their resilience to distri

Revitalizing Dense Material Segmentation: Stabilized Vision Transformers and the Generalization Paradox

Model ReleasesDGX agent

arXiv:2605.23747v1 Announce Type: new Abstract: Material segmentation, the pixel-wise classification of physical surface properties, remains a challenging problem in computer vision, requiring physico

Super-Linear: A Lightweight Pretrained Mixture of Linear Experts for Time Series Forecasting

ApplicationsDGX agent

arXiv:2509.15105v3 Announce Type: replace Abstract: Time series forecasting (TSF) is critical in domains like energy, finance, healthcare, and logistics, requiring models that generalize across divers

The Efficiency Frontier: A Unified Framework for Cost-Performance Optimization in LLM Context Management

ResearchDGX agent

arXiv:2605.23071v1 Announce Type: new Abstract: Large language models (LLMs) increasingly rely on long-context processing, but expanding context windows introduces substantial computational and financ

Uncovering the Latent Potential of Deep Intermediate Representations

TutorialsDGX agent

arXiv:2605.23033v1 Announce Type: cross Abstract: Foundational Models pretrained on huge amount of data learn representations that evolve across depth, forming a hierarchy of embeddings with distinct

ZipMoE: Efficient On-Device MoE Serving via Lossless Compression and Cache-Affinity Scheduling

Local AiDGX agent

arXiv:2601.21198v2 Announce Type: replace-cross Abstract: While Mixture-of-Experts (MoE) architectures substantially bolster the expressive power of large-language models, their prohibitive memory foo

24 May 2026

DeepSeek says it will lower V4 Pro API prices by 75% to 0.435/1M input and 0.87/1M output tokens, making permanent the discount prices set to expire on May 31 (Bloomberg)

Model ReleasesDGX agent

Bloomberg: DeepSeek says it will lower V4 Pro API prices by 75% to 0.435/1M input and 0.87/1M output tokens, making permanent the discount prices set to expire on May 31 — DeepSeek said it will make p

Sooo many people dont know this but Hermes creator @NousResearch has its own sub program... the FREE account is HUGE because it gives you ac…

ResearchDGX agent

Sooo many people dont know this but Hermes creator @NousResearch has its own sub program... the FREE account is HUGE because it gives you access to all of their FREE models there are paid subs and if

23 May 2026

Adaptive RBF-KAN: A Comparative Evaluation of Dynamic Shape Parameters in Kolmogorov-Arnold Networks

Model ReleasesDGX agent

arXiv:2605.21534v1 Announce Type: cross Abstract: Kolmogorov-Arnold Networks (KANs) approximate multivariate functions using learnable univariate edge functions, typically parameterized by B-spline ba

Automatic Contextual Audio Denoising

Model ReleasesDGX agent

arXiv:2605.22262v1 Announce Type: cross Abstract: Audio context determines which sound components and sources are relevant and which can be perceived as irrelevant (noise) by listeners. For example, t

Benchmarking Machine Learning Architectures for Antimicrobial Stewardship in Pediatric ICUs

TutorialsDGX agent

arXiv:2605.22611v1 Announce Type: new Abstract: Antimicrobial stewardship (AMS) is critical in pediatric intensive care units (PICUs), where diagnostic uncertainty often drives broad-spectrum antibiot

Do Deep Ensembles Actually Capture Uncertainty in Graph Neural Networks?

Model ReleasesDGX agent

arXiv:2605.22593v1 Announce Type: new Abstract: While deep ensembles are widely considered to be the default method for uncertainty quantification in deep learning, their effectiveness for graph-struc

Ex-GraphRAG: Interpretable Evidence Routing for Graph-Augmented LLMs

ResearchDGX agent

arXiv:2605.21994v1 Announce Type: new Abstract: GraphRAG conditions language models on subgraphs retrieved from knowledge graphs, encoded via message-passing GNNs. Because these encoders entangle node

From Snapshots to Trajectories: Learning Single-Cell Gene Expression Dynamics via Conditional Flow Matching

SafetyDGX agent

arXiv:2605.22340v1 Announce Type: new Abstract: Single-cell RNA sequencing (scRNA-seq) provides high-dimensional profiles of cellular states, enabling data-driven modeling of cellular dynamics over ti

Hybrid Kolmogorov-Arnold Network and XGBoost Framework for Week-Ahead Price Forecasting in Australia's National Electricity Market

Model ReleasesDGX agent

arXiv:2605.22387v1 Announce Type: new Abstract: Accurate electricity price forecasting (EPF) is essential for market participants to support operational planning and risk management, yet remains chall

IKNO: Infinite-order Kernel Neural Operators

Model ReleasesDGX agent

arXiv:2605.22182v1 Announce Type: new Abstract: Neural operators have achieved significant success in modern scientific computing due to their flexibility and strong generalization capabilities. Exist

Neural Flow Operators can Approximate any Operator: Abstract Frameworks and Universal Approcimations

ResearchDGX agent

arXiv:2605.22557v1 Announce Type: new Abstract: We introduce an abstract neural flow framework for neural networks and neural operators. The framework contains two continuous-depth models, namely neur

On Statistical Estimation of Edge-Reinforced Random Walks

ResearchDGX agent

arXiv:2503.06115v2 Announce Type: replace-cross Abstract: Reinforced random walks (RRWs), including vertex-reinforced random walks (VRRWs) and edge-reinforced random walks (ERRWs), model random walks

One LR Doesn't Fit All: Heavy-Tail Guided Layerwise Learning Rates for LLMs

Model ReleasesDGX agent

arXiv:2605.22297v1 Announce Type: new Abstract: Learning rate configuration is a fundamental aspect of modern deep learning. The prevailing practice of applying a uniform learning rate across all laye

Protein Thoughts: Interpretable Reasoning with Tree of Thoughts and Embedding-Space Flow Matching for Protein-Protein Interaction Discovery

Model ReleasesDGX agent

arXiv:2605.21522v1 Announce Type: cross Abstract: Protein-protein interactions (PPIs) govern nearly all cellular processes, yet computational methods for identifying binding partners typically produce

Representation Gap: Explaining the Unreasonable Effectiveness of Neural Networks from a Geometric Perspective

Model ReleasesDGX agent

arXiv:2605.21692v1 Announce Type: new Abstract: Characterizing precisely the asymptotic generalization error of neural networks using parameters that can be estimated efficiently is a crucial problem

Short-Term-to-Long-Term Memory Transfer for Knowledge Graphs under Partial Observability

Model ReleasesDGX agent

arXiv:2605.22142v1 Announce Type: new Abstract: Reinforcement learning under partial observability requires deciding what information to retain, yet most memory-based approaches do not explicitly mode

The case against me below is completely intellectually dishonest, filled with lies and misrepresentations, wrong about almost literally ever…

Model ReleasesDGX agent

The case against me below is completely intellectually dishonest, filled with lies and misrepresentations, wrong about almost literally everything it says—a textbook example of propaganda: - I didn’t

The Neural Compiler: Program-to-Network Translation for Hybrid Scientific Machine Learning

ResearchDGX agent

arXiv:2605.22498v1 Announce Type: new Abstract: Scientific machine learning often requires combining known physics with unknown parameters or correction terms learned from data. Existing approaches ei

this has been my experience as well. there definitely were improvements, specifically wrt shell based computer use, but also regressions, es…

Model ReleasesDGX agent

this has been my experience as well. there definitely were improvements, specifically wrt shell based computer use, but also regressions, especially in the last 3 version bumps of flicker and gerperte

← Previous
1…550551552553554…1061
Next →