AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,588 results
1 Jul 2026

One Shot vs. Iterative: Rethinking Pruning Strategies for Model Compression

Model ReleasesDGX agent

arXiv:2508.13836v2 Announce Type: replace-cross Abstract: Pruning is a core technique for compressing neural networks to improve computational efficiency. This process is typically approached in two w

Preserve the Hard, Regenerate the Rest: Uncertainty-Guided Synthetic Training Data Augmentation with Diffusion Models

AgentsDGX agent

arXiv:2606.31603v1 Announce Type: cross Abstract: Semantic segmentation models struggle with data sparsity and rare or visually diverse regions, e.g., dense regions or small objects in aerial or auton

Relational and Sequential Conformal Inference for Energy Time Series over Graphs via Foundation Models

SafetyDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.31804v1 Announce Type: new Abstract: Accurate energy demand forecasting is essential for the reliable operation and planning of modern sustainable energy systems. Spatial-temporal graph neu

Sources: Meta plans a cloud infrastructure business that will sell access to AI compute and models, to compete with AWS, Azure, and Google Cloud; META jumps 8%+ (Bloomberg)

IndustryDGX agent

Bloomberg: Sources: Meta plans a cloud infrastructure business that will sell access to AI compute and models, to compete with AWS, Azure, and Google Cloud; META jumps 8%+ — Meta Platforms Inc. is dev

T-QPM: Enabling Temporal Out-Of-Distribution Detection and Domain Generalization for Vision-Language Models in Open-World

ResearchDGX agent

arXiv:2603.18481v2 Announce Type: replace Abstract: Out-of-distribution (OOD) detection remains a critical challenge in open-world learning, where models must adapt to evolving data distributions. Whi

Together AI, which offers access to open-source models, raised 800M led by Saudi Aramco's Prosperity7 at an 8.3B valuation, taking its total funding to $1.3B (Niko Gallogly/New York Times)

IndustryDGX agent

Niko Gallogly / New York Times: Together AI, which offers access to open-source models, raised 800M led by Saudi Aramco's Prosperity7 at an 8.3B valuation, taking its total funding to 1.3B — Together

TwelveLabs raises $100M to bring superintelligence to AI video models

IndustryDGX agent

TwelveLabs Inc., the developer of generative artificial intelligence foundation models that can understand videos like humans, today announced it has raised 100 million in early funding to expand beyo

UniTac: A Unified Multimodal Model for Cross-Sensor Tactile Understanding and Generation

SafetyDGX agent

arXiv:2606.31451v1 Announce Type: cross Abstract: Unified multimodal models (UMMs) have shown great promise in integrating understanding and generation across diverse modalities. However, existing res

Visual Semantic Entropy: Do Vision Language Models Recognize Visual Ambiguity?

ResearchDGX agent

arXiv:2606.31407v1 Announce Type: cross Abstract: Vision-language models can produce confident answers on visually ambiguous inputs, resulting in biased predictions. Common entropy-based methods, such

30 Jun 2026

A causal modeling perspective on decision theory

SafetyDGX agent

arXiv:2606.29911v1 Announce Type: new Abstract: Decision theory provides a formal framework for how agents should make choices under uncertainty, drawing on ideas from philosophy, probability, and cau

AERMANI-VLM: Structured Prompting and Reasoning for Aerial Manipulation with Vision Language Models

SafetyDGX agent

arXiv:2511.01472v2 Announce Type: replace Abstract: The rapid progress of vision--language models (VLMs) has sparked growing interest in robotic control, where natural language can express the operati

Be Faithful When Response: Returning Fluent and Grounded Answers for Vision-Language Models Reinforcement Learning

SafetyDGX agent

arXiv:2606.29984v1 Announce Type: new Abstract: Reinforcement Learning (RL) is an important paradigm for improving the reasoning capabilities of Vision-Language Models (VLMs). However, directly applyi

Can regularization based JEPA (e.g. SIGReg) scale and compete with SOTA foundation models (DINO)? Here is the answer: yes and with 10x less …

ResearchDGX agent

Can regularization based JEPA (e.g. SIGReg) scale and compete with SOTA foundation models (DINO)? Here is the answer: yes and with 10x less data. VISReg (slight variation of SIGReg) competes with DINO

CaresAI at CT-DEB26: Detecting Dosing Errors In Clinical Trials Using Domain-Specific Transformer Embeddings and Classification Models

SafetyDGX agent

arXiv:2606.30236v1 Announce Type: new Abstract: Medication errors, particularly dosing errors in clinical trials (CT), can lead to patient harm, adverse drug events and worse patient outcomes. Dosing

Diffusion Models Observe Only Gradients: A Geometric Perspective on Score Matching Errors

ResearchDGX agent

arXiv:2606.06179v2 Announce Type: replace-cross Abstract: Score-based diffusion models are typically trained by minimizing the L^2 score matching error, and standard theoretical analyses rely on this

DreamForge-World 0.1 Preview: A Low-Compute Real-Time Controllable World Model

HardwareDGX agent

arXiv:2606.30292v1 Announce Type: cross Abstract: We present DreamForge-World 0.1 Preview, a preview foundational world model for real-time interactive world simulation. The system adapts the LongLive

DuoMem: Towards Capable On-Device Memory Agents via Dual-Space Distillation

Model ReleasesDGX agent

arXiv:2606.29961v1 Announce Type: cross Abstract: Large Language Model (LLM)-based agents can solve complex procedural tasks by interacting with environments over multiple turns, but this ability typi

Em-ergence of the em-dash: a population-level rise in em-dash frequency in medRxiv preprints at the dawn of the large-language-model era

ResearchDGX agent

arXiv:2606.29540v1 Announce Type: cross Abstract: Large language models (LLMs) can leave subtle stylistic traces in assisted text; one of the most cited is the em-dash (Unicode U+2014). Yet no one has

Evolutional Math: Cross-Validated Island-Model Genetic Programming for Interpretable Symbolic Regression on Small, Wide Datasets

Model ReleasesDGX agent

arXiv:2606.28381v1 Announce Type: cross Abstract: Symbolic regression via genetic programming routinely fails on small, wide datasets - a regime common in clinical-trial monitoring, biostatistics, and

FDM-MFVT: Few-step Sampling Diffusion Model for Mask-Free Virtual Try-On

SafetyDGX agent

arXiv:2606.29319v1 Announce Type: new Abstract: Image-based Virtual Try-On (IVTON) has greatly advanced through diffusion models, yet existing methods require many sampling steps and depend on masks w

From Word Sequences to Behavioral Sequences: Adapting Modeling and Evaluation Paradigms for Longitudinal NLP

ResearchDGX agent

arXiv:2601.07988v2 Announce Type: replace-cross Abstract: While NLP typically treats documents as independent and unordered samples, in longitudinal studies, this assumption rarely holds: documents ar

General and Efficient Steering of Diffusion Models

SafetyDGX agent

arXiv:2602.11395v2 Announce Type: replace-cross Abstract: Steering diffusion models toward conditions unseen during training typically requires either retraining with conditional inputs or per-step gr

HDDPM: Heteroscedastic Denoising Diffusion Probabilistic Model for Quantitative Low-Count Brain PET Recovery

Local AiDGX agent

arXiv:2606.28513v1 Announce Type: cross Abstract: Positron emission tomography (PET) seeks to balance diagnostic quality with ra-diation dose. Low-count PET noise is non-Gaussian, non-stationary, and

IMCBench: A benchmark for multimodal LLMs in Image-grounded Medical Conversations

Model ReleasesDGX agent

arXiv:2606.28556v1 Announce Type: new Abstract: Recent advances in large language models and vision-language models have enabled reasoning over multimodal data, offering opportunities for clinical app

Improvement of Robot's Simultaneous Localization and Mapping Using an Effective Transformation to Achieve Linear Model

Local AiDGX agent

arXiv:2606.28475v1 Announce Type: cross Abstract: Nowadays mobile robots have wide engineering applications. Simultaneous localization and mapping (SLAM) is an important task of these robots. The majo

Interpretable Inverse Design of Metal-Organic Frameworks with Large Language Model Agents

AgentsDGX agent

arXiv:2606.29459v1 Announce Type: cross Abstract: Inverse design of metal-organic frameworks (MOFs) requires searching a combinatorially vast space where property labels are expensive and most machine

Mechanistically Eliciting Latent Behaviors in Language Models

SafetyDGX agent

arXiv:2606.29604v1 Announce Type: cross Abstract: We aim to discover diverse, generalizable perturbations of LLM internals that can surface hidden behavioral modes. Such perturbations could help resha

Momentum Guidance: Plug-and-Play Guidance for Flow Models

ResearchDGX agent

arXiv:2602.20360v2 Announce Type: replace-cross Abstract: Flow-based generative methods offer a simple and effective framework for high-fidelity generation, yet pretrained flow models are rarely used

Multimodal and Multiscale Spatial-Temporal Semantic Search and Recommendation with AI Foundation Models

Local AiDGX agent

arXiv:2606.28369v1 Announce Type: cross Abstract: Semantic search and recommendation of similar documents, such as news and reports about unusual environmental events (e.g., a dead whale washed ashore

Pondering the Way: Spatial-perceiving World Action Model for Embodied Navigation

SafetyDGX agent

arXiv:2606.29908v1 Announce Type: cross Abstract: Existing world model-based planners for visual navigation typically follow a verification-centric paradigm, decoupling goal intent from trajectory syn

Predicting Metastatic Risk from Primary Tissue Architecture via Distance-Aware Spatial Modeling

ResearchDGX agent

arXiv:2606.28676v1 Announce Type: cross Abstract: Predicting the risk of distant metastasis from primary tumor tissue histology is a critical yet challenging task in computational pathology. Multiple

Recommended reading if you are looking to scale with open models in production.

ApplicationsDGX agent

This post likely recommends resources or reading materials for practitioners looking to deploy and scale open-source language models in production environments. DAIR.AI, a research organization focuse

See our full model rankings: http://cursor.com/evals

ToolsDGX agent

Cursor published a comprehensive ranking of AI models, likely evaluating them across performance metrics, capabilities, and use cases relevant to their code editor platform. The rankings help users un

SGD Provably Prioritizes a Shortcut Spurious Feature in the XOR Model

SafetyDGX agent

arXiv:2606.30444v1 Announce Type: cross Abstract: Neural networks are known to be susceptible to over-reliance on spurious correlations. However, the precise mechanism by which models exploit shortcut

Simplifying Flow Matching Transformations with Low-Rank Mixture Models

TutorialsDGX agent

arXiv:2606.29724v1 Announce Type: new Abstract: Normalizing flows are powerful generative models that learn an invertible mapping between complex data distributions and simple latent distributions, ty

SOTAlign: Semi-Supervised Alignment of Unimodal Vision and Language Models via Optimal Transport

SafetyDGX agent

arXiv:2602.23353v2 Announce Type: replace-cross Abstract: The Platonic Representation Hypothesis posits that neural networks trained on different modalities converge toward a shared statistical model

SWE-fficiency: Can Language Models Optimize Real-World Repositories on Real Workloads?

Model ReleasesDGX agent

arXiv:2511.06090v3 Announce Type: replace-cross Abstract: Optimizing the performance of large-scale software repositories demands expertise in code reasoning and software engineering (SWE) to reduce r

TAP-VLA: Tactile Annotation Prompting for Vision Language Action Models

SafetyDGX agent

arXiv:2606.29089v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models demonstrate impressive reasoning over visual, semantic, and spatial task variations by leveraging large-scale vision

The Surprising Effectiveness of Video Diffusion Models for Hand Motion Reconstruction

ResearchDGX agent

arXiv:2606.30308v1 Announce Type: new Abstract: 4D hand motion reconstruction from egocentric video is bottlenecked by clear limitations of existing methods: image-based pipelines depend on a detector

Towards Spatial Trace with Reasoning in Vision-Language Models for Robotics

Model ReleasesDGX agent

arXiv:2512.13660v3 Announce Type: replace-cross Abstract: Spatial tracing, as a fundamental embodied interaction ability for robots, is inherently challenging as it requires multi-step metric-grounded

We heard your feedback. You want to go faster. Introducing GLM 5.2 Fast The same model and quality as GLM 5.2 standard, now at 140 tok/s Fli…

ToolsDGX agent

Fireworks AI announced GLM 5.2 Fast, an optimized version of their GLM 5.2 model that maintains the same quality and capabilities as the standard version while delivering significantly faster inferenc

When Does Synthetic CT Transfer? A Label-Free Donor/Host Diagnostic for Medical Vision-Language Model Routing on Real Lung CT

ResearchDGX agent

arXiv:2606.29232v1 Announce Type: new Abstract: A synthetic measurement of model competence is useful only if it survives the move to real data, yet the real labels that would verify it are exactly wh

29 Jun 2026

A Gaussian Perspective for Distributional Discrepancy in Generative Diffusion Models

ResearchDGX agent

arXiv:2601.13602v3 Announce Type: replace-cross Abstract: This paper introduces an analytical approach to quantifying and optimizing the distributional discrepancy in generative diffusion models. For

An Approach to Enriching Surgical Video Datasets for Fine-Grained Spatial-Temporal Understanding of Vision-Language Models

ResearchDGX agent

arXiv:2604.00784v2 Announce Type: replace Abstract: Surgical video understanding is a crucial prerequisite for advancing Computer-Assisted Surgery. While vision-language models (VLMs) have recently be

Are Time-Series Foundation Models Ready for E-Nose Data? An Empirical Assessment of Their Embeddings

ApplicationsDGX agent

arXiv:2606.27672v1 Announce Type: new Abstract: Inspired by advances in natural language processing and computer vision, 'time-series foundation models' (TSFMs) have recently been introduced with the

Class-frequency Guided Noise Schedule for Diffusion Models

ResearchDGX agent

arXiv:2606.27696v1 Announce Type: cross Abstract: In this paper, we are the first to examine the correlations between class frequency and the multi-scale noise schedule within diffusion models. For sc

DFM: Difference Feature Modeling with Text-Guided Gated Contrastive Loss for Remote Sensing Image Change Captioning

TutorialsDGX agent

arXiv:2606.27410v1 Announce Type: cross Abstract: The primary goal of Remote Sensing Image Change Captioning (RSICC) is to automatically generate descriptions of changes between remote sensing images

DIM-WAM: World-Action Modeling with Diverse Historical Event Memory

Local AiDGX agent

arXiv:2606.27677v1 Announce Type: cross Abstract: World-action models have shown promising robot-manipulation performance by jointly predicting future visual states and actions. However, existing meth

End-to-End Dynamic Sparsity for Resource-Adaptive LLM Inference

Model ReleasesDGX agent

arXiv:2606.27743v1 Announce Type: cross Abstract: Large Language Models (LLMs) inference is typically deployed under a static resource assumption, where models execute a fixed computational graph rega

Odyssey: Constructing Verifiable Local Truth-Preserving Foundation Models

Local AiDGX agent

arXiv:2606.27593v1 Announce Type: new Abstract: We introduce a categorical framework called ODYSSEY for constructing verifiable, local truth-preserving foundation models as compositions of foundries:

Ornith-1.0: Self-Scaffolding LLMs for Agentic Coding

Model ReleasesDGX agent

Ornith-1.0: Self-Scaffolding LLMs for Agentic Coding This is an interesting new open weights (MIT licensed) model, the first model release from DeepReinforce. [...] with variants including 9B Dense, 3

Perceptual 3D Simulation With Physical World Modeling

Local AiDGX agent

arXiv:2606.27575v1 Announce Type: new Abstract: Predicting how a scene will evolve after a desired 3D transformation from images is a central goal in vision, graphics, and robotics. Yet unlike ideal s

Physics-Informed Neural Network with Transfer Learning for State Estimation in Lithium-Ion Batteries using the Single Particle Model with Electrolyte

TutorialsDGX agent

arXiv:2606.28220v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) have emerged as a powerful tool for solving nonlinear partial differential equations (PDEs), including battery

PRISON: Unmasking the Criminal Potential of Large Language Models

SafetyDGX agent

arXiv:2506.16150v4 Announce Type: replace-cross Abstract: As large language models (LLMs) advance, concerns about their misconduct in complex social contexts intensify. Existing research overlooked th

Rheos: Modelling Continuous Motion Dynamics in Hierarchical 3D Scene Graphs

ResearchDGX agent

arXiv:2603.20239v2 Announce Type: replace-cross Abstract: 3D Scene Graphs (3DSGs) provide hierarchical, multi-resolution abstractions that encode the geometric and semantic structure of an environment

SelectAnyTree: A Promptable Instance Segmentation Model for 3D Forest LiDAR Point Clouds

Local AiDGX agent

arXiv:2606.27491v1 Announce Type: new Abstract: Automated instance segmentation of forest LiDAR point clouds is increasingly critical as forest monitoring moves toward scalable, detailed, 3D measureme

Sources: component and supplier lists, and photos of iPhone 18 Pro models, are among files a ransomware group stole from Apple's Indian supplier Tata (Reuters)

IndustryDGX agent

Reuters: Sources: component and supplier lists, and photos of iPhone 18 Pro models, are among files a ransomware group stole from Apple's Indian supplier Tata — Sensitive lists of components and suppl

SpikeVLA: Vision-Language-Action Models with Spiking Neural Networks

SafetyDGX agent

arXiv:2606.27807v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have become a dominant paradigm for embodied intelligence. However, most existing approaches are built on large-scal

This is a neat idea around doing model routing and sub-agent delegation while ensuring cache hits on accumulated context for all agents. It …

AgentsDGX agent

This is a neat idea around doing model routing and sub-agent delegation while ensuring cache hits on accumulated context for all agents. It makes a lot of sense: you want to ensure that all subagents

VGB for Masked Diffusion Model: Efficient Test-time Scaling for Reward Satisfaction and Sample Editing

ResearchDGX agent

arXiv:2606.28301v1 Announce Type: new Abstract: Inference-time scaling is a promising paradigm to improve generative models, especially when outputs must satisfy structural constraints or optimize dow

← Previous
1…175176177178179…1010
Next →