AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,545 results
23 Jul 2026

OpenSkillRisk: Benchmarking Agent Safety When Using Real-World Risky Third-Party Skills

Model ReleasesDGX agent

arXiv:2607.20121v1 Announce Type: new Abstract: LLM-based agents leverage third-party skills to extend their capabilities in open-world scenarios. However, third-party skills can introduce extra secur

Opto-ViT-v2: Noise-Resilient On-Chip Fine-Tuning for Photonic Near-Sensor Vision Transformer Accelerators

Model ReleasesDGX agent

arXiv:2607.19421v1 Announce Type: cross Abstract: Silicon-photonic (SiPh) accelerators have emerged as a promising platform for Vision Transformer (ViT) inference by performing matrix multiplications

Outperforming Fable 5 at half the price: meet model synthesis, a new server-side tool on DigitalOcean Inference Engine


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases
DGX agent

Anyone building with AI eventually hits the same tradeoff: how to get the most intelligence per dollar, the right model at the right cost for each task. That’s what DigitalOcean Inference Engine is bu

[Paper] SLAI T-Rex: Full-Parameter Post-training of the DeepSeek-V4 Family on Ascend SuperPOD

Model ReleasesDGX agent

Full-parameter post-training of trillion-parameter-scale MoE models introduces substantial system-level challenges for large-scale distributed training, including severe memory pressure, non-overlappe

PathAgentBench: Benchmarking Evidence-Seeking Vision-Language Models on Whole-Slide Pathology Image

Model ReleasesDGX agent

arXiv:2607.19261v1 Announce Type: new Abstract: Whole-slide image (WSI) diagnosis requires identifying diagnostically relevant regions, examining them across magnifications, and integrating multi-scal

PathReportEval: A Systematic Benchmark for Pathology Report Generation

Model ReleasesDGX agent

arXiv:2607.18448v1 Announce Type: cross Abstract: Pathology report generation from whole-slide images (WSIs) is a rapidly growing multimodal learning problem, yet progress is difficult to measure beca

PerfAgent: Profiler-Guided Iterative Refinement for Repository-Level Code Optimization

Model ReleasesDGX agent

arXiv:2607.19653v1 Announce Type: cross Abstract: Large language model (LLM) agents now perform well on correctness-oriented repository-level tasks, including SWE-Bench issue resolution and feature im

PG-KINN: A Physics-Informed Petrov-Galerkin Kolmogorov-Arnold Network for Solving Forward and Inverse PDEs

Model ReleasesDGX agent

arXiv:2607.20378v1 Announce Type: new Abstract: Physics-informed learning of partial differential equations (PDEs) has been dominated by multilayer perceptrons (MLPs), whose spectral bias and dense pa

PhenSPINE: A Standardized Benchmark for Spine Pathology Diagnosis

Model ReleasesDGX agent

arXiv:2607.19696v1 Announce Type: cross Abstract: The accurate diagnosis of spinal pathologies depends heavily on radiological interpretation, yet automated systems are hindered by the lack of diverse

Physics-Aware Complex-Valued State Space Model with Scattering-Prior Feature Modulation for PolSAR Image Classification

Model ReleasesDGX agent

arXiv:2607.19787v1 Announce Type: cross Abstract: Polarimetric synthetic aperture radar (PolSAR) image classification is a representative task for physics-aware GeoAI, where land-cover semantics are c

PN-QNN: Harnessing Physical Noise as a Native Regularizer in Photonic Hybrid Quantum Neural Networks

Model ReleasesDGX agent

arXiv:2607.20045v1 Announce Type: cross Abstract: Physical noise in near-term quantum hardware is usually treated as a nuisance to suppress. We ask whether it can instead act as a hardware-native regu

Point Ladder Tuning: Parameter-Efficient Hierarchical Adaptation for 3D Point Cloud Understanding

Model ReleasesDGX agent

arXiv:2607.19171v1 Announce Type: new Abstract: Fine-tuning pre-trained point-cloud backbones typically updates all parameters, resulting in substantial computation and memory overhead. More important

Point-Selection Fine-Tuning Framework for Robust Point Cloud Classification

Model ReleasesDGX agent

arXiv:2607.19711v1 Announce Type: new Abstract: Noisy and corrupted points can substantially degrade point cloud recognition performance, especially under challenging corruption settings. In particula

Post-Training in Time Series Foundation Models: A Unifying Framework

Model ReleasesDGX agent

arXiv:2607.20002v1 Announce Type: cross Abstract: Time series foundation models (TSFMs) have emerged as general-purpose models for time series analysis, but pretraining alone is often insufficient for

PRiSM: Prototype Regularization for Few-Shot VLMs

Model ReleasesDGX agent

arXiv:2607.17820v2 Announce Type: replace Abstract: Training-free few-shot adaptation methods have gained significant attention recently in the context of Vision-language Models (VLMs). Yet, current b

Prober.ai: Gated Inquiry-Based Feedback via LLM-Constrained Personas for Argumentative Writing Development

Model ReleasesDGX agent

arXiv:2605.05598v2 Announce Type: replace Abstract: The proliferation of large language models (LLMs) in educational settings has paradoxically undermined the cognitive processes they purport to suppo

Profile-Graph Memory for LLM Agents: Implicit Cross-Entity Traversal through Narrative Profiles

Model ReleasesDGX agent

arXiv:2607.19359v1 Announce Type: new Abstract: Long-term memory is essential for LLM agents that interact across sessions, yet current memory benchmarks primarily evaluate single-hop recall, leaving

PSA on Laguna S-2.1 - Use the updated chat template and GGUF

Model ReleasesDGX agent

Link to their official GGUF repo: https://huggingface.co/poolside/Laguna-S-2.1-GGUF/tree/main All the GGUFs received this fix 5ish hours ago - correct yarn_attn_factor to 1.0 (llama.cpp derives mscale

Pushing the Frontier of Full-Song Generation: Hierarchical Autoregressive Planning Meets Flow-Matching Rendering

Model ReleasesDGX agent

arXiv:2607.20253v1 Announce Type: cross Abstract: In this report, we present a unified song generation framework capable of producing high-quality full-length music from lyrics, text descriptions, and

Quantile Transfer for Reliable Operating Point Selection in Visual Place Recognition

Model ReleasesDGX agent

arXiv:2602.04401v4 Announce Type: replace Abstract: Visual Place Recognition (VPR) is a key component for localization in Global Navigation Satellite System (GNSS)-denied environments, but its perform

Rarity-Aware Discrete Diffusion with Spatially Consistent Decoding for Photo-Realistic Image Super-Resolution

Model ReleasesDGX agent

arXiv:2607.17612v2 Announce Type: replace Abstract: Continuous diffusion models have become the dominant paradigm for photo-realistic image Super-Resolution (SR), but they typically formulate reconstr

Read or Ignore? A Unified Benchmark for Typographic-Attack Robustness and Text Recognition in Vision-Language Models

Model ReleasesDGX agent

arXiv:2512.11899v2 Announce Type: replace Abstract: Large vision-language models (LVLMs) are vulnerable to typographic attacks, where misleading text inserted into an image can override visual underst

Reading and Steering Representations of Materials-Science Mechanisms in an Open-Weight Language Model

Model ReleasesDGX agent

arXiv:2607.20058v1 Announce Type: new Abstract: Large language models can answer scientific questions, yet a correct output does not reveal whether the model represents or uses the governing physics.

Recovering Clinical Utility Under Differential Privacy: Empirical Validation of Adaptive Federated Aggregation on Heterogeneous Cardiovascular Datasets

Model ReleasesDGX agent

arXiv:2607.19403v1 Announce Type: cross Abstract: Validating federated learning frameworks on real clinical data is an essential step between proof-of-concept demonstrations in controlled synthetic en

Recti-Q: Feature-Space Rectification for Out-of-Distribution-Robust Quantized Perception in Edge Robotics

Model ReleasesDGX agent

arXiv:2607.18540v1 Announce Type: new Abstract: Robotic perception pipelines increasingly rely on large vision backbones deployed on SWaP-constrained edge platforms, making post-training quantization

Reducing Learner Redundancy in Boosting via Residual Orthogonalization

Model ReleasesDGX agent

arXiv:2606.17567v2 Announce Type: replace Abstract: While sequential residual fitting is the bedrock of standard boosting frameworks, it inherently breeds learner redundancy by repeatedly revisiting c

ReFace: Reorganizing Facial Spatiotemporal Representations for Improved Pain Assessment

Model ReleasesDGX agent

arXiv:2607.19722v1 Announce Type: new Abstract: Automatic pain assessment from facial video remains challenging due to the spatial heterogeneity of pain-related facial cues. This study proposes ReFace

Reference-Free Evaluation of Reasoning in Open-Ended Question Answering

Model ReleasesDGX agent

arXiv:2607.19678v1 Announce Type: cross Abstract: AI-generated answers in high-stakes domains are often fluent but difficult to verify, especially when they contain multi-step reasoning rather than a

Reinforcement Learning for Large Language Model Selective Evidence Adoption from Contaminated Retrieval Results

Model ReleasesDGX agent

arXiv:2607.20090v1 Announce Type: cross Abstract: Retrieval-augmented large language models frequently face contexts that interleave useful evidence with misleading statements or instruction-like cont

Reliability-Aware Hard--Soft Physics-Informed Neural Networks for Robust Learning of Challenging Partial Differential Equations

Model ReleasesDGX agent

arXiv:2607.19377v1 Announce Type: cross Abstract: Physics-informed neural networks (PINNs) provide a mesh-free framework for solving partial differential equations, but their training is often affecte

RoboInter1.5: A Holistic Intermediate Representation Suite for Embodied World Modeling and Robotic Manipulation

Model ReleasesDGX agent

arXiv:2607.18709v2 Announce Type: replace Abstract: Existing robot datasets remain expensive to curate, embodiment-specific, and insufficiently annotated with the fine-grained structure required for g

ROMS-IMLE: A Minimalist Approach to Competitive Single-Step Generative Modelling

Model ReleasesDGX agent

arXiv:2607.19332v1 Announce Type: cross Abstract: Generative models have undergone many generations of evolution, from VAEs/GANs to diffusion/flow matching. Along the way, the underlying techniques ha

RS-RIE-Bench: Benchmarking Reasoning-Guided Remote Sensing Image Editing

Model ReleasesDGX agent

arXiv:2607.20197v1 Announce Type: new Abstract: Remote sensing image editing aims to modify remote sensing images according to natural language instructions while preserving geographic rules and senso

Running Qwen 3.6 35B MoE (Q4_K_M) on a Zeus (Xiaomi 12 Pro, 12GB RAM)

Model ReleasesDGX agent

Shoutout to this awesome guy - https://www.reddit.com/r/LLM/s/IDUyU3v9ap Thanks to his project, BigMoeOnEdge https://github.com/Helldez/BigMoeOnEdge, I managed to successfully run a 35B MoE model on j

Runway launches Runway Media Router, which it says is the first built specifically for generative media, as it expands from AI video to AI infrastructure (Rebecca Bellan/TechCrunch)

Model ReleasesDGX agent

Rebecca Bellan / TechCrunch: Runway launches Runway Media Router, which it says is the first built specifically for generative media, as it expands from AI video to AI infrastructure — Runway no longe

Safe Remediation as Risk-Constrained Intervention Decision in Microservice Systems

Model ReleasesDGX agent

arXiv:2607.20005v1 Announce Type: new Abstract: In modern IT operations (IT-Ops), the cost of an incorrect repair often exceeds the cost of no action at all. Yet existing automated remediation systems

Seeing Before Generating: Object Perception Enhances Single-View 3D Reconstruction

Model ReleasesDGX agent

arXiv:2607.18630v1 Announce Type: new Abstract: The relationship between object perception and reconstruction is well established in human vision, yet remains underexplored in computer vision. In this

Self Gradient Forcing: Native Long Video Extrapolation

Model ReleasesDGX agent

arXiv:2607.20368v1 Announce Type: new Abstract: Recent autoregressive video diffusion methods are increasingly built upon Self Forcing, where the student is trained on histories produced by its own ro

SenWorld: A Digital-Twin Simulation for Generating Context-Rich Evaluation Data

Model ReleasesDGX agent

arXiv:2607.19949v1 Announce Type: new Abstract: Smartphone personal assistants reason over longitudinal personal data, yet evaluating them requires context-rich evaluation data whose correct answers a

Simultaneous Speech-to-Speech Translation Without Aligned Data

Model ReleasesDGX agent

arXiv:2602.11072v2 Announce Type: replace Abstract: Simultaneous speech translation requires translating source speech into a target language in real-time while handling non-monotonic word dependencie

Single-Teacher View Augmentation: Enhancing Knowledge Distillation with Student-Guided Perturbations

Model ReleasesDGX agent

arXiv:2607.11557v2 Announce Type: replace Abstract: Knowledge distillation (KD) typically relies on the fixed perspective of a single teacher, limiting the diversity of supervisory signals. While mult

SLAI T-Rex: Full-Parameter Post-training of the DeepSeek-V4 Family on Ascend SuperPOD

Model ReleasesDGX agent

arXiv:2607.20145v1 Announce Type: cross Abstract: Full-parameter post-training of trillion-parameter-scale MoE models introduces substantial system-level challenges for large-scale distributed trainin

Small, Free, and Effective: Orchestrating Open-Weight Small Language Models to Outperform Single LLM for Malware Analysis

Model ReleasesDGX agent

arXiv:2607.20216v1 Announce Type: cross Abstract: Malware analysis demands rapid interpretation of complex detonation reports spanning filesystem, network, and process behaviours. While large language

Solar Open 2 Technical Report

Model ReleasesDGX agent

arXiv:2607.20062v1 Announce Type: new Abstract: We present Solar Open 2, a 250B-A15B Mixture-of-Experts language model built for long-horizon agentic tasks, scaled up from Solar Open 1 (Solar Open 100

Sophisticated Policies from Epistemic Priors

Model ReleasesDGX agent

arXiv:2607.19518v1 Announce Type: new Abstract: Sophisticated Inference is a variant of active inference often associated with recursive belief modeling and tree search. We argue that its central comp

Spectral-LSH: Sub-Quadratic Prompt Compression via Krylov-Projected Locality-Sensitive Hashing

Model ReleasesDGX agent

arXiv:2607.19368v1 Announce Type: new Abstract: Long-prompt inference remains expensive because prefill attention scales quadratically with sequence length. We propose Spectral-LSH, a training-free pr

Start Customizing NVIDIA Nemotron 3 Nano with Prime Intellect Lab in Minutes

Model ReleasesDGX agent

Prime Intellect Lab offers a streamlined, hosted reinforcement‑learning workflow that lets users customize the NVIDIA Nemotron 3 Nano in minutes. The process establishes a baseline, trains the model o

Stateful Guardrails for Multi-Turn LLM Systems: A Conversational Risk Accumulation Framework

Model ReleasesDGX agent

arXiv:2607.19361v1 Announce Type: cross Abstract: Most safety guardrails for large language models (LLMs) evaluate each prompt-response pair in isolation, which misses failures that arise only over a

Statistical Inference for Rank Allocation in Low-Rank Adaptation

Model ReleasesDGX agent

arXiv:2607.20205v1 Announce Type: cross Abstract: Low-rank adaptation (LoRA) has become a widely used parameter-efficient fine-tuning method for large language models. Since different modules and laye

Statistically Grounded Sparse-Feature Interventions for Activation-Space Control in Large Language Models

Model ReleasesDGX agent

arXiv:2607.19364v1 Announce Type: new Abstract: Activation steering offers a lightweight alternative to fine-tuning for behavioral control of large language models, but SAE-based steering methods ofte

STN-TGAT: Top-K Portfolio Construction via Prior-Guided Graph Attention with Learnable Soft-Threshold Sparsification

Model ReleasesDGX agent

arXiv:2607.19385v1 Announce Type: new Abstract: This paper tackles the problem of stock ranking and portfolio construction under realistic investment settings by jointly modeling temporal dynamics and

Strength-Parity Ensembling with Parameter-Isolated Experts for Multi-Task Affect Recognition

Model ReleasesDGX agent

arXiv:2607.16290v2 Announce Type: replace Abstract: Leading entries on the multi-task track of the 11th ABAW challenge rely on heavy ensembling, yet which member is worth adding to an already strong e

StrokeSeg2: Stroke Lesion Segmentation in Clinical Research Workflows

Model ReleasesDGX agent

arXiv:2607.19901v1 Announce Type: new Abstract: Deep learning frameworks like nnU-Net achieve state-of-theart brain lesion segmentation performance but remain difficult to deploy in clinical research

SUM: Unified Geometric Surgery on Spatio-Temporal Adaptation Vectors for Federated Class Incremental Learning

Model ReleasesDGX agent

arXiv:2607.19384v1 Announce Type: new Abstract: Real-world intelligent systems often require both distributed collaboration across data-isolated clients and continual adaptation to evolving tasks. Thi

SynGallery: A Synthetic Gallery of Real Paintings for Instance-Level Artwork Recognition

Model ReleasesDGX agent

arXiv:2607.18907v1 Announce Type: new Abstract: Instance-level artwork recognition requires matching a handheld visitor photograph to a specific work in a large museum collection. This is challenging

Task Competence Is Not Instruction Following: Evaluating Instruction-Conflicting Behavior in Small Language Models

Model ReleasesDGX agent

arXiv:2607.19608v1 Announce Type: new Abstract: Instruction tuning is meant to make language models follow user requests, yet it is unclear whether small models comply when an instruction conflicts wi

The Anatomy of a Truth Direction: Knowledge-Dependent Dimensionality, a Relational Law, and a Convergent Category Geometry in Small Language Models

Model ReleasesDGX agent

arXiv:2607.16741v2 Announce Type: replace Abstract: Burger et al. (2024) demonstrated that truth representations in large language models are universal across statement polarity but reside within a mu

The Blessing of Dimensionality: How Near-Orthogonality in High-Dimensional Spaces Explains Temporal Portability

Model ReleasesDGX agent

arXiv:2607.20301v1 Announce Type: cross Abstract: Fine-tuning has been widely used to adapt large language models (LLMs) for domain-specific tasks. Parameter efficient fine-tuning (PEFT) methods such

The Blueprint: How Voicify makes AI-enabled ordering a delight for customers

Model ReleasesDGX agent

Welcome to The Blueprint, a new feature where we highlight how Google Cloud customers are tackling unique and common challenges across industries using the latest AI and cloud technologies. We hope to

The Chronos Vulnerability: A Taxonomy of Temporal Persistence and Memory-Based Deception in Agentic AI

Model ReleasesDGX agent

arXiv:2607.19433v1 Announce Type: new Abstract: The transition from stateless generative models in artificial intelligence to stateful, autonomous agents represents an architectural evolution that, wh

← Previous
1…7576777879…376
Next →