AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,563 results
Model Releases

RoboInter1.5: A Holistic Intermediate Representation Suite for Embodied World Modeling and Robotic Manipulation

DGX agent

arXiv:2607.18709v2 Announce Type: replace Abstract: Existing robot datasets remain expensive to curate, embodiment-specific, and insufficiently annotated with the fine-grained structure required for g

model-releasesarxiv-cs-ro
23 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

ROMS-IMLE: A Minimalist Approach to Competitive Single-Step Generative Modelling

DGX agent

arXiv:2607.19332v1 Announce Type: cross Abstract: Generative models have undergone many generations of evolution, from VAEs/GANs to diffusion/flow matching. Along the way, the underlying techniques ha

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

RS-RIE-Bench: Benchmarking Reasoning-Guided Remote Sensing Image Editing

DGX agent

arXiv:2607.20197v1 Announce Type: new Abstract: Remote sensing image editing aims to modify remote sensing images according to natural language instructions while preserving geographic rules and senso

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

Running Qwen 3.6 35B MoE (Q4_K_M) on a Zeus (Xiaomi 12 Pro, 12GB RAM)

DGX agent

Shoutout to this awesome guy - https://www.reddit.com/r/LLM/s/IDUyU3v9ap Thanks to his project, BigMoeOnEdge https://github.com/Helldez/BigMoeOnEdge, I managed to successfully run a 35B MoE model on j

model-releasesr-localllama
23 Jul 2026
Model Releases

Runway launches Runway Media Router, which it says is the first built specifically for generative media, as it expands from AI video to AI infrastructure (Rebecca Bellan/TechCrunch)

DGX agent

Rebecca Bellan / TechCrunch: Runway launches Runway Media Router, which it says is the first built specifically for generative media, as it expands from AI video to AI infrastructure — Runway no longe

model-releasestechmeme
23 Jul 2026
Model Releases

Safe Remediation as Risk-Constrained Intervention Decision in Microservice Systems

DGX agent

arXiv:2607.20005v1 Announce Type: new Abstract: In modern IT operations (IT-Ops), the cost of an incorrect repair often exceeds the cost of no action at all. Yet existing automated remediation systems

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

Seeing Before Generating: Object Perception Enhances Single-View 3D Reconstruction

DGX agent

arXiv:2607.18630v1 Announce Type: new Abstract: The relationship between object perception and reconstruction is well established in human vision, yet remains underexplored in computer vision. In this

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

Self Gradient Forcing: Native Long Video Extrapolation

DGX agent

arXiv:2607.20368v1 Announce Type: new Abstract: Recent autoregressive video diffusion methods are increasingly built upon Self Forcing, where the student is trained on histories produced by its own ro

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

SenWorld: A Digital-Twin Simulation for Generating Context-Rich Evaluation Data

DGX agent

arXiv:2607.19949v1 Announce Type: new Abstract: Smartphone personal assistants reason over longitudinal personal data, yet evaluating them requires context-rich evaluation data whose correct answers a

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

Simultaneous Speech-to-Speech Translation Without Aligned Data

DGX agent

arXiv:2602.11072v2 Announce Type: replace Abstract: Simultaneous speech translation requires translating source speech into a target language in real-time while handling non-monotonic word dependencie

model-releasesarxiv-cs-cl
23 Jul 2026
Model Releases

Single-Teacher View Augmentation: Enhancing Knowledge Distillation with Student-Guided Perturbations

DGX agent

arXiv:2607.11557v2 Announce Type: replace Abstract: Knowledge distillation (KD) typically relies on the fixed perspective of a single teacher, limiting the diversity of supervisory signals. While mult

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

SLAI T-Rex: Full-Parameter Post-training of the DeepSeek-V4 Family on Ascend SuperPOD

DGX agent

arXiv:2607.20145v1 Announce Type: cross Abstract: Full-parameter post-training of trillion-parameter-scale MoE models introduces substantial system-level challenges for large-scale distributed trainin

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

Small, Free, and Effective: Orchestrating Open-Weight Small Language Models to Outperform Single LLM for Malware Analysis

DGX agent

arXiv:2607.20216v1 Announce Type: cross Abstract: Malware analysis demands rapid interpretation of complex detonation reports spanning filesystem, network, and process behaviours. While large language

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

Solar Open 2 Technical Report

DGX agent

arXiv:2607.20062v1 Announce Type: new Abstract: We present Solar Open 2, a 250B-A15B Mixture-of-Experts language model built for long-horizon agentic tasks, scaled up from Solar Open 1 (Solar Open 100

model-releasesarxiv-cs-cl
23 Jul 2026
Model Releases

Sophisticated Policies from Epistemic Priors

DGX agent

arXiv:2607.19518v1 Announce Type: new Abstract: Sophisticated Inference is a variant of active inference often associated with recursive belief modeling and tree search. We argue that its central comp

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

Spectral-LSH: Sub-Quadratic Prompt Compression via Krylov-Projected Locality-Sensitive Hashing

DGX agent

arXiv:2607.19368v1 Announce Type: new Abstract: Long-prompt inference remains expensive because prefill attention scales quadratically with sequence length. We propose Spectral-LSH, a training-free pr

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

Start Customizing NVIDIA Nemotron 3 Nano with Prime Intellect Lab in Minutes

DGX agent

Prime Intellect Lab offers a streamlined, hosted reinforcement‑learning workflow that lets users customize the NVIDIA Nemotron 3 Nano in minutes. The process establishes a baseline, trains the model o

model-releasesnvidia-developer
23 Jul 2026
Model Releases

Stateful Guardrails for Multi-Turn LLM Systems: A Conversational Risk Accumulation Framework

DGX agent

arXiv:2607.19361v1 Announce Type: cross Abstract: Most safety guardrails for large language models (LLMs) evaluate each prompt-response pair in isolation, which misses failures that arise only over a

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

Statistical Inference for Rank Allocation in Low-Rank Adaptation

DGX agent

arXiv:2607.20205v1 Announce Type: cross Abstract: Low-rank adaptation (LoRA) has become a widely used parameter-efficient fine-tuning method for large language models. Since different modules and laye

model-releasesarxiv-cs-lg
23 Jul 2026
Model Releases

Statistically Grounded Sparse-Feature Interventions for Activation-Space Control in Large Language Models

DGX agent

arXiv:2607.19364v1 Announce Type: new Abstract: Activation steering offers a lightweight alternative to fine-tuning for behavioral control of large language models, but SAE-based steering methods ofte

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

STN-TGAT: Top-K Portfolio Construction via Prior-Guided Graph Attention with Learnable Soft-Threshold Sparsification

DGX agent

arXiv:2607.19385v1 Announce Type: new Abstract: This paper tackles the problem of stock ranking and portfolio construction under realistic investment settings by jointly modeling temporal dynamics and

model-releasesarxiv-cs-lg
23 Jul 2026
Model Releases

Strength-Parity Ensembling with Parameter-Isolated Experts for Multi-Task Affect Recognition

DGX agent

arXiv:2607.16290v2 Announce Type: replace Abstract: Leading entries on the multi-task track of the 11th ABAW challenge rely on heavy ensembling, yet which member is worth adding to an already strong e

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

StrokeSeg2: Stroke Lesion Segmentation in Clinical Research Workflows

DGX agent

arXiv:2607.19901v1 Announce Type: new Abstract: Deep learning frameworks like nnU-Net achieve state-of-theart brain lesion segmentation performance but remain difficult to deploy in clinical research

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

SUM: Unified Geometric Surgery on Spatio-Temporal Adaptation Vectors for Federated Class Incremental Learning

DGX agent

arXiv:2607.19384v1 Announce Type: new Abstract: Real-world intelligent systems often require both distributed collaboration across data-isolated clients and continual adaptation to evolving tasks. Thi

model-releasesarxiv-cs-lg
23 Jul 2026
Model Releases

SynGallery: A Synthetic Gallery of Real Paintings for Instance-Level Artwork Recognition

DGX agent

arXiv:2607.18907v1 Announce Type: new Abstract: Instance-level artwork recognition requires matching a handheld visitor photograph to a specific work in a large museum collection. This is challenging

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

Task Competence Is Not Instruction Following: Evaluating Instruction-Conflicting Behavior in Small Language Models

DGX agent

arXiv:2607.19608v1 Announce Type: new Abstract: Instruction tuning is meant to make language models follow user requests, yet it is unclear whether small models comply when an instruction conflicts wi

model-releasesarxiv-cs-cl
23 Jul 2026
Model Releases

The Anatomy of a Truth Direction: Knowledge-Dependent Dimensionality, a Relational Law, and a Convergent Category Geometry in Small Language Models

DGX agent

arXiv:2607.16741v2 Announce Type: replace Abstract: Burger et al. (2024) demonstrated that truth representations in large language models are universal across statement polarity but reside within a mu

model-releasesarxiv-cs-lg
23 Jul 2026
Model Releases

The Blessing of Dimensionality: How Near-Orthogonality in High-Dimensional Spaces Explains Temporal Portability

DGX agent

arXiv:2607.20301v1 Announce Type: cross Abstract: Fine-tuning has been widely used to adapt large language models (LLMs) for domain-specific tasks. Parameter efficient fine-tuning (PEFT) methods such

model-releasesarxiv-cs-cl
23 Jul 2026
Model Releases

The Blueprint: How Voicify makes AI-enabled ordering a delight for customers

DGX agent

Welcome to The Blueprint, a new feature where we highlight how Google Cloud customers are tackling unique and common challenges across industries using the latest AI and cloud technologies. We hope to

model-releasesgoogle-cloud-ai
23 Jul 2026
Model Releases

The Chronos Vulnerability: A Taxonomy of Temporal Persistence and Memory-Based Deception in Agentic AI

DGX agent

arXiv:2607.19433v1 Announce Type: new Abstract: The transition from stateless generative models in artificial intelligence to stateful, autonomous agents represents an architectural evolution that, wh

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

The first known runaway AI agent - or a very bad marketing stunt?

DGX agent

The first known runaway AI agent - or a very bad marketing stunt? Martin Alderson's commentary on the OpenAI accidental cyberattack against Hugging Face includes a couple of details I hadn't considere

model-releasessimon-willison
23 Jul 2026
Model Releases

The first router built for generative media is here.

DGX agent

The first router built for generative media is here. Runway launches AI model router as generative media gets crowded https://techcrunch.com/2026/07/23/runway-bets-on-ai-model-routing-as-generative-me

model-releasescristobal-valenzuela--x
23 Jul 2026
Model Releases

The Maskability Index: Predicting Task-Objective Alignment in Pretrained Language Models

DGX agent

arXiv:2607.20265v1 Announce Type: cross Abstract: Large-scale pretrained language models such as T5 and BERT have demonstrated strong capabilities for generating structured knowledge. However, their p

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

The World Model Remembers, the Actor Forgets: Dream Rehearsal for Continual Model-Based RL

DGX agent

arXiv:2607.19749v1 Announce Type: cross Abstract: Model-based reinforcement-learning agents of the DreamerV3 family forget catastrophically when trained on task sequences, even when an unbounded repla

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

Think Sparse, Predict Dense: Continuous Thought Machines for Image Super-Resolution

DGX agent

arXiv:2607.18856v1 Announce Type: new Abstract: Continuous Thought Machines introduce an internal temporal dimension in which neuron-level histories and synchronization-derived representations evolve

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

Toward a Vision-Language Foundation Model for Medical Data: Multimodal Dataset and Benchmarks for Vietnamese PET/CT Report Generation

DGX agent

arXiv:2509.24739v4 Announce Type: replace Abstract: Vision-Language Foundation Models (VLMs), trained on large-scale multimodal datasets, have driven significant advances in Artificial Intelligence (A

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

Toward Anthropomorphic Dialogue: A Closed-Loop Framework for Human-Like Chat Generation, Evaluation, and Preference Alignment

DGX agent

arXiv:2607.17191v2 Announce Type: replace Abstract: Human-like private chat requires more than fluent response generation: a system must preserve persona, relationship, memory, bounded knowledge, medi

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

Train the Model, Not the Reader: Decodability Supervision for Verifiable Activation Explanations

DGX agent

arXiv:2607.20379v1 Announce Type: new Abstract: Natural-language autoencoders score explanations of hidden activations by reconstruction: an explanation is deemed faithful if the activation can be reg

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

Trained a 32B FLUX.2 LoRA on a 24GB AMD 7900 XTX, native ROCm on Windows — full guide + patches

DGX agent

TL;DR: Everyone says QLoRA past ~13B is dead on a 24GB card. I got the full 32B FLUX.2 dev transformer QLoRA-training resident on the GPU on a 7900 XTX under native ROCm on Windows (no ZLUDA, no CUDA

model-releasesr-stablediffusion
23 Jul 2026
Model Releases

Trend strength predicts when generative foundation models win: a power-controlled benchmark, a mechanism, and an actionable selection rule

DGX agent

arXiv:2607.19383v1 Announce Type: cross Abstract: Pretrained generative foundation models cast forecasting as conditional generation from a learned predictive distribution and forecast unseen series z

model-releasesarxiv-cs-lg
23 Jul 2026
Model Releases

TriAgent: Divergence-Aware Multi-Agent Committees for Cost-Efficient Financial Sentiment Analysis

DGX agent

arXiv:2607.19794v1 Announce Type: new Abstract: Production LLM-based financial sentiment analysis faces a structural cost trap: most queries are trivially classifiable, yet expensive cloud reasoners p

model-releasesarxiv-cs-cl
23 Jul 2026
Model Releases

Trusted Multi-View Deep Learning Classification of Fetal Congenital Heart Disease with Feature-level and Decision-level Fusion

DGX agent

arXiv:2606.15265v2 Announce Type: replace Abstract: Congenital heart disease (CHD) refers to the abnormal anatomical structure caused by the abnormal development of the heart and great vessels during

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

Understanding Generative AI-mediated User Engagement with Academic Library Resources

DGX agent

arXiv:2607.20328v1 Announce Type: cross Abstract: This study empirically analyzed generative AI as an emerging discovery pathway to academic library resources. Utilizing web analytics from August 2023

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

Unified Prediction and Planning via Conflict-Aware Disjoint Parameter Training

DGX agent

arXiv:2607.19971v1 Announce Type: new Abstract: Accurate motion prediction of surrounding agents and safe motion planning are two closely coupled key tasks for social robot navigation in crowded envir

model-releasesarxiv-cs-ro
23 Jul 2026
Model Releases

Universality Reconsidered: Rethinking the Validation of Foundation Models for General-Purpose 3D Medical Segmentation

DGX agent

arXiv:2602.07643v2 Announce Type: replace Abstract: Foundation models have emerged as a transformative paradigm in 3D medical imaging, with the promise of unified quantitative analysis across diverse

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

Unlearning as Distribution Restoration: A Controlled Counterfactual Study, a Validated Selective Screen, and the Limits of Oracle-Free Certification

DGX agent

arXiv:2607.19442v1 Announce Type: cross Abstract: Machine unlearning is commonly evaluated by matching a retrained oracle on trained probes. In a controlled nonce-fact testbed with a matched retrainin

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

v0.32.3

DGX agent

What's Changed mlx update by @dhiltgen in #17332 model/parsers: finalize incomplete GLM tool calls by @dhiltgen in #17250 docs: update retirements by @mxyng in #17289 model: align Laguna with upstream

model-releasesollama-releases
23 Jul 2026
Model Releases

VG3S: Visual Geometry Grounded Gaussian Splatting for Semantic Occupancy Prediction

DGX agent

arXiv:2603.06210v2 Announce Type: replace-cross Abstract: 3D semantic occupancy prediction has become a crucial perception task for comprehensive scene understanding in autonomous driving. While recen

model-releasesarxiv-cs-ro
23 Jul 2026
← Previous
1…96979899100…471
Next →