AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
All
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,585 results
Model Releases

Multi-Narrow Transformation as a Single-Model Ensemble: Boundary Conditions, Mechanisms, and Failure Modes

DGX agent

arXiv:2605.11530v1 Announce Type: new Abstract: Single-model ensembles (SMEs) have attracted attention as a way to approximate some of the benefits of deep ensembles within a single network. However,

model-releasesarxiv-cs-lg
13 May 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Multi-Task Representation Learning for Conservative Linear Bandits

DGX agent

arXiv:2605.12176v1 Announce Type: new Abstract: This paper presents the Constrained Multi-Task Representation Learning (CMTRL) framework for linear bandits. We consider T linear bandit tasks in a d di

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

MuonQ: Enhancing Low-Bit Muon Quantization via Directional Fidelity Optimization

DGX agent

arXiv:2605.11396v1 Announce Type: new Abstract: The Muon optimizer has emerged as a compelling alternative to Adam for training large language models, achieving remarkable computational savings throug

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Mythos Preview is the first AI model to complete both of AISI's cyber ranges, which measure models' cyberattack capabilities; GPT-5.5 solved only one of them (AI Security Institute)

DGX agent

AI Security Institute: Mythos Preview is the first AI model to complete both of AISI's cyber ranges, which measure models' cyberattack capabilities; GPT-5.5 solved only one of them — In February 2026,

model-releasestechmeme
13 May 2026
Model Releases

Nautilus: From One Prompt to Plug-and-Play Robot Learning

DGX agent

arXiv:2605.11665v1 Announce Type: new Abstract: Robot learning research is fragmented across policy families, benchmark suites, and real robots; each implementation is entangled with the others in a c

model-releasesarxiv-cs-ro
13 May 2026
Model Releases

NavOL: Navigation Policy with Online Imitation Learning

DGX agent

arXiv:2605.11762v1 Announce Type: new Abstract: Learning robust navigation policies remains a core challenge in robotics. Offline imitation learning suffers from distribution shift and compounding err

model-releasesarxiv-cs-ro
13 May 2026
Model Releases

Neural ARFIMA model for forecasting BRIC exchange rates with long memory

DGX agent

arXiv:2509.06697v2 Announce Type: replace-cross Abstract: Accurate forecasting of exchange rates remains a persistent challenge, particularly for emerging economies such as Brazil, Russia, India, and

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

No More, No Less: Task Alignment in Terminal Agents

DGX agent

arXiv:2605.12233v1 Announce Type: new Abstract: Terminal agents are increasingly capable of executing complex, long-horizon tasks autonomously from a single user prompt. To do so, they must interpret

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Not How Many, But Which: Parameter Placement in Low-Rank Adaptation

DGX agent

arXiv:2605.12207v1 Announce Type: cross Abstract: We study the extit{parameter placement problem}: given a fixed budget of k trainable entries within the B matrix of a LoRA adapter (A frozen), does th

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

@nvidia Nemotron native support in Deep Agents 0.6 #interrupt #langchain

DGX agent

NVIDIA Nemotron models received native support integration in Deep Agents version 0.6, enabling improved language model capabilities within the LangChain framework. This update allows developers to le

model-releasesharrison-chase--x
13 May 2026
Model Releases

Our continued commitment to Chromebooks, and looking ahead

DGX agent

At the Android Show yesterday, we introduced Googlebooks: a new category of premium laptops built with Gemini’s helpfulness at the core. Designed for Gemini Intelligence, Googlebooks will give persona

model-releasesgoogle-cloud-ai
13 May 2026
Model Releases

Output Composability of QLoRA PEFT Modules for Plug-and-Play Attribute-Controlled Text Generation

DGX agent

arXiv:2605.12345v1 Announce Type: new Abstract: Parameter-efficient fine-tuning (PEFT) techniques offer task-specific fine-tuning at a fraction of the cost of full fine-tuning, but require separate fi

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

Overcoming Dynamics-Blindness: Training-Free Pace-and-Path Correction for VLA Models

DGX agent

arXiv:2605.11459v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models achieve remarkable flexibility and generalization beyond classical control paradigms. However, most prevailing VLA

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

Overtrained, Not Misaligned

DGX agent

arXiv:2605.12199v1 Announce Type: new Abstract: Emergent misalignment (EM), where fine-tuning on a narrow task (like insecure code) causes broad misalignment across unrelated domains, was first demons

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Overview of the MedHopQA track at BioCreative IX: track description, participation and evaluation of systems for multi-hop medical question answering

DGX agent

arXiv:2605.12313v1 Announce Type: new Abstract: Multi-hop question answering (QA) remains a significant challenge in the biomedical domain, requiring systems to integrate information across multiple s

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

Parameter-Efficient Adaptation of Pre-Trained Vision Foundation Models for Active and Passive Seismic Data Denoising

DGX agent

arXiv:2605.10953v1 Announce Type: cross Abstract: The demand for high-resolution subsurface imaging and continuous Earth monitoring has driven rapid growth in active and passive seismic data from dens

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

Patterns behind Chaos: Forecasting Data Movement for Efficient Large-Scale MoE LLM Inference

DGX agent

arXiv:2510.05497v5 Announce Type: replace-cross Abstract: Large-scale Mixture of Experts (MoE) Large Language Models (LLMs) have recently become the frontier open-weight models, achieving remarkable m

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

PD-4DGS:Progressive Decomposition of 4D Gaussian Splatting for Bandwidth-Adaptive Dynamic Scene Streaming

DGX agent

arXiv:2605.11427v1 Announce Type: new Abstract: 4D Gaussian Splatting (4DGS) enables high-quality dynamic novel view synthesis, yet current models remain monolithic bitstreams that clients must downlo

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

Picasso: Holistic Scene Reconstruction with Physics-Constrained Sampling

DGX agent

arXiv:2602.08058v2 Announce Type: replace Abstract: In the presence of occlusions and measurement noise, geometrically accurate scene reconstructions -- which fit the sensor data -- can still be physi

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

POP: Prior-Fitted First-Order Optimization Policies

DGX agent

arXiv:2602.15473v2 Announce Type: replace Abstract: Gradient-based optimizers are highly sensitive to design choices in their adaptive learning rate mechanisms. To address this limitation, we introduc

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

PoseBridge: Bridging the Skeletonization Gap for Zero-Shot Skeleton-Based Action Recognition

DGX agent

arXiv:2605.11497v1 Announce Type: new Abstract: Zero-shot skeleton-based action recognition (ZSSAR) is typically treated as a skeleton-text alignment problem: encode joint-coordinate sequences, align

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

Posterior Contraction Rates for Sparse Kolmogorov-Arnold Networks in Anisotropic Besov Spaces

DGX agent

arXiv:2605.11652v1 Announce Type: cross Abstract: We study posterior contraction rates for sparse Bayesian Kolmogorov-Arnold networks (KANs) over anisotropic Besov spaces, providing a statistical foun

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Predicting Psychological Well-Being from Spontaneous Speech using LLMs

DGX agent

arXiv:2605.11303v1 Announce Type: new Abstract: We investigate the use of Large Language Models (LLMs) for zero-shot prediction of Ryff Psychological Well-Being (PWB) scores from spontaneous speech. U

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

Premover: Fast Vision-Language-Action Control by Acting Before Instructions Are Complete

DGX agent

arXiv:2605.12160v1 Announce Type: new Abstract: Vision-Language-Action (VLA) policies are typically evaluated as if the user had finished typing or speaking before the robot begins acting. In real dep

model-releasesarxiv-cs-ro
13 May 2026
Model Releases

PreScam: A Benchmark for Predicting Scam Progression from Early Conversations

DGX agent

arXiv:2605.12243v1 Announce Type: new Abstract: Conversational scams, such as romance and investment scams, are emerging as a major form of online fraud. Unlike one-shot scam lures such as fake lotter

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

PresentAgent-2: Towards Generalist Multimodal Presentation Agents

DGX agent

arXiv:2605.11363v1 Announce Type: cross Abstract: Presentation generation is moving beyond static slide creation toward end-to-end presentation video generation with research grounding, multimodal med

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

PRISM: Pareto-Efficient Retrieval over Intent-Aware Structured Memory for Long-Horizon Agents

DGX agent

arXiv:2605.12260v1 Announce Type: new Abstract: Long-horizon language agents accumulate conversation history far faster than any fixed context window can hold, making memory management critical to bot

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

PRISM: : Planning and Reasoning with Intent in Simulated Embodied Environments

DGX agent

arXiv:2605.11534v1 Announce Type: new Abstract: When an LLM-based embodied agent fails at a household task, the culprit could be misidentified objects, forgotten sub-goals, or poor action sequencing -

model-releasesarxiv-cs-ro
13 May 2026
Model Releases

Probabilistic Calibration Is a Trainable Capability in Language Models

DGX agent

arXiv:2605.11845v1 Announce Type: new Abstract: Language models are increasingly used in settings where outputs must satisfy user-specified randomness constraints, yet their generation probabilities a

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

Probabilistic Computers for Neural Quantum States

DGX agent

arXiv:2512.24558v2 Announce Type: replace-cross Abstract: Neural quantum states efficiently represent many-body wavefunctions with neural networks, but the cost of Monte Carlo sampling limits their sc

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Probing Non-Equilibrium Grain Boundary Dynamics with XPCS and Domain-Adaptive Machine Learning

DGX agent

arXiv:2605.12194v1 Announce Type: cross Abstract: Grain-boundary (GB) dynamics control the stability, mechanical, and functional response of nanocrystalline materials, but direct experimental access t

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Procedural-skill SFT across capacity tiers: A W-Shaped pre-SFT Trajectory and Regime-Asymmetric Mechanism on 0.8B-4B Qwen3.5 Models

DGX agent

arXiv:2605.11907v1 Announce Type: new Abstract: We measure procedural-skill SFT contribution across three Qwen3.5 dense scales (0.8B, 2B, 4B) on a 200-task / 40-skill holdout, with Claude Haiku 4.5 as

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Provably Data-driven Multiple Hyper-parameter Tuning with Structured Loss Function

DGX agent

arXiv:2602.02406v2 Announce Type: replace-cross Abstract: Data-driven algorithm design automates hyperparameter tuning, but its statistical foundations remain limited because model performance can dep

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

PVLM: Parsing-Aware Vision Language Model with Dynamic Contrastive Learning for Zero-Shot Deepfake Attribution

DGX agent

arXiv:2504.14129v4 Announce Type: replace Abstract: The challenge of tracing the source attribution of forged faces has gained significant attention due to the rapid advancement of generative models.

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

QuIDE: Mastering the Quantized Intelligence Trade-off via Active Optimization

DGX agent

arXiv:2605.10959v1 Announce Type: new Abstract: There is currently no unified metric for evaluating the efficiency of quantized neural networks. We propose QuIDE, built around the Intelligence Index I

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Quite excited about llama-eval, a proposed eval tool for llama.cpp. Could be a nice step toward more comparable community evals 🎉 https://g…

DGX agent

Llama-eval is a proposed evaluation tool for llama.cpp designed to standardize and improve comparability of community-run evaluations. The tool aims to address inconsistencies in how different users b

model-releasesgeorgi-gerganov--x
13 May 2026
Model Releases

Qwen-Scope: Turning Sparse Features into Development Tools for Large Language Models

DGX agent

arXiv:2605.11887v1 Announce Type: new Abstract: Large language models have achieved remarkable capabilities across diverse tasks, yet their internal decision-making processes remain largely opaque, li

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

🚀Qwen3.6-Plus is on Nous Portal now and FREE for a limited time. Hermes Agent, here we go!! ⚡️ @NousResearch

DGX agent

🚀Qwen3.6-Plus is on Nous Portal now and FREE for a limited time. Hermes Agent, here we go!! ⚡️ @NousResearch Qwen 3.6 Plus by @Alibaba_Qwen is now FREE for a limited time on Nous Portal! Nous Portal i

model-releasesqwen--x
13 May 2026
Model Releases

READ: Recurrent Adapter with Partial Video-Language Alignment for Parameter-Efficient Transfer Learning in Low-Resource Video-Language Modeling

DGX agent

arXiv:2312.06950v3 Announce Type: replace-cross Abstract: Fully fine-tuning pretrained large-scale transformer models has become a popular paradigm for video-language modeling tasks, such as temporal

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

Really curious when Gemini is going to join the Cowork & Codex race to build a local app that isn’t just for developers. Antigravity hasn’t …

DGX agent

Really curious when Gemini is going to join the Cowork & Codex race to build a local app that isn’t just for developers. Antigravity hasn’t posted updates to X in a month, and remains very software fo

model-releasesethan-mollick--x
13 May 2026
Model Releases

Reconsidering the energy efficiency of spiking neural networks

DGX agent

arXiv:2409.08290v4 Announce Type: replace-cross Abstract: Spiking Neural Networks (SNNs) promise higher energy efficiency over conventional Quantized Artificial Neural Networks (QNNs) due to their eve

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Reconstructing Sepsis Trajectories from Clinical Case Reports using LLMs: the Textual Time Series Corpus for Sepsis

DGX agent

arXiv:2504.12326v3 Announce Type: replace Abstract: Clinical case reports and discharge summaries may be the most complete and accurate summarization of patient encounters, yet they are finalized, i.e

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

Rethinking LLMOps for Fraud and AML: Building a Compliance-Grade LLM Serving Stack

DGX agent

arXiv:2605.11232v1 Announce Type: cross Abstract: Fraud detection and anti-money-laundering (AML) compliance are high-value domains for large language models (LLMs), but their serving requirements dif

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Revisiting Shadow Detection from a Vision-Language Perspective

DGX agent

arXiv:2605.11771v1 Announce Type: new Abstract: Shadow detection is commonly formulated as a vision-driven dense prediction problem, where models rely primarily on pixel-wise visual supervision to dis

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

Reviving In-domain Fine-tuning Methods for Source-Free Cross-domain Few-shot Learning

DGX agent

arXiv:2605.11659v1 Announce Type: new Abstract: Cross-Domain Few-Shot Learning (CDFSL) aims to adapt large-scale pretrained models to specialized target domains with limited samples, yet the few-shot

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

Robust Promptable Video Object Segmentation

DGX agent

arXiv:2605.12006v1 Announce Type: new Abstract: The performance of promptable video object segmentation (PVOS) models substantially degrades under input corruptions, which prevents PVOS deployment in

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

ROMER: Expert Replacement and Router Calibration for Robust MoE LLMs on Analog Compute-in-Memory Systems

DGX agent

arXiv:2605.11800v1 Announce Type: cross Abstract: Large language models (LLMs) with mixture-of-experts (MoE) architectures achieve remarkable scalability by sparsely activating a subset of experts per

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

Routers Learn the Geometry of Their Experts: Geometric Coupling in Sparse Mixture-of-Experts

DGX agent

arXiv:2605.12476v1 Announce Type: cross Abstract: Sparse Mixture-of-Experts (SMoE) models enable scaling language models efficiently, but training them remains challenging, as routing can collapse ont

model-releasesarxiv-cs-cl
13 May 2026
← Previous
1…322323324325326…471
Next →