AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,507 results
28 Apr 2026

MetaGAI: A Large-Scale and High-Quality Benchmark for Generative AI Model and Data Card Generation

Model ReleasesDGX agent

arXiv:2604.23539v1 Announce Type: new Abstract: The rapid proliferation of Generative AI necessitates rigorous documentation standards for transparency and governance. However, manual creation of Mode

Mind the Gap: Evaluating Model- and Agentic-Level Vulnerabilities in LLMs with Action Graphs

Model ReleasesDGX agent

arXiv:2509.04802v3 Announce Type: replace Abstract: As large language models increasingly deployed into agentic systems, existing methods face critical gaps in observing, assessing, and mitigating dep

Mixture of Heterogeneous Grouped Experts for Language Modeling

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.23108v1 Announce Type: cross Abstract: Large Language Models (LLMs) based on Mixture-of-Experts (MoE) are pivotal in industrial applications for their ability to scale performance efficient

MLorc: Momentum Low-rank Compression for Memory Efficient Large Language Model Adaptation

Model ReleasesDGX agent

arXiv:2506.01897v5 Announce Type: replace Abstract: With increasing size of large language models (LLMs), full-parameter fine-tuning imposes substantial memory demands. To alleviate this, we propose a

Mobile-R1: Towards Interactive Capability for VLM-Based Mobile Agent via Systematic Training

Model ReleasesDGX agent

arXiv:2506.20332v4 Announce Type: replace Abstract: Vision-language model-based mobile agents have gained the ability to understand complex instructions and mobile screenshots, benefiting from reinfor

Move-Then-Operate: Behavioral Phasing for Human-Like Robotic Manipulation

Model ReleasesDGX agent

arXiv:2604.23620v1 Announce Type: new Abstract: We present Move-Then-Operate, a Vision language action framework that explicitly decouples robotic manipulation into two distinct behavioral phases: coa

MTRouter: Cost-Aware Multi-Turn LLM Routing with History-Model Joint Embeddings

Model ReleasesDGX agent

arXiv:2604.23530v1 Announce Type: cross Abstract: Multi-turn, long-horizon tasks are increasingly common for large language models (LLMs), but solving them typically requires many sequential model inv

Multi-Dimensional Evaluation of Sustainable City Trips with LLM-as-a-Judge and Human-in-the-Loop

Model ReleasesDGX agent

arXiv:2604.24158v1 Announce Type: new Abstract: Evaluating nuanced conversational travel recommendations is challenging when human annotations are costly and standard metrics ignore stakeholder-centri

Multi-View Synergistic Learning with Vision-Language Adaption for Low-Resource Biomedical Image Classification

Model ReleasesDGX agent

arXiv:2604.23977v1 Announce Type: new Abstract: Accurate biomedical image classification under low-resource conditions remains challenging due to limited annotations, subtle inter-class visual differe

MuSS: A Large-Scale Dataset and Cinematic Narrative Benchmark for Multi-Shot Subject-to-Video Generation

Model ReleasesDGX agent

arXiv:2604.23789v1 Announce Type: new Abstract: While video foundation models excel at single-shot generation, real-world cinematic storytelling inherently relies on complex multi-shot sequencing. Fur

Nearly Optimal Subdata Selection

Model ReleasesDGX agent

arXiv:2604.23930v1 Announce Type: cross Abstract: When, in terms of the number of data points, the size of a dataset exceeds available computing resources, or when labeling is expensive, an attractive

Nemotron-3-Nano-Omni-30B-A3B-Reasoning, New model?

Model ReleasesDGX agent

NVIDIA Nemotron 3 Nano Omni is a multimodal large language model that unifies video, audio, image, and text understanding for enterprise Q&A, summarization, transcription, and document intelligence, w

Nemotron 3 Nano Omni is available locally on Ollama! This requires the latest Ollama 0.22 release.

Model ReleasesDGX agent

Nemotron 3 Nano Omni is available locally on Ollama! This requires the latest Ollama 0.22 release. Meet Nemotron 3 Nano Omni 👋 Our latest addition to the Nemotron family is the highest efficiency, ope

Nemotron 3 Nano Omni is now in LM Studio! A new 30B multi-modal MoE from @nvidia Supports Image input, reasoning, and tool use Requires ~25G…

Model ReleasesDGX agent

Nemotron 3 Nano Omni is now in LM Studio! A new 30B multi-modal MoE from @nvidia Supports Image input, reasoning, and tool use Requires ~25GB to run locally 🔥🚀 https://lmstudio.ai/models/nemotron-3-om

NeuroClaw Technical Report

Model ReleasesDGX agent

arXiv:2604.24696v1 Announce Type: new Abstract: Agentic artificial intelligence systems promise to accelerate scientific workflows, but neuroimaging poses unique challenges: heterogeneous modalities (

new in ml-intern: you can now actually see what's going on inside added native metric logging + trackio integration. every training run the …

Model ReleasesDGX agent

new in ml-intern: you can now actually see what's going on inside added native metric logging + trackio integration. every training run the agent kicks off now has live curves you can watch in real ti

No Test Cases, No Problem: Distillation-Driven Code Generation for Scientific Workflows

Model ReleasesDGX agent

arXiv:2604.23106v1 Announce Type: cross Abstract: Existing multi-agent Large Language Model (LLM) frameworks for code generation typically use execution feedback and improve iteratively using Input/Ou

Not All Directions Matter: Towards Structured and Task-Aware Low-Rank Model Adaptation

Model ReleasesDGX agent

arXiv:2603.14228v2 Announce Type: replace Abstract: Low-Rank Adaptation (LoRA) has become a cornerstone of parameter-efficient fine-tuning (PEFT). Yet, its efficacy is hampered by two fundamental limi

NowSecure launches Mobile App Risk Intelligence to expose hidden AI in third-party apps

Model ReleasesDGX agent

Mobile app security company NowSecure Inc. today announced Mobile App Risk Intelligence, new capabilities that are designed to give enterprises evidence-based visibility into third-party mobile apps a

Nvidia introduces Nemotron 3 Nano Omni with vision and speech for powerful agentic AI use

Model ReleasesDGX agent

Nvidia Corp. today launched a powerful reasoning artificial intelligence model that unifies text, vision and speech, capable of acting as the “brains” of faster, smarter agentic AI applications. Dubbe

Nvidia launches Nemotron 3 Nano Omni, an open multimodal model with a 30B-A3B hybrid MoE architecture; the Nemotron 3 family saw 50M+ downloads in the past year (Kyt Dotson/SiliconANGLE)

Model ReleasesDGX agent

Kyt Dotson / SiliconANGLE: Nvidia launches Nemotron 3 Nano Omni, an open multimodal model with a 30B-A3B hybrid MoE architecture; the Nemotron 3 family saw 50M+ downloads in the past year — Nvidia Cor

NVIDIA Launches Nemotron 3 Nano Omni Model, Unifying Vision, Audio and Language for up to 9x More Efficient AI Agents

Model ReleasesDGX agent

AI agent systems today juggle separate models for vision, speech and language — losing time and context as they pass data from one model to the other. Unveiled today, NVIDIA Nemotron 3 Nano Omni is an

@NVIDIA Nemotron 3 Nano Omni is now on Together AI. Enterprise multimodal AI — video, audio, image, documents & text — optimized for speed a…

Model ReleasesDGX agent

@NVIDIA Nemotron 3 Nano Omni is now on Together AI. Enterprise multimodal AI — video, audio, image, documents & text — optimized for speed and scale. ✅ ~3B active params, 9x higher throughput ✅ Fully

NVIDIA Nemotron 3 Nano Omni model now available on Amazon SageMaker JumpStart

Model ReleasesDGX agent

Today, we are excited to announce the day zero availability of NVIDIA Nemotron 3 Nano Omni on Amazon SageMaker JumpStart. In this post, we walk through the model architecture and key capabilities of N

NVIDIA Nemotron™ 3 Nano Omni Now Deployed on Vultr

Model ReleasesDGX agent

NVIDIA Nemotron 3 Nano Omni, a compact multimodal AI model, is now available for deployment on Vultr's cloud infrastructure, enabling developers to run efficient vision and language tasks at scale. Th

NVIDIA Nemotron 3 Nano Omni Powers Multimodal Agent Reasoning in a Single Efficient Open Model

Model ReleasesDGX agent

NVIDIA Nemotron 3 Nano Omni is a 30B hybrid mixture-of-experts model that brings multimodal perception and reasoning into a single system, natively supporting text, image, video, and audio inputs whil

ODE-GS: Latent ODEs for Dynamic Scene Extrapolation with 3D Gaussian Splatting

Model ReleasesDGX agent

arXiv:2506.05480v4 Announce Type: replace-cross Abstract: We introduce ODE-GS, a novel approach that integrates 3D Gaussian Splatting with latent neural ordinary differential equations (ODEs) to enabl

OLaPh: Optimal Language Phonemizer

Model ReleasesDGX agent

arXiv:2509.20086v2 Announce Type: replace Abstract: Phonemization is a critical component in text-to-speech synthesis. Traditional approaches rely on deterministic transformations and lexica, while ne

OmniSch: A Multimodal PCB Schematic Benchmark For Structured Diagram Visual Reasoning

Model ReleasesDGX agent

arXiv:2604.00270v2 Announce Type: replace Abstract: Recent large multimodal models (LMMs) have made rapid progress in visual grounding, document understanding, and diagram reasoning tasks. However, th

OmniShotCut: Holistic Relational Shot Boundary Detection with Shot-Query Transformer

Model ReleasesDGX agent

arXiv:2604.24762v1 Announce Type: new Abstract: Shot Boundary Detection (SBD) aims to automatically identify shot changes and divide a video into coherent shots. While SBD was widely studied in the li

On-Device Vision Training, Deployment, and Inference on a Thumb-Sized Microcontroller

Model ReleasesDGX agent

arXiv:2604.23012v1 Announce Type: cross Abstract: This paper presents a complete, end-to-end on-device vision machine learning pipeline, comprising data acquisition, two-layer CNN training with Adam o

On the Surprising Effectiveness of a Single Global Merging in Decentralized Learning

Model ReleasesDGX agent

arXiv:2507.06542v4 Announce Type: replace Abstract: Decentralized learning provides a scalable alternative to parameter-server-based training, yet its performance is often hindered by limited peer-to-

OntoLogX: Ontology-Guided Knowledge Graph Extraction from Cybersecurity Logs with Large Language Models

Model ReleasesDGX agent

arXiv:2510.01409v2 Announce Type: replace Abstract: System logs represent a valuable source of Cyber Threat Intelligence (CTI), capturing attacker behaviors, exploited vulnerabilities, and traces of m

Open Models & a potential looming Haiku-apocalypse? 🚀 chart comparison UI courtesy of @OpenRouter there's a large class of problems where y…

Model ReleasesDGX agent

Open Models & a potential looming Haiku-apocalypse? 🚀 chart comparison UI courtesy of @OpenRouter there's a large class of problems where you don't need frontier intelligence. you want cheap, fast, an

OpenAI models, Codex, and Managed Agents come to AWS

Model ReleasesDGX agent

OpenAI announced the availability of its models, including Codex, and managed agent capabilities on Amazon Web Services (AWS) infrastructure. This integration enables AWS customers to access OpenAI's

OpenClaw 2026.4.26 🦞 🎙️ Google Live Talk 🦙 Better Ollama/local models 🧳 Bring over Claude + Hermes setups 🔐 One-command Matrix E2EE Big…

Model ReleasesDGX agent

OpenClaw 2026.4.26 🦞 🎙️ Google Live Talk 🦙 Better Ollama/local models 🧳 Bring over Claude + Hermes setups 🔐 One-command Matrix E2EE Big release. Local models eat well. https://github.com/openclaw/open

Optimal Experimental Design for Reliable Learning of History-Dependent Constitutive Laws

Model ReleasesDGX agent

arXiv:2603.12365v2 Announce Type: replace-cross Abstract: History-dependent constitutive models serve as macroscopic closures for the aggregated effects of micromechanics. Their parameters are typical

OptProver: Bridging Olympiad and Optimization through Continual Training in Formal Theorem Proving

Model ReleasesDGX agent

arXiv:2604.23712v1 Announce Type: cross Abstract: Recent advances in formal theorem proving have focused on Olympiad-level mathematics, leaving undergraduate domains largely unexplored. Optimization,

Parameter Efficiency Is Not Memory Efficiency: Rethinking Fine-Tuning for On-Device LLM Adaptation

Model ReleasesDGX agent

arXiv:2604.22783v1 Announce Type: cross Abstract: Parameter-Efficient Fine-Tuning (PEFT) has become the standard for adapting large language models (LLMs). In this work we challenge the wide-spread as

Parameter-Efficient Multi-Task Learning via Progressive Task-Specific Adaptation

Model ReleasesDGX agent

arXiv:2509.19602v2 Announce Type: replace Abstract: Parameter-efficient fine-tuning methods have emerged as a promising solution for adapting pre-trained models to various downstream tasks. While thes

ParkingScenes: A Structured Dataset for End-to-End Autonomous Parking in Simulation Scenes

Model ReleasesDGX agent

arXiv:2604.22835v1 Announce Type: cross Abstract: Autonomous parking remains a critical yet challenging task in intelligent driving systems, particularly within constrained urban environments where ma

Patching LLM Like Software: A Lightweight Method for Improving Safety Policy in Large Language Models

Model ReleasesDGX agent

arXiv:2511.08484v2 Announce Type: replace Abstract: We propose patching for large language models (LLMs) like software versions, a lightweight and modular approach for addressing safety vulnerabilitie

Patterns vs. Patients: Evaluating LLMs against Mental Health Professionals on Personality Disorder Diagnosis through First-Person Narratives

Model ReleasesDGX agent

arXiv:2512.20298v2 Announce Type: replace-cross Abstract: Growing reliance on LLMs for psychiatric self-assessment raises questions about their ability to interpret qualitative patient narratives. Thi

PDF-WuKong: A Large Multimodal Model for Efficient Long PDF Reading with End-to-End Sparse Sampling

Model ReleasesDGX agent

arXiv:2410.05970v3 Announce Type: replace-cross Abstract: Multimodal document understanding is a challenging task to process and comprehend large amounts of textual and visual information. Recent adva

PExA: Parallel Exploration Agent for Complex Text-to-SQL

Model ReleasesDGX agent

arXiv:2604.22934v1 Announce Type: new Abstract: LLM-based agents for text-to-SQL often struggle with latency-performance trade-off, where performance improvements come at the cost of latency or vice v

PhysCodeBench: Benchmarking Physics-Aware Symbolic Simulation of 3D Scenes via Self-Corrective Multi-Agent Refinement

Model ReleasesDGX agent

arXiv:2604.23580v1 Announce Type: cross Abstract: Physics-aware symbolic simulation of 3D scenes is critical for robotics, embodied AI, and scientific computing, requiring models to understand natural

Pi + local models are definitely really cool! Short demo to clean up my Desktop: > terminal 1: llama-server -hf unsloth/Qwen3.5-9B-GGUF:UD-Q…

Model ReleasesDGX agent

Pi + local models are definitely really cool! Short demo to clean up my Desktop: > terminal 1: llama-server -hf unsloth/Qwen3.5-9B-GGUF:UD-Q4_K_XL > terminal 2: simply type 'pi' and start talking to i

PivotMerge: Bridging Heterogeneous Multimodal Pre-training via Post-Alignment Model Merging

Model ReleasesDGX agent

arXiv:2604.22823v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) rely on multimodal pre-training over diverse data sources, where different datasets often induce complementar

PoseX: AI Defeats Physics Approaches on Protein-Ligand Cross Docking

Model ReleasesDGX agent

arXiv:2505.01700v3 Announce Type: replace Abstract: Existing protein-ligand docking studies typically focus on the self-docking scenario, which is less practical in real applications. Moreover, some s

Post-agentic: Encouraging firms to use agentic tools (Claude Code/Lovable/N8N...) markedly improves startup growth & productivity. Pre-agent…

Model ReleasesDGX agent

Post-agentic: Encouraging firms to use agentic tools (Claude Code/Lovable/N8N...) markedly improves startup growth & productivity. Pre-agentic: Encouraging firms to use a GPT4 advisor has uneven effec

PRAXIS: Integrating Program Analysis with Observability for Root-Cause Analysis

Model ReleasesDGX agent

arXiv:2512.22113v2 Announce Type: replace-cross Abstract: Unresolved production cloud incidents cost an average of over $2M per hour. This paper introduces PRAXIS, an orchestrator that manages and dep

Predicting one-year clinical instability and mortality in heart failure patients using sequence modeling

Model ReleasesDGX agent

arXiv:2511.16839v3 Announce Type: replace-cross Abstract: Heart failure (HF) discharge planning depends on identifying patients at risk of deterioration or death, yet accurate prediction from routinel

Predicting Wind Loads on Container Ships in Harbor Environments through Multi-Fidelity Modeling

Model ReleasesDGX agent

arXiv:2604.22882v1 Announce Type: new Abstract: Modern container ships face higher wind loads due to increased windage areas, making accurate predictions of wind loads essential for mooring design. Ex

Pref-CTRL: Preference Driven LLM Alignment using Representation Editing

Model ReleasesDGX agent

arXiv:2604.23543v1 Announce Type: cross Abstract: Test-time alignment methods offer a promising alternative to fine-tuning by steering the outputs of large language models (LLMs) at inference time wit

Primitive Recursion without Composition: Dynamical Characterizations, from Neural Networks to Polynomial ODEs

Model ReleasesDGX agent

arXiv:2604.24356v1 Announce Type: cross Abstract: What do recurrent neural networks, polynomial ODEs, and discrete polynomial maps each bring to computation, and what do they lack? All three operate o

Probe-Based Data Attribution: Discovering and Mitigating Undesirable Behaviors in LLM Post-Training

Model ReleasesDGX agent

arXiv:2602.11079v3 Announce Type: replace-cross Abstract: We propose probe-based data attribution, a method that traces behavioral changes in post-trained language models to responsible training datap

Progressive Approximation in Deep Residual Networks: Theory and Validation

Model ReleasesDGX agent

arXiv:2604.24154v1 Announce Type: cross Abstract: The Universal Approximation Theorem (UAT) guarantees universal function approximation but does not explain how residual models distribute approximatio

Projected Attainable Speed Space: A Driving Efficiency Metric Connecting Instantaneous Evaluation to Travel Time

Model ReleasesDGX agent

arXiv:2604.24295v1 Announce Type: new Abstract: Inefficient driving behaviors, such as overly conservative yielding, remain a key obstacle to deployment of autonomous vehicles (AVs). Instantaneous dri

Psychologically-Grounded Graph Modeling for Interpretable Depression Detection

Model ReleasesDGX agent

arXiv:2604.24126v1 Announce Type: new Abstract: Automatic depression detection from conversational interactions holds significant promise for scalable screening but remains hindered by severe data sca

Putting AI to work: AWS unveils agentic enhancements for Connect and Quick alongside new alliance with OpenAI

Model ReleasesDGX agent

In an indication of how rapidly the world of artificial intelligence is evolving, Amazon Web Services Inc. today unveiled an expanded portfolio of offerings designed to move agentic AI up the software

← Previous
1…307308309310311…376
Next →