AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,553 results
Model Releases

Mixture of Heterogeneous Grouped Experts for Language Modeling

DGX agent

arXiv:2604.23108v1 Announce Type: cross Abstract: Large Language Models (LLMs) based on Mixture-of-Experts (MoE) are pivotal in industrial applications for their ability to scale performance efficient

model-releasesarxiv-cs-ai
28 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

MLorc: Momentum Low-rank Compression for Memory Efficient Large Language Model Adaptation

DGX agent

arXiv:2506.01897v5 Announce Type: replace Abstract: With increasing size of large language models (LLMs), full-parameter fine-tuning imposes substantial memory demands. To alleviate this, we propose a

model-releasesarxiv-cs-lg
28 Apr 2026
Model Releases

Mobile-R1: Towards Interactive Capability for VLM-Based Mobile Agent via Systematic Training

DGX agent

arXiv:2506.20332v4 Announce Type: replace Abstract: Vision-language model-based mobile agents have gained the ability to understand complex instructions and mobile screenshots, benefiting from reinfor

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Move-Then-Operate: Behavioral Phasing for Human-Like Robotic Manipulation

DGX agent

arXiv:2604.23620v1 Announce Type: new Abstract: We present Move-Then-Operate, a Vision language action framework that explicitly decouples robotic manipulation into two distinct behavioral phases: coa

model-releasesarxiv-cs-ro
28 Apr 2026
Model Releases

MTRouter: Cost-Aware Multi-Turn LLM Routing with History-Model Joint Embeddings

DGX agent

arXiv:2604.23530v1 Announce Type: cross Abstract: Multi-turn, long-horizon tasks are increasingly common for large language models (LLMs), but solving them typically requires many sequential model inv

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Multi-Dimensional Evaluation of Sustainable City Trips with LLM-as-a-Judge and Human-in-the-Loop

DGX agent

arXiv:2604.24158v1 Announce Type: new Abstract: Evaluating nuanced conversational travel recommendations is challenging when human annotations are costly and standard metrics ignore stakeholder-centri

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Multi-View Synergistic Learning with Vision-Language Adaption for Low-Resource Biomedical Image Classification

DGX agent

arXiv:2604.23977v1 Announce Type: new Abstract: Accurate biomedical image classification under low-resource conditions remains challenging due to limited annotations, subtle inter-class visual differe

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

MuSS: A Large-Scale Dataset and Cinematic Narrative Benchmark for Multi-Shot Subject-to-Video Generation

DGX agent

arXiv:2604.23789v1 Announce Type: new Abstract: While video foundation models excel at single-shot generation, real-world cinematic storytelling inherently relies on complex multi-shot sequencing. Fur

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

Nearly Optimal Subdata Selection

DGX agent

arXiv:2604.23930v1 Announce Type: cross Abstract: When, in terms of the number of data points, the size of a dataset exceeds available computing resources, or when labeling is expensive, an attractive

model-releasesarxiv-cs-lg
28 Apr 2026
Model Releases

Nemotron-3-Nano-Omni-30B-A3B-Reasoning, New model?

DGX agent

NVIDIA Nemotron 3 Nano Omni is a multimodal large language model that unifies video, audio, image, and text understanding for enterprise Q&A, summarization, transcription, and document intelligence, w

model-releasesr-localllama
28 Apr 2026
Model Releases

Nemotron 3 Nano Omni is available locally on Ollama! This requires the latest Ollama 0.22 release.

DGX agent

Nemotron 3 Nano Omni is available locally on Ollama! This requires the latest Ollama 0.22 release. Meet Nemotron 3 Nano Omni 👋 Our latest addition to the Nemotron family is the highest efficiency, ope

model-releasesollama--x
28 Apr 2026
Model Releases

Nemotron 3 Nano Omni is now in LM Studio! A new 30B multi-modal MoE from @nvidia Supports Image input, reasoning, and tool use Requires ~25G…

DGX agent

Nemotron 3 Nano Omni is now in LM Studio! A new 30B multi-modal MoE from @nvidia Supports Image input, reasoning, and tool use Requires ~25GB to run locally 🔥🚀 https://lmstudio.ai/models/nemotron-3-om

model-releaseslm-studio--x
28 Apr 2026
Model Releases

NeuroClaw Technical Report

DGX agent

arXiv:2604.24696v1 Announce Type: new Abstract: Agentic artificial intelligence systems promise to accelerate scientific workflows, but neuroimaging poses unique challenges: heterogeneous modalities (

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

new in ml-intern: you can now actually see what's going on inside added native metric logging + trackio integration. every training run the …

DGX agent

new in ml-intern: you can now actually see what's going on inside added native metric logging + trackio integration. every training run the agent kicks off now has live curves you can watch in real ti

model-releasesclem-delangue--x
28 Apr 2026
Model Releases

No Test Cases, No Problem: Distillation-Driven Code Generation for Scientific Workflows

DGX agent

arXiv:2604.23106v1 Announce Type: cross Abstract: Existing multi-agent Large Language Model (LLM) frameworks for code generation typically use execution feedback and improve iteratively using Input/Ou

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Not All Directions Matter: Towards Structured and Task-Aware Low-Rank Model Adaptation

DGX agent

arXiv:2603.14228v2 Announce Type: replace Abstract: Low-Rank Adaptation (LoRA) has become a cornerstone of parameter-efficient fine-tuning (PEFT). Yet, its efficacy is hampered by two fundamental limi

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

NowSecure launches Mobile App Risk Intelligence to expose hidden AI in third-party apps

DGX agent

Mobile app security company NowSecure Inc. today announced Mobile App Risk Intelligence, new capabilities that are designed to give enterprises evidence-based visibility into third-party mobile apps a

model-releasessiliconangle
28 Apr 2026
Model Releases

Nvidia introduces Nemotron 3 Nano Omni with vision and speech for powerful agentic AI use

DGX agent

Nvidia Corp. today launched a powerful reasoning artificial intelligence model that unifies text, vision and speech, capable of acting as the “brains” of faster, smarter agentic AI applications. Dubbe

model-releasessiliconangle
28 Apr 2026
Model Releases

Nvidia launches Nemotron 3 Nano Omni, an open multimodal model with a 30B-A3B hybrid MoE architecture; the Nemotron 3 family saw 50M+ downloads in the past year (Kyt Dotson/SiliconANGLE)

DGX agent

Kyt Dotson / SiliconANGLE: Nvidia launches Nemotron 3 Nano Omni, an open multimodal model with a 30B-A3B hybrid MoE architecture; the Nemotron 3 family saw 50M+ downloads in the past year — Nvidia Cor

model-releasestechmeme
28 Apr 2026
Model Releases

NVIDIA Launches Nemotron 3 Nano Omni Model, Unifying Vision, Audio and Language for up to 9x More Efficient AI Agents

DGX agent

AI agent systems today juggle separate models for vision, speech and language — losing time and context as they pass data from one model to the other. Unveiled today, NVIDIA Nemotron 3 Nano Omni is an

model-releasesnvidia-blog
28 Apr 2026
Model Releases

@NVIDIA Nemotron 3 Nano Omni is now on Together AI. Enterprise multimodal AI — video, audio, image, documents & text — optimized for speed a…

DGX agent

@NVIDIA Nemotron 3 Nano Omni is now on Together AI. Enterprise multimodal AI — video, audio, image, documents & text — optimized for speed and scale. ✅ ~3B active params, 9x higher throughput ✅ Fully

model-releasestogether-ai--x
28 Apr 2026
Model Releases

NVIDIA Nemotron 3 Nano Omni model now available on Amazon SageMaker JumpStart

DGX agent

Today, we are excited to announce the day zero availability of NVIDIA Nemotron 3 Nano Omni on Amazon SageMaker JumpStart. In this post, we walk through the model architecture and key capabilities of N

model-releasesaws-ml-blog
28 Apr 2026
Model Releases

NVIDIA Nemotron™ 3 Nano Omni Now Deployed on Vultr

DGX agent

NVIDIA Nemotron 3 Nano Omni, a compact multimodal AI model, is now available for deployment on Vultr's cloud infrastructure, enabling developers to run efficient vision and language tasks at scale. Th

model-releasesvultr
28 Apr 2026
Model Releases

NVIDIA Nemotron 3 Nano Omni Powers Multimodal Agent Reasoning in a Single Efficient Open Model

DGX agent

NVIDIA Nemotron 3 Nano Omni is a 30B hybrid mixture-of-experts model that brings multimodal perception and reasoning into a single system, natively supporting text, image, video, and audio inputs whil

model-releasesnvidia-developer
28 Apr 2026
Model Releases

ODE-GS: Latent ODEs for Dynamic Scene Extrapolation with 3D Gaussian Splatting

DGX agent

arXiv:2506.05480v4 Announce Type: replace-cross Abstract: We introduce ODE-GS, a novel approach that integrates 3D Gaussian Splatting with latent neural ordinary differential equations (ODEs) to enabl

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

OLaPh: Optimal Language Phonemizer

DGX agent

arXiv:2509.20086v2 Announce Type: replace Abstract: Phonemization is a critical component in text-to-speech synthesis. Traditional approaches rely on deterministic transformations and lexica, while ne

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

OmniSch: A Multimodal PCB Schematic Benchmark For Structured Diagram Visual Reasoning

DGX agent

arXiv:2604.00270v2 Announce Type: replace Abstract: Recent large multimodal models (LMMs) have made rapid progress in visual grounding, document understanding, and diagram reasoning tasks. However, th

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

OmniShotCut: Holistic Relational Shot Boundary Detection with Shot-Query Transformer

DGX agent

arXiv:2604.24762v1 Announce Type: new Abstract: Shot Boundary Detection (SBD) aims to automatically identify shot changes and divide a video into coherent shots. While SBD was widely studied in the li

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

On-Device Vision Training, Deployment, and Inference on a Thumb-Sized Microcontroller

DGX agent

arXiv:2604.23012v1 Announce Type: cross Abstract: This paper presents a complete, end-to-end on-device vision machine learning pipeline, comprising data acquisition, two-layer CNN training with Adam o

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

On the Surprising Effectiveness of a Single Global Merging in Decentralized Learning

DGX agent

arXiv:2507.06542v4 Announce Type: replace Abstract: Decentralized learning provides a scalable alternative to parameter-server-based training, yet its performance is often hindered by limited peer-to-

model-releasesarxiv-cs-lg
28 Apr 2026
Model Releases

OntoLogX: Ontology-Guided Knowledge Graph Extraction from Cybersecurity Logs with Large Language Models

DGX agent

arXiv:2510.01409v2 Announce Type: replace Abstract: System logs represent a valuable source of Cyber Threat Intelligence (CTI), capturing attacker behaviors, exploited vulnerabilities, and traces of m

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Open Models & a potential looming Haiku-apocalypse? 🚀 chart comparison UI courtesy of @OpenRouter there's a large class of problems where y…

DGX agent

Open Models & a potential looming Haiku-apocalypse? 🚀 chart comparison UI courtesy of @OpenRouter there's a large class of problems where you don't need frontier intelligence. you want cheap, fast, an

model-releasesharrison-chase--x
28 Apr 2026
Model Releases

OpenAI models, Codex, and Managed Agents come to AWS

DGX agent

OpenAI announced the availability of its models, including Codex, and managed agent capabilities on Amazon Web Services (AWS) infrastructure. This integration enables AWS customers to access OpenAI's

model-releasesopenai
28 Apr 2026
Model Releases

OpenClaw 2026.4.26 🦞 🎙️ Google Live Talk 🦙 Better Ollama/local models 🧳 Bring over Claude + Hermes setups 🔐 One-command Matrix E2EE Big…

DGX agent

OpenClaw 2026.4.26 🦞 🎙️ Google Live Talk 🦙 Better Ollama/local models 🧳 Bring over Claude + Hermes setups 🔐 One-command Matrix E2EE Big release. Local models eat well. https://github.com/openclaw/open

model-releasesollama--x
28 Apr 2026
Model Releases

Optimal Experimental Design for Reliable Learning of History-Dependent Constitutive Laws

DGX agent

arXiv:2603.12365v2 Announce Type: replace-cross Abstract: History-dependent constitutive models serve as macroscopic closures for the aggregated effects of micromechanics. Their parameters are typical

model-releasesarxiv-cs-lg
28 Apr 2026
Model Releases

OptProver: Bridging Olympiad and Optimization through Continual Training in Formal Theorem Proving

DGX agent

arXiv:2604.23712v1 Announce Type: cross Abstract: Recent advances in formal theorem proving have focused on Olympiad-level mathematics, leaving undergraduate domains largely unexplored. Optimization,

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Parameter Efficiency Is Not Memory Efficiency: Rethinking Fine-Tuning for On-Device LLM Adaptation

DGX agent

arXiv:2604.22783v1 Announce Type: cross Abstract: Parameter-Efficient Fine-Tuning (PEFT) has become the standard for adapting large language models (LLMs). In this work we challenge the wide-spread as

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Parameter-Efficient Multi-Task Learning via Progressive Task-Specific Adaptation

DGX agent

arXiv:2509.19602v2 Announce Type: replace Abstract: Parameter-efficient fine-tuning methods have emerged as a promising solution for adapting pre-trained models to various downstream tasks. While thes

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

ParkingScenes: A Structured Dataset for End-to-End Autonomous Parking in Simulation Scenes

DGX agent

arXiv:2604.22835v1 Announce Type: cross Abstract: Autonomous parking remains a critical yet challenging task in intelligent driving systems, particularly within constrained urban environments where ma

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Patching LLM Like Software: A Lightweight Method for Improving Safety Policy in Large Language Models

DGX agent

arXiv:2511.08484v2 Announce Type: replace Abstract: We propose patching for large language models (LLMs) like software versions, a lightweight and modular approach for addressing safety vulnerabilitie

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Patterns vs. Patients: Evaluating LLMs against Mental Health Professionals on Personality Disorder Diagnosis through First-Person Narratives

DGX agent

arXiv:2512.20298v2 Announce Type: replace-cross Abstract: Growing reliance on LLMs for psychiatric self-assessment raises questions about their ability to interpret qualitative patient narratives. Thi

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

PDF-WuKong: A Large Multimodal Model for Efficient Long PDF Reading with End-to-End Sparse Sampling

DGX agent

arXiv:2410.05970v3 Announce Type: replace-cross Abstract: Multimodal document understanding is a challenging task to process and comprehend large amounts of textual and visual information. Recent adva

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

PExA: Parallel Exploration Agent for Complex Text-to-SQL

DGX agent

arXiv:2604.22934v1 Announce Type: new Abstract: LLM-based agents for text-to-SQL often struggle with latency-performance trade-off, where performance improvements come at the cost of latency or vice v

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

PhysCodeBench: Benchmarking Physics-Aware Symbolic Simulation of 3D Scenes via Self-Corrective Multi-Agent Refinement

DGX agent

arXiv:2604.23580v1 Announce Type: cross Abstract: Physics-aware symbolic simulation of 3D scenes is critical for robotics, embodied AI, and scientific computing, requiring models to understand natural

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Pi + local models are definitely really cool! Short demo to clean up my Desktop: > terminal 1: llama-server -hf unsloth/Qwen3.5-9B-GGUF:UD-Q…

DGX agent

Pi + local models are definitely really cool! Short demo to clean up my Desktop: > terminal 1: llama-server -hf unsloth/Qwen3.5-9B-GGUF:UD-Q4_K_XL > terminal 2: simply type 'pi' and start talking to i

model-releasesclem-delangue--x
28 Apr 2026
Model Releases

PivotMerge: Bridging Heterogeneous Multimodal Pre-training via Post-Alignment Model Merging

DGX agent

arXiv:2604.22823v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) rely on multimodal pre-training over diverse data sources, where different datasets often induce complementar

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

PoseX: AI Defeats Physics Approaches on Protein-Ligand Cross Docking

DGX agent

arXiv:2505.01700v3 Announce Type: replace Abstract: Existing protein-ligand docking studies typically focus on the self-docking scenario, which is less practical in real applications. Moreover, some s

model-releasesarxiv-cs-lg
28 Apr 2026
Model Releases

Post-agentic: Encouraging firms to use agentic tools (Claude Code/Lovable/N8N...) markedly improves startup growth & productivity. Pre-agent…

DGX agent

Post-agentic: Encouraging firms to use agentic tools (Claude Code/Lovable/N8N...) markedly improves startup growth & productivity. Pre-agentic: Encouraging firms to use a GPT4 advisor has uneven effec

model-releasesethan-mollick--x
28 Apr 2026
← Previous
1…385386387388389…470
Next →