AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
All
85,136Total entries
1Added by human
85,135Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,778 results
Model Releases

LLMs + Persona-Plug = Personalized LLMs

DGX agent

arXiv:2409.11901v2 Announce Type: replace Abstract: Personalization plays a critical role in numerous language tasks and applications, since users with the same requirements may prefer diverse outputs

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

Long Live Fine-Tuning: Task-Specific Transformers Outperform Zero-Shot LLMs for Misinformation Response Classification on Reddit

Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2606.04274v1 Announce Type: new Abstract: As large language models (LLMs) become default tools for online information verification, an implicit assumption follows them: that scale and general ca

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

Longer Context, Deeper Thinking: Uncovering the Role of Long-Context Ability in Reasoning

DGX agent

arXiv:2505.17315v2 Announce Type: replace Abstract: Recent language models exhibit strong reasoning capabilities, yet the influence of long-context capacity on reasoning remains underexplored. In this

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Look closely. There’s more in the Showcase.

DGX agent

OpenAI's developer account posted this message on X (formerly Twitter), likely encouraging developers to explore additional features, updates, or resources available in OpenAI's Showcase platform or d

model-releasesopenai--x
4 Jun 2026
Model Releases

LoopMoE: Unifying Iterative Computation with Mixture-of-Experts for Language Modeling

DGX agent

arXiv:2606.04438v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) and looped architectures scale models along two orthogonal axes, namely parameter capacity and effective depth. However, main

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

M^3Eval: Multi-Modal Memory Evaluation through Cognitively-Grounded Video Tasks

DGX agent

arXiv:2606.05008v1 Announce Type: cross Abstract: As multi-modal models advance towards long-form video understanding, memory emerges as a critical capability. Despite substantial efforts in developin

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

MedForge: Interpretable Medical Deepfake Detection via Forgery-aware Reasoning

DGX agent

arXiv:2603.18577v2 Announce Type: replace Abstract: Text-guided image editors can now manipulate authentic medical scans with high fidelity, enabling lesion implantation/removal that threatens clinica

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

MemoryDocDataSet: A Benchmark for Joint Conversational Memory and Long Document Reasoning

DGX agent

arXiv:2606.04442v1 Announce Type: cross Abstract: AI systems increasingly need to combine two demanding capabilities: navigating multi-session conversation history and performing deep reading comprehe

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

MesaNet: Sequence Modeling by Locally Optimal Test-Time Training

DGX agent

arXiv:2506.05233v2 Announce Type: replace-cross Abstract: Sequence modeling is currently dominated by causal transformer architectures that use softmax self-attention. Although widely adopted, transfo

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

MeshTok: Efficient Multi-Scale Tokenization for Scalable PDE Transformers

DGX agent

arXiv:2606.04366v1 Announce Type: new Abstract: Conventional patchified Transformers operate on uniform spatial partitions, distributing computational effort evenly across the domain irrespective of l

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

Metric-Aware Hybrid Forecasting for the CTF4Science Lorenz Challenge

DGX agent

arXiv:2606.04191v1 Announce Type: cross Abstract: We describe our approach to the CTF4Science Lorenz challenge, a benchmark that mixes short-horizon forecasting, long-time distribution matching, and t

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

MimeLens: Position-Agnostic Content-Type Detection for Binary Fragments

DGX agent

arXiv:2606.04171v1 Announce Type: cross Abstract: File-type classification underlies many workflows like malware triage, forensic carving, packet inspection, and storage indexing. Learned systems such

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

MineXplore: An Open-Source Reinforcement Learning Exploration Benchmark for GNSS-Denied Underground Environment

DGX agent

arXiv:2606.04569v1 Announce Type: new Abstract: Underground mines present extreme conditions for autonomous robot navigation: GPS is denied, lighting is degraded, and tunnel topology is loop-rich and

model-releasesarxiv-cs-ro
4 Jun 2026
Model Releases

Most AI pipelines are only as good as the data we provide them with, and that usually means PDFs or other unstructured documents. Contracts,…

DGX agent

Most AI pipelines are only as good as the data we provide them with, and that usually means PDFs or other unstructured documents. Contracts, invoices, reports... All have special layout, language, and

model-releasesjerry-liu--x
4 Jun 2026
Model Releases

Multi-Column RBF Neural Network Using Adaptive and Non-Adaptive Particle Swarm Optimization

DGX agent

arXiv:2606.05150v1 Announce Type: cross Abstract: The radial basis function neural network (RBFN) trained with a gradient descending algorithm provides an effective fully connected structure in both s

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Multi-SPIN: Multi-Access Speculative Inference for Cooperative Token Generation at the Edge

DGX agent

arXiv:2606.04581v1 Announce Type: cross Abstract: Speculative inference (SPIN) was originally developed as an efficient architecture to accelerate Large Language Models (LLMs). In this work, we propos

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Need to Know: Contextual-Integrity-Grounded Query Rewriting for Privacy-Conscious LLM Delegation

DGX agent

arXiv:2606.04067v1 Announce Type: cross Abstract: As LLMs become increasingly woven into everyday workflows, user queries sent to cloud hosted LLMs routinely mix task-essential content with task non-e

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Nemotron 3 Ultra!

DGX agent

Nemotron 3 Ultra is an advanced language model released by NVIDIA, representing an improvement over previous versions in the Nemotron series with enhanced capabilities for various NLP tasks. The annou

model-releasesclem-delangue--x
4 Jun 2026
Model Releases

Nemotron 3 Ultra (550B-A55B) is here - our strongest open-weight model and full training recipe to date. Heavy emphasis on real-world infere…

DGX agent

Nemotron 3 Ultra (550B-A55B) is here - our strongest open-weight model and full training recipe to date. Heavy emphasis on real-world inference efficiency for long-context agentic workloads. Everythin

model-releasesclem-delangue--x
4 Jun 2026
Model Releases

Nemotron 3 Ultra now available on AI Gateway

DGX agent

Nemotron 3 Ultra, NVIDIA's advanced language model, is now accessible through Vercel's AI Gateway, enabling developers to integrate this model into their applications alongside other LLM options. The

model-releasesvercel-blog
4 Jun 2026
Model Releases

Nemotron 3.5 ASR is built for streaming multilingual speech recognition and voice agents. One 0.6B checkpoint. 40 language-locales. Sub-100m…

DGX agent

Nemotron 3.5 ASR is built for streaming multilingual speech recognition and voice agents. One 0.6B checkpoint. 40 language-locales. Sub-100ms latency. Cache-aware FastConformer carries context forward

model-releasestogether-ai--x
4 Jun 2026
Model Releases

Nemotron 3.5 Content Safety: Customizable Multimodal Safety for Global Enterprise AI

DGX agent

Nemotron 3.5 Content Safety is NVIDIA's multimodal safety solution designed for enterprise AI applications, offering customizable safeguards for both text and image inputs across different global cont

model-releaseshugging-face
4 Jun 2026
Model Releases

Neural Galerkin Normalizing Flows for Bayesian Inference of Diffusions with Inaccessible Boundaries

DGX agent

arXiv:2606.04324v1 Announce Type: new Abstract: One of the primary challenges in Bayesian inference on the parameters of a diffusion model from discrete observations is the unavailability of an analyt

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

New Benchmarking Shows Limited Generalization Power of TCR Antigenic Epitope Prediction Models

DGX agent

arXiv:2606.04994v1 Announce Type: new Abstract: Accurate computational prediction of T cell receptor (TCR) antigen specificity would transform the study of T cell biology and enable scalable immune en

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

New course on serving LLMs efficiently -- how do you serve models to many concurrent users at low latency and reasonable cost? This short co…

DGX agent

New course on serving LLMs efficiently -- how do you serve models to many concurrent users at low latency and reasonable cost? This short course is built with @RedHat and taught by @cedricclyburn. Eff

model-releasesandrew-ng--x
4 Jun 2026
Model Releases

NEW: NVIDIA ships 550B MoE open model for long-running agents. Very exciting times to see more open models to support local long-running cod…

DGX agent

NEW: NVIDIA ships 550B MoE open model for long-running agents. Very exciting times to see more open models to support local long-running coding agents. Today we're shipping Nemotron 3 Ultra. A 550B Mo

model-releasesdair-ai--x
4 Jun 2026
Model Releases

NextMotionQA: Benchmarking and Judging Human Motion Understanding with Vision-Language Models

DGX agent

arXiv:2606.04773v1 Announce Type: cross Abstract: Reliable evaluation of human motion understanding is fundamental to advancing embodied AI, robotics, and animation. However, existing benchmarks suffe

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

NoRA: Evaluating Grounded Reasonableness in Visual First-person Normative Action Reasoning

DGX agent

arXiv:2606.04806v1 Announce Type: cross Abstract: LLMs and agentic systems are increasingly deployed in social environments, making normative competence critical for safe and appropriate behavior. How

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Not All Errors Are Equal: Consequence-Aware Reasoning Compute Allocation

DGX agent

arXiv:2606.04402v1 Announce Type: new Abstract: Modern reasoning models can allocate different amounts of test-time computation, such as thinking tokens, model calls, or compute budget, to different t

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

@nvidia @nebiustf More info on the new Nemotron 3 Ultra: https://x.com/NVIDIAAI/status/2062521325076299981?s=20

DGX agent

@nvidia @nebiustf More info on the new Nemotron 3 Ultra: https://x.com/NVIDIAAI/status/2062521325076299981?s=20 Today we're shipping Nemotron 3 Ultra. A 550B MoE frontier-intelligence open model built

model-releasesnous-research--x
4 Jun 2026
Model Releases

@nvidia @nebiustf Setup guide: http://hermes-agent.nousresearch.com/docs/guides/run-nemotron-3-ultra-free Sign up for Nous Portal: http://po…

DGX agent

This is a setup guide for running Nemotron-3 Ultra, NVIDIA's open-source language model, through Nous Research's platform. The guide directs users to sign up for the Nous Portal and access documentati

model-releasesnous-research--x
4 Jun 2026
Model Releases

NVIDIA Nemotron 3 Ultra is on Fireworks, day zero. Nemotron Ultra is an open model for frontier reasoning and orchestration in long-running …

DGX agent

NVIDIA Nemotron 3 Ultra is on Fireworks, day zero. Nemotron Ultra is an open model for frontier reasoning and orchestration in long-running autonomous agents. Think use cases like coding agents, deep

model-releasesfireworks-ai--x
4 Jun 2026
Model Releases

NVIDIA Nemotron 3 Ultra now available on Amazon SageMaker JumpStart

DGX agent

The search results primarily discuss the Nemotron 3 Nano Omni model rather than Nemotron 3 Ultra. However, I found a recent NVIDIA blog post reference that indicates Nemotron 3 Ultra is available thro

model-releasesaws-ml-blog
4 Jun 2026
Model Releases

NVIDIA Nemotron 3 Ultra Powers Faster, More Efficient Reasoning for Long-Running Agents

DGX agent

NVIDIA's Nemotron 3 Ultra is a 550B-parameter Mixture-of-Experts model with 55B active parameters, optimized for orchestrating complex, long-running agent workflows by combining frontier reasoning and

model-releasesnvidia-developer
4 Jun 2026
Model Releases

NVIDIA Nemotron 3.5 Content Safety Now Available on Vultr

DGX agent

Deploy NVIDIA Nemotron 3.5 Content Safety on Vultr Cloud GPU with Day Zero support for multimodal AI moderation, custom policy enforcement, multilingual safety workflows, and scalable enterprise AI go

model-releasesvultr
4 Jun 2026
Model Releases

NVIDIA’s Nemotron 3 Ultra is available on Ollama’s cloud! Try it 👇 Claude Code: ollama launch claude --model nemotron-3-ultra:cloud Hermes …

DGX agent

NVIDIA’s Nemotron 3 Ultra is available on Ollama’s cloud! Try it 👇 Claude Code: ollama launch claude --model nemotron-3-ultra:cloud Hermes Agent: ollama launch hermes --model nemotron-3-ultra:cloud Op

model-releasesollama--x
4 Jun 2026
Model Releases

OckBench: Measuring the Efficiency of LLM Reasoning

DGX agent

arXiv:2511.05722v3 Announce Type: replace-cross Abstract: Large language models (LLMs) such as GPT-5 and Gemini 3 have pushed the frontier of automated reasoning and code generation. Yet current bench

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Offroad launches with $7M to automate identity security with AI agents

DGX agent

Offroad Inc. launched today with 7 million in funding to build what it calls an agentic identity security team, using artificial intelligence agents to investigate and remediate access risks across hu

model-releasessiliconangle
4 Jun 2026
Model Releases

ollama run gemma4:12b Gemma 4 12B is updated on Ollama, and available across all platforms! Try it on: Claude Code ollama launch claude --mo…

DGX agent

ollama run gemma4:12b Gemma 4 12B is updated on Ollama, and available across all platforms! Try it on: Claude Code ollama launch claude --model gemma4:12b Hermes Agent ollama launch hermes --model gem

model-releasesollama--x
4 Jun 2026
Model Releases

On the Relationship Between CoCoA and ADMM for Distributed Empirical Risk Minimization

DGX agent

arXiv:2502.00470v3 Announce Type: replace-cross Abstract: Distributed empirical risk minimization (ERM) is often studied through two influential yet seemingly separate families of methods: CoCoA-type

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

On TV Tokyo’s WBS (@wbs_tvtokyo) tonight I’ll be discussing Sakana AI’s upcoming 1T parameter model project, supported by METI’s GENIAC init…

DGX agent

On TV Tokyo’s WBS (@wbs_tvtokyo) tonight I’ll be discussing Sakana AI’s upcoming 1T parameter model project, supported by METI’s GENIAC initiative. We are scaling up to build Japan’s first 1T paramete

model-releasesdavid-ha--x
4 Jun 2026
Model Releases

Online Skill Learning for Web Agents via State-Grounded Dynamic Retrieval

DGX agent

arXiv:2606.04391v1 Announce Type: new Abstract: Language agents increasingly rely on reusable skills to improve multi-step web automation across related tasks. A growing line of work studies online sk

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Optical-Guided Neural Collapse for SAR Few-Shot Class Incremental Learning

DGX agent

arXiv:2606.04528v1 Announce Type: cross Abstract: Few-shot class-incremental learning (FSCIL) in synthetic aperture radar imagery presents unique challenges due to severe data scarcity and SAR-specifi

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Our internal data shows Claude is accelerating AI development—a possible path to recursive self-improvement, or AI autonomously building a m…

DGX agent

Our internal data shows Claude is accelerating AI development—a possible path to recursive self-improvement, or AI autonomously building a more capable successor. It’s happening faster than we thought

model-releasesboris-cherny--x
4 Jun 2026
Model Releases

Our team is at CVPR 2026 if you want to come say hi :)

DGX agent

Our team is at CVPR 2026 if you want to come say hi :) We're presenting ParseBench at CVPR 2026! ParseBench is the most comprehensive document understanding benchmark for VLMs. ✅ It contains 2k pages

model-releasesjerry-liu--x
4 Jun 2026
Model Releases

Outstanding paper on long-horizon agents. (bookmark it) Similar to humans, how do you make agents persist on a difficult task, and how is th…

DGX agent

Outstanding paper on long-horizon agents. (bookmark it) Similar to humans, how do you make agents persist on a difficult task, and how is that useful? And which models today work well on this? This ne

model-releasesdair-ai--x
4 Jun 2026
Model Releases

Parameter-Efficient Fine-Tuning with Learnable Rank

DGX agent

arXiv:2606.04325v1 Announce Type: new Abstract: Low-Rank Adaptation (LoRA) is a popular parameter-efficient fine-tuning (PEFT) method that restricts weight updates to low-rank adapters, introducing a

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

PE-MHL: Physics-Encoded Modular Hybrid Layers for Scalable Learning of Complex Systems

DGX agent

arXiv:2606.04290v1 Announce Type: new Abstract: Hybrid models that combine physics-based and data-driven components have shown strong potential for achieving accuracy and interpretability in control a

model-releasesarxiv-cs-lg
4 Jun 2026
← Previous
1…217218219220221…475
Next →