AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
All
85,136Total entries
1Added by human
85,135Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,770 results
Model Releases

MesaNet: Sequence Modeling by Locally Optimal Test-Time Training

DGX agent

arXiv:2506.05233v2 Announce Type: replace-cross Abstract: Sequence modeling is currently dominated by causal transformer architectures that use softmax self-attention. Although widely adopted, transfo

model-releasesarxiv-cs-ai
4 Jun 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

MeshTok: Efficient Multi-Scale Tokenization for Scalable PDE Transformers

DGX agent

arXiv:2606.04366v1 Announce Type: new Abstract: Conventional patchified Transformers operate on uniform spatial partitions, distributing computational effort evenly across the domain irrespective of l

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

Metric-Aware Hybrid Forecasting for the CTF4Science Lorenz Challenge

DGX agent

arXiv:2606.04191v1 Announce Type: cross Abstract: We describe our approach to the CTF4Science Lorenz challenge, a benchmark that mixes short-horizon forecasting, long-time distribution matching, and t

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

MimeLens: Position-Agnostic Content-Type Detection for Binary Fragments

DGX agent

arXiv:2606.04171v1 Announce Type: cross Abstract: File-type classification underlies many workflows like malware triage, forensic carving, packet inspection, and storage indexing. Learned systems such

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

MineXplore: An Open-Source Reinforcement Learning Exploration Benchmark for GNSS-Denied Underground Environment

DGX agent

arXiv:2606.04569v1 Announce Type: new Abstract: Underground mines present extreme conditions for autonomous robot navigation: GPS is denied, lighting is degraded, and tunnel topology is loop-rich and

model-releasesarxiv-cs-ro
4 Jun 2026
Model Releases

Most AI pipelines are only as good as the data we provide them with, and that usually means PDFs or other unstructured documents. Contracts,…

DGX agent

Most AI pipelines are only as good as the data we provide them with, and that usually means PDFs or other unstructured documents. Contracts, invoices, reports... All have special layout, language, and

model-releasesjerry-liu--x
4 Jun 2026
Model Releases

Multi-Column RBF Neural Network Using Adaptive and Non-Adaptive Particle Swarm Optimization

DGX agent

arXiv:2606.05150v1 Announce Type: cross Abstract: The radial basis function neural network (RBFN) trained with a gradient descending algorithm provides an effective fully connected structure in both s

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Multi-SPIN: Multi-Access Speculative Inference for Cooperative Token Generation at the Edge

DGX agent

arXiv:2606.04581v1 Announce Type: cross Abstract: Speculative inference (SPIN) was originally developed as an efficient architecture to accelerate Large Language Models (LLMs). In this work, we propos

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Need to Know: Contextual-Integrity-Grounded Query Rewriting for Privacy-Conscious LLM Delegation

DGX agent

arXiv:2606.04067v1 Announce Type: cross Abstract: As LLMs become increasingly woven into everyday workflows, user queries sent to cloud hosted LLMs routinely mix task-essential content with task non-e

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Nemotron 3 Ultra!

DGX agent

Nemotron 3 Ultra is an advanced language model released by NVIDIA, representing an improvement over previous versions in the Nemotron series with enhanced capabilities for various NLP tasks. The annou

model-releasesclem-delangue--x
4 Jun 2026
Model Releases

Nemotron 3 Ultra (550B-A55B) is here - our strongest open-weight model and full training recipe to date. Heavy emphasis on real-world infere…

DGX agent

Nemotron 3 Ultra (550B-A55B) is here - our strongest open-weight model and full training recipe to date. Heavy emphasis on real-world inference efficiency for long-context agentic workloads. Everythin

model-releasesclem-delangue--x
4 Jun 2026
Model Releases

Nemotron 3 Ultra now available on AI Gateway

DGX agent

Nemotron 3 Ultra, NVIDIA's advanced language model, is now accessible through Vercel's AI Gateway, enabling developers to integrate this model into their applications alongside other LLM options. The

model-releasesvercel-blog
4 Jun 2026
Model Releases

Nemotron 3.5 ASR is built for streaming multilingual speech recognition and voice agents. One 0.6B checkpoint. 40 language-locales. Sub-100m…

DGX agent

Nemotron 3.5 ASR is built for streaming multilingual speech recognition and voice agents. One 0.6B checkpoint. 40 language-locales. Sub-100ms latency. Cache-aware FastConformer carries context forward

model-releasestogether-ai--x
4 Jun 2026
Model Releases

Nemotron 3.5 Content Safety: Customizable Multimodal Safety for Global Enterprise AI

DGX agent

Nemotron 3.5 Content Safety is NVIDIA's multimodal safety solution designed for enterprise AI applications, offering customizable safeguards for both text and image inputs across different global cont

model-releaseshugging-face
4 Jun 2026
Model Releases

Neural Galerkin Normalizing Flows for Bayesian Inference of Diffusions with Inaccessible Boundaries

DGX agent

arXiv:2606.04324v1 Announce Type: new Abstract: One of the primary challenges in Bayesian inference on the parameters of a diffusion model from discrete observations is the unavailability of an analyt

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

New Benchmarking Shows Limited Generalization Power of TCR Antigenic Epitope Prediction Models

DGX agent

arXiv:2606.04994v1 Announce Type: new Abstract: Accurate computational prediction of T cell receptor (TCR) antigen specificity would transform the study of T cell biology and enable scalable immune en

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

New course on serving LLMs efficiently -- how do you serve models to many concurrent users at low latency and reasonable cost? This short co…

DGX agent

New course on serving LLMs efficiently -- how do you serve models to many concurrent users at low latency and reasonable cost? This short course is built with @RedHat and taught by @cedricclyburn. Eff

model-releasesandrew-ng--x
4 Jun 2026
Model Releases

NEW: NVIDIA ships 550B MoE open model for long-running agents. Very exciting times to see more open models to support local long-running cod…

DGX agent

NEW: NVIDIA ships 550B MoE open model for long-running agents. Very exciting times to see more open models to support local long-running coding agents. Today we're shipping Nemotron 3 Ultra. A 550B Mo

model-releasesdair-ai--x
4 Jun 2026
Model Releases

NextMotionQA: Benchmarking and Judging Human Motion Understanding with Vision-Language Models

DGX agent

arXiv:2606.04773v1 Announce Type: cross Abstract: Reliable evaluation of human motion understanding is fundamental to advancing embodied AI, robotics, and animation. However, existing benchmarks suffe

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

NoRA: Evaluating Grounded Reasonableness in Visual First-person Normative Action Reasoning

DGX agent

arXiv:2606.04806v1 Announce Type: cross Abstract: LLMs and agentic systems are increasingly deployed in social environments, making normative competence critical for safe and appropriate behavior. How

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Not All Errors Are Equal: Consequence-Aware Reasoning Compute Allocation

DGX agent

arXiv:2606.04402v1 Announce Type: new Abstract: Modern reasoning models can allocate different amounts of test-time computation, such as thinking tokens, model calls, or compute budget, to different t

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

@nvidia @nebiustf More info on the new Nemotron 3 Ultra: https://x.com/NVIDIAAI/status/2062521325076299981?s=20

DGX agent

@nvidia @nebiustf More info on the new Nemotron 3 Ultra: https://x.com/NVIDIAAI/status/2062521325076299981?s=20 Today we're shipping Nemotron 3 Ultra. A 550B MoE frontier-intelligence open model built

model-releasesnous-research--x
4 Jun 2026
Model Releases

@nvidia @nebiustf Setup guide: http://hermes-agent.nousresearch.com/docs/guides/run-nemotron-3-ultra-free Sign up for Nous Portal: http://po…

DGX agent

This is a setup guide for running Nemotron-3 Ultra, NVIDIA's open-source language model, through Nous Research's platform. The guide directs users to sign up for the Nous Portal and access documentati

model-releasesnous-research--x
4 Jun 2026
Model Releases

NVIDIA Nemotron 3 Ultra is on Fireworks, day zero. Nemotron Ultra is an open model for frontier reasoning and orchestration in long-running …

DGX agent

NVIDIA Nemotron 3 Ultra is on Fireworks, day zero. Nemotron Ultra is an open model for frontier reasoning and orchestration in long-running autonomous agents. Think use cases like coding agents, deep

model-releasesfireworks-ai--x
4 Jun 2026
Model Releases

NVIDIA Nemotron 3 Ultra now available on Amazon SageMaker JumpStart

DGX agent

The search results primarily discuss the Nemotron 3 Nano Omni model rather than Nemotron 3 Ultra. However, I found a recent NVIDIA blog post reference that indicates Nemotron 3 Ultra is available thro

model-releasesaws-ml-blog
4 Jun 2026
Model Releases

NVIDIA Nemotron 3 Ultra Powers Faster, More Efficient Reasoning for Long-Running Agents

DGX agent

NVIDIA's Nemotron 3 Ultra is a 550B-parameter Mixture-of-Experts model with 55B active parameters, optimized for orchestrating complex, long-running agent workflows by combining frontier reasoning and

model-releasesnvidia-developer
4 Jun 2026
Model Releases

NVIDIA Nemotron 3.5 Content Safety Now Available on Vultr

DGX agent

Deploy NVIDIA Nemotron 3.5 Content Safety on Vultr Cloud GPU with Day Zero support for multimodal AI moderation, custom policy enforcement, multilingual safety workflows, and scalable enterprise AI go

model-releasesvultr
4 Jun 2026
Model Releases

NVIDIA’s Nemotron 3 Ultra is available on Ollama’s cloud! Try it 👇 Claude Code: ollama launch claude --model nemotron-3-ultra:cloud Hermes …

DGX agent

NVIDIA’s Nemotron 3 Ultra is available on Ollama’s cloud! Try it 👇 Claude Code: ollama launch claude --model nemotron-3-ultra:cloud Hermes Agent: ollama launch hermes --model nemotron-3-ultra:cloud Op

model-releasesollama--x
4 Jun 2026
Model Releases

OckBench: Measuring the Efficiency of LLM Reasoning

DGX agent

arXiv:2511.05722v3 Announce Type: replace-cross Abstract: Large language models (LLMs) such as GPT-5 and Gemini 3 have pushed the frontier of automated reasoning and code generation. Yet current bench

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Offroad launches with $7M to automate identity security with AI agents

DGX agent

Offroad Inc. launched today with 7 million in funding to build what it calls an agentic identity security team, using artificial intelligence agents to investigate and remediate access risks across hu

model-releasessiliconangle
4 Jun 2026
Model Releases

ollama run gemma4:12b Gemma 4 12B is updated on Ollama, and available across all platforms! Try it on: Claude Code ollama launch claude --mo…

DGX agent

ollama run gemma4:12b Gemma 4 12B is updated on Ollama, and available across all platforms! Try it on: Claude Code ollama launch claude --model gemma4:12b Hermes Agent ollama launch hermes --model gem

model-releasesollama--x
4 Jun 2026
Model Releases

On the Relationship Between CoCoA and ADMM for Distributed Empirical Risk Minimization

DGX agent

arXiv:2502.00470v3 Announce Type: replace-cross Abstract: Distributed empirical risk minimization (ERM) is often studied through two influential yet seemingly separate families of methods: CoCoA-type

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

On TV Tokyo’s WBS (@wbs_tvtokyo) tonight I’ll be discussing Sakana AI’s upcoming 1T parameter model project, supported by METI’s GENIAC init…

DGX agent

On TV Tokyo’s WBS (@wbs_tvtokyo) tonight I’ll be discussing Sakana AI’s upcoming 1T parameter model project, supported by METI’s GENIAC initiative. We are scaling up to build Japan’s first 1T paramete

model-releasesdavid-ha--x
4 Jun 2026
Model Releases

Online Skill Learning for Web Agents via State-Grounded Dynamic Retrieval

DGX agent

arXiv:2606.04391v1 Announce Type: new Abstract: Language agents increasingly rely on reusable skills to improve multi-step web automation across related tasks. A growing line of work studies online sk

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Optical-Guided Neural Collapse for SAR Few-Shot Class Incremental Learning

DGX agent

arXiv:2606.04528v1 Announce Type: cross Abstract: Few-shot class-incremental learning (FSCIL) in synthetic aperture radar imagery presents unique challenges due to severe data scarcity and SAR-specifi

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Our internal data shows Claude is accelerating AI development—a possible path to recursive self-improvement, or AI autonomously building a m…

DGX agent

Our internal data shows Claude is accelerating AI development—a possible path to recursive self-improvement, or AI autonomously building a more capable successor. It’s happening faster than we thought

model-releasesboris-cherny--x
4 Jun 2026
Model Releases

Our team is at CVPR 2026 if you want to come say hi :)

DGX agent

Our team is at CVPR 2026 if you want to come say hi :) We're presenting ParseBench at CVPR 2026! ParseBench is the most comprehensive document understanding benchmark for VLMs. ✅ It contains 2k pages

model-releasesjerry-liu--x
4 Jun 2026
Model Releases

Outstanding paper on long-horizon agents. (bookmark it) Similar to humans, how do you make agents persist on a difficult task, and how is th…

DGX agent

Outstanding paper on long-horizon agents. (bookmark it) Similar to humans, how do you make agents persist on a difficult task, and how is that useful? And which models today work well on this? This ne

model-releasesdair-ai--x
4 Jun 2026
Model Releases

Parameter-Efficient Fine-Tuning with Learnable Rank

DGX agent

arXiv:2606.04325v1 Announce Type: new Abstract: Low-Rank Adaptation (LoRA) is a popular parameter-efficient fine-tuning (PEFT) method that restricts weight updates to low-rank adapters, introducing a

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

PE-MHL: Physics-Encoded Modular Hybrid Layers for Scalable Learning of Complex Systems

DGX agent

arXiv:2606.04290v1 Announce Type: new Abstract: Hybrid models that combine physics-based and data-driven components have shown strong potential for achieving accuracy and interpretability in control a

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

PersistBench: When Should Long-Term Memories Be Forgotten by LLMs?

DGX agent

arXiv:2602.01146v2 Announce Type: replace Abstract: Conversational assistants are increasingly integrating long-term memory with large language models (LLMs). This persistence of memories, e.g., the u

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Physics-Informed Video Generation via Mixture-of-Experts Latent Alignment

DGX agent

arXiv:2606.04737v1 Announce Type: new Abstract: Large-scale video generation models have made remarkable progress in semantic consistency and visual quality, producing videos that are increasingly coh

model-releasesarxiv-cs-cv
4 Jun 2026
Model Releases

Plan, Watch, Recover: A Benchmark and Architectures for Proactive Procedural Assistance

DGX agent

arXiv:2606.04970v1 Announce Type: cross Abstract: We envision a proactive multi-modal assistant system which gives users real-time step-by-step guidance on a procedural task, autonomously deciding ext

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

PoliticsBench: Benchmarking Political Values in Large Language Models with Multi-Turn Roleplay

DGX agent

arXiv:2603.23841v2 Announce Type: replace-cross Abstract: While Large Language Models (LLMs) are increasingly used as primary sources of information, their potential for political bias may impact thei

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Prompt-Level Distillation: A Non-Parametric Alternative to Model Fine-Tuning for Efficient Reasoning

DGX agent

arXiv:2602.21103v2 Announce Type: replace Abstract: Advanced reasoning typically requires Chain-of-Thought prompting, which is accurate but incurs prohibitive latency and substantial test-time inferen

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

Proof-Carrying Agent Actions: Model-Agnostic Runtime Governance for Heterogeneous Agent Systems

DGX agent

arXiv:2606.04104v1 Announce Type: cross Abstract: Agent systems execute through runtimes with very different control points: local coding tools, framework SDKs, managed agent platforms, API gateways,

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Provably Reduced Sample Cost in Prior-Guided Hyperparameter Optimization

DGX agent

arXiv:2606.04866v1 Announce Type: new Abstract: Large-scale hyperparameter optimization (HPO) in automated machine learning (AutoML) consumes substantial computational resources, raising growing conce

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

Pseudospectral Bounds for Transient Amplification in Coupled Gradient Descent

DGX agent

arXiv:2606.04031v1 Announce Type: new Abstract: Coupled gradient descent--where the update of one parameter block depends on another--underlies bilevel optimization, two-time-scale stochastic approxim

model-releasesarxiv-cs-lg
4 Jun 2026
← Previous
1…216217218219220…475
Next →