AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
85,188Total entries
1Added by human
85,187Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,805 results
Model Releases

Many research labs only consider inference efficiency after the fact. Step 3.7 Flash is a 196B MoE model, and built for inference from the s…

DGX agent

Many research labs only consider inference efficiency after the fact. Step 3.7 Flash is a 196B MoE model, and built for inference from the start by @StepFun_ai. Multi-Matrix Factorization Attention (M

model-releasesfireworks-ai--x
1 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

MAVEN: Improving Generalization in Agentic Tool Calling

DGX agent

arXiv:2605.30738v1 Announce Type: new Abstract: Generalization across agentic tool-calling environments remains a central challenge for reliable agentic reasoning systems. Although large language mode

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

MedFact: Benchmarking the Fact-Checking Capabilities of Large Language Models on Chinese Medical Texts

DGX agent

arXiv:2509.12440v3 Announce Type: replace-cross Abstract: Deploying Large Language Models (LLMs) in medical applications requires fact-checking capabilities to ensure patient safety and regulatory com

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Mellum2 Technical Report

DGX agent

arXiv:2605.31268v1 Announce Type: new Abstract: We present Mellum 2, an open-weight 12B-parameter Mixture-of-Experts (MoE) language model with 2.5B active parameters per token. Mellum 2 is a general-p

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

Memory-Bound but Not Bandwidth-Limited: The Physical AI Inference Gap in Batch-1 LLM Decode

DGX agent

arXiv:2605.30571v1 Announce Type: cross Abstract: Physical AI systems, including robots, autonomous vehicles, embodied agents and edge copilots, often run a different inference workload from cloud LLM

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Memory by Design: Probabilistic Sequence Layers

DGX agent

arXiv:2605.31163v1 Announce Type: cross Abstract: We introduce the design-model framework: a way to derive efficient recurrent sequence maps from explicit assumptions about memory. A design model writ

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

Merge launches Agent Handler for Employees as an IT gatekeeper for workplace AI agents

DGX agent

Merge API Inc., a platform provider delivering connective infrastructure for artificial intelligence to business data and tools, launched Agent Handler for Employees, easing the strain on information

model-releasessiliconangle
1 Jun 2026
Model Releases

MIMO: Multilingual Information Retrieval via Monolingual Objectives

DGX agent

arXiv:2605.31171v1 Announce Type: cross Abstract: Multilingual Information Retrieval (MLIR) reflects real-world search environments in which queries and relevant documents may appear in different lang

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

MineExplorer: Evaluating Open-World Exploration of MLLM Agents in Minecraft

DGX agent

arXiv:2605.30931v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have shown strong capabilities in perception, reasoning, and action generation. However, their ability to susta

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

.@MiniMax_AI M3 model is available on Ollama's Cloud! In partnership with MiniMax, the M3 model on Ollama's Cloud is US-based with zero data…

DGX agent

.@MiniMax_AI M3 model is available on Ollama's Cloud! In partnership with MiniMax, the M3 model on Ollama's Cloud is US-based with zero data retention. Try M3 on coding and agentic tasks: Claude Code:

model-releasesollama--x
1 Jun 2026
Model Releases

MLIPilot: LLM-Driven Auto-Research for Machine-Learned Interatomic Potentials

DGX agent

arXiv:2605.30889v1 Announce Type: cross Abstract: Constructing production-quality machine-learned interatomic potentials (MLIPs) requires balancing accuracy, dynamical stability, and computational thr

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

MosaicLeaks:Privacy Risks in Querying-in-the-Open for Deep Research Agents

DGX agent

arXiv:2605.30727v1 Announce Type: new Abstract: Deep research agents increasingly combine private local documents with external tools like web retrieval, creating a privacy risk: an agent's external q

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

MultiPriv: Benchmarking Individual-Level Privacy Reasoning in Vision-Language Models

DGX agent

arXiv:2511.16940v3 Announce Type: replace Abstract: Modern Vision-Language Models (VLMs) pose significant individual-level privacy risks by linking fragmented multimodal data to identifiable individua

model-releasesarxiv-cs-cv
1 Jun 2026
Model Releases

Nemotron 3 Ultra: Frontier smart. 5X faster. 30% cheaper. 💚💚💚

DGX agent

Nemotron 3 Ultra is NVIDIA's latest language model featuring significant improvements in speed (5X faster) and cost efficiency (30% cheaper) compared to previous versions, positioning it as a frontier

model-releasesjeremy-howard--x
1 Jun 2026
Model Releases

NeUQI: Near-Optimal Uniform Quantization Parameter Initialization for Low-Bit LLMs

DGX agent

arXiv:2505.17595v4 Announce Type: replace-cross Abstract: Large language models (LLMs) achieve impressive performance across domains but face significant challenges when deployed on consumer-grade GPU

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

Neuro-symbolic Syntactic Parsing: Shaping a Neural Network with the CYK Algorithm

DGX agent

arXiv:2605.31421v1 Announce Type: cross Abstract: In this paper, we show the possibility of a direct injection of algorithms into neural network architecture. We focus on a complex algorithm, that is,

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

NGDBench: Towards Neural Graph Data Management

DGX agent

arXiv:2603.05529v2 Announce Type: replace-cross Abstract: Data critical to real-world decision-making is increasingly found within organizations. Such data is heterogeneous, constantly evolving, and o

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Not All Synthetic Data Is Yours to Learn From

DGX agent

arXiv:2605.31126v1 Announce Type: cross Abstract: Can a language model improve from plain text sampled from itself, with no prompts, no teacher, no verifier, and no reward model? Yes, but only when th

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

nuReasoning: A Reasoning-Centric Dataset and Benchmark for Long-Tail Autonomous Driving

DGX agent

arXiv:2605.31572v1 Announce Type: new Abstract: Reasoning is essential for autonomous driving (AD) in long-tail scenarios, where vehicles must apply commonsense knowledge, understand spatial relations

model-releasesarxiv-cs-cv
1 Jun 2026
Model Releases

Nvidia launches Nemotron 3 Ultra, a 550B-parameter MoE open model; Artificial Analysis: it's the smartest open US model but trails the Chinese model Kimi K2.6 (Maximilian Schreiner/The Decoder)

DGX agent

Maximilian Schreiner / The Decoder: Nvidia launches Nemotron 3 Ultra, a 550B-parameter MoE open model; Artificial Analysis: it's the smartest open US model but trails the Chinese model Kimi K2.6 — It

model-releasestechmeme
1 Jun 2026
Model Releases

OBCache: Optimal Brain KV Cache Pruning for Efficient Long-Context LLM Inference

DGX agent

arXiv:2510.07651v2 Announce Type: replace-cross Abstract: Large language models (LLMs) with extended context windows enable powerful applications but impose significant memory overhead, as caching all

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Omni-Supervised Motion Editing: Balancing Change and Invariance through Positive-Negative Learning

DGX agent

arXiv:2605.30969v1 Announce Type: new Abstract: Text-based human motion editing aims to modify existing motion sequences according to natural language instructions while maintaining the consistency of

model-releasesarxiv-cs-cv
1 Jun 2026
Model Releases

On-Device Generative AI for GDPR-Compliant Visual Monitoring: Natural Language Alerts from Local Object Detection

DGX agent

arXiv:2605.30544v1 Announce Type: new Abstract: Visual monitoring systems that rely on cloud-based AI inference expose raw image data to external services, creating fundamental tensions with the data-

model-releasesarxiv-cs-cv
1 Jun 2026
Model Releases

On the Robustness of Multilingual Text Embedding Rankings Across Learning Tasks, Languages, and Benchmark Datasets

DGX agent

arXiv:2605.31142v1 Announce Type: cross Abstract: Large-scale multilingual text embedding models play crucial role in both research and industry, yet their behavior in language-specific, multi-task se

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

One of the big reasons for the current lack of patriotism and pride in our nation’s history is that about 40 years ago our most prominent st…

DGX agent

One of the big reasons for the current lack of patriotism and pride in our nation’s history is that about 40 years ago our most prominent storytellers in Hollywood just basically stopped telling stori

model-releaseselon-musk--x
1 Jun 2026
Model Releases

One of the new, buzzy jobs in Silicon Valley is the AI Forward Deployed Engineer (FDE), an engineer who is embedded within a client organiza…

DGX agent

One of the new, buzzy jobs in Silicon Valley is the AI Forward Deployed Engineer (FDE), an engineer who is embedded within a client organization to help customize solutions, such as building and tunin

model-releasesandrew-ng--x
1 Jun 2026
Model Releases

OpenAI frontier models and Codex are now generally available on AWS, giving enterprises a new way to build on Amazon Bedrock with OpenAI thr…

DGX agent

OpenAI frontier models and Codex are now generally available on AWS, giving enterprises a new way to build on Amazon Bedrock with OpenAI through the security, compliance, and governance workflows they

model-releasesopenai--x
1 Jun 2026
Model Releases

OpenAI models and Codex on Amazon Bedrock are now generally available

DGX agent

OpenAI frontier models GPT-5.5 and GPT-5.4, and Codex, the OpenAI coding agent, are now generally available on Amazon Bedrock. AWS customers can access these latest OpenAI models through the same Amaz

model-releasesaws-ml-blog
1 Jun 2026
Model Releases

Pairwise Reference Alignment as a Model-Level Ordinal Observable

DGX agent

arXiv:2605.30758v1 Announce Type: new Abstract: Pairwise preference data is widely used in language-model evaluation and alignment, often for model ranking, reward modeling, or preference optimization

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

Palo Alto Networks says Mythos found 24+ critical bugs using $1M+ in tokens; Anthropic subsidizes Mythos but some companies plan to boost their Mythos budgets (Aaron Holmes/The Information)

DGX agent

Aaron Holmes / The Information: Palo Alto Networks says Mythos found 24+ critical bugs using $1M+ in tokens; Anthropic subsidizes Mythos but some companies plan to boost their Mythos budgets — When Pa

model-releasestechmeme
1 Jun 2026
Model Releases

Parameter-free Dynamic Regret: Time-varying Movement Costs, Delayed Feedback, and Memory

DGX agent

arXiv:2602.06902v2 Announce Type: replace Abstract: In this paper, we study dynamic regret in unconstrained online convex optimization (OCO) with movement costs. Specifically, we generalize the standa

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

PhyDrawGen: Physically Grounded Diagram Generation from Natural Language

DGX agent

arXiv:2605.30512v1 Announce Type: new Abstract: Generating physics diagrams from text requires strict adherence to physical laws. While current generative models produce visually plausible outputs, th

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Physically Viable World Models: A Case for Query-Conditioned Embodied AI

DGX agent

arXiv:2605.30542v1 Announce Type: new Abstract: World models for embodied AI must be physically viable: constructed to answer intervention queries by representing the physical structure governing acti

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Physics Enhanced Deep Surrogates for the Phonon Boltzmann Transport Equation

DGX agent

arXiv:2512.05976v3 Announce Type: replace-cross Abstract: Designing materials with controlled heat flow at the nano-scale is central to advances in microelectronics, thermoelectrics, and energy-conver

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

PInVerify: An Offline Embodied Benchmark for Active Instance Verification

DGX agent

arXiv:2605.30639v1 Announce Type: cross Abstract: Embodied agents have made strong progress in navigating to target objects, but reaching the goal vicinity does not guarantee that the agent has found

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Plain Transformers are Surprisingly Powerful Link Predictors

DGX agent

arXiv:2602.01553v2 Announce Type: replace-cross Abstract: Link prediction is a core challenge in graph machine learning, demanding models that capture rich and complex topological dependencies. While

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

PRISM: Progressive Reasoning through Iterative Slot Memory for Vision

DGX agent

arXiv:2605.30942v1 Announce Type: new Abstract: Modern vision models process images in a single feed-forward pass, which limits their ability to recover missing evidence or refine uncertain representa

model-releasesarxiv-cs-cv
1 Jun 2026
Model Releases

Probabilistic Precipitation Nowcasting with Rectified Flow Transformers

DGX agent

arXiv:2605.31204v1 Announce Type: new Abstract: Accurate weather forecasts are essential across various domains and are safety-critical in extreme weather conditions. Compared to simulation-based fore

model-releasesarxiv-cs-cv
1 Jun 2026
Model Releases

Probing Collision Grounding in Vision-Language Models for Safe Human-Robot Collaboration

DGX agent

arXiv:2605.31196v1 Announce Type: cross Abstract: Safe human--robot collaboration requires more than visual description: a monitor must determine whether the robot body is safely separated, already co

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Probing the Prompt KV Cache: Where It Becomes Dispensable

DGX agent

arXiv:2605.30574v1 Announce Type: new Abstract: Prior KV cache compression schemes empirically demonstrate that the prompt cache is partially redundant during decoding, dropping or summarising entries

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

QASM-Eval: A Dataset to Train and Evaluate LLMs on OpenQASM-3 Beyond Quantum Circuits

DGX agent

arXiv:2605.30358v1 Announce Type: new Abstract: Quantum computing remains in the Noisy Intermediate-Scale Quantum (NISQ) era, where the performance is highly constrained to noise. Addressing the limit

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

Quantifying the Uncertainty of Foundation Models with Singular Value Ensembles

DGX agent

arXiv:2601.22068v2 Announce Type: replace Abstract: Foundation models have become a dominant paradigm in machine learning, achieving remarkable performance across diverse tasks through large-scale pre

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

Query-focused and Memory-aware Reranker for Long Context Processing

DGX agent

arXiv:2602.12192v3 Announce Type: replace Abstract: Built upon the existing analysis of retrieval heads in large language models, we propose an alternative reranking framework that trains models to es

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

QVGGT: Post-Training Quantized Visual Geometry Grounded Transformer

DGX agent

arXiv:2605.31124v1 Announce Type: new Abstract: Estimating 3D attributes directly from images has advanced rapidly with the Visual Geometry Grounded Transformer (VGGT), which predicts camera parameter

model-releasesarxiv-cs-cv
1 Jun 2026
Model Releases

Qwen 3.7 Plus now available on AI Gateway

DGX agent

Vercel has made Qwen 3.7 Plus, an AI model, available through its AI Gateway service, allowing developers to access this model via Vercel's platform. This addition expands the model options available

model-releasesvercel-blog
1 Jun 2026
Model Releases

Randomized Feasibility Methods for Constrained Optimization with Adaptive Step Sizes

DGX agent

arXiv:2601.20076v2 Announce Type: replace-cross Abstract: We consider minimizing an objective function subject to constraints defined by the intersection of lower-level sets of convex functions. We st

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

Re-examining Low Rank adaptation for private LLM fine-tuning

DGX agent

arXiv:2510.01137v3 Announce Type: replace Abstract: Privacy is a central concern when fine-tuning large language models (LLMs) on sensitive data, and differentially private stochastic gradient descent

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

Read more about all the fun ways we used AI to bring I/O to life this year: https://blog.google/innovation-and-ai/technology/ai/io-2026-goog…

DGX agent

Google's I/O 2026 event showcased various AI applications and innovations developed by Google to enhance the conference experience. The post highlights creative implementations of AI technology integr

model-releasesgoogle-ai--x
1 Jun 2026
← Previous
1…241242243244245…476
Next →