AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
All
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,106 results
Model Releases

Looking for a faster specialized model for your Agent Work? @nvidia Nemotron 3.5 Lightning (30B MoE, 3B active params) is now live on Firewo…

DGX agent

Looking for a faster specialized model for your Agent Work? @nvidia Nemotron 3.5 Lightning (30B MoE, 3B active params) is now live on Fireworks. It’s distilled from NVIDIA Nemotron 3 Ultra to be your

model-releasesfireworks-ai--x
11 Aug 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

LoRSA: Toward Generalizable Parameter-Efficient Fine-Tuning for Biomedical Downstream Tasks

DGX agent

arXiv:2608.07749v1 Announce Type: cross Abstract: Parameter-efficient fine-tuning enables the adaptation of vision foundation models to biomedical tasks under limited computational resources, but a si

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Luth-2: New State-of-the-Art French Small Language Models

DGX agent

Hey everyone, Today we release Luth-2-0.8B and Luth2-2-2B, two non-reasoning models that set a new state of the art for French across a wide variety of tasks for their size 🚀 A few notable scores on F

model-releasesr-localllama
11 Aug 2026
Model Releases

MADBench: A Benchmark for Modality-Aware Audio Deepfake Detection

DGX agent

arXiv:2608.09593v1 Announce Type: cross Abstract: Recent advances in speech synthesis and audio generation have made high-fidelity acoustic forgery low-cost and difficult to attribute, enabling a real

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Made by Google 2026: all the Pixel news and announcements

DGX agent

Google has revealed a bunch of new Pixel devices ahead of its Made by Google event. The colorful Pixel 11 lineup comes with upgraded cameras and performance, with the Pro models offering a built-in LE

model-releasesthe-verge-ai
11 Aug 2026
Model Releases

Marrying Optimal Transport and ODEs for Unified Continuous-Time 4D Reconstruction and Tracking

DGX agent

arXiv:2608.09613v1 Announce Type: new Abstract: Existing unified 4D reconstruction and point tracking approaches typically rely on heuristic interpolations or just predict at integer timestamps, lacki

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

MasDrift: Benchmarking Authorization Preservation Across Multi-Agent Architectures

DGX agent

arXiv:2608.07556v1 Announce Type: cross Abstract: Multi-agent systems (MAS) decompose long-horizon tasks across supervisors and subagents, but delegated goals do not necessarily carry their original a

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Matching Accuracy, Different Geometry: Evolution Strategies vs GRPO in LLM Post-Training

DGX agent

arXiv:2604.01499v2 Announce Type: replace Abstract: Evolution Strategies (ES) have emerged as a scalable gradient-free alternative to reinforcement learning based LLM fine-tuning, but it remains uncle

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

MateInfoUB: A Real-World Benchmark for Testing LLMs in Competitive, Multilingual, and Multimodal Educational Tasks

DGX agent

arXiv:2507.03162v2 Announce Type: replace-cross Abstract: The rapid advancement of Large Language Models (LLMs) has transformed various domains, particularly computer science (CS) education. These mod

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Math-Vision Diagrams: A Comprehensive Benchmark for Evaluating LLM Mathematical Diagram Generation Capabilities

DGX agent

arXiv:2608.08964v1 Announce Type: new Abstract: The generation of mathematically precise diagrams from tex- tual prompts has emerged as a critical yet underexplored capability of Large Language Models

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

MathShikkha: A Controlled Study of Answer-Only and Chain-of-Thought Supervision for Bangla Mathematical Reasoning in Small Language Models

DGX agent

arXiv:2608.08503v1 Announce Type: new Abstract: Mathematical reasoning remains challenging in low-resource languages such as Bangla. We study whether teacher-generated Bangla Chain-of-Thought (CoT) su

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Matrix-free Neural Preconditioner for the Dirac Operator in Lattice Gauge Theory

DGX agent

arXiv:2509.10378v2 Announce Type: replace-cross Abstract: Linear systems arise in generating samples and in calculating observables in lattice quantum chromodynamics~(QCD). Solving the Hermitian posit

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

Matryoshka Language Model Suites

DGX agent

arXiv:2608.09703v1 Announce Type: new Abstract: Training a language model suite classically requires training each model separately and serving them independently. We improve both training and inferen

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Mawqif-v2: An Arabic Benchmark Dataset for Cross-Target Stance Detection

DGX agent

arXiv:2608.09539v1 Announce Type: new Abstract: Publicly available Arabic datasets for target-specific stance detection remain limited, particularly for evaluating cross-target generalization. This pa

model-releasesarxiv-cs-cl
11 Aug 2026
Model Releases

MCIF: Multimodal Crosslingual Instruction-Following Benchmark from Scientific Talks

DGX agent

arXiv:2507.19634v4 Announce Type: replace-cross Abstract: Recent advances in large language models have laid the foundation for multimodal LLMs (MLLMs), which unify text, speech, and vision within a s

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Measuring the Tokenization Premium: A Cost Audit for Underserved Language Communities

DGX agent

arXiv:2608.09046v1 Announce Type: new Abstract: Large language models are increasingly deployed as general-purpose educational and technical assistance systems, but their underlying infrastructure doe

model-releasesarxiv-cs-cl
11 Aug 2026
Model Releases

Measuring the Wrong Thing: Internal Harmfulness Scores Anti-Rank Successful Jailbreaks

DGX agent

arXiv:2608.09624v1 Announce Type: cross Abstract: Internal safety scores judge a prompt before any text is generated, and they are validated by how well they separate harmful prompts from benign ones.

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Mechanistic Interpretability-Guided Selective Fine-Tuning of Vision-Language Models for Centimeter-Level Flood Depth Estimation

DGX agent

arXiv:2608.07562v1 Announce Type: new Abstract: Urban flooding poses an escalating threat to transportation infrastructure, yet no operational system provides real-time, street-level flood-depth estim

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

MedPixel: A Unified Pixel-Language Model for Medical Reasoning and Segmentation

DGX agent

arXiv:2608.09818v1 Announce Type: cross Abstract: Reliable medical image understanding requires models to connect clinical language and visual reasoning with pixel-level grounding. Yet medical vision-

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

MELLON - Multimodal Enhanced LLM for Online Navigation

DGX agent

arXiv:2608.09121v1 Announce Type: new Abstract: Web navigation agents are capable of addressing various types of tasks on different websites. Current baselines on web navigation are either unimodal or

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

MemeMind: Reference-Guided Trace Construction for Offline Context Optimization

DGX agent

arXiv:2608.09316v1 Announce Type: new Abstract: Offline context optimization improves an agent by revising its instructions and examples while keeping the model frozen. This approach learns from rollo

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

Memorization Dynamics in Knowledge Distillation for Language Models

DGX agent

arXiv:2601.15394v2 Announce Type: replace Abstract: Knowledge Distillation (KD) is increasingly adopted to transfer capabilities from large language models to smaller ones, offering significant improv

model-releasesarxiv-cs-cl
11 Aug 2026
Model Releases

Memory-Efficient Activation Checkpointing with Sliding Window and Hirschberg's Algorithm for 0/1 Knapsack Solving in PyTorch

DGX agent

arXiv:2608.08740v1 Announce Type: new Abstract: Activation checkpointing minimizes the runtime of neural networks under a given memory budget, by selecting which intermediate tensors to store and whic

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

MetaSpace: Metamorphic Testing for Spatial Cognition in Embodied Agents

DGX agent

arXiv:2608.07533v1 Announce Type: new Abstract: An embodied agent is an intelligent entity that interacts with its environment through a physical body. Currently, the evaluation of embodied agents pri

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Mind the Hook: Source-Level Auditing of Privacy Defenses in Retrieval-Augmented Generation

DGX agent

arXiv:2608.09001v1 Announce Type: cross Abstract: Black-box privacy scores for retrieval-augmented generation (RAG) are difficult to interpret unless the audited defense's active pipeline hook is know

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

MiraMind: Benchmarking Reliable Mental Health Reasoning beyond Answer Accuracy

DGX agent

arXiv:2512.09636v3 Announce Type: replace Abstract: Mental-health reasoning with large language models (LLMs) is an evidence-constrained judgment problem: models must transform limited, subjective, an

model-releasesarxiv-cs-cl
11 Aug 2026
Model Releases

☁️Mistral is bringing together the inference infrastructure, open models, and long-term commitments Europe needs to control its AI future, a…

DGX agent

☁️Mistral is bringing together the inference infrastructure, open models, and long-term commitments Europe needs to control its AI future, and setting a roadmap for the world. 🧵: https://mistral.ai/ne

model-releasesmistral-ai--x
11 Aug 2026
Model Releases

MMArch: Benchmarking Multimodal Reasoning Grounded in Architectural Evidence

DGX agent

arXiv:2608.09281v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) perform strongly on engineering imagery, yet existing benchmarks mostly test drawing recognition, information e

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Model Discovery Agent: LLM-assisted Bayesian experiment design for data-efficient discovery of mechanistic world models

DGX agent

arXiv:2608.09696v1 Announce Type: new Abstract: Predicting the answer to interventional ``what if'' questions --- the outcome of an action never taken --- requires a mechanistic, causal model, not a c

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

MonitorBench: A Comprehensive Benchmark for Chain-of-Thought Monitorability in Large Language Models

DGX agent

arXiv:2603.28590v3 Announce Type: replace Abstract: Large language models (LLMs) can generate chains of thought (CoTs) that are not always causally responsible for their final outputs. When such a mis

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

MoRSE: Task-Oriented Multi-Agent System with Mixture of Role-Subtask Experts

DGX agent

arXiv:2608.09251v1 Announce Type: cross Abstract: Large language model-based multi-agent systems have recently shown strong potential for complex, long-horizon tasks. However, existing methods mainly

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

MOSAIC: Adversarial Co-evolution of Specialist Heuristics and Problem Instances for LLM-based Automated Heuristic Design

DGX agent

arXiv:2608.07544v1 Announce Type: cross Abstract: Automated heuristic design (AHD) with large language models (LLMs) has produced strong heuristics for combinatorial optimization problems (COPs). Yet

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

MPISuperRes-PnP: A Super-Resolution Zero-Shot Plug-and-Play Reconstruction Algorithm for Magnetic Particle Imaging

DGX agent

arXiv:2608.09672v1 Announce Type: new Abstract: Magnetic Particle Imaging (MPI) is an emerging medical imaging modality. MPI is based on the non-linear response of magnetic nanoparticles to an applied

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

MRBench: A Comprehensive Benchmark for Human Motion-Text Retrieval

DGX agent

arXiv:2608.07993v1 Announce Type: new Abstract: Human motion-text retrieval provides a rigorous means of assessing cross-modal alignment. Prevailing benchmarks are dominated by homogeneous indoor moti

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

MSP-Net: Manifold-Guided Spectral Prompt Network for Hyperspectral Object Tracking

DGX agent

arXiv:2608.09575v1 Announce Type: new Abstract: Hyperspectral object tracking leverages abundant spectral information to provide unique advantages for target discrimination in complex scenes. However,

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

Multi-Agent Reinforcement Learning via Agent-Specific Preference

DGX agent

arXiv:2608.08604v1 Announce Type: new Abstract: Multi-agent reinforcement learning (MARL) is a powerful framework for solving complex collaborative tasks, but it relies heavily on well-defined global

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

Multilingual Agent-Based World Modeling for Social Science

DGX agent

arXiv:2512.07195v2 Announce Type: replace-cross Abstract: Multi-agent role-playing has recently shown promise for studying social behavior with language agents, but existing simulations are mostly mon

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Muse-Glimmer 30B Hits ~280 t/s in Real Production Coding

DGX agent

These numbers were captured during a real feature implementation task in Next.js and Nest.js (adding a theme switching system across components). The structural predictability of UI/state refactoring

model-releasesr-localllama
11 Aug 2026
Model Releases

My conversation with @ericvishria of Benchmark. Eric has spent a decade investing across software and hardware, backing companies like Firew…

DGX agent

My conversation with @ericvishria of Benchmark. Eric has spent a decade investing across software and hardware, backing companies like Fireworks, Sierra, Sunday Robotics, and Cerebras. This one is abo

model-releasessonya-huang--x
11 Aug 2026
Model Releases

NBA_Streaming: A Large-Scale Benchmark for Fine-Grained Basketball Commentary Generation in Continuous Streams

DGX agent

arXiv:2608.09200v1 Announce Type: new Abstract: Live basketball commentary generation requires determining when an event is sufficiently observable and describing it before subsequent events unfold. H

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

Nemotron 3.5 Lightning is available in LM Studio! The model is 30B MoE (3B active), can run very fast, and is trained for high volume agenti…

DGX agent

Nemotron 3.5 Lightning is available in LM Studio! The model is 30B MoE (3B active), can run very fast, and is trained for high volume agentic use cases. Model page: https://lmstudio.ai/models/nvidia/n

model-releaseslm-studio--x
11 Aug 2026
Model Releases

Nemotron 3.5 Lightning is available on Together AI through Dedicated Model Inference, giving teams reserved capacity and predictable perform…

DGX agent

Nemotron 3.5 Lightning is available on Together AI through Dedicated Model Inference, giving teams reserved capacity and predictable performance for high-volume agent workloads. Start building: https:

model-releasestogether-ai--x
11 Aug 2026
Model Releases

Neural Operators for Immersed-Boundary Soft Swimmers Locomotion

DGX agent

arXiv:2608.07722v1 Announce Type: new Abstract: High-fidelity immersed-boundary simulation resolves the coupled motion of a deforming swimmer and its surrounding flow, but the resulting cost limits re

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

NeuroAda: Activating Each Neuron's Potential for Parameter-Efficient Fine-Tuning

DGX agent

arXiv:2510.18940v2 Announce Type: replace-cross Abstract: Existing parameter-efficient fine-tuning (PEFT) methods primarily fall into two categories: addition-based and selective in-situ adaptation. T

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

NeuroGuard: Neural Gradient Update Aware of Representation Damage

DGX agent

arXiv:2608.08068v1 Announce Type: new Abstract: Long-tailed class-incremental learning (LT-CIL) must learn new classes from imbalanced streams while retaining old classes. Existing methods mainly chan

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

Neurosymbolic Discovery of Algebraic Graph Constructions

DGX agent

arXiv:2608.08118v1 Announce Type: new Abstract: There are several methods for searching for graphs with prescribed properties, such as SAT solvers and specialized generators. These methods return the

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

NL2SHACL-Bench: A Benchmark Suite for Natural Language to SHACL Translation

DGX agent

arXiv:2608.07530v1 Announce Type: new Abstract: SHACL is a core technology for validating the conformance of RDF knowledge graphs (KGs). Yet, authoring SHACL shapes requires technical expertise that m

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

No Unique Minimizer, No Problem: On the Consistency of Robust Neural Classifiers

DGX agent

arXiv:2608.08489v1 Announce Type: new Abstract: Neural network classifiers trained by cross-entropy minimization are highly sensitive to label noise and adversarial contamination. While robust alterna

model-releasesarxiv-cs-lg
11 Aug 2026
← Previous
1…1011121314…461
Next →