AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,292 results
Model Releases

MOONSHOT : A Framework for Multi-Objective Pruning of Vision and Large Language Models

DGX agent

arXiv:2604.13287v1 Announce Type: new Abstract: Weight pruning is a common technique for compressing large neural networks. We focus on the challenging post-training one-shot setting, where a pre-trai

model-releasesarxiv-cs-lg
16 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

More on my blog, including results from the previously secret 'flamingo on a unicycle' test https://simonwillison.net/2026/Apr/16/qwen-beats…

DGX agent

Simon Willison discusses results from a 'flamingo on a unicycle' test on his blog, likely comparing AI model performance including Qwen. The post appears to reference previously undisclosed or unconve

model-releasessimon-willison--x
16 Apr 2026
Model Releases

Mosaic: An Extensible Framework for Composing Rule-Based and Learned Motion Planners

DGX agent

arXiv:2604.13853v1 Announce Type: new Abstract: Safe and explainable motion planning remains a central challenge in autonomous driving. While rule-based planners offer predictable and explainable beha

model-releasesarxiv-cs-ro
16 Apr 2026
Model Releases

Mozilla launches Thunderbolt AI client with focus on self-hosted infrastructure

DGX agent

Thunderbolt is a new open-source AI client from Mozilla-owned MZLA Technologies aimed at enterprises who want to run self-hosted chatbots on their own infrastructure. The platform allows organizations

model-releasesars-technica
16 Apr 2026
Model Releases

Mozilla launches Thunderbolt, an open-source AI client for users and businesses who want to run their own self-hosted AI infrastructure, available on GitHub (Kyle Orland/Ars Technica)

DGX agent

Kyle Orland / Ars Technica: Mozilla launches Thunderbolt, an open-source AI client for users and businesses who want to run their own self-hosted AI infrastructure, available on GitHub — Mozilla is th

model-releasestechmeme
16 Apr 2026
Model Releases

MulDimIF: A Multi-Dimensional Constraint Framework for Evaluating and Improving Instruction Following in Large Language Models

DGX agent

arXiv:2505.07591v2 Announce Type: replace Abstract: Instruction following refers to the ability of large language models (LLMs) to generate outputs that satisfy all specified constraints. Existing res

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Multi-Task LLM with LoRA Fine-Tuning for Automated Cancer Staging and Biomarker Extraction

DGX agent

arXiv:2604.13328v1 Announce Type: new Abstract: Pathology reports serve as the definitive record for breast cancer staging, yet their unstructured format impedes large-scale data curation. While Large

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

New insane model from Jackrong on @huggingface 🤯 Qwen3.5-9B-GLM5.1-Distill-v1 🧠 Distilled on GLM-5.1 reasoning ⚙️ Deeper thinking than bas…

DGX agent

New insane model from Jackrong on @huggingface 🤯 Qwen3.5-9B-GLM5.1-Distill-v1 🧠 Distilled on GLM-5.1 reasoning ⚙️ Deeper thinking than base model 🧪 Benchmarks coming soon ✅ Fits on 8GB VRAM ✍️ New mod

model-releasesclem-delangue--x
16 Apr 2026
Model Releases

New ways to create personalized images in the Gemini app

DGX agent

Gemini has been updated to generate highly personalized images by accessing a user's Google Photos and personal preferences. This new feature allows users to simply request scenes featuring themselves

model-releasesgoogle-ai
16 Apr 2026
Model Releases

No Need for Space Gear — Capcom’s ‘PRAGMATA’ Joins GeForce NOW on Launch Day

DGX agent

Head straight for orbit with GeForce NOW — no space helmet required. PRAGMATA, Capcom’s long-awaited sci-fi action adventure, touches down on GeForce NOW the same day it launches worldwide. The futuri

model-releasesnvidia-blog
16 Apr 2026
Model Releases

Olfactory pursuit: catching a moving odor source in complex flows

DGX agent

arXiv:2604.13121v1 Announce Type: new Abstract: Locating and intercepting a moving target from possibly delayed, intermittent sensory signals is a paradigmatic problem in decision-making under uncerta

model-releasesarxiv-cs-ro
16 Apr 2026
Model Releases

🎬 Ollama Gemma Day Recap: SGLang at the Ollama Gemma 4 Party in Palo Alto 🍾 Last night, @ollama hosted a packed Gemma Day at the Palo Alto…

DGX agent

🎬 Ollama Gemma Day Recap: SGLang at the Ollama Gemma 4 Party in Palo Alto 🍾 Last night, @ollama hosted a packed Gemma Day at the Palo Alto office alongside the @GoogleDeepMind Gemma team. SGLang was i

model-releasesollama--x
16 Apr 2026
Model Releases

Online learning with noisy side observations

DGX agent

arXiv:2604.13740v1 Announce Type: new Abstract: We propose a new partial-observability model for online learning problems where the learner, besides its own loss, also observes some noisy feedback abo

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

OpenAI launches GPT-Rosalind, an AI model for life sciences research, including drug discovery, as a research preview for customers such as Moderna and Amgen (Megan Morrone/Axios)

DGX agent

Megan Morrone / Axios: OpenAI launches GPT-Rosalind, an AI model for life sciences research, including drug discovery, as a research preview for customers such as Moderna and Amgen — OpenAI announced

model-releasestechmeme
16 Apr 2026
Model Releases

OpenAI’s big Codex update is a direct shot at Claude Code

DGX agent

OpenAI is beefing up its agentic coding and development system, Codex, with a suite of updates that let it use your computer, generate images, and remember from past experiences. The package of update

model-releasesthe-verge-ai
16 Apr 2026
Model Releases

OPTED: Open Preprocessed Trachoma Eye Dataset Using Zero-Shot SAM 3 Segmentation

DGX agent

arXiv:2603.06885v2 Announce Type: replace Abstract: Trachoma remains the leading infectious cause of blindness worldwide, with Sub-Saharan Africa bearing over 85% of the global burden and Ethiopia alo

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

Optimization with SpotOptim

DGX agent

arXiv:2604.13672v1 Announce Type: new Abstract: The `spotoptim` package implements surrogate-model-based optimization of expensive black-box functions in Python. Building on two decades of Sequential

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Opus 4.7 feels more intelligent, agentic, and precise than 4.6. It took a few days for me to learn how to work with it effectively, to fully…

DGX agent

Opus 4.7 feels more intelligent, agentic, and precise than 4.6. It took a few days for me to learn how to work with it effectively, to fully take advantage of its new capabilities. Will post a few mor

model-releasesboris-cherny--x
16 Apr 2026
Model Releases

Opus 4.7 is a model I’ve loved working with in Claude Code. It’s more agentic and instruction following but also incredibly smart and creati…

DGX agent

Opus 4.7 is a model I’ve loved working with in Claude Code. It’s more agentic and instruction following but also incredibly smart and creative. I think it takes a slight adjustment to get used to, but

model-releasesthariq--x
16 Apr 2026
Model Releases

Opus 4.7 is in Claude Code today. It's more agentic, more precise, and a lot better at long-running work. It carries context across sessions…

DGX agent

Opus 4.7 is in Claude Code today. It's more agentic, more precise, and a lot better at long-running work. It carries context across sessions and handles ambiguity much better. Introducing Claude Opus

model-releasesboris-cherny--x
16 Apr 2026
Model Releases

Opus 4.7 is now supported in Hermes Agent 🚀🚀

DGX agent

Opus 4.7 is now supported in Hermes Agent 🚀🚀 Introducing Claude Opus 4.7, our most capable Opus model yet. It handles long-running tasks with more rigor, follows instructions more precisely, and verif

model-releasesnous-research--x
16 Apr 2026
Model Releases

Ordinary Least Squares is a Special Case of Transformer

DGX agent

arXiv:2604.13656v1 Announce Type: new Abstract: The statistical essence of the Transformer architecture has long remained elusive: Is it a universal approximator, or a neural network version of known

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Our Life Sciences model series is available as a research preview starting today for qualified customers including @Amgen, @moderna_tx, the …

DGX agent

Our Life Sciences model series is available as a research preview starting today for qualified customers including @Amgen, @moderna_tx, the @AllenInstitute, and @thermofisher Scientific through ChatGP

model-releasesopenai--x
16 Apr 2026
Model Releases

Out of Context: Reliability in Multimodal Anomaly Detection Requires Contextual Inference

DGX agent

arXiv:2604.13252v1 Announce Type: new Abstract: Anomaly detection aims to identify observations that deviate from expected behavior. Because anomalous events are inherently sparse, most frameworks are

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Parameter-efficient Quantum Multi-task Learning

DGX agent

arXiv:2604.13560v1 Announce Type: new Abstract: Multi-task learning (MTL) improves generalization and data efficiency by jointly learning related tasks through shared representations. In the widely us

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Parameter-Free Non-Ergodic Extragradient Algorithms for Solving Monotone Variational Inequalities

DGX agent

arXiv:2604.07662v2 Announce Type: replace-cross Abstract: Monotone variational inequalities (VIs) provide a unifying framework for convex minimization, equilibrium computation, and convex-concave sadd

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Parameter Importance is Not Static: Evolving Parameter Isolation for Supervised Fine-Tuning

DGX agent

arXiv:2604.14010v1 Announce Type: cross Abstract: Supervised Fine-Tuning (SFT) of large language models often suffers from task interference and catastrophic forgetting. Recent approaches alleviate th

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

PatchPoison: Poisoning Multi-View Datasets to Degrade 3D Reconstruction

DGX agent

arXiv:2604.13153v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) has recently enabled highly photorealistic 3D reconstruction from casually captured multi-view images. However, this access

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

PBE-UNet: A light weight Progressive Boundary-Enhanced U-Net with Scale-Aware Aggregation for Ultrasound Image Segmentation

DGX agent

arXiv:2604.13791v1 Announce Type: new Abstract: Accurate lesion segmentation in ultrasound images is essential for preventive screening and clinical diagnosis, yet remains challenging due to low contr

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

Peer-Predictive Self-Training for Language Model Reasoning

DGX agent

arXiv:2604.13356v1 Announce Type: new Abstract: Mechanisms for continued self-improvement of language models without external supervision remain an open challenge. We propose Peer-Predictive Self-Trai

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

PersonaVLM: Long-Term Personalized Multimodal LLMs

DGX agent

arXiv:2604.13074v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) serve as daily assistants for millions. However, their ability to generate responses aligned with individual pr

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Predicting Time Pressure of Powered Two-Wheeler Riders for Proactive Safety Interventions

DGX agent

arXiv:2601.03173v2 Announce Type: replace Abstract: Time pressure critically influences risky maneuvers and crash proneness among powered two-wheeler riders, yet its prediction remains underexplored i

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Qwen 3.6 is here, and open-source! Run it locally with improved agentic coding capabilities. Try it with Claude Code: ollama launch claude -…

DGX agent

Qwen 3.6 is here, and open-source! Run it locally with improved agentic coding capabilities. Try it with Claude Code: ollama launch claude --model qwen3.6 Try it with OpenClaw: ollama launch openclaw

model-releasesollama--x
16 Apr 2026
Model Releases

Qwen3.6-35B-A3B on my laptop drew me a better pelican than Claude Opus 4.7

DGX agent

For anyone who has been (inadvisably) taking my pelican riding a bicycle benchmark seriously as a robust way to test models, here are pelicans from this morning's two big model releases - Qwen3.6-35B-

model-releasessimon-willison
16 Apr 2026
Model Releases

RAG or Learning? Understanding the Limits of LLM Adaptation under Continuous Knowledge Drift in the Real World

DGX agent

arXiv:2604.05096v2 Announce Type: replace Abstract: Large language models (LLMs) acquire most of their knowledge during pretraining, which ties them to a fixed snapshot of the world and makes adaptati

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

RANDPOL: Parameter-Efficient End-to-End Quadruped Locomotion via Randomized Policy Learning

DGX agent

arXiv:2505.19054v2 Announce Type: replace Abstract: Modern learning-based locomotion controllers typically rely on fully trainable deep neural networks with a large number of parameters. This paper st

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Rare Event Analysis via Stochastic Optimal Control

DGX agent

arXiv:2604.13213v1 Announce Type: cross Abstract: Rare events such as conformational changes in biomolecules, phase transitions, and chemical reactions are central to the behavior of many physical sys

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

ReConText3D: Replay-based Continual Text-to-3D Generation

DGX agent

arXiv:2604.13730v1 Announce Type: new Abstract: Continual learning enables models to acquire new knowledge over time while retaining previously learned capabilities. However, its application to text-t

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

Red Skills or Blue Skills? A Dive Into Skills Published on ClawHub

DGX agent

arXiv:2604.13064v1 Announce Type: new Abstract: Skill ecosystems have emerged as an increasingly important layer in Large Language Model (LLM) agent systems, enabling reusable task packaging, public d

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Replit Agent 4 is even smarter now with Claude Opus 4.7! 50% off for a limited time. Go try it now ↓

DGX agent

Replit Agent 4 has been upgraded to use Claude Opus 4.7, an advanced AI model, enhancing its code generation and problem-solving capabilities. The company is offering a 50% discount for a limited time

model-releasesreplit--x
16 Apr 2026
Model Releases

Response.

DGX agent

Response. Hey Ethan! Sean here, PM on http://Claude.ai - thanks for the feedback. This isn't a router, this is the model being trained to decide when to think based on the context -- we've been runnin

model-releasesethan-mollick--x
16 Apr 2026
Model Releases

Reward Design for Physical Reasoning in Vision-Language Models

DGX agent

arXiv:2604.13993v1 Announce Type: cross Abstract: Physical reasoning over visual inputs demands tight integration of visual perception, domain knowledge, and multi-step symbolic inference. Yet even st

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Reward Hacking in the Era of Large Models: Mechanisms, Emergent Misalignment, Challenges

DGX agent

arXiv:2604.13602v1 Announce Type: new Abstract: Reinforcement Learning from Human Feedback (RLHF) and related alignment paradigms have become central to steering large language models (LLMs) and multi

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

RiskWebWorld: A Realistic Interactive Benchmark for GUI Agents in E-commerce Risk Management

DGX agent

arXiv:2604.13531v1 Announce Type: cross Abstract: Graphical User Interface (GUI) agents show strong capabilities for automating web tasks, but existing interactive benchmarks primarily target benign,

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

ROBOGATE: Adaptive Failure Discovery for Safe Robot Policy Deployment via Two-Stage Boundary-Focused Sampling

DGX agent

arXiv:2603.22126v3 Announce Type: replace Abstract: Deploying learned robot manipulation policies in industrial settings requires rigorous pre-deployment validation, yet exhaustive testing across high

model-releasesarxiv-cs-ro
16 Apr 2026
Model Releases

Robust Low-Rank Tensor Completion based on M-product with Weighted Correlated Total Variation and Sparse Regularization

DGX agent

arXiv:2604.13525v1 Announce Type: cross Abstract: The robust low-rank tensor completion problem addresses the challenge of recovering corrupted high-dimensional tensor data with missing entries, outli

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Robust Reward Modeling for Large Language Models via Causal Decomposition

DGX agent

arXiv:2604.13833v1 Announce Type: new Abstract: Reward models are central to aligning large language models, yet they often overfit to spurious cues such as response length and overly agreeable tone.

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

ROSE: Retrieval-Oriented Segmentation Enhancement

DGX agent

arXiv:2604.14147v1 Announce Type: new Abstract: Existing segmentation models based on multimodal large language models (MLLMs), such as LISA, often struggle with novel or emerging entities due to thei

model-releasesarxiv-cs-cv
16 Apr 2026
← Previous
1…428429430431432…465
Next →