AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,323 results
3 Aug 2026

POSSE-kNN: Pathwise Out-of-Bag Selected Subspace Ensembles for Binary Classification

Model ReleasesDGX agent

arXiv:2211.11278v3 Announce Type: replace-cross Abstract: Nearest neighbour classification is attractive for tabular data, but its performance can deteriorate when a fixed query centred neighbourhood

Predicting Steel Fatigue Life from Micrographs Using Physics-Informed Deep Learning

Model ReleasesDGX agent

arXiv:2607.28695v1 Announce Type: cross Abstract: Here is the plain text version optimized for arXiv's submission form. Custom macros (like CV and SI) have been converted to standard text/math so they

Question about Quant versus Size.

Model ReleasesDGX agent

Sorry if this is asked a lot, but I was wondering if there is any clear winner on the Quantization versus Model Size debate? I can run Qwen3.6 27b at Q8, Laguna at Q6, and the new Deepseek Flash at Q3


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Qwen 3.8 Max and MiniMax-H3 within hours of each other

Model ReleasesDGX agent

Simon Willison noted that Qwen 3.8 Max and MiniMax‑H3 were released within hours of each other. The MiniMax team announced that MiniMax‑H3 is now publicly available on Hugging Face (https://huggingfac

Qwen3.8-Max matches Kimi K3 and DeepSeek V4 Flash

Model ReleasesDGX agent

Qwen3.8-Max (2.4T) is another massive contribution to the open weight community. On benchmarks, it performs closely to Kimi K3 and DeepSeek V4 flash across all categories and is better at coding and s

RareSense: Rarity-Aware Similarity Search for Anomaly Retrieval in Transactional Data

Model ReleasesDGX agent

arXiv:2607.28879v1 Announce Type: cross Abstract: Similarity search over sparse set-valued data is often dominated by frequent background attributes because classical measures such as Jaccard, cosine,

RayViT: Ray-Conditioned Visual Representations for Viewpoint-Robust Imitation Learning

Model ReleasesDGX agent

arXiv:2607.29622v1 Announce Type: cross Abstract: Visual imitation learning enables robots to acquire visuomotor skills directly from images, yet RGB observations lack explicit geometric cues, making

Real-world mainframe modernization with AI: A safe, scalable path from mainframe to cloud

Model ReleasesDGX agent

For too long, enterprises with legacy mainframe estates have been faced with a high-stakes dilemma: continue maintaining their mainframes, essentially kicking the modernization can down the road (they

Receding-Horizon Next-Best-View Planner for Autonomous Leaf Surface Reconstruction

Model ReleasesDGX agent

arXiv:2607.28995v1 Announce Type: new Abstract: Accurate plant leaf modeling is fundamental to downstream tasks such as plant growth monitoring, and phenotyping for yield estimation. Autonomous roboti

Reflected UAS: Corrected Deterministic Stability and Direct CTMC Drift Calculation

Model ReleasesDGX agent

arXiv:2607.28688v1 Announce Type: cross Abstract: We analyze Reflected UAS routing for heterogeneous multi-server queues at fixed parameters under subcritical load. The deterministic surrogate is a re

[RELEASE] SupraBrain-50M-v0.1

Model ReleasesDGX agent

Hey there! So today we're releasing SupraBrain-50M, a hybrid language model that combines Gated DeltaNet linear recurrence with Sliding-Window Attention and Surprise-Gated update mechanisms to deliver

ReLoop-UME: Recurrent Depth with Learnable Retrieval Registers for Universal Multimodal Embedding

Model ReleasesDGX agent

arXiv:2607.28751v1 Announce Type: new Abstract: Universal multimodal embedding (UME) maps heterogeneous multimodal inputs into a shared embedding space. Existing UME models either form embeddings thro

Retrieval-Driven Training-Free AI-Generated Video Attribution

Model ReleasesDGX agent

arXiv:2607.28955v1 Announce Type: cross Abstract: AI-generated videos are becoming increasingly realistic and difficult to distinguish from authentic ones, which facilitates malicious misuse and poses

Revisiting Multi-Permutation Equivariance through the Lens of Irreducible Representations

Model ReleasesDGX agent

arXiv:2410.06665v4 Announce Type: replace-cross Abstract: This paper explores the characterization of equivariant linear layers for representations of permutations and related groups. Unlike tradition

Rolling With Resistance: Preference-Optimized LLM Counselors Can Trade Goal Persistence for Relational Attunement in Motivational Interviewing

Model ReleasesDGX agent

arXiv:2607.28814v1 Announce Type: cross Abstract: In Motivational Interviewing (MI), a client's sustain talk (arguments for the status quo) calls for the counselor to roll with resistance, a move that

RynnBrain 1.1: Towards More Capable and Generalizable Embodied Foundation Model

Model ReleasesDGX agent

arXiv:2607.17977v2 Announce Type: replace Abstract: We present RynnBrain 1.1, a family of embodied foundation models spanning 2B, 9B, and 122B-A10B scales. Trained with a unified spatio-temporal and p

Safe Vision Language Action Models via Barrier Enhanced Flow Matching

Model ReleasesDGX agent

arXiv:2607.29569v1 Announce Type: new Abstract: This article presents a modular inference framework that integrates Flow Matching generative models with formal Control Barrier Function (CBF) safety gu

Safety, or Just Capability? A Validity Audit of Agent-Safety Benchmarks

Model ReleasesDGX agent

arXiv:2607.28685v1 Announce Type: new Abstract: Agent-safety benchmarks measure different behaviors, and their scores get quoted interchangeably as an agent's safety. We treat four of them (R-Judge, I

SAM+D: Parameter-Efficient Dimensional Lifting of SAM-Family Models via Depth-Routed LoRA and Depth Shifting

Model ReleasesDGX agent

arXiv:2607.29033v1 Announce Type: new Abstract: Existing methods for adapting 2D foundation models such as SAM to 3D volumes either process slices independently---ignoring inter-slice context---or req

Scaling Properties of Text Conditioning in Visual Generation

Model ReleasesDGX agent

arXiv:2607.29679v1 Announce Type: new Abstract: We study empirical scaling properties for text conditioning in visual generation. Such properties have rarely been measured because diffusion loss does

SciFigPlag-Bench: A Benchmark for Provenance-Aware Scientific Figure Plagiarism Detection

Model ReleasesDGX agent

arXiv:2607.29124v1 Announce Type: new Abstract: Scientific figures often encode the visual evidence behind scientific findings, yet figure plagiarism remains underexplored as a benchmarked multimodal

SciToolAgent-Evo: An Ontology-Aware Self-Evolving Agent for Open-World Scientific Tool Acquisition

Model ReleasesDGX agent

arXiv:2607.28692v1 Announce Type: new Abstract: Large language model (LLM) agents have been increasingly adopted in scientific research for organizing and invoking specialized computational tools. How

SEDR-Seq2P: A Lightweight Dilated Residual Sequence-to-Point Network for Multi-Task Industrial NILM

Model ReleasesDGX agent

arXiv:2607.28693v1 Announce Type: cross Abstract: Industrial NILM remains challenging because measurement noise and widespread concurrent machine operation reduce the generalization of models tuned on

SeekBrain: An Autonomous Multi-Agent System for Accelerating Neuroscience Discovery

Model ReleasesDGX agent

arXiv:2607.29347v1 Announce Type: cross Abstract: Modern neuroscience relies on integrating multi-scale, multimodal datasets to uncover the neural principles underlying intelligence. However, analytic

Sign compression for Muon: SignMuon, MuonSign, and the Limits of Error Feedback

Model ReleasesDGX agent

arXiv:2607.29674v1 Announce Type: cross Abstract: SignMuon compresses the Muon update to one bit per parameter by taking its elementwise sign, providing the most direct way to run a matrix-aware optim

SILVA Networks as Structured Implicit Layers and Vector Attractors via Dynamic Interaction Fields

Model ReleasesDGX agent

arXiv:2607.28989v1 Announce Type: new Abstract: Many learning problems require representations that reconcile direct input, nearby structure, and broader context. In implicit neural layers, these infl

Simulation Code Generation for Fluid Systems using Large Language Models: Benchmarking Models and Prompting Strategies

Model ReleasesDGX agent

arXiv:2607.29389v1 Announce Type: new Abstract: Large language models (LLMs) have demonstrated a strong ability to generate syntactically correct code from natural-language specifications. In this stu

So-Fake: Benchmarking and Explaining Social Media Image Forgery Detection

Model ReleasesDGX agent

arXiv:2505.18660v5 Announce Type: replace Abstract: Recent advances in AI-powered generative models have enabled the creation of increasingly realistic synthetic images, posing significant risks to in

Speculative decoding with deepseek v4 flash 0731?

Model ReleasesDGX agent

Has anyone figured out how to enable speculative decoding with deepseek v4 flash 0731 on llamacpp? I’m on the right release for llamacpp (b10228 or earlier) and running am17an’s draft model with unslo

.@ssankar says Palantir was able to make Nvidia's Nemotron Ultra model 'better than frontier': 'I literally almost felt gaslit when, within …

Model ReleasesDGX agent

.@ssankar says Palantir was able to make Nvidia's Nemotron Ultra model 'better than frontier': 'I literally almost felt gaslit when, within 24 hours of getting Nemotron up with no post-training, this

Step-Level Visual Grounding Faithfulness Predicts Out-of-Distribution Generalization in Long-Horizon Vision-Language Models

Model ReleasesDGX agent

arXiv:2603.06828v2 Announce Type: replace-cross Abstract: We uncover a behavioral law of long-horizon vision-language models: models that maintain temporally grounded beliefs generalize better. Standa

StraightDP: Geometry-Aware Differential Privacy for Rectified-Flow Transformers

Model ReleasesDGX agent

arXiv:2607.29100v1 Announce Type: cross Abstract: Differentially private (DP) training of text-conditioned generative models suffers a utility cliff at strong privacy. We revisit this problem through

SULAND v2: A Refined RGB Dataset and Deep Learning Object Detection Benchmark for UAV/UGV-Based SUrface LANDmine Detection Under Domain Shift

Model ReleasesDGX agent

arXiv:2607.28996v1 Announce Type: new Abstract: RGB imagery offers a practical, low-cost option for Unmanned Aerial/Ground Vehicle (UAV/UGV) survey support in surface-landmine detection, but object de

TAPR: Enhancing LLM Performance with a Task-Aware Prompt Rewriter

Model ReleasesDGX agent

arXiv:2607.28657v1 Announce Type: new Abstract: Large Language Models (LLMs) often require carefully crafted prompts to unlock their full potential, which can be a barrier for non-expert users. This w

Teaching Video Generators to Remember: Eliciting Dynamic Memory for Out-of-Sight State Evolution

Model ReleasesDGX agent

arXiv:2605.25333v2 Announce Type: replace Abstract: Video world models should maintain evolving states when evidence is unobserved, yet current generators often freeze hidden states upon interruption.

Thanks for the recognition. We'll keep building! 🚀

Model ReleasesDGX agent

Thanks for the recognition. We'll keep building! 🚀 Big news: Qwen3.8-Max by @Alibaba_Qwen just landed at #4 on the Frontend Code Arena leaderboard with a score of 1,668! With 1,668 points, Qwen3.8-Max

The Asymmetric Effects of Knowledge Distillation on Bias in Small Language Models

Model ReleasesDGX agent

arXiv:2607.28639v1 Announce Type: cross Abstract: We show that knowledge distillation in small instruction-tuned language models has asymmetric effects on bias. On unambiguous tasks (BBQ-disambig), re

The Chinese labs everyone lumps together are making four pretty different bets. I work at one of them.

Model ReleasesDGX agent

Every time a model drops from a Chinese lab the thread fills with people who already know who made it, and the guess is usually Alibaba. There was a thread here recently asking what separates the open

The Grokked Illusion: True Equilibrium Mitigates Catastrophic Forgetting

Model ReleasesDGX agent

arXiv:2607.29503v1 Announce Type: new Abstract: While neural networks are typically evaluated by their training and test performance, these metrics do not reveal how robust a learned representation is

The Morphological Core of Dungan: A Two-Dialect Finite-State Model and a Multi-Genre Evaluation

Model ReleasesDGX agent

arXiv:2607.28766v1 Announce Type: new Abstract: Dungan, a Sinitic language of Central Asia written in a Cyrillic-based script, is described in detail in the grammatical literature, yet the quantitativ

The Parts Are Greater Than the Sum: Automated Task Sequencing for Efficient Training of Multi-Policy LLMs

Model ReleasesDGX agent

arXiv:2607.29601v1 Announce Type: new Abstract: Parameter-Efficient Fine-Tuning (PEFT) commonly adapts large language models using a single shared Low-Rank Adapter (LoRA). This shared optimization spa

The result is a faster, more natural conversation with ChatGPT Voice from the moment a session starts. How we built it: https://openai.com/i…

Model ReleasesDGX agent

OpenAI has redesigned the ChatGPT Voice stack—from client to model—to enable continuous audio streaming, allowing GPT‑Live to listen while speaking without interruption. The new architecture supports

The results span sphere packing, coding theory, group theory, quantum complexity, lattice cryptography, extremal combinatorics, and more. Am…

Model ReleasesDGX agent

The results span sphere packing, coding theory, group theory, quantum complexity, lattice cryptography, extremal combinatorics, and more. Among them: establishing the existence of non-sofic groups and

To Add Is Machine, To Delete Is Human: Measuring and Mitigating Deletion Avoidance in LLM Code Editing

Model ReleasesDGX agent

arXiv:2607.28887v1 Announce Type: cross Abstract: Large language models increasingly write and repair production code, yet evidence is mounting that their test-passing patches leave codebases harder t

Tokenizer-Agnostic Engram Module

Model ReleasesDGX agent

arXiv:2607.29065v1 Announce Type: new Abstract: Deepseek's Engram, a conditional memory module, was introduced to trade-off storage versus reasoning in large language models. However, the module relie

Towards bridging the gap: Systematic sim-to-real transfer for diverse legged robots

Model ReleasesDGX agent

arXiv:2509.06342v2 Announce Type: replace Abstract: Legged robots must achieve both robust locomotion and energy efficiency to be practical in real-world environments. Yet controllers trained in simul

Translation with Thought: Difficulty-Adaptive Reasoning via Reinforcement Learning for Multi-Domain Machine Translation

Model ReleasesDGX agent

arXiv:2607.29287v1 Announce Type: cross Abstract: Multi-domain machine translation (MDMT) poses a unique challenge due to varying levels of linguistic complexity across domains. Inspired by human tran

Tri-Space Operational Control of Redundant Multilink and Hybrid Cable-Driven Parallel Robots Using an Iterative-Learning based Reactive Approach

Model ReleasesDGX agent

arXiv:2607.29500v1 Announce Type: new Abstract: Cable-Driven Parallel Robots (CDPRs) are a type of parallel mechanism in which cables are used as actuators. Due to the two levels of redundancy and num

two weeks ago i went on @swyx's pod and said some things that i... should not have said. a lot has happened since then, i owe you all an apo…

Model ReleasesDGX agent

two weeks ago i went on @swyx's pod and said some things that i... should not have said. a lot has happened since then, i owe you all an apology. i'm sorry that i was right about every single thing. a

V4-Flash-0731 - vibes after first weekend of use

Model ReleasesDGX agent

Spent way too much time with V4-Flash-0731 this weekend and wanted to share my vibes as briefly as possible. I sent it through a bit of real-work and some of my personal benchmarks. My quick thoughts

Validation Evidence in LLM Repair Agents: How Much of What Passes Actually Tests the Bug?

Model ReleasesDGX agent

arXiv:2607.28871v1 Announce Type: cross Abstract: When a repair agent runs a test and sees it pass, the result is treated as evidence about the reported defect. We measure how often that treatment is

Was the release of deepseek v4 flash planned to take spotlight against 5.6 luna?

Model ReleasesDGX agent

Id figured since they first emailed people about api price changes coming mid july then delayed the v4 flash release to late july, I wonder if they delayed it for the sake of stealing spotlight from o

WebCoderBench: Benchmarking Web Application Generation with Comprehensive and Interpretable Evaluation Metrics

Model ReleasesDGX agent

arXiv:2601.02430v3 Announce Type: replace-cross Abstract: Web applications (web apps) have become a key arena for large language models (LLMs) to demonstrate their code generation capabilities and com

We’re releasing the manuscripts, formal Lean certificates, and reasoning walkthroughs so mathematicians can examine these results and build …

Model ReleasesDGX agent

We’re releasing the manuscripts, formal Lean certificates, and reasoning walkthroughs so mathematicians can examine these results and build on their ideas. https://openai.com/index/ten-advances-in-mat

What Is Missing in Surgical Risk Stratification and Outcome Prediction: A Scoping Review of End-to-End Machine Learning Approaches

Model ReleasesDGX agent

arXiv:2607.29090v1 Announce Type: new Abstract: Postoperative adverse events, including mortality and morbidity, remain a major global burden, many of which are preventable through early identificatio

White House invites AI companies to review its new AI safety framework

Model ReleasesDGX agent

Cybersecurity chiefs at the White House have reportedly finalized the outline of a forthcoming framework that will enable artificial intelligence companies to voluntarily submit their latest frontier

Why It Hurts: Identifying the Drivers of Negative Thoughts in Emotional Support Conversations

Model ReleasesDGX agent

arXiv:2607.28648v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used for emotional support tasks, such as negative thought reframing. This task relies on modifying cogn

WitCert: Sound Runtime Risk Observability and Gating for KV-Cache Quantization

Model ReleasesDGX agent

arXiv:2607.28699v1 Announce Type: cross Abstract: KV-cache quantization is validated today by offline benchmark averages; a deployed system cannot tell whether compression is damaging the request it i

You probably used @FireworksAI_HQ this week. The cool part is that you just did not know you did. 🎆 Open-source models are free, sure... bu…

Model ReleasesDGX agent

You probably used @FireworksAI_HQ this week. The cool part is that you just did not know you did. 🎆 Open-source models are free, sure... but the hard part is tailoring them so they perform best on you

You shouldn't need a vision model to know your PDF has checkboxes. LiteParse can now pull structured data directly from your PDFs: form fiel…

Model ReleasesDGX agent

You shouldn't need a vision model to know your PDF has checkboxes. LiteParse can now pull structured data directly from your PDFs: form field values, checkbox states, annotations, embedded images, vec

← Previous
1…4243444546…373
Next →