AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
All
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,553 results
Model Releases

minAction.net: Energy-First Neural Architecture Design -- From Biological Principles to Systematic Validation

DGX agent

arXiv:2604.24805v1 Announce Type: new Abstract: Modern machine learning optimizes for accuracy without explicitly accounting for internal computational cost, even though physical and biological system

model-releasesarxiv-cs-lg
29 Apr 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Minimax Generalized Cross-Entropy

DGX agent

arXiv:2603.19874v3 Announce Type: replace-cross Abstract: Loss functions play a central role in supervised classification. Cross-entropy (CE) is widely used, whereas the mean absolute error (MAE) loss

model-releasesarxiv-cs-lg
29 Apr 2026
Model Releases

Mitigating Coordinate Prediction Bias from Positional Encoding Failures

DGX agent

arXiv:2510.22102v2 Announce Type: replace-cross Abstract: While Multimodal Large Language Models (MLLMs) excel at general vision-language tasks, precise coordinate prediction remains a significant cha

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

MMLANDMARKS: a Cross-View Instance-Level Benchmark for Geo-Spatial Understanding

DGX agent

arXiv:2512.17492v2 Announce Type: replace Abstract: Geo-spatial analysis of our world benefits from a multimodal approach, as every single geographic location can be described in numerous ways (images

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

Model-Harness-Task fit is very real. The world's best agents carefully tailor the harness around the model to take advantage of each model's…

DGX agent

Model-Harness-Task fit is very real. The world's best agents carefully tailor the harness around the model to take advantage of each model's unique intelligence & capabilities 🧠 Today we released Harn

model-releasesharrison-chase--x
29 Apr 2026
Model Releases

Nautile-370M: Spectral Memory Meets Attention in a Small Reasoning Model

DGX agent

arXiv:2604.24809v1 Announce Type: new Abstract: We present Nautile-370M, a 371-million-parameter small language model designed for efficient reasoning under strict parameter and inference budgets. Nau

model-releasesarxiv-cs-lg
29 Apr 2026
Model Releases

Nemotron 3 Nano Omni: Efficient and Open Multimodal Intelligence

DGX agent

arXiv:2604.24954v1 Announce Type: cross Abstract: We introduce Nemotron 3 Nano Omni, the latest model in the Nemotron multimodal series and the first to natively support audio inputs alongside text, i

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

NUBO: A Transparent Python Package for Bayesian Optimization

DGX agent

arXiv:2305.06709v4 Announce Type: replace Abstract: NUBO, short for Newcastle University Bayesian Optimisation, is a Bayesian optimization framework for the optimization of expensive-to-evaluate black

model-releasesarxiv-cs-lg
29 Apr 2026
Model Releases

Odysseys: Benchmarking Web Agents on Realistic Long Horizon Tasks

DGX agent

arXiv:2604.24964v1 Announce Type: cross Abstract: Existing web agent benchmarks have largely converged on short, single-site tasks that frontier models are approaching saturation on. However, real wor

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

OMHBench: Benchmarking Balanced and Grounded Omni-Modal Multi-Hop Reasoning

DGX agent

arXiv:2508.16198v3 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have increasingly supported omni-modal processing across text, vision, and speech. However, existing evalua

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

One Perturbation, Two Failure Modes: Probing VLM Safety via Embedding-Guided Typographic Perturbations

DGX agent

arXiv:2604.25102v1 Announce Type: new Abstract: Typographic prompt injection exploits vision language models' (VLMs) ability to read text rendered in images, posing a growing threat as VLMs power auto

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

OneThinker: All-in-one Reasoning Model for Image and Video

DGX agent

arXiv:2512.03043v3 Announce Type: replace Abstract: Reinforcement learning (RL) has recently achieved remarkable success in eliciting visual reasoning within Multimodal Large Language Models (MLLMs).

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

OpenAI DevDay is back. San Francisco September 29

DGX agent

OpenAI announced the return of DevDay, their developer conference, scheduled for September 29 in San Francisco. The event typically features product announcements, demonstrations of new capabilities,

model-releasesopenai--x
29 Apr 2026
Model Releases

Origin-Destination Demand Prediction: An Urban Radiation and Attraction Perspective

DGX agent

arXiv:2412.00167v2 Announce Type: replace Abstract: In recent years, origin-destination (OD) demand prediction has gained significant attention for its profound implications in urban development. Exis

model-releasesarxiv-cs-lg
29 Apr 2026
Model Releases

Pentagon's Digital and AI Chief Cameron Stanley confirms the DoD is expanding its Google Gemini use, saying 'overreliance on one vendor is never a good thing' (Seema Mody/CNBC)

DGX agent

Seema Mody / CNBC: Pentagon's Digital and AI Chief Cameron Stanley confirms the DoD is expanding its Google Gemini use, saying “overreliance on one vendor is never a good thing” — Pentagon AI chief Ca

model-releasestechmeme
29 Apr 2026
Model Releases

Personalization Toolkit: Training Free Personalization of Large Vision Language Models

DGX agent

arXiv:2502.02452v4 Announce Type: replace Abstract: Personalization of Large Vision-Language Models (LVLMs) involves customizing models to recognize specific users or object instances and to generate

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

Phase-Associative Memory: Sequence Modeling in Complex Hilbert Space

DGX agent

arXiv:2604.05030v2 Announce Type: replace Abstract: Experiments probing natural language processing by both humans and LLMs suggest that the meaning of a semantic expression is indeterminate prior to

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

PolyKV: A Shared Asymmetrically-Compressed KV Cache Pool for Multi-Agent LLM Inference

DGX agent

arXiv:2604.24971v1 Announce Type: cross Abstract: We present PolyKV, a system in which multiple concurrent inference agents share a single, asymmetrically compressed KV cache pool. Rather than allocat

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

PortraVec: Image-Based Portrait Vectorization with Text-Guided Manipulation

DGX agent

arXiv:2410.04182v2 Announce Type: replace Abstract: While portrait sketch generation is a special task in sketch synthesis, most existing methods are pixel-based, limiting their interpretability and e

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

Praxy Voice: Voice-Prompt Recovery + BUPS for Commercial-Class Indic TTS from a Frozen Non-Indic Base at Zero Commercial-Training-Data Cost

DGX agent

arXiv:2604.25441v1 Announce Type: cross Abstract: Commercial TTS systems produce near-native Indic audio, but the best open-source bases (Chatterbox, Indic Parler-TTS, IndicF5) trail them on measured

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

Prior-Aligned Data Cleaning for Tabular Foundation Models

DGX agent

arXiv:2604.25154v1 Announce Type: new Abstract: Tabular Foundation Models (TFMs) achieve state-of-the-art zero-shot accuracy on small tabular datasets by meta-learning over synthetic data-generating p

model-releasesarxiv-cs-lg
29 Apr 2026
Model Releases

Proactive agent that thinks and acts like you. Multiplayer AI Brain for teams. Proper GUI for commanding 50 agents. http://Sauna.ai is all t…

DGX agent

Proactive agent that thinks and acts like you. Multiplayer AI Brain for teams. Proper GUI for commanding 50 agents. http://Sauna.ai is all three. Sauna goes live today. First 2000 people, use access c

model-releasesswyx--x
29 Apr 2026
Model Releases

PSI-Bench: Towards Clinically Grounded and Interpretable Evaluation of Depression Patient Simulators

DGX agent

arXiv:2604.25840v1 Announce Type: new Abstract: Patient simulators are gaining traction in mental health training by providing scalable exposure to complex and sensitive patient interactions. Simulati

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

PSP: An Interpretable Per-Dimension Accent Benchmark for Indic Text-to-Speech

DGX agent

arXiv:2604.25476v1 Announce Type: cross Abstract: Standard text-to-speech (TTS) evaluation measures intelligibility (WER, CER) and overall naturalness (MOS, UTMOS) but does not quantify accent. A synt

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

QB-LIF: Learnable-Scale Quantized Burst Neurons for Efficient SNNs

DGX agent

arXiv:2604.25688v1 Announce Type: new Abstract: Binary spike coding enables sparse and event-driven computation in spiking neural networks (SNNs), yet its 1-bit-per-timestep representation fundamental

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

QCalEval: Benchmarking Vision-Language Models for Quantum Calibration Plot Understanding

DGX agent

arXiv:2604.25884v1 Announce Type: cross Abstract: Quantum computing calibration depends on interpreting experimental data, and calibration plots provide the most universal human-readable representatio

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

Quantifying and Mitigating Socially Desirable Responding in LLMs: A Desirability-Matched Graded Forced-Choice Psychometric Study

DGX agent

arXiv:2602.17262v2 Announce Type: replace Abstract: Human self-report questionnaires are increasingly used in NLP to benchmark and audit large language models (LLMs), from persona consistency to safet

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

Query-Efficient Quantum Approximate Optimization via Graph-Conditioned Trust Regions

DGX agent

arXiv:2604.24803v1 Announce Type: new Abstract: In low-depth implementations of the Quantum Approximate Optimization Algorithm (QAOA), the dominant cost is often the number of objective evaluations ra

model-releasesarxiv-cs-lg
29 Apr 2026
Model Releases

Qwen 3.5 from @Alibaba_Qwen is now available on @FireworksAI_HQ Training Platform across the Managed and Training API workflows. Try SFT, DP…

DGX agent

Qwen 3.5 from @Alibaba_Qwen is now available on @FireworksAI_HQ Training Platform across the Managed and Training API workflows. Try SFT, DPO, RL with smart defaults or your own custom loss function w

model-releasesfireworks-ai--x
29 Apr 2026
Model Releases

Qwen Introduced FlashQLA

DGX agent

FlashQLA is a high-performance linear attention kernel library built on TileLang developed by Alibaba's Qwen team. The introduction of FlashQLA represents an optimization technology designed to improv

model-releasesr-localllama
29 Apr 2026
Model Releases

RbtAct: Rebuttal as Supervision for Actionable Review Feedback Generation

DGX agent

arXiv:2603.09723v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly used across the scientific workflow, including to draft peer-review reports. However, many AI-generate

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

RCProb: Probabilistic Rule Extraction for Efficient Simplification of Tree Ensembles

DGX agent

arXiv:2604.25304v1 Announce Type: new Abstract: Tree ensembles are widely used in industrial machine learning due to their strong predictive performance and efficient training procedures. However, as

model-releasesarxiv-cs-lg
29 Apr 2026
Model Releases

Read the DeepSeek V4 Pro quickstart https://docs.together.ai/docs/deepseek-v4-quickstart

DGX agent

DeepSeek V4 Pro is a language model available through Together AI's platform, with official quickstart documentation provided to help users get started with the model. The quickstart guide likely cove

model-releasestogether-ai--x
29 Apr 2026
Model Releases

Recreated the app with ml intern to compare. Went faster than in cowork, some things are better (ex direct integration in the desktop reachy…

DGX agent

Recreated the app with ml intern to compare. Went faster than in cowork, some things are better (ex direct integration in the desktop reachy mini app), some things are worse (had to debut the install)

model-releasesclem-delangue--x
29 Apr 2026
Model Releases

Representation Paradigms in AI-based 3D Radiological Image Reconstruction: A Systematic Review

DGX agent

arXiv:2504.11349v3 Announce Type: replace Abstract: The demand for high-quality medical imaging in clinical practice and assisted diagnosis has made 3D image reconstruction in radiological imaging a k

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

Residual-loss Anomaly Analysis of Physics-Informed Neural Networks: An Inverse Method for Change-point Detection in Nonlinear Dynamical Systems with Regime Switching

DGX agent

arXiv:2604.25655v1 Announce Type: cross Abstract: Nonlinear dynamical systems with regime transitions are typically described by ordinary differential equations with jumping parameters parameters. Tra

model-releasesarxiv-cs-lg
29 Apr 2026
Model Releases

SARE: Sample-wise Adaptive Reasoning for Training-free Fine-grained Visual Recognition

DGX agent

arXiv:2603.17729v3 Announce Type: replace Abstract: Recent advances in Large Vision-Language Models (LVLMs) have enabled training-free Fine-Grained Visual Recognition (FGVR). However, effectively expl

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

SARU: A Shadow-Aware and Removal Unified Framework for Remote Sensing Images with New Benchmarks

DGX agent

arXiv:2604.25432v1 Announce Type: new Abstract: Shadows are a prevalent problem in remote sensing imagery (RSI), degrading visual quality and severely limiting the performance of downstream tasks like

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

Scaling Probabilistic Transformer via Efficient Cross-Scale Hyperparameter Transfer

DGX agent

arXiv:2604.25409v1 Announce Type: new Abstract: Probabilistic Transformer (PT), a white-box probabilistic model for contextual word representation, has demonstrated substantial similarity to standard

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

SecureScan: An AI-Driven Multi-Layer Framework for Malware and Phishing Detection Using Logistic Regression and Threat Intelligence Integration

DGX agent

arXiv:2602.10750v2 Announce Type: replace-cross Abstract: The growing sophistication of modern malware and phishing campaigns has diminished the effectiveness of traditional signature-based intrusion

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

Self-DACE++: Robust Low-Light Enhancement via Efficient Adaptive Curve Estimation

DGX agent

arXiv:2604.25367v1 Announce Type: new Abstract: In this paper, we present Self-DACE++, an improved unsupervised and lightweight framework for Low-Light Image Enhancement (LLIE), building upon our prev

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

Shearlet Neural Operators for Anisotropic-Shock-Dominated and Multi-scale parametric partial differential equations

DGX agent

arXiv:2604.25181v1 Announce Type: new Abstract: Neural operators have emerged as powerful data-driven surrogates for learning solution operators of parametric partial differential equations (PDEs). Ho

model-releasesarxiv-cs-lg
29 Apr 2026
Model Releases

SIEVES: Selective Prediction Generalizes through Visual Evidence Scoring

DGX agent

arXiv:2604.25855v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) achieve ever-stronger performance on visual-language tasks. Even as traditional visual question answering bench

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

SnapMLA: Efficient Long-Context MLA Decoding via Hardware-Aware FP8 Quantized Pipelining

DGX agent

arXiv:2602.10718v3 Announce Type: replace-cross Abstract: While FP8 attention has shown substantial promise in innovations like FlashAttention-3, its integration into the decoding phase of the DeepSee

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

So excited for this - in the past four months I've written all my code using Deep Agents with: - Opus 4.6 - GPT-5.4 - GLM-5 - Kimi 2.6 - Opu…

DGX agent

So excited for this - in the past four months I've written all my code using Deep Agents with: - Opus 4.6 - GPT-5.4 - GLM-5 - Kimi 2.6 - Opus 4.7 - GPT-5.5 Currently using 5.5 as my driver with Opus s

model-releasesharrison-chase--x
29 Apr 2026
Model Releases

Soft-TransFormers for Continual Learning

DGX agent

arXiv:2411.16073v3 Announce Type: replace-cross Abstract: Inspired by the Well-initialized Lottery Ticket Hypothesis (WLTH), we introduce Soft-Transformer (Soft-TF), a parameter-efficient framework fo

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

Stay tuned for registration details https://openai.com/index/devday-2026/

DGX agent

OpenAI announced an upcoming DevDay 2026 event and indicated that registration details would be shared at a later time. The announcement was made via OpenAI's X (Twitter) account, directing interested

model-releasesopenai--x
29 Apr 2026
Model Releases

Still wondering how you can use Codex for (almost) everything? Codex can help with more of the work that supports the work, from organizing …

DGX agent

Still wondering how you can use Codex for (almost) everything? Codex can help with more of the work that supports the work, from organizing research to making spreadsheets, decks, and summaries. Media

model-releasesopenai--x
29 Apr 2026
← Previous
1…378379380381382…470
Next →