AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,520 results
29 Apr 2026

Google Photos launches an AI try-on feature for clothes you already have

Model ReleasesDGX agent

Google Photos is launching a new AI-powered feature you can use to virtually try on clothes you already have. Using the photos in your gallery, Google will create a virtual 'wardrobe,' allowing you to

Google says paid subscriptions reached 350M in Q1, up 25M QoQ, driven by YouTube and Google One, while Gemini Enterprise paid MAUs grew 40% QoQ (Sarah Perez/TechCrunch)

Model ReleasesDGX agent

Sarah Perez / TechCrunch: Google says paid subscriptions reached 350M in Q1, up 25M QoQ, driven by YouTube and Google One, while Gemini Enterprise paid MAUs grew 40% QoQ — Google has added another 25

GPT-Image-2 in the Wild: A Twitter Dataset of Self-Reported AI-Generated Images from the First Week of Deployment


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model ReleasesDGX agent

arXiv:2604.25370v1 Announce Type: new Abstract: The release of GPT-image-2 by OpenAI marks a watershed moment in AI-generated imagery: the boundary between photographic reality and synthetic content h

Gradient-Direction Sensitivity Reveals Linear-Centroid Coupling Hidden by Optimizer Trajectories

Model ReleasesDGX agent

arXiv:2604.25143v1 Announce Type: new Abstract: We show that replacing the rolling SVD of AdamW updates with a rolling SVD of loss gradients changes the diagnostic by 1-2 orders of magnitude. Performi

GraphPL: Leveraging GNN for Efficient and Robust Modalities Imputation in Patchwork Learning

Model ReleasesDGX agent

arXiv:2604.25352v1 Announce Type: new Abstract: Current research on distributed multi-modal learning typically assumes that clients can access complete information across all modalities, which may not

HANDFUL: Sequential Grasp-Conditioned Dexterous Manipulation with Resource Awareness

Model ReleasesDGX agent

arXiv:2604.25126v1 Announce Type: new Abstract: Dexterous robot hands offer rich opportunities for multifunctional manipulation, where a robot must execute multiple skills in sequence while maintainin

Hidden States as Early Signals: Step-level Trace Evaluation and Pruning for Efficient Test-Time Scaling

Model ReleasesDGX agent

arXiv:2601.09093v2 Announce Type: replace Abstract: Large Language Models (LLMs) can enhance reasoning capabilities through test-time scaling by generating multiple traces. However, the combination of

Hierarchical Reinforcement Learning for the Dynamic VNE with Alternatives Problem

Model ReleasesDGX agent

arXiv:2512.05207v2 Announce Type: replace-cross Abstract: Virtual Network Embedding (VNE) is a key enabler of network slicing, yet most formulations assume that each Virtual Network Request (VNR) has

HuM-Eval: A Coarse-to-Fine Framework for Human-Centric Video Evaluation

Model ReleasesDGX agent

arXiv:2604.25361v1 Announce Type: new Abstract: Video generation models have developed rapidly in recent years, where generating natural human motion plays a pivotal role. However, accurately evaluati

I built a TPU-native medical Q&A fine-tuning pipeline using Gemma 3 with Keras and JAX The project fine-tunes Gemma-3 on medical dialogue da…

Model ReleasesDGX agent

I built a TPU-native medical Q&A fine-tuning pipeline using Gemma 3 with Keras and JAX The project fine-tunes Gemma-3 on medical dialogue data from ChatDoctor and evaluates it on MedMCQA, with a large

i love that the team does stuff like this

Model ReleasesDGX agent

i love that the team does stuff like this Codex is not like claude code. if you know the limit is going to end, like last 10 to 8%, give an very long run task, and even after the limit got ever, it wi

I released LLM 0.32a0 this morning, a major backwards-compatible refactor of my LLM Python library and CLI tool for working with language mo…

Model ReleasesDGX agent

I released LLM 0.32a0 this morning, a major backwards-compatible refactor of my LLM Python library and CLI tool for working with language models - the new changes should help LLM work better with reas

IBM Granite just released two multilingual embedding models with 97M and 311M parameters 🤏🏻 ModernBERT-based, 200+ languages, 32K context,…

Model ReleasesDGX agent

IBM Granite just released two multilingual embedding models with 97M and 311M parameters 🤏🏻 ModernBERT-based, 200+ languages, 32K context, and built for retrieval, search, similarity, and code. And...

Image Classification via Random Dilated Convolution with Multi-Branch Feature Extraction and Context Excitation

Model ReleasesDGX agent

arXiv:2604.25188v1 Announce Type: new Abstract: Image classification remains a fundamental yet challenging task in computer vision, particularly when fine-grained feature extraction and background noi

Images Amplify Misinformation Sharing in Vision-Language Models

Model ReleasesDGX agent

arXiv:2505.13302v2 Announce Type: replace Abstract: As language and vision-language models (VLMs) become central to information access and online interaction, concerns grow about their potential to am

IMO DeepSeek v4 demonstrated utter confidence and competence by not benchmaxxing, not focusing on some BS final run cost, not even spending …

Model ReleasesDGX agent

IMO DeepSeek v4 demonstrated utter confidence and competence by not benchmaxxing, not focusing on some BS final run cost, not even spending inference-optimal compute. just showed up, demonstrated SOTA

Improving LLM Predictions via Inter-Layer Structural Encoders

Model ReleasesDGX agent

arXiv:2603.22665v2 Announce Type: replace Abstract: The standard practice in Large Language Models (LLMs) is to base predictions on final-layer representations. However, intermediate layers encode com

Incompressible Knowledge Probes: Estimating Black-Box LLM Parameter Counts via Factual Capacity

Model ReleasesDGX agent

arXiv:2604.24827v1 Announce Type: new Abstract: Closed-source frontier labs do not disclose parameter counts, and the standard alternative -- inference economics -- carries 2imes+ uncertainty from har

Injecting Measurement Information Yields a Fast and Noise-Robust Diffusion-Based Inverse Problem Solver

Model ReleasesDGX agent

arXiv:2508.02964v3 Announce Type: replace Abstract: Diffusion models have been firmly established as principled zero-shot solvers for linear and nonlinear inverse problems, owing to their powerful ima

Interpretable Fuzzy Modeling Reveals Population-Level Representation Differences in P300 Brain Computer Interfaces Across Neurodivergent and Neurotypical Cohorts

Model ReleasesDGX agent

arXiv:2604.24765v1 Announce Type: cross Abstract: P300-based brain-computer interfaces (BCIs) are widely used for communication, but population heterogeneity may alter the neural patterns available fo

🚀 Introducing FlashQLA: high-performance linear attention kernels built on TileLang. ⚡ 2–3× forward speedup. 2× backward speedup. 💻 Purpos…

Model ReleasesDGX agent

🚀 Introducing FlashQLA: high-performance linear attention kernels built on TileLang. ⚡ 2–3× forward speedup. 2× backward speedup. 💻 Purpose-built for agentic AI on your personal devices. 💡Key insights

Introducing remote agents in Vibe and Mistral Medium 3.5. You can now launch remote agents in the cloud, including from the CLI or Le Chat. …

Model ReleasesDGX agent

Introducing remote agents in Vibe and Mistral Medium 3.5. You can now launch remote agents in the cloud, including from the CLI or Le Chat. Plus, new Work mode in Le Chat for complex, multi-step tasks

Investigation into In-Context Learning Capabilities of Transformers

Model ReleasesDGX agent

arXiv:2604.25858v1 Announce Type: new Abstract: Transformers have demonstrated a strong ability for in-context learning (ICL), enabling models to solve previously unseen tasks using only example input

Is This Just Fantasy? Language Model Representations Reflect Human Judgments of Event Plausibility

Model ReleasesDGX agent

arXiv:2507.12553v3 Announce Type: replace Abstract: Language models (LMs) are used for a diverse range of tasks, from question answering to writing fantastical stories. In order to reliably accomplish

Its a bit frustrating, because Gemini 3.1 Pro is an excellent model and can deliver really good results. But here is GPT-5.5 Pro for compari…

Model ReleasesDGX agent

Its a bit frustrating, because Gemini 3.1 Pro is an excellent model and can deliver really good results. But here is GPT-5.5 Pro for comparison. Sadly, it took this assignment very seriously and ethic

jina-embeddings-v5-text: Task-Targeted Embedding Distillation

Model ReleasesDGX agent

arXiv:2602.15547v2 Announce Type: replace Abstract: Text embedding models are widely used for semantic similarity tasks, including information retrieval, clustering, and classification. General-purpos

KinDER: A Physical Reasoning Benchmark for Robot Learning and Planning

Model ReleasesDGX agent

arXiv:2604.25788v1 Announce Type: new Abstract: Robotic systems that interact with the physical world must reason about kinematic and dynamic constraints imposed by their own embodiment, their environ

Large Language Models Explore by Latent Distilling

Model ReleasesDGX agent

arXiv:2604.24927v1 Announce Type: new Abstract: Generating diverse responses is crucial for test-time scaling of large language models (LLMs), yet standard stochastic sampling mostly yields surface-le

Last week we announced DeepSeek-V4. Today we’re sharing a closer look at DeepSeek-V4 Pro on Together AI: 512K context, controllable reasonin…

Model ReleasesDGX agent

Last week we announced DeepSeek-V4. Today we’re sharing a closer look at DeepSeek-V4 Pro on Together AI: 512K context, controllable reasoning modes, and cached-input pricing for long-context workloads

Lightweight Real-Time Rendering Parameter Optimization via XGBoost-Driven Lookup Tables

Model ReleasesDGX agent

arXiv:2604.25178v1 Announce Type: new Abstract: Achieving a desirable balance between rendering quality and real-time performance is a long-standing challenge in modern game and rendering engines, par

Liquid Neural Network Models for Natural Gas Spot Price Time-Series Forecasting

Model ReleasesDGX agent

arXiv:2604.24788v1 Announce Type: new Abstract: Natural gas is undoubtedly an essential component of the global energy system. Accurate short-term forecasting of natural gas price is challenging due t

Live from @DeepLearningAI conference: our CPO Or Dagan is taking the stage, explaining how we got SOTA on Browsecomp-Plus with the Maestro a…

Model ReleasesDGX agent

AI21 Labs announced that their Chief Product Officer Or Dagan presented at the DeepLearning.AI conference, discussing how the company achieved state-of-the-art results on the Browsecomp-Plus benchmark

LLM 0.32a0 is a major backwards-compatible refactor

Model ReleasesDGX agent

I just released LLM 0.32a0, an alpha release of my LLM Python library and CLI tool for accessing LLMs, with some consequential changes that I've been working towards for quite a while. Previous versio

LLM-ReSum: A Framework for LLM Reflective Summarization through Self-Evaluation

Model ReleasesDGX agent

arXiv:2604.25665v1 Announce Type: new Abstract: Reliable evaluation of large language model (LLM)-generated summaries remains an open challenge, particularly across heterogeneous domains and document

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization

Model ReleasesDGX agent

arXiv:2604.25130v1 Announce Type: new Abstract: Evaluating long document summaries remains the primary bottleneck in summarization research. Existing metrics correlate weakly with human judgments and

Lookout launches mobile-native tool to expose shadow AI on enterprise devices

Model ReleasesDGX agent

Cybersecurity company Lookout Inc. today announced the launch of Lookout AI Visibility & Governance, a new mobile-native solution designed to provide organizations with the visibility needed to discov

M^3-VQA: A Benchmark for Multimodal, Multi-Entity, Multi-Hop Visual Question Answering

Model ReleasesDGX agent

arXiv:2604.25122v1 Announce Type: new Abstract: We present M^3-VQA, a novel knowledge-based Visual Question Answering (VQA) benchmark, to enhance the evaluation of multimodal large language models (ML

MGSM-Pro: A Simple Strategy for Robust Multilingual Mathematical Reasoning Evaluation

Model ReleasesDGX agent

arXiv:2601.21225v2 Announce Type: replace Abstract: Large language models have made substantial progress in mathematical reasoning. However, benchmark development for multilingual evaluation has lagge

MICo-150K: A Comprehensive Dataset Advancing Multi-Image Composition

Model ReleasesDGX agent

arXiv:2512.07348v2 Announce Type: replace Abstract: In controllable image generation, synthesizing coherent and consistent images from multiple reference inputs, i.e., Multi-Image Composition (MICo),

minAction.net: Energy-First Neural Architecture Design -- From Biological Principles to Systematic Validation

Model ReleasesDGX agent

arXiv:2604.24805v1 Announce Type: new Abstract: Modern machine learning optimizes for accuracy without explicitly accounting for internal computational cost, even though physical and biological system

Minimax Generalized Cross-Entropy

Model ReleasesDGX agent

arXiv:2603.19874v3 Announce Type: replace-cross Abstract: Loss functions play a central role in supervised classification. Cross-entropy (CE) is widely used, whereas the mean absolute error (MAE) loss

Mitigating Coordinate Prediction Bias from Positional Encoding Failures

Model ReleasesDGX agent

arXiv:2510.22102v2 Announce Type: replace-cross Abstract: While Multimodal Large Language Models (MLLMs) excel at general vision-language tasks, precise coordinate prediction remains a significant cha

MMLANDMARKS: a Cross-View Instance-Level Benchmark for Geo-Spatial Understanding

Model ReleasesDGX agent

arXiv:2512.17492v2 Announce Type: replace Abstract: Geo-spatial analysis of our world benefits from a multimodal approach, as every single geographic location can be described in numerous ways (images

Model-Harness-Task fit is very real. The world's best agents carefully tailor the harness around the model to take advantage of each model's…

Model ReleasesDGX agent

Model-Harness-Task fit is very real. The world's best agents carefully tailor the harness around the model to take advantage of each model's unique intelligence & capabilities 🧠 Today we released Harn

Nautile-370M: Spectral Memory Meets Attention in a Small Reasoning Model

Model ReleasesDGX agent

arXiv:2604.24809v1 Announce Type: new Abstract: We present Nautile-370M, a 371-million-parameter small language model designed for efficient reasoning under strict parameter and inference budgets. Nau

Nemotron 3 Nano Omni: Efficient and Open Multimodal Intelligence

Model ReleasesDGX agent

arXiv:2604.24954v1 Announce Type: cross Abstract: We introduce Nemotron 3 Nano Omni, the latest model in the Nemotron multimodal series and the first to natively support audio inputs alongside text, i

NUBO: A Transparent Python Package for Bayesian Optimization

Model ReleasesDGX agent

arXiv:2305.06709v4 Announce Type: replace Abstract: NUBO, short for Newcastle University Bayesian Optimisation, is a Bayesian optimization framework for the optimization of expensive-to-evaluate black

Odysseys: Benchmarking Web Agents on Realistic Long Horizon Tasks

Model ReleasesDGX agent

arXiv:2604.24964v1 Announce Type: cross Abstract: Existing web agent benchmarks have largely converged on short, single-site tasks that frontier models are approaching saturation on. However, real wor

OMHBench: Benchmarking Balanced and Grounded Omni-Modal Multi-Hop Reasoning

Model ReleasesDGX agent

arXiv:2508.16198v3 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have increasingly supported omni-modal processing across text, vision, and speech. However, existing evalua

One Perturbation, Two Failure Modes: Probing VLM Safety via Embedding-Guided Typographic Perturbations

Model ReleasesDGX agent

arXiv:2604.25102v1 Announce Type: new Abstract: Typographic prompt injection exploits vision language models' (VLMs) ability to read text rendered in images, posing a growing threat as VLMs power auto

OneThinker: All-in-one Reasoning Model for Image and Video

Model ReleasesDGX agent

arXiv:2512.03043v3 Announce Type: replace Abstract: Reinforcement learning (RL) has recently achieved remarkable success in eliciting visual reasoning within Multimodal Large Language Models (MLLMs).

OpenAI DevDay is back. San Francisco September 29

Model ReleasesDGX agent

OpenAI announced the return of DevDay, their developer conference, scheduled for September 29 in San Francisco. The event typically features product announcements, demonstrations of new capabilities,

Origin-Destination Demand Prediction: An Urban Radiation and Attraction Perspective

Model ReleasesDGX agent

arXiv:2412.00167v2 Announce Type: replace Abstract: In recent years, origin-destination (OD) demand prediction has gained significant attention for its profound implications in urban development. Exis

Pentagon's Digital and AI Chief Cameron Stanley confirms the DoD is expanding its Google Gemini use, saying 'overreliance on one vendor is never a good thing' (Seema Mody/CNBC)

Model ReleasesDGX agent

Seema Mody / CNBC: Pentagon's Digital and AI Chief Cameron Stanley confirms the DoD is expanding its Google Gemini use, saying “overreliance on one vendor is never a good thing” — Pentagon AI chief Ca

Personalization Toolkit: Training Free Personalization of Large Vision Language Models

Model ReleasesDGX agent

arXiv:2502.02452v4 Announce Type: replace Abstract: Personalization of Large Vision-Language Models (LVLMs) involves customizing models to recognize specific users or object instances and to generate

Phase-Associative Memory: Sequence Modeling in Complex Hilbert Space

Model ReleasesDGX agent

arXiv:2604.05030v2 Announce Type: replace Abstract: Experiments probing natural language processing by both humans and LLMs suggest that the meaning of a semantic expression is indeterminate prior to

PolyKV: A Shared Asymmetrically-Compressed KV Cache Pool for Multi-Agent LLM Inference

Model ReleasesDGX agent

arXiv:2604.24971v1 Announce Type: cross Abstract: We present PolyKV, a system in which multiple concurrent inference agents share a single, asymmetrically compressed KV cache pool. Rather than allocat

PortraVec: Image-Based Portrait Vectorization with Text-Guided Manipulation

Model ReleasesDGX agent

arXiv:2410.04182v2 Announce Type: replace Abstract: While portrait sketch generation is a special task in sketch synthesis, most existing methods are pixel-based, limiting their interpretability and e

Praxy Voice: Voice-Prompt Recovery + BUPS for Commercial-Class Indic TTS from a Frozen Non-Indic Base at Zero Commercial-Training-Data Cost

Model ReleasesDGX agent

arXiv:2604.25441v1 Announce Type: cross Abstract: Commercial TTS systems produce near-native Indic audio, but the best open-source bases (Chatterbox, Indic Parler-TTS, IndicF5) trail them on measured

Prior-Aligned Data Cleaning for Tabular Foundation Models

Model ReleasesDGX agent

arXiv:2604.25154v1 Announce Type: new Abstract: Tabular Foundation Models (TFMs) achieve state-of-the-art zero-shot accuracy on small tabular datasets by meta-learning over synthetic data-generating p

← Previous
1…301302303304305…376
Next →