AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
All
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,553 results
Model Releases

From Local to Global: Revisiting Structured Pruning Paradigms for Large Language Models

DGX agent

arXiv:2510.18030v2 Announce Type: replace Abstract: Structured pruning is a practical approach to deploying large language models (LLMs) efficiently, as it yields compact, hardware-friendly architectu

model-releasesarxiv-cs-cl
29 Apr 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Frontier Coding Agents Can Now Implement an AlphaZero Self-Play Machine Learning Pipeline For Connect Four That Performs Comparably to an External Solver

DGX agent

arXiv:2604.25067v1 Announce Type: cross Abstract: Forecasting when AI systems will become capable of meaningfully accelerating AI research is a central challenge for AI safety. Existing benchmarks mea

model-releasesarxiv-cs-lg
29 Apr 2026
Model Releases

G-Loss: Graph-Guided Fine-Tuning of Language Models

DGX agent

arXiv:2604.25853v1 Announce Type: new Abstract: Traditional loss functions, including cross-entropy, contrastive, triplet, and su pervised contrastive losses, used for fine-tuning pre-trained language

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

GAIA-v2-LILT: Multilingual Adaptation of Agent Benchmark beyond Translation

DGX agent

arXiv:2604.24929v1 Announce Type: new Abstract: Agent benchmarks remain largely English-centric, while their multilingual versions are often built with machine translation (MT) and limited post-editin

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

Gemini now can create documents, and it is a nice start, but not up to the frontier yet, as you can see from my 'LBO of Hogwarts' test. Powe…

DGX agent

Gemini now can create documents, and it is a nice start, but not up to the frontier yet, as you can see from my 'LBO of Hogwarts' test. PowerPoints are substantially worse than NotebookLM, spreadsheet

model-releasesethan-mollick--x
29 Apr 2026
Model Releases

General Motors is adding Gemini to four million cars

DGX agent

General Motors is planning to bring Google's Gemini AI assistant to around four million vehicles across the US. Model year 2022 and newer Cadillac, Chevrolet, Buick, and GMC vehicles with Google built

model-releasesthe-verge-ai
29 Apr 2026
Model Releases

Generalizable Human Gaussian Splatting via Multi-view Semantic Consistency

DGX agent

arXiv:2604.25466v1 Announce Type: new Abstract: Recently, generalizable human Gaussian splatting from sparse-view inputs has been actively studied for the photorealistic human rendering. Most existing

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

Genie Sim 3.0 : A High-Fidelity Comprehensive Simulation Platform for Humanoid Robot

DGX agent

arXiv:2601.02078v2 Announce Type: replace Abstract: The development of robust and generalizable robot learning models is critically contingent upon the availability of large-scale, diverse training da

model-releasesarxiv-cs-ro
29 Apr 2026
Model Releases

Golden RPG: Confidence-Adaptive Region-Aware Noise for Compositional Text-to-Image Generation

DGX agent

arXiv:2604.25314v1 Announce Type: new Abstract: Compositional text-to-image (T2I) generation requires a model to honour multiple sub-prompts that describe distinct image regions. Recent work shows tha

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

Google Photos launches an AI try-on feature for clothes you already have

DGX agent

Google Photos is launching a new AI-powered feature you can use to virtually try on clothes you already have. Using the photos in your gallery, Google will create a virtual 'wardrobe,' allowing you to

model-releasesthe-verge-ai
29 Apr 2026
Model Releases

Google says paid subscriptions reached 350M in Q1, up 25M QoQ, driven by YouTube and Google One, while Gemini Enterprise paid MAUs grew 40% QoQ (Sarah Perez/TechCrunch)

DGX agent

Sarah Perez / TechCrunch: Google says paid subscriptions reached 350M in Q1, up 25M QoQ, driven by YouTube and Google One, while Gemini Enterprise paid MAUs grew 40% QoQ — Google has added another 25

model-releasestechmeme
29 Apr 2026
Model Releases

GPT-Image-2 in the Wild: A Twitter Dataset of Self-Reported AI-Generated Images from the First Week of Deployment

DGX agent

arXiv:2604.25370v1 Announce Type: new Abstract: The release of GPT-image-2 by OpenAI marks a watershed moment in AI-generated imagery: the boundary between photographic reality and synthetic content h

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

Gradient-Direction Sensitivity Reveals Linear-Centroid Coupling Hidden by Optimizer Trajectories

DGX agent

arXiv:2604.25143v1 Announce Type: new Abstract: We show that replacing the rolling SVD of AdamW updates with a rolling SVD of loss gradients changes the diagnostic by 1-2 orders of magnitude. Performi

model-releasesarxiv-cs-lg
29 Apr 2026
Model Releases

GraphPL: Leveraging GNN for Efficient and Robust Modalities Imputation in Patchwork Learning

DGX agent

arXiv:2604.25352v1 Announce Type: new Abstract: Current research on distributed multi-modal learning typically assumes that clients can access complete information across all modalities, which may not

model-releasesarxiv-cs-lg
29 Apr 2026
Model Releases

HANDFUL: Sequential Grasp-Conditioned Dexterous Manipulation with Resource Awareness

DGX agent

arXiv:2604.25126v1 Announce Type: new Abstract: Dexterous robot hands offer rich opportunities for multifunctional manipulation, where a robot must execute multiple skills in sequence while maintainin

model-releasesarxiv-cs-ro
29 Apr 2026
Model Releases

Hidden States as Early Signals: Step-level Trace Evaluation and Pruning for Efficient Test-Time Scaling

DGX agent

arXiv:2601.09093v2 Announce Type: replace Abstract: Large Language Models (LLMs) can enhance reasoning capabilities through test-time scaling by generating multiple traces. However, the combination of

model-releasesarxiv-cs-lg
29 Apr 2026
Model Releases

Hierarchical Reinforcement Learning for the Dynamic VNE with Alternatives Problem

DGX agent

arXiv:2512.05207v2 Announce Type: replace-cross Abstract: Virtual Network Embedding (VNE) is a key enabler of network slicing, yet most formulations assume that each Virtual Network Request (VNR) has

model-releasesarxiv-cs-lg
29 Apr 2026
Model Releases

HuM-Eval: A Coarse-to-Fine Framework for Human-Centric Video Evaluation

DGX agent

arXiv:2604.25361v1 Announce Type: new Abstract: Video generation models have developed rapidly in recent years, where generating natural human motion plays a pivotal role. However, accurately evaluati

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

I built a TPU-native medical Q&A fine-tuning pipeline using Gemma 3 with Keras and JAX The project fine-tunes Gemma-3 on medical dialogue da…

DGX agent

I built a TPU-native medical Q&A fine-tuning pipeline using Gemma 3 with Keras and JAX The project fine-tunes Gemma-3 on medical dialogue data from ChatDoctor and evaluates it on MedMCQA, with a large

model-releasesfrancois-chollet--x
29 Apr 2026
Model Releases

i love that the team does stuff like this

DGX agent

i love that the team does stuff like this Codex is not like claude code. if you know the limit is going to end, like last 10 to 8%, give an very long run task, and even after the limit got ever, it wi

model-releasessam-altman--x
29 Apr 2026
Model Releases

I released LLM 0.32a0 this morning, a major backwards-compatible refactor of my LLM Python library and CLI tool for working with language mo…

DGX agent

I released LLM 0.32a0 this morning, a major backwards-compatible refactor of my LLM Python library and CLI tool for working with language models - the new changes should help LLM work better with reas

model-releasessimon-willison--x
29 Apr 2026
Model Releases

IBM Granite just released two multilingual embedding models with 97M and 311M parameters 🤏🏻 ModernBERT-based, 200+ languages, 32K context,…

DGX agent

IBM Granite just released two multilingual embedding models with 97M and 311M parameters 🤏🏻 ModernBERT-based, 200+ languages, 32K context, and built for retrieval, search, similarity, and code. And...

model-releasesjeremy-howard--x
29 Apr 2026
Model Releases

Image Classification via Random Dilated Convolution with Multi-Branch Feature Extraction and Context Excitation

DGX agent

arXiv:2604.25188v1 Announce Type: new Abstract: Image classification remains a fundamental yet challenging task in computer vision, particularly when fine-grained feature extraction and background noi

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

Images Amplify Misinformation Sharing in Vision-Language Models

DGX agent

arXiv:2505.13302v2 Announce Type: replace Abstract: As language and vision-language models (VLMs) become central to information access and online interaction, concerns grow about their potential to am

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

IMO DeepSeek v4 demonstrated utter confidence and competence by not benchmaxxing, not focusing on some BS final run cost, not even spending …

DGX agent

IMO DeepSeek v4 demonstrated utter confidence and competence by not benchmaxxing, not focusing on some BS final run cost, not even spending inference-optimal compute. just showed up, demonstrated SOTA

model-releasesswyx--x
29 Apr 2026
Model Releases

Improving LLM Predictions via Inter-Layer Structural Encoders

DGX agent

arXiv:2603.22665v2 Announce Type: replace Abstract: The standard practice in Large Language Models (LLMs) is to base predictions on final-layer representations. However, intermediate layers encode com

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

Incompressible Knowledge Probes: Estimating Black-Box LLM Parameter Counts via Factual Capacity

DGX agent

arXiv:2604.24827v1 Announce Type: new Abstract: Closed-source frontier labs do not disclose parameter counts, and the standard alternative -- inference economics -- carries 2imes+ uncertainty from har

model-releasesarxiv-cs-lg
29 Apr 2026
Model Releases

Injecting Measurement Information Yields a Fast and Noise-Robust Diffusion-Based Inverse Problem Solver

DGX agent

arXiv:2508.02964v3 Announce Type: replace Abstract: Diffusion models have been firmly established as principled zero-shot solvers for linear and nonlinear inverse problems, owing to their powerful ima

model-releasesarxiv-cs-lg
29 Apr 2026
Model Releases

Interpretable Fuzzy Modeling Reveals Population-Level Representation Differences in P300 Brain Computer Interfaces Across Neurodivergent and Neurotypical Cohorts

DGX agent

arXiv:2604.24765v1 Announce Type: cross Abstract: P300-based brain-computer interfaces (BCIs) are widely used for communication, but population heterogeneity may alter the neural patterns available fo

model-releasesarxiv-cs-lg
29 Apr 2026
Model Releases

🚀 Introducing FlashQLA: high-performance linear attention kernels built on TileLang. ⚡ 2–3× forward speedup. 2× backward speedup. 💻 Purpos…

DGX agent

🚀 Introducing FlashQLA: high-performance linear attention kernels built on TileLang. ⚡ 2–3× forward speedup. 2× backward speedup. 💻 Purpose-built for agentic AI on your personal devices. 💡Key insights

model-releasesqwen--x
29 Apr 2026
Model Releases

Introducing remote agents in Vibe and Mistral Medium 3.5. You can now launch remote agents in the cloud, including from the CLI or Le Chat. …

DGX agent

Introducing remote agents in Vibe and Mistral Medium 3.5. You can now launch remote agents in the cloud, including from the CLI or Le Chat. Plus, new Work mode in Le Chat for complex, multi-step tasks

model-releasesmistral-ai--x
29 Apr 2026
Model Releases

Investigation into In-Context Learning Capabilities of Transformers

DGX agent

arXiv:2604.25858v1 Announce Type: new Abstract: Transformers have demonstrated a strong ability for in-context learning (ICL), enabling models to solve previously unseen tasks using only example input

model-releasesarxiv-cs-lg
29 Apr 2026
Model Releases

Is This Just Fantasy? Language Model Representations Reflect Human Judgments of Event Plausibility

DGX agent

arXiv:2507.12553v3 Announce Type: replace Abstract: Language models (LMs) are used for a diverse range of tasks, from question answering to writing fantastical stories. In order to reliably accomplish

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

Its a bit frustrating, because Gemini 3.1 Pro is an excellent model and can deliver really good results. But here is GPT-5.5 Pro for compari…

DGX agent

Its a bit frustrating, because Gemini 3.1 Pro is an excellent model and can deliver really good results. But here is GPT-5.5 Pro for comparison. Sadly, it took this assignment very seriously and ethic

model-releasesethan-mollick--x
29 Apr 2026
Model Releases

jina-embeddings-v5-text: Task-Targeted Embedding Distillation

DGX agent

arXiv:2602.15547v2 Announce Type: replace Abstract: Text embedding models are widely used for semantic similarity tasks, including information retrieval, clustering, and classification. General-purpos

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

KinDER: A Physical Reasoning Benchmark for Robot Learning and Planning

DGX agent

arXiv:2604.25788v1 Announce Type: new Abstract: Robotic systems that interact with the physical world must reason about kinematic and dynamic constraints imposed by their own embodiment, their environ

model-releasesarxiv-cs-ro
29 Apr 2026
Model Releases

Large Language Models Explore by Latent Distilling

DGX agent

arXiv:2604.24927v1 Announce Type: new Abstract: Generating diverse responses is crucial for test-time scaling of large language models (LLMs), yet standard stochastic sampling mostly yields surface-le

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

Last week we announced DeepSeek-V4. Today we’re sharing a closer look at DeepSeek-V4 Pro on Together AI: 512K context, controllable reasonin…

DGX agent

Last week we announced DeepSeek-V4. Today we’re sharing a closer look at DeepSeek-V4 Pro on Together AI: 512K context, controllable reasoning modes, and cached-input pricing for long-context workloads

model-releasestogether-ai--x
29 Apr 2026
Model Releases

Lightweight Real-Time Rendering Parameter Optimization via XGBoost-Driven Lookup Tables

DGX agent

arXiv:2604.25178v1 Announce Type: new Abstract: Achieving a desirable balance between rendering quality and real-time performance is a long-standing challenge in modern game and rendering engines, par

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

Liquid Neural Network Models for Natural Gas Spot Price Time-Series Forecasting

DGX agent

arXiv:2604.24788v1 Announce Type: new Abstract: Natural gas is undoubtedly an essential component of the global energy system. Accurate short-term forecasting of natural gas price is challenging due t

model-releasesarxiv-cs-lg
29 Apr 2026
Model Releases

Live from @DeepLearningAI conference: our CPO Or Dagan is taking the stage, explaining how we got SOTA on Browsecomp-Plus with the Maestro a…

DGX agent

AI21 Labs announced that their Chief Product Officer Or Dagan presented at the DeepLearning.AI conference, discussing how the company achieved state-of-the-art results on the Browsecomp-Plus benchmark

model-releasesai21-labs--x
29 Apr 2026
Model Releases

LLM 0.32a0 is a major backwards-compatible refactor

DGX agent

I just released LLM 0.32a0, an alpha release of my LLM Python library and CLI tool for accessing LLMs, with some consequential changes that I've been working towards for quite a while. Previous versio

model-releasessimon-willison
29 Apr 2026
Model Releases

LLM-ReSum: A Framework for LLM Reflective Summarization through Self-Evaluation

DGX agent

arXiv:2604.25665v1 Announce Type: new Abstract: Reliable evaluation of large language model (LLM)-generated summaries remains an open challenge, particularly across heterogeneous domains and document

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization

DGX agent

arXiv:2604.25130v1 Announce Type: new Abstract: Evaluating long document summaries remains the primary bottleneck in summarization research. Existing metrics correlate weakly with human judgments and

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

Lookout launches mobile-native tool to expose shadow AI on enterprise devices

DGX agent

Cybersecurity company Lookout Inc. today announced the launch of Lookout AI Visibility & Governance, a new mobile-native solution designed to provide organizations with the visibility needed to discov

model-releasessiliconangle
29 Apr 2026
Model Releases

M^3-VQA: A Benchmark for Multimodal, Multi-Entity, Multi-Hop Visual Question Answering

DGX agent

arXiv:2604.25122v1 Announce Type: new Abstract: We present M^3-VQA, a novel knowledge-based Visual Question Answering (VQA) benchmark, to enhance the evaluation of multimodal large language models (ML

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

MGSM-Pro: A Simple Strategy for Robust Multilingual Mathematical Reasoning Evaluation

DGX agent

arXiv:2601.21225v2 Announce Type: replace Abstract: Large language models have made substantial progress in mathematical reasoning. However, benchmark development for multilingual evaluation has lagge

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

MICo-150K: A Comprehensive Dataset Advancing Multi-Image Composition

DGX agent

arXiv:2512.07348v2 Announce Type: replace Abstract: In controllable image generation, synthesizing coherent and consistent images from multiple reference inputs, i.e., Multi-Image Composition (MICo),

model-releasesarxiv-cs-cv
29 Apr 2026
← Previous
1…377378379380381…470
Next →