AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
All
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,611 results
Model Releases

How Far Are We From True Auto-Research?

DGX agent

arXiv:2605.19156v1 Announce Type: new Abstract: Recent auto-research systems can produce complete papers, but feasibility is not the same as quality, and the field still lacks a systematic study of ho

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

How Ramp engineers accelerate code review with Codex

DGX agent
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Ramp engineers use OpenAI's Codex to streamline and accelerate their code review process, leveraging AI-assisted code understanding and analysis. The implementation likely demonstrates how Codex helps

model-releasesopenai
20 May 2026
Model Releases

Hybrid-LoRA: Bridging Full Fine-Tuning and Low-Rank Adaptation for Post-Training

DGX agent

arXiv:2605.18822v1 Announce Type: cross Abstract: Post-training has become essential for adapting large language models (LLMs) to complex downstream behaviors, including instruction following, prefere

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

I am starting to have trouble paying attention to even interesting information if it is written in Claude or ChatGPT house style. I think so…

DGX agent

I am starting to have trouble paying attention to even interesting information if it is written in Claude or ChatGPT house style. I think some is the sameness of the rhythm rather than obvious tics: C

model-releasesethan-mollick--x
20 May 2026
Model Releases

I don't have much to say about this year's Google I/O because I prefer to write about products that have shipped, not just 'coming soon' ann…

DGX agent

I don't have much to say about this year's Google I/O because I prefer to write about products that have shipped, not just 'coming soon' announcements - but here are some notes on Gemini Spark and Ant

model-releasessimon-willison--x
20 May 2026
Model Releases

i don't think i need cloud models anymore

DGX agent

i don't think i need cloud models anymore MTP speedup Qwen by 2.5x in Atomic Chat Dense vs MoE models on 2x RTX 5090 Qwen3.6 27B: 51 → 117 tps +137% Qwen3.6 35B-A3B: 218 → 267 tps +25% MTP drafts seve

model-releasesclem-delangue--x
20 May 2026
Model Releases

iGSP:Implicit Gradient Subspace Projection for Efficient Continual Learning of Vision-Language Models

DGX agent

arXiv:2605.19301v1 Announce Type: new Abstract: Vision-Language Models require efficient adaptation to continually emerging downstream tasks. While Parameter-Efficient Fine-Tuning mitigates catastroph

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

IMLJD: A Computational Dataset for Indian Matrimonial Litigation Analysis

DGX agent

arXiv:2605.19346v1 Announce Type: cross Abstract: We present IMLJD, an open dataset of 3,613 Indian court judgments covering matrimonial disputes under IPC Section 498A, the Protection of Women from D

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

In-Context Learning Operates as Concept Subspace Learning

DGX agent

arXiv:2605.18830v1 Announce Type: new Abstract: Regression and Bayesian accounts of in-context learning (ICL) explain how demonstrations can induce predictors, while mechanistic analyses often identif

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

Information Processing Capacity of Stationary Physical Systems: Theory, Data-efficient Estimation Methods, and Photonic Demonstration

DGX agent

arXiv:2605.19152v1 Announce Type: cross Abstract: Physical computing systems provide a promising route toward hardware-native machine learning, but their computational capabilities remain difficult to

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

INSHAPE: Instance-Level Shapelets for Interpretable Time-Series Classification

DGX agent

arXiv:2605.20088v1 Announce Type: cross Abstract: Discovering shapelets -- i.e., discriminative temporal patterns within time series -- has been widely studied to address the inherent complexity of ti

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Introducing Agent Executor, Google’s distributed Agent Runtime

DGX agent

As models and harnesses improve, agents are taking on increasingly complex tasks that can run for hours or even days. But as we push agents to do more, this has surfaced a new operational problem: lon

model-releasesgoogle-cloud-ai
20 May 2026
Model Releases

Introducing: Cohere Command A+ We’ve created our most powerful LLM yet, optimized it to run on as little hardware as possible, and released …

DGX agent

Cohere announced Command A+, their most advanced large language model to date, designed with optimization for efficient hardware requirements to enable broader deployment and accessibility. The model

model-releasescohere--x
20 May 2026
Model Releases

It's been *almost* a bit quiet around LLM architecture releases in the past two weeks 😅 Interesting tidbit is the parallel block design. Vi…

DGX agent

It's been *almost* a bit quiet around LLM architecture releases in the past two weeks 😅 Interesting tidbit is the parallel block design. Via the Cmd-A the tech report 'equivalent performance but signi

model-releasessebastian-raschka--x
20 May 2026
Model Releases

it’s in gemini, just create it in ai studio. oh, that’s for your personal google one account. for workspace you need gemini business. no, no…

DGX agent

it’s in gemini, just create it in ai studio. oh, that’s for your personal google one account. for workspace you need gemini business. no, not gemini advanced, that’s ai pro now. unless you need ai ult

model-releasessimon-willison--x
20 May 2026
Model Releases

JAXenstein: Accelerated Benchmarking for First-Person Environments

DGX agent

arXiv:2605.19926v1 Announce Type: new Abstract: The progression of reinforcement learning algorithms have been driven by challenging benchmarks. The rate in which a researcher can iterate on a problem

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

K-Quantization and its Impact on Output Performance

DGX agent

arXiv:2605.19645v1 Announce Type: new Abstract: Recent advancements in large language models (LLMs) have shown their remarkable capacities in many NLP tasks. However, their substantial size often pres

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

KappaPlace: Learning Hyperspherical Uncertainty for Visual Place Recognition via Prototype-Anchored Supervision

DGX agent

arXiv:2605.19435v1 Announce Type: cross Abstract: Visual Place Recognition (VPR) is critical for autonomous navigation, yet state-of-the-art methods lack well-calibrated uncertainty estimation. Standa

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Learn-by-Wire Training Control Governance: Bounded Autonomous Training Under Stress for Stability and Efficiency

DGX agent

arXiv:2605.19008v1 Announce Type: new Abstract: Modern language-model training is increasingly exposed to instability, degraded runs, and wasted compute, especially under aggressive learning-rate, sca

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Learning-Accelerated Optimization-based Trajectory Planning for Cooperative Aerial-Ground Handover Missions

DGX agent

arXiv:2605.19562v1 Announce Type: cross Abstract: This paper presents a learning-augmented trajectory planning framework for cooperative unmanned aerial vehicle (UAV) and unmanned ground vehicle (UGV)

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

Learning Efficient Guardrails for Compliance

DGX agent

arXiv:2510.03485v2 Announce Type: replace Abstract: Autonomous web agents are increasingly deployed for long-horizon tasks, yet their ability to adhere to real-world policies remains critically undere

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Learning When to Adapt

DGX agent

arXiv:2605.19028v1 Announce Type: new Abstract: Low-rank adaptation (LoRA) is a widely used parameter-efficient fine-tuning method, yet its learned correction is static: the same low-rank update is ap

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

Lens Privacy Sealing: A New Benchmark and Method for Physical Privacy-Preserving Action Recognition

DGX agent

arXiv:2605.19578v1 Announce Type: cross Abstract: RGB camera-based surveillance systems enable human action recognition for public safety and healthcare, yet raise serious privacy concerns. Existing m

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Less Back-and-Forth: A Comparative Study of Structured Prompting

DGX agent

arXiv:2605.20149v1 Announce Type: cross Abstract: Large language models (LLMs) are widely used for open-ended tasks, but underspecified prompts can lead to low-quality answers and additional interacti

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Library Hallucinations in LLM-Generated Code: A Risk Analysis Grounded in Developer Queries

DGX agent

arXiv:2509.22202v3 Announce Type: replace-cross Abstract: Large language models (LLMs) now play a central role in code generation, yet they continue to hallucinate, frequently inventing non-existent l

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

LIFT and PLACE: A Simple, Stable, and Effective Knowledge Distillation Framework for Lightweight Diffusion Models

DGX agent

arXiv:2605.19729v1 Announce Type: cross Abstract: We demonstrate that in knowledge distillation for diffusion models, the teacher network's highly complex denoising process - stemming from its substan

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Lightweight and Fast Backdoor Model Detection

DGX agent

arXiv:2605.18907v1 Announce Type: cross Abstract: Deep neural networks (DNN), despite their remarkable performance, are highly vulnerable to backdoor attacks. Existing defenses mainly rely on activati

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

LLM Benchmark Datasets Should Be Contamination-Resistant

DGX agent

arXiv:2605.19999v1 Announce Type: cross Abstract: Benchmark datasets are critical for reproducible, reliable, and discriminative evaluation of LLMs. However, recent studies reveal that many benchmark

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

LLMEval-Logic: A Solver-Verified Chinese Benchmark for Logical Reasoning of LLMs with Adversarial Hardening

DGX agent

arXiv:2605.19597v1 Announce Type: new Abstract: Evaluating large language models (LLMs) on natural-language logical reasoning is essential because rule-governed tasks require conclusions to follow str

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

LMM-Track4D: Eliciting 4D Dynamic Reasoning in LMMs via Trajectory-Grounded Dialogue

DGX agent

arXiv:2605.19390v1 Announce Type: new Abstract: Recent large multimodal models (LMMs) have become increasingly capable on image and video understanding, yet still struggle to sustain 4D continuous spa

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

LWiAI Podcast #245 - TML-Interaction, Claude For Legal, Sam Altman on Stand

DGX agent

This podcast episode covers three main topics: TML-Interaction (likely a new AI model or technical development), the application of Claude AI in legal settings and use cases, and Sam Altman's testimon

model-releaseslast-week-in-ai
20 May 2026
Model Releases

Lying Is Just a Phase: The Hidden Alignment Transition in Language Model Scaling

DGX agent

arXiv:2605.18838v1 Announce Type: cross Abstract: Scaling laws predict loss from compute but not how capabilities interact. We measure the coupling between reasoning and truthfulness across 63 base mo

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Lynx: Enabling Efficient MoE Inference through Dynamic Batch-Aware Expert Selection

DGX agent

arXiv:2411.08982v3 Announce Type: replace Abstract: Selective parameter activation provided by Mixture-of-Expert (MoE) models have made them a popular choice in modern foundational models. However, Mo

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

m3BERT: A Modern, Multi-lingual, Matryoshka Bidirectional Encoder

DGX agent

arXiv:2605.19568v1 Announce Type: new Abstract: Embedding models are pivotal in industrial information retrieval systems like search and advertising. However, existing pretrained models often exhibit

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

MAM-CLIP: Vision-Language Pretraining on Mammography Atlases for BI-RADS Classification

DGX agent

arXiv:2605.19359v1 Announce Type: new Abstract: Deep learning methods have demonstrated promising results in predicting BI-RADS scores from mammography images. However, the interpretation of these ima

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

Managed Agents through the Gemini API is @GoogleAI's response to Anthropic Managed Agents Since it's powered by the new Antigravity agent bu…

DGX agent

Managed Agents through the Gemini API is @GoogleAI's response to Anthropic Managed Agents Since it's powered by the new Antigravity agent built on Gemini 3.5 Flash, it is the most cost-effective gener

model-releasesjerry-liu--x
20 May 2026
Model Releases

MANGO: Meta-Adaptive Network Gradient Optimization for Online Continual Learning

DGX agent

arXiv:2605.19080v1 Announce Type: cross Abstract: In Online Continual Learning (OCL), a neural network sequentially learns from a non-stationary data stream in a single-pass with access only to a limi

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Mask-to-Correct^+: Leveraging Retriever Diversity for Masking-guided Faithful Fact Correction

DGX agent

arXiv:2605.18776v1 Announce Type: cross Abstract: The rapid spread of misinformation on social media highlights the need for robust, automated fact correction frameworks. However, existing works rely

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Mathematical Reasoning in Large Language Models: Benchmarks, Architectures, Evaluation, and Open Challenges

DGX agent

arXiv:2605.19723v1 Announce Type: cross Abstract: Mathematical reasoning is essential for problem-solving in education, science, and industry, serving as a crucial benchmark for evaluating artificial

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Measuring Safety Alignment Effects in Autonomous Security Agents

DGX agent

arXiv:2605.19722v1 Announce Type: cross Abstract: Do stock safety-aligned language models and their uncensored or abliterated derivatives behave differently when run as autonomous security agents? Sin

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

MedFM-Robust: Benchmarking Robustness of Medical Foundation Models

DGX agent

arXiv:2605.19027v1 Announce Type: new Abstract: Medical foundation models (MedFMs) have emerged as transformative tools in healthcare, demonstrating capabilities across diverse clinical applications.

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

Meet Stable Audio 3.0, the model family built for artistic experimentation with open-weight models

DGX agent

Stable Audio 3.0 is Stability AI's latest generative audio model designed to create music and soundscapes from textual prompts. The release includes four models ranging from 459M to 2.7B parameters, w

model-releasesstability-ai
20 May 2026
Model Releases

// Memory as a Model // The paper augments any LLM with a separate trained memory model that stores, retrieves, and integrates facts on its …

DGX agent

// Memory as a Model // The paper augments any LLM with a separate trained memory model that stores, retrieves, and integrates facts on its behalf. It decouples memory updates from base-model weight u

model-releasesdair-ai--x
20 May 2026
Model Releases

MetaRA: Metamorphic Robustness Assessment for Multimodal Large Language Model-based Visual Question Answering Systems

DGX agent

arXiv:2605.19307v1 Announce Type: new Abstract: Visual Question Answering (VQA), as the representative multimodal task, serves as a key benchmark for evaluating the reasoning capabilities of Multimoda

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

Microsoft Senior AI developer just showed how they build AI agents with Claude at Microsoft. 34-minutes. free. By Microsoft team Opus 4.7 + …

DGX agent

Microsoft Senior AI developer just showed how they build AI agents with Claude at Microsoft. 34-minutes. free. By Microsoft team Opus 4.7 + 1,400+ pre-built MCP tools plug Claude into agent → give it

model-releasesboris-cherny--x
20 May 2026
Model Releases

MIRO: MultI-Reward cOnditioned pretraining improves T2I quality and efficiency

DGX agent

arXiv:2510.25897v2 Announce Type: replace Abstract: The default paradigm of post-training text-to-image generators includes post-hoc selection of generated images, and subsequent training with one rew

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

MixRea: Benchmarking Explicit-Implicit Reasoning in Large Language Models

DGX agent

arXiv:2605.20128v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly integrated into high-stakes decision-making. Inspired by the theory of inattentional blindness in human co

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

MotionMERGE: A Multi-granular Framework for Human Motion Editing, Reasoning, Generation, and Explanation

DGX agent

arXiv:2605.18956v1 Announce Type: new Abstract: Recent motion-language models unify tasks like comprehension and generation but operate at a coarse granularity, lacking fine-grained understanding and

model-releasesarxiv-cs-cv
20 May 2026
← Previous
1…287288289290291…472
Next →