AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,603 results
Model Releases

Hierarchical Transformer Preconditioning for Interactive Physics Simulation

DGX agent

arXiv:2605.13343v1 Announce Type: cross Abstract: Neural preconditioners for real-time physics simulation offer promising data-driven priors, but they often fail to capture long-range couplings effici

model-releasesarxiv-cs-lg
14 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

High-Dimensional Analysis of Bootstrap Ensemble Classifiers

DGX agent

arXiv:2505.14587v2 Announce Type: replace-cross Abstract: Bootstrap methods have long been the cornerstone of ensemble learning in machine learning. This paper presents a theoretical analysis of boots

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

High-Rate Quantized Matrix Multiplication II

DGX agent

arXiv:2605.13768v1 Announce Type: cross Abstract: This is the second part of the work investigating quantized matrix multiplication (MatMul). In part I we considered the case of calibration-free quant

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

HLS-Seek: QoR-Aware Code Generation for High-Level Synthesis via Proxy Comparative Reward Reinforcement Learning

DGX agent

arXiv:2605.13536v1 Announce Type: cross Abstract: High-Level Synthesis (HLS) compiles algorithmic C/C++ descriptions into hardware, with Quality of Results (QoR) -- latency and resource utilization --

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

(How) Do Large Language Models Understand High-Level Message Sequence Charts?

DGX agent

arXiv:2605.13773v1 Announce Type: cross Abstract: Large Language Models (LLMs) are being employed widely to automate tasks across the software development life-cycle. It is, however, unclear whether t

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

How to Interpret Agent Behavior

DGX agent

arXiv:2605.13625v1 Announce Type: new Abstract: Autonomous agents such as Claude Code and Codex now operate for hours or even days. Understanding their runtime behavior has become critical for downstr

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

How Well Do Large-Scale Chemical Language Models Transfer to Downstream Tasks?

DGX agent

arXiv:2602.11618v4 Announce Type: replace Abstract: Chemical Language Models (CLMs) pre-trained on large scale molecular data are widely used for molecular property prediction. However, the common bel

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

I was going to say LangChain Labs is the most exciting thing we’ve announced recently, but with so many launches it’s hard to justify… but s…

DGX agent

I was going to say LangChain Labs is the most exciting thing we’ve announced recently, but with so many launches it’s hard to justify… but still incredibly thrilled to see this get started!! Checkout

model-releasesharrison-chase--x
14 May 2026
Model Releases

Identifying the nonlinear string dynamics with port-Hamiltonian neural networks

DGX agent

arXiv:2605.12785v1 Announce Type: new Abstract: Hybrid machine learning combines physical knowledge with data-driven models to enhance interpretability and performance. In this context, Port-Hamiltoni

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

ImageAttributionBench: How Far Are We from Generalizable Attribution?

DGX agent

arXiv:2605.12967v1 Announce Type: new Abstract: The rapid advancement of generative AI has enabled the creation of highly realistic and diverse synthetic images, posing critical challenges for image p

model-releasesarxiv-cs-cv
14 May 2026
Model Releases

Implicit Behavioral Decoding from Next-Step Spike Forecasts at Population Scale

DGX agent

arXiv:2605.12999v1 Announce Type: cross Abstract: Closed-loop brain-computer interfaces often require both a forecast of upcoming neural population activity and a readout of the animal's behavioral st

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

Imposing Boundary Conditions on Neural Operators via Learned Function Extensions

DGX agent

arXiv:2602.04923v2 Announce Type: replace Abstract: Neural operators have emerged as powerful surrogates for the solution of partial differential equations (PDEs), yet their ability to handle general,

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

IndicMedDialog: A Parallel Multi-Turn Medical Dialogue Dataset for Accessible Healthcare in Indic Languages

DGX agent

arXiv:2605.13292v1 Announce Type: cross Abstract: Most existing medical dialogue systems operate in a single-turn question--answering paradigm or rely on template-based datasets, limiting conversation

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Inducing Overthink: Hierarchical Genetic Algorithm-based DoS Attack on Black-Box Large Language Reasoning Models

DGX agent

arXiv:2605.13338v1 Announce Type: cross Abstract: Large Reasoning Models (LRMs) are increasingly integrated into systems requiring reliable multi-step inference, yet this growing dependence exposes ne

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Inference-Time Machine Unlearning via Gated Activation Redirection

DGX agent

arXiv:2605.12765v1 Announce Type: new Abstract: Large Language Models memorize vast amounts of training data, raising concerns regarding privacy, copyright infringement, and safety. Machine unlearning

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

ISOMORPH: A Supply Chain Digital Twin for Simulation, Dataset Generation, and Forecasting Benchmarks

DGX agent

arXiv:2605.12768v1 Announce Type: cross Abstract: Open time-series forecasting (TSF) benchmarks cover retail, energy, weather, and traffic, but supply-chain logistics remains underserved. We introduce

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

KamonBench: A Grammar-Based Dataset for Evaluating Compositional Factor Recovery in Vision-Language Models

DGX agent

arXiv:2605.13322v1 Announce Type: new Abstract: Kamon (family crests) are an important part of Japanese culture and a natural test case for compositional visual recognition: each crest combines a smal

model-releasesarxiv-cs-cv
14 May 2026
Model Releases

Kernel-based guarantees for nonlinear parametric models in Bayesian optimization

DGX agent

arXiv:2605.13160v1 Announce Type: cross Abstract: Modern Bayesian optimization and adaptive sampling methods increasingly rely on nonlinear parametric models, yet theoretical guarantees for such model

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

Kimi K2.6 and DeepSeek V4 Pro are now GA on @FireworksAI_HQ on Foundry + PTU support in the US Data Zone—predictable performance, data resid…

DGX agent

Kimi K2.6 and DeepSeek V4 Pro are now GA on @FireworksAI_HQ on Foundry + PTU support in the US Data Zone—predictable performance, data residency, enterprise SLAs. Frontier open models. Azure controls.

model-releasesfireworks-ai--x
14 May 2026
Model Releases

Kimi K2.6 is now open-weight #1 on Finance Agent Benchmark V2.

DGX agent

Kimi K2.6 is now open-weight #1 on Finance Agent Benchmark V2. Can AI do the job of a financial analyst? We just released V2 of our Finance Agent Benchmark and tested the frontier models. The results

model-releaseskimi-moonshot--x
14 May 2026
Model Releases

Kiwi-Edit: Versatile Video Editing via Instruction and Reference Guidance

DGX agent

arXiv:2603.02175v4 Announce Type: replace-cross Abstract: Instruction-based video editing has witnessed rapid progress, yet current methods often struggle with precise visual control, as natural langu

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

LangChain 在 Interrupt 大会上发布了底层数据库 SmithDB 和自动化排障引擎 LangSmith Engine。 Agent 运行会产生海量 trace(执行轨迹),把旧数据库撑到了瓶颈。新底座 SmithDB 放弃了本地磁盘,全面转向对象存储,将核心查询…

DGX agent

LangChain 在 Interrupt 大会上发布了底层数据库 SmithDB 和自动化排障引擎 LangSmith Engine。 Agent 运行会产生海量 trace(执行轨迹),把旧数据库撑到了瓶颈。新底座 SmithDB 放弃了本地磁盘,全面转向对象存储,将核心查询速度拉高了 15 倍。 底座换新后,LangSmith Engine 顺势接管了查 Bug 的体力活。它在后台持续监控生

model-releasesharrison-chase--x
14 May 2026
Model Releases

Language Model Goal Selection Differs from Humans' in a Self-Directed Learning Task

DGX agent

arXiv:2603.03295v2 Announce Type: replace-cross Abstract: Whether in agentic workflows, social studies, or chat settings, large language models (LLMs) are increasingly being asked to replace humans in

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Large Language Models Lack Temporal Awareness of Medical Knowledge

DGX agent

arXiv:2605.13045v1 Announce Type: new Abstract: The existing methods for evaluating the medical knowledge of Large Language Models (LLMs) are largely based on atemporal examination-style benchmarks, w

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

LeanSearch v2: Global Premise Retrieval for Lean 4 Theorem Proving

DGX agent

arXiv:2605.13137v1 Announce Type: cross Abstract: Proving theorems in Lean 4 often requires identifying a scattered set of library lemmas whose joint use enables a concise proof -- a task we call glob

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Learning a Continue-Thinking Token for Enhanced Test-Time Scaling

DGX agent

arXiv:2506.11274v2 Announce Type: replace-cross Abstract: Test-time scaling has emerged as an effective approach for improving language model performance by utilizing additional compute at inference t

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

Learning Responsibility-Attributed Adversarial Scenarios for Testing Autonomous Vehicles

DGX agent

arXiv:2605.13751v1 Announce Type: new Abstract: Establishing trustworthy safety assurance for autonomous driving systems (ADSs) requires evidence that failures arise from avoidable system deficiencies

model-releasesarxiv-cs-ro
14 May 2026
Model Releases

LENS: Multi-level Evaluation of Multimodal Reasoning with Large Language Models

DGX agent

arXiv:2505.15616v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have achieved significant advances in integrating visual and linguistic information, yet their ability to r

model-releasesarxiv-cs-cv
14 May 2026
Model Releases

Lifelong Learning in Vision-Language Models: Enhanced EWC with Cross-Modal Knowledge Retention

DGX agent

arXiv:2605.12789v1 Announce Type: new Abstract: Large language-vision models (LVLMs) such as CLIP, Flamingo, and BLIP have revolutionized AI by enabling understanding across textual and visual modalit

model-releasesarxiv-cs-ro
14 May 2026
Model Releases

LIFT: Last-Mile Fine-Tuning for Table Explicitation

DGX agent

arXiv:2605.13424v1 Announce Type: new Abstract: We propose last-mile fine-tuning, or Lift, a pipeline in which a pre-trained large language model extracts an initial table from unstructured clipboard

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

Like any AI dev not employed by a closed lab, we share the ambition that at least 10x more should be capable of training frontier models in …

DGX agent

Like any AI dev not employed by a closed lab, we share the ambition that at least 10x more should be capable of training frontier models in 2026. And like we all know, some 10x leaps can't survive a c

model-releasesfireworks-ai--x
14 May 2026
Model Releases

LLMs as annotators of credibility assessment in Danish asylum decisions: evaluating classification performance and errors beyond aggregated metrics

DGX agent

arXiv:2605.13412v1 Announce Type: cross Abstract: Off-the-shelf large language models (LLMs) are increasingly used to automate text annotation, yet their effectiveness remains underexplored for underr

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Looks like @NVIDIAAI has released official getting started guide for hermes on DGX Spark! Great way for new folks to get started and believe…

DGX agent

NVIDIA AI has released an official getting started guide for Hermes on DGX Spark, providing new users with foundational resources and instructions for deploying or working with Hermes on NVIDIA's DGX

model-releasesnous-research--x
14 May 2026
Model Releases

LoRA-Mixer: Coordinate Modular LoRA Experts Through Serial Attention Routing

DGX agent

arXiv:2507.00029v2 Announce Type: replace-cross Abstract: Recent attempts to combine low-rank adaptation (LoRA) with mixture-of-experts (MoE) for multi-task adaptation of Large Language Models (LLMs)

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Low-Rank Adapters Initialization via Gradient Surgery for Continual Learning

DGX agent

arXiv:2605.12752v1 Announce Type: new Abstract: LoRA is widely adopted for continual fine-tuning of Large Language Models due to its parameter efficiency, modularity across tasks, and compatibility wi

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

LY Corp launches a bid with Bain Capital to buy Kakaku.com, which operates restaurant review and booking site Tabelog, for 4B, challenging EQT's 3.75B offer (Nikkei Asia)

DGX agent

Nikkei Asia: LY Corp launches a bid with Bain Capital to buy Kakaku.com, which operates restaurant review and booking site Tabelog, for 4B, challenging EQT's 3.75B offer — TOKYO — The operator of the

model-releasestechmeme
14 May 2026
Model Releases

Many-Shot CoT-ICL: Making In-Context Learning Truly Learn

DGX agent

arXiv:2605.13511v1 Announce Type: cross Abstract: In-context learning (ICL) adapts large language models (LLMs) to new tasks by conditioning on demonstrations in the prompt without parameter updates.

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Mechanistic Evidence for Spectral Structures in Prior-Data Fitted Networks

DGX agent

arXiv:2601.21731v2 Announce Type: replace Abstract: Prior-Data Fitted Networks (PFNs) enable amortized Bayesian inference in a single forward pass, yet their internal representations remain opaque. It

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

MedCore: Boundary-Preserving Medical Core Pruning for MedSAM

DGX agent

arXiv:2605.13688v1 Announce Type: new Abstract: Medical segmentation foundation models such as SAM and MedSAM provide strong prompt-driven segmentation, but their image encoders are still too large fo

model-releasesarxiv-cs-cv
14 May 2026
Model Releases

MedOpenClaw and MedFlowBench: Auditing Medical Agents in Full-Study Workflows

DGX agent

arXiv:2603.24649v2 Announce Type: replace Abstract: Medical imaging benchmarks often evaluate VLMs on pre-selected 2D images, slices, crops, or patches, making evaluation closer to visual recognition.

model-releasesarxiv-cs-cv
14 May 2026
Model Releases

Meet Kimi Web Bridge - Kimi's browser extension. Agent can now interact with websites like a human: search, scroll, click, type and complete…

DGX agent

Meet Kimi Web Bridge - Kimi's browser extension. Agent can now interact with websites like a human: search, scroll, click, type and complete tasks. Supports Kimi Code CLI, Claude Code, Cursor, Codex,

model-releaseskimi-moonshot--x
14 May 2026
Model Releases

Meta observation: DeepSeek is still king of the active-parameter ratio

DGX agent

DeepSeek maintains the highest efficiency in terms of active parameters relative to total model size, outperforming competitors in the ratio of parameters actually used during inference versus total t

model-releasessebastian-raschka--x
14 May 2026
Model Releases

Microsoft starts canceling Claude Code licenses

DGX agent

Microsoft first started opening up access to Claude Code in December, inviting thousands of its own developers to use Anthropic's AI coding tool daily. It was part of an effort to get project managers

model-releasesthe-verge-ai
14 May 2026
Model Releases

MindVLA-U1: VLA Beats VA with Unified Streaming Architecture for Autonomous Driving

DGX agent

arXiv:2605.12624v1 Announce Type: cross Abstract: Autonomous driving has progressed from modular pipelines toward end-to-end unification, and Vision-Language-Action (VLA) models are a natural extensio

model-releasesarxiv-cs-cv
14 May 2026
Model Releases

Mixed neural posterior estimation for simulators with discrete and continuous parameters

DGX agent

arXiv:2605.13551v1 Announce Type: new Abstract: Neural Posterior Estimation (NPE) enables rapid parameter inference for complex simulators with intractable likelihoods. NPE trains an inference network

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

MMCL-Bench: Multimodal Context Learning from Visual Rules, Procedures, and Evidence

DGX agent

arXiv:2605.12703v1 Announce Type: cross Abstract: We introduce MMCL-Bench, a benchmark for multimodal context learning: learning task-local rules, procedures, and empirical patterns from visual or mix

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

MobiBench: Multi-Branch, Modular Benchmark for Mobile GUI Agents

DGX agent

arXiv:2512.12634v3 Announce Type: replace Abstract: Mobile GUI Agents, AI agents capable of interacting with mobile applications on behalf of users, have the potential to transform human computer inte

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Multi-Armed Sampling Problem and the End of Exploration

DGX agent

arXiv:2507.10797v2 Announce Type: replace Abstract: This paper introduces the framework of multi-armed sampling, which serves as the sampling counterpart to the optimization problem of multi-armed ban

model-releasesarxiv-cs-lg
14 May 2026
← Previous
1…316317318319320…471
Next →