AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,566 results
14 May 2026

GeomHair: Reconstruction of Hair Strands from Colorless 3D Scans

Model ReleasesDGX agent

arXiv:2505.05376v3 Announce Type: replace Abstract: We propose a novel method that reconstructs hair strands directly from colorless 3D scans by leveraging multi-modal hair orientation extraction. Hai

GHGbench: A Unified Multi-Entity, Multi-Task Benchmark for Carbon Emission Prediction

Model ReleasesDGX agent

arXiv:2605.13743v1 Announce Type: new Abstract: Open datasets and benchmarks for entity-level carbon-emission prediction remain fragmented across access, scale, granularity, and evaluation. We introdu

GraphIP-Bench: How Hard Is It to Steal a Graph Neural Network, and Can We Stop It?

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.12827v1 Announce Type: cross Abstract: Graph neural networks (GNNs) deployed as cloud services can be stolen through model-extraction attacks, which train a surrogate from query responses t

Grid-Orch: An LLM-Powered Orchestrator for Distribution Grid Simulation and Analytics

Model ReleasesDGX agent

arXiv:2605.12728v1 Announce Type: cross Abstract: The power distribution engineering workforce faces a projected shortage of up to 1.5 million engineers by 2030, creating urgent demand for more access

GuardMarkGS: Unified Ownership Tracing and Edit Deterrence for 3D Gaussian Splatting

Model ReleasesDGX agent

arXiv:2605.12919v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) is becoming a practical representation for novel view synthesis, but its growing adoption, together with rapid advances in

Guide, Think, Act: Interactive Embodied Reasoning in Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2605.13632v1 Announce Type: cross Abstract: In this paper, we propose GTA-VLA(Guide, Think, Act), an interactive Vision-Language-Action (VLA) framework that enables spatially steerable embodied

GUIGuard-Bench: Toward a General Evaluation for Privacy-Preserving GUI Agents

Model ReleasesDGX agent

arXiv:2601.18842v3 Announce Type: replace-cross Abstract: As GUI agents increasingly rely on screenshots to perceive and operate digital environments, they may inadvertently expose sensitive informati

HADAR-Based Thermal Infrared Hyperspectral Image Restoration

Model ReleasesDGX agent

arXiv:2605.13664v1 Announce Type: new Abstract: Thermal-infrared (TIR) hyperspectral imagery (HSI) provides critical scene information for various applications. However, its practical utility is sever

HCSG: Human-Centric Semantic-Geometric Reasoning for Vision-Language Navigation

Model ReleasesDGX agent

arXiv:2605.13321v1 Announce Type: new Abstract: VLN has achieved remarkable progress by scaling data and model capacity. However, the assumption of a static environment breaks down in real-world indoo

Hessian Matching for Machine-Learned Coarse-Grained Molecular Dynamics

Model ReleasesDGX agent

arXiv:2605.12823v1 Announce Type: new Abstract: Coarse-grained (CG) molecular dynamics enables simulations of atomic systems such as biomolecules at timescales inaccessible to all-atom (AA) methods, b

Hierarchical Attacks for Multi-Modal Multi-Agent Reasoning

Model ReleasesDGX agent

arXiv:2605.13213v1 Announce Type: new Abstract: Multi-modal multi-agent systems (MM-MAS) have gained increasing attention for their capacity to enable complex reasoning and coordination across diverse

Hierarchical Transformer Preconditioning for Interactive Physics Simulation

Model ReleasesDGX agent

arXiv:2605.13343v1 Announce Type: cross Abstract: Neural preconditioners for real-time physics simulation offer promising data-driven priors, but they often fail to capture long-range couplings effici

High-Dimensional Analysis of Bootstrap Ensemble Classifiers

Model ReleasesDGX agent

arXiv:2505.14587v2 Announce Type: replace-cross Abstract: Bootstrap methods have long been the cornerstone of ensemble learning in machine learning. This paper presents a theoretical analysis of boots

High-Rate Quantized Matrix Multiplication II

Model ReleasesDGX agent

arXiv:2605.13768v1 Announce Type: cross Abstract: This is the second part of the work investigating quantized matrix multiplication (MatMul). In part I we considered the case of calibration-free quant

HLS-Seek: QoR-Aware Code Generation for High-Level Synthesis via Proxy Comparative Reward Reinforcement Learning

Model ReleasesDGX agent

arXiv:2605.13536v1 Announce Type: cross Abstract: High-Level Synthesis (HLS) compiles algorithmic C/C++ descriptions into hardware, with Quality of Results (QoR) -- latency and resource utilization --

(How) Do Large Language Models Understand High-Level Message Sequence Charts?

Model ReleasesDGX agent

arXiv:2605.13773v1 Announce Type: cross Abstract: Large Language Models (LLMs) are being employed widely to automate tasks across the software development life-cycle. It is, however, unclear whether t

How to Interpret Agent Behavior

Model ReleasesDGX agent

arXiv:2605.13625v1 Announce Type: new Abstract: Autonomous agents such as Claude Code and Codex now operate for hours or even days. Understanding their runtime behavior has become critical for downstr

How Well Do Large-Scale Chemical Language Models Transfer to Downstream Tasks?

Model ReleasesDGX agent

arXiv:2602.11618v4 Announce Type: replace Abstract: Chemical Language Models (CLMs) pre-trained on large scale molecular data are widely used for molecular property prediction. However, the common bel

I was going to say LangChain Labs is the most exciting thing we’ve announced recently, but with so many launches it’s hard to justify… but s…

Model ReleasesDGX agent

I was going to say LangChain Labs is the most exciting thing we’ve announced recently, but with so many launches it’s hard to justify… but still incredibly thrilled to see this get started!! Checkout

Identifying the nonlinear string dynamics with port-Hamiltonian neural networks

Model ReleasesDGX agent

arXiv:2605.12785v1 Announce Type: new Abstract: Hybrid machine learning combines physical knowledge with data-driven models to enhance interpretability and performance. In this context, Port-Hamiltoni

ImageAttributionBench: How Far Are We from Generalizable Attribution?

Model ReleasesDGX agent

arXiv:2605.12967v1 Announce Type: new Abstract: The rapid advancement of generative AI has enabled the creation of highly realistic and diverse synthetic images, posing critical challenges for image p

Implicit Behavioral Decoding from Next-Step Spike Forecasts at Population Scale

Model ReleasesDGX agent

arXiv:2605.12999v1 Announce Type: cross Abstract: Closed-loop brain-computer interfaces often require both a forecast of upcoming neural population activity and a readout of the animal's behavioral st

Imposing Boundary Conditions on Neural Operators via Learned Function Extensions

Model ReleasesDGX agent

arXiv:2602.04923v2 Announce Type: replace Abstract: Neural operators have emerged as powerful surrogates for the solution of partial differential equations (PDEs), yet their ability to handle general,

IndicMedDialog: A Parallel Multi-Turn Medical Dialogue Dataset for Accessible Healthcare in Indic Languages

Model ReleasesDGX agent

arXiv:2605.13292v1 Announce Type: cross Abstract: Most existing medical dialogue systems operate in a single-turn question--answering paradigm or rely on template-based datasets, limiting conversation

Inducing Overthink: Hierarchical Genetic Algorithm-based DoS Attack on Black-Box Large Language Reasoning Models

Model ReleasesDGX agent

arXiv:2605.13338v1 Announce Type: cross Abstract: Large Reasoning Models (LRMs) are increasingly integrated into systems requiring reliable multi-step inference, yet this growing dependence exposes ne

Inference-Time Machine Unlearning via Gated Activation Redirection

Model ReleasesDGX agent

arXiv:2605.12765v1 Announce Type: new Abstract: Large Language Models memorize vast amounts of training data, raising concerns regarding privacy, copyright infringement, and safety. Machine unlearning

ISOMORPH: A Supply Chain Digital Twin for Simulation, Dataset Generation, and Forecasting Benchmarks

Model ReleasesDGX agent

arXiv:2605.12768v1 Announce Type: cross Abstract: Open time-series forecasting (TSF) benchmarks cover retail, energy, weather, and traffic, but supply-chain logistics remains underserved. We introduce

KamonBench: A Grammar-Based Dataset for Evaluating Compositional Factor Recovery in Vision-Language Models

Model ReleasesDGX agent

arXiv:2605.13322v1 Announce Type: new Abstract: Kamon (family crests) are an important part of Japanese culture and a natural test case for compositional visual recognition: each crest combines a smal

Kernel-based guarantees for nonlinear parametric models in Bayesian optimization

Model ReleasesDGX agent

arXiv:2605.13160v1 Announce Type: cross Abstract: Modern Bayesian optimization and adaptive sampling methods increasingly rely on nonlinear parametric models, yet theoretical guarantees for such model

Kimi K2.6 and DeepSeek V4 Pro are now GA on @FireworksAI_HQ on Foundry + PTU support in the US Data Zone—predictable performance, data resid…

Model ReleasesDGX agent

Kimi K2.6 and DeepSeek V4 Pro are now GA on @FireworksAI_HQ on Foundry + PTU support in the US Data Zone—predictable performance, data residency, enterprise SLAs. Frontier open models. Azure controls.

Kimi K2.6 is now open-weight #1 on Finance Agent Benchmark V2.

Model ReleasesDGX agent

Kimi K2.6 is now open-weight #1 on Finance Agent Benchmark V2. Can AI do the job of a financial analyst? We just released V2 of our Finance Agent Benchmark and tested the frontier models. The results

Kiwi-Edit: Versatile Video Editing via Instruction and Reference Guidance

Model ReleasesDGX agent

arXiv:2603.02175v4 Announce Type: replace-cross Abstract: Instruction-based video editing has witnessed rapid progress, yet current methods often struggle with precise visual control, as natural langu

LangChain 在 Interrupt 大会上发布了底层数据库 SmithDB 和自动化排障引擎 LangSmith Engine。 Agent 运行会产生海量 trace(执行轨迹),把旧数据库撑到了瓶颈。新底座 SmithDB 放弃了本地磁盘,全面转向对象存储,将核心查询…

Model ReleasesDGX agent

LangChain 在 Interrupt 大会上发布了底层数据库 SmithDB 和自动化排障引擎 LangSmith Engine。 Agent 运行会产生海量 trace(执行轨迹),把旧数据库撑到了瓶颈。新底座 SmithDB 放弃了本地磁盘,全面转向对象存储,将核心查询速度拉高了 15 倍。 底座换新后,LangSmith Engine 顺势接管了查 Bug 的体力活。它在后台持续监控生

Language Model Goal Selection Differs from Humans' in a Self-Directed Learning Task

Model ReleasesDGX agent

arXiv:2603.03295v2 Announce Type: replace-cross Abstract: Whether in agentic workflows, social studies, or chat settings, large language models (LLMs) are increasingly being asked to replace humans in

Large Language Models Lack Temporal Awareness of Medical Knowledge

Model ReleasesDGX agent

arXiv:2605.13045v1 Announce Type: new Abstract: The existing methods for evaluating the medical knowledge of Large Language Models (LLMs) are largely based on atemporal examination-style benchmarks, w

LeanSearch v2: Global Premise Retrieval for Lean 4 Theorem Proving

Model ReleasesDGX agent

arXiv:2605.13137v1 Announce Type: cross Abstract: Proving theorems in Lean 4 often requires identifying a scattered set of library lemmas whose joint use enables a concise proof -- a task we call glob

Learning a Continue-Thinking Token for Enhanced Test-Time Scaling

Model ReleasesDGX agent

arXiv:2506.11274v2 Announce Type: replace-cross Abstract: Test-time scaling has emerged as an effective approach for improving language model performance by utilizing additional compute at inference t

Learning Responsibility-Attributed Adversarial Scenarios for Testing Autonomous Vehicles

Model ReleasesDGX agent

arXiv:2605.13751v1 Announce Type: new Abstract: Establishing trustworthy safety assurance for autonomous driving systems (ADSs) requires evidence that failures arise from avoidable system deficiencies

LENS: Multi-level Evaluation of Multimodal Reasoning with Large Language Models

Model ReleasesDGX agent

arXiv:2505.15616v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have achieved significant advances in integrating visual and linguistic information, yet their ability to r

Lifelong Learning in Vision-Language Models: Enhanced EWC with Cross-Modal Knowledge Retention

Model ReleasesDGX agent

arXiv:2605.12789v1 Announce Type: new Abstract: Large language-vision models (LVLMs) such as CLIP, Flamingo, and BLIP have revolutionized AI by enabling understanding across textual and visual modalit

LIFT: Last-Mile Fine-Tuning for Table Explicitation

Model ReleasesDGX agent

arXiv:2605.13424v1 Announce Type: new Abstract: We propose last-mile fine-tuning, or Lift, a pipeline in which a pre-trained large language model extracts an initial table from unstructured clipboard

Like any AI dev not employed by a closed lab, we share the ambition that at least 10x more should be capable of training frontier models in …

Model ReleasesDGX agent

Like any AI dev not employed by a closed lab, we share the ambition that at least 10x more should be capable of training frontier models in 2026. And like we all know, some 10x leaps can't survive a c

LLMs as annotators of credibility assessment in Danish asylum decisions: evaluating classification performance and errors beyond aggregated metrics

Model ReleasesDGX agent

arXiv:2605.13412v1 Announce Type: cross Abstract: Off-the-shelf large language models (LLMs) are increasingly used to automate text annotation, yet their effectiveness remains underexplored for underr

Looks like @NVIDIAAI has released official getting started guide for hermes on DGX Spark! Great way for new folks to get started and believe…

Model ReleasesDGX agent

NVIDIA AI has released an official getting started guide for Hermes on DGX Spark, providing new users with foundational resources and instructions for deploying or working with Hermes on NVIDIA's DGX

LoRA-Mixer: Coordinate Modular LoRA Experts Through Serial Attention Routing

Model ReleasesDGX agent

arXiv:2507.00029v2 Announce Type: replace-cross Abstract: Recent attempts to combine low-rank adaptation (LoRA) with mixture-of-experts (MoE) for multi-task adaptation of Large Language Models (LLMs)

Low-Rank Adapters Initialization via Gradient Surgery for Continual Learning

Model ReleasesDGX agent

arXiv:2605.12752v1 Announce Type: new Abstract: LoRA is widely adopted for continual fine-tuning of Large Language Models due to its parameter efficiency, modularity across tasks, and compatibility wi

LY Corp launches a bid with Bain Capital to buy Kakaku.com, which operates restaurant review and booking site Tabelog, for 4B, challenging EQT's 3.75B offer (Nikkei Asia)

Model ReleasesDGX agent

Nikkei Asia: LY Corp launches a bid with Bain Capital to buy Kakaku.com, which operates restaurant review and booking site Tabelog, for 4B, challenging EQT's 3.75B offer — TOKYO — The operator of the

Many-Shot CoT-ICL: Making In-Context Learning Truly Learn

Model ReleasesDGX agent

arXiv:2605.13511v1 Announce Type: cross Abstract: In-context learning (ICL) adapts large language models (LLMs) to new tasks by conditioning on demonstrations in the prompt without parameter updates.

Mechanistic Evidence for Spectral Structures in Prior-Data Fitted Networks

Model ReleasesDGX agent

arXiv:2601.21731v2 Announce Type: replace Abstract: Prior-Data Fitted Networks (PFNs) enable amortized Bayesian inference in a single forward pass, yet their internal representations remain opaque. It

MedCore: Boundary-Preserving Medical Core Pruning for MedSAM

Model ReleasesDGX agent

arXiv:2605.13688v1 Announce Type: new Abstract: Medical segmentation foundation models such as SAM and MedSAM provide strong prompt-driven segmentation, but their image encoders are still too large fo

MedOpenClaw and MedFlowBench: Auditing Medical Agents in Full-Study Workflows

Model ReleasesDGX agent

arXiv:2603.24649v2 Announce Type: replace Abstract: Medical imaging benchmarks often evaluate VLMs on pre-selected 2D images, slices, crops, or patches, making evaluation closer to visual recognition.

Meet Kimi Web Bridge - Kimi's browser extension. Agent can now interact with websites like a human: search, scroll, click, type and complete…

Model ReleasesDGX agent

Meet Kimi Web Bridge - Kimi's browser extension. Agent can now interact with websites like a human: search, scroll, click, type and complete tasks. Supports Kimi Code CLI, Claude Code, Cursor, Codex,

Meta observation: DeepSeek is still king of the active-parameter ratio

Model ReleasesDGX agent

DeepSeek maintains the highest efficiency in terms of active parameters relative to total model size, outperforming competitors in the ratio of parameters actually used during inference versus total t

Microsoft starts canceling Claude Code licenses

Model ReleasesDGX agent

Microsoft first started opening up access to Claude Code in December, inviting thousands of its own developers to use Anthropic's AI coding tool daily. It was part of an effort to get project managers

MindVLA-U1: VLA Beats VA with Unified Streaming Architecture for Autonomous Driving

Model ReleasesDGX agent

arXiv:2605.12624v1 Announce Type: cross Abstract: Autonomous driving has progressed from modular pipelines toward end-to-end unification, and Vision-Language-Action (VLA) models are a natural extensio

Mixed neural posterior estimation for simulators with discrete and continuous parameters

Model ReleasesDGX agent

arXiv:2605.13551v1 Announce Type: new Abstract: Neural Posterior Estimation (NPE) enables rapid parameter inference for complex simulators with intractable likelihoods. NPE trains an inference network

MMCL-Bench: Multimodal Context Learning from Visual Rules, Procedures, and Evidence

Model ReleasesDGX agent

arXiv:2605.12703v1 Announce Type: cross Abstract: We introduce MMCL-Bench, a benchmark for multimodal context learning: learning task-local rules, procedures, and empirical patterns from visual or mix

MobiBench: Multi-Branch, Modular Benchmark for Mobile GUI Agents

Model ReleasesDGX agent

arXiv:2512.12634v3 Announce Type: replace Abstract: Mobile GUI Agents, AI agents capable of interacting with mobile applications on behalf of users, have the potential to transform human computer inte

Multi-Armed Sampling Problem and the End of Exploration

Model ReleasesDGX agent

arXiv:2507.10797v2 Announce Type: replace Abstract: This paper introduces the framework of multi-armed sampling, which serves as the sampling counterpart to the optimization problem of multi-armed ban

N-vium: Mixture-of-Exits Transformer for Accelerated Exact Generation

Model ReleasesDGX agent

arXiv:2605.13190v1 Announce Type: cross Abstract: Improving the inference efficiency of autoregressive transformers typically means reducing FLOPs per token, usually through approximations that degrad

← Previous
1…251252253254255…377
Next →