AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent
84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,612 results
1 Jun 2026

Harness Updating Is Not Harness Benefit: Disentangling Evolution Capabilities in Self-Evolving LLM Agents

Model ReleasesDGX agent

arXiv:2605.30621v1 Announce Type: new Abstract: LLM agents are increasingly deployed as systems built around editable external harnesses, including prompts, skills, memories and tools, that shape task

Hedging on the Frontier: Learning New Tasks with Few Samples

Model ReleasesDGX agent

arXiv:2605.30997v1 Announce Type: cross Abstract: When a learner faces a new task with few samples, it must leverage any available side information. In practice, this often comes in the form of model

HERMES: Towards Efficient and Verifiable Mathematical Reasoning in LLMs

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2511.18760v2 Announce Type: replace Abstract: Informal mathematics has been central to modern large language model (LLM) reasoning, offering flexibility and efficient construction of arguments.

How Far Can You Grow? Characterizing the Extrapolation Frontier of Graph Generative Models for Materials Science

Model ReleasesDGX agent

arXiv:2602.09309v2 Announce Type: replace-cross Abstract: Every generative model for crystalline materials harbors a critical structure size beyond which its outputs become unreliable; we call this th

How Trustpilot built a real-time architecture for data enrichment using Gemma

Model ReleasesDGX agent

Processing millions of user reviews in real-time, under strict latency and cost constraints, is no easy task. Trustpilot has been doing exactly that with custom machine learning since long before larg

HypoSpace: A Diagnostic Benchmark for Set-Valued Hypothesis Generation under Underdetermination and Sublinear Coverage Bounds

Model ReleasesDGX agent

arXiv:2510.15614v3 Announce Type: replace Abstract: Many scientific problems are underdetermined: multiple distinct hypotheses are equally consistent with the same observations. In such settings, effe

i was today years old when i learned that claude code deletes your session traces after a month

Model ReleasesDGX agent

Claude Code automatically deletes session traces after one month, a feature that was apparently not widely known among users. This retention and deletion policy is part of Claude's data management pra

Identifiable Equivariant Networks are Layerwise Equivariant

Model ReleasesDGX agent

arXiv:2601.21645v2 Announce Type: replace Abstract: We investigate the relation between end-to-end equivariance and layerwise equivariance in deep neural networks. We prove the following: For a networ

ImmigrationQA: A Source-Grounded Dataset and Small-Model Adaptation for U.S. Immigration Law

Model ReleasesDGX agent

arXiv:2605.30589v1 Announce Type: cross Abstract: U.S. immigration law spans thousands of pages of official policy, federal regulations, and procedural guidance that change frequently and carry high s

Improving Small Language Models for Code Generation with Reinforcement Learning from Verification Feedback

Model ReleasesDGX agent

arXiv:2605.30478v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) trains language models using programmatically checkable signals such as unit-test outcomes, enab

In May, we integrated 11 new models spanning image, 3D, audio, video, and multimodal. The highlights: → Krea 2 — style-first image generatio…

Model ReleasesDGX agent

In May, we integrated 11 new models spanning image, 3D, audio, video, and multimodal. The highlights: → Krea 2 — style-first image generation, live as a Partner Node on day one. Competes on how the fr

Inconsistency-Aware Minimization: Improving Generalization with Unlabeled Data

Model ReleasesDGX agent

arXiv:2605.31324v1 Announce Type: cross Abstract: Estimating the generalization gap and developing optimization methods that improve generalization are crucial for deep learning models, for both theor

Inference-Free Multimodal Learned Sparse Retrieval for Production-Scale Visual Document Search

Model ReleasesDGX agent

arXiv:2605.30917v1 Announce Type: cross Abstract: As large-scale visual-document corpora such as arXiv papers and enterprise PDFs continue to grow, visual-document retrieval has gained increasing atte

Intel touts 130-plus edge design wins for Series 3 and launches OpenVINO Physical AI framework

Model ReleasesDGX agent

Intel Corp. today announced that more than 130 design engagements for its Series 3 processor family for edge artificial intelligence and edge computing designs and also unveiled a new open-source fram

👏👏 Introducing Qwen3.7-Plus — a multimodal agent model that unifies vision and language into one versatile agent foundation. ✅ Multimodal …

Model ReleasesDGX agent

👏👏 Introducing Qwen3.7-Plus — a multimodal agent model that unifies vision and language into one versatile agent foundation. ✅ Multimodal interactive hybrid agent: unified GUI & CLI operation across v

Introducing the GKE standby buffer: Improve node startup times without blowing your budget

Model ReleasesDGX agent

Application owners and platform engineers have long faced a difficult choice: spend excessively by over-provisioning to guarantee quick startups, or minimize costs but endure slow cold starts. We are

Inversion-Free Natural Gradient Descent on Riemannian Manifolds

Model ReleasesDGX agent

arXiv:2604.02969v2 Announce Type: replace-cross Abstract: The natural gradient method is a central tool for statistical optimisation, but its broader application is hindered by the assumption of a Euc

just a small zoom out on the vibe shift: in Feb 2025 @soumithchintala was talking about his dream of personal, local, private agents, most p…

Model ReleasesDGX agent

just a small zoom out on the vibe shift: in Feb 2025 @soumithchintala was talking about his dream of personal, local, private agents, most people didn't believe him. it's June 2026 and @pewdiepie has

KernelCraft: Benchmarking for Agentic Close-to-Metal Kernel Generation on Emerging Hardware

Model ReleasesDGX agent

arXiv:2603.08721v2 Announce Type: replace-cross Abstract: New AI accelerators with novel instruction set architectures (ISAs) often require developers to manually craft low-level kernels, a time-consu

Knowledge Boundary Probing and Demand-Guided Intervention for LLM-Based Power System Code Generation

Model ReleasesDGX agent

arXiv:2605.31478v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used to automate power-system analysis, but many utilities and energy-research labs require on-premise s

LangMap: A Human-Verified Benchmark for Hierarchical Open-Vocabulary Goal Navigation

Model ReleasesDGX agent

arXiv:2602.02220v2 Announce Type: replace Abstract: Language-conditioned goal navigation (LGN) requires agents to locate user-specified targets without step-by-step guidance. However, existing benchma

Language Models Learn Constructional Semantics, Not To Mention Syntax: Investigating LM Understanding of Paired-Focus Constructions

Model ReleasesDGX agent

arXiv:2605.31586v1 Announce Type: cross Abstract: Grasping the semantics of rare constructions (form-meaning pairings) has been shown to be a challenging problem that has currently only been solved by

Last week we revamped Liteparse to be the fastest PDF parser out there ⚡️ An underrated part of liteparse is it doesn't just give you text. …

Model ReleasesDGX agent

Last week we revamped Liteparse to be the fastest PDF parser out there ⚡️ An underrated part of liteparse is it doesn't just give you text. It gives you bounding boxes that a coding agent can use to p

Learning-Based Navigation for Indoor Mobile Robots

Model ReleasesDGX agent

arXiv:2605.30468v1 Announce Type: new Abstract: This paper presents a learning-based navigation framework for indoor mobile robots. The proposed method combines a supervised neural global planner, tra

Learning Multi-Agent Coordination via Sheaf-ADMM

Model ReleasesDGX agent

arXiv:2605.31005v1 Announce Type: new Abstract: We present a differentiable optimization framework for multi-agent coordination. An input is decomposed into overlapping local views, each processed by

Learning Randomized Reductions

Model ReleasesDGX agent

arXiv:2412.18134v4 Announce Type: replace Abstract: Randomized self-reductions (RSRs) express f(x) using f evaluated at random correlated points, enabling self-correcting programs, instance-hiding pro

Learning Whom to Trust: Market-Feedback Adaptive Retrieval for Frozen LLMs in Event-Driven Financial RAG

Model ReleasesDGX agent

arXiv:2605.31201v1 Announce Type: new Abstract: Financial retrieval-augmented generation (RAG) systems typically rank evidence by textual relevance, but in financial markets the useful evidence source

LegSegNet: A Public Deep Learning System for Lower Extremity CT Tissue Segmentation and Quantification

Model ReleasesDGX agent

arXiv:2605.30829v1 Announce Type: new Abstract: Lower extremity computed tomography (CT) contains clinically relevant information for body composition analysis, sarcopenia assessment, and musculoskele

Linear Ordering Problem: Time for a Change

Model ReleasesDGX agent

arXiv:2605.31051v1 Announce Type: cross Abstract: The Linear Ordering Problem (LOP) is a fundamental combinatorial optimization problem with important applications in areas such as economics, social c

LLM Bias Evaluation: Gender, Racial, and Age Disparities in Occupational and Crime Scenarios

Model ReleasesDGX agent

arXiv:2409.14583v4 Announce Type: replace Abstract: LLM bias evaluation is critical as large language models (LLMs) increasingly influence high-stakes decisions. This paper provides a comprehensive as

LLMs Without Deep Neural Networks: New Architecture, Benefits and Case Study

Model ReleasesDGX agent

arXiv:2605.30385v1 Announce Type: cross Abstract: The purpose of this article is to provide validation to my deep neural network alternative in the context of LLMs. Very recently, there has been a sig

LongDS-Bench: On the Failure of Long-Horizon Agentic Data Analysis

Model ReleasesDGX agent

arXiv:2605.30434v1 Announce Type: cross Abstract: Real-world data analysis is inherently iterative, yet existing benchmarks mostly evaluate isolated or short interactive tasks, leaving agents' ability

Lots of companies are in the 'encourage AI adoption' phase, whether teaching them ChatGPT/Claude or (sigh) tokenmaxxing. That dodges the har…

Model ReleasesDGX agent

Lots of companies are in the 'encourage AI adoption' phase, whether teaching them ChatGPT/Claude or (sigh) tokenmaxxing. That dodges the harder problems of firm leadership: What do you want people to

MAAT: Multi-phase Adapter-Aware Targeted Unlearning

Model ReleasesDGX agent

arXiv:2605.30514v1 Announce Type: cross Abstract: Machine unlearning evaluation is structurally skewed: Why-type questions, which probe causal and relational knowledge, comprise less than 0.06% of Cou

MADS: Model-Aware Diverse Core Set Selection for Instruction Tuning

Model ReleasesDGX agent

arXiv:2605.30857v1 Announce Type: new Abstract: Instruction fine-tuning is employed to enhance the instruction-following ability of large language models (LLMs). As the amount of instruction fine-tuni

Many research labs only consider inference efficiency after the fact. Step 3.7 Flash is a 196B MoE model, and built for inference from the s…

Model ReleasesDGX agent

Many research labs only consider inference efficiency after the fact. Step 3.7 Flash is a 196B MoE model, and built for inference from the start by @StepFun_ai. Multi-Matrix Factorization Attention (M

MAVEN: Improving Generalization in Agentic Tool Calling

Model ReleasesDGX agent

arXiv:2605.30738v1 Announce Type: new Abstract: Generalization across agentic tool-calling environments remains a central challenge for reliable agentic reasoning systems. Although large language mode

MedFact: Benchmarking the Fact-Checking Capabilities of Large Language Models on Chinese Medical Texts

Model ReleasesDGX agent

arXiv:2509.12440v3 Announce Type: replace-cross Abstract: Deploying Large Language Models (LLMs) in medical applications requires fact-checking capabilities to ensure patient safety and regulatory com

Mellum2 Technical Report

Model ReleasesDGX agent

arXiv:2605.31268v1 Announce Type: new Abstract: We present Mellum 2, an open-weight 12B-parameter Mixture-of-Experts (MoE) language model with 2.5B active parameters per token. Mellum 2 is a general-p

Memory-Bound but Not Bandwidth-Limited: The Physical AI Inference Gap in Batch-1 LLM Decode

Model ReleasesDGX agent

arXiv:2605.30571v1 Announce Type: cross Abstract: Physical AI systems, including robots, autonomous vehicles, embodied agents and edge copilots, often run a different inference workload from cloud LLM

Memory by Design: Probabilistic Sequence Layers

Model ReleasesDGX agent

arXiv:2605.31163v1 Announce Type: cross Abstract: We introduce the design-model framework: a way to derive efficient recurrent sequence maps from explicit assumptions about memory. A design model writ

Merge launches Agent Handler for Employees as an IT gatekeeper for workplace AI agents

Model ReleasesDGX agent

Merge API Inc., a platform provider delivering connective infrastructure for artificial intelligence to business data and tools, launched Agent Handler for Employees, easing the strain on information

MIMO: Multilingual Information Retrieval via Monolingual Objectives

Model ReleasesDGX agent

arXiv:2605.31171v1 Announce Type: cross Abstract: Multilingual Information Retrieval (MLIR) reflects real-world search environments in which queries and relevant documents may appear in different lang

MineExplorer: Evaluating Open-World Exploration of MLLM Agents in Minecraft

Model ReleasesDGX agent

arXiv:2605.30931v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have shown strong capabilities in perception, reasoning, and action generation. However, their ability to susta

.@MiniMax_AI M3 model is available on Ollama's Cloud! In partnership with MiniMax, the M3 model on Ollama's Cloud is US-based with zero data…

Model ReleasesDGX agent

.@MiniMax_AI M3 model is available on Ollama's Cloud! In partnership with MiniMax, the M3 model on Ollama's Cloud is US-based with zero data retention. Try M3 on coding and agentic tasks: Claude Code:

MLIPilot: LLM-Driven Auto-Research for Machine-Learned Interatomic Potentials

Model ReleasesDGX agent

arXiv:2605.30889v1 Announce Type: cross Abstract: Constructing production-quality machine-learned interatomic potentials (MLIPs) requires balancing accuracy, dynamical stability, and computational thr

MosaicLeaks:Privacy Risks in Querying-in-the-Open for Deep Research Agents

Model ReleasesDGX agent

arXiv:2605.30727v1 Announce Type: new Abstract: Deep research agents increasingly combine private local documents with external tools like web retrieval, creating a privacy risk: an agent's external q

MultiPriv: Benchmarking Individual-Level Privacy Reasoning in Vision-Language Models

Model ReleasesDGX agent

arXiv:2511.16940v3 Announce Type: replace Abstract: Modern Vision-Language Models (VLMs) pose significant individual-level privacy risks by linking fragmented multimodal data to identifiable individua

Nemotron 3 Ultra: Frontier smart. 5X faster. 30% cheaper. 💚💚💚

Model ReleasesDGX agent

Nemotron 3 Ultra is NVIDIA's latest language model featuring significant improvements in speed (5X faster) and cost efficiency (30% cheaper) compared to previous versions, positioning it as a frontier

NeUQI: Near-Optimal Uniform Quantization Parameter Initialization for Low-Bit LLMs

Model ReleasesDGX agent

arXiv:2505.17595v4 Announce Type: replace-cross Abstract: Large language models (LLMs) achieve impressive performance across domains but face significant challenges when deployed on consumer-grade GPU

Neuro-symbolic Syntactic Parsing: Shaping a Neural Network with the CYK Algorithm

Model ReleasesDGX agent

arXiv:2605.31421v1 Announce Type: cross Abstract: In this paper, we show the possibility of a direct injection of algorithms into neural network architecture. We focus on a complex algorithm, that is,

NGDBench: Towards Neural Graph Data Management

Model ReleasesDGX agent

arXiv:2603.05529v2 Announce Type: replace-cross Abstract: Data critical to real-world decision-making is increasingly found within organizations. Such data is heterogeneous, constantly evolving, and o

Not All Synthetic Data Is Yours to Learn From

Model ReleasesDGX agent

arXiv:2605.31126v1 Announce Type: cross Abstract: Can a language model improve from plain text sampled from itself, with no prompts, no teacher, no verifier, and no reward model? Yes, but only when th

nuReasoning: A Reasoning-Centric Dataset and Benchmark for Long-Tail Autonomous Driving

Model ReleasesDGX agent

arXiv:2605.31572v1 Announce Type: new Abstract: Reasoning is essential for autonomous driving (AD) in long-tail scenarios, where vehicles must apply commonsense knowledge, understand spatial relations

Nvidia launches Nemotron 3 Ultra, a 550B-parameter MoE open model; Artificial Analysis: it's the smartest open US model but trails the Chinese model Kimi K2.6 (Maximilian Schreiner/The Decoder)

Model ReleasesDGX agent

Maximilian Schreiner / The Decoder: Nvidia launches Nemotron 3 Ultra, a 550B-parameter MoE open model; Artificial Analysis: it's the smartest open US model but trails the Chinese model Kimi K2.6 — It

OBCache: Optimal Brain KV Cache Pruning for Efficient Long-Context LLM Inference

Model ReleasesDGX agent

arXiv:2510.07651v2 Announce Type: replace-cross Abstract: Large language models (LLMs) with extended context windows enable powerful applications but impose significant memory overhead, as caching all

Omni-Supervised Motion Editing: Balancing Change and Invariance through Positive-Negative Learning

Model ReleasesDGX agent

arXiv:2605.30969v1 Announce Type: new Abstract: Text-based human motion editing aims to modify existing motion sequences according to natural language instructions while maintaining the consistency of

On-Device Generative AI for GDPR-Compliant Visual Monitoring: Natural Language Alerts from Local Object Detection

Model ReleasesDGX agent

arXiv:2605.30544v1 Announce Type: new Abstract: Visual monitoring systems that rely on cloud-based AI inference expose raw image data to external services, creating fundamental tensions with the data-

On the Robustness of Multilingual Text Embedding Rankings Across Learning Tasks, Languages, and Benchmark Datasets

Model ReleasesDGX agent

arXiv:2605.31142v1 Announce Type: cross Abstract: Large-scale multilingual text embedding models play crucial role in both research and industry, yet their behavior in language-specific, multi-task se

One of the big reasons for the current lack of patriotism and pride in our nation’s history is that about 40 years ago our most prominent st…

Model ReleasesDGX agent

One of the big reasons for the current lack of patriotism and pride in our nation’s history is that about 40 years ago our most prominent storytellers in Hollywood just basically stopped telling stori

← Previous
1…188189190191192…377
Next →