AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
All
85,188Total entries
1Added by human
85,187Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,797 results
Model Releases

Go-UT-Bench: A Fine-Tuning Dataset for LLM-Based Unit Test Generation in Go

DGX agent

arXiv:2511.10868v2 Announce Type: replace Abstract: Training data imbalance poses a major challenge for code LLMs. Most available data heavily over represents raw opensource code while underrepresenti

model-releasesarxiv-cs-lg
1 Jun 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Goldfish: Monolingual Language Models for 350 Languages

DGX agent

arXiv:2408.10441v3 Announce Type: replace Abstract: For many low-resource languages, the only available language models are large multilingual models trained on many languages simultaneously. Despite

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

Graph-Conditioned Mixture of Graph Neural Network Experts for Traffic Forecasting

DGX agent

arXiv:2605.30486v1 Announce Type: cross Abstract: Spatio-temporal forecasting on sensor graphs is commonly tackled with a single backbone architecture applied uniformly across all nodes, although grap

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

GraphARC: A Comprehensive Benchmark for Graph-Based Abstract Reasoning

DGX agent

arXiv:2605.31031v1 Announce Type: new Abstract: Relational reasoning lies at the heart of intelligence, but existing benchmarks are typically confined to formats such as grids or text. We introduce Gr

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

GUI-C^2: Coarse-to-Fine GUI Grounding via Difficulty-Aware Reinforcement Learning

DGX agent

arXiv:2605.30884v1 Announce Type: new Abstract: Existing agentic reinforcement learning methods for GUI grounding have limitations at two levels. At the data level, current approaches typically treat

model-releasesarxiv-cs-cv
1 Jun 2026
Model Releases

Harness Updating Is Not Harness Benefit: Disentangling Evolution Capabilities in Self-Evolving LLM Agents

DGX agent

arXiv:2605.30621v1 Announce Type: new Abstract: LLM agents are increasingly deployed as systems built around editable external harnesses, including prompts, skills, memories and tools, that shape task

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Hedging on the Frontier: Learning New Tasks with Few Samples

DGX agent

arXiv:2605.30997v1 Announce Type: cross Abstract: When a learner faces a new task with few samples, it must leverage any available side information. In practice, this often comes in the form of model

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

HERMES: Towards Efficient and Verifiable Mathematical Reasoning in LLMs

DGX agent

arXiv:2511.18760v2 Announce Type: replace Abstract: Informal mathematics has been central to modern large language model (LLM) reasoning, offering flexibility and efficient construction of arguments.

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

How Far Can You Grow? Characterizing the Extrapolation Frontier of Graph Generative Models for Materials Science

DGX agent

arXiv:2602.09309v2 Announce Type: replace-cross Abstract: Every generative model for crystalline materials harbors a critical structure size beyond which its outputs become unreliable; we call this th

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

How Trustpilot built a real-time architecture for data enrichment using Gemma

DGX agent

Processing millions of user reviews in real-time, under strict latency and cost constraints, is no easy task. Trustpilot has been doing exactly that with custom machine learning since long before larg

model-releasesgoogle-cloud-ai
1 Jun 2026
Model Releases

HypoSpace: A Diagnostic Benchmark for Set-Valued Hypothesis Generation under Underdetermination and Sublinear Coverage Bounds

DGX agent

arXiv:2510.15614v3 Announce Type: replace Abstract: Many scientific problems are underdetermined: multiple distinct hypotheses are equally consistent with the same observations. In such settings, effe

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

i was today years old when i learned that claude code deletes your session traces after a month

DGX agent

Claude Code automatically deletes session traces after one month, a feature that was apparently not widely known among users. This retention and deletion policy is part of Claude's data management pra

model-releasesclem-delangue--x
1 Jun 2026
Model Releases

Identifiable Equivariant Networks are Layerwise Equivariant

DGX agent

arXiv:2601.21645v2 Announce Type: replace Abstract: We investigate the relation between end-to-end equivariance and layerwise equivariance in deep neural networks. We prove the following: For a networ

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

ImmigrationQA: A Source-Grounded Dataset and Small-Model Adaptation for U.S. Immigration Law

DGX agent

arXiv:2605.30589v1 Announce Type: cross Abstract: U.S. immigration law spans thousands of pages of official policy, federal regulations, and procedural guidance that change frequently and carry high s

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Improving Small Language Models for Code Generation with Reinforcement Learning from Verification Feedback

DGX agent

arXiv:2605.30478v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) trains language models using programmatically checkable signals such as unit-test outcomes, enab

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

In May, we integrated 11 new models spanning image, 3D, audio, video, and multimodal. The highlights: → Krea 2 — style-first image generatio…

DGX agent

In May, we integrated 11 new models spanning image, 3D, audio, video, and multimodal. The highlights: → Krea 2 — style-first image generation, live as a Partner Node on day one. Competes on how the fr

model-releasescomfyui--x
1 Jun 2026
Model Releases

Inconsistency-Aware Minimization: Improving Generalization with Unlabeled Data

DGX agent

arXiv:2605.31324v1 Announce Type: cross Abstract: Estimating the generalization gap and developing optimization methods that improve generalization are crucial for deep learning models, for both theor

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Inference-Free Multimodal Learned Sparse Retrieval for Production-Scale Visual Document Search

DGX agent

arXiv:2605.30917v1 Announce Type: cross Abstract: As large-scale visual-document corpora such as arXiv papers and enterprise PDFs continue to grow, visual-document retrieval has gained increasing atte

model-releasesarxiv-cs-cv
1 Jun 2026
Model Releases

Intel touts 130-plus edge design wins for Series 3 and launches OpenVINO Physical AI framework

DGX agent

Intel Corp. today announced that more than 130 design engagements for its Series 3 processor family for edge artificial intelligence and edge computing designs and also unveiled a new open-source fram

model-releasessiliconangle
1 Jun 2026
Model Releases

👏👏 Introducing Qwen3.7-Plus — a multimodal agent model that unifies vision and language into one versatile agent foundation. ✅ Multimodal …

DGX agent

👏👏 Introducing Qwen3.7-Plus — a multimodal agent model that unifies vision and language into one versatile agent foundation. ✅ Multimodal interactive hybrid agent: unified GUI & CLI operation across v

model-releasesqwen--x
1 Jun 2026
Model Releases

Introducing the GKE standby buffer: Improve node startup times without blowing your budget

DGX agent

Application owners and platform engineers have long faced a difficult choice: spend excessively by over-provisioning to guarantee quick startups, or minimize costs but endure slow cold starts. We are

model-releasesgoogle-cloud-ai
1 Jun 2026
Model Releases

Inversion-Free Natural Gradient Descent on Riemannian Manifolds

DGX agent

arXiv:2604.02969v2 Announce Type: replace-cross Abstract: The natural gradient method is a central tool for statistical optimisation, but its broader application is hindered by the assumption of a Euc

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

just a small zoom out on the vibe shift: in Feb 2025 @soumithchintala was talking about his dream of personal, local, private agents, most p…

DGX agent

just a small zoom out on the vibe shift: in Feb 2025 @soumithchintala was talking about his dream of personal, local, private agents, most people didn't believe him. it's June 2026 and @pewdiepie has

model-releasesswyx--x
1 Jun 2026
Model Releases

KernelCraft: Benchmarking for Agentic Close-to-Metal Kernel Generation on Emerging Hardware

DGX agent

arXiv:2603.08721v2 Announce Type: replace-cross Abstract: New AI accelerators with novel instruction set architectures (ISAs) often require developers to manually craft low-level kernels, a time-consu

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

Knowledge Boundary Probing and Demand-Guided Intervention for LLM-Based Power System Code Generation

DGX agent

arXiv:2605.31478v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used to automate power-system analysis, but many utilities and energy-research labs require on-premise s

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

LangMap: A Human-Verified Benchmark for Hierarchical Open-Vocabulary Goal Navigation

DGX agent

arXiv:2602.02220v2 Announce Type: replace Abstract: Language-conditioned goal navigation (LGN) requires agents to locate user-specified targets without step-by-step guidance. However, existing benchma

model-releasesarxiv-cs-cv
1 Jun 2026
Model Releases

Language Models Learn Constructional Semantics, Not To Mention Syntax: Investigating LM Understanding of Paired-Focus Constructions

DGX agent

arXiv:2605.31586v1 Announce Type: cross Abstract: Grasping the semantics of rare constructions (form-meaning pairings) has been shown to be a challenging problem that has currently only been solved by

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Last week we revamped Liteparse to be the fastest PDF parser out there ⚡️ An underrated part of liteparse is it doesn't just give you text. …

DGX agent

Last week we revamped Liteparse to be the fastest PDF parser out there ⚡️ An underrated part of liteparse is it doesn't just give you text. It gives you bounding boxes that a coding agent can use to p

model-releasesjerry-liu--x
1 Jun 2026
Model Releases

Learning-Based Navigation for Indoor Mobile Robots

DGX agent

arXiv:2605.30468v1 Announce Type: new Abstract: This paper presents a learning-based navigation framework for indoor mobile robots. The proposed method combines a supervised neural global planner, tra

model-releasesarxiv-cs-ro
1 Jun 2026
Model Releases

Learning Multi-Agent Coordination via Sheaf-ADMM

DGX agent

arXiv:2605.31005v1 Announce Type: new Abstract: We present a differentiable optimization framework for multi-agent coordination. An input is decomposed into overlapping local views, each processed by

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

Learning Randomized Reductions

DGX agent

arXiv:2412.18134v4 Announce Type: replace Abstract: Randomized self-reductions (RSRs) express f(x) using f evaluated at random correlated points, enabling self-correcting programs, instance-hiding pro

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

Learning Whom to Trust: Market-Feedback Adaptive Retrieval for Frozen LLMs in Event-Driven Financial RAG

DGX agent

arXiv:2605.31201v1 Announce Type: new Abstract: Financial retrieval-augmented generation (RAG) systems typically rank evidence by textual relevance, but in financial markets the useful evidence source

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

LegSegNet: A Public Deep Learning System for Lower Extremity CT Tissue Segmentation and Quantification

DGX agent

arXiv:2605.30829v1 Announce Type: new Abstract: Lower extremity computed tomography (CT) contains clinically relevant information for body composition analysis, sarcopenia assessment, and musculoskele

model-releasesarxiv-cs-cv
1 Jun 2026
Model Releases

Linear Ordering Problem: Time for a Change

DGX agent

arXiv:2605.31051v1 Announce Type: cross Abstract: The Linear Ordering Problem (LOP) is a fundamental combinatorial optimization problem with important applications in areas such as economics, social c

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

LLM Bias Evaluation: Gender, Racial, and Age Disparities in Occupational and Crime Scenarios

DGX agent

arXiv:2409.14583v4 Announce Type: replace Abstract: LLM bias evaluation is critical as large language models (LLMs) increasingly influence high-stakes decisions. This paper provides a comprehensive as

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

LLMs Without Deep Neural Networks: New Architecture, Benefits and Case Study

DGX agent

arXiv:2605.30385v1 Announce Type: cross Abstract: The purpose of this article is to provide validation to my deep neural network alternative in the context of LLMs. Very recently, there has been a sig

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

LongDS-Bench: On the Failure of Long-Horizon Agentic Data Analysis

DGX agent

arXiv:2605.30434v1 Announce Type: cross Abstract: Real-world data analysis is inherently iterative, yet existing benchmarks mostly evaluate isolated or short interactive tasks, leaving agents' ability

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Lots of companies are in the 'encourage AI adoption' phase, whether teaching them ChatGPT/Claude or (sigh) tokenmaxxing. That dodges the har…

DGX agent

Lots of companies are in the 'encourage AI adoption' phase, whether teaching them ChatGPT/Claude or (sigh) tokenmaxxing. That dodges the harder problems of firm leadership: What do you want people to

model-releasesethan-mollick--x
1 Jun 2026
Model Releases

MAAT: Multi-phase Adapter-Aware Targeted Unlearning

DGX agent

arXiv:2605.30514v1 Announce Type: cross Abstract: Machine unlearning evaluation is structurally skewed: Why-type questions, which probe causal and relational knowledge, comprise less than 0.06% of Cou

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

MADS: Model-Aware Diverse Core Set Selection for Instruction Tuning

DGX agent

arXiv:2605.30857v1 Announce Type: new Abstract: Instruction fine-tuning is employed to enhance the instruction-following ability of large language models (LLMs). As the amount of instruction fine-tuni

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

Many research labs only consider inference efficiency after the fact. Step 3.7 Flash is a 196B MoE model, and built for inference from the s…

DGX agent

Many research labs only consider inference efficiency after the fact. Step 3.7 Flash is a 196B MoE model, and built for inference from the start by @StepFun_ai. Multi-Matrix Factorization Attention (M

model-releasesfireworks-ai--x
1 Jun 2026
Model Releases

MAVEN: Improving Generalization in Agentic Tool Calling

DGX agent

arXiv:2605.30738v1 Announce Type: new Abstract: Generalization across agentic tool-calling environments remains a central challenge for reliable agentic reasoning systems. Although large language mode

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

MedFact: Benchmarking the Fact-Checking Capabilities of Large Language Models on Chinese Medical Texts

DGX agent

arXiv:2509.12440v3 Announce Type: replace-cross Abstract: Deploying Large Language Models (LLMs) in medical applications requires fact-checking capabilities to ensure patient safety and regulatory com

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Mellum2 Technical Report

DGX agent

arXiv:2605.31268v1 Announce Type: new Abstract: We present Mellum 2, an open-weight 12B-parameter Mixture-of-Experts (MoE) language model with 2.5B active parameters per token. Mellum 2 is a general-p

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

Memory-Bound but Not Bandwidth-Limited: The Physical AI Inference Gap in Batch-1 LLM Decode

DGX agent

arXiv:2605.30571v1 Announce Type: cross Abstract: Physical AI systems, including robots, autonomous vehicles, embodied agents and edge copilots, often run a different inference workload from cloud LLM

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Memory by Design: Probabilistic Sequence Layers

DGX agent

arXiv:2605.31163v1 Announce Type: cross Abstract: We introduce the design-model framework: a way to derive efficient recurrent sequence maps from explicit assumptions about memory. A design model writ

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

Merge launches Agent Handler for Employees as an IT gatekeeper for workplace AI agents

DGX agent

Merge API Inc., a platform provider delivering connective infrastructure for artificial intelligence to business data and tools, launched Agent Handler for Employees, easing the strain on information

model-releasessiliconangle
1 Jun 2026
Model Releases

MIMO: Multilingual Information Retrieval via Monolingual Objectives

DGX agent

arXiv:2605.31171v1 Announce Type: cross Abstract: Multilingual Information Retrieval (MLIR) reflects real-world search environments in which queries and relevant documents may appear in different lang

model-releasesarxiv-cs-ai
1 Jun 2026
← Previous
1…239240241242243…475
Next →