AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,545 results
Model Releases

GLI-AL: A Multi-Modal Glioma MRI Label Resource with Unified Anatomy-Lesion Labels

DGX agent

arXiv:2607.22135v1 Announce Type: new Abstract: Existing BraTS-GLI datasets provide a widely used benchmark for adult glioma MRI segmentation, but their task definition focuses on tumor subregions and

model-releasesarxiv-cs-cv
27 Jul 2026
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Great technical paper from Google. Great read on why context beats scale for agents working against unfamiliar APIs. (bookmark it) GPU kerne…

DGX agent

Great technical paper from Google. Great read on why context beats scale for agents working against unfamiliar APIs. (bookmark it) GPU kernel optimization has KernelBench to hillclimb on. TPUs had not

model-releasesdair-ai--x
27 Jul 2026
Model Releases

Grok Build with Grok 4.5 stands far ahead on the efficiency frontier....completely alone inside the chart’s most attractive quadrant It deli…

DGX agent

Grok Build with Grok 4.5 stands far ahead on the efficiency frontier....completely alone inside the chart’s most attractive quadrant It delivers top-tier coding-agent performance while using only arou

model-releaseselon-musk--x
27 Jul 2026
Model Releases

Ground Truth First: A Longitudinal Evaluation Instrument for Agent Memory, and the Tenure Crossover in Memory-Architecture Rankings

DGX agent

arXiv:2607.21962v1 Announce Type: new Abstract: Benchmarks for LLM-agent memory typically generate conversations first and extract answer keys afterwards -- with documented label-error and contaminati

model-releasesarxiv-cs-cl
27 Jul 2026
Model Releases

Happy to have @FireworksAI_HQ as our day0 launch partner and bring Kimi K3 to more developers. With Fireworks, you can now deploy and fine-t…

DGX agent

Happy to have @FireworksAI_HQ as our day0 launch partner and bring Kimi K3 to more developers. With Fireworks, you can now deploy and fine-tune the 2.8T Kimi K3 model with just a few clicks! Kimi K3 i

model-releaseskimi-moonshot--x
27 Jul 2026
Model Releases

Hyperball May Not Be a Free Lunch

DGX agent

arXiv:2607.22444v1 Announce Type: new Abstract: For scale-invariant deep networks, Hyperball-style optimizers have shown strong performance in large-scale training by fixing the norms of matrix-valued

model-releasesarxiv-cs-lg
27 Jul 2026
Model Releases

I ran the 35B agentic comparison someone asked for (stock vs Ornith vs KAT-Coder, 120 runs)

DGX agent

Someone in the comments of my 27B post-train bakeoff asked for the 35B version, so I ran it. Same setup as last time: fresh Coder workspaces on my k8s cluster, each driving my own agent (Hermes) headl

model-releasesr-localllama
27 Jul 2026
Model Releases

IFCLoRA: Topology-Aware Rank Allocation for Parameter-Efficient Fine-Tuning

DGX agent

arXiv:2607.22251v1 Announce Type: new Abstract: Low-Rank Adaptation (LoRA) is a widely used parameter-efficient fine-tuning method for large language models, but its performance depends strongly on ho

model-releasesarxiv-cs-lg
27 Jul 2026
Model Releases

Improving Large Vision-Language Models' Understanding for Flow Field Data

DGX agent

arXiv:2507.18311v3 Announce Type: replace Abstract: Large Vision-Language Models (LVLMs) have shown impressive capabilities across a range of tasks that integrate visual and textual understanding, suc

model-releasesarxiv-cs-cv
27 Jul 2026
Model Releases

Indexing: the Beginning and the End

DGX agent

arXiv:2607.22361v1 Announce Type: new Abstract: We study information bottlenecks in modern deep-learning architectures -- RNNs, softmax transformers, linear-attention transformers and state-space mode

model-releasesarxiv-cs-lg
27 Jul 2026
Model Releases

InteractComp: Evaluating Search Agents With Ambiguous Queries

DGX agent

arXiv:2510.24668v2 Announce Type: replace Abstract: Language agents have demonstrated remarkable potential in web search and information retrieval. However, many search-agent benchmarks assume that us

model-releasesarxiv-cs-cl
27 Jul 2026
Model Releases

Interpretable Anomaly and Drift Detection with Gaussian Mixture Models

DGX agent

arXiv:2607.16811v2 Announce Type: replace Abstract: We revisit Gaussian Mixture Models (GMMs) as a lightweight, interpretable tool for anomaly detection and, in particular, for detecting distributiona

model-releasesarxiv-cs-lg
27 Jul 2026
Model Releases

Interpretable EEG biomarkers with bag-of-waves: Spatial and temporal waveform dictionaries for low-data regimes

DGX agent

arXiv:2607.22508v1 Announce Type: new Abstract: Electroencephalography (EEG) is widely used to diagnose neurological conditions, but its analysis usually relies on either predefined spectral features

model-releasesarxiv-cs-lg
27 Jul 2026
Model Releases

IR275K: A Benchmark for Infrared Multi-Frame Super-Resolution Toward Efficient Remote Sensing

DGX agent

arXiv:2607.22380v1 Announce Type: new Abstract: Efficient processing is becoming increasingly important in infrared remote sensing, where satellite constellations produce large volumes of observations

model-releasesarxiv-cs-cv
27 Jul 2026
Model Releases

J-CoT: Chain-of-Thought in J-Space

DGX agent

arXiv:2607.21981v1 Announce Type: new Abstract: Chain-of-thought prompting improves language-model reasoning by carrying intermediate states across successive computation steps. However, relying on na

model-releasesarxiv-cs-cl
27 Jul 2026
Model Releases

k{appa}-LoRA: Condition Numbers Reveal Which LoRA Matrices Worth Updating

DGX agent

arXiv:2607.22489v1 Announce Type: new Abstract: Low-Rank Adaptation (LoRA) has become a widely adopted technique for efficient neural network fine-tuning, decomposing model updates into low-rank matri

model-releasesarxiv-cs-lg
27 Jul 2026
Model Releases

Kat Coder 2.5 is insane. Especially considering I ran it at Q4_K_M

DGX agent

I tested Kat Coder 2.5 with this prompt: Create a spaceship game inspired by Star Fox using vanilla Three.js and HTML. It should have at least five levels, keyboard and mouse controls, enemies, and a

model-releasesr-localllama
27 Jul 2026
Model Releases

Khondo: A Multimodal Benchmark for Document Packet Splitting of Bangla Forms

DGX agent

arXiv:2607.21780v1 Announce Type: new Abstract: Document packets, multiple documents concatenated into a single file, are common in government and administrative workflows, yet splitting them into the

model-releasesarxiv-cs-cl
27 Jul 2026
Model Releases

Kimi K3 is live on Fireworks. Day 0, inference and training. US-hosted, and zero data retention. This is the first frontier open model in th…

DGX agent

Kimi K3 is live on Fireworks. Day 0, inference and training. US-hosted, and zero data retention. This is the first frontier open model in the 3 trillion parameter class. It sports 1M context, native v

model-releasesfireworks-ai--x
27 Jul 2026
Model Releases

Kimi K3 is now available for inference & training in @FireworksAI_HQ. Crazy how easy they make it to tune frontier open models like K3 using…

DGX agent

Kimi K3 is now available for inference & training in @FireworksAI_HQ. Crazy how easy they make it to tune frontier open models like K3 using LoRA adapters. Best time to figure out how to own your inte

model-releasesdair-ai--x
27 Jul 2026
Model Releases

Kimi K3 is now available on Ollama’s cloud. To use it with Claude Code, run: ollama launch claude --model kimi-k3:cloud Currently Kimi K3 re…

DGX agent

Kimi K3 is now available on Ollama’s cloud. To use it with Claude Code, run: ollama launch claude --model kimi-k3:cloud Currently Kimi K3 requires a Pro or Max subscription, and consumes extra usage c

model-releasesollama--x
27 Jul 2026
Model Releases

Kimi K3 weights drop today. We're deploying on A100s, H200s and B300s this week and the A100 math is already rough

DGX agent

tldr; we are going to host K3 on A100s (yes, thats correct, we'll try to see if it holds up), H200s & B300s - expect results for A100s & H200s this week while we setup the B300 cluster this weekend &

model-releasesr-localllama
27 Jul 2026
Model Releases

@Kimi_Moonshot K3 on Together AI is built for long-running agent workflows: → 2.8T parameters and a 1M context window → Native vision for sc…

DGX agent

@Kimi_Moonshot K3 on Together AI is built for long-running agent workflows: → 2.8T parameters and a 1M context window → Native vision for screenshot-guided coding → Repository navigation and terminal

model-releasestogether-ai--x
27 Jul 2026
Model Releases

Language-Aware Distillation for Multilingual Instruction-Following Speech LLMs with ASR-Only Supervision

DGX agent

arXiv:2603.07025v2 Announce Type: replace Abstract: Speech Large Language Models (LLMs) that understand and follow instructions in many languages are useful for real-world interaction, but are difficu

model-releasesarxiv-cs-cl
27 Jul 2026
Model Releases

Latent PDE mapping for efficient physics-informed learning across geometries with limited data

DGX agent

arXiv:2607.22215v1 Announce Type: new Abstract: In this study, we introduce latent PDE mapping, a broadly applicable physics-informed learning technique designed to enable efficient geometric generali

model-releasesarxiv-cs-lg
27 Jul 2026
Model Releases

Layer-wise LoRA fine-tuning: a similarity metric approach

DGX agent

arXiv:2602.05988v2 Announce Type: replace Abstract: Pre-training Large Language Models (LLMs) on web-scale datasets becomes fundamental for advancing general-purpose AI. In contrast, enhancing their p

model-releasesarxiv-cs-lg
27 Jul 2026
Model Releases

LeAct: Learning to Reason from Expert Actions

DGX agent

arXiv:2607.21856v1 Announce Type: cross Abstract: Modern reasoning models depend on reasoning data, today sourced from human annotations or distilled from stronger LLMs. However, a rich and largely un

model-releasesarxiv-cs-cl
27 Jul 2026
Model Releases

Learning What Matters: Supervising Sparse Attention Routing with Causal Evidence Sets

DGX agent

arXiv:2607.21692v1 Announce Type: cross Abstract: Sparse attention reduces the cost of long contexts by allowing each query to read only selected parts of the input. These selectors are often trained

model-releasesarxiv-cs-cl
27 Jul 2026
Model Releases

Ling-3.0-flash weights: SGLang says day-0, vLLM says when they land, llama.cpp closed the 2.6 request as not_planned

DGX agent

Some Ling-3.0-flash threads here last week ended on the same two questions with no real answer, so I went through the repos. State as of writing, with links so you can check instead of taking my word

model-releasesr-localllama
27 Jul 2026
Model Releases

LLM-Based Visual Explanation Evaluation Framework for Assessing the Explainability of Facial Skin Disease Classification Models

DGX agent

arXiv:2606.16794v2 Announce Type: replace Abstract: This study proposes a domain-specific LLM-based Visual Explanation Evaluation Framework for assessing visual attention explanations in facial skin d

model-releasesarxiv-cs-cv
27 Jul 2026
Model Releases

LMEB: Long-horizon Memory Embedding Benchmark

DGX agent

arXiv:2603.12572v5 Announce Type: replace Abstract: Memory embeddings are crucial for memory-augmented systems, such as OpenClaw, but their evaluation is underexplored in current text embedding benchm

model-releasesarxiv-cs-cl
27 Jul 2026
Model Releases

Local-Global Geometric Insights for Graph Neural Networks via Entropic Curvature

DGX agent

arXiv:2607.22381v1 Announce Type: new Abstract: Curvature notions on graphs, particularly Ollivier-Ricci and Forman, have emerged as powerful tools for addressing fundamental issues in Graph Neural Ne

model-releasesarxiv-cs-lg
27 Jul 2026
Model Releases

love this frame. calls to mind the role of the hippocampus in human navigation (via place cells and grid cells), and how navigation is, in a…

DGX agent

love this frame. calls to mind the role of the hippocampus in human navigation (via place cells and grid cells), and how navigation is, in a sense, what makes agents *agents* vs plain old LLM calls in

model-releasesyohei-nakajima--x
27 Jul 2026
Model Releases

Medical-Checklist: Assessing the Comprehension of Medical Images by Multimodal Models

DGX agent

arXiv:2607.21998v1 Announce Type: new Abstract: This paper introduces a new benchmark test, Medical-Checklist, for assessing medical multimodal models. The recent advancements in multimodal models hav

model-releasesarxiv-cs-cv
27 Jul 2026
Model Releases

Microsoft introduces MAI-Cyber-1-Flash, an AI model trained for cybersecurity, and launches Perception, an agentic security system to patch vulnerabilities (New York Times)

DGX agent

New York Times: Microsoft introduces MAI-Cyber-1-Flash, an AI model trained for cybersecurity, and launches Perception, an agentic security system to patch vulnerabilities — As some executives fret ov

model-releasestechmeme
27 Jul 2026
Model Releases

Modernizing the skies: NOAA and Google Cloud collaborate to advance weather forecasting

DGX agent

The National Oceanic and Atmospheric Administration (NOAA) is embarking on a transformative journey to redefine how we understand and predict patterns in the Earth’s atmosphere that affect the weather

model-releasesgoogle-cloud-ai
27 Jul 2026
Model Releases

MoE^2-LoRA: When MoE Models Meet MoE-style Low-Rank Adaptation

DGX agent

arXiv:2607.21978v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) architectures have been widely adopted in large language models, yet parameter-efficient fine-tuning (PEFT) for MoE models rema

model-releasesarxiv-cs-cl
27 Jul 2026
Model Releases

moonshotai/Kimi-K3

DGX agent

moonshotai/Kimi-K3 As promised earlier this month, Moonshot have released the weights for their excellent 2.8 trillion parameter Kimi K3. They're a hefty 1.56TB on Hugging Face. Kimi introduced their

model-releasessimon-willison
27 Jul 2026
Model Releases

Neural Feature Governance: Extending Atom Prevalence

DGX agent

arXiv:2607.21671v1 Announce Type: new Abstract: Neural network compression and interpretability remain open challenges in modern deep learn- ing, where billion-parameter architectures deliver impressi

model-releasesarxiv-cs-lg
27 Jul 2026
Model Releases

Nexus connects to the agentic harnesses your teams already use, whether that’s Claude Code, Codex, OpenCode, or your own custom tooling. It …

DGX agent

Nexus connects to the agentic harnesses your teams already use, whether that’s Claude Code, Codex, OpenCode, or your own custom tooling. It gives you: → Intelligent routing, automatically matching eac

model-releasesfireworks-ai--x
27 Jul 2026
Model Releases

Nice little insights on doing autoresearch with coding agents. Hand a coding agent a dataset, an eval script, one editable file, and no supe…

DGX agent

Nice little insights on doing autoresearch with coding agents. Hand a coding agent a dataset, an eval script, one editable file, and no supervision. That's autoresearch and it tries to optimize the nu

model-releasesdair-ai--x
27 Jul 2026
Model Releases

Nifer is insane. 700t/s with Qwen 3.6 35B (no thinking). Purpose build for RTX5090. Full 250k context too.

DGX agent

I just managed to get it running on windows and this thing is fucking insane. I get around 550-720t/s depending on task at hand. Previously to get to such numbers i would have to do batching and agent

model-releasesr-localllama
27 Jul 2026
Model Releases

NVIDIA Nemotron 3 Ultra Leads Open Models on Accuracy and Efficiency in Agentic RTL Coding

DGX agent

NVIDIA’s Nemotron 3 Ultra, when paired with the ACE‑RTL agent, achieves a 97.1 % average pass rate on the CVDP benchmark across nine RTL task categories—surpassing GLM 5.2 and Kimi K2.6 while using up

model-releasesnvidia-developer
27 Jul 2026
Model Releases

NWaaS: A Non-Intrusive and Privacy-Preserving Watermarking-as-a-Service System with Adaptive Resource Scheduling

DGX agent

arXiv:2507.18036v2 Announce Type: replace-cross Abstract: Securing intellectual property (IP) in Machine Learning as a Service is critical yet challenging. While deep neural network watermarking serve

model-releasesarxiv-cs-cv
27 Jul 2026
Model Releases

Offline Vision-Language Navigation with Geometric Goal Localization for Outdoor Environments

DGX agent

arXiv:2607.22226v1 Announce Type: new Abstract: Foundation-model-based vision-language navigation (VLN) has advanced autonomous robot navigation by enabling robots to interpret natural-language instru

model-releasesarxiv-cs-ro
27 Jul 2026
Model Releases

On FrontierCode 1.1 Extended, our benchmark for real-world engineering tasks that grades mergeability and quality, Kimi K3 scores 58.2% with…

DGX agent

On FrontierCode 1.1 Extended, our benchmark for real-world engineering tasks that grades mergeability and quality, Kimi K3 scores 58.2% with a 63.6% pass rate. Within Devin, it excels on reproducing b

model-releasescognition-ai--x
27 Jul 2026
Model Releases

One Hand Watches The Other: Dynamic Multi-Agent Cooperation for Sample-Efficient Bimanual Manipulation in Dynamic Environments

DGX agent

arXiv:2607.22119v1 Announce Type: cross Abstract: Multi-stream robot manipulation policies achieve unparalleled sample efficiency and generalization by modeling actions relative to environmental refer

model-releasesarxiv-cs-lg
27 Jul 2026
Model Releases

Opaque Epistemic Mediation: How LLM Deployment Configurations Shape the Validation of Pseudo-Science

DGX agent

arXiv:2607.22513v1 Announce Type: cross Abstract: Commercial large language models are increasingly used as knowledge references, yet their stance on contested scientific claims is neither stable nor

model-releasesarxiv-cs-cl
27 Jul 2026
← Previous
1…8384858687…470
Next →