AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,098 results
10 Aug 2026

Configure and run Fireworks training jobs from your coding agent. Install the Training Skill in one line, describe your goal, and Claude Cod…

Model ReleasesDGX agent

Configure and run Fireworks training jobs from your coding agent. Install the Training Skill in one line, describe your goal, and Claude Code, Codex, or Cursor helps choose the method, validate data,

Control-Anchored Residual Flow Matching Conditioned on Gene Geometry for Virtual Cell Perturbation Modeling

Model ReleasesDGX agent

arXiv:2608.06824v1 Announce Type: cross Abstract: A central task in virtual cell modeling is predicting single-cell transcriptional responses to unseen genetic perturbations and drug combinations, and

Corma launches with $60M in funding for defensive cybersecurity AI


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases
DGX agent

Defensive cybersecurity startup Corma Labs Ltd. today announced it has raised 60 million in seed funding to build a foundation model purpose-built for security defense. Founded in 2025, Corma runs off

Coupling Planning with Episodic Memory in LLM Agents for Software Issue Resolution

Model ReleasesDGX agent

arXiv:2608.06811v1 Announce Type: cross Abstract: Resolving a real software issue with a large language model (LLM) agent is a long repair episode, often tens to hundreds of steps spanning exploration

Critical Acclaim Orientation in Large Language Models: Evidence from Film Preference Elicitation

Model ReleasesDGX agent

arXiv:2608.06955v1 Announce Type: new Abstract: Large language models (LLMs) are trained on corpora that contain expressions of human judgment about films, books, music, and more. Yet whether LLMs sys

CrossTracer: Cross-Embodiment Navigation via VLA Model Reasoning and Trace Residuals Adapting

Model ReleasesDGX agent

arXiv:2608.06688v1 Announce Type: new Abstract: Vision-language-action (VLA) models provide strong semantic priors for robot navigation, but they often ignore embodiment-specific mobility constraints.

Cryptanalytic Extraction of Isolated Bias-Free GLU Feed-Forward Blocks by Antipodal Separation

Model ReleasesDGX agent

arXiv:2608.06631v1 Announce Type: cross Abstract: Cryptanalytic extraction has been demonstrated for ReLU networks, for networks using componentwise activations such as GELU or SiLU, and for a Transfo

Cyber resilience becomes the ultimate measure of security success as AI reshapes the CISO role

Model ReleasesDGX agent

CISO role evolution is accelerating as cyber resilience emerges as a defining benchmark of security success and AI forces enterprises to rethink not just how they defend against attacks, but how quick

CyberForge: Verified Vulnerability Injection at Repository Level for Cybersecurity Agent Training

Model ReleasesDGX agent

arXiv:2608.06471v1 Announce Type: cross Abstract: Despite recent advances, frontier large language model (LLM) agents remain limited in discovering and patching complex vulnerabilities in real-world s

DATAREEL: Automated Data-Driven Video Story Generation with Animations

Model ReleasesDGX agent

arXiv:2604.25220v2 Announce Type: replace Abstract: Data videos combine animated visualizations with synchronized narration to communicate quantitative information and are widely used in journalism, e

Deep Evidential Regression for Sparse Forest Height Estimation from Multimodal Satellite Imagery

Model ReleasesDGX agent

arXiv:2608.06406v1 Announce Type: new Abstract: Accurate estimation of forest height from satellite imagery is essential for applications such as carbon accounting, biodiversity monitoring, and ecosys

DeepSeek V4 Flash 0731 is the ‘killer app’ that is going to sell A LOT of DGX Sparks

Model ReleasesDGX agent

Having a ‘Killer Application’ that everyone wants to use helps sell hardware, plain and simple. DeepSeek V4 Flash 0731 isn’t an app of course, but I think it’s going to be the major catalyst for getti

Density-aware Hierarchical Clustering Based on Element-Categorized Connection Subgraphs

Model ReleasesDGX agent

arXiv:2608.06990v1 Announce Type: cross Abstract: Clustering is a fundamental data mining technique for pattern recognition through unsupervised learning. Among various clustering methods, hierarchica

Diffusion LLMs as Targets and Adversaries: Mechanistic Safety Exploits

Model ReleasesDGX agent

arXiv:2608.07430v1 Announce Type: cross Abstract: Diffusion Large Language Models (DLLMs) replace autoregressive next-token prediction with iterative parallel denoising, yet their internal safety mech

Divergent Response Modes in Frontier Language Models Under Steering Pressure

Model ReleasesDGX agent

arXiv:2608.06578v1 Announce Type: new Abstract: Frontier language models are trained using distinct data, objectives, and safety pipelines. Whether these differences produce measurably different behav

Do AI Personas Grow? Analyzing and Benchmarking Personality Evolution in LLM Agents After Life Events

Model ReleasesDGX agent

arXiv:2608.06485v1 Announce Type: cross Abstract: Personality-conditioned LLM agents (PC-Agents) are increasingly used in emotional support, social simulation, and role-playing, motivating the develop

Do Audio Language Models Use Paralinguistic Evidence? Counterfactual Audits for Response Evaluation

Model ReleasesDGX agent

arXiv:2608.06718v1 Announce Type: new Abstract: Audio-language models (ALMs) are increasingly used as judges for speech-to-speech systems, but a judge that receives audio may not actually use paraling

ECAD: Expanding Class-Agnostic Detection Beyond Thing-Centric Objectness

Model ReleasesDGX agent

arXiv:2608.06841v1 Announce Type: new Abstract: Object detection is a fundamental task in visual perception, providing structured region representations for recognition, grounding, reasoning, and inte

EchoVLA: Robotic Vision-Language-Action Model with Synergistic Declarative Memory for Mobile Manipulation

Model ReleasesDGX agent

arXiv:2511.18112v3 Announce Type: replace Abstract: Recent progress in Vision-Language-Action (VLA) models has enabled embodied agents to interpret multimodal instructions and perform complex tasks. H

ED-CSP: Crystal Structure Prediction from Electron Diffraction

Model ReleasesDGX agent

arXiv:2608.06448v1 Announce Type: cross Abstract: Recovering a periodic 3D crystal structure from sparse, unindexed electron diffraction (ED) observations is a challenging generative inverse problem.

Edge Sparsification via Temporal Forman-Ricci Curvature for Dynamic Graph Learning

Model ReleasesDGX agent

arXiv:2608.07158v1 Announce Type: new Abstract: Temporal graph learning has become essential for analyzing real-world systems whose interactions continuously evolve over time, including financial tran

EMAS: Stabilizing Multi-Agent System Evolution through Evidence-Guided Revision

Model ReleasesDGX agent

arXiv:2608.07196v1 Announce Type: new Abstract: Many methods for automated multi-agent system design optimize prompts and topologies during an initial design stage and then deploy the resulting system

Evaluating XAI Support From A Hierarchical Reinforcement Learning Policy in Human-Agent Collaboration

Model ReleasesDGX agent

arXiv:2608.06381v1 Announce Type: cross Abstract: Explainable AI (XAI) has shown promise for human-agent collaboration, yet results rely on hand-crafted policies in custom environments, limiting gener

Explanation Stability of Test-Time Adaptation in Computational Pathology: A Large-Scale Benchmark

Model ReleasesDGX agent

arXiv:2608.07062v1 Announce Type: new Abstract: Test-time adaptation (TTA) has become a practical way to adapt deployed models to unlabeled target data, a setting that is especially relevant in comput

Fairis: Fairness-Aware Aggregation with Provable Influence Containment against Fairness Poisoning Attacks in Collaborative Machine Learning

Model ReleasesDGX agent

arXiv:2608.06469v1 Announce Type: cross Abstract: Collaborative machine learning among financial institutions must be both group-fair and robust against deliberate adversarial manipulation. Existing f

Fast and Accurate: An Adaptive VLA Inference Framework through Environment-aware Model Selection

Model ReleasesDGX agent

arXiv:2608.06434v1 Announce Type: cross Abstract: Embodied intelligence demands both long-horizon reasoning and real-time closed-loop responsiveness. Recent dual-system Vision-Language-Action (VLA) ar

Faster Query-Key Learning Sharpens Attention in Self-Attention Models

Model ReleasesDGX agent

arXiv:2608.06776v1 Announce Type: new Abstract: A standard self-attention layer consists of two interacting circuits: the query-key circuit that governs attention allocation, and the output-value circ

Finding Usable Weight Mechanisms with Tiled SVD

Model ReleasesDGX agent

arXiv:2608.06969v1 Announce Type: new Abstract: The dominant approach to mechanistic interpretability trains proxy dictionaries such as sparse autoencoders and labels features from max-activating text

FinRank: An Evidence-Grounded Benchmark for Financial Question Answering and Retrieval over SEC Filings

Model ReleasesDGX agent

arXiv:2608.07400v1 Announce Type: new Abstract: Financial question answering is typically evaluated by answer correctness, yet in SEC filings a plausible and even numerically correct answer can be gro

Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing

Model ReleasesDGX agent

arXiv:2608.07437v1 Announce Type: new Abstract: Reliable hypothesis testing is the foundation of many empirical scientific claims. Large language model (LLM) agents are increasingly used to automate t

Frontier performance you can actually own. Proud to help power DeepSeek-V4-Flash on Ollama's cloud, with the fastest hosted performance avai…

Model ReleasesDGX agent

Frontier performance you can actually own. Proud to help power DeepSeek-V4-Flash on Ollama's cloud, with the fastest hosted performance available. Open weights, zero data retention, U.S. & EU hosting.

Game-Theoretic Inverse Reinforcement Learning for Modeling Competitive Human Driving: A Cut-in Prediction Study

Model ReleasesDGX agent

arXiv:2608.06445v1 Announce Type: cross Abstract: Capturing the strategic decision-making inherent in competitive human driving is critical for autonomous vehicle safety and traffic simulation. This s

Generative Embedding Benchmark: How Much Information Survives in a Dense Embedding?

Model ReleasesDGX agent

arXiv:2608.06972v1 Announce Type: new Abstract: Embeddings have emerged as a standard representational interface linking foundation models with downstream systems. Most embedding benchmarks assess rep

Geo-Spatial Concept Probing of Large Language Models: Abstraction, Compositionality, and Grounding

Model ReleasesDGX agent

arXiv:2608.07353v1 Announce Type: cross Abstract: Understanding concepts is fundamental to generalization. Despite their impressive performance on a wide range of tasks, Large Language Models (LLMs) s

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks

Model ReleasesDGX agent

arXiv:2608.07411v1 Announce Type: new Abstract: In the context of geodata, existing Large Language Models have often been studied in a homogeneous setting, which has considerably limited insights into

Georeferencing Non-Gazetteered Place Names using Biological Specimen Records

Model ReleasesDGX agent

arXiv:2608.06884v1 Announce Type: cross Abstract: Biological specimen records collected by natural history institutions constitute a rich source of temporal geographic knowledge, capturing biodiversit

Google named a Leader in The Forrester Wave™: AI Platforms, Q3 2026

Model ReleasesDGX agent

At Google Cloud, we help organizations of all sizes build and operationalize complex agentic workflows with total confidence. By combining world-class AI research with an open, fully integrated AI pla

Got Meta's new Muse Glimmer 30B running on my MacBook (M3 Max, 96GG) and tested the serving options available so far. Fastest right now: Oll…

Model ReleasesDGX agent

Got Meta's new Muse Glimmer 30B running on my MacBook (M3 Max, 96GG) and tested the serving options available so far. Fastest right now: Ollama's MLX engine (DFlash included) at ~29 tok/s. Tuned llama

GPT 5.6 Sol High and X-High (Web Chat) feels severely nerfed since 08/06/2026 update

Model ReleasesDGX agent

GPT-5.6 Sol High and X-High (Web Chat) feels severely nerfed since 08/06/2026 update (https://openai.com/index/improving-gpt-5-6-sol-in-chatgpt) I use it for a pretty complex Unreal Engine 5 project (

GraphVerse: A Comprehensive Visual Graph Reasoning Benchmark for Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2608.06769v1 Announce Type: new Abstract: Recent Multimodal Large Language Models (MLLMs) have achieved remarkable progress across diverse vision-language tasks, creating an urgent need for more

GRASP: Reinforcing Language Model Anonymizers with Group Relative Policy Optimization

Model ReleasesDGX agent

arXiv:2608.06526v1 Announce Type: new Abstract: Large language models can infer sensitive personal attributes, such as age, location, and occupation, from ordinary text, turning everyday writing into

Grok Imagine is focused on professional usefulness, fun for consumers and overall ease-of-use

Model ReleasesDGX agent

Grok Imagine is focused on professional usefulness, fun for consumers and overall ease-of-use Grok 4.5 just ranked #1 on Design Arena for Daily Usage - measuring the average unique users actually inte

HarnessSafe: Evaluating Safety Across Persistent Carriers in Agent Harnesses

Model ReleasesDGX agent

arXiv:2608.06984v1 Announce Type: cross Abstract: Modern agent harnesses persist state across tasks and sessions through persistent carriers like memory, skills, tools, and shared artifacts. However,

Hidden Gauge Controls Feature Specialization in ReLU Networks

Model ReleasesDGX agent

arXiv:2608.06766v1 Announce Type: cross Abstract: Training changes a network's predictions while allocating task-relevant structure across its internal units. In an overparameterized ReLU network, sev

HLSmith: An Expert-Guided Agentic Framework for C/C++-to-HLS Translation

Model ReleasesDGX agent

arXiv:2608.06791v1 Announce Type: cross Abstract: Application-specific FPGA accelerators offer substantial performance and energy-efficiency gains across many application domains, but developing them

Hyperbolic Graph Embedders for Link Prediction and Topology Reconstruction

Model ReleasesDGX agent

arXiv:2608.07029v1 Announce Type: new Abstract: Hyperbolic embeddings provide compact geometric representations of complex networks in hyperbolic spaces, but systematic comparisons of methods develope

I asked OPUS 5 to make a video about what it's like to be an LLM

Model ReleasesDGX agent

Full prompt I gave to Claude Opus 5: can you use whatever resources you like, and python, to generate a short 'youtube poop' video and render it using ffmpeg ? can you put more of a personal spin on i

I compared GGUF quants of Qwen3.6 27B to NVFP4, AWQ, AutoRound, and FP8

Model ReleasesDGX agent

There's an interactive chart and some extra data in the blog post if you're interested. There are plenty of KL-divergence benchmarks for GGUF models, but most of them compare one GGUF quant against an

I Seek You in Videos: Identity-Conditioned Queries for Person-Centric Video Reasoning

Model ReleasesDGX agent

arXiv:2608.07417v1 Announce Type: cross Abstract: Real-world video reasoning often involves multimodal, multi-source inputs, whereas existing video reasoning tasks typically assume a simplified video-

I think this shows 2 key skills w/ AI 1) compute allocation - for most jobs there's not a list of 'X most important problems', you have to d…

Model ReleasesDGX agent

I think this shows 2 key skills w/ AI 1) compute allocation - for most jobs there's not a list of 'X most important problems', you have to decide what problems are worth it 2) thought partnership - @_

I trained a 1B-parameter LLM from scratch on 20B tokens for about $200

Model ReleasesDGX agent

A few months ago, I had the idea of making a LLM from scratch as a personal project (for learning and partly for improving my resume). Since I learned a lot from other posts on here over the past year

IB-RL: Isolated Bilateral Reinforcement Learning for Strategic Dialogue Agents

Model ReleasesDGX agent

arXiv:2608.06735v1 Announce Type: new Abstract: Reinforcement learning (RL) has achieved strong results in improving large language models (LLMs) on tasks with stationary, verifiable rewards, such as

inclusionAI/Ling-3.0-tiny · 8B A1.3B MoE· Hugging Face

Model ReleasesDGX agent

Looks like the Ling team open weighted a much smaller version of the Ling-3.0-flash they open weighted a few days ago. It's 8B params with 1.3B active, and seems to fall between the 4B and 8-12B Qwen

Index SLM Technical Report

Model ReleasesDGX agent

arXiv:2607.09885v2 Announce Type: replace Abstract: We present Index-1.9B, a series of open small language models developed at Bilibili. The series comprises four models: Index-1.9B-Base, a foundation

InsertFuse: A Unified Framework for Multi-Category Reference-Guided Image Insertion

Model ReleasesDGX agent

arXiv:2608.06490v1 Announce Type: new Abstract: We present InsertFuse, a unified framework for multi-category reference-guided image insertion. Its key idea is to decouple category-specific expertise

Introducing Muse Glimmer

Model ReleasesDGX agent

Introducing Muse Glimmer Meta are back in the open weights game! Muse Glimmer is a brand new 30B model under a clean Apache 2.0 license (a step up from the janky Llama licenses of old). They claim to

Introducing Muse Glimmer, an open-weight 30B-parameter model optimized for local, always-on agent workflows. Muse Glimmer delivers strong pe…

Model ReleasesDGX agent

Introducing Muse Glimmer, an open-weight 30B-parameter model optimized for local, always-on agent workflows. Muse Glimmer delivers strong performance on key agentic use cases and benchmarks compared w

Introducing Muse Glimmer: an open-weight model optimized for always-on local agent workflows

Model ReleasesDGX agent

Hi r/LocalLLaMA 👋 Today we’re excited to release Muse Glimmer, a 30B open-weight model built specifically for local agent workflows. We’re releasing the weights to the community under a permissive Apa

Introducing the Developer Device Platform for agentic mobile app development

Model ReleasesDGX agent

Most enterprises connect with their customers through a device. Whether it’s using a mobile app to order a product, contact customer service, view content, or manage their account, the customer experi

Iterative Training of Physics-Informed Neural Networks with Fourier-enhanced Features

Model ReleasesDGX agent

arXiv:2510.19399v2 Announce Type: replace Abstract: Spectral bias, the tendency of neural networks to learn low-frequency features first, is a well-known issue with many training algorithms for physic

← Previous
1…1213141516…369
Next →