AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,504 results
4 Aug 2026

Kimi K3 full model running on 16x GB10 cluster at 20+tps

Model ReleasesDGX agent

Kimi K3 full model running on 16x GB10 cluster at 20+tps average (llama-benchy coherent corpus) 38tps peak, 750tps prefill. This is the first run of full k3 with dspark on my cluster. I will be doing

Leak It: A Probabilistic Approach to Training-Data Extraction from Black-Box Language Models

ResearchDGX agent

arXiv:2608.00144v1 Announce Type: cross Abstract: Membership inference (MIA) on language models is usually summarised by an aggregate ROC-AUC, but such evaluations are confounded: model-free blind bas

MDTD-ArtIR: Benchmarking Image Editing and Restoration Models for Art Image Restoration under Texture-Overlay Degradations

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2608.00736v1 Announce Type: new Abstract: Restoring severely degraded visual media still remains a formidable challenge, as existing methods often hallucinate unnatural textures and contents, st

Model-Agnostic FDR Control via Group Gaussian Mirror and Permutation SHAP

ApplicationsDGX agent

arXiv:2608.00989v1 Announce Type: cross Abstract: Most FDR-controlled feature selection methods are designed for coordinate-wise hypotheses, where each feature has a single weight or importance score.

Robust Watermarks Meet Backdoored Models: Evading Diffusion Semantic Watermarks via Stealthy Backdoor

Model ReleasesDGX agent

arXiv:2608.00543v1 Announce Type: cross Abstract: Although semantic watermarking is considered a promising safeguard for images generated by Latent Diffusion Models (LDMs), the reliance of the waterma

RSVideo: Are Your Vision-Language Models Ready for Remote Sensing Videos?

Model ReleasesDGX agent

arXiv:2608.02039v1 Announce Type: new Abstract: Remote-sensing videos enable real-time observation of changes in target attributes, short-term activities, and scene evolution. They record motion, acti

SCALP: Semi-Supervised Statistical Shape Modeling from Imperfect 3D Photogrammetry via Landmark-Anchored Spectral Warp

ApplicationsDGX agent

arXiv:2608.00187v1 Announce Type: new Abstract: Correspondence-based statistical shape modeling (SSM) is vital for population-level morphometric analysis, but conventional pipelines assume clean, full

SelfWAM: A Self-Grounded Unified World Action Model for Fast Robot Control

SafetyDGX agent

arXiv:2608.00725v1 Announce Type: new Abstract: World Action Models (WAMs) improve robot policy learning by jointly modeling actions and future observations. However, conditioning future prediction on

SG-WAM: Self-Guided World Modeling in Geometry-Aware Policy Space

SafetyDGX agent

arXiv:2608.01397v1 Announce Type: cross Abstract: World Action Models (WAMs) couple action generation with prediction of future states. Their effectiveness depends on whether future dynamics are model

SPECTRA: Band-Routed Embedding and Stage-Wise LoRA for Cross-Sensor Fine-Tuning of Geospatial Foundation Models

Model ReleasesDGX agent

arXiv:2608.01751v1 Announce Type: new Abstract: Geospatial foundation models (GeoFMs), pretrained on large-scale geospatial data such as Earth observation (EO), climate, and weather data, have shown p

Structured Memory for Edge Language Models: Persistent Context and Corpus Retrieval via O(1) SSM State Injection

Model ReleasesDGX agent

arXiv:2608.02560v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) imposes a prefill cost proportional to retrieved context length, and -- with Transformer backbones -- a KV-cache th

Today, we’re launching Alpamayo 2 Super, our frontier open reasoning model for autonomous vehicles. Beyond seeing, Alpamayo understands and …

SafetyDGX agent

Today, we’re launching Alpamayo 2 Super, our frontier open reasoning model for autonomous vehicles. Beyond seeing, Alpamayo understands and reasons through the complex world - thinks before it acts. I

Understanding Synergistic Interactions among Pathology Foundation Models via Adaptive Fusion

ResearchDGX agent

arXiv:2608.01370v1 Announce Type: new Abstract: Pathology foundation models (PFMs) provide strong tile-level representations via self-supervised pre-training on large-scale pathology images. Yet, PFMs

3 Aug 2026

AFAIK the most significant breakthrough since 2017 besides scaling old ideas was broadening from base models into larger systems that incorp…

Model ReleasesDGX agent

AFAIK the most significant breakthrough since 2017 besides scaling old ideas was broadening from base models into larger systems that incorporate symbol/manipulating entities like harnesses, tools, an

An internal version of our next major model produced 10 new results on long-standing open problems in mathematics and theoretical computer s…

Model ReleasesDGX agent

An internal version of our next major model produced 10 new results on long-standing open problems in mathematics and theoretical computer science, using roughly $2,000 worth of tokens at GPT-5.6 Sol

CAER: Conflict-Aware Evidence Routing with Dual Prefix Experts for Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2607.28991v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have demonstrated remarkable capabilities in multimodal understanding and generation. However, when textual inp

Epistemic-aware Vision-Language Foundation Model for Fetal Ultrasound Interpretation

ResearchDGX agent

arXiv:2510.12953v4 Announce Type: replace-cross Abstract: Recent medical vision-language models have shown promise on tasks such as VQA, report generation, and anomaly detection. However, most are ada

Faster but Different: Diagnosing and Controlling Content Drift in Accelerated Multimodal Diffusion Language Models

SafetyDGX agent

arXiv:2607.29079v1 Announce Type: new Abstract: Training-free acceleration makes diffusion-based multimodal large language models (dMLLMs) more deployable, but it may silently change generated content

In-situ Autoguidance: Eliciting Self-Correction in Diffusion Models

SafetyDGX agent

arXiv:2510.17136v2 Announce Type: replace Abstract: The generation of high-quality, diverse, and prompt-aligned images is a central goal in image-generating diffusion models. The popular classifier-fr

// Model or Harness // Great paper if you are building with agents in production. (bookmark it) It organizes 41 agent failure modes by the i…

AgentsDGX agent

// Model or Harness // Great paper if you are building with agents in production. (bookmark it) It organizes 41 agent failure modes by the interaction they originate in. Each mode gets assigned to an

MOT-SR: Multi-Objective Tool-Augmented Scientific Equation Discovery with Large Language Models

Local AiDGX agent

arXiv:2607.29561v1 Announce Type: cross Abstract: Symbolic Regression (SR) aims to discover analytical equations from observational data and plays a central role in scientific modeling. While recent L

Parameter-Efficient Fine-Tuning for Spiking Point Cloud Models

Model ReleasesDGX agent

arXiv:2607.29048v1 Announce Type: new Abstract: Spiking Neural Networks (SNNs) offer energy-efficient solutions for point cloud analysis on resource-constrained devices through event-driven computatio

PiDDM: Physics-Informed Differentiable Degradation Modeling for Lithium-Ion Battery State-of-Health Prediction

ResearchDGX agent

arXiv:2607.29095v1 Announce Type: new Abstract: Accurate prediction of lithium-ion battery state of health (SOH) is essential for reliable energy storage operation. However, purely data-driven models

Reasoning in Real World Clinical Care: Why Large Language Models Are Not Yet Safe for Autonomous Clinical Decision Support

SafetyDGX agent

arXiv:2607.28677v1 Announce Type: new Abstract: LLM now pass medical licensing examinations and, in curated cases, can rival physicians at diagnostic reasoning. These developments have accelerated the

.@ssankar says Palantir was able to make Nvidia's Nemotron Ultra model 'better than frontier': 'I literally almost felt gaslit when, within …

Model ReleasesDGX agent

.@ssankar says Palantir was able to make Nvidia's Nemotron Ultra model 'better than frontier': 'I literally almost felt gaslit when, within 24 hours of getting Nemotron up with no post-training, this

ST-WAM: Semantic-Temporal World Action Model for Robust Manipulation under Visual Distribution Shifts

ApplicationsDGX agent

arXiv:2607.28993v1 Announce Type: cross Abstract: World Action Models (WAMs) have emerged as a promising paradigm by jointly modeling robot actions and future visual dynamics. However, their reliance

The Asymmetric Effects of Knowledge Distillation on Bias in Small Language Models

Model ReleasesDGX agent

arXiv:2607.28639v1 Announce Type: cross Abstract: We show that knowledge distillation in small instruction-tuned language models has asymmetric effects on bias. On unambiguous tasks (BBQ-disambig), re

This is a wild result. Locus, the automated research system from @intology, post-trained Qwen3 base models that beat the official human-tune…

ApplicationsDGX agent

This is a wild result. Locus, the automated research system from @intology, post-trained Qwen3 base models that beat the official human-tuned Qwen3 1.7B Instruct release. SoTA on PostTrainBench! The m

UltraSAM3: A Concept-Driven Foundation Model for Universal Ultrasound Image Segmentation

AgentsDGX agent

arXiv:2607.29200v1 Announce Type: new Abstract: Ultrasound imaging has become increasingly widespread in clinical practice due to its portability, low cost and real-time capability, making ultrasound

2 Aug 2026

I built an open-source LLM Gateway to route and fallback between local Ollama models and cloud APIs

Model ReleasesDGX agent

Hey r/ollama 👋 If you run Ollama locally alongside cloud endpoints for agent workflows, Cursor/Windsurf, or custom scripts, managing API switching, failover logic, and context limits can get messy fas

31 Jul 2026

As shocking as the Kimi K3 release. Massive performance gain was just with post-training Model is 3x smaller than GLM 5.2 (10x smaller than …

Model ReleasesDGX agent

As shocking as the Kimi K3 release. Massive performance gain was just with post-training Model is 3x smaller than GLM 5.2 (10x smaller than K3) & works on a MacBook / Spark This is Q1 flagship (Opus 4

Auditing Question-Order Effects in Large Language Models with the QQ Equality: Mechanism Characterization and a Saturation Caveat

Model ReleasesDGX agent

arXiv:2607.17219v2 Announce Type: replace Abstract: Question-order effects in human survey data have been reported to approximately satisfy the QQ (quantum question) equality, a parameter-free predict

Bridging the Gap in Ophthalmic AI: MM-Retinal-Reason Dataset and OphthaReason Model toward Dynamic Multimodal Reasoning

TutorialsDGX agent

arXiv:2508.16129v3 Announce Type: replace Abstract: Multimodal large language models (MLLMs) have recently demonstrated remarkable reasoning abilities with reinforcement learning paradigm. Although se

Deepseek V4 Flash is now ~#2 open weight model to Kimi K3 and >50x cheaper

Model ReleasesDGX agent

https://preview.redd.it/h7zv5tb3tmgh1.png?width=2854&format=png&auto=webp&s=507380e8f862c18f10f7c5c84da9e8d1c59139b0 Deepseek's new flash model is unexpectedly cheap and high-performing across useful

Introducing Qwen-Audio-3.0-ASR-Flash: More context-aware. Stronger domain-term recognition. 🚀Our latest ASR model upgrades: • Context consi…

Model ReleasesDGX agent

Introducing Qwen-Audio-3.0-ASR-Flash: More context-aware. Stronger domain-term recognition. 🚀Our latest ASR model upgrades: • Context consistency • Domain-term recognition • Custom hotwords • Speech p

LLM2Vec-Gen: Generative Embeddings from Large Language Models

Model ReleasesDGX agent

arXiv:2603.10913v3 Announce Type: replace Abstract: Fine-tuning LLM-based text embedders via contrastive learning maps inputs and outputs into a new representational space, discarding the LLM's output

Neural Network-Assisted CLEAN for Channel Modeling in Low-SNR Regimes

Model ReleasesDGX agent

arXiv:2607.27450v1 Announce Type: new Abstract: Accurate multipath parameter estimation is critical for modern wireless communication systems, particularly in challenging low-SNR environments. Traditi

OpenAI’s entire growth loop: new models, price cuts, and Tibo’s token resets

Model ReleasesDGX agent

OpenAI’s entire growth loop: new models, price cuts, and Tibo’s token resets major price cuts today: *80% drop for GPT-5.6 Luna, now 0.20 per million input tokens and 1.20 per million output *20% drop

Position, Not Provenance: Separating Reasoning Mediation from Sycophancy in Medical Vision-Language Models

ResearchDGX agent

arXiv:2607.27304v1 Announce Type: cross Abstract: Medical vision-language models (VLMs) generate chain-of-thought (CoT) reasoning before answering clinical questions, but whether this reasoning causal

QQWorld: Quantile-Quantile Matching for World Model Regularization

SafetyDGX agent

arXiv:2607.28415v1 Announce Type: cross Abstract: Latent world models enable efficient planning by predicting future states in a compact representation space, but their performance depends critically

State-Dependent Safety Failures in Multi-Turn Language Model Interaction

SafetyDGX agent

arXiv:2603.15684v2 Announce Type: replace-cross Abstract: Safety alignment in large language models is typically evaluated under isolated queries, yet real-world use is inherently multi-turn. Although

What’s new in AI infrastructure and orchestration this month

Model ReleasesDGX agent

At Google, AI is a soup-to-nuts endeavor. Obviously, we make leading AI models like Gemini and Nano Banana. We incorporate AI into the tools you use every day (think Gmail, BigQuery, AlloyDB, Google C

Why are AI model tests always the same generic prompts?

Model ReleasesDGX agent

Okay, hear me out. Why is it that every time a new model comes out, all the tests I see are 'make a car game,' 'make a website,' or something equally generic, usually from a prompt that's barely a lin

30 Jul 2026

Anatomy Contextualized Adaption of CT Foundation Models

SafetyDGX agent

arXiv:2607.27154v1 Announce Type: new Abstract: CT vision-language foundation models have demonstrated promising performance across downstream tasks, but are typically trained with whole-volume repres

Dual Inversion for Text-to-Image Diffusion Models: From Both Prompt and Noise Perspectives

SafetyDGX agent

arXiv:2607.26735v1 Announce Type: new Abstract: Prompt inversion, as a typical reverse engineering technique, enables text-to-image (T2I) diffusion models to generate the desired target images without

From Found to Designed: Concepts as a Design Axis for Large Language Models

SafetyDGX agent

arXiv:2607.26825v1 Announce Type: new Abstract: Large language models (LLMs) encode rich concept-like information, but represent it implicitly through distributed statistical associations rather than

Rad-JEPA 3D: Radiology Joint-Embedding Predictive Model for 3D Computed Tomography

Model ReleasesDGX agent

arXiv:2607.26196v1 Announce Type: new Abstract: Self-supervised pretraining is central to 3D medical image analysis, where unlabeled CT volumes are abundant but expert annotations are scarce. Yet exis

See2Think: Do Multimodal Models Really Use Intermediate Visual States?

ApplicationsDGX agent

arXiv:2607.26769v1 Announce Type: new Abstract: Multimodal large language models increasingly use sketches, annotations, tools, and intermediate images during reasoning, but it remains unclear whether

Try Again, Don't Look Back: Blind Resampling Outperforms Self-Repair in Small Code Models

Local AiDGX agent

arXiv:2607.26117v1 Announce Type: cross Abstract: Self-repair - returning a failed program to the model together with its test output and asking for a correction - is a standard component of code agen

Two Calls Beat Five Agents: Evaluating Multi-Agent Pipelines Against Self-Refinement for Local Language Models

Model ReleasesDGX agent

arXiv:2607.26922v1 Announce Type: new Abstract: Multi-agent LLM pipeline systems break down the task among multiple roles for better reasoning, but are benchmarked mainly with large-scale commercial m

What Can Latent World Models Know? Physical Parameter Identifiability in Multimodal Predictive Representations

Model ReleasesDGX agent

arXiv:2607.27017v1 Announce Type: new Abstract: A central premise of latent world models is that predicting the future forces a representation to internalize the physics of its environment. Which phys

29 Jul 2026

Accurate structural modeling of chemically diverse molecular interfaces with Vilya-2

ResearchDGX agent

arXiv:2607.25156v1 Announce Type: new Abstract: Structure-prediction networks built on co-evolutionary statistics have transformed protein-based drug discovery, yet their accuracy does not extend to p

BREAKING: SpaceXAI's newly released Grok Voice Think Fast 2.0 beats voice models from OpenAI, Google, Alibaba, and DeepSlate in the Artifici…

Model ReleasesDGX agent

SpaceXAI has released its new Grok Voice Think Fast 2.0, which on the Artificial Analysis Speech‑to‑Speech benchmark outperformed leading models from OpenAI, Google, Alibaba and DeepSlate. The claim w

CaRE Compute-aware Remasking Evaluation Protocol for Masked Diffusion Language Models

ResearchDGX agent

arXiv:2607.24763v1 Announce Type: new Abstract: Masked diffusion language models (MDLMs) are advancing rapidly, yet the evaluation standards needed to reliably interpret their progress have not kept p

Empirical Evaluation of Out-Of-Distribution Performance of Tabular Foundation Models

ApplicationsDGX agent

arXiv:2607.26000v1 Announce Type: cross Abstract: Tabular Foundation Models (TFMs) have emerged as novel approaches for tabular predictive tasks, demonstrating competitive predictive performance to en

GrocLM: Grocery Category Recommendation in E-Commerce with Large Language Models

Model ReleasesDGX agent

arXiv:2607.24764v1 Announce Type: new Abstract: The rapid growth of online grocery shopping requires recommendation systems that capture cyclical purchasing behavior and diverse user intents. Traditio

Med-SegLens: Latent-Level Model Diffing for Interpretable Medical Image Segmentation

SafetyDGX agent

arXiv:2602.10508v2 Announce Type: replace Abstract: Modern segmentation models achieve strong predictive performance but remain largely opaque, limiting our ability to diagnose failures, understand da

Physics of Language Models: Part 4.1, Architecture Design and the Magic of Canon Layers

ApplicationsDGX agent

arXiv:2512.17351v2 Announce Type: replace Abstract: Understanding architectural differences in language models is challenging, especially at academic-scale pretraining (e.g., 1.3B parameters, 100B tok

SAM3D-Guided Object-Centric Representation Alignment for Vision-Language-Action Models

Local AiDGX agent

arXiv:2607.25912v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have shown strong potential for general robot manipulation, but most existing models rely on 2D visual-language ba

Sharpness-aware Model Merging with Salience Recovery for LLM-based Cross-Domain Sequential Recommendation

Model ReleasesDGX agent

arXiv:2607.25366v1 Announce Type: cross Abstract: LLM-based Cross-Domain Sequential Recommendation (CDSR) leverages LLMs to enhance target performance via deep semantic reasoning, alleviating the depe

← Previous
1…99100101102103…1009
Next →