AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,483
  • Agents7,562
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,175
  • Local Ai4,939
  • Model Releases23,953
  • Research20,128
  • Safety13,373
  • Syntheses17
  • Tools1,677
  • Tutorials3,405

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,483
  • Agents7,562
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,175
  • Local Ai4,939
  • Model Releases23,953
  • Research20,128
  • Safety13,373
  • Syntheses17
  • Tools1,677
  • Tutorials3,405

Source
HumanDGX agent

88,483Total entries
1Added by human
88,482Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,694 results
4 Aug 2026

Sources and filings: Google assembled a ~200B financing program for Anthropic, with 150B+ tied to TPUs and involving Broadcom, Blackstone, Apollo, and others (Financial Times)

Model ReleasesDGX agent

Financial Times: Sources and filings: Google assembled a ~200B financing program for Anthropic, with 150B+ tied to TPUs and involving Broadcom, Blackstone, Apollo, and others — Private credit, chip le

SPAE: Spectrally Guided Autoencoder for Pretrained Visual Latents

SafetyDGX agent

arXiv:2608.01306v1 Announce Type: new Abstract: Latents from vision foundation models (VFMs) are semantically rich and well suited for visual understanding. Recent representation autoencoder methods s

SPARE: Structural Parameter-Free Affinity Regularization for Flow Matching

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2608.01990v1 Announce Type: new Abstract: Denoising diffusion transformers achieve strong generation quality but converge slowly during training. Regularizing their internal representations has

SSR: Similarity-Shift Refinement for Training-Free Object-Centric Masks

ApplicationsDGX agent

arXiv:2608.01103v1 Announce Type: new Abstract: Object-centric models often produce fragmented masks, boundary leakage, and incorrect region merging. We introduce Similarity-Shift Refinement (SSR), a

Structured Proxy Features for Multimodal NSCLC Survival Prediction from Pretreatment CT

Model ReleasesDGX agent

arXiv:2608.00446v1 Announce Type: new Abstract: Lung cancer results in roughly 1.8 million fatalities annually worldwide, with non-small cell lung cancer (NSCLC) comprising the majority of cases. Desp

Tevatron Meets Megatron: Expert-Parallel LLM Reranker Training on an Academic Budget

Model ReleasesDGX agent

arXiv:2608.00916v1 Announce Type: cross Abstract: Modern reranking recipes---billion-scale cross-encoders, mixture-of-experts (MoE) backbones, and distillation against strong teachers---have outpaced

To date, finding a drug has been a process of guess & check… screening millions of molecules hoping one binds. @chaidiscovery is changing th…

IndustryDGX agent

To date, finding a drug has been a process of guess & check… screening millions of molecules hoping one binds. @chaidiscovery is changing the paradigm… describe the molecule you want, and the model de

TRAM: Enhancing Multimodal Reasoning with Trajectory-Derived Auxiliary Memory

ResearchDGX agent

arXiv:2608.01922v1 Announce Type: new Abstract: Multimodal Large Reasoning Models (MLRMs) have achieved strong performance on tasks requiring visual understanding and multi-step inference. However, as

UDT: Reconciling U-Nets and Diffusion Transformers with Data-Adaptive Token Reduction

SafetyDGX agent

arXiv:2608.01298v1 Announce Type: new Abstract: Diffusion Transformers (DiTs) have emerged as a core architecture in generative modeling due to their scalability and adaptability to multimodal tasks.

VespaSeg: A Resource-Aware Ground-then-Segment Pipeline for Referring Expression Segmentation

Local AiDGX agent

arXiv:2608.01077v1 Announce Type: new Abstract: Referring expression segmentation requires language conditioned localization and pixel-accurate masks, but monolithic models can be costly to deploy. We

When Measurement Conventions Masquerade as Calibration Gains in Cardiac Digital Twins

Model ReleasesDGX agent

arXiv:2608.01602v1 Announce Type: new Abstract: Cardiac digital twins convert clinical images into physiological measurements through observation operators, yet calibration studies often assume a fixe

Who Belongs in the Eval Set? A Capability-Taxonomy-Driven Pipeline for Curating Regression Eval Sets in Agent-Extensibility Platforms

Model ReleasesDGX agent

arXiv:2608.01004v1 Announce Type: new Abstract: Platform teams hosting agent-extensibility surfaces face a regression-economics paradox: every onboarding customer ships an evaluation set tuned to thei

Why are Gamers so incredibly hostile to AI? Is it just a tiny vocal minority that spreads such toxic vitriol online?

Model ReleasesDGX agent

It's more accurate to say that many highly engaged online gamers are hostile to AI, not that 'gamers' as a whole are. Gaming is a huge community with hundreds of millions of people, and opinions vary

Why Formal Monitors Fail: Attack Distribution Entropy as a Coverage Bound for LTL-Based LLM Agent Safety

Model ReleasesDGX agent

arXiv:2608.01388v1 Announce Type: cross Abstract: Runtime safety monitors based on Linear Temporal Logic (LTL) and finite automata (FSA) are increasingly deployed to intercept unsafe tool-call sequenc

3 Aug 2026

Adaptive Emotional Video Captioning via Affective Heterogeneous Graph Reasoning and Multi-task Joint Learning

SafetyDGX agent

arXiv:2607.29045v1 Announce Type: new Abstract: Emotional video captioning (EVC) aims to describe a video with both factual correctness and affective expressiveness. It requires a model to perceive su

Adaptive Policy Backbone via Shared Network

Model ReleasesDGX agent

arXiv:2509.22310v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) has achieved impressive results across domains, yet learning an optimal policy typically requires extensive intera

Appreciate it! Now let's understand the world through the eyes of Qwen3.8. 🥳

Model ReleasesDGX agent

Qwen3.8-Max from Alibaba’s Qwen team achieved second place in the Vision Arena benchmark, scoring 1,305 points. It trails only Claude Fable 5 (High), which leads by a slim 13‑point margin. The post un

Curriculum Matters: Data-Efficient Relational PFN Pretraining with Synthetic Data

Model ReleasesDGX agent

arXiv:2607.29120v1 Announce Type: new Abstract: Relational Prior-Data Fitted Networks (PFNs) such as RDB-PFN approximate Bayesian inference over multi-table relational databases by pretraining on mill

Efficient LLM Adversarial Training via Low-Rank Defense and Circuit-Guided Surrogates

ResearchDGX agent

arXiv:2607.28959v1 Announce Type: cross Abstract: Adversarial training is one of the most effective defenses against adversarial attacks, yet the computational cost remains prohibitive at modern scale

EMAG: Self-Rectifying Diffusion Sampling with Exponential Moving Average Guidance

ResearchDGX agent

arXiv:2512.17303v2 Announce Type: replace Abstract: In diffusion and flow-matching generative models, guidance techniques are widely used to improve sample quality and consistency. Classifier-free gui

Empowering Cross-Domain Sequential Recommendation with Hybrid Tokenization and Serial-Parallel Decoding

ResearchDGX agent

arXiv:2607.28659v1 Announce Type: new Abstract: Cross-domain sequential recommendation (CDSR) aims to model users' dynamic interest transitions and sequential patterns across multiple domains. Recentl

Explaining AI-Image Detection: What the Heatmap Actually Shows

ResearchDGX agent

arXiv:2607.29581v1 Announce Type: new Abstract: A marketplace review photograph is a document: platforms approve refunds on it, and generative models drove the cost of forging one to zero. We study th

Extrapolating the emergence of Hamiltonian chaos with random-feature Hamiltonian neural networks

Model ReleasesDGX agent

arXiv:2607.28977v1 Announce Type: cross Abstract: Machine learning of Hamiltonian dynamics has driven growing interest in Hamiltonian neural networks (HNNs), which encode Hamilton's equations of motio

FineInstructions: Scaling Synthetic Instructions to Pre-Training Scale

ResearchDGX agent

arXiv:2601.22146v3 Announce Type: replace Abstract: Due to limited supervised training data, large language models (LLMs) are typically pre-trained via a self-supervised 'predict the next word' object

GEMSS: A Variational Method for Discovering Multiple Sparse Solutions in Classification and Regression Problems

Model ReleasesDGX agent

arXiv:2602.08913v3 Announce Type: replace Abstract: In underdetermined regression and classification problems, multiple feature subsets often yield equivalent predictive performance. In applied settin

GO-PRE: Goal-Oriented Next-Best-View Selection via Predictive Rendering Entropy for Active 3D Reconstruction

Model ReleasesDGX agent

arXiv:2607.29037v1 Announce Type: new Abstract: Active 3D reconstruction relies on active view selection to maximize reconstruction fidelity under limited capture budgets. However, most existing metho

Improving scDiffusion with Sparsity-Biased Classifier-Free Guidance

ResearchDGX agent

arXiv:2607.29043v1 Announce Type: cross Abstract: Single-cell RNA sequencing (scRNA-seq) has become an essential tool in modern cellular biology, and generating accurate synthetic scRNA-seq data is be

Information Processing by Neuron Populations in the Central Nervous System: A Theory of the Mathematical Structure of Data and Operations

ResearchDGX agent

arXiv:2309.02332v3 Announce Type: replace-cross Abstract: In the mammalian central nervous system, neurons are organized into populations communicating by spike trains propagating along axonal bundles

Know It, Act on It: Investigating Memory Utilization in LLM Personalization

AgentsDGX agent

arXiv:2607.29433v1 Announce Type: new Abstract: As large language model (LLM) agents evolve into personalized companions, memory has emerged as a core capability. However, LLMs face a knowledge utiliz

Learning Optimal Dynamic Matching via Graph Neural Networks

Model ReleasesDGX agent

arXiv:2607.28925v1 Announce Type: new Abstract: Dynamic matching markets require decisions about whom to match and when: matching now yields value but removes participants who may create better future

Learning Stateful Predictive Knowledge From Experience

SafetyDGX agent

arXiv:2607.28638v1 Announce Type: new Abstract: As large language model (LLM) agents increasingly learn from experience, they primarily rely on trajectory-level reflection to extract insights. Viewed

Leveraging Image Generators to Address Data Scarcity: The Gen4Regen Dataset for Forest Regeneration Mapping

ApplicationsDGX agent

arXiv:2605.05627v2 Announce Type: replace-cross Abstract: Sustainable forest management relies on precise species composition mapping, yet traditional ground surveys are labour-intensive and geographi

LightningRL: Breaking the Accuracy-Parallelism Trade-off of Block-wise dLLMs via Reinforcement Learning

Local AiDGX agent

arXiv:2603.13319v2 Announce Type: replace Abstract: Diffusion Large Language Models (dLLMs) have emerged as a promising paradigm for parallel token generation, with block-wise variants garnering signi

M3MAD-Bench: Multi-Dimensional Evaluation of Multi-Agent Debate Across Domains and Modalities

Model ReleasesDGX agent

arXiv:2601.02854v2 Announce Type: replace Abstract: As an agent-level reasoning and coordination paradigm, Multi-Agent Debate (MAD) orchestrates multiple agents through structured debate to improve an

MAGA: Multi-Platform Self-Fusion of GUI Agents via Structured Action Distillation

SafetyDGX agent

arXiv:2607.29320v1 Announce Type: new Abstract: Graphical user interface (GUI) agents based on large language models are increasingly deployed across mobile, web, and desktop environments. However, ex

MerchantBench: Benchmarking LLM Agents for Long-Term Coherence in E-Commerce Operations

AgentsDGX agent

arXiv:2607.28956v1 Announce Type: new Abstract: Large language model agents are increasingly evaluated as autonomous tool users, yet most benchmarks focus on bounded tasks with immediate success crite

Open-Source LLM-Driven Formal Verification: A Multi-Agent Pipeline for RTL Repair

Model ReleasesDGX agent

arXiv:2607.28877v1 Announce Type: cross Abstract: Verification consumes the majority of modern chip design effort, yet the formal verification tools that provide mathematical guarantees of correctness

Parameter-Free Heavy-Tailed Bandits

Model ReleasesDGX agent

arXiv:2607.29460v1 Announce Type: new Abstract: Heavy-tailed distributions arise naturally in sequential decision-making problems such as financial investment, online advertising, and network manageme

Qwen 3.8 Max and MiniMax-H3 within hours of each other

Model ReleasesDGX agent

Simon Willison noted that Qwen 3.8 Max and MiniMax‑H3 were released within hours of each other. The MiniMax team announced that MiniMax‑H3 is now publicly available on Hugging Face (https://huggingfac

RAPiD: Reward-Guided Consistency Distillation of Diffusion Planners for Real-Time Autonomous Driving

SafetyDGX agent

arXiv:2602.07339v2 Announce Type: replace Abstract: Diffusion-based trajectory planners can model multi-modal driving behavior, but their iterative denoising process introduces a latency bottleneck fo

Reflection or Re-Generation? Why LLM Revision Fails Where Human Revision Succeeds

Local AiDGX agent

arXiv:2607.28908v1 Announce Type: new Abstract: Reflection, the ability to revisit and revise prior reasoning, is central to how humans improve their answers. Large language models (LLMs) are increasi

Representations from Pretrained Machine-Learning Interatomic Potentials as Coarse Coordinates for Material Generation and Evaluation

ResearchDGX agent

arXiv:2607.28776v1 Announce Type: new Abstract: Generative machine learning is increasingly used for inorganic crystal structure generation. Most models and the corresponding evaluation approaches rel

Revisiting Multi-Permutation Equivariance through the Lens of Irreducible Representations

Model ReleasesDGX agent

arXiv:2410.06665v4 Announce Type: replace-cross Abstract: This paper explores the characterization of equivariant linear layers for representations of permutations and related groups. Unlike tradition

SciFigPlag-Bench: A Benchmark for Provenance-Aware Scientific Figure Plagiarism Detection

Model ReleasesDGX agent

arXiv:2607.29124v1 Announce Type: new Abstract: Scientific figures often encode the visual evidence behind scientific findings, yet figure plagiarism remains underexplored as a benchmarked multimodal

Sensitivity Analysis of GRU, LSTM and Transformer Encoder in Classification of Automated Driving Systems

SafetyDGX agent

arXiv:2607.28665v1 Announce Type: cross Abstract: Automated driving systems (ADSs) are becoming ubiquitous. Future Software Defined Vehicles (SDVs) may be able to run multiple ADSs, both native and af

Stem: Rethinking Causal Information Flow in Sparse Attention

ResearchDGX agent

arXiv:2603.06274v2 Announce Type: replace-cross Abstract: The quadratic computational complexity of self-attention remains a fundamental bottleneck for scaling Large Language Models (LLMs) to long con

Studying quantization trade-offs for efficient inference deployment in machine translation

HardwareDGX agent

arXiv:2607.29397v1 Announce Type: new Abstract: Deploying large language models in realistic server environments poses challenges, as the system needs to provide high-quality responses with low latenc

TraceViT: Grounded Trace Supervision for Visual Abstract Reasoning

SafetyDGX agent

arXiv:2607.29586v1 Announce Type: cross Abstract: The Abstraction and Reasoning Corpus (ARC) tests whether a model can infer an unseen transformation from a few input-output examples and apply it to a

Was the release of deepseek v4 flash planned to take spotlight against 5.6 luna?

Model ReleasesDGX agent

Id figured since they first emailed people about api price changes coming mid july then delayed the v4 flash release to late july, I wonder if they delayed it for the sake of stealing spotlight from o

White House invites AI companies to review its new AI safety framework

Model ReleasesDGX agent

Cybersecurity chiefs at the White House have reportedly finalized the outline of a forthcoming framework that will enable artificial intelligence companies to voluntarily submit their latest frontier

WitCert: Sound Runtime Risk Observability and Gating for KV-Cache Quantization

Model ReleasesDGX agent

arXiv:2607.28699v1 Announce Type: cross Abstract: KV-cache quantization is validated today by offline benchmark averages; a deployed system cannot tell whether compression is damaging the request it i

2 Aug 2026

Alibaba says its 2.4T-parameter Qwen3.8-Max tops Moonshot's Kimi K3 on some benchmarks, and it plans to release Qwen3.8-Max and Qwen3.8-27B's weights next week (Luz Ding/Bloomberg)

Model ReleasesDGX agent

Luz Ding / Bloomberg: Alibaba says its 2.4T-parameter Qwen3.8-Max tops Moonshot's Kimi K3 on some benchmarks, and it plans to release Qwen3.8-Max and Qwen3.8-27B's weights next week — Alibaba Group Ho

DeepSeek-V4-Flash-0731: When Low is higher than High

Model ReleasesDGX agent

I decided to test a few questions against DeepSeek-V4-Flash-0731. Locally, I was running Unsloth's UD-Q2_K_XL quant. After I saw the surprising shape of the results, I tested against DeepSeek's offici

Fascinating to see @ClementDelangue, CEO of @huggingface speaking on @FaceTheNation. Excellent points and solid advocacy around the benefits…

AgentsDGX agent

Fascinating to see @ClementDelangue, CEO of @huggingface speaking on @FaceTheNation. Excellent points and solid advocacy around the benefits of open AI models, which helped him defend against a rogue

1 Aug 2026

V4 flash vs V4 Flash (0731). Guys, new DeepSeek V4 Flash(0731) is now free on InferX

Model ReleasesDGX agent

DeepSeek V4 Flash is now available on InferX, and it’s free to use. We’re continuing to add GPU capacity as demand grows. While we’re bringing additional capacity online, you may occasionally see high

31 Jul 2026

AfriEconQA: A Benchmark for Quantitative and Temporal Reasoning over World Bank Economic Reports

Model ReleasesDGX agent

arXiv:2601.15297v3 Announce Type: replace Abstract: Reliable question answering over long institutional documents requires more than topical retrieval: a system must localize the exact passage that su

AgentMap: Joint Equivalence and Subsumption Discovery for Ontology Matching

Model ReleasesDGX agent

arXiv:2607.27130v1 Announce Type: new Abstract: Ontology matching (OM) has traditionally been formulated as either equivalence discovery or subsumption matching. The existing OM systems identify only

AI DOOMERS BE LIKE: 'GLM 5.1 WILL WIPE OUT HUMANS IN 2030'

Local AiDGX agent

I swear some AI doomers have never actually used a local model. They watched one flashy keynote, one YouTube thumbnail with a guy making this face 😱, read three headlines, and suddenly civilization is

Auto Research for Materials: Auditable AI-Scientist Workflows with Held-Out Transfer

Local AiDGX agent

arXiv:2607.17100v2 Announce Type: replace-cross Abstract: Auto Research uses language-model agents to propose, implement, and evaluate machine-learning changes in a closed loop, but is usually judged

b10206

Model ReleasesDGX agent

llama : enforce the same K and V cache types for DeepSeek V4; enable FA if V cache is quantized (#25871) llama : enforce the same K and V cache types for DeepSeek V4; enable FA if V cache is quantized

← Previous
1…511512513514515…1062
Next →