AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,914
  • Agents7,747
  • Applications5,533
  • Concepts5
  • Hardware1,920
  • Industry6,196
  • Local Ai5,094
  • Model Releases24,727
  • Research20,781
  • Safety13,739
  • Syntheses17
  • Tools1,678
  • Tutorials3,477

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,914
  • Agents7,747
  • Applications5,533
  • Concepts5
  • Hardware1,920
  • Industry6,196
  • Local Ai5,094
  • Model Releases24,727
  • Research20,781
  • Safety13,739
  • Syntheses17
  • Tools1,678
  • Tutorials3,477

Source
HumanDGX agent

Content type
AllBlog
90,914Total entries
1Added by human
90,913Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,676 results
Model Releases

Adaptive Policy Backbone via Shared Network

DGX agent

arXiv:2509.22310v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) has achieved impressive results across domains, yet learning an optimal policy typically requires extensive intera

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

Appreciate it! Now let's understand the world through the eyes of Qwen3.8. 🥳

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

Qwen3.8-Max from Alibaba’s Qwen team achieved second place in the Vision Arena benchmark, scoring 1,305 points. It trails only Claude Fable 5 (High), which leads by a slim 13‑point margin. The post un

model-releasesqwen--x
3 Aug 2026
Model Releases

Curriculum Matters: Data-Efficient Relational PFN Pretraining with Synthetic Data

DGX agent

arXiv:2607.29120v1 Announce Type: new Abstract: Relational Prior-Data Fitted Networks (PFNs) such as RDB-PFN approximate Bayesian inference over multi-table relational databases by pretraining on mill

model-releasesarxiv-cs-lg
3 Aug 2026
Research

Efficient LLM Adversarial Training via Low-Rank Defense and Circuit-Guided Surrogates

DGX agent

arXiv:2607.28959v1 Announce Type: cross Abstract: Adversarial training is one of the most effective defenses against adversarial attacks, yet the computational cost remains prohibitive at modern scale

researcharxiv-cs-ai
3 Aug 2026
Research

EMAG: Self-Rectifying Diffusion Sampling with Exponential Moving Average Guidance

DGX agent

arXiv:2512.17303v2 Announce Type: replace Abstract: In diffusion and flow-matching generative models, guidance techniques are widely used to improve sample quality and consistency. Classifier-free gui

researcharxiv-cs-cv
3 Aug 2026
Research

Empowering Cross-Domain Sequential Recommendation with Hybrid Tokenization and Serial-Parallel Decoding

DGX agent

arXiv:2607.28659v1 Announce Type: new Abstract: Cross-domain sequential recommendation (CDSR) aims to model users' dynamic interest transitions and sequential patterns across multiple domains. Recentl

researcharxiv-cs-ai
3 Aug 2026
Research

Explaining AI-Image Detection: What the Heatmap Actually Shows

DGX agent

arXiv:2607.29581v1 Announce Type: new Abstract: A marketplace review photograph is a document: platforms approve refunds on it, and generative models drove the cost of forging one to zero. We study th

researcharxiv-cs-cv
3 Aug 2026
Model Releases

Extrapolating the emergence of Hamiltonian chaos with random-feature Hamiltonian neural networks

DGX agent

arXiv:2607.28977v1 Announce Type: cross Abstract: Machine learning of Hamiltonian dynamics has driven growing interest in Hamiltonian neural networks (HNNs), which encode Hamilton's equations of motio

model-releasesarxiv-cs-lg
3 Aug 2026
Research

FineInstructions: Scaling Synthetic Instructions to Pre-Training Scale

DGX agent

arXiv:2601.22146v3 Announce Type: replace Abstract: Due to limited supervised training data, large language models (LLMs) are typically pre-trained via a self-supervised 'predict the next word' object

researcharxiv-cs-cl
3 Aug 2026
Model Releases

GEMSS: A Variational Method for Discovering Multiple Sparse Solutions in Classification and Regression Problems

DGX agent

arXiv:2602.08913v3 Announce Type: replace Abstract: In underdetermined regression and classification problems, multiple feature subsets often yield equivalent predictive performance. In applied settin

model-releasesarxiv-cs-lg
3 Aug 2026
Model Releases

GO-PRE: Goal-Oriented Next-Best-View Selection via Predictive Rendering Entropy for Active 3D Reconstruction

DGX agent

arXiv:2607.29037v1 Announce Type: new Abstract: Active 3D reconstruction relies on active view selection to maximize reconstruction fidelity under limited capture budgets. However, most existing metho

model-releasesarxiv-cs-cv
3 Aug 2026
Research

Improving scDiffusion with Sparsity-Biased Classifier-Free Guidance

DGX agent

arXiv:2607.29043v1 Announce Type: cross Abstract: Single-cell RNA sequencing (scRNA-seq) has become an essential tool in modern cellular biology, and generating accurate synthetic scRNA-seq data is be

researcharxiv-cs-ai
3 Aug 2026
Research

Information Processing by Neuron Populations in the Central Nervous System: A Theory of the Mathematical Structure of Data and Operations

DGX agent

arXiv:2309.02332v3 Announce Type: replace-cross Abstract: In the mammalian central nervous system, neurons are organized into populations communicating by spike trains propagating along axonal bundles

researcharxiv-cs-ai
3 Aug 2026
Agents

Know It, Act on It: Investigating Memory Utilization in LLM Personalization

DGX agent

arXiv:2607.29433v1 Announce Type: new Abstract: As large language model (LLM) agents evolve into personalized companions, memory has emerged as a core capability. However, LLMs face a knowledge utiliz

agentsarxiv-cs-cl
3 Aug 2026
Model Releases

Learning Optimal Dynamic Matching via Graph Neural Networks

DGX agent

arXiv:2607.28925v1 Announce Type: new Abstract: Dynamic matching markets require decisions about whom to match and when: matching now yields value but removes participants who may create better future

model-releasesarxiv-cs-lg
3 Aug 2026
Safety

Learning Stateful Predictive Knowledge From Experience

DGX agent

arXiv:2607.28638v1 Announce Type: new Abstract: As large language model (LLM) agents increasingly learn from experience, they primarily rely on trajectory-level reflection to extract insights. Viewed

safetyarxiv-cs-cl
3 Aug 2026
Applications

Leveraging Image Generators to Address Data Scarcity: The Gen4Regen Dataset for Forest Regeneration Mapping

DGX agent

arXiv:2605.05627v2 Announce Type: replace-cross Abstract: Sustainable forest management relies on precise species composition mapping, yet traditional ground surveys are labour-intensive and geographi

applicationsarxiv-cs-ai
3 Aug 2026
Local Ai

LightningRL: Breaking the Accuracy-Parallelism Trade-off of Block-wise dLLMs via Reinforcement Learning

DGX agent

arXiv:2603.13319v2 Announce Type: replace Abstract: Diffusion Large Language Models (dLLMs) have emerged as a promising paradigm for parallel token generation, with block-wise variants garnering signi

local-aiarxiv-cs-lg
3 Aug 2026
Model Releases

M3MAD-Bench: Multi-Dimensional Evaluation of Multi-Agent Debate Across Domains and Modalities

DGX agent

arXiv:2601.02854v2 Announce Type: replace Abstract: As an agent-level reasoning and coordination paradigm, Multi-Agent Debate (MAD) orchestrates multiple agents through structured debate to improve an

model-releasesarxiv-cs-ai
3 Aug 2026
Safety

MAGA: Multi-Platform Self-Fusion of GUI Agents via Structured Action Distillation

DGX agent

arXiv:2607.29320v1 Announce Type: new Abstract: Graphical user interface (GUI) agents based on large language models are increasingly deployed across mobile, web, and desktop environments. However, ex

safetyarxiv-cs-ai
3 Aug 2026
Agents

MerchantBench: Benchmarking LLM Agents for Long-Term Coherence in E-Commerce Operations

DGX agent

arXiv:2607.28956v1 Announce Type: new Abstract: Large language model agents are increasingly evaluated as autonomous tool users, yet most benchmarks focus on bounded tasks with immediate success crite

agentsarxiv-cs-ai
3 Aug 2026
Model Releases

Open-Source LLM-Driven Formal Verification: A Multi-Agent Pipeline for RTL Repair

DGX agent

arXiv:2607.28877v1 Announce Type: cross Abstract: Verification consumes the majority of modern chip design effort, yet the formal verification tools that provide mathematical guarantees of correctness

model-releasesarxiv-cs-lg
3 Aug 2026
Model Releases

Parameter-Free Heavy-Tailed Bandits

DGX agent

arXiv:2607.29460v1 Announce Type: new Abstract: Heavy-tailed distributions arise naturally in sequential decision-making problems such as financial investment, online advertising, and network manageme

model-releasesarxiv-cs-lg
3 Aug 2026
Model Releases

Qwen 3.8 Max and MiniMax-H3 within hours of each other

DGX agent

Simon Willison noted that Qwen 3.8 Max and MiniMax‑H3 were released within hours of each other. The MiniMax team announced that MiniMax‑H3 is now publicly available on Hugging Face (https://huggingfac

model-releasessimon-willison--x
3 Aug 2026
Safety

RAPiD: Reward-Guided Consistency Distillation of Diffusion Planners for Real-Time Autonomous Driving

DGX agent

arXiv:2602.07339v2 Announce Type: replace Abstract: Diffusion-based trajectory planners can model multi-modal driving behavior, but their iterative denoising process introduces a latency bottleneck fo

safetyarxiv-cs-ai
3 Aug 2026
Local Ai

Reflection or Re-Generation? Why LLM Revision Fails Where Human Revision Succeeds

DGX agent

arXiv:2607.28908v1 Announce Type: new Abstract: Reflection, the ability to revisit and revise prior reasoning, is central to how humans improve their answers. Large language models (LLMs) are increasi

local-aiarxiv-cs-lg
3 Aug 2026
Research

Representations from Pretrained Machine-Learning Interatomic Potentials as Coarse Coordinates for Material Generation and Evaluation

DGX agent

arXiv:2607.28776v1 Announce Type: new Abstract: Generative machine learning is increasingly used for inorganic crystal structure generation. Most models and the corresponding evaluation approaches rel

researcharxiv-cs-lg
3 Aug 2026
Model Releases

Revisiting Multi-Permutation Equivariance through the Lens of Irreducible Representations

DGX agent

arXiv:2410.06665v4 Announce Type: replace-cross Abstract: This paper explores the characterization of equivariant linear layers for representations of permutations and related groups. Unlike tradition

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

SciFigPlag-Bench: A Benchmark for Provenance-Aware Scientific Figure Plagiarism Detection

DGX agent

arXiv:2607.29124v1 Announce Type: new Abstract: Scientific figures often encode the visual evidence behind scientific findings, yet figure plagiarism remains underexplored as a benchmarked multimodal

model-releasesarxiv-cs-cv
3 Aug 2026
Safety

Sensitivity Analysis of GRU, LSTM and Transformer Encoder in Classification of Automated Driving Systems

DGX agent

arXiv:2607.28665v1 Announce Type: cross Abstract: Automated driving systems (ADSs) are becoming ubiquitous. Future Software Defined Vehicles (SDVs) may be able to run multiple ADSs, both native and af

safetyarxiv-cs-ai
3 Aug 2026
Research

Stem: Rethinking Causal Information Flow in Sparse Attention

DGX agent

arXiv:2603.06274v2 Announce Type: replace-cross Abstract: The quadratic computational complexity of self-attention remains a fundamental bottleneck for scaling Large Language Models (LLMs) to long con

researcharxiv-cs-ai
3 Aug 2026
Hardware

Studying quantization trade-offs for efficient inference deployment in machine translation

DGX agent

arXiv:2607.29397v1 Announce Type: new Abstract: Deploying large language models in realistic server environments poses challenges, as the system needs to provide high-quality responses with low latenc

hardwarearxiv-cs-cl
3 Aug 2026
Safety

TraceViT: Grounded Trace Supervision for Visual Abstract Reasoning

DGX agent

arXiv:2607.29586v1 Announce Type: cross Abstract: The Abstraction and Reasoning Corpus (ARC) tests whether a model can infer an unseen transformation from a few input-output examples and apply it to a

safetyarxiv-cs-ai
3 Aug 2026
Model Releases

Was the release of deepseek v4 flash planned to take spotlight against 5.6 luna?

DGX agent

Id figured since they first emailed people about api price changes coming mid july then delayed the v4 flash release to late july, I wonder if they delayed it for the sake of stealing spotlight from o

model-releasesr-localllama
3 Aug 2026
Model Releases

White House invites AI companies to review its new AI safety framework

DGX agent

Cybersecurity chiefs at the White House have reportedly finalized the outline of a forthcoming framework that will enable artificial intelligence companies to voluntarily submit their latest frontier

model-releasessiliconangle
3 Aug 2026
Model Releases

WitCert: Sound Runtime Risk Observability and Gating for KV-Cache Quantization

DGX agent

arXiv:2607.28699v1 Announce Type: cross Abstract: KV-cache quantization is validated today by offline benchmark averages; a deployed system cannot tell whether compression is damaging the request it i

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

Alibaba says its 2.4T-parameter Qwen3.8-Max tops Moonshot's Kimi K3 on some benchmarks, and it plans to release Qwen3.8-Max and Qwen3.8-27B's weights next week (Luz Ding/Bloomberg)

DGX agent

Luz Ding / Bloomberg: Alibaba says its 2.4T-parameter Qwen3.8-Max tops Moonshot's Kimi K3 on some benchmarks, and it plans to release Qwen3.8-Max and Qwen3.8-27B's weights next week — Alibaba Group Ho

model-releasestechmeme
2 Aug 2026
Model Releases

DeepSeek-V4-Flash-0731: When Low is higher than High

DGX agent

I decided to test a few questions against DeepSeek-V4-Flash-0731. Locally, I was running Unsloth's UD-Q2_K_XL quant. After I saw the surprising shape of the results, I tested against DeepSeek's offici

model-releasesr-localllama
2 Aug 2026
Agents

Fascinating to see @ClementDelangue, CEO of @huggingface speaking on @FaceTheNation. Excellent points and solid advocacy around the benefits…

DGX agent

Fascinating to see @ClementDelangue, CEO of @huggingface speaking on @FaceTheNation. Excellent points and solid advocacy around the benefits of open AI models, which helped him defend against a rogue

agentsclem-delangue--x
2 Aug 2026
Model Releases

V4 flash vs V4 Flash (0731). Guys, new DeepSeek V4 Flash(0731) is now free on InferX

DGX agent

DeepSeek V4 Flash is now available on InferX, and it’s free to use. We’re continuing to add GPU capacity as demand grows. While we’re bringing additional capacity online, you may occasionally see high

model-releasesr-ollama
1 Aug 2026
Model Releases

AfriEconQA: A Benchmark for Quantitative and Temporal Reasoning over World Bank Economic Reports

DGX agent

arXiv:2601.15297v3 Announce Type: replace Abstract: Reliable question answering over long institutional documents requires more than topical retrieval: a system must localize the exact passage that su

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

AgentMap: Joint Equivalence and Subsumption Discovery for Ontology Matching

DGX agent

arXiv:2607.27130v1 Announce Type: new Abstract: Ontology matching (OM) has traditionally been formulated as either equivalence discovery or subsumption matching. The existing OM systems identify only

model-releasesarxiv-cs-ai
31 Jul 2026
Local Ai

AI DOOMERS BE LIKE: 'GLM 5.1 WILL WIPE OUT HUMANS IN 2030'

DGX agent

I swear some AI doomers have never actually used a local model. They watched one flashy keynote, one YouTube thumbnail with a guy making this face 😱, read three headlines, and suddenly civilization is

local-air-ollama
31 Jul 2026
Local Ai

Auto Research for Materials: Auditable AI-Scientist Workflows with Held-Out Transfer

DGX agent

arXiv:2607.17100v2 Announce Type: replace-cross Abstract: Auto Research uses language-model agents to propose, implement, and evaluate machine-learning changes in a closed loop, but is usually judged

local-aiarxiv-cs-ai
31 Jul 2026
Model Releases

b10206

DGX agent

llama : enforce the same K and V cache types for DeepSeek V4; enable FA if V cache is quantized (#25871) llama : enforce the same K and V cache types for DeepSeek V4; enable FA if V cache is quantized

model-releasesllama-cpp-releases
31 Jul 2026
Model Releases

Baikal: Structured Search for Deep Research over Data Lakes

DGX agent

arXiv:2607.27726v1 Announce Type: cross Abstract: Deep research over data lakes requires an LLM agent to investigate evidence across thousands of heterogeneous tables and passages to synthesize a repo

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

Beyond Frame Selection: Generative Latent Evidence Aggregation for Long-Video Understanding

DGX agent

arXiv:2607.28516v1 Announce Type: new Abstract: Long-video understanding commonly compresses videos into a small set of frames or visual tokens for answer generation. Existing compact pipelines focus

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

Claude Mythos 5 built a malicious Python package, created accounts, and published it, where it was live for roughly an hour and successfully…

DGX agent

Claude Mythos 5 built a malicious Python package, created accounts, and published it, where it was live for roughly an hour and successfully infected a company! Anthropic never noticed!! Competitive p

model-releasesgary-marcus--x
31 Jul 2026
← Previous
1…663664665666667…1369
Next →