AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
All
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,332 results
Model Releases

Beyond KV Reconstruction: Functional Reconstruction for MLA Draft Models in Speculative Decoding

DGX agent

arXiv:2607.27269v1 Announce Type: new Abstract: Multi-head latent attention (MLA) is increasingly important for long-context LLM inference because compact latent states replace the growing key-value (

model-releasesarxiv-cs-lg
31 Jul 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Beyond Sentiment: Structured Information Extraction from Financial News

DGX agent

arXiv:2607.28496v1 Announce Type: new Abstract: Financial sentiment analysis has become a standard component in news-driven stock prediction, yet it reduces rich, multi-dimensional news articles to a

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

Beyond Similarity: Grounded Agentic Extraction and Expert-Adjudicated Evaluation of Intertextuality in Classical Chinese Histories

DGX agent

arXiv:2607.27595v1 Announce Type: new Abstract: Computational approaches to intertextuality have advanced from string matching to neural retrieval, yet their outputs, similarity scores and parallel-pa

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

Beyond the Bidirectional Promise: Re-evaluating the Robustness of Diffusion Language Models

DGX agent

arXiv:2607.27386v1 Announce Type: cross Abstract: Diffusion Language Models (DLMs) offer a compelling alternative to autoregressive (AR) generation by enabling bidirectional context and iterative refi

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

BlindPSNR: A No-Reference Fidelity Predictor for Low-Light Image Enhancement

DGX agent

arXiv:2607.27628v1 Announce Type: new Abstract: Low-light image enhancement (LLIE) methods involve tunable parameters that are typically fixed, often leading to performance degradation when applied ac

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

Bridging AI and Energy Forecasting: An Autonomous Workflow with Customized Toolkit

DGX agent

arXiv:2307.07191v3 Announce Type: replace Abstract: Energy forecasting is crucial for the power grid, but fundamentally different from general time series analysis: it highly relies on covariates like

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

Bunraku: Turning a Single Illustration into an Editable Live2D Character

DGX agent

arXiv:2607.27348v1 Announce Type: new Abstract: Live2D is the dominant 2D character-animation format for anime characters and virtual avatars, representing each character as a stack of RGBA layers dri

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

Can Agents Deceive? Evaluating Reasoning and Deception in ParliamentBench using a Social Deduction Game

DGX agent

arXiv:2607.28146v1 Announce Type: new Abstract: As large language models (LLMs) are deployed as agents in high-stakes settings, such as medical and legal systems, understanding their deceptive capabil

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

Can LVLMs Uncover the Truth Behind Visual Illusions? An Analysis of Perceptual and Reasoning Capabilities

DGX agent

arXiv:2607.27747v1 Announce Type: new Abstract: Large Vision Language Models have integrated reasoning capabilities, elevating cognitive performance to new levels. However, existing evaluations either

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

Chem World: A Large-Scale Benchmark and Physics-Informed Framework for Trustworthy Chemical Property Prediction

DGX agent

arXiv:2607.28079v1 Announce Type: new Abstract: Chemical property prediction plays a critical role in accelerating scientific discovery in chemistry, materials science, and drug development. However,

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers

DGX agent

arXiv:2607.28611v1 Announce Type: new Abstract: Visual generation increasingly requires high-resolution images, long videos, and multimodal context, making the quadratic cost of full attention prohibi

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

ChronoMem: Version Control and Semantic Rollback for Large Language Model Agent Memory

DGX agent

arXiv:2607.27773v1 Announce Type: new Abstract: LLM agents increasingly rely on long-term memory to support multi-session interaction and personalization. However, existing agent memory systems are de

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

Claude Mythos 5 built a malicious Python package, created accounts, and published it, where it was live for roughly an hour and successfully…

DGX agent

Claude Mythos 5 built a malicious Python package, created accounts, and published it, where it was live for roughly an hour and successfully infected a company! Anthropic never noticed!! Competitive p

model-releasesgary-marcus--x
31 Jul 2026
Model Releases

ClawTrack: Towards Trace-Level Evaluation and Improvement of Real-World Autonomous Agents

DGX agent

arXiv:2607.28037v1 Announce Type: new Abstract: As LLM-based agents are deployed in complex, multi-step workflows, a critical evaluation gap has emerged: most existing benchmarks judge only final outc

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

ClinLens: Towards Long-Horizon Coding Agents for Longitudinal Multimodal Clinical Data Science

DGX agent

arXiv:2607.26155v1 Announce Type: new Abstract: Clinical data-science agents must transform heterogeneous longitudinal records into auditable analyses, yet existing benchmarks largely isolate medical

model-releasesarxiv-cs-ai
31 Jul 2026
Model Releases

Collusion with Competitive Marginals: Price-Level Audits Are Blind by Construction

DGX agent

arXiv:2607.26385v1 Announce Type: cross Abstract: Empirical work on algorithmic collusion asks one question of the data: are prices supracompetitive? We show this can be answered 'no' by a conspiracy

model-releasesarxiv-cs-ai
31 Jul 2026
Model Releases

Comparison of a Parametric Physics-Informed Neural Network and a Tensorial Reduced-Order Model for the Shallow-Water Dam-Break Problem

DGX agent

arXiv:2607.27433v1 Announce Type: cross Abstract: We develop two parametric data-driven reduced models: a physics-informed neural network (PINN) and a non-intrusive tensorial reduced-order model (TROM

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

Continuous-time reinforcement learning for optimal switching over multiple regimes

DGX agent

arXiv:2512.04697v3 Announce Type: replace-cross Abstract: This paper studies the continuous-time reinforcement learning (RL) for optimal switching problems across multiple regimes. We consider a type

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

CoRE-UIR: Prior-guided common and residual experts for efficient all-in-one remote sensing image restoration

DGX agent

arXiv:2607.27898v1 Announce Type: new Abstract: Remote sensing images acquired by unmanned aerial vehicles (UAVs) and satellites are often degraded by adverse weather, illumination variation, and imag

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

Critical attention scaling in long-context transformers

DGX agent

arXiv:2510.05554v2 Announce Type: replace Abstract: As large language models scale to longer contexts, attention layers suffer from a fundamental pathology: attention scores collapse toward uniformity

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

Cross-Embodiment Transfer via Behavior-Aligned Representations

DGX agent

arXiv:2607.27549v1 Announce Type: cross Abstract: Recent progress in large-scale imitation learning for robot manipulation has been driven by leveraging datasets across a wide range of robot embodimen

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

CXR-Retrieve: Compositional Text-to-Image Retrieval in Chest Radiography

DGX agent

arXiv:2607.27779v1 Announce Type: new Abstract: Large chest radiography archives are difficult to search because most studies are paired only with free-text reports rather than structured clinical ann

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

deepseek-ai/DeepSeek-V4-Flash-0731

DGX agent

deepseek-ai/DeepSeek-V4-Flash-0731 The latest release in DeepSeek's V4 family, 'with substantially enhanced agentic capabilities'. It's 304 billion parameters - 167GB on Hugging Face - but it appears

model-releasessimon-willison
31 Jul 2026
Model Releases

DeepSeek rolls out the official V4 Flash API in public beta, touting enhanced agent capabilities and benchmark scores 'far surpassing' V4 Pro Preview (Newley Purnell/Bloomberg)

DGX agent

Newley Purnell / Bloomberg: DeepSeek rolls out the official V4 Flash API in public beta, touting enhanced agent capabilities and benchmark scores “far surpassing” V4 Pro Preview — China's DeepSeek rol

model-releasestechmeme
31 Jul 2026
Model Releases

DeepSeek V4 Flash 0731 in Hermes Agent and one prompt, took 32 minutes and cost 0.07$, this model is so cheap to the point where 2 dollars c…

DGX agent

**DeepSeek V4 Flash 0731 Performance Test** On July 31 2026, a single prompt executed via the Hermes Agent on DeepSeek V4 Flash 0731 completed in 32 minutes and incurred an estimated cost of 0.07 USD.

model-releasesnous-research--x
31 Jul 2026
Model Releases

DeepSeek-V4-Flash-0731 unsloth gguf on A100

DGX agent

A100 with 40gb VRAM: 162GB Q8_K_XL ~16.1 tok/s generation Only 15.8GB of 40GB VRAM used with all experts on CPU NOTE just tested coding on linux box DeepSeek-V4-Flash-0731 runs losslessly on the singl

model-releasesr-localllama
31 Jul 2026
Model Releases

DeepSeek v4 Flash for DS4 (DwarfStar) GGUF w/ DSpark MTP Head

DGX agent

I'm an avid user of Deepseek v4 Flash via antirez's DS4 DwarfStar inference engine, and so when the new checkpoint dropped, the first thing I did was rent a cloud box and spin up a quantization for us

model-releasesr-localllama
31 Jul 2026
Model Releases

DeepSeek V4 Flash GA ranks the same as Sonnet 5 and Grok 4.5 on DeepSWE

DGX agent

Source: https://x.com/deepseek_ai/status/2083084415157022911 & https://deepswe.datacurve.ai/ just combined data view. DeepSeek claims, not verified by DeepSWE yet. submitted by /u/sdexca [link] [comme

model-releasesr-localllama
31 Jul 2026
Model Releases

DeepSeek v4 Flash has a nice bump in Capability

DGX agent

DeepSeek V4 Flash: Preview → 2026-07-31 Benchmark Preview 0731 Δ Terminal Bench* 56.9 82.7 +25.8 Toolathlon 51.8 70.3 +18.5 NL2Repo — 54.2 new Cybergym — 76.7 new DeepSWE — 54.4 new Agent Last Exam —

model-releasesr-localllama
31 Jul 2026
Model Releases

Deepseek V4 Flash is now ~#2 open weight model to Kimi K3 and >50x cheaper

DGX agent

https://preview.redd.it/h7zv5tb3tmgh1.png?width=2854&format=png&auto=webp&s=507380e8f862c18f10f7c5c84da9e8d1c59139b0 Deepseek's new flash model is unexpectedly cheap and high-performing across useful

model-releasesr-localllama
31 Jul 2026
Model Releases

🚀 DeepSeek-V4-Flash Official API is now LIVE in public beta! 🔷 We’ve massively upgraded its Agent capabilities—benchmark scores are now fa…

DGX agent

🚀 DeepSeek-V4-Flash Official API is now LIVE in public beta! 🔷 We’ve massively upgraded its Agent capabilities—benchmark scores are now far surpassing the V4-Pro-Preview. Check out the massive perform

model-releasesdeepseek--x
31 Jul 2026
Model Releases

Deepseek V4 Flash on SlopCodeBench

DGX agent

While waiting for some of the quants to drop, I load the API with $50 and ran it on SlopCodeBench Just vibe reading the results it seems like Opus 4.8 < Deepseek < Opus 5 https://github.com/michaelasp

model-releasesr-localllama
31 Jul 2026
Model Releases

DeepSeek V4 Flash scores 50 on the Artificial Analysis Intelligence Index, matching Gemini 3.6 Flash and up 10 points from the preview launch in April (Artificial Analysis)

DGX agent

Artificial Analysis: DeepSeek V4 Flash scores 50 on the Artificial Analysis Intelligence Index, matching Gemini 3.6 Flash and up 10 points from the preview launch in April — DeepSeek V4 Flash 0731 is

model-releasestechmeme
31 Jul 2026
Model Releases

Dense Supervision, Sparse Updates: On the Sparsity and Geometry of On-Policy Distillation

DGX agent

arXiv:2606.13657v3 Announce Type: replace Abstract: On-policy distillation (OPD) has recently become a prominent post-training recipe by combining two desirable ingredients: on-policy student-generate

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

Dimensionality and Measurement Precision in HLE's Multiple-Choice Subset

DGX agent

arXiv:2607.27420v1 Announce Type: cross Abstract: Humanity's Last Exam (HLE) is widely used to evaluate frontier language models. HLE organizes its questions into eight subject-domain categories, whos

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

Divergence Decoding: Training-Free Capability Fusion

DGX agent

arXiv:2607.27248v1 Announce Type: cross Abstract: While large language models excel in reasoning, these generalists often lack knowledge for specialized scientific domains. Conversely, domain models~(

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

DoTime: A Synthetic Benchmark Generator for Interventional and Counterfactual Time Series

DGX agent

arXiv:2607.27263v1 Announce Type: new Abstract: Most benchmarks for causal inference over time series are observational, small, or domain-specific, leaving interventional and counterfactual estimation

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

DS@GT ARC at ImageCLEFmedical 2026: Architectural Diversity for Concept Detection and Foundation-Model Scaling for Caption Prediction in Medical Image Analysis

DGX agent

arXiv:2607.27763v1 Announce Type: new Abstract: We describe the DS@GT submissions to the ImageCLEFmedical Caption 2026 challenge, which continues a long-running benchmark on the ROCOv2 dataset with tw

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

ECG-InterpBench: Benchmarking the Interpretability of ECG Foundation Models with Matched-Scale Sparse Autoencoders

DGX agent

arXiv:2607.27404v1 Announce Type: new Abstract: Existing benchmarks for electrocardiogram foundation models primarily evaluate downstream predictive performance, providing limited insight into whether

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

Echoverse: Deep, Evolving Environments for Training Computer-Use Agents at Scale

DGX agent

arXiv:2607.28074v1 Announce Type: cross Abstract: Computer-use agents learn from what their actions change, so training one needs applications it can act on, break and reset. The applications that mat

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

EEG-EditBench: Probing Visual Information in EEG-Image Retrieval Models with Controlled Image Edits

DGX agent

arXiv:2607.27857v1 Announce Type: new Abstract: Recent EEG-to-image retrieval models have achieved strong performance in identifying viewed images from semantically diverse candidates. Yet such succes

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

Efficient LLMs with AMP: Attention Heads and MLP Pruning

DGX agent

arXiv:2504.21174v2 Announce Type: replace Abstract: Deep learning drives a new wave in computing systems and triggers the automation of increasingly complex problems. In particular, Large Language Mod

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

EgoGVAE: Ego-body Mesh Reconstruction via Guided Variational Autoencoder

DGX agent

arXiv:2607.27755v1 Announce Type: new Abstract: We address the problem of recovering the full-body mesh from only the head pose. This task has become essential for various applications based on head-m

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

EHGCN: Hierarchical Euclidean-Hyperbolic Fusion via Motion-Aware GCN for Hybrid Event Stream Perception

DGX agent

arXiv:2504.16616v4 Announce Type: replace Abstract: Event cameras, characterized by microsecond temporal resolution and very High Dynamic Range (HDR), emit high-speed event streams for perception task

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

EMBL AI Librarian: Life-Sciences Knowledge Layer for AI Agents

DGX agent

arXiv:2607.28229v1 Announce Type: new Abstract: The web is increasingly accessed by AI agents rather than humans. Every agent needs knowledge, especially in the life-sciences, where agentic pipelines

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

Epistemic diversity across language models mitigates knowledge collapse

DGX agent

arXiv:2512.15011v3 Announce Type: replace Abstract: Artificial intelligence (AI) increasingly generates the very content used to train future AI systems. This feedback loop can degrade model quality,

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

eta-OPSD: Deriving with Policy Optimization, Training with Self-Distillation

DGX agent

arXiv:2607.28582v1 Announce Type: new Abstract: On-policy self-distillation (OPSD) is a promising approach to improve reasoning language models, but it remains brittle in practice: making it work reli

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

Evidence-Ledger Adjudication for Claim-Evidence Traceability

DGX agent

arXiv:2607.26512v1 Announce Type: new Abstract: AI agents can draft claims faster than authors can check whether the cited or retrieved evidence supports them. We study evidence-ledger adjudication: a

model-releasesarxiv-cs-ai
31 Jul 2026
← Previous
1…5657585960…466
Next →