AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
All
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,315 results
Model Releases

SciFigPlag-Bench: A Benchmark for Provenance-Aware Scientific Figure Plagiarism Detection

DGX agent

arXiv:2607.29124v1 Announce Type: new Abstract: Scientific figures often encode the visual evidence behind scientific findings, yet figure plagiarism remains underexplored as a benchmarked multimodal

model-releasesarxiv-cs-cv
3 Aug 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

SciToolAgent-Evo: An Ontology-Aware Self-Evolving Agent for Open-World Scientific Tool Acquisition

DGX agent

arXiv:2607.28692v1 Announce Type: new Abstract: Large language model (LLM) agents have been increasingly adopted in scientific research for organizing and invoking specialized computational tools. How

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

SEDR-Seq2P: A Lightweight Dilated Residual Sequence-to-Point Network for Multi-Task Industrial NILM

DGX agent

arXiv:2607.28693v1 Announce Type: cross Abstract: Industrial NILM remains challenging because measurement noise and widespread concurrent machine operation reduce the generalization of models tuned on

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

SeekBrain: An Autonomous Multi-Agent System for Accelerating Neuroscience Discovery

DGX agent

arXiv:2607.29347v1 Announce Type: cross Abstract: Modern neuroscience relies on integrating multi-scale, multimodal datasets to uncover the neural principles underlying intelligence. However, analytic

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

Sign compression for Muon: SignMuon, MuonSign, and the Limits of Error Feedback

DGX agent

arXiv:2607.29674v1 Announce Type: cross Abstract: SignMuon compresses the Muon update to one bit per parameter by taking its elementwise sign, providing the most direct way to run a matrix-aware optim

model-releasesarxiv-cs-lg
3 Aug 2026
Model Releases

SILVA Networks as Structured Implicit Layers and Vector Attractors via Dynamic Interaction Fields

DGX agent

arXiv:2607.28989v1 Announce Type: new Abstract: Many learning problems require representations that reconcile direct input, nearby structure, and broader context. In implicit neural layers, these infl

model-releasesarxiv-cs-lg
3 Aug 2026
Model Releases

Simulation Code Generation for Fluid Systems using Large Language Models: Benchmarking Models and Prompting Strategies

DGX agent

arXiv:2607.29389v1 Announce Type: new Abstract: Large language models (LLMs) have demonstrated a strong ability to generate syntactically correct code from natural-language specifications. In this stu

model-releasesarxiv-cs-lg
3 Aug 2026
Model Releases

So-Fake: Benchmarking and Explaining Social Media Image Forgery Detection

DGX agent

arXiv:2505.18660v5 Announce Type: replace Abstract: Recent advances in AI-powered generative models have enabled the creation of increasingly realistic synthetic images, posing significant risks to in

model-releasesarxiv-cs-cv
3 Aug 2026
Model Releases

Speculative decoding with deepseek v4 flash 0731?

DGX agent

Has anyone figured out how to enable speculative decoding with deepseek v4 flash 0731 on llamacpp? I’m on the right release for llamacpp (b10228 or earlier) and running am17an’s draft model with unslo

model-releasesr-localllama
3 Aug 2026
Model Releases

.@ssankar says Palantir was able to make Nvidia's Nemotron Ultra model 'better than frontier': 'I literally almost felt gaslit when, within …

DGX agent

.@ssankar says Palantir was able to make Nvidia's Nemotron Ultra model 'better than frontier': 'I literally almost felt gaslit when, within 24 hours of getting Nemotron up with no post-training, this

model-releasesclem-delangue--x
3 Aug 2026
Model Releases

Step-Level Visual Grounding Faithfulness Predicts Out-of-Distribution Generalization in Long-Horizon Vision-Language Models

DGX agent

arXiv:2603.06828v2 Announce Type: replace-cross Abstract: We uncover a behavioral law of long-horizon vision-language models: models that maintain temporally grounded beliefs generalize better. Standa

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

StraightDP: Geometry-Aware Differential Privacy for Rectified-Flow Transformers

DGX agent

arXiv:2607.29100v1 Announce Type: cross Abstract: Differentially private (DP) training of text-conditioned generative models suffers a utility cliff at strong privacy. We revisit this problem through

model-releasesarxiv-cs-cv
3 Aug 2026
Model Releases

SULAND v2: A Refined RGB Dataset and Deep Learning Object Detection Benchmark for UAV/UGV-Based SUrface LANDmine Detection Under Domain Shift

DGX agent

arXiv:2607.28996v1 Announce Type: new Abstract: RGB imagery offers a practical, low-cost option for Unmanned Aerial/Ground Vehicle (UAV/UGV) survey support in surface-landmine detection, but object de

model-releasesarxiv-cs-cv
3 Aug 2026
Model Releases

TAPR: Enhancing LLM Performance with a Task-Aware Prompt Rewriter

DGX agent

arXiv:2607.28657v1 Announce Type: new Abstract: Large Language Models (LLMs) often require carefully crafted prompts to unlock their full potential, which can be a barrier for non-expert users. This w

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

Teaching Video Generators to Remember: Eliciting Dynamic Memory for Out-of-Sight State Evolution

DGX agent

arXiv:2605.25333v2 Announce Type: replace Abstract: Video world models should maintain evolving states when evidence is unobserved, yet current generators often freeze hidden states upon interruption.

model-releasesarxiv-cs-cv
3 Aug 2026
Model Releases

Thanks for the recognition. We'll keep building! 🚀

DGX agent

Thanks for the recognition. We'll keep building! 🚀 Big news: Qwen3.8-Max by @Alibaba_Qwen just landed at #4 on the Frontend Code Arena leaderboard with a score of 1,668! With 1,668 points, Qwen3.8-Max

model-releasesqwen--x
3 Aug 2026
Model Releases

The Asymmetric Effects of Knowledge Distillation on Bias in Small Language Models

DGX agent

arXiv:2607.28639v1 Announce Type: cross Abstract: We show that knowledge distillation in small instruction-tuned language models has asymmetric effects on bias. On unambiguous tasks (BBQ-disambig), re

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

The Chinese labs everyone lumps together are making four pretty different bets. I work at one of them.

DGX agent

Every time a model drops from a Chinese lab the thread fills with people who already know who made it, and the guess is usually Alibaba. There was a thread here recently asking what separates the open

model-releasesr-localllama
3 Aug 2026
Model Releases

The Grokked Illusion: True Equilibrium Mitigates Catastrophic Forgetting

DGX agent

arXiv:2607.29503v1 Announce Type: new Abstract: While neural networks are typically evaluated by their training and test performance, these metrics do not reveal how robust a learned representation is

model-releasesarxiv-cs-lg
3 Aug 2026
Model Releases

The Morphological Core of Dungan: A Two-Dialect Finite-State Model and a Multi-Genre Evaluation

DGX agent

arXiv:2607.28766v1 Announce Type: new Abstract: Dungan, a Sinitic language of Central Asia written in a Cyrillic-based script, is described in detail in the grammatical literature, yet the quantitativ

model-releasesarxiv-cs-cl
3 Aug 2026
Model Releases

The Parts Are Greater Than the Sum: Automated Task Sequencing for Efficient Training of Multi-Policy LLMs

DGX agent

arXiv:2607.29601v1 Announce Type: new Abstract: Parameter-Efficient Fine-Tuning (PEFT) commonly adapts large language models using a single shared Low-Rank Adapter (LoRA). This shared optimization spa

model-releasesarxiv-cs-lg
3 Aug 2026
Model Releases

The result is a faster, more natural conversation with ChatGPT Voice from the moment a session starts. How we built it: https://openai.com/i…

DGX agent

OpenAI has redesigned the ChatGPT Voice stack—from client to model—to enable continuous audio streaming, allowing GPT‑Live to listen while speaking without interruption. The new architecture supports

model-releasesopenai--x
3 Aug 2026
Model Releases

The results span sphere packing, coding theory, group theory, quantum complexity, lattice cryptography, extremal combinatorics, and more. Am…

DGX agent

The results span sphere packing, coding theory, group theory, quantum complexity, lattice cryptography, extremal combinatorics, and more. Among them: establishing the existence of non-sofic groups and

model-releasesopenai--x
3 Aug 2026
Model Releases

To Add Is Machine, To Delete Is Human: Measuring and Mitigating Deletion Avoidance in LLM Code Editing

DGX agent

arXiv:2607.28887v1 Announce Type: cross Abstract: Large language models increasingly write and repair production code, yet evidence is mounting that their test-passing patches leave codebases harder t

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

Tokenizer-Agnostic Engram Module

DGX agent

arXiv:2607.29065v1 Announce Type: new Abstract: Deepseek's Engram, a conditional memory module, was introduced to trade-off storage versus reasoning in large language models. However, the module relie

model-releasesarxiv-cs-cl
3 Aug 2026
Model Releases

Towards bridging the gap: Systematic sim-to-real transfer for diverse legged robots

DGX agent

arXiv:2509.06342v2 Announce Type: replace Abstract: Legged robots must achieve both robust locomotion and energy efficiency to be practical in real-world environments. Yet controllers trained in simul

model-releasesarxiv-cs-ro
3 Aug 2026
Model Releases

Translation with Thought: Difficulty-Adaptive Reasoning via Reinforcement Learning for Multi-Domain Machine Translation

DGX agent

arXiv:2607.29287v1 Announce Type: cross Abstract: Multi-domain machine translation (MDMT) poses a unique challenge due to varying levels of linguistic complexity across domains. Inspired by human tran

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

Tri-Space Operational Control of Redundant Multilink and Hybrid Cable-Driven Parallel Robots Using an Iterative-Learning based Reactive Approach

DGX agent

arXiv:2607.29500v1 Announce Type: new Abstract: Cable-Driven Parallel Robots (CDPRs) are a type of parallel mechanism in which cables are used as actuators. Due to the two levels of redundancy and num

model-releasesarxiv-cs-ro
3 Aug 2026
Model Releases

two weeks ago i went on @swyx's pod and said some things that i... should not have said. a lot has happened since then, i owe you all an apo…

DGX agent

two weeks ago i went on @swyx's pod and said some things that i... should not have said. a lot has happened since then, i owe you all an apology. i'm sorry that i was right about every single thing. a

model-releasesswyx--x
3 Aug 2026
Model Releases

V4-Flash-0731 - vibes after first weekend of use

DGX agent

Spent way too much time with V4-Flash-0731 this weekend and wanted to share my vibes as briefly as possible. I sent it through a bit of real-work and some of my personal benchmarks. My quick thoughts

model-releasesr-localllama
3 Aug 2026
Model Releases

Validation Evidence in LLM Repair Agents: How Much of What Passes Actually Tests the Bug?

DGX agent

arXiv:2607.28871v1 Announce Type: cross Abstract: When a repair agent runs a test and sees it pass, the result is treated as evidence about the reported defect. We measure how often that treatment is

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

Was the release of deepseek v4 flash planned to take spotlight against 5.6 luna?

DGX agent

Id figured since they first emailed people about api price changes coming mid july then delayed the v4 flash release to late july, I wonder if they delayed it for the sake of stealing spotlight from o

model-releasesr-localllama
3 Aug 2026
Model Releases

WebCoderBench: Benchmarking Web Application Generation with Comprehensive and Interpretable Evaluation Metrics

DGX agent

arXiv:2601.02430v3 Announce Type: replace-cross Abstract: Web applications (web apps) have become a key arena for large language models (LLMs) to demonstrate their code generation capabilities and com

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

We’re releasing the manuscripts, formal Lean certificates, and reasoning walkthroughs so mathematicians can examine these results and build …

DGX agent

We’re releasing the manuscripts, formal Lean certificates, and reasoning walkthroughs so mathematicians can examine these results and build on their ideas. https://openai.com/index/ten-advances-in-mat

model-releasesopenai--x
3 Aug 2026
Model Releases

What Is Missing in Surgical Risk Stratification and Outcome Prediction: A Scoping Review of End-to-End Machine Learning Approaches

DGX agent

arXiv:2607.29090v1 Announce Type: new Abstract: Postoperative adverse events, including mortality and morbidity, remain a major global burden, many of which are preventable through early identificatio

model-releasesarxiv-cs-lg
3 Aug 2026
Model Releases

White House invites AI companies to review its new AI safety framework

DGX agent

Cybersecurity chiefs at the White House have reportedly finalized the outline of a forthcoming framework that will enable artificial intelligence companies to voluntarily submit their latest frontier

model-releasessiliconangle
3 Aug 2026
Model Releases

Why It Hurts: Identifying the Drivers of Negative Thoughts in Emotional Support Conversations

DGX agent

arXiv:2607.28648v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used for emotional support tasks, such as negative thought reframing. This task relies on modifying cogn

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

WitCert: Sound Runtime Risk Observability and Gating for KV-Cache Quantization

DGX agent

arXiv:2607.28699v1 Announce Type: cross Abstract: KV-cache quantization is validated today by offline benchmark averages; a deployed system cannot tell whether compression is damaging the request it i

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

You probably used @FireworksAI_HQ this week. The cool part is that you just did not know you did. 🎆 Open-source models are free, sure... bu…

DGX agent

You probably used @FireworksAI_HQ this week. The cool part is that you just did not know you did. 🎆 Open-source models are free, sure... but the hard part is tailoring them so they perform best on you

model-releasesfireworks-ai--x
3 Aug 2026
Model Releases

You shouldn't need a vision model to know your PDF has checkboxes. LiteParse can now pull structured data directly from your PDFs: form fiel…

DGX agent

You shouldn't need a vision model to know your PDF has checkboxes. LiteParse can now pull structured data directly from your PDFs: form field values, checkbox states, annotations, embedded images, vec

model-releasesllamaindex--x
3 Aug 2026
Model Releases

Alibaba says its 2.4T-parameter Qwen3.8-Max tops Moonshot's Kimi K3 on some benchmarks, and it plans to release Qwen3.8-Max and Qwen3.8-27B's weights next week (Luz Ding/Bloomberg)

DGX agent

Luz Ding / Bloomberg: Alibaba says its 2.4T-parameter Qwen3.8-Max tops Moonshot's Kimi K3 on some benchmarks, and it plans to release Qwen3.8-Max and Qwen3.8-27B's weights next week — Alibaba Group Ho

model-releasestechmeme
2 Aug 2026
Model Releases

All other models on the Portal remain 20% discounted, aside from GPT-5.6 Terra and Luna which are 50% off. https://x.com/NousResearch/status…

DGX agent

All other models on the Portal remain 20% discounted, aside from GPT-5.6 Terra and Luna which are 50% off. https://x.com/NousResearch/status/2080039066771337475?s=20 All models are now 20% off for a l

model-releasesnous-research--x
2 Aug 2026
Model Releases

Another day and another full frontier model running on your computer. Been teething DeepSeek V4 Flash on over 60 employees at The Zero Human…

DGX agent

Another day and another full frontier model running on your computer. Been teething DeepSeek V4 Flash on over 60 employees at The Zero Human Company for a few hours and it is stunning. Between Kimi K3

model-releasesclem-delangue--x
2 Aug 2026
Model Releases

b10224

DGX agent

ggml-webgpu: add support for f16 repeat (#26307) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFramew

model-releasesllama-cpp-releases
2 Aug 2026
Model Releases

b10225

DGX agent

model : load MiMo V2 MTP tensors only if used (#26412) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XC

model-releasesllama-cpp-releases
2 Aug 2026
Model Releases

b10226

DGX agent

sycl: fix classification of iGPUs (#26105) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFramework Li

model-releasesllama-cpp-releases
2 Aug 2026
Model Releases

b10227

DGX agent

chat : add qwen3 specialized parser (#26252) Add tagged thinking tool parser chat : refactor and add permute helper cont : add support for <tool_call> omission cont : update tool delimiters cont : add

model-releasesllama-cpp-releases
2 Aug 2026
Model Releases

b10228

DGX agent

DeepseekV4 MTP + DSpark (#25784) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFramework Linux: Ubunt

model-releasesllama-cpp-releases
2 Aug 2026
← Previous
1…5253545556…465
Next →