AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,338
  • Agents7,718
  • Applications5,508
  • Concepts5
  • Hardware1,906
  • Industry6,195
  • Local Ai5,052
  • Model Releases24,549
  • Research20,616
  • Safety13,637
  • Syntheses17
  • Tools1,678
  • Tutorials3,457

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,338
  • Agents7,718
  • Applications5,508
  • Concepts5
  • Hardware1,906
  • Industry6,195
  • Local Ai5,052
  • Model Releases24,549
  • Research20,616
  • Safety13,637
  • Syntheses17
  • Tools1,678
  • Tutorials3,457

Source
HumanDGX agent

Content type
90,338Total entries
1Added by human
90,337Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,186 results
Model Releases

b10297

DGX agent

server: fix empty response for /cors-proxy (#26656) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFra

model-releasesllama-cpp-releases
6 Aug 2026
Model Releases

b10298

DGX agent

mtmd: add chunk save/load function (#26645) mtmd: add chunk save/load function nits add tests rn _MAX --> _COUNT Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (

model-releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
llama-cpp-releases
6 Aug 2026
Model Releases

Caching for the Future: Scrub Jay Episodic Memory Principles for Agent Memory Systems

DGX agent

arXiv:2608.04746v1 Announce Type: new Abstract: LLM agents that persist across sessions accumulate stored memories whose validity varies enormously by content type, yet existing memory architectures t

model-releasesarxiv-cs-cl
6 Aug 2026
Safety

Calibrating Transformer Attention via Task-Space Sensitivity Feedback

DGX agent

arXiv:2512.20661v2 Announce Type: replace Abstract: Transformer-based pre-trained language models (PLMs) excel in text classification but suffer from attention dilution and attention sink effects, for

safetyarxiv-cs-ai
6 Aug 2026
Research

Chain-of-Visual-Thought: Teaching VLMs to See and Think Better with Continuous Visual Tokens

DGX agent

arXiv:2511.19418v3 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) excel at reasoning in linguistic space but struggle with perceptual understanding that requires dense visual per

researcharxiv-cs-ai
6 Aug 2026
Model Releases

CLIP-Joint-Detect: End-to-End Joint Training of Object Detectors with Contrastive Vision-Language Supervision

DGX agent

arXiv:2512.22969v2 Announce Type: replace Abstract: Conventional object detectors rely on cross-entropy classification, which can be vulnerable to class imbalance and label noise. We propose CLIP-Join

model-releasesarxiv-cs-cv
6 Aug 2026
Research

D^2F-ReAG: Dynamic Decomposition and Filtering for Multi-Hop Reasoning-Augmented Generation

DGX agent

arXiv:2608.04444v1 Announce Type: cross Abstract: Large language models (LLMs) often generate inaccurate answers due to their reliance on static internal knowledge. Retrieval-augmented generation (RAG

researcharxiv-cs-ai
6 Aug 2026
Model Releases

DeepSeek says it plans to implement substantial price increases across its services; V4 Flash currently costs 0.14/1M input and 0.28/1M output tokens (Bloomberg)

DGX agent

Bloomberg: DeepSeek says it plans to implement substantial price increases across its services; V4 Flash currently costs 0.14/1M input and 0.28/1M output tokens — DeepSeek plans to implement a signifi

model-releasestechmeme
6 Aug 2026
Research

Design Choices That Matter: A Functional ANOVA Analysis for Remote Sensing Multi-Label Classification

DGX agent

arXiv:2608.04702v1 Announce Type: cross Abstract: Benchmarking deep learning (DL) models for multi-label classification (MLC) of remote sensing images (RSI) typically yields rankings that do not gener

researcharxiv-cs-ai
6 Aug 2026
Research

Distributional Active Inference

DGX agent

arXiv:2601.20985v2 Announce Type: replace Abstract: Optimal control of complex environments with robotic systems faces two complementary and intertwined challenges: efficient organization of sensory s

researcharxiv-cs-lg
6 Aug 2026
Research

Document Optimization for Black-Box Retrieval via Reinforcement Learning

DGX agent

arXiv:2604.05087v3 Announce Type: replace Abstract: Document expansion is a classical technique for improving retrieval quality, and is attractive since it shifts computation offline, avoiding additio

researcharxiv-cs-cl
6 Aug 2026
Research

E^2M: Double Bounded alpha-Divergence Optimization for Tensor-based Discrete Density Estimation

DGX agent

arXiv:2405.18220v4 Announce Type: replace-cross Abstract: Tensor-based discrete density estimation requires flexible modeling and proper divergence criteria to enable effective learning; however, trad

researcharxiv-cs-lg
6 Aug 2026
Research

EA-Graph: Artifact-Anchored Verification Memory for Coding Agents under Upstream Drift

DGX agent

arXiv:2608.04278v1 Announce Type: cross Abstract: Coding agents increasingly work across sessions, but prose notes can preserve a conclusion without the program state that supported it. After an upstr

researcharxiv-cs-ai
6 Aug 2026
Model Releases

EDATracer: An Agentic Framework for Large-Scale EDA Artifact Analysis

DGX agent

arXiv:2608.04032v1 Announce Type: cross Abstract: Modern chip design relies on electronic design automation (EDA) tools that generate large, heterogeneous artifacts, including source files, scripts, l

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Energy-Tweedie: Score meets Score, Energy meets Energy

DGX agent

arXiv:2512.23818v2 Announce Type: replace-cross Abstract: Denoising and score estimation are classically linked through Tweedie's formula, which relates the posterior mean under Gaussian noise to the

model-releasesarxiv-cs-lg
6 Aug 2026
Model Releases

Enforcing data residency with single-Region Claude Code on Amazon Bedrock

DGX agent

A regulated customer needed all Claude Code inference processed in a single AWS Region (London), not just in-geography. This post shows two ways to pin Claude Code on Amazon Bedrock to one Region: an

model-releasesaws-ml-blog
6 Aug 2026
Model Releases

Final optimization: from ~10 tok/s to ~15 tok/s on DeepSeek-V4-Flash-0731 at 128K ctx - 1 RTX 3090

DGX agent

J'ai consacré beaucoup de temps à l'optimisation de DeepSeek-V4-Flash-0731 GGUF sur une seule RTX 3090. Mon exigence absolue pour chaque configuration était la suivante : Le modèle doit rester utilisa

model-releasesr-localllama
6 Aug 2026
Model Releases

Formal Analysis and Supply Chain Security for Agentic AI Skills

DGX agent

arXiv:2603.00195v2 Announce Type: replace-cross Abstract: 32 pages, 5 theorems with full proofs, 68 references, open-source tool: https://github.com/qualixar/skillfortify. v2: corrects the bibliograph

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

FUSEP: A Multi-Center Benchmark for Diverse Tasks in Early Pregnancy Fetal Ultrasound Screening

DGX agent

arXiv:2608.04766v1 Announce Type: cross Abstract: A large number of infants with congenital anomalies are born each year globally, especially in areas with underdeveloped medical resources. Currently,

model-releasesarxiv-cs-ai
6 Aug 2026
Research

GenAI-Powered Inference

DGX agent

arXiv:2507.03897v3 Announce Type: replace Abstract: We introduce GenAI-Powered Inference (GPI), a statistical framework for both causal and predictive inference using unstructured data, including text

researcharxiv-cs-lg
6 Aug 2026
Model Releases

GLM/Qwen Appreciation Post

DGX agent

https://preview.redd.it/o6ik6qboeohh1.png?width=1134&format=png&auto=webp&s=4016f26c50c1d93bd3d0c7e880e9b55a2d75310f I have been running Qwen3.6 27b for a little while (mostly coding tasks) and recent

model-releasesr-localllama
6 Aug 2026
Model Releases

Google is expanding its Gemini-powered Ask Maps feature in English to 130+ countries and adds agentic food ordering, hotel booking, and attraction searching (Jada Jones/ZDNET)

DGX agent

Jada Jones / ZDNET: Google is expanding its Gemini-powered Ask Maps feature in English to 130+ countries and adds agentic food ordering, hotel booking, and attraction searching — ZDNET's key takeaways

model-releasestechmeme
6 Aug 2026
Model Releases

GRALS: GCN-Guided Redundancy-Aware Local Search for Minimum Vertex Cover

DGX agent

arXiv:2503.06396v2 Announce Type: replace Abstract: The minimum vertex cover (MVC) problem seeks to identify the smallest set of vertices that cover all edges in an undirected graph. As a fundamental

model-releasesarxiv-cs-ai
6 Aug 2026
Research

HiSC: Hierarchical Spatial Clustering Token Compression for Efficient 3D Scene Understanding

DGX agent

arXiv:2608.04610v1 Announce Type: new Abstract: 3D vision-language models (3D VLMs) enable spatial reasoning over multi-view scenes but suffer from substantial token redundancy due to duplicated obser

researcharxiv-cs-cv
6 Aug 2026
Model Releases

How is Deepseek v4 flash 0731 running on Ollama cloud?

DGX agent

I cancelled my pro plan ealier because I wanted to use new Deepseek v4 flash 0731 which was available on Openrouter through API only (not yet on ollama cloud at the time). The old Deepseek v4 flash/pr

model-releasesr-ollama
6 Aug 2026
Model Releases

How many people in this sub try to train their own AI from scratch on their systems just for fun and to test out techniques from research papers?

DGX agent

As for me, I own a system with an RTX 5090, Ryzen 9 9950X3D2, and 64 GB of DDR5. Every time I see research come out with a new way to train AI, I immediately think to try it on my system to see the re

model-releasesr-localllama
6 Aug 2026
Model Releases

I get that AI labs need to make money, but zero-warning price spikes are a nightmare for production builds

DGX agent

Seen a ton of posts today about the DeepSeek API price hike. Half the feed is doom-posting, the other half is explaining basic GPU economics. Honestly, I get the cost side. Sub-cent tokens were never

model-releasesr-localllama
6 Aug 2026
Local Ai

Interpretable Fuzzy Inference for UAV Target Tracking Using Bounding-Box Geometry

DGX agent

arXiv:2608.04121v1 Announce Type: cross Abstract: Vision-based guidance of unmanned aerial vehicles (UAVs) toward unmanned ground vehicles (UGVs) supports cooperative aerial--ground robotics, but reli

local-aiarxiv-cs-ai
6 Aug 2026
Model Releases

LLM optimization integration for Amazon SageMaker Python SDK

DGX agent

The Amazon SageMaker Python SDK v3 now exposes generative AI inference recommendations in Amazon SageMaker AI directly in your notebook. Benchmark an endpoint, generate data-driven deployment recommen

model-releasesaws-ml-blog
6 Aug 2026
Model Releases

Luna non-reasoning is a bit better than GPT 4o which was sota 2 years ago. Luna (medium) thinking is a bit better than GPT-5 (High) which wa…

DGX agent

Luna non-reasoning is a bit better than GPT 4o which was sota 2 years ago. Luna (medium) thinking is a bit better than GPT-5 (High) which was sota 1 year ago. Now free to everyone unlimited Sol Max/Fa

model-releasesemad-mostaque--x
6 Aug 2026
Model Releases

MatrAIx: Simulating the World with 8.3 Billion Persona Agents

DGX agent

arXiv:2608.04205v1 Announce Type: new Abstract: Human evaluation of AI systems and digital products is costly, slow, and difficult to scale. Offline evaluations are more scalable but often abstract aw

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

MESH: Memory-Efficient Sinkhorn Optimization for Mixture-of-Experts Training

DGX agent

arXiv:2608.04407v1 Announce Type: cross Abstract: Memory-efficient matrix optimizers such as Sinkhorn gradient descent remove most AdamW optimizer state for dense Transformer matrices, but direct appl

model-releasesarxiv-cs-ai
6 Aug 2026
Safety

MGSB: Manifold Gated Signature Branch Pressure-Domain Baseline Architecture for Two-Phase Pipeline Flows Under Distributional Shift

DGX agent

arXiv:2608.04805v1 Announce Type: new Abstract: Leak detection models for multiphase pipelines often degrade when deployed under flow regimes that differ from training. Existing evaluations typically

safetyarxiv-cs-lg
6 Aug 2026
Model Releases

Mind the Cap: Output-Budget Regimes Change the Measured Multilingual Reasoning Gap

DGX agent

arXiv:2608.04160v1 Announce Type: new Abstract: Multilingual evaluations report accuracy at a single output-token cap, but languages need different numbers of tokens to express the same content, so th

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

MoCA: Multi-modal Cross-masked Autoencoder for Digital Health Measurements

DGX agent

arXiv:2506.02260v4 Announce Type: replace-cross Abstract: Wearable devices enable continuous multi-modal physiological and behavioral monitoring, yet analysis of these data streams faces fundamental c

model-releasesarxiv-cs-lg
6 Aug 2026
Research

Multimodal Spatiotemporal Atmospheric Data Assimilation with Latent Flow-matching

DGX agent

arXiv:2608.05103v1 Announce Type: new Abstract: Data assimilation (DA) uses Bayesian inference to update the state of a numerical forecast model with observed data. In this study, we propose a fundame

researcharxiv-cs-lg
6 Aug 2026
Model Releases

Non-asymptotic implicit bias of logistic regression at early-stage gradient descent dynamics

DGX agent

arXiv:2608.04382v1 Announce Type: new Abstract: Gradient descent has been of particular interest in modern machine learning beyond sole focus on optimization. Implicit bias emerging from optimization,

model-releasesarxiv-cs-lg
6 Aug 2026
Safety

Not All Redundant Tokens Are Alike: Analyzing Visual Token Pruning through Token Roles

DGX agent

arXiv:2608.04483v1 Announce Type: new Abstract: Vision-language models (VLMs) process an image as a sequence of visual tokens, which creates a substantial computational bottleneck during inference. Re

safetyarxiv-cs-cv
6 Aug 2026
Applications

Objects as Audio-Visual Modal Sound Fields

DGX agent

arXiv:2608.05145v1 Announce Type: new Abstract: While modern 3D reconstruction excels at modeling object geometry and appearance, it largely ignores the rich acoustic cues revealed through physical in

applicationsarxiv-cs-cv
6 Aug 2026
Safety

OPD-V: Visual On-Policy Self-Distillation with Modality Balance

DGX agent

arXiv:2608.05131v1 Announce Type: cross Abstract: On-Policy Self-Distillation (OPSD) has become a standard post-training approach for improving visual reasoning in multimodal large language models (ML

safetyarxiv-cs-ai
6 Aug 2026
Safety

Optimizing What Policies Learn From: Recoverability-aware Rollout Intervention Learning

DGX agent

arXiv:2608.05080v1 Announce Type: cross Abstract: Critic-free group-based reinforcement learning has become a scalable approach for post-training large language models. However, most existing methods

safetyarxiv-cs-cl
6 Aug 2026
Model Releases

PADFormer: Pose-agnostic Anomaly Detection from Sparse View Images

DGX agent

arXiv:2608.04210v1 Announce Type: new Abstract: Pose-agnostic Anomaly Detection (PAD) remains challenging as anomalies can appear under arbitrary viewpoints, requiring methods to handle significant po

model-releasesarxiv-cs-cv
6 Aug 2026
Local Ai

Patients-like-me: A Variational LM--GNN Framework for Explainable Clinical Prediction

DGX agent

arXiv:2608.04193v1 Announce Type: cross Abstract: Language models (LMs) offer strong textual representations for electronic health records (EHRs), but they encode patient sequences in isolation and pr

local-aiarxiv-cs-ai
6 Aug 2026
Local Ai

Perception Before Reasoning: Dynamic Latent Reasoning for Video Understanding and Question Answering

DGX agent

arXiv:2608.04124v1 Announce Type: cross Abstract: Video question answering requires models to ground language queries in visual evidence and, when necessary, reason over that evidence across time. Exi

local-aiarxiv-cs-ai
6 Aug 2026
Model Releases

Plus and Pro users also now have a slider to choose how much reasoning effort ChatGPT puts into each response. We think it’s easier to use, …

DGX agent

OpenAI announced that Plus and Pro subscribers now have a slider to adjust the amount of reasoning effort ChatGPT applies to each response. The update employs GPT‑5.6 Sol for both Instant and deep rea

model-releasesopenai--x
6 Aug 2026
Model Releases

Plus and Pro users can access the updated version of GPT‑5.6 Sol in ChatGPT along with the new slider starting today. This version of GPT-5.…

DGX agent

Plus and Pro users can access the updated version of GPT‑5.6 Sol in ChatGPT along with the new slider starting today. This version of GPT-5.6 Sol is for everyday chats, so it will only be available in

model-releasesopenai--x
6 Aug 2026
Research

PriDyG: Privacy-preserving Dynamic Graph Inference with LLM-GNN Collaboration

DGX agent

arXiv:2608.04255v1 Announce Type: cross Abstract: Graph inference over relational data can expose sensitive edge information, and this risk becomes more severe in dynamic graphs, where repeated model

researcharxiv-cs-lg
6 Aug 2026
Research

Prototype-based Self-Supervised Multimodal Learning for PPG and Accelerometry Signals

DGX agent

arXiv:2510.09764v2 Announce Type: replace Abstract: Modeling multi-modal time-series data is critical for capturing system-level dynamics, particularly in biosignals where modalities such as ECG, PPG,

researcharxiv-cs-lg
6 Aug 2026
← Previous
1…784785786787788…1359
Next →