AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,566
  • Agents7,790
  • Applications5,565
  • Concepts5
  • Hardware1,943
  • Industry6,214
  • Local Ai5,132
  • Model Releases24,957
  • Research20,928
  • Safety13,836
  • Syntheses17
  • Tools1,680
  • Tutorials3,499

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,566
  • Agents7,790
  • Applications5,565
  • Concepts5
  • Hardware1,943
  • Industry6,214
  • Local Ai5,132
  • Model Releases24,957
  • Research20,928
  • Safety13,836
  • Syntheses17
  • Tools1,680
  • Tutorials3,499

Source
HumanDGX agent

Content type
91,566Total entries
1Added by human
91,565Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
66,232 results
Model Releases

DS4 flash 0731 - Acquarium Panel Failure - Q3_K_XL Unsloth

DGX agent

https://preview.redd.it/1a39x4zivqgh1.png?width=1550&format=png&auto=webp&s=de591c039cc18782a6b5d8e402fdc1594be05132 start C:llmllamam5uildinllama-server.exe --model 'H:UD-Q3_K_XLDeepSeek-V4-Flash-073

model-releasesr-localllama
1 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Possible to create accurate medieval woodcut style art?

DGX agent

Wondering if it's possible to actually produce ai art works that are indistinguishable from authentic medieval woodcut illustrations like the one attached. All the AI attempts I've seen at re creating

model-releasesr-stablediffusion
1 Aug 2026
Model Releases

Qwen 3.6 27B Q5 on 3x2080ti: 55tps with llama.cpp. Can I squeeze out more?

DGX agent

CPU: Threadripper 3970X RAM: 128GB DDR4 GPUs: 3x2080ti 11GB The current best parameters to run it: llama-server --model Qwen3.6-27B-Q5_K_S.gguf --n-gpu-layers 999 --split-mode tensor --flash-attn on -

model-releasesr-localllama
1 Aug 2026
Model Releases

Albilich: Steerable Proof-State Orchestration for LLM-Based Mathematical Research with CAS Integration

DGX agent

arXiv:2607.27705v1 Announce Type: cross Abstract: Large language models can contribute useful ideas to mathematical research, yet long-horizon proof attempts remain difficult to coordinate, evaluate,

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

Anthropic discloses that Claude hacked three organizations during internal tests

DGX agent

Three of Anthropic PBC’s large language models carried out successful cyberattacks during routine internal tests. The company detailed the breaches on Thursday. A few days earlier, rival OpenAI Group

model-releasessiliconangle
31 Jul 2026
Research

Articulated Object Reconstruction from Rest-State Observation

DGX agent

arXiv:2607.27749v1 Announce Type: new Abstract: Building interactive digital twins requires recovering both 3D geometry and the kinematic structures that govern how objects articulate. Yet existing me

researcharxiv-cs-cv
31 Jul 2026
Model Releases

b10212

DGX agent

llama : load MTP tensors only if they are really used (#26296) llama : load MTP tensors only if they are really used llama : skip loading MTP (if not used) in remaining models that support MTP Co-auth

model-releasesllama-cpp-releases
31 Jul 2026
Model Releases

Back from the Future: Key-Value Cache Management by Counter-Causal Surprise

DGX agent

arXiv:2607.27600v1 Announce Type: new Abstract: Key-value (KV) cache management through compression and eviction strategies has emerged as an important research direction in recent years. Computationa

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

Beyond Sentiment: Structured Information Extraction from Financial News

DGX agent

arXiv:2607.28496v1 Announce Type: new Abstract: Financial sentiment analysis has become a standard component in news-driven stock prediction, yet it reduces rich, multi-dimensional news articles to a

model-releasesarxiv-cs-cl
31 Jul 2026
Applications

CACHE-UK: A Stability-Aware Memory Editor for Sequentially Updated Quantized LLMs in Finance

DGX agent

arXiv:2607.28292v1 Announce Type: new Abstract: Large Language Models (LLMs) deployed in dynamic financial environments face a critical challenge: maintaining factual accuracy as market conditions, re

applicationsarxiv-cs-cl
31 Jul 2026
Model Releases

Can LVLMs Uncover the Truth Behind Visual Illusions? An Analysis of Perceptual and Reasoning Capabilities

DGX agent

arXiv:2607.27747v1 Announce Type: new Abstract: Large Vision Language Models have integrated reasoning capabilities, elevating cognitive performance to new levels. However, existing evaluations either

model-releasesarxiv-cs-cl
31 Jul 2026
Safety

Class-Aware Reinforcement Learning for Counterfactual Explanation Generation

DGX agent

arXiv:2607.27905v1 Announce Type: new Abstract: Counterfactual explanations (CFEs) enhance the interpretability of black-box models by generating alternative instances with adjusted feature values tha

safetyarxiv-cs-lg
31 Jul 2026
Research

Contrastive Concept Importance: Explaining Pairwise Class Decisions Through Automatically Extracted Concept Representations

DGX agent

arXiv:2607.27904v1 Announce Type: new Abstract: Concept-based explanations are a prevalent way to explain the decisions of complex black-box methods through semantically meaningful, human-interpretabl

researcharxiv-cs-lg
31 Jul 2026
Research

Deep learning-based hierarchical insect classification using camera trap imagery

DGX agent

arXiv:2607.28005v1 Announce Type: new Abstract: Declining insect populations make reliable biodiversity monitoring increasingly urgent, yet monitoring of insect biodiversity is hampered by a lack of s

researcharxiv-cs-cv
31 Jul 2026
Model Releases

DeepSeek v4 Flash for DS4 (DwarfStar) GGUF w/ DSpark MTP Head

DGX agent

I'm an avid user of Deepseek v4 Flash via antirez's DS4 DwarfStar inference engine, and so when the new checkpoint dropped, the first thing I did was rent a cloud box and spin up a quantization for us

model-releasesr-localllama
31 Jul 2026
Model Releases

Dense Supervision, Sparse Updates: On the Sparsity and Geometry of On-Policy Distillation

DGX agent

arXiv:2606.13657v3 Announce Type: replace Abstract: On-policy distillation (OPD) has recently become a prominent post-training recipe by combining two desirable ingredients: on-policy student-generate

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

DoTime: A Synthetic Benchmark Generator for Interventional and Counterfactual Time Series

DGX agent

arXiv:2607.27263v1 Announce Type: new Abstract: Most benchmarks for causal inference over time series are observational, small, or domain-specific, leaving interventional and counterfactual estimation

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

Exploring Structures in Physics Problems: Can AI Agents Discover Statistical Mechanical Mappings?

DGX agent

arXiv:2607.26367v1 Announce Type: new Abstract: An important skill in theoretical physics is to recognize when a new problem can be transformed into a known model. We study this skill as an AI-agent t

model-releasesarxiv-cs-ai
31 Jul 2026
Model Releases

Fantastic Adaptive Taxonomies and How to Use Them

DGX agent

arXiv:2607.16387v2 Announce Type: replace-cross Abstract: An agent system's execution traces record how it fails, and procedures that improve such a system without changing model weights (trajectory s

model-releasesarxiv-cs-ai
31 Jul 2026
Safety

FiRE: Enhancing MLLMs with Fine-Grained Context Learning for Complex Image Retrieval

DGX agent

arXiv:2607.27959v1 Announce Type: new Abstract: Due to their strong generalizable multimodal processing and reasoning capabilities, Multimodal Large Language Models (MLLMs) have demonstrated significa

safetyarxiv-cs-cv
31 Jul 2026
Research

Hallucinations Leave a Grounding Signature:Verifier-Guided Decoding for Selective Object Correction

DGX agent

arXiv:2607.27823v1 Announce Type: new Abstract: Large vision-language models (LVLMs) often hallucinate objects that are absent from an image. Despite recent progress, existing mitigation methods still

researcharxiv-cs-cv
31 Jul 2026
Local Ai

I built a hybrid Transformer–SSM LLM agent with a local CLI, active control, and run receipts

DGX agent

I’m one of the builders of LOLM, a hybrid Transformer–SSM model and agent system from Qira. The model separates surface token processing from persistent latent-state tracking. An NFET controller can s

local-air-ollama
31 Jul 2026
Model Releases

I predict DeepSeek V4 Flash 0731's Artificial Analysis score to be 57 ± 1 point (Kimi K3 Level)

DGX agent

Deepseek's new model V4 Flash 0731 is much better, I (Claude lol) did a bit of linear regression with a leave one out style verification to predict its AA Score, and that puts it at Kimi K3 level, whi

model-releasesr-localllama
31 Jul 2026
Model Releases

I switched my https://agent.datasette.io instance to Luna (it was previously on Gemini 3.1 Flash-Lite - Luna is cheaper now) - you can sign …

DGX agent

Simon Willison switched his Datasette Agent instance from Gemini 3.1 Flash‑Lite to GPT‑5.6 “Luna” after a recent 80% price drop. He reports the new model is significantly faster and automatically gene

model-releasessimon-willison--x
31 Jul 2026
Model Releases

It’s been a busy couple of weeks! ICYMI, here’s the recap ⬇️ — Gemini Robotics 2 from @GoogleDeepmind brings whole-body intelligence to robo…

DGX agent

It’s been a busy couple of weeks! ICYMI, here’s the recap ⬇️ — Gemini Robotics 2 from @GoogleDeepmind brings whole-body intelligence to robots — Gemini 3.5 Flash-Lite is our fastest, most cost-effecti

model-releasesgoogle-ai--x
31 Jul 2026
Model Releases

KernelGenBench: A Multi-Source and Multi-Chip Benchmark for LLM-based Kernel Generation

DGX agent

arXiv:2607.27231v1 Announce Type: cross Abstract: Large language models (LLMs) have significantly increased the demand for efficient accelerator kernels, but kernel development remains a highly specia

model-releasesarxiv-cs-lg
31 Jul 2026
Research

LAST: The Last Query Token Guides Visual Token Pruning for Edge-Cloud Collaborative MLLM Inference

DGX agent

arXiv:2607.27952v1 Announce Type: new Abstract: Multimodal foundation models are reshaping edge-cloud visual intelligence from task-specific feature pipelines into token-based interfaces, where edge d

researcharxiv-cs-cv
31 Jul 2026
Model Releases

LoRA Scaffolded Policy Optimization (LSPO): A Sampling-Time Low-Rank Scaffold for Recovering Reinforcement-Learning Gradient on Zero-Reward Cliff Prompts

DGX agent

arXiv:2607.27787v1 Announce Type: new Abstract: Reinforcement learning from verifiable rewards (RLVR) for mathematical reasoning suffers from a structural blind spot: on 'cliff' prompts-those on which

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

Minimum VRAM GPU to run DeepSeek-V4-Flash-0731 Q4_K_XL at around 30 t/s ?

DGX agent

Hello guys, I'm curious about running DeepSeek-V4-Flash-0731 locally. Since it’s a Mixture of Experts (MoE) model with only 13B active parameters, I was hoping the VRAM requirements might be manageabl

model-releasesr-localllama
31 Jul 2026
Model Releases

MMHBench: A Multi-Perspective Benchmark for Mental Health Understanding in Long-Form Videos

DGX agent

arXiv:2607.27895v1 Announce Type: cross Abstract: Mental health understanding in long-form videos requires nuanced reasoning over observable behavior, interpersonal context, and latent psychological s

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

MOON2.0: Dynamic Modality-balanced Multimodal Representation Learning for E-commerce Product Understanding

DGX agent

arXiv:2511.12449v3 Announce Type: replace Abstract: Recent Multimodal Large Language Models (MLLMs) have significantly advanced e-commerce product understanding. However, they still face three challen

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

PCAP-LM: An LLM-Native Text Representation for TLS Bulk Traffic Analysis

DGX agent

arXiv:2607.28100v1 Announce Type: cross Abstract: Large language models (LLMs) offer powerful reasoning capabilities for network traffic analysis, but standard capture formats and their textual equiva

model-releasesarxiv-cs-cl
31 Jul 2026
Safety

Policy Gradient Steering: Interventions from Behavioral Objectives

DGX agent

arXiv:2607.27574v1 Announce Type: new Abstract: Activation steering has emerged in large language models as a lightweight alternative for dynamically changing a model's behavior at inference time. How

safetyarxiv-cs-lg
31 Jul 2026
Model Releases

Private Face Recognition Training Dataset Publication via Identity-Decoupled and Geometry-Preserving Face Distillation

DGX agent

arXiv:2607.27764v1 Announce Type: new Abstract: Publishing private face recognition~(FR) training datasets is privacy-sensitive because faces expose identity information. Private FR training dataset p

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

RedFlow: Redirect Failure into Action-Level Corrections for Flow-matching VLA Policy

DGX agent

arXiv:2607.27782v1 Announce Type: new Abstract: Flow-matching Vision-Language-Action (VLA) policies have shown strong potential for robotic manipulation but often suffer from compounding errors caused

model-releasesarxiv-cs-ro
31 Jul 2026
Research

S-Avatar: Diffusion-Guided Gaussian Head Avatars from a Single Image

DGX agent

arXiv:2607.28164v1 Announce Type: new Abstract: We propose S-Avatar, a novel method for generating photorealistic 3D head avatars from a single image using a diffusion-guided 3D model generation modul

researcharxiv-cs-cv
31 Jul 2026
Model Releases

Shared Symbolic Backbones for Physically Consistent Multi-Output Symbolic Regression

DGX agent

arXiv:2607.26528v1 Announce Type: cross Abstract: Symbolic regression provides analytical expressions, but it is usually applied one output at a time. This is limiting in process systems, where state

model-releasesarxiv-cs-ai
31 Jul 2026
Model Releases

Some deepseek-v4-flash 20260731 opinion review

DGX agent

First of all, I want to apologize if it's off-topic or in the wrong format. Having tried Deepseek Flash with reasoning high on a conceptually difficult task, involving Machine Learning classifiers and

model-releasesr-localllama
31 Jul 2026
Research

Theia: Large-Scale Multimodal Captioning and Automated Validation of the Incidents1M Dataset for Data-Free Distillation

DGX agent

arXiv:2607.28269v1 Announce Type: new Abstract: The deployment of Vision-Language Models (VLMs) in critical domains like disaster management requires high-quality multimodal datasets, especially for t

researcharxiv-cs-cv
31 Jul 2026
Model Releases

Thinking Once Is Enough: Intermediate-Layer Evidence Routing for High-Resolution VQA

DGX agent

arXiv:2607.27830v1 Announce Type: new Abstract: High-resolution visual question answering (HR-VQA) is often treated as a problem of insufficient evidence acquisition, where failing multimodal large la

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

This will happen frequently as AI becomes smarter and more agentic

DGX agent

This will happen frequently as AI becomes smarter and more agentic In a review of our cybersecurity evaluations, we found three incidents in which a Claude model reached the internet from within or wh

model-releaseselon-musk--x
31 Jul 2026
Agents

ThreatForest: Multi-Agent Attack Tree Generation with Pluggable TTP Framework Mapping

DGX agent

arXiv:2607.27528v1 Announce Type: cross Abstract: Threat modeling is essential for secure software development, yet manual analysis of cloud-native architectures is slow and demands scarce security ex

agentsarxiv-cs-cl
31 Jul 2026
Tutorials

Towards Robust Monocular Depth Estimation in Non-Lambertian Surfaces

DGX agent

arXiv:2408.06083v2 Announce Type: replace Abstract: In the field of monocular depth estimation (MDE), many models with excellent zero-shot performance in general scenes emerge recently. However, these

tutorialsarxiv-cs-cv
31 Jul 2026
Model Releases

Towards Unified Multimodal Misinformation Detection in Social Media: A Benchmark Dataset and Baseline

DGX agent

arXiv:2509.25991v3 Announce Type: replace-cross Abstract: Detecting deceptive multimodal content on social media has become an increasingly important problem. Two major types of deception dominate: hu

model-releasesarxiv-cs-cv
31 Jul 2026
Agents

Training Skills Like Parameters via Self-Supervised Semantic Diffusion

DGX agent

arXiv:2607.27557v1 Announce Type: new Abstract: While Large Language Models (LLMs) demonstrate remarkable general instruction-following capabilities, they often fall short of human experts in highly s

agentsarxiv-cs-cl
31 Jul 2026
Model Releases

Very interesting paper on recursive self-improvement. The whole stack is released. Machine learning engineering gives recursive self-improve…

DGX agent

Very interesting paper on recursive self-improvement. The whole stack is released. Machine learning engineering gives recursive self-improvement a concrete, executable testbed. OpenMLE is an open full

model-releasesdair-ai--x
31 Jul 2026
Research

VETO: Towards Protecting Images From Frontier AI Editing

DGX agent

arXiv:2607.27292v1 Announce Type: new Abstract: The rise of powerful, accessible image-editing models such as FLUX.2 has brought high-fidelity editing within broad reach. Their capabilities now extend

researcharxiv-cs-cv
31 Jul 2026
Model Releases

We have an idea for dinner 2 already :) But what ideas are you all interested in? - technical topics like rl envs, continual learning, cloud…

DGX agent

We have an idea for dinner 2 already :) But what ideas are you all interested in? - technical topics like rl envs, continual learning, cloud agents, world models - general startup / company building f

model-releasesjerry-liu--x
31 Jul 2026
← Previous
1…506507508509510…1380
Next →