AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
All
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,323 results
Model Releases

Boosting Generalizable Depth Estimation in Endoscopy by Mixture of Lightweight Experts and Intrinsic Image Alignment

DGX agent

arXiv:2608.00415v1 Announce Type: new Abstract: Depth estimation is a significant task for 3D perception in endoscopic surgeries. However, illumination interference and feature diversity in various en

model-releasesarxiv-cs-cv
4 Aug 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Breaking the Horizontal Prior: From Long-Tailed Orientation Bias to Roll-Robust Monocular Depth Estimation

DGX agent

arXiv:2608.00678v1 Announce Type: new Abstract: Despite recent advances in Monocular Depth Estimation, state-of-the-art depth foundation models remain vulnerable to robustness issues. Particularly, ev

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

Bridging the English-Arabic Medical Knowledge Gap: Targeted Low-Rank Adaptation via Causal Layer Selection

DGX agent

arXiv:2608.00207v1 Announce Type: new Abstract: Large Language Models (LLMs) perform strongly in English medical tasks but degrade substantially in Arabic, a gap widely attributed to limited training

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

BRiG-AFA: Bellman Risk-to-Go Learning for Non-Myopic Active Feature Acquisition

DGX agent

arXiv:2608.02305v1 Announce Type: new Abstract: Active feature acquisition (AFA) asks which unobserved feature to measure next for each test instance under a budget. Greedy rules are easy to train but

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

CADENA: Stepwise CAD Reverse Engineering

DGX agent

arXiv:2608.00799v1 Announce Type: new Abstract: Computer-Aided Design (CAD) underpins modern engineering, yet converting existing shapes into editable models still demands substantial expert effort. M

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

Can You Trust the Confidence? ConfBench for Vision-Language Models on Document Extraction

DGX agent

arXiv:2608.01792v1 Announce Type: cross Abstract: Intelligent document processing (IDP) with vision-language models (VLMs) hinges on confidence scores trustworthy enough to route extractions between a

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Capability Provenance in Language Models: A Case Study in Social Reasoning

DGX agent

arXiv:2606.19625v2 Announce Type: replace Abstract: We use training-data attribution as an interpretable tool for capability discovery, mapping which regions of the pretraining corpus support social-r

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

CAPEval: A Decoupled Caption Evaluation across Understanding and Generation

DGX agent

arXiv:2608.02589v1 Announce Type: new Abstract: Captions serve as a primary supervision signal for both multimodal understanding and text-to-image generation. However, previous evaluations treat the c

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

CAPMix: Robust KPI Anomaly Detection for AIOps in Noisy and Dynamic Environments

DGX agent

arXiv:2509.06419v2 Announce Type: replace Abstract: Time-series anomaly detection is crucial in AIOps for maintaining large-scale service reliability. In production, streams of Key Performance Indicat

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

Certifying Plans under Model Mismatch: A Trilemma for Reachability from Scarce Data

DGX agent

arXiv:2608.02453v1 Announce Type: new Abstract: Sim-to-real policies are designed under nominal dynamics, but target-system trials may yield only a few isolated one-step transitions. We study pre-exec

model-releasesarxiv-cs-ro
4 Aug 2026
Model Releases

ChainVLA: Chaining Vision-Language-Action Queries through a Unified Execution State for Long-Horizon Manipulation

DGX agent

arXiv:2608.02326v1 Announce Type: new Abstract: Humans perform long-horizon manipulation by retaining knowledge of what earlier actions have established while continuously adapting the motion underway

model-releasesarxiv-cs-ro
4 Aug 2026
Model Releases

ChaosProbe: A Neurochaotic Lens on Frozen Transformer Input-Embedding Spaces

DGX agent

arXiv:2608.01968v1 Announce Type: new Abstract: Transformer models are most often understood through what they do: their benchmark performance, generation quality, or behavior on downstream tasks. Yet

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

CHOW-SLAM: Compact Hybrid Representation with Complementary Overlap Window Optimization for RGB-D SLAM

DGX agent

arXiv:2608.01914v1 Announce Type: new Abstract: Simultaneous localization and mapping (SLAM) based on Neural Radiance Fields (NeRF) enables dense, continuous scene reconstruction. However, existing sy

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

Cluster-Aware Over-the-Air Federated Learning with Energy-Harvesting Devices: From Global Training to Model Personalization

DGX agent

arXiv:2608.01426v1 Announce Type: new Abstract: Federated learning (FL) enables distributed optimization and learning across decentralized edge devices while preserving data privacy, but its performan

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

CompanionBench: A Theory-Anchored, Real-World-Grounded Benchmark for AI Emotional Companionship

DGX agent

arXiv:2608.02046v1 Announce Type: new Abstract: LLM companions are deployed at scale in personally consequential settings, yet poorly evaluated. Existing benchmarks use hand-authored scenarios and pro

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Company approved 128GB Mac for research proposal, best model?

DGX agent

I‘m doing a research proposal at my company about running local LLMs to replace daily coding models. Qwen 3.6 27B (or 3.8 potentially) is widely seen as the best model in that 20-60GB space, is that s

model-releasesr-localllama
4 Aug 2026
Model Releases

CoNav-UAV: Cooperative Dual-Altitude Aerial Navigation via Stackelberg Learning

DGX agent

arXiv:2608.01802v1 Announce Type: cross Abstract: Target-oriented vision-and-language navigation (VLN) on aerial platforms is attracting growing attention for missions such as disaster rescue, infrast

model-releasesarxiv-cs-ro
4 Aug 2026
Model Releases

Conditional Deep Levy Models for Exotic Derivatives: History-Aware Path Generation and P-Q Payoff Diagnostics

DGX agent

arXiv:2509.13374v2 Announce Type: replace-cross Abstract: We develop and audit a history-aware financial path generator based on Denoising Levy Probabilistic Models (DLPMs) for conditional equity-inde

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

Conservation laws determine what physical learning remembers

DGX agent

arXiv:2608.00097v1 Announce Type: cross Abstract: Physical learning rules such as equilibrium propagation (EP), coupled learning (CL), and adjoint coupled learning (AL) train resistive networks throug

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

Control Under Compression: Reliability Frontiers for Tool-Using Agents

DGX agent

arXiv:2608.01056v1 Announce Type: cross Abstract: Tool-using language-model agents are governed not only by task prompts but also by persistent system-side instructions that specify tools, arguments,

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Cooperative Coevolution for Resource-Constrained Agentic LLM Post-Training

DGX agent

arXiv:2608.02391v1 Announce Type: cross Abstract: Tool-using large language model (LLM) agents produce long, multi-turn trajectories, making gradient-based post-training memory-intensive. Evolution st

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

CORTIVA: Candidate-Score Fusion of Complementary Visual Teachers for EEG- and MEG-to-Image Retrieval

DGX agent

arXiv:2608.01355v1 Announce Type: new Abstract: Decoding visual experience from non-invasive brain activity is central to neuroscience and brain-computer interfaces. Functional magnetic resonance imag

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

Cost-Effective Automated Judging of Natural-Language Mathematical Proofs

DGX agent

arXiv:2608.00004v1 Announce Type: new Abstract: Grading natural-language mathematical proofs is a recurring cost in evaluating math-reasoning systems, and frontier LLM judges are expensive. We ask whe

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

CoSynFlow: Conformal Symplectic Neural Flows for Cross-System Prediction of Dissipative Hamiltonian Dynamics

DGX agent

arXiv:2608.00571v1 Announce Type: new Abstract: Learning solution operators for differential equations is a central problem in scientific machine learning. However, many neural operator methods optimi

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

CRAFT: Compression via Recursive Adaptive Fusion of Video Tokens for Vision-Language Models

DGX agent

arXiv:2608.01644v1 Announce Type: new Abstract: In video understanding, vision-language models (VLMs) must ingest massive numbers of visual tokens, causing the computational and memory cost of the pre

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

CRIP: Channel Level Representation Injection for Personalized One-Shot Federated Learning

DGX agent

arXiv:2608.02222v1 Announce Type: new Abstract: One-shot federated learning (OSFL) has emerged as a promising collaborative model learning framework with only a single round of communication, offering

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

CrossLex: A Source-Grounded Benchmark for Cross-Jurisdictional Legal Reasoning in Large Language Models

DGX agent

arXiv:2608.01292v1 Announce Type: new Abstract: Legal reasoning is inherently jurisdiction-dependent: the same facts can call for different legal rules and yield different conclusions across legal sys

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

CrossProjection: Geometric Grounding Beyond Viewpoint Change in Architectural Drawings

DGX agent

arXiv:2608.00473v1 Announce Type: cross Abstract: Architectural drawings violate the usual assumption behind multi-view reasoning: plans and sections are cuts, while elevations are facade projections,

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Crushing the Evidence: A Dual-Penalty Evasion Framework for Fooling White-Box Explainable AI Auditors

DGX agent

arXiv:2608.00566v1 Announce Type: new Abstract: Post-hoc model explainers such as LIME, SHAP, and Integrated Gradients are widely deployed to audit models in high-stakes sensitive domains, including f

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

CultureVidBench: Benchmarking Cultural Understanding in Text-to-Video Generation

DGX agent

arXiv:2608.01942v1 Announce Type: cross Abstract: Text-to-video (T2V) generation models have advanced rapidly, yet their ability to represent diverse cultural contexts remains underexplored. Existing

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

CurveShift: Is Agent Progress Scalar? Separating Level from Shape

DGX agent

arXiv:2608.00355v1 Announce Type: new Abstract: Progress in large language models is often summarized using a single scalar measure, such as a time horizon, a latent ability estimate, or an aggregate

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Data-Driven Pinball-Loss Selection for Vertically Distributed Elastic-Net SVMs

DGX agent

arXiv:2608.00949v1 Announce Type: new Abstract: The pinball-loss support vector machine is robust, but its asymmetry parameter is usually fixed in advance. We propose a data-driven elastic-net support

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

DE-NER : Zero-shot Named Entity Recognition via Dialogue Elicitation of Large Language Models

DGX agent

arXiv:2608.00538v1 Announce Type: new Abstract: Recent advancements of zero-shot Named Entity Recognition (NER) establish strong baselines by formulating sequence labeling into question answering wher

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

DeBERTa-Sentinel: Toward Transparent and Trustworthy Detection of AI-Generated Text

DGX agent

arXiv:2608.01046v1 Announce Type: new Abstract: The rapid spread of large language models (LLMs) across the web raises concerns about misinformation, academic integrity, automated content manipulation

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

DeCLIP: Decoupled Prompting for Multi-Label Class-Incremental Learning with CLIP

DGX agent

arXiv:2509.23335v3 Announce Type: replace Abstract: Multi-label class-incremental learning (MLCIL) continuously expands the label space while recognizing multiple co-occurring categories, making catas

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

Decoupling semantics from vision: A framework for faithful visual-text compression evaluation

DGX agent

arXiv:2608.01848v1 Announce Type: new Abstract: Recent visual-text compression (VTC) methods, typified by DeepSeek-OCR, report impressive high token compression ratios for long-context modeling tasks

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

Decrease the power limit of your 5090 to at least 480W - the performance penalty for inference is negligible.

DGX agent

I run my inference machine in the living room, so noise and heat output are a significant concern. Ran a quick test using my daily driver model (Qwen 3.6-27b) and at 480W, the card outputs only 2.1% l

model-releasesr-localllama
4 Aug 2026
Model Releases

Deep Learning for Cyber Threat Detection and Mitigation in Healthcare-IoT

DGX agent

arXiv:2608.00118v1 Announce Type: cross Abstract: Cybersecurity is a fundamental requirement for protecting wearable devices used in healthcare Internet of Things (H-IoT) systems. Security failures in

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

Deep Learning for Retinal Degeneration Assessment: A Comprehensive Analysis of the MARIO Challenge

DGX agent

arXiv:2506.02976v4 Announce Type: replace Abstract: The MARIO challenge, held at MICCAI 2024, focused on advancing the automated detection and monitoring of age-related macular degeneration (AMD) thro

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

[Deepseek-V4-Flash-0731] Full 1M context on a single RTX5090 + DDR5 Desktop Setup with VLLM CPU/Ram Offloading, ~800 tps pp & 15+ tps decode [Agentic Coding]

DGX agent

First of all, obviously I took some help from AI to type this post and this is the topic that enabled me to accomplish all that: https://old.reddit.com/r/LocalLLaMA/comments/1veow4b/deepseek_v4flash_2

model-releasesr-localllama
4 Aug 2026
Model Releases

DeepSeek-V4-Flash-0731 is Ollama's fastest growing model ever in token usage. We are scaling capacity in US & Europe. On Ollama, this model …

DGX agent

DeepSeek-V4-Flash-0731 is Ollama's fastest growing model ever in token usage. We are scaling capacity in US & Europe. On Ollama, this model runs with high performance (100tps+) and zero data retention

model-releasesollama--x
4 Aug 2026
Model Releases

DeepSeek V4 Flash 0731 (Q4) now reaches 1,328 tok/s prefill and ~29 tok/s decode on one RTX PRO 6000

DGX agent

I've been working on speeding up DeepSeek-V4-Flash-0731 in Krasis and have now got the long-prompt prefill quite a bit faster on a single RTX PRO 6000 96GB. These are timing-disabled internal Krasis r

model-releasesr-localllama
4 Aug 2026
Model Releases

Deepseek V4 flash 0731 ranks #21 on Agent Arena

DGX agent

https://preview.redd.it/522fsdwvtdhh1.png?width=1200&format=png&auto=webp&s=6a6cf7a467514167a8193029dbd20fb3a9ba4f6c It ranks lower than both Sonnet 4.6 and Luna. I'd wager Luna costs in the same ball

model-releasesr-localllama
4 Aug 2026
Model Releases

Deepseek V4 Flash 2-bit quant is the first model I can run locally that achieves 100% in this SQL benchmark

DGX agent

I really like to use this one SQL benchmark when testing new models. I had another post some time ago with my benchmarks, but I decided to post a new one because of how well Deepseek did. I like the b

model-releasesr-localllama
4 Aug 2026
Model Releases

DeepSeek V4 Flash is now live on Together AI. Frontier agent performance is getting dramatically cheaper. DSV4 on Together AI brings a major…

DGX agent

DeepSeek V4 Flash is now live on Together AI. Frontier agent performance is getting dramatically cheaper. DSV4 on Together AI brings a major jump in coding, tool use, and long-running agent performanc

model-releasestogether-ai--x
4 Aug 2026
Model Releases

DeepSeek v4 Flash vs. Qwen3.6-27B, 3.5-122B, and Gemma 4 31B Benchmark

DGX agent

Just wanted to share my agentic coding benchmark run of DSv4F 0731 at both High and Low reasoning efforts (not Max)... I ran a 109-question subset of Aider Polyglot (the JS/C++/Python languages), base

model-releasesr-localllama
4 Aug 2026
Model Releases

DeepSurvey-Bench: Evaluating Academic Value of Automatically Generated Scientific Surveys

DGX agent

arXiv:2601.15307v2 Announce Type: replace-cross Abstract: The rapid development of automated survey generation technology has made it increasingly important to establish a comprehensive benchmark to e

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Defending Membership Inference Attacks via Privacy-aware Sparsity Tuning

DGX agent

arXiv:2410.06814v2 Announce Type: replace Abstract: Over-parameterized models are typically vulnerable to membership inference attacks, which aim to determine whether a specific sample is included in

model-releasesarxiv-cs-lg
4 Aug 2026
← Previous
1…4344454647…466
Next →