AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
All
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,315 results
Model Releases

Certifying Plans under Model Mismatch: A Trilemma for Reachability from Scarce Data

DGX agent

arXiv:2608.02453v1 Announce Type: new Abstract: Sim-to-real policies are designed under nominal dynamics, but target-system trials may yield only a few isolated one-step transitions. We study pre-exec

model-releasesarxiv-cs-ro
4 Aug 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

ChainVLA: Chaining Vision-Language-Action Queries through a Unified Execution State for Long-Horizon Manipulation

DGX agent

arXiv:2608.02326v1 Announce Type: new Abstract: Humans perform long-horizon manipulation by retaining knowledge of what earlier actions have established while continuously adapting the motion underway

model-releasesarxiv-cs-ro
4 Aug 2026
Model Releases

ChaosProbe: A Neurochaotic Lens on Frozen Transformer Input-Embedding Spaces

DGX agent

arXiv:2608.01968v1 Announce Type: new Abstract: Transformer models are most often understood through what they do: their benchmark performance, generation quality, or behavior on downstream tasks. Yet

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

CHOW-SLAM: Compact Hybrid Representation with Complementary Overlap Window Optimization for RGB-D SLAM

DGX agent

arXiv:2608.01914v1 Announce Type: new Abstract: Simultaneous localization and mapping (SLAM) based on Neural Radiance Fields (NeRF) enables dense, continuous scene reconstruction. However, existing sy

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

Cluster-Aware Over-the-Air Federated Learning with Energy-Harvesting Devices: From Global Training to Model Personalization

DGX agent

arXiv:2608.01426v1 Announce Type: new Abstract: Federated learning (FL) enables distributed optimization and learning across decentralized edge devices while preserving data privacy, but its performan

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

CompanionBench: A Theory-Anchored, Real-World-Grounded Benchmark for AI Emotional Companionship

DGX agent

arXiv:2608.02046v1 Announce Type: new Abstract: LLM companions are deployed at scale in personally consequential settings, yet poorly evaluated. Existing benchmarks use hand-authored scenarios and pro

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Company approved 128GB Mac for research proposal, best model?

DGX agent

I‘m doing a research proposal at my company about running local LLMs to replace daily coding models. Qwen 3.6 27B (or 3.8 potentially) is widely seen as the best model in that 20-60GB space, is that s

model-releasesr-localllama
4 Aug 2026
Model Releases

CoNav-UAV: Cooperative Dual-Altitude Aerial Navigation via Stackelberg Learning

DGX agent

arXiv:2608.01802v1 Announce Type: cross Abstract: Target-oriented vision-and-language navigation (VLN) on aerial platforms is attracting growing attention for missions such as disaster rescue, infrast

model-releasesarxiv-cs-ro
4 Aug 2026
Model Releases

Conditional Deep Levy Models for Exotic Derivatives: History-Aware Path Generation and P-Q Payoff Diagnostics

DGX agent

arXiv:2509.13374v2 Announce Type: replace-cross Abstract: We develop and audit a history-aware financial path generator based on Denoising Levy Probabilistic Models (DLPMs) for conditional equity-inde

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

Conservation laws determine what physical learning remembers

DGX agent

arXiv:2608.00097v1 Announce Type: cross Abstract: Physical learning rules such as equilibrium propagation (EP), coupled learning (CL), and adjoint coupled learning (AL) train resistive networks throug

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

Control Under Compression: Reliability Frontiers for Tool-Using Agents

DGX agent

arXiv:2608.01056v1 Announce Type: cross Abstract: Tool-using language-model agents are governed not only by task prompts but also by persistent system-side instructions that specify tools, arguments,

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Cooperative Coevolution for Resource-Constrained Agentic LLM Post-Training

DGX agent

arXiv:2608.02391v1 Announce Type: cross Abstract: Tool-using large language model (LLM) agents produce long, multi-turn trajectories, making gradient-based post-training memory-intensive. Evolution st

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

CORTIVA: Candidate-Score Fusion of Complementary Visual Teachers for EEG- and MEG-to-Image Retrieval

DGX agent

arXiv:2608.01355v1 Announce Type: new Abstract: Decoding visual experience from non-invasive brain activity is central to neuroscience and brain-computer interfaces. Functional magnetic resonance imag

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

Cost-Effective Automated Judging of Natural-Language Mathematical Proofs

DGX agent

arXiv:2608.00004v1 Announce Type: new Abstract: Grading natural-language mathematical proofs is a recurring cost in evaluating math-reasoning systems, and frontier LLM judges are expensive. We ask whe

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

CoSynFlow: Conformal Symplectic Neural Flows for Cross-System Prediction of Dissipative Hamiltonian Dynamics

DGX agent

arXiv:2608.00571v1 Announce Type: new Abstract: Learning solution operators for differential equations is a central problem in scientific machine learning. However, many neural operator methods optimi

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

CRAFT: Compression via Recursive Adaptive Fusion of Video Tokens for Vision-Language Models

DGX agent

arXiv:2608.01644v1 Announce Type: new Abstract: In video understanding, vision-language models (VLMs) must ingest massive numbers of visual tokens, causing the computational and memory cost of the pre

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

CRIP: Channel Level Representation Injection for Personalized One-Shot Federated Learning

DGX agent

arXiv:2608.02222v1 Announce Type: new Abstract: One-shot federated learning (OSFL) has emerged as a promising collaborative model learning framework with only a single round of communication, offering

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

CrossLex: A Source-Grounded Benchmark for Cross-Jurisdictional Legal Reasoning in Large Language Models

DGX agent

arXiv:2608.01292v1 Announce Type: new Abstract: Legal reasoning is inherently jurisdiction-dependent: the same facts can call for different legal rules and yield different conclusions across legal sys

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

CrossProjection: Geometric Grounding Beyond Viewpoint Change in Architectural Drawings

DGX agent

arXiv:2608.00473v1 Announce Type: cross Abstract: Architectural drawings violate the usual assumption behind multi-view reasoning: plans and sections are cuts, while elevations are facade projections,

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Crushing the Evidence: A Dual-Penalty Evasion Framework for Fooling White-Box Explainable AI Auditors

DGX agent

arXiv:2608.00566v1 Announce Type: new Abstract: Post-hoc model explainers such as LIME, SHAP, and Integrated Gradients are widely deployed to audit models in high-stakes sensitive domains, including f

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

CultureVidBench: Benchmarking Cultural Understanding in Text-to-Video Generation

DGX agent

arXiv:2608.01942v1 Announce Type: cross Abstract: Text-to-video (T2V) generation models have advanced rapidly, yet their ability to represent diverse cultural contexts remains underexplored. Existing

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

CurveShift: Is Agent Progress Scalar? Separating Level from Shape

DGX agent

arXiv:2608.00355v1 Announce Type: new Abstract: Progress in large language models is often summarized using a single scalar measure, such as a time horizon, a latent ability estimate, or an aggregate

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Data-Driven Pinball-Loss Selection for Vertically Distributed Elastic-Net SVMs

DGX agent

arXiv:2608.00949v1 Announce Type: new Abstract: The pinball-loss support vector machine is robust, but its asymmetry parameter is usually fixed in advance. We propose a data-driven elastic-net support

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

DE-NER : Zero-shot Named Entity Recognition via Dialogue Elicitation of Large Language Models

DGX agent

arXiv:2608.00538v1 Announce Type: new Abstract: Recent advancements of zero-shot Named Entity Recognition (NER) establish strong baselines by formulating sequence labeling into question answering wher

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

DeBERTa-Sentinel: Toward Transparent and Trustworthy Detection of AI-Generated Text

DGX agent

arXiv:2608.01046v1 Announce Type: new Abstract: The rapid spread of large language models (LLMs) across the web raises concerns about misinformation, academic integrity, automated content manipulation

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

DeCLIP: Decoupled Prompting for Multi-Label Class-Incremental Learning with CLIP

DGX agent

arXiv:2509.23335v3 Announce Type: replace Abstract: Multi-label class-incremental learning (MLCIL) continuously expands the label space while recognizing multiple co-occurring categories, making catas

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

Decoupling semantics from vision: A framework for faithful visual-text compression evaluation

DGX agent

arXiv:2608.01848v1 Announce Type: new Abstract: Recent visual-text compression (VTC) methods, typified by DeepSeek-OCR, report impressive high token compression ratios for long-context modeling tasks

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

Decrease the power limit of your 5090 to at least 480W - the performance penalty for inference is negligible.

DGX agent

I run my inference machine in the living room, so noise and heat output are a significant concern. Ran a quick test using my daily driver model (Qwen 3.6-27b) and at 480W, the card outputs only 2.1% l

model-releasesr-localllama
4 Aug 2026
Model Releases

Deep Learning for Cyber Threat Detection and Mitigation in Healthcare-IoT

DGX agent

arXiv:2608.00118v1 Announce Type: cross Abstract: Cybersecurity is a fundamental requirement for protecting wearable devices used in healthcare Internet of Things (H-IoT) systems. Security failures in

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

Deep Learning for Retinal Degeneration Assessment: A Comprehensive Analysis of the MARIO Challenge

DGX agent

arXiv:2506.02976v4 Announce Type: replace Abstract: The MARIO challenge, held at MICCAI 2024, focused on advancing the automated detection and monitoring of age-related macular degeneration (AMD) thro

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

[Deepseek-V4-Flash-0731] Full 1M context on a single RTX5090 + DDR5 Desktop Setup with VLLM CPU/Ram Offloading, ~800 tps pp & 15+ tps decode [Agentic Coding]

DGX agent

First of all, obviously I took some help from AI to type this post and this is the topic that enabled me to accomplish all that: https://old.reddit.com/r/LocalLLaMA/comments/1veow4b/deepseek_v4flash_2

model-releasesr-localllama
4 Aug 2026
Model Releases

DeepSeek-V4-Flash-0731 is Ollama's fastest growing model ever in token usage. We are scaling capacity in US & Europe. On Ollama, this model …

DGX agent

DeepSeek-V4-Flash-0731 is Ollama's fastest growing model ever in token usage. We are scaling capacity in US & Europe. On Ollama, this model runs with high performance (100tps+) and zero data retention

model-releasesollama--x
4 Aug 2026
Model Releases

DeepSeek V4 Flash 0731 (Q4) now reaches 1,328 tok/s prefill and ~29 tok/s decode on one RTX PRO 6000

DGX agent

I've been working on speeding up DeepSeek-V4-Flash-0731 in Krasis and have now got the long-prompt prefill quite a bit faster on a single RTX PRO 6000 96GB. These are timing-disabled internal Krasis r

model-releasesr-localllama
4 Aug 2026
Model Releases

Deepseek V4 flash 0731 ranks #21 on Agent Arena

DGX agent

https://preview.redd.it/522fsdwvtdhh1.png?width=1200&format=png&auto=webp&s=6a6cf7a467514167a8193029dbd20fb3a9ba4f6c It ranks lower than both Sonnet 4.6 and Luna. I'd wager Luna costs in the same ball

model-releasesr-localllama
4 Aug 2026
Model Releases

Deepseek V4 Flash 2-bit quant is the first model I can run locally that achieves 100% in this SQL benchmark

DGX agent

I really like to use this one SQL benchmark when testing new models. I had another post some time ago with my benchmarks, but I decided to post a new one because of how well Deepseek did. I like the b

model-releasesr-localllama
4 Aug 2026
Model Releases

DeepSeek V4 Flash is now live on Together AI. Frontier agent performance is getting dramatically cheaper. DSV4 on Together AI brings a major…

DGX agent

DeepSeek V4 Flash is now live on Together AI. Frontier agent performance is getting dramatically cheaper. DSV4 on Together AI brings a major jump in coding, tool use, and long-running agent performanc

model-releasestogether-ai--x
4 Aug 2026
Model Releases

DeepSeek v4 Flash vs. Qwen3.6-27B, 3.5-122B, and Gemma 4 31B Benchmark

DGX agent

Just wanted to share my agentic coding benchmark run of DSv4F 0731 at both High and Low reasoning efforts (not Max)... I ran a 109-question subset of Aider Polyglot (the JS/C++/Python languages), base

model-releasesr-localllama
4 Aug 2026
Model Releases

DeepSurvey-Bench: Evaluating Academic Value of Automatically Generated Scientific Surveys

DGX agent

arXiv:2601.15307v2 Announce Type: replace-cross Abstract: The rapid development of automated survey generation technology has made it increasingly important to establish a comprehensive benchmark to e

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Defending Membership Inference Attacks via Privacy-aware Sparsity Tuning

DGX agent

arXiv:2410.06814v2 Announce Type: replace Abstract: Over-parameterized models are typically vulnerable to membership inference attacks, which aim to determine whether a specific sample is included in

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

DeltaFlow: Noise-Adaptive Bidirectional Gated Delta Networks for Embedded Language Flows

DGX agent

arXiv:2608.01240v1 Announce Type: new Abstract: Embedded Language Flows (ELF) rely primarily on full non-causal attention for iterative denoising, repeatedly incurring quadratic sequence-mixing cost a

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Deployment-Ready UWB Localization for Industrial Ground Robots with Automatic Anchor Calibration and Terrain-Aware Fusion

DGX agent

arXiv:2607.15807v2 Announce Type: replace Abstract: Ultra-Wideband (UWB) ranging has become a viable option for industrial Autonomous Mobile Robot (AMR) localization due to improved accuracy and low c

model-releasesarxiv-cs-ro
4 Aug 2026
Model Releases

Detail Continuation over a Trustworthy Coarse Scale for Autoregressive Super-Resolution

DGX agent

arXiv:2608.01823v1 Announce Type: new Abstract: Hallucination remains a persistent challenge in generative super-resolution (GSR), where reconstructed results may contain visually plausible yet weakly

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

DiffusionGemma Technical Report

DGX agent

arXiv:2608.00146v1 Announce Type: new Abstract: We introduce DiffusionGemma, an experimental open-weight language model that uses discrete diffusion to generate text at exceptionally high speed. Rathe

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Disentangling Visuo-Tactile Foresight: Oracle-Guided Interface Discovery for World Action Models

DGX agent

arXiv:2608.00547v1 Announce Type: new Abstract: Contact-rich manipulation remains challenging because successful control depends on physical interaction cues that are often weakly observable from visi

model-releasesarxiv-cs-ro
4 Aug 2026
Model Releases

Divergent large language model predictions from convergent representations in ambiguous word pairs

DGX agent

arXiv:2608.01816v1 Announce Type: new Abstract: In this work we investigate how decoder-only transformers resolve lexical ambiguity through layer-by-layer analysis of three models spanning three param

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

DLLM-TTS: Block Discrete Diffusion Language Model for Text-to-Speech Synthesis

DGX agent

arXiv:2608.00011v1 Announce Type: new Abstract: Current text-to-speech systems face a trade-off: autoregres- sive codec language models produce highly intelligible speech but require large-scale model

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Do Maps Still Matter for Machines: Revisiting the Role of Choropleth Maps in Foundation Model Spatial Understanding

DGX agent

arXiv:2607.17999v2 Announce Type: replace-cross Abstract: Spatial understanding is crucial for foundation models (FMs), and maps have long helped humans organize and reason about geographic informatio

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

Do Static Embeddings Add Value to Hybrid Dutch Retrieval?

DGX agent

arXiv:2608.02112v1 Announce Type: new Abstract: Embedding benchmarks measure standalone model quality, but they do not establish whether a low-cost retriever contributes complementary ranking informat

model-releasesarxiv-cs-lg
4 Aug 2026
← Previous
1…4243444546…465
Next →