AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,503 results
Model Releases

COSI-Lab: Conference Living Lab for Modeling Multi-Perspective Multimodal Social Intention

DGX agent

arXiv:2607.28649v1 Announce Type: cross Abstract: COSI-Lab presents a multimodal, multi-sensor dataset of an interdisciplinary scientific workshop containing 32 academics at an international conferenc

model-releasesarxiv-cs-ai
3 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Distilling Knowledge from Large Language Models into Lightweight Reinforcement Learning Agents for Autonomous Cyber Operations

DGX agent

arXiv:2607.28826v1 Announce Type: new Abstract: Autonomous Cyber Operations (ACO) are increasingly important for defending enterprise networks as cyber threats continue to evolve in sophistication. AC

model-releasesarxiv-cs-lg
3 Aug 2026
Tutorials

FBFM: A Training-Free Asynchronous Feedback Mechanism for Flow-Matching in World-Action Models Execution

DGX agent

arXiv:2607.29235v1 Announce Type: cross Abstract: Although world-action models (WAMs) enhance long-horizon robot control by predicting visual evolution before acting, long-horizon reliability demands

tutorialsarxiv-cs-ai
3 Aug 2026
Model Releases

Ling-3.0-flash is another potential model to test before qwen3.8 27b

DGX agent

I tested Ling-3.0-flash with hard bugs and it fixed bugs that qwen3.6-27b could not. This models speed faster than deepseek v4 flash but almost the same level as (old) deepseek v4 flash. Note: hard bu

model-releasesr-localllama
3 Aug 2026
Model Releases

You shouldn't need a vision model to know your PDF has checkboxes. LiteParse can now pull structured data directly from your PDFs: form fiel…

DGX agent

You shouldn't need a vision model to know your PDF has checkboxes. LiteParse can now pull structured data directly from your PDFs: form field values, checkbox states, annotations, embedded images, vec

model-releasesllamaindex--x
3 Aug 2026
Industry

Instead of limiting the progress of AI companies or preventing them from releasing models as proposed in the bipartisan AI Kill Switch Act, …

DGX agent

Instead of limiting the progress of AI companies or preventing them from releasing models as proposed in the bipartisan AI Kill Switch Act, Hugging Face CEO Clément Delague says he would rather see Co

industryclem-delangue--x
2 Aug 2026
Safety

LLMs can know a task is impossible and still optimize it anyway. Ask whether to walk or drive to a car wash 50 meters away, and some models …

DGX agent

LLMs can know a task is impossible and still optimize it anyway. Ask whether to walk or drive to a car wash 50 meters away, and some models focus on distance while missing that the car itself must rea

safetygary-marcus--x
2 Aug 2026
Model Releases

Benchmarking Foundation and Large Language Models for Few-Shot Medical Image Segmentation

DGX agent

arXiv:2607.27856v1 Announce Type: new Abstract: Few-shot medical image segmentation (FS-MIS) aims to segment novel regions of interest (ROIs) from a few annotated support examples. Despite rapid progr

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

DeepSeek V4 Flash 0731 in Hermes Agent and one prompt, took 32 minutes and cost 0.07$, this model is so cheap to the point where 2 dollars c…

DGX agent

**DeepSeek V4 Flash 0731 Performance Test** On July 31 2026, a single prompt executed via the Hermes Agent on DeepSeek V4 Flash 0731 completed in 32 minutes and incurred an estimated cost of 0.07 USD.

model-releasesnous-research--x
31 Jul 2026
Safety

Eco3S: Complex Socio-Economic System Simulation via Agent-Based Models

DGX agent

arXiv:2607.26588v1 Announce Type: new Abstract: The rapid development of large language models (LLMs) has renewed interest in agent-based modeling (ABM). However, current LLM-based ABM research faces

safetyarxiv-cs-ai
31 Jul 2026
Model Releases

I have trained a model to predict my blood sugar [P]

DGX agent

It's an encoder-only transformer that consumes past(blood glucose + carbs + insulin) and future(carbs + insulin) and predicts future blood glucose for the next 2 hours. Announced meals and boluses/bas

model-releasesr-machinelearning
31 Jul 2026
Model Releases

ORCA-bench: How Ready Are Language Model Agents for Oncall?

DGX agent

arXiv:2607.28545v1 Announce Type: new Abstract: Large language models can write, patch, and search code, but oncall root cause analysis (RCA) demands something different: reasoning over noisy metrics,

model-releasesarxiv-cs-cl
31 Jul 2026
Research

ShadowDancer: Teaching Video World Models Any Action by Learning Unified Dynamics Representations from a Video and Its Shadow

DGX agent

arXiv:2607.28362v1 Announce Type: new Abstract: We present ShadowDancer, a novel approach to any-action, frame-level control of interactive video world models. The obstacle is representational: existi

researcharxiv-cs-cv
31 Jul 2026
Model Releases

Sources: OpenAI demoed a new 'Astra' AI model family to US policymakers and regulators this week, touting its improved abilities to complete long-running tasks (The Information)

DGX agent

The Information: Sources: OpenAI demoed a new “Astra” AI model family to US policymakers and regulators this week, touting its improved abilities to complete long-running tasks — OpenAI is preparing t

model-releasestechmeme
31 Jul 2026
Research

(Towards) Scalable Reliable Automated Evaluation with Large Language Models

DGX agent

arXiv:2607.28282v1 Announce Type: new Abstract: Evaluating the quality and relevance of textual outputs from Large Language Models (LLMs) remains challenging and resource-intensive. Existing automated

researcharxiv-cs-cl
31 Jul 2026
Model Releases

We've gotten some great medium sized models lately (DSV4 Flash 0731, Inkling Small, Laguna S 2.1, Step 3.7 Flash) but does anybody else want to see some new 70-80b contenders?

DGX agent

I can run the mediums, but sometimes I want a faster option that's smarter than Qwen 27B/35B. On my hardware I get like 500 to 800 tok/s prefill and 16 to 22 tok/s gen on ~120B class models, which is

model-releasesr-localllama
31 Jul 2026
Safety

Do Unified Multimodal Models Think in One Space? A Lens Through Cross-Branch Steering

DGX agent

arXiv:2607.26411v1 Announce Type: new Abstract: Unified multimodal models (UMMs) aim to integrate understanding and generation within a single architecture, yet it remains unclear whether these capabi

safetyarxiv-cs-cv
30 Jul 2026
Research

Equilibrium Training of Energy-Based Models with Parallel Trajectory Tempering

DGX agent

arXiv:2607.27077v1 Announce Type: new Abstract: Energy-Based Models (EBMs) provide an interpretable framework for generative modeling of scientific data, but poor Markov Chain Monte Carlo mixing often

researcharxiv-cs-lg
30 Jul 2026
Model Releases

GPT-5.6 found optimizations that 'reduced end-to-end serving costs by 20%' for OpenAI to serve that model Presumably that's billions of doll…

DGX agent

GPT-5.6 found optimizations that 'reduced end-to-end serving costs by 20%' for OpenAI to serve that model Presumably that's billions of dollars a month in savings at this point? Codex analysed product

model-releasessimon-willison--x
30 Jul 2026
Local Ai

MedARC: Training-Free Adaptive Redundancy Compression of Visual Tokens for 3D Medical Vision-Language Models

DGX agent

arXiv:2607.26554v1 Announce Type: new Abstract: Integrating 3D medical images with vision-language models (VLMs) holds substantial promise for computer-aided diagnosis. However, volumetric images gene

local-aiarxiv-cs-cv
30 Jul 2026
Safety

one of the top cybersecurity models, post-trained from open weights by @depthfirstlabs on @FireworksAI_HQ long-horizon RL is as much an infr…

DGX agent

one of the top cybersecurity models, post-trained from open weights by @depthfirstlabs on @FireworksAI_HQ long-horizon RL is as much an infra problem as a research one: 100+ turn rollouts, async/pipel

safetyfireworks-ai--x
30 Jul 2026
Model Releases

WildShadowRemover: In-the-Wild Video Shadow Removal via Detail-Preserving Video Diffusion Models

DGX agent

arXiv:2607.26203v1 Announce Type: new Abstract: Video shadow removal in the wild remains challenging due to complex illumination, diverse shadow appearances, and limited training data. Despite its imp

model-releasesarxiv-cs-cv
30 Jul 2026
Local Ai

Beyond Counts: A Distributional Robustness Margin For Pathology Foundation Models

DGX agent

arXiv:2607.25497v1 Announce Type: cross Abstract: Pathology foundation models are approaching clinical deployment, yet remain vulnerable to systematic non-biological variation across centres. Differen

local-aiarxiv-cs-ai
29 Jul 2026
Model Releases

BREAKING: Grok 4.5 just claimed the top spot on the new HighWalk Benchmark. The independent test measures how well AI models update real tec…

DGX agent

BREAKING: Grok 4.5 just claimed the top spot on the new HighWalk Benchmark. The independent test measures how well AI models update real technical specifications from 46 Laravel commits — heavy on cod

model-releaseselon-musk--x
29 Jul 2026
Applications

Can Deep Generative Models Reproduce Non-Stationary Gaussian Random Fields?

DGX agent

arXiv:2607.25929v1 Announce Type: cross Abstract: Deep generative models (DGMs) are widely used for complex high-dimensional data and increasingly applied to spatial and spatio-temporal modeling. Thei

applicationsarxiv-cs-lg
29 Jul 2026
Model Releases

Diffusion Model-based Parameter Estimation in Dynamic Power Systems

DGX agent

arXiv:2411.10431v3 Announce Type: replace Abstract: Parameter estimation, which represents a classical inverse problem, is often ill-posed as different parameter combinations can yield identical outpu

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Dual-Domain Manifold Modeling for Hyperspectral Image Fusion

DGX agent

arXiv:2607.25338v1 Announce Type: new Abstract: Achieving a coherent integration of spectral richness and spatial fidelity remains a central objective in hyperspectral image fusion. However, existing

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Forensic Reproducibility Audit of a Radiology Vision-Language Model Benchmark: From Intended Protocol to Released Artifact

DGX agent

arXiv:2607.25589v1 Announce Type: cross Abstract: Medical-imaging AI benchmarks combine datasets, DICOM rendering, prompts, provider APIs, automated labels, statistical code, manuscripts, and reposito

model-releasesarxiv-cs-cl
29 Jul 2026
Model Releases

Inspect India Evals: An Open Benchmarking Framework for Evaluating Large Language Models in the Indian Linguistic and Cultural Context

DGX agent

arXiv:2607.25375v1 Announce Type: new Abstract: India is a vast nation of over 1.4 billion people, varied by hundreds of diverse and locally specific traditions and cultures and 22 officially recogniz

model-releasesarxiv-cs-cl
29 Jul 2026
Safety

Medical world models in healthcare: foundations, applications, and challenges for trustworthy clinical translation

DGX agent

arXiv:2607.25242v1 Announce Type: new Abstract: Medical world models offer a framework for extending medical artificial intelligence beyond static prediction by representing evolving patient states an

safetyarxiv-cs-cv
29 Jul 2026
Model Releases

Memory for Large Language Models

DGX agent

arXiv:2607.25380v1 Announce Type: new Abstract: Memory has evolved into a foundational architectural dimension in large language models (LLMs), shifting from an implicit byproduct of computation to a

model-releasesarxiv-cs-cl
29 Jul 2026
Model Releases

MyoCardBench: A Real-World Data Benchmark for Evaluating Large Language Models in Clinically Authentic Cardiovascular Care Scenarios

DGX agent

arXiv:2607.25186v1 Announce Type: new Abstract: Background: Most medical large language model (LLM) benchmarks focus on examination knowledge or isolated tasks and may not reflect the longitudinal, mu

model-releasesarxiv-cs-cl
29 Jul 2026
Model Releases

PerceptionBench: Evaluating Atomic Visual Perception in Multimodal Large Language Models

DGX agent

arXiv:2607.24957v1 Announce Type: new Abstract: We introduce PerceptionBench, a benchmark specifically designed to evaluate the atomic visual perception capabilities of Multimodal Large Language Model

model-releasesarxiv-cs-cv
29 Jul 2026
Model Releases

RDQ: Residual Distribution Quantization for Large Language Models

DGX agent

arXiv:2607.10137v2 Announce Type: replace Abstract: Post-training quantization (PTQ) of large language models degrades sharply below 4-bit precision. We identify the root cause as residual stream dist

model-releasesarxiv-cs-lg
29 Jul 2026
Model Releases

RRS-10K: A Multitask Vision-Language Model Benchmark for Rare Remote Sensing Image Interpretation

DGX agent

arXiv:2607.24810v1 Announce Type: new Abstract: Vision-language models (VLMs) have achieved strong performance on general remote sensing tasks. However, their capability for rare scenes remains insuff

model-releasesarxiv-cs-ai
29 Jul 2026
Research

Act, Think or Abstain: Complexity-Aware Adaptive Inference for Vision-Language-Action Models

DGX agent

arXiv:2603.05147v2 Announce Type: replace Abstract: Current research on Vision-Language-Action (VLA) models predominantly focuses on enhancing generalization through reasoning techniques. While effect

researcharxiv-cs-cv
28 Jul 2026
Safety

Concept-based Visual Counterfactual Explanations with Diffusion Models

DGX agent

arXiv:2607.22544v1 Announce Type: new Abstract: Visual counterfactual explanations aim to answer 'what minimal change to this image would flip the model's prediction?', and are increasingly important

safetyarxiv-cs-ai
28 Jul 2026
Local Ai

Development of Vision-Language Model-based GNSS Spoofing Detection for Autonomous Vehicle Navigation

DGX agent

arXiv:2607.23962v1 Announce Type: new Abstract: Autonomous vehicles (AVs) depend on Global Navigation Satellite Systems (GNSS) for localization and navigation, making them vulnerable to spoofing attac

local-aiarxiv-cs-cv
28 Jul 2026
Research

Exact Evaluation of the Accuracy of Diffusion Models for Inverse Problems with Gaussian Data Distributions

DGX agent

arXiv:2507.07008v2 Announce Type: replace Abstract: Used as priors for Bayesian inverse problems, diffusion models have recently attracted considerable attention in the literature. Their flexibility a

researcharxiv-cs-lg
28 Jul 2026
Model Releases

Guiding Language Models to Be More Empathetic: Culturally Sensitive Mental Health Advice Generation Through Human-LLM Collaboration

DGX agent

arXiv:2607.23538v1 Announce Type: new Abstract: Despite recent advances in large language models (LLMs), their ability to generate empathetic mental health counseling responses in low-resource languag

model-releasesarxiv-cs-cl
28 Jul 2026
Model Releases

How OpenAI hacked HuggingFace. What we know. Hugging Face proved that open platforms and open models can still win those battles when the al…

DGX agent

How OpenAI hacked HuggingFace. What we know. Hugging Face proved that open platforms and open models can still win those battles when the alternative is locked-down systems that refuse to assist their

model-releasesclem-delangue--x
28 Jul 2026
Safety

Hybrid AI-Physical Modeling for Penetration Bias Correction in X-band InSAR DEMs: A Greenland Case Study

DGX agent

arXiv:2504.08909v2 Announce Type: replace Abstract: Digital elevation models derived from Interferometric Synthetic Aperture Radar (InSAR) data over glacial and snow-covered regions often exhibit syst

safetyarxiv-cs-ai
28 Jul 2026
Safety

IJCB-AFMFR 2026: Competition on Adapting Foundation Models for Face Recognition Using Synthetic Training Data

DGX agent

arXiv:2607.24422v1 Announce Type: new Abstract: This paper presents a summary of the Competition on Adapting Foundation Models for Face Recognition Using Synthetic Training Data (AFMFR), held at the 2

safetyarxiv-cs-cv
28 Jul 2026
Research

LEDOM: Reverse Language Model

DGX agent

arXiv:2507.01335v4 Announce Type: replace-cross Abstract: Autoregressive language models are trained exclusively left-to-right. We explore the complementary factorization, training right-to-left at sc

researcharxiv-cs-ai
28 Jul 2026
Model Releases

LIBMoE: A Library for comprehensive benchmarking Mixture of Experts in Large Language Models

DGX agent

arXiv:2411.00918v5 Announce Type: replace-cross Abstract: Mixture of experts (MoE) architectures have become a cornerstone for scaling up and are a key component in most large language models such as

model-releasesarxiv-cs-ai
28 Jul 2026
Hardware

OmniScope: Modality-Decoupled Token Compression for Omnimodal Large Language Models

DGX agent

arXiv:2607.23193v1 Announce Type: new Abstract: Existing token compression methods for omnimodal large language models typically rely on one modality to determine what to retain in the other. We show

hardwarearxiv-cs-cv
28 Jul 2026
Model Releases

Stability of AI Governance Systems: A Coupled Dynamics Model of Public Trust and Social Disruptions

DGX agent

arXiv:2603.20248v2 Announce Type: replace-cross Abstract: AI systems are increasingly entrenched in public governance, yet scholarship lacks formal tools to determine when deviations of public trust i

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

TRE: Training-Free Hallucination Detection for Diffusion Language Models

DGX agent

arXiv:2607.22661v1 Announce Type: new Abstract: Diffusion large language models (D-LLMs) have recently gained increasing attention, yet their reliability is significantly hindered by the hallucination

model-releasesarxiv-cs-ai
28 Jul 2026
← Previous
1…9899100101102…1261
Next →