AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,564 results
Model Releases

Performance of large language models in the optical diagnosis of colorectal polyps

DGX agent

arXiv:2608.07543v1 Announce Type: cross Abstract: Background and Study Aims: Accurate optical diagnosis of colorectal polyps guides resection strategy and surveillance, with multimodal large language

model-releasesarxiv-cs-ai
11 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

ReMIND: Orchestrating Modular Large Language Models for Controllable Serendipity A REM-Inspired System Design for Emergent Creative Ideation

DGX agent

arXiv:2601.07121v3 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly used not only for problem solving but also for creative ideation; however, generating ideas that

researcharxiv-cs-ai
11 Aug 2026
Research

Renormalising Generative Models for Active Inference: Foundations, Derivations, and Verification

DGX agent

arXiv:2608.09512v1 Announce Type: new Abstract: Active inference offers a unified framework for perception, learning, and action, but scaling discrete active-inference models to rich spatial and tempo

researcharxiv-cs-ai
11 Aug 2026
Hardware

Route AI Agent Workloads Across Models with NVIDIA NeMo Switchyard

DGX agent

NVIDIA NeMo Switchyard is a routing platform that directs AI agent workloads to the most suitable specialized or frontier model for each step of a task, balancing performance, cost, and latency. It of

hardwarenvidia-developer
11 Aug 2026
Industry

Source: Trajectory, founded by ex-DeepMind, Apple, OpenAI, and Meta staffers to build continual learning models, raised 40M led by Sequoia at a 300M valuation (Stephanie Palazzolo/The Information)

DGX agent

Stephanie Palazzolo / The Information: Source: Trajectory, founded by ex-DeepMind, Apple, OpenAI, and Meta staffers to build continual learning models, raised 40M led by Sequoia at a 300M valuation —

industrytechmeme
11 Aug 2026
Model Releases

Toward Mask Annotation-Free Surgical Instrument Segmentation from Endoscopic Images Using Text-Prompted Segment Anything Model 3 (SAM3)

DGX agent

arXiv:2608.08844v1 Announce Type: new Abstract: Surgical instrument segmentation is a fundamental task for computer-assisted interventions, yet most existing methods rely on pixel-level annotations or

model-releasesarxiv-cs-cv
11 Aug 2026
Research

CellWorld: From Gene-Level Reconstruction to Latent Cell Prediction in Spatial Transcriptomics Foundation Models

DGX agent

arXiv:2608.06659v1 Announce Type: new Abstract: This paper shows that latent-space predictive pretraining can provide a scalable route to foundation models for spatial transcriptomics. Existing spatia

researcharxiv-cs-ai
10 Aug 2026
Research

Convergence of Diffusion Models Under the Manifold Hypothesis in High-Dimensions

DGX agent

arXiv:2409.18804v3 Announce Type: replace-cross Abstract: Denoising Diffusion Probabilistic Models (DDPM) are powerful state-of-the-art methods used to generate synthetic data from high-dimensional da

researcharxiv-cs-lg
10 Aug 2026
Research

Depth-Wise Probing and Pruning of the Planning Token in a Driving Vision-Language-Action Model

DGX agent

arXiv:2608.07361v1 Announce Type: cross Abstract: Vision-language-action (VLA) models route driving decisions through a deep language model, but it is unclear how much of that depth the action itself

researcharxiv-cs-cv
10 Aug 2026
Model Releases

Faster Query-Key Learning Sharpens Attention in Self-Attention Models

DGX agent

arXiv:2608.06776v1 Announce Type: new Abstract: A standard self-attention layer consists of two interacting circuits: the query-key circuit that governs attention allocation, and the output-value circ

model-releasesarxiv-cs-lg
10 Aug 2026
Research

How Molecular Generative Models Organize Molecular Identity

DGX agent

arXiv:2608.06956v1 Announce Type: new Abstract: Generative models for matter are often evaluated as samplers over output representations, and their latent spaces are commonly used as proxies for navig

researcharxiv-cs-lg
10 Aug 2026
Model Releases

Native Long Video Understanding Models locally?

DGX agent

I've been building a personal project and wanted to check with the community on multi-modal inputs since I can't find a lot of material around this online. Ultimately I'm trying to build something tha

model-releasesr-localllama
10 Aug 2026
Research

Separating Decision-Rule Misalignment from Readout-Coverage Limitations in Speech Language Models

DGX agent

arXiv:2608.06409v1 Announce Type: cross Abstract: Speech language models are increasingly evaluated on paralinguistic tasks by the accuracy of prompted answers, but answer accuracy combines failures a

researcharxiv-cs-ai
10 Aug 2026
Industry

In-depth look at OpenAI's model training, dangerous decisions, and cluelessness before the HuggingFace hack; despite delaying Astra, OpenAI still doesn't get it (Zvi Mowshowitz/Don't Worry About the Vase)

DGX agent

Zvi Mowshowitz / Don't Worry About the Vase: In-depth look at OpenAI's model training, dangerous decisions, and cluelessness before the HuggingFace hack; despite delaying Astra, OpenAI still doesn't g

industrytechmeme
9 Aug 2026
Model Releases

是我见过OCR效果最好的一个了,人眼都难以分辨的,它能处理的十分准确,而且速度非常的快。llamaindex 真不愧是文档解析界的一哥啊。

DGX agent

是我见过OCR效果最好的一个了,人眼都难以分辨的,它能处理的十分准确,而且速度非常的快。llamaindex 真不愧是文档解析界的一哥啊。 The best 'raw' frontier model for document parsing is gemini 3 flash, but the issue is that since then the flash models have gotte

model-releasesjerry-liu--x
9 Aug 2026
Model Releases

Now we have a timeline of the OpenAI accidental attack against Hugging Face

DGX agent

My comment on Now we have a timeline of the OpenAI accidental attack against Hugging Face — Hacker News.I think one of the most interesting details here might be tucked away in that first bulletin poi

model-releasessimon-willison
8 Aug 2026
Model Releases

AI Playing Business Games: Benchmarking Large Language Models on Managerial Decision-Making in Dynamic Simulations

DGX agent

arXiv:2509.26331v2 Announce Type: replace Abstract: The rapid advancement of LLMs sparked significant interest in their potential to augment or automate managerial functions. One of the most recent tr

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

An Early Warning of Emerging Biosecurity Risks in Frontier LLMs

DGX agent

arXiv:2607.18056v2 Announce Type: replace Abstract: Frontier large language models (LLMs) are increasingly integrated into scientific workflows, yet their growing biological capabilities may outpace c

model-releasesarxiv-cs-cl
7 Aug 2026
Model Releases

Clinician input steers AI toward accurate and harmful recommendations

DGX agent

arXiv:2603.14158v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are entering clinical workflows, yet evaluations rarely assess how clinician reasoning shapes model behavior duri

model-releasesarxiv-cs-lg
7 Aug 2026
Safety

DreamGuard: Efficient Runtime Guardrail for LLM Agents via Risk-Aware World Model

DGX agent

arXiv:2608.05695v1 Announce Type: new Abstract: As large language model (LLM) agents increasingly invoke external tools and interact with real-world systems, unsafe actions may cause irreversible cons

safetyarxiv-cs-ai
7 Aug 2026
Safety

LAWM-3D: Learning 3D-Aware Latent Actions from Human Videos for Generalizable Robot World Models

DGX agent

arXiv:2608.05706v1 Announce Type: new Abstract: World models enable agents to perform forward rollout and planning without real-world interaction. However, their application in open-world embodied int

safetyarxiv-cs-cv
7 Aug 2026
Research

On the Anisotropy of Score-Based Generative Models

DGX agent

arXiv:2510.22899v2 Announce Type: replace Abstract: We investigate the role of network architecture in shaping the inductive biases of modern score-based generative models. To this end, we introduce t

researcharxiv-cs-lg
7 Aug 2026
Model Releases

The Low Frequency Trap: Video Language Models Fail at Simple Event Bookkeeping

DGX agent

arXiv:2608.06361v1 Announce Type: new Abstract: Real-world video benchmarks provide broad coverage, but their fixed clips entangle event count, rate, duration, and visual complexity, making failure mo

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Toward Deployable Bangla Sign Language Recognition with Expert-Validated Data and a Lightweight Attention-Based Model

DGX agent

arXiv:2608.06252v1 Announce Type: cross Abstract: Deaf and hard-of-hearing people in Bangladesh communicate mainly through Bangla Sign Language (BdSL). Automatic BdSL recognition on personal devices c

model-releasesarxiv-cs-ai
7 Aug 2026
Applications

Uncertainty-Aware World Model for Aerial Image-Goal Navigation

DGX agent

arXiv:2608.05597v1 Announce Type: new Abstract: Aerial image-goal navigation requires an unmanned aerial vehicle (UAV) to reach a target location specified by a goal image. Existing world-model-based

applicationsarxiv-cs-cv
7 Aug 2026
Safety

XEWorld: Can Action-Conditioned World Models Generalize to Unseen Robot Embodiments?

DGX agent

arXiv:2608.05799v1 Announce Type: cross Abstract: Action-conditioned world models are promising learned simulators for robotic manipulation, yet evaluating them exclusively on training robots fails to

safetyarxiv-cs-cv
7 Aug 2026
Local Ai

AMA: MiniMax H3 Team — Ask us anything about our open video generation model, training, and future plans

DGX agent

https://preview.redd.it/kihat320ashh1.png?width=1672&format=png&auto=webp&s=a7ccc40ba3fb229ac7ebf57e8e6a314e0ee45646 Hi r/StableDiffusion! u/New-Requirement1419 -> dacongya (Head of H3 Researcher) u/A

local-air-stablediffusion
6 Aug 2026
Research

Convergence and Stability Analysis of Self-Consuming Generative Models with Heterogeneous Human Curation

DGX agent

arXiv:2511.09002v3 Announce Type: replace-cross Abstract: Self-consuming generative models have received significant attention over the last few years. In this paper, we study a self-consuming generat

researcharxiv-cs-lg
6 Aug 2026
Research

Do LLMs Know What Is Private Internally? Probing and Steering Contextual Privacy Norms in Large Language Model Representations

DGX agent

arXiv:2604.00209v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly deployed in high-stakes settings, yet they frequently violate contextual privacy by disclosing private

researcharxiv-cs-cl
6 Aug 2026
Model Releases

EuroExec: Frontier Language Models Fall Short of Expert Judgment on European Executive Decision Tasks

DGX agent

arXiv:2608.04549v1 Announce Type: cross Abstract: Frontier LLMs are increasingly put to use on open-ended complex questions, different in nature from the ones they are typically evaluated on. We dedic

model-releasesarxiv-cs-ai
6 Aug 2026
Research

Evaluating Theory of Mind in Reasoning Models: Robustness over Reasoning

DGX agent

arXiv:2608.04646v1 Announce Type: new Abstract: Large language models (LLMs) have recently shown strong performance on Theory of Mind (ToM) tests, prompting debate about the nature and validity of the

researcharxiv-cs-cl
6 Aug 2026
Research

Explicit Language Memory for Long-Horizon Planning in Vision-Language-Action Models

DGX agent

arXiv:2608.04765v1 Announce Type: cross Abstract: Vision-language-action (VLA) models provide a unified paradigm for connecting visual perception, language understanding, and robotic control. However,

researcharxiv-cs-ai
6 Aug 2026
Tools

FLUX 3 is now live on Together AI. @bfl_ai’s new multimodal model generates video and synchronized audio together, with up to 20-second clip…

DGX agent

FLUX 3 is now live on Together AI. @bfl_ai’s new multimodal model generates video and synchronized audio together, with up to 20-second clips, multiple shots, and control from text, images, or keyfram

toolstogether-ai--x
6 Aug 2026
Model Releases

Helping Music Co-Creation Agents 'Listen' Well: Hierarchical Self-Supervised World Models for Understanding and Generation

DGX agent

arXiv:2608.04378v1 Announce Type: cross Abstract: Collaborative music agents need internal representations rich enough to support both understanding and generation, yet flexible enough for a workflow

model-releasesarxiv-cs-lg
6 Aug 2026
Research

Koopman-Based Nonlinear Identification and Model Predictive Control of a Turbofan Engine

DGX agent

arXiv:2604.01730v2 Announce Type: replace Abstract: This paper investigates Koopman operator-based approaches for multivariable control of a two-spool turbofan engine. A physics-based component-level

researcharxiv-cs-lg
6 Aug 2026
Model Releases

LiveXiv -- A Multi-Modal Live Benchmark Based on Arxiv Papers Content

DGX agent

arXiv:2410.10783v4 Announce Type: replace Abstract: The large-scale training of multi-modal models on data scraped from the web has shown outstanding utility in infusing these models with the required

model-releasesarxiv-cs-cv
6 Aug 2026
Research

MarsCast: Transfer Learning of AI Weather Foundation Models to Planetary Atmospheres

DGX agent

arXiv:2608.05054v1 Announce Type: cross Abstract: We investigate the transferability of Earth weather foundation models to planetary atmospheres by adapting the GraphCast graph neural weather forecast

researcharxiv-cs-ai
6 Aug 2026
Model Releases

Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models

DGX agent

arXiv:2608.04633v1 Announce Type: new Abstract: Recent Vision-Language-Action (VLA) methods improve generalization by aligning their representations with 3D scene geometry. However, these methods are

model-releasesarxiv-cs-ro
6 Aug 2026
Model Releases

The death of SLMs?

DGX agent

I love to see these impressive models coming out that compete with the giants from companies like Z.ai, Moonshot, Alibaba, etc. A win for the open source/weight community is always welcome. While I am

model-releasesr-localllama
6 Aug 2026
Safety

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation

DGX agent

arXiv:2608.04436v1 Announce Type: new Abstract: Text-to-image (T2I) models can produce visually compelling images, yet they remain limited on open-world tasks that require complex semantic understandi

safetyarxiv-cs-cv
6 Aug 2026
Research

When Absence Is Evidence: Evaluating Completeness-Sensitive Negative Reasoning in Large Language Models

DGX agent

arXiv:2608.04591v1 Announce Type: cross Abstract: Large language models (LLMs) are often asked whether something is absent from a record, list, or retrieved context. Yet non-observation licenses a neg

researcharxiv-cs-ai
6 Aug 2026
Local Ai

Any-OPD: Heterogeneous On-Policy Distillation for Flow-Matching Models via Representation-Space Bridging

DGX agent

arXiv:2608.03316v1 Announce Type: cross Abstract: On-policy distillation, in which a teacher corrects samples that the student itself generates, presupposes that the two models speak the same language

local-aiarxiv-cs-cv
5 Aug 2026
Model Releases

ArtECulture: Benchmarking Culture-Conditioned Visual Emotion Understanding in Multimodal Large Language Models

DGX agent

arXiv:2608.03358v1 Announce Type: new Abstract: Existing visual emotion understanding methods typically ignore cultural variations in emotional perception. We introduce culture-conditioned visual emot

model-releasesarxiv-cs-cl
5 Aug 2026
Safety

BOW: Training Language Models to Reason Over Plausible Next Words

DGX agent

arXiv:2506.13502v3 Announce Type: replace Abstract: Next-word prediction (NWP) trains language models against a single observed continuation, even though many contexts admit multiple plausible next wo

safetyarxiv-cs-cl
5 Aug 2026
Model Releases

Can Large Language Models Recover Semantic Optimization Opportunities That Compilers Miss?

DGX agent

arXiv:2608.03983v1 Announce Type: cross Abstract: Optimizing compilers miss profitable transformations when their enabling semantics are absent from the analyzed program representation. We ask whether

model-releasesarxiv-cs-ai
5 Aug 2026
Research

Can Training Logs Make Model Comparisons More Precise?

DGX agent

arXiv:2608.02705v1 Announce Type: cross Abstract: Comparing stochastically trained models requires estimating both a performance difference and its uncertainty from repeated runs. We study whether tra

researcharxiv-cs-ai
5 Aug 2026
Applications

Cross-Model KV Cache Transfer in LLM Families: A Closed-Form Linear Mapping for Prefill Reuse

DGX agent

arXiv:2608.03893v1 Announce Type: new Abstract: Production deployments often swap between different-sized models in a family for cost-quality cascading, mid-conversation switching, and routing, and ea

applicationsarxiv-cs-lg
5 Aug 2026
Model Releases

Cura 1T: Specialized Model for Agentic Healthcare

DGX agent

arXiv:2607.15314v2 Announce Type: replace Abstract: Healthcare AI agents handle patient consultation, clinical reasoning over text and images, interactive diagnosis, and electronic health record (EHR)

model-releasesarxiv-cs-ai
5 Aug 2026
← Previous
1…155156157158159…1262
Next →