AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlog
84,661Total entries
1Added by human
84,660Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,597 results
Model Releases

Decoding-Level Taboo: A Diagnostic Stress Test for LLM Robustness

DGX agent

arXiv:2608.09900v1 Announce Type: new Abstract: Large language model evaluations typically focus on performance under nominal conditions, creating an illusion of capability where models comfortably wa

model-releasesarxiv-cs-cl
11 Aug 2026
Applications
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Distilling Physical Priors into Streaming World Models

DGX agent

arXiv:2608.07981v1 Announce Type: new Abstract: Streaming world models predict future visual states online while maintaining physically coherent dynamics over long horizons. However, their rollouts of

applicationsarxiv-cs-cv
11 Aug 2026
Safety

Distilling Vision-Language Models for Robust Traffic Sign Perception in Autonomous Vehicles

DGX agent

arXiv:2608.08815v1 Announce Type: new Abstract: Traffic sign recognition (TSR) models based on deep neural networks achieve strong clean-data performance but remain vulnerable to physically realizable

safetyarxiv-cs-lg
11 Aug 2026
Safety

DreOPD: Degraded-Reference Extrapolative On-Policy Distillation for Flow-matching Models

DGX agent

arXiv:2608.09233v1 Announce Type: cross Abstract: Flow-matching models are now a mainstream method to image generation, but its adaptation to diverse downstream scenarios typically relies on post-trai

safetyarxiv-cs-cv
11 Aug 2026
Safety

Energy-Structured Latent World Models with Neural Time Fields for Physically Constistent Open-World Motion Planning

DGX agent

arXiv:2608.09876v1 Announce Type: cross Abstract: Physically consistent motion planning remains a fundamental challenge in embodied AI, as generated trajectories must strictly conform to real-world ex

safetyarxiv-cs-ai
11 Aug 2026
Local Ai

General OOD Detection via Model-aware and Subspace-aware Variable Priority

DGX agent

arXiv:2512.13003v2 Announce Type: replace-cross Abstract: Out-of-distribution (OOD) detection is essential for determining when a supervised model encounters inputs that differ meaningfully from its t

local-aiarxiv-cs-lg
11 Aug 2026
Local Ai

HarnessWAM: Bridging Prediction and Deliberation in World Action Models

DGX agent

arXiv:2608.09516v1 Announce Type: new Abstract: World Action Models (WAMs) jointly learn environmental dynamics and robot actions, introducing priors over physical evolution into embodied control. How

local-aiarxiv-cs-ro
11 Aug 2026
Model Releases

InfoOps Bench: A live information operations safety benchmark

DGX agent

arXiv:2607.28503v3 Announce Type: replace Abstract: In this paper we present an active, constantly updated AI benchmark which measures the integrity of frontier language models against being co-opted

model-releasesarxiv-cs-ai
11 Aug 2026
Applications

Length-MAX Tokenizer for Language Models

DGX agent

arXiv:2511.20849v2 Announce Type: replace-cross Abstract: We introduce a new tokenizer for language models that minimizes the average tokens per character, thereby reducing the number of tokens needed

applicationsarxiv-cs-ai
11 Aug 2026
Model Releases

LLM within MCP Matters: Measuring Inefficient Resource Utilization Driven by LLMs

DGX agent

arXiv:2608.08467v1 Announce Type: new Abstract: The Model Context Protocol (MCP) standardizes how servers expose data and tools to Large Language Models (LLMs). A common server design embeds frequentl

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

LLMVisor: A Real-Time Latency Attribution Model for Multi-Tenant LLM Serving

DGX agent

arXiv:2608.08382v1 Announce Type: new Abstract: As LLM inference shifts to multi-tenant GPU clusters, co-batching improves throughput but obscures per-tenant usage and limits control. Enabling fractio

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Mechanistic Interpretability-Guided Selective Fine-Tuning of Vision-Language Models for Centimeter-Level Flood Depth Estimation

DGX agent

arXiv:2608.07562v1 Announce Type: new Abstract: Urban flooding poses an escalating threat to transportation infrastructure, yet no operational system provides real-time, street-level flood-depth estim

model-releasesarxiv-cs-cv
11 Aug 2026
Safety

Persuasive and Compliant Tendencies Predict Group Decision-Making in Humans and Language Models

DGX agent

arXiv:2608.08199v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly involved in group decision-making with other LLMs and humans. Yet it remains unclear whether their influen

safetyarxiv-cs-ai
11 Aug 2026
Agents

Population-Scalable Multi-Agent World Modeling

DGX agent

arXiv:2608.08600v1 Announce Type: cross Abstract: World models have recently achieved impressive progress in visual prediction and interactive generation, but extending them to multi-agent environment

agentsarxiv-cs-ai
11 Aug 2026
Applications

Reading Cognition as Decisions Unfold in Words: A Factorized Inverse Decision Model

DGX agent

arXiv:2608.09222v1 Announce Type: new Abstract: Inverse decision modeling infers latent properties of decision processes from observed behavior, but existing formulations rely primarily on action traj

applicationsarxiv-cs-cl
11 Aug 2026
Research

SC^{2}-WM: A Self-Correcting World Model with Closed-Loop Feedback for Vision-and-Language Navigation in Continuous Environments

DGX agent

arXiv:2608.07548v1 Announce Type: cross Abstract: Vision-and-Language Navigation in Continuous Environments (VLN-CE) requires agents to make fine-grained navigation decisions under partial observabili

researcharxiv-cs-cv
11 Aug 2026
Research

Sign Language Recognition Using Original and Synthetic Depth Image Based Point Cloud Data Models

DGX agent

arXiv:2608.09400v1 Announce Type: cross Abstract: Research regarding the sign language recognition mostly relies on RGB images, whileas sign language datasets that provide depth images are limited. Po

researcharxiv-cs-ai
11 Aug 2026
Model Releases

Time Present and Time Past: Benchmarking Large Language Models on Temporally Evolving Document Understanding

DGX agent

arXiv:2608.08512v1 Announce Type: new Abstract: Evolving documents, such as laws, tax codes, and software documentation, are amended, replaced, and sometimes reverted over time, so a question has diff

model-releasesarxiv-cs-ai
11 Aug 2026
Safety

Topographic Constraints Shape Brain-Like Component Structure in Auditory Models

DGX agent

arXiv:2509.24039v2 Announce Type: replace-cross Abstract: If topography is a fundamental feature of the brain, it should influence both how neurons are arranged in space (i.e. explain brain maps) and

safetyarxiv-cs-ai
11 Aug 2026
Model Releases

1M context with 17 GB model in 24 GB VRAM: 'for the first time I was able to load a context of almost 1M tokens and extract 7 needles from various parts of the text'

DGX agent

https://preview.redd.it/xxjh11f38jih1.png?width=1852&format=png&auto=webp&s=76850ed51e29a8bc86c2ca718d4320075eed4363 Just wanted to share a user report that I found to be very interesting. Some person

model-releasesr-localllama
10 Aug 2026
Local Ai

A MARL Centered Reference Architecture for Large Language Model Augmentation in Smart Manufacturing

DGX agent

arXiv:2608.07148v1 Announce Type: new Abstract: Modern manufacturing imposes six coupled demands on adaptive control: local decisions with global consequences, partial observability, nonstationarity,

local-aiarxiv-cs-ai
10 Aug 2026
Safety

Gated-BEPO: Confidence-Gated Bellman Credit Assignment for Large Language Model Agents

DGX agent

arXiv:2608.06861v1 Announce Type: new Abstract: Training large language model agents in long-horizon environments requires assigning credit from sparse terminal outcomes to individual actions. Existin

safetyarxiv-cs-ai
10 Aug 2026
Research

HRDiT: Training-Free High-Resolution Image Generation with Off-the-Shelf Diffusion Transformer Models

DGX agent

arXiv:2608.07003v1 Announce Type: new Abstract: Training-free text-to-high-resolution image generation has recently attracted growing research attention. However, existing studies on this task primari

researcharxiv-cs-cv
10 Aug 2026
Hardware

If Anthropic begin shipping chips, will NVIDIA begin shipping frontier models? Who will win?

DGX agent

On August 10 2026 at 1:39 AM UTC, user Itamar Friedman (@itamar_mar) posted a short tweet asking whether Anthropic’s potential launch of its own chips would prompt NVIDIA to release frontier AI models

hardwareitamar-friedman--x
10 Aug 2026
Model Releases

Muse Glimmer on 1/2 AMD v620

DGX agent

Hey. Just tried it on my old ass gpus 😄 Surprisingly Tensor Split is working on 2 gpus almost doubling PP (wonder how it will work with 4 gpus) Q6 — 1 GPU llama-server --model <MODEL_DIR>/Muse-Glimmer

model-releasesr-localllama
10 Aug 2026
Research

Prune Once: Retraining-Free Task-Agnostic Pruning for Vision-Language Models

DGX agent

arXiv:2608.06901v1 Announce Type: new Abstract: Vision-language models (VLMs) have achieved remarkable generalization across diverse multimodal tasks through large-scale pre-training, yet their rapidl

researcharxiv-cs-cv
10 Aug 2026
Model Releases

Science Edge Evaluation: SEE the Missing Step Toward Real Scientific Discovery

DGX agent

arXiv:2608.06931v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly involved in scientific discovery, yet it remains unclear whether they can support complex real laboratory

model-releasesarxiv-cs-ai
10 Aug 2026
Research

Unsupervised Adaptation of PDE Foundation Models

DGX agent

arXiv:2608.07053v1 Announce Type: new Abstract: Pretrained partial differential equation (PDE) foundation models can generalize across different equations, but adapting them to unseen PDE systems typi

researcharxiv-cs-ai
10 Aug 2026
Tools

Voyage AI models now run natively on Fireworks, the first and only dedicated inference platform @VoyageAI by @MongoDB has partnered with. Em…

DGX agent

Voyage AI models now run natively on Fireworks, the first and only dedicated inference platform @VoyageAI by @MongoDB has partnered with. Embed, retrieve, rerank, generate: your full retrieval pipelin

toolsfireworks-ai--x
10 Aug 2026
Industry

Hugging Face is getting one new repository every seven seconds. That is roughly 12,000 new models and datasets a day, every day. Clem Delang…

DGX agent

Hugging Face is getting one new repository every seven seconds. That is roughly 12,000 new models and datasets a day, every day. Clem Delangue @ClementDelangue said it on stage at AMD's developer conf

industryclem-delangue--x
9 Aug 2026
Research

AegisShield: Democratizing Cyber Threat Modeling with Generative AI

DGX agent

arXiv:2509.10482v2 Announce Type: replace-cross Abstract: The increasing sophistication of technology systems makes traditional threat modeling hard to scale, especially for small organizations with l

researcharxiv-cs-ai
7 Aug 2026
Model Releases

Beyond Full-Model Rollback: AuroSFT for Adapter-State Multi-Task Fine-Tuning

DGX agent

arXiv:2608.05250v1 Announce Type: new Abstract: Multi-task supervised fine-tuning (SFT) often casts a heterogeneous data mixture as a single optimization problem, even though different tasks may reach

model-releasesarxiv-cs-lg
7 Aug 2026
Model Releases

Can Open-Weight LLMs Produce Kernel-Verified Coq Proofs? A Pilot Study

DGX agent

arXiv:2608.05420v1 Announce Type: cross Abstract: Large language models (LLMs) can generate text that resembles a mathematical proof, but resemblance does not establish correctness. A formal proof che

model-releasesarxiv-cs-lg
7 Aug 2026
Applications

Do Tabular Foundation Models Agree with Themselves?

DGX agent

arXiv:2608.06004v1 Announce Type: new Abstract: Tabular Foundation Models (TFMs) are currently the best approach to tabular prediction problems. They are constructed as transformers that approximate t

applicationsarxiv-cs-lg
7 Aug 2026
Research

How Far Do Simple Transformations Translate Across Text Embedding Models?

DGX agent

arXiv:2608.05980v1 Announce Type: new Abstract: We investigate whether simple transformations can translate representations across heterogeneous text embedding models. Understanding how independently

researcharxiv-cs-lg
7 Aug 2026
Safety

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use

DGX agent

arXiv:2608.05738v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have become the dominant recipe for generalist manipulation, yet they are almost universally trained by behavior clo

safetyarxiv-cs-ro
7 Aug 2026
Research

KVAE: Family of Tokenizers for Multimodal Generative Models

DGX agent

arXiv:2608.05798v1 Announce Type: new Abstract: Latent diffusion modeling (LDM), a prominent paradigm, utilizes tokenizers to map input signal to compressed representation. This dependency positions t

researcharxiv-cs-cv
7 Aug 2026
Model Releases

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs

DGX agent

arXiv:2608.05246v1 Announce Type: new Abstract: Existing personalized LLM benchmarks primarily rely on textual personas or isolated behavioral signals, providing limited evaluation of cross-domain beh

model-releasesarxiv-cs-ai
7 Aug 2026
Tutorials

Mapping Similarity Spaces across Embedding Models with Synthetic Query Probing

DGX agent

arXiv:2608.05857v1 Announce Type: new Abstract: Retrieval-Augmented Generation systems rely on similarity scores to retrieve relevant content, yet scores are not directly comparable across embedding m

tutorialsarxiv-cs-cl
7 Aug 2026
Tutorials

Mean-Field Dynamics of Chain-of-Thought Reasoning in Large Language Models

DGX agent

arXiv:2608.05152v1 Announce Type: cross Abstract: Large language models (LLMs) with chain-of-thought reasoning have been widely applied in recent years, and theoretical explanations of their behavior

tutorialsarxiv-cs-ai
7 Aug 2026
Safety

PhyLatent: Learning Dynamics-Relevant Representations for JEPA World Models

DGX agent

arXiv:2608.05720v1 Announce Type: new Abstract: We propose PhyLatent, a dynamics-relevant training objective for JointEmbedding Predictive Architecture (JEPA) world models. Our key observation is that

safetyarxiv-cs-cv
7 Aug 2026
Model Releases

Simulator-Grounded Large Language Models for Industrial Causal Reasoning: Tool-Use, Structured Injection, and Plant-Portable Retrieval for Wastewater Treatment Decision Support

DGX agent

arXiv:2608.05151v1 Announce Type: cross Abstract: Wastewater operators need answers grounded in how their plant's variables interact and how fast effects propagate, not in generic pretraining text, wh

model-releasesarxiv-cs-ai
7 Aug 2026
Industry

Sources: ByteDance is pretraining an AI model with up to 10T parameters, roughly 3x larger than Kimi K3 and larger than the 8T estimate for Anthropic's Mythos 5 (Financial Times)

DGX agent

Financial Times: Sources: ByteDance is pretraining an AI model with up to 10T parameters, roughly 3x larger than Kimi K3 and larger than the 8T estimate for Anthropic's Mythos 5 — TikTok owner trainin

industrytechmeme
7 Aug 2026
Research

The em-dash em-beds in Congress: A population-level rise in em-dash frequency in U.S. congressional press releases at the dawn of the large-language-model era, 2021-2025

DGX agent

arXiv:2608.05889v1 Announce Type: cross Abstract: Large language models (LLMs) can leave small stylistic traces in text written with their help. The most discussed is the em-dash (U+2014), especially

researcharxiv-cs-ai
7 Aug 2026
Model Releases

this talk by openai researchers going through hugging face incident is totally insane, so much to unpack openai only realized it was their a…

DGX agent

this talk by openai researchers going through hugging face incident is totally insane, so much to unpack openai only realized it was their agent who hacked hugging face infra while asking hf to revoke

model-releasesswyx--x
7 Aug 2026
Local Ai

TRW: TRACE-RealWorld---An Auditable Consistency Contract for World Models as Materialized Views

DGX agent

arXiv:2607.21910v2 Announce Type: replace Abstract: World models let agents plan against predicted physical state, but that state drifts; re-observation is costly and delayed, and repair can fail. We

local-aiarxiv-cs-ai
7 Aug 2026
Model Releases

Vorch-Director: Interactive World Story Model via Noise-Aware Error Rectification

DGX agent

arXiv:2608.05776v1 Announce Type: new Abstract: Autoregressive continuation provides a natural path toward minute-scale audio-visual generation by repeatedly extending a short-window generator conditi

model-releasesarxiv-cs-cv
7 Aug 2026
Model Releases

What Current AI Benchmarks Leave Unmeasured: Modality, Search, Citations, and Implications (for Safety Evaluations)

DGX agent

arXiv:2608.06202v1 Announce Type: cross Abstract: Large language model (LLM) benchmark evaluations are routinely used to support claims about model safety, reliability, and deployment readiness. Yet m

model-releasesarxiv-cs-ai
7 Aug 2026
← Previous
1…183184185186187…1263
Next →