AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
All
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,272 results
Model Releases

LoRetta: A Foundation Model and Extensive Dataset for Global-Scale Remote Sensing Dense Image Matching

DGX agent

arXiv:2608.04106v1 Announce Type: new Abstract: Dense image matching establishes pixel-wise correspondences and underpins broad applications in computer vision and photogrammetry. However, extending d

model-releasesarxiv-cs-cv
6 Aug 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Luna non-reasoning is a bit better than GPT 4o which was sota 2 years ago. Luna (medium) thinking is a bit better than GPT-5 (High) which wa…

DGX agent

Luna non-reasoning is a bit better than GPT 4o which was sota 2 years ago. Luna (medium) thinking is a bit better than GPT-5 (High) which was sota 1 year ago. Now free to everyone unlimited Sol Max/Fa

model-releasesemad-mostaque--x
6 Aug 2026
Model Releases

MathDebugger: Detecting and Diagnosing Errors in Synthetic Mathematical Data

DGX agent

arXiv:2502.19058v2 Announce Type: replace Abstract: Synthetic mathematical data has become an important resource for scaling the reasoning capabilities of large language models, yet errors in generate

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

MatrAIx: Simulating the World with 8.3 Billion Persona Agents

DGX agent

arXiv:2608.04205v1 Announce Type: new Abstract: Human evaluation of AI systems and digital products is costly, slow, and difficult to scale. Offline evaluations are more scalable but often abstract aw

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

MediRec: Enhancing Chinese Medication Recommendation with Explainable Clinical Reasoning

DGX agent

arXiv:2510.21084v3 Announce Type: replace-cross Abstract: Large language models (LLMs) have shown strong potential for clinical decision support through their advanced language understanding and reaso

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

MERaLiON-GR: Speech Gender Recognition Model for English and SEA Languages

DGX agent

arXiv:2608.04433v1 Announce Type: cross Abstract: We present MERaLiON-GR, a speech gender recognition system that performs binary classification (female / male) on English and Southeast Asian (SEA) la

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

MESH: Memory-Efficient Sinkhorn Optimization for Mixture-of-Experts Training

DGX agent

arXiv:2608.04407v1 Announce Type: cross Abstract: Memory-efficient matrix optimizers such as Sinkhorn gradient descent remove most AdamW optimizer state for dense Transformer matrices, but direct appl

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Meta’s Muse Spark 1.1 hacked an external organization during cybersecurity test

DGX agent

A large language model developed by Meta Platforms Inc. hacked a third party organization during a cybersecurity evaluation. The Facebook parent disclosed the incident on Wednesday without specifying

model-releasessiliconangle
6 Aug 2026
Model Releases

Mimir: A Neuro-Symbolic Memory System with Dynamic Grounding for Embodied Agents in Interactive Environments

DGX agent

arXiv:2608.04933v1 Announce Type: new Abstract: Long-horizon embodied task requires agents to act under partial observability while preserving both scene belief and execution progress. Flat histories

model-releasesarxiv-cs-ro
6 Aug 2026
Model Releases

Mind the Cap: Output-Budget Regimes Change the Measured Multilingual Reasoning Gap

DGX agent

arXiv:2608.04160v1 Announce Type: new Abstract: Multilingual evaluations report accuracy at a single output-token cap, but languages need different numbers of tokens to express the same content, so th

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models

DGX agent

arXiv:2608.04633v1 Announce Type: new Abstract: Recent Vision-Language-Action (VLA) methods improve generalization by aligning their representations with 3D scene geometry. However, these methods are

model-releasesarxiv-cs-ro
6 Aug 2026
Model Releases

Mirendil taps AI Hypercomputer TPUs and GPUs for pre- and post-training applications

DGX agent

Nearly every major AI lab uses Google Cloud infrastructure, including for training of models, inference for agents, and new frontier research. Google Cloud also continues to be the platform of choice

model-releasesgoogle-cloud-ai
6 Aug 2026
Model Releases

MobileWAM: Bridging World Action Models to Mobile Manipulation with Chain-of-Foresight

DGX agent

arXiv:2608.04657v1 Announce Type: new Abstract: World action models (WAMs) built on video generation backbones are a rising recipe for robot learning, yet remain confined to tabletop manipulation. Mob

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

MoCA: Multi-modal Cross-masked Autoencoder for Digital Health Measurements

DGX agent

arXiv:2506.02260v4 Announce Type: replace-cross Abstract: Wearable devices enable continuous multi-modal physiological and behavioral monitoring, yet analysis of these data streams faces fundamental c

model-releasesarxiv-cs-lg
6 Aug 2026
Model Releases

Modality Agreement- and Conflict-Aware Prototype Hypergraph Learning for Multimodal Intent Understanding

DGX agent

arXiv:2608.04054v1 Announce Type: cross Abstract: Multimodal intent recognition requires understanding not only what textual, acoustic, and visual signals share, but also how they disagree. Such disag

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Monte Carlo Tree Search for Table-to-Multimodal Report Generation

DGX agent

arXiv:2608.04071v1 Announce Type: new Abstract: Automatically generating professional multimodal reports comprising both textual analysis and visual charts from structured tabular data is a critical c

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

MOON3.0: Reasoning-aware Multimodal Representation Learning for E-commerce Product Understanding

DGX agent

arXiv:2604.00513v3 Announce Type: replace-cross Abstract: With the rapid growth of e-commerce, exploring general representations rather than task-specific ones has attracted increasing attention. Alth

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Neural Diversity Regularizes Hallucinations in Language Models

DGX agent

arXiv:2510.20690v3 Announce Type: replace-cross Abstract: Language models continue to hallucinate despite increases in parameters, compute, and data. We propose neural diversity -- decorrelated parall

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

NOLLI: A Difficulty-Calibrated Puzzle Benchmark for Diagnosing the English-Korean Performance Gap

DGX agent

arXiv:2608.04397v1 Announce Type: new Abstract: We introduce NOLLI, a procedurally generated English-Korean puzzle benchmark designed to diagnose where Korean performance gaps arise. It comprises 15 p

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

Non-asymptotic implicit bias of logistic regression at early-stage gradient descent dynamics

DGX agent

arXiv:2608.04382v1 Announce Type: new Abstract: Gradient descent has been of particular interest in modern machine learning beyond sole focus on optimization. Implicit bias emerging from optimization,

model-releasesarxiv-cs-lg
6 Aug 2026
Model Releases

Not Truly Multilingual: Script Consistency as a Missing Dimension in VLM Evaluation

DGX agent

arXiv:2606.17188v3 Announce Type: replace-cross Abstract: Current multilingual evaluations for Vision-Language Models (VLMs) assume a one-to-one mapping between language and orthography, overlooking b

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

NuclearDiffusion: Text-to-Image Foundation Models for Learning Nuclear Energy Concepts

DGX agent

arXiv:2608.04030v1 Announce Type: cross Abstract: Generative artificial intelligence (AI) has transformed text-to-image synthesis, yet its ability to represent specialized engineering domains remains

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

nvidia/NVIDIA-Nemotron-Parse-2.0 · Hugging Face

DGX agent

NVIDIA Nemotron Parse 2.0 transforms document images into structured, machine-readable representations with text, layout classes, bounding boxes, and reading-order information. Given a Red, Green, Blu

model-releasesr-localllama
6 Aug 2026
Model Releases

nvidias nemotron omni only loads its text half on a mac, so i wrote the vision and audio towers in mlx

DGX agent

nvidias nemotron omni is open weights and it sees, hears and reasons. theres already a 4bit mlx quant on hugging face but only the text backbone loads with standard mlx tooling. the model card says it

model-releasesr-localllama
6 Aug 2026
Model Releases

🟩 NVIDIA's whole speech stack just went local. ASR + TTS + codec, quantized to GGUF, running on-device via NeMo-Speech.cpp

DGX agent

🐦‍⬛ Magpie-TTS Multilingual 🦜 Nemotron Speech Streaming EN 0.6B 🦜 Nemotron-3.5 ASR Streaming 🦜 Parakeet CTC 1.1B 🦜 Parakeet TDT 0.6B v3 🥦 NanoCodec Merged PR https://huggingface.co/nvidia/magpie_tts_m

model-releasesr-localllama
6 Aug 2026
Model Releases

OmniEdit-Bench: A Comprehensive Benchmark for Instruction-based Video Editing

DGX agent

arXiv:2608.05049v1 Announce Type: new Abstract: Instruction-based video editing (IVE) is an emerging field with broad applications, yet evaluating editing models remains challenging. Existing benchmar

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

OmniRouting: A Semantic-Coupled Multimodal Benchmark for Constraint-Aware Spatial Reasoning in PCB Routing

DGX agent

arXiv:2608.04434v1 Announce Type: new Abstract: Recent large language models (LLMs) have demonstrated remarkable progress in constraint-aware navigation, maze reasoning, and graph reasoning. However,

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

OmniVR: Joint Video-Audio Conditional Generation for Restoring Degraded Historical Films

DGX agent

arXiv:2608.04224v1 Announce Type: new Abstract: Historical films suffer from co-occurring visual and audio degradations---blur, noise, flicker, hiss, clipping, and dropout---yet existing methods resto

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

On the Effectiveness of Adaptation Strategies for VLM-Based Federated Learning in Remote Sensing

DGX agent

arXiv:2608.04791v1 Announce Type: new Abstract: Federated learning (FL) enables collaborative training of deep learning models across decentralized image archives without requiring data centralization

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

One of the things the pelican benchmark is still useful for is visually representing (to a tiny extent) the improvements in a single model f…

DGX agent

One of the things the pelican benchmark is still useful for is visually representing (to a tiny extent) the improvements in a single model family Here's Meta AI's Spark (8th April), Spark 1.1 (9th Jul

model-releasessimon-willison--x
6 Aug 2026
Model Releases

One Surrogate to Fool Them All: Universal, Transferable, and Targeted Adversarial Attacks with CLIP

DGX agent

arXiv:2505.19840v3 Announce Type: replace-cross Abstract: Deep Neural Networks (DNNs) have achieved widespread success yet remain prone to adversarial attacks. Typically, such attacks either involve f

model-releasesarxiv-cs-lg
6 Aug 2026
Model Releases

Open models give teams more room to run agent loops, with API prices at a fraction of GPT-5.6 Sol and Claude Fable 5. That matters as planni…

DGX agent

Open models give teams more room to run agent loops, with API prices at a fraction of GPT-5.6 Sol and Claude Fable 5. That matters as planning, tool calls, retries, and long contexts compound token us

model-releasestogether-ai--x
6 Aug 2026
Model Releases

OpenAI debuts Agent Plugins, an open standard for bundling skills and MCP servers, and says Amazon, Cursor, Microsoft, and Vercel are on its steering committee (Zac Hall/9to5Mac)

DGX agent

Zac Hall / 9to5Mac: OpenAI debuts Agent Plugins, an open standard for bundling skills and MCP servers, and says Amazon, Cursor, Microsoft, and Vercel are on its steering committee — OpenAI's GPT-5 tur

model-releasestechmeme
6 Aug 2026
Model Releases

OpenAI updates the default model for free users to GPT-5.6 Luna, adds unlimited text chats for free users, rolls out an improved GPT-5.6 Sol version, and more (Herb Scribner/Axios)

DGX agent

Herb Scribner / Axios: OpenAI updates the default model for free users to GPT-5.6 Luna, adds unlimited text chats for free users, rolls out an improved GPT-5.6 Sol version, and more — OpenAI on Thursd

model-releasestechmeme
6 Aug 2026
Model Releases

PADFormer: Pose-agnostic Anomaly Detection from Sparse View Images

DGX agent

arXiv:2608.04210v1 Announce Type: new Abstract: Pose-agnostic Anomaly Detection (PAD) remains challenging as anomalies can appear under arbitrary viewpoints, requiring methods to handle significant po

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

Persistent Object Narratives for Token-Efficient Video Language Models

DGX agent

arXiv:2608.04866v1 Announce Type: new Abstract: Video large language models (Video-LLMs) have made strong progress in open-ended video understanding. However, their visual interfaces remain token-inte

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

Personalized Federated Sparse Adaptation of Time-Series Foundation Models

DGX agent

arXiv:2608.04695v1 Announce Type: cross Abstract: Federated adaptation of time-series foundation models (TSFMs) is attractive for building energy forecasting because meter data are private, distribute

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Physics-informed reduced-order modelling with equivariant spectral submanifolds

DGX agent

arXiv:2608.04239v1 Announce Type: new Abstract: Spectral submanifold (SSM) reduction has emerged as a mathematically principled route to reliable nonlinear reduced-order models, capturing dynamics bey

model-releasesarxiv-cs-lg
6 Aug 2026
Model Releases

PhysMind: From Video to Executable Worlds for Training-Free Physical Reasoning

DGX agent

arXiv:2608.04575v1 Announce Type: cross Abstract: Reliable physical reasoning from video requires understanding how objects move, interact, and respond to interventions. Existing vision-language model

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

PICopilot: An LLM-based Agentic Framework for Assisting Photonic Integrated Circuit Design via Script Generation

DGX agent

arXiv:2608.01791v2 Announce Type: replace-cross Abstract: The rapid development of photonic integrated circuits (PICs) is shifting the design flow from traditional graphical user interface (GUI)-based

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Plus and Pro users also now have a slider to choose how much reasoning effort ChatGPT puts into each response. We think it’s easier to use, …

DGX agent

OpenAI announced that Plus and Pro subscribers now have a slider to adjust the amount of reasoning effort ChatGPT applies to each response. The update employs GPT‑5.6 Sol for both Instant and deep rea

model-releasesopenai--x
6 Aug 2026
Model Releases

Plus and Pro users can access the updated version of GPT‑5.6 Sol in ChatGPT along with the new slider starting today. This version of GPT-5.…

DGX agent

Plus and Pro users can access the updated version of GPT‑5.6 Sol in ChatGPT along with the new slider starting today. This version of GPT-5.6 Sol is for everyday chats, so it will only be available in

model-releasesopenai--x
6 Aug 2026
Model Releases

Predict, Then Retrieve: Cross-Instance Future-State Retrieval from Video Prefixes

DGX agent

arXiv:2608.04426v1 Announce Type: cross Abstract: We introduce Predictive State Retrieval (PSR), a task in which a model observes a short video prefix and a temporal question about an object's future

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

Privacy-Preserving Action Recognition: Taxonomy, Methods, and Privacy-Utility Trade-offs

DGX agent

arXiv:2608.04501v1 Announce Type: new Abstract: Video surveillance in public safety, healthcare, and smart environments has made continuous human monitoring routine, raising real risks to personal ide

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

Protoreasoning in Tiny Transformers

DGX agent

arXiv:2608.04980v1 Announce Type: cross Abstract: We show that tiny transformers can profitably employ a simple form of Chain of Thought, which we call protoreasoning, allowing us to study step-by-ste

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Radar4D-VLM: Proposal-Grounded Temporal 4D Radar Reasoning Across Frozen Language Models

DGX agent

arXiv:2608.04130v1 Announce Type: new Abstract: Vision-language models for autonomous driving primarily rely on cameras and LiDAR, leaving 4D radar largely unexplored as a standalone perceptual modali

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

Reading Between the Frames: Interpreting Implicit and Non-literal Meaning in Social Media Videos

DGX agent

arXiv:2608.04939v1 Announce Type: new Abstract: Social media videos often communicate meanings that go beyond their visible actions, captions, or speech. A mundane clip may become humorous, ironic, or

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

ReGround: Restoring Visual Grounding in Multi-Step Reasoning through Self-Diagnosis and Visual Re-Examination

DGX agent

arXiv:2608.04385v1 Announce Type: new Abstract: Vision-Language Models (VLMs) often lose visual grounding during multi-step reasoning: as reasoning chains grow longer, later inference steps rely incre

model-releasesarxiv-cs-cv
6 Aug 2026
← Previous
1…3132333435…464
Next →