AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
84,460 results
6 Aug 2026

Multimodal Alignment Through Joint Kernel Entropic Gromov--Wasserstein Optimal Transport

SafetyDGX agent

arXiv:2608.04234v1 Announce Type: cross Abstract: We study the problem of aligning data from multiple modalities into a shared representation space, focusing on settings where strong pretrained unimod

Multimodal Spatiotemporal Atmospheric Data Assimilation with Latent Flow-matching

ResearchDGX agent

arXiv:2608.05103v1 Announce Type: new Abstract: Data assimilation (DA) uses Bayesian inference to update the state of a numerical forecast model with observed data. In this study, we propose a fundame

MultiPathFormer: Towards a Foundation Model for Multipath Wireless Propagation

Local AiDGX agent

arXiv:2608.05076v1 Announce Type: cross Abstract: Recent advances in machine learning have enabled training of wireless foundation models, which aim to support tasks such as channel estimation, beam p

Content type
AllBlogX PostPaperYouTubeRedditGitHub

muSync-GS: Physics-Synchronized Driving Video Synthesis for Weather and Geometric Road Hazards

SafetyDGX agent

arXiv:2608.04412v1 Announce Type: new Abstract: High-quality driving data are essential for autonomous-driving systems and generative world models. However, rare and safety-critical scenarios involvin

MVTOP: Multi-View Transformer-based Object Pose-Estimation

ResearchDGX agent

arXiv:2508.03243v2 Announce Type: replace Abstract: We present MVTOP, a novel transformer-based method for multi-view rigid object pose estimation. Through an early fusion of the view-specific feature

Neighborhood-Aware Dual Biomedical Entity Linking

ResearchDGX agent

arXiv:2608.04144v1 Announce Type: cross Abstract: Biomedical entity linking grounds mentions in clinical and scientific text to entities in a curated knowledge base (KB) with ontological structure, wh

NeuMoSync: End-to-End Neuromodulatory Control for Plasticity and Adaptability in Continual Learning

TutorialsDGX agent

arXiv:2608.04358v1 Announce Type: new Abstract: Continual learning (CL) requires models to learn tasks sequentially, yet deep neural networks often suffer from plasticity loss and poor knowledge trans

Neural Diversity Regularizes Hallucinations in Language Models

Model ReleasesDGX agent

arXiv:2510.20690v3 Announce Type: replace-cross Abstract: Language models continue to hallucinate despite increases in parameters, compute, and data. We propose neural diversity -- decorrelated parall

Neurocomputational Mechanisms of Syntactic Transfer in Bilingual Sentence Production

ApplicationsDGX agent

arXiv:2601.18056v2 Announce Type: replace Abstract: We discuss the benefits of incorporating oscillatory neural mechanisms into the study of bilingual production errors and their traditionally documen

NeuroPB: Scaling Neural Decoding with Pretrained Behavioral Representations

ResearchDGX agent

arXiv:2608.04389v1 Announce Type: new Abstract: Decoding continuous motor trajectories from neural activity is essential for developing practical brain-computer interfaces (BCIs). However, current neu

New details on OpenAI/Hugging Face attack emerge as security industry debates AI agent controls

AgentsDGX agent

How fast is artificial intelligence advancing? Behind the scenes at OpenAI Group PBC, AI agents are fluent, technically precise and occasionally profane in their extensive conversations…with each othe

NodeJEPA: Structure-Conditioned Latent Prediction for Node-Level Graph Self-Supervised Learning

TutorialsDGX agent

arXiv:2608.04381v1 Announce Type: cross Abstract: Self-supervised learning on graphs is largely shaped by contrastive methods that depend on carefully designed augmentations, and by generative methods

NOLLI: A Difficulty-Calibrated Puzzle Benchmark for Diagnosing the English-Korean Performance Gap

Model ReleasesDGX agent

arXiv:2608.04397v1 Announce Type: new Abstract: We introduce NOLLI, a procedurally generated English-Korean puzzle benchmark designed to diagnose where Korean performance gaps arise. It comprises 15 p

Non-asymptotic implicit bias of logistic regression at early-stage gradient descent dynamics

Model ReleasesDGX agent

arXiv:2608.04382v1 Announce Type: new Abstract: Gradient descent has been of particular interest in modern machine learning beyond sole focus on optimization. Implicit bias emerging from optimization,

Non-Stationary Inventory Control with Lead Times

SafetyDGX agent

arXiv:2602.05799v2 Announce Type: replace-cross Abstract: We study non-stationary single-item, periodic-review inventory control problems in which the demand distribution is unknown and may change ove

Nonparametric Goodness-of-fit Testing under Covariate Shift

ResearchDGX agent

arXiv:2608.04860v1 Announce Type: cross Abstract: This paper develops procedures for nonparametric goodness-of-fit testing under covariate shift, where labelled data are drawn from a source population

Not All Redundant Tokens Are Alike: Analyzing Visual Token Pruning through Token Roles

SafetyDGX agent

arXiv:2608.04483v1 Announce Type: new Abstract: Vision-language models (VLMs) process an image as a sequence of visual tokens, which creates a substantial computational bottleneck during inference. Re

Not Every Divergence Should Be Suppressed: Counterfactual Recoverability in On-Policy Distillation

SafetyDGX agent

arXiv:2608.04408v1 Announce Type: cross Abstract: On-policy distillation (OPD) supervises student-visited trajectories, yet divergence-based rules cannot determine whether an erroneous prefix remains

Not Truly Multilingual: Script Consistency as a Missing Dimension in VLM Evaluation

Model ReleasesDGX agent

arXiv:2606.17188v3 Announce Type: replace-cross Abstract: Current multilingual evaluations for Vision-Language Models (VLMs) assume a one-to-one mapping between language and orthography, overlooking b

NSF-HRPT: Neural Semantic Field meets Hierarchical Risk Perception Tree for Safety-Critical Scenario Assessment

SafetyDGX agent

arXiv:2608.04776v1 Announce Type: new Abstract: The ability to accurately assess and anticipate risks in safety-critical scenarios is crucial for autonomous driving systems. While existing research ha

NuclearDiffusion: Text-to-Image Foundation Models for Learning Nuclear Energy Concepts

Model ReleasesDGX agent

arXiv:2608.04030v1 Announce Type: cross Abstract: Generative artificial intelligence (AI) has transformed text-to-image synthesis, yet its ability to represent specialized engineering domains remains

nvidia/NVIDIA-Nemotron-Parse-2.0 · Hugging Face

Model ReleasesDGX agent

NVIDIA Nemotron Parse 2.0 transforms document images into structured, machine-readable representations with text, layout classes, bounding boxes, and reading-order information. Given a Red, Green, Blu

nvidias nemotron omni only loads its text half on a mac, so i wrote the vision and audio towers in mlx

Model ReleasesDGX agent

nvidias nemotron omni is open weights and it sees, hears and reasons. theres already a 4bit mlx quant on hugging face but only the text backbone loads with standard mlx tooling. the model card says it

🟩 NVIDIA's whole speech stack just went local. ASR + TTS + codec, quantized to GGUF, running on-device via NeMo-Speech.cpp

Model ReleasesDGX agent

🐦‍⬛ Magpie-TTS Multilingual 🦜 Nemotron Speech Streaming EN 0.6B 🦜 Nemotron-3.5 ASR Streaming 🦜 Parakeet CTC 1.1B 🦜 Parakeet TDT 0.6B v3 🥦 NanoCodec Merged PR https://huggingface.co/nvidia/magpie_tts_m

Objects as Audio-Visual Modal Sound Fields

ApplicationsDGX agent

arXiv:2608.05145v1 Announce Type: new Abstract: While modern 3D reconstruction excels at modeling object geometry and appearance, it largely ignores the rich acoustic cues revealed through physical in

OctoLong: Mid-Training On Cross-Repository Code Contexts Enhances Long-Context Modeling

AgentsDGX agent

arXiv:2608.05141v1 Announce Type: new Abstract: Context lengths of language models (LMs) have dramatically increased, driven by the demands for in-context learning, self-improvement, and long-horizon

ODRA: Synthesizing Cognitive Behavioral Therapy Sessions with Structured Chain-Of-Thought and Dynamic Patient Resistance

SafetyDGX agent

arXiv:2608.04524v1 Announce Type: new Abstract: Synthetic generation of Cognitive Behavioral Therapy (CBT) sessions is challenged by two competing demands: adhering to strict therapeutic structure whi

OMG i wrote some of the original work on what is to be neurosymbolic in 2001 and dude who probably hasn’t read that work is trying to school…

SafetyDGX agent

OMG i wrote some of the original work on what is to be neurosymbolic in 2001 and dude who probably hasn’t read that work is trying to school me on the definition 🤦‍♂️ coding harness and tools calls ar

OmniEdit-Bench: A Comprehensive Benchmark for Instruction-based Video Editing

Model ReleasesDGX agent

arXiv:2608.05049v1 Announce Type: new Abstract: Instruction-based video editing (IVE) is an emerging field with broad applications, yet evaluating editing models remains challenging. Existing benchmar

OmniRouting: A Semantic-Coupled Multimodal Benchmark for Constraint-Aware Spatial Reasoning in PCB Routing

Model ReleasesDGX agent

arXiv:2608.04434v1 Announce Type: new Abstract: Recent large language models (LLMs) have demonstrated remarkable progress in constraint-aware navigation, maze reasoning, and graph reasoning. However,

OmniVR: Joint Video-Audio Conditional Generation for Restoring Degraded Historical Films

Model ReleasesDGX agent

arXiv:2608.04224v1 Announce Type: new Abstract: Historical films suffer from co-occurring visual and audio degradations---blur, noise, flicker, hiss, clipping, and dropout---yet existing methods resto

On Hamming-Lipschitz Type Stability of the Subdominant (Minmax) Ultrametric: Theory and Simple Proofs

ResearchDGX agent

arXiv:2608.04014v1 Announce Type: cross Abstract: The subdominant (minmax) ultrametric is a canonical tree-structured summary of a dissimilarity matrix, arising equivalently as the ultrametric induced

On July 24, the Black at Mila community hosted Prof. Vukosi Marivate @vukosi (University of Pretoria, co-founder of Masakhane, Lelapa AI, an…

SafetyDGX agent

On July 24, the Black at Mila community hosted Prof. Vukosi Marivate @vukosi (University of Pretoria, co-founder of Masakhane, Lelapa AI, and Deep Learning Indaba) for a powerful session on the future

On MUON optimization: From non-convergence to an error analysis with Polar Express and the Newton-Schulz polynomial from implementations

ResearchDGX agent

arXiv:2608.04607v1 Announce Type: cross Abstract: Stochastic gradient descent (SGD) optimization methods are the standard instruments for the training of deep neural networks (DNNs). In many relevant

On the Effectiveness of Adaptation Strategies for VLM-Based Federated Learning in Remote Sensing

Model ReleasesDGX agent

arXiv:2608.04791v1 Announce Type: new Abstract: Federated learning (FL) enables collaborative training of deep learning models across decentralized image archives without requiring data centralization

On The Suitability of Differential Dataflow For Datalog Interpretation In Highly Dynamic Settings

ResearchDGX agent

arXiv:2308.04214v2 Announce Type: replace-cross Abstract: In the domain of knowledge representation and reasoning within AI, datalog engines play an ever-increasingly crucial role. The crux of their o

One of the things the pelican benchmark is still useful for is visually representing (to a tiny extent) the improvements in a single model f…

Model ReleasesDGX agent

One of the things the pelican benchmark is still useful for is visually representing (to a tiny extent) the improvements in a single model family Here's Meta AI's Spark (8th April), Spark 1.1 (9th Jul

One Surrogate to Fool Them All: Universal, Transferable, and Targeted Adversarial Attacks with CLIP

Model ReleasesDGX agent

arXiv:2505.19840v3 Announce Type: replace-cross Abstract: Deep Neural Networks (DNNs) have achieved widespread success yet remain prone to adversarial attacks. Typically, such attacks either involve f

OneDayAgent: Towards a Long-Horizon Harness for Autonomous Agents

AgentsDGX agent

arXiv:2608.05013v1 Announce Type: cross Abstract: LLM agents are increasingly applied to open-ended everyday requests that span work, study, and life. These tasks are long-horizon, cross-environment,

Ooredoo, Nvidia, Nokia, and Indosat launch Zankore, Indonesia's first dedicated AI compute and neocloud platform; Ooredoo, which owns a 49% stake, commits $800M (Melissa Hancock/Fortune)

HardwareDGX agent

Melissa Hancock / Fortune: Ooredoo, Nvidia, Nokia, and Indosat launch Zankore, Indonesia's first dedicated AI compute and neocloud platform; Ooredoo, which owns a 49% stake, commits $800M — Ooredoo Gr

OPD-V: Visual On-Policy Self-Distillation with Modality Balance

SafetyDGX agent

arXiv:2608.05131v1 Announce Type: cross Abstract: On-Policy Self-Distillation (OPSD) has become a standard post-training approach for improving visual reasoning in multimodal large language models (ML

Open models give teams more room to run agent loops, with API prices at a fraction of GPT-5.6 Sol and Claude Fable 5. That matters as planni…

Model ReleasesDGX agent

Open models give teams more room to run agent loops, with API prices at a fraction of GPT-5.6 Sol and Claude Fable 5. That matters as planning, tool calls, retries, and long contexts compound token us

OpenAI debuts Agent Plugins, an open standard for bundling skills and MCP servers, and says Amazon, Cursor, Microsoft, and Vercel are on its steering committee (Zac Hall/9to5Mac)

Model ReleasesDGX agent

Zac Hall / 9to5Mac: OpenAI debuts Agent Plugins, an open standard for bundling skills and MCP servers, and says Amazon, Cursor, Microsoft, and Vercel are on its steering committee — OpenAI's GPT-5 tur

OpenAI files a motion to dismiss Apple's lawsuit accusing the AI company of stealing trade secrets, saying that the iPhone maker's allegations are meritless (Bloomberg)

IndustryDGX agent

Bloomberg: OpenAI files a motion to dismiss Apple's lawsuit accusing the AI company of stealing trade secrets, saying that the iPhone maker's allegations are meritless — OpenAI is asking a federal jud

OpenAI is giving ChatGPT free users unlimited text chats

IndustryDGX agent

OpenAI is making a big change for ChatGPT users on its free and Go tiers: Starting next week, users on those tiers will be able to have unlimited text chats with the chatbot, according to OpenAI. Righ

OpenAI says Apple’s trade secrets lawsuit is ‘rotten to its core’

IndustryDGX agent

OpenAI has asked a federal judge to toss out Apple's landmark lawsuit accusing the ChatGPT maker of stealing trade secrets, describing the allegations as 'meritless.' In a motion filed yesterday to di

OpenAI updates the default model for free users to GPT-5.6 Luna, adds unlimited text chats for free users, rolls out an improved GPT-5.6 Sol version, and more (Herb Scribner/Axios)

Model ReleasesDGX agent

Herb Scribner / Axios: OpenAI updates the default model for free users to GPT-5.6 Luna, adds unlimited text chats for free users, rolls out an improved GPT-5.6 Sol version, and more — OpenAI on Thursd

Optimal Constrained sc-LTL Planning in MDPs via Switching Policies

SafetyDGX agent

arXiv:2608.05021v1 Announce Type: new Abstract: We study the synthesis of optimal policies for planning problems on Markov decision processes with both objectives and safety constraints specified in c

Optimal Training-Time Scaling in Gradual Adaptation

ResearchDGX agent

arXiv:2608.04927v1 Announce Type: new Abstract: In gradual adaptation, how should the training time on each task change as the number of intermediate tasks increases? We study this question for overpa

Optimizing the Preconditioner: A Black-box Online-to-Nonconvex Conversion with Static Regret Minimization Oracles

ResearchDGX agent

arXiv:2607.17607v2 Announce Type: replace Abstract: We study whether stochastic nonconvex optimization can be reduced to ordinary static regret minimization in online convex optimization in a black-bo

Optimizing What Policies Learn From: Recoverability-aware Rollout Intervention Learning

SafetyDGX agent

arXiv:2608.05080v1 Announce Type: cross Abstract: Critic-free group-based reinforcement learning has become a scalable approach for post-training large language models. However, most existing methods

ORACLE: A Multi-Objective Reinforcement Learning-Based Analog Circuit Design Optimizer with Large Language Models-Guided Exploration

ResearchDGX agent

arXiv:2608.04999v1 Announce Type: cross Abstract: Analog circuit design automation using reinforcement learning (RL) has emerged as a promising approach for reducing manual effort. However, many exist

Our Black Hat talk on the OpenAI-Hugging Face incident is now live on youtube. This is a watershed moment for the industry. I encourage all …

ToolsDGX agent

Our Black Hat talk on the OpenAI-Hugging Face incident is now live on youtube. This is a watershed moment for the industry. I encourage all defenders to watch, consider how attack dynamics will immine

Out-Of-The-Loop Multi-Fidelity Bayesian Optimization

ApplicationsDGX agent

arXiv:2608.04113v1 Announce Type: cross Abstract: Black-box optimization is a ubiquitous problem in science and engineering, often dealing with expensive objective functions with cheaper lower-fidelit

OutLangSplat: 3D Language Gaussian Splatting for UAV Outdoor Scenes

SafetyDGX agent

arXiv:2608.04560v1 Announce Type: new Abstract: 3D Language Gaussian Splatting embeds open-vocabulary language features into 3D Gaussian Splatting, providing an efficient explicit representation for t

Outlook where frontier AI is headed next 18 months: The AI reasoning training + harness loop works if you can produce enough data and reason…

AgentsDGX agent

Outlook where frontier AI is headed next 18 months: The AI reasoning training + harness loop works if you can produce enough data and reasoning traces (via verifiers). Proven with code and math result

Overcoming Statistical Bias in Action-Controllable World Models

SafetyDGX agent

arXiv:2608.04653v1 Announce Type: new Abstract: Action-conditioned world models aim to predict how visual environments evolve under an agent's actions. Yet future frames are often highly predictable f

PADFormer: Pose-agnostic Anomaly Detection from Sparse View Images

Model ReleasesDGX agent

arXiv:2608.04210v1 Announce Type: new Abstract: Pose-agnostic Anomaly Detection (PAD) remains challenging as anomalies can appear under arbitrary viewpoints, requiring methods to handle significant po

Patients-like-me: A Variational LM--GNN Framework for Explainable Clinical Prediction

Local AiDGX agent

arXiv:2608.04193v1 Announce Type: cross Abstract: Language models (LMs) offer strong textual representations for electronic health records (EHRs), but they encode patient sequences in isolation and pr

Perception Before Reasoning: Dynamic Latent Reasoning for Video Understanding and Question Answering

Local AiDGX agent

arXiv:2608.04124v1 Announce Type: cross Abstract: Video question answering requires models to ground language queries in visual evidence and, when necessary, reason over that evidence across time. Exi

← Previous
1…8687888990…1408
Next →