AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
Human
84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
84,605 results
31 Jul 2026

MIND: Multimodal Intent-Driven Network via Diffusion Transformers for Medical Image Fusion

SafetyDGX agent

arXiv:2607.28565v1 Announce Type: new Abstract: Medical image fusion aims to integrate complementary information from diverse imaging modalities to support clinical diagnosis. Existing methods typical

MiniMax H3 discussion

Model ReleasesDGX agent

https://x.com/MiniMax_AI/status/2083008095488516262 (says it's coming in a few days) it can do text-to-image, and image editing everyone seems to be mostly hyped about the video generation part (I am

MiniMax H3: Open-weight multimodel video model

Model ReleasesDGX agent

Just saw this posted by Fal.ai and then by Hailuo themselves, the next video model will be open weight released! Here's the blurb and link to to the feature post: Today, we're launching MiniMax H3, a

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Minimax-H3 video model released, open weights coming in the next few days

Model ReleasesDGX agent

https://x.com/MiniMax_AI/status/2083006198828417501?s=20 Quote from their article: Today, we're launching MiniMax H3, a general-purpose multimodal generation model. H3 understands unified context acro

Minimum VRAM GPU to run DeepSeek-V4-Flash-0731 Q4_K_XL at around 30 t/s ?

Model ReleasesDGX agent

Hello guys, I'm curious about running DeepSeek-V4-Flash-0731 locally. Since it’s a Mixture of Experts (MoE) model with only 13B active parameters, I was hoping the VRAM requirements might be manageabl

MixFrag: Fragility-Guided Mixed-Precision Post-Training Quantization for Vision Transformers

ResearchDGX agent

arXiv:2607.28589v1 Announce Type: new Abstract: Post-training quantization (PTQ) has emerged as an effective solution for deploying Vision Transformers (ViTs) on resource-constrained devices. However,

MMAC: A Massive Multi-dimensional Benchmark for Audio Captioning

Model ReleasesDGX agent

arXiv:2607.27109v2 Announce Type: cross Abstract: With the development of audio large language models (AudioLLMs), audio captioning needs to move from brief descriptions toward open-ended and fine-gra

MMHBench: A Multi-Perspective Benchmark for Mental Health Understanding in Long-Form Videos

Model ReleasesDGX agent

arXiv:2607.27895v1 Announce Type: cross Abstract: Mental health understanding in long-form videos requires nuanced reasoning over observable behavior, interpersonal context, and latent psychological s

MMOOC: A Comprehensive Benchmark for Out-of-Context Evaluation in Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2607.27637v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have achieved strong performance on a wide range of vision-language tasks, but often fail under imperfect or sh

mmRadarTwin: A Measurement-Calibrated Signal-Level Digital Twin Platform for Indoor mmWave Radar

ResearchDGX agent

arXiv:2607.28108v1 Announce Type: new Abstract: Indoor mmWave radar perception is difficult to reproduce because measured range-angle responses depend on scene geometry, material response, multipath,

Model-Driven Requirements Configuration with Three-Valued Uncertainty Scoring

AgentsDGX agent

arXiv:2607.26220v1 Announce Type: cross Abstract: Context: Large Language Models (LLMs) offer natural-language flexibility for automated requirements elicitation but frequently generate structurally i

Modeling Decisions in Blockchain Analytics: A Leakage-Aware Evaluation of Tree-Based vs. Sequential Models

ResearchDGX agent

arXiv:2607.27350v1 Announce Type: new Abstract: Sybil bots are Ethereum actors that imitate legitimate users to extract airdrop rewards or influence governance. Recent Sybil detection methods increasi

Models for minimalist RAG: B1ade 335M Embedding and 1B Parameter Small Language Models

Model ReleasesDGX agent

arXiv:2607.27506v1 Announce Type: new Abstract: Language and embedding models used in RAG systems are conventionally assumed to require large-scale pretraining and explicit grounding supervision. We p

MonoVoc: Decoupling Geometry and Semantics for Lightweight Monocular Open-Vocabulary 3D Gaussians

ResearchDGX agent

arXiv:2607.28300v1 Announce Type: new Abstract: Open vocabulary 3D scene understanding is essential for next-generation interactive systems, empowering users to intuitively query and navigate reconstr

MOON2.0: Dynamic Modality-balanced Multimodal Representation Learning for E-commerce Product Understanding

Model ReleasesDGX agent

arXiv:2511.12449v3 Announce Type: replace Abstract: Recent Multimodal Large Language Models (MLLMs) have significantly advanced e-commerce product understanding. However, they still face three challen

More Data, Worse Decisions? Preference Reversals in Neural Networks under Gram Incompatibility

ResearchDGX agent

arXiv:2607.27255v1 Announce Type: cross Abstract: Neural networks increasingly combine data across populations, time periods, and operating conditions to improve generalization. This raises a reliabil

MORFES: A Benchmark for Productive Inflectional Competence in Modern Greek

Model ReleasesDGX agent

arXiv:2607.28274v1 Announce Type: new Abstract: Modern Greek is a richly inflected language, yet the language models built for it are evaluated mainly on factual knowledge, and no benchmark is dedicat

Morphological Detection and Classification of Microplastics and Nanoplastics Emerged from Consumer Products by Deep Learning

ResearchDGX agent

arXiv:2409.13688v2 Announce Type: replace Abstract: Plastic pollution presents an escalating global issue, impacting health and environmental systems, with micro- and nanoplastics found across mediums

MPEcho: A Melody and Phoneme-Aware Generative Framework for Controllable Cover Song Generation

SafetyDGX agent

arXiv:2607.26698v1 Announce Type: cross Abstract: Cover song generation (CSG) should preserve the melodic and linguistic content of a reference song while recreating the remaining musical components.

MPIE-Bench: Benchmarking Anatomically Plausible Multi-Person Interaction Editing

Model ReleasesDGX agent

arXiv:2607.27616v1 Announce Type: new Abstract: Text-to-image and personalized editing models now synthesize high-fidelity single-subject images with ease. Yet placing multiple named people into share

MRD: Using Physically Based Differentiable Rendering to Probe Vision Models for 3D Scene Understanding

ResearchDGX agent

arXiv:2512.12307v5 Announce Type: replace Abstract: While deep learning methods have achieved impressive success in many vision benchmarks, it remains difficult to understand and explain the represent

MSCM-net: A hyperspectral image classiffcation method based on multi-scale convolution and Mamba

Model ReleasesDGX agent

arXiv:2607.28277v1 Announce Type: new Abstract: Hyperspectral imaging is widely used in remote sensing and engineering. Therefore, research on its classification methods is crucial. While CNN and Tran

MSGNN: A Spectral Graph Neural Network Based on a Novel Magnetic Signed Laplacian

ApplicationsDGX agent

arXiv:2209.00546v5 Announce Type: replace-cross Abstract: Signed and directed networks are ubiquitous in real-world applications. However, there has been relatively little work proposing spectral grap

MUGEN: A Unified Framework for Efficient Motion Understanding and Generation

SafetyDGX agent

arXiv:2607.27581v1 Announce Type: new Abstract: Grounding human motion in language, and language in motion, is a central step toward physical AI systems that can understand, generate, and communicate

MUL-T: Decoding Spatial Cellular Architecture in Multiplexed Tissue Images

ResearchDGX agent

arXiv:2607.28030v1 Announce Type: cross Abstract: Understanding tissue organisation in multiplexed imaging requires modelling both cellular phenotypes and their spatial context. Existing approaches ty

Multi-Agent Debate Strategies: Survey, Taxonomy, and Challenges

AgentsDGX agent

arXiv:2607.26212v1 Announce Type: cross Abstract: Multi-Agent Debate (MAD) is a promising paradigm for improving the accuracy and robustness of Large Language Model (LLM)-based agentic systems. It ena

Multi-channel Uplift Policy Learning

SafetyDGX agent

arXiv:2607.28182v1 Announce Type: new Abstract: E-commerce platforms must allocate fixed marketing budgets across multiple channels to maximize business utility. However, standard predict-then-optimiz

MultivationBench: A Benchmark for Multimodal Sequential Motivation Reasoning

Model ReleasesDGX agent

arXiv:2607.26465v1 Announce Type: new Abstract: Multimodal Large Language Models have sparked significant interest due to their potential for social intelligence; however, their ability to perform seq

Nanoparticle Networks for Neuromorphic Computing

ResearchDGX agent

arXiv:2607.27844v1 Announce Type: cross Abstract: Physical computing leverages complex dynamical systems for energy-efficient data processing. In this work, we present a neuromorphic architecture base

Neat work on long-horizon agents. Splitting a hard task across agents is typically how standard multi-agent work. The usual design lets them…

Model ReleasesDGX agent

Neat work on long-horizon agents. Splitting a hard task across agents is typically how standard multi-agent work. The usual design lets them exchange findings only at phase boundaries, through staged

Negative controls reveal volume-driven confounding in radiomics and imaging foundation model features

ResearchDGX agent

arXiv:2607.28423v1 Announce Type: new Abstract: Radiomics and imaging foundation models promise non-invasive biomarkers of tumour biology, yet predictive signatures may reflect tumour volume or acquis

Neural Network Approximation of Solutions to Fractional Parabolic Partial Differential Equations

ResearchDGX agent

arXiv:2607.27781v1 Announce Type: cross Abstract: We establish a dimension-efficient neural network approximation theory for solutions to fractional parabolic equations with lower-order drift and pote

Neural Network-Assisted CLEAN for Channel Modeling in Low-SNR Regimes

Model ReleasesDGX agent

arXiv:2607.27450v1 Announce Type: new Abstract: Accurate multipath parameter estimation is critical for modern wireless communication systems, particularly in challenging low-SNR environments. Traditi

New research from Microsoft. This one is on training computer-use agents at scale. Recent pipelines generate synthetic environments in bulk,…

Model ReleasesDGX agent

New research from Microsoft. This one is on training computer-use agents at scale. Recent pipelines generate synthetic environments in bulk, which moved the bottleneck from how many exist to what is i

NMINE: Normalized Mutual Information Neural Estimation

ResearchDGX agent

arXiv:2607.27710v1 Announce Type: new Abstract: Mutual information is a general measure of statistical dependence that captures both linear and nonlinear relationships between random variables. For co

Noisy Data is Destructive to Reinforcement Learning with Verifiable Rewards

TutorialsDGX agent

arXiv:2603.16140v2 Announce Type: replace Abstract: Reinforcement learning with verifiable rewards (RLVR) has driven recent capability advances of large language models across various domains. Recent

Now Suddenly too many choices for DGX Spark with Qwen 3.5 122B . What would be the next upgrade?

Model ReleasesDGX agent

Laguna 2.1 at NVFP4 Deepseek v4 at Q2 Inkling-Small at IQ3 Which models you guys running now ? How it compares to 122b? Upcoming in few days : Ling 3.0 124B (Could be new king) LongCat 69B A3B ( very

Now You Have My Healthy Attention: A U-DiT for Brain-MRI Inpainting

Local AiDGX agent

arXiv:2607.27974v1 Announce Type: new Abstract: The ASNR-MICCAI BraTS Local Synthesis (Inpainting) task asks for the anatomically plausible completion of healthy brain tissue within a masked region of

Nscale buys AI infrastructure optimization startup Anyscale for reported $1.65B

HardwareDGX agent

Data center builder Nscale Global Holdings Ltd. today announced plans to acquire Anyscale Inc., a venture-backed provider of artificial intelligence software. The terms of the deal were not disclosed.

NVIDIA Video Codec SDK 13.1: Zero-Copy Transcode, AV1 B-Frames, and Frame-Accurate Seek

HardwareDGX agent

NVIDIA Video Codec SDK 13.1 adds AV1 hierarchical reference mode supporting up to 31 B‑frames and efficient iterative tuning that delivers significant bitrate savings in CQ and VBR modes. It enhances

Objective-Aligned Direct Answer SFT for Robust Multi-Frame Medical VQA

Model ReleasesDGX agent

arXiv:2607.27566v1 Announce Type: new Abstract: Multi-frame medical VQA appears to reward increasingly complex adaptation: controller-style inference, localization-aware reranking, static hard-negativ

ObjectStream: Latent Objects as Memory Anchors for Streaming Video Understanding

HardwareDGX agent

arXiv:2607.28312v1 Announce Type: new Abstract: Streaming video understanding requires models to continuously retain useful visual evidence before future questions are known. Existing approaches prima

ODEWorld: A Continuous Predictive Architecture via Physical-Time Flow

SafetyDGX agent

arXiv:2607.27924v1 Announce Type: cross Abstract: In the physical world we inhabit, space and time are fundamentally continuous. However, existing machine learning paradigms for world modeling are lar

OK, GPT-5.6 Luna is a bit of a beast. Given the 80% price drop today I decided to try it in Datasette Agent, and it's furiously quick and ge…

Model ReleasesDGX agent

OK, GPT-5.6 Luna is a bit of a beast. Given the 80% price drop today I decided to try it in Datasette Agent, and it's furiously quick and generates all the SQL, HTML and JavaScript (for Datasette Apps

On a joint simultaneous learning of relevant feature subsets and subspaces in regression-like problems

Model ReleasesDGX agent

arXiv:2607.28080v1 Announce Type: cross Abstract: We extend a recently introduced Entropy-Optimal Manifold Clustering (EOMC) to allow for a joint simultaneous identification of subsets and subspaces o

On-Policy and Off-Policy Learning for Large Action Spaces

SafetyDGX agent

arXiv:2607.28408v1 Announce Type: new Abstract: This thesis studies policy learning in interactive systems where an agent observes a context, selects an action from a very large set, and receives part

On the Rate of Convergence of Kolmogorov-Arnold Network Regression Estimators

ResearchDGX agent

arXiv:2509.19830v3 Announce Type: replace Abstract: Kolmogorov-Arnold Networks (KANs) approximate multivariate functions by composing univariate transformations through additive or multiplicative aggr

One Future, Every Robot: Label-Efficient Collective-State Prediction with Decentralized JEPA

AgentsDGX agent

arXiv:2607.28443v1 Announce Type: new Abstract: Can every robot in a swarm predict the same future collective state from only local observations and bandwidth-limited messages? We formulate this as de

One of our big findings in our study at Procter and Gamble was that AI blurred the lines between jobs. Now OpenAI has a similar finding. Org…

ApplicationsDGX agent

One of our big findings in our study at Procter and Gamble was that AI blurred the lines between jobs. Now OpenAI has a similar finding. Organizational boundaries are becoming porous, the walls thinni

One Patch Is Enough: Reinforcement-Optimized Visual Token Grounding for MLLM-Based Scene Text Spotting

Model ReleasesDGX agent

arXiv:2607.27902v1 Announce Type: new Abstract: Scene text spotting requires high-precision alignment between textual recognition and spatial localization. While visual-token grounding has emerged as

One Run Is Not an Idea: The Implementation Lottery in Automated Research

AgentsDGX agent

arXiv:2607.26587v1 Announce Type: cross Abstract: Automated research systems use experimental scores both to deliver artifacts and to decide which ideas to retain, transfer, and pursue. Yet one run sc

OneShot: Index-in-Ranking with Neural Scoring for Large-Scale Retrieval

SafetyDGX agent

arXiv:2607.27475v1 Announce Type: cross Abstract: In modern recommendation systems, retrieval serves as a primary stage responsible for filtering billions of candidate items down to thousands prior to

Open Source Ternary LLM Engine in Rust/CUDA for Quantization, Serving, and Training of models on consumer GPUs, called Tritium (Apache 2.0)

Model ReleasesDGX agent

This post was not written by a clanker. Hey guys, I'm a comp sci major who wanted to introduce a cool project I built for quantizing models to ternary (1.58 bit) with as minimal of loss as possible, a

OPENAI IS GOING TO TAKE THIS ENTIRE MARKET DOWN WITH IT And you don't have to own a single share to get hurt. What I'm about to explain shou…

HardwareDGX agent

OPENAI IS GOING TO TAKE THIS ENTIRE MARKET DOWN WITH IT And you don't have to own a single share to get hurt. What I'm about to explain should worry anybody who thinks they're diversified: OpenAI is a

OpenAI says its models now have more than 1B active users and are used by more than 2M businesses (Katherine Hamilton/Wall Street Journal)

IndustryDGX agent

Katherine Hamilton / Wall Street Journal: OpenAI says its models now have more than 1B active users and are used by more than 2M businesses — The announcement comes after OpenAI said earlier this week

OpenAI’s entire growth loop: new models, price cuts, and Tibo’s token resets

Model ReleasesDGX agent

OpenAI’s entire growth loop: new models, price cuts, and Tibo’s token resets major price cuts today: *80% drop for GPT-5.6 Luna, now 0.20 per million input tokens and 1.20 per million output *20% drop

OPLD: On-Policy Latent Distillation for Multimodal Reasoning

SafetyDGX agent

arXiv:2607.28154v1 Announce Type: new Abstract: Interleaved multimodal Chain-of-Thought (CoT) improves visual reasoning by incorporating auxiliary visual evidence into intermediate reasoning. However,

Optimal Realistic Local AI for Most

Model ReleasesDGX agent

So you’ve got a 3090 or maybe even a 5090? Or more likely a 4060 8GB Ti. You wanna try local AI, you don’t know what it can/can’t do. 1) Install the best model you can. If you have a 3090 or a 5090, t

Optimizing production agents with Amazon Bedrock AgentCore Observability

AgentsDGX agent

As your AI agents move from prototype to production, the challenge shifts from getting them to work to keeping them fast and efficient. Learn how to use Amazon Bedrock AgentCore Observability and Amaz

Optimizing Regret

SafetyDGX agent

arXiv:2607.18866v2 Announce Type: replace-cross Abstract: Building on the identity that expected regret equals the covariance between costs and decisions, this paper develops a derivative theory of th

← Previous
1…147148149150151…1411
Next →