AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent
84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,612 results
4 Jun 2026

Flow Matching Calibration for Simulation-Based Inference under Model Misspecification

Model ReleasesDGX agent

arXiv:2509.23385v5 Announce Type: replace-cross Abstract: Simulation-based inference (SBI) is transforming experimental sciences by enabling parameter estimation in complex non-linear models from simu

Food-R1: A Unified Multi-Task Food Vision-Language Model with Reinforcement Learning

Model ReleasesDGX agent

arXiv:2606.04986v1 Announce Type: new Abstract: Recent studies have explored Vision-Language Models (VLMs) for food analysis. However, most existing methods rely primarily on supervised fine-tuning (S

foom!

Model ReleasesDGX agent

foom! Our internal data shows Claude is accelerating AI development—a possible path to recursive self-improvement, or AI autonomously building a more capable successor. It’s happening faster than we t


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Founders Fund launches a TV-style game show featuring A-list founders and investors, including Sam Altman and Palmer Luckey, playing a game of Mafia (Tom Dotan/Newcomer)

Model ReleasesDGX agent

Tom Dotan / Newcomer: Founders Fund launches a TV-style game show featuring A-list founders and investors, including Sam Altman and Palmer Luckey, playing a game of Mafia — Do people want to watch the

From Symbolic to Geometric: Enabling Spatial Reasoning in Large Language Models

Model ReleasesDGX agent

arXiv:2606.04381v1 Announce Type: cross Abstract: Recent large language models (LLMs) often appear to exhibit spatial reasoning ability; however, this capability is largely symbolic, arising from patt

From Untrusted Input to Trusted Memory: A Systematic Study of Memory Poisoning Attacks in LLM Agents

Model ReleasesDGX agent

arXiv:2606.04329v1 Announce Type: cross Abstract: Memory is a core component of AI agents, enabling them to accumulate knowledge across interactions and improve performance. However, persistent memory

GeM-NR: Geometry-Aware Multi-View Editing for Nonrigid Scene Changes

Model ReleasesDGX agent

arXiv:2606.05142v1 Announce Type: cross Abstract: Recent developments in multi-view image editing with generative models have brought us a step closer toward general 3D content generation and customiz

GENEB: Why Genomic Models Are Hard to Compare

Model ReleasesDGX agent

arXiv:2606.04525v1 Announce Type: new Abstract: Progress in genomic foundation models is difficult to assess due to fragmented benchmarks, incompatible evaluation protocols, and task-specific reportin

Geometry-Aware Hallucination Detection in Large Language Models

Model ReleasesDGX agent

arXiv:2601.06196v3 Announce Type: replace-cross Abstract: Large language models (LLMs) frequently generate factually incorrect or unsupported content, commonly referred to as hallucinations. Prior wor

Geometry Gaussians: Decoupling Appearance and Geometry in Gaussian Splatting

Model ReleasesDGX agent

arXiv:2606.05124v1 Announce Type: cross Abstract: After the success of 3D Gaussian Splatting (3DGS) for novel view synthesis, many works have explored how to also use it for geometric surface represen

Geometry-Preserving Unsupervised Alignment for Heterogeneous Foundation Models

Model ReleasesDGX agent

arXiv:2606.04385v1 Announce Type: new Abstract: Foundation models have driven rapid progress in computer vision, yet the two dominant paradigms, vision-language foundation models (VLMs) and vision-onl

GPT-5.5 dominates $1,500 LLM hacking test while Gemini refuses to even try

Model ReleasesDGX agent

A security researcher spent 1,500 running 13+ AI models against a deliberately vulnerable app, with GPT-5.5 achieving a 70% solve rate while Gemini refused to engage almost entirely. The test app cont

Gradient estimators for parameter inference in discrete stochastic kinetic models

Model ReleasesDGX agent

arXiv:2604.02121v2 Announce Type: replace-cross Abstract: Stochastic kinetic models are ubiquitous in physics, yet inferring their parameters from experimental data remains challenging. For determinis

Graph Set Transformer

Model ReleasesDGX agent

arXiv:2606.05116v1 Announce Type: new Abstract: We introduce the Graph Set Transformer (GST), a neural network architecture for learning on sets of graphs, designed for tasks in which per-element pred

Gravity-Aware Hierarchical Routing for Lightweight SensorLLM on Human Activity Recognition

Model ReleasesDGX agent

arXiv:2606.04019v1 Announce Type: cross Abstract: Recent studies on sensor-language alignment have shown that two-stage frameworks can improve the semantic modeling ability of wearable-sensor human ac

GroupToM-Bench: Benchmarking Group Theory of Mind and Nonlinear Social Emergence in MLLMs

Model ReleasesDGX agent

arXiv:2606.04184v1 Announce Type: new Abstract: True general intelligence requires not only a model of the physical world but also a social world model: the capacity to infer how individual mental sta

HD-DinoMoE: A Class-Aware Hierarchical Dual Mixture-of-Experts Network for Scleral Anomaly Segmentation in Complex Acquisition Scenarios

Model ReleasesDGX agent

arXiv:2606.04888v1 Announce Type: new Abstract: Traditional Chinese Medicine (TCM) ocular inspection provides empirical cues for assessing scleral surface anomalies, but its clinical use remains subje

Hierarchical Self-Supervised Adversarial Training for Robust Vision Models in Histopathology

Model ReleasesDGX agent

arXiv:2503.10629v2 Announce Type: replace Abstract: Adversarial attacks pose significant challenges for vision models in critical fields like healthcare, where reliability is essential. Although adver

High-Quality Entity Segmentation and Grounding

Model ReleasesDGX agent

arXiv:2402.02555v2 Announce Type: replace-cross Abstract: In this work, we propose ESG, a pipeline for high-quality entity segmentation and grounding supported by a new dataset EntitySeg. At first, th

Highlighting recent advances in multi-GPU and tensor parallel support in llama.cpp Over the last few months llama.cpp maintainers and engine…

Model ReleasesDGX agent

Highlighting recent advances in multi-GPU and tensor parallel support in llama.cpp Over the last few months llama.cpp maintainers and engineers from NVIDIA collaborated to improve the multi-GPU perfor

HighTide: An Agent-Curated Open-Source VLSI Benchmark Suite

Model ReleasesDGX agent

arXiv:2606.04126v1 Announce Type: cross Abstract: We introduce HighTide, an evolving AI-assisted benchmark suite. Specifically, the contributions are: (i) a diverse open-source suite spanning multiple

HORIZON: Recoverability-Governed Curriculum for Physical-Domain Scaling

Model ReleasesDGX agent

arXiv:2606.05143v1 Announce Type: new Abstract: Scaling robust robot policies requires more than broader randomization, because physical-domain experience must remain organized and learnable throughou

How dynamic workflows allow Claude Code to handle whole new types of tasks https://x.com/trq212/status/2061907337154367865

Model ReleasesDGX agent

Dynamic workflows in Claude Code enable the model to handle complex, multi-step tasks by allowing execution flows to adapt based on intermediate results rather than following fixed paths. This capabil

How to Fine-Tune Nemotron 3.5 ASR for Your Language, Domain, or Accent

Model ReleasesDGX agent

This guide explains how to adapt NVIDIA's Nemotron 3.5 Automatic Speech Recognition (ASR) model to specific languages, domains, or accents through fine-tuning techniques. It likely covers the fine-tun

https://ollama.com/library/nemotron-3-ultra

Model ReleasesDGX agent

Nemotron-3-Ultra is a large language model available through Ollama's model library, likely representing an advanced iteration in NVIDIA's Nemotron model series optimized for performance and capabilit

Hyper-ICL: Attention Calibration with Hyperbolic Anchor Distillation for Multimodal In-Context Learning

Model ReleasesDGX agent

arXiv:2606.04434v1 Announce Type: new Abstract: Multimodal In-Context Learning (ICL) has emerged as a practical inference paradigm for Multimodal Large Language Models, where a small set of interleave

I am hooked on Dynamic Workflows! The idea of generating harnesses on the fly is so compelling that I reverse-engineered it for my agent orc…

Model ReleasesDGX agent

I am hooked on Dynamic Workflows! The idea of generating harnesses on the fly is so compelling that I reverse-engineered it for my agent orchestrator. And then I built a monitoring dashboard (as an HT

Identifying and Correcting Label Noise for Robust GNNs via Influence Contradiction

Model ReleasesDGX agent

arXiv:2601.17469v2 Announce Type: replace Abstract: Graph Neural Networks (GNNs) have shown remarkable capabilities in learning from graph-structured data with various applications such as social anal

Iliad (Troy) trailer made by Grok Imagine 1.5, which was just released

Model ReleasesDGX agent

Elon Musk shared a trailer for 'Iliad (Troy)' created using Grok Imagine 1.5, Xai's newly released text-to-image generation model. The post demonstrates the capabilities of the latest version of Grok'

Imagine Before You Draw: Visual Prompt Engineering for Image Generation

Model ReleasesDGX agent

arXiv:2606.04457v1 Announce Type: new Abstract: Incorporating visual semantic representations as an intermediate step before image generation can reduce the modeling difficulty between text and images

Impostor: An Agent-Curated Benchmark for Realistic AIGC Manipulation Localization

Model ReleasesDGX agent

arXiv:2606.04545v1 Announce Type: new Abstract: Recent advances in generative image editing have improved the realism and controllability of localized image manipulation, raising new challenges for im

In policy paper, OpenAI diverges from White House on AI safety

Model ReleasesDGX agent

OpenAI Group PBC’s newly released proposal for how advanced artificial intelligence should be regulated differs slightly from the Trump administration’s executive order, also released this week. Relea

InstantRetouch: Efficient and High-Fidelity Instruction-Guided Image Retouching with Bilateral Space

Model ReleasesDGX agent

arXiv:2606.05071v1 Announce Type: new Abstract: Language-guided photo retouching aims to adjust color and tone while preserving geometry and texture. Recently, diffusion-based retouching shows a super

Interfaze: The Future of AI is built on Task-Specific Small Models

Model ReleasesDGX agent

arXiv:2602.04101v2 Announce Type: replace Abstract: We present Interfaze, a native hybrid model that fuses task-specific deep neural networks (CNNs and DNNs) directly into a transformer decoder throug

Introducing NVIDIA Nemotron 3 Ultra. A frontier smart open model built for long-running agents that need to plan, reason, use tools and keep…

Model ReleasesDGX agent

Introducing NVIDIA Nemotron 3 Ultra. A frontier smart open model built for long-running agents that need to plan, reason, use tools and keep working across complex coding, research and enterprise work

Introducing two NVIDIA Nemotron models on Together AI: Nemotron 3 Ultra for high-throughput agentic workloads and Nemotron 3.5 ASR for low-l…

Model ReleasesDGX agent

Introducing two NVIDIA Nemotron models on Together AI: Nemotron 3 Ultra for high-throughput agentic workloads and Nemotron 3.5 ASR for low-latency multilingual speech recognition. AI natives can now b

Invariant Gradient Alignment for Robust Reasoning Distillation

Model ReleasesDGX agent

arXiv:2606.05025v1 Announce Type: cross Abstract: Large language models (LLMs) suffer from shortcut learning: they systematically fail on out-of-distribution (OOD) inputs whose semantic surface differ

It's TIME: Towards the Next Generation of Time Series Forecasting Benchmarks

Model ReleasesDGX agent

arXiv:2602.12147v4 Announce Type: replace Abstract: Time series foundation models (TSFMs) are revolutionizing the forecasting landscape from specific dataset modeling to generalizable task evaluation.

Knowledge Index of Noah's Ark

Model ReleasesDGX agent

arXiv:2606.05104v1 Announce Type: new Abstract: Knowledge benchmarks for LLMs face three issues: scaling-driven designs that do not operationalize disciplinary representativeness; flat-payment annotat

LCSHBench: A Multilingual, Consensus-Grounded Benchmark for Library of Congress Subject Heading Assignment

Model ReleasesDGX agent

arXiv:2606.04382v1 Announce Type: cross Abstract: Automated subject cataloging assigns controlledvocabulary headings to bibliographic records, but LCSH has no standard public benchmark. We introduce L

LDA-1B: Scaling Latent Dynamics Action Model via Universal Embodied Data Ingestion

Model ReleasesDGX agent

arXiv:2602.12215v2 Announce Type: replace Abstract: Recent robot foundation models largely rely on large-scale behavior cloning, which imitates expert actions but discards transferable dynamics knowle

LDARNet: DNA Adaptive Representation Network with Learnable Tokenization for Genomic Modeling

Model ReleasesDGX agent

arXiv:2606.04552v1 Announce Type: new Abstract: Genomic foundation models increasingly adopt large language model architectures, yet almost universally rely on fixed tokenization schemes such as k-mer

Learning Long Range Spatio-Temporal Representations over Continuous Time Dynamic Graphs with State Space Models

Model ReleasesDGX agent

arXiv:2606.04672v1 Announce Type: cross Abstract: Continuous-time dynamic graphs (CTDGs) provide a richer framework to capture fine-grained temporal patterns in evolving relational data. Long-range in

Learning When to Act or Refuse: Guarding Agentic Reasoning Models for Safe Multi-Step Tool Use

Model ReleasesDGX agent

arXiv:2603.03205v2 Announce Type: replace Abstract: Agentic language models operate in a fundamentally different safety regime than chat models: they must plan, call tools, and execute long-horizon ac

Leaving aside the question of consciousness, the Ted Chiang piece has a reasonable point about moral atrophy if you let AI make choices. But…

Model ReleasesDGX agent

Leaving aside the question of consciousness, the Ted Chiang piece has a reasonable point about moral atrophy if you let AI make choices. But it is also interesting in light of the fact that repeated r

LifeSide: Benchmarking Agents as Lifelong Digital Companions

Model ReleasesDGX agent

arXiv:2606.04660v1 Announce Type: new Abstract: Lifelong digital companions must integrate cross-session cues, continually update their understanding of users, and adapt to shifting privacy boundaries

LiftQuant: Continuous Bit-Width LLM via Dimensional Lifting and Projection

Model ReleasesDGX agent

arXiv:2606.04050v1 Announce Type: cross Abstract: Existing quantization methods are fundamentally limited by rigid, integer-based bit-widths (e.g., 2, 3-bit), resulting in a ``deployment gap' where La

LimiX-2M: Mitigating Low-Rank Collapse and Attention Bottlenecks in Tabular Foundation Models

Model ReleasesDGX agent

arXiv:2606.04485v1 Announce Type: new Abstract: Tabular foundation models (TFMs) increasingly rival tree ensembles, but their performance is often compute-inefficient: with standard affine scalar toke

Listen to the OpenAI Podcast on— Spotify https://open.spotify.com/episode/3ca5s3o53D5xcEKmKgLLGj?si=4a9a555641fa4293 Apple https://podcasts.…

Model ReleasesDGX agent

Listen to the OpenAI Podcast on— Spotify https://open.spotify.com/episode/3ca5s3o53D5xcEKmKgLLGj?si=4a9a555641fa4293 Apple https://podcasts.apple.com/us/podcast/how-a-reasoning-model-cracked-an-80-yea

Literature-Guided Minimax Optimization of Virtual Epilepsy Neurostimulation

Model ReleasesDGX agent

arXiv:2606.04339v1 Announce Type: new Abstract: Computational models of epilepsy promise patient-specific treatment design, but most optimization workflows still search for parameters that perform wel

LLMs + Persona-Plug = Personalized LLMs

Model ReleasesDGX agent

arXiv:2409.11901v2 Announce Type: replace Abstract: Personalization plays a critical role in numerous language tasks and applications, since users with the same requirements may prefer diverse outputs

Long Live Fine-Tuning: Task-Specific Transformers Outperform Zero-Shot LLMs for Misinformation Response Classification on Reddit

Model ReleasesDGX agent

arXiv:2606.04274v1 Announce Type: new Abstract: As large language models (LLMs) become default tools for online information verification, an implicit assumption follows them: that scale and general ca

Longer Context, Deeper Thinking: Uncovering the Role of Long-Context Ability in Reasoning

Model ReleasesDGX agent

arXiv:2505.17315v2 Announce Type: replace Abstract: Recent language models exhibit strong reasoning capabilities, yet the influence of long-context capacity on reasoning remains underexplored. In this

Look closely. There’s more in the Showcase.

Model ReleasesDGX agent

OpenAI's developer account posted this message on X (formerly Twitter), likely encouraging developers to explore additional features, updates, or resources available in OpenAI's Showcase platform or d

LoopMoE: Unifying Iterative Computation with Mixture-of-Experts for Language Modeling

Model ReleasesDGX agent

arXiv:2606.04438v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) and looped architectures scale models along two orthogonal axes, namely parameter capacity and effective depth. However, main

M^3Eval: Multi-Modal Memory Evaluation through Cognitively-Grounded Video Tasks

Model ReleasesDGX agent

arXiv:2606.05008v1 Announce Type: cross Abstract: As multi-modal models advance towards long-form video understanding, memory emerges as a critical capability. Despite substantial efforts in developin

MedForge: Interpretable Medical Deepfake Detection via Forgery-aware Reasoning

Model ReleasesDGX agent

arXiv:2603.18577v2 Announce Type: replace Abstract: Text-guided image editors can now manipulate authentic medical scans with high fidelity, enabling lesion implantation/removal that threatens clinica

MemoryDocDataSet: A Benchmark for Joint Conversational Memory and Long Document Reasoning

Model ReleasesDGX agent

arXiv:2606.04442v1 Announce Type: cross Abstract: AI systems increasingly need to combine two demanding capabilities: navigating multi-session conversation history and performing deep reading comprehe

MesaNet: Sequence Modeling by Locally Optimal Test-Time Training

Model ReleasesDGX agent

arXiv:2506.05233v2 Announce Type: replace-cross Abstract: Sequence modeling is currently dominated by causal transformer architectures that use softmax self-attention. Although widely adopted, transfo

MeshTok: Efficient Multi-Scale Tokenization for Scalable PDE Transformers

Model ReleasesDGX agent

arXiv:2606.04366v1 Announce Type: new Abstract: Conventional patchified Transformers operate on uniform spatial partitions, distributing computational effort evenly across the domain irrespective of l

← Previous
1…169170171172173…377
Next →