AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,566 results
30 Jun 2026

MUSE: Unlocking Timestep as Native Task Steering for One-Step Dense Prediction

Model ReleasesDGX agent

arXiv:2606.30370v1 Announce Type: new Abstract: Monocular dense prediction has recently seen remarkable success by repurposing pre-trained diffusion models. This opens a promising yet challenging aven

MuseBench: Benchmarking Intent-Level Audiovisual Arts Understanding in MLLMs

Model ReleasesDGX agent

arXiv:2606.30026v1 Announce Type: cross Abstract: Audiovisual arts encompass diverse creative disciplines, including cinema, visual arts, stage performance, and game design, where artistic meaning ari

my process for writing right now is to do some engineering work, talk to a bunch of people about it, brainstorm and research with Claude, wr…


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases
DGX agent

my process for writing right now is to do some engineering work, talk to a bunch of people about it, brainstorm and research with Claude, write a post, give 1 or 2 talks on it, rewrite the post, give

Nano Banana 2 Lite

Model ReleasesDGX agent

Nano Banana 2 Lite Also known as Gemini 3.1 Flash Lite Image (gemini-3.1-flash-lite-image in their API), this is the 'fastest and cheapest Gemini image model, engineered for velocity and scale'. I use

Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image) now on AI Gateway

Model ReleasesDGX agent

Vercel has announced the availability of Nano Banana 2 Lite, a lightweight image variant featuring Google's Gemini 3.1 Flash model, through its AI Gateway service. This release likely provides develop

Nemotron-Labs-Diffusion-Image: Advancing Masked Discrete Diffusion for High-Resolution Image Synthesis

Model ReleasesDGX agent

arXiv:2606.29814v1 Announce Type: new Abstract: We propose Nemotron-Labs-Diffusion-Image, a state-of-the-art masked discrete diffusion model (MDM) for high-resolution text-to-image synthesis. Compared

// Neural procedural memory // Good paper on agent memory beyond prompt retrieval. NPM stores procedural skills as activation steering vecto…

Model ReleasesDGX agent

// Neural procedural memory // Good paper on agent memory beyond prompt retrieval. NPM stores procedural skills as activation steering vectors distilled from contrastive historical experience. Textual

Neural Subspace Reallocation: Continual Learning as Retrieval-Based Subspace Memory Management

Model ReleasesDGX agent

arXiv:2606.30067v1 Announce Type: cross Abstract: We introduce Neural Subspace Reallocation (NSR), which reframes continual learning as memory management over parameter subspaces. Instead of treating

Never Skip a Batch: Dense Learning of Temporal GNNs via Adaptive Pseudo-Supervision

Model ReleasesDGX agent

arXiv:2505.12526v2 Announce Type: replace Abstract: Temporal graph networks suffer from irregular supervision in realworld dynamic graphs, as most minibatches contain few labeled events. The lack of l

NEW on Hugging Face: Hardware filters 🖥️ A new Hardware filter on the Models page results to models that fit a specific GPU, CPU, or Apple …

Model ReleasesDGX agent

NEW on Hugging Face: Hardware filters 🖥️ A new Hardware filter on the Models page results to models that fit a specific GPU, CPU, or Apple Silicon chip, so you only see what will actually run on your

Nonlinear mixture model motivated subspace clustering

Model ReleasesDGX agent

arXiv:2606.29261v1 Announce Type: cross Abstract: We derive the linear union-of-subspaces (UoS) model for subspace clustering (SC) from the nonlinear mixture model (NMM) used in blind source separatio

Not-quite-human tastes: the stylized omnivorousness of LLM survey surrogates

Model ReleasesDGX agent

arXiv:2606.30085v1 Announce Type: new Abstract: Large-language models have proven to be remarkable if inconsistent parrots of public attitudes and opinions. The extent to which LLMs are able to produc

Notes (and a Pelican) on Claude Sonnet 5 - the new tokenizer makes it ~1.4x more expensive for English, ~1.33x more expensive for Spanish bu…

Model ReleasesDGX agent

Notes (and a Pelican) on Claude Sonnet 5 - the new tokenizer makes it ~1.4x more expensive for English, ~1.33x more expensive for Spanish but roughly the same price for Simplified Mandarin https://sim

NVIDIA BioNeMo Agent Toolkit Brings Accelerated AI to Life Sciences Researchers in Claude Science

Model ReleasesDGX agent

Life sciences has entered an era of computational scale, and for more than a decade, NVIDIA has built the full GPU-accelerated computing stack — spanning hardware, frameworks, libraries, models, micro

Obliviate: Erasing Concepts from Autoregressive Image Generation Models

Model ReleasesDGX agent

arXiv:2606.28643v1 Announce Type: new Abstract: The widespread adoption of generative AI models has intensified concerns about misuse, including the creation of unsafe or disturbing imagery. To mitiga

OmniCoT: A Benchmark for Global and Multi-Step Panoramic Reasoning

Model ReleasesDGX agent

arXiv:2606.30378v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have demonstrated promising spatial reasoning capabilities, while these abilities remain underexplored in the e

On FrontierCode (Extended), our benchmark for real-world engineering tasks that grades mergeability and quality, Sonnet 5 scores 53.8% and h…

Model ReleasesDGX agent

On FrontierCode (Extended), our benchmark for real-world engineering tasks that grades mergeability and quality, Sonnet 5 scores 53.8% and has a 57.6% pass rate (higher than Opus 4.8). Note: These rel

On the Policy Gradient Foundations of Group Relative Policy Optimization: Credit Assignment, Gradient Sparsity, and Rank Collapse

Model ReleasesDGX agent

arXiv:2606.29238v1 Announce Type: new Abstract: Group Relative Policy Optimization (GRPO) eliminates the learned critic in PPO by using the mean reward of grouped rollouts as a baseline. We provide a

On the Vulnerability of Parameter-Level Defenses to Model Merging

Model ReleasesDGX agent

arXiv:2606.30360v1 Announce Type: cross Abstract: The training-free integration of expert models via model merging has exposed significant security risks, enabling free-riders to combine specialized m

One Scene, Two Depths: Probing Geometric Ambiguity in Monocular Foundation Models

Model ReleasesDGX agent

arXiv:2606.29600v1 Announce Type: cross Abstract: A faithful 3D world representation should account for layered geometry, where a single camera ray may contain multiple visible and geometrically valid

OP3DSG: Open-Vocabulary Part-Aware 3D Scene Graph Generation for Real-World Environments

Model ReleasesDGX agent

arXiv:2606.29786v1 Announce Type: new Abstract: 3D scene graphs (3DSGs) provide a compact and structured abstraction of 3D environments. Although advances in foundation models have enabled open-vocabu

Optimizing Expert-Designed Crystal Graph Networks for Band-Gap Prediction with an Autonomous LLM Research Loop

Model ReleasesDGX agent

arXiv:2606.29717v1 Announce Type: cross Abstract: Predicting a material's properties from its structure is a central, fast-advancing problem in computational materials science. A decade of work has pr

OptiMUS-0.3: Using Large Language Models to Model and Solve Optimization Problems at Scale

Model ReleasesDGX agent

arXiv:2407.19633v4 Announce Type: replace Abstract: Optimization problems are pervasive in sectors from manufacturing and distribution to healthcare. However, most such problems are still solved heuri

ORCA: Open-ended Response Correctness Assessment for Audio Question Answering

Model ReleasesDGX agent

arXiv:2512.09066v2 Announce Type: replace-cross Abstract: Reliable assessment of the abilities of large audio language models (LALMs) is essential to advancing the state of the art. As benchmarks rapi

Ornith-1.0-35B is now available in claude code through hf-claude

Model ReleasesDGX agent

Ornith-1.0-35B, a 35-billion parameter model, has been made available for use through Claude Code via Hugging Face integration. This announcement indicates expanded model availability and integration

OSWorld2.0: Benchmarking Computer Use Agents on Long-Horizon Real-World Tasks

Model ReleasesDGX agent

arXiv:2606.29537v1 Announce Type: new Abstract: Existing computer-use benchmarks fail to capture the realism, complexity, and long-horizon demands of real-world computer use, limiting their ability to

Parametric Skills

Model ReleasesDGX agent

arXiv:2606.30015v1 Announce Type: new Abstract: Since intelligence fundamentally relies on efficient skill acquisition (Chollet, 2019), the ability to leverage skills is critical. For LLMs, skills, ma

PCGD: Physics-Guided Conditional Graph Diffusion for TCAD Device Simulation

Model ReleasesDGX agent

arXiv:2606.29272v1 Announce Type: new Abstract: Technology computer-aided design (TCAD) semiconductor device simulation is fundamentally constrained by the high computational cost of iteratively solvi

Perforce launches Agentic Gateway to govern AI agents and cut token costs

Model ReleasesDGX agent

Perforce Software Inc. today expanded its Perforce Intelligence lineup with an agentic gateway for managing artificial intelligence agents, an autonomous testing platform driven by natural language an

PGE-SAM: Prompt-Guided Feature Enhancement for Interactive Segmentation under Degradation

Model ReleasesDGX agent

arXiv:2606.30477v1 Announce Type: new Abstract: Segment Anything Model (SAM) has revolutionized promptable image segmentation with strong zero-shot generalization. However, its performance degrades su

Pie launches with $19.5M to bring AI marketing to small businesses

Model ReleasesDGX agent

Pie Tech Inc., a startup using artificial intelligence to provide growth tools for small businesses, today officially launched with an announcement that it has raised 19.5 million in new funding to ex

PlantExpertVQA: A Visual Question Answering Dataset for Benchmarking Vision-Language Models in Plant Science

Model ReleasesDGX agent

arXiv:2508.17117v3 Announce Type: replace-cross Abstract: Existing plant-disease datasets target classification and detection, leaving vision-language models unable to support interactive, reasoning-b

Pointer-CAD v2: Plan-Then-Construct CAD Generation with Dimension-Aware Parametric Precision

Model ReleasesDGX agent

arXiv:2606.29301v1 Announce Type: new Abstract: Computer-aided design (CAD) plays a fundamental role in modern manufacturing by providing the high precision required for industrial production. Recent

PolicyGuard: A Dialogue-Grounded Sub-Agent Verifier for Policy Adherence in LLM Agents

Model ReleasesDGX agent

arXiv:2606.29225v1 Announce Type: new Abstract: LLM agents handle user requests on behalf of organizations through tool calls and must follow the company policies stated in their system prompts. Prior

Pooled Leaderboards Hide System-Specific Winners: A Reporting-Protocol Audit of Offline Root-Cause Analysis Benchmarks

Model ReleasesDGX agent

arXiv:2606.29159v1 Announce Type: new Abstract: Offline root-cause-analysis (RCA) benchmarks commonly rank methods by a single pooled top-1 accuracy across multiple subsystems, and engineers often rea

PoseShield: Neural Collision Fields for Human Self-Collision Resolution

Model ReleasesDGX agent

arXiv:2606.29686v1 Announce Type: new Abstract: Self-collision remains a persistent challenge in SMPL-based human pose estimation and motion generation. Under extreme articulations or stochastic motio

Position: RL Researchers Need to Distinguish Between Solving Simulators and Using Simulators as a Proxy

Model ReleasesDGX agent

arXiv:2606.28433v1 Announce Type: new Abstract: One goal in reinforcement learning (RL) research is to understand general-purpose sequential decision-making, using benchmark simulators as a proxy for

Post-training for Efficient Communication via Convention Formation

Model ReleasesDGX agent

arXiv:2508.06482v2 Announce Type: replace-cross Abstract: Humans communicate with increasing efficiency in multi-turn interactions, by adapting their language and forming ad-hoc conventions. In contra

Primary ICD Category Prediction using LLM-based Probing

Model ReleasesDGX agent

arXiv:2606.28798v1 Announce Type: new Abstract: Objective: ICD codes are central to reimbursement, research, and population health surveillance, yet automated coding systems often struggle to integrat

Probabilistic Approach to Black-Box Binary Optimization with Budget Constraints: Application to Sensor Placement

Model ReleasesDGX agent

arXiv:2406.05830v2 Announce Type: replace-cross Abstract: This paper presents a fully probabilistic approach for solving optimal experimental design problems under budget constraints. The experimental

Progressive Self-Supervised Learning with Individualized Community Assignment for Brain Network Analysis

Model ReleasesDGX agent

arXiv:2606.29695v1 Announce Type: new Abstract: Brain networks exhibit a modular community structure that varies across individuals and neurological conditions. However, existing self-supervised learn

Projected Exploitability Descent for Nash Equilibrium Computation in Multiplayer Imperfect-Information Games

Model ReleasesDGX agent

arXiv:2606.29169v1 Announce Type: cross Abstract: Many important games have more than two players and imperfect information. Existing approaches for computing Nash equilibrium, the central game-theore

Proteus: Automated Adversarial Robustness Testing for Audio Deepfake Detectors

Model ReleasesDGX agent

arXiv:2606.29544v1 Announce Type: cross Abstract: We present Proteus, a framework developed at Resemble AI for automated robustness testing of our audio deepfake detection system. Given a detector, Pr

Pushing Forward Pareto Frontiers of Proactive Agents with Behavioral Agentic Optimization

Model ReleasesDGX agent

arXiv:2602.11351v2 Announce Type: replace Abstract: Proactive large language model (LLM) agents aim to actively plan, query, and interact over multiple turns, enabling efficient task completion beyond

Qwen publishes new work on RL coding agents. (bookmark it) The idea is to continually build a verification system that co-evolves with AI ag…

Model ReleasesDGX agent

Qwen publishes new work on RL coding agents. (bookmark it) The idea is to continually build a verification system that co-evolves with AI agents. LLMs suffer from all sorts of reward hacking issues. T

Qwen-RobotNav Technical Report: A Scalable Navigation Model Designed for an Agentic Navigation System

Model ReleasesDGX agent

arXiv:2606.18112v3 Announce Type: replace-cross Abstract: Agentic navigation systems require a base navigation model whose observation strategy can be externally reconfigured at inference time, becaus

RA-QA: A Benchmarking System for Respiratory Audio Question Answering Under Real-World Heterogeneity

Model ReleasesDGX agent

arXiv:2602.18452v3 Announce Type: replace-cross Abstract: As conversational multimodal AI tools are increasingly adopted to process patient data for health assessment, robust benchmarks are needed to

Randomized neural operator for parametric PDEs with fast training and conformal uncertainty quantification

Model ReleasesDGX agent

arXiv:2606.29440v1 Announce Type: new Abstract: Repeatedly solving parametric PDEs is essential for uncertainty quantification, design optimization and inverse problems, but conventional neural operat

RankGraph-2: Lifecycle Co-Design for Billion-Node Graph Learning in Recommendation

Model ReleasesDGX agent

arXiv:2606.18379v2 Announce Type: replace-cross Abstract: Graph-based retrieval at billion-node scale requires jointly solving three tightly coupled problems -- graph construction, representation lear

Reachability Guarantees for Cart-Pole Swing-Up and Stabilization

Model ReleasesDGX agent

arXiv:2606.28627v1 Announce Type: cross Abstract: The cart-pole swing-up is a canonical benchmark for nonlinear control of underactuated systems, yet an end-to-end guarantee linking the global swing-u

Read the full Aston Martin F1 team interview with @aidangomez here: https://www.astonmartinf1.com/en-GB/news/feature/perspectives-aidan-gome…

Model ReleasesDGX agent

This post links to a full interview with Aidan Gomez conducted by the Aston Martin F1 team, published on their official website under their 'Perspectives' feature section. The interview likely covers

Recursive Self-Evolving Agents via Held-Out Selection

Model ReleasesDGX agent

arXiv:2606.28374v1 Announce Type: new Abstract: LLM agents are increasingly improved without weight updates by evolving a natural-language artifact, such as reflections, workflows, playbooks, cheatshe

Redefining Maritime Anomaly Detection via Equation-Grounded Synthetic Anomalies

Model ReleasesDGX agent

arXiv:2606.29721v1 Announce Type: cross Abstract: Maritime anomaly detection is essential for ensuring maritime safety, security, and efficient traffic management at sea, with Automatic Identification

RefAlign: Representation Alignment for Reference-to-Video Generation

Model ReleasesDGX agent

arXiv:2603.25743v2 Announce Type: replace Abstract: Reference-to-video (R2V) generation is a controllable video synthesis paradigm that constrains the generation process using both text prompts and re

Rehearsed Multi-Agent Live Product Demonstrations with Real-Time Voice Question Answering

Model ReleasesDGX agent

arXiv:2606.30294v1 Announce Type: new Abstract: Live product demonstrations are a recurring, high-cost activity in software organizations: a human presenter must select features, dispatch the correspo

Reliability-Prioritized Fine-Grained Generation in Multimodal Large

Model ReleasesDGX agent

arXiv:2606.29573v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) are increasingly expected to generate fine-grained descriptions of visual content. However, we observe and theo

REPAIR-Bench: A Benchmark for Robot Error Perception And Interaction Recovery

Model ReleasesDGX agent

arXiv:2606.29937v1 Announce Type: new Abstract: Understanding how users perceive and respond to robot failures is essential for building robust and trustworthy robot systems. Prior work, however, (i)

Reported Confidence in LLMs Tracks Commitment More Than Correctness

Model ReleasesDGX agent

arXiv:2606.29490v1 Announce Type: cross Abstract: Confidence is an estimate of the probability that a chosen answer is correct. Verbal confidence reports are widely used as uncertainty measures in lar

Representational Depth of Evaluation Awareness Shifts With Scale in Open-Weight Language Models

Model ReleasesDGX agent

arXiv:2606.29196v1 Announce Type: cross Abstract: Do language models know when they are being tested? This question matters for AI safety: a model that recognises an evaluation context could alter its

Research Entity Extraction and Topic Detection from UKRI Grant Proposals

Model ReleasesDGX agent

arXiv:2606.30304v1 Announce Type: cross Abstract: This paper presents preliminary findings from a UKRI-funded Metascience project comparing three LLM-based approaches, GPT-4o, Mistral, and a bespoke a

← Previous
1…120121122123124…377
Next →