AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent
84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,612 results
1 Jun 2026

One of the new, buzzy jobs in Silicon Valley is the AI Forward Deployed Engineer (FDE), an engineer who is embedded within a client organiza…

Model ReleasesDGX agent

One of the new, buzzy jobs in Silicon Valley is the AI Forward Deployed Engineer (FDE), an engineer who is embedded within a client organization to help customize solutions, such as building and tunin

OpenAI frontier models and Codex are now generally available on AWS, giving enterprises a new way to build on Amazon Bedrock with OpenAI thr…

Model ReleasesDGX agent

OpenAI frontier models and Codex are now generally available on AWS, giving enterprises a new way to build on Amazon Bedrock with OpenAI through the security, compliance, and governance workflows they

OpenAI models and Codex on Amazon Bedrock are now generally available


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model ReleasesDGX agent

OpenAI frontier models GPT-5.5 and GPT-5.4, and Codex, the OpenAI coding agent, are now generally available on Amazon Bedrock. AWS customers can access these latest OpenAI models through the same Amaz

Pairwise Reference Alignment as a Model-Level Ordinal Observable

Model ReleasesDGX agent

arXiv:2605.30758v1 Announce Type: new Abstract: Pairwise preference data is widely used in language-model evaluation and alignment, often for model ranking, reward modeling, or preference optimization

Palo Alto Networks says Mythos found 24+ critical bugs using $1M+ in tokens; Anthropic subsidizes Mythos but some companies plan to boost their Mythos budgets (Aaron Holmes/The Information)

Model ReleasesDGX agent

Aaron Holmes / The Information: Palo Alto Networks says Mythos found 24+ critical bugs using $1M+ in tokens; Anthropic subsidizes Mythos but some companies plan to boost their Mythos budgets — When Pa

Parameter-free Dynamic Regret: Time-varying Movement Costs, Delayed Feedback, and Memory

Model ReleasesDGX agent

arXiv:2602.06902v2 Announce Type: replace Abstract: In this paper, we study dynamic regret in unconstrained online convex optimization (OCO) with movement costs. Specifically, we generalize the standa

PhyDrawGen: Physically Grounded Diagram Generation from Natural Language

Model ReleasesDGX agent

arXiv:2605.30512v1 Announce Type: new Abstract: Generating physics diagrams from text requires strict adherence to physical laws. While current generative models produce visually plausible outputs, th

Physically Viable World Models: A Case for Query-Conditioned Embodied AI

Model ReleasesDGX agent

arXiv:2605.30542v1 Announce Type: new Abstract: World models for embodied AI must be physically viable: constructed to answer intervention queries by representing the physical structure governing acti

Physics Enhanced Deep Surrogates for the Phonon Boltzmann Transport Equation

Model ReleasesDGX agent

arXiv:2512.05976v3 Announce Type: replace-cross Abstract: Designing materials with controlled heat flow at the nano-scale is central to advances in microelectronics, thermoelectrics, and energy-conver

PInVerify: An Offline Embodied Benchmark for Active Instance Verification

Model ReleasesDGX agent

arXiv:2605.30639v1 Announce Type: cross Abstract: Embodied agents have made strong progress in navigating to target objects, but reaching the goal vicinity does not guarantee that the agent has found

Plain Transformers are Surprisingly Powerful Link Predictors

Model ReleasesDGX agent

arXiv:2602.01553v2 Announce Type: replace-cross Abstract: Link prediction is a core challenge in graph machine learning, demanding models that capture rich and complex topological dependencies. While

PRISM: Progressive Reasoning through Iterative Slot Memory for Vision

Model ReleasesDGX agent

arXiv:2605.30942v1 Announce Type: new Abstract: Modern vision models process images in a single feed-forward pass, which limits their ability to recover missing evidence or refine uncertain representa

Probabilistic Precipitation Nowcasting with Rectified Flow Transformers

Model ReleasesDGX agent

arXiv:2605.31204v1 Announce Type: new Abstract: Accurate weather forecasts are essential across various domains and are safety-critical in extreme weather conditions. Compared to simulation-based fore

Probing Collision Grounding in Vision-Language Models for Safe Human-Robot Collaboration

Model ReleasesDGX agent

arXiv:2605.31196v1 Announce Type: cross Abstract: Safe human--robot collaboration requires more than visual description: a monitor must determine whether the robot body is safely separated, already co

Probing the Prompt KV Cache: Where It Becomes Dispensable

Model ReleasesDGX agent

arXiv:2605.30574v1 Announce Type: new Abstract: Prior KV cache compression schemes empirically demonstrate that the prompt cache is partially redundant during decoding, dropping or summarising entries

QASM-Eval: A Dataset to Train and Evaluate LLMs on OpenQASM-3 Beyond Quantum Circuits

Model ReleasesDGX agent

arXiv:2605.30358v1 Announce Type: new Abstract: Quantum computing remains in the Noisy Intermediate-Scale Quantum (NISQ) era, where the performance is highly constrained to noise. Addressing the limit

Quantifying the Uncertainty of Foundation Models with Singular Value Ensembles

Model ReleasesDGX agent

arXiv:2601.22068v2 Announce Type: replace Abstract: Foundation models have become a dominant paradigm in machine learning, achieving remarkable performance across diverse tasks through large-scale pre

Query-focused and Memory-aware Reranker for Long Context Processing

Model ReleasesDGX agent

arXiv:2602.12192v3 Announce Type: replace Abstract: Built upon the existing analysis of retrieval heads in large language models, we propose an alternative reranking framework that trains models to es

QVGGT: Post-Training Quantized Visual Geometry Grounded Transformer

Model ReleasesDGX agent

arXiv:2605.31124v1 Announce Type: new Abstract: Estimating 3D attributes directly from images has advanced rapidly with the Visual Geometry Grounded Transformer (VGGT), which predicts camera parameter

Qwen 3.7 Plus now available on AI Gateway

Model ReleasesDGX agent

Vercel has made Qwen 3.7 Plus, an AI model, available through its AI Gateway service, allowing developers to access this model via Vercel's platform. This addition expands the model options available

Randomized Feasibility Methods for Constrained Optimization with Adaptive Step Sizes

Model ReleasesDGX agent

arXiv:2601.20076v2 Announce Type: replace-cross Abstract: We consider minimizing an objective function subject to constraints defined by the intersection of lower-level sets of convex functions. We st

Re-examining Low Rank adaptation for private LLM fine-tuning

Model ReleasesDGX agent

arXiv:2510.01137v3 Announce Type: replace Abstract: Privacy is a central concern when fine-tuning large language models (LLMs) on sensitive data, and differentially private stochastic gradient descent

Read more about all the fun ways we used AI to bring I/O to life this year: https://blog.google/innovation-and-ai/technology/ai/io-2026-goog…

Model ReleasesDGX agent

Google's I/O 2026 event showcased various AI applications and innovations developed by Google to enhance the conference experience. The post highlights creative implementations of AI technology integr

Recognizing Co-Speech Gestures in-the-Wild

Model ReleasesDGX agent

arXiv:2605.31589v1 Announce Type: new Abstract: While humans naturally gesture during speech, only a sparse subset of these movements are visually depictive and semantically linked to specific spoken

ReTabAD: A Benchmark for Restoring Semantic Context in Tabular Anomaly Detection

Model ReleasesDGX agent

arXiv:2510.02060v2 Announce Type: replace Abstract: In tabular anomaly detection (AD), textual semantics often carry critical signals, as the definition of an anomaly is closely tied to domain-specifi

SAEmnesia: Erasing Concepts in Diffusion Models with Supervised Sparse Autoencoders

Model ReleasesDGX agent

arXiv:2509.21379v3 Announce Type: replace-cross Abstract: Concept unlearning in diffusion models is hampered by feature splitting, where concepts are distributed across many latent features, making th

Safe Equilibrium Policy Optimization for Strategic Agent Policies

Model ReleasesDGX agent

arXiv:2605.30854v1 Announce Type: cross Abstract: Language models fine-tuned with reinforcement learning typically optimize for task reward, ignoring multi-agent strategic structure. Because these age

Same Patient, Different Words, Different Diagnosis? Evaluating Semantic Stability in Clinical LLMs

Model ReleasesDGX agent

arXiv:2605.30646v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used in clinical applications. However, their behavior remains highly sensitive to subtle linguistic var

SAW-Bench: Learning Situated Awareness in the Real World

Model ReleasesDGX agent

arXiv:2602.16682v2 Announce Type: replace Abstract: A core aspect of human perception is situated awareness, the ability to relate ourselves to the surrounding physical environment and reason over pos

Scaling Conversational Hungarian ASR: The BEA-Dialogue+ Corpus

Model ReleasesDGX agent

arXiv:2605.31469v1 Announce Type: cross Abstract: Conversational automatic speech recognition in Hungarian is constrained by the limited amount of publicly available dialogue-style training data. The

Scaling Multi-Hop Training Data via Graph-Constrained Path Selection

Model ReleasesDGX agent

arXiv:2605.31238v1 Announce Type: new Abstract: Endowing large language models with compositional reasoning over specialized documents requires multi-hop training data at scale, where such data rarely

Self-Tuning Regularization for Image Scanning Microscopy

Model ReleasesDGX agent

arXiv:2605.31426v1 Announce Type: cross Abstract: Image Scanning Microscopy (ISM) is a fluorescence imaging technique that combines detector-array acquisition and computational reconstruction to achie

Send a SCOUT First: Pre-hoc Reasoning for Adaptive Detector Allocation in Prompt-Injection Defense

Model ReleasesDGX agent

arXiv:2605.30837v1 Announce Type: cross Abstract: Prompt-injection detectors are heterogeneous: each is strong on a different slice of attacks, and none is always reliable. Yet existing systems still

Sequential Least-Squares Estimators with Fast Randomized Sketching for Linear Statistical Models

Model ReleasesDGX agent

arXiv:2509.06856v2 Announce Type: replace-cross Abstract: We propose a novel randomized framework for the estimation problem of large-scale linear statistical models, namely Sequential Least-Squares E

Sequential Subspace Noise Injection Prevents Accuracy Collapse in Certified Unlearning

Model ReleasesDGX agent

arXiv:2601.05134v2 Announce Type: replace Abstract: Certified unlearning based on differential privacy offers strong guarantees but remains largely impractical: the noisy fine-tuning approaches propos

SERA: Soft-Verified Efficient Repository Agents

Model ReleasesDGX agent

arXiv:2601.20789v3 Announce Type: replace Abstract: Open-weight coding agents should hold a fundamental advantage over closed-source systems because they can specialize to private codebases, encoding

SimulCost: A Cost-Aware Benchmark and Toolkit for Automating Physics Simulations with LLMs

Model ReleasesDGX agent

arXiv:2603.20253v2 Announce Type: replace-cross Abstract: Evaluating LLM agents for scientific tasks has focused on token costs while ignoring tool-use costs like simulation time and experimental reso

Skill Availability and Presentation Granularity in Large-Language-Model Agents: A Controlled SkillsBench Study

Model ReleasesDGX agent

arXiv:2605.31408v1 Announce Type: cross Abstract: Skill documents provide procedural knowledge to large-language-model agents at inference time. This article studies whether the presentation granulari

Smaller and Faster 3DGS via Post-Training Dictionary Learning

Model ReleasesDGX agent

arXiv:2605.30396v1 Announce Type: cross Abstract: 3D Gaussian Splatting (3DGS) is a promising neural scene representation for real-time rendering, but trained models often suffer from large memory foo

So much great work lately from Nvidia, the 'King of American Open-source AI'! - Crossed 1,000 total public repositories on @huggingface (820…

Model ReleasesDGX agent

So much great work lately from Nvidia, the 'King of American Open-source AI'! - Crossed 1,000 total public repositories on @huggingface (820 models, 249 datasets & 57 spaces) & almost 60,000 followers

Social welfare optimisation under institutional reward and punishment

Model ReleasesDGX agent

arXiv:2605.31330v1 Announce Type: cross Abstract: Institutional incentives are widely used to promote cooperation among autonomous, self-regarding agents, from human societies to multi-agent and AI sy

SOCO: Benchmarking Semantic Object Correspondence in Vision Foundation Models

Model ReleasesDGX agent

arXiv:2605.31597v1 Announce Type: new Abstract: Measuring structured object understanding in vision foundation models remains challenging due to inconsistent evaluation protocols and limited part-leve

Softsign: Smooth Sign in Your Optimizer For Better Parameter Heterogeneity Handling

Model ReleasesDGX agent

arXiv:2605.31371v1 Announce Type: new Abstract: Sign-based and LMO-inspired optimizers have recently attracted substantial attention in deep learning due to their strong performance and low memory foo

SpatialAct: Probing Spatial Reasoning-to-Action Capabilities of VLM Agents in 3D Scenes

Model ReleasesDGX agent

arXiv:2605.31148v1 Announce Type: cross Abstract: Humans can effortlessly perceive spatial layouts, form cognitive representations, reason about spatial relations, and translate such reasoning into ac

SPM-Bench: Benchmarking Large Language Models for Scanning Probe Microscopy

Model ReleasesDGX agent

arXiv:2602.22971v2 Announce Type: replace Abstract: As LLMs achieved breakthroughs in general reasoning, their proficiency in specialized scientific domains reveals pronounced gaps in existing benchma

Steering LLMs? Actually, Sparse Autoencoders can outperform simple baselines

Model ReleasesDGX agent

arXiv:2605.31183v1 Announce Type: cross Abstract: Sparse Autoencoders (SAEs) have been seen as a promising avenue for exploring the internals of Large Language Models (LLMs) and for steering model out

Structured interactions improve distributed coordination beyond model scaling in a real-world multi-robot system

Model ReleasesDGX agent

arXiv:2605.30383v1 Announce Type: cross Abstract: Scaling individual robot capabilities is common but costly. Here we investigate a system-level design question in real-world multi-robot coordination:

SVI-Bench: A Dynamic Microworld for Strategic Video Intelligence

Model ReleasesDGX agent

arXiv:2605.31529v1 Announce Type: new Abstract: True video intelligence demands more than recognizing what is visible: it requires reasoning about why events unfold, predicting what would change under

Symbolic Intermediaries as a Linguistic-Numerical Interface for LLM-Driven Geometric Reasoning

Model ReleasesDGX agent

arXiv:2505.17607v3 Announce Type: replace Abstract: Large Language Models (LLMs) display reasoning capabilities over linguistic and symbolic objects but have limited capabilities to directly interpret

TabCausal: Pretraining Across Causal Environments for Tabular Causal Discovery

Model ReleasesDGX agent

arXiv:2605.31156v1 Announce Type: new Abstract: Causal discovery aims to recover directed causal relations from observational and interventional data, providing a basis for mechanistic understanding a

TAGA: A Tangent-Based Reactive Approach for Socially Compliant Robot Navigation Around Human Groups

Model ReleasesDGX agent

arXiv:2503.21168v3 Announce Type: replace Abstract: Robots navigating human-populated environments must avoid collisions while respecting the social structure of crowds, particularly the implicit boun

Target-Agnostic Calibration under Distribution Shift with Frequency-Aware Gradient Rectification

Model ReleasesDGX agent

arXiv:2508.19830v2 Announce Type: replace-cross Abstract: Real-world model deployments inevitably encounter distribution shifts, rendering the confidence estimates of deep neural networks highly unrel

Targeted Speaker Poisoning Framework in Zero-Shot Text-to-Speech

Model ReleasesDGX agent

arXiv:2603.07551v2 Announce Type: replace-cross Abstract: Zero-shot Text-to-Speech (TTS) voice cloning poses severe privacy risks, demanding the removal of specific speaker identities from trained TTS

TaxoBell: Gaussian Box Embeddings for Self-Supervised Taxonomy Expansion

Model ReleasesDGX agent

arXiv:2601.09633v2 Announce Type: replace Abstract: Taxonomies form the backbone of structured knowledge representation across diverse domains, enabling applications such as e-commerce and semantic se

TeachObs: A Human-Validated Benchmark for Multimodal Teaching Observation and Model Evaluation

Model ReleasesDGX agent

arXiv:2605.30673v1 Announce Type: new Abstract: Classroom videos contain observable teaching practices, but their pedagogical and visual signals are rarely organized in forms suitable for model evalua

The fully-managed Remote MCP Server for AlloyDB is now Generally Available

Model ReleasesDGX agent

AI agents possess incredible reasoning capabilities and can perform increasingly complex actions. But the reliability of agentic outcomes depends entirely on the quality of the context they can access

The Geometry of Activity Cliffs: Representation Dependence and Multi-Scale Characterization of Activity Landscapes

Model ReleasesDGX agent

arXiv:2605.30831v1 Announce Type: cross Abstract: Activity cliffs, structurally similar compounds with large potency differences, are widely treated as intrinsic features of chemical datasets. We argu

The Illusion of Generalization in Tabular Language Models

Model ReleasesDGX agent

arXiv:2602.04031v2 Announce Type: replace Abstract: Tabular Language Models (TLMs) have been claimed to achieve strong generalization for tabular prediction. We conduct a systematic re-evaluation of T

The Regularizing Power of Language-Training Deepfake Detectors

Model ReleasesDGX agent

arXiv:2605.31192v1 Announce Type: new Abstract: Recently, thanks to the advent of Multimodal-LLMs, deepfake detectors are striving not only to be generalizable but also interpretable. We propose that

The Surface You Test Is Not the Surface That Breaks

Model ReleasesDGX agent

arXiv:2605.30454v1 Announce Type: cross Abstract: Tool-augmented LLM agents are vulnerable to prompt injection: a third party who controls part of the agent's context can plant instructions that the a

← Previous
1…189190191192193…377
Next →