AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent
84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,585 results
20 May 2026

Distance-Aware Muon: Adaptive Step Scaling for Normalized Optimization

Model ReleasesDGX agent

arXiv:2605.18999v1 Announce Type: new Abstract: Muon and related normalized optimizers decouple the choice of update direction from the choice of step scale, but their practical performance remains se

Distilling Linearized Behavior for Effective Task Arithmetic

Model ReleasesDGX agent

arXiv:2605.18993v1 Announce Type: cross Abstract: Task vector composition has emerged as a promising paradigm for editing pre-trained models, enabling model merging through addition and unlearning thr

Distribution-Free Uncertainty Quantification for Continuous AI Agent Evaluation

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.19779v1 Announce Type: new Abstract: We adapt split conformal prediction and adaptive conformal inference (ACI) to continuous AI agent evaluation, providing distribution-free coverage guara

Distributional Energy-Based Models for Uncertainty-Aware Structured LLM Reasoning

Model ReleasesDGX agent

arXiv:2605.18871v1 Announce Type: cross Abstract: When Large Language Models produce structured outputs such as travel plans, code solutions, or multi-step proofs, individual reasoning steps may appea

Distributionally Robust Control via Stein Variational Inference for Contact-Rich Manipulation

Model ReleasesDGX agent

arXiv:2605.19029v1 Announce Type: new Abstract: Reliable robotic manipulation requires control policies that can accurately represent and adapt to uncertainty arising from contact-rich interactions. M

DLEBench: Evaluating Small-scale Object Editing Ability for Instruction-based Image Editing Model

Model ReleasesDGX agent

arXiv:2602.23622v2 Announce Type: replace-cross Abstract: Significant progress has been made in the field of Instruction-based Image Editing Models (IIEMs). However, while these models demonstrate pla

DMN: A Compositional Framework for Jailbreaking Multimodal LLMs with Multi-Image Inputs

Model ReleasesDGX agent

arXiv:2605.18915v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) are vulnerable to jailbreak attacks, which can elicit harmful responses from MLLMs. Many MLLMs support multi-

DocQT: Improving Document Forgery Localization Robustness via Diverse JPEG Quantization Tables

Model ReleasesDGX agent

arXiv:2605.19688v1 Announce Type: new Abstract: Document manipulation localization models achieve strong performance on public benchmarks yet fail to generalize to operational document workflows. We i

Does Code Cleanliness Affect Coding Agents? A Controlled Minimal-Pair Study

Model ReleasesDGX agent

arXiv:2605.20049v1 Announce Type: cross Abstract: As autonomous coding agents see rapid adoption, their evaluation has primarily focused on task completion rates holding the target codebase fixed. Thi

Dual-Prompt CLIP with Hybrid Visual Encoders for Occluded Person Re-Identification

Model ReleasesDGX agent

arXiv:2605.19527v1 Announce Type: new Abstract: Occluded person re-identification focuses on matching partially visible pedestrians across multiple camera views. However, occlusions disrupt body-regio

DualView: Adaptive Local-Global Fusion for Multi-Hop Document Reranking

Model ReleasesDGX agent

arXiv:2605.18767v1 Announce Type: cross Abstract: Multi-hop question answering requires aggregating information from multiple documents, a critical capability for knowledge-intensive applications. A f

Dynamic Model Merging Made Slim

Model ReleasesDGX agent

arXiv:2605.18904v1 Announce Type: cross Abstract: Model merging enables the reuse of fine-tuned models without joint training or access to original data. Dynamic merging further improves flexibility b

DynaTrain: Fast Online Parallelism Switching for Elastic LLM Training

Model ReleasesDGX agent

arXiv:2605.18815v1 Announce Type: new Abstract: Modern large language model (LLM) training is inherently dynamic: resource fluctuations, RLHF phase shifts, and cluster elasticity continually reshape t

EgoBabyVLM: Benchmarking Cross-Modal Learning from Naturalistic Egocentric Video Data

Model ReleasesDGX agent

arXiv:2605.19130v1 Announce Type: cross Abstract: Children acquire language grounding with remarkable robustness from limited visuo-linguistic input in ways that surpass today's best large multimodal

EgoCoT-Bench: Benchmarking Grounded and Verifiable Operation-Centric Chain of Thought Reasoning for MLLMs

Model ReleasesDGX agent

arXiv:2605.19559v1 Announce Type: cross Abstract: The rapid development of Multimodal Large Language Models (MLLMs) has led to growing interest in egocentric video understanding, specifically the abil

EgoTraj: Real-World Egocentric Human Trajectory Dataset for Multimodal Prediction

Model ReleasesDGX agent

arXiv:2605.19004v1 Announce Type: new Abstract: Accurately forecasting human trajectories from an egocentric perspective plays a central role in applications such as humanoid robotics, wearable sensin

Eliminating Inductive Bias in Reward Models with Information-Theoretic Guidance

Model ReleasesDGX agent

arXiv:2512.23461v2 Announce Type: replace-cross Abstract: Reward models (RMs) are essential in reinforcement learning from human feedback (RLHF) to align large language models (LLMs) with human values

Embedding by Elicitation: Dynamic Representations for Bayesian Optimization of System Prompts

Model ReleasesDGX agent

arXiv:2605.19093v1 Announce Type: new Abstract: System prompts are a central control mechanism in modern AI systems, shaping behavior across conversations, tasks, and user populations. Yet they are di

Emergence of Frontier Superposition: Mobius attractor and Cascade Supervision

Model ReleasesDGX agent

arXiv:2605.18820v1 Announce Type: cross Abstract: Superposition allows Transformers to reason in depth, carrying an entire reasoning frontier in parallel through a bounded-depth forward pass instead o

Enabling Real-Time Colonoscopic Polyp Segmentation on Commodity CPUs via Ultra-Lightweight Architecture

Model ReleasesDGX agent

arXiv:2602.04381v2 Announce Type: replace-cross Abstract: Real-time polyp segmentation is essential for early colorectal cancer detection, yet clinical deployment remains blocked by GPU dependency. We

EngiAI: A Multi-Agent Framework and Benchmark Suite for LLM-Driven Engineering Design

Model ReleasesDGX agent

arXiv:2605.19743v1 Announce Type: new Abstract: Large Language Model (LLM) agents are increasingly applied to engineering design tasks, yet existing evaluation frameworks do not adequately address mul

Entry-level guide to the use of large language models for medical research

Model ReleasesDGX agent

arXiv:2410.18856v4 Announce Type: replace Abstract: Frontier large language models (LLMs), such as GPT-5, Claude 4.5, Gemini 3, Llama 4, and DeepSeek-R1, represent a transformative class of AI tools c

EUPHORIA: Efficient Universal Planning via Hybrid Optimization for Robust Industrial Robotic Assembly

Model ReleasesDGX agent

arXiv:2605.18872v1 Announce Type: cross Abstract: Robotic assembly in architectural construction faces a persistent bottleneck: existing planners are either highly specialized, requiring prohibitive r

EVA-0: Test-Time Model Evolution with Only Two Forward Passes per Sample

Model ReleasesDGX agent

arXiv:2605.18867v1 Announce Type: cross Abstract: Test-time model evolution offers a promising way for deployed models to improve from unlabeled test-time experience, yet most existing methods depend

Evaluating the Utility of Personal Health Records in Personalized Health AI

Model ReleasesDGX agent

arXiv:2605.18937v1 Announce Type: new Abstract: Patient-managed Personal Health Records (PHRs) promises to empower patients to better understand their health; but information in the record is complex,

EventPrune: Cascaded Event-Assisted Token Pruning for Efficient First-Person Dynamic Spatial Reasoning

Model ReleasesDGX agent

arXiv:2605.19506v1 Announce Type: new Abstract: First-person dynamic spatial reasoning requires models to track continuous motion and precise geometric structure, but the quadratic attention cost of T

EviTrack: Selection over Sampling for Delayed Disambiguation

Model ReleasesDGX agent

arXiv:2605.19283v1 Announce Type: cross Abstract: Sequential prediction is challenging in regimes of delayed disambiguation, where early observations are ambiguous and multiple latent explanations rem

Explainable Wastewater Digital Twins: Adaptive Context-Conditioned Structured Simulators with Self-Falsifying Decision Support

Model ReleasesDGX agent

arXiv:2605.19826v1 Announce Type: new Abstract: Operators of safety-critical industrial processes increasingly rely on digital twins to screen control interventions, but such simulators rarely carry c

Exposing Functional Fusion: A New Class of Strategic Backdoor in Dynamic Prompt Architectures

Model ReleasesDGX agent

arXiv:2605.19478v1 Announce Type: cross Abstract: Existing ViT backdoor attacks based on backbone-overwriting full-tuning are computationally expensive and inflict performance degradation. This has fo

Fast and Featureless Node Representation Learning with Partial Pairwise Supervision

Model ReleasesDGX agent

arXiv:2605.19916v1 Announce Type: cross Abstract: We introduce Contrastive FUSE, a fast and unified framework for scalable node representation learning in graphs with partially available pairwise node

Fast and Lightweight Backdoor Detection via Head Random Probing

Model ReleasesDGX agent

arXiv:2605.18908v1 Announce Type: cross Abstract: Deep neural networks (DNNs) remain critically vulnerable to backdoor attacks. Existing post-training detectors often require clean or surrogate data,

Fast-BEV++: Fast by Algorithm, Deployable by Design

Model ReleasesDGX agent

arXiv:2512.08237v3 Announce Type: replace Abstract: The advancement of vision-only Bird's-Eye-View (BEV) perception, a core paradigm for cost-effective autonomous driving, is hindered by the long-stan

Federated Learning for ICD Classification with Lightweight Models and Pretrained Embeddings

Model ReleasesDGX agent

arXiv:2507.03122v2 Announce Type: replace-cross Abstract: This study investigates the feasibility and performance of federated learning (FL) for multi-label ICD code classification using clinical note

FedMental: Evaluating Federated Learning for Mental Health Detection from Social Media Data

Model ReleasesDGX agent

arXiv:2605.18936v1 Announce Type: cross Abstract: Social media text data are often used to train Machine Learning (ML) models to identify users exhibiting high-risk mental health behaviors. However, s

Finding Structure in Continual Learning

Model ReleasesDGX agent

arXiv:2602.04555v2 Announce Type: replace Abstract: Learning from a stream of tasks usually pits plasticity against stability: acquiring new knowledge often causes catastrophic forgetting of past info

Fine-Grained Benchmark Generation for Comprehensive Evaluation of Foundation Models

Model ReleasesDGX agent

arXiv:2605.18824v1 Announce Type: cross Abstract: Evaluation of foundation models often rely on aggregate scores from benchmarks that lack comprehensive coverage and metadata for a fine-grained evalua

Fine-tuned a distilbert with ml-intern today for the first time. The procedure is really straightforward, almost unexpectedly so. - Found a …

Model ReleasesDGX agent

Fine-tuned a distilbert with ml-intern today for the first time. The procedure is really straightforward, almost unexpectedly so. - Found a few datasets relevant to the task (prompt injection detectio

Fine-tuning Large Language Model for Automated Algorithm Design

Model ReleasesDGX agent

arXiv:2507.10614v2 Announce Type: replace-cross Abstract: The integration of large language models (LLMs) into automated algorithm design has shown promising potential. A prevalent approach embeds LLM

FineBench: Benchmarking and Enhancing Vision-Language Models for Fine-grained Human Activity Understanding

Model ReleasesDGX agent

arXiv:2605.19846v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have demonstrated remarkable capabilities in general video understanding, yet they often struggle with the fine-grained

First-Passage Prediction of Grokking Delay: ACalibrated Law under AdamW with Causal Validation

Model ReleasesDGX agent

arXiv:2605.18845v1 Announce Type: cross Abstract: We give the first quantitative prediction of grokking delay under AdamW. Treating the delay as a first-passage time, we derive a closed-form law T_gro

FLUIDSPLAT: Reconstructing Physical Fields from Sparse Sensors via Gaussian Primitives

Model ReleasesDGX agent

arXiv:2605.18866v1 Announce Type: cross Abstract: Reconstructing continuous flow fields from sparse surface-mounted sensors is central to aerodynamic design, flow control, and digital-twin instrumenta

FLUXtrapolation: A benchmark on extrapolating ecosystem fluxes

Model ReleasesDGX agent

arXiv:2605.19812v1 Announce Type: cross Abstract: We introduce FLUXtrapolation, a benchmark for extrapolating ecosystem fluxes under progressively harder distribution shifts. Ecosystem fluxes are cent

for anyone curious, this was the result of many experiments bouncing around but this version uses... initial ideation: chatgpt for ideas/@re…

Model ReleasesDGX agent

for anyone curious, this was the result of many experiments bouncing around but this version uses... initial ideation: chatgpt for ideas/@replit for quick build final repo build: opus 4.7 prompting cl

For centuries, the scientific method has been our best tool for progress. But today, there’s so much data out there that it’s impossible for…

Model ReleasesDGX agent

For centuries, the scientific method has been our best tool for progress. But today, there’s so much data out there that it’s impossible for any one researcher to connect all the dots. We want to fix

Forward launches Predict to verify network changes before they reach production

Model ReleasesDGX agent

Network verification company Forward Inc. today launched Forward Predict, a new capability that lets network teams test proposed changes against a digital twin of their production network before deplo

From Llama to Cria: Scaling Down Neural Networks via Neuron-Level Spectral Structural Importance Evaluation

Model ReleasesDGX agent

arXiv:2605.18860v1 Announce Type: cross Abstract: This paper proposes a neuron pruning framework based on neuron-level spectral structural importance evaluation. Given a trained neural network, we rec

From Prompts to Pavement Through Time: Temporal Grounding in Agentic Scene-to-Plan Reasoning

Model ReleasesDGX agent

arXiv:2605.19824v1 Announce Type: new Abstract: Recent attempts to support high-level scene interpretation and planning in Autonomous Vehicles (AVs) using ensembles of Large Language Models (LLMs) and

From SGD to Muon: Adaptive Optimization via Schatten-p Norms

Model ReleasesDGX agent

arXiv:2605.19781v1 Announce Type: new Abstract: Modern optimizers, like Muon, impose matrix-wise geometry constraints on their updates. These matrix-wise constraints can be unified under Linear Minimi

From Simple to Complex: Curriculum-Guided Physics-Informed Neural Networks via Gaussian Mixture Models

Model ReleasesDGX agent

arXiv:2605.19263v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) offer a mesh-free framework for solving partial differential equations (PDEs), yet training often suffers from

From Sparsity to Simplicity: Enabling Simpler Sequential Replacements via Sparse Attention Distillation

Model ReleasesDGX agent

arXiv:2605.18865v1 Announce Type: cross Abstract: Self-attention serves as the core foundation of large-scale transformer pretraining, but its quadratic token interaction cost makes inference expensiv

fun read on real tradeoffs & design decisions we debated when designing Engine for the data scale that customers produce one common thread i…

Model ReleasesDGX agent

fun read on real tradeoffs & design decisions we debated when designing Engine for the data scale that customers produce one common thread is that we’re pretty strong supporters of just giving the age

Gemini 3.5 announce.

Model ReleasesDGX agent

Google introduced Gemini 3.5, its latest family of models combining frontier intelligence with action capabilities, representing a major leap forward in building more capable, intelligent agents. The

General Lower Bounds for Differentially Private Federated Learning with Arbitrary Public-Transcript Interactions

Model ReleasesDGX agent

arXiv:2605.19813v1 Announce Type: new Abstract: We prove a general lower bound for differentially private federated learning protocols with arbitrary public-transcript interactions. The protocol may u

Generalization Bounds of Surrogate Policies for Combinatorial Optimization Problems

Model ReleasesDGX agent

arXiv:2407.17200v3 Announce Type: replace-cross Abstract: Many real-world decision problems require solving, again and again, combinatorial optimization instances drawn from a common distribution. A r

GeoX: Mastering Geospatial Reasoning Through Self-Play and Verifiable Rewards

Model ReleasesDGX agent

arXiv:2605.20006v1 Announce Type: new Abstract: Geospatial reasoning requires solving image-grounded problems over the complex spatial structure of a scene. However, developing this capability is hind

GoLongRL: Capability-Oriented Long Context Reinforcement Learning with Multitask Alignment

Model ReleasesDGX agent

arXiv:2605.19577v1 Announce Type: new Abstract: We present GoLongRL, a fully open-source, capability-oriented post-training recipe for long-context reinforcement learning with verifiable rewards (RLVR

Google I/O, Gemini Spark, Antigravity

Model ReleasesDGX agent

It's hard to find much to write about Google I/O this year because I have a policy of not writing about anything that I can't try out myself, and a lot of the big announcements are 'coming soon'. I ac

Google says it is testing new ad formats in search results and AI Mode, including Conversational Discovery ads, Highlighted Answers, and AI-powered Shopping ads (Anu Adegbola/Search Engine Land)

Model ReleasesDGX agent

Anu Adegbola / Search Engine Land: Google says it is testing new ad formats in search results and AI Mode, including Conversational Discovery ads, Highlighted Answers, and AI-powered Shopping ads — Go

Google Search’s AI evolution includes more ads

Model ReleasesDGX agent

Google's AI-powered Search era apparently also extends to its ads. Now, when you search for a product, Google's Gemini AI chatbot will surface relevant items and generate a 'custom explainer' about wh

Got to play with a little of this before launch as well. My experience as a social scientist was that it was more bioscience focused right n…

Model ReleasesDGX agent

Got to play with a little of this before launch as well. My experience as a social scientist was that it was more bioscience focused right now, but I think Google has been the leading lab in releasing

← Previous
1…227228229230231…377
Next →