AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent
84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,585 results
20 May 2026

MTraining: Distributed Dynamic Sparse Attention for Efficient Ultra-Long Context Training

Model ReleasesDGX agent

arXiv:2510.18830v2 Announce Type: replace Abstract: The adoption of long context windows has become a standard feature in Large Language Models (LLMs), as extended contexts significantly enhance their

Multi-axis Analysis of Image Manipulation Localization

Model ReleasesDGX agent

arXiv:2605.20174v1 Announce Type: new Abstract: Advanced image editing software enables easy creation of highly convincing image manipulations, which has been made even more accessible in recent years

MVI-Bench: A Comprehensive Benchmark for Evaluating Robustness to Misleading Visual Inputs in LVLMs

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2511.14159v2 Announce Type: replace Abstract: Evaluating the robustness of Large Vision-Language Models (LVLMs) is essential for their continued development and responsible deployment in real-wo

Next-Acceleration-Scale Prediction for Autoregressive MRI Reconstruction

Model ReleasesDGX agent

arXiv:2605.19354v1 Announce Type: cross Abstract: MRI reconstruction is an inherently ill-posed inverse problem, since incomplete measurements admit many plausible solutions. This ambiguity becomes mo

NGL: Natural Garment Language for Training-Free Sewing Pattern Estimation

Model ReleasesDGX agent

arXiv:2602.20700v2 Announce Type: replace Abstract: Estimating sewing patterns from images is a practical approach for creating high-quality 3D garments, but it remains challenging due to the scarcity

No Hard Negatives Required: Concept Centric Learning Leads to Compositionality without Degrading Zero-shot Capabilities of Contrastive Models

Model ReleasesDGX agent

arXiv:2603.25722v2 Announce Type: replace Abstract: Contrastive vision-language (V&L) models remain a popular choice for various applications. However, several limitations have emerged, most notably t

Nonlinearity as Rank: Generative Low-Rank Adapter with Radial Basis Functions

Model ReleasesDGX agent

arXiv:2602.05709v2 Announce Type: replace Abstract: Low-rank adaptation (LoRA) approximates the update of a pretrained weight matrix using the product of two low-rank matrices. However, standard LoRA

Not All Tokens Are Worth Caching: Learning Semantic-Aware Eviction for LLM Prefix Caches

Model ReleasesDGX agent

arXiv:2605.18825v1 Announce Type: new Abstract: Prefix caching is a key optimization in Large Language Model (LLM) serving, reusing attention Key-Value (KV) states across requests with shared prompt p

OmniGUI: Benchmarking GUI Agents in Omni-Modal Smartphone Environments

Model ReleasesDGX agent

arXiv:2605.18758v1 Announce Type: cross Abstract: Current benchmarks for graphical user interface (GUI) agents predominantly rely on static screenshots. However, real-world smartphone interaction rout

On the Provable Suboptimality of Momentum SGD in Nonstationary Stochastic Optimization

Model ReleasesDGX agent

arXiv:2601.12238v4 Announce Type: replace-cross Abstract: In this paper, we provide a comprehensive theoretical analysis of Stochastic Gradient Descent (SGD) and its momentum variants (Polyak Heavy-Ba

OpenCompass: A Universal Evaluation Platform for Large Language Models

Model ReleasesDGX agent

arXiv:2605.19276v1 Announce Type: new Abstract: In recent years, the field of artificial intelligence has undergone a paradigm shift from task-specific small-scale models to general-purpose large lang

Operationalizing Document AI: A Microservice Architecture for OCR and LLM Pipelines in Production

Model ReleasesDGX agent

arXiv:2605.18818v1 Announce Type: new Abstract: Academic research tends to focus on new models for document understanding creating a wide gap in the literature between model definition and running mod

Optimal Reconstruction from Linear Queries

Model ReleasesDGX agent

arXiv:2605.19625v1 Announce Type: new Abstract: We study the problem of reconstructing an unknown point in R^d from approximate linear queries. This setting arises naturally in applications ranging fr

optimize_anything: A Universal API for Optimizing any Text Parameter

Model ReleasesDGX agent

arXiv:2605.19633v1 Announce Type: cross Abstract: Can a single LLM-based optimization system match specialized tools across fundamentally different domains? We show that when optimization problems are

P2DNav: Panorama-to-Downview Reasoning for Zero-shot Vision-and-Language Navigation

Model ReleasesDGX agent

arXiv:2605.19634v1 Announce Type: cross Abstract: Vision-and-language navigation (VLN) requires an embodied agent to ground natural-language instructions into executable navigation actions in unseen e

Passive Construction Site Safety Monitoring via Persona-Scaffolded Adversarial Chain-of-Thought VLM Verification

Model ReleasesDGX agent

arXiv:2605.19869v1 Announce Type: cross Abstract: Construction remains the deadliest industry sector in the United States, with 1,055 fatal worker injuries recorded in 2023, and the majority preventab

PEPL: Precision-Enhanced Pseudo-Labeling for Fine-Grained Image Classification in Semi-Supervised Learning

Model ReleasesDGX agent

arXiv:2409.03192v2 Announce Type: replace Abstract: Fine-grained image classification has witnessed significant advancements with the advent of deep learning and computer vision technologies. However,

✨ Personal AI is the next computing platform. AI is shifting from something you access to something you build with, locally, at the edge, an…

Model ReleasesDGX agent

✨ Personal AI is the next computing platform. AI is shifting from something you access to something you build with, locally, at the edge, and across systems. We’re unlocking new possibilities for deve

Physics-in-the-Loop: A Hybrid Agentic Architecture for Validated CAD Engineering Design

Model ReleasesDGX agent

arXiv:2605.19717v1 Announce Type: new Abstract: Large Language Models (LLMs) can generate Computer-Aided Design (CAD), yet lack physical comprehension required for reliable engineering design. Instead

Physics-Informed Graph Neural Network Surrogates for Turbulent Nanoparticle Dispersion in Dental Clinical Environments

Model ReleasesDGX agent

arXiv:2605.19589v1 Announce Type: new Abstract: Dental aerosol procedures produce sub-50 micrometre nuclei that can remain airborne for long periods in enclosed clinics, creating pathways for airborne

PhyWorld: Physics-Faithful World Model for Video Generation

Model ReleasesDGX agent

arXiv:2605.19242v1 Announce Type: cross Abstract: World simulators can provide safe and scalable environments for training Physical AI systems before real-world deployment. Large video generation mode

PixVerve: Advancing Native UHR Image Generation to 100MP with a Large-Scale High-Quality Dataset

Model ReleasesDGX agent

arXiv:2605.20147v1 Announce Type: new Abstract: Text-to-Image (T2I) models have recently seen notable progress around 1K and 2K resolution. With the extreme desire for better visual experience and the

PlantTraitNet: An Uncertainty-Aware Multimodal Framework for Global-Scale Plant Trait Inference from Citizen Science Data

Model ReleasesDGX agent

arXiv:2511.06943v3 Announce Type: replace-cross Abstract: Global plant maps of plant traits, such as leaf nitrogen or plant height, are essential for understanding ecosystem processes, including the c

POLAR-Bench: A Diagnostic Benchmark for Privacy-Utility Trade-offs in LLM Agents

Model ReleasesDGX agent

arXiv:2605.19127v1 Announce Type: new Abstract: LLM agents increasingly have access to private user data and act on the user's behalf when interacting with third-party systems. The user defines what m

Position: The Turing-Completeness of Real-World Autoregressive Transformers Relies Heavily on Context Management

Model ReleasesDGX agent

arXiv:2605.19514v1 Announce Type: new Abstract: Many works make the eye-catching claim that Transformers are Turing-complete. However, the literature often conflates two distinct settings: (i) a fixed

PrAda: Few-Shot Visual Adaptation for Text-Prompted Segmentation

Model ReleasesDGX agent

arXiv:2605.19623v1 Announce Type: new Abstract: Segmenting images is critical for visual understanding but demands extensive pixel-level annotations. Foundational models have enabled new paradigms for

Precision Tracked Transformer via Kalman Filtering, Kriging and Process Noise

Model ReleasesDGX agent

arXiv:2605.18832v1 Announce Type: cross Abstract: The Transformer is the foundational building block of modern AI, yet offers no principled handling of uncertainty, which is prevalent in real applicat

Preferences Order, Ratings Anchor: From Fused Expert Aesthetic Ground Truth to Self-Distillation

Model ReleasesDGX agent

arXiv:2605.19776v1 Announce Type: new Abstract: Pairwise preferences and pointwise ratings are the two dominant annotation protocols in image aesthetic assessment (IAA), yet existing benchmarks adopt

PRISM: A Benchmark for Programmatic Spatial-Temporal Reasoning

Model ReleasesDGX agent

arXiv:2605.19382v1 Announce Type: new Abstract: Programmatic video generation through code offers geometric precision and temporal coherence beyond pixel-level diffusion models, yet rigorously evaluat

Program Evaluation with Remotely Sensed Outcomes

Model ReleasesDGX agent

arXiv:2411.10959v4 Announce Type: replace-cross Abstract: We study causal inference in experiments and quasi-experiments, where the economic outcome is imperfectly measured by a remotely sensed variab

ProJo4D: Progressive Joint Optimization for Sparse-View Inverse Physics Estimation

Model ReleasesDGX agent

arXiv:2506.05317v3 Announce Type: replace Abstract: Neural rendering has advanced significantly in 3D reconstruction and novel view synthesis, and integrating physics into these frameworks opens new a

Prompting language influences diagnostic reasoning and accuracy of large language models

Model ReleasesDGX agent

arXiv:2605.19173v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly explored for clinical decision support, yet most evaluations are conducted in English, leaving their relia

PromptRad: Knowledge-Enhanced Multi-Label Prompt-Tuning for Low-Resource Radiology Report Labeling

Model ReleasesDGX agent

arXiv:2605.20052v1 Announce Type: cross Abstract: Automatic report labeling facilitates the identification of clinical findings from unstructured text and enables large-scale annotation for medical im

Protein Autoregressive Modeling via Multiscale Structure Generation

Model ReleasesDGX agent

arXiv:2602.04883v2 Announce Type: replace-cross Abstract: We present protein autoregressive modeling (PAR), the first multi-scale autoregressive framework for protein backbone generation via coarse-to

Provable Fairness Repair for Deep Neural Networks

Model ReleasesDGX agent

arXiv:2605.19549v1 Announce Type: cross Abstract: Deep neural networks (DNNs) are suffering from ethical issues such as individual discrimination. In response, extensive NN repair techniques have been

Quantifying the Generalization Gap in Seizure Detection: A Large-Scale Empirical Benchmark via the SzCORE Challenge

Model ReleasesDGX agent

arXiv:2505.18191v2 Announce Type: replace-cross Abstract: Reliable automatic seizure detection from long-term electroencephalography (EEG) remains an unsolved challenge, as current models often fail t

Quantum Machine Learning for Cyber-Physical Anomaly Detection in Unmanned Aerial Vehicles: A Leakage-Free Evaluation with Proxy-Audited Feature Sets

Model ReleasesDGX agent

arXiv:2605.19233v1 Announce Type: cross Abstract: Unmanned aerial vehicles (UAVs) are cyber-physical systems whose attack surface spans networked avionics and on-board sensor fusion: a compromised GPS

RE-VLM: Event-Augmented Vision-Language Model for Scene Understanding

Model ReleasesDGX agent

arXiv:2605.19329v1 Announce Type: cross Abstract: Conventional vision-language models (VLMs) struggle to interpret scenes captured under adverse conditions (e.g., low light, high dynamic range, or fas

ReacTOD: Bounded Neuro-Symbolic Agentic NLU for Zero-Shot Dialogue State Tracking

Model ReleasesDGX agent

arXiv:2605.19077v1 Announce Type: cross Abstract: Task-oriented dialogue systems -- handling transactions, reservations, and service requests -- require predictable behavior, yet the moderately-sized

Real-World On-Vehicle Evaluation of Embedding-Based Anomaly Detection

Model ReleasesDGX agent

arXiv:2605.19744v1 Announce Type: new Abstract: Detecting anomalies in traffic scenes is crucial for ensuring safety in autonomous driving, yet collecting representative anomalous data remains challen

Rebalancing Reference Frame Dominance to Improve Motion in Image-to-Video Models

Model ReleasesDGX agent

arXiv:2605.19398v1 Announce Type: cross Abstract: Image-to-video models often generate videos that remain overly static, compared to text-to-video models. While prior approaches mitigate this issue by

RECIPE: Procedural Planning via Grounding in Instructional Video

Model ReleasesDGX agent

arXiv:2605.19976v1 Announce Type: new Abstract: Visual planning asks a model to generate the remaining steps of a procedure in natural language given a partial video context and a goal. Progress on th

RecoAtlas: From Semantic Plausibility to Set-Level Utility in LLM Recommendation Agents

Model ReleasesDGX agent

arXiv:2605.18805v1 Announce Type: cross Abstract: LLM recommendation agents increasingly produce structured recommendation reports: sets of items accompanied by natural-language justifications. Yet ex

Recursive Entropic Risk Optimization in Discounted MDPs: Sample Complexity Bounds with a Generative Model

Model ReleasesDGX agent

arXiv:2506.00286v3 Announce Type: replace-cross Abstract: We study risk-sensitive reinforcement learning in finite discounted MDPs with recursive entropic risk measures (ERM), where the risk parameter

Replacement Learning: Training Neural Networks with Fewer Parameters

Model ReleasesDGX agent

arXiv:2605.19533v1 Announce Type: new Abstract: End-to-end training with full-depth backpropagation remains the dominant paradigm for optimizing deep neural networks, but its efficiency deteriorates a

Reporting from Google I/O 2026 with the four biggest themes from one of the biggest AI labs in the world. 🎤 Voice AI as an interface Google…

Model ReleasesDGX agent

Reporting from Google I/O 2026 with the four biggest themes from one of the biggest AI labs in the world. 🎤 Voice AI as an interface Google and Samsung announced new Gemini-powered glasses with Gentle

Resilient Byzantine Agreement with Predictions

Model ReleasesDGX agent

arXiv:2605.19452v1 Announce Type: cross Abstract: This paper studies the Byzantine Agreement problem where the nodes have access to a predictor that flags nodes for suspicion of faulty (Byzantine) beh

Rethinking How to Remember: Beyond Atomic Facts in Lifelong LLM Agent Memory

Model ReleasesDGX agent

arXiv:2605.19952v1 Announce Type: new Abstract: To enable reliable long-term interaction, LLM agents require a memory system that can faithfully store, efficiently retrieve, and deeply reason over acc

Retrieval-Augmented Generation for Natural Language Processing: A Survey

Model ReleasesDGX agent

arXiv:2407.13193v4 Announce Type: replace Abstract: Large language models (LLMs) have achieved strong empirical performance in various fields, benefiting from their huge amount of parameters that stor

Rewriting History: A Recipe for Interventional Analyses to Study Data Effects on Model Behavior

Model ReleasesDGX agent

arXiv:2510.14261v2 Announce Type: replace Abstract: We present an experimental recipe for studying the relationship between training data and language model (LM) behavior. We outline steps for interve

RoboJailBench: Benchmarking Adversarial Attacks and Defenses in Embodied Robotic Agents

Model ReleasesDGX agent

arXiv:2605.19328v1 Announce Type: cross Abstract: Recent advances in Vision-Language Models (VLMs) facilitate a new class of embodied AI systems, where these models are integrated into physical platfo

Robots that learn to evaluate models of collective behavior

Model ReleasesDGX agent

arXiv:2604.07303v2 Announce Type: replace Abstract: Understanding and modeling animal behavior is essential for studying collective motion, decision-making, and bio-inspired robotics. Yet, evaluating

Robust Basis Spline Decoupling for the Compression of Transformer Models

Model ReleasesDGX agent

arXiv:2605.18794v1 Announce Type: cross Abstract: Decoupling is a powerful modeling paradigm for representing multivariate functions as compositions of linear transformations and univariate nonlinear

SAGA: A Sequence-Adaptive Generative Architecture for Multi-Horizon Probabilistic Forecasting with Adaptive Temporal Conformal Prediction

Model ReleasesDGX agent

arXiv:2605.19014v1 Announce Type: new Abstract: Microsimulation models used by ministries of finance and central banks rely on parametric processes for lifetime earnings that capture only first and se

ScheduleFree+: Scaling Learning-Rate-Free & Schedule-Free Learning to Large Language Models

Model ReleasesDGX agent

arXiv:2605.19095v1 Announce Type: cross Abstract: Schedule-Free Learning has shown promise as a practical anytime training method for machine learning, showing success across dozens of standard benchm

SciCustom: A Framework for Custom Evaluation of Scientific Capabilities in Large Language Models

Model ReleasesDGX agent

arXiv:2605.19357v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly applied to scientific research, yet existing evaluations often fail to reflect the fine-grained capabiliti

Search Self-play: Pushing the Frontier of Agent Capability without Supervision

Model ReleasesDGX agent

arXiv:2510.18821v3 Announce Type: replace Abstract: Reinforcement learning with verifiable rewards (RLVR) has become the mainstream technique for training LLM agents. However, RLVR highly depends on w

Self-Filtered Distillation with LLMs-generated Trust Indicators for Reliable Patent Classification

Model ReleasesDGX agent

arXiv:2510.05431v4 Announce Type: replace Abstract: Organizing large-scale patent corpora according to classification schemes is a core information management task that determines the accuracy and eff

Self-improving AI is a big deal! As a first step, I've been exploring how much of the post-training can be automated. Here is a first post o…

Model ReleasesDGX agent

Self-improving AI is a big deal! As a first step, I've been exploring how much of the post-training can be automated. Here is a first post on how I am using @FireworksAI_HQ Agent to automate LLM fine-

Semantic-Enriched Latent Visual Reasoning

Model ReleasesDGX agent

arXiv:2605.19342v1 Announce Type: new Abstract: Multimodal latent-space reasoning aims to replace explicit thinking with images by performing visual reasoning directly in a compact latent space. Howev

← Previous
1…229230231232233…377
Next →