AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,376
  • Agents7,554
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,170
  • Local Ai4,930
  • Model Releases23,883
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,376
  • Agents7,554
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,170
  • Local Ai4,930
  • Model Releases23,883
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

88,376Total entries
1Added by human
88,375Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,602 results
26 May 2026

OASIS: Observation-Action Space Alignment via SE(3) Trajectory Prediction for Robotic Manipulation

SafetyDGX agent

arXiv:2605.25829v1 Announce Type: cross Abstract: Recent vision-language-action (VLA) models and world action models (WAMs) advance robotic manipulation by enriching intermediate representations with

OSDTW: Optimal Shared Depth and Task Weighting for Long-Tailed Recognition

Model ReleasesDGX agent

arXiv:2605.24969v1 Announce Type: cross Abstract: Long-tailed recognition suffers from a persistent head--tail trade-off: improving tail performance often degrades head accuracy and can increase train

Parameter Efficient Multi-Class Intelligent Scheduling for Multimodal Online Distributed Industrial Anomaly Detection

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.23984v1 Announce Type: cross Abstract: Industrial anomaly detection has attracted significant attention as a fundamental challenge in industrial systems. The rapid advancement of heterogene

PDEInvBench: A Comprehensive Dataset and Design Space Exploration of Neural Networks for PDE Inverse Problems

Model ReleasesDGX agent

arXiv:2605.25353v1 Announce Type: new Abstract: Inverse problems in partial differential equations (PDEs) involve estimating the physical parameters of a system from observed spatiotemporal solution f

Physen-Noise2Noise: Physics-Guided Self-Supervised Defocus Deblurring with Bias Correction under Low-Light Conditions

Model ReleasesDGX agent

arXiv:2605.24590v1 Announce Type: cross Abstract: Low-light, long-exposure defocus deblurring remains a challenging problem due to the simultaneous presence of severe blur and complex biased noise. Ex

Pragmatic Reasoning improves LLM Code Generation

Model ReleasesDGX agent

arXiv:2502.15835v5 Announce Type: replace-cross Abstract: Pragmatic reasoning helps interlocutors infer intended meaning from ambiguous or underspecified messages by considering shared context and cou

QUEST: Training Frontier Deep Research Agents with Fully Synthetic Tasks

Model ReleasesDGX agent

arXiv:2605.24218v1 Announce Type: new Abstract: Deep research agents extend the role of search engines from retrieving keyword-matched pages to synthesizing knowledge, fundamentally changing how human

Reading, Not Thinking: Understanding and Bridging the Modality Gap When Text Becomes Pixels in Multimodal LLMs

SafetyDGX agent

arXiv:2603.09095v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) can process text presented as images, yet they often perform worse than when the same content is provided a

Residual Drift Dominates Contradiction in Multi-Turn Constraint Reasoning

Model ReleasesDGX agent

arXiv:2605.23940v1 Announce Type: new Abstract: How do multi-turn reasoning systems fail? The expected answer is logical contradiction, in which the system's maintained state becomes unsatisfiable. We

RoboManipBaselines: A Unified Framework for Imitation Learning in Robotic Manipulation across Real and Simulation Environments

Model ReleasesDGX agent

arXiv:2509.17057v3 Announce Type: replace Abstract: We present RoboManipBaselines, an open-source software framework for imitation learning research in robotic manipulation. The framework supports the

Robust Fuzzy Multi-view Learning under View Conflict

Model ReleasesDGX agent

arXiv:2605.24475v1 Announce Type: cross Abstract: Trusted multi-view classification aims to deliver reliable fusion for accurate predictions and has recently attracted substantial attention in both ac

Spectral Probe-Circuits: A Three-Step Recipe for Identifying Attention-Head Circuits in Pretrained Transformers

Model ReleasesDGX agent

arXiv:2605.24059v1 Announce Type: cross Abstract: We present a three-step recipe for identifying attention-head circuits in pretrained transformers. A per-head spectral signal -- the time-integrated p

TapSampling: Inference-Time Sampling with a Task-Progress-Understanding Verifier for Robotic Manipulation

SafetyDGX agent

arXiv:2605.25547v1 Announce Type: new Abstract: Existing embodied control research demonstrates remarkable performance improvements by scaling training data and model size. We instead explore inferenc

The Concept Allocation Zone: Tracking How Concepts Form Across Transformer Depth

SafetyDGX agent

arXiv:2605.24856v1 Announce Type: cross Abstract: Concept formation in transformer language models is depth-extended, not a single-layer event: concepts emerge gradually across a contiguous region of

Uni-DPO: A Unified Paradigm for Dynamic Preference Optimization of LLMs

Model ReleasesDGX agent

arXiv:2506.10054v4 Announce Type: replace-cross Abstract: Direct Preference Optimization (DPO) has emerged as a cornerstone of reinforcement learning from human feedback (RLHF) due to its simplicity a

v0.13.0 of Exo just dropped, and it's one I've been looking forward to for a while. tl;dr swapping out anthropic for @ollama cloud dropped m…

Model ReleasesDGX agent

v0.13.0 of Exo just dropped, and it's one I've been looking forward to for a while. tl;dr swapping out anthropic for @ollama cloud dropped my usage costs from 30/day to just 20 per MONTH with no notic

VectorArk: Learning Practical Image Vectorization with Rounded Polygon Representation

ApplicationsDGX agent

arXiv:2605.24398v1 Announce Type: cross Abstract: Recent vision-language model (VLM)-based approaches have achieved impressive results on image vectorization tasks. However, they are typically evaluat

WhoSaidIt: Human-LLM Collaborative Annotation for Text-Based Multilingual Speaker-Attribute Classification

Model ReleasesDGX agent

arXiv:2605.26070v1 Announce Type: new Abstract: Annotating speaker attributes from text is inherently ambiguous, particularly in multilingual settings where demographic and social cues are implicit an

25 May 2026

AI Evaluation Should Require Standardized Item-Level Data Releases

Model ReleasesDGX agent

arXiv:2604.03244v2 Announce Type: replace Abstract: This position paper argues that standardized item-level benchmark data should become the default infrastructure for AI evaluation. Current evaluatio

ALIVE: Awakening LLM Reasoning via Adversarial Learning and Instructive Verbal Evaluation

SafetyDGX agent

arXiv:2602.05472v2 Announce Type: replace Abstract: The quest for expert-level reasoning in Large Language Models (LLMs) has been hampered by a persistent extit{reward bottleneck}: traditional reinfor

Any-Dimensional Invariant Universality

ResearchDGX agent

arXiv:2605.23156v1 Announce Type: new Abstract: Several machine learning models are defined for inputs of any size, such as graphs with different numbers of nodes and point clouds containing varying n

BarrierSteer: LLM Safety via Learning Barrier Steering

SafetyDGX agent

arXiv:2602.20102v2 Announce Type: replace-cross Abstract: Despite the strong performance of large language models (LLMs) across diverse tasks, their susceptibility to adversarial attacks and unsafe co

Contrastive Distribution Matching for Amortized Sequential Monte Carlo in Discrete Diffusion

SafetyDGX agent

arXiv:2605.23346v1 Announce Type: new Abstract: Discrete diffusion models have emerged as powerful frameworks for generating structured categorical data. However, efficiently sampling from reward-tilt

DDX-TRACE: A Benchmark for Medical Diagnostic Trajectories in VLMs

Model ReleasesDGX agent

arXiv:2605.23629v1 Announce Type: new Abstract: Medical diagnosis is not a single prediction from a fully specified vignette. It is a sequential workup: clinicians decide what evidence to obtain, revi

Decoupling Spatio-Temporal Adapter for Fine-Grained Badminton Action Localization

Model ReleasesDGX agent

arXiv:2605.23355v1 Announce Type: new Abstract: Temporal Action Localization (TAL) has been extensively studied in generic video understanding, while fine-grained sports scenarios, such as professiona

Enhancing Blood Cells Classification using Hybrid Quantum Neural Networks

ResearchDGX agent

arXiv:2605.23324v1 Announce Type: new Abstract: Accurate classification of microscopic blood cells is still a critical task in medical image analysis, where subtle variations and limited data can chal

It’ll be interesting to see if the post training for this uses a multiple of the compute of pretraining as cursor did when they tuned Kimi a…

IndustryDGX agent

It’ll be interesting to see if the post training for this uses a multiple of the compute of pretraining as cursor did when they tuned Kimi as the base model Grok foundation model V9-Medium (1.5T) has

Lipschitz Optimization for Formal Verification of Homographies

Model ReleasesDGX agent

arXiv:2605.23203v1 Announce Type: cross Abstract: The adoption of vision neural networks in regulated industries requires formal robustness guarantees, especially in safety-critical domains such as he

MedExpMem: Adapting Experience Memory for Differential Diagnosis

Model ReleasesDGX agent

arXiv:2605.22872v1 Announce Type: cross Abstract: Experienced physicians develop diagnostic expertise through clinical practice, acquiring not only disease knowledge but also the ability to differenti

Pointwise Metrics Mislead: An Evaluation Protocol for Multimodal Inverse Problems

Model ReleasesDGX agent

arXiv:2605.22891v1 Announce Type: new Abstract: Evaluation in scientific reconstruction is dominated by pointwise metrics - RMSE, MAE, per-event resolution - under the implicit assumption that lower e

Reinforcement Learning for Microcanonical Graph Ensemble with Assortativity Constraints

Model ReleasesDGX agent

arXiv:2605.23285v1 Announce Type: cross Abstract: How network structure determines function is a fundamental question, and it can be investigated by graph ensembles with precisely controlled structura

Robust LLM Watermarking with Minimal Semantic Distortion for IP Protection

ResearchDGX agent

arXiv:2605.23175v1 Announce Type: cross Abstract: Proprietary large language models (LLMs) face risks of intellectual property (IP) violation, as adversaries can replicate an LLM by collecting input-o

Semantically Structured Mixture-of-Experts for Compositional Robotic Manipulation

Model ReleasesDGX agent

arXiv:2605.23477v1 Announce Type: new Abstract: Diffusion-based policies have established a new standard for precise robotic manipulation but face a critical scalability bottleneck: high-performance m

SemEval-2026 Task 6: CLARITY -- Unmasking Political Question Evasions

Model ReleasesDGX agent

arXiv:2603.14027v2 Announce Type: replace Abstract: Political speakers often avoid answering questions directly while maintaining the appearance of responsiveness. Despite its importance for public di

STAMBRIDGE: Spectral-Temporal Amplitude-aware Mid-Feature Bridge for EEG Visual Decoding

Model ReleasesDGX agent

arXiv:2605.23137v1 Announce Type: cross Abstract: Electroencephalography (EEG) visual decoding remains challenging due to the modality gap between low-SNR neural signals and highly structured vision--

Targeted Regularization for Causal Effect Estimation with Exponential Dispersion Family Outcomes

Model ReleasesDGX agent

arXiv:2502.07295v2 Announce Type: replace Abstract: Neural Networks (NNs) for causal effect estimation have shown strong empirical performance, yet endowing them with desirable semiparametric properti

The TIME Machine: On The Power of Motion for Efficient Perception

TutorialsDGX agent

arXiv:2605.23045v1 Announce Type: cross Abstract: Video representation learning has seen tremendous progress in recent years. This has been driven by many factors, including the scale of training and

Using Ensemble Diffusion to Estimate Uncertainty for End-to-End Autonomous Driving

Model ReleasesDGX agent

arXiv:2506.00560v2 Announce Type: replace-cross Abstract: End-to-end planning systems for autonomous driving are rapidly improving, especially in closed-loop simulation environments like CARLA. Many s

VideoTemp-o3: Harmonizing Temporal Grounding and Video Understanding in Agentic Thinking-with-Videos

Model ReleasesDGX agent

arXiv:2602.07801v4 Announce Type: replace-cross Abstract: In long-video understanding, conventional uniform frame sampling often fails to capture key visual evidence, leading to degraded performance a

VINS-120K: Ultra High-Resolution Image Editing with A Large-Scale Dataset

Model ReleasesDGX agent

arXiv:2605.23518v1 Announce Type: new Abstract: Directly editing ultra-high-resolution (UHR) images is valuable but underexplored, primarily due to the lack of high-quality data and the challenge in m

When Is Next-Token Prediction Useful? Marginalization, Ergodicity, Mixture Identifiability, Local Sufficiency, RAG, Tools, and Programming

AgentsDGX agent

arXiv:2605.23278v1 Announce Type: new Abstract: Language models trained on observed sequences are often described as learning the conditional distribution of the next token given previous tokens. This

24 May 2026

Dual 3090s?

Local AiDGX agent

A discussion about difficulties getting Ollama to fully utilize dual NVIDIA RTX 3090 GPUs when running various language models including 8B and 70B parameter models. Users report that despite Ollama r

23 May 2026

A Boundary-Layer Mechanism for One-Third Scaling in Online Softmax Classification

Model ReleasesDGX agent

arXiv:2605.22341v1 Announce Type: new Abstract: Hard-label classification is usually trained with smooth surrogate losses, most prominently softmax cross-entropy. We isolate an asymptotic mechanism by

Alike Parts: A Feature-Informed Approach to Local and Global Prototype Explanations

Model ReleasesDGX agent

arXiv:2605.21646v1 Announce Type: new Abstract: Prototype-based explanations offer an intuitive, example-based approach to support the interpretability of machine learning black box classifiers but of

AsymFLUX.2-klein-9B is all about textures

Local AiDGX agent

AsymFLUX.2-klein-9B is a pixel-space text-to-image model finetuned from FLUX.2-klein-base-9B using the AsymFlow method. The model is noted for producing sharp textures with strong adherence and compos

Asymmetric Virtual Memory Paging for Hybrid Mamba-Transformer Inference

ResearchDGX agent

arXiv:2605.22416v1 Announce Type: new Abstract: Hybrid language models like Jamba mix attention layers with State Space Models (SSMs), creating two memory cache types with opposite profiles: Key-Value

Flashlight: PyTorch Compiler Extensions to Accelerate Attention Variants

ResearchDGX agent

arXiv:2511.02043v4 Announce Type: replace Abstract: Attention is a fundamental building block of large language models (LLMs), so there have been many efforts to implement it efficiently. For example,

Holomorphic Neural ODEs with Kolmogorov-Arnold Networks for Interpretable Discovery of Complex Dynamics

Model ReleasesDGX agent

arXiv:2605.22235v1 Announce Type: new Abstract: Complex dynamical systems governed by holomorphic maps such as z^2 + c exhibit fractal boundaries with extreme sensitivity to initial conditions. Accura

Reasoning through Verifiable Forecast Actions: Consistency-Grounded RL for Financial LLMs

Model ReleasesDGX agent

arXiv:2605.21975v1 Announce Type: new Abstract: Financial markets are characterized by extreme non-stationarity, low signal-to-noise ratios, and strong dependence on external information such as news,

Rule-State Inference (RSI): A Bayesian Framework for Compliance Monitoring in Rule-Governed Domains

Model ReleasesDGX agent

arXiv:2603.21610v2 Announce Type: replace Abstract: Compliance monitoring in rule-governed domains (tax administration, clinical protocol adherence, environmental regulation) faces three structural ob

Same Architecture, Different Capacity: Optimizer-Induced Spectral Scaling Laws

ResearchDGX agent

arXiv:2605.21803v1 Announce Type: new Abstract: Scaling laws have made language-model performance predictable from model size, data, and compute, but they typically treat the optimizer as a fixed trai

Self-Supervised ConvLSTM for Fermi Large Area Telescope Transient Detection

Model ReleasesDGX agent

arXiv:2605.22112v1 Announce Type: cross Abstract: We present a framework for detecting transient gamma-ray phenomena in a controlled environment by combining end-to-end simulations of the Fermi-LAT sk

Skill Weaving: Efficient LLM Improvement via Modular Skillpacks

AgentsDGX agent

arXiv:2605.22205v1 Announce Type: cross Abstract: Large language models increasingly require specialization across diverse domains, yet existing approaches struggle to balance multi-domain capacities

Stabilising Explainability Fragility in Cybersecurity AI: The Impact and Mitigation of Multicollinearity in Public Benchmark Datasets

Model ReleasesDGX agent

arXiv:2605.22529v1 Announce Type: new Abstract: This paper investigates a unexplored yet impactful vulnerability in AI explainability used in intrusion detection (IDS): multicollinearity-induced insta

The Distillation Game: Adaptive Attacks & Efficient Defenses

ResearchDGX agent

arXiv:2605.22737v1 Announce Type: new Abstract: Distillation attacks create a deployment trade-off for model providers: the same outputs that make a model more useful can also make it easier to imitat

The Secretary Problem with a Stochastic Precursor

Model ReleasesDGX agent

arXiv:2605.22653v1 Announce Type: cross Abstract: In learning-augmented online algorithms, predictions are usually valued for what they say: a value estimate, a solution, or an algorithmic recommendat

TONIC: Token-Centric Semantic Communication for Task-Oriented Wireless Systems

ResearchDGX agent

arXiv:2605.21553v1 Announce Type: new Abstract: Tokens are becoming the basic units through which foundation models represent and process information for understanding and inference. However, traditio

22 May 2026

AesFormer: Transform Everyday Photos into Beautiful Memories

Model ReleasesDGX agent

arXiv:2605.22126v1 Announce Type: new Abstract: In everyday photography, aesthetically appealing moments are often captured with structural flaws (e.g., composition, camera viewpoint, or pose) that ex

AgroVG: A Large-Scale Multi-Source Benchmark for Agricultural Visual Grounding

Model ReleasesDGX agent

arXiv:2605.22034v1 Announce Type: new Abstract: Visual grounding, the task of localizing objects described by natural-language expressions, is a foundational capability for agricultural AI systems, en

Attacking the Spike: On the Transferability and Security of Spiking Neural Networks to Adversarial Examples

ResearchDGX agent

arXiv:2209.03358v5 Announce Type: replace-cross Abstract: Spiking neural networks (SNNs) have attracted much attention for their high energy efficiency and recent advances in classification performanc

← Previous
1…469470471472473…1061
Next →