AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,376
  • Agents7,554
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,170
  • Local Ai4,930
  • Model Releases23,883
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,376
  • Agents7,554
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,170
  • Local Ai4,930
  • Model Releases23,883
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

88,376Total entries
1Added by human
88,375Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,602 results
22 May 2026

Cursor Composer 2.5's is 3–18x cheaper than Opus 4.7 in Claude Code (medium reasoning), and 5–32x cheaper than GPT-5.5 in Codex (medium) bas…

Model ReleasesDGX agent

Cursor Composer 2.5's is 3–18x cheaper than Opus 4.7 in Claude Code (medium reasoning), and 5–32x cheaper than GPT-5.5 in Codex (medium) based on API pricing This low Cost per Task isn't just driven b

DecQ: Detail-Condensing Queries for Enhanced Reconstruction and Generation in Representation Autoencoders

ResearchDGX agent

arXiv:2605.22777v1 Announce Type: new Abstract: Representation Autoencoders (RAEs) leverage frozen vision foundation models (VFMs) as tokenizer encoders, providing robust high-level representations th

Efficient Agentic Reasoning Through Self-Regulated Simulative Planning

Model Releases
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2605.22138v1 Announce Type: cross Abstract: How should an agent decide when and how to plan? A dominant approach builds agents as reactive policies with adaptive computation (e.g., chain-of-thou

EventGait: Towards Robust Gait Recognition with Event Streams

Model ReleasesDGX agent

arXiv:2605.22139v1 Announce Type: new Abstract: Gait recognition enables non-intrusive, privacy-preserving identification but suffers in uncontrolled environments due to illumination and motion sensit

ForeSplat: Optimization-Aware Foresight for Feed-Forward 3D Gaussian Splatting

ResearchDGX agent

arXiv:2605.22020v1 Announce Type: new Abstract: Feed-forward 3D Gaussian Splatting (3DGS) models offer fast single-pass reconstruction,but scaling them to match per-scene optimization quality is funda

From Recognition to Reasoning: Benchmarking and Enhancing MLLMs on Real-World Receipt Document Understanding

Model ReleasesDGX agent

arXiv:2605.22413v1 Announce Type: new Abstract: Extracting structured information from visual documents (Visual Information Extraction, VIE) is a cornerstone of business automation. While recent Multi

GazePrior: Zero-Shot AR/VR Eye Tracking via Learned 3D Gaze Reconstruction

ResearchDGX agent

arXiv:2605.22359v1 Announce Type: new Abstract: Eye tracking (ET) is a foundational technology for advanced AR/VR applications. However, training ET models for every new ET device is challenging: real

I truly miss the age of science and transparency in AI. Especially given how much money and political power and governance and scientific un…

Model ReleasesDGX agent

I truly miss the age of science and transparency in AI. Especially given how much money and political power and governance and scientific understanding is at stake. We don’t know for example • How man

LLM Readiness Harness: Evaluation, Observability, and CI Gates for LLM/RAG Applications

Model ReleasesDGX agent

arXiv:2603.27355v2 Announce Type: replace-cross Abstract: We present a readiness harness for LLM and RAG applications that turns evaluation into a deployment decision workflow. The system combines aut

Look-Closer-Then-Diagnose: Confidence-Aware Ultrasound VQA via Active Zooming

Local AiDGX agent

arXiv:2605.21652v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have significantly advanced medical visual question answering, yet their performance in ultrasound remains suboptimal. In

Making the Discrete Continuous: Synthetic RAW Augmentations for Fine-Grained Evaluation of Person Detection Performance in Low Light

SafetyDGX agent

arXiv:2605.22455v1 Announce Type: new Abstract: Real-world deployment of AI vision models is both fueled and limited by the data available for training and testing. Real datasets are sparse and uneven

OSS: Open Suturing Skills Vision-Based Assessment Challenge 2024-2025

Model ReleasesDGX agent

arXiv:2605.22200v1 Announce Type: new Abstract: Achieving high levels of surgical skill through effective training is essential for optimal patient outcomes. Automated, data-driven skill assessment ho

Pattern-and-root inflectional morphology: the Arabic broken plural

ResearchDGX agent

arXiv:2605.22310v1 Announce Type: new Abstract: We present a substantially implemented model of description of the inflectional morphology of Arabic nouns, with special attention to the management of

QuantSR+: Pushing the Limit of Quantized Image Super-Resolution Networks

SafetyDGX agent

arXiv:2605.22351v1 Announce Type: new Abstract: Low-bit quantization is widely used to compress super-resolution (SR) models and reduce storage and computation costs for deployment on resource-limited

Residual Skill Optimization for Text-to-SQL Ensembles

Model ReleasesDGX agent

arXiv:2605.21792v1 Announce Type: new Abstract: Text-to-SQL ensembles improve over single-candidate generation by drawing multiple SQL candidates and selecting one, but their effectiveness is bounded

SegGuidedNet: Sub-Region-Aware Attention Supervision for Interpretable Brain Tumor Segmentation

Model ReleasesDGX agent

arXiv:2605.22572v1 Announce Type: new Abstract: Accurate segmentation of brain tumour sub-regions from multi-parametric MRI is critical for treatment planning yet remains challenging due to morphologi

Self-Policy Distillation via Capability-Selective Subspace Projection

SafetyDGX agent

arXiv:2605.22675v1 Announce Type: new Abstract: Self-distillation bootstraps large language models (LLMs) by training on their own generations. However, existing methods either rely on external signal

SiameseNorm: Breaking the Barrier to Reconciling Pre/Post-Norm

Model ReleasesDGX agent

arXiv:2602.08064v2 Announce Type: replace-cross Abstract: The long-standing tension between Pre- and Post-Norm remains an open problem in Transformer architecture, reflecting a fundamental trade-off b

Tokenization with Split Trees

Model ReleasesDGX agent

arXiv:2605.22705v1 Announce Type: new Abstract: We introduce Tokenization with Split Trees (ToaST), a subword tokenization method that directly optimizes compression under a new recursive inference pr

Towards Clinically Interpretable Ophthalmic VQA via Spatially-Grounded Lesion Evidence

Model ReleasesDGX agent

arXiv:2605.22414v1 Announce Type: new Abstract: Visual Question Answering (VQA) holds great promise for clinical support, particularly in ophthalmology, where retinal fundus photography is essential f

Ultra-High-Definition Image Quality Assessment via Graph Representation Learning

Model ReleasesDGX agent

arXiv:2605.22192v1 Announce Type: new Abstract: Blind image quality assessment (BIQA) for ultrahighdefinition (UHD) images remains challenging because native-resolution inference is computationally ex

Vector Policy Optimization: Training for Diversity Improves Test-Time Search

SafetyDGX agent

arXiv:2605.22817v1 Announce Type: cross Abstract: Language models must now generalize out of the box to novel environments and work inside inference-scaling search procedures, such as AlphaEvolve, tha

21 May 2026

A Semantic and Occlusion-Aware GM-PHD Filter

AgentsDGX agent

arXiv:2605.20666v1 Announce Type: new Abstract: This paper proposes a new birth model including semantic information derived from deep learning to create an occlusion-aware Gaussian Mixture Probabilit

ACL-Verbatim: hallucination-free question answering for research

Model ReleasesDGX agent

arXiv:2605.21102v1 Announce Type: new Abstract: Academic researchers need efficient and reliable methods for collecting high-quality information from trusted sources, but modern tools for AI-assisted

AnimeAdapter: Fine-grained and Consistent Zero-shot Anime Character Generation

Model ReleasesDGX agent

arXiv:2605.20237v1 Announce Type: new Abstract: We present a lightweight appearance adapter for Stable Diffusion that enables controllable and consistent anime character generation under diverse editi

Conditioning Gaussian Processes on Almost Anything

ApplicationsDGX agent

arXiv:2605.21041v1 Announce Type: cross Abstract: Gaussian processes (GPs) offer a principled probabilistic model over functions, but exact inference is restricted to the linear-Gaussian regime. We es

Cowork Productivity Assistant:Qwen3.7-Max serves as your advanced coworker for real-world productivity.

ApplicationsDGX agent

Qwen3.7-Max is an advanced AI model designed to function as a virtual coworker that enhances real-world productivity tasks. The model likely leverages Alibaba's Qwen technology to assist with work-rel

Distribution-Aware Reward: Reinforcement Learning over Predictive Distributions for LLM Regression

SafetyDGX agent

arXiv:2605.20740v1 Announce Type: cross Abstract: Large language models can predict real-valued quantities from heterogeneous inputs such as text, code, and molecular strings, but most training object

Distributional Alignment as a Criterion for Designing Task Vectors in In-Context Learning

SafetyDGX agent

arXiv:2605.20730v1 Announce Type: new Abstract: In-context learning (ICL) allows large language models (LLMs) to adapt to new tasks through demonstrations, yet it suffers from escalating inference cos

Epistemic Uncertainty Quantification for Pre-trained VLMs via Riemannian Flow Matching

ResearchDGX agent

arXiv:2601.21662v2 Announce Type: replace Abstract: Vision-Language Models (VLMs) are typically deterministic in nature and lack intrinsic mechanisms to quantify epistemic uncertainty, which reflects

FedCoE: Bridging Generalization and Personalization via Federated Coordinated Dual-level MoEs

Model ReleasesDGX agent

arXiv:2605.21264v1 Announce Type: new Abstract: Federated Learning (FL) has emerged as a promising paradigm for privacy-preserving distributed learning. However, existing FL methods face a fundamental

FlowLong: Inference-time Long Video Generation via Manifold-constrained Tweedie Matching

SafetyDGX agent

arXiv:2605.20910v1 Announce Type: new Abstract: Extending the generation horizon of video diffusion models to long sequences remains a long-standing and important challenge. Existing training-free app

Free-Grained Hierarchical Visual Recognition

Model ReleasesDGX agent

arXiv:2510.14737v3 Announce Type: replace Abstract: Hierarchical image recognition seeks to predict class labels along a semantic taxonomy, from broad categories to specific ones, typically under the

GSA-YOLO: A High-Efficiency Framework via Structured Sparsity and Adaptive Knowledge Distillation for Real-Time X-ray Security Inspection

ResearchDGX agent

arXiv:2605.20669v1 Announce Type: new Abstract: X-ray security inspection requires accurate real-time detection of prohibited items, but existing models often struggle to balance the challenges of sev

Humanoid Whole-Body Manipulation via Active Spatial Brain and Generalizable Action Cerebellum

Model ReleasesDGX agent

arXiv:2605.21133v1 Announce Type: new Abstract: In this paper, we explore spatial-aware humanoid whole-body manipulation task. Compared with tabletop settings, this task poses two key challenges: 1) S

Lip sync and Lora

Local AiDGX agent

LoRA models can be used with image-to-video workflows to create automatic lip-syncing effects by controlling camera movements and instructing the model to synchronize audio with mouth movements. This

Long-Context Reasoning Through Proxy-Based Chain-of-Thought Tuning

ResearchDGX agent

arXiv:2605.20201v1 Announce Type: new Abstract: Recent large language models support inputs of up to 10 million tokens, yet they perform poorly on long-context tasks that require complex reasoning. Su

LT2: Linear-Time Looped Transformers

TutorialsDGX agent

arXiv:2605.20670v1 Announce Type: new Abstract: Looped Transformers (LT) have emerged as a powerful architecture by iterating their layers multiple times before decoding the final token. However, pair

LTX 2.3 + LTX Director is a Huge improvement

Local AiDGX agent

LTX-2.3 is a multimodal video generation model developed by Lightricks that generates synchronized audio and video in a single forward pass at resolutions up to 4K at 50 frames per second. The model i

Markovian Circuit Tracing for Transformer State Dynamic

Model ReleasesDGX agent

arXiv:2605.20824v1 Announce Type: new Abstract: Many sequence computations are easier to study as movement through internal states than as isolated local circuits. We introduce Markovian Circuit Traci

Modality-Decoupled Online Recursive Editing

Local AiDGX agent

arXiv:2605.20273v1 Announce Type: new Abstract: Online model editing for multimodal large language models (MLLMs) requires assimilating a stream of corrections under tight compute and memory budgets.

Modular Multimodal Classification Without Fine-Tuning: A Simple Compositional Approach

ResearchDGX agent

arXiv:2605.20674v1 Announce Type: new Abstract: We introduce CoMET, extit{extbf{C}omposing extbf{M}odality extbf{E}ncoders with extbf{T}abular foundation models}, a simple yet highly competitive metho

RadProPoser: Probabilistic Radar Tensor Human Pose Estimation That Knows Its Limits

Model ReleasesDGX agent

arXiv:2508.03578v2 Announce Type: replace Abstract: Radar-based human pose estimation enables privacy-preserving motion tracking for ambient intelligence, yet the noisy nature of radar sensing makes u

Reinforcing Human Behavior Simulation via Verbal Feedback

Model ReleasesDGX agent

arXiv:2605.20506v1 Announce Type: cross Abstract: Humans learn social norms and behaviors from verbal feedback (e.g., a parent saying 'that was rude' or a friend explaining 'here's why that hurt'). Ye

Resolving Long-Tail Ambiguity in Unsupervised 3D Point Cloud Segmentation with Language Priors

Model ReleasesDGX agent

arXiv:2605.20737v1 Announce Type: new Abstract: Existing approaches for unsupervised 3D point cloud segmentation predominantly rely on a purely visual similarity-based learning-by-clustering paradigm,

Reviving Error Correction in Modern Deep Time-Series Forecasting

ResearchDGX agent

arXiv:2605.21088v1 Announce Type: new Abstract: Modern deep-learning models have achieved remarkable success in time-series forecasting. Yet, their performance degrades in long-term prediction due to

Runtime-Certified Bounded-Error Quantized Attention

Model ReleasesDGX agent

arXiv:2605.20868v1 Announce Type: new Abstract: KV cache quantization reduces the memory cost of long-context LLM inference, but introduces approximation error that is typically validated only empiric

Semantic Granularity Navigation in Image Editing

Local AiDGX agent

arXiv:2605.21190v1 Announce Type: new Abstract: Despite the generative capabilities of diffusion and flow models, real-image editing remains constrained by a persistent trade-off between semantic edit

SR-Ground: Image Quality Grounding for Super-Resolved Content

ResearchDGX agent

arXiv:2605.21244v1 Announce Type: new Abstract: Super-Resolution (SR) has advanced rapidly in recent years, with diffusion-based models achieving unprecedented fidelity at the cost of introducing new

Strategy-Induct: Task-Level Strategy Induction for Instruction Generation

ResearchDGX agent

arXiv:2605.20924v1 Announce Type: new Abstract: Designing effective task-level prompts is crucial for improving the performance of Large Language Models (LLMs). While prior work on instruction inducti

The Devil is in the Condition Numbers: Why is GLU Better than non-GLU Structure?

ResearchDGX agent

arXiv:2605.20749v1 Announce Type: new Abstract: Gated Linear Units (GLU) and their variants are widely adopted in modern open-source large language model architectures and consistently outperform thei

The General Theory of Localization Methods

Local AiDGX agent

arXiv:2605.20635v1 Announce Type: new Abstract: This paper proposes a general machine learning framework called the localization method, which is fundamentally built on two core concepts: localization

This is the way

IndustryDGX agent

This is the way Elon Musk on why the Model S and Model X succeeded: “Those cars were designed with love. Every part of it, inside and outside, even things people couldn’t see - we put there because we

Training Language Agents to Learn from Experience

Model ReleasesDGX agent

arXiv:2605.20477v1 Announce Type: cross Abstract: Language agents can adapt from experience in interactive environments, but current reflection-based methods can only self-correct within a single task

VDFP: Video Deflickering with Flicker-banding Priors

Model ReleasesDGX agent

arXiv:2605.21079v1 Announce Type: new Abstract: Capturing digital screens with smartphones frequently induces severe banding due to hardware synchronization mismatches. Existing video restoration meth

WCXB: A Multi-Type Web Content Extraction Benchmark

Model ReleasesDGX agent

arXiv:2605.21097v1 Announce Type: new Abstract: Web content extraction - isolating a page's main content from surrounding boilerplate - is a prerequisite for search indexing, retrieval-augmented gener

What if Agents Could Imagine? Reinforcing Open-Vocabulary HOI Comprehension through Generation

AgentsDGX agent

arXiv:2602.11499v2 Announce Type: replace Abstract: Multimodal Large Language Models have shown promising capabilities in bridging visual and textual reasoning, yet their reasoning capabilities in Ope

When Irregularity Helps: A Subclass Analysis of Inductive Bias in Neural Morphology

Model ReleasesDGX agent

arXiv:2605.20558v1 Announce Type: new Abstract: Neural morphological generation systems often achieve high aggregate accuracy on benchmark datasets, yet such performance can conceal systematic errors

20 May 2026

A Framework for Evaluating Zero-Shot Image Generation in Concept-based Explainability

ResearchDGX agent

arXiv:2605.19855v1 Announce Type: cross Abstract: Concept-based Explainable Artificial Intelligence (XAI) interprets deep learning models using human-understandable visual features (e.g., textures or

A Methodology for Selecting and Composing Runtime Architecture Patterns for Production LLM Agents

AgentsDGX agent

arXiv:2605.20173v1 Announce Type: new Abstract: Production LLM agents combine stochastic model outputs with deterministic software systems, yet the boundary between the two is rarely treated as a firs

← Previous
1…470471472473474…1061
Next →