AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlog
91,060Total entries
1Added by human
91,059Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,794 results
Model Releases

Stabilising Explainability Fragility in Cybersecurity AI: The Impact and Mitigation of Multicollinearity in Public Benchmark Datasets

DGX agent

arXiv:2605.22529v1 Announce Type: new Abstract: This paper investigates a unexplored yet impactful vulnerability in AI explainability used in intrusion detection (IDS): multicollinearity-induced insta

model-releasesarxiv-cs-lg
23 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Research

The Distillation Game: Adaptive Attacks & Efficient Defenses

DGX agent

arXiv:2605.22737v1 Announce Type: new Abstract: Distillation attacks create a deployment trade-off for model providers: the same outputs that make a model more useful can also make it easier to imitat

researcharxiv-cs-lg
23 May 2026
Model Releases

The Secretary Problem with a Stochastic Precursor

DGX agent

arXiv:2605.22653v1 Announce Type: cross Abstract: In learning-augmented online algorithms, predictions are usually valued for what they say: a value estimate, a solution, or an algorithmic recommendat

model-releasesarxiv-cs-lg
23 May 2026
Research

TONIC: Token-Centric Semantic Communication for Task-Oriented Wireless Systems

DGX agent

arXiv:2605.21553v1 Announce Type: new Abstract: Tokens are becoming the basic units through which foundation models represent and process information for understanding and inference. However, traditio

researcharxiv-cs-lg
23 May 2026
Model Releases

AesFormer: Transform Everyday Photos into Beautiful Memories

DGX agent

arXiv:2605.22126v1 Announce Type: new Abstract: In everyday photography, aesthetically appealing moments are often captured with structural flaws (e.g., composition, camera viewpoint, or pose) that ex

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

AgroVG: A Large-Scale Multi-Source Benchmark for Agricultural Visual Grounding

DGX agent

arXiv:2605.22034v1 Announce Type: new Abstract: Visual grounding, the task of localizing objects described by natural-language expressions, is a foundational capability for agricultural AI systems, en

model-releasesarxiv-cs-cv
22 May 2026
Research

Attacking the Spike: On the Transferability and Security of Spiking Neural Networks to Adversarial Examples

DGX agent

arXiv:2209.03358v5 Announce Type: replace-cross Abstract: Spiking neural networks (SNNs) have attracted much attention for their high energy efficiency and recent advances in classification performanc

researcharxiv-cs-cv
22 May 2026
Model Releases

Cursor Composer 2.5's is 3–18x cheaper than Opus 4.7 in Claude Code (medium reasoning), and 5–32x cheaper than GPT-5.5 in Codex (medium) bas…

DGX agent

Cursor Composer 2.5's is 3–18x cheaper than Opus 4.7 in Claude Code (medium reasoning), and 5–32x cheaper than GPT-5.5 in Codex (medium) based on API pricing This low Cost per Task isn't just driven b

model-releaseselon-musk--x
22 May 2026
Research

DecQ: Detail-Condensing Queries for Enhanced Reconstruction and Generation in Representation Autoencoders

DGX agent

arXiv:2605.22777v1 Announce Type: new Abstract: Representation Autoencoders (RAEs) leverage frozen vision foundation models (VFMs) as tokenizer encoders, providing robust high-level representations th

researcharxiv-cs-cv
22 May 2026
Model Releases

Efficient Agentic Reasoning Through Self-Regulated Simulative Planning

DGX agent

arXiv:2605.22138v1 Announce Type: cross Abstract: How should an agent decide when and how to plan? A dominant approach builds agents as reactive policies with adaptive computation (e.g., chain-of-thou

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

EventGait: Towards Robust Gait Recognition with Event Streams

DGX agent

arXiv:2605.22139v1 Announce Type: new Abstract: Gait recognition enables non-intrusive, privacy-preserving identification but suffers in uncontrolled environments due to illumination and motion sensit

model-releasesarxiv-cs-cv
22 May 2026
Research

ForeSplat: Optimization-Aware Foresight for Feed-Forward 3D Gaussian Splatting

DGX agent

arXiv:2605.22020v1 Announce Type: new Abstract: Feed-forward 3D Gaussian Splatting (3DGS) models offer fast single-pass reconstruction,but scaling them to match per-scene optimization quality is funda

researcharxiv-cs-cv
22 May 2026
Model Releases

From Recognition to Reasoning: Benchmarking and Enhancing MLLMs on Real-World Receipt Document Understanding

DGX agent

arXiv:2605.22413v1 Announce Type: new Abstract: Extracting structured information from visual documents (Visual Information Extraction, VIE) is a cornerstone of business automation. While recent Multi

model-releasesarxiv-cs-cv
22 May 2026
Research

GazePrior: Zero-Shot AR/VR Eye Tracking via Learned 3D Gaze Reconstruction

DGX agent

arXiv:2605.22359v1 Announce Type: new Abstract: Eye tracking (ET) is a foundational technology for advanced AR/VR applications. However, training ET models for every new ET device is challenging: real

researcharxiv-cs-cv
22 May 2026
Model Releases

I truly miss the age of science and transparency in AI. Especially given how much money and political power and governance and scientific un…

DGX agent

I truly miss the age of science and transparency in AI. Especially given how much money and political power and governance and scientific understanding is at stake. We don’t know for example • How man

model-releasesgary-marcus--x
22 May 2026
Model Releases

LLM Readiness Harness: Evaluation, Observability, and CI Gates for LLM/RAG Applications

DGX agent

arXiv:2603.27355v2 Announce Type: replace-cross Abstract: We present a readiness harness for LLM and RAG applications that turns evaluation into a deployment decision workflow. The system combines aut

model-releasesarxiv-cs-cl
22 May 2026
Local Ai

Look-Closer-Then-Diagnose: Confidence-Aware Ultrasound VQA via Active Zooming

DGX agent

arXiv:2605.21652v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have significantly advanced medical visual question answering, yet their performance in ultrasound remains suboptimal. In

local-aiarxiv-cs-cv
22 May 2026
Safety

Making the Discrete Continuous: Synthetic RAW Augmentations for Fine-Grained Evaluation of Person Detection Performance in Low Light

DGX agent

arXiv:2605.22455v1 Announce Type: new Abstract: Real-world deployment of AI vision models is both fueled and limited by the data available for training and testing. Real datasets are sparse and uneven

safetyarxiv-cs-cv
22 May 2026
Model Releases

OSS: Open Suturing Skills Vision-Based Assessment Challenge 2024-2025

DGX agent

arXiv:2605.22200v1 Announce Type: new Abstract: Achieving high levels of surgical skill through effective training is essential for optimal patient outcomes. Automated, data-driven skill assessment ho

model-releasesarxiv-cs-cv
22 May 2026
Research

Pattern-and-root inflectional morphology: the Arabic broken plural

DGX agent

arXiv:2605.22310v1 Announce Type: new Abstract: We present a substantially implemented model of description of the inflectional morphology of Arabic nouns, with special attention to the management of

researcharxiv-cs-cl
22 May 2026
Safety

QuantSR+: Pushing the Limit of Quantized Image Super-Resolution Networks

DGX agent

arXiv:2605.22351v1 Announce Type: new Abstract: Low-bit quantization is widely used to compress super-resolution (SR) models and reduce storage and computation costs for deployment on resource-limited

safetyarxiv-cs-cv
22 May 2026
Model Releases

Residual Skill Optimization for Text-to-SQL Ensembles

DGX agent

arXiv:2605.21792v1 Announce Type: new Abstract: Text-to-SQL ensembles improve over single-candidate generation by drawing multiple SQL candidates and selecting one, but their effectiveness is bounded

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

SegGuidedNet: Sub-Region-Aware Attention Supervision for Interpretable Brain Tumor Segmentation

DGX agent

arXiv:2605.22572v1 Announce Type: new Abstract: Accurate segmentation of brain tumour sub-regions from multi-parametric MRI is critical for treatment planning yet remains challenging due to morphologi

model-releasesarxiv-cs-cv
22 May 2026
Safety

Self-Policy Distillation via Capability-Selective Subspace Projection

DGX agent

arXiv:2605.22675v1 Announce Type: new Abstract: Self-distillation bootstraps large language models (LLMs) by training on their own generations. However, existing methods either rely on external signal

safetyarxiv-cs-cl
22 May 2026
Model Releases

SiameseNorm: Breaking the Barrier to Reconciling Pre/Post-Norm

DGX agent

arXiv:2602.08064v2 Announce Type: replace-cross Abstract: The long-standing tension between Pre- and Post-Norm remains an open problem in Transformer architecture, reflecting a fundamental trade-off b

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

Tokenization with Split Trees

DGX agent

arXiv:2605.22705v1 Announce Type: new Abstract: We introduce Tokenization with Split Trees (ToaST), a subword tokenization method that directly optimizes compression under a new recursive inference pr

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

Towards Clinically Interpretable Ophthalmic VQA via Spatially-Grounded Lesion Evidence

DGX agent

arXiv:2605.22414v1 Announce Type: new Abstract: Visual Question Answering (VQA) holds great promise for clinical support, particularly in ophthalmology, where retinal fundus photography is essential f

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

Ultra-High-Definition Image Quality Assessment via Graph Representation Learning

DGX agent

arXiv:2605.22192v1 Announce Type: new Abstract: Blind image quality assessment (BIQA) for ultrahighdefinition (UHD) images remains challenging because native-resolution inference is computationally ex

model-releasesarxiv-cs-cv
22 May 2026
Safety

Vector Policy Optimization: Training for Diversity Improves Test-Time Search

DGX agent

arXiv:2605.22817v1 Announce Type: cross Abstract: Language models must now generalize out of the box to novel environments and work inside inference-scaling search procedures, such as AlphaEvolve, tha

safetyarxiv-cs-cl
22 May 2026
Agents

A Semantic and Occlusion-Aware GM-PHD Filter

DGX agent

arXiv:2605.20666v1 Announce Type: new Abstract: This paper proposes a new birth model including semantic information derived from deep learning to create an occlusion-aware Gaussian Mixture Probabilit

agentsarxiv-cs-ro
21 May 2026
Model Releases

ACL-Verbatim: hallucination-free question answering for research

DGX agent

arXiv:2605.21102v1 Announce Type: new Abstract: Academic researchers need efficient and reliable methods for collecting high-quality information from trusted sources, but modern tools for AI-assisted

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

AnimeAdapter: Fine-grained and Consistent Zero-shot Anime Character Generation

DGX agent

arXiv:2605.20237v1 Announce Type: new Abstract: We present a lightweight appearance adapter for Stable Diffusion that enables controllable and consistent anime character generation under diverse editi

model-releasesarxiv-cs-cv
21 May 2026
Applications

Conditioning Gaussian Processes on Almost Anything

DGX agent

arXiv:2605.21041v1 Announce Type: cross Abstract: Gaussian processes (GPs) offer a principled probabilistic model over functions, but exact inference is restricted to the linear-Gaussian regime. We es

applicationsarxiv-cs-lg
21 May 2026
Applications

Cowork Productivity Assistant:Qwen3.7-Max serves as your advanced coworker for real-world productivity.

DGX agent

Qwen3.7-Max is an advanced AI model designed to function as a virtual coworker that enhances real-world productivity tasks. The model likely leverages Alibaba's Qwen technology to assist with work-rel

applicationsqwen--x
21 May 2026
Safety

Distribution-Aware Reward: Reinforcement Learning over Predictive Distributions for LLM Regression

DGX agent

arXiv:2605.20740v1 Announce Type: cross Abstract: Large language models can predict real-valued quantities from heterogeneous inputs such as text, code, and molecular strings, but most training object

safetyarxiv-cs-cl
21 May 2026
Safety

Distributional Alignment as a Criterion for Designing Task Vectors in In-Context Learning

DGX agent

arXiv:2605.20730v1 Announce Type: new Abstract: In-context learning (ICL) allows large language models (LLMs) to adapt to new tasks through demonstrations, yet it suffers from escalating inference cos

safetyarxiv-cs-cl
21 May 2026
Research

Epistemic Uncertainty Quantification for Pre-trained VLMs via Riemannian Flow Matching

DGX agent

arXiv:2601.21662v2 Announce Type: replace Abstract: Vision-Language Models (VLMs) are typically deterministic in nature and lack intrinsic mechanisms to quantify epistemic uncertainty, which reflects

researcharxiv-cs-lg
21 May 2026
Model Releases

FedCoE: Bridging Generalization and Personalization via Federated Coordinated Dual-level MoEs

DGX agent

arXiv:2605.21264v1 Announce Type: new Abstract: Federated Learning (FL) has emerged as a promising paradigm for privacy-preserving distributed learning. However, existing FL methods face a fundamental

model-releasesarxiv-cs-lg
21 May 2026
Safety

FlowLong: Inference-time Long Video Generation via Manifold-constrained Tweedie Matching

DGX agent

arXiv:2605.20910v1 Announce Type: new Abstract: Extending the generation horizon of video diffusion models to long sequences remains a long-standing and important challenge. Existing training-free app

safetyarxiv-cs-cv
21 May 2026
Model Releases

Free-Grained Hierarchical Visual Recognition

DGX agent

arXiv:2510.14737v3 Announce Type: replace Abstract: Hierarchical image recognition seeks to predict class labels along a semantic taxonomy, from broad categories to specific ones, typically under the

model-releasesarxiv-cs-cv
21 May 2026
Research

GSA-YOLO: A High-Efficiency Framework via Structured Sparsity and Adaptive Knowledge Distillation for Real-Time X-ray Security Inspection

DGX agent

arXiv:2605.20669v1 Announce Type: new Abstract: X-ray security inspection requires accurate real-time detection of prohibited items, but existing models often struggle to balance the challenges of sev

researcharxiv-cs-cv
21 May 2026
Model Releases

Humanoid Whole-Body Manipulation via Active Spatial Brain and Generalizable Action Cerebellum

DGX agent

arXiv:2605.21133v1 Announce Type: new Abstract: In this paper, we explore spatial-aware humanoid whole-body manipulation task. Compared with tabletop settings, this task poses two key challenges: 1) S

model-releasesarxiv-cs-ro
21 May 2026
Local Ai

Lip sync and Lora

DGX agent

LoRA models can be used with image-to-video workflows to create automatic lip-syncing effects by controlling camera movements and instructing the model to synchronize audio with mouth movements. This

local-air-stablediffusion
21 May 2026
Research

Long-Context Reasoning Through Proxy-Based Chain-of-Thought Tuning

DGX agent

arXiv:2605.20201v1 Announce Type: new Abstract: Recent large language models support inputs of up to 10 million tokens, yet they perform poorly on long-context tasks that require complex reasoning. Su

researcharxiv-cs-cl
21 May 2026
Tutorials

LT2: Linear-Time Looped Transformers

DGX agent

arXiv:2605.20670v1 Announce Type: new Abstract: Looped Transformers (LT) have emerged as a powerful architecture by iterating their layers multiple times before decoding the final token. However, pair

tutorialsarxiv-cs-lg
21 May 2026
Local Ai

LTX 2.3 + LTX Director is a Huge improvement

DGX agent

LTX-2.3 is a multimodal video generation model developed by Lightricks that generates synchronized audio and video in a single forward pass at resolutions up to 4K at 50 frames per second. The model i

local-air-stablediffusion
21 May 2026
Model Releases

Markovian Circuit Tracing for Transformer State Dynamic

DGX agent

arXiv:2605.20824v1 Announce Type: new Abstract: Many sequence computations are easier to study as movement through internal states than as isolated local circuits. We introduce Markovian Circuit Traci

model-releasesarxiv-cs-lg
21 May 2026
Local Ai

Modality-Decoupled Online Recursive Editing

DGX agent

arXiv:2605.20273v1 Announce Type: new Abstract: Online model editing for multimodal large language models (MLLMs) requires assimilating a stream of corrections under tight compute and memory budgets.

local-aiarxiv-cs-lg
21 May 2026
← Previous
1…610611612613614…1371
Next →