AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,403
  • Agents7,557
  • Applications5,411
  • Concepts5
  • Hardware1,836
  • Industry6,170
  • Local Ai4,931
  • Model Releases23,900
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,403
  • Agents7,557
  • Applications5,411
  • Concepts5
  • Hardware1,836
  • Industry6,170
  • Local Ai4,931
  • Model Releases23,900
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

88,403Total entries
1Added by human
88,402Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,624 results
14 May 2026

Loopholing Discrete Diffusion: Deterministic Bypass of the Sampling Wall

ResearchDGX agent

arXiv:2510.19304v3 Announce Type: replace Abstract: Discrete diffusion models offer a promising alternative to autoregressive generation through parallel decoding, but they suffer from a sampling wall

MedOpenClaw and MedFlowBench: Auditing Medical Agents in Full-Study Workflows

Model ReleasesDGX agent

arXiv:2603.24649v2 Announce Type: replace Abstract: Medical imaging benchmarks often evaluate VLMs on pre-selected 2D images, slices, crops, or patches, making evaluation closer to visual recognition.

MinT: Managed Infrastructure for Training and Serving Millions of LLMs

SafetyDGX agent

arXiv:2605.13779v1 Announce Type: cross Abstract: We present MindLab Toolkit (MinT), a managed infrastructure system for Low-Rank Adaptation (LoRA) post-training and online serving. MinT targets a set

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

MobiBench: Multi-Branch, Modular Benchmark for Mobile GUI Agents

Model ReleasesDGX agent

arXiv:2512.12634v3 Announce Type: replace Abstract: Mobile GUI Agents, AI agents capable of interacting with mobile applications on behalf of users, have the potential to transform human computer inte

Multi-Agent Systems in Emergency Departments: Validation Study on a ED Digital Twin

AgentsDGX agent

arXiv:2605.13345v1 Announce Type: new Abstract: Emergency departments (ED) face challenges in patient care and resource management. We propose to explore optimization strategies in a realistic and fle

Multi-Armed Sampling Problem and the End of Exploration

Model ReleasesDGX agent

arXiv:2507.10797v2 Announce Type: replace Abstract: This paper introduces the framework of multi-armed sampling, which serves as the sampling counterpart to the optimization problem of multi-armed ban

Multimodal Graph-based Classification of Esophageal Motility Disorders

TutorialsDGX agent

arXiv:2605.13623v1 Announce Type: new Abstract: Diagnosing esophageal motility disorders pose significant challenges due to the complexity of high-resolution impedance manometry (HRIM) data and variab

ODRPO: Ordinal Decompositions of Discrete Rewards for Robust Policy Optimization

SafetyDGX agent

arXiv:2605.12667v1 Announce Type: cross Abstract: The alignment of Large Language Models (LLMs) utilizes Reinforcement Learning from AI Feedback (RLAIF) for non-verifiable domains such as long-form qu

On Hallucinations in Inverse Problems: Fundamental Limits and Provable Assessment Methods

ResearchDGX agent

arXiv:2605.13146v1 Announce Type: cross Abstract: Artificial intelligence (AI) has transformed imaging inverse problems, from medical diagnostics to Earth observation. Yet deep neural networks can pro

Pattern-Enhanced RT-DETR for Multi-Class Battery Detection

Model ReleasesDGX agent

arXiv:2605.13670v1 Announce Type: new Abstract: Accurate and efficient battery detection is increasingly important for applications in electronic waste recycling, industrial quality control, and autom

PolySHAP: Extending KernelSHAP with Interaction-Informed Polynomial Regression

Model ReleasesDGX agent

arXiv:2601.18608v3 Announce Type: replace Abstract: Shapley values have emerged as a central game-theoretic tool in explainable AI (XAI). However, computing Shapley values exactly requires 2^d game ev

Preserve-Then-Quantize: Balancing Rank Budgets for Quantization Error Reconstruction in LLMs

Model ReleasesDGX agent

arXiv:2602.02001v2 Announce Type: replace-cross Abstract: Quantization Error Reconstruction (QER) reduces accuracy loss in Post-Training Quantization (PTQ) by approximating weights as mathbf{W} approx

Procedural Refinement by LLM-driven Algorithmic Debugging for ARC-AGI-2

Model ReleasesDGX agent

arXiv:2603.20334v2 Announce Type: replace-cross Abstract: In high-complexity abstract reasoning, a system must infer a latent rule from a few examples or structured observations and apply it to unseen

PwC expands Anthropic alliance, will train 30,000 staff on Claude

Model ReleasesDGX agent

PricewaterhouseCoopers LLP and Anthropic PBC today announced a major expansion of their alliance, as the consulting giant is set to roll out Claude Code and Claude Cowork across its workforce and trai

RealICU: Do LLM Agents Understand Long-Context ICU Data? A Benchmark Beyond Behavior Imitation

Model ReleasesDGX agent

arXiv:2605.13542v1 Announce Type: new Abstract: Intensive care units (ICU) generate long, dense and evolving streams of clinical information, where physicians must repeatedly reassess patient states u

Realtime-VLA FLASH: Speculative Inference Framework for Diffusion-based VLAs

ApplicationsDGX agent

arXiv:2605.13778v1 Announce Type: cross Abstract: Diffusion-based vision-language-action models (dVLAs) are promising for embodied intelligence but are fundamentally limited in real-time deployment by

Reducing Bias and Variance: Generative Semantic Guidance and Bi-Layer Ensemble for Image Clustering

Model ReleasesDGX agent

arXiv:2605.12961v1 Announce Type: new Abstract: Image clustering aims to partition unlabeled image datasets into distinct groups. A core aspect of this task is constructing and leveraging prior knowle

Revisiting Reinforcement Learning with Verifiable Rewards from a Contrastive Perspective

Model ReleasesDGX agent

arXiv:2605.12969v1 Announce Type: cross Abstract: RLVR has become a widely adopted paradigm for improving LLMs' reasoning capabilities, and GRPO is one of its most representative algorithms. In this p

Reward-Weighted On-Policy Distillation with an Open Property-Equivalence Verifier for NL-to-SVA Generation

SafetyDGX agent

arXiv:2605.13501v1 Announce Type: cross Abstract: LLM-based generation of SystemVerilog Assertions (SVA) is often reported as nearing saturation, with the strongest specialized model reaching {sim}76%

Rigel3D: Rig-aware Latents for Animation-Ready 3D Asset Generation

ResearchDGX agent

arXiv:2605.13129v1 Announce Type: cross Abstract: Recent 3D generative models can synthesize high-quality assets, but their outputs are typically static: they lack the skeletal rigs, joint hierarchies

RISED: A Pre-Deployment Safety Evaluation Framework for Clinical AI Decision-Support Systems

Model ReleasesDGX agent

arXiv:2605.12895v1 Announce Type: cross Abstract: Aggregate accuracy metrics dominate the evaluation of clinical AI decision-support systems but do not detect deployment-phase failures of input reliab

RoSplat: Robust Feed-Forward Pixel-wise Gaussian Splatting for Varying Input Views and High-Resolution Rendering

Model ReleasesDGX agent

arXiv:2605.13093v1 Announce Type: new Abstract: Generalizable 3D Gaussian Splatting has recently emerged as an efficient approach for novel-view synthesis, enabling feed-forward synthesis from only a

scShapeBench: Discovering geometry from high dimensional scRNAseq data

Model ReleasesDGX agent

arXiv:2605.12662v1 Announce Type: new Abstract: High-dimensional point cloud data arise across many scientific domains, especially single-cell biology. The shapes or topologies of these datasets deter

Simulating Students or Sycophantic Problem Solving? On Misconception Faithfulness of LLM Simulators

ResearchDGX agent

arXiv:2605.12748v1 Announce Type: cross Abstract: Large language models (LLMs) can fluently generate student-like responses, making them attractive as simulated students for training and evaluating AI

Skill-Aligned Annotation for Reliable Evaluation in Text-to-Image Generation

ResearchDGX agent

arXiv:2605.13223v1 Announce Type: new Abstract: Text-to-image (T2I) generation has advanced rapidly, making reliable evaluation critical as performance differences between models narrow. Existing eval

Spatiotemporal downscaling and nowcasting of urban land surface temperatures with deep neural networks

SafetyDGX agent

arXiv:2605.13566v1 Announce Type: new Abstract: Land Surface Temperature (LST) is a key variable for various applications, such as urban climate and ecology studies. Yet, existing satellite-derived LS

SSDA: Bridging Spectral and Structural Gaps via Dual Adaptation for Vision-Based Time Series Forecasting

ApplicationsDGX agent

arXiv:2605.12550v1 Announce Type: cross Abstract: Large vision models (LVMs) have recently proven to be surprisingly effective time series forecasters, simply by rendering temporal data as images. Thi

The Diffusion Encoder

ResearchDGX agent

arXiv:2605.13399v1 Announce Type: new Abstract: We construct a new kind of encoder, leveraging the expressive power of diffusion models. In a traditional variational autoencoder, the encoder and decod

The Readability Spectrum: Patterns, Issues, and Prompt Effects in LLM-Generated Code

ResearchDGX agent

arXiv:2605.13280v1 Announce Type: cross Abstract: As Large Language Models (LLMs) are transforming software development, the functional quality of generated code has become a central focus, leaving re

The WidthWall: A Strict Expressivity Hierarchy for Hypergraph Neural Networks

TutorialsDGX agent

arXiv:2605.13690v1 Announce Type: cross Abstract: Hypergraphs provide a natural framework to model higher-order interactions in scientific, social, and biological systems. Hypergraph neural networks (

ToolMol: Evolutionary Agentic Framework for Multi-objective Drug Discovery

AgentsDGX agent

arXiv:2605.12784v1 Announce Type: new Abstract: Advances in large language models (LLMs) have recently opened new and promising avenues for small-molecule drug discovery. Yet existing LLM-based approa

TouchAnything: A Dataset and Framework for Bimanual Tactile Estimation from Egocentric Video

Model ReleasesDGX agent

arXiv:2605.13083v1 Announce Type: new Abstract: Egocentric human video data, which captures rich human-environment interactions and can be collected at scale, has become a key driver of embodied intel

Tracing Persona Vectors Through LLM Pretraining

SafetyDGX agent

arXiv:2605.13329v1 Announce Type: cross Abstract: How large language models internally represent high-level behaviors is a core interpretability question with direct relevance to AI safety: it determi

Uncertainty-Driven Anomaly Detection for Psychotic Relapse Using Smartwatches: Forecasting and Multi-Task Learning Fusion

Model ReleasesDGX agent

arXiv:2605.13816v1 Announce Type: new Abstract: Digital phenotyping enables continuous passive monitoring of behavior and physiology, offering a promising paradigm for early detection of psychotic rel

UniJEPA: Enhancing Robot Policy via Unified Continuous and Discrete Representation Learning

SafetyDGX agent

arXiv:2510.10642v3 Announce Type: replace-cross Abstract: Building generalist robot policies that can handle diverse tasks in open-ended environments is a central challenge in robotics. To leverage kn

Vector-Quantized Discrete Latent Factors Meet Financial Priors: Dynamic Cross-Sectional Stock Ranking Prediction for Portfolio Construction

ResearchDGX agent

arXiv:2605.13407v1 Announce Type: new Abstract: Predicting cross-sectional stock returns is challenging due to low signal-to-noise ratios and evolving market regimes. Classical factor models offer int

VERA-MH Concept Paper

Model ReleasesDGX agent

arXiv:2510.15297v4 Announce Type: replace-cross Abstract: We introduce VERA-MH (Validation of Ethical and Responsible AI in Mental Health), an automated evaluation of the safety of AI chatbots used in

Where Does Reasoning Break? Step-Level Hallucination Detection via Hidden-State Transport Geometry

Local AiDGX agent

arXiv:2605.13772v1 Announce Type: cross Abstract: Large language models hallucinate during multi-step reasoning, but most existing detectors operate at the trace level: they assign one confidence scor

Yield Curves Dynamics Using Variational Autoencoders Under No-arbitrage

ResearchDGX agent

arXiv:2605.12764v1 Announce Type: cross Abstract: This paper introduces a physics-informed generative framework that resolves the fundamental conflict between the statistical flexibility of deep learn

13 May 2026

A Comparative Study of Federated Learning Aggregation Strategies under Homogeneous and Heterogeneous Data Distributions

Model ReleasesDGX agent

arXiv:2605.11010v1 Announce Type: new Abstract: Federated Learning has emerged as a transformative paradigm for collaborative machine learning across distributed environments. However, its performance

A New Technique for AI Explainability using Feature Association Map

Model ReleasesDGX agent

arXiv:2605.12350v1 Announce Type: new Abstract: Lack of transparency in AI systems poses challenges in critical real-life applications. It is important to be able to explain the decisions of an AI sys

A Switching System Theory of Q-Learning with Linear Function Approximation

Model ReleasesDGX agent

arXiv:2605.11021v1 Announce Type: new Abstract: This paper develops a switching-system interpretation of Q-learning with linear function approximation (LFA) based on the joint spectral radius (JSR). W

AlphaGRPO: Unlocking Self-Reflective Multimodal Generation in UMMs via Decompositional Verifiable Reward

SafetyDGX agent

arXiv:2605.12495v1 Announce Type: new Abstract: In this paper, we propose AlphaGRPO, a novel framework that applies Group Relative Policy Optimization (GRPO) to AR-Diffusion Unified Multimodal Models

Assessment of cloud and associated radiation fields from a GAN stochastic cloud subcolumn generator

SafetyDGX agent

arXiv:2605.11968v1 Announce Type: cross Abstract: Modern Earth System Models (ESMs) operate on horizontal scales far larger than typical cloud features, requiring stochastic subcolumn generators to re

Asymmetric Advantage Modulation Calibrates Entropy Dynamics in RLVR

SafetyDGX agent

arXiv:2604.04894v2 Announce Type: replace Abstract: Reinforcement learning with verifiable rewards (RLVR) has substantially improved the reasoning ability of large language models (LLMs), but it often

Beyond Manual Curation: Augmenting Targeted Protein Degradation Databases via Agentic Literature Extraction Workflows

AgentsDGX agent

arXiv:2605.11221v1 Announce Type: cross Abstract: Predictive models in biomedicine depend on structured assay data locked in the text, tables, and supplements of primary publications. This bottleneck

Beyond Similarity: Temporal Operator Attention for Time Series Analysis

ResearchDGX agent

arXiv:2605.11287v1 Announce Type: new Abstract: A persistent paradox in time-series forecasting is that structurally simple MLP and linear models often outperform high-capacity Transformers. We argue

Breaking Down and Building Up: Mixture of Skill-Based Vision-and-Language Navigation Agents

Model ReleasesDGX agent

arXiv:2508.07642v3 Announce Type: replace-cross Abstract: Vision-and-Language Navigation (VLN) poses significant challenges for agents to interpret natural language instructions and navigate complex 3

Breaking extit{Winner-Takes-All}: Cooperative Policy Optimization Improves Diverse LLM Reasoning

Model ReleasesDGX agent

arXiv:2605.11461v1 Announce Type: cross Abstract: Reinforcement learning with verifiers (RLVR) has become a central paradigm for improving LLM reasoning, yet popular group-based optimization algorithm

Calibrated Multimodal Representation Learning with Missing Modalities

Model ReleasesDGX agent

arXiv:2511.12034v2 Announce Type: replace Abstract: Multimodal representation learning harmonizes distinct modalities by aligning them into a unified latent space. Recent research generalizes traditio

Caraman at SemEval-2026 Task 8: Three-Stage Multi-Turn Retrieval with Query Rewriting, Hybrid Search, and Cross-Encoder Reranking

Model ReleasesDGX agent

arXiv:2605.12028v1 Announce Type: new Abstract: We describe our system for SemEval-2026 Task 8 (MTRAGEval), participating in Task A (Retrieval) across four English-language domains. Our approach emplo

DreamAvoid: Critical-Phase Test-Time Dreaming to Avoid Failures in VLA Policies

AgentsDGX agent

arXiv:2605.11750v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models are often brittle in fine-grained manipulation, where minor action errors during the critical phases can rapidly e

Evaluating the Pre-Consultation Ability of LLMs using Diagnostic Guidelines

Model ReleasesDGX agent

arXiv:2601.03627v3 Announce Type: replace Abstract: We introduce EPAG, a benchmark dataset and framework designed for Evaluating the Pre-consultation Ability of LLMs using diagnostic Guidelines. LLMs

Evolutionary Task Discovery: Advancing Reasoning Frontiers via Skill Composition and Complexity Scaling

ResearchDGX agent

arXiv:2605.11666v1 Announce Type: new Abstract: The reasoning frontier of Large Language Models (LLMs) has advanced significantly through modern post-training paradigms (e.g., Reinforcement Learning f

Fast MoE Inference via Predictive Prefetching and Expert Replication

HardwareDGX agent

arXiv:2605.11537v1 Announce Type: new Abstract: The Mixture of Experts (MoE) architecture has become a fundamental building block in state-of-the-art large language models (LLMs), improving domain-spe

FIS-DiT: Breaking the Few-Step Video Inference Barrier via Training-Free Frame Interleaved Sparsity

ResearchDGX agent

arXiv:2605.11869v1 Announce Type: new Abstract: While the overall inference latency of Video Diffusion Transformers (DiTs) can be substantially reduced through model distillation, per-step inference l

From Imagined Futures to Executable Actions: Mixture of Latent Actions for Robot Manipulation

SafetyDGX agent

arXiv:2605.12167v1 Announce Type: cross Abstract: Video generation models offer a promising imagination mechanism for robot manipulation by predicting long-horizon future observations, but effectively

From Web to Pixels: Bringing Agentic Search into Visual Perception

Model ReleasesDGX agent

arXiv:2605.12497v1 Announce Type: new Abstract: Visual perception connects high-level semantic understanding to pixel-level perception, but most existing settings assume that the decisive evidence for

Generative Diffusion Prior Distillation for Long-Context Knowledge Transfer

Model ReleasesDGX agent

arXiv:2605.11414v1 Announce Type: new Abstract: While traditional time-series classifiers assume full sequences at inference, practical constraints (latency and cost) often limit inputs to partial pre

GeomHerd: A Forward-looking Herding Quantification via Ricci Flow Geometry on Agent Interactive Simulations

Model ReleasesDGX agent

arXiv:2605.11645v1 Announce Type: cross Abstract: Herding -- where agents align their behaviors and act collectively -- is a central driver of market fragility and systemic risk. Existing approaches t

← Previous
1…559560561562563…1061
Next →