AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

85,188Total entries
1Added by human
85,187Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,037 results
23 Jun 2026

Asymptotic Signal Subspace Recovery in Softmax Attention Models

ResearchDGX agent

arXiv:2606.22406v1 Announce Type: new Abstract: Attention mechanisms have demonstrated remarkable empirical success in identifying relevant information from large collections of tokens, yet the theore

Combinatorial Sparse PCA Beyond the Spiked Identity Model

ApplicationsDGX agent

arXiv:2603.02607v2 Announce Type: replace-cross Abstract: Sparse PCA is one of the most well-studied problems in high-dimensional statistics. In this problem, we are given samples from a distribution

DamageArbiter: A Multimodal Arbitration Framework for Disaster Damage Assessment from Street-View Imagery

ResearchDGX agent

arXiv:2603.14837v2 Announce Type: replace Abstract: Analyzing street-view imagery with computer vision models offers a promising approach for rapid, hyperlocal disaster damage assessment, but existing

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

DBT-Bleed: Dual-Branch Temporal Modeling with Key-Frame Selection for Surgical Bleeding Detection

SafetyDGX agent

arXiv:2606.22829v1 Announce Type: new Abstract: Intraoperative Adverse Events (IAEs) detection is critical for improving surgical safety, with bleeding being among the most frequent events across many

DiT-Reward: Generative Representations for Text-to-Image Reward Modeling

SafetyDGX agent

arXiv:2606.23626v1 Announce Type: new Abstract: Can representations learned for image generation also support the evaluation of generated images? We study text-to-image reward prediction as a downstre

Federated Learning for Global Carbon Emission Forecasting: A Hybrid Time-Series Approach with Statistical and Neural Models

ResearchDGX agent

arXiv:2606.22618v1 Announce Type: new Abstract: Climate change, primarily driven by carbon dioxide (CO2) emissions, requires accurate forecasting tools to support effective mitigation policies and sus

From Discrete Plans to Real-World Execution: A World-Model-Driven Framework for Execution-Aware Multi-Agent Path Finding

AgentsDGX agent

arXiv:2511.21886v2 Announce Type: replace Abstract: Multi-Agent Path Finding (MAPF) studies how to coordinate multiple agents to reach their goals without collisions and underpins a range of large-sca

Imitation from Heterogeneous Demonstrations using Grounded Latent-Action World Models

SafetyDGX agent

arXiv:2606.21672v1 Announce Type: cross Abstract: Imitation learning has emerged as a powerful paradigm for learning visuomotor policies, but its generalisation and stability are limited by the scale

MMOU: A Massive Multi-Task Omni Understanding and Reasoning Benchmark for Long and Complex Real-World Videos

Model ReleasesDGX agent

arXiv:2603.14145v2 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) have shown strong performance in visual and audio understanding when evaluated in isolation. However,

MS-rPPG: Multi-spectral State Space Model for Remote Photoplethysmography in Driver Monitoring Systems

ApplicationsDGX agent

arXiv:2606.21115v1 Announce Type: new Abstract: Remote photoplethysmography (rPPG) is a camera-based technique for measuring physiological signals, particularly cardiac activity. From the remotely mea

MV-WAM: Manifold-Aware World Action Model with Value Augmentation

SafetyDGX agent

arXiv:2606.21088v1 Announce Type: new Abstract: Achieving robust and generalizable manipulation across diverse environments remains a fundamental challenge in embodied robotics. Recent world action mo

NegAS: Negative Label Guided Attention and Scoring for Out-of-Distribution Object Detection with Vision-Language Models

SafetyDGX agent

arXiv:2606.22537v1 Announce Type: new Abstract: Out-of-Distribution (OOD) detection is essential for ensuring the robustness and reliability of object detection systems deployed in safety-critical app

NeuPAN: Direct Point Robot Navigation with End-to-End Model-based Learning

AgentsDGX agent

arXiv:2403.06828v4 Announce Type: replace Abstract: Navigating a nonholonomic robot in a cluttered, unknown environment requires accurate perception and precise motion control for real-time collision

Predicting High-Risk Colorectal Polyps in African Americans Using Pre-Colonoscopy Clinical Features: Machine Learning Model Development and Temporal Validation

ResearchDGX agent

arXiv:2606.21492v1 Announce Type: new Abstract: Risk stratification for advanced colorectal polyps typically relies on colonoscopy and/or pathology findings. However, there is growing interest in whet

Predicting Immune Biomarkers with MultiModal Mixture-of-Expert Pathology Foundation Models Empowers Precision Oncology

ResearchDGX agent

arXiv:2606.18123v2 Announce Type: replace Abstract: Predicting immune biomarkers associated with the tumor immune microenvironment (TIME) is critical for advancing precision oncology, yet existing app

Protein contacts are already in the attention: a single-forward-pass alternative to the Categorical Jacobian

Model ReleasesDGX agent

arXiv:2606.21876v1 Announce Type: new Abstract: The Categorical Jacobian (CJ) of Zhang et al. (2024) reads protein contacts from a language model by perturbing every residue with every alternative ami

Provable Learning of Random Hierarchy Models and Hierarchical Shallow-to-Deep Chaining

TutorialsDGX agent

arXiv:2601.19756v2 Announce Type: replace Abstract: The empirical success of deep learning is often attributed to deep networks' ability to exploit hierarchical structure in data, constructing increas

RARM: Confidence-Gated Progress Reward Modeling for RL in Manipulation

ApplicationsDGX agent

arXiv:2606.22027v1 Announce Type: new Abstract: Reinforcement learning for robot manipulation is often bottlenecked by reward design, especially in long-horizon tasks: sparse success rewards provide w

Robot Self-Improvement via Human-Video Dynamics Models

SafetyDGX agent

arXiv:2606.21406v1 Announce Type: cross Abstract: A central question in robot learning is how to acquire skills from the kinds of data that humans learn from: passive observation, embodied practice, a

Self-Curriculum Model-based Reinforcement Learning for Shape Control of Deformable Linear Objects

SafetyDGX agent

arXiv:2602.21816v2 Announce Type: replace Abstract: Precise shape control of Deformable Linear Objects (DLOs) is crucial in robotic applications such as industrial and medical fields. However, existin

SimAC: A Simple Anti-Customization Method for Protecting Face Privacy against Text-to-Image Synthesis of Diffusion Models

ResearchDGX agent

arXiv:2312.07865v4 Announce Type: replace Abstract: Despite the success of diffusion-based customization methods on visual content creation, increasing concerns have been raised about such techniques

Toward Non-Expert Customized Congestion Control: Large Language Model-Assisted CCA Code Generation with eBPF Deployment

ApplicationsDGX agent

arXiv:2601.22461v2 Announce Type: replace-cross Abstract: General-purpose congestion control algorithms (CCAs) are designed to achieve general congestion control goals, but they may not meet the speci

Verifiable, private AI: Google Cloud expands Confidential Computing frontiers

Model ReleasesDGX agent

Protecting sensitive data used with AI is a critical part of our commitment to providing advanced and secure cloud infrastructure. Confidential Computing cryptographically protects data in use in hard

Zero-order Parameter-free Optimization for LMO-based Methods: Novel Approach for Efficient Fine-tuning

Model ReleasesDGX agent

arXiv:2606.14970v2 Announce Type: replace Abstract: Fine-tuning large language models (LLMs) has become a central application of modern optimization, enabling pretrained models to adapt to diverse dow

22 Jun 2026

Here is the prompt method behind this AR try-on app. The trick is not a magic prompt. It is the architecture of the prompt, and it works acr…

Model ReleasesDGX agent

Here is the prompt method behind this AR try-on app. The trick is not a magic prompt. It is the architecture of the prompt, and it works across GLM-5.2 and other frontier models. Full prompt: http://c

Use Case 1: Autonomous ML Research Can an AI autonomously improve another AI’s training recipe? We tasked Fugu Ultra with improving a small …

Model ReleasesDGX agent

Use Case 1: Autonomous ML Research Can an AI autonomously improve another AI’s training recipe? We tasked Fugu Ultra with improving a small GPT model using AutoResearch. Over 14 hours on a single H100

19 Jun 2026

Over the last two weeks, both the U.S. Government and Anthropic took significant actions that demonstrated their power to control access to …

Model ReleasesDGX agent

Over the last two weeks, both the U.S. Government and Anthropic took significant actions that demonstrated their power to control access to AI by restricting what others can do with frontier models. T

11 Jun 2026

DeceptionX: Explainable Deception Detection with Multimodal Large Language Models

ApplicationsDGX agent

arXiv:2606.11385v1 Announce Type: new Abstract: Deception detection is a critical and highly challenging task within affective computing and behavioral analysis. Existing deep learning methods typical

From Prompts to Tokens: Internalizing Causal Supervision in Vision-Language Model for Multi-Image Causal Reasoning

ResearchDGX agent

arXiv:2606.11745v1 Announce Type: cross Abstract: Visual causal reasoning is essential for understanding and intervening in the physical world, requiring identification of causal variables from visual

iPack: Intuitive Bin Packing with Large Language Models

ResearchDGX agent

arXiv:2503.08445v2 Announce Type: replace Abstract: Robotics and automation are increasingly influential in logistics but remain largely confined to traditional warehouses. In grocery retail, advancem

LakeFM: Toward a Foundation Model for Aquatic Ecosystems Using Irregular Multivariate Multi-depth Time Series Data

ApplicationsDGX agent

arXiv:2606.11268v1 Announce Type: new Abstract: Understanding and forecasting lake dynamics is critical for monitoring water quality and ecosystem health across lakes and reservoirs. While machine lea

Least-Action-Guided Diffusion for Physical Extrapolation

Model ReleasesDGX agent

arXiv:2606.11277v1 Announce Type: new Abstract: Reliable extrapolation remains a central challenge for generative models in computational physics, because models trained over finite ranges of time, pa

LUCID: Learning Embodiment-Agnostic Intent Models from Unstructured Human Videos for Scalable Dexterous Robot Skill Acquisition

SafetyDGX agent

arXiv:2606.11628v1 Announce Type: cross Abstract: The most widely-adopted robot learning pipelines today learn skills from robot demonstrations or structured human data, which are expensive to collect

MentisOculi: Revealing the Limits of Reasoning with Mental Imagery

ResearchDGX agent

arXiv:2602.02465v2 Announce Type: replace Abstract: Frontier models are transitioning from multimodal large language models (MLLMs) that merely ingest visual information to unified multimodal models (

My Chemical Harness: Evolutionary Molecular Design over Synthetic Pathways with Large Language Model Agents

Local AiDGX agent

arXiv:2606.11256v1 Announce Type: cross Abstract: Designing molecules with target properties is most useful when candidate structures are accompanied by feasible synthetic routes. We introduce My Chem

ProcessThinker: Enhancing Multi-modal Large Language Models Reasoning via Rollout-based Process Reward

SafetyDGX agent

arXiv:2606.11209v1 Announce Type: cross Abstract: Visual question answering increasingly requires multi-step reasoning. Recent post-training with reinforcement learning under verifiable rewards (RLVR)

Robust Privacy: Inference-Stage Privacy through Certified Robustness

Model ReleasesDGX agent

arXiv:2601.17360v2 Announce Type: replace-cross Abstract: An adversary observing a model's released prediction can infer sensitive attributes of the queried input, or even reconstruct representatives

Seeing Before Colliding: Anticipatory Safe RL with Frozen Vision-Language Models

SafetyDGX agent

arXiv:2606.11266v1 Announce Type: new Abstract: The cost signal that constrained-RL algorithms optimize against is almost always reactive: the simulator emits a non-zero cost only after a collision ha

Teaching Diffusion to Speculate Left-to-Right

Model ReleasesDGX agent

arXiv:2606.11552v1 Announce Type: new Abstract: Large language models (LLMs) achieve remarkable performance across a wide range of tasks, but their autoregressive decoding process incurs substantial i

VietMed-MCQ: A Consistency-Filtered Data Synthesis Framework for Vietnamese Traditional Medicine Evaluation

Model ReleasesDGX agent

arXiv:2601.03792v2 Announce Type: replace Abstract: Large Language Models (LLMs) have demonstrated remarkable proficiency in general medical domains. However, their performance significantly degrades

10 Jun 2026

ABot-Earth 0.5: Generative 3D Earth Model

ApplicationsDGX agent

arXiv:2606.09967v1 Announce Type: new Abstract: We present ABot-Earth 0.5, a generative 3D framework designed to synthesize vast, seamless 3D environments from ubiquitous, geospatially referenced sate

Advancing Wood Identification in the Philippines: Utilizing the Xylorix Platform for Efficient AI Model Development and Deployment for Five Key Species

ResearchDGX agent

arXiv:2606.10876v1 Announce Type: new Abstract: Illegal logging and timber trade continue to pose significant challenges in the Philippines, where accurate wood species identification is essential for

Beyond Explaining Predictions: Logic-Based Explanations for Confidence in Machine Learning Models

ResearchDGX agent

arXiv:2606.10347v1 Announce Type: new Abstract: Machine learning is increasingly used in critical domains, where both predictions and their associated confidence levels influence important decisions.

Compositional Generative Modeling from Decentralized Data

ResearchDGX agent

arXiv:2606.10153v1 Announce Type: new Abstract: Learning the compositional nature of the physical world requires joint observation of interacting factors. However, because practical data is often dece

DiffusionGemma

Model ReleasesDGX agent

DiffusionGemma Last May Google briefly released an experimental Gemini Diffusion model. I tried the preview at the time and recorded it running at 857 tokens/second. It was an exciting model, but Goog

From Patches to Patients: A study of the tile-to-slide performance transferability in Digital Pathology

Model ReleasesDGX agent

arXiv:2606.10778v1 Announce Type: new Abstract: Foundation Models (FMs) have recently redefined the state-of-the-art in histopathology by providing robust representations for whole-slide image (WSI) a

Less Context, Better Agents: Efficient Context Engineering for Long-Horizon Tool-Using LLM Agents

Model ReleasesDGX agent

arXiv:2606.10209v1 Announce Type: new Abstract: Large language models deployed as autonomous agents for enterprise workflows face a key challenge: verbose tool responses from enterprise systems can ca

TENP: Trapezoidal Expert Neuron Pruning For Mixture-of-Experts

Model ReleasesDGX agent

arXiv:2606.09885v1 Announce Type: new Abstract: Mixture-of-Experts large language models (LLMs) scale efficiently through sparse activation, yet their deployment is fundamentally constrained by the la

Training Set Augmentation and Biology-Aware Harmonization Improve Radiomic Models for Lung Cancer Prediction in Indeterminate Nodules

ResearchDGX agent

arXiv:2412.16758v3 Announce Type: replace-cross Abstract: CT radiomics-based machine learning has potential to predict lung cancer in pulmonary nodules (PNs) earlier than standard-of-care methods. Low

Transformer Based Model for Spatiotemporal Feature Learning in EEG Emotion Recognition

Local AiDGX agent

arXiv:2606.10718v1 Announce Type: cross Abstract: Electroencephalography (EEG) is a widely adopted technique for monitoring brain activity, offering valuable insights into neurological states due to i

Using the YOLOv12 Model for Verifying the Correct Color Sequence of Wires in Network Cables (Patch Cords) on the Production Line

ApplicationsDGX agent

arXiv:2606.10699v1 Announce Type: cross Abstract: In the production process of network cables, ensuring the correct color sequence of wire pairs inside the standard connector plays a critical role in

UXBench: Benchmarking User Experience in AI Assistants

Model ReleasesDGX agent

arXiv:2606.09570v2 Announce Type: replace Abstract: As AI assistants serve millions of users daily, evaluating user experience (UX) beyond general model capability has become increasingly important. W

9 Jun 2026

A generalizable 3D framework and model for self-supervised learning in medical imaging

ResearchDGX agent

arXiv:2501.11755v2 Announce Type: replace-cross Abstract: Current self-supervised learning methods for 3D medical imaging rely on simple pretext formulations and organ- or modality-specific datasets,

A Regret Minimization Framework on Preference Learning in Large Language Models

ResearchDGX agent

arXiv:2606.09124v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has enabled progress on reasoning-intensive tasks by relying on task-specific verifiers that provi

A Survey on Large Language Model-Based Game Agents

AgentsDGX agent

arXiv:2404.02039v5 Announce Type: replace Abstract: Game environments provide rich, controllable settings that stimulate many aspects of real-world complexity. As such, game agents offer a valuable te

Activation Steering Induces Emergent Misalignment: A More Comprehensive Evaluation

Model ReleasesDGX agent

arXiv:2606.08682v1 Announce Type: cross Abstract: Activation steering has emerged as a popular inference-time technique for modulating the behavior of large language models (LLMs). By constructing a s

Anthropic says these topics are too dangerous to let its Fable 5 model talk about

IndustryDGX agent

Anthropic has implemented safeguards on Fable 5 that restrict responses on cybersecurity, biology, and chemistry topics due to their advanced capabilities in these areas. When users ask about these se

Back to Point: Exploring Point-Language Models for Zero-Shot 3D Anomaly Detection

ResearchDGX agent

arXiv:2603.21511v2 Announce Type: replace Abstract: Zero-shot (ZS) 3D anomaly detection is crucial for reliable industrial inspection, as it enables detecting and localizing defects without requiring

CamoSAM2: SAM2-oriented Prompt Auto-Refinement for Video Camouflaged Object Detection

Model ReleasesDGX agent

arXiv:2504.00375v2 Announce Type: replace Abstract: The Segment Anything Model 2 (SAM2), a prompt-guided video foundation model, has remarkably performed in video object segmentation, drawing signific

Claude Fable 5: Available on Google Cloud

Model ReleasesDGX agent

Claude Fable 5, Anthropic’s latest frontier model, is now generally available on Google Cloud. This launch is the latest proof point of our ongoing commitment to bring the industry's latest models str

← Previous
1…231232233234235…1018
Next →