AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,537 results
23 Jun 2026

LUMINA-26: Low-Light Understanding for Modeling and Interpreting Night-time Actions

Model ReleasesDGX agent

arXiv:2606.23118v1 Announce Type: new Abstract: Low-light human action recognition remains a challenging problem due to poor illumination, amplified noise, motion ambiguity, and diverse real-world sce

Measuring Intent Comprehension in LLMs

Model ReleasesDGX agent

arXiv:2506.16584v3 Announce Type: replace-cross Abstract: People judge interactions with large language models (LLMs) as successful when outputs match what they want, not what they type. Yet LLMs are

MemoryVAM: Integrating Memory into Video Action Model for Robot Manipulation

SafetyDGX agent

arXiv:2606.20679v1 Announce Type: cross Abstract: Video-world-model policies learn action-relevant representations by predicting future observations. However, they condition on only a short observatio

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

muMatch: Foundation Models for Semi-supervised Learning and Domain Adaptation in EM

ResearchDGX agent

arXiv:2606.21605v1 Announce Type: new Abstract: Vision foundation models have substantially advanced computer vision, enabling state-of-the-art performance in zero- and few-shot settings. They have be

OmniNWM: Omniscient Driving Navigation World Models

SafetyDGX agent

arXiv:2510.18313v5 Announce Type: replace Abstract: Autonomous driving world models are expected to work effectively across three core dimensions: state, action, and reward. However, existing methods

PG-MAP: Joint MAP Optimization for Inference-Time Alignment of Diffusion and Flow-Matching Models

SafetyDGX agent

arXiv:2606.22958v1 Announce Type: cross Abstract: Inference-time alignment of pretrained text-to-image models is typically performed along a single control axis, such as classifier-free guidance, atte

Pose Anything Anywhere:Model-free Object Poses from Arbitrary References

SafetyDGX agent

arXiv:2606.23634v1 Announce Type: new Abstract: Estimating the 6D pose of unseen objects is a fundamental yet challenging problem for open-world robotics and embodied perception. Model-based methods a

SamatNext v0.2-B: An Exploratory Study of RMS-Normalized Hybrid Decoders for Curriculum Retention in Small Code Models

Model ReleasesDGX agent

arXiv:2606.22248v1 Announce Type: new Abstract: Standard autoregressive Transformer decoders can often exhibit substantial forgetting under sequential fine-tuning on shifting curriculum distributions.

Skeleton-to-Image Encoding: Enabling Skeleton Representation Learning via Vision-Pretrained Models

ResearchDGX agent

arXiv:2603.05963v2 Announce Type: replace Abstract: Recent advances in large-scale pretrained vision models have demonstrated impressive capabilities across a wide range of downstream tasks, including

Stationary Robust Mean-Field Games under Model Mismatches

SafetyDGX agent

arXiv:2606.22579v1 Announce Type: new Abstract: Deploying multi-agent reinforcement learning (MARL) in the real world is often limited by model mismatches between the training simulators and the true

Steer, Don't Solve: Training Small Critic Models for Large Code Agents

Model ReleasesDGX agent

arXiv:2606.21811v1 Announce Type: cross Abstract: End-to-end code agent training is resource-intensive and plateaus on the strategy-level reasoning needed to resolve code issues, since jointly optimiz

TaLK: Text-attributed Graph Dataset Distillation via Coupling Language Model with Graph-Aware Kernel

ApplicationsDGX agent

arXiv:2606.22975v1 Announce Type: new Abstract: Text-attributed graphs (TAGs) are widely used in many real-world domains, and learning on TAGs requires jointly modeling text semantics and graph struct

Understanding Latent Flow Models for Tabular Data Synthesis: Targets, Paths, and Sampling

ResearchDGX agent

arXiv:2606.20878v1 Announce Type: new Abstract: Synthetic tabular data enables microdata sharing in regulated domains, yet deploying continuous-time generative models requires balancing analytical uti

UniFS: Unified Fast-to-Slow Hierarchical Architecture for Vision-Language-Action Models

ResearchDGX agent

arXiv:2606.22794v1 Announce Type: new Abstract: Mainstream Fast-Slow dual system vision-language-action models decouple a high-frequency action expert from a low-frequency vision-language model for ef

Unlocking In-Context Learning in Audio-Language Models from Decentralized Medical Audio

ResearchDGX agent

arXiv:2606.23243v1 Announce Type: new Abstract: Clinical audio diagnosis in low-resource settings requires models that identify conditions from minimal examples without large annotated corpora. We pro

VeriBound: PAC-Bayesian Generalization Bounds for Process Reward Models Trained with Formal Verification Tools

ResearchDGX agent

arXiv:2606.20740v1 Announce Type: cross Abstract: Process Reward Models (PRMs) provide step-level verification for Large Language Model (LLM) reasoning, yet their training data acquisition remains a b

Words as Difference Makers: How Large Language Models Determine Causal Structure in Text

TutorialsDGX agent

arXiv:2606.22430v1 Announce Type: cross Abstract: Because large language models (LLMs) are impressively successful in predicting text, it appears that they must have access to a 'world model' represen

You can try it out on your own images here (~1.3GB model download in your browser) https://simonw.github.io/moebius-web/

ToolsDGX agent

This post highlights an interactive web-based demo of the Moebius model, which users can test directly in their browser with approximately 1.3GB of model data downloaded locally. The tool appears to b

20 Jun 2026

Pretty remarkable what’s happening with open weights AI right now. We’re seeing models achieve SOTA results on specific tasks, and getting c…

IndustryDGX agent

Pretty remarkable what’s happening with open weights AI right now. We’re seeing models achieve SOTA results on specific tasks, and getting close to frontier on some areas of coding and other domains.

11 Jun 2026

A Judge-Aware Ranking Framework for Evaluating Large Language Models without Ground Truth

ResearchDGX agent

arXiv:2601.21817v3 Announce Type: replace-cross Abstract: Evaluating large language models (LLMs) on open-ended tasks without ground-truth labels is increasingly done via the LLM-as-a-judge paradigm.

Afrispeech Semantics: Evaluating Audio Semantic Reasoning in Spoken Language Models Across Domains and Accents

ResearchDGX agent

arXiv:2606.11219v1 Announce Type: cross Abstract: Audio language models (ALMs) are increasingly used for speech-based understanding, yet their ability to perform semantic reasoning beyond transcriptio

Beyond Fully Random Masking: Attention-Guided Denoising and Optimization for Diffusion Language Models

ResearchDGX agent

arXiv:2606.12273v1 Announce Type: new Abstract: Diffusion large language models (dLLMs) offer an efficient alternative to autoregressive models through parallel decoding, yet existing post-training me

Cross-Layer Discrete Concept Discovery for Interpreting Language Models

ResearchDGX agent

arXiv:2506.20040v3 Announce Type: replace-cross Abstract: Interpreting language models remains challenging due to the existence of residual stream, which linearly mixes and duplicates features across

Damage-TriageFormer: A Foundation-Model Framework for Typology-Based Building Damage Assessment from Mono-Temporal Imagery

Model ReleasesDGX agent

arXiv:2606.12248v1 Announce Type: new Abstract: Decision-relevant building damage assessment is critical for prioritizing resources and recovery after a disaster, yet most automated methods either fla

Frozen Foundation-Model Embeddings Discard Small-Lesion Signal in Chest Radiography: Implications for Pre-Deployment Evaluation

Local AiDGX agent

arXiv:2606.11606v1 Announce Type: new Abstract: Frozen vision-transformer (ViT) foundation-model embeddings increasingly serve as the substrate for downstream chest-radiography (CXR) pipelines, yet wh

Lius: Translation Model Based Instructional Lingustic Using Continual Instruction Tuning In Kupang Malay

ResearchDGX agent

arXiv:2606.11786v1 Announce Type: new Abstract: Large Language Models (LLMs) offer new potential for translation tasks but often experience performance degradation when handling low-resource languages

OmniLoc: A Geometry-Aware Foundation Model for Anchor-Free UE Localization Across Diverse Indoor Environments

Model ReleasesDGX agent

arXiv:2606.11490v1 Announce Type: new Abstract: Indoor localization from wireless measurements remains challenging in large-scale deployments due to substantial variation in building geometry, the set

Physics-informed generative AI for semiconductor manufacturing: Enforcing hard physical constraints in generative models by construction

AgentsDGX agent

arXiv:2606.11247v1 Announce Type: cross Abstract: Generative models are increasingly used to propose designs, data, and control actions for physical systems, yet many such systems are governed by hard

Pretrained self-supervised speech models can recognize unseen consonants

ResearchDGX agent

arXiv:2606.11542v1 Announce Type: cross Abstract: Modern pretrained self-supervised automatic speech recognition models are trained on large-scale audio data to encode speech into contextualized repre

PRInTS: Reward Modeling for Long-Horizon Information Seeking

AgentsDGX agent

arXiv:2511.19314v2 Announce Type: replace Abstract: Information-seeking is a core capability for AI agents, requiring them to gather and reason over tool-generated information across long trajectories

Re-evaluating Confidence Remasking in Masked Diffusion Language Models

ResearchDGX agent

arXiv:2606.12232v1 Announce Type: new Abstract: Masked diffusion language models (dLLMs) have recently emerged as a competitive alternative to autoregressive language models, with the promise of faste

Time-Series Foundation Model Embeddings for Remaining Useful Life Estimation

Model ReleasesDGX agent

arXiv:2606.11990v1 Announce Type: cross Abstract: Remaining Useful Life (RUL) prediction is essential for industrial predictive maintenance, yet many learning-based approaches rely on extensive featur

Towards Fully Automated Exam Grading: Fairness-Aware Recognition of Handwritten Answers with Foundation Models

Model ReleasesDGX agent

arXiv:2606.11477v1 Announce Type: cross Abstract: Correcting handwritten exams by hand is time-consuming and error-prone, particularly for large cohorts, while fully digital exams tend to force a dida

10 Jun 2026

AI Serving Platform That Adapts to Your Model

ApplicationsDGX agent

Databricks offers an AI serving platform designed to flexibly accommodate different machine learning models and their specific requirements. The platform likely provides features for deploying, scalin

ASTRA-sim 3.0: Next-Level Distributed Machine Learning Simulations via High-Fidelity GPU and Infrastructure Modeling

HardwareDGX agent

arXiv:2606.10440v1 Announce Type: cross Abstract: Distributed machine learning (ML) is a key paradigm for today's large-scale artificial intelligence applications. As model inference arises as an impo

Breaking the Curse of Dimensionality: Diffusion Models Efficiently Learn Low-Dimensional Distributions

TutorialsDGX agent

arXiv:2409.02426v5 Announce Type: replace-cross Abstract: Despite their empirical success across a wide range of generative tasks, the fundamental principles underlying the ability of diffusion models

Flow-DPPO: Divergence Proximal Policy Optimization for Flow Matching Models

SafetyDGX agent

arXiv:2606.11025v1 Announce Type: new Abstract: Recent work has demonstrated that online reinforcement learning (RL) can substantially improve the quality and alignment of flow matching models for ima

GRAFT: Gain-Recalibrated Adapters for Transformer-Based Neural Population Activity Modeling

ResearchDGX agent

arXiv:2606.11066v1 Announce Type: new Abstract: Neural population activity models can recover rich temporal structure from binned spikes, but their read-in and readout layers often remain tied to a fi

I asked 8 AI models (including Fable 5) for their world cup predictions. Going to keep an updated leaderboard based on match results to see …

ToolsDGX agent

I asked 8 AI models (including Fable 5) for their world cup predictions. Going to keep an updated leaderboard based on match results to see which AI model performed the best! Launching tomorrow, right

In addition to transparency, I now believe frontier models should face mandatory third-party testing for cyber, bio, and autonomy risks—with…

SafetyDGX agent

In addition to transparency, I now believe frontier models should face mandatory third-party testing for cyber, bio, and autonomy risks—with the power to block or revoke deployment of models that pose

Instruction Finetuning DeepSeek-R1-8B Model Using LoRA and NEFTune

Model ReleasesDGX agent

arXiv:2606.10392v1 Announce Type: new Abstract: Financial named-entity recognition (NER) is essential for translating unstructured financial reports and news into structured knowledge graphs. However,

K-Forcing: Joint Next-K-Token Decoding via Push-Forward Language Modeling

ApplicationsDGX agent

arXiv:2606.10820v1 Announce Type: cross Abstract: Autoregressive (AR) language modeling is the dominant paradigm for text generation, yet its sequential token-by-token decoding makes inference memory-

MIND-V: Hierarchical World Model for Long-Horizon Robotic Manipulation with RL-based Physical Alignment

SafetyDGX agent

arXiv:2512.06628v3 Announce Type: replace-cross Abstract: Scalable embodied intelligence is constrained by the scarcity of diverse, long-horizon robotic manipulation data. Existing video world models

MMD Guidance: Training-Free Distribution Adaptation for Diffusion Models via Maximum Mean Discrepancy Guidance

SafetyDGX agent

arXiv:2601.08379v2 Announce Type: replace-cross Abstract: Pre-trained diffusion models have emerged as powerful generative priors for both unconditional and conditional sample generation, yet their ou

Model-Based Diffusion Sampling for Predictive Control in Offline Decision Making

ResearchDGX agent

arXiv:2512.08280v3 Announce Type: replace-cross Abstract: Offline decision-making via diffusion models often produces trajectories that are misaligned with system dynamics, limiting their reliability

NOVA: Symbolic Regression Discovery of Interpretable Car-Following and Lane-Change Models with Driver Heterogeneity

Model ReleasesDGX agent

arXiv:2606.10583v1 Announce Type: cross Abstract: We present NOVA, an autonomous symbolic regression framework that identifies interpretable car-following and lane-change structures from raw trajector

ParaBridge: Bridging Paralinguistic Perception and Dialogue Behavior in Speech Language Models

SafetyDGX agent

arXiv:2606.10581v1 Announce Type: new Abstract: Speech carries more information than just words: a child's voice, a fearful tone, or a noisy background should all lead a sufficiently competent spoken-

SPDM: Geometry-Modulated State Space Modeling with Manifold Constraints for Time Series Forecasting

Model ReleasesDGX agent

arXiv:2606.09917v1 Announce Type: new Abstract: Multivariate time series forecasting requires capturing the continuously evolving correlation structure among interacting variables. Existing state-spac

They should name the next model Halo 6 then ninja gaiden 7 cover the entire xbox catalogue

AgentsDGX agent

This post suggests naming future models after Xbox game franchises, specifically proposing 'Halo 6' and 'Ninja Gaiden 7' as model names that would reference the broader Xbox catalog. The comment appea

When Do Autoregressive Sequence Models Forecast Physical Wavefields? A Controlled Study on Synthetic Seismograms

Model ReleasesDGX agent

arXiv:2606.10868v1 Announce Type: new Abstract: Long-horizon autoregressive forecasting of oscillatory physical signals, such as seismograms, gravitational-wave strain, and similar wavefields is limit

9 Jun 2026

Anthropic played the media like a fiddle. From “untold catastrophe” to “check our latest model”, in two months and a day 🙄 (cc @tomfriedman…

SafetyDGX agent

Anthropic played the media like a fiddle. From “untold catastrophe” to “check our latest model”, in two months and a day 🙄 (cc @tomfriedman) This is the scary phase of AI — a model deemed so powerful

Artificial Intelligence for Mathematical Reasoning: An Integrated Survey of Language Models, Neuro-symbolic Systems, and Verified Discovery

Model ReleasesDGX agent

arXiv:2606.08728v1 Announce Type: new Abstract: Mathematical reasoning has long served as a stringent test of machine intelligence; over the past decade, it has moved from a niche problem within NLP t

Beyond Point Estimates: Benchmarking Uncertainty Quantification Methods on the AION-1 Astronomical Foundation Model

Local AiDGX agent

arXiv:2606.07771v1 Announce Type: cross Abstract: Foundation models for astronomical surveys offer powerful learned representations that can be transferred to downstream regression tasks such as galax

BLM-SGAN: Bidirectional Language Modeling for Semantic-Spatial Text-to-Image Generation

ResearchDGX agent

arXiv:2606.08847v1 Announce Type: cross Abstract: Despite the success of image generation from text descriptions, it still faces challenges that are difficult to overcome in domains such as natural la

ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning

SafetyDGX agent

arXiv:2606.08088v1 Announce Type: new Abstract: Reinforcement Learning from Verifiable Rewards (RLVR) has recently become a key paradigm for improving the reasoning abilities of Large Language Models

CT-VAM: A Cerebello-Thalamic-Inspired Vision-Action Model for Efficient Visuomotor Control

Local AiDGX agent

arXiv:2606.09572v1 Announce Type: cross Abstract: Vision-language-action models have shown strong promise for robot manipulation, yet raw language is primarily needed to specify task intent rather tha

DALE-CT: Depth-Aware Foundation Models for Computed Tomography

ResearchDGX agent

arXiv:2606.07775v1 Announce Type: new Abstract: Recent breakthroughs in self-supervised learning (SSL), such as the Latent-Euclidean Joint-Embedding Predictive Architecture (LeJEPA), alongside success

DeepMine-Mamba: Mitigating Information Dilution in Mamba-Based State Space Models for Document Image Binarization

Model ReleasesDGX agent

arXiv:2606.08781v1 Announce Type: new Abstract: Document image binarization aims to separate foreground text from degraded backgrounds while preserving thin, broken, and low-contrast strokes. Although

DN-Hypo-Pipeline: An AI-Driven Workflow for Hypothesis Generation via Large Language Models and Scientific Explanations

ResearchDGX agent

arXiv:2606.08532v1 Announce Type: new Abstract: A scientific hypothesis is the first step in research and undergoes experimental validation, yet it also reflects a deep understanding of and reasoning

Evaluating the Representation Space of Diffusion Models via Self-Supervised Principles

ResearchDGX agent

arXiv:2606.09718v1 Announce Type: cross Abstract: Diffusion models have demonstrated remarkable generative capabilities and have also emerged as powerful self-supervised representation learners, yet t

← Previous
1…130131132133134…1009
Next →