AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

85,136Total entries
1Added by human
85,135Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,002 results
3 Jun 2026

Reinforcement Learning from Cross-domain Videos with Video Prediction Model

AgentsDGX agent

arXiv:2606.03201v1 Announce Type: cross Abstract: Reinforcement learning from expert videos across visually distinct domains is challenging due to the absence of reward signals and the presence of dom

Small RL Controller, Large Language Model: RL-Guided Adaptive Sampling for Test-Time Scaling

ResearchDGX agent

arXiv:2606.03102v1 Announce Type: new Abstract: Test-time scaling improves the reasoning performance of large language models but incurs substantial cost in both total computation and latency. Existin

Speedrunning Tabular Foundation Model Pretraining

HardwareDGX agent

arXiv:2606.03681v1 Announce Type: new Abstract: Pretraining cost is a major bottleneck for research on tabular foundation models, slowing the iteration cycle for new architectures, priors, and optimiz

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Traj-Evolve: A Self-Evolving Multi-Agent System for Patient Trajectory Modeling in Lung Cancer Early Detection

AgentsDGX agent

arXiv:2606.02812v1 Announce Type: new Abstract: Modeling patient trajectories from longitudinal electronic health records (EHRs) requires reasoning over sparse, noisy, and long-context multimodal sequ

Typhoon: Towards an Effective Task-Specific Masking Strategy for Pre-trained Language Models

ResearchDGX agent

arXiv:2303.15619v2 Announce Type: replace-cross Abstract: The choice of which tokens to mask is a central, under-examined design decision in masked language modeling (MLM). Standard pretraining masks

You can use Hermes Desktop with Ollama using local or cloud models. Get started 👇👇👇

Local AiDGX agent

You can use Hermes Desktop with Ollama using local or cloud models. Get started 👇👇👇 The next evolution of Hermes Agent is here! Introducing Hermes Desktop: everything you love about Hermes, now native

2 Jun 2026

A live blog of Microsoft Build 2026, where the company is set to unveil a Copilot 'super app', a new reasoning model, and Windows improvements for developers (Engadget)

IndustryDGX agent

Engadget: A live blog of Microsoft Build 2026, where the company is set to unveil a Copilot “super app”, a new reasoning model, and Windows improvements for developers — We're expecting the presentati

AI Governance Maturity Model: Matrix, Assessment, and Roadmap

IndustryDGX agent

This Databricks resource presents a framework for assessing and advancing an organization's AI governance capabilities through a maturity model structure. It likely provides a matrix for evaluating cu

AI will significantly disrupt IT consultancies as AI labs build their own advisory arms and execs expect more outcome-based pricing over hourly billing models (Stephen Foley/Financial Times)

IndustryDGX agent

Stephen Foley / Financial Times: AI will significantly disrupt IT consultancies as AI labs build their own advisory arms and execs expect more outcome-based pricing over hourly billing models — Accent

An Enigma of Artificial Reason: Investigating the Production-Evaluation Gap in Large Reasoning Models

SafetyDGX agent

arXiv:2606.01462v1 Announce Type: new Abstract: Studies of human reasoning have shown that people are typically stronger at evaluating reasoning than producing it from scratch. In contrast, large reas

An interview with Sam Altman on OpenAI's massive Stargate data center project in Saline, Michigan, coding models being the biggest driver of AI demand, and more (CNBC)

IndustryDGX agent

CNBC: An interview with Sam Altman on OpenAI's massive Stargate data center project in Saline, Michigan, coding models being the biggest driver of AI demand, and more — Following is the unofficial tra

Arcee needs more attention that it gets! There aren't a lot of great American open-source AI model companies and they're one of them! https:…

IndustryDGX agent

Arcee is highlighted as a notable American open-source AI model company that deserves greater recognition within the AI community. The post emphasizes the scarcity of strong domestic competitors in th

ART: Attention Run-time Termination for Efficient Large Language Model Decoding

ResearchDGX agent

arXiv:2606.00024v1 Announce Type: new Abstract: Long-context decoding in Large Language Models (LLMs) is severely constrained by the memory bandwidth required to fetch the extensive Key-Value (KV) cac

Behavior-Invariant Task Representation Learning with Transformer-based World Models for Offline Meta-Reinforcement Learning

SafetyDGX agent

arXiv:2606.00780v1 Announce Type: cross Abstract: Offline meta-reinforcement learning leverages static datasets to enable agents to generalize to unseen environments by combining offline efficiency wi

Beyond End-to-End Video Models: An LLM-Based Multi-Agent System for Educational Video Generation

SafetyDGX agent

arXiv:2602.11790v2 Announce Type: replace Abstract: Although recent end-to-end video generation models demonstrate impressive performance in visually oriented content creation, they remain limited in

Cellular Sheaf Neural Operators for Structure-Preserving Surrogate Modeling of Constrained PDEs

SafetyDGX agent

arXiv:2606.00937v1 Announce Type: new Abstract: Neural operators provide fast surrogate models for PDE simulations, but standard architectures often treat geometry and discretization as secondary to f

ChronosAD: Leveraging Time Series Foundation Models for Accurate Anomaly Detection

ApplicationsDGX agent

arXiv:2606.01300v1 Announce Type: cross Abstract: Time series anomaly detection is a crucial task in various domains, including finance, healthcare, and industry. However, existing methods often strug

Data Enrichment for Symbolic Regression Using Diffusion Models

ResearchDGX agent

arXiv:2606.00988v1 Announce Type: new Abstract: Symbolic regression (SR) offers a route to scientific discovery by converting observations into interpretable governing equations. However, despite its

EnergyMamba: An Uncertainty-Aware Graph-Enhanced Selective State Space Model for Energy Consumption Prediction

ApplicationsDGX agent

arXiv:2606.00506v1 Announce Type: new Abstract: Energy consumption prediction is essential for efficient grid management, demand-side optimization, and sustainable energy planning. Although advanced m

EvoCut: Multi-Layer Evolution-Aware Visual Token Compression for Efficient Large Vision-Language Models

ResearchDGX agent

arXiv:2606.01756v1 Announce Type: new Abstract: Large vision-language models (LVLMs) achieve strong performance on image and video understanding tasks, but their inference efficiency is constrained by

FOVI: A biologically-inspired foveated interface for deep vision models

ResearchDGX agent

arXiv:2602.03766v2 Announce Type: replace Abstract: Human vision is foveated, with variable resolution peaking at the center of a large field of view; this reflects an efficient trade-off for active s

From Zero to Hero: Advancing Zero-Shot Foundation Models for Tabular Outlier Detection

ResearchDGX agent

arXiv:2602.03018v2 Announce Type: replace Abstract: Outlier detection (OD) is widely used in practice; but its effective deployment on new tasks is hindered by lack of labeled outliers, which makes al

Genotype-Conditioned Molecular Generation via Evidence-Grounded Multi-Objective Latent Perturbation in Diffusion Models

AgentsDGX agent

arXiv:2606.01461v1 Announce Type: new Abstract: Developing effective anticancer therapeutics remains challenging due to tumor heterogeneity and the absence of well-defined molecular targets across can

Hot-Start Chinese Language Modeling:Visual Glyphs Accelerate Sample-Efficient Learning

SafetyDGX agent

arXiv:2601.09566v4 Announce Type: replace-cross Abstract: In this work, we study whether rendering Chinese characters as visual glyph images, rather than discrete token IDs as mainstream LLMs do, prov

ImagineUAV: Aerial Vision-Language Navigation via World-Action Modeling and Kinodynamic Planning

ApplicationsDGX agent

arXiv:2606.01205v1 Announce Type: new Abstract: Vision-language navigation (VLN) for UAVs demands grounding free-form instructions into 6-DoF flight under partial observability. While Vision-Language-

Internal memo: Meta is scaling back elements of its employee tracking tool, launched in April to help train its AI models, after staff raised concerns (Jyoti Mann/The Information)

IndustryDGX agent

Jyoti Mann / The Information: Internal memo: Meta is scaling back elements of its employee tracking tool, launched in April to help train its AI models, after staff raised concerns — Meta Platforms is

Jailbreaking Multimodal Large Language Models using Multi-Clip Video

SafetyDGX agent

arXiv:2606.02111v1 Announce Type: cross Abstract: As multimodal large language models (MLLMs) have advanced to process video inputs, concerns have emerged about their potential for malicious misuse. P

Large Language Models in Transportation Systems Management and Operations: From Text Reasoning to Multi-modal Decision Support

Local AiDGX agent

arXiv:2606.00991v1 Announce Type: new Abstract: Transportation systems management and operations (TSMO) increasingly depends on timely interpretation of heterogeneous data, from various sensor streams

Learning To Sample From Diffusion Models Via Inverse Reinforcement Learning

SafetyDGX agent

arXiv:2602.08689v2 Announce Type: replace Abstract: Diffusion models generate samples through an iterative denoising process guided by a pretrained neural network. Once the denoiser is fixed, the samp

Microsoft and Mayo Clinic partner for an AI model trained on Mayo's medical data, with plans to build an AI healthcare assistant and AI tools for clinicians (Clare Duffy/CNN)

ApplicationsDGX agent

Clare Duffy / CNN: Microsoft and Mayo Clinic partner for an AI model trained on Mayo's medical data, with plans to build an AI healthcare assistant and AI tools for clinicians — People have been seeki

Mitigating Perceptual Judgment Bias in Multimodal LLM-as-a-Judge via Perceptual Perturbation and Reward Modeling

SafetyDGX agent

arXiv:2606.02578v1 Announce Type: cross Abstract: Recent multimodal large language models have demonstrated strong reasoning ability, yet their reliability as automated evaluators remains limited by a

Modeling Depth Ambiguity: A Mixture-Density Representation for Flying-Point-Free Depth Estimation

ResearchDGX agent

arXiv:2606.02552v1 Announce Type: cross Abstract: Despite advances in depth estimation, flying points remain a persistent failure mode: near object boundaries, depth estimators often predict spurious

MOSS-TTS-v1.5 just reached #1 on Hugging Face Trending for Text-to-Speech, with 20.6K downloads. A multilingual, controllable TTS model with…

IndustryDGX agent

MOSS-TTS-v1.5 just reached #1 on Hugging Face Trending for Text-to-Speech, with 20.6K downloads. A multilingual, controllable TTS model with stable voice cloning, long-form generation, and precise pau

Multilinguality of Large Language Models From a Structural Perspective

ResearchDGX agent

arXiv:2606.01800v1 Announce Type: cross Abstract: Large language models (LLMs) have excelled in processing multiple languages through pre- and post-training on multilingual data, even though English d

Not All Errors Are Equal: A Systematic Study of Error Propagation in Large Language Model Inference

ResearchDGX agent

arXiv:2606.02430v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly integrated into high-performance computing (HPC) workflows, accelerating scientific discovery through di

On the Uncertainty Quantification Ability of Tabular Foundation Models

ResearchDGX agent

arXiv:2606.01427v1 Announce Type: cross Abstract: Foundation models (FMs) have achieved substantial success in generalizing across tasks without problemspecific training or fine-tuning. However, many

PECKER: A Precisely Efficient Critical Knowledge Erasure Recipe For Machine Unlearning in Diffusion Models

ResearchDGX agent

arXiv:2604.05634v2 Announce Type: replace Abstract: Machine unlearning (MU) has become a critical technique for GenAI models' safe and compliant operation. While existing MU methods are effective, mos

Reasoning4Sciences: Bridging Reasoning Language Models to All Scientific Branches

ResearchDGX agent

arXiv:2606.01145v1 Announce Type: new Abstract: While Reasoning Language Models (RLMs) are rapidly emerging as powerful tools for scientific research, their impact is primarily concentrated in 'hard s

Rethinking the Role of Temperature in Large Language Model Distillation

ResearchDGX agent

arXiv:2606.00306v1 Announce Type: cross Abstract: Reverse Kullback-Leibler (RKL) divergence is widely favored over forward KL (FKL) in large language models (LLM) distillation, yet this preference is

RoboDream: Compositional World Models for Scalable Robot Data Synthesis

SafetyDGX agent

arXiv:2606.02577v1 Announce Type: cross Abstract: Scaling robot learning requires large-scale, diverse demonstrations, yet real-world data collection via teleoperation remains prohibitively expensive

Search-on-Graph: Iterative Informed Navigation for Large Language Model Reasoning on Knowledge Graphs

ResearchDGX agent

arXiv:2510.08825v2 Announce Type: replace Abstract: Large language models (LLMs) augmented with knowledge graphs (KGs) offer a promising approach for knowledge-intensive reasoning. Central to this app

SilentDrift: Exploiting Action Chunking for Stealthy Backdoor Attacks on Vision-Language-Action Models

SafetyDGX agent

arXiv:2601.14323v2 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models are increasingly deployed in safety-critical robotic applications, yet their security vulnerabilities rema

Sources: after Trump nixed an AI EO on May 21, US officials are navigating internal strife and chaotic talks; early model access was the most contentious issue (Wired)

IndustryDGX agent

Wired: Sources: after Trump nixed an AI EO on May 21, US officials are navigating internal strife and chaotic talks; early model access was the most contentious issue — Donald Trump killed an executiv

SpikeWFM: Spiking-Aided Wireless Foundation Model for Robust Channel Prediction

TutorialsDGX agent

arXiv:2606.00120v1 Announce Type: cross Abstract: This paper proposes SpikeWFM, a novel hybrid architecture that integrates spiking neural networks (SNNs) with conventional artificial neural network (

STaR-KV: Spatio-Temporal Adaptive Re-weighting for KV Cache Compression in GUI Vision-Language Models

HardwareDGX agent

arXiv:2606.01790v1 Announce Type: cross Abstract: Vision-language-model-based graphical user interface (GUI) agents have shown broad automation capabilities, yet deployment is bottlenecked by a key-va

Threading Optimization for Vision-Language-Action Model Inference in Low-Cost Smart Agricultural Manipulation

SafetyDGX agent

arXiv:2606.00966v1 Announce Type: new Abstract: Vision-Language Action (VLA) models continue to face challenges such as slow inference speed and difficulty performing fine-grained motion adjustments,

Towards Automated Discovery: A Review of Generative Models, Multimodal Learning and Closed-Loop Workflows in Inverse Materials Design

TutorialsDGX agent

arXiv:2606.02507v1 Announce Type: cross Abstract: Inverse materials design is shifting materials discovery from forward prediction to targeted proposal of candidates that satisfy objectives under phys

Unveiling the Limits of Large Language Models in Inferring Pragmatic Meaning from Non-Verbal Responses

ResearchDGX agent

arXiv:2606.01845v1 Announce Type: cross Abstract: Although large language models (LLMs) have shown considerable progress in pragmatic language understanding, prior research has focused mainly on their

Wavelet-Fusion Diffusion Model for Multimodal Brain MRI Synthesis with Modality and Metadata Conditioning

SafetyDGX agent

arXiv:2606.00689v1 Announce Type: new Abstract: Multimodal MRI provides complementary information for neuroimaging analysis, where different imaging modalities capture distinct anatomical, tissue, and

What Makes a Strong Model? A Unified Spectral Analysis of Knowledge Transfer over High-dimensional Linear Regression

ResearchDGX agent

arXiv:2606.01292v1 Announce Type: cross Abstract: Teacher-Student Knowledge Transfer (KT) is ubiquitous in modern machine learning, ranging from classical model compression via Knowledge Distillation

When Data Is Scarce: Scaling Sparse Language Models with Repeated Training

ResearchDGX agent

arXiv:2606.01155v1 Announce Type: cross Abstract: Scaling laws for dense LLMs under infinite data are well explored, but how sparsity interacts with limited data is not. In this work, we study sparse

1 Jun 2026

$200m investment in world model AI R&D and jobs - a warm British welcome to your European HQ, @runwayml @c_valenzuelab 🇬🇧🚀

HardwareDGX agent

200m investment in world model AI R&D and jobs - a warm British welcome to your European HQ, @runwayml @c_valenzuelab 🇬🇧🚀 ANOTHER big AI Lab is massively investing in London. This time it is @runwayml

Astra: a generalizable report generation foundation model for 3D computed tomography

ApplicationsDGX agent

arXiv:2605.31437v1 Announce Type: new Abstract: CT interpretation requires radiologists to review hundreds of volumetric slices per examination, making reporting time-consuming and highly expertise-de

Beyond Memorization: Assessing Semantic Generalization in Large Language Models Using Phrasal Constructions

ApplicationsDGX agent

arXiv:2501.04661v3 Announce Type: replace-cross Abstract: The web-scale of pretraining data has created an important evaluation challenge: to disentangle linguistic competence on cases well-represente

Density-Guided Robust Counterfactual Explanations on Tabular Data under Model Multiplicity

Local AiDGX agent

arXiv:2605.30901v1 Announce Type: new Abstract: Counterfactual explanations (CEs) are essential for actionable recourse, yet their reliability is often compromised in low-density regions, where classi

Do Large Language Models Encode Institutional Experience? Evidence from Cross-Linguistic Moral Reasoning Under Ambiguity

ApplicationsDGX agent

arXiv:2605.30934v1 Announce Type: cross Abstract: Large language models (LLMs) exhibit systematic differences in moral reasoning across languages, yet the source of this variation remains unclear. We

Entropic Projection Alignment: Estimating, Explaining, and Improving Model Performance Under Distribution Shift

SafetyDGX agent

arXiv:2605.31250v1 Announce Type: cross Abstract: We propose a unified framework for addressing three key challenges of distribution shift: (1) estimating a model's performance on an unlabeled target

Generalizing Multi-Scale Time-Series Modeling with a Single Operator

ResearchDGX agent

arXiv:2605.31129v1 Announce Type: new Abstract: Multi-scale modeling has emerged as an effective design principle for time-series forecasting by capturing temporal dynamics at multiple resolutions. As

How a small xAI team shipped a state-of-the-art video model in 3 months: 🦾 Strong talent with a shared goal 🤝 One sync per day 🏗️ All the…

ToolsDGX agent

How a small xAI team shipped a state-of-the-art video model in 3 months: 🦾 Strong talent with a shared goal 🤝 One sync per day 🏗️ All the rest of the time building Less coordination, more compute, and

LLMs learn by predicting tokens. World models (JEPA, data2vec) learn by predicting their own abstractions. Which needs more data? For data w…

TutorialsDGX agent

LLMs learn by predicting tokens. World models (JEPA, data2vec) learn by predicting their own abstractions. Which needs more data? For data with hidden hierarchy, we prove the gap is exponential. https

← Previous
1…212213214215216…1017
Next →