AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

85,136Total entries
1Added by human
85,135Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,002 results
25 May 2026

The Grok Build team is one of the most productive teams I have worked with. I need something to adopt my diverse workflow from model trainin…

ApplicationsDGX agent

The Grok Build team is one of the most productive teams I have worked with. I need something to adopt my diverse workflow from model training to validating the Cybercab manufacturing process. Their da

V-VLAPS: Value-Guided Planning for Vision-Language-Action Models

SafetyDGX agent

arXiv:2601.00969v2 Announce Type: replace-cross Abstract: Vision-language-action (VLA) models provide strong action priors for robotic manipulation, but their reactive behavior can fail under distribu

We built Aleph as one of the first generalist text-based editing models. Aleph 2.0 takes it further: visual prompting for precise control, l…

Industry
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

We built Aleph as one of the first generalist text-based editing models. Aleph 2.0 takes it further: visual prompting for precise control, longer videos (up to 30 seconds), and yes, multi-shot support

24 May 2026

> I’m now in the LeCun/Marcus camp on LLMs > real programming agents will need world models > not some RLVR shit it’s over

SafetyDGX agent

> I’m now in the LeCun/Marcus camp on LLMs > real programming agents will need world models > not some RLVR shit it’s over The Eternal Sloptember https://geohot.github.io//blog/jekyll/update/2026/05/2

23 May 2026

ARC-STAR: Auditable Post-Hoc Correction for PDE Foundation Models

SafetyDGX agent

arXiv:2605.22222v1 Announce Type: new Abstract: Partial differential equation (PDE) foundation models are pretrained networks that forecast how physical fields like velocity and pressure evolve from a

Beyond One-Size-Fits-All: Adaptive Subgraph Denoising for Zero-Shot Graph Learning with Large Language Models

SafetyDGX agent

arXiv:2603.02938v2 Announce Type: replace Abstract: Graph-based tasks in the zero-shot setting remain a significant challenge due to data scarcity and the inability of traditional Graph Neural Network

Causal Discovery in Structural VAR Models Under Equal Noise Variance

SafetyDGX agent

arXiv:2605.21846v1 Announce Type: cross Abstract: Causal discovery from multivariate time series is challenging when causal effects may occur both across time and within the same sampling interval. Th

CogAdapt: Transferring Clinical ECG Foundation Models to Wearable Cognitive Load Assessment via Lead Adaptation

ResearchDGX agent

arXiv:2605.22774v1 Announce Type: new Abstract: Real-time cognitive load assessment is essential for adaptive human-computer interaction but remains challenging due to limited labeled data and poor cr

DualOptim+: Bridging Shared and Decoupled Optimizer States for Better Machine Unlearning in Large Language Models

SafetyDGX agent

arXiv:2605.21539v1 Announce Type: new Abstract: We propose DualOptim+, a novel optimization framework for improving machine unlearning in large language models. It introduces a base state to capture c

Image diffusion models like Flux natively output at 1k resolution, but what if we want to generate much higher resolution images (6k+)? SEGA…

IndustryDGX agent

Image diffusion models like Flux natively output at 1k resolution, but what if we want to generate much higher resolution images (6k+)? SEGA modifies the RoPE encodings during the diffusion process to

Learning Mixture Models via Efficient High-dimensional Sparse Fourier Transforms

ResearchDGX agent

arXiv:2601.05157v2 Announce Type: replace-cross Abstract: In this work, we give a {rm poly}(d,k) time and sample algorithm for efficiently learning the parameters of a mixture of k spherical distribut

Noise Schedule Design for Diffusion Models: An Optimal Control Perspective

ResearchDGX agent

arXiv:2605.21911v1 Announce Type: new Abstract: We develop a principled framework for analyzing and designing noise schedules in diffusion models. We show that one can recast this design problem as an

Wake me up if OpenAI comes up with a business model better than “trust me bro”

SafetyDGX agent

Gary Marcus critiques OpenAI's business model sustainability, suggesting it lacks transparency or clear long-term viability beyond investor confidence. The post implies skepticism about OpenAI's path

22 May 2026

Accelerating LLM Inference with Prompt Caching for Open‑Source Models on Databricks

IndustryDGX agent

This article describes techniques for using prompt caching to improve the inference speed and efficiency of open-source large language models when deployed on Databricks' platform. Prompt caching redu

Accelerating Vision Foundation Models with Drop-in Depthwise Convolution

ResearchDGX agent

arXiv:2605.22132v1 Announce Type: new Abstract: Pretrained vision foundation models deliver strong performance across tasks with limited fine-tuning. However, their Vision Transformer (ViT) backbones

And it is the model that makes it so it is the labs, as opposed to every other software vendor, can ultimately be the sole provider of produ…

ApplicationsDGX agent

And it is the model that makes it so it is the labs, as opposed to every other software vendor, can ultimately be the sole provider of products. Their post-training and their harnesses and their contr

Broken Memories: Detecting and Mitigating Memorization in Diffusion Models with Degraded Generations

ResearchDGX agent

arXiv:2605.22050v1 Announce Type: new Abstract: While diffusion models excel at generating high-quality images, their tendency to memorize training data poses significant privacy and copyright risks.

Enhancing Gaze Reasoning in Vision Foundation Models for Gaze Following

ResearchDGX agent

arXiv:2605.22607v1 Announce Type: new Abstract: Gaze following requires both scene understanding and gaze reasoning to localize the gaze target of an in-scene person. Recently, vision foundation model

Enhancing Visual Token Representations for Video Large Language Models via Training-Free Spatial-Temporal Pooling and Gridding

ResearchDGX agent

arXiv:2605.22078v1 Announce Type: cross Abstract: Recent advances in Multimodal Large Language Models (MLLMs) have significantly advanced video understanding tasks, yet challenges remain in efficientl

Excited to share our newly improved website for @xai. We overhauled every page to better showcase our various models and products, and help …

IndustryDGX agent

Excited to share our newly improved website for @xai. We overhauled every page to better showcase our various models and products, and help developers, enterprises and users get started quickly. Now l

Focusing Where Vision Matters: Selective Training for Large Vision Language Models via Visual Information Gain

SafetyDGX agent

arXiv:2602.17186v2 Announce Type: replace Abstract: Large Vision Language Models (LVLMs) have achieved remarkable progress, yet they often suffer from language bias, producing answers without relying

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model

SafetyDGX agent

arXiv:2605.22671v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models often suffer from performance degradation under distribution shifts, as they struggle to learn generalized behavior

i suspect that this is a wild underestimate, both ignoring the costs in developing the model and ignoring the fact that many questions may w…

SafetyDGX agent

i suspect that this is a wild underestimate, both ignoring the costs in developing the model and ignoring the fact that many questions may well have been posed that didn’t succeed. If this is true, us

Improving 3D Labeling in Self-Driving by Inferring Vehicle Information using Vision Language Models

ResearchDGX agent

arXiv:2605.21747v1 Announce Type: new Abstract: We present an approach to improve 3D vehicle labeling in self-driving applications through zero-shot inference of vehicle information, leveraging Vehicl

Interpreting and Enhancing Emotional Circuits in Large Vision-Language Models via Cross-Modal Information Flow

ResearchDGX agent

arXiv:2605.21980v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) represent a significant leap towards empathetic agents, demonstrating remarkable capabilities in emotion understand

More Context, Larger Models, or Moral Knowledge? A Systematic Study of Schwartz Value Detection in Political Texts

ResearchDGX agent

arXiv:2605.22641v1 Announce Type: new Abstract: Detecting Schwartz values in political text is difficult because implicit cues often depend on surrounding arguments and fine-grained distinctions betwe

Most models are only evaluated on a fraction of the benchmarks out there. ArtifactLinker, our new system, predicts which ones would set a ne…

IndustryDGX agent

Most models are only evaluated on a fraction of the benchmarks out there. ArtifactLinker, our new system, predicts which ones would set a new state-of-the-art on benchmarks hosted on @HuggingFace, the

Open weight models running on open source harnesses solve this problem. One of these days, American companies will wake up to the same solut…

AgentsDGX agent

Open weight models running on open source harnesses solve this problem. One of these days, American companies will wake up to the same solution we pioneered decades ago that Chinese teams are now runn

PolycubeNet: A Dual-latent Diffusion Model for Polycube-Based Hexahedral Mesh Generation

ResearchDGX agent

arXiv:2605.20274v1 Announce Type: cross Abstract: Hexahedral meshes are widely used in simulation pipelines, yet automatic generation remains challenging for complex CAD geometries. Polycube-based hex

Sources and documents detail Satya Nadella's effort to revamp Microsoft's senior leadership, creating a startup-style operating model to compete in the AI race (Ashley Stewart/Business Insider)

IndustryDGX agent

Ashley Stewart / Business Insider: Sources and documents detail Satya Nadella's effort to revamp Microsoft's senior leadership, creating a startup-style operating model to compete in the AI race — CEO

Sources: David Sacks told President Trump that federal reviews of AI models before release would slow down innovation and hurt the US in its AI race with China (Politico)

IndustryDGX agent

Politico: Sources: David Sacks told President Trump that federal reviews of AI models before release would slow down innovation and hurt the US in its AI race with China — But during a conversation wi

Zuby’s first time experiencing FSD was in Saphire Model Y AI4 on version 14.3.2 in Austin. His reaction is GOLD.

IndustryDGX agent

Zuby experienced Tesla's Full Self-Driving (FSD) for the first time in a Sapphire Model Y running software version 14.3.2 in Austin, generating a notably positive reaction. The anecdote was shared by

21 May 2026

AI-Augmented Surveys: Leveraging Large Language Models and Surveys for Opinion Prediction

ResearchDGX agent

arXiv:2305.09620v4 Announce Type: replace Abstract: Nationally representative surveys track public opinion, yet they ask only a limited set of questions each year, limiting its potential to capture hi

Bridging Language Models and Financial Analysis

ApplicationsDGX agent

arXiv:2503.22693v2 Announce Type: replace-cross Abstract: The rapid advancements in Large Language Models (LLMs) have unlocked transformative possibilities in natural language processing, particularly

Causal Discovery from Heteroscedastic Stochastic Dynamical Systems under Imperfect Physical Models

ApplicationsDGX agent

arXiv:2602.04907v2 Announce Type: replace Abstract: Causal discovery is a data-driven paradigm for analyzing complex systems, while physics-based models, such as ordinary differential equations (ODEs)

Color has RGB. Sound has MP3. We just gave smell its own system. Today, we’re announcing Sense1, the first model to map activation across al…

AgentsDGX agent

Sense1 is announced as the first model capable of mapping olfactory activation patterns across multiple dimensions, establishing a standardized system for smell comparable to how RGB defines color and

Conflict-Aware Additive Guidance for Flow Models under Compositional Rewards

ResearchDGX agent

arXiv:2605.20758v1 Announce Type: cross Abstract: Inference-time guided sampling steers state-of-the-art diffusion and flow models without fine-tuning by interpreting the generation process as a contr

Correcting Stochastic Update Bias in Preconditioned Language Model Optimizers

SafetyDGX agent

arXiv:2605.20756v1 Announce Type: new Abstract: Preconditioned optimizers are central to language model training, but their stochastic update rules are usually treated as direct approximations to popu

Enhancing Speech Large Language Models through Reinforced Behavior Alignment

SafetyDGX agent

arXiv:2509.03526v2 Announce Type: replace Abstract: The recent advancements of Large Language Models (LLMs) have spurred considerable research interest in extending their linguistic capabilities beyon

For anyone who isn't sure, this is how you release a model and talk about the performance. Not 3-5 cherry-picked benchmarks.

TutorialsDGX agent

For anyone who isn't sure, this is how you release a model and talk about the performance. Not 3-5 cherry-picked benchmarks. Performance:Qwen3.7-Max performs strongly across benchmarks in coding agent

GROW: Aligning GRPO with State-Action Modeling for Open-World VLM Agents

SafetyDGX agent

arXiv:2605.20246v1 Announce Type: new Abstract: Recently, vision-language model (VLM) agents have shown promising progress in open-world tasks, where successful task completion often requires multiple

How Many Human Survey Respondents is a Large Language Model Worth? An Uncertainty Quantification Perspective

ResearchDGX agent

arXiv:2502.17773v5 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly used to simulate survey responses, but synthetic data can be misaligned with the human populatio

I was happy to join the @ScienceBoard_UN podcast to discuss deception among frontier AI models and the need to manage its potential global i…

SafetyDGX agent

I was happy to join the @ScienceBoard_UN podcast to discuss deception among frontier AI models and the need to manage its potential global impacts and risks. 🎙️What happens when AI learns to lie? Scie

indicative although there are very important confounders here that have nothing to do with models https://x.com/swyx/status/2055231013253472…

ToolsDGX agent

indicative although there are very important confounders here that have nothing to do with models https://x.com/swyx/status/2055231013253472418?s=20 also i think the publicly disclosed revenue time se

iReasoner: Trajectory-Aware Intrinsic Reasoning Supervision for Self-Evolving Large Multimodal Models

ResearchDGX agent

arXiv:2601.05877v3 Announce Type: replace Abstract: Recent work shows that large multimodal models (LMMs) can self-improve from unlabeled data via self-play and intrinsic feedback. Yet existing self-e

Musical Attention Transformer: Music Generation Using a Music-Specific Attention Model

ResearchDGX agent

arXiv:2605.21081v1 Announce Type: cross Abstract: This study aims to enhance the quality of music generation using Transformers by incorporating meta-information. While Transformer-based approaches ar

nor do we know how the (new) model works nor how it does on anything else nor how it was trained. scientists wait for facts; cheerleaders (o…

SafetyDGX agent

nor do we know how the (new) model works nor how it does on anything else nor how it was trained. scientists wait for facts; cheerleaders (over and over) rush to judgments that have often been wrong.

PointACT: Vision-Language-Action Models with Multi-Scale Point-Action Interaction

SafetyDGX agent

arXiv:2605.21414v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have shown strong potential for general-purpose robotic manipulation by leveraging large pretrained vision-languag

Reimagining ML Operations with Agent Skills: a new maturity model for on-call

AgentsDGX agent

This article presents a maturity model for ML operations that leverages agent skills to improve on-call practices and incident response workflows. It likely discusses how organizations can evolve thei

SDM: A Powerful Tool for Evaluating Model Robustness

ResearchDGX agent

arXiv:2605.20308v1 Announce Type: new Abstract: Gradient-based attacks are important methods for evaluating model robustness. However, since the proposal of APGD, it has been difficult for such method

Stable Audio 3.0 is now Day-0 supported in ComfyUI. Open-weight music models (fully licensed data)—from quick SFX and short tracks to longer…

Local AiDGX agent

Stable Audio 3.0 is now Day-0 supported in ComfyUI. Open-weight music models (fully licensed data)—from quick SFX and short tracks to longer, more musical pieces—inside the workflows you already use.

Towards Context-Invariant Safety Alignment for Large Language Models

SafetyDGX agent

arXiv:2605.20994v1 Announce Type: new Abstract: Preference-based post-training aligns LLMs with human intent, yet safety behavior often remains brittle. A model may refuse a harmful request in a stand

while everyone seems to be focused on text, vision, and audio, i'm excited to be backing a company building a foundational model for smell …

AgentsDGX agent

while everyone seems to be focused on text, vision, and audio, i'm excited to be backing a company building a foundational model for smell 👃 that's right, they've: - mapped how olfactory receptors rea

You Don't Need Attention: Gated Convolutional Modeling for Watch-Based Fall Detection

Local AiDGX agent

arXiv:2605.20275v1 Announce Type: new Abstract: Existing deep learning approaches for wearable fall detection systems rely on self-attention mechanisms that impose quadratic computational overhead, di

20 May 2026

3D Modeling and Automated Measurement of Concrete Cracks via Segment Anything Refinement and Visual Inertial LiDAR Fusion

Local AiDGX agent

arXiv:2501.09203v2 Announce Type: replace Abstract: Visual-Spatial Systems has become increasingly essential in concrete crack inspection. However, existing methods often lacks adaptability to diverse

a general-purpose model solved a major open problem in mathematics. we'll be saying this a lot over the coming years, but this is a kinda bi…

IndustryDGX agent

a general-purpose model solved a major open problem in mathematics. we'll be saying this a lot over the coming years, but this is a kinda big milestone. i'm very excited for AI to greatly extend our u

Artificial Phantasia: Emergent Mental Imagery in Large Language Models

ResearchDGX agent

arXiv:2509.23108v2 Announce Type: replace Abstract: Can visual imagery be driven solely by language? This idea goes against cognitive science's traditional view that visual mental imagery is only poss

BabyMamba-HAR: Lightweight Selective State Space Models for Efficient Human Activity Recognition on Resource Constrained Devices

Local AiDGX agent

arXiv:2602.09872v2 Announce Type: replace Abstract: Human activity recognition (HAR) on resource constrained devices requires high accuracy across diverse sensor setups. Selective state space models (

Boosting Text-to-Image Diffusion Models via Core Token Attention-Based Seed Selection

SafetyDGX agent

arXiv:2605.19532v1 Announce Type: new Abstract: Text-to-image diffusion models can synthesize high-quality images, yet the outcome is notoriously sensitive to the random seed: different initial seeds

Breaking Modality Heterogeneity in Low-Bit Quantization for Large Vision-Language Models

ResearchDGX agent

arXiv:2605.19929v1 Announce Type: cross Abstract: Low-bit post-training quantization (PTQ) is a pivotal technique for deploying Vision-Language Models (VLMs) on resource-constrained devices. However,

← Previous
1…215216217218219…1017
Next →