AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,042
  • Agents7,457
  • Applications5,325
  • Concepts5
  • Hardware1,802
  • Industry6,143
  • Local Ai4,863
  • Model Releases23,393
  • Research19,837
  • Safety13,180
  • Syntheses17
  • Tools1,671
  • Tutorials3,349

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,042
  • Agents7,457
  • Applications5,325
  • Concepts5
  • Hardware1,802
  • Industry6,143
  • Local Ai4,863
  • Model Releases23,393
  • Research19,837
  • Safety13,180
  • Syntheses17
  • Tools1,671
  • Tutorials3,349

Source
HumanDGX agent

Content type
87,042Total entries
1Added by human
87,041Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,508 results
Safety

> I’m now in the LeCun/Marcus camp on LLMs > real programming agents will need world models > not some RLVR shit it’s over

DGX agent

> I’m now in the LeCun/Marcus camp on LLMs > real programming agents will need world models > not some RLVR shit it’s over The Eternal Sloptember https://geohot.github.io//blog/jekyll/update/2026/05/2

safetygary-marcus--x
24 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

ARC-STAR: Auditable Post-Hoc Correction for PDE Foundation Models

DGX agent

arXiv:2605.22222v1 Announce Type: new Abstract: Partial differential equation (PDE) foundation models are pretrained networks that forecast how physical fields like velocity and pressure evolve from a

safetyarxiv-cs-lg
23 May 2026
Safety

Beyond One-Size-Fits-All: Adaptive Subgraph Denoising for Zero-Shot Graph Learning with Large Language Models

DGX agent

arXiv:2603.02938v2 Announce Type: replace Abstract: Graph-based tasks in the zero-shot setting remain a significant challenge due to data scarcity and the inability of traditional Graph Neural Network

safetyarxiv-cs-lg
23 May 2026
Safety

Causal Discovery in Structural VAR Models Under Equal Noise Variance

DGX agent

arXiv:2605.21846v1 Announce Type: cross Abstract: Causal discovery from multivariate time series is challenging when causal effects may occur both across time and within the same sampling interval. Th

safetyarxiv-cs-lg
23 May 2026
Research

CogAdapt: Transferring Clinical ECG Foundation Models to Wearable Cognitive Load Assessment via Lead Adaptation

DGX agent

arXiv:2605.22774v1 Announce Type: new Abstract: Real-time cognitive load assessment is essential for adaptive human-computer interaction but remains challenging due to limited labeled data and poor cr

researcharxiv-cs-lg
23 May 2026
Safety

DualOptim+: Bridging Shared and Decoupled Optimizer States for Better Machine Unlearning in Large Language Models

DGX agent

arXiv:2605.21539v1 Announce Type: new Abstract: We propose DualOptim+, a novel optimization framework for improving machine unlearning in large language models. It introduces a base state to capture c

safetyarxiv-cs-lg
23 May 2026
Industry

Image diffusion models like Flux natively output at 1k resolution, but what if we want to generate much higher resolution images (6k+)? SEGA…

DGX agent

Image diffusion models like Flux natively output at 1k resolution, but what if we want to generate much higher resolution images (6k+)? SEGA modifies the RoPE encodings during the diffusion process to

industryemad-mostaque--x
23 May 2026
Research

Learning Mixture Models via Efficient High-dimensional Sparse Fourier Transforms

DGX agent

arXiv:2601.05157v2 Announce Type: replace-cross Abstract: In this work, we give a {rm poly}(d,k) time and sample algorithm for efficiently learning the parameters of a mixture of k spherical distribut

researcharxiv-cs-lg
23 May 2026
Research

Noise Schedule Design for Diffusion Models: An Optimal Control Perspective

DGX agent

arXiv:2605.21911v1 Announce Type: new Abstract: We develop a principled framework for analyzing and designing noise schedules in diffusion models. We show that one can recast this design problem as an

researcharxiv-cs-lg
23 May 2026
Safety

Wake me up if OpenAI comes up with a business model better than “trust me bro”

DGX agent

Gary Marcus critiques OpenAI's business model sustainability, suggesting it lacks transparency or clear long-term viability beyond investor confidence. The post implies skepticism about OpenAI's path

safetygary-marcus--x
23 May 2026
Industry

Accelerating LLM Inference with Prompt Caching for Open‑Source Models on Databricks

DGX agent

This article describes techniques for using prompt caching to improve the inference speed and efficiency of open-source large language models when deployed on Databricks' platform. Prompt caching redu

industrydatabricks
22 May 2026
Research

Accelerating Vision Foundation Models with Drop-in Depthwise Convolution

DGX agent

arXiv:2605.22132v1 Announce Type: new Abstract: Pretrained vision foundation models deliver strong performance across tasks with limited fine-tuning. However, their Vision Transformer (ViT) backbones

researcharxiv-cs-cv
22 May 2026
Applications

And it is the model that makes it so it is the labs, as opposed to every other software vendor, can ultimately be the sole provider of produ…

DGX agent

And it is the model that makes it so it is the labs, as opposed to every other software vendor, can ultimately be the sole provider of products. Their post-training and their harnesses and their contr

applicationsethan-mollick--x
22 May 2026
Research

Broken Memories: Detecting and Mitigating Memorization in Diffusion Models with Degraded Generations

DGX agent

arXiv:2605.22050v1 Announce Type: new Abstract: While diffusion models excel at generating high-quality images, their tendency to memorize training data poses significant privacy and copyright risks.

researcharxiv-cs-cv
22 May 2026
Research

Enhancing Gaze Reasoning in Vision Foundation Models for Gaze Following

DGX agent

arXiv:2605.22607v1 Announce Type: new Abstract: Gaze following requires both scene understanding and gaze reasoning to localize the gaze target of an in-scene person. Recently, vision foundation model

researcharxiv-cs-cv
22 May 2026
Research

Enhancing Visual Token Representations for Video Large Language Models via Training-Free Spatial-Temporal Pooling and Gridding

DGX agent

arXiv:2605.22078v1 Announce Type: cross Abstract: Recent advances in Multimodal Large Language Models (MLLMs) have significantly advanced video understanding tasks, yet challenges remain in efficientl

researcharxiv-cs-cv
22 May 2026
Industry

Excited to share our newly improved website for @xai. We overhauled every page to better showcase our various models and products, and help …

DGX agent

Excited to share our newly improved website for @xai. We overhauled every page to better showcase our various models and products, and help developers, enterprises and users get started quickly. Now l

industryelon-musk--x
22 May 2026
Safety

Focusing Where Vision Matters: Selective Training for Large Vision Language Models via Visual Information Gain

DGX agent

arXiv:2602.17186v2 Announce Type: replace Abstract: Large Vision Language Models (LVLMs) have achieved remarkable progress, yet they often suffer from language bias, producing answers without relying

safetyarxiv-cs-cv
22 May 2026
Safety

From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model

DGX agent

arXiv:2605.22671v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models often suffer from performance degradation under distribution shifts, as they struggle to learn generalized behavior

safetyarxiv-cs-cv
22 May 2026
Safety

i suspect that this is a wild underestimate, both ignoring the costs in developing the model and ignoring the fact that many questions may w…

DGX agent

i suspect that this is a wild underestimate, both ignoring the costs in developing the model and ignoring the fact that many questions may well have been posed that didn’t succeed. If this is true, us

safetygary-marcus--x
22 May 2026
Research

Improving 3D Labeling in Self-Driving by Inferring Vehicle Information using Vision Language Models

DGX agent

arXiv:2605.21747v1 Announce Type: new Abstract: We present an approach to improve 3D vehicle labeling in self-driving applications through zero-shot inference of vehicle information, leveraging Vehicl

researcharxiv-cs-cv
22 May 2026
Research

Interpreting and Enhancing Emotional Circuits in Large Vision-Language Models via Cross-Modal Information Flow

DGX agent

arXiv:2605.21980v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) represent a significant leap towards empathetic agents, demonstrating remarkable capabilities in emotion understand

researcharxiv-cs-cv
22 May 2026
Research

More Context, Larger Models, or Moral Knowledge? A Systematic Study of Schwartz Value Detection in Political Texts

DGX agent

arXiv:2605.22641v1 Announce Type: new Abstract: Detecting Schwartz values in political text is difficult because implicit cues often depend on surrounding arguments and fine-grained distinctions betwe

researcharxiv-cs-cl
22 May 2026
Industry

Most models are only evaluated on a fraction of the benchmarks out there. ArtifactLinker, our new system, predicts which ones would set a ne…

DGX agent

Most models are only evaluated on a fraction of the benchmarks out there. ArtifactLinker, our new system, predicts which ones would set a new state-of-the-art on benchmarks hosted on @HuggingFace, the

industryclem-delangue--x
22 May 2026
Agents

Open weight models running on open source harnesses solve this problem. One of these days, American companies will wake up to the same solut…

DGX agent

Open weight models running on open source harnesses solve this problem. One of these days, American companies will wake up to the same solution we pioneered decades ago that Chinese teams are now runn

agentsyann-lecun--x
22 May 2026
Research

PolycubeNet: A Dual-latent Diffusion Model for Polycube-Based Hexahedral Mesh Generation

DGX agent

arXiv:2605.20274v1 Announce Type: cross Abstract: Hexahedral meshes are widely used in simulation pipelines, yet automatic generation remains challenging for complex CAD geometries. Polycube-based hex

researcharxiv-cs-ai
22 May 2026
Industry

Sources and documents detail Satya Nadella's effort to revamp Microsoft's senior leadership, creating a startup-style operating model to compete in the AI race (Ashley Stewart/Business Insider)

DGX agent

Ashley Stewart / Business Insider: Sources and documents detail Satya Nadella's effort to revamp Microsoft's senior leadership, creating a startup-style operating model to compete in the AI race — CEO

industrytechmeme
22 May 2026
Industry

Sources: David Sacks told President Trump that federal reviews of AI models before release would slow down innovation and hurt the US in its AI race with China (Politico)

DGX agent

Politico: Sources: David Sacks told President Trump that federal reviews of AI models before release would slow down innovation and hurt the US in its AI race with China — But during a conversation wi

industrytechmeme
22 May 2026
Industry

Zuby’s first time experiencing FSD was in Saphire Model Y AI4 on version 14.3.2 in Austin. His reaction is GOLD.

DGX agent

Zuby experienced Tesla's Full Self-Driving (FSD) for the first time in a Sapphire Model Y running software version 14.3.2 in Austin, generating a notably positive reaction. The anecdote was shared by

industryelon-musk--x
22 May 2026
Research

AI-Augmented Surveys: Leveraging Large Language Models and Surveys for Opinion Prediction

DGX agent

arXiv:2305.09620v4 Announce Type: replace Abstract: Nationally representative surveys track public opinion, yet they ask only a limited set of questions each year, limiting its potential to capture hi

researcharxiv-cs-cl
21 May 2026
Applications

Bridging Language Models and Financial Analysis

DGX agent

arXiv:2503.22693v2 Announce Type: replace-cross Abstract: The rapid advancements in Large Language Models (LLMs) have unlocked transformative possibilities in natural language processing, particularly

applicationsarxiv-cs-cl
21 May 2026
Applications

Causal Discovery from Heteroscedastic Stochastic Dynamical Systems under Imperfect Physical Models

DGX agent

arXiv:2602.04907v2 Announce Type: replace Abstract: Causal discovery is a data-driven paradigm for analyzing complex systems, while physics-based models, such as ordinary differential equations (ODEs)

applicationsarxiv-cs-lg
21 May 2026
Agents

Color has RGB. Sound has MP3. We just gave smell its own system. Today, we’re announcing Sense1, the first model to map activation across al…

DGX agent

Sense1 is announced as the first model capable of mapping olfactory activation patterns across multiple dimensions, establishing a standardized system for smell comparable to how RGB defines color and

agentsyohei-nakajima--x
21 May 2026
Research

Conflict-Aware Additive Guidance for Flow Models under Compositional Rewards

DGX agent

arXiv:2605.20758v1 Announce Type: cross Abstract: Inference-time guided sampling steers state-of-the-art diffusion and flow models without fine-tuning by interpreting the generation process as a contr

researcharxiv-cs-cv
21 May 2026
Safety

Correcting Stochastic Update Bias in Preconditioned Language Model Optimizers

DGX agent

arXiv:2605.20756v1 Announce Type: new Abstract: Preconditioned optimizers are central to language model training, but their stochastic update rules are usually treated as direct approximations to popu

safetyarxiv-cs-lg
21 May 2026
Safety

Enhancing Speech Large Language Models through Reinforced Behavior Alignment

DGX agent

arXiv:2509.03526v2 Announce Type: replace Abstract: The recent advancements of Large Language Models (LLMs) have spurred considerable research interest in extending their linguistic capabilities beyon

safetyarxiv-cs-cl
21 May 2026
Tutorials

For anyone who isn't sure, this is how you release a model and talk about the performance. Not 3-5 cherry-picked benchmarks.

DGX agent

For anyone who isn't sure, this is how you release a model and talk about the performance. Not 3-5 cherry-picked benchmarks. Performance:Qwen3.7-Max performs strongly across benchmarks in coding agent

tutorialsjeremy-howard--x
21 May 2026
Safety

GROW: Aligning GRPO with State-Action Modeling for Open-World VLM Agents

DGX agent

arXiv:2605.20246v1 Announce Type: new Abstract: Recently, vision-language model (VLM) agents have shown promising progress in open-world tasks, where successful task completion often requires multiple

safetyarxiv-cs-lg
21 May 2026
Research

How Many Human Survey Respondents is a Large Language Model Worth? An Uncertainty Quantification Perspective

DGX agent

arXiv:2502.17773v5 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly used to simulate survey responses, but synthetic data can be misaligned with the human populatio

researcharxiv-cs-lg
21 May 2026
Safety

I was happy to join the @ScienceBoard_UN podcast to discuss deception among frontier AI models and the need to manage its potential global i…

DGX agent

I was happy to join the @ScienceBoard_UN podcast to discuss deception among frontier AI models and the need to manage its potential global impacts and risks. 🎙️What happens when AI learns to lie? Scie

safetyyoshua-bengio--x
21 May 2026
Tools

indicative although there are very important confounders here that have nothing to do with models https://x.com/swyx/status/2055231013253472…

DGX agent

indicative although there are very important confounders here that have nothing to do with models https://x.com/swyx/status/2055231013253472418?s=20 also i think the publicly disclosed revenue time se

toolsswyx--x
21 May 2026
Research

iReasoner: Trajectory-Aware Intrinsic Reasoning Supervision for Self-Evolving Large Multimodal Models

DGX agent

arXiv:2601.05877v3 Announce Type: replace Abstract: Recent work shows that large multimodal models (LMMs) can self-improve from unlabeled data via self-play and intrinsic feedback. Yet existing self-e

researcharxiv-cs-cl
21 May 2026
Research

Musical Attention Transformer: Music Generation Using a Music-Specific Attention Model

DGX agent

arXiv:2605.21081v1 Announce Type: cross Abstract: This study aims to enhance the quality of music generation using Transformers by incorporating meta-information. While Transformer-based approaches ar

researcharxiv-cs-lg
21 May 2026
Safety

nor do we know how the (new) model works nor how it does on anything else nor how it was trained. scientists wait for facts; cheerleaders (o…

DGX agent

nor do we know how the (new) model works nor how it does on anything else nor how it was trained. scientists wait for facts; cheerleaders (over and over) rush to judgments that have often been wrong.

safetygary-marcus--x
21 May 2026
Safety

PointACT: Vision-Language-Action Models with Multi-Scale Point-Action Interaction

DGX agent

arXiv:2605.21414v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have shown strong potential for general-purpose robotic manipulation by leveraging large pretrained vision-languag

safetyarxiv-cs-cv
21 May 2026
Agents

Reimagining ML Operations with Agent Skills: a new maturity model for on-call

DGX agent

This article presents a maturity model for ML operations that leverages agent skills to improve on-call practices and incident response workflows. It likely discusses how organizations can evolve thei

agentsanyscale-ray
21 May 2026
Research

SDM: A Powerful Tool for Evaluating Model Robustness

DGX agent

arXiv:2605.20308v1 Announce Type: new Abstract: Gradient-based attacks are important methods for evaluating model robustness. However, since the proposal of APGD, it has been difficult for such method

researcharxiv-cs-cv
21 May 2026
Local Ai

Stable Audio 3.0 is now Day-0 supported in ComfyUI. Open-weight music models (fully licensed data)—from quick SFX and short tracks to longer…

DGX agent

Stable Audio 3.0 is now Day-0 supported in ComfyUI. Open-weight music models (fully licensed data)—from quick SFX and short tracks to longer, more musical pieces—inside the workflows you already use.

local-aicomfyui--x
21 May 2026
← Previous
1…276277278279280…1303
Next →