AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

85,136Total entries
1Added by human
85,135Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,002 results
6 Jun 2026

GOTabPFN: From Feature Ordering to Compact Tokenization for Tabular Foundation Models on High-Dimensional Data

TutorialsDGX agent

arXiv:2606.05441v1 Announce Type: cross Abstract: We investigate how to make small tabular foundation models effective for High-Dimensional, Low-Sample Size (HDLSS) tabular prediction without retraini

Gradient descent at the Edge of Stability: free energy model and kinetic description of the two-layer network

ResearchDGX agent

arXiv:2606.05326v1 Announce Type: cross Abstract: We study the dynamics of gradient descent in the Edge of Stability regime, where the learning rate is large enough to induce persistent oscillations i

Plug-and-Play Guidance for Discrete Diffusion Models via Gradient-Informed Logit Correction

ResearchDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.06303v1 Announce Type: cross Abstract: Controllable generation with discrete diffusion models is often hindered by high computational overhead or the need for retraining. In this paper, we

The Score Hamiltonian: Mapping Diffusion Models to Adiabatic Transport

ResearchDGX agent

arXiv:2606.05217v1 Announce Type: cross Abstract: We exhibit an exact correspondence between sampling with score-based diffusion models and adiabatic transport of ground states for a family of Schrodi

This analysis from yesterday looks even more apt today — especially after we heard that the model providers have gone to Washington in searc…

AgentsDGX agent

This analysis from yesterday looks even more apt today — especially after we heard that the model providers have gone to Washington in search of handouts. ⚠️Why didn’t the hyperscalers wait until afte

5 Jun 2026

A Model of Multi-turn Human Persuadability Using Probabilistic Belief Tracing

ResearchDGX agent

arXiv:2606.05330v1 Announce Type: new Abstract: Large language models can shift human beliefs across high-stakes domains, but most persuasion studies rely on pre/post belief change. These endpoint mea

Adversarial Attacks Already Tell the Answer: Directional Bias-Guided Test-time Defense for Vision-Language Models

SafetyDGX agent

arXiv:2606.06186v1 Announce Type: new Abstract: Vision-Language Models (VLMs), such as CLIP, have shown strong zero-shot generalization but remain highly vulnerable to adversarial perturbations, posin

Amortized Nonlinear Model Predictive Control

ResearchDGX agent

arXiv:2606.05840v1 Announce Type: cross Abstract: Nonlinear Model Predictive Control requires solving a constrained nonlinear program (NLP) in real-time at every sampling instant, a computational bott

Can Language Models Learn to Listen?

TutorialsDGX agent

arXiv:2308.10897v2 Announce Type: replace Abstract: We present a framework for generating appropriate facial responses from a listener in dyadic social interactions based on the speaker's words. Given

Diff-CA: Separating Common and Salient Factors with Diffusion Models

TutorialsDGX agent

arXiv:2606.06120v1 Announce Type: new Abstract: Contrastive Analysis aims to separate factors that are common between two data distributions from those that are salient to only one of them. Existing c

Epistemic Injustice in Language Models: An Audit of Pretraining Filters and Guardrails

SafetyDGX agent

arXiv:2606.05936v1 Announce Type: new Abstract: Modern language models rely on pretraining filters to remove undesirable content from training corpora and inference-time guardrails to suppress undesir

Flash-WAM: Modality-Aware Distillation for World Action Models

HardwareDGX agent

arXiv:2606.05254v1 Announce Type: cross Abstract: World-action models (WAMs) jointly generate future video and robot actions through iterative diffusion, achieving strong performance on manipulation b

Introducing Ideogram 4 from @ideogram_ai on Together AI, an open image model built for design with strong text rendering, layout control, an…

ApplicationsDGX agent

Introducing Ideogram 4 from @ideogram_ai on Together AI, an open image model built for design with strong text rendering, layout control, and native 2K image generation. AI natives can now use Ideogra

Learning Self-Correction in Vision-Language Models via Rollout Augmentation

TutorialsDGX agent

arXiv:2602.08503v2 Announce Type: replace-cross Abstract: Self-correction is essential for solving complex reasoning problems in vision-language models (VLMs). However, existing reinforcement learning

Leveraging Large Language Models for Generating Research Topic Ontologies: A Multi-Disciplinary Study

ResearchDGX agent

arXiv:2508.20693v2 Announce Type: replace-cross Abstract: Ontologies and taxonomies of research fields are critical for managing and organising scientific knowledge, as they facilitate efficient class

MS-DKC: A Dataset Knowledge Card Framework for Designing and Adapting Medical Image Segmentation Models

ResearchDGX agent

arXiv:2606.06103v1 Announce Type: new Abstract: Medical image segmentation is often framed as a search for stronger architectures, but this can obscure a more fundamental question: what does the datas

OpenAI confirms it will comply with President Trump's EO that asks AI companies to allow the US government to assess their models' capabilities before release (Michael Considine/CNBC)

IndustryDGX agent

Michael Considine / CNBC: OpenAI confirms it will comply with President Trump's EO that asks AI companies to allow the US government to assess their models' capabilities before release — OpenAI has co

ReCache: Learning Budget-Aware Caching Schedules for Diffusion Models via REINFORCE

SafetyDGX agent

arXiv:2606.06060v1 Announce Type: new Abstract: Modern diffusion models generate high-quality images and videos, but their iterative denoising process makes inference expensive. Feature caching accele

Temporal Preference Concepts and their Functions in a Large Language Model

Local AiDGX agent

arXiv:2606.05194v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly being deployed to make decisions that require trading off near-term gains against long-term consequences

Wow Ideogram-4.0 is immediately going into my @ComfyUI library of models. Ideogram even joins the top 10 all-around image leaderboard joinin…

Local AiDGX agent

Wow Ideogram-4.0 is immediately going into my @ComfyUI library of models. Ideogram even joins the top 10 all-around image leaderboard joining Microsoft, Google, Grok and OpenAI. In the Image Arena: op

4 Jun 2026

1/ 🔥 @NoPriorsPod x @LatentSpacePod chat with @SatyaNadella at @Microsoft Build. He has the sharpest mental models of any public company CE…

AgentsDGX agent

1/ 🔥 @NoPriorsPod x @LatentSpacePod chat with @SatyaNadella at @Microsoft Build. He has the sharpest mental models of any public company CEO I've interviewed. $MSFT is at its heart still a tools compa

A Systematic Analysis of Linguistic Features in AI-Generated Text Detection Across Domains and Models

ResearchDGX agent

arXiv:2606.04177v1 Announce Type: cross Abstract: Interpretable linguistic features offer a promising approach for explaining why a given text appears machine-generated, particularly for non-expert us

Activation Steering of Video Generation Models via Reduced-Order Linear Optimal Control

SafetyDGX agent

arXiv:2606.04775v1 Announce Type: cross Abstract: Text-to-video (T2V) models trained on large-scale web data can generate undesired content, motivating interventions that reduce harmful outputs withou

Beyond Symmetric Alignment: Spectral Diagnostics of Modality Imbalance in Vision-Language Models in the Medical Domain

SafetyDGX agent

arXiv:2606.04613v1 Announce Type: new Abstract: Vision-Language Models (VLMs) struggle when applied to medical image-text data, yet the tools available to diagnose this failure remain limited. Existin

Bounded Hyperbolic Tangent: A Stable and Efficient Alternative to Pre-Layer Normalization in Large Language Models

ResearchDGX agent

arXiv:2601.09719v3 Announce Type: replace-cross Abstract: Pre-Layer Normalization (Pre-LN) is the de facto choice for large language models (LLMs) and is crucial for stable pretraining and effective t

Culturally Grounded Personas in Large Language Models: Characterization and Alignment with Socio-Psychological Value Frameworks

SafetyDGX agent

arXiv:2601.22396v2 Announce Type: replace-cross Abstract: Despite the growing utility of Large Language Models (LLMs) for simulating human behavior, the extent to which these synthetic personas accura

Dynamic Infilling Anchors for Format-Constrained Generation in Diffusion Large Language Models

ResearchDGX agent

arXiv:2606.04535v1 Announce Type: cross Abstract: Diffusion large language models (dLLMs) offer bidirectional attention and parallel generation, enabling them to exploit global context and naturally s

Geometry-Aware Distillation for Prompt Tuning Biomedical Vision-Language Models

SafetyDGX agent

arXiv:2606.04922v1 Announce Type: cross Abstract: Current prompt-based and adapter-based tuning of vision-language models (VLMs) is attractive for medical imaging, where clinical data sensitivity favo

GeoMin: Data-Efficient Semi-Supervised RLVR via Geometric Distribution Modeling

TutorialsDGX agent

arXiv:2606.04516v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) significantly advances LLM reasoning, yet it faces a dilemma: standard supervised scaling is thr

How Users Understand Robot Foundation Model Performance through Task Success Rates and Beyond

ResearchDGX agent

arXiv:2602.03920v2 Announce Type: replace Abstract: Robot Foundation Models (RFMs) represent a promising approach to developing general-purpose home robots. Given the broad capabilities of RFMs, users

Introducing Magenta RealTime 2 🎺 - Open model for live music generation - Just 2.4B parameters, perfect for on-device - Low latency control…

Local AiDGX agent

Introducing Magenta RealTime 2 🎺 - Open model for live music generation - Just 2.4B parameters, perfect for on-device - Low latency control - Control with audio, MIDI, and text We're releasing it with

It seems like @GaryMarcus was right: the AI revenue models are imploding

SafetyDGX agent

It seems like @GaryMarcus was right: the AI revenue models are imploding Sam Altman said AI budgeting has recently become a 'huge issue' for some companies, something that 'never came up' earlier this

Locally is now LM Studio’s mobile app. And today we're bringing LM Link to iPhone. Use your largest local models over a secure, end-to-end e…

Local AiDGX agent

Locally is now LM Studio’s mobile app. And today we're bringing LM Link to iPhone. Use your largest local models over a secure, end-to-end encrypted connection, anywhere you go. Download the app now:

MAD: Mapping-Aware World Models for Agile Quadrotor Flight

Local AiDGX agent

arXiv:2606.04534v1 Announce Type: new Abstract: Agile quadrotor flight in cluttered scenes requires more than a reactive mapping from a depth image to a control command: the vehicle must remember whic

MaskForge: Structure-Aware Adaptive Attacks for Jailbreaking Diffusion Large Language Models

SafetyDGX agent

arXiv:2606.04027v1 Announce Type: cross Abstract: Diffusion large language models (dLLMs) generate text by iteratively denoising partially masked sequences under bidirectional context, exposing a safe

Q&A with Satya Nadella on Microsoft's competitive position, MAI models, OpenAI, the software business, GitHub Copilot, Project Solara, data centers, and more (Ben Thompson/Stratechery)

IndustryDGX agent

Ben Thompson / Stratechery: Q&A with Satya Nadella on Microsoft's competitive position, MAI models, OpenAI, the software business, GitHub Copilot, Project Solara, data centers, and more — An interview

Read the Trace, Steer the Path: Trajectory-Aware Reinforcement Learning for Diffusion Language Models

ResearchDGX agent

arXiv:2606.04396v1 Announce Type: new Abstract: Diffusion large language models (dLLMs) generate responses by iteratively unmasking and revising many positions in parallel. This process leaves a rich

ReSGA: A Large Tail Risk Model for Learning Value-at-Risk and Expected Shortfall

ResearchDGX agent

arXiv:2606.04576v1 Announce Type: cross Abstract: Learning Value-at-Risk (VaR) and Expected Shortfall (ES) is important for managing financial risks effectively. Existing approaches with limited param

SAID: Accelerating Diffusion-Based Language Models via Scaffold-Aware Iterative Decoding

ResearchDGX agent

arXiv:2606.04974v1 Announce Type: new Abstract: Diffusion large language models (DLLMs) enable non-autoregressive generation by iteratively denoising corrupted token sequences with bidirectional conte

Scaling Novel Graph Generation via Lightweight Structure-Guided Autoregressive Models

HardwareDGX agent

arXiv:2606.04287v1 Announce Type: cross Abstract: Generating realistic and diverse graphs is a key problem in machine learning, with applications in molecular discovery, circuit design, cybersecurity,

The Mechanistic Emergence of Symbol Grounding in Language Models

ApplicationsDGX agent

arXiv:2510.13796v3 Announce Type: replace Abstract: Symbol grounding (Harnad, 1990) describes how symbols such as words acquire their meanings by connecting to real-world sensorimotor experiences. Rec

T^star: Progressive Block Scaling for Masked Diffusion Language Models Through Trajectory Aware Reinforcement Learning

ResearchDGX agent

arXiv:2601.11214v5 Announce Type: replace Abstract: We present T^star, a simple TraceRL-based training curriculum for progressive block-size scaling in masked diffusion language models (MDMs). Startin

3 Jun 2026

An Attention-Based Denoising Model for Diffusion Weighted Imaging

ResearchDGX agent

arXiv:2606.03903v1 Announce Type: new Abstract: Diffusion-weighted imaging (DWI) is used for whole-body cancer screening, but it typically requires a long acquisition time. When the scan time is reduc

Are Common Substructures Transferable? Riemannian Graph Foundation Model with Neural Vector Bundles

ResearchDGX agent

arXiv:2606.03270v1 Announce Type: cross Abstract: Foundation models have sparked a revolution via a pretraining-adaptation paradigm, with recent efforts extending this success to graphs. Unlike other

At this point, I suspect you could put endpoints named 0pus 4.8 & GPT 5.S in your apps powered by open-source models and it would get massiv…

IndustryDGX agent

At this point, I suspect you could put endpoints named 0pus 4.8 & GPT 5.S in your apps powered by open-source models and it would get massive usage without people complaining. The power of 'frontier'

Causal Evidence of Stack Representations in Modeling Counter Languages Using Transformers

TutorialsDGX agent

arXiv:2606.03398v1 Announce Type: cross Abstract: Formal languages have proven to be effective conduits to understand the inner mechanisms of transformers. Past work has shown that transformers traine

CTR-Sink: Attention Sink for Language Models in Click-Through Rate Prediction

ResearchDGX agent

arXiv:2508.03668v2 Announce Type: replace Abstract: Click-Through Rate (CTR) prediction, a core task in recommendation systems, estimates user click likelihood using historical behavioral data. Modeli

Day 2 at #MSBuild is about what it takes to move beyond generic foundation models. Think customization, inference performance, and getting p…

ApplicationsDGX agent

Day 2 at #MSBuild is about what it takes to move beyond generic foundation models. Think customization, inference performance, and getting production-ready AI deployed at scale. @chahvivi will lead a

Distill-then-Replace: Efficient Task-Specific Hybrid Attention Model Construction

Local AiDGX agent

arXiv:2601.11667v2 Announce Type: replace-cross Abstract: Transformer architectures deliver state-of-the-art accuracy via dense full-attention, but their quadratic time and memory complexity with resp

Fundamental’s Large Tabular Model NEXUS is now available on Amazon SageMaker JumpStart

TutorialsDGX agent

Fundamental's NEXUS is a foundation model designed specifically for tabular data, trained on billions of tabular datasets using Amazon SageMaker HyperPod to understand non-linear relationships across

GeoAlign: Beyond Semantics with State-Guided Spatial Alignment in VLA Models

SafetyDGX agent

arXiv:2606.03240v1 Announce Type: new Abstract: Current Vision--Language--Action (VLA) models often optimize for semantic grounding, whereas executable manipulation requires geometry-aware spatial ali

GRZO: Group-Relative Zeroth-Order Optimization for Large Language Model Fine-Tuning

HardwareDGX agent

arXiv:2606.02857v1 Announce Type: cross Abstract: Zeroth-order (ZO) optimization is a memory-efficient alternative to backpropagation for fine-tuning large language models, but its deployment is limit

Ideogram 4.0 is now natively supported on ComfyUI @ideogram_ai v4.0 is an open-weight 9.3B text-to-image foundation model. It is exclusively…

Local AiDGX agent

Ideogram 4.0 is now natively supported on ComfyUI @ideogram_ai v4.0 is an open-weight 9.3B text-to-image foundation model. It is exclusively trained on structured JSON caption datasets for precise sce

I'm finally launching this agent tomorrow! Will be free, open source, and powered exclusively by open models on @togethercompute. Will also …

AgentsDGX agent

I'm finally launching this agent tomorrow! Will be free, open source, and powered exclusively by open models on @togethercompute. Will also drop a full guide on how it works! Building an agent that ca

Introducing Ideogram 4.0: the best open image model in the world. Think it. Make it. Own it. Download the weights, fine-tune on your own dat…

IndustryDGX agent

Introducing Ideogram 4.0: the best open image model in the world. Think it. Make it. Own it. Download the weights, fine-tune on your own data, and run it on your hardware. Live on every Ideogram plan

Learning Unmasking Policies for Diffusion Language Models

SafetyDGX agent

arXiv:2512.09106v4 Announce Type: replace Abstract: Diffusion (Large) Language Models (dLLMs) now match the downstream performance of their autoregressive counterparts on many tasks, while holding the

Most people, including really accomplished people, don't have an accurate mental model of how LLMs operate (and why would they?) You see thi…

ApplicationsDGX agent

Most people, including really accomplished people, don't have an accurate mental model of how LLMs operate (and why would they?) You see this in wide beliefs that AI is just copying from known sources

Oscillatory State-Space Models as Inductive Biases for Physics-Informed Neural PDE Solvers

TutorialsDGX agent

arXiv:2606.02623v1 Announce Type: cross Abstract: Solving time-dependent partial differential equations (PDEs) is an important problem in computational science and engineering. Physics-informed neural

Physical Plausibility Reasoning via HCM-GRPO: Empowering Compact Model for Superior Performance

SafetyDGX agent

arXiv:2511.10055v2 Announce Type: replace Abstract: The performance of image generation has been significantly improved in recent years. However, the study of image screening is rare, and its performa

ReciNet: Reciprocal Space-Aware Long-Range Modeling for Crystalline Property Prediction

Local AiDGX agent

arXiv:2502.02748v4 Announce Type: replace Abstract: Predicting properties of crystals from their structures is a fundamental yet challenging task in materials science. Unlike molecules, crystal struct

← Previous
1…211212213214215…1017
Next →