AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlog
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,428 results
Research

Diffusion-warm sampling of the XY model enables fast thermalization at scale

DGX agent

arXiv:2606.30773v1 Announce Type: cross Abstract: We introduce a novel technique for scalable sampling of spin-system states with continuous symmetries using diffusion models. By applying our approach

researcharxiv-cs-lg
1 Jul 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Industry

GLM 5.2 just became the first open-source model to lead a category on APEX-SWE. It scored a 55.3% Pass@1 on Integration, the top score we've…

DGX agent

GLM 5.2 just became the first open-source model to lead a category on APEX-SWE. It scored a 55.3% Pass@1 on Integration, the top score we've recorded for any model, open or closed source. On the overa

industryclem-delangue--x
1 Jul 2026
Local Ai

Large Databases Need Small, Open-Weight Language Models

DGX agent

arXiv:2606.31808v1 Announce Type: new Abstract: Language model systems built around proprietary APIs often operate on a token-based cost model. This becomes prohibitively expensive in the context of l

local-aiarxiv-cs-ai
1 Jul 2026
Model Releases

Learning to Deny: Action Denial in Multimodal Large Language Models

DGX agent

arXiv:2606.31187v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have rapidly advanced video understanding, achieving strong zero-shot and few-shot recognition across standard

model-releasesarxiv-cs-cv
1 Jul 2026
Model Releases

LightOnOCR: A 1B End-to-End Multilingual Vision-Language Model for State-of-the-Art OCR

DGX agent

arXiv:2601.14251v2 Announce Type: replace Abstract: We present LightOnOCR-2-1B, a 1B-parameter end-to-end multilingual vision--language model that converts document images (e.g., PDFs) into clean, nat

model-releasesarxiv-cs-cv
1 Jul 2026
Safety

Long-term Traffic Simulation via Structured Autoregressive Modeling

DGX agent

arXiv:2606.31209v1 Announce Type: new Abstract: Interactive traffic simulation is a vital world model for autonomous driving. A central challenge in long-horizon simulation is modeling sustained multi

safetyarxiv-cs-ai
1 Jul 2026
Applications

MemLearner: Learning to Query Context memory for Video World Models

DGX agent

arXiv:2606.31734v1 Announce Type: new Abstract: Video World Models are interactive video generation models that predict future world states based on user actions and history video frames. A critical c

applicationsarxiv-cs-cv
1 Jul 2026
Model Releases

Modeling Cell-Cycle-Aware Single-Cell Drug Perturbation Responses

DGX agent

arXiv:2606.30695v1 Announce Type: cross Abstract: Single-cell drug perturbation models should predict not only transcriptional response magnitude, but also whether a treatment alters the proliferative

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

Predictable GRPO: A Closed-Form Model of Training Dynamics

DGX agent

arXiv:2606.30789v1 Announce Type: new Abstract: Group Relative Policy Optimization (GRPO) has become a standard tool for improving the reasoning ability of large language models, yet its training dyna

model-releasesarxiv-cs-lg
1 Jul 2026
Model Releases

Probing Stylistic Appropriation using Large Language Models: An Evaluation Framework for Copyright Infringement under EU Law

DGX agent

arXiv:2606.31250v1 Announce Type: cross Abstract: Large language models (LLM) trained on web-scale corpora generate output that may infringe copyright, yet existing technical safeguards focus narrowly

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

Revisiting Parameter Redundancy in Vision-Language-Action Models: Insights from VLM-to-VLA Adaptation

DGX agent

arXiv:2606.31382v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have made significant strides in embodied intelligence by integrating the powerful representations of pre-trained Vi

model-releasesarxiv-cs-ro
1 Jul 2026
Research

Unsupervised Thermodynamics of Molecular Diffusion Models: Action-Operator Semantics and Auditable Free-Energy Readout

DGX agent

arXiv:2606.30687v1 Announce Type: cross Abstract: Diffusion models are increasingly utilized for modeling molecular structures and conformational ensembles, yet the thermodynamic meaning of their lear

researcharxiv-cs-ai
1 Jul 2026
Model Releases

When Does Learning to Stop Help? A Cost-Aware Study of Early Exits in Reasoning Models

DGX agent

arXiv:2606.30852v1 Announce Type: new Abstract: Reasoning models spend different amounts of useful computation across instances, but it remains unclear when a learned stopping rule improves over simpl

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

Wisdom Of The (AI) Crowd: Investigating Artificial Swarm Intelligence In Large Language Models

DGX agent

arXiv:2606.31404v1 Announce Type: new Abstract: Human swarm intelligence demonstrates remarkable collective accuracy but faces scalability constraints in cost, coordination, and time. We investigate w

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

WorldRoamBench: An Open-World Benchmark for Long-Horizon Stability of Interactive World Models

DGX agent

arXiv:2606.31672v1 Announce Type: cross Abstract: Despite rapid progress in interactive world models (IWMs), existing benchmarks evaluate action following only at trajectory level and ignore memory an

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

Z-1: Efficient Reinforcement Learning for Vision-Language-Action Models

DGX agent

arXiv:2606.31846v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models offer a promising framework for robotic manipulation by connecting language instructions, visual observations, and

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

Attractor States Emerge in Multi-Turn LLM Conversations

DGX agent

arXiv:2606.30571v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used in open-ended multi-agent settings, but the long-run dynamics of model--model interaction remain po

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

Can AI Draw Science? A Benchmark for Evaluating Scientific Figure Generation by Text-to-Image and Multimodal Models

DGX agent

arXiv:2606.28406v1 Announce Type: cross Abstract: Text-to-image and multimodal generative models are increasingly used to produce scientific figures such as mechanism diagrams, experimental-design sch

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

CAREBench: A Child-Safety Risk Benchmark for Language Models

DGX agent

arXiv:2606.29685v1 Announce Type: new Abstract: How can we evaluate whether frontier AI systems recognize child-safety risks before they escalate into explicit harm? Existing child safety evaluations

model-releasesarxiv-cs-lg
30 Jun 2026
Model Releases

DataComp-VLM: Improved Open Datasets for Vision-Language Models

DGX agent

arXiv:2606.28551v1 Announce Type: cross Abstract: Building performant Vision-Language Models (VLMs) requires carefully curating large-scale training datasets, yet the community lacks systematic benchm

model-releasesarxiv-cs-cl
30 Jun 2026
Safety

Do Models Read What They Write? Causal Registers in Scratchpad Reasoning

DGX agent

arXiv:2606.29522v1 Announce Type: cross Abstract: A central hope behind process supervision is that models can expose intermediate variables that matter for their later behavior. For this to help with

safetyarxiv-cs-cl
30 Jun 2026
Research

Do We Still Need Fine Tuning? Turkish Sentiment Analysis in the Era of Large Language Model

DGX agent

arXiv:2606.29614v1 Announce Type: cross Abstract: This study examines whether supervised fine-tuning remains necessary for Turkish sentiment analysis in the era of large language models. We compare cl

researcharxiv-cs-ai
30 Jun 2026
Model Releases

Early Estimation of Language to Latent Alignment in Diffusion Models

DGX agent

arXiv:2512.08505v2 Announce Type: replace Abstract: Conditional diffusion models frequently suffer from language-image misalignments. Due to the ambiguity of intermediate noise corrupted latents, asse

model-releasesarxiv-cs-cv
30 Jun 2026
Safety

FutureNav: Unified World-Action Modeling for Vision-and-Language Navigation

DGX agent

arXiv:2606.30367v1 Announce Type: new Abstract: Vision-and-language navigation (VLN) in continuous environments requires an agent to ground instructions in egocentric observations while maintaining sp

safetyarxiv-cs-ro
30 Jun 2026
Applications

Gradient Boosted Mixed Models: Flexible Estimation of Mean and Variance Components for Clustered Data

DGX agent

arXiv:2511.00217v2 Announce Type: replace-cross Abstract: We introduce Gradient Boosted Mixed Models (GBMixed), a framework which extends boosting to clustered data by jointly modeling the mean and va

applicationsarxiv-cs-lg
30 Jun 2026
Safety

HERO: Improving the Reliability and Sensitivity of Generative Model Evaluation Using Historical Data

DGX agent

arXiv:2606.29784v1 Announce Type: cross Abstract: Reliable generative AI models critically rely on expert human annotations to evaluate output quality, yet these 'gold' labels are expensive to collect

safetyarxiv-cs-ai
30 Jun 2026
Model Releases

JuZhou 1.0 Technical Report: The First Edge-Native Text-to-Image Foundation Model Trained Entirely on China-Developed AI Accelerators

DGX agent

arXiv:2606.28421v1 Announce Type: cross Abstract: Text-to-image (T2I) diffusion models typically require substantial computational resources and cloud infrastructure, posing significant challenges for

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

KnowsTFM: Knowledge-Informed Fine-Tuning of Small Tabular Foundation Models

DGX agent

arXiv:2606.30258v1 Announce Type: cross Abstract: Tabular foundation models have advanced deep learning for tabular data by delivering strong default performance across many small and medium tasks. Ye

model-releasesarxiv-cs-ai
30 Jun 2026
Safety

MM-Nav: Multi-View VLA Model for Robust Visual Navigation via Multi-Expert Learning

DGX agent

arXiv:2510.03142v2 Announce Type: replace-cross Abstract: Visual navigation policy is widely regarded as a promising direction, as it mimics humans by using egocentric visual observations for navigati

safetyarxiv-cs-cv
30 Jun 2026
Research

Physics-Informed Distillation of Diffusion Models for PDE-Constrained Generation

DGX agent

arXiv:2505.22391v2 Announce Type: replace-cross Abstract: Modeling physical systems in a generative manner offers several advantages, including the ability to handle partial observations, generate div

researcharxiv-cs-ai
30 Jun 2026
Model Releases

RoboGaze: Evaluating Robot World Models via Structured Vision-Language Analysis

DGX agent

arXiv:2606.28385v1 Announce Type: cross Abstract: Recent advances in robot world models enable synthetic video generation for embodied prediction and planning. However, evaluating these videos is chal

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Statistically Indistinguishable, Operationally Distinct: A Formal Barrier for Tabular Foundation Models

DGX agent

arXiv:2606.29091v1 Announce Type: cross Abstract: Tabular foundation models cannot reason about data produced by running systems without access to the rules that govern them. We make this statement fa

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

SVC-Probe: A Framework for Evaluating Perturbation Generalization in Spatial Foundation-Model Embeddings

DGX agent

arXiv:2606.28465v1 Announce Type: cross Abstract: This work examines perturbation generalization in spatial foundation-model embeddings derived from fluorescence microscopy images. Although these mode

model-releasesarxiv-cs-ai
30 Jun 2026
Research

t-STEP: An interpretable model for Total Electron Content predictions and irregularities estimations

DGX agent

arXiv:2606.29644v1 Announce Type: new Abstract: Earth system infrastructures relying on satellite-based technologies, such as Global Positioning System (GPS) communications, are affected by ionospheri

researcharxiv-cs-lg
30 Jun 2026
Model Releases

𝗚𝗟𝗠-𝟱.𝟮 (the latest open weights model) is having an Enterprise moment, and it is not an exaggeration.🚀 🔥 We have been impressed by h…

DGX agent

𝗚𝗟𝗠-𝟱.𝟮 (the latest open weights model) is having an Enterprise moment, and it is not an exaggeration.🚀 🔥 We have been impressed by how strongly GLM-5.2 is pushing long-horizon performance .. not just

model-releasesclem-delangue--x
30 Jun 2026
Safety

Vision-Language-Action Models: Experimental Insights from a Real-World UR5 Platform

DGX agent

arXiv:2606.30456v1 Announce Type: cross Abstract: This project investigates whether recent Vision-Language-Action (VLA) models can be transferred from controlled research benchmarks to a real-world ro

safetyarxiv-cs-cv
30 Jun 2026
Safety

WoVR: World Models as Reliable Simulators for Post-Training VLA Policies with RL

DGX agent

arXiv:2602.13977v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) promises to unlock capabilities beyond imitation learning for Vision--Language--Action (VLA) models, but its requi

safetyarxiv-cs-ai
30 Jun 2026
Model Releases

Aloe-Vision: Robust Vision-Language Models for Healthcare

DGX agent

arXiv:2606.27500v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) specialized in healthcare are emerging as a promising research direction due to their potential impact in clinica

model-releasesarxiv-cs-cl
29 Jun 2026
Model Releases

Do Speech Emphasis Models Generalize across Languages and Emotions?

DGX agent

arXiv:2606.27717v1 Announce Type: cross Abstract: Prosodic emphasis varies across languages, emotions, and speaking styles, yet existing emphasis detection models are largely trained and evaluated on

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

MobileManiBench: Simplifying Model Verification for Mobile Manipulation

DGX agent

arXiv:2602.05233v2 Announce Type: replace Abstract: Vision-language-action models have advanced robotic manipulation but remain constrained by reliance on the large, teleoperation-collected datasets d

model-releasesarxiv-cs-ro
29 Jun 2026
Model Releases

MultiHashFormer: Hash-based Generative Language Models

DGX agent

arXiv:2606.28057v1 Announce Type: cross Abstract: Language models (LMs) represent tokens using embedding matrices that scale linearly with the vocabulary size. To constrain the parameter footprint, pr

model-releasesarxiv-cs-ai
29 Jun 2026
Safety

ReWorld: Learning Better Representations for World Action Models

DGX agent

arXiv:2606.27504v1 Announce Type: new Abstract: World Action Models (WAMs) model future environment evolution under action conditioning, offering a scalable paradigm for autonomous driving. However, e

safetyarxiv-cs-cv
29 Jun 2026
Model Releases

StereoVLA: Enhancing Vision-Language-Action Models with Stereo Vision

DGX agent

arXiv:2512.21970v2 Announce Type: replace Abstract: While Vision-Language-Action (VLA) models excel in generalist manipulation, they often lack fine-grained spatial awareness and show limited viewpoin

model-releasesarxiv-cs-ro
29 Jun 2026
Research

The Weakest Link Tells It All: Outcome-Supervised Process Reward Modeling via Learnable Credit Assignment

DGX agent

arXiv:2606.27739v1 Announce Type: new Abstract: Process reward models (PRMs) enhance the reasoning capabilities of large language models (LLMs) by providing fine-grained feedback, yet training PRMs ty

researcharxiv-cs-lg
29 Jun 2026
Model Releases

Your AI Travel Agent Would Book You a Bullfight: An Agentic Benchmark for Implicit Animal Welfare in Frontier AI Models

DGX agent

arXiv:2606.18142v3 Announce Type: replace Abstract: AI agents are moving from advisors to actors, booking travel, planning menus, and running procurement on behalf of users. Existing benchmarks for AI

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

Grok 4.5, based on our 1.5T V9 foundation model, with Cursor data added in supplemental training, is now in private beta at SpaceX & Tesla. …

DGX agent

Grok 4.5, based on our 1.5T V9 foundation model, with Cursor data added in supplemental training, is now in private beta at SpaceX & Tesla. Early evals show performance close to, perhaps exceeding Opu

model-releaseselon-musk--x
28 Jun 2026
Model Releases

And the idea that frontier open weights models will continue to be released was fragile even before the current de facto licensing regime. h…

DGX agent

And the idea that frontier open weights models will continue to be released was fragile even before the current de facto licensing regime. https://x.com/emollick/status/2067669551685218638?s=20 Is the

model-releasesethan-mollick--x
27 Jun 2026
Model Releases

VCs are now sharing screenshots in group chats of Claude discouraging investment in open-source AI infra startups and models. Obviously ther…

DGX agent

VCs are now sharing screenshots in group chats of Claude discouraging investment in open-source AI infra startups and models. Obviously there is an absolute EXPLOSION of pitches in inference companies

model-releasesclem-delangue--x
27 Jun 2026
← Previous
1…7980818283…1259
Next →