AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
48,975 results
Model Releases

Restoring Linguistic Grounding in VLA Models via Train-Free Attention Recalibration

DGX agent

arXiv:2603.06001v2 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models enable robots to perform manipulation tasks directly from natural language instructions and are increasing

model-releasesarxiv-cs-ai
3 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Unlocking Speech-Text Compositional Powers: Instruction-Following Speech Language Models without Instruction Tuning

DGX agent

arXiv:2607.02214v1 Announce Type: new Abstract: Instruction tuning for speech language models (SLMs) is substantially more challenging than for text-based large language models (LLMs), as it requires

researcharxiv-cs-cl
3 Jul 2026
Model Releases

VLSA: Vision-Language-Action Models with Plug-and-Play Safety Constraint Layer

DGX agent

arXiv:2512.11891v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) models have demonstrated remarkable capabilities in generalizing across diverse robotic manipulation tasks. However, de

model-releasesarxiv-cs-ro
3 Jul 2026
Model Releases

Large language models replicate and predict human cooperation across experiments in game theory

DGX agent

arXiv:2511.04500v3 Announce Type: replace Abstract: Large language models (LLMs) are increasingly deployed as decision-making agents in high-stakes domains and as imitators of human behavior in the so

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

MEPA: Multi-Scale Representation Alignment for Visual Autoregressive Modeling with Mixture of Experts

DGX agent

arXiv:2607.00371v1 Announce Type: cross Abstract: Visual AutoRegressive modeling (VAR) has pioneered a coarse-to-fine multi-scale autoregressive generative paradigm, demonstrating strong capabilities

model-releasesarxiv-cs-ai
2 Jul 2026
Safety

Predicting LLM Reasoning Performance with Small Proxy Model

DGX agent

arXiv:2509.21013v4 Announce Type: replace-cross Abstract: Given the prohibitive cost of pre-training large language models, it is essential to leverage smaller proxy models to optimize datasets before

safetyarxiv-cs-ai
2 Jul 2026
Model Releases

RetailSMV: Exocentric vs. Egocentric Adaptation of Foundation Video World Models in Retail

DGX agent

arXiv:2607.00310v1 Announce Type: cross Abstract: Foundation video diffusion models are increasingly viewed as world simulators for embodied agents, yet their pretraining on internet-scale generic vid

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

Citation Discipline in Spec-Driven Development: A Cross-Model Empirical Study of Output Determinism and Automated Hallucination Detection in LLM-Generated Code

DGX agent

arXiv:2606.30689v1 Announce Type: cross Abstract: Spec-Driven Development (SDD) frameworks guide Large Language Model (LLM)-powered code generation through formal specifications, yet they differ funda

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

Curvature-Guided Module Localization for Low-Rank Detoxification of Backdoored Large Language Models

DGX agent

arXiv:2606.30899v1 Announce Type: cross Abstract: Backdoor attacks pose a serious threat to large language models (LLMs) by causing otherwise benign systems to produce attacker-specified malicious beh

model-releasesarxiv-cs-ai
1 Jul 2026
Research

Diffusion-warm sampling of the XY model enables fast thermalization at scale

DGX agent

arXiv:2606.30773v1 Announce Type: cross Abstract: We introduce a novel technique for scalable sampling of spin-system states with continuous symmetries using diffusion models. By applying our approach

researcharxiv-cs-lg
1 Jul 2026
Local Ai

Large Databases Need Small, Open-Weight Language Models

DGX agent

arXiv:2606.31808v1 Announce Type: new Abstract: Language model systems built around proprietary APIs often operate on a token-based cost model. This becomes prohibitively expensive in the context of l

local-aiarxiv-cs-ai
1 Jul 2026
Model Releases

Learning to Deny: Action Denial in Multimodal Large Language Models

DGX agent

arXiv:2606.31187v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have rapidly advanced video understanding, achieving strong zero-shot and few-shot recognition across standard

model-releasesarxiv-cs-cv
1 Jul 2026
Model Releases

LightOnOCR: A 1B End-to-End Multilingual Vision-Language Model for State-of-the-Art OCR

DGX agent

arXiv:2601.14251v2 Announce Type: replace Abstract: We present LightOnOCR-2-1B, a 1B-parameter end-to-end multilingual vision--language model that converts document images (e.g., PDFs) into clean, nat

model-releasesarxiv-cs-cv
1 Jul 2026
Safety

Long-term Traffic Simulation via Structured Autoregressive Modeling

DGX agent

arXiv:2606.31209v1 Announce Type: new Abstract: Interactive traffic simulation is a vital world model for autonomous driving. A central challenge in long-horizon simulation is modeling sustained multi

safetyarxiv-cs-ai
1 Jul 2026
Applications

MemLearner: Learning to Query Context memory for Video World Models

DGX agent

arXiv:2606.31734v1 Announce Type: new Abstract: Video World Models are interactive video generation models that predict future world states based on user actions and history video frames. A critical c

applicationsarxiv-cs-cv
1 Jul 2026
Model Releases

Modeling Cell-Cycle-Aware Single-Cell Drug Perturbation Responses

DGX agent

arXiv:2606.30695v1 Announce Type: cross Abstract: Single-cell drug perturbation models should predict not only transcriptional response magnitude, but also whether a treatment alters the proliferative

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

Predictable GRPO: A Closed-Form Model of Training Dynamics

DGX agent

arXiv:2606.30789v1 Announce Type: new Abstract: Group Relative Policy Optimization (GRPO) has become a standard tool for improving the reasoning ability of large language models, yet its training dyna

model-releasesarxiv-cs-lg
1 Jul 2026
Model Releases

Probing Stylistic Appropriation using Large Language Models: An Evaluation Framework for Copyright Infringement under EU Law

DGX agent

arXiv:2606.31250v1 Announce Type: cross Abstract: Large language models (LLM) trained on web-scale corpora generate output that may infringe copyright, yet existing technical safeguards focus narrowly

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

Revisiting Parameter Redundancy in Vision-Language-Action Models: Insights from VLM-to-VLA Adaptation

DGX agent

arXiv:2606.31382v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have made significant strides in embodied intelligence by integrating the powerful representations of pre-trained Vi

model-releasesarxiv-cs-ro
1 Jul 2026
Research

Unsupervised Thermodynamics of Molecular Diffusion Models: Action-Operator Semantics and Auditable Free-Energy Readout

DGX agent

arXiv:2606.30687v1 Announce Type: cross Abstract: Diffusion models are increasingly utilized for modeling molecular structures and conformational ensembles, yet the thermodynamic meaning of their lear

researcharxiv-cs-ai
1 Jul 2026
Model Releases

When Does Learning to Stop Help? A Cost-Aware Study of Early Exits in Reasoning Models

DGX agent

arXiv:2606.30852v1 Announce Type: new Abstract: Reasoning models spend different amounts of useful computation across instances, but it remains unclear when a learned stopping rule improves over simpl

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

Wisdom Of The (AI) Crowd: Investigating Artificial Swarm Intelligence In Large Language Models

DGX agent

arXiv:2606.31404v1 Announce Type: new Abstract: Human swarm intelligence demonstrates remarkable collective accuracy but faces scalability constraints in cost, coordination, and time. We investigate w

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

WorldRoamBench: An Open-World Benchmark for Long-Horizon Stability of Interactive World Models

DGX agent

arXiv:2606.31672v1 Announce Type: cross Abstract: Despite rapid progress in interactive world models (IWMs), existing benchmarks evaluate action following only at trajectory level and ignore memory an

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

Z-1: Efficient Reinforcement Learning for Vision-Language-Action Models

DGX agent

arXiv:2606.31846v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models offer a promising framework for robotic manipulation by connecting language instructions, visual observations, and

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

Attractor States Emerge in Multi-Turn LLM Conversations

DGX agent

arXiv:2606.30571v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used in open-ended multi-agent settings, but the long-run dynamics of model--model interaction remain po

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

Can AI Draw Science? A Benchmark for Evaluating Scientific Figure Generation by Text-to-Image and Multimodal Models

DGX agent

arXiv:2606.28406v1 Announce Type: cross Abstract: Text-to-image and multimodal generative models are increasingly used to produce scientific figures such as mechanism diagrams, experimental-design sch

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

CAREBench: A Child-Safety Risk Benchmark for Language Models

DGX agent

arXiv:2606.29685v1 Announce Type: new Abstract: How can we evaluate whether frontier AI systems recognize child-safety risks before they escalate into explicit harm? Existing child safety evaluations

model-releasesarxiv-cs-lg
30 Jun 2026
Model Releases

DataComp-VLM: Improved Open Datasets for Vision-Language Models

DGX agent

arXiv:2606.28551v1 Announce Type: cross Abstract: Building performant Vision-Language Models (VLMs) requires carefully curating large-scale training datasets, yet the community lacks systematic benchm

model-releasesarxiv-cs-cl
30 Jun 2026
Safety

Do Models Read What They Write? Causal Registers in Scratchpad Reasoning

DGX agent

arXiv:2606.29522v1 Announce Type: cross Abstract: A central hope behind process supervision is that models can expose intermediate variables that matter for their later behavior. For this to help with

safetyarxiv-cs-cl
30 Jun 2026
Research

Do We Still Need Fine Tuning? Turkish Sentiment Analysis in the Era of Large Language Model

DGX agent

arXiv:2606.29614v1 Announce Type: cross Abstract: This study examines whether supervised fine-tuning remains necessary for Turkish sentiment analysis in the era of large language models. We compare cl

researcharxiv-cs-ai
30 Jun 2026
Model Releases

Early Estimation of Language to Latent Alignment in Diffusion Models

DGX agent

arXiv:2512.08505v2 Announce Type: replace Abstract: Conditional diffusion models frequently suffer from language-image misalignments. Due to the ambiguity of intermediate noise corrupted latents, asse

model-releasesarxiv-cs-cv
30 Jun 2026
Safety

FutureNav: Unified World-Action Modeling for Vision-and-Language Navigation

DGX agent

arXiv:2606.30367v1 Announce Type: new Abstract: Vision-and-language navigation (VLN) in continuous environments requires an agent to ground instructions in egocentric observations while maintaining sp

safetyarxiv-cs-ro
30 Jun 2026
Applications

Gradient Boosted Mixed Models: Flexible Estimation of Mean and Variance Components for Clustered Data

DGX agent

arXiv:2511.00217v2 Announce Type: replace-cross Abstract: We introduce Gradient Boosted Mixed Models (GBMixed), a framework which extends boosting to clustered data by jointly modeling the mean and va

applicationsarxiv-cs-lg
30 Jun 2026
Safety

HERO: Improving the Reliability and Sensitivity of Generative Model Evaluation Using Historical Data

DGX agent

arXiv:2606.29784v1 Announce Type: cross Abstract: Reliable generative AI models critically rely on expert human annotations to evaluate output quality, yet these 'gold' labels are expensive to collect

safetyarxiv-cs-ai
30 Jun 2026
Model Releases

JuZhou 1.0 Technical Report: The First Edge-Native Text-to-Image Foundation Model Trained Entirely on China-Developed AI Accelerators

DGX agent

arXiv:2606.28421v1 Announce Type: cross Abstract: Text-to-image (T2I) diffusion models typically require substantial computational resources and cloud infrastructure, posing significant challenges for

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

KnowsTFM: Knowledge-Informed Fine-Tuning of Small Tabular Foundation Models

DGX agent

arXiv:2606.30258v1 Announce Type: cross Abstract: Tabular foundation models have advanced deep learning for tabular data by delivering strong default performance across many small and medium tasks. Ye

model-releasesarxiv-cs-ai
30 Jun 2026
Safety

MM-Nav: Multi-View VLA Model for Robust Visual Navigation via Multi-Expert Learning

DGX agent

arXiv:2510.03142v2 Announce Type: replace-cross Abstract: Visual navigation policy is widely regarded as a promising direction, as it mimics humans by using egocentric visual observations for navigati

safetyarxiv-cs-cv
30 Jun 2026
Research

Physics-Informed Distillation of Diffusion Models for PDE-Constrained Generation

DGX agent

arXiv:2505.22391v2 Announce Type: replace-cross Abstract: Modeling physical systems in a generative manner offers several advantages, including the ability to handle partial observations, generate div

researcharxiv-cs-ai
30 Jun 2026
Model Releases

RoboGaze: Evaluating Robot World Models via Structured Vision-Language Analysis

DGX agent

arXiv:2606.28385v1 Announce Type: cross Abstract: Recent advances in robot world models enable synthetic video generation for embodied prediction and planning. However, evaluating these videos is chal

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Statistically Indistinguishable, Operationally Distinct: A Formal Barrier for Tabular Foundation Models

DGX agent

arXiv:2606.29091v1 Announce Type: cross Abstract: Tabular foundation models cannot reason about data produced by running systems without access to the rules that govern them. We make this statement fa

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

SVC-Probe: A Framework for Evaluating Perturbation Generalization in Spatial Foundation-Model Embeddings

DGX agent

arXiv:2606.28465v1 Announce Type: cross Abstract: This work examines perturbation generalization in spatial foundation-model embeddings derived from fluorescence microscopy images. Although these mode

model-releasesarxiv-cs-ai
30 Jun 2026
Research

t-STEP: An interpretable model for Total Electron Content predictions and irregularities estimations

DGX agent

arXiv:2606.29644v1 Announce Type: new Abstract: Earth system infrastructures relying on satellite-based technologies, such as Global Positioning System (GPS) communications, are affected by ionospheri

researcharxiv-cs-lg
30 Jun 2026
Safety

Vision-Language-Action Models: Experimental Insights from a Real-World UR5 Platform

DGX agent

arXiv:2606.30456v1 Announce Type: cross Abstract: This project investigates whether recent Vision-Language-Action (VLA) models can be transferred from controlled research benchmarks to a real-world ro

safetyarxiv-cs-cv
30 Jun 2026
Safety

WoVR: World Models as Reliable Simulators for Post-Training VLA Policies with RL

DGX agent

arXiv:2602.13977v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) promises to unlock capabilities beyond imitation learning for Vision--Language--Action (VLA) models, but its requi

safetyarxiv-cs-ai
30 Jun 2026
Model Releases

Aloe-Vision: Robust Vision-Language Models for Healthcare

DGX agent

arXiv:2606.27500v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) specialized in healthcare are emerging as a promising research direction due to their potential impact in clinica

model-releasesarxiv-cs-cl
29 Jun 2026
Model Releases

Do Speech Emphasis Models Generalize across Languages and Emotions?

DGX agent

arXiv:2606.27717v1 Announce Type: cross Abstract: Prosodic emphasis varies across languages, emotions, and speaking styles, yet existing emphasis detection models are largely trained and evaluated on

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

MobileManiBench: Simplifying Model Verification for Mobile Manipulation

DGX agent

arXiv:2602.05233v2 Announce Type: replace Abstract: Vision-language-action models have advanced robotic manipulation but remain constrained by reliance on the large, teleoperation-collected datasets d

model-releasesarxiv-cs-ro
29 Jun 2026
Model Releases

MultiHashFormer: Hash-based Generative Language Models

DGX agent

arXiv:2606.28057v1 Announce Type: cross Abstract: Language models (LMs) represent tokens using embedding matrices that scale linearly with the vocabulary size. To constrain the parameter footprint, pr

model-releasesarxiv-cs-ai
29 Jun 2026
← Previous
1…6061626364…1021
Next →