AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,490 results
23 Jun 2026

Interleaved Speech Language Models Latently Work In Text

ResearchDGX agent

arXiv:2606.22473v1 Announce Type: cross Abstract: Speech language models (SLMs) have been extensively studied, with the common paradigm incorporating text data and pre-trained text LMs. A leading appr

IOI: Decoupling Kinematics and Physics for Interactive World Models

Model ReleasesDGX agent

arXiv:2606.23296v1 Announce Type: new Abstract: Developing generalist embodied agents requires interactive environments providing visually realistic feedback and accurate action-conditioned dynamics.

Learning a Normal World Model for Few-Shot Boundary-Calibrated Abnormality Detection

Model ReleasesDGX agent

arXiv:2606.22261v1 Announce Type: new Abstract: Abnormality detection in complex systems faces two practical barriers: abnormal labels are scarce, and binary labels do not quantify how far an event ha

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Localizing and Editing Knowledge in Large Audio-Language Models

Model ReleasesDGX agent

arXiv:2603.14343v2 Announce Type: replace Abstract: Large Audio-Language Models (LALMs) have shown strong performance in speech understanding, making speech a natural interface for accessing factual i

MEDLAYXPLAIN: Benchmarking the Expert-Lay Gap in Medical Vision-Language Models

Model ReleasesDGX agent

arXiv:2606.21194v1 Announce Type: new Abstract: Medical Vision-Language Models (Med-VLMs) achieve strong expert-level performance, yet their ability to generate patient-accessible descriptions remains

NAC: Neural Action Codec for Vision-Language-Action Models

ApplicationsDGX agent

arXiv:2606.21372v1 Announce Type: cross Abstract: Vision-language-action (VLA) models rely on discrete action tokenizers to bridge continuous robot control and autoregressive sequence modeling, yet ex

On the Expressive Power of Weight Quantization in Large Language Models

ResearchDGX agent

arXiv:2606.22249v1 Announce Type: new Abstract: In recent years, weight quantization that encodes the learnable parameters of large language models in an n-bit format has garnered significant attentio

Perturbation-Based Uncertainty for Failure Detection in Vision-Language-Action Models

Local AiDGX agent

arXiv:2606.20754v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have shown strong performance in robotic manipulation, but reliable uncertainty quantification remains challenging,

Policy4OOD: A Knowledge-Guided World Model for Policy Intervention Simulation against the Opioid Overdose Crisis

Model ReleasesDGX agent

arXiv:2602.12373v2 Announce Type: replace Abstract: The opioid epidemic remains one of the most severe public health crises in the United States, yet evaluating policy interventions before implementat

rumors of new models being delayed

Model ReleasesDGX agent

rumors of new models being delayed 🚨 SCOOP(s): - GPT-5.6 has been delayed and will no longer release this week. New target is ~mid-July. - DeepMind are not satisfied with the current state of 3.5 Pro

SafeDojo: Safe Reinforcement Learning for VLA via Interactive World Model

SafetyDGX agent

arXiv:2606.20698v1 Announce Type: new Abstract: Safe control is a prerequisite for real-world embodied intelligence, for which safe reinforcement learning has emerged as a promising paradigm. However,

Self-Evolving Cognitive Framework via Causal World Modeling for Embodied Scientific Intelligence

ResearchDGX agent

arXiv:2606.22449v1 Announce Type: cross Abstract: Current embodied world models are primarily optimized for predictive objectives, limiting their ability to generalize under distribution shifts and re

SkyJEPA: Learning Long-Horizon World Models for Zero-Shot Sim-to-Real Control of Quadrotors

ApplicationsDGX agent

arXiv:2606.23444v1 Announce Type: cross Abstract: Accurate dynamics models are critical for informed decision-making in robotic systems, particularly for agile aerial vehicles operating under uncertai

SVD-Surgeon: Optimal Singular-Value Surgery for Large Language Model Compression

Model ReleasesDGX agent

arXiv:2606.23568v1 Announce Type: new Abstract: Large language models (LLMs) achieve remarkable performance across a wide range of tasks, but their deployment is constrained by substantial memory and

Temper-Then-Tilt: Principled Unlearning for Generative Models through Tempering and Classifier Guidance

Model ReleasesDGX agent

arXiv:2602.10217v2 Announce Type: replace Abstract: We study machine unlearning in large generative models by framing the task as density ratio estimation to a target distribution rather than supervis

The Energy Consumption of Transformer Fine-Tuning: A Roofline-Inspired Scaling Model

ResearchDGX agent

arXiv:2606.23546v1 Announce Type: new Abstract: Transformer-based models underpin modern natural language processing but incur rapidly growing computational and energy costs. As training scales in bot

VDAWorld: World Modelling via VLM-Directed Abstraction and Simulation

AgentsDGX agent

arXiv:2512.11061v2 Announce Type: replace Abstract: Generative video models, a leading approach to world modelling, face fundamental limitations. They often violate physical and logical rules, lack in

Vera: A Layered Diffusion Model for Content-Preserving Video Editing

Model ReleasesDGX agent

arXiv:2606.23610v1 Announce Type: new Abstract: Video diffusion models have enabled remarkable progress in video generation and editing. However, content preservation remains a core challenge: existin

What Shapes Emergent Misalignment? Insights from Training Dynamics, Model Priors, and Data

Local AiDGX agent

arXiv:2606.20814v1 Announce Type: cross Abstract: Emergent misalignment (EM) is a phenomenon in which models generalize with narrow fine-tuning, leading to broad (yet uneven) misalignment across evalu

22 Jun 2026

GLM-5.2 has been the most popular new model on Fireworks this past week. @ArtificialAnlys confirms why: #3 overall on GDPval-AA (1524 Elo), …

Model ReleasesDGX agent

GLM-5.2 has been the most popular new model on Fireworks this past week. @ArtificialAnlys confirms why: #3 overall on GDPval-AA (1524 Elo), #1 open weights by 116 points. Interest is showing no signs

In a rare joint statement, Five Eyes leaders warn AI models capable of taking down governments and businesses are mere months away, urging leaders to 'act now' (Sarah Basford Canales/The Guardian)

Model ReleasesDGX agent

Sarah Basford Canales / The Guardian: In a rare joint statement, Five Eyes leaders warn AI models capable of taking down governments and businesses are mere months away, urging leaders to “act now” —

OpenAI unveils an updated GPT-5.5-Cyber model, launches the Patch the Planet initiative in partnership with Trail of Bits to fix open source bugs, and more (Lily Hay Newman/Wired)

Model ReleasesDGX agent

Lily Hay Newman / Wired: OpenAI unveils an updated GPT-5.5-Cyber model, launches the Patch the Planet initiative in partnership with Trail of Bits to fix open source bugs, and more — Amid concerns abo

Three announcements from our keynote at Compile, including how we're training a new model with SpaceX.

Model ReleasesDGX agent

Cursor announced three major initiatives at the Compile keynote, including a partnership with SpaceX to train a new AI model. The announcement highlights Cursor's expansion efforts and collaboration w

21 Jun 2026

I released a softmax-free attention model at GPT-2 Medium scale (~354M params, 11.5B tokens): structural sparsity + tile-skipping kernels for long-context VRAM savings. Open weights + custom Triton kernels [R]

Model ReleasesDGX agent

A researcher released an open-source softmax-free attention model at GPT-2 Medium scale (354M parameters trained on 11.5B tokens) that uses structural sparsity and tile-skipping kernels to reduce VRAM

11 Jun 2026

4DP-QA: Scalable QA for 4D Perception in Vision Language Models

Model ReleasesDGX agent

arXiv:2606.11568v1 Announce Type: new Abstract: Despite recent advances, Vision Language Models (VLMs) still struggle to grasp the dynamics of the world. We note that the ability to reason about a 4D

APEX: A Network-Native Time-Series Foundation Model for Forecasting and Anomaly Detection for Wireless Edge Operations

Model ReleasesDGX agent

arXiv:2606.11553v1 Announce Type: new Abstract: Generic time-series foundation models transfer poorly to wireless network telemetry whose signals are bursty, zero-inflated, and coupled across protocol

Benchmarking Large Language Models for Safety Data Extraction

Model ReleasesDGX agent

arXiv:2606.11204v1 Announce Type: new Abstract: Accurate extraction of structured information from Safety Data Sheets (SDS) remains challenging in industrial safety due to heterogeneous document forma

GLACIER: A Multimodal Student-Teacher Foundation Model for Molecular Property Prediction

SafetyDGX agent

arXiv:2606.11382v1 Announce Type: new Abstract: Deep learning models facilitate the discovery of molecules with tailored properties among billions of candidate compounds. However, the computational bu

Intermittent time series forecasting: local vs global models

SafetyDGX agent

arXiv:2601.14031v2 Announce Type: replace-cross Abstract: Forecasting intermittent time series, which contain zeros, is a crucial challenge in supply chains as inventory policies require probabilistic

LifeSentence: Language models can encode human life course trajectories from longitudinal panel data

Model ReleasesDGX agent

arXiv:2606.11220v1 Announce Type: new Abstract: Forecasting human life outcomes is important to gain insights into how individuals attain long and healthy lives. Conventional statistical approaches yi

Modelling magnetic material properties with uncertainty-aware neural networks

Model ReleasesDGX agent

arXiv:2606.11870v1 Announce Type: cross Abstract: Machine learning is increasingly applied to accelerate the discovery of novel materials by exploring large compositional and structural design spaces.

Ouroboros-Spatial: Closing the Data-Model Loop for Spatial Reasoning

ResearchDGX agent

arXiv:2606.11719v1 Announce Type: cross Abstract: Spatial reasoning remains a persistent challenge for multimodal large language models (MLLMs). Existing approaches largely rely on large-scale, static

PLUME: Probabilistic Latent Unified World Modeling and Parameter Estimation for Multi-Finger Manipulation

Model ReleasesDGX agent

arXiv:2606.11396v1 Announce Type: new Abstract: Dexterous manipulation with multi-finger hands can be sensitive to physical parameters such as object shape, pose, and friction coefficients. While simu

RAIL: Rethinking Auditory Intelligence in Large Audio-Language Models with a CHC-Grounded Benchmark

Model ReleasesDGX agent

arXiv:2606.11260v1 Announce Type: cross Abstract: Humans process rich auditory environments through tightly integrated cognitive capabilities such as audio perception, audio reasoning, and memory. Des

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning

Model ReleasesDGX agent

arXiv:2606.11816v1 Announce Type: cross Abstract: Forecasting real-world events requires language-model agents to reason under uncertainty from incomplete, time-bounded information. Yet evaluating whe

10 Jun 2026

Anthropic backtracks on a policy limiting Claude Fable 5's ability to develop other AI models, after significant backlash from the AI research community (Maxwell Zeff/Wired)

Model ReleasesDGX agent

Maxwell Zeff / Wired: Anthropic backtracks on a policy limiting Claude Fable 5's ability to develop other AI models, after significant backlash from the AI research community — The company changed cou

Cohere Transcribe, our open-source speech recognition model, is #1 on the new @huggingface Far-Field ASR benchmark.

Model ReleasesDGX agent

Cohere has released Transcribe, an open-source speech recognition model that achieved the top ranking on Hugging Face's newly established Far-Field Automatic Speech Recognition (ASR) benchmark. The mo

Deployment-Time Memorization in Foundation-Model Agents

Model ReleasesDGX agent

arXiv:2606.10062v1 Announce Type: new Abstract: Foundation-model agents are increasingly long-lived systems that remember users across interactions, making memorization an explicit deployment-time fun

Does Normalization Choice Matter for Causal Large Time-Series Models?

ApplicationsDGX agent

arXiv:2606.09954v1 Announce Type: cross Abstract: Large models for time-series forecasting have been emerged as a promising paradigm for training models on heterogeneous collections of signals. These

Earth-OneVision: Extending Remote Sensing Multimodal Large Language Models to More Sensor Modalities and Tasks

Model ReleasesDGX agent

arXiv:2606.10819v1 Announce Type: cross Abstract: RS-MLLMs enable natural-language understanding and spatial reasoning over earth observation imagery. However, existing models support only a narrow ra

From Observation to Intervention: A Causal Audit of Expert Importance in Mixture-of-Experts Models

Model ReleasesDGX agent

arXiv:2606.10703v1 Announce Type: cross Abstract: Interpretability methods routinely use population-level summary statistics over observed model behaviour to license claims about the effects of target

KCSAT-ML: Probing Reasoning Models with Nationwide-Cohort Human Difficulty

Model ReleasesDGX agent

arXiv:2606.10403v1 Announce Type: new Abstract: Math reasoning benchmarks have proliferated, yet most lack a per-item difficulty signal grounded in actual human performance. We introduce KCSAT-ML, a d

LIBERO-Occ: Evaluating and Improving Vision-Language-Action Models under Scene-Induced Occlusion via Viewpoint Imagination

Model ReleasesDGX agent

arXiv:2606.10862v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models achieve strong performance on standard manipulation benchmarks, but most evaluations assume that task-relevant obj

Parametric Knowledge is Not All You Need: Toward Honest Large Language Models via Retrieval of Pretraining Data

Model ReleasesDGX agent

arXiv:2601.21218v2 Announce Type: replace Abstract: Large language models (LLMs) are highly capable of answering questions, but they are often unaware of their own knowledge boundary, i.e., knowing wh

ReflectiChain: Epistemic Grounding in LLM-Driven World Models for Supply Chain Resilience

Model ReleasesDGX agent

arXiv:2606.10359v1 Announce Type: new Abstract: AI agents in supply chains face a fundamental epistemic gap: large language models (LLMs) interpret policies but lack physical grounding, while reinforc

Rod models in continuum and soft robot control: a review

ApplicationsDGX agent

arXiv:2407.05886v3 Announce Type: replace Abstract: Continuum and soft robots can transform automation tasks requiring compliant interaction in constrained or unstructured environments, including heal

Vehicle Prediction Model for Enhanced MPC Path Tracking in Formula Student Driverless

AgentsDGX agent

arXiv:2606.10732v1 Announce Type: new Abstract: Autonomous race cars, such as in Formula Student Driverless, operate close to their physical handling limits. The resulting highly nonlinear vehicle beh

When Metrics Disagree: A Meta-Analysis of Knowledge-Graph-Completion Model Benchmarking

ResearchDGX agent

arXiv:2606.10287v1 Announce Type: cross Abstract: Evaluating Knowledge Graph Completion (KGC) models remains challenging because standard assessment relies on isolated rank-based metrics such as MRR,

9 Jun 2026

Bayesian Optimization of a Multi-Product Chemical Reactor Using Composite Models and Partial Physics Knowledge

Model ReleasesDGX agent

arXiv:2606.08611v1 Announce Type: cross Abstract: We study data-driven real-time economic optimization of a multi-product chemical reactor when no reliable first-principles model is available beyond a

Benchmarking Empirical Privacy Protection for Adaptations of Large Language Models

Model ReleasesDGX agent

arXiv:2606.09401v1 Announce Type: new Abstract: Recent work has applied differential privacy (DP) to adapt large language models (LLMs) for sensitive applications, offering theoretical guarantees. How

BLUE: Toward Better Language Use in Efficient Vision-Language-Action Models for Autonomous Driving

Model ReleasesDGX agent

arXiv:2606.08684v1 Announce Type: new Abstract: We present BLUE, a minimal method for better language use in vision-language-action (VLA) models for autonomous driving (AD). Through extensive analysis

@cohere Nice! Great to see another open source model released. 🙌

Model ReleasesDGX agent

Cohere announced the release of another open source model, receiving positive reception from the community. The post was shared on X (formerly Twitter) and highlights Cohere's continued contribution t

Dendrograms of Mixing Measures for Softmax-Gated Gaussian Mixture of Experts: Consistency Without Model Sweeps

Model ReleasesDGX agent

arXiv:2510.12744v2 Announce Type: replace-cross Abstract: We develop a unified statistical framework for softmax-gated Gaussian mixture of experts (SGMoE) that addresses three long-standing obstacles

DIYHealth Suite: Dataset, Model, and Benchmark for Health Management at Home

Model ReleasesDGX agent

arXiv:2606.07542v1 Announce Type: cross Abstract: Generative AI is reshaping healthcare, yet most existing advances rely on hospital-grade devices, which limits their accessibility and potential for h

Dream-Tac: A Unified Tactile World Action Model for Contact-Rich Robot Manipulation

SafetyDGX agent

arXiv:2606.08737v1 Announce Type: new Abstract: World action models inherit the predictive capability of world models, enabling action generation to be guided by anticipated future observations. Howev

From `May' to `Is': Certainty Distortion in Language Model Rewriting

Model ReleasesDGX agent

arXiv:2606.07951v1 Announce Type: cross Abstract: Humans increasingly turn to Language Models (LMs) in ways that shape beliefs and drive decisions, including discussing, rewriting, and summarizing inf

GraphLoRA: Structure-Aware Low-Rank Adaptation for Large Language Model Recommendation

Model ReleasesDGX agent

arXiv:2606.07526v1 Announce Type: cross Abstract: Large Language Models (LLMs) have shown strong potential for recommendation (LLMRec) due to their powerful reasoning and generalization abilities. How

Introducing Gemma 4 12B: a unified, encoder-free multimodal model

Model ReleasesDGX agent

Gemma 4 12B is a unified, encoder-free multimodal model designed to bring high-performance intelligence to laptops and released under an Apache 2.0 license. It eliminates separate encoders by projecti

LEAF: Growing Trees Without Branching for Speech-Aware Large Language Model Post-Training

Model ReleasesDGX agent

arXiv:2606.07610v1 Announce Type: cross Abstract: State-of-the-art GRPO-style methods for speech-aware large language model post-training suffer from coarse credit assignment, broadcasting the same te

MLingualFC: Evaluating Jailbreak Vulnerabilities in Multilingual Vision-Language Models

Model ReleasesDGX agent

arXiv:2606.07706v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have demonstrated strong performance across multimodal tasks, yet their safety robustness remains an open challenge. Whi

← Previous
1…8384858687…1009
Next →