AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,490 results
30 Jun 2026

The food delivery leader in China just dropped an open weights 1.6 trillion parameter model while half the USA was sleeping … 🫨 🇨🇳

Model ReleasesDGX agent

The food delivery leader in China just dropped an open weights 1.6 trillion parameter model while half the USA was sleeping … 🫨 🇨🇳 Meituan, China's largest food delivery platform, open-sourced a 1.6 t

Toward Secure and Reliable PDDL Formalization of Large Language Models with Planner-in-the-Loop Feedback

Model ReleasesDGX agent

arXiv:2606.29700v1 Announce Type: new Abstract: Planning often requires symbolic specifications that are both executable and verifiable. For large language models deployed in autonomous or decision-su

Towards Evaluating Data Priors for Tabular Foundation Models

ResearchDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.29241v1 Announce Type: new Abstract: Data-generating priors are a central component of tabular foundation models because they define the task distribution used during pretraining. However,

X-Mind: Efficient Visual Chain-of-Thought via Predictive World Model for End-to-End Driving

SafetyDGX agent

arXiv:2606.28758v1 Announce Type: cross Abstract: Predicting future states is essential for autonomous agents, yet current Vision-Language-Action (VLA) models fundamentally lack this capability, relyi

You can embed this model practically anywhere - like a chrome extension This is transformersjs + rampart for real-time PII removal in your b…

ApplicationsDGX agent

You can embed this model practically anywhere - like a chrome extension This is transformersjs + rampart for real-time PII removal in your browser (personal email still blurred) You can see I'm toggli

29 Jun 2026

A Comprehensive Survey on World Models for Embodied AI

AgentsDGX agent

arXiv:2510.16732v3 Announce Type: replace Abstract: Embodied AI requires agents that perceive, act, and anticipate how actions reshape future world states. World models serve as internal simulators th

AirGroundBench: Probing Spatial Intelligence in Multimodal Large Models under Heterogeneous Multi-View Embodied Collaboration

Model ReleasesDGX agent

arXiv:2606.28049v1 Announce Type: new Abstract: In recent years, multimodal large language models (MLLMs) have shown strong potential for embodied intelligence, yet their ability to maintain geometric

CalBrief: A Pilot Diagnostic Benchmark for Evidence-Calibrated Scientific Briefing with Large Language Models

Model ReleasesDGX agent

arXiv:2606.27383v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as research assistants, yet it remains unclear whether they can calibrate research takeaways to the

Continual Learning for Sequential Personalization of Small Language Models: A Stability Monitoring Analysis

Local AiDGX agent

arXiv:2606.27634v1 Announce Type: new Abstract: Small Language Models (SLMs) are increasingly being considered for deployment on edge devices such as laptops, enabling private, low-latency, and locall

DeepSeek details DSpark, a speculative decoding framework for its V4 models, saying it speeds up AI inference by up to 85% and was tested on Gemma and Qwen (Ben Jiang/South China Morning Post)

Model ReleasesDGX agent

Ben Jiang / South China Morning Post: DeepSeek details DSpark, a speculative decoding framework for its V4 models, saying it speeds up AI inference by up to 85% and was tested on Gemma and Qwen — Chin

Foundation vs. Specialized Models: Evaluating Catastrophic Forgetting in Continual Time Series Forecasting

ApplicationsDGX agent

arXiv:2510.00809v3 Announce Type: replace Abstract: While Time Series Foundation Models (TSFMs) excel in zero-shot tasks, their behavior under continual fine tuning is poorly understood. We present th

From Signals to Transfer: A Factorised Study of Probe-Based Uncertainty Estimation in Large Language Models

Model ReleasesDGX agent

arXiv:2606.27679v1 Announce Type: cross Abstract: Probe-based uncertainty estimation (UE) has emerged as a prominent approach to detect hallucinations in Large Language Models (LLMs) by learning uncer

Hippocampus-DETR: An Explicit Memory Object Detection Framework Based on Hippocampus Modeling

ResearchDGX agent

arXiv:2606.27831v1 Announce Type: cross Abstract: This paper addresses the lack of explicit memory mechanisms in current object detection models and proposes Hippocampus-DETR, a novel detection framew

Large Language Model Teaches Visual Students: Cross-Modality Transfer of Fine-Grained Conceptual Knowledge

ResearchDGX agent

arXiv:2606.27527v1 Announce Type: cross Abstract: Large Language Models (LLMs) possess broad conceptual knowledge acquired through large-scale text pretraining, yet their potential to supervise models

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments

Model ReleasesDGX agent

arXiv:2606.27537v1 Announce Type: new Abstract: Video generation models aspire to simulate dynamic environments, and several benchmarks now evaluate memory consistency across frames. However, most ass

MVPruner: Dynamic Token Pruning for Accelerating Multi-view Vision-Language Models in Autonomous Driving

Model ReleasesDGX agent

arXiv:2606.27660v1 Announce Type: new Abstract: Vision-Language Models (VLMs) improve generalization and interpretability in autonomous driving but suffer from efficiency issues due to long visual tok

RAE-NWM: Navigation World Model in Dense Visual Representation Space

TutorialsDGX agent

arXiv:2603.09241v2 Announce Type: replace Abstract: Visual navigation requires agents to reach goals in complex environments through perception and planning. World models address this task by simulati

RECAST: Model Reconstruction via Counterfactual-Aware Wasserstein Geometry under Limited Data

SafetyDGX agent

arXiv:2606.27948v1 Announce Type: new Abstract: Counterfactual explanations (CFs) help understand machine learning models by identifying minimal input changes that would lead to alternative model outc

SemCityLoc: Aerial 6DoF Localization Using Semantic 3D City Models

Model ReleasesDGX agent

arXiv:2606.27444v1 Announce Type: new Abstract: Aerial 6DoF localization typically relies on precise GNSS signals or radiometrically rich 3D reconstructions, limiting scalability and on-board deployme

SurgXBench: Explainable Vision-Language Model Benchmark for Surgery

Model ReleasesDGX agent

arXiv:2505.10764v4 Announce Type: replace Abstract: Innovations in digital intelligence are transforming robotic surgery with more informed decision-making. Real-time awareness of surgical instrument

This is smart from Cline. They just launched ClinePass, which makes it easy to access the latest open-weight models like GLM 5.2, Kimi k2.7-…

Model ReleasesDGX agent

This is smart from Cline. They just launched ClinePass, which makes it easy to access the latest open-weight models like GLM 5.2, Kimi k2.7-code, Mimo 2.5, Deepseek v4 pro, Minimax M3, and more. Alway

We have seen multi model harnesses for cheaper & faster tasks What about for the hardest challenges? What about open source? Proud to share …

IndustryDGX agent

We have seen multi model harnesses for cheaper & faster tasks What about for the hardest challenges? What about open source? Proud to share the latest update our Zenith harness, taking models you can

27 Jun 2026

Sakana Fugu Technical Report Instead of training one larger model, Sakana AI trains an orchestrator that reads each query and dynamically ro…

Model ReleasesDGX agent

Sakana Fugu Technical Report Instead of training one larger model, Sakana AI trains an orchestrator that reads each query and dynamically routes or composes GPT-5.5, Gemini-3.1-Pro, Claude Opus 4.8 an

26 Jun 2026

Can Large Language Models Reliably Code Qualitative Humanitarian Data? A Benchmark Study Against Human Expert Adjudication

Model ReleasesDGX agent

arXiv:2606.26541v1 Announce Type: new Abstract: Data from affected populations are crucial for informing humanitarian response, but their value depends on timely and consistent interpretation of nuanc

Don't Settle at the Mode! Mitigating Diversity Collapse in Pretrained Flow Models via Feature Self-Guidance

SafetyDGX agent

arXiv:2606.27371v1 Announce Type: new Abstract: State-of-the-art flow models generate stunning images from text or image prompts. However, they suffer from diversity collapse when generating multiple

From Structure to Synergy: A Survey of Vision-Language Perception Paradigm Evolution in Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2606.26196v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have recently made remarkable progress in unifying vision-language understanding and reasoning, especially fo

Heterogeneous Neural Predictivity from Language Models During Naturalistic Comprehension

Local AiDGX agent

arXiv:2606.26880v1 Announce Type: new Abstract: Language-model representations provide structured, high-dimensional annotations of naturalistic language stimuli and can serve as informative neural pre

NuclearQAv2: A Structured Benchmark for Evaluating Domain-Science Competence in Large Language Models

Model ReleasesDGX agent

arXiv:2606.27047v1 Announce Type: cross Abstract: Large language models (LLMs) have demonstrated strong performance across a wide range of tasks, but ensuring their reliability in highly technical dom

Refusal Lives Downstream of Persona in Chat Models

Model ReleasesDGX agent

arXiv:2606.26161v1 Announce Type: new Abstract: Linear directions in activation space have been identified for both refusal and persona traits in instruction-tuned chat models, but the two have been s

SSM Adapters via Hankel Reduced-order Modeling: Injection Site Determines Task Suitability in Long-Context Fine-Tuning

Model ReleasesDGX agent

arXiv:2606.26290v1 Announce Type: cross Abstract: While parameter-efficient fine-tuning (PEFT) typically targets attention projectors, its efficacy for tasks requiring sequential state accumulation re

Unconventional AI debuts oscillator-based Un-0 model series

Model ReleasesDGX agent

Unconventional AI Inc. has developed an artificial intelligence architecture that could improve the power efficiency of image generation models. The technology is the basis of a new neural network ser

Where Larger Models Excel: The Primacy of Constraint-Guided Reasoning

ResearchDGX agent

arXiv:2606.26108v1 Announce Type: new Abstract: Larger language models consistently outperform smaller ones on reasoning benchmarks, yet the reasoning differences underlying this gap remain underexplo

25 Jun 2026

Adaptive Oscillatory Inductive Bias for Modeling Sharp Prosodic Dynamics in Diffusion-Based TTS

SafetyDGX agent

arXiv:2606.25424v1 Announce Type: cross Abstract: Diffusion-based text-to-speech (TTS) models have achieved significant improvements in speech quality. However, modeling sharp prosodic transitions and

Agent-as-a-Router: Agentic Model Routing for Coding Tasks

Model ReleasesDGX agent

arXiv:2606.22902v2 Announce Type: replace Abstract: Real-world users typically have access to multiple Large Language Models (LLMs) from different providers, and these LLMs often excel at distinct dom

Beyond Next-Observation Prediction: Agent-Authored World Modeling for Sequential Decision Making

SafetyDGX agent

arXiv:2606.25421v1 Announce Type: new Abstract: Recent studies on world modeling for Large Language Model (LLM) agents typically formulate the learning objective as next-observation prediction. Howeve

How Small Can 6G Reason? Scaling Tiny-to-Small Language Models for AI-Native Networks

Model ReleasesDGX agent

arXiv:2603.02156v2 Announce Type: replace-cross Abstract: Emerging 6G visions, reflected in ongoing standardization efforts within 3GPP, IETF, ETSI, ITU-T, and the O-RAN Alliance, increasingly charact

MIMFlow: Integrating Masked Image Modeling with Normalizing Flows for End-to-End Image Generation

ResearchDGX agent

arXiv:2606.26016v1 Announce Type: new Abstract: Normalizing Flows (NFs) are powerful generative models capable of exact density estimation and sampling. However, their strict invertibility often force

PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models

Model ReleasesDGX agent

arXiv:2606.25442v1 Announce Type: new Abstract: Safety alignment of large language models (LLMs) typically depends on high-quality supervision data, such as safe demonstrations or preference pairs. Ho

RotRNN: Modelling Long Sequences with Rotations

ResearchDGX agent

arXiv:2407.07239v3 Announce Type: replace Abstract: Linear recurrent neural networks, such as State Space Models (SSMs) and Linear Recurrent Units (LRUs), have recently shown state-of-the-art performa

The underrated part of this announcement is that Fireworks has been quietly great behind the scenes helping us eval and serve these models B…

Model ReleasesDGX agent

The underrated part of this announcement is that Fireworks has been quietly great behind the scenes helping us eval and serve these models Big kudos to the Fireworks team! Kimi K2.7 Code and GLM 5.2 a

Toward Low-Latency Vision-Language Models with Doubly-Correct Predictions in Egocentric Visual Understanding

Model ReleasesDGX agent

arXiv:2606.25160v1 Announce Type: cross Abstract: The rapid rise of Vision-Language Models (VLMs) in egocentric visual understanding has made low-latency inference in human-robot collaborative (HRC) t

When Do Conservation Laws Survive Learned Representations? Certified Horizons for Latent World Models

SafetyDGX agent

arXiv:2606.24945v1 Announce Type: new Abstract: We ask a representation-learning question about physical world models: when does a conservation law remain certifiable after a model learns a latent rep

Yuvion VL: A Multimodal Foundation Model for Adversarial Content and AI Safety

SafetyDGX agent

arXiv:2606.25034v1 Announce Type: new Abstract: General-purpose models often struggle to reliably identify and understand real-world multimodal risks, largely due to the inherent multimodal adversaria

24 Jun 2026

A Physics-Informed Fourier-Wavelet Transformer for Multiscale Computational Fluid Dynamics Surrogate Modeling

Model ReleasesDGX agent

arXiv:2606.24696v1 Announce Type: cross Abstract: Physics-informed surrogate models can accelerate computational fluid dynamics simulations. However, many existing methods reproduce global flow patter

A specialized reasoning large language model for accelerating rare disease diagnosis: a randomized AI physician assistance trial

Model ReleasesDGX agent

arXiv:2606.24510v1 Announce Type: new Abstract: Rare diseases affect millions of individuals worldwide, yet timely diagnosis remains a major public health challenge due to scarcity of specialized clin

DREAM: Dense Retrieval Embeddings via Autoregressive Modeling

ResearchDGX agent

arXiv:2606.24667v1 Announce Type: new Abstract: Dense retrieval embedding models are a fundamental component of modern retrieval-based AI systems. Most dense retrievers are trained with contrastive ob

Flood Mapping from RGB imagery using a Vision Foundation Model

ResearchDGX agent

arXiv:2606.24120v1 Announce Type: new Abstract: Timely, high-resolution maps of flood extent around settlements are essential for emergency response and damage assessment. We consider airborne RGB ima

GeoT2V-Bench: Benchmarking 3D Consistency in Text-to-Video Models via 3D Reconstruction

Model ReleasesDGX agent

arXiv:2606.24829v1 Announce Type: new Abstract: Camera-prompted text-to-video (T2V) models are increasingly used to synthesize virtual camera captures, such as orbiting objects or moving through stati

GLM-5.2 is now available in Cursor. The model has performed strongly on OpenRouter's Cursor usage rankings over the past week. Would love to…

Model ReleasesDGX agent

GLM-5.2 is now available in Cursor. The model has performed strongly on OpenRouter's Cursor usage rankings over the past week. Would love to hear comparisons of the experience using BYOK (GLM Coding P

Performance and Interpretability of Convolutional, Transformer, and Hybrid Deep Learning Models in Colorectal Histology Classification

Model ReleasesDGX agent

arXiv:2606.23744v1 Announce Type: cross Abstract: Deep learning has become an important tool in computational pathology, enabling automated analysis of histopathological images. While convolutional ne

PORTER: Language-Grounded Event Representations for Portable Structured EHR Foundation Models

ResearchDGX agent

arXiv:2606.24102v1 Announce Type: new Abstract: Most electronic health record (EHR) foundation models encode clinical events as discrete event tokens from a fixed vocabulary and therefore cannot direc

RetiSEM: Generalising Causal Models for Fragmented Biomedical Data

Model ReleasesDGX agent

arXiv:2606.24488v1 Announce Type: cross Abstract: Learning causal models from fragmented biomedical data is challenging because clinical, molecular, and imaging variables are often incomplete or not j

Towards Fast and Effective Long Video Understanding of Multimodal Large Language Models via Adaptive Quasi-Gaussian Sampling

Model ReleasesDGX agent

arXiv:2606.24187v1 Announce Type: new Abstract: Long video understanding remains a daunting challenge for Multimodal Large Language Models (MLLMs) due to the excessive computation and memory footprint

23 Jun 2026

A polarity-aware multi-relational model for the signed interaction prediction in biological networks

ApplicationsDGX agent

arXiv:2407.07357v4 Announce Type: replace Abstract: Predicting signed interactions in biological networks is crucial for understanding drug mechanisms and facilitating drug repurposing. While deep gra

BioMatrix: Towards a Comprehensive Biological Foundation Model Spanning the Modality Matrix of Sequences, Structures, and Language

ResearchDGX agent

arXiv:2606.22138v1 Announce Type: cross Abstract: We present BioMatrix, the first multimodal foundation model that natively integrates sequences, structures, and natural language for both molecules an

Black-Box Continual Learning for Vision-Language Models

Model ReleasesDGX agent

arXiv:2606.22999v1 Announce Type: new Abstract: The rapid deployment of Vision-Language Models (VLMs) in dynamic environments necessitates the ability to learn continuously without forgetting. However

Decoupling the Declarative from the Procedural in Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2606.21496v1 Announce Type: cross Abstract: Deploying generalist robotic agents in the real world requires transferable skills. Specifically, a policy trained to clone a behavior from object-spe

Discrete State Diffusion Models: A Sample Complexity Perspective

ResearchDGX agent

arXiv:2510.10854v3 Announce Type: replace Abstract: Diffusion models have demonstrated remarkable performance in generating high-dimensional samples across domains such as vision, language, and the sc

Distributional Regression with Tabular Foundation Models: Evaluating Probabilistic Predictions via Proper Scoring Rules

ResearchDGX agent

arXiv:2603.08206v5 Announce Type: replace Abstract: Modern tabular foundation models such as TabPFN and TabICL naturally produce full predictive distributions, while the benchmarks used to evaluate th

Do Activation Monitors Survive Model Updates? Benchmarking, Predicting, and Repairing Activation-Monitor Staleness

SafetyDGX agent

arXiv:2606.15980v2 Announce Type: replace Abstract: Activation monitors -- lightweight probes trained on a language model's internal representations -- are an increasingly common layer in deployment s

← Previous
1…8283848586…1009
Next →