AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,405 results
2 Jun 2026

ProbeScale: Probing Analysis to Optimize Neural Scaling Laws for Efficient Small Language Model Inference

Model ReleasesDGX agent

arXiv:2606.01806v1 Announce Type: cross Abstract: Small Language Models (SLMs) offer a balance between capability and computational feasibility. Neural scaling laws inform their optimal training, sugg

RA-LWLM: Retrieval-Augmented In-Context Localization with Wireless Foundation Models

Local AiDGX agent

arXiv:2606.01899v1 Announce Type: cross Abstract: Wireless localization is a fundamental capability of sixth-generation (6G) networks. Conventional model-based methods require accurate modeling of the

Revise, Don't Freeze: Sampler-Matched Training for Self-Correcting Masked Diffusion Language Models

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.01026v1 Announce Type: new Abstract: Masked diffusion language models (MDLMs) re-predict every position at each denoising step, but standard samplers commit tokens once revealed, leaving th

RoboSemanticBench: Diagnosing Semantic Grounding in Action Prediction for VLA Models

Model ReleasesDGX agent

arXiv:2606.02277v1 Announce Type: new Abstract: Vision-language-action (VLA) models are built on the premise that semantic understanding from pretrained language or vision-language backbones should gu

Scaling Parallel Sequence Models to Foundation-Scale Vision Encoders

Model ReleasesDGX agent

arXiv:2606.00746v1 Announce Type: new Abstract: Vision foundation models are bottlenecked by the quadratic cost of self-attention, which limits usable resolution and increases the cost of large-scale

Simple Recipe Works: Vision-Language-Action Models are Natural Continual Learners with Reinforcement Learning

Model ReleasesDGX agent

arXiv:2603.11653v2 Announce Type: replace Abstract: Continual Reinforcement Learning (CRL) for Vision-Language-Action (VLA) models is a promising direction toward self-improving embodied agents that c

SimSD: Simple Speculative Decoding in Diffusion Language Models

ResearchDGX agent

arXiv:2606.02544v1 Announce Type: cross Abstract: Diffusion large language models (dLLMs) have recently emerged as a promising alternative to autoregressive (AR) LLMs, offering faster inference throug

Tiny Recursive Models for Solving the J2-Perturbed Lambert Problem

Model ReleasesDGX agent

arXiv:2606.00895v1 Announce Type: cross Abstract: This paper presents a fast, recursive neural solver for the J2-perturbed Lambert problem based on Tiny Recursive Models (TRM), termed the TRM-Perturbe

Towards 3D-Aware Video Diffusion Models: Render-Free Human Motion Control with Mesh Tokenization

ResearchDGX agent

arXiv:2606.02000v1 Announce Type: cross Abstract: Diffusion models have shown remarkable success in video generation. However, whether such models are truly aware of the 3D structure underlying visual

Towards Interactive Video World Modeling: Frontiers, Challenges, Benchmarks, and Future Trends

AgentsDGX agent

arXiv:2606.01164v1 Announce Type: new Abstract: With rapid development of large language models and diffusion-based content generation, world modeling has attracted increasing research attention, bene

Understanding-Enhanced Model Collaboration for Long-Tailed Egocentric Mistake Detection

Local AiDGX agent

arXiv:2606.02120v1 Announce Type: cross Abstract: In this report, we address the problem of determining whether a user performs an action incorrectly from egocentric video data. To this end, we propos

Verifying Meta-Awareness via Predictive Rewards in Reasoning Models

ResearchDGX agent

arXiv:2510.03259v2 Announce Type: replace-cross Abstract: Recent research on reasoning models explores the meta-awareness of language models, including their ability to determine optimal thinking dura

Why Financial Institutions Are Converging on Transaction Foundation Models to Build Their Own Intelligence

ApplicationsDGX agent

Financial institutions have spent years building AI: fraud models, credit models, recommendation engines and risk systems. While this sprawl of task-specific models has been effective, it’s also const

1 Jun 2026

Chinese AI developer MiniMax launches M3, a new coding model that it says rivals Opus 4.7, costing 0.12 per 1M input tokens, compared with 5 for Opus 4.7 (Juro Osawa/The Information)

Model ReleasesDGX agent

Juro Osawa / The Information: Chinese AI developer MiniMax launches M3, a new coding model that it says rivals Opus 4.7, costing 0.12 per 1M input tokens, compared with 5 for Opus 4.7 — Chinese AI dev

Divergence Decoding: Inference-Time Unlearning via Auxiliary Models

ResearchDGX agent

arXiv:2605.31293v1 Announce Type: new Abstract: Large Language Models (LLMs) frequently memorize sensitive training data thereby creating significant privacy and copyright risks. Addressing these risk

ERGeoBench:A Comprehensive Benchmark for Embodied Reasoning and Geo-localization in Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2605.31251v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) have shown strong potential as embodied agents, yet embodied geo-localization remains underexplored due to th

ExpGraph: Model-Agnostic Experience Learning with Graph-Structured Memory for LLM Agents

Model ReleasesDGX agent

arXiv:2605.30712v1 Announce Type: new Abstract: Large language model (LLM) agents have shown strong capabilities in reasoning, tool use, and multi-step interaction, but they often solve tasks from scr

Generating Graph-like Rules for Knowledge Graph Reasoning via Diffusion Models

Model ReleasesDGX agent

arXiv:2605.30747v1 Announce Type: new Abstract: Logical rules constitute a cornerstone of knowledge graph (KG) reasoning, valued for their interpretability and ability to model relational patterns. Ho

In @latentspacepod podcast, I shared my view on video generation, world models, LLMs, agents, continual learning and where the next frontier…

AgentsDGX agent

In @latentspacepod podcast, I shared my view on video generation, world models, LLMs, agents, continual learning and where the next frontier is. 1. Video models get most of their intelligence from lan

👏👏 Introducing Qwen3.7-Plus — a multimodal agent model that unifies vision and language into one versatile agent foundation. ✅ Multimodal …

Model ReleasesDGX agent

👏👏 Introducing Qwen3.7-Plus — a multimodal agent model that unifies vision and language into one versatile agent foundation. ✅ Multimodal interactive hybrid agent: unified GUI & CLI operation across v

MedFact: Benchmarking the Fact-Checking Capabilities of Large Language Models on Chinese Medical Texts

Model ReleasesDGX agent

arXiv:2509.12440v3 Announce Type: replace-cross Abstract: Deploying Large Language Models (LLMs) in medical applications requires fact-checking capabilities to ensure patient safety and regulatory com

Model page: https://ollama.com/library/minimax-m3

Local AiDGX agent

Minimax-M3 is a model available through Ollama's model library, accessible at ollama.com/library/minimax-m3. The model was promoted or announced via Ollama's official X (Twitter) account, indicating i

Modeling a digital twin of a food supply chain using BigQuery Graph

SafetyDGX agent

The example of a growing restaurant Imagine you are running a restaurant chain. You just can't physically feel and touch things to know how your business operates. You need tools and a digital replica

Nvidia releasesCosmos3-Super-Text2Image model . 64 billion paramteres

HardwareDGX agent

Cosmos3-Super-Text2Image is a 64 billion parameter omnimodal world model capable of generating high-quality images from text inputs as part of NVIDIA's Cosmos 3 foundation model platform. The model is

SOCO: Benchmarking Semantic Object Correspondence in Vision Foundation Models

Model ReleasesDGX agent

arXiv:2605.31597v1 Announce Type: new Abstract: Measuring structured object understanding in vision foundation models remains challenging due to inconsistent evaluation protocols and limited part-leve

TeachObs: A Human-Validated Benchmark for Multimodal Teaching Observation and Model Evaluation

Model ReleasesDGX agent

arXiv:2605.30673v1 Announce Type: new Abstract: Classroom videos contain observable teaching practices, but their pedagogical and visual signals are rarely organized in forms suitable for model evalua

29 May 2026

Adaptive Targeted Dynamic Chunking for Tokenization-Free Hierarchical Model

ResearchDGX agent

arXiv:2605.30080v1 Announce Type: new Abstract: Tokenization-free hierarchical models are emerging as a promising alternative to traditional Large Language Models (LLMs), addressing inherent preproces

Another proof point for the open-weights thesis. From @RampLabs: 'If we built this again, we'd lean more on open-weight models.' Ramp pointe…

Model ReleasesDGX agent

Another proof point for the open-weights thesis. From @RampLabs: 'If we built this again, we'd lean more on open-weight models.' Ramp pointed 10K agents at their own backend. Kimi K2.6 and DeepSeek V4

Bayesian model selection and misspecification testing in imaging inverse problems only from noisy and partial measurements

ResearchDGX agent

arXiv:2510.27663v3 Announce Type: replace-cross Abstract: Modern imaging techniques heavily rely on Bayesian statistical models to address difficult image reconstruction and restoration tasks. This pa

Controlling the Risk of Corrupted Contexts for Language Models via Early-Exiting

ResearchDGX agent

arXiv:2510.02480v3 Announce Type: replace Abstract: Large language models (LLMs) can be influenced by harmful or irrelevant context, which can significantly harm model performance on downstream tasks.

Embodied3DBench: Benchmarking Low-Level Embodied Spatial Intelligence of Vision Language Models

Model ReleasesDGX agent

arXiv:2605.29074v1 Announce Type: new Abstract: Are current Vision Language Models (VLMs) ready to comprehend and reason about complex embodied interactions in 3D environments? We introduce Embodied3D

MATNet: Multi-Level Fusion Transformer-Based Model for Day-Ahead PV Generation Forecasting

Model ReleasesDGX agent

arXiv:2306.10356v3 Announce Type: replace-cross Abstract: Accurate forecasting of renewable generation is crucial to facilitate the integration of Renewable Energy Sources into the power system. Focus

MELD: Mel-Spectrogram-Based Speech Language Modeling with Discrete Latent Variables

ResearchDGX agent

arXiv:2605.29859v1 Announce Type: cross Abstract: Recent speech language models rely on encoders that are optimized separately from autoregressive models. Since these encoders are unaware of the downs

MENTOR: Efficient Multimodal-Conditioned Tuning for Autoregressive Vision Generation Models

Model ReleasesDGX agent

arXiv:2507.09574v3 Announce Type: replace-cross Abstract: Recent text-to-image models produce high-quality results but still struggle with precise visual control, balancing multimodal inputs, and requ

Rethinking Stepwise Model Routing: A Cost-Efficient Table Reasoning Perspective

ResearchDGX agent

arXiv:2605.29319v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) achieve strong performance on table reasoning tasks but incur substantial inference cost due to long reasoning traces. Ste

The Chain Holds, the Answer Folds: Trace-Answer Dissociation in Reasoning Models Under Adversarial Pressure

Model ReleasesDGX agent

arXiv:2605.29087v1 Announce Type: new Abstract: Reasoning models are evaluated on single-turn benchmarks but deployed in multi-turn dialogue, where users push back on correct answers. Under sustained

UA-Legal-Bench: A Benchmark for Evaluating Large Language Models on Ukrainian Legal Reasoning

Model ReleasesDGX agent

arXiv:2605.29170v1 Announce Type: cross Abstract: Legal NLP benchmarks are overwhelmingly English-centric, leaving failure modes in morphologically rich, non-Latin-script languages undetected. We intr

VLAConf: Calibrated Task-Success Confidence for Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2605.29605v1 Announce Type: new Abstract: Confidence estimation for Vision-Language-Action (VLA) models is essential for robots to perform manipulation tasks in the open world, providing crucial

28 May 2026

A Methodology to Assess Power Modeling in Energy-Aware Federated Learning on Heterogeneous Mobile Devices

ResearchDGX agent

arXiv:2605.27601v1 Announce Type: cross Abstract: Estimating CPU power on heterogeneous ARM-based commodity devices is challenging due to limited access to CPU's voltage domains. As a result, state-of

Apple working to cram massive Gemini model into iPhone to power new Siri

Model ReleasesDGX agent

Apple is reportedly working to distill knowledge and skills from Google's larger Gemini model into a smaller version that could run on iPhones. The new Siri will use a tiered system where simple tasks

Continuous Diffusion Models Can Obey Formal Syntax

TutorialsDGX agent

arXiv:2602.12468v2 Announce Type: replace Abstract: Diffusion language models offer a promising alternative to autoregressive models due to their global, non-causal generation process, but their conti

Debate with Images: Detecting Deceptive Behaviors in Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2512.00349v2 Announce Type: replace Abstract: Are frontier AI systems becoming more capable? Certainly. Yet such progress is not an unalloyed blessing but rather a Trojan horse: behind their per

Dynamic Topic Modeling with a Higher-Order Hypergraphical Representation

ResearchDGX agent

arXiv:2605.28269v1 Announce Type: new Abstract: Dynamic topic modeling is widely used to analyze evolving trends in scientific literature, medical records, and social media. Traditional topic models r

Faster Thermal Profiling of a Lunar Rover with Machine Learning Adapted Finite Difference Model

AgentsDGX agent

arXiv:2605.27651v1 Announce Type: new Abstract: Autonomous space systems operating in extreme thermal environments require accurate and efficient thermal modeling to support both pre-mission system de

Forget to Know, Remember to Use: Context-Aware Unlearning for Large Language Models

ResearchDGX agent

arXiv:2510.17620v2 Announce Type: replace Abstract: Large language models may encode sensitive information or outdated knowledge that needs to be removed, to ensure responsible and compliant model res

Harness-Bench: Measuring Harness Effects across Models in Realistic Agent Workflows

Model ReleasesDGX agent

arXiv:2605.27922v1 Announce Type: new Abstract: LLM agents are increasingly deployed as executable systems that use tools, modify workspaces, and produce concrete artifacts. In such workflows, perform

NUCLEUS-MoE: Unified Model of Pool Boiling for Liquid Cooling

ResearchDGX agent

arXiv:2605.27722v1 Announce Type: new Abstract: Two-phase boiling enables heat transfer rates an order of magnitude higher than single-phase cooling, but it remains difficult to model due to the stron

Parameter-Efficient Generative Modeling with Controlled Vector Fields

Model ReleasesDGX agent

arXiv:2605.28267v1 Announce Type: new Abstract: We introduce a continuous-time generative modeling framework, motivated by the Chow-Rashevskii theorem, that builds expressive flows from a small set of

Probing for Knowledge Attribution in Large Language Models

Model ReleasesDGX agent

arXiv:2602.22787v2 Announce Type: replace-cross Abstract: Large language model (LLM) hallucinations, meaning fluent but factually incorrect generations, fall into two types: faithfulness violations, w

Prompting Is All You Need: Multi-view Prompting Large Language Models for Aspect-Based Sentiment Analysis

Model ReleasesDGX agent

arXiv:2605.28058v1 Announce Type: new Abstract: Recent work explored the capabilities of Large Language Models (LLMs) in Aspect-Based Sentiment Analysis (ABSA) through few-shot prompting, requiring su

Sentence Curve Language Models

Local AiDGX agent

arXiv:2602.01807v3 Announce Type: replace Abstract: Language models (LMs) are a central component of modern AI systems, and diffusion language models (DLMs) have recently emerged as a competitive alte

Turning Video Models into Generalist Robot Policies

SafetyDGX agent

arXiv:2605.27817v1 Announce Type: cross Abstract: Video generative models have emerged as a promising robotics backbone, capable of generating videos that depict the completion of complex tasks across

We're selectively releasing the Paris 2.0 weights and partnering with researchers and teams interested in diffusion-based video models, worl…

IndustryDGX agent

We're selectively releasing the Paris 2.0 weights and partnering with researchers and teams interested in diffusion-based video models, world models, and embodied agents. The model is on Hugging Face:

Why Gaussian Diffusion Models Fail on Discrete Data and How to Prevent It?

TutorialsDGX agent

arXiv:2604.02028v2 Announce Type: replace Abstract: Diffusion models have become a standard approach for generative modeling in continuous domains, yet their application to discrete data remains chall

27 May 2026

From Scores to Gibbs Correctors: Accelerating Uniform-Rate Discrete Diffusion Models

ResearchDGX agent

arXiv:2605.27352v1 Announce Type: new Abstract: Discrete diffusion models have achieved strong empirical performance in text and other symbolic domains, but, especially for uniform-rate models, they o

Gemini Embedding 2: A Native Multimodal Embedding Model from Gemini

Model ReleasesDGX agent

arXiv:2605.27295v1 Announce Type: new Abstract: We introduce Gemini Embedding 2, a native multimodal embedding model that allows embedding video, audio, image, and text modalities in a unified represe

KARMA: Karma-Aligned Reward Model Adaptation

SafetyDGX agent

arXiv:2605.26738v1 Announce Type: new Abstract: Human communication depends on implicit social signals where effectiveness is shaped by tone, context, and conversational norms rather than semantic con

Large Language Model-Powered Query-Driven Event Timeline Summarization in Industrial Search

Model ReleasesDGX agent

arXiv:2605.27066v1 Announce Type: new Abstract: Understanding how events evolve over time is essential for search engines handling queries about trending news. We present QDET (Query-Driven Event Time

Learning When to Think While Listening in Large Audio-Language Models

Model ReleasesDGX agent

arXiv:2605.27190v1 Announce Type: cross Abstract: Recent advances in Large Audio-Language Models (LALMs) have made real-time, streaming spoken interaction increasingly practical. In this setting, reas

LiveK12Bench: Have Large Multimodal Models Truly Conquered High School-level Examinations?

Model ReleasesDGX agent

arXiv:2605.26781v1 Announce Type: new Abstract: Advanced Large Multimodal Models (LMMs) have demonstrated impressive performance in K-12 reasoning tasks, exhibiting great promise as intelligent tutors

← Previous
1…6667686970…1007
Next →