AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
59,904 results
10 Jul 2026

DeltaV: Thinking with Visual State Updates in Unified Large Multimodal Models

Model ReleasesDGX agent

arXiv:2607.08434v1 Announce Type: new Abstract: Current Unified Large Multimodal Models (ULMMs) support interleaved multimodal reasoning through textual reasoning and intermediate visual states, but t

Do You Need a Frontier Model as a Citation Verifier? Benchmarking Rubric LLMs for Deep-Research Source Attribution

Model ReleasesDGX agent

arXiv:2607.08700v1 Announce Type: new Abstract: Reinforcement learning increasingly relies on an LLM judge to score each rubric criterion, and that judge acts as the reward model during training. Befo

Improving Ad-hoc Search Effectiveness for Conversational Information Retrieval via Model Merging

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2607.08540v1 Announce Type: cross Abstract: Conversational information retrieval is challenging since it requires the consideration of the conversation history which potentially gives rise to to

When LLMs Agree, Are They Right? Auditing Self-Consistency and Cross-Model Agreement as Confidence Signals

Model ReleasesDGX agent

arXiv:2607.08065v1 Announce Type: new Abstract: LLM-as-judge (Zheng et al., 2023) is increasingly the default for evaluating AI systems in enterprise pipelines, often scaled to ensembles (Verga et al.

When Thinking Hurts: Epistemic Signals in the Reasoning Chains of Visual Language Models

Model ReleasesDGX agent

arXiv:2607.08059v1 Announce Type: cross Abstract: Uncertainty quantification for visual language models (VLMs) conventionally targets the answer token distribution. We provide the first three-family e

9 Jul 2026

Co-LMLM: Continuous-Query Limited Memory Language Models

Model ReleasesDGX agent

arXiv:2607.07707v1 Announce Type: cross Abstract: Limited memory language models (LMLMs) externalize factual knowledge during pretraining to a knowledge base (KB), rather than memorizing it in their w

Does Bielik Know What It Doesn't Know? Activation Dispersion Separates Entity Familiarity from Factual Reliability Across Model Scale

ResearchDGX agent

arXiv:2607.07670v1 Announce Type: new Abstract: Large language models hallucinate most about entities they have never seen. We ask whether a model's activations betray entity familiarity before a sing

FMMC: Harnessing the Power of Foundation Models for Accurate Material Classification

Model ReleasesDGX agent

arXiv:2603.17390v2 Announce Type: replace Abstract: Material classification has emerged as a critical task in computer vision and graphics, supporting the assignment of accurate material properties to

MedPMC: A Systematic Framework for Scaling High-Fidelity Medical Multimodal Data for Foundation Models

Model ReleasesDGX agent

arXiv:2607.07673v1 Announce Type: new Abstract: Medicine is inherently multimodal, requiring clinicians to synthesize information across diverse data streams. Yet the development of multimodal foundat

Muse Spark 1.1 available in new Meta Model API. Somewhere near Opus-4.8/GPT-5.5 level. 1M context window! Computer-use capabilities sound gr…

Model ReleasesDGX agent

Muse Spark 1.1 available in new Meta Model API. Somewhere near Opus-4.8/GPT-5.5 level. 1M context window! Computer-use capabilities sound great: Write scripts when automation is faster, click when dir

Sol, Terra, and Luna, our GPT‑5.6 family of models, are starting to roll out now in ChatGPT, Codex, and the API.

Model ReleasesDGX agent

OpenAI has announced the rollout of its GPT-5.6 family of models, consisting of Sol, Terra, and Luna variants, across ChatGPT, Codex, and API platforms. These models represent the latest iteration in

The upcoming wave of SpaceXAI Grok updates is insane Grok 4.5: The 1.5T foundation model is being refined almost daily, and its context wind…

Model ReleasesDGX agent

The upcoming wave of SpaceXAI Grok updates is insane Grok 4.5: The 1.5T foundation model is being refined almost daily, and its context window is expected to jump to 1M tokens, possibly as soon as nex

8 Jul 2026

A good voice model should be enjoyable to talk to, and GPT-Live is a great conversationalist with a more natural and defined personality tha…

Model ReleasesDGX agent

GPT-Live is OpenAI's voice model designed to be an engaging conversational partner with natural speech and a distinct personality. The model prioritizes making interactions enjoyable for users through

AirflowAttack: Thermal-Airflow Adversarial Perturbations against Infrared Remote-Sensing Vision-Language Models

Model ReleasesDGX agent

arXiv:2607.06485v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly deployed on infrared (IR) remote sensing imagery in security-critical settings, yet their adversarial r

From Foundation to Application: Improving VLA Models in Practice

Model ReleasesDGX agent

arXiv:2607.06403v1 Announce Type: new Abstract: Despite recent progress of VLA foundation models, the disparity between laboratory conditions and real-world applications continues to impede their prac

GPT-Live makes talking with AI feel like having a real conversation. It’s also our smartest voice model yet. https://openai.com/index/introd…

Model ReleasesDGX agent

OpenAI introduced GPT-Live, a voice-based AI model designed to provide conversational interactions that feel more natural and human-like than previous versions. The model represents an advancement in

Hierarchical Acoustic-Semantic Modeling: Modality Separation and Semantic Coherence for Full-Duplex SLMs

Model ReleasesDGX agent

arXiv:2607.06540v1 Announce Type: new Abstract: Developing seamless, high-performance, native intelligent full-duplex Spoken Language Models (SLMs) remains a critical challenge and long-standing goal

Introducing GPT-Live, a new generation of voice models for natural human-AI interaction. Rolling out in ChatGPT starting today. You’ll want …

Model ReleasesDGX agent

GPT-Live is a new generation of voice model developed by OpenAI designed to enable more natural human-AI voice interactions. The model began rolling out to ChatGPT users starting on the date of this a

OpenAI launches GPT-Live, a new generation of voice models built on a full-duplex architecture, meaning they can listen and speak at the same time (OpenAI)

Model ReleasesDGX agent

OpenAI: OpenAI launches GPT-Live, a new generation of voice models built on a full-duplex architecture, meaning they can listen and speak at the same time — A new generation of voice models for natura

Refiant goes where rivals only promised with a 10 million-token AI model

Model ReleasesDGX agent

Artificial intelligence optimization startup Refiant Inc. today launched Protea, a suite of long-context AI models led by a 10 million-token context window that the company says ranks among the larges

VisCoP: Visual Probing for Video Domain Adaptation of Vision Language Models

Model ReleasesDGX agent

arXiv:2510.13808v2 Announce Type: replace Abstract: Large Vision Language Models (VLMs) excel at general visual reasoning but experience significant performance degradation when deployed in novel doma

7 Jul 2026

Alibaba's Qwen models have made it an AI powerhouse, but the company has struggled to turn their global popularity into a profitable business (New York Times)

Model ReleasesDGX agent

New York Times: Alibaba's Qwen models have made it an AI powerhouse, but the company has struggled to turn their global popularity into a profitable business — The Chinese company's models have won ov

Dashboard2Code: Evaluating Multimodal Models on Reconstructing Interactive Dashboards

Model ReleasesDGX agent

arXiv:2607.04727v1 Announce Type: cross Abstract: Automatic data visualization generation has advanced rapidly with multi-modal large language models, yet existing efforts largely focus on static char

Do Diabetic Foot Ulcer Segmentation Models Generalize? A Cross-Dataset Benchmark of CNN and Transformer Architectures

Model ReleasesDGX agent

arXiv:2607.02555v1 Announce Type: new Abstract: Deep learning models for diabetic foot ulcer (DFU) segmentation routinely report high accuracy, but they are almost always trained and tested on the sam

LILAC: Layer-Wise Independent LoRAs and Cascaded Conditioning for Multi-Concept Customization of Diffusion Models

Model ReleasesDGX agent

arXiv:2607.04801v1 Announce Type: new Abstract: Personalizing text-to-image diffusion models to render several specific subjects in a coherent image remains challenging: the model must preserve each s

Metronome: Bound the Cache, Keep the Beat for Real-Time Interaction Model Serving

Model ReleasesDGX agent

arXiv:2607.02640v1 Announce Type: cross Abstract: Real-time interaction models -- Moshi, MiniCPM-o, Qwen-Omni -- turn serving into a periodic real-time task: on every frame a session ingests streaming

One Framework for All: Cross-Modal Membership Inference for Generative Models

ApplicationsDGX agent

arXiv:2607.04339v1 Announce Type: cross Abstract: Large generative models across text-to-text, text-to-image, and image-to-text modalities have been shown to pose significant privacy risks. One fundam

Predicting Biased Human Decision-Making with Large Language Models in Conversational Settings

Model ReleasesDGX agent

arXiv:2601.11049v2 Announce Type: replace-cross Abstract: We examine whether large language models (LLMs) can predict biased decision-making in conversational settings, and whether their predictions c

RABBiT: Rapidly adaptive BOLD foundation model via brain-tuning for accurate zero-shot and few-shot prediction of speech-elicited responses in the brain

Model ReleasesDGX agent

arXiv:2607.05171v1 Announce Type: new Abstract: Language understanding in the brain is context-dependent, varying across experimental stimuli and individuals, which makes it difficult to build computa

Reconstruction-Anchored Diffusion Model for Text-to-Motion Generation

Model ReleasesDGX agent

arXiv:2601.14788v2 Announce Type: replace Abstract: Diffusion models have seen widespread adoption for text-driven human motion generation and related tasks due to their impressive generative capabili

TACO: TActile World Model as a Self-COrrector forScalable VLA Post-Training

Local AiDGX agent

arXiv:2607.02840v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have shown promising generalization in robotic manipulation, but they still struggle with contact-rich tasks, where

Topology-Driven Transferability Estimation for 3D Medical Vision Foundation Models

Model ReleasesDGX agent

arXiv:2607.04199v1 Announce Type: new Abstract: The growing number of medical vision foundation models highlights the need for effective model selection. However, mainstream selection methods rely on

Variable Bit-width Quantization: Learning Per-Group Precision for 'Bigger-but-Smaller' Language Models

Model ReleasesDGX agent

arXiv:2607.02893v1 Announce Type: cross Abstract: Low-bit quantization shrinks language models but treats precision as a single global hyper-parameter: every weight uses the same bit-width. We introdu

WSA_1: a 3D-Centric World-Spatial-Action Model for Generalizable Robot Control

Model ReleasesDGX agent

arXiv:2607.03941v1 Announce Type: new Abstract: Recent advances in embodied AI have established robot foundation models (RFMs) as the dominant approach for generalist robotic systems to date. By lever

6 Jul 2026

Revisiting ASR Error Correction with Specialized Models

ResearchDGX agent

Language models play a central role in automatic speech recognition (ASR), yet most methods rely on text-only models unaware of ASR error patterns. Recently, large language models (LLMs) have been app

3 Jul 2026

BRIDGE: Predicting Human Task Completion Time From Model Performance

Model ReleasesDGX agent

arXiv:2602.07267v2 Announce Type: replace Abstract: Evaluating the real-world capabilities of AI systems requires grounding benchmark performance in human-interpretable measures of task difficulty. Ex

Locality-Aware Continual Unlearning for Diffusion Models

Model ReleasesDGX agent

arXiv:2512.02657v2 Announce Type: replace-cross Abstract: Real-world deployment of text-to-image diffusion models requires continual concept removal as new privacy, copyright, or safety obligations ar

PhysMani: Physics-principled 3D World Model for Dynamic Object Manipulation

Model ReleasesDGX agent

arXiv:2607.01938v1 Announce Type: cross Abstract: Manipulating fast and dynamically moving targets in unstructured 3D environments remains challenging for embodied AI. Existing visual-language-action

2 Jul 2026

Auditing Forgetting in Limited Memory Language Models

Model ReleasesDGX agent

arXiv:2607.00605v1 Announce Type: cross Abstract: Limited Memory Language Models (LMLMs) externalize factual knowledge to a database to enable deletion-based unlearning without retraining. Existing ev

From Structural Equation Modelling to Double Machine Learning: Robustness Analysis for Survey-Based Research

Model ReleasesDGX agent

arXiv:2607.00512v1 Announce Type: new Abstract: Structural equation modelling (SEM) is widely used in survey-based business and information systems research to assess latent constructs and theory-driv

MoHallBench: A Benchmark for Motion Hallucination in Video Large Language Models

Model ReleasesDGX agent

arXiv:2607.01117v1 Announce Type: new Abstract: Video Large Language Models (VideoLLMs) have shown strong progress in video understanding, yet they still suffer from hallucinations that are inconsiste

Revisiting Autoregressive Models for Generative Image Classification

SafetyDGX agent

arXiv:2603.19122v2 Announce Type: replace Abstract: Class-conditional generative models have emerged as accurate and robust classifiers, with diffusion models demonstrating clear advantages over other

SocialOmni: Benchmarking Audio-Visual Social Interactivity in Omni Models

Model ReleasesDGX agent

arXiv:2603.16859v2 Announce Type: replace Abstract: Omni-modal large language models (OLMs) redefine human-machine interaction by natively integrating audio, vision, and text. However, existing OLM be

Testing Frontier Large Language Models' Physics Literacy in Parallel Physical Worlds

Model ReleasesDGX agent

arXiv:2607.00276v1 Announce Type: cross Abstract: Current large-language-model (LLM) physics benchmarks are usually scored by answer accuracy, which cannot distinguish genuine reasoning from recall of

UniDrive-WM: Unified Understanding, Planning and Generation World Model for Autonomous Driving

Model ReleasesDGX agent

arXiv:2601.04453v4 Announce Type: replace Abstract: World models have become central to autonomous driving, where accurate scene understanding and future prediction are crucial for safe control. Recen

Valdi: Value Diffusion World Models

ResearchDGX agent

arXiv:2607.00917v1 Announce Type: cross Abstract: World models can enable Model Predictive Control (MPC), but this requires dynamics prediction that is both fast enough for online use and expressive e

1 Jul 2026

AdaJEPA: An Adaptive Latent World Model

ResearchDGX agent

arXiv:2606.32026v1 Announce Type: cross Abstract: Latent world models enable planning from high-dimensional observations by predicting future states in a compact latent space. However, these models ar

Beyond Binary Instrument QA: Probing Instrument Grounding in Music Audio-Language Models

Model ReleasesDGX agent

arXiv:2606.31338v1 Announce Type: cross Abstract: Recent music audio-language models achieve high accuracy on instrument question-answering benchmarks, but it remains unclear whether this reflects rob

CoLT: Teaching Multi-Modal Models to Think with Chain of Latent Thoughts

Model ReleasesDGX agent

arXiv:2606.31986v1 Announce Type: new Abstract: Chain-of-thought (CoT) reasoning has enabled multi-modal large language models (MLLMs) to tackle complex visual reasoning tasks by generating explicit i

Really confused by all the excitement I see in my timeline for a nerfed model. Never seen anything like it. So many will end up very disappo…

Model ReleasesDGX agent

Really confused by all the excitement I see in my timeline for a nerfed model. Never seen anything like it. So many will end up very disappointed. Time to rethink how to build around frontier and open

The Bidirectional Process Reward Model

Model ReleasesDGX agent

arXiv:2508.01682v3 Announce Type: replace Abstract: Process Reward Models (PRMs), which assign fine-grained scores to intermediate reasoning steps within a solution trajectory, have emerged as a promi

You really need to benchmark models for your use case. As soon as judgements & decisions stack on top of each other, the differences between…

Model ReleasesDGX agent

You really need to benchmark models for your use case. As soon as judgements & decisions stack on top of each other, the differences between models amplifies, and no standard benchmark will tell you t

30 Jun 2026

AURORA: Asymmetry and Update-Induced Rotation for Robust Hallucination Detection in Large Language Models

Model ReleasesDGX agent

arXiv:2606.29545v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities across a wide range of natural language processing tasks. However, their tendency

BrepLLM: Enabling Large Language Models to Understand Boundary Representations

Model ReleasesDGX agent

arXiv:2512.16413v2 Announce Type: replace Abstract: Current token-sequence-based Large Language Models (LLMs) struggle to directly process 3D Boundary Representation (B-rep) models that contain comple

ExploreVLA: Dense World Modeling and Exploration for End-to-End Autonomous Driving

SafetyDGX agent

arXiv:2604.02714v2 Announce Type: replace Abstract: End-to-end autonomous driving models based on Vision-Language-Action (VLA) architectures have shown promising results by learning driving policies t

Fine-Tuning General-Purpose Large Language Models for Agricultural Applications:A Reproducible Framework and Evaluation Protocol Based on Qwen3-8B

Model ReleasesDGX agent

arXiv:2606.28992v1 Announce Type: cross Abstract: General-purpose large language models (LLMs) have demonstrated strong abilities in opendomain question answering, information extraction, and text gen

Introducing Claude Sonnet 5 on AWS: Anthropic’s most capable Sonnet model

Model ReleasesDGX agent

Today, we’re excited to announce the availability of Anthropic’s most advanced Sonnet model, Claude Sonnet 5, on Amazon Bedrock and Claude Platform on AWS. Claude Sonnet 5 is the first Sonnet model of

Morphing into Hybrid Attention Models

Model ReleasesDGX agent

arXiv:2606.30562v1 Announce Type: new Abstract: Hybrid attention models improve long-context efficiency by retaining only a subset of full-attention layers and replacing the remaining layers with line

PlantExpertVQA: A Visual Question Answering Dataset for Benchmarking Vision-Language Models in Plant Science

Model ReleasesDGX agent

arXiv:2508.17117v3 Announce Type: replace-cross Abstract: Existing plant-disease datasets target classification and detection, leaving vision-language models unable to support interactive, reasoning-b

Qwen-RobotNav Technical Report: A Scalable Navigation Model Designed for an Agentic Navigation System

Model ReleasesDGX agent

arXiv:2606.18112v3 Announce Type: replace-cross Abstract: Agentic navigation systems require a base navigation model whose observation strategy can be externally reconfigured at inference time, becaus

← Previous
1…4748495051…999
Next →