AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,522 results
9 Jul 2026

HIVE: Understanding Post-Hallucination Reasoning in Vision Language Models

ResearchDGX agent

arXiv:2607.07507v1 Announce Type: cross Abstract: Hallucinations in vision language models (VLMs) are commonly treated as semantic errors, yet they often arise from partial or ambiguous visual evidenc

Object Search in Partially-Known Environments via LLM-informed Model-based Planning and Prompt Selection

ResearchDGX agent

arXiv:2603.23800v2 Announce Type: replace-cross Abstract: We present a novel LLM-informed model-based planning framework, and a novel prompt selection method, for object search in partially-known envi

Try Grok 4.5 for free, an all new Opus-class model that is fast and low cost. Great for real-world coding and engineering tasks.

ApplicationsDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Grok 4.5 is a new Opus-class AI model released by xAI that offers improved speed and cost efficiency compared to previous versions. The model is positioned as particularly well-suited for real-world c

8 Jul 2026

BlueMagpie-TTS: A Token-Efficient Tokenizer, Language Model, and TTS for Taiwanese-Accent Code-Switching Speech

Model ReleasesDGX agent

arXiv:2607.06054v1 Announce Type: cross Abstract: Off-the-shelf TTS systems are poorly adapted to Taiwanese Mandarin. Their accent defaults to other Mandarin variants, their tokenizers over-segment co

Decision Protocols in Multi-Agent Large Language Model Conversations

AgentsDGX agent

arXiv:2607.05477v1 Announce Type: cross Abstract: Improving the task performance of Large Language Models (LLMs) is essential, yet scaling these models faces significant challenges such as diminishing

Factory (@FactoryAI) CEO Matan Grinberg (@matanSF) says 'In the next 12 months, 90% of tokens will be going to open models.' Enterprises are…

ApplicationsDGX agent

Factory (@FactoryAI) CEO Matan Grinberg (@matanSF) says 'In the next 12 months, 90% of tokens will be going to open models.' Enterprises are already shifting rapidly.. 'At the beginning of the year, <

Lift3D-VLA: Lifting VLA Models to 3D Geometry and Dynamics-Aware Manipulation

ApplicationsDGX agent

arXiv:2607.06564v1 Announce Type: cross Abstract: Recently, Vision-Language-Action (VLA) models have demonstrated strong generalization across diverse tasks. However, effective robotic manipulation in

Love partnering with baseten to make sure everyone can use open weight models in deep agents

AgentsDGX agent

This post discusses Baseten's partnership efforts to democratize access to open-weight models for use in AI agents, making advanced model capabilities available to a broader audience. The initiative a

RMISC: A Large-scale Real-world Multivariate Corpus for Time Series Foundation Models

ApplicationsDGX agent

arXiv:2607.06504v1 Announce Type: new Abstract: Recent years have witnessed the emergence of multivariate modeling using time series foundation models (TSFMs), which achieve advanced zero-shot general

Scene Graph Thinking: Reinforcing Structured Visual Reasoning for Multimodal Large Language Models

ResearchDGX agent

arXiv:2607.05716v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have demonstrated strong perception and reasoning capabilities. However, most existing models focus on isolated

The Large Cancer Assistant (LCA): A Model-Agnostic Orchestration Framework for Scalable Clinical Decision Support in Oncology

SafetyDGX agent

arXiv:2607.06531v1 Announce Type: new Abstract: - Objective: Multimodal deep learning models in oncology are currently limited by monolithic designs that rigidly couple data ingestion, clinical routin

this should be a huge update in your model of software engineering: rewrites can be good, cheap and fast of course most apps are not as test…

TutorialsDGX agent

this should be a huge update in your model of software engineering: rewrites can be good, cheap and fast of course most apps are not as testable and verifiable as Bun, but the models will continue to

We will continue to make refinements to the Grok Build harness and the 1.5T foundation model almost every day in response to user requests. …

IndustryDGX agent

We will continue to make refinements to the Grok Build harness and the 1.5T foundation model almost every day in response to user requests. The 2T model will finish training this month and be availabl

When Does Tool Use Increase the Expressive Power of Finite-Precision Recurrent Models?

AgentsDGX agent

arXiv:2607.06155v1 Announce Type: cross Abstract: Modern sequence models are increasingly deployed as agents that interleave token generation with calls to external tools. We give an exact, architectu

7 Jul 2026

A Preliminary Study on Explaining Risk of Code Changes using LLM-Based Prediction Models

ResearchDGX agent

arXiv:2607.02782v1 Announce Type: cross Abstract: Predictions by machine learning (ML) and artificial intelligence (AI) models are often received skeptically unless they are paired with intelligible e

ClinOCR-Bench: A Comprehensive Clinical Scanned Document Dataset for Optical Character Recognition Model Evaluation

Model ReleasesDGX agent

arXiv:2607.03650v1 Announce Type: cross Abstract: Extracting textual information from scanned medical documents, such as external laboratory reports and manually filled forms, has been a major challen

CONFLUX: A Latent Diusion Model for 3D Chest-CT Synthesis with RL Post-Training

SafetyDGX agent

arXiv:2607.02998v1 Announce Type: cross Abstract: Controllable generative models of 3D medical images can synthesize volumes with specified clinical attributes, but this demands samples that are simul

Decomposed Prompting Does Not Fix Knowledge Gaps, But Helps Models Say 'I Don't Know'

SafetyDGX agent

arXiv:2602.04853v2 Announce Type: replace Abstract: Large language models often struggle to recognize their knowledge limits in closed-book question answering, leading to confident hallucinations. Whi

Does It Fail to See or Fail to Know? Attributing Errors in Vision-Language Models

ResearchDGX agent

arXiv:2607.04683v1 Announce Type: cross Abstract: Vision-language models (VLMs) perform well on visual question answering with high-quality images but struggle when questions require knowledge beyond

dOPSD: On-Policy Self-Distillation for Diffusion Language Models

SafetyDGX agent

arXiv:2607.04428v1 Announce Type: cross Abstract: Diffusion large language models (dLLMs) generate text by iteratively denoising a masked sequence, offering a parallel alternative to autoregressive mo

Evaluating Large Language Models for Antisemitic Incident Classification

Model ReleasesDGX agent

arXiv:2607.04890v1 Announce Type: new Abstract: Addressing hate and violence in society requires timely detection of hateful events from public reporting, but automated identification of hateful event

FOI-O: An NZ-first ontology and verification methods package for Freedom of Information process modelling

AgentsDGX agent

arXiv:2607.02947v1 Announce Type: cross Abstract: Public official-information request records contain process signals. They can support research, workflow review, and human-supervised agent help. Yet

Geometric Causal Models

ApplicationsDGX agent

arXiv:2607.05153v1 Announce Type: cross Abstract: Scientists often seek to draw causal inferences from structured data that is not independently and identically distributed, such as spatial data, netw

GeoSAM-Lite: A Lightweight Foundation Model for Onboard Remote Sensing Segmentation

ResearchDGX agent

arXiv:2607.03760v1 Announce Type: new Abstract: The deployment of large-scale foundation models like Segment Anything Model (SAM) on resource-constrained Earth observation platforms is hindered by pro

In-span learning: adapting reduced-order models using their own predictions

ResearchDGX agent

arXiv:2607.02937v1 Announce Type: new Abstract: Reduced-order models compress high-dimensional dynamics into low-dimensional representations that can be evaluated rapidly, but they lose accuracy when

Joint Velocity Slope Diffusion Prior for Structurally Constrained Velocity Model Building

Local AiDGX agent

arXiv:2607.04982v1 Announce Type: cross Abstract: High-resolution velocity models are crucial for reservoir characterization and subsurface delineation. However, the band limited nature of our surface

Kairos: A Regret-Aware Native World-Action Model Stack for Physical AI

Local AiDGX agent

arXiv:2606.16533v3 Announce Type: replace Abstract: We introduce extbf{Kairos}, a regret-aware native world-action model stack for Physical AI. Kairos is motivated by the view that a physical world mo

Lacuna Inc. at SemEval-2026 Task 4: Structurally Gated State-Space Models for Disentangling Narrative Similarity

SafetyDGX agent

arXiv:2607.03482v1 Announce Type: new Abstract: In this paper, we present the Invariant-Variant Disentangled State-Space Model (IVD-SSM), our submission to SemEval-2026 Task 4 on Narrative Story Simil

Language models guide symbolic equation discovery by controlling search

TutorialsDGX agent

arXiv:2607.04156v1 Announce Type: new Abstract: Scientific equation discovery must combine broad domain priors with strict numerical testing. Symbolic regression supplies numerical grounding but faces

Large Language Models Develop Novel Social Biases Through Adaptive Exploration

SafetyDGX agent

arXiv:2511.06148v4 Announce Type: replace-cross Abstract: As large language models (LLMs) are adopted into frameworks that grant them the capacity to make real decisions, it is increasingly important

Meta’s new Muse Image model can pull other Instagram users into AI photos

IndustryDGX agent

Meta is launching the first AI image generation model made by its Superintelligence Labs division. The Muse Image model now powers the image-making tools across the Meta AI app, Instagram, and WhatsAp

Microsoft is reportedly ditching OpenAI’s and Anthropic’s AI models in favor of its own to cut costs

IndustryDGX agent

Microsoft Corp. is reportedly transitioning away from using OpenAI Group PBC’s and Anthropic PBC’s most advanced artificial intelligence models in favor of its own — increasingly leaning on the new Mi

MoP-JEPA: Hard-Assigned Predictor Mixtures for Stochastic JEPA World Models

ResearchDGX agent

arXiv:2607.05238v1 Announce Type: new Abstract: JEPA world models predict the next latent state with a single deterministic predictor trained by latent regression. We show that this fails structurally

One Prompt, Many Sounds: Modeling Listener Variability in LLM-Based Equalization

Model ReleasesDGX agent

arXiv:2601.09448v3 Announce Type: replace-cross Abstract: Conventional audio equalization is a static process that requires manual and cumbersome adjustments to adapt to changing listening contexts (e

OSF: On Pre-training and Scaling of Sleep Foundation Models

Model ReleasesDGX agent

arXiv:2603.00190v2 Announce Type: replace-cross Abstract: Polysomnography (PSG) provides the gold standard for sleep assessment but suffers from substantial heterogeneity across recording devices and

Out-of-Distribution Detection in Molecular Complexes via Diffusion Models for Irregular Graphs

ResearchDGX agent

arXiv:2512.18454v3 Announce Type: replace Abstract: Predictive machine learning models generally excel on in-distribution data, but their performance degrades on out-of-distribution (OOD) inputs. Reli

Pathways of Visual Information Flow in Vision-Language Models

ResearchDGX agent

arXiv:2607.03358v1 Announce Type: cross Abstract: We study how visual information is routed in vision-language models (VLMs). Using causal patching on controlled synthetic and natural datasets, we fin

PixelPilot: Scalable Vision-Language-Action Models for End-to-End Autonomous Driving

SafetyDGX agent

arXiv:2607.04637v1 Announce Type: new Abstract: Vision-Language-Action Models (VLAs), which leverage the advanced reasoning capabilities of Vision-Language Models (VLMs), show promising generalization

Pooling-Based Context Modeling for Convolution-Free Deep Image Prior

Model ReleasesDGX agent

arXiv:2607.02952v1 Announce Type: cross Abstract: Convolutional Neural Networks (CNNs) achieve strong denoising performance by exploiting spatial context from neighboring pixels. Deep Image Prior (DIP

Responsibility Distribution Estimation in Ego-View Accident Videos with Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2607.03591v1 Announce Type: cross Abstract: Recent studies on multimodal traffic accident understanding have mainly relied on infrastructure-camera footage, satellite imagery, or structured cras

See the Emotion: A Facial Emoji Proxy Modeling for EEG Emotion Recognition

ResearchDGX agent

arXiv:2607.02912v1 Announce Type: new Abstract: Despite the high accuracy of EEG-based emotion recognition, existing models remain opaque 'black boxes', lacking semantic grounding between abstract neu

SOV-CAD: Stepwise Orthographic Views Guided CAD Modeling Sequence Reconstruction

SafetyDGX agent

arXiv:2607.04119v1 Announce Type: cross Abstract: Reconstructing Computer-Aided Design (CAD) modeling sequences from images is crucial for preserving design intent and supporting parametric editing. H

Specific Domain Ontology Construction Using Large Language Models

Model ReleasesDGX agent

arXiv:2606.20691v1 Announce Type: cross Abstract: Ontologies are useful structures to organize and maintain information that can be understood both by humans and systems. However, since their manual c

StageCraft: Execution Aware Mitigation of Distractor and Obstruction Failures in VLA Models

SafetyDGX agent

arXiv:2603.20659v2 Announce Type: replace Abstract: Large scale pre-training on text and image data along with diverse robot demonstrations has helped Vision Language Action models (VLAs) to generaliz

Verifier-free Test-Time Sampling for Vision-Language-Action Models

TutorialsDGX agent

arXiv:2510.05681v2 Announce Type: replace-cross Abstract: Vision-Language-Action models (VLAs) have demonstrated remarkable performance in robot control. However, they remain fundamentally limited in

VISOR: Visual Input-based Steering for Output Redirection in Vision-Language Models

SafetyDGX agent

arXiv:2508.08521v2 Announce Type: replace-cross Abstract: Vision Language Models (VLMs) are increasingly being used in a broad range of applications, bringing their security and behavioral control to

Weak-to-Strong Generalization via Direct On-Policy Distillation

SafetyDGX agent

arXiv:2607.05394v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) is a powerful recipe for improving language-model reasoning, but it is expensive to repeat on ev

Worldscape-MoE: A Unified Mixture-of-Experts World Model for Scalable Heterogeneous Action Control

ResearchDGX agent

arXiv:2607.03964v1 Announce Type: cross Abstract: World models are rapidly becoming a core infrastructure for embodied intelligence and interactive agents: they provide controllable simulators in whic

6 Jul 2026

'Alex really hit the nail on the head, by sending data to a company that has extremely smart models, you are really giving up your business'…

ToolsDGX agent

'Alex really hit the nail on the head, by sending data to a company that has extremely smart models, you are really giving up your business's recipe for them to copy.' Our CEO @vipulved when asked abo

'Most enterprise workflows don't require frontier AI models. Outside of coding, tasks like summarization, document generation, and briefing …

ApplicationsDGX agent

'Most enterprise workflows don't require frontier AI models. Outside of coding, tasks like summarization, document generation, and briefing can often be handled effectively by lower-cost open-source m

With inference scale and research scale to drive efficiency, and ever-improving frontier models as brains, this would be a way for the Labs …

ApplicationsDGX agent

With inference scale and research scale to drive efficiency, and ever-improving frontier models as brains, this would be a way for the Labs to undercut even open weights models. Some companies will st

3 Jul 2026

A rubric-based controlled comparison of frontier language models on expert-authored clinical reasoning tasks

Model ReleasesDGX agent

arXiv:2607.02175v1 Announce Type: new Abstract: Multiple-choice medical benchmarks are increasingly saturated, and recent rubric-based evaluations such as HealthBench have shown that open-ended clinic

DRIFTLENS: Measuring Memory-Induced Reasoning Drift in Personalized Language Models

ResearchDGX agent

arXiv:2607.02374v1 Announce Type: new Abstract: Personalization changes what a model says to a user; we show that it can also change the reasoning trajectory used to justify the response. Modern LLMs

LearNAT: Learning NL2SQL with AST-guided Task Decomposition for Large Language Models

Model ReleasesDGX agent

arXiv:2504.02327v2 Announce Type: replace Abstract: Natural Language to SQL (NL2SQL) aims to translate natural language queries into executable SQL statements, offering non-expert users intuitive acce

Navigating the Alignment-Calibration Trade-off: A Pareto-Superior Frontier via Model Merging

SafetyDGX agent

arXiv:2510.17426v3 Announce Type: replace-cross Abstract: The 'alignment tax' of post-training is typically framed as a drop in task accuracy. We show it also involves a severe loss of calibration, ma

Regularized Variational and Spectral Log-Density-Ratio Estimation in the Gaussian Location Model

ResearchDGX agent

arXiv:2607.01895v1 Announce Type: new Abstract: We study ridge-regularized log-density-ratio estimation in the Gaussian location model with a common covariance matrix. By affine invariance, the model

ThreadWeaver: Adaptive Threading for Efficient Parallel Reasoning in Language Models

ResearchDGX agent

arXiv:2512.07843v2 Announce Type: replace-cross Abstract: Scaling inference-time computation has enabled Large Language Models (LLMs) to achieve strong reasoning performance, but their inherently sequ

VLAFlow: A Unified Training Framework for Vision-Language-Action Models via Co-training and Future Latent Alignment

SafetyDGX agent

arXiv:2607.01586v1 Announce Type: cross Abstract: Vision-language-action models (VLAs) have recently advanced robotic manipulation, yet the effects of different robot-data pre-training paradigms remai

2 Jul 2026

A Text-Steerable Instrument for Sketching Procedural Soundscapes via Language Models

Model ReleasesDGX agent

arXiv:2607.00309v1 Announce Type: cross Abstract: We present a real-time musical interface that converts natural-language scene descriptions into evolving procedural soundscapes. A performer types a p

A Unified Benchmark for RCM-Constrained Visual Servoing: Modeling-Controller Interaction and Robustness Analysis in Laparoscopic Robots

Model ReleasesDGX agent

arXiv:2607.00030v1 Announce Type: new Abstract: In robot-assisted laparoscopic minimally invasive surgery (MIS), accurate enforcement of the remote center of motion (RCM) constraint is critical for sa

← Previous
1…127128129130131…1009
Next →