AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,565 results
8 Jun 2026

Narrative violation: according to @Stanford research, local models can answer 71.3% of real-world chat and reasoning queries accurately, up …

ApplicationsDGX agent

Narrative violation: according to @Stanford research, local models can answer 71.3% of real-world chat and reasoning queries accurately, up from 23.2% in 2023. Obviously at a fraction of the cost and

Phonetic Error Analysis of Raw Waveform Acoustic Models

ResearchDGX agent

arXiv:2606.07030v1 Announce Type: cross Abstract: We analyse error patterns of raw waveform acoustic models on TIMIT phone recognition beyond the overall phone error rate (PER). PER is decomposed acro

Robots Need More than VLA and World Models

SafetyDGX agent

arXiv:2606.06556v1 Announce Type: new Abstract: Generalist robot intelligence is often framed as a policy-scaling problem: collect more robot demonstrations, train larger Vision-Language-Action (VLA)

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Should You Use Your Large Language Model to Explore or Exploit?

AgentsDGX agent

arXiv:2502.00225v4 Announce Type: replace-cross Abstract: We evaluate the ability of the current generation of large language models (LLMs) to help a decision-making agent facing an exploration-exploi

Small Language Model Agents Enable Efficient and High-Quality Knowledge Mining

AgentsDGX agent

arXiv:2510.01427v3 Announce Type: replace Abstract: At the core of Deep Research is knowledge mining, the task of extracting structured information from massive unstructured text in response to user i

The Dark Regulome: Disentangling Predictability from Regulation in Genomic Foundation Models

Local AiDGX agent

arXiv:2606.06834v1 Announce Type: new Abstract: High-grade gliomas integrate into neural circuits through functional synapses with neurons, raising the question of which noncoding elements shape synap

VALUEFLOW: Toward Pluralistic and Steerable Value-based Alignment in Large Language Models

SafetyDGX agent

arXiv:2602.03160v2 Announce Type: replace Abstract: Aligning Large Language Models (LLMs) with the diverse spectrum of human values remains a central challenge: preference-based methods often fail to

6 Jun 2026

RREDCoT: Segment-Level Reward Redistribution for Reasoning Models

SafetyDGX agent

arXiv:2606.06475v1 Announce Type: cross Abstract: Recent advancements in reasoning language models have been driven by Reinforcement Learning (RL) fine-tuning. Most often, these rely on the Group Rela

5 Jun 2026

A Survey on Diffusion Language Models

ResearchDGX agent

arXiv:2508.10875v3 Announce Type: replace Abstract: Diffusion Language Models (DLMs) are rapidly emerging as a powerful and promising alternative to the dominant autoregressive (AR) paradigm. By gener

Biomazon: A Multimodal Dataset for 3D Forest Structure and Biomass Modeling in the Amazon Basin

Model ReleasesDGX agent

arXiv:2606.05368v1 Announce Type: new Abstract: Accurate, spatially explicit characterization of tropical forest structure is essential for carbon accounting and ecosystem monitoring, yet most ML pipe

DRIFT: A Residual Flow Adapter for Decoding Continuous Outputs in Vision-Language Models

ResearchDGX agent

arXiv:2606.05758v1 Announce Type: new Abstract: Many modern vision-language models (VLMs) build on autoregressive decoding of discrete tokens. While text-based output interfaces enable scalable pretra

Let It Be Simple: One-Step Action Generation for Vision-Language-Action Models

SafetyDGX agent

arXiv:2606.05737v1 Announce Type: new Abstract: Diffusion-based vision-language-action (VLA) models often inherit the image-generation view: actions are generated by iterative denoising. We argue that

PLAN-S: Bridging Planning with Latent Style Dynamics for Autonomous Driving World Models

AgentsDGX agent

arXiv:2606.06014v1 Announce Type: cross Abstract: Latent world models (LWMs) have strengthened end-to-end autonomous driving by forecasting compact scene dynamics for downstream planning. However, exi

Self-Augmenting Retrieval for Diffusion Language Models

TutorialsDGX agent

arXiv:2606.06474v1 Announce Type: new Abstract: Discrete diffusion language models generate text by iteratively denoising an entire response in parallel. At each step, they predict tentative tokens fo

4 Jun 2026

ADAPTOOD: Uncertainty-Aware Fine-Tuning for Out-of-Distribution ECG Time Series Models

TutorialsDGX agent

arXiv:2606.04164v1 Announce Type: cross Abstract: Data samples used for training often differ from those encountered during fine-tuning and deployment, and while ML models show promise, their performa

Beyond Pixel Histories: World Models with Persistent 3D State

ResearchDGX agent

arXiv:2603.03482v2 Announce Type: replace-cross Abstract: Interactive world models continually generate video by responding to a user's actions, enabling open-ended generation capabilities. However, e

Can VLMs Predict Future States? Bootstrapping World Models from Inverse Dynamics

TutorialsDGX agent

arXiv:2506.06006v3 Announce Type: replace-cross Abstract: Can unified vision-language models (VLMs) perform forward dynamics prediction (FDP), i.e., predicting the future state (in image form) given t

Critical Hugging Face Transformers flaw ran attacker code on a routine model load

IndustryDGX agent

Pluto Security Inc. today disclosed a critical remote code execution vulnerability in Hugging Face Inc.’s Transformers library that allowed attacker-controlled artificial intelligence models to run ar

Global Sketch-Based Watermarking for Diffusion Language Models

SafetyDGX agent

arXiv:2606.04486v1 Announce Type: cross Abstract: Watermarking methods for language models have been studied extensively in the autoregressive setting, where tokens are generated sequentially. These w

https://ollama.com/library/nemotron-3-ultra

Model ReleasesDGX agent

Nemotron-3-Ultra is a large language model available through Ollama's model library, likely representing an advanced iteration in NVIDIA's Nemotron model series optimized for performance and capabilit

On Forgetting and Stability of Score-based Generative models

ResearchDGX agent

arXiv:2601.21868v2 Announce Type: replace-cross Abstract: Understanding the stability and long-time behavior of generative models is a fundamental problem in modern machine learning. This paper provid

Supportive Token Revealing for Fast Diffusion Language Model Decoding

ResearchDGX agent

arXiv:2606.04236v1 Announce Type: cross Abstract: Discrete diffusion language models can generate text efficiently by updating multiple masked positions in parallel, but this parallelism introduces a

ViewMask-1-to-3: Multi-View Consistent Image Generation via Multimodal Discrete Diffusion Models

ResearchDGX agent

arXiv:2512.14099v3 Announce Type: replace Abstract: Motivated by discrete diffusion's success in language-vision modeling, we explore its potential for multi-view generation, a task dominated by conti

we're taking an early bet on open models specifically, because they're SO much cheaper one point of reference: an app outputting 10M tokens/…

AgentsDGX agent

we're taking an early bet on open models specifically, because they're SO much cheaper one point of reference: an app outputting 10M tokens/day costs roughly 250/day on Opus 4.8 versus ~24/day for Min

Worker Utility as Hysteresis: A Preisach Model of Transaction Acceptance in Gig Labour Markets

ResearchDGX agent

arXiv:2606.04916v1 Announce Type: new Abstract: Worker utility is not observed -- only its consequence is. Each gig transaction produces a single bit: accepted or rejected. We argue this structure poi

3 Jun 2026

A Sparse Bayesian Learning Algorithm for Estimation of Interaction Kernels in Motsch-Tadmor Model

ResearchDGX agent

arXiv:2505.07068v2 Announce Type: replace-cross Abstract: In this paper, we investigate the data-driven identification of asymmetric interaction kernels in the Motsch-Tadmor model based on observed tr

An Improved Method for Personalizing Diffusion Models

ResearchDGX agent

arXiv:2407.05312v2 Announce Type: replace Abstract: Diffusion models have demonstrated impressive image generation capabilities. Personalized approaches, such as textual inversion and Dreambooth, enha

Attention, May I Have Your Decision? Localizing Generative Choices in Diffusion Models

Local AiDGX agent

arXiv:2604.06052v2 Announce Type: replace Abstract: Text-to-image diffusion models exhibit remarkable generative capabilities, yet their internal operations remain opaque, particularly when handling p

Beyond Semantics: Modeling Factual and Affective Perceptual Experiences from Vision-Language Data

ResearchDGX agent

arXiv:2606.03345v1 Announce Type: cross Abstract: We present P-Topics (Perception Topics) modeling, a novel problem for understanding how images are perceived affectively and across cultures. The goal

BYORn: Bootstrap Your Own Responses to Defend Large Vision-Language Models Against Backdoor Attacks

ResearchDGX agent

arXiv:2606.02947v1 Announce Type: cross Abstract: Supervised fine-tuning is the predominant approach for adapting autoregressive vision-language models to downstream tasks. Recent work has shown that

CoralBay: A Self-Supervised CT Foundation Model

Model ReleasesDGX agent

arXiv:2606.03888v1 Announce Type: new Abstract: Self-supervised learning has enabled large-scale pre-training on 2D natural images, producing general-purpose visual representations that transfer effec

DMT-CBT: Longitudinal Therapeutic State Modeling for CBT Counseling

Local AiDGX agent

arXiv:2606.03132v1 Announce Type: new Abstract: Large language models (LLMs) have shown growing potential for Cognitive Behavioral Therapy (CBT) counseling. However, most existing approaches still for

Do Explanations Increase the Risk of Decision Logic Leakage? Explanation-Guided Stealing of Graph Models

SafetyDGX agent

arXiv:2506.03087v2 Announce Type: replace-cross Abstract: Graph Neural Networks (GNNs) have become essential tools for analyzing graph-structured data in domains such as drug discovery and financial a

FlowGuard: Flow Matching for Identity-Independent Detection of Data-Free Model Stealing Attacks on Energy System Intrusion Detection Systems

ResearchDGX agent

arXiv:2606.03430v1 Announce Type: cross Abstract: Artificial Intelligence (AI)-based Intrusion Detection Systems (IDS) deployed in energy infrastructure are vulnerable to model theft attacks, which al

Generalizing Graph Foundation Models via Hyperbolic Retrieval-Augmented Generation

Local AiDGX agent

arXiv:2606.03307v1 Announce Type: cross Abstract: Graph foundation models (GFMs) emerged as a dominant paradigm in graph representation learning by leveraging large-scale pre-training for cross-domain

GeoSem-WAM: Geometry- and Semantic-Aware World Action Models

ResearchDGX agent

arXiv:2606.03188v1 Announce Type: new Abstract: Recent World Action Models (WAMs) have demonstrated impressive capabilities in embodied decision-making. However, whether their effectiveness stems from

GuidedBridge: Training-freely Improving Bridge Models with Prior Guidance

ResearchDGX agent

arXiv:2606.03119v1 Announce Type: cross Abstract: Guidance methods, such as classifier-free guidance (CFG) and auto-guidance (AG), have advanced noise-to-data generation in diffusion models. Recently,

KBQA-R1: Reinforcing Large Language Models for Knowledge Base Question Answering

SafetyDGX agent

arXiv:2512.10999v3 Announce Type: replace Abstract: Knowledge Base Question Answering (KBQA) challenges models to bridge the gap between natural language and strict knowledge graph schemas by generati

Language Models Compare Quantities Using Number-specific and Unit-specific Heuristics

ResearchDGX agent

arXiv:2606.03982v1 Announce Type: new Abstract: Quantities with measurement units, such as 110 cm and 1.2 m, require language models (LMs) to combine a numeral with a symbolic unit scale. Here, we stu

Language Models Need Sleep: Learning to Self-Modify and Consolidate Memories

SafetyDGX agent

arXiv:2606.03979v1 Announce Type: cross Abstract: The past few decades have witnessed significant advances in the design of machine learning algorithms, from early studies on task-specific shallow mod

Neural Attention Search Linear: Towards Adaptive Token-Level Hybrid Attention Models

ResearchDGX agent

arXiv:2602.03681v2 Announce Type: replace Abstract: The quadratic computational complexity of softmax transformers has become a bottleneck in long-context scenarios. In contrast, linear attention mode

NewtPhys: Do Foundation Models Understand Newtonian Physics?

ApplicationsDGX agent

arXiv:2606.03986v1 Announce Type: new Abstract: Previous work has evaluated physics reasoning in foundation models using synthetic or semi-synthetic scenes and visual question-answering tasks. However

Optimal Design and Analytical Modeling of a Soft Fin-Ray Effect Gripper Finger Using the Finite Rigid Elements Method

ResearchDGX agent

arXiv:2606.03798v1 Announce Type: new Abstract: Fin Ray-inspired soft grippers offer a promising solution for gently handling delicate, irregular objects, especially in agriculture. The objective of t

Partially Observable Adversarial Patch Attacks on Vision-Language-Action Models in Robotics

Local AiDGX agent

arXiv:2606.03556v1 Announce Type: new Abstract: Vision-language-action (VLA) models are gaining attention in robotics, yet their robustness to adversarial attacks remains largely unexplored. Existing

q0: Primitives for Hyper-Epoch Pretraining

Model ReleasesDGX agent

arXiv:2606.03938v1 Announce Type: cross Abstract: Multi-epoch training is becoming the standard now that compute is growing faster than the supply of high-quality text. But pretraining a single model

SAMatcher: Co-Visibility Modeling with Segment Anything for Robust Feature Matching

Local AiDGX agent

arXiv:2606.03406v1 Announce Type: new Abstract: Reliable correspondence estimation is a fundamental problem in image processing, underpinning applications such as Structure from Motion, visual localiz

See, Infer, Intervene: Proactive World Modeling for Goal-Oriented Social Intelligence

Model ReleasesDGX agent

arXiv:2606.03371v1 Announce Type: new Abstract: Multimodal retail agents should not only recognize what a customer is doing, but also decide whether and how to assist before an explicit request is mad

Spatial Transcriptomics-Guided Alignment Enhances Molecular Profiling in Pathology Foundation Model

Model ReleasesDGX agent

arXiv:2606.03644v1 Announce Type: new Abstract: Comprehensive molecular profiling is essential for modern precision oncology but remains hindered by prohibitive costs, specimen exhaustion, and protrac

TGV-KV: Text-Grounded KV Eviction for Vision-Language Models

SafetyDGX agent

arXiv:2606.03075v1 Announce Type: new Abstract: Vision-Language Models (VLMs) inherit the auto-regressive generation paradigm and cache the keys and values (KV) of all previous tokens to accelerate in

Validation-Gated Multi-Agent Governance for Online Adaptation of Thermal-Hydraulic Surrogate Models under Operating-Regime Shift

SafetyDGX agent

arXiv:2606.03321v1 Announce Type: new Abstract: Artificial-intelligence surrogates can support second-by-second thermal-hydraulic forecasting, but models selected and frozen offline may become conditi

2 Jun 2026

A Theoretical Framework for Statistical Evaluability of Generative Models

ResearchDGX agent

arXiv:2604.05324v2 Announce Type: replace Abstract: Statistical evaluation aims to estimate the generalization performance of a model using held-out i.i.d. test data sampled from the ground-truth dist

Annotation-Informed Block-Sparse Bayesian Modeling for cis-Expression Prediction

Local AiDGX agent

arXiv:2606.00483v1 Announce Type: cross Abstract: Genotype-based cis-expression prediction depends on accurately modeling local regulatory architecture. We present block-sparse Bayesian sparse linear

BlockGen: Flexible Blockwise Sequence Modeling with Hybrid Samplers

ResearchDGX agent

arXiv:2606.02241v1 Announce Type: new Abstract: Is the uniform-state diffusion framework a more powerful paradigm for discrete diffusion? Recent studies indicate that this may be the case. In combinat

Characterization of Multi-Model Agentic AI Systems on General Tasks via Trace-Driven Simulation

Model ReleasesDGX agent

arXiv:2606.01725v1 Announce Type: new Abstract: Agentic AI completes tasks through iterative planning, tool use, and reasoning based on observed outcomes. Despite its popularity, its system-level beha

Coupling Language Models with Physics-based Simulation for Synthesis of Inorganic Materials

ApplicationsDGX agent

arXiv:2606.00315v1 Announce Type: new Abstract: Modern generative machine learning (ML) models can propose novel inorganic crystalline materials with targeted properties; however, synthesis planning o

Custom model training is bringing enterprise AI from experimentation to production

ApplicationsDGX agent

As artificial intelligence moves from proof of concept into enterprise production, custom model training on governed data is emerging as the critical unlock for organizations that need domain-specific

Diffusion Models for Hyperspectral Image Analysis: A Comprehensive Review

ResearchDGX agent

arXiv:2505.11158v4 Announce Type: replace-cross Abstract: Hyperspectral image (HSI) analysis plays a critical role in remote sensing, agriculture, and environmental monitoring. However, traditional me

Early Prediction of Liver Cirrhosis Up to Two Years in Advance: A Machine Learning Study Benchmarking Against the FIB-4 and APRI Scores

Model ReleasesDGX agent

arXiv:2601.00175v2 Announce Type: replace Abstract: Objective: Develop and evaluate machine learning (ML) models for predicting incident liver cirrhosis (LC) one and two years prior to diagnosis using

Entropy Minimization without Model Collapse: Mitigating Prediction Bias in Medical Imaging

SafetyDGX agent

arXiv:2606.02339v1 Announce Type: cross Abstract: Entropy minimization (EM) is the dominant objective for test-time adaptation, yet its failure mode, model collapse, remains poorly understood. In this

Extreme Low-Bit Inference in Reasoning Models: Failure Modes and Targeted Recovery

ResearchDGX agent

arXiv:2606.02011v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) rely on long reasoning traces, making inference expensive. While low-bit quantization reduces per-token decoding cost, we

← Previous
1…155156157158159…1010
Next →