AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,537 results
25 May 2026

Beyond VLM-Based Rewards: Diffusion-Native Latent Reward Modeling

SafetyDGX agent

arXiv:2602.11146v2 Announce Type: replace-cross Abstract: Preference optimization for diffusion and flow-matching models relies on reward functions that are both discriminatively robust and computatio

Differences in Typological Alignment in Language Models' Treatment of Differential Argument Marking

SafetyDGX agent

arXiv:2602.17653v2 Announce Type: replace Abstract: Recent work has shown that language models (LMs) trained on synthetic corpora can exhibit typological preferences that resemble cross-linguistic reg

Dithering Defense: Adversarial Robustness of Vision Foundation Models via Multi-Level Floyd-Steinberg Dithering

ResearchDGX agent

arXiv:2605.23065v1 Announce Type: cross Abstract: Vision foundation models are widely used as frozen backbones across many downstream tasks, making them a single point of failure under adversarial att

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

GEM-4D: Geometry-Enhanced Video World Models for Robot Manipulation

ApplicationsDGX agent

arXiv:2605.22882v1 Announce Type: new Abstract: Video world models can generate realistic futures from a single instruction, but they often fail to preserve consistent point-level motion over time. As

How Human-Like Are Large Language Models? A Register-Aware Linguistic Evaluation Framework

ApplicationsDGX agent

arXiv:2605.23651v1 Announce Type: new Abstract: While factual correctness and task-performance have been in focus of Large Language Model (LLM) research for a long time, the fundamental question of ho

Learnability-Informed Fine-Tuning of Diffusion Language Models

TutorialsDGX agent

arXiv:2605.22939v1 Announce Type: new Abstract: We aim to improve the reasoning capabilities of diffusion language models (DLMs). While SFT is a popular post-training recipe for autoregressive models,

VAMP-Diff: VampPrior Latent Diffusion for Photoplethysmography Modeling

TutorialsDGX agent

arXiv:2605.22851v1 Announce Type: cross Abstract: Photoplethysmography (PPG) has become a ubiquitous physiological signal; however, current generative models still struggle to preserve realistic wavef

What Does the Server See? Understanding Privacy Leakage from Large Language Models in Split Inference

ResearchDGX agent

arXiv:2605.23158v1 Announce Type: cross Abstract: The deployment of large language models (LLMs) on resource-constrained devices remains challenging, spurring interest in split inference, where models

23 May 2026

A Diffusive Classification Loss for Learning Energy-based Generative Models

ResearchDGX agent

arXiv:2601.21025v3 Announce Type: replace-cross Abstract: Score-based generative models have recently achieved remarkable success. While they are usually parameterized by the score, an alternative way

Bringing Stability to Diffusion: Decomposing and Reducing Variance of Training Masked Diffusion Models

ResearchDGX agent

arXiv:2511.18159v2 Announce Type: replace Abstract: Masked diffusion models (MDMs) are a promising alternative to autoregressive models (ARMs), but they suffer from inherently much higher training var

CellFluxRL: Biologically-Constrained Virtual Cell Modeling via Reinforcement Learning

ResearchDGX agent

arXiv:2603.21743v4 Announce Type: replace Abstract: Building virtual cells with generative models to simulate cellular behavior in silico is emerging as a promising paradigm for accelerating drug disc

DecepChain: Inducing Deceptive Reasoning in Large Language Models

SafetyDGX agent

arXiv:2510.00319v2 Announce Type: replace Abstract: Large Language Models (LLMs) have been demonstrating strong reasoning capability with their chain-of-thoughts (CoT), which are routinely used by hum

MDM-Prime-v2: Binary Encoding and Index Shuffling Enable Scaling of Diffusion Language Models

TutorialsDGX agent

arXiv:2603.16077v3 Announce Type: replace Abstract: Masked diffusion models (MDM) exhibit superior generalization when learned using a Partial masking scheme (Prime). This approach converts tokens int

Soft Bayesian Context Tree Models for Real-Valued Time Series

ResearchDGX agent

arXiv:2601.11079v2 Announce Type: replace Abstract: This paper proposes the soft Bayesian context tree model (Soft-BCT), which is a novel BCT model for real-valued time series. The Soft-BCT considers

22 May 2026

Boiling the Frog: A Multi-Turn Benchmark for Agentic Safety

Model ReleasesDGX agent

arXiv:2605.22643v1 Announce Type: new Abstract: Background. Traditional safety benchmarks for language models evaluate generated text: whether a model outputs toxic language, reproduces bias, or follo

Circle-RoPE: Cone-like Decoupled Rotary Positional Embedding for Large Vision-Language Models

SafetyDGX agent

arXiv:2505.16416v3 Announce Type: replace Abstract: Rotary Position Embedding (RoPE) is widely adopted in large language models, but when applied to vision-language models (VLMs) it couples text and i

Comparing LLM and Fine-Tuned Model Performance on NVDRS Circumstance Extraction with Varying Prompt Complexity

Model ReleasesDGX agent

arXiv:2605.21845v1 Announce Type: new Abstract: Suicide is a leading cause of death in the United States, and understanding the circumstances that precede it requires extracting structured information

GesVLA: Gesture-Aware Vision-Language-Action Model Embedded Representations

ApplicationsDGX agent

arXiv:2605.22812v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have shown strong potential for general-purpose robot manipulation by unifying perception and action. However, exi

I built a free demo for Pixal3D (Tencent new image-to-3D model)

Local AiDGX agent

Pixal3D is a Tencent image-to-3D model that generates high-fidelity 3D assets from a single image by explicitly lifting pixel features into 3D through back-projection to establish direct pixel-to-3D c

Open Materials 2024 (OMat24) Inorganic Materials Dataset and Models

ResearchDGX agent

arXiv:2410.12771v2 Announce Type: replace-cross Abstract: The ability to discover new materials with desirable properties is critical for numerous applications from helping mitigate climate change to

PALS: Power-Aware LLM Serving for Mixture-of-Experts Models

HardwareDGX agent

arXiv:2605.21427v1 Announce Type: new Abstract: Large language model (LLM) inference has become a dominant workload in modern data centers, driving significant GPU utilization and energy consumption.

SDGBiasBench: Benchmarking and Mitigating Vision--Language Models' Biases in Sustainable Development Goals

Model ReleasesDGX agent

arXiv:2605.21919v1 Announce Type: new Abstract: Assessing progress toward the Sustainable Development Goals (SDGs) requires multi-step reasoning over visual cues, contextual knowledge, and development

Today, Zyphra Research is sharing fundamental work extending Equilibrium Propagation beyond Energy-Based Models to biologically realistic ne…

IndustryDGX agent

Today, Zyphra Research is sharing fundamental work extending Equilibrium Propagation beyond Energy-Based Models to biologically realistic neuron models. A step toward more efficient AI, local learning

UniSD: Towards a Unified Self-Distillation Framework for Large Language Models

SafetyDGX agent

arXiv:2605.06597v2 Announce Type: replace Abstract: Self-distillation (SD) offers a promising path for adapting large language models (LLMs) without relying on stronger external teachers. However, SD

21 May 2026

Deeper Thought, Weaker Aim: Understanding and Mitigating Perceptual Impairment during Reasoning in Multimodal Large Language Models

ResearchDGX agent

arXiv:2603.14184v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) often suffer from perceptual impairments under extended reasoning modes, particularly in visual question an

Ensemble RL through Classifier Models: Enhancing Risk-Return Trade-offs in Trading Strategies

ResearchDGX agent

arXiv:2502.17518v2 Announce Type: replace Abstract: This paper presents a comprehensive study on the use of ensemble Reinforcement Learning (RL) models in financial trading strategies, leveraging clas

EvalMORAAL: Interpretable Chain-of-Thought and LLM-as-Judge Evaluation for Moral Alignment in Large Language Models

SafetyDGX agent

arXiv:2510.05942v3 Announce Type: replace Abstract: We present EvalMORAAL, a transparent chain-of-thought (CoT) framework that uses two scoring methods (log-probabilities and direct ratings) plus a mo

Explainable AI: Context-Aware Layer-Wise Integrated Gradients for Explaining Transformer Models

Local AiDGX agent

arXiv:2602.16608v2 Announce Type: replace Abstract: Transformer models achieve state-of-the-art performance across domains and tasks, yet their deeply layered representations make their predictions di

Grok 4.3 stands out for being the most intelligent model in its price range

IndustryDGX agent

Grok 4.3 stands out for being the most intelligent model in its price range We built a live job board of all the frontier labs hiring right now. Plus a way to see news, funding rounds, compare model r

Hybrid Machine Learning Model for Forest Height Estimation from TanDEM-X and Landsat Data

ResearchDGX agent

arXiv:2605.20997v1 Announce Type: new Abstract: Integrating machine learning (ML) with physical models (PM) has emerged as a promising way of retrieving geophysical parameters from remote sensing data

Lighting-aware Unified Model for Instance Segmentation

ApplicationsDGX agent

arXiv:2605.20436v1 Announce Type: new Abstract: Foundation models like the Segment Anything Model (SAM) demonstrate impressive zero-shot generalization but frequently degrade under diverse real-world

MedicalBench: Evaluating Large Language Models Toward Improved Medical Concept Extraction

Model ReleasesDGX agent

arXiv:2605.20197v1 Announce Type: new Abstract: Medical concept extraction from electronic health records underpins many downstream applications, yet remains challenging because medically meaningful c

Miller-Index-Based Latent Crystallographic Fracture Plane Reasoning with Vision-Language Models

ApplicationsDGX agent

arXiv:2605.20416v1 Announce Type: new Abstract: We study whether multimodal large language models (MLLMs) can leverage crystallographic plane indices (Miller indices) as a structured latent representa

Neural Estimation of Pairwise Mutual Information in Masked Discrete Sequence Models

ResearchDGX agent

arXiv:2605.20187v1 Announce Type: new Abstract: Understanding dependencies between variables is critical for interpretability and efficient generation in masked diffusion models (MDMs), yet these mode

PulseCol: Periodically Refreshed Column-Sparse Attention for Accelerating Diffusion Language Models

HardwareDGX agent

arXiv:2605.20813v1 Announce Type: new Abstract: Inference in diffusion large language models (dLLMs) is computationally expensive, as full self-attention must be repeatedly executed at each step of th

Q-ARVD: Quantizing Autoregressive Video Diffusion Models

ResearchDGX agent

arXiv:2605.21072v1 Announce Type: new Abstract: Autoregressive video diffusion models (ARVDs) have emerged as a promising architecture for streaming video generation, paving the way for real-time inte

RISE: Reliable Improvement in Self-Evolving Vision-Language Models

ResearchDGX agent

arXiv:2605.20914v1 Announce Type: new Abstract: Vision-language models (VLMs) have achieved strong multimodal reasoning capabilities, but further improving them still relies heavily on large-scale hum

Synchronization and Turn-Taking in Full-Duplex Speech Dialogue Models

SafetyDGX agent

arXiv:2605.20356v1 Announce Type: new Abstract: Full-duplex spoken dialogue models (SDMs) can listen and speak simultaneously, enabling interaction dynamics closer to human conversation than turn-base

TabPFN Extensions for Interpretable Geotechnical Modelling

Model ReleasesDGX agent

arXiv:2603.21033v2 Announce Type: replace-cross Abstract: Geotechnical site characterisation relies on sparse, heterogeneous borehole data, where uncertainty quantification and interpretability matter

Under Pressure: Emotional Framing Induces Measurable Behavioral Shifts and Structured Internal Geometry in Small Language Models

Model ReleasesDGX agent

arXiv:2605.20202v1 Announce Type: new Abstract: I study whether emotionally framed evaluation follow-ups change both the behavior and the calm-relative internal representations of small, locally deplo

20 May 2026

Can Large Language Models Reliably Correct Errors in Low-Resource ASR? A Contamination-Aware Case Study on West Frisian

Model ReleasesDGX agent

arXiv:2605.19711v1 Announce Type: new Abstract: Automatic speech recognition (ASR) has improved substantially in recent years, yet performance remains limited for low-resource languages. Large languag

CoLD: Counterfactually-Guided Length Debiasing for Process Reward Models in Mathematical Reasoning

SafetyDGX agent

arXiv:2507.15698v2 Announce Type: replace-cross Abstract: Process Reward Models (PRMs) play a central role in evaluating and guiding multi-step reasoning in large language models (LLMs), especially fo

Command A+ from @cohere is out now :) its our best model yet and its open source apache 2.0

IndustryDGX agent

Cohere has released Command A+, their latest and most advanced language model to date, under the Apache 2.0 open-source license. The model is now publicly available for use and development by the comm

Composition of Memory Experts for Diffusion World Models

ApplicationsDGX agent

arXiv:2605.18813v1 Announce Type: cross Abstract: World models aim to predict plausible futures consistent with past observations, a capability central to planning and decision-making in reinforcement

Fine-tuning language encoding models on slow fMRI improves prediction for fast ECoG

ResearchDGX agent

arXiv:2605.19224v1 Announce Type: new Abstract: Neuroscientists have recently turned to intracranial brain recording methods, like electrocorticography (ECoG), for human experiments because of the fin

Grok Build 0.1 is available on Vercel AI Gateway. xAI's beta coding model in Grok Build CLI. 𝚖𝚘𝚍𝚎𝚕: '𝚡𝚊𝚒/𝚐𝚛𝚘𝚔-𝚋𝚞𝚒𝚕𝚍-𝟶.𝟷' …

IndustryDGX agent

Grok Build 0.1, xAI's beta coding model, is now available on Vercel AI Gateway for developers to access and integrate into their applications. The model can be accessed via the identifier 'xai/grok-bu

Hybrid Training for Vision-Language-Action Models

AgentsDGX agent

arXiv:2510.00600v2 Announce Type: replace-cross Abstract: Using Large Language Models to produce intermediate thoughts, a.k.a. Chain-of-thought (CoT), before providing an answer has been a successful

Language models struggle with compartmentalization

TutorialsDGX agent

arXiv:2605.19284v1 Announce Type: new Abstract: In the training data used by large language models (LLMs), the same latent concept is often presented in multiple distinct ways: the same facts appear i

Lightweight and Fast Backdoor Model Detection

Model ReleasesDGX agent

arXiv:2605.18907v1 Announce Type: cross Abstract: Deep neural networks (DNN), despite their remarkable performance, are highly vulnerable to backdoor attacks. Existing defenses mainly rely on activati

Markov Chain Decoders Overcome the Heavy-Tail Limitations of Lipschitz Generative Models

ResearchDGX agent

arXiv:2605.18931v1 Announce Type: cross Abstract: Heavy-tailed distributions are prevalent in performance evaluation, network traffic, and risk modeling. This behavior poses a fundamental challenge fo

MSAlign: Aligning Molecule and Mass Spectra Foundation Models for Metabolite Identification

Model ReleasesDGX agent

arXiv:2605.19752v1 Announce Type: new Abstract: Accurately identifying metabolites i.e. small molecules from mass spectrometry data remains a core challenge in metabolomics, with broad applications in

Quantized Machine Learning Models for Medical Imaging in Low-Resource Healthcare Settings

ApplicationsDGX agent

arXiv:2605.19207v1 Announce Type: cross Abstract: Deep learning models have shown strong performance in medical image analysis, but deploying them in low-resource clinical environments remains difficu

Scaling Evaluation-time Compute with Reasoning Models as Evaluators

ResearchDGX agent

arXiv:2503.19877v2 Announce Type: replace Abstract: As language model (LM) outputs get more and more natural, it is becoming more difficult than ever to evaluate their quality. Simultaneously, increas

🌍Today we release Mosaic, a probabilistic weather model that shifts the Pareto frontier of ML weather forecasting. It matches the skill of …

HardwareDGX agent

🌍Today we release Mosaic, a probabilistic weather model that shifts the Pareto frontier of ML weather forecasting. It matches the skill of state-of-the-art models while generating a 24-member, 10-day

Transformers Linearly Represent Highly Structured World Models

ResearchDGX agent

arXiv:2605.18847v1 Announce Type: cross Abstract: Do transformers, when trained on sequential reasoning traces, build internal models of the underlying task? And if so, does the structure of those int

When Preference Labels Fall Short: Aligning Diffusion Models from Real Data

SafetyDGX agent

arXiv:2605.19839v1 Announce Type: new Abstract: Preference alignment aims to guide generative models by learning from comparisons between preferred and non-preferred samples. In practice, most existin

When Tabular Foundation Models Meet Strategic Tabular Data: A Prior Alignment Approach

SafetyDGX agent

arXiv:2605.19662v1 Announce Type: new Abstract: Tabular foundation models based on pretrained prior-data fitted networks~(PFNs) have shown strong generalization on diverse tabular tasks, but they are

19 May 2026

Adversarial Attacks on Downstream Weather Forecasting Models: Application to Tropical Cyclone Trajectory Prediction

ResearchDGX agent

arXiv:2510.10140v2 Announce Type: replace Abstract: Deep learning-based weather forecasting (DLWF) models leverage past weather observations to generate future forecasts, supporting a wide range of do

Attention Hijacking: Response Manipulation Across Queries in Vision-Language Models

ResearchDGX agent

arXiv:2605.17310v1 Announce Type: cross Abstract: Existing adversarial attacks on vision-language models (VLMs) can steer model outputs toward attacker-specified target responses, but their effectiven

CBT-Audio: Evaluating Audio Language Models for Patient-Side Distress Intensity Estimation in CBT Session Recordings

TutorialsDGX agent

arXiv:2605.17370v1 Announce Type: new Abstract: Cognitive behavioural therapy is widely used to help patients understand and manage psychological distress. It is often delivered through spoken convers

← Previous
1…135136137138139…1009
Next →