AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,565 results
26 May 2026

How Neural Reward Models Learn Features for Policy Optimization: A Single-Index Analysis

SafetyDGX agent

arXiv:2605.24749v1 Announce Type: cross Abstract: Reward modeling is not only a prediction problem: in KL-regularized policy optimization, the learned reward is exponentiated to define the deployed po

Improving Ensemble CAPE Forecasts with a Diffusion Model Incorporating Aerosol Information

SafetyDGX agent

arXiv:2605.24009v1 Announce Type: cross Abstract: Convective available potential energy (CAPE) is an important variable for forecasting severe weather and understanding deep convection and precipitati

Know You Before You Speak: User-State Modeling for LLM Personalization in Multi-Turn Conversation

ApplicationsDGX agent

arXiv:2605.24647v1 Announce Type: new Abstract: Personalized dialogue requires more than recalling explicit user histories: systems also need to infer hidden user states that evolve through interactio

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

MultiPUFFIN: A Multimodal Domain-Constrained Foundation Model for Molecular Property Prediction of Small Molecules

TutorialsDGX agent

arXiv:2603.00857v2 Announce Type: replace-cross Abstract: MultiPUFFIN is a domain-informed multimodal foundation model for predicting thermophysical properties of small molecules, addressing a critica

Paris 2.0: A Decentralized Diffusion Model for Video Generation

HardwareDGX agent

arXiv:2605.26064v1 Announce Type: cross Abstract: We present Paris 2.0, the first video generation model pre-trained through decentralized computation. Its training recipe builds upon Paris 1.0 (arXiv

PathWise: Planning through World Model for Automated Heuristic Design via Self-Evolving LLMs

SafetyDGX agent

arXiv:2601.20539v3 Announce Type: replace Abstract: Large Language Models (LLMs) have enabled automated heuristic design (AHD) for combinatorial optimization problems (COPs), but existing frameworks'

Polynomial Context-Truncation Sensitivity in Autoregressive Language Models: Sequential Wyner-Ziv Bounds for KV Cache Compression

SafetyDGX agent

arXiv:2605.25085v1 Announce Type: cross Abstract: We study the rate-distortion limits of online KV cache compression in autoregressive language models, formulating it as sequential Wyner-Ziv source co

Repeated Sequences Reveal Gaps between Large Language Models and Natural Language

ResearchDGX agent

arXiv:2605.24850v1 Announce Type: new Abstract: Evaluating whether large language models (LLMs) capture the structure of natural language beyond local fluency remains an open challenge. Existing evalu

SEIDM: A Safe and Efficient Intelligent Driver Model for Autonomous Driving Behavior

SafetyDGX agent

arXiv:2605.23915v1 Announce Type: cross Abstract: The Intelligent Driver Model (IDM) is a cornerstone of Adaptive Cruise Control (ACC), valued for its interpretable parameters and effectiveness in car

SliceWorld: A Predictive and Controllable World-State Model for CT Report Generation

SafetyDGX agent

arXiv:2605.24371v1 Announce Type: cross Abstract: CT report generation (CTRG) requires models to summarize three-dimensional anatomical context and pathological findings from hundreds of axial slices.

SpecPrune-VLA: Accelerating Vision-Language-Action Models via Action-Aware Self-Speculative Pruning

Local AiDGX agent

arXiv:2509.05614v3 Announce Type: replace-cross Abstract: Pruning is a typical acceleration technique for compute-bound models by removing computation on unimportant values. Recently, it has been appl

Summoning the Oracle to Slay It: Mitigating Look-Ahead Bias in Financial Backtesting with Large Language Models

SafetyDGX agent

arXiv:2605.24564v1 Announce Type: new Abstract: Backtesting large language models (LLMs) on historical financial data is unreliable because pre-training cuts off after the events happened. An LLM trai

Teaching large language models to reason like expert diagnosticians

Model ReleasesDGX agent

arXiv:2509.12194v2 Announce Type: replace Abstract: Differential diagnosis is an iterative process that integrates patient information with broader medical knowledge. Clinical case series such as the

The Path Matters: Learning a Token-Commitment Policy for Diffusion Language Models

SafetyDGX agent

arXiv:2605.24697v1 Announce Type: cross Abstract: Diffusion large language models promise faster generation by refining many token positions in parallel, but this parallelism introduces a hidden contr

X-DiffVLA: X-Embodied Diffusion Action Heads for Vision-Language-Action Models

SafetyDGX agent

arXiv:2605.25044v1 Announce Type: new Abstract: Learning universal policies from cross-embodied data remains a fundamental challenge in robotics. Although Vision-Language-Action (VLA) models are pre-t

25 May 2026

Assessing Predictive Models for Fairness Based on Movement Patterns

SafetyDGX agent

arXiv:2605.23234v1 Announce Type: new Abstract: Assessing the spatial fairness of predictive models involves establishing whether they are statistically penalizing (favoring) individuals associated wi

Built a local MCP memory server that uses Ollama to give AI coding assistants persistent memory — no cloud, no API keys, Tool and Model Agnostic

Local AiDGX agent

A developer created a local Model Context Protocol (MCP) memory server that integrates Ollama to provide AI coding assistants with persistent memory capabilities while maintaining complete privacy and

Diffusion Domain Expansion: Learning to Coordinate Pre-trained Diffusion Models

ResearchDGX agent

arXiv:2605.23275v1 Announce Type: new Abstract: In this paper, we propose Diffusion Domain Expansion (DDE), a method that efficiently extends pre-trained diffusion models to generate larger objects an

Energy-Guided Generative Modeling for Low-Energy Molecular Structure Discovery

ResearchDGX agent

arXiv:2512.22597v2 Announce Type: replace Abstract: Exploring molecular energy landscapes and identifying ground-state conformations are central challenges in computational chemistry. However, generat

LACY: A Vision-Language Model-based Language-Action Cycle for Self-Improving Robotic Manipulation

AgentsDGX agent

arXiv:2511.02239v2 Announce Type: replace-cross Abstract: Learning generalizable policies for robotic manipulation increasingly relies on large-scale models that map language instructions to actions (

Learned Relay Representations for Forward-Thinking Discrete Diffusion Models

TutorialsDGX agent

arXiv:2605.22967v1 Announce Type: new Abstract: When Masked Diffusion Models (MDMs) generate sequences through iterative refinement, the rich internal computation over masked positions is discarded, f

Point Tracking Improves World Action Models

SafetyDGX agent

arXiv:2605.23856v1 Announce Type: new Abstract: Robot policy learning benefits from world-action models that capture environment dynamics, but pixel-level prediction entangles dynamics with nuisance f

SCOPE: Simulating Cross-game Operations in Playable Environments for FPS World Models

Local AiDGX agent

arXiv:2605.23345v1 Announce Type: new Abstract: Interactive world models for first-person shooter (FPS) games must resolve high-frequency overlapping control signals at every frame without disrupting

UniReg: A Universal Model for Controllable CT Image Registration

SafetyDGX agent

arXiv:2503.12868v2 Announce Type: replace Abstract: Learning-based medical image registration has matched the accuracy of conventional methods while offering superior computational efficiency. However

USIM and U0: A Vision-Language-Action Dataset and Model for General Underwater Robots

ResearchDGX agent

arXiv:2510.07869v4 Announce Type: replace Abstract: Underwater environments pose unique challenges for robotic navigation and manipulation. While existing research has primarily focused on task-specif

23 May 2026

Decision Potential Surface: A Theoretical and Practical Approximation of Large Language Model Decision Boundary

ResearchDGX agent

arXiv:2510.03271v2 Announce Type: replace Abstract: Decision boundary, the subspace of inputs where a machine learning model assigns equal classification probabilities to two classes, is pivotal in re

Equilibrium Propagation and Hamiltonian Inference in the Diffusive Fitzhugh-Nagumo Model

ResearchDGX agent

arXiv:2605.21568v1 Announce Type: new Abstract: In this work, we extend the Equilibrium Propagation framework to skew-gradient systems and show an equivalence between deep Energy-Based Models and Hami

Relational Linear Properties in Language Models: An Empirical Investigation

ResearchDGX agent

arXiv:2605.22532v1 Announce Type: new Abstract: Linear properties are ubiquitous in the representations of language models; however, testing them experimentally remains a challenging task. This work f

six years ago world (cognitive) models were the centerpiece of my essay The Next Decade in AI. their time is finally coming.

SafetyDGX agent

six years ago world (cognitive) models were the centerpiece of my essay The Next Decade in AI. their time is finally coming. Demis Hassabis on the limit in today’s AI: language can describe the world,

World models have existed for years (though not in LLMs); I take them to be explicit representation of objects, places, events, mechanisms e…

SafetyDGX agent

World models have existed for years (though not in LLMs); I take them to be explicit representation of objects, places, events, mechanisms etc you can reason over. Chess computers have them (board, pi

22 May 2026

Ablate-to-Validate: Are Vision-Language Models Really Using Continuous Thought Tokens?

ResearchDGX agent

arXiv:2605.21642v1 Announce Type: new Abstract: Vision-language models (VLMs) are increasingly augmented with continuous or latent non-textual tokens intended to support 'visual thinking.' Despite imp

Enhancing Multimodal Large Language Models for Safety-Critical Driving Video Analysis

SafetyDGX agent

arXiv:2605.22185v1 Announce Type: new Abstract: Recent advancements in Multimodal Large Language Models (MLLMs) have demonstrated impressive capabilities in general visual understanding. However, thei

EvoVid: Temporal-Centric Self-Evolution for Video Large Language Models

AgentsDGX agent

arXiv:2605.21931v1 Announce Type: new Abstract: Recent Video Large Language Models (Video-LLMs) have demonstrated strong capabilities in video reasoning through reinforcement learning (RL). However, e

General Agentic Planning Through Simulative Reasoning with World Models

AgentsDGX agent

arXiv:2507.23773v3 Announce Type: replace-cross Abstract: What does it mean to plan? Current agentic systems, whether scaffolded workflows or end-to-end policies, rely on reactive decision-making: sel

VSAS-Bench: Real-Time Evaluation of Visual Streaming Assistant Models

ResearchDGX agent

Streaming vision-language models (VLMs) continuously generate responses given an instruction prompt and an online stream of input frames. This is a core mechanism for real-time visual assistants. Exis

21 May 2026

Compute Only Once: UG-Separation for Efficient Large Recommendation Models

ResearchDGX agent

arXiv:2602.10455v2 Announce Type: replace-cross Abstract: Driven by scaling laws, recommender systems increasingly rely on larger-scale models to capture complex feature interactions and user behavior

Deformba: Vision State Space Model with Adaptive State Fusion

ResearchDGX agent

arXiv:2605.21308v1 Announce Type: new Abstract: State Space Models (SSMs) have emerged as a powerful and efficient alternative to Transformers, demonstrating linear-time complexity and exceptional seq

DNACHUNKER: Learnable Tokenization for DNA Language Models

Local AiDGX agent

arXiv:2601.03019v4 Announce Type: replace-cross Abstract: DNA language models are increasingly used to represent genomic sequence, yet their effectiveness depends critically on how raw nucleotides are

Dual Quaternion Based Contact Modeling for Fast and Smooth Collision Recovery of Quadrotors

ResearchDGX agent

arXiv:2603.14698v3 Announce Type: replace Abstract: Unmanned aerial vehicles (UAVs) operating in cluttered environments require efficient and accurate impact modeling to maintain stability post collis

Efficient numeracy in language models through single-token number embeddings

TutorialsDGX agent

arXiv:2510.06824v2 Announce Type: replace Abstract: To drive progress in science and engineering, large language models (LLMs) must be able to process large amounts of numerical data and solve long ca

FEAT: A Linear-Complexity Foundation Model for Extremely Large Structured Data

SafetyDGX agent

arXiv:2603.16513v3 Announce Type: replace Abstract: Structured data is widely used in domains such as healthcare, finance, and scientific data management. Recent studies on structured data foundation

LASH: Adaptive Semantic Hybridization for Black-Box Jailbreaking of Large Language Models

SafetyDGX agent

arXiv:2605.21362v1 Announce Type: new Abstract: Jailbreak attacks expose a persistent gap between the intended safety behavior of aligned large language models and their behavior under adversarial pro

Let EEG Models Learn EEG

TutorialsDGX agent

arXiv:2605.21280v1 Announce Type: new Abstract: High-fidelity EEG generation is critical for alleviating data scarcity and addressing privacy constraints in large-scale neural modeling. Despite recent

MAPS: A Synthetic Dataset for Probing Vision Models in a Controlled 3D Scene Space

ApplicationsDGX agent

arXiv:2605.20549v1 Announce Type: new Abstract: Modern vision models achieve strong performance on standard benchmarks, yet their aggregate accuracy reveals little about which scene properties drive t

Mechanisms of Misgeneralization in Physical Sequence Modeling

Local AiDGX agent

arXiv:2605.20299v1 Announce Type: new Abstract: Generative sequence models are often trained to plan motion in physical domains, from robotics to mechanical simulations. When constructing a dataset to

Modeling and Control of a Pneumatic Morphing Soft Quadrotor based on the SOFA Framework for Dynamic Soft Robotic Simulation

ResearchDGX agent

arXiv:2605.21031v1 Announce Type: new Abstract: This article presents a novel SOFA based finite element method for the soft body modeling and the corresponding dynamic simulation and control of a pneu

On the Regularity and Generalization of One-Step Wasserstein-guided Generative Models for PDE-Induced Measures

TutorialsDGX agent

arXiv:2605.21388v1 Announce Type: new Abstract: Despite the remarkable empirical success of generative models, the available theory on their statistical accuracy in scientific computing remains largel

PaintCopilot: Modeling Painting as Autonomous Artistic Continuation

Local AiDGX agent

arXiv:2605.20941v1 Announce Type: new Abstract: We present PaintCopilot, a co-creative neural painting assistant that models painting as an open-ended autoregressive artistic behavior conditioned on e

Quantum reservoir computing in Jaynes-Cummings models: Nonlinear memory and time-series prediction

Model ReleasesDGX agent

arXiv:2510.00171v2 Announce Type: replace-cross Abstract: We investigate quantum reservoir computing (QRC) using a hybrid qubit-boson system described by the Jaynes-Cummings (JC) Hamiltonian and its d

Robust Subspace-Constrained Quadratic Models for Low-Dimensional Structure Learning

ResearchDGX agent

arXiv:2605.20300v1 Announce Type: new Abstract: In this paper, we propose a robust subspace-constrained quadratic model (SCQM) for learning low-dimensional structure from high-dimensional data. Buildi

SynCB: A Synergy Concept-Based Model with Dynamic Routing Between Concepts and Complementary Neural Branches

SafetyDGX agent

arXiv:2605.20908v1 Announce Type: new Abstract: Concept-based (CB) models provide interpretability and support test-time human intervention, while standard neural networks (NN) offer strong task perfo

Ten years ago, I walked into Tesla as an intern — January 2015, back when all we built was the Model S. I never imagined I'd one day lead th…

IndustryDGX agent

Ten years ago, I walked into Tesla as an intern — January 2015, back when all we built was the Model S. I never imagined I'd one day lead the Fremont factory where we are building some of the most ama

Tippett-minimum Fusion of Representation-space Diffusion Models for Multi-Encoder Out-of-Distribution Detection

Model ReleasesDGX agent

arXiv:2605.20502v1 Announce Type: cross Abstract: We address out-of-distribution (OOD) detection across the full spectrum of distribution shifts -- global domain changes, semantic divergence, texture

20 May 2026

A Logistic Regression Model to Predict Malaria Severity in Children

AgentsDGX agent

arXiv:2605.18900v1 Announce Type: cross Abstract: One of the main causes of death around the globe is malaria. Researchers have sought to develop predictive models for malaria outbreaks based on meteo

AffectVerse: Emotional World Models for Multimodal Affective Computing

ResearchDGX agent

arXiv:2605.19950v1 Announce Type: new Abstract: Humans infer emotions by integrating observed multimodal cues with expectations about how affective states may unfold. Existing multimodal large languag

Awakening the Hydra: Stabilizing Multi-Concept Backdoor Injection in Text-to-Image Diffusion Models

ResearchDGX agent

arXiv:2605.19698v1 Announce Type: cross Abstract: Text-to-image diffusion models are increasingly developed through open-source reuse and repeated downstream fine-tuning, where reused checkpoints are

Conflict-Resilient Multi-Agent Reasoning via Signed Graph Modeling

Model ReleasesDGX agent

arXiv:2605.19418v1 Announce Type: new Abstract: LLM-based multi-agent systems (MAS) have demonstrated strong reasoning and decision-making capabilities that consistently surpass those of single LLM ag

Data-centric Design of Learning-based Surgical Gaze Perception Models in Multi-Task Simulation

TutorialsDGX agent

arXiv:2602.09259v2 Announce Type: replace Abstract: In robot-assisted minimally invasive surgery (RMIS), reduced haptic feedback and depth cues increase reliance on expert visual perception, motivatin

Extreme Self-Preference in Language Models

SafetyDGX agent

arXiv:2509.26464v2 Announce Type: replace Abstract: Self-preference is a fundamental feature of biological organisms. Since large language models (LLMs) lack sentience, they might be expected to avoid

FlowErase-RL: Rethinking Concept Erasure as Reward Optimization in Flow Matching Models

SafetyDGX agent

arXiv:2605.19739v1 Announce Type: new Abstract: Recent advances in flow matching models have significantly improved text-to-image generation quality, but also introduce growing safety risks due to the

← Previous
1…158159160161162…1010
Next →