AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,588 results
Model Releases

PEDESTRIANQA: A Benchmark for Vision-Language Models on Pedestrian Intention and Trajectory Prediction

DGX agent

arXiv:2605.24562v1 Announce Type: cross Abstract: Pedestrian intention and trajectory prediction are critical for the safe deployment of autonomous driving systems, directly influencing navigation dec

model-releasesarxiv-cs-ai
26 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

PiXTime: A Model for Federated Time Series Forecasting with Heterogeneous Data across Nodes

DGX agent

arXiv:2601.05613v2 Announce Type: replace-cross Abstract: While collaborative forecasting on distributed time series is highly desirable, directly pooling localized datasets is often impractical due t

model-releasesarxiv-cs-ai
26 May 2026
Safety

Reason--Imagine--Act: Closed-Loop LLM Decision Making with World Models for Autonomous Driving

DGX agent

arXiv:2605.24004v1 Announce Type: new Abstract: Large language models (LLMs) are promising for autonomous driving, but semantics-only decision policies can yield physically unsafe behavior in dynamic

safetyarxiv-cs-ai
26 May 2026
Research

SAE-FD: Sparse Autoencoder Feature Distillation for Continual Learning of Large Language Models

DGX agent

arXiv:2605.25525v1 Announce Type: new Abstract: Continual learning enables large language models to adapt to evolving tasks without retraining from scratch, yet catastrophic forgetting remains a centr

researcharxiv-cs-lg
26 May 2026
Model Releases

Structural Abstraction as an Inductive Bias for Non-Stationary Language Model Training

DGX agent

arXiv:2603.17198v2 Announce Type: replace-cross Abstract: A foundational principle in cognitive science holds that intelligent agents do not learn by storing experiences as isolated instances, but by

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

TimeSpot: Benchmarking Geo-Temporal Understanding in Vision-Language Models in Real-World Settings

DGX agent

arXiv:2603.06687v2 Announce Type: replace-cross Abstract: Geo-temporal understanding, the ability to infer location, time, and contextual properties from visual input alone, underpins applications suc

model-releasesarxiv-cs-cl
26 May 2026
Safety

Universal Boosts, Specific Suppressors: Sparse Autoencoder Steering of Medical Vision-Language Models

DGX agent

arXiv:2605.24977v1 Announce Type: cross Abstract: Medical vision-language models (VLMs) often hallucinate findings when generating chest X-ray reports: they fabricate findings that are not present in

safetyarxiv-cs-cl
26 May 2026
Model Releases

VeriTrace: Evolving Mental Models for Deep Research Agents

DGX agent

arXiv:2605.26081v1 Announce Type: new Abstract: Deep research agents face vast, interdependent, and pervasively uncertain information. Existing systems explore what evolving intermediate representatio

model-releasesarxiv-cs-ai
26 May 2026
Safety

Beyond VLM-Based Rewards: Diffusion-Native Latent Reward Modeling

DGX agent

arXiv:2602.11146v2 Announce Type: replace-cross Abstract: Preference optimization for diffusion and flow-matching models relies on reward functions that are both discriminatively robust and computatio

safetyarxiv-cs-ai
25 May 2026
Safety

Differences in Typological Alignment in Language Models' Treatment of Differential Argument Marking

DGX agent

arXiv:2602.17653v2 Announce Type: replace Abstract: Recent work has shown that language models (LMs) trained on synthetic corpora can exhibit typological preferences that resemble cross-linguistic reg

safetyarxiv-cs-cl
25 May 2026
Research

Dithering Defense: Adversarial Robustness of Vision Foundation Models via Multi-Level Floyd-Steinberg Dithering

DGX agent

arXiv:2605.23065v1 Announce Type: cross Abstract: Vision foundation models are widely used as frozen backbones across many downstream tasks, making them a single point of failure under adversarial att

researcharxiv-cs-ai
25 May 2026
Applications

GEM-4D: Geometry-Enhanced Video World Models for Robot Manipulation

DGX agent

arXiv:2605.22882v1 Announce Type: new Abstract: Video world models can generate realistic futures from a single instruction, but they often fail to preserve consistent point-level motion over time. As

applicationsarxiv-cs-cv
25 May 2026
Applications

How Human-Like Are Large Language Models? A Register-Aware Linguistic Evaluation Framework

DGX agent

arXiv:2605.23651v1 Announce Type: new Abstract: While factual correctness and task-performance have been in focus of Large Language Model (LLM) research for a long time, the fundamental question of ho

applicationsarxiv-cs-cl
25 May 2026
Tutorials

Learnability-Informed Fine-Tuning of Diffusion Language Models

DGX agent

arXiv:2605.22939v1 Announce Type: new Abstract: We aim to improve the reasoning capabilities of diffusion language models (DLMs). While SFT is a popular post-training recipe for autoregressive models,

tutorialsarxiv-cs-cl
25 May 2026
Tutorials

VAMP-Diff: VampPrior Latent Diffusion for Photoplethysmography Modeling

DGX agent

arXiv:2605.22851v1 Announce Type: cross Abstract: Photoplethysmography (PPG) has become a ubiquitous physiological signal; however, current generative models still struggle to preserve realistic wavef

tutorialsarxiv-cs-lg
25 May 2026
Research

What Does the Server See? Understanding Privacy Leakage from Large Language Models in Split Inference

DGX agent

arXiv:2605.23158v1 Announce Type: cross Abstract: The deployment of large language models (LLMs) on resource-constrained devices remains challenging, spurring interest in split inference, where models

researcharxiv-cs-cl
25 May 2026
Research

A Diffusive Classification Loss for Learning Energy-based Generative Models

DGX agent

arXiv:2601.21025v3 Announce Type: replace-cross Abstract: Score-based generative models have recently achieved remarkable success. While they are usually parameterized by the score, an alternative way

researcharxiv-cs-lg
23 May 2026
Research

Bringing Stability to Diffusion: Decomposing and Reducing Variance of Training Masked Diffusion Models

DGX agent

arXiv:2511.18159v2 Announce Type: replace Abstract: Masked diffusion models (MDMs) are a promising alternative to autoregressive models (ARMs), but they suffer from inherently much higher training var

researcharxiv-cs-lg
23 May 2026
Research

CellFluxRL: Biologically-Constrained Virtual Cell Modeling via Reinforcement Learning

DGX agent

arXiv:2603.21743v4 Announce Type: replace Abstract: Building virtual cells with generative models to simulate cellular behavior in silico is emerging as a promising paradigm for accelerating drug disc

researcharxiv-cs-lg
23 May 2026
Safety

DecepChain: Inducing Deceptive Reasoning in Large Language Models

DGX agent

arXiv:2510.00319v2 Announce Type: replace Abstract: Large Language Models (LLMs) have been demonstrating strong reasoning capability with their chain-of-thoughts (CoT), which are routinely used by hum

safetyarxiv-cs-lg
23 May 2026
Tutorials

MDM-Prime-v2: Binary Encoding and Index Shuffling Enable Scaling of Diffusion Language Models

DGX agent

arXiv:2603.16077v3 Announce Type: replace Abstract: Masked diffusion models (MDM) exhibit superior generalization when learned using a Partial masking scheme (Prime). This approach converts tokens int

tutorialsarxiv-cs-lg
23 May 2026
Research

Soft Bayesian Context Tree Models for Real-Valued Time Series

DGX agent

arXiv:2601.11079v2 Announce Type: replace Abstract: This paper proposes the soft Bayesian context tree model (Soft-BCT), which is a novel BCT model for real-valued time series. The Soft-BCT considers

researcharxiv-cs-lg
23 May 2026
Model Releases

Boiling the Frog: A Multi-Turn Benchmark for Agentic Safety

DGX agent

arXiv:2605.22643v1 Announce Type: new Abstract: Background. Traditional safety benchmarks for language models evaluate generated text: whether a model outputs toxic language, reproduces bias, or follo

model-releasesarxiv-cs-cl
22 May 2026
Safety

Circle-RoPE: Cone-like Decoupled Rotary Positional Embedding for Large Vision-Language Models

DGX agent

arXiv:2505.16416v3 Announce Type: replace Abstract: Rotary Position Embedding (RoPE) is widely adopted in large language models, but when applied to vision-language models (VLMs) it couples text and i

safetyarxiv-cs-cv
22 May 2026
Model Releases

Comparing LLM and Fine-Tuned Model Performance on NVDRS Circumstance Extraction with Varying Prompt Complexity

DGX agent

arXiv:2605.21845v1 Announce Type: new Abstract: Suicide is a leading cause of death in the United States, and understanding the circumstances that precede it requires extracting structured information

model-releasesarxiv-cs-cl
22 May 2026
Applications

GesVLA: Gesture-Aware Vision-Language-Action Model Embedded Representations

DGX agent

arXiv:2605.22812v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have shown strong potential for general-purpose robot manipulation by unifying perception and action. However, exi

applicationsarxiv-cs-cv
22 May 2026
Local Ai

I built a free demo for Pixal3D (Tencent new image-to-3D model)

DGX agent

Pixal3D is a Tencent image-to-3D model that generates high-fidelity 3D assets from a single image by explicitly lifting pixel features into 3D through back-projection to establish direct pixel-to-3D c

local-air-stablediffusion
22 May 2026
Research

Open Materials 2024 (OMat24) Inorganic Materials Dataset and Models

DGX agent

arXiv:2410.12771v2 Announce Type: replace-cross Abstract: The ability to discover new materials with desirable properties is critical for numerous applications from helping mitigate climate change to

researcharxiv-cs-ai
22 May 2026
Hardware

PALS: Power-Aware LLM Serving for Mixture-of-Experts Models

DGX agent

arXiv:2605.21427v1 Announce Type: new Abstract: Large language model (LLM) inference has become a dominant workload in modern data centers, driving significant GPU utilization and energy consumption.

hardwarearxiv-cs-ai
22 May 2026
Model Releases

SDGBiasBench: Benchmarking and Mitigating Vision--Language Models' Biases in Sustainable Development Goals

DGX agent

arXiv:2605.21919v1 Announce Type: new Abstract: Assessing progress toward the Sustainable Development Goals (SDGs) requires multi-step reasoning over visual cues, contextual knowledge, and development

model-releasesarxiv-cs-cv
22 May 2026
Industry

Today, Zyphra Research is sharing fundamental work extending Equilibrium Propagation beyond Energy-Based Models to biologically realistic ne…

DGX agent

Today, Zyphra Research is sharing fundamental work extending Equilibrium Propagation beyond Energy-Based Models to biologically realistic neuron models. A step toward more efficient AI, local learning

industryemad-mostaque--x
22 May 2026
Safety

UniSD: Towards a Unified Self-Distillation Framework for Large Language Models

DGX agent

arXiv:2605.06597v2 Announce Type: replace Abstract: Self-distillation (SD) offers a promising path for adapting large language models (LLMs) without relying on stronger external teachers. However, SD

safetyarxiv-cs-cl
22 May 2026
Research

Deeper Thought, Weaker Aim: Understanding and Mitigating Perceptual Impairment during Reasoning in Multimodal Large Language Models

DGX agent

arXiv:2603.14184v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) often suffer from perceptual impairments under extended reasoning modes, particularly in visual question an

researcharxiv-cs-cv
21 May 2026
Research

Ensemble RL through Classifier Models: Enhancing Risk-Return Trade-offs in Trading Strategies

DGX agent

arXiv:2502.17518v2 Announce Type: replace Abstract: This paper presents a comprehensive study on the use of ensemble Reinforcement Learning (RL) models in financial trading strategies, leveraging clas

researcharxiv-cs-lg
21 May 2026
Safety

EvalMORAAL: Interpretable Chain-of-Thought and LLM-as-Judge Evaluation for Moral Alignment in Large Language Models

DGX agent

arXiv:2510.05942v3 Announce Type: replace Abstract: We present EvalMORAAL, a transparent chain-of-thought (CoT) framework that uses two scoring methods (log-probabilities and direct ratings) plus a mo

safetyarxiv-cs-cl
21 May 2026
Local Ai

Explainable AI: Context-Aware Layer-Wise Integrated Gradients for Explaining Transformer Models

DGX agent

arXiv:2602.16608v2 Announce Type: replace Abstract: Transformer models achieve state-of-the-art performance across domains and tasks, yet their deeply layered representations make their predictions di

local-aiarxiv-cs-cl
21 May 2026
Industry

Grok 4.3 stands out for being the most intelligent model in its price range

DGX agent

Grok 4.3 stands out for being the most intelligent model in its price range We built a live job board of all the frontier labs hiring right now. Plus a way to see news, funding rounds, compare model r

industryelon-musk--x
21 May 2026
Research

Hybrid Machine Learning Model for Forest Height Estimation from TanDEM-X and Landsat Data

DGX agent

arXiv:2605.20997v1 Announce Type: new Abstract: Integrating machine learning (ML) with physical models (PM) has emerged as a promising way of retrieving geophysical parameters from remote sensing data

researcharxiv-cs-cv
21 May 2026
Applications

Lighting-aware Unified Model for Instance Segmentation

DGX agent

arXiv:2605.20436v1 Announce Type: new Abstract: Foundation models like the Segment Anything Model (SAM) demonstrate impressive zero-shot generalization but frequently degrade under diverse real-world

applicationsarxiv-cs-cv
21 May 2026
Model Releases

MedicalBench: Evaluating Large Language Models Toward Improved Medical Concept Extraction

DGX agent

arXiv:2605.20197v1 Announce Type: new Abstract: Medical concept extraction from electronic health records underpins many downstream applications, yet remains challenging because medically meaningful c

model-releasesarxiv-cs-cl
21 May 2026
Applications

Miller-Index-Based Latent Crystallographic Fracture Plane Reasoning with Vision-Language Models

DGX agent

arXiv:2605.20416v1 Announce Type: new Abstract: We study whether multimodal large language models (MLLMs) can leverage crystallographic plane indices (Miller indices) as a structured latent representa

applicationsarxiv-cs-lg
21 May 2026
Research

Neural Estimation of Pairwise Mutual Information in Masked Discrete Sequence Models

DGX agent

arXiv:2605.20187v1 Announce Type: new Abstract: Understanding dependencies between variables is critical for interpretability and efficient generation in masked diffusion models (MDMs), yet these mode

researcharxiv-cs-lg
21 May 2026
Hardware

PulseCol: Periodically Refreshed Column-Sparse Attention for Accelerating Diffusion Language Models

DGX agent

arXiv:2605.20813v1 Announce Type: new Abstract: Inference in diffusion large language models (dLLMs) is computationally expensive, as full self-attention must be repeatedly executed at each step of th

hardwarearxiv-cs-cl
21 May 2026
Research

Q-ARVD: Quantizing Autoregressive Video Diffusion Models

DGX agent

arXiv:2605.21072v1 Announce Type: new Abstract: Autoregressive video diffusion models (ARVDs) have emerged as a promising architecture for streaming video generation, paving the way for real-time inte

researcharxiv-cs-cv
21 May 2026
Research

RISE: Reliable Improvement in Self-Evolving Vision-Language Models

DGX agent

arXiv:2605.20914v1 Announce Type: new Abstract: Vision-language models (VLMs) have achieved strong multimodal reasoning capabilities, but further improving them still relies heavily on large-scale hum

researcharxiv-cs-cv
21 May 2026
Safety

Synchronization and Turn-Taking in Full-Duplex Speech Dialogue Models

DGX agent

arXiv:2605.20356v1 Announce Type: new Abstract: Full-duplex spoken dialogue models (SDMs) can listen and speak simultaneously, enabling interaction dynamics closer to human conversation than turn-base

safetyarxiv-cs-cl
21 May 2026
Model Releases

TabPFN Extensions for Interpretable Geotechnical Modelling

DGX agent

arXiv:2603.21033v2 Announce Type: replace-cross Abstract: Geotechnical site characterisation relies on sparse, heterogeneous borehole data, where uncertainty quantification and interpretability matter

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Under Pressure: Emotional Framing Induces Measurable Behavioral Shifts and Structured Internal Geometry in Small Language Models

DGX agent

arXiv:2605.20202v1 Announce Type: new Abstract: I study whether emotionally framed evaluation follow-ups change both the behavior and the calm-relative internal representations of small, locally deplo

model-releasesarxiv-cs-cl
21 May 2026
← Previous
1…169170171172173…1263
Next →