AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,435 results
Model Releases

PEDESTRIANQA: A Benchmark for Vision-Language Models on Pedestrian Intention and Trajectory Prediction

DGX agent

arXiv:2605.24562v1 Announce Type: cross Abstract: Pedestrian intention and trajectory prediction are critical for the safe deployment of autonomous driving systems, directly influencing navigation dec

model-releasesarxiv-cs-ai
26 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

PiXTime: A Model for Federated Time Series Forecasting with Heterogeneous Data across Nodes

DGX agent

arXiv:2601.05613v2 Announce Type: replace-cross Abstract: While collaborative forecasting on distributed time series is highly desirable, directly pooling localized datasets is often impractical due t

model-releasesarxiv-cs-ai
26 May 2026
Safety

Reason--Imagine--Act: Closed-Loop LLM Decision Making with World Models for Autonomous Driving

DGX agent

arXiv:2605.24004v1 Announce Type: new Abstract: Large language models (LLMs) are promising for autonomous driving, but semantics-only decision policies can yield physically unsafe behavior in dynamic

safetyarxiv-cs-ai
26 May 2026
Research

SAE-FD: Sparse Autoencoder Feature Distillation for Continual Learning of Large Language Models

DGX agent

arXiv:2605.25525v1 Announce Type: new Abstract: Continual learning enables large language models to adapt to evolving tasks without retraining from scratch, yet catastrophic forgetting remains a centr

researcharxiv-cs-lg
26 May 2026
Model Releases

Structural Abstraction as an Inductive Bias for Non-Stationary Language Model Training

DGX agent

arXiv:2603.17198v2 Announce Type: replace-cross Abstract: A foundational principle in cognitive science holds that intelligent agents do not learn by storing experiences as isolated instances, but by

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

TimeSpot: Benchmarking Geo-Temporal Understanding in Vision-Language Models in Real-World Settings

DGX agent

arXiv:2603.06687v2 Announce Type: replace-cross Abstract: Geo-temporal understanding, the ability to infer location, time, and contextual properties from visual input alone, underpins applications suc

model-releasesarxiv-cs-cl
26 May 2026
Safety

Universal Boosts, Specific Suppressors: Sparse Autoencoder Steering of Medical Vision-Language Models

DGX agent

arXiv:2605.24977v1 Announce Type: cross Abstract: Medical vision-language models (VLMs) often hallucinate findings when generating chest X-ray reports: they fabricate findings that are not present in

safetyarxiv-cs-cl
26 May 2026
Model Releases

VeriTrace: Evolving Mental Models for Deep Research Agents

DGX agent

arXiv:2605.26081v1 Announce Type: new Abstract: Deep research agents face vast, interdependent, and pervasively uncertain information. Existing systems explore what evolving intermediate representatio

model-releasesarxiv-cs-ai
26 May 2026
Safety

Beyond VLM-Based Rewards: Diffusion-Native Latent Reward Modeling

DGX agent

arXiv:2602.11146v2 Announce Type: replace-cross Abstract: Preference optimization for diffusion and flow-matching models relies on reward functions that are both discriminatively robust and computatio

safetyarxiv-cs-ai
25 May 2026
Safety

Differences in Typological Alignment in Language Models' Treatment of Differential Argument Marking

DGX agent

arXiv:2602.17653v2 Announce Type: replace Abstract: Recent work has shown that language models (LMs) trained on synthetic corpora can exhibit typological preferences that resemble cross-linguistic reg

safetyarxiv-cs-cl
25 May 2026
Research

Dithering Defense: Adversarial Robustness of Vision Foundation Models via Multi-Level Floyd-Steinberg Dithering

DGX agent

arXiv:2605.23065v1 Announce Type: cross Abstract: Vision foundation models are widely used as frozen backbones across many downstream tasks, making them a single point of failure under adversarial att

researcharxiv-cs-ai
25 May 2026
Applications

GEM-4D: Geometry-Enhanced Video World Models for Robot Manipulation

DGX agent

arXiv:2605.22882v1 Announce Type: new Abstract: Video world models can generate realistic futures from a single instruction, but they often fail to preserve consistent point-level motion over time. As

applicationsarxiv-cs-cv
25 May 2026
Applications

How Human-Like Are Large Language Models? A Register-Aware Linguistic Evaluation Framework

DGX agent

arXiv:2605.23651v1 Announce Type: new Abstract: While factual correctness and task-performance have been in focus of Large Language Model (LLM) research for a long time, the fundamental question of ho

applicationsarxiv-cs-cl
25 May 2026
Tutorials

Learnability-Informed Fine-Tuning of Diffusion Language Models

DGX agent

arXiv:2605.22939v1 Announce Type: new Abstract: We aim to improve the reasoning capabilities of diffusion language models (DLMs). While SFT is a popular post-training recipe for autoregressive models,

tutorialsarxiv-cs-cl
25 May 2026
Tutorials

VAMP-Diff: VampPrior Latent Diffusion for Photoplethysmography Modeling

DGX agent

arXiv:2605.22851v1 Announce Type: cross Abstract: Photoplethysmography (PPG) has become a ubiquitous physiological signal; however, current generative models still struggle to preserve realistic wavef

tutorialsarxiv-cs-lg
25 May 2026
Research

What Does the Server See? Understanding Privacy Leakage from Large Language Models in Split Inference

DGX agent

arXiv:2605.23158v1 Announce Type: cross Abstract: The deployment of large language models (LLMs) on resource-constrained devices remains challenging, spurring interest in split inference, where models

researcharxiv-cs-cl
25 May 2026
Research

A Diffusive Classification Loss for Learning Energy-based Generative Models

DGX agent

arXiv:2601.21025v3 Announce Type: replace-cross Abstract: Score-based generative models have recently achieved remarkable success. While they are usually parameterized by the score, an alternative way

researcharxiv-cs-lg
23 May 2026
Research

Bringing Stability to Diffusion: Decomposing and Reducing Variance of Training Masked Diffusion Models

DGX agent

arXiv:2511.18159v2 Announce Type: replace Abstract: Masked diffusion models (MDMs) are a promising alternative to autoregressive models (ARMs), but they suffer from inherently much higher training var

researcharxiv-cs-lg
23 May 2026
Research

CellFluxRL: Biologically-Constrained Virtual Cell Modeling via Reinforcement Learning

DGX agent

arXiv:2603.21743v4 Announce Type: replace Abstract: Building virtual cells with generative models to simulate cellular behavior in silico is emerging as a promising paradigm for accelerating drug disc

researcharxiv-cs-lg
23 May 2026
Safety

DecepChain: Inducing Deceptive Reasoning in Large Language Models

DGX agent

arXiv:2510.00319v2 Announce Type: replace Abstract: Large Language Models (LLMs) have been demonstrating strong reasoning capability with their chain-of-thoughts (CoT), which are routinely used by hum

safetyarxiv-cs-lg
23 May 2026
Tutorials

MDM-Prime-v2: Binary Encoding and Index Shuffling Enable Scaling of Diffusion Language Models

DGX agent

arXiv:2603.16077v3 Announce Type: replace Abstract: Masked diffusion models (MDM) exhibit superior generalization when learned using a Partial masking scheme (Prime). This approach converts tokens int

tutorialsarxiv-cs-lg
23 May 2026
Research

Soft Bayesian Context Tree Models for Real-Valued Time Series

DGX agent

arXiv:2601.11079v2 Announce Type: replace Abstract: This paper proposes the soft Bayesian context tree model (Soft-BCT), which is a novel BCT model for real-valued time series. The Soft-BCT considers

researcharxiv-cs-lg
23 May 2026
Model Releases

Boiling the Frog: A Multi-Turn Benchmark for Agentic Safety

DGX agent

arXiv:2605.22643v1 Announce Type: new Abstract: Background. Traditional safety benchmarks for language models evaluate generated text: whether a model outputs toxic language, reproduces bias, or follo

model-releasesarxiv-cs-cl
22 May 2026
Safety

Circle-RoPE: Cone-like Decoupled Rotary Positional Embedding for Large Vision-Language Models

DGX agent

arXiv:2505.16416v3 Announce Type: replace Abstract: Rotary Position Embedding (RoPE) is widely adopted in large language models, but when applied to vision-language models (VLMs) it couples text and i

safetyarxiv-cs-cv
22 May 2026
Model Releases

Comparing LLM and Fine-Tuned Model Performance on NVDRS Circumstance Extraction with Varying Prompt Complexity

DGX agent

arXiv:2605.21845v1 Announce Type: new Abstract: Suicide is a leading cause of death in the United States, and understanding the circumstances that precede it requires extracting structured information

model-releasesarxiv-cs-cl
22 May 2026
Applications

GesVLA: Gesture-Aware Vision-Language-Action Model Embedded Representations

DGX agent

arXiv:2605.22812v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have shown strong potential for general-purpose robot manipulation by unifying perception and action. However, exi

applicationsarxiv-cs-cv
22 May 2026
Research

Open Materials 2024 (OMat24) Inorganic Materials Dataset and Models

DGX agent

arXiv:2410.12771v2 Announce Type: replace-cross Abstract: The ability to discover new materials with desirable properties is critical for numerous applications from helping mitigate climate change to

researcharxiv-cs-ai
22 May 2026
Hardware

PALS: Power-Aware LLM Serving for Mixture-of-Experts Models

DGX agent

arXiv:2605.21427v1 Announce Type: new Abstract: Large language model (LLM) inference has become a dominant workload in modern data centers, driving significant GPU utilization and energy consumption.

hardwarearxiv-cs-ai
22 May 2026
Model Releases

SDGBiasBench: Benchmarking and Mitigating Vision--Language Models' Biases in Sustainable Development Goals

DGX agent

arXiv:2605.21919v1 Announce Type: new Abstract: Assessing progress toward the Sustainable Development Goals (SDGs) requires multi-step reasoning over visual cues, contextual knowledge, and development

model-releasesarxiv-cs-cv
22 May 2026
Safety

UniSD: Towards a Unified Self-Distillation Framework for Large Language Models

DGX agent

arXiv:2605.06597v2 Announce Type: replace Abstract: Self-distillation (SD) offers a promising path for adapting large language models (LLMs) without relying on stronger external teachers. However, SD

safetyarxiv-cs-cl
22 May 2026
Research

Deeper Thought, Weaker Aim: Understanding and Mitigating Perceptual Impairment during Reasoning in Multimodal Large Language Models

DGX agent

arXiv:2603.14184v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) often suffer from perceptual impairments under extended reasoning modes, particularly in visual question an

researcharxiv-cs-cv
21 May 2026
Research

Ensemble RL through Classifier Models: Enhancing Risk-Return Trade-offs in Trading Strategies

DGX agent

arXiv:2502.17518v2 Announce Type: replace Abstract: This paper presents a comprehensive study on the use of ensemble Reinforcement Learning (RL) models in financial trading strategies, leveraging clas

researcharxiv-cs-lg
21 May 2026
Safety

EvalMORAAL: Interpretable Chain-of-Thought and LLM-as-Judge Evaluation for Moral Alignment in Large Language Models

DGX agent

arXiv:2510.05942v3 Announce Type: replace Abstract: We present EvalMORAAL, a transparent chain-of-thought (CoT) framework that uses two scoring methods (log-probabilities and direct ratings) plus a mo

safetyarxiv-cs-cl
21 May 2026
Local Ai

Explainable AI: Context-Aware Layer-Wise Integrated Gradients for Explaining Transformer Models

DGX agent

arXiv:2602.16608v2 Announce Type: replace Abstract: Transformer models achieve state-of-the-art performance across domains and tasks, yet their deeply layered representations make their predictions di

local-aiarxiv-cs-cl
21 May 2026
Research

Hybrid Machine Learning Model for Forest Height Estimation from TanDEM-X and Landsat Data

DGX agent

arXiv:2605.20997v1 Announce Type: new Abstract: Integrating machine learning (ML) with physical models (PM) has emerged as a promising way of retrieving geophysical parameters from remote sensing data

researcharxiv-cs-cv
21 May 2026
Applications

Lighting-aware Unified Model for Instance Segmentation

DGX agent

arXiv:2605.20436v1 Announce Type: new Abstract: Foundation models like the Segment Anything Model (SAM) demonstrate impressive zero-shot generalization but frequently degrade under diverse real-world

applicationsarxiv-cs-cv
21 May 2026
Model Releases

MedicalBench: Evaluating Large Language Models Toward Improved Medical Concept Extraction

DGX agent

arXiv:2605.20197v1 Announce Type: new Abstract: Medical concept extraction from electronic health records underpins many downstream applications, yet remains challenging because medically meaningful c

model-releasesarxiv-cs-cl
21 May 2026
Applications

Miller-Index-Based Latent Crystallographic Fracture Plane Reasoning with Vision-Language Models

DGX agent

arXiv:2605.20416v1 Announce Type: new Abstract: We study whether multimodal large language models (MLLMs) can leverage crystallographic plane indices (Miller indices) as a structured latent representa

applicationsarxiv-cs-lg
21 May 2026
Research

Neural Estimation of Pairwise Mutual Information in Masked Discrete Sequence Models

DGX agent

arXiv:2605.20187v1 Announce Type: new Abstract: Understanding dependencies between variables is critical for interpretability and efficient generation in masked diffusion models (MDMs), yet these mode

researcharxiv-cs-lg
21 May 2026
Hardware

PulseCol: Periodically Refreshed Column-Sparse Attention for Accelerating Diffusion Language Models

DGX agent

arXiv:2605.20813v1 Announce Type: new Abstract: Inference in diffusion large language models (dLLMs) is computationally expensive, as full self-attention must be repeatedly executed at each step of th

hardwarearxiv-cs-cl
21 May 2026
Research

Q-ARVD: Quantizing Autoregressive Video Diffusion Models

DGX agent

arXiv:2605.21072v1 Announce Type: new Abstract: Autoregressive video diffusion models (ARVDs) have emerged as a promising architecture for streaming video generation, paving the way for real-time inte

researcharxiv-cs-cv
21 May 2026
Research

RISE: Reliable Improvement in Self-Evolving Vision-Language Models

DGX agent

arXiv:2605.20914v1 Announce Type: new Abstract: Vision-language models (VLMs) have achieved strong multimodal reasoning capabilities, but further improving them still relies heavily on large-scale hum

researcharxiv-cs-cv
21 May 2026
Safety

Synchronization and Turn-Taking in Full-Duplex Speech Dialogue Models

DGX agent

arXiv:2605.20356v1 Announce Type: new Abstract: Full-duplex spoken dialogue models (SDMs) can listen and speak simultaneously, enabling interaction dynamics closer to human conversation than turn-base

safetyarxiv-cs-cl
21 May 2026
Model Releases

TabPFN Extensions for Interpretable Geotechnical Modelling

DGX agent

arXiv:2603.21033v2 Announce Type: replace-cross Abstract: Geotechnical site characterisation relies on sparse, heterogeneous borehole data, where uncertainty quantification and interpretability matter

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Under Pressure: Emotional Framing Induces Measurable Behavioral Shifts and Structured Internal Geometry in Small Language Models

DGX agent

arXiv:2605.20202v1 Announce Type: new Abstract: I study whether emotionally framed evaluation follow-ups change both the behavior and the calm-relative internal representations of small, locally deplo

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Can Large Language Models Reliably Correct Errors in Low-Resource ASR? A Contamination-Aware Case Study on West Frisian

DGX agent

arXiv:2605.19711v1 Announce Type: new Abstract: Automatic speech recognition (ASR) has improved substantially in recent years, yet performance remains limited for low-resource languages. Large languag

model-releasesarxiv-cs-cl
20 May 2026
Safety

CoLD: Counterfactually-Guided Length Debiasing for Process Reward Models in Mathematical Reasoning

DGX agent

arXiv:2507.15698v2 Announce Type: replace-cross Abstract: Process Reward Models (PRMs) play a central role in evaluating and guiding multi-step reasoning in large language models (LLMs), especially fo

safetyarxiv-cs-ai
20 May 2026
Applications

Composition of Memory Experts for Diffusion World Models

DGX agent

arXiv:2605.18813v1 Announce Type: cross Abstract: World models aim to predict plausible futures consistent with past observations, a capability central to planning and decision-making in reinforcement

applicationsarxiv-cs-ai
20 May 2026
← Previous
1…134135136137138…1030
Next →