AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,435 results
Safety

MidSteer: Optimal Affine Framework for Steering Generative Models

DGX agent

arXiv:2605.05220v2 Announce Type: replace-cross Abstract: Steering intermediate representations has emerged as a powerful strategy for controlling generative models, particularly in post-deployment al

safetyarxiv-cs-ai
2 Jun 2026
Model Releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Model Parallelism With Subnetwork Data Parallelism

DGX agent

arXiv:2507.09029v5 Announce Type: replace-cross Abstract: Pre-training large neural networks at scale imposes heavy memory demands on accelerators and often requires costly communication. We introduce

model-releasesarxiv-cs-ai
2 Jun 2026
Research

Scaling Pre-training to One Hundred Billion Data for Vision Language Models

DGX agent

arXiv:2502.07617v2 Announce Type: replace Abstract: We provide an empirical investigation of the potential of pre-training vision-language models on an unprecedented scale: 100 billion examples. We fi

researcharxiv-cs-cv
2 Jun 2026
Research

ScaRF-SLAM: Scale-Consistent Reconstruction with Feed-Forward Models and Classical Visual SLAM

DGX agent

arXiv:2606.00307v1 Announce Type: new Abstract: Recent works have explored unifying SLAM with geometric foundation models (GFMs). However, directly using GFM predictions for tracking is highly sensiti

researcharxiv-cs-ro
2 Jun 2026
Research

The Entropic Signature of Class Speciation in Diffusion Models

DGX agent

arXiv:2602.09651v2 Announce Type: replace-cross Abstract: Diffusion models do not recover semantic structure uniformly over time. Instead, samples transition from semantic ambiguity to class commitmen

researcharxiv-cs-lg
2 Jun 2026
Safety

THRD: A Training-Free Multi-Turn Defense Framework for Jailbreak Attacks on Large Language Models

DGX agent

arXiv:2606.01738v1 Announce Type: cross Abstract: Multi-turn jailbreak attacks pose a growing threat to LLMs by exploiting conversational dynamics such as gradual escalation and cross-turn coordinatio

safetyarxiv-cs-ai
2 Jun 2026
Safety

Towards Understanding Modality Interaction in Multimodal Language Models via Partial Information Decomposition

DGX agent

arXiv:2606.00959v1 Announce Type: new Abstract: Understanding modality interaction in multimodal large language models (MLLMs) is central to reliable deployment. We introduce Partial Information Decom

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

VLBM: Variational Latent Basis Modeling for OOD Robust Multivariate Time Series Forecasting

DGX agent

arXiv:2606.02138v1 Announce Type: cross Abstract: Out of distribution (OOD) events in multivariate time series forecasting are rare but often dominate real world risk, making average case forecasting

model-releasesarxiv-cs-ai
2 Jun 2026
Research

BadBone: Backdoor Attacks Against Backbone Models in Visual Prompt Learning

DGX agent

arXiv:2605.31246v1 Announce Type: cross Abstract: Prompt learning is a new machine learning paradigm that has attracted ample attention due to its simplicity and proven efficacy. Despite its growing a

researcharxiv-cs-cv
1 Jun 2026
Research

Clustering Guided Domain-Specific Pretrained Foundation Model Very High-Resolution Arctic Remote Sensing

DGX agent

arXiv:2605.30467v1 Announce Type: new Abstract: This study introduces a novel Arctic-focused remote sensing foundation model (RSFM) by combining diversity-aware regional-scale image curation with mask

researcharxiv-cs-cv
1 Jun 2026
Safety

COFT: Counterfactual-Conformal Decoding for Fair Chain-of-Thought Reasoning in Large Language Models

DGX agent

arXiv:2605.30641v1 Announce Type: cross Abstract: Large language models (LLMs) can reveal and amplify societal biases during chain-of-thought (CoT) generation. We present COFT (Chain of Fair Thought),

safetyarxiv-cs-ai
1 Jun 2026
Agents

CoMem: Context Management with A Decoupled Long-Context Model

DGX agent

arXiv:2605.30842v1 Announce Type: new Abstract: Context management enables agentic models to solve long-horizon tasks through iterative summarization of previous interaction histories. However, this p

agentsarxiv-cs-lg
1 Jun 2026
Research

DEM: A Distilled Explanation Model for Interpretable Anomaly Detection in Physiological Sensor Networks

DGX agent

arXiv:2605.31007v1 Announce Type: cross Abstract: Anomaly detection in physiological sensor data from Wireless Body Area Networks (WBANs) can be caused by sensor faults, network disruptions, or missin

researcharxiv-cs-ai
1 Jun 2026
Research

Flow Equivariant World Models: Memory for Partially Observed Dynamic Environments

DGX agent

arXiv:2601.01075v2 Announce Type: replace-cross Abstract: Embodied systems experience the world as 'a symphony of flows': a combination of many continuous streams of sensory input coupled to self-moti

researcharxiv-cs-ai
1 Jun 2026
Tutorials

Guidance for Low-Level Perceptual Editing in Unconditional Diffusion Models

DGX agent

arXiv:2605.31162v1 Announce Type: new Abstract: Unconditional diffusion models offer powerful generative priors, yet steering them toward aesthetically enhanced outputs remains largely unexplored. We

tutorialsarxiv-cs-cv
1 Jun 2026
Research

HiPPO Zoo: Explicit Memory Mechanisms for Interpretable State Space Models

DGX agent

arXiv:2602.21340v2 Announce Type: replace Abstract: Representing the past in a compressed, efficient, and informative manner is a central problem for systems trained on sequential data. The HiPPO fram

researcharxiv-cs-lg
1 Jun 2026
Research

Immuno-VLM: Immunizing Large Vision-Language Models via Generative Semantic Antibodies for Open-World Trustworthiness

DGX agent

arXiv:2605.30745v1 Announce Type: new Abstract: Large Vision-Language Models have achieved unprecedented success in zero-shot recognition by aligning visual features with broad semantic concepts. Howe

researcharxiv-cs-cv
1 Jun 2026
Tutorials

Lumos-Nexus: Efficient Frequency Bridging with Homogeneous Latent Space for Video Unified Models

DGX agent

arXiv:2605.31603v1 Announce Type: cross Abstract: Connector-based video unified models have demonstrated strong capability in instruction-grounded video synthesis, but integrating a large high-fidelit

tutorialsarxiv-cs-ai
1 Jun 2026
Research

Protein Language Model Embeddings Improve Generalization of Implicit Transfer Operators

DGX agent

arXiv:2602.11216v2 Announce Type: replace Abstract: Molecular dynamics (MD) is a central computational tool in physics, chemistry, and biology, enabling quantitative prediction of experimental observa

researcharxiv-cs-lg
1 Jun 2026
Research

Semantic Triplet Restoration: A Novel Protocol for Hierarchical Table Understanding in Large Language Models

DGX agent

arXiv:2605.31550v1 Announce Type: new Abstract: Table question answering requires models to recover semantic relations encoded implicitly by two-dimensional layout, merged cells, and hierarchical head

researcharxiv-cs-cl
1 Jun 2026
Research

SLAP: The Semantic Least Action Principle for Variational Video-Language Modeling

DGX agent

arXiv:2605.30750v1 Announce Type: new Abstract: In the era of Large Video-Language Models (LVLMs), the computational necessity of sparse frame sampling creates a fundamental ``temporal gap'', renderin

researcharxiv-cs-cv
1 Jun 2026
Local Ai

YARD: Y-Architecture Register Decoding for Efficient Hallucination Mitigation in Large Vision-Language Models

DGX agent

arXiv:2605.31429v1 Announce Type: new Abstract: Contrastive decoding (CD) seeks to mitigate hallucinations in Large Vision-Language Models (LVLMs) by contrasting the output distributions of a standard

local-aiarxiv-cs-cv
1 Jun 2026
Safety

ActTraitBench: Quantifying the Knowledge-Decision Gap in Large Language Models via Human-Grounded Behavioral Validation

DGX agent

arXiv:2605.29791v1 Announce Type: new Abstract: While Large Language Models (LLMs) can convincingly simulate personas in explicit self-reports, they often deviate in implicit behavioral decisions, rev

safetyarxiv-cs-cl
29 May 2026
Model Releases

AutoSizer: Automatic Sizing of Analog and Mixed-Signal Circuits via Large Language Model (LLM) Agents

DGX agent

arXiv:2602.02849v2 Announce Type: replace Abstract: The design of Analog and Mixed-Signal (AMS) integrated circuits remains heavily reliant on expert knowledge, with transistor sizing a major bottlene

model-releasesarxiv-cs-ai
29 May 2026
Safety

Causal-JEPA: Learning World Models through Object-Level Latent Masking

DGX agent

arXiv:2602.11389v2 Announce Type: replace Abstract: World models require robust relational understanding to support prediction, reasoning, and control. While object-centric representations provide a u

safetyarxiv-cs-ai
29 May 2026
Research

Data filtering methods for training language models

DGX agent

arXiv:2605.29807v1 Announce Type: cross Abstract: Data quality is a critical factor in the effectiveness of machine learning models. Label errors, present even in widely used benchmarks, introduce noi

researcharxiv-cs-ai
29 May 2026
Agents

Dissociative Identity: Language Model Agents Lack Grounding for Reputation Mechanisms

DGX agent

arXiv:2605.30169v1 Announce Type: cross Abstract: As autonomous language model agents proliferate, forming an emerging agentic web with real-world consequences, what credibility signals can you use to

agentsarxiv-cs-ai
29 May 2026
Research

DVSM: Decoder-only View Synthesis Model Done Right

DGX agent

arXiv:2605.29891v1 Announce Type: new Abstract: Recent Large View Synthesis Models (LVSMs) advocate an encoder-decoder architecture that separates reconstruction and rendering into distinct networks.

researcharxiv-cs-cv
29 May 2026
Safety

Emergent Semantic Representations in World Models through Physical Interaction without Linguistic Supervision

DGX agent

arXiv:2605.28865v1 Announce Type: cross Abstract: What does a world model learn from physical exploration, without any linguistic supervision? We argue the answer is organized by a single principle: t

safetyarxiv-cs-ai
29 May 2026
Agents

Estimating the Empowerment of Language Model Agents

DGX agent

arXiv:2509.22504v3 Announce Type: replace Abstract: As language model (LM) agents become increasingly capable and adopted in real-world applications, there is a growing need for scalable evaluation fr

agentsarxiv-cs-ai
29 May 2026
Research

Fingerprinting Inference Systems of Large Language Models

DGX agent

arXiv:2605.29979v1 Announce Type: cross Abstract: The behavior of LLMs does not depend solely on the model itself. Components of the inference system, such as the inference engine, attention backend,

researcharxiv-cs-lg
29 May 2026
Safety

GRPO is Secretly a Process Reward Model

DGX agent

arXiv:2509.21154v4 Announce Type: replace-cross Abstract: Process reward models (PRMs) allow for fine-grained credit assignment in reinforcement learning (RL), and seemingly contrast with outcome rewa

safetyarxiv-cs-ai
29 May 2026
Safety

Harnessing non-adversarial robustness in large language models

DGX agent

arXiv:2605.29816v1 Announce Type: new Abstract: The work presents an approach for addressing the challenge of robustness in Large Language Models (LLMs) to alterations and potential errors caused by s

safetyarxiv-cs-ai
29 May 2026
Research

Joint Model and Data Sparsification via the Marginal Likelihood

DGX agent

arXiv:2605.29908v1 Announce Type: cross Abstract: Sparse recovery in linear systems underpins applications from signal processing to high-dimensional regression. Sparse Bayesian Learning, grounded in

researcharxiv-cs-lg
29 May 2026
Tutorials

Large Depth Completion Model from Sparse Observations

DGX agent

arXiv:2605.30115v1 Announce Type: new Abstract: This work presents the Large Depth Completion Model (LDCM), a simple, effective, and robust framework for single-view metric depth estimation with spars

tutorialsarxiv-cs-cv
29 May 2026
Model Releases

Long-Context Modeling with Dynamic Hierarchical Sparse Attention for Memory-Constrained LLM Inference

DGX agent

arXiv:2510.24606v2 Announce Type: replace Abstract: The quadratic cost of attention limits the scalability of long-context LLMs, especially under limited hardware memory budgets. While attention is of

model-releasesarxiv-cs-cl
29 May 2026
Research

Micro-Macro Retrieval: Reducing Long-Form Hallucination in Large Language Models

DGX agent

arXiv:2605.28828v1 Announce Type: cross Abstract: Large Language Models (LLMs) achieve impressive performance across many tasks but remain prone to hallucination, especially in long-form generation wh

researcharxiv-cs-ai
29 May 2026
Safety

Position: Stop Chasing the C-index when Evaluating Survival Analysis Models

DGX agent

arXiv:2506.02075v2 Announce Type: replace-cross Abstract: The current state of evaluation in survival analysis is plagued by the persistent use of evaluation metrics in ways that are misaligned with t

safetyarxiv-cs-lg
29 May 2026
Research

Reliable Reasoning with Large Language Models via Preference-Based Maximum Satisfiability

DGX agent

arXiv:2605.29687v1 Announce Type: new Abstract: Large Language Models (LLMs) excel at understanding natural language but struggle with optimisation tasks involving multiple constraints and user-define

researcharxiv-cs-ai
29 May 2026
Research

Riemannian AmbientFlow: Towards Simultaneous Manifold Learning and Generative Modeling from Corrupted Data

DGX agent

arXiv:2601.18728v2 Announce Type: replace Abstract: Modern generative modeling methods have demonstrated strong performance in learning complex data distributions from clean samples. In many scientifi

researcharxiv-cs-lg
29 May 2026
Research

SM2ITH: Safe Mobile Manipulation with Interactive Human Prediction via Task-Hierarchical Bilevel Model Predictive Control

DGX agent

arXiv:2511.17798v2 Announce Type: replace Abstract: Mobile manipulators are designed to perform complex sequences of navigation and manipulation tasks in human-centered environments. While recent opti

researcharxiv-cs-ro
29 May 2026
Applications

Specialty-Specific Medical Language Model for Immune-Mediated Diseases

DGX agent

arXiv:2605.28838v1 Announce Type: cross Abstract: Extracting detailed clinical information from free-text medical narratives remains a practical challenge for researchers and healthcare systems. Termi

applicationsarxiv-cs-ai
29 May 2026
Tutorials

SRUG: Shadow-Guided Relightable Urban Scene with Generation Model

DGX agent

arXiv:2605.24700v2 Announce Type: replace Abstract: Creating relightable urban scenes from images or videos is widely useful but highly ill-posed. Urban environments are typically unbounded and extend

tutorialsarxiv-cs-cv
29 May 2026
Model Releases

TaxDistill: Improving Metagenomic Taxonomic Annotation via Distilled Genomic Foundation Models

DGX agent

arXiv:2605.28868v1 Announce Type: cross Abstract: Metagenomic taxonomic annotation aims to identify the microbial origins of DNA fragments in environmental samples. Traditional methods that rely on se

model-releasesarxiv-cs-ai
29 May 2026
Safety

VLA-Trace: Diagnosing Vision-Language-Action Models through Representation and Behavior Tracing

DGX agent

arXiv:2605.30117v1 Announce Type: new Abstract: Understanding how Vision-Language-Action (VLA) models transform multimodal knowledge into embodied control remains an open challenge. We present VLA-Tra

safetyarxiv-cs-ai
29 May 2026
Model Releases

Are Large Pre-trained Vision Language Models Effective Construction Safety Inspectors?

DGX agent

arXiv:2508.11011v2 Announce Type: replace Abstract: Construction safety inspections typically involve a human inspector identifying safety concerns on-site. With the rise of powerful Vision Language M

model-releasesarxiv-cs-cv
28 May 2026
Safety

Auditable Decision Models with Learned Abstention and Real-Time Steering

DGX agent

arXiv:2605.27768v1 Announce Type: new Abstract: Production AI systems often operate with incomplete, conflicting, or insufficient evidence. Forced classifiers collapse such cases into action labels, w

safetyarxiv-cs-ai
28 May 2026
Model Releases

Can Large Language Models Handle Discourse Particles? A Case Study of Colloquial Malay

DGX agent

arXiv:2605.28782v1 Announce Type: new Abstract: Discourse particles, such as extit{well} and extit{kind of}, are crucial components that enable LLMs to ``speak'' more like humans. They are used to con

model-releasesarxiv-cs-cl
28 May 2026
← Previous
1…157158159160161…1030
Next →