AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,565 results
2 Jun 2026

FIRM: Federated In-client Regularized Multi-objective Alignment for Large Language Models

SafetyDGX agent

arXiv:2511.16992v3 Announce Type: replace Abstract: Aligning Large Language Models (LLMs) with human values often involves balancing multiple, conflicting objectives such as helpfulness and harmlessne

ForesightKV: Optimizing KV Cache Eviction for Reasoning Models by Learning Long-Term Contribution

ResearchDGX agent

arXiv:2602.03203v2 Announce Type: replace Abstract: Recently, large language models (LLMs) have shown remarkable reasoning abilities by producing long reasoning traces. However, as the sequence length

FreqLite: A Lightweight Frequency-Decomposed Linear Model with Adaptive Reversible Normalization for Robust Long-Term Time-Series Forecasting

HardwareDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.01339v1 Announce Type: cross Abstract: Long-term time-series forecasting needs models that are accurate yet efficient enough for commodity hardware. Lightweight linear forecasters are remar

From Unfamiliar to Familiar: Detecting Pre-training Data via Gradient Deviations in Large Language Models

Model ReleasesDGX agent

arXiv:2603.04828v2 Announce Type: replace Abstract: Pre-training data detection for LLMs is essential for addressing copyright concerns and mitigating benchmark contamination. Existing methods mainly

Heterogeneous Decentralized Diffusion Models

Local AiDGX agent

arXiv:2603.06741v2 Announce Type: replace-cross Abstract: Training frontier-scale diffusion models often requires substantial computational resources concentrated in tightly-coupled clusters, limiting

Hyperbolic and Evidence-Prioritized Experts for Large Vision-Language Models

ResearchDGX agent

arXiv:2606.00275v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) have demonstrated impressive performance on multimodal tasks through scaled architectures and extensive training.

Improving Visual Grounding in Remote Sensing via Cluster-Guided Refinement and Model Ensemble Voting

ResearchDGX agent

arXiv:2606.00556v1 Announce Type: new Abstract: Visual grounding aims to locate image regions that correspond to natural language descriptions and is a key component of interpretable vision systems. I

KnowledgeBerg: Evaluating Systematic Knowledge Coverage and Compositional Reasoning in Large Language Models

Model ReleasesDGX agent

arXiv:2604.17621v2 Announce Type: replace Abstract: Many real-world questions appear deceptively simple yet implicitly demand two capabilities: (i) systematic coverage of a bounded knowledge universe

Large-scale Uncertainty Quantification for Latent Variable Models Using Subsampling Markov Chain Monte Carlo

Model ReleasesDGX agent

arXiv:2606.00309v1 Announce Type: new Abstract: Stochastic gradient Langevin dynamics combined with Gibbs updates (SGLD--Gibbs) provides a highly scalable approach to approximate Bayesian inference in

LLMSynthor: Macro-Aligned Micro-Records Synthesis with Large Language Models

ApplicationsDGX agent

arXiv:2505.14752v3 Announce Type: replace Abstract: Macro-aligned micro-records are crucial for credible simulations in social science and urban studies. For example, epidemic models are only reliable

lmfaoooo at SemEval-2026 Task 1: Humor Is an Audience. Preference Modeling for Constrained Humor Generation

ResearchDGX agent

arXiv:2606.00022v1 Announce Type: cross Abstract: Humor generation remains difficult not only because producing fluent, novel jokes is hard, but because 'funny' is audience-dependent and supervision i

M^3 Scaling Law: Optimizing Multi-Epoch, Multi-Lingual, and Multi-Stage Training for Low-Resource Language Models

ResearchDGX agent

arXiv:2410.12325v2 Announce Type: replace Abstract: In this paper, we study a fundamental design problem in pretraining Large Language Models (LLMs) for low-resource language regimes. Existing works a

Mechanistic Diagnostics of Spatial Lexical Bias in Multimodal Large Language Model Spatial Reasoning

SafetyDGX agent

arXiv:2606.01914v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) remain unreliable on spatial multiple-choice questions, and their failures are often attributed to poorly atten

Meta-Black-Box Optimization with Ensemble Surrogate Modeling for Robustness-Accuracy Trade-off within SAEA

SafetyDGX agent

arXiv:2606.00862v1 Announce Type: cross Abstract: Surrogate-assisted evolutionary algorithms (SAEAs) have been widely used for expensive black-box optimization problems. However, their reliance on rig

MidSteer: Optimal Affine Framework for Steering Generative Models

SafetyDGX agent

arXiv:2605.05220v2 Announce Type: replace-cross Abstract: Steering intermediate representations has emerged as a powerful strategy for controlling generative models, particularly in post-deployment al

Model Parallelism With Subnetwork Data Parallelism

Model ReleasesDGX agent

arXiv:2507.09029v5 Announce Type: replace-cross Abstract: Pre-training large neural networks at scale imposes heavy memory demands on accelerators and often requires costly communication. We introduce

Scaling Pre-training to One Hundred Billion Data for Vision Language Models

ResearchDGX agent

arXiv:2502.07617v2 Announce Type: replace Abstract: We provide an empirical investigation of the potential of pre-training vision-language models on an unprecedented scale: 100 billion examples. We fi

ScaRF-SLAM: Scale-Consistent Reconstruction with Feed-Forward Models and Classical Visual SLAM

ResearchDGX agent

arXiv:2606.00307v1 Announce Type: new Abstract: Recent works have explored unifying SLAM with geometric foundation models (GFMs). However, directly using GFM predictions for tracking is highly sensiti

The Entropic Signature of Class Speciation in Diffusion Models

ResearchDGX agent

arXiv:2602.09651v2 Announce Type: replace-cross Abstract: Diffusion models do not recover semantic structure uniformly over time. Instead, samples transition from semantic ambiguity to class commitmen

THRD: A Training-Free Multi-Turn Defense Framework for Jailbreak Attacks on Large Language Models

SafetyDGX agent

arXiv:2606.01738v1 Announce Type: cross Abstract: Multi-turn jailbreak attacks pose a growing threat to LLMs by exploiting conversational dynamics such as gradual escalation and cross-turn coordinatio

Towards Understanding Modality Interaction in Multimodal Language Models via Partial Information Decomposition

SafetyDGX agent

arXiv:2606.00959v1 Announce Type: new Abstract: Understanding modality interaction in multimodal large language models (MLLMs) is central to reliable deployment. We introduce Partial Information Decom

VLBM: Variational Latent Basis Modeling for OOD Robust Multivariate Time Series Forecasting

Model ReleasesDGX agent

arXiv:2606.02138v1 Announce Type: cross Abstract: Out of distribution (OOD) events in multivariate time series forecasting are rare but often dominate real world risk, making average case forecasting

1 Jun 2026

BadBone: Backdoor Attacks Against Backbone Models in Visual Prompt Learning

ResearchDGX agent

arXiv:2605.31246v1 Announce Type: cross Abstract: Prompt learning is a new machine learning paradigm that has attracted ample attention due to its simplicity and proven efficacy. Despite its growing a

Clustering Guided Domain-Specific Pretrained Foundation Model Very High-Resolution Arctic Remote Sensing

ResearchDGX agent

arXiv:2605.30467v1 Announce Type: new Abstract: This study introduces a novel Arctic-focused remote sensing foundation model (RSFM) by combining diversity-aware regional-scale image curation with mask

COFT: Counterfactual-Conformal Decoding for Fair Chain-of-Thought Reasoning in Large Language Models

SafetyDGX agent

arXiv:2605.30641v1 Announce Type: cross Abstract: Large language models (LLMs) can reveal and amplify societal biases during chain-of-thought (CoT) generation. We present COFT (Chain of Fair Thought),

CoMem: Context Management with A Decoupled Long-Context Model

AgentsDGX agent

arXiv:2605.30842v1 Announce Type: new Abstract: Context management enables agentic models to solve long-horizon tasks through iterative summarization of previous interaction histories. However, this p

DEM: A Distilled Explanation Model for Interpretable Anomaly Detection in Physiological Sensor Networks

ResearchDGX agent

arXiv:2605.31007v1 Announce Type: cross Abstract: Anomaly detection in physiological sensor data from Wireless Body Area Networks (WBANs) can be caused by sensor faults, network disruptions, or missin

Flow Equivariant World Models: Memory for Partially Observed Dynamic Environments

ResearchDGX agent

arXiv:2601.01075v2 Announce Type: replace-cross Abstract: Embodied systems experience the world as 'a symphony of flows': a combination of many continuous streams of sensory input coupled to self-moti

Guidance for Low-Level Perceptual Editing in Unconditional Diffusion Models

TutorialsDGX agent

arXiv:2605.31162v1 Announce Type: new Abstract: Unconditional diffusion models offer powerful generative priors, yet steering them toward aesthetically enhanced outputs remains largely unexplored. We

HiPPO Zoo: Explicit Memory Mechanisms for Interpretable State Space Models

ResearchDGX agent

arXiv:2602.21340v2 Announce Type: replace Abstract: Representing the past in a compressed, efficient, and informative manner is a central problem for systems trained on sequential data. The HiPPO fram

Immuno-VLM: Immunizing Large Vision-Language Models via Generative Semantic Antibodies for Open-World Trustworthiness

ResearchDGX agent

arXiv:2605.30745v1 Announce Type: new Abstract: Large Vision-Language Models have achieved unprecedented success in zero-shot recognition by aligning visual features with broad semantic concepts. Howe

Lumos-Nexus: Efficient Frequency Bridging with Homogeneous Latent Space for Video Unified Models

TutorialsDGX agent

arXiv:2605.31603v1 Announce Type: cross Abstract: Connector-based video unified models have demonstrated strong capability in instruction-grounded video synthesis, but integrating a large high-fidelit

Protein Language Model Embeddings Improve Generalization of Implicit Transfer Operators

ResearchDGX agent

arXiv:2602.11216v2 Announce Type: replace Abstract: Molecular dynamics (MD) is a central computational tool in physics, chemistry, and biology, enabling quantitative prediction of experimental observa

Real-time multilingual ASR using rolling buffers and monolingual models [P]

ResearchDGX agent

This research discusses a technique for enabling real-time multilingual automatic speech recognition (ASR) on edge devices by dynamically switching between compact monolingual models rather than using

Semantic Triplet Restoration: A Novel Protocol for Hierarchical Table Understanding in Large Language Models

ResearchDGX agent

arXiv:2605.31550v1 Announce Type: new Abstract: Table question answering requires models to recover semantic relations encoded implicitly by two-dimensional layout, merged cells, and hierarchical head

SLAP: The Semantic Least Action Principle for Variational Video-Language Modeling

ResearchDGX agent

arXiv:2605.30750v1 Announce Type: new Abstract: In the era of Large Video-Language Models (LVLMs), the computational necessity of sparse frame sampling creates a fundamental ``temporal gap'', renderin

YARD: Y-Architecture Register Decoding for Efficient Hallucination Mitigation in Large Vision-Language Models

Local AiDGX agent

arXiv:2605.31429v1 Announce Type: new Abstract: Contrastive decoding (CD) seeks to mitigate hallucinations in Large Vision-Language Models (LVLMs) by contrasting the output distributions of a standard

29 May 2026

ActTraitBench: Quantifying the Knowledge-Decision Gap in Large Language Models via Human-Grounded Behavioral Validation

SafetyDGX agent

arXiv:2605.29791v1 Announce Type: new Abstract: While Large Language Models (LLMs) can convincingly simulate personas in explicit self-reports, they often deviate in implicit behavioral decisions, rev

AutoSizer: Automatic Sizing of Analog and Mixed-Signal Circuits via Large Language Model (LLM) Agents

Model ReleasesDGX agent

arXiv:2602.02849v2 Announce Type: replace Abstract: The design of Analog and Mixed-Signal (AMS) integrated circuits remains heavily reliant on expert knowledge, with transistor sizing a major bottlene

Causal-JEPA: Learning World Models through Object-Level Latent Masking

SafetyDGX agent

arXiv:2602.11389v2 Announce Type: replace Abstract: World models require robust relational understanding to support prediction, reasoning, and control. While object-centric representations provide a u

Data filtering methods for training language models

ResearchDGX agent

arXiv:2605.29807v1 Announce Type: cross Abstract: Data quality is a critical factor in the effectiveness of machine learning models. Label errors, present even in widely used benchmarks, introduce noi

Dissociative Identity: Language Model Agents Lack Grounding for Reputation Mechanisms

AgentsDGX agent

arXiv:2605.30169v1 Announce Type: cross Abstract: As autonomous language model agents proliferate, forming an emerging agentic web with real-world consequences, what credibility signals can you use to

DVSM: Decoder-only View Synthesis Model Done Right

ResearchDGX agent

arXiv:2605.29891v1 Announce Type: new Abstract: Recent Large View Synthesis Models (LVSMs) advocate an encoder-decoder architecture that separates reconstruction and rendering into distinct networks.

Emergent Semantic Representations in World Models through Physical Interaction without Linguistic Supervision

SafetyDGX agent

arXiv:2605.28865v1 Announce Type: cross Abstract: What does a world model learn from physical exploration, without any linguistic supervision? We argue the answer is organized by a single principle: t

Estimating the Empowerment of Language Model Agents

AgentsDGX agent

arXiv:2509.22504v3 Announce Type: replace Abstract: As language model (LM) agents become increasingly capable and adopted in real-world applications, there is a growing need for scalable evaluation fr

Fingerprinting Inference Systems of Large Language Models

ResearchDGX agent

arXiv:2605.29979v1 Announce Type: cross Abstract: The behavior of LLMs does not depend solely on the model itself. Components of the inference system, such as the inference engine, attention backend,

GRPO is Secretly a Process Reward Model

SafetyDGX agent

arXiv:2509.21154v4 Announce Type: replace-cross Abstract: Process reward models (PRMs) allow for fine-grained credit assignment in reinforcement learning (RL), and seemingly contrast with outcome rewa

Harnessing non-adversarial robustness in large language models

SafetyDGX agent

arXiv:2605.29816v1 Announce Type: new Abstract: The work presents an approach for addressing the challenge of robustness in Large Language Models (LLMs) to alterations and potential errors caused by s

Joint Model and Data Sparsification via the Marginal Likelihood

ResearchDGX agent

arXiv:2605.29908v1 Announce Type: cross Abstract: Sparse recovery in linear systems underpins applications from signal processing to high-dimensional regression. Sparse Bayesian Learning, grounded in

Large Depth Completion Model from Sparse Observations

TutorialsDGX agent

arXiv:2605.30115v1 Announce Type: new Abstract: This work presents the Large Depth Completion Model (LDCM), a simple, effective, and robust framework for single-view metric depth estimation with spars

Long-Context Modeling with Dynamic Hierarchical Sparse Attention for Memory-Constrained LLM Inference

Model ReleasesDGX agent

arXiv:2510.24606v2 Announce Type: replace Abstract: The quadratic cost of attention limits the scalability of long-context LLMs, especially under limited hardware memory budgets. While attention is of

Micro-Macro Retrieval: Reducing Long-Form Hallucination in Large Language Models

ResearchDGX agent

arXiv:2605.28828v1 Announce Type: cross Abstract: Large Language Models (LLMs) achieve impressive performance across many tasks but remain prone to hallucination, especially in long-form generation wh

Position: Stop Chasing the C-index when Evaluating Survival Analysis Models

SafetyDGX agent

arXiv:2506.02075v2 Announce Type: replace-cross Abstract: The current state of evaluation in survival analysis is plagued by the persistent use of evaluation metrics in ways that are misaligned with t

Reliable Reasoning with Large Language Models via Preference-Based Maximum Satisfiability

ResearchDGX agent

arXiv:2605.29687v1 Announce Type: new Abstract: Large Language Models (LLMs) excel at understanding natural language but struggle with optimisation tasks involving multiple constraints and user-define

Riemannian AmbientFlow: Towards Simultaneous Manifold Learning and Generative Modeling from Corrupted Data

ResearchDGX agent

arXiv:2601.18728v2 Announce Type: replace Abstract: Modern generative modeling methods have demonstrated strong performance in learning complex data distributions from clean samples. In many scientifi

SM2ITH: Safe Mobile Manipulation with Interactive Human Prediction via Task-Hierarchical Bilevel Model Predictive Control

ResearchDGX agent

arXiv:2511.17798v2 Announce Type: replace Abstract: Mobile manipulators are designed to perform complex sequences of navigation and manipulation tasks in human-centered environments. While recent opti

Specialty-Specific Medical Language Model for Immune-Mediated Diseases

ApplicationsDGX agent

arXiv:2605.28838v1 Announce Type: cross Abstract: Extracting detailed clinical information from free-text medical narratives remains a practical challenge for researchers and healthcare systems. Termi

SRUG: Shadow-Guided Relightable Urban Scene with Generation Model

TutorialsDGX agent

arXiv:2605.24700v2 Announce Type: replace Abstract: Creating relightable urban scenes from images or videos is widely useful but highly ill-posed. Urban environments are typically unbounded and extend

TaxDistill: Improving Metagenomic Taxonomic Annotation via Distilled Genomic Foundation Models

Model ReleasesDGX agent

arXiv:2605.28868v1 Announce Type: cross Abstract: Metagenomic taxonomic annotation aims to identify the microbial origins of DNA fragments in environmental samples. Traditional methods that rely on se

VLA-Trace: Diagnosing Vision-Language-Action Models through Representation and Behavior Tracing

SafetyDGX agent

arXiv:2605.30117v1 Announce Type: new Abstract: Understanding how Vision-Language-Action (VLA) models transform multimodal knowledge into embodied control remains an open challenge. We present VLA-Tra

← Previous
1…156157158159160…1010
Next →