AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,115
  • Agents7,313
  • Applications5,228
  • Concepts5
  • Hardware1,762
  • Industry6,105
  • Local Ai4,756
  • Model Releases22,759
  • Research19,333
  • Safety12,889
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,115
  • Agents7,313
  • Applications5,228
  • Concepts5
  • Hardware1,762
  • Industry6,105
  • Local Ai4,756
  • Model Releases22,759
  • Research19,333
  • Safety12,889
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlog
85,115Total entries
1Added by human
85,114Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,986 results
Safety

FIRM: Federated In-client Regularized Multi-objective Alignment for Large Language Models

DGX agent

arXiv:2511.16992v3 Announce Type: replace Abstract: Aligning Large Language Models (LLMs) with human values often involves balancing multiple, conflicting objectives such as helpfulness and harmlessne

safetyarxiv-cs-lg
2 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Research

ForesightKV: Optimizing KV Cache Eviction for Reasoning Models by Learning Long-Term Contribution

DGX agent

arXiv:2602.03203v2 Announce Type: replace Abstract: Recently, large language models (LLMs) have shown remarkable reasoning abilities by producing long reasoning traces. However, as the sequence length

researcharxiv-cs-cl
2 Jun 2026
Hardware

FreqLite: A Lightweight Frequency-Decomposed Linear Model with Adaptive Reversible Normalization for Robust Long-Term Time-Series Forecasting

DGX agent

arXiv:2606.01339v1 Announce Type: cross Abstract: Long-term time-series forecasting needs models that are accurate yet efficient enough for commodity hardware. Lightweight linear forecasters are remar

hardwarearxiv-cs-ai
2 Jun 2026
Model Releases

From Unfamiliar to Familiar: Detecting Pre-training Data via Gradient Deviations in Large Language Models

DGX agent

arXiv:2603.04828v2 Announce Type: replace Abstract: Pre-training data detection for LLMs is essential for addressing copyright concerns and mitigating benchmark contamination. Existing methods mainly

model-releasesarxiv-cs-cl
2 Jun 2026
Local Ai

Heterogeneous Decentralized Diffusion Models

DGX agent

arXiv:2603.06741v2 Announce Type: replace-cross Abstract: Training frontier-scale diffusion models often requires substantial computational resources concentrated in tightly-coupled clusters, limiting

local-aiarxiv-cs-ai
2 Jun 2026
Research

Hyperbolic and Evidence-Prioritized Experts for Large Vision-Language Models

DGX agent

arXiv:2606.00275v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) have demonstrated impressive performance on multimodal tasks through scaled architectures and extensive training.

researcharxiv-cs-ai
2 Jun 2026
Research

Improving Visual Grounding in Remote Sensing via Cluster-Guided Refinement and Model Ensemble Voting

DGX agent

arXiv:2606.00556v1 Announce Type: new Abstract: Visual grounding aims to locate image regions that correspond to natural language descriptions and is a key component of interpretable vision systems. I

researcharxiv-cs-cv
2 Jun 2026
Model Releases

KnowledgeBerg: Evaluating Systematic Knowledge Coverage and Compositional Reasoning in Large Language Models

DGX agent

arXiv:2604.17621v2 Announce Type: replace Abstract: Many real-world questions appear deceptively simple yet implicitly demand two capabilities: (i) systematic coverage of a bounded knowledge universe

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Large-scale Uncertainty Quantification for Latent Variable Models Using Subsampling Markov Chain Monte Carlo

DGX agent

arXiv:2606.00309v1 Announce Type: new Abstract: Stochastic gradient Langevin dynamics combined with Gibbs updates (SGLD--Gibbs) provides a highly scalable approach to approximate Bayesian inference in

model-releasesarxiv-cs-lg
2 Jun 2026
Applications

LLMSynthor: Macro-Aligned Micro-Records Synthesis with Large Language Models

DGX agent

arXiv:2505.14752v3 Announce Type: replace Abstract: Macro-aligned micro-records are crucial for credible simulations in social science and urban studies. For example, epidemic models are only reliable

applicationsarxiv-cs-lg
2 Jun 2026
Research

lmfaoooo at SemEval-2026 Task 1: Humor Is an Audience. Preference Modeling for Constrained Humor Generation

DGX agent

arXiv:2606.00022v1 Announce Type: cross Abstract: Humor generation remains difficult not only because producing fluent, novel jokes is hard, but because 'funny' is audience-dependent and supervision i

researcharxiv-cs-ai
2 Jun 2026
Research

M^3 Scaling Law: Optimizing Multi-Epoch, Multi-Lingual, and Multi-Stage Training for Low-Resource Language Models

DGX agent

arXiv:2410.12325v2 Announce Type: replace Abstract: In this paper, we study a fundamental design problem in pretraining Large Language Models (LLMs) for low-resource language regimes. Existing works a

researcharxiv-cs-cl
2 Jun 2026
Safety

Mechanistic Diagnostics of Spatial Lexical Bias in Multimodal Large Language Model Spatial Reasoning

DGX agent

arXiv:2606.01914v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) remain unreliable on spatial multiple-choice questions, and their failures are often attributed to poorly atten

safetyarxiv-cs-cl
2 Jun 2026
Safety

Meta-Black-Box Optimization with Ensemble Surrogate Modeling for Robustness-Accuracy Trade-off within SAEA

DGX agent

arXiv:2606.00862v1 Announce Type: cross Abstract: Surrogate-assisted evolutionary algorithms (SAEAs) have been widely used for expensive black-box optimization problems. However, their reliance on rig

safetyarxiv-cs-lg
2 Jun 2026
Safety

MidSteer: Optimal Affine Framework for Steering Generative Models

DGX agent

arXiv:2605.05220v2 Announce Type: replace-cross Abstract: Steering intermediate representations has emerged as a powerful strategy for controlling generative models, particularly in post-deployment al

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

Model Parallelism With Subnetwork Data Parallelism

DGX agent

arXiv:2507.09029v5 Announce Type: replace-cross Abstract: Pre-training large neural networks at scale imposes heavy memory demands on accelerators and often requires costly communication. We introduce

model-releasesarxiv-cs-ai
2 Jun 2026
Research

Scaling Pre-training to One Hundred Billion Data for Vision Language Models

DGX agent

arXiv:2502.07617v2 Announce Type: replace Abstract: We provide an empirical investigation of the potential of pre-training vision-language models on an unprecedented scale: 100 billion examples. We fi

researcharxiv-cs-cv
2 Jun 2026
Research

ScaRF-SLAM: Scale-Consistent Reconstruction with Feed-Forward Models and Classical Visual SLAM

DGX agent

arXiv:2606.00307v1 Announce Type: new Abstract: Recent works have explored unifying SLAM with geometric foundation models (GFMs). However, directly using GFM predictions for tracking is highly sensiti

researcharxiv-cs-ro
2 Jun 2026
Research

The Entropic Signature of Class Speciation in Diffusion Models

DGX agent

arXiv:2602.09651v2 Announce Type: replace-cross Abstract: Diffusion models do not recover semantic structure uniformly over time. Instead, samples transition from semantic ambiguity to class commitmen

researcharxiv-cs-lg
2 Jun 2026
Safety

THRD: A Training-Free Multi-Turn Defense Framework for Jailbreak Attacks on Large Language Models

DGX agent

arXiv:2606.01738v1 Announce Type: cross Abstract: Multi-turn jailbreak attacks pose a growing threat to LLMs by exploiting conversational dynamics such as gradual escalation and cross-turn coordinatio

safetyarxiv-cs-ai
2 Jun 2026
Safety

Towards Understanding Modality Interaction in Multimodal Language Models via Partial Information Decomposition

DGX agent

arXiv:2606.00959v1 Announce Type: new Abstract: Understanding modality interaction in multimodal large language models (MLLMs) is central to reliable deployment. We introduce Partial Information Decom

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

VLBM: Variational Latent Basis Modeling for OOD Robust Multivariate Time Series Forecasting

DGX agent

arXiv:2606.02138v1 Announce Type: cross Abstract: Out of distribution (OOD) events in multivariate time series forecasting are rare but often dominate real world risk, making average case forecasting

model-releasesarxiv-cs-ai
2 Jun 2026
Research

BadBone: Backdoor Attacks Against Backbone Models in Visual Prompt Learning

DGX agent

arXiv:2605.31246v1 Announce Type: cross Abstract: Prompt learning is a new machine learning paradigm that has attracted ample attention due to its simplicity and proven efficacy. Despite its growing a

researcharxiv-cs-cv
1 Jun 2026
Research

Clustering Guided Domain-Specific Pretrained Foundation Model Very High-Resolution Arctic Remote Sensing

DGX agent

arXiv:2605.30467v1 Announce Type: new Abstract: This study introduces a novel Arctic-focused remote sensing foundation model (RSFM) by combining diversity-aware regional-scale image curation with mask

researcharxiv-cs-cv
1 Jun 2026
Safety

COFT: Counterfactual-Conformal Decoding for Fair Chain-of-Thought Reasoning in Large Language Models

DGX agent

arXiv:2605.30641v1 Announce Type: cross Abstract: Large language models (LLMs) can reveal and amplify societal biases during chain-of-thought (CoT) generation. We present COFT (Chain of Fair Thought),

safetyarxiv-cs-ai
1 Jun 2026
Agents

CoMem: Context Management with A Decoupled Long-Context Model

DGX agent

arXiv:2605.30842v1 Announce Type: new Abstract: Context management enables agentic models to solve long-horizon tasks through iterative summarization of previous interaction histories. However, this p

agentsarxiv-cs-lg
1 Jun 2026
Research

DEM: A Distilled Explanation Model for Interpretable Anomaly Detection in Physiological Sensor Networks

DGX agent

arXiv:2605.31007v1 Announce Type: cross Abstract: Anomaly detection in physiological sensor data from Wireless Body Area Networks (WBANs) can be caused by sensor faults, network disruptions, or missin

researcharxiv-cs-ai
1 Jun 2026
Research

Flow Equivariant World Models: Memory for Partially Observed Dynamic Environments

DGX agent

arXiv:2601.01075v2 Announce Type: replace-cross Abstract: Embodied systems experience the world as 'a symphony of flows': a combination of many continuous streams of sensory input coupled to self-moti

researcharxiv-cs-ai
1 Jun 2026
Tutorials

Guidance for Low-Level Perceptual Editing in Unconditional Diffusion Models

DGX agent

arXiv:2605.31162v1 Announce Type: new Abstract: Unconditional diffusion models offer powerful generative priors, yet steering them toward aesthetically enhanced outputs remains largely unexplored. We

tutorialsarxiv-cs-cv
1 Jun 2026
Research

HiPPO Zoo: Explicit Memory Mechanisms for Interpretable State Space Models

DGX agent

arXiv:2602.21340v2 Announce Type: replace Abstract: Representing the past in a compressed, efficient, and informative manner is a central problem for systems trained on sequential data. The HiPPO fram

researcharxiv-cs-lg
1 Jun 2026
Research

Immuno-VLM: Immunizing Large Vision-Language Models via Generative Semantic Antibodies for Open-World Trustworthiness

DGX agent

arXiv:2605.30745v1 Announce Type: new Abstract: Large Vision-Language Models have achieved unprecedented success in zero-shot recognition by aligning visual features with broad semantic concepts. Howe

researcharxiv-cs-cv
1 Jun 2026
Tutorials

Lumos-Nexus: Efficient Frequency Bridging with Homogeneous Latent Space for Video Unified Models

DGX agent

arXiv:2605.31603v1 Announce Type: cross Abstract: Connector-based video unified models have demonstrated strong capability in instruction-grounded video synthesis, but integrating a large high-fidelit

tutorialsarxiv-cs-ai
1 Jun 2026
Research

Protein Language Model Embeddings Improve Generalization of Implicit Transfer Operators

DGX agent

arXiv:2602.11216v2 Announce Type: replace Abstract: Molecular dynamics (MD) is a central computational tool in physics, chemistry, and biology, enabling quantitative prediction of experimental observa

researcharxiv-cs-lg
1 Jun 2026
Research

Real-time multilingual ASR using rolling buffers and monolingual models [P]

DGX agent

This research discusses a technique for enabling real-time multilingual automatic speech recognition (ASR) on edge devices by dynamically switching between compact monolingual models rather than using

researchr-machinelearning
1 Jun 2026
Research

Semantic Triplet Restoration: A Novel Protocol for Hierarchical Table Understanding in Large Language Models

DGX agent

arXiv:2605.31550v1 Announce Type: new Abstract: Table question answering requires models to recover semantic relations encoded implicitly by two-dimensional layout, merged cells, and hierarchical head

researcharxiv-cs-cl
1 Jun 2026
Research

SLAP: The Semantic Least Action Principle for Variational Video-Language Modeling

DGX agent

arXiv:2605.30750v1 Announce Type: new Abstract: In the era of Large Video-Language Models (LVLMs), the computational necessity of sparse frame sampling creates a fundamental ``temporal gap'', renderin

researcharxiv-cs-cv
1 Jun 2026
Local Ai

YARD: Y-Architecture Register Decoding for Efficient Hallucination Mitigation in Large Vision-Language Models

DGX agent

arXiv:2605.31429v1 Announce Type: new Abstract: Contrastive decoding (CD) seeks to mitigate hallucinations in Large Vision-Language Models (LVLMs) by contrasting the output distributions of a standard

local-aiarxiv-cs-cv
1 Jun 2026
Safety

ActTraitBench: Quantifying the Knowledge-Decision Gap in Large Language Models via Human-Grounded Behavioral Validation

DGX agent

arXiv:2605.29791v1 Announce Type: new Abstract: While Large Language Models (LLMs) can convincingly simulate personas in explicit self-reports, they often deviate in implicit behavioral decisions, rev

safetyarxiv-cs-cl
29 May 2026
Model Releases

AutoSizer: Automatic Sizing of Analog and Mixed-Signal Circuits via Large Language Model (LLM) Agents

DGX agent

arXiv:2602.02849v2 Announce Type: replace Abstract: The design of Analog and Mixed-Signal (AMS) integrated circuits remains heavily reliant on expert knowledge, with transistor sizing a major bottlene

model-releasesarxiv-cs-ai
29 May 2026
Safety

Causal-JEPA: Learning World Models through Object-Level Latent Masking

DGX agent

arXiv:2602.11389v2 Announce Type: replace Abstract: World models require robust relational understanding to support prediction, reasoning, and control. While object-centric representations provide a u

safetyarxiv-cs-ai
29 May 2026
Research

Data filtering methods for training language models

DGX agent

arXiv:2605.29807v1 Announce Type: cross Abstract: Data quality is a critical factor in the effectiveness of machine learning models. Label errors, present even in widely used benchmarks, introduce noi

researcharxiv-cs-ai
29 May 2026
Agents

Dissociative Identity: Language Model Agents Lack Grounding for Reputation Mechanisms

DGX agent

arXiv:2605.30169v1 Announce Type: cross Abstract: As autonomous language model agents proliferate, forming an emerging agentic web with real-world consequences, what credibility signals can you use to

agentsarxiv-cs-ai
29 May 2026
Research

DVSM: Decoder-only View Synthesis Model Done Right

DGX agent

arXiv:2605.29891v1 Announce Type: new Abstract: Recent Large View Synthesis Models (LVSMs) advocate an encoder-decoder architecture that separates reconstruction and rendering into distinct networks.

researcharxiv-cs-cv
29 May 2026
Safety

Emergent Semantic Representations in World Models through Physical Interaction without Linguistic Supervision

DGX agent

arXiv:2605.28865v1 Announce Type: cross Abstract: What does a world model learn from physical exploration, without any linguistic supervision? We argue the answer is organized by a single principle: t

safetyarxiv-cs-ai
29 May 2026
Agents

Estimating the Empowerment of Language Model Agents

DGX agent

arXiv:2509.22504v3 Announce Type: replace Abstract: As language model (LM) agents become increasingly capable and adopted in real-world applications, there is a growing need for scalable evaluation fr

agentsarxiv-cs-ai
29 May 2026
Research

Fingerprinting Inference Systems of Large Language Models

DGX agent

arXiv:2605.29979v1 Announce Type: cross Abstract: The behavior of LLMs does not depend solely on the model itself. Components of the inference system, such as the inference engine, attention backend,

researcharxiv-cs-lg
29 May 2026
Safety

GRPO is Secretly a Process Reward Model

DGX agent

arXiv:2509.21154v4 Announce Type: replace-cross Abstract: Process reward models (PRMs) allow for fine-grained credit assignment in reinforcement learning (RL), and seemingly contrast with outcome rewa

safetyarxiv-cs-ai
29 May 2026
Safety

Harnessing non-adversarial robustness in large language models

DGX agent

arXiv:2605.29816v1 Announce Type: new Abstract: The work presents an approach for addressing the challenge of robustness in Large Language Models (LLMs) to alterations and potential errors caused by s

safetyarxiv-cs-ai
29 May 2026
← Previous
1…197198199200201…1271
Next →