AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
85,136Total entries
1Added by human
85,135Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,811 results
Research

Dual-Pathway Circuits of Object Hallucination in Vision-Language Models

DGX agent

arXiv:2605.13156v1 Announce Type: new Abstract: Vision-language models (VLMs) have demonstrated remarkable capabilities in bridging visual perception and natural language understanding, enabling a wid

researcharxiv-cs-cv
14 May 2026
Research
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Generative Modeling by Minimizing the Wasserstein-2 Loss

DGX agent

arXiv:2406.13619v4 Announce Type: replace-cross Abstract: This paper develops a generative model by minimizing the second-order Wasserstein loss (the W_2 loss) through a distribution-dependent ordinar

researcharxiv-cs-lg
14 May 2026
Local Ai

GRIP-VLM: Group-Relative Importance Pruning for Efficient Vision-Language Models

DGX agent

arXiv:2605.13375v1 Announce Type: cross Abstract: In Vision-Language Models (VLMs), processing a massive number of visual tokens incurs prohibitive computational overhead. While recent training-aware

local-aiarxiv-cs-ai
14 May 2026
Local Ai

Learning to See What You Need: Gaze Attention for Multimodal Large Language Models

DGX agent

arXiv:2605.13080v1 Announce Type: new Abstract: When humans describe a visual scene, they do not process the entire image uniformly; instead, they selectively fixate on regions relevant to their inten

local-aiarxiv-cs-cv
14 May 2026
Applications

MILM: Large Language Models for Multimodal Irregular Time Series with Informative Sampling

DGX agent

arXiv:2605.13711v1 Announce Type: new Abstract: Multimodal irregular time series (MITS) consist of asynchronous and irregularly sampled observations from heterogeneous numerical and textual channels.

applicationsarxiv-cs-lg
14 May 2026
Research

Multi-Modal World Model for Physical Robot Interactions: Simultaneous Visual and Tactile Predictions for Enhanced Accuracy

DGX agent

arXiv:2304.11193v2 Announce Type: replace-cross Abstract: Predicting the outcomes of robotic actions, often referred to as learning a world model, in complex environments remains a fundamental challen

researcharxiv-cs-ai
14 May 2026
Research

Multitask Multimodal Fusion with Tabular Foundation Models for Peak and Durability Prediction of Pertussis Booster Response

DGX agent

arXiv:2605.12852v1 Announce Type: new Abstract: Pertussis booster vaccination produces immune responses that vary widely across individuals in both peak magnitude and long-term durability. These two p

researcharxiv-cs-lg
14 May 2026
Research

On the Limits of Latent Reuse in Diffusion Models

DGX agent

arXiv:2605.13448v1 Announce Type: cross Abstract: Diffusion models are often trained in low-dimensional latent spaces, which are then reused for related but shifted datasets. In this work, we study wh

researcharxiv-cs-lg
14 May 2026
Safety

Pretraining Language Models with Subword Regularization: An Empirical Study of BPE Dropout in Low-Resource NLP

DGX agent

arXiv:2605.13436v1 Announce Type: cross Abstract: Subword regularization methods such as BPE dropout are typically applied only during fine-tuning, while pretraining is usually done with deterministic

safetyarxiv-cs-lg
14 May 2026
Research

SceneGraphVLM: Dynamic Scene Graph Generation from Video with Vision-Language Models

DGX agent

arXiv:2605.13667v1 Announce Type: new Abstract: Scene graph generation provides a compact structured representation for visual perception, but accurate and fast graph prediction from images and videos

researcharxiv-cs-cv
14 May 2026
Research

TimelineReasoner: Advancing Timeline Summarization with Large Reasoning Models

DGX agent

arXiv:2605.12518v1 Announce Type: cross Abstract: The proliferation of online news poses a challenge to extracting structured timelines from unstructured content. While recent studies have shown that

researcharxiv-cs-ai
14 May 2026
Safety

What to Ignore, What to React: Visually Robust RL Fine-Tuning of VLA Models

DGX agent

arXiv:2605.13105v1 Announce Type: new Abstract: Reinforcement learning (RL) fine-tuning has shown promise for Vision-Language-Action (VLA) models in robotic manipulation, but deployment-time visual sh

safetyarxiv-cs-ro
14 May 2026
Research

When to Think Fast and Slow? AMOR: Adaptive Entropy Gate for Hybrid Models

DGX agent

arXiv:2602.13215v2 Announce Type: replace Abstract: Recurrent-attention hybrids aim to combine the efficiency of recurrence with the expressivity of attention, but existing approaches typically apply

researcharxiv-cs-ai
14 May 2026
Safety

A Survey of On-Policy Distillation for Large Language Models

DGX agent

arXiv:2604.00626v2 Announce Type: replace-cross Abstract: As Large Language Models (LLMs) continue to grow in both capability and cost, transferring frontier capabilities into smaller, deployable stud

safetyarxiv-cs-cl
13 May 2026
Safety

Cluster-Aware Neural Collapse Prompt Tuning for Long-Tailed Generalization of Vision-Language Models

DGX agent

arXiv:2605.11939v1 Announce Type: new Abstract: Prompt learning has emerged as an efficient alternative to fine-tuning pre-trained vision-language models (VLMs). Despite its promise, current methods s

safetyarxiv-cs-cv
13 May 2026
Safety

Diffusion-State Policy Optimization for Masked Diffusion Language Models

DGX agent

arXiv:2602.06462v3 Announce Type: replace Abstract: Masked diffusion language models generate text through iterative masked-token filling, but terminal-only rewards on final completions provide coarse

safetyarxiv-cs-cl
13 May 2026
Research

Do Language Models Encode Knowledge of Linguistic Constraint Violations?

DGX agent

arXiv:2605.12055v1 Announce Type: new Abstract: Large Language Models (LLMs) achieve strong linguistic performance, yet their internal mechanisms for producing these predictions remain unclear. We inv

researcharxiv-cs-cl
13 May 2026
Research

DP-{lambda}CGD: Efficient Noise Correlation for Differentially Private Model Training

DGX agent

arXiv:2601.22334v2 Announce Type: replace Abstract: Differentially private stochastic gradient descent (DP-SGD) is the gold standard for training machine learning models with formal differential priva

researcharxiv-cs-lg
13 May 2026
Applications

Efficient LLM-based Advertising via Model Compression and Parallel Verification

DGX agent

arXiv:2605.11582v1 Announce Type: new Abstract: Large language models (LLMs) have shown remarkable potential in advertising scenarios such as ad creative generation and targeted advertising. However,

applicationsarxiv-cs-cl
13 May 2026
Safety

Enhancing Target-Guided Proactive Dialogue Systems via Conversational Scenario Modeling and Intent-Keyword Bridging

DGX agent

arXiv:2605.11964v1 Announce Type: new Abstract: A target-guided proactive dialogue system aims to steer conversations proactively toward pre-defined targets, such as designated keywords or specific to

safetyarxiv-cs-cl
13 May 2026
Applications

From Token to Token Pair: Efficient Prompt Compression for Large Language Models in Clinical Prediction

DGX agent

arXiv:2605.11774v1 Announce Type: new Abstract: By processing electronic health records (EHRs) as natural language sequences, large language models (LLMs) have shown potential in clinical prediction t

applicationsarxiv-cs-cl
13 May 2026
Safety

Gradient-Free Noise Optimization for Reward Alignment in Generative Models

DGX agent

arXiv:2605.11347v1 Announce Type: cross Abstract: Existing reward alignment methods for diffusion and flow models rely on multi-step stochastic trajectories, making them difficult to extend to determi

safetyarxiv-cs-cv
13 May 2026
Agents

Joint Learning of Hierarchical Neural Options and Abstract World Model

DGX agent

arXiv:2602.02799v2 Announce Type: replace Abstract: Building agents that can perform new skills by composing existing skills is a long-standing goal of AI agent research. Towards this end, we investig

agentsarxiv-cs-lg
13 May 2026
Safety

Model-based Bootstrap of Controlled Markov Chains

DGX agent

arXiv:2605.12410v1 Announce Type: cross Abstract: We propose and analyze a model-based bootstrap for transition kernels in finite controlled Markov chains (CMCs) with possibly nonstationary or history

safetyarxiv-cs-lg
13 May 2026
Applications

Model-Level GNN Explanations via Rule-to-Graph Readout for Logit Reconstruction

DGX agent

arXiv:2503.09051v2 Announce Type: replace Abstract: We propose a novel model-level GNN explanation framework that shifts the explanation target from class-wise rule extraction to rule-based logit reco

applicationsarxiv-cs-lg
13 May 2026
Research

OTT-Vid: Optimal Transport Temporal Token Compression for Video Large Language Models

DGX agent

arXiv:2605.11803v1 Announce Type: new Abstract: As Video Large Language Models (Video-LLMs) scale to longer and more complex videos, their inference cost grows rapidly due to the large volume of visua

researcharxiv-cs-cv
13 May 2026
Local Ai

Rank Is Not Capacity: Spectral Occupancy for Latent Graph Models

DGX agent

arXiv:2605.11142v1 Announce Type: new Abstract: Graph representation learning has become a standard approach for analyzing networked data, with latent embeddings widely used for link prediction, commu

local-aiarxiv-cs-lg
13 May 2026
Safety

Red-Teaming Text-to-Image Models via In-Context Experience Replay and Semantic-Preserving Prompt Rewriting

DGX agent

arXiv:2411.16769v3 Announce Type: replace-cross Abstract: Understanding the capabilities of text-to-image (T2I) models in harmful content generation is essential to safety and compliance. However, hum

safetyarxiv-cs-cl
13 May 2026
Safety

Simpson's Paradox in Behavioral Curves: How Aggregation Distorts Parametric Models of User Dynamics

DGX agent

arXiv:2605.11017v1 Announce Type: new Abstract: Behavioral curve modeling -- fitting parametric functions to engagement-versus-exposure data -- is standard practice in recommendation, advertising, and

safetyarxiv-cs-lg
13 May 2026
Safety

The Evaluation Differential: When Frontier AI Models Recognise They Are Being Tested

DGX agent

arXiv:2605.11496v1 Announce Type: cross Abstract: Recent published evidence from frontier laboratories shows that contemporary AI models can recognise evaluation contexts, latently represent them, and

safetyarxiv-cs-lg
13 May 2026
Tutorials

APCD: Adaptive Path-Contrastive Decoding for Reliable Large Language Model Generation

DGX agent

arXiv:2605.09492v1 Announce Type: cross Abstract: Large language models (LLMs) often suffer from hallucinations due to error accumulation in autoregressive decoding, where suboptimal early token choic

tutorialsarxiv-cs-ai
12 May 2026
Research

ATAAT: Adaptive Threat-Aware Adversarial Tuning Framework against Backdoor Attacks on Vision-Language-Action Models

DGX agent

arXiv:2605.08612v1 Announce Type: new Abstract: Addressing the escalating security vulnerabilities in Vision-Language-Action (VLA) models, this study investigates backdoor attacks targeting the visual

researcharxiv-cs-ro
12 May 2026
Agents

AtteConDA: Attention-Based Conflict Suppression in Multi-Condition Diffusion Models and Synthetic Data Augmentation

DGX agent

arXiv:2605.09425v1 Announce Type: cross Abstract: Recent conditional image generation methods can improve controllability by generating images that are faithful to conditions such as sketches, human p

agentsarxiv-cs-ai
12 May 2026
Applications

Biosignal Fingerprinting: A Cross-Modal PPG-ECG Foundation Model

DGX agent

arXiv:2605.09579v1 Announce Type: cross Abstract: Cardiovascular disease remains the leading cause of global mortality, yet scalable cardiac monitoring is hindered by the gap between diagnostic-rich E

applicationsarxiv-cs-ai
12 May 2026
Research

Can Muon Fine-tune Adam-Pretrained Models?

DGX agent

arXiv:2605.10468v1 Announce Type: new Abstract: Muon has emerged as an efficient alternative to Adam for pretraining, yet remains underused for fine-tuning. A key obstacle is that most open models are

researcharxiv-cs-lg
12 May 2026
Safety

Composing Policy Gradients and Prompt Optimization for Language Model Programs

DGX agent

arXiv:2508.04660v2 Announce Type: replace Abstract: Group Relative Policy Optimization (GRPO) has proven to be an effective tool for post-training language models (LMs). However, AI systems are increa

safetyarxiv-cs-cl
12 May 2026
Local Ai

Detector-Empowered Video Large Language Model for Efficient Spatio-Temporal Grounding

DGX agent

arXiv:2512.06673v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) are rapidly expanding from general video understanding to finer-grained understanding such as spatio-tempor

local-aiarxiv-cs-cv
12 May 2026
Local Ai

DP-LAC: Lightweight Adaptive Clipping for Differentially Private Federated Fine-tuning of Language Models

DGX agent

arXiv:2605.10272v1 Announce Type: cross Abstract: Federated learning (FL) enables the collaborative training of large-scale language models (LLMs) across edge devices while keeping user data on-device

local-aiarxiv-cs-ai
12 May 2026
Safety

EvoStreaming: Your Offline Video Model Is a Natively Streaming Assistant

DGX agent

arXiv:2605.10343v1 Announce Type: cross Abstract: Streaming video understanding demands more than watching longer videos: assistants must decide when to speak in real time, balancing responsiveness ag

safetyarxiv-cs-ai
12 May 2026
Research

FERA: Uncertainty-Aware Federated Reasoning for Large Language Models

DGX agent

arXiv:2605.10082v1 Announce Type: new Abstract: Large language models (LLMs) exhibit strong reasoning capabilities when guided by high-quality demonstrations, yet such data is often distributed across

researcharxiv-cs-cl
12 May 2026
Applications

Forecasting Source Stability in Scientific Experiments using Temporal Learning Models: A Case Study from Tritium Monitoring

DGX agent

arXiv:2605.08140v1 Announce Type: cross Abstract: The Karlsruhe Tritium Neutrino Experiment (KATRIN) aims to measure the absolute neutrino mass with unprecedented sensitivity, requiring precise monito

applicationsarxiv-cs-ai
12 May 2026
Local Ai

Fully Decentralized Cooperative Multi-Agent Reinforcement Learning is A Context Modeling Problem

DGX agent

arXiv:2509.15519v2 Announce Type: replace Abstract: This paper studies fully decentralized cooperative multi-agent reinforcement learning, where each agent solely observes the states, its local action

local-aiarxiv-cs-lg
12 May 2026
Research

Functional Subspace, where language models can use vector algebra to solve problems

DGX agent

arXiv:2602.01687v2 Announce Type: replace-cross Abstract: Large language models (LLMs) were invented for natural language tasks such as translation, but they have proved that they can perform highly c

researcharxiv-cs-ai
12 May 2026
Agents

GenCellAgent: Generalizable, Training-Free Cellular Image Segmentation via Large Language Model Agents

DGX agent

arXiv:2510.13896v2 Announce Type: replace-cross Abstract: Cellular image segmentation is essential for quantitative biology yet remains difficult due to heterogeneous modalities, morphological variabi

agentsarxiv-cs-ai
12 May 2026
Research

Generative Giants, Retrieval Weaklings: Why do Multimodal Large Language Models Fail at Multimodal Retrieval?

DGX agent

arXiv:2512.19115v2 Announce Type: replace Abstract: Despite the remarkable success of multimodal large language models (MLLMs) in generative tasks, we observe that they exhibit a counterintuitive defi

researcharxiv-cs-cv
12 May 2026
Safety

Hierarchical Causal Abduction: A Foundation Framework for Explainable Model Predictive Control

DGX agent

arXiv:2605.10624v1 Announce Type: new Abstract: Model Predictive Control (MPC) is widely used to operate safety-critical infrastructure by predicting future trajectories and optimizing control actions

safetyarxiv-cs-ai
12 May 2026
Research

Lakestream: A Consistent and Brokerless Data Plane for Large Foundation Model Training

DGX agent

arXiv:2605.09994v1 Announce Type: cross Abstract: Modern Large Foundation Model (LFM) training has transformed the data pipeline from a static ingestion layer into a dynamic component that must co-evo

researcharxiv-cs-lg
12 May 2026
Safety

LAQuant: A Simple Overhead-free Large Reasoning Model Quantization by Layer-wise Lookahead Loss

DGX agent

arXiv:2605.08755v1 Announce Type: new Abstract: Large reasoning models (LRMs) reach competition-level math and coding accuracy via long autoregressive decoding, making per-token decoding cost a primar

safetyarxiv-cs-lg
12 May 2026
← Previous
1…218219220221222…1038
Next →