AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
84,661Total entries
1Added by human
84,660Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,435 results
Research

From Average Sensitivity to Small-Loss Regret Bounds under Random-Order Model

DGX agent

arXiv:2602.09457v2 Announce Type: replace-cross Abstract: We study online learning in the random-order model, where the multiset of loss functions is chosen adversarially but revealed in a uniformly r

researcharxiv-cs-lg
11 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Hardware

GATO: GPU-Accelerated and Batched Trajectory Optimization for Scalable Edge Model Predictive Control

DGX agent

arXiv:2510.07625v2 Announce Type: replace Abstract: While Model Predictive Control (MPC) delivers strong performance across robotics applications, solving the underlying (batches of) nonlinear traject

hardwarearxiv-cs-ro
11 May 2026
Safety

HEART: Hyperspherical Embedding Alignment via Kent-Representation Traversal in Diffusion Models

DGX agent

arXiv:2605.07973v1 Announce Type: new Abstract: Text-to-image diffusion models can generate visually stunning images, yet, controlling what appears and how it appears, remains surprisingly difficult,

safetyarxiv-cs-cv
11 May 2026
Research

Memory-Efficient Looped Transformer: Decoupling Compute from Memory in Looped Language Models

DGX agent

arXiv:2605.07721v1 Announce Type: cross Abstract: Recurrent LLM architectures have emerged as a promising approach for improving reasoning, as they enable multi-step computation in the embedding space

researcharxiv-cs-ai
11 May 2026
Research

Mixture of Masters: Sparse Chess Language Models with Player Routing

DGX agent

arXiv:2602.04447v2 Announce Type: replace-cross Abstract: Modern chess language models are dense transformers trained on millions of games played by thousands of high-rated individuals. However, these

researcharxiv-cs-ai
11 May 2026
Safety

Modality Gap-Driven Subspace Alignment Training Paradigm For Multimodal Large Language Models

DGX agent

arXiv:2602.07026v2 Announce Type: replace-cross Abstract: Despite the success of multimodal contrastive learning in aligning visual and linguistic representations, a persistent geometric anomaly, the

safetyarxiv-cs-ai
11 May 2026
Applications

Neural CDEs as Correctors for Learned Time Series Models

DGX agent

arXiv:2512.12116v3 Announce Type: replace Abstract: Learned time-series models, whether continuous or discrete, are widely used for forecasting the states of dynamical systems but suffer from error ac

applicationsarxiv-cs-lg
11 May 2026
Safety

One Token Per Frame: Reconsidering Visual Bandwidth in World Models for VLA Policy

DGX agent

arXiv:2605.07931v1 Announce Type: cross Abstract: Vision-language-action (VLA) models increasingly rely on auxiliary world modules to plan over long horizons, yet how such modules should be parameteri

safetyarxiv-cs-ai
11 May 2026
Safety

SARA: Semantically Adaptive Relational Alignment for Video Diffusion Models

DGX agent

arXiv:2605.07800v1 Announce Type: new Abstract: Recent video diffusion models (VDMs) synthesize visually convincing clips, yet still drop entities, mis-bind attributes, and weaken the interactions spe

safetyarxiv-cs-cv
11 May 2026
Research

Skip-It? Theoretical Conditions for Layer Skipping in Vision-Language Models

DGX agent

arXiv:2509.25584v2 Announce Type: replace Abstract: Vision-language models achieve incredible performance across a wide range of tasks, but their large size makes inference costly. Recent work has sho

researcharxiv-cs-ai
11 May 2026
Safety

TextLDM: Language Modeling with Continuous Latent Diffusion

DGX agent

arXiv:2605.07748v1 Announce Type: new Abstract: Diffusion Transformers (DiT) trained with flow matching in a VAE latent space have unified visual generation across images and videos. A natural next st

safetyarxiv-cs-cl
11 May 2026
Research

TimeLesSeg: Unified Contrast-Agnostic Cross-Sectional and Longitudinal MS Lesion Segmentation via a Stochastic Generative Model

DGX agent

arXiv:2605.07955v1 Announce Type: cross Abstract: Multiple sclerosis (MS) expresses substantial clinical and radiological heterogeneity, which poses significant challenges for automatic lesion segment

researcharxiv-cs-ai
11 May 2026
Local Ai

Topology-Enhanced Alignment for Large Language Models: Trajectory Topology Loss and Topological Preference Optimization

DGX agent

arXiv:2605.07172v1 Announce Type: new Abstract: Alignment of large language models (LLMs) via SFT and RLHF/DPO typically ignores the global geometry of the representation space, relying instead on loc

local-aiarxiv-cs-cl
11 May 2026
Research

Understanding Performance Collapse in Layer-Pruned Large Language Models via Decision Representation Transitions

DGX agent

arXiv:2605.07271v1 Announce Type: cross Abstract: Layer pruning efficiently reduces Large Language Model (LLM) computational costs but often triggers sudden performance collapse. Existing representati

researcharxiv-cs-ai
11 May 2026
Applications

Vaporizer: Breaking Watermarking Schemes for Large Language Model Outputs

DGX agent

arXiv:2605.07481v1 Announce Type: cross Abstract: In this paper, we investigate the recent state-of-the-art schemes for watermarking large language models (LLMs) outputs. These techniques are claimed

applicationsarxiv-cs-ai
11 May 2026
Research

VITA-QinYu: Expressive Spoken Language Model for Role-Playing and Singing

DGX agent

arXiv:2605.06765v1 Announce Type: cross Abstract: Human speech conveys expressiveness beyond linguistic content, including personality, mood, or performance elements, such as a comforting tone or humm

researcharxiv-cs-ai
11 May 2026
Research

Adapting Medical Vision Foundation Models for Volumetric Medical Image Segmentation via Active Learning and Selective Semi-supervised Fine-tuning

DGX agent

arXiv:2509.10784v3 Announce Type: replace-cross Abstract: Medical vision foundation models remain limited in downstream tasks, particularly volumetric medical image segmentation. While fine-tuning on

researcharxiv-cs-cv
7 May 2026
Model Releases

Are Multimodal LLMs Ready for Clinical Dermatology? A Real-World Evaluation in Dermatology

DGX agent

arXiv:2605.04098v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have demonstrated promise on publicly available dermatology benchmarks. However, benchmark performance may not

model-releasesarxiv-cs-cv
7 May 2026
Research

Are you with me? A Framework for Detecting Mental Model Discrepancies in Task-Based Team Dialogues

DGX agent

arXiv:2605.03149v1 Announce Type: new Abstract: Humans typically use natural language to update teammates on task states. Since not all updates are communicated, discrepancies arise between the team m

researcharxiv-cs-ai
7 May 2026
Safety

Efficiency of Parallel and Restart Exploration Strategies in Model Free Stochastic Simulations

DGX agent

arXiv:2503.03565v3 Announce Type: replace-cross Abstract: We analyze the efficiency of parallelization and restart mechanisms for stochastic simulations in model-free settings, where the underlying sy

safetyarxiv-cs-lg
7 May 2026
Applications

Exploring Clustering Capability of Inpainting Model Embeddings for Pattern-based Individual Identification

DGX agent

arXiv:2605.04904v1 Announce Type: new Abstract: In this paper, we explore deep learning techniques for individual identification of animals based on their skin patterns. Individual identification is c

applicationsarxiv-cs-cv
7 May 2026
Local Ai

Geometry-Aware State Space Model: A New Paradigm for Whole-Slide Image Representation

DGX agent

arXiv:2605.05164v1 Announce Type: new Abstract: Accurate analysis of histopathological images is critical for disease diagnosis and treatment planning. Whole-slide images (WSIs), which digitize tissue

local-aiarxiv-cs-cv
7 May 2026
Model Releases

Gradients with Respect to Semantics Preserving Embeddings Tell the Uncertainty of Large Language Models

DGX agent

arXiv:2605.04638v1 Announce Type: new Abstract: Uncertainty quantification (UQ) is an important technique for ensuring the trustworthiness of LLMs, given their tendency to hallucinate. Existing state-

model-releasesarxiv-cs-cl
7 May 2026
Research

Local Intrinsic Dimension Unveils Hallucinations in Diffusion Models

DGX agent

arXiv:2605.05026v1 Announce Type: new Abstract: Diffusion models are prone to generating structural hallucinations - samples that match the statistical properties of the training data yet defy underly

researcharxiv-cs-cv
7 May 2026
Safety

Misaligned by Reward: Socially Undesirable Preferences in LLMs

DGX agent

arXiv:2605.05003v1 Announce Type: new Abstract: Reward models are a key component of large language model alignment, serving as proxies for human preferences during training. However, existing evaluat

safetyarxiv-cs-cl
7 May 2026
Safety

ReasoningGuard: Safeguarding Large Reasoning Models with Inference-time Safety Aha Moments

DGX agent

arXiv:2508.04204v2 Announce Type: replace Abstract: Large Reasoning Models (LRMs) have demonstrated impressive performance in reasoning-intensive tasks, but they remain vulnerable to harmful content g

safetyarxiv-cs-cl
7 May 2026
Safety

UAV-VL-R1: Generalizing Vision-Language Models via Supervised Fine-Tuning and Multi-Stage GRPO for UAV Visual Reasoning

DGX agent

arXiv:2508.11196v2 Announce Type: replace Abstract: Recent advances in vision-language models (VLMs) have demonstrated strong generalization in natural image tasks. However, their performance often de

safetyarxiv-cs-cv
7 May 2026
Research

When Relations Break: Analyzing Relation Hallucination in Vision-Language Model Under Rotation and Noise

DGX agent

arXiv:2605.05045v1 Announce Type: cross Abstract: Vision-language models (VLMs) achieve strong multimodal performance but remain prone to relation hallucination, which requires accurate reasoning over

researcharxiv-cs-cl
7 May 2026
Research

3D Human Face Reconstruction with 3DMM face model from RGB image

DGX agent

arXiv:2605.03996v1 Announce Type: new Abstract: Nowadays as convolution neural networks demonstrate its powerful problem-solving ability in the area of image processing, efforts have been made to reco

researcharxiv-cs-cv
6 May 2026
Research

Can Multimodal Large Language Models Understand Pathologic Movements? A Pilot Study on Seizure Semiology

DGX agent

arXiv:2605.03352v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have demonstrated robust capabilities in recognizing everyday human activities, yet their potential for analyzi

researcharxiv-cs-cv
6 May 2026
Tutorials

Direct Simultaneous Translation Activation for Large Audio-Language Models

DGX agent

arXiv:2509.15692v2 Announce Type: replace-cross Abstract: Simultaneous speech-to-text translation (Simul-S2TT) aims to translate speech into target text in real time, outputting translations while rec

tutorialsarxiv-cs-cl
6 May 2026
Local Ai

HeadQ: Model-Visible Distortion and Score-Space Correction for KV-Cache Quantization

DGX agent

arXiv:2605.03562v1 Announce Type: new Abstract: KV-cache quantizers usually optimize storage-space reconstruction, even though attention reads keys through logits and values through attention-weighted

local-aiarxiv-cs-lg
6 May 2026
Model Releases

IRIS: Intent Resolution via Inference-time Saccades for Open-Ended VQA in Large Vision-Language Models

DGX agent

arXiv:2602.16138v2 Announce Type: replace Abstract: We introduce IRIS (Intent Resolution via Inference-time Saccades), a novel training-free approach that uses eye-tracking data in real-time to resolv

model-releasesarxiv-cs-cv
6 May 2026
Safety

Large Language Models are Universal Reasoners for Visual Generation

DGX agent

arXiv:2605.04040v1 Announce Type: new Abstract: Text-to-image generation has advanced rapidly with diffusion models, progressing from CLIP and T5 conditioning to unified systems where a single LLM bac

safetyarxiv-cs-cv
6 May 2026
Model Releases

LightSBB-M: Bridging Schrodinger and Bass for Generative Diffusion Modeling

DGX agent

arXiv:2601.19312v2 Announce Type: replace Abstract: The Schrodinger Bridge and Bass (SBB) formulation, which jointly controls drift and volatility, is an established extension of the classical Schrodi

model-releasesarxiv-cs-lg
6 May 2026
Model Releases

MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports

DGX agent

arXiv:2605.03103v1 Announce Type: new Abstract: Semi-structured information extraction (IE) from OCR-derived clinical reports is crucial for efficiently reconstructing patients' longitudinal medical h

model-releasesarxiv-cs-cl
6 May 2026
Safety

Neuron-Anchored Rule Extraction for Large Language Models via Contrastive Hierarchical Ablation

DGX agent

arXiv:2605.03058v1 Announce Type: new Abstract: A key goal of explainable AI (XAI) is to express the decision logic of large language models (LLMs) in symbolic form and link it to internal mechanisms.

safetyarxiv-cs-lg
6 May 2026
Tutorials

PRISM-CTG: A Foundation Model for Cardiotocography Analysis with Multi-View SSL

DGX agent

arXiv:2605.02917v1 Announce Type: new Abstract: Supervised deep learning models for automated CTG analysis are typically constrained by narrowly curated labelled datasets and limited patient cohorts,

tutorialsarxiv-cs-lg
6 May 2026
Research

Quaternion Wavelet-Conditioned Diffusion Models for Image Super-Resolution

DGX agent

arXiv:2505.00334v3 Announce Type: replace Abstract: Image Super-Resolution is a fundamental problem in computer vision with broad applications spacing from medical imaging to satellite analysis. The a

researcharxiv-cs-cv
6 May 2026
Safety

Resource-Efficient Reinforcement for Reasoning Large Language Models via Dynamic One-Shot Policy Refinement

DGX agent

arXiv:2602.00815v2 Announce Type: replace Abstract: Large language models (LLMs) have exhibited remarkable performance on complex reasoning tasks, with reinforcement learning under verifiable rewards

safetyarxiv-cs-ai
6 May 2026
Research

Revisiting Graph-Tokenizing Large Language Models: A Systematic Evaluation of Graph Token Understanding

DGX agent

arXiv:2605.03514v1 Announce Type: new Abstract: The remarkable success of large language models (LLMs) has motivated researchers to adapt them as universal predictors for various graph tasks. As a wid

researcharxiv-cs-cl
6 May 2026
Research

Structural Ranking of the Cognitive Plausibility of Computational Models of Analogy and Metaphors with the Minimal Cognitive Grid

DGX agent

arXiv:2605.01359v1 Announce Type: new Abstract: In this paper, we employ the Minimal Cognitive Grid (MCG), a framework created to evaluate the cognitive plausibility of artificial systems, to offer a

researcharxiv-cs-ai
6 May 2026
Model Releases

Tenability and Weak Semantics: Modeling Non-uniform Defense -- Extended Version

DGX agent

arXiv:2605.02024v1 Announce Type: new Abstract: In Dung-style abstract argumentation, various semantics capture notions of acceptability of arguments. The admissibility semantics capture the notion th

model-releasesarxiv-cs-ai
6 May 2026
Safety

The Design and Composition of Structural Causal Decision Processes

DGX agent

arXiv:2605.02681v1 Announce Type: cross Abstract: We present two new classes of causal models of decision-making agents. Our approach is motivated by the needs of modeling the economics of computing s

safetyarxiv-cs-ai
6 May 2026
Safety

Towards Safer Large Reasoning Models by Promoting Safety Decision-Making before Chain-of-Thought Generation

DGX agent

arXiv:2603.17368v2 Announce Type: replace Abstract: Large reasoning models (LRMs) achieved remarkable performance via chain-of-thought (CoT), but recent studies showed that such enhanced reasoning cap

safetyarxiv-cs-ai
6 May 2026
Research

A framework for analyzing concept representations in neural models

DGX agent

arXiv:2605.01381v1 Announce Type: new Abstract: Understanding how neural models represent human-interpretable concepts is challenging. Prior work has explored linear concept subspaces from diverse per

researcharxiv-cs-cl
5 May 2026
Safety

A Unified Multi-Dynamics Framework for Perception-Oriented Modeling in Tendon-Driven Continuum Robots

DGX agent

arXiv:2511.18088v2 Announce Type: replace Abstract: Tendon-driven continuum robots offer intrinsically safe and contact-rich interactions owing to their kinematic redundancy and structural compliance.

safetyarxiv-cs-ro
5 May 2026
Research

Active Reasoning Vision-Language Models via Sequential Experimental Design

DGX agent

arXiv:2605.01345v1 Announce Type: new Abstract: Visual perception in modern Vision-Language Models (VLMs) is constrained by a fundamental perceptual bandwidth bottleneck: a broad field of view inevita

researcharxiv-cs-cv
5 May 2026
← Previous
1…190191192193194…1030
Next →