AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,457
  • Agents7,399
  • Applications5,302
  • Concepts5
  • Hardware1,786
  • Industry6,117
  • Local Ai4,835
  • Model Releases23,193
  • Research19,715
  • Safety13,094
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,457
  • Agents7,399
  • Applications5,302
  • Concepts5
  • Hardware1,786
  • Industry6,117
  • Local Ai4,835
  • Model Releases23,193
  • Research19,715
  • Safety13,094
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
86,457Total entries
1Added by human
86,456Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
50,764 results
Safety

Step-level Denoising-time Diffusion Alignment with Multiple Objectives

DGX agent

arXiv:2604.14379v1 Announce Type: cross Abstract: Reinforcement learning (RL) has emerged as a powerful tool for aligning diffusion models with human preferences, typically by optimizing a single rewa

safetyarxiv-cs-cv
17 Apr 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

StoryCoder: Narrative Reformulation for Structured Reasoning in LLM Code Generation

DGX agent

arXiv:2604.14631v1 Announce Type: new Abstract: Effective code generation requires both model capability and a problem representation that carefully structures how models reason and plan. Existing app

safetyarxiv-cs-cl
17 Apr 2026
Research

Threshold Differential Attention for Sink-Free, Ultra-Sparse, and Non-Dispersive Language Modeling

DGX agent

arXiv:2601.12145v2 Announce Type: replace Abstract: Softmax attention struggles with long contexts due to structural limitations: the strict sum-to-one constraint forces attention sinks on irrelevant

researcharxiv-cs-lg
17 Apr 2026
Model Releases

A Study of Failure Modes in Two-Stage Human-Object Interaction Detection

DGX agent

arXiv:2604.13448v1 Announce Type: new Abstract: Human-object interaction (HOI) detection aims to detect interactions between humans and objects in images. While recent advances have improved performan

model-releasesarxiv-cs-cv
16 Apr 2026
Safety

Asymmetric-Loss-Guided Hybrid CNN-BiLSTM-Attention Model for Industrial RUL Prediction with Interpretable Failure Heatmaps

DGX agent

arXiv:2604.13459v1 Announce Type: new Abstract: Turbofan engine degradation under sustained operational stress necessitates robust prognostic systems capable of accurately estimating the Remaining Use

safetyarxiv-cs-lg
16 Apr 2026
Safety

Beyond Arrow's Impossibility: Fairness as an Emergent Property of Multi-Agent Collaboration

DGX agent

arXiv:2604.13705v1 Announce Type: new Abstract: Fairness in language models is typically studied as a property of a single, centrally optimized model. As large language models become increasingly agen

safetyarxiv-cs-cl
16 Apr 2026
Safety

Beyond Conservative Automated Driving in Multi-Agent Scenarios via Coupled Model Predictive Control and Deep Reinforcement Learning

DGX agent

arXiv:2604.13891v1 Announce Type: new Abstract: Automated driving at unsignalized intersections is challenging due to complex multi-vehicle interactions and the need to balance safety and efficiency.

safetyarxiv-cs-ro
16 Apr 2026
Research

Dual-Enhancement Product Bundling: Bridging Interactive Graph and Large Language Model

DGX agent

arXiv:2604.14030v1 Announce Type: new Abstract: Product bundling boosts e-commerce revenue by recommending complementary item combinations. However, existing methods face two critical challenges: (1)

researcharxiv-cs-cl
16 Apr 2026
Model Releases

From Feelings to Metrics: Understanding and Formalizing How Users Vibe-Test LLMs

DGX agent

arXiv:2604.14137v1 Announce Type: new Abstract: Evaluating LLMs is challenging, as benchmark scores often fail to capture models' real-world usefulness. Instead, users often rely on ``vibe-testing'':

model-releasesarxiv-cs-cl
16 Apr 2026
Agents

Hierarchical DLO Routing with Reinforcement Learning and In-Context Vision-language Models

DGX agent

arXiv:2510.19268v2 Announce Type: replace-cross Abstract: Long-horizon routing tasks of deformable linear objects (DLOs), such as cables and ropes, are common in industrial assembly lines and everyday

agentsarxiv-cs-lg
16 Apr 2026
Model Releases

LongCoT: Benchmarking Long-Horizon Chain-of-Thought Reasoning

DGX agent

arXiv:2604.14140v1 Announce Type: new Abstract: As language models are increasingly deployed for complex autonomous tasks, their ability to reason accurately over longer horizons becomes critical. An

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

MedRCube: A Multidimensional Framework for Fine-Grained and In-Depth Evaluation of MLLMs in Medical Imaging

DGX agent

arXiv:2604.13756v1 Announce Type: new Abstract: The potential of Multimodal Large Language Models (MLLMs) in domain of medical imaging raise the demands of systematic and rigorous evaluation framework

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Multi-Task LLM with LoRA Fine-Tuning for Automated Cancer Staging and Biomarker Extraction

DGX agent

arXiv:2604.13328v1 Announce Type: new Abstract: Pathology reports serve as the definitive record for breast cancer staging, yet their unstructured format impedes large-scale data curation. While Large

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Synthesizing Instruction-Tuning Datasets with Contrastive Decoding

DGX agent

arXiv:2604.13538v1 Announce Type: new Abstract: Using responses generated by high-performing large language models (LLMs) for instruction tuning has become a widely adopted approach. However, the exis

model-releasesarxiv-cs-cl
16 Apr 2026
Research

Training-Free Semantic Multi-Object Tracking with Vision-Language Models

DGX agent

arXiv:2604.14074v1 Announce Type: new Abstract: Semantic Multi-Object Tracking (SMOT) extends multi-object tracking with semantic outputs such as video summaries, instance-level captions, and interact

researcharxiv-cs-cv
16 Apr 2026
Model Releases

TREX: Automating LLM Fine-tuning via Agent-Driven Tree-based Exploration

DGX agent

arXiv:2604.14116v1 Announce Type: cross Abstract: While Large Language Models (LLMs) have empowered AI research agents to perform isolated scientific tasks, automating complex, real-world workflows, s

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

ViBES: A Conversational Agent with Behaviorally-Intelligent 3D Virtual Body

DGX agent

arXiv:2512.14234v2 Announce Type: replace Abstract: Human communication is inherently multimodal and social: words, prosody, and body language jointly carry intent. Yet most prior systems model human

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

Beyond Scores: Diagnostic LLM Evaluation via Fine-Grained Abilities

DGX agent

arXiv:2604.12191v1 Announce Type: new Abstract: Current evaluations of large language models aggregate performance across diverse tasks into single scores. This obscures fine-grained ability variation

model-releasesarxiv-cs-ai
15 Apr 2026
Safety

Causal Diffusion Models for Counterfactual Outcome Distributions in Longitudinal Data

DGX agent

arXiv:2604.12992v1 Announce Type: cross Abstract: Predicting counterfactual outcomes in longitudinal data, where sequential treatment decisions heavily depend on evolving patient states, is critical y

safetyarxiv-cs-lg
15 Apr 2026
Local Ai

CREG: Compass Relational Evidence Graph for Characterizing Directional Structure in VLM Spatial-Reasoning Attribution

DGX agent

arXiv:2603.20475v3 Announce Type: replace Abstract: Standard attribution heatmaps show where a vision-language model (VLM) focuses, but they do not reveal whether the recovered evidence is organized b

local-aiarxiv-cs-cv
15 Apr 2026
Research

Does Visual Token Pruning Improve Calibration? An Empirical Study on Confidence in MLLMs

DGX agent

arXiv:2604.12035v1 Announce Type: new Abstract: Visual token pruning is a widely used strategy for efficient inference in multimodal large language models (MLLMs), but existing work mainly evaluates i

researcharxiv-cs-cv
15 Apr 2026
Model Releases

Dual-Modality Anchor-Guided Filtering for Test-time Prompt Tuning

DGX agent

arXiv:2604.12403v1 Announce Type: new Abstract: Test-Time Prompt Tuning (TPT) adapts vision-language models using augmented views, but its effectiveness is hindered by the challenge of determining whi

model-releasesarxiv-cs-cv
15 Apr 2026
Model Releases

Efficient Adversarial Training via Criticality-Aware Fine-Tuning

DGX agent

arXiv:2604.12780v1 Announce Type: cross Abstract: Vision Transformer (ViT) models have achieved remarkable performance across various vision tasks, with scalability being a key advantage when applied

model-releasesarxiv-cs-ai
15 Apr 2026
Applications

GGD-SLAM: Monocular 3DGS SLAM Powered by Generalizable Motion Model for Dynamic Environments

DGX agent

arXiv:2604.12837v1 Announce Type: new Abstract: Visual SLAM algorithms achieve significant improvements through the exploration of 3D Gaussian Splatting (3DGS) representations, particularly in generat

applicationsarxiv-cs-ro
15 Apr 2026
Applications

Interpretable Relational Inference with LLM-Guided Symbolic Dynamics Modeling

DGX agent

arXiv:2604.12806v1 Announce Type: new Abstract: Inferring latent interaction structures from observed dynamics is a fundamental inverse problem in many-body interacting systems. Most neural approaches

applicationsarxiv-cs-lg
15 Apr 2026
Safety

Meet Dynamic Individual Preferences: Resolving Conflicting Human Value with Paired Fine-Tuning

DGX agent

arXiv:2604.12479v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) have significantly improved the alignment of models with general human preferences. However, a major cha

safetyarxiv-cs-cl
15 Apr 2026
Model Releases

Mitigating Shortcut Learning via Feature Disentanglement in Medical Imaging: A Benchmark Study

DGX agent

arXiv:2602.18502v2 Announce Type: replace Abstract: Although deep learning models in medical imaging often achieve excellent classification performance, they can rely on shortcut learning, exploiting

model-releasesarxiv-cs-cv
15 Apr 2026
Research

Mixed-Integer vs. Continuous Model Predictive Control for Binary Thrusters: A Comparative Study

DGX agent

arXiv:2603.19796v3 Announce Type: replace-cross Abstract: Binary on/off thrusters are commonly used for spacecraft attitude and position control during proximity operations. However, their discrete na

researcharxiv-cs-ro
15 Apr 2026
Agents

OVAL: Open-Vocabulary Augmented Memory Model for Lifelong Object Goal Navigation

DGX agent

arXiv:2604.12872v1 Announce Type: new Abstract: Object Goal Navigation (ObjectNav) refers to an agent navigating to an object in an unseen environment, which is an ability often required in the accomp

agentsarxiv-cs-ro
15 Apr 2026
Research

Representation geometry shapes task performance in vision-language modeling for CT enterography

DGX agent

arXiv:2604.13021v1 Announce Type: cross Abstract: Computed tomography (CT) enterography is a primary imaging modality for assessing inflammatory bowel disease (IBD), yet the representational choices t

researcharxiv-cs-ai
15 Apr 2026
Tutorials

SubFlow: Sub-mode Conditioned Flow Matching for Diverse One-Step Generation

DGX agent

arXiv:2604.12273v1 Announce Type: cross Abstract: Flow matching has emerged as a powerful generative framework, with recent few-step methods achieving remarkable inference acceleration. However, we id

tutorialsarxiv-cs-cv
15 Apr 2026
Model Releases

TCL: Enabling Fast and Efficient Cross-Hardware Tensor Program Optimization via Continual Learning

DGX agent

arXiv:2604.12891v1 Announce Type: new Abstract: Deep learning (DL) compilers rely on cost models and auto-tuning to optimize tensor programs for target hardware. However, existing approaches depend on

model-releasesarxiv-cs-lg
15 Apr 2026
Model Releases

TriFit: Trimodal Fusion with Protein Dynamics for Mutation Fitness Prediction

DGX agent

arXiv:2604.12026v1 Announce Type: new Abstract: Predicting the functional impact of single amino acid substitutions (SAVs) is central to understanding genetic disease and engineering therapeutic prote

model-releasesarxiv-cs-lg
15 Apr 2026
Local Ai

Uncertainty Guided Exploratory Trajectory Optimization for Sampling-Based Model Predictive Control

DGX agent

arXiv:2604.12149v1 Announce Type: new Abstract: Trajectory optimization depends heavily on initialization. In particular, sampling-based approaches are highly sensitive to initial solutions, and limit

local-aiarxiv-cs-ro
15 Apr 2026
Agents

Agentic Driving Coach: Robustness and Determinism of Agentic AI-Powered Human-in-the-Loop Cyber-Physical Systems

DGX agent

arXiv:2604.11705v1 Announce Type: new Abstract: Foundation models, including large language models (LLMs), are increasingly used for human-in-the-loop (HITL) cyber-physical systems (CPS) because found

agentsarxiv-cs-ai
14 Apr 2026
Model Releases

Audio-Omni: Extending Multi-modal Understanding to Versatile Audio Generation and Editing

DGX agent

arXiv:2604.10708v1 Announce Type: cross Abstract: Recent progress in multimodal models has spurred rapid advances in audio understanding, generation, and editing. However, these capabilities are typic

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

Bridging Linguistic Gaps: Cross-Lingual Mapping in Pre-Training and Dataset for Enhanced Multilingual LLM Performance

DGX agent

arXiv:2604.10590v1 Announce Type: cross Abstract: Multilingual Large Language Models (LLMs) struggle with cross-lingual tasks due to data imbalances between high-resource and low-resource languages, a

safetyarxiv-cs-ai
14 Apr 2026
Safety

Bridging the RGB-IR Gap: Consensus and Discrepancy Modeling for Text-Guided Multispectral Detection

DGX agent

arXiv:2604.11234v1 Announce Type: new Abstract: Text-guided multispectral object detection uses text semantics to guide semantic-aware cross-modal interaction between RGB and IR for more robust percep

safetyarxiv-cs-cv
14 Apr 2026
Safety

ConfigSpec: Profiling-Based Configuration Selection for Distributed Edge--Cloud Speculative LLM Serving

DGX agent

arXiv:2604.09722v1 Announce Type: cross Abstract: Speculative decoding enables collaborative Large Language Model (LLM) inference across cloud and edge by separating lightweight token drafting from he

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

From UAV Imagery to Agronomic Reasoning: A Multimodal LLM Benchmark for Plant Phenotyping

DGX agent

arXiv:2604.09907v1 Announce Type: cross Abstract: To improve crop genetics, high-throughput, effective and comprehensive phenotyping is a critical prerequisite. While such tasks were traditionally per

model-releasesarxiv-cs-ai
14 Apr 2026
Research

Gypscie: A Cross-Platform AI Artifact Management System

DGX agent

arXiv:2604.10311v1 Announce Type: new Abstract: Artificial Intelligence (AI) models, encompassing both traditional machine learning (ML) and more advanced approaches such as deep learning and large la

researcharxiv-cs-ai
14 Apr 2026
Research

LDEPrompt: Layer-importance guided Dual Expandable Prompt Pool for Pre-trained Model-based Class-Incremental Learning

DGX agent

arXiv:2604.11091v1 Announce Type: new Abstract: Prompt-based class-incremental learning methods typically construct a prompt pool consisting of multiple trainable key-prompts and perform instance-leve

researcharxiv-cs-cv
14 Apr 2026
Model Releases

LottieGPT: Tokenizing Vector Animation for Autoregressive Generation

DGX agent

arXiv:2604.11792v1 Announce Type: new Abstract: Despite rapid progress in video generation, existing models are incapable of producing vector animation, a dominant and highly expressive form of multim

model-releasesarxiv-cs-cv
14 Apr 2026
Agents

MapATM: Enhancing HD Map Construction through Actor Trajectory Modeling

DGX agent

arXiv:2604.11081v1 Announce Type: new Abstract: High-definition (HD) mapping tasks, which perform lane detections and predictions, are extremely challenging due to non-ideal conditions such as view oc

agentsarxiv-cs-cv
14 Apr 2026
Model Releases

MEMENTO: Teaching LLMs to Manage Their Own Context

DGX agent

arXiv:2604.09852v1 Announce Type: new Abstract: Reasoning models think in long, unstructured streams with no mechanism for compressing or organizing their own intermediate state. We introduce MEMENTO:

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

RationalRewards: Reasoning Rewards Scale Visual Generation Both Training and Test Time

DGX agent

arXiv:2604.11626v1 Announce Type: new Abstract: Most reward models for visual generation reduce rich human judgments to a single unexplained score, discarding the reasoning that underlies preference.

model-releasesarxiv-cs-ai
14 Apr 2026
Hardware

SCNO: Spiking Compositional Neural Operator -- Towards a Neuromorphic Foundation Model for Nuclear PDE Solving

DGX agent

arXiv:2604.11625v1 Announce Type: cross Abstract: Neural operators have emerged as powerful surrogates for partial differential equation (PDE) solvers, yet they are typically trained as monolithic mod

hardwarearxiv-cs-ai
14 Apr 2026
Model Releases

Sign Language Recognition in the Age of LLMs

DGX agent

arXiv:2604.11225v1 Announce Type: cross Abstract: Recent Vision Language Models (VLMs) have demonstrated strong performance across a wide range of multimodal reasoning tasks. This raises the question

model-releasesarxiv-cs-cl
14 Apr 2026
← Previous
1…271272273274275…1058
Next →