AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
Human
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,851 results
30 Apr 2026

SWAN: World-Aware Adaptive Multimodal Networks for Runtime Variations

AgentsDGX agent

arXiv:2604.26181v1 Announce Type: new Abstract: Multimodal deep neural networks deployed in realistic environments must contend with runtime variations: changes in modality quality, overall input comp

Swap distance minimization shapes the order of subject, object and verb in languages of the world

ResearchDGX agent

arXiv:2604.26726v1 Announce Type: new Abstract: Languages of the world vary concerning the order of subject, object and verb. The most frequent dominant orders are SOV and SVO, and researchers have ta

SWE-Edit: Rethinking Code Editing for Efficient SWE-Agent

Model ReleasesDGX agent

arXiv:2604.26102v1 Announce Type: cross Abstract: Large language model agents have achieved remarkable progress on software engineering tasks, yet current approaches suffer from a fundamental context

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

SynSur: An end-to-end generative pipeline for synthetic industrial surface defect generation and detection

ResearchDGX agent

arXiv:2604.26633v1 Announce Type: cross Abstract: The bottleneck in learning-based industrial defect detection is often limited not by model capacity, but by the scarcity of labeled defect data: defec

Talent or Luck? Evaluating Attribution Bias in Large Language Models

SafetyDGX agent

arXiv:2505.22910v2 Announce Type: replace Abstract: When a student fails an exam, do we tend to blame their effort or the test's difficulty? Attribution, defined as how reasons are assigned to event o

TAP into the Patch Tokens: Leveraging Vision Foundation Model Features for AI-Generated Image Detection

Model ReleasesDGX agent

arXiv:2604.26772v1 Announce Type: new Abstract: Recent methods demonstrate that large-scale pretrained models, such as CLIP vision transformers, effectively detect AI-generated images (AIGIs) from uns

Tatemae: Detecting Alignment Faking via Tool Selection in LLMs

SafetyDGX agent

arXiv:2604.26511v1 Announce Type: cross Abstract: Alignment faking (AF) occurs when an LLM strategically complies with training objectives to avoid value modification, reverting to prior preferences o

TDD Governance for Multi-Agent Code Generation via Prompt Engineering

AgentsDGX agent

arXiv:2604.26615v1 Announce Type: cross Abstract: Large language models (LLMs) accelerate software development but often exhibit instability, non-determinism, and weak adherence to development discipl

Teaching LLM to be Persuasive: Reward-Enhanced Policy Optimization for Alignment from Heterogeneous Rewards

SafetyDGX agent

arXiv:2510.04214v3 Announce Type: replace Abstract: We deploy large language models (LLMs) as business development (BD) agents for persuasive price negotiation in online travel agencies (OTAs). The ag

Tell Model Where to Look: Mitigating Hallucinations in MLLMs by Vision-Guided Attention

Local AiDGX agent

arXiv:2511.20032v3 Announce Type: replace Abstract: Visual attention serves as the primary mechanism through which MLLMs interpret visual information; however, its limited localization capability ofte

Test-Time Safety Alignment

SafetyDGX agent

arXiv:2604.26167v1 Announce Type: cross Abstract: Recent work has shown that a model's input word embeddings can serve as effective control variables for steering its behavior toward outputs that sati

Text Style Transfer with Machine Translation for Graphic Designs

SafetyDGX agent

arXiv:2604.26361v1 Announce Type: cross Abstract: Globalization of graphic designs such as those used in marketing materials and magazines is increasingly important for communication to broad audience

Text-Utilization for Encoder-dominated Speech Recognition Models

ResearchDGX agent

arXiv:2604.26514v1 Announce Type: cross Abstract: This paper investigates efficient methods for utilizing text-only data to improve speech recognition, focusing on encoder-dominated models that facili

The Alignment Flywheel: A Governance-Centric Hybrid MAS for Architecture-Agnostic Safety

SafetyDGX agent

arXiv:2603.02259v2 Announce Type: replace-cross Abstract: Multi-agent systems provide mature methodologies for role decomposition, coordination, and normative governance, capabilities that remain esse

The Bandit's Blind Spot: The Critical Role of User State Representation in Recommender Systems

ResearchDGX agent

arXiv:2604.26651v1 Announce Type: cross Abstract: With the increasing availability of online information, recommender systems have become an important tool for many web-based systems. Due to the conti

The devil is in the details: Enhancing Video Virtual Try-On via Keyframe-Driven Details Injection

ResearchDGX agent

arXiv:2512.20340v3 Announce Type: replace Abstract: Although diffusion transformer (DiT)-based video virtual try-on (VVT) has made significant progress in synthesizing realistic videos, existing metho

The Dual Role of Abstracting over the Irrelevant in Symbolic Explanations: Cognitive Effort vs. Understanding

ResearchDGX agent

arXiv:2602.03467v2 Announce Type: replace Abstract: Explanations are central to human cognition, yet AI systems often produce outputs that are difficult to understand. While symbolic AI offers a trans

The False Resonance: A Critical Examination of Emotion Embedding Similarity for Speech Generation Evaluation

SafetyDGX agent

arXiv:2604.26347v1 Announce Type: cross Abstract: Objective metrics for emotional expressiveness are vital for speech generation, particularly in expressive synthesis and voice conversion requiring em

The Fools are Certain; the Wise are Doubtful: Exploring LLM Confidence in Code Completion

ResearchDGX agent

arXiv:2508.16131v2 Announce Type: replace-cross Abstract: Code completion entails the task of providing missing tokens given a surrounding context. It can boost developer productivity while providing

The hidden risks of temporal resampling in clinical reinforcement learning

ApplicationsDGX agent

arXiv:2602.06603v3 Announce Type: replace Abstract: Reinforcement learning (RL) is a type of artificial intelligence for making optimal choices. In healthcare, researchers generally use offline RL (OR

The Prompt Engineering Report Distilled: Quick Start Guide for Life Sciences

Model ReleasesDGX agent

arXiv:2509.11295v2 Announce Type: replace Abstract: Developing effective prompts demands significant cognitive investment to generate reliable, high-quality responses from Large Language Models (LLMs)

The Serial Scaling Hypothesis

ResearchDGX agent

arXiv:2507.12549v4 Announce Type: replace Abstract: While machine learning has advanced through massive parallelization, we identify a critical blind spot: some problems are fundamentally sequential.

The Unseen Adversaries: Robust and Generalized Defense Against Adversarial Patches

Model ReleasesDGX agent

arXiv:2604.26317v1 Announce Type: new Abstract: The vulnerabilities of deep neural networks against singularities have raised serious concerns regarding their deployment in the physical world. One of

Theory-Grounded Evaluation Exposes the Authorship Gap in LLM Personalization

Model ReleasesDGX agent

arXiv:2604.26460v1 Announce Type: new Abstract: Stylistic personalization - making LLMs write in a specific individual's style, rather than merely adapting to task preferences - lacks evaluation groun

Thinking with Drafting: Optical Decompression via Logical Reconstruction

Model ReleasesDGX agent

arXiv:2602.11731v2 Announce Type: replace Abstract: Existing multimodal large language models have achieved high-fidelity visual perception and exploratory visual generation. However, a precision para

Three-Step Nav: A Hierarchical Global-Local Planner for Zero-Shot Vision-and-Language Navigation

AgentsDGX agent

arXiv:2604.26946v1 Announce Type: new Abstract: Breakthrough progress in vision-based navigation through unknown environments has been achieved by using multimodal large language models (MLLMs). These

Through a Compressed Lens: Investigating The Impact of Quantization on Factual Knowledge Recall

ResearchDGX agent

arXiv:2505.13963v3 Announce Type: replace Abstract: Quantization methods are widely used to accelerate inference and streamline the deployment of large language models (LLMs). Although quantization's

TildeOpen LLM: Leveraging Curriculum Learning to Achieve Equitable Language Representation

Model ReleasesDGX agent

arXiv:2603.08182v2 Announce Type: replace-cross Abstract: Large language models often underperform in many European languages due to the dominance of English and a few high-resource languages in train

Time Blindness: Why Video-Language Models Can't See What Humans Can?

Model ReleasesDGX agent

arXiv:2505.24867v2 Announce Type: replace-cross Abstract: Recent advances in vision-language models (VLMs) have made impressive strides in understanding spatio-temporal relationships in videos. Howeve

Time series classification with random convolution kernels: pooling operators and input representations matter

Model ReleasesDGX agent

arXiv:2409.01115v5 Announce Type: replace Abstract: This article presents a new approach based on MiniRocket, called SelF-Rocket, for fast time series classification (TSC). Unlike existing approaches

TimeMM: Time-as-Operator Spectral Filtering for Dynamic Multimodal Recommendation

ApplicationsDGX agent

arXiv:2604.26247v1 Announce Type: cross Abstract: Multimodal recommendation improves user modeling by integrating collaborative signals with heterogeneous item content. In real applications, user inte

TinyR1-32B-Preview: Boosting Accuracy with Branch-Merge Distillation

Model ReleasesDGX agent

arXiv:2503.04872v3 Announce Type: replace-cross Abstract: The challenge of reducing the size of Large Language Models (LLMs) while maintaining their performance has gained significant attention. Howev

TLPO: Token-Level Policy Optimization for Mitigating Language Confusion in Large Language Models

Local AiDGX agent

arXiv:2604.26553v1 Announce Type: cross Abstract: Large language models (LLMs) demonstrate strong multilingual capabilities, yet often fail to consistently generate responses in the intended language,

ToolPRM: Fine-Grained Inference Scaling of Structured Outputs for Function Calling

ResearchDGX agent

arXiv:2510.14703v2 Announce Type: replace Abstract: Large language models (LLMs) excel at function calling, but inference scaling has been explored mainly for unstructured generation. We propose an in

Topology-Aware Representation Alignment for Semi-Supervised Vision-Language Learning

SafetyDGX agent

arXiv:2604.26370v1 Announce Type: new Abstract: Vision-language models have shown strong performance, but they often generalize poorly to specialized domains. While semi-supervised vision-language lea

Towards Redundancy Reduction in Diffusion Models for Efficient Video Super-Resolution

TutorialsDGX agent

arXiv:2509.23980v2 Announce Type: replace Abstract: Diffusion models have recently shown promising results for video super-resolution (VSR). However, directly adapting generative diffusion models to V

Training Computer Use Agents to Assess the Usability of Graphical User Interfaces

AgentsDGX agent

arXiv:2604.26020v1 Announce Type: cross Abstract: Usability testing with experts and potential users can assess the effectiveness, efficiency, and user satisfaction of graphical user interfaces (GUIs)

Training-Free Adaptation of New-Generation LLMs using Legacy Clinical Models

Model ReleasesDGX agent

arXiv:2601.03423v3 Announce Type: replace-cross Abstract: Adapting language models to the clinical domain through continued pretraining and instruction tuning requires costly retraining for each new m

Training-Free Loosely Speculative Decoding: Accepting Semantically Correct Drafts Beyond Exact Match

Model ReleasesDGX agent

arXiv:2511.22972v3 Announce Type: replace Abstract: Large language models (LLMs) achieve strong performance across diverse tasks but suffer from high inference latency due to their autoregressive gene

Translating Under Pressure: Domain-Aware LLMs for Crisis Communication

SafetyDGX agent

arXiv:2604.26597v1 Announce Type: cross Abstract: Timely and reliable multilingual communication is critical during natural and human-induced disasters, but developing effective solutions for crisis c

Tree-of-Evidence: Efficient 'System 2' Search for Faithful Multimodal Grounding

ApplicationsDGX agent

arXiv:2604.07692v2 Announce Type: replace Abstract: Large Multimodal Models (LMMs) achieve state-of-the-art performance in high-stakes domains like healthcare, yet their reasoning remains opaque. Curr

Tree-of-Text: A Tree-based Prompting Framework for Table-to-Text Generation in the Sports Domain

ResearchDGX agent

arXiv:2604.26501v1 Announce Type: cross Abstract: Generating sports game reports from structured tables is a complex table-to-text task that demands both precise data interpretation and fluent narrati

Turning the TIDE: Cross-Architecture Distillation for Diffusion Large Language Models

ResearchDGX agent

arXiv:2604.26951v1 Announce Type: cross Abstract: Diffusion large language models (dLLMs) offer parallel decoding and bidirectional context, but state-of-the-art dLLMs require billions of parameters f

U-FaceBP: Uncertainty-aware Bayesian Ensemble Deep Learning for Face Video-based Blood Pressure Estimation

ResearchDGX agent

arXiv:2412.10679v3 Announce Type: replace Abstract: Blood pressure (BP) measurement is crucial for daily health assessment. Remote photoplethysmography (rPPG), which extracts pulse waves from face vid

Uncertainty-Aware Information Pursuit for Interpretable and Reliable Medical Image Analysis

SafetyDGX agent

arXiv:2506.16742v3 Announce Type: replace Abstract: To be adopted in safety-critical domains like medical image analysis, AI systems must provide human-interpretable decisions. Variational Information

Uncertainty-Aware Pedestrian Attribute Recognition via Evidential Deep Learning

ApplicationsDGX agent

arXiv:2604.26873v1 Announce Type: new Abstract: We propose UAPAR, an Uncertainty-Aware Pedestrian Attribute Recognition framework. To the best of our knowledge, this is the first EDL-based uncertainty

Uncertainty-Aware Predictive Safety Filters for Probabilistic Neural Network Dynamics

SafetyDGX agent

arXiv:2604.26836v1 Announce Type: new Abstract: Predictive safety filters (PSFs) leverage model predictive control to enforce constraint satisfaction during deep reinforcement learning (RL) exploratio

Uncertainty-Aware Reward Discounting for Mitigating Reward Hacking

SafetyDGX agent

arXiv:2604.26360v1 Announce Type: cross Abstract: Reinforcement learning (RL) systems typically optimize scalar reward functions that assume precise and reliable evaluation of outcomes. However, real-

Understanding DNNs in Feature Interaction Models: A Dimensional Collapse Perspective

ResearchDGX agent

arXiv:2604.26489v1 Announce Type: new Abstract: DNNs have gained widespread adoption in feature interaction recommendation models. However, there has been a longstanding debate on their roles. On one

Unified 4D World Action Modeling from Video Priors with Asynchronous Denoising

ResearchDGX agent

arXiv:2604.26694v1 Announce Type: cross Abstract: We propose X-WAM, a Unified 4D World Model that unifies real-time robotic action execution and high-fidelity 4D world synthesis (video + 3D reconstruc

Unifying Runtime Monitoring Approaches for Safety-Critical Machine Learning: Application to Vision-Based Landing

SafetyDGX agent

arXiv:2604.26411v1 Announce Type: new Abstract: Runtime monitoring is essential to ensure the safety of ML applications in safety-critical domains. However, current research is fragmented, with indepe

Unifying Sparse Attention with Hierarchical Memory for Scalable Long-Context LLM Serving

Local AiDGX agent

arXiv:2604.26837v1 Announce Type: new Abstract: Long-context LLM serving is bottlenecked by the cost of attending over ever-growing KV caches. Dynamic sparse attention promises relief by accessing onl

Unsupervised Graph Modeling for Anomaly Detection in Accounting Subject Relationships

ResearchDGX agent

arXiv:2604.26216v1 Announce Type: new Abstract: This paper addresses the problem of anomaly detection in accounting subject association structures, proposing a structured modeling and unsupervised dis

Value-Guided Iterative Refinement and the DIQ-H Benchmark for Evaluating VLM Robustness

Model ReleasesDGX agent

arXiv:2512.03992v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) are essential for embodied AI and safety-critical applications, such as robotics and autonomous systems. However

Variable Elimination in Hybrid Factor Graphs for Discrete-Continuous Inference & Estimation

ApplicationsDGX agent

arXiv:2601.00545v3 Announce Type: replace Abstract: Many problems in robotics involve both continuous and discrete components, and modeling them together for estimation tasks has been a long standing

Verified Critical Step Optimization for LLM Agents

SafetyDGX agent

arXiv:2602.03412v2 Announce Type: replace Abstract: As large language model agents tackle increasingly complex long-horizon tasks, effective post-training becomes critical. Prior work faces fundamenta

Vertex Features for Neural Global Illumination

ResearchDGX agent

arXiv:2508.07852v2 Announce Type: replace-cross Abstract: Recent research on learnable neural representations has been widely adopted in the field of 3D scene reconstruction and neural rendering appli

Vibe Check: Understanding the Effects of LLM-Based Conversational Agents' Personality and Alignment on User Perceptions in Goal-Oriented Tasks

SafetyDGX agent

arXiv:2509.09870v2 Announce Type: replace-cross Abstract: Large language models (LLMs) enable conversational agents (CAs) to express distinctive personalities, raising new questions about how such des

ViBE: Visual-to-M/EEG Brain Encoding via Spatio-Temporal VAE and Distribution-Aligned Projection

SafetyDGX agent

arXiv:2604.26218v1 Announce Type: new Abstract: Brain encoding models not only serve to decipher how visual stimuli are transformed into neural responses, but also represent a critical step toward vis

ViCrop-Det: Spatial Attention Entropy Guided Cropping for Training-Free Small-Object Detection

Local AiDGX agent

arXiv:2604.26806v1 Announce Type: cross Abstract: Transformer-based architectures have established a dominant paradigm in global semantic perception; however, they remain fundamentally constrained by

← Previous
1…813814815816817…998
Next →