AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
Human
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,851 results
28 Apr 2026

Unified Multi-Foundation-Model Slide Representation for Pan-Cancer Recognition and Text-Guided Tumor Localization

Local AiDGX agent

arXiv:2604.22846v1 Announce Type: new Abstract: The expanding ecosystem of pathology foundation models has produced powerful but fragmented tile-level representations, limiting their use in clinical t

Universal Approximation of Operators with Transformers and Neural Integral Operators

ResearchDGX agent

arXiv:2409.00841v3 Announce Type: replace Abstract: We study the universal approximation properties of transformers and neural integral operators for operators in Banach spaces. In particular, we show

Universal approximation property of Banach space-valued random feature models including random neural networks

TutorialsDGX agent
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2312.08410v5 Announce Type: replace Abstract: We introduce a Banach space-valued extension of random feature learning, a data-driven supervised machine learning technique for large-scale kernel

Unleashing the Agility of Wheeled-Legged Robots for High-Dynamic Reflexive Obstacle Evasion

ApplicationsDGX agent

arXiv:2604.23761v1 Announce Type: new Abstract: Wheeled-legged robots combine the energy efficiency of wheeled locomotion with the terrain adaptability of legged systems, making them promising platfor

Unrealized Expectations: Comparing AI Methods vs Classical Algorithms for Maximum Independent Set

Local AiDGX agent

arXiv:2502.03669v3 Announce Type: replace-cross Abstract: AI methods, such as generative models and reinforcement learning, have recently been applied to combinatorial optimization (CO) problems, espe

UNSEEN: A Cross-Stack LLM Unlearning Defense against AR-LLM Social Engineering Attacks

SafetyDGX agent

arXiv:2604.23141v1 Announce Type: cross Abstract: Emerging AR-LLM-based Social Engineering attack (e.g., SEAR) is at the edge of posing great threats to real-world social life. In such AR-LLM-SE attac

Unstable Rankings in Bayesian Deep Learning Evaluation

ResearchDGX agent

arXiv:2604.23102v1 Announce Type: new Abstract: Standard evaluations of Bayesian deep learning methods assume that metric estimates are reliable, but we show this assumption fails under data scarcity.

Unveiling the Backdoor Mechanism Hidden Behind Catastrophic Overfitting in Fast Adversarial Training

ResearchDGX agent

arXiv:2604.24350v1 Announce Type: cross Abstract: Fast Adversarial Training (FAT) has attracted significant attention due to its efficiency in enhancing neural network robustness against adversarial a

UpstreamQA: A Modular Framework for Explicit Reasoning on Video Question Answering Tasks

Model ReleasesDGX agent

arXiv:2604.23145v1 Announce Type: cross Abstract: Video Question Answering (VideoQA) demands models that jointly reason over spatial, temporal, and linguistic cues. However, the task's inherent comple

Urban Flood Observations (UFO): A hand-labeled training and validation dataset of post-flood inundation

ResearchDGX agent

arXiv:2604.23066v1 Announce Type: new Abstract: Urban flooding affects lives and infrastructure worldwide. Mapping inundation in complex urban environments from satellite imagery remains challenging d

Usable Agent Discovery for Decentralized AI Systems

AgentsDGX agent

arXiv:2604.23080v1 Announce Type: cross Abstract: Large-scale agentic systems run on distributed infrastructures where many software agents share physical hosts and are discovered via peer-to-peer mec

Using Language Models as Closed-Loop High-Level Planners for Robotics Applications: A Brief Overview and Benchmarks

ResearchDGX agent

arXiv:2511.07410v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) and Vision Language Models (VLMs) have become popular tools for embodied high-level planning. However, their depl

Utility-Aware Data Pricing: Token-Level Quality and Empirical Training Gain for LLMs

SafetyDGX agent

arXiv:2604.22893v1 Announce Type: cross Abstract: Traditional data valuation methods based on ``row-count imes quality coefficient'' paradigms fail to capture the nuanced, nonlinear contributions that

V-GRPO: Online Reinforcement Learning for Denoising Generative Models Is Easier than You Think

SafetyDGX agent

arXiv:2604.23380v1 Announce Type: cross Abstract: Aligning denoising generative models with human preferences or verifiable rewards remains a key challenge. While policy-gradient online reinforcement

V-SEAM: Visual Semantic Editing and Attention Modulating for Causal Interpretability of Vision-Language Models

Model ReleasesDGX agent

arXiv:2509.14837v2 Announce Type: replace Abstract: Recent advances in causal interpretability have extended from language models to vision-language models (VLMs), seeking to reveal their internal mec

Value Alignment Tax: Measuring Value Trade-offs in LLM Alignment

SafetyDGX agent

arXiv:2602.12134v2 Announce Type: replace Abstract: Existing work on value alignment typically characterizes value relations statically, ignoring how alignment interventions, such as prompting, fine-t

VAMP-Net: An Interpretable Multi-Path Network of Genomic Permutation-Invariant Set Attention and Quality-Aware 1D-CNN for MTB Drug Resistance

ResearchDGX agent

arXiv:2512.21786v2 Announce Type: replace Abstract: Genomic prediction of drug resistance in Mycobacterium tuberculosis is often hindered by complex epistatic interactions and variable sequencing qual

VAPO: End-to-end Slide-Enhanced Speech Recognition with Omni-modal Large Language Models

Model ReleasesDGX agent

arXiv:2510.08618v2 Announce Type: replace-cross Abstract: Omni-modal large language models (OLLMs) offer a promising end-to-end solution for slide-enhanced speech recognition due to their inherent mul

Variational Grey-Box Dynamics Matching

ApplicationsDGX agent

arXiv:2602.17477v3 Announce Type: replace Abstract: Deep generative models such as flow matching and diffusion models have shown great potential in learning complex distributions and dynamical systems

VDLF-Net: Variational Feature Fusion for Adaptive and Few-Shot Visual Learning

ResearchDGX agent

arXiv:2604.23641v1 Announce Type: new Abstract: This paper introduces VDLF-Net, which attaches a compact VAE to a multi-scale CNN backbone. Latent vectors and softmax-gate support the backbone feature

Verifying Quantized GNNs With Readout Is Decidable But Highly Intractable

SafetyDGX agent

arXiv:2510.08045v2 Announce Type: replace-cross Abstract: We introduce a logical language for reasoning about quantized aggregate-combine graph neural networks with global readout (ACR-GNNs). We provi

VeriLLMed: Interactive Visual Debugging of Medical Large Language Models with Knowledge Graphs

ApplicationsDGX agent

arXiv:2604.23356v1 Announce Type: new Abstract: Large language models (LLMs) show promise in medical diagnosis, but real-world deployment remains challenging due to high-stakes clinical decisions and

Vibe Medicine: Redefining Biomedical Research Through Human-AI Co-Work

AgentsDGX agent

arXiv:2604.23674v1 Announce Type: new Abstract: With the emergence of large language models (LLMs) and AI agent frameworks, the human-AI co-work paradigm known as Vibe Coding is changing how people co

Viewport-Unaware Blind Omnidirectional Image Quality Assessment: A Unified and Generalized Approach

ResearchDGX agent

arXiv:2604.23953v1 Announce Type: cross Abstract: Blind omnidirectional image quality assessment (BOIQA) presents a great challenge to the visual quality assessment community, due to different storage

Vision-Based Lane Following and Traffic Sign Recognition for Resource-Constrained Autonomous Vehicles

Local AiDGX agent

arXiv:2604.22872v1 Announce Type: new Abstract: Autonomous vehicles (AVs) rely on real-time perception systems to understand road environments and ensure safe navigation. However, implementing reliabl

Vision-Language-Action in Robotics: A Survey of Datasets, Benchmarks, and Data Engines

SafetyDGX agent

arXiv:2604.23001v1 Announce Type: cross Abstract: Despite remarkable progress in Vision--Language--Action (VLA) models, a central bottleneck remains underexamined: the data infrastructure that underli

Vision-Language-Action Safety: Threats, Challenges, Evaluations, and Mechanisms

SafetyDGX agent

arXiv:2604.23775v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models are emerging as a unified substrate for embodied intelligence. This shift raises a new class of safety challenges, s

Visual Funnel: Resolving Contextual Blindness in Multimodal Large Language Models

ResearchDGX agent

arXiv:2512.10362v2 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) demonstrate impressive reasoning capabilities, but often fail to perceive fine-grained visual details

VitaminP: cross-modal learning enables whole-cell segmentation from routine histology

ResearchDGX agent

arXiv:2604.23799v1 Announce Type: new Abstract: Accurate whole-cell and nuclear segmentation is essential for precision pathology and spatial omics, yet routine hematoxylin and eosin (H&E) staining pr

Voxify3D: Pixel Art Meets Volumetric Rendering

SafetyDGX agent

arXiv:2512.07834v2 Announce Type: replace Abstract: Voxel art is a distinctive stylization widely used in games and digital media, yet automated generation from 3D meshes remains challenging due to co

VS-DDPM: Efficient Low-Cost Diffusion Model for Medical Modality Translation

ResearchDGX agent

arXiv:2604.22942v1 Announce Type: cross Abstract: Diffusion models produce high-quality synthetic data but suffer from slow inference. We propose 3D Variable-Step Denoising Diffusion Probabilistic Mod

Weakly Supervised Multicenter Nancy Index Scoring in Ulcerative Colitis Using Foundation Models

ResearchDGX agent

arXiv:2604.23706v1 Announce Type: new Abstract: Histologic assessment of ulcerative colitis (UC) activity is an important endpoint in clinical trials and routine care, but manual grading with indices

WeatherSeg: Weather-Robust Image Segmentation using Teacher-Student Dual Learning and Classifier-Updating Attention

AgentsDGX agent

arXiv:2604.22824v1 Announce Type: cross Abstract: WeatherSeg, an advanced semi-supervised segmentation framework, addresses autonomous driving's environmental perception challenges in adverse weather

WebSerial Vision Training for Microcontrollers: A Browser-Based Companion to On-Device CNN Training

Model ReleasesDGX agent

arXiv:2604.22834v1 Announce Type: new Abstract: This paper presents webmcu-vision-web, a single-file, zero-install browser application for end-to-end TinyML vision model training and deployment on the

Well-Conditioned Oblivious Perturbations in Linear Space

ResearchDGX agent

arXiv:2604.23193v1 Announce Type: cross Abstract: Perturbing a deterministic n-dimensional matrix with small Gaussian noise is a cornerstone of smoothed analysis of algorithms [Spielman and Teng, JACM

What Did They Mean? How LLMs Resolve Ambiguous Social Situations across Perspectives and Roles

Model ReleasesDGX agent

arXiv:2604.23942v1 Announce Type: cross Abstract: People increasingly turn to large language models (LLMs) to interpret ambiguous social situations: a delayed text reply, an unusually cold supervisor,

What Drives Compositional Generalization? The Importance of Continuous Training Objectives in Visual Generative Models

ResearchDGX agent

arXiv:2510.03075v3 Announce Type: replace-cross Abstract: Compositional generalization, the ability to generate novel combinations of known concepts, is a key ingredient for visual generative models.

What Prompts Don't Say: Understanding and Managing Underspecification in LLM Prompts

ResearchDGX agent

arXiv:2505.13360v3 Announce Type: replace Abstract: Prompt underspecification is a common challenge when interacting with LLMs. In this paper, we present an in-depth analysis of this problem, showing

What Understanding Means in AI-Laden Astronomy

ApplicationsDGX agent

arXiv:2601.10038v2 Announce Type: replace-cross Abstract: Artificial intelligence is rapidly transforming astronomical research, yet the scientific community has largely treated this transformation as

When AI reviews science: Can we trust the referee?

TutorialsDGX agent

arXiv:2604.23593v1 Announce Type: new Abstract: The volume of scientific submissions continues to climb, outpacing the capacity of qualified human referees and stretching editorial timelines. At the s

When Annotators Agree but Labels Disagree: The Projection Problem in Stance Detection

SafetyDGX agent

arXiv:2603.24231v2 Announce Type: replace Abstract: Stance detection is nearly always formulated as classifying text into Favor, Against, or Neutral. This convention was inherited from debate analysis

When Chain-of-Thought Fails, the Solution Hides in the Hidden States

ResearchDGX agent

arXiv:2604.23351v1 Announce Type: cross Abstract: Whether intermediate reasoning is computationally useful or merely explanatory depends on whether chain-of-thought (CoT) tokens contain task-relevant

When Context Sticks: Studying Interference in In-Context Learning

ResearchDGX agent

arXiv:2604.23371v1 Announce Type: new Abstract: This paper investigates context stickiness in in-context learning (ICL), a phenomenon where earlier examples in a prompt interfere with a transformer's

When Corrective Hints Hurt: Prompt Design in Reasoner-Guided Repair of LLM Overcaution on Entailed Negations under OWL~2~DL

Model ReleasesDGX agent

arXiv:2604.23398v1 Announce Type: new Abstract: We report a reproducible error pattern in GPT-5.4 on OWL~2~DL compliance queries: the model frequently answers ``unknown'' when the reasoner-entailed an

When Does Removing LayerNorm Help? Activation Bounding as a Regime-Dependent Implicit Regularizer

Model ReleasesDGX agent

arXiv:2604.23434v1 Announce Type: cross Abstract: Dynamic Tanh (DyT) removes LayerNorm by bounding activations with a learned tanh(alpha x). We show that this bounding is a regime-dependent implicit r

When PINNs Go Wrong: Pseudo-Time Stepping Against Spurious Solutions

ResearchDGX agent

arXiv:2604.23528v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) provide a promising machine learning framework for solving partial differential equations, but their training o

When Policies Cannot Be Retrained: A Unified Closed-Form View of Post-Training Steering in Offline Reinforcement Learning

SafetyDGX agent

arXiv:2604.22873v1 Announce Type: cross Abstract: Offline reinforcement learning (RL) can learn effective policies from fixed datasets, but deployment objectives may change after training, and in many

When Silence Matters: The Impact of Irrelevant Audio on Text Reasoning in Large Audio-Language Models

ApplicationsDGX agent

arXiv:2510.00626v3 Announce Type: replace-cross Abstract: Large audio-language models (LALMs) unify speech and text processing, but their robustness in noisy real-world settings remains underexplored.

When to Commit? Towards Variable-Size Self-Contained Blocks for Discrete Diffusion Language Models

ResearchDGX agent

arXiv:2604.23994v1 Announce Type: cross Abstract: Discrete diffusion language models (dLLMs) enable parallel token updates with bidirectional attention, yet practical generation typically adopts block

When VLMs 'Fix' Students: Identifying and Penalizing Over-Correction in the Evaluation of Multi-line Handwritten Math OCR

Model ReleasesDGX agent

arXiv:2604.22774v1 Announce Type: cross Abstract: Accurate transcription of handwritten mathematics is crucial for educational AI systems, yet current benchmarks fail to evaluate this capability prope

Why AI Harms Can't Be Fixed One Identity at a Time: What 5300 Incident Reports Reveal About Intersectionality

ResearchDGX agent

arXiv:2604.24519v1 Announce Type: cross Abstract: AI risk assessment is the primary tool for identifying harms caused by AI systems. These include intersectional harms, which arise from the interactio

Why Architecture Choice Matters in Symbolic Regression

ResearchDGX agent

arXiv:2604.23256v1 Announce Type: cross Abstract: Symbolic regression discovers mathematical formulas from data. Some methods fix a tree of operators, assign learnable weights, and train by gradient d

WildLIFT: Lifting monocular drone video to 3D for species-agnostic wildlife monitoring

ResearchDGX agent

arXiv:2604.24718v1 Announce Type: new Abstract: Monocular RGB cameras mounted on drones are widely used for wildlife monitoring, yet most analytical pipelines remain confined to two-dimensional image

WinkTPG: An Execution Framework for Multi-Agent Path Finding Using Temporal Reasoning

AgentsDGX agent

arXiv:2508.01495v2 Announce Type: replace Abstract: Planning collision-free paths for a large group of agents is a challenging problem in many real-world applications. While recent advances in Multi-A

WISE-FM:Operation-Aware, Engineering-Informed Foundation Model for Multi-Task Well Design

Model ReleasesDGX agent

arXiv:2604.23767v1 Announce Type: new Abstract: Deploying machine learning models across diverse well portfolios requires generalisation to wells with design parameters outside the training distributi

World-R1: Reinforcing 3D Constraints for Text-to-Video Generation

SafetyDGX agent

arXiv:2604.24764v1 Announce Type: new Abstract: Recent video foundation models demonstrate impressive visual synthesis but frequently suffer from geometric inconsistencies. While existing methods atte

X-NegoBox: An Explainable Privacy-Budget Negotiation Framework for Secure Peer-to-Peer Energy Data Exchange

AgentsDGX agent

arXiv:2604.24326v1 Announce Type: cross Abstract: The decentralization of modern energy systems is transforming consumers into prosumers who continuously exchange data with aggregators, peers, and mar

XGRAG: A Graph-Native Framework for Explaining KG-based Retrieval-Augmented Generation

SafetyDGX agent

arXiv:2604.24623v1 Announce Type: new Abstract: Graph-based Retrieval-Augmented Generation (GraphRAG) extends traditional RAG by using knowledge graphs (KGs) to give large language models (LLMs) a str

XITE: Cross-lingual Interpolation for Transfer using Embeddings

ResearchDGX agent

arXiv:2604.23589v1 Announce Type: new Abstract: Facilitating cross-lingual transfer in multilingual language models remains a critical challenge. Towards this goal, we propose an embedding-based data

xOffense: An Autonomous Multi-Agent Framework for Penetration Testing with Domain-Adapted Large Language Models

Model ReleasesDGX agent

arXiv:2509.13021v2 Announce Type: replace-cross Abstract: This work introduces xOffense, an AI-driven, multi-agent penetration testing framework that shifts the process from labor-intensive, expert-dr

← Previous
1…840841842843844…998
Next →