AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
Human
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,851 results
4 Aug 2026

Understanding Synergistic Interactions among Pathology Foundation Models via Adaptive Fusion

ResearchDGX agent

arXiv:2608.01370v1 Announce Type: new Abstract: Pathology foundation models (PFMs) provide strong tile-level representations via self-supervised pre-training on large-scale pathology images. Yet, PFMs

UniHEAR: Unified Heterogeneous-Source Attentive Retrieval for Knowledge-Based Visual Question Answering

ResearchDGX agent

arXiv:2608.01147v1 Announce Type: cross Abstract: Knowledge-Based Visual Question Answering (KB-VQA) requires retrieving relevant entity knowledge from external sources to answer visually grounded que

UniMoCa: Unifying Motion and Camera Controls as Visual Proxies for Faithful Human Video Generation

ResearchDGX agent
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2608.01944v1 Announce Type: new Abstract: Controlling human motion and camera movement is essential for faithful human-oriented video generation, yet remains challenging in multi-person scenes w

UniqueSplat: View-conditioned 3D Gaussian Splatting for Generalizable 3D Reconstruction

TutorialsDGX agent

arXiv:2608.02145v1 Announce Type: new Abstract: In this paper, we propose UniqueSplat, a view-conditioned feed-forward 3D Gaussian Splatting model to reconstruct customized 3D radiance fields for each

UniSim-SLAM: Feed-Forward SLAM with Unified Sim(3) Optimization

ResearchDGX agent

arXiv:2608.01706v1 Announce Type: new Abstract: Recent geometric foundation models enable feed-forward inference for SLAM, but their predictions are strongly dependent on the input view set, which lea

Unleashing the Potential of Large Language Models: A Blueprint for Real-Time, Enterprise-Ready Deployments

SafetyDGX agent

arXiv:2608.00419v1 Announce Type: cross Abstract: Large language models deployed in real-time, regulated settings face knowledge staleness, catastrophic forgetting, hallucination, and weak feedback lo

Unleashing the Power of Text: Text-Guided Flow Matching for Image Fusion under Complex Degradations

TutorialsDGX agent

arXiv:2608.00530v1 Announce Type: new Abstract: Infrared-visible image fusion under realistic degradation scenarios is a challenging task, as degradations not only cause a loss of reliable modality-sp

Unpacking Hateful Memes: Presupposed Context and False Claims

ResearchDGX agent

arXiv:2510.09935v2 Announce Type: replace Abstract: While memes are often humorous, they are frequently used to disseminate hate, causing serious harm to individuals and society. Current approaches to

Unsupervised Multidomain Approaches to Named Entity Recognition with Small Datasets

ResearchDGX agent

arXiv:2608.00984v1 Announce Type: new Abstract: This paper explores the challenges and the methodologies associated with learning quality representations in scenarios with unlabelled small or limited

UOT-IR: Structured Routing of High-Polyphony Symbolic Music into Fixed-Budget Representations

ResearchDGX agent

arXiv:2608.00576v1 Announce Type: cross Abstract: High-polyphony symbolic music is increasingly used in generation, analysis, and arrangement, yet many downstream tasks require bounded representations

UpliftBench: Revealing Outcome-Regime and Objective Mismatch in Uplift Evaluation

Model ReleasesDGX agent

arXiv:2608.00915v1 Announce Type: new Abstract: Uplift modeling (conditional-average-treatment-effect estimation) drives personalized targeting, yet published uplift benchmarks frequently disagree on

Upper-Expectile Multi-Step Q-Learning for Off-Policy Reinforcement Learning

SafetyDGX agent

arXiv:2608.02034v1 Announce Type: new Abstract: Multi-step returns accelerate reward propagation in off-policy reinforcement learning, but couple the evaluation of each decision to the suboptimal logg

Using Lower-Bound Representations for Trajectory Similarity Learning

ApplicationsDGX agent

arXiv:2608.01039v1 Announce Type: cross Abstract: Trajectory similarity learning is fundamental to efficient trajectory retrieval under complex distance measures. Existing learning-based methods typic

Using Non-Lipschitz Signum-based Functions for Distributed Optimization and Machine Learning: Trade-off Between Con-vergence Rate and Optimality Gap

ResearchDGX agent

arXiv:2608.01220v1 Announce Type: cross Abstract: In recent years, the prevalence of large-scale data-sets and the demand for sophisti-cated learning models have necessitated the development of effici

USP-Mamba: Unmixing-Derived Spectral and Structural Prompting for Hyperspectral Image Super-Resolution

SafetyDGX agent

arXiv:2608.02401v1 Announce Type: new Abstract: Hyperspectral image super-resolution aims to reconstruct high-resolution imagery while preserving dense spectral information. Recently, Mamba-based mode

V-Mem: Modality-Routed Retrieval for Long-Term Multimodal Agentic Memory

AgentsDGX agent

arXiv:2608.01543v1 Announce Type: cross Abstract: Interaction between users and LLM agents is increasingly multimodal: conversations interleave text with images, and a later question may target either

VARPose: Flexible 2D Pose Densification via Visual Autoregressive Modeling for Enhanced 3D Lifting

ResearchDGX agent

arXiv:2608.02214v1 Announce Type: new Abstract: Visual AutoRegressive Modeling (VAR) has excelled in natural image generation via next-scale prediction, but its use on topology-structured data like hu

VaRS-Doc: Interpretation-Aware Variant Representations via Latent Self-Probing for Visual Document Retrieval

ApplicationsDGX agent

arXiv:2608.01211v1 Announce Type: new Abstract: Visual document retrieval has recently become increasingly important in applications such as enterprise search, scientific literature discovery, and ret

VC-Tooler: Learning Compositional and Adaptive Visual Tool Use

AgentsDGX agent

arXiv:2608.02217v1 Announce Type: new Abstract: Agentic multimodal reasoning extends passive image understanding by allowing VLMs to actively acquire and refine visual evidence through visual tool int

Verification Without Sufficiency: Per-Chunk Filtering Fails on Multi-Hop RAG, and Decomposition Repairs It

ResearchDGX agent

arXiv:2608.00585v1 Announce Type: new Abstract: Verification for retrieval-augmented generation usually scores each retrieved chunk and drops the ones that fail. We show this cannot work for multi-hop

Verifier-Induced Support Reshaping in On-Policy Optimization

SafetyDGX agent

arXiv:2608.00220v1 Announce Type: cross Abstract: We show that on-policy reinforcement learning with verifiable rewards (RLVR) can improve the current objective while making successful behaviors for l

VertiAKD: Adaptive Off-Road Kinodynamics on Vertically Challenging Terrain

Local AiDGX agent

arXiv:2608.00945v1 Announce Type: new Abstract: Off-road mobility requires autonomous mobile robots to generalize across heterogeneous vehicle fleets and continuously changing terrain conditions. Exis

VespaSeg: A Resource-Aware Ground-then-Segment Pipeline for Referring Expression Segmentation

Local AiDGX agent

arXiv:2608.01077v1 Announce Type: new Abstract: Referring expression segmentation requires language conditioned localization and pixel-accurate masks, but monolithic models can be costly to deploy. We

VGER: Voxel-Guided Global Event Ranking for Event Cloud Attribution

ResearchDGX agent

arXiv:2608.01470v1 Announce Type: new Abstract: Event cameras produce sparse and asynchronous event streams that provide rich spatio-temporal information for efficient perception. Recent advances in e

Video Models as Native 4D Renderers: World-Grounded Conditioning from Animated Mesh

Model ReleasesDGX agent

arXiv:2608.00094v1 Announce Type: new Abstract: Pretrained video diffusion models can act as renderers when the desired scene state is already specified by an animated mesh, a camera trajectory, and a

Visualising Information Flow in Word Embeddings with Diffusion Tensor Imaging

ResearchDGX agent

arXiv:2601.05713v2 Announce Type: replace Abstract: Understanding how large language models (LLMs) represent natural language is a central challenge in natural language processing (NLP) research. Many

VLAGuard: A Framework for Evaluating and Mitigating Physical Attention Hijacking in Vision-Language-Action Robots within Wireless Sensor Networks

SafetyDGX agent

arXiv:2608.01028v1 Announce Type: new Abstract: Deploying Vision-Language-Action (VLA) robots as mobile edge nodes within wireless sensor networks (WSNs) requires robust protection against physical ad

Volcanic Clouds Detection through QCNN and Geostationary Satellite Multispectral Imagery

SafetyDGX agent

arXiv:2608.00072v1 Announce Type: new Abstract: Recent advances in quantum computing are opening new possibilities for Earth Observation (EO) data analysis. Quantum machine learning (QML) approaches o

VR3D: View-Robust 3D Representation Learning for Aerial-Ground Person Re-Identification

Model ReleasesDGX agent

arXiv:2608.02598v1 Announce Type: new Abstract: Aerial-ground person re-identification is a challenging task due to cross-platform viewpoint variations, which cause severe occlusion and geometric defo

WAM-Diff2: Hierarchical AR-to-Diffusion Distillation for Highly Efficient Autonomous Driving VLA

SafetyDGX agent

arXiv:2608.01035v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have emerged as a prominent paradigm for end-to-end autonomous driving; however, their efficient deployment is sev

Wasserstein mixing time of the unadjusted Langevin algorithm

SafetyDGX agent

arXiv:2608.02430v1 Announce Type: cross Abstract: We provide new estimates in Wasserstein distance for the asymptotic bias of the unadjusted Langevin algorithm, in the classical setting of log-smooth

Weights or Skills? A Survey of Robot-Learning Techniques: from Action-Predicting Weights to Robots that Write their Own Skills

SafetyDGX agent

arXiv:2608.01851v1 Announce Type: new Abstract: Robot learning is splitting into two bets: policies that bake competence into frozen weights (vision-language-action, or VLA, models), and agents that w

WHALE: A Scalable Unified Model for Recommendation with Wukong-HSTU Architecture

ApplicationsDGX agent

arXiv:2607.17017v2 Announce Type: replace-cross Abstract: As scalability becomes increasingly important in recommendation modeling, recent architectures have advanced the modeling of two broad sources

What Carries the Signal in Pathology Foundation-Model Atlases? A Patient-Level Controlled Benchmark in Breast Cancer

Model ReleasesDGX agent

arXiv:2608.00105v1 Announce Type: new Abstract: Pathology foundation models are reported to encode molecular programmes in tissue morphology, but the evidence is usually a cohort-wide ranked gene list

What Could the Agent See at 19:05? Generating Temporal Enterprise Scenarios from Real Research and Replaying Them to Evaluate Agents

AgentsDGX agent

arXiv:2608.01042v1 Announce Type: cross Abstract: Enterprise AI agents act across many apps whose data changes continuously, so an answer is correct only relative to what data existed and who could se

What Makes Position Zero Special? A Mechanistic Study of Position Zero Attention Sinks in LLMs

Model ReleasesDGX agent

arXiv:2603.06591v2 Announce Type: replace-cross Abstract: Transformers frequently allocate disproportionate attention to specific tokens, a phenomenon known as attention sinks. Causal large language m

What Transfers from Text to Vision? Capability Scaling Laws and Transfer Dynamics for VLMs

Model ReleasesDGX agent

arXiv:2608.00013v1 Announce Type: new Abstract: Choosing the right large language model (LLM) backbone is the most consequential decision when building a vision-language model (VLM), yet it remains fu

When Collaboration Becomes a Trigger: Collective Evidence-Threshold Backdoors in Multi-Agent Systems

AgentsDGX agent

arXiv:2608.01085v1 Announce Type: cross Abstract: LLM-based multi-agent systems (MAS) extend LLM capabilities through iterative communication and shared contexts. However, this collaboration introduce

When Differential Privacy Meets Wireless Federated Learning: An Improved Analysis for Privacy and Convergence

TutorialsDGX agent

arXiv:2603.19040v2 Announce Type: replace Abstract: Differentially private wireless federated learning (DPWFL) is a promising framework for protecting sensitive user data. However, foundational questi

When Do Surrogate Updates Improve Decisions? A Local Theory of Trajectory-Wise Transfer

ResearchDGX agent

arXiv:2608.01130v1 Announce Type: new Abstract: A broad range of models face the mismatch where they are updated through trajectory losses but are evaluated by downstream task reward. Here, a trajecto

When Extreme Darkness Meets Motion Blur: MeanFlow for Unified RAW Restoration

ResearchDGX agent

arXiv:2608.01720v1 Announce Type: new Abstract: Extremely low-light RAW enhancement aims to recover severely attenuated sensor signals, yet existing methods often focus on illumination and noise while

When LLM Essays Outscore Student Essays: What a Korean Writing Rubric Rewards and Where Readers Disagree

ResearchDGX agent

arXiv:2601.19913v4 Announce Type: replace Abstract: LLMs now help students plan, draft, and revise essays. Educational assessment therefore faces a basic question: how should student and LLM writing b

When May a Model Replace the Experiment? Audits, Licenses, and the Price of Trust in Surrogate-Driven Design

SafetyDGX agent

arXiv:2608.01378v1 Announce Type: new Abstract: Design campaigns in chemistry, materials science, and machine learning share a bottleneck: determining how good a candidate truly is requires an expensi

When Measurement Conventions Masquerade as Calibration Gains in Cardiac Digital Twins

Model ReleasesDGX agent

arXiv:2608.01602v1 Announce Type: new Abstract: Cardiac digital twins convert clinical images into physiological measurements through observation operators, yet calibration studies often assume a fixe

When Only the Final Text Survives: Implicit Execution Tracing for Multi-Agent Auditing

AgentsDGX agent

arXiv:2603.17445v5 Announce Type: replace-cross Abstract: When a multi-agent system produces an incorrect or harmful answer, who is accountable if execution logs and agent identifiers are unavailable?

When Prompts Control Robots: Prompt Injection Attacks in Multi-Agent Robotic Systems

AgentsDGX agent

arXiv:2608.00747v1 Announce Type: new Abstract: Large language models are increasingly integrated into autonomous robotic systems for task planning and control, but this integration exposes them to pr

When Replanning Becomes the Bottleneck: Budgeted Replanning for Embodied Agents

ResearchDGX agent

arXiv:2608.01428v1 Announce Type: cross Abstract: Embodied agents replan frequently to recover from execution drift, partial observability, and coordination hazards, but each LLM-based replanning call

When Retrieval Helps and Distracts: Evaluating Evidence-Generating LLMs for Biomedical Claim Verification

Model ReleasesDGX agent

arXiv:2608.01409v1 Announce Type: new Abstract: Biomedical fact-checking systems must do more than predict whether a claim is supported, contradicted, or unaddressed: they should also produce evidence

When Words Divide: Diachronic Ideological Polarization in Political Discourse on Social Media

ResearchDGX agent

arXiv:2608.01176v1 Announce Type: new Abstract: Political polarization has become a defining feature of online discourse, yet its long-term evolution remains poorly understood. We present a longitudin

Where did the ambiguity go? Examining how multimodal models interpret polysemous words

ResearchDGX agent

arXiv:2608.00410v1 Announce Type: cross Abstract: Human language is highly polysemous. Many common words (e.g., 'bank' or 'palm') carry several distinct meanings that shape what humans communicate and

Where Does Generative Difficulty Reside? An Empirical Study of Target Representations

Local AiDGX agent

arXiv:2608.00626v1 Announce Type: new Abstract: The target representation defines the distribution an image generator must learn, yet it is often treated as an interchangeable interface. This assumpti

Which Modality Decides? Counterfactual Modality Attribution for Multimodal LLMs

SafetyDGX agent

arXiv:2608.00076v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) increasingly support high-stakes decision making by combining complementary information from images and text. W

Who Belongs in the Eval Set? A Capability-Taxonomy-Driven Pipeline for Curating Regression Eval Sets in Agent-Extensibility Platforms

Model ReleasesDGX agent

arXiv:2608.01004v1 Announce Type: new Abstract: Platform teams hosting agent-extensibility surfaces face a regression-economics paradox: every onboarding customer ships an evaluation set tuned to thei

Who Is Responsible? Self-Adaptation Under Multiple Concurrent Uncertainties With Unknown Sources in Complex ROS-Based Systems

ResearchDGX agent

arXiv:2504.20477v4 Announce Type: replace Abstract: Robotic systems increasingly operate in dynamic, unpredictable environments, where tightly coupled sensors and software modules increase the probabi

Who Should Be Generated? Justifying Demographic Targets in Open-Ended Generation

SafetyDGX agent

arXiv:2608.02551v1 Announce Type: cross Abstract: Fairness evaluation concerns not only what a model produces, but also what its outputs ought to be compared against. When a model generates 'a CEO in

Why Does Action Chunking Improve Behavioral Cloning Performance in Robotic Control?

SafetyDGX agent

arXiv:2608.02547v1 Announce Type: new Abstract: Action chunking---predicting and executing multiple actions instead of a single action---has proven to be a critical component for learning effective ro

Why Formal Monitors Fail: Attack Distribution Entropy as a Coverage Bound for LTL-Based LLM Agent Safety

Model ReleasesDGX agent

arXiv:2608.01388v1 Announce Type: cross Abstract: Runtime safety monitors based on Linear Temporal Logic (LTL) and finite automata (FSA) are increasingly deployed to intercept unsafe tool-call sequenc

Why Large Language Models Fail at Tabular Prediction

Model ReleasesDGX agent

arXiv:2608.02412v1 Announce Type: new Abstract: Large language models (LLMs) have become the default tool for a remarkable range of tasks, yet they have had conspicuously little success at one of the

Why LLMs Give In: Conversational Factors and Reasoning Behind Medical Sycophancy

ResearchDGX agent

arXiv:2608.01017v1 Announce Type: new Abstract: A language model that abandons a correct medical answer under user pushback is more dangerous than one that was simply wrong, because it lends the credi

WiFuse: An Attention Mechanism for Human Activity Recognition using Fused CSI Amplitude and Delay-Doppler Channel Features

ResearchDGX agent

arXiv:2608.00642v1 Announce Type: new Abstract: Recently, Wi-Fi sensing has played a significant role in Human Activity Recognition (HAR), as it enables the detection of various activities using only

← Previous
1…104105106107108…998
Next →