AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

Content type
85,202Total entries
1Added by human
85,201Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
60,292 results
Research

Trustworthy Federated Label Distribution Learning under Annotation Quality Disparity

DGX agent

arXiv:2605.04827v1 Announce Type: new Abstract: Label Distribution Learning (LDL) models supervision as an instance-wise probability distribution, enabling fine-grained learning under inherent ambigui

researcharxiv-cs-lg
7 May 2026
Model Releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

TSCG: Deterministic Tool-Schema Compilation for Agentic LLM Deployments

DGX agent

arXiv:2605.04107v1 Announce Type: cross Abstract: Production agent frameworks (OpenAI Function Calling, Anthropic Tool Use, MCP) transmit tool schemas as JSON, a format designed for machine parsing, n

model-releasesarxiv-cs-cl
7 May 2026
Model Releases

UAV as Urban Construction Change Monitor: A New Benchmark and Change Captioning Model

DGX agent

arXiv:2605.04409v1 Announce Type: new Abstract: Remote Sensing Image Change Captioning (RSICC) aims to generate spatially grounded natural language descriptions of scene evolution from bi-temporal ima

model-releasesarxiv-cs-cv
7 May 2026
Safety

UAV-VL-R1: Generalizing Vision-Language Models via Supervised Fine-Tuning and Multi-Stage GRPO for UAV Visual Reasoning

DGX agent

arXiv:2508.11196v2 Announce Type: replace Abstract: Recent advances in vision-language models (VLMs) have demonstrated strong generalization in natural image tasks. However, their performance often de

safetyarxiv-cs-cv
7 May 2026
Model Releases

UFAL-CUNI at SemEval-2026 Task 11: An Efficient Modular Neuro-symbolic Method for Syllogistic Reasoning

DGX agent

arXiv:2605.04941v1 Announce Type: new Abstract: This paper describes our system submitted to SemEval-2026 Task 11: Disentangling Content and Formal Reasoning in Large Language Models. We present an ef

model-releasesarxiv-cs-cl
7 May 2026
Safety

UI2Code^N: UI-to-Code Generation as Interactive Visual Optimization

DGX agent

arXiv:2511.08195v3 Announce Type: replace Abstract: UI-to-code aims to translate UI screenshots into executable front-end code. Despite progress with vision-language models (VLMs), most existing metho

safetyarxiv-cs-cv
7 May 2026
Safety

ULF-Loc: Unbiased Landmark Feature for Robust Visual Localization with 3D Gaussian Splatting

DGX agent

arXiv:2605.04730v1 Announce Type: new Abstract: Visual localization is a core technology for augmented reality and autonomous navigation. Recent methods combine the efficient rendering of 3D Gaussian

safetyarxiv-cs-cv
7 May 2026
Safety

Uncertainty-Aware Exploratory Direct Preference Optimization for Multimodal Large Language Models

DGX agent

arXiv:2605.04874v1 Announce Type: cross Abstract: Direct Preference Optimization (DPO) has proven to be an effective solution for mitigating hallucination in Multimodal Large Language Models (MLLMs) b

safetyarxiv-cs-cl
7 May 2026
Local Ai

Uncovering Cross-Objective Interference in Multi-Objective Alignment

DGX agent

arXiv:2602.06869v2 Announce Type: replace Abstract: We study a persistent failure mode in multi-objective alignment for large language models (LLMs): training improves performance on only a subset of

local-aiarxiv-cs-cl
7 May 2026
Tutorials

Understanding In-Context Learning for Nonlinear Regression with Transformers: Attention as Featurizer

DGX agent

arXiv:2605.05176v1 Announce Type: new Abstract: Pre-trained transformers are able to learn from examples provided as part of the prompt without any weight updates, a remarkable ability known as in-con

tutorialsarxiv-cs-lg
7 May 2026
Research

Understanding LoRA as Knowledge Memory: An Empirical Analysis

DGX agent

arXiv:2603.01097v2 Announce Type: replace Abstract: Continuous knowledge updating for pre-trained large language models (LLMs) is increasingly necessary yet remains challenging. Although inference-tim

researcharxiv-cs-lg
7 May 2026
Research

Understanding Transformers through the Lens of Pavlovian Conditioning

DGX agent

arXiv:2508.08289v2 Announce Type: replace Abstract: Transformer architectures have revolutionized artificial intelligence (AI) through their attention mechanisms, yet the computational principles unde

researcharxiv-cs-lg
7 May 2026
Research

Undetectable Backdoors in Model Parameters: Hiding Sparse Secrets in High Dimensions

DGX agent

arXiv:2605.04209v1 Announce Type: cross Abstract: We present Sparse Backdoor, a supply-chain attack that plants a provably undetectable backdoor in pre-trained image classifiers, including convolution

researcharxiv-cs-lg
7 May 2026
Model Releases

Unified Framework of Distributional Regret in Multi-Armed Bandits and Reinforcement Learning

DGX agent

arXiv:2605.05102v1 Announce Type: new Abstract: We study the distribution of regret in stochastic multi-armed bandits and episodic reinforcement learning through a unified framework. We formalize a di

model-releasesarxiv-cs-lg
7 May 2026
Safety

Unifying Dynamical Systems and Graph Theory to Mechanistically Understand Computation in Neural Networks

DGX agent

arXiv:2605.03598v2 Announce Type: cross Abstract: Understanding how biological and artificial neural networks implement computation from connectivity is a central problem in neuroscience and machine l

safetyarxiv-cs-ai
7 May 2026
Safety

UniMoCo: Unified Modality Completion for Robust Multi-Modal Embeddings

DGX agent

arXiv:2505.11815v2 Announce Type: replace Abstract: Current vision-language models have been explored for multi-modal embedding tasks like information retrieval. However, they face significant challen

safetyarxiv-cs-cv
7 May 2026
Research

Unintended Negative Impacts of Promotional Language in Patent Evaluation

DGX agent

arXiv:2605.04926v1 Announce Type: new Abstract: Promotional language has been increasingly used to aid the communication of innovative ideas in science. Yet, less is known about its role in the contex

researcharxiv-cs-cl
7 May 2026
Research

UniPCB: A Generation-Assisted Detection Framework for PCB Defect Inspection

DGX agent

arXiv:2605.04635v1 Announce Type: new Abstract: Printed Circuit Board (PCB) defect inspection faces two compounding challenges: scarce and imbalanced defect samples that limit model training, and insu

researcharxiv-cs-cv
7 May 2026
Local Ai

UniVer: A Unified Perspective for Multi-step and Multi-draft Speculative Decoding

DGX agent

arXiv:2605.04543v1 Announce Type: new Abstract: Speculative decoding accelerates Large Language Models via draft-then-verify, where verification can be framed as an Optimal Transport (OT) problem. Exi

local-aiarxiv-cs-cl
7 May 2026
Tutorials

Unsat Core Prediction through Polarity-Aware Representation Learning over Clause-Literal Hypergraphs

DGX agent

arXiv:2605.04819v1 Announce Type: new Abstract: Graph neural networks have been widely used in Boolean satisfiability (SAT) tasks to learn structural information from SAT formulas. The goal of these s

tutorialsarxiv-cs-lg
7 May 2026
Safety

Using Common Random Numbers for Simulation-based Planning with Rollouts

DGX agent

arXiv:2605.04732v1 Announce Type: new Abstract: Simulation-based planning with rollouts is a widely-deployed technique for decision making in stochastic environments. The primary instrument of simulat

safetyarxiv-cs-lg
7 May 2026
Local Ai

Validity-Calibrated Reasoning Distillation

DGX agent

arXiv:2605.04078v1 Announce Type: new Abstract: Reasoning distillation aims to transfer multi-step reasoning capabilities from large language models to smaller, more efficient ones. While recent metho

local-aiarxiv-cs-lg
7 May 2026
Safety

Variance Matters: Improving Domain Adaptation via Stratified Sampling

DGX agent

arXiv:2512.05226v2 Announce Type: replace Abstract: Domain shift remains a key challenge in deploying machine learning models to the real world. Unsupervised domain adaptation (UDA) aims to address th

safetyarxiv-cs-lg
7 May 2026
Research

VC-FeS: Viewpoint-Conditioned Feature Selection for Vehicle Re-identification in Thermal Vision

DGX agent

arXiv:2605.04750v1 Announce Type: new Abstract: Identification of less-articulated objects using single-channel images, such as thermal images, is important in many applications, such as surveillance.

researcharxiv-cs-cv
7 May 2026
Model Releases

VCBench: Benchmarking LLMs in Venture Capital

DGX agent

arXiv:2509.14448v2 Announce Type: replace Abstract: Benchmarks such as SWE-bench and ARC-AGI demonstrate how shared datasets accelerate progress toward artificial general intelligence (AGI). We introd

model-releasesarxiv-cs-ai
7 May 2026
Tutorials

Velox: Learning Representations of 4D Geometry and Appearance

DGX agent

arXiv:2605.04527v1 Announce Type: new Abstract: We introduce a framework for learning latent representations of 4D objects which are descriptive, faithfully capturing object geometry and appearance; c

tutorialsarxiv-cs-cv
7 May 2026
Model Releases

Vision-EKIPL: External Knowledge-Infused Policy Learning for Visual Reasoning

DGX agent

arXiv:2506.06856v3 Announce Type: replace Abstract: Visual reasoning is crucial for understanding complex multimodal data and advancing Artificial General Intelligence. Existing methods enhance the re

model-releasesarxiv-cs-cv
7 May 2026
Research

Visual Disentangled Diffusion Autoencoders: Scalable Counterfactual Generation for Foundation Models

DGX agent

arXiv:2601.21851v2 Announce Type: replace Abstract: Foundation models, despite their robust zero-shot capabilities, remain vulnerable to spurious correlations and 'Clever Hans' strategies. Existing mi

researcharxiv-cs-lg
7 May 2026
Model Releases

VL-UniTrack: A Unified Framework with Visual-Language Prompts for UAV-Ground Visual Tracking

DGX agent

arXiv:2605.04574v1 Announce Type: new Abstract: UAV-ground visual tracking (UGVT) aims to simultaneously track the same object from both the UAV and the ground view. However, existing two-stream metho

model-releasesarxiv-cs-cv
7 May 2026
Research

Vol-Mark: A Watermark for 3D Medical Volume Data Via Cubic Difference Expansion and Contrastive Learning

DGX agent

arXiv:2605.04705v1 Announce Type: cross Abstract: Today, advances in medical technology extensively utilize 3D volume data for accurate and efficient diagnostics. However, sharing these data across ne

researcharxiv-cs-lg
7 May 2026
Model Releases

VTAgent: Agentic Keyframe Anchoring for Evidence-Aware Video TextVQA

DGX agent

arXiv:2605.04870v1 Announce Type: new Abstract: Video text-based visual question answering (Video TextVQA) aims to answer questions by reasoning over visual textual content appearing in videos. Despit

model-releasesarxiv-cs-cv
7 May 2026
Model Releases

Wasserstein-Aligned Localisation for VLM-Based Distributional OOD Detection in Medical Imaging

DGX agent

arXiv:2605.05161v1 Announce Type: new Abstract: Zero-shot anomaly localisation via vision-language models (VLMs) offers a compelling approach for rare pathology detection, yet its performance is funda

model-releasesarxiv-cs-cv
7 May 2026
Research

What Can Be Recovered Under Sparse Adversarial Corruption? Assumption-Free Theory for Linear Measurements

DGX agent

arXiv:2510.24215v4 Announce Type: replace-cross Abstract: Recovery from linear measurements under sparse adversarial corruption is typically formulated as an exact-recovery problem: one seeks structur

researcharxiv-cs-lg
7 May 2026
Model Releases

What Happens Inside Agent Memory? Circuit Analysis from Emergence to Diagnosis

DGX agent

arXiv:2605.03354v1 Announce Type: new Abstract: Agent memory failures are silent: an LLM-based agent can produce a fluent response even when it fails to extract, retain, or retrieve the information ne

model-releasesarxiv-cs-ai
7 May 2026
Local Ai

What Matters in Practical Learned Image Compression

DGX agent

arXiv:2605.05148v1 Announce Type: new Abstract: One of the major differentiators unlocked by learned codecs relative to their hard-coded traditional counterparts is their ability to be optimized direc

local-aiarxiv-cs-cv
7 May 2026
Agents

What You Think is What You See: Driving Exploration in VLM Agents via Visual-Linguistic Curiosity

DGX agent

arXiv:2605.03782v1 Announce Type: new Abstract: To navigate partially observable visual environments, recent VLM agents increasingly internalize world modeling capabilities into their policies via exp

agentsarxiv-cs-ai
7 May 2026
Hardware

When Agents Handle Secrets: A Survey of Confidential Computing for Agentic AI

DGX agent

arXiv:2605.03213v1 Announce Type: cross Abstract: Agentic AI systems, specifically LLM-driven agents that plan, invoke tools, maintain persistent memory, and delegate tasks to peer agents via protocol

hardwarearxiv-cs-ai
7 May 2026
Research

When Does Gene Regulatory Network Inference Break? A Controlled Diagnostic Study of Causal and Correlational Methods on Single-Cell Data

DGX agent

arXiv:2605.04930v1 Announce Type: new Abstract: Despite theoretical advantages, causal methods for Gene Regulatory Network (GRN) inference from single-cell RNA-seq data consistently fail to match or o

researcharxiv-cs-lg
7 May 2026
Research

When Engineering Outruns Intelligence: Rethinking Instruction-Guided Navigation

DGX agent

arXiv:2507.20021v3 Announce Type: replace-cross Abstract: Recent ObjectNav systems credit large language models (LLMs) for sizable zero-shot gains, yet it remains unclear how much comes from language

researcharxiv-cs-lg
7 May 2026
Safety

When Life Gives You BC, Make Q-functions: Extracting Q-values from Behavior Cloning for On-Robot Reinforcement Learning

DGX agent

arXiv:2605.05172v1 Announce Type: new Abstract: Behavior Cloning (BC) has emerged as a highly effective paradigm for robot learning. However, BC lacks a self-guided mechanism for online improvement af

safetyarxiv-cs-ro
7 May 2026
Applications

When LLMs get significantly worse: A statistical approach to detect model degradations

DGX agent

arXiv:2602.10144v2 Announce Type: replace-cross Abstract: Minimizing the inference cost and latency of foundation models has become a crucial area of research. Optimization approaches include theoreti

applicationsarxiv-cs-lg
7 May 2026
Research

When Relations Break: Analyzing Relation Hallucination in Vision-Language Model Under Rotation and Noise

DGX agent

arXiv:2605.05045v1 Announce Type: cross Abstract: Vision-language models (VLMs) achieve strong multimodal performance but remain prone to relation hallucination, which requires accurate reasoning over

researcharxiv-cs-cl
7 May 2026
Safety

Why Expert Alignment Is Hard: Evidence from Subjective Evaluation

DGX agent

arXiv:2605.04972v1 Announce Type: new Abstract: Aligning large language models with expert judgment is especially difficult in subjective evaluation tasks, where experts may disagree, rely on tacit cr

safetyarxiv-cs-cl
7 May 2026
Research

Why Geometric Continuity Emerges in Deep Neural Networks: Residual Connections and Rotational Symmetry Breaking

DGX agent

arXiv:2605.04971v1 Announce Type: cross Abstract: Weight matrices in deep networks exhibit geometric continuity -- principal singular vectors of adjacent layers point in similar directions. While this

researcharxiv-cs-cl
7 May 2026
Applications

YOTOnet: Zero-Shot Cross-Domain Fault Diagnosis via Domain-Conditioned Mixture of Experts

DGX agent

arXiv:2605.04528v1 Announce Type: new Abstract: Mechanical equipment forms the critical backbone of modern industrial production, yet domain shift severely limits the generalization of deep learning b

applicationsarxiv-cs-lg
7 May 2026
Model Releases

12 Angry AI Agents: Evaluating Multi-Agent LLM Decision-Making Through Cinematic Jury Deliberation

DGX agent

arXiv:2605.01986v1 Announce Type: new Abstract: What if the twelve jurors of Sidney Lumet's 12 Angry Men (1957) were not men, but large language models? Would the one juror who disagrees still be able

model-releasesarxiv-cs-ai
6 May 2026
Research

3D Human Face Reconstruction with 3DMM face model from RGB image

DGX agent

arXiv:2605.03996v1 Announce Type: new Abstract: Nowadays as convolution neural networks demonstrate its powerful problem-solving ability in the area of image processing, efforts have been made to reco

researcharxiv-cs-cv
6 May 2026
Research

4RC: 4D Reconstruction via Conditional Querying Anytime and Anywhere

DGX agent

arXiv:2602.10094v2 Announce Type: replace Abstract: We present 4RC, a unified feed-forward framework for 4D reconstruction from monocular videos. Unlike existing approaches that typically decouple mot

researcharxiv-cs-cv
6 May 2026
← Previous
1…957958959960961…1257
Next →