AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,510
  • Agents7,405
  • Applications5,305
  • Concepts5
  • Hardware1,789
  • Industry6,120
  • Local Ai4,835
  • Model Releases23,219
  • Research19,716
  • Safety13,102
  • Syntheses17
  • Tools1,670
  • Tutorials3,327

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,510
  • Agents7,405
  • Applications5,305
  • Concepts5
  • Hardware1,789
  • Industry6,120
  • Local Ai4,835
  • Model Releases23,219
  • Research19,716
  • Safety13,102
  • Syntheses17
  • Tools1,670
  • Tutorials3,327

Source
Human
86,510Total entries
1Added by human
86,509Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
61,498 results
2 Jul 2026

Personalization as Inverse Planning: Learning Latent Design Intents for Agentic Slide Generation via Structural Denoising

SafetyDGX agent

arXiv:2607.00407v1 Announce Type: new Abstract: Slide design requires personalizing both deck themes and page layouts. Yet, current AI agent-based methods struggle with fine-grained, page-level design

Personalized Object Identification and Localization via In-Context Inference with Vision-Language Models

Local AiDGX agent

arXiv:2607.00357v1 Announce Type: new Abstract: Personalized object localization (POL) localizes an object instance in a query image based on a few reference images with bounding-box annotations and a

PETIMOT: A Novel Framework for Inferring Protein Motions from Sparse Data Using SE(3)-Equivariant Graph Neural Networks

ResearchDGX agent
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2504.02839v2 Announce Type: replace-cross Abstract: Proteins move and deform to ensure their biological functions. Despite significant progress in protein structure prediction, approximating con

Phantom References: Hallucinated Citations That Survive Peer Review at Top-Tier Conferences

ResearchDGX agent

arXiv:2607.00738v1 Announce Type: cross Abstract: Large language models can generate polished scientific text that includes unsupported claims, allowing hallucinations to enter the archival record. As

PHREEQC-MCQ-200: A Diagnostic Benchmark for Tool-Augmented Scientific Simulator Agents

Model ReleasesDGX agent

arXiv:2607.00436v1 Announce Type: new Abstract: Large language model agents are increasingly connected to scientific software, yet it remains unclear when tool access makes scientific computation more

PixelEyes: Decoupling Perception and Reasoning for Pinpoint Visual Evidence Seeking

Model ReleasesDGX agent

arXiv:2607.00115v1 Announce Type: new Abstract: This paper explores multi-turn visual reasoning and observes that MLLMs repeatedly fail to localize the target, leading to long, redundant trajectories.

Planning over MAPF Agent Dependencies via Multi-Dependency PIBT

AgentsDGX agent

arXiv:2603.23405v2 Announce Type: replace-cross Abstract: Modern Multi-Agent Path Finding (MAPF) algorithms must plan for hundreds to thousands of agents in congested environments within a second, req

Play Like Champions: Counterfactual Feedback Generation in Latent Space

ResearchDGX agent

arXiv:2607.00190v1 Announce Type: cross Abstract: Recent advances in reinforcement learning have produced superhuman agents across a wide range of competitive games. As a byproduct, researchers have b

PorTEXTO: A European Portuguese Benchmark for Visual Text Extraction

Model ReleasesDGX agent

arXiv:2606.19096v2 Announce Type: replace Abstract: European Portuguese (pt-PT) is largely absent from OCR benchmarks, which skew toward high-resource languages. The few benchmarks that cover pt-PT fo

Post-Training Pruning for Diffusion Transformers

Model ReleasesDGX agent

arXiv:2607.00927v1 Announce Type: cross Abstract: Diffusion Transformers (DiTs) have demonstrated impressive performance in image generation but suffer from substantial computational overhead and reso

PRA-RAG: Provably Robust Aggregation in Retrieval-Augmented Generation against Retrieval Corruption

ResearchDGX agent

arXiv:2607.00012v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) enhances Large Language Models (LLMs) by incorporating external knowledge, effectively mitigating their inherent

Predicting Lethal Outcome (Cause) And Understanding Key Biomarkers Linked With Acute Myocardial Infarction Using Deep Artificial Neural Network And Ensemble Of Machine Learning Methodologies

ResearchDGX agent

arXiv:2607.00472v1 Announce Type: cross Abstract: Cardiovascular disease is still one of the main causes of death around the world. Acute myocardial infarction (MI), or heart attack, claims millions o

Predicting LLM Reasoning Performance with Small Proxy Model

SafetyDGX agent

arXiv:2509.21013v4 Announce Type: replace-cross Abstract: Given the prohibitive cost of pre-training large language models, it is essential to leverage smaller proxy models to optimize datasets before

[Preprint] Dynamic Modeling, Gait Synthesis, and Control of a Novel Subsurface Bore Propagator

ResearchDGX agent

arXiv:2607.00569v1 Announce Type: new Abstract: In this article, we present dynamic modeling, gait synthesis, and feedback control design for a modular novel subsurface robot, designed for human-free

Prior-Anchored Debiasing for Long-Tailed Multi-Organ Pathology Report Generation

SafetyDGX agent

arXiv:2607.00499v1 Announce Type: new Abstract: Automated pathology report generation from Whole Slide Images (WSIs) has attracted increasing attention in digital pathology. However, existing methods

PRISM: Prioritized Channel Importance with Semi-supervised Domain Adaptation for Cross-Subject EEG Emotion Recognition

SafetyDGX agent

arXiv:2607.00358v1 Announce Type: new Abstract: Electroencephalogram (EEG) captures endogenous brain activity with high temporal fidelity and holds substantial promise for precise emotion decoding. Ho

PRISM-VO: Scale-Aware Visual Odometry Using Photometric Plenoptic Bundle Adjustment

ResearchDGX agent

arXiv:2607.00176v1 Announce Type: new Abstract: We introduce PRISM-VO, a novel pure optimization-based sparse photometric visual odometry framework for focused plenoptic cameras. The core of PRISM-VO

Privacy-Preserving Depth-Only Open-Vocabulary 3D Semantic Segmentation Via Uncertainty-Guided Test-Time Optimization

ApplicationsDGX agent

arXiv:2607.00978v1 Announce Type: new Abstract: Privacy-preserving perception is a critical requirement for deploying 3D scene understanding systems in real-world indoor environments, yet it remains u

Progressive Pose-Guided 4D Animal Reconstruction from Monocular Video

ResearchDGX agent

arXiv:2607.00157v1 Announce Type: new Abstract: Reconstructing 4D animals from monocular videos is challenging due to large inter-species variation, complex articulations, and the lack of reliable tem

Prompt Optimization for User Simulation in Conversational Recommender Systems: A Multi-Objective Framework

SafetyDGX agent

arXiv:2607.00010v1 Announce Type: cross Abstract: Conversational recommender systems (CRSs) are a core component of next-generation intelligent recommender systems because they enable users to activel

Prompt2Effect: Training-Free Image-to-Video Model Specialization via LoRA Generation

SafetyDGX agent

arXiv:2606.13971v2 Announce Type: replace Abstract: While personalizing Image-to-Video (I2V) diffusion models with specific visual effects is increasingly demanded for high-end generation, current pra

Prompting GPT-5 on Scrum Certification Questions: An Empirical Accuracy Study

Model ReleasesDGX agent

arXiv:2607.00049v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used in Agile Software Development for documentation, coaching, and training. As practitioners adopt the

Prototype Language Models

Local AiDGX agent

arXiv:2607.00510v1 Announce Type: new Abstract: Knowing which training examples drive outputs is fundamental to auditing, correcting, and understanding language models, yet for modern LLMs this remain

Prototype Memory-Guided Training-Free Anomaly Classification and Localization in Prenatal Ultrasound

ResearchDGX agent

arXiv:2607.00744v1 Announce Type: cross Abstract: Prenatal anomaly classification and localization is of critical importance for fetal health and pregnancy management. Although ultrasound (US) is the

QCA: Query- and Content-Aware Keyframe Selection for Long Video Understanding

ResearchDGX agent

arXiv:2607.00983v1 Announce Type: new Abstract: Video understanding is often plagued by severe temporal redundancy, where processing dense frame sequences is both semantically inefficient and computat

Quadratic Programming Approach for Nash Equilibrium Computation in Multiplayer Imperfect-Information Games

ResearchDGX agent

arXiv:2509.25618v2 Announce Type: replace-cross Abstract: There has been significant recent progress in algorithms for approximation of Nash equilibrium in large two-player zero-sum imperfect-informat

QuaMoE-DRF: Proactive Beam and Rate Adaptation via Multimodal Dynamic Radio Map Forecasting in ISAC Networks

Model ReleasesDGX agent

arXiv:2607.00974v1 Announce Type: cross Abstract: Static radio maps provide location-dependent propagation priors, but they cannot capture short-term blockage caused by moving objects. Direct sensing-

Quantifying the Affective Gap: A Zero-Shot Evaluation of LLMs on Fine-Grained Emotion Taxonomies

Model ReleasesDGX agent

arXiv:2607.00968v1 Announce Type: new Abstract: Emotion recognition in natural language is a foundational challenge in affective computing, with critical implications for human-computer interaction, m

Quantitative Understanding of PDF Fits and their Uncertainties

ResearchDGX agent

arXiv:2512.24116v3 Announce Type: replace-cross Abstract: Parton Distribution Functions (PDFs) play a central role in describing experimental data at colliders and provide insight into the structure o

Quantum vs. Classical Machine Learning: A Unified Empirical Comparison

Model ReleasesDGX agent

arXiv:2607.01197v1 Announce Type: new Abstract: Quantum computing has emerged as a promising computational paradigm for machine learning (ML), with the potential to offer computational advantages over

QuasiMoTTo: Quasi-Monte Carlo Test-Time Scaling

SafetyDGX agent

arXiv:2607.01179v1 Announce Type: cross Abstract: Scaling inference compute, by generating many parallel attempts per problem, is a costly but reliable lever for improving language model capabilities.

Radial Interaction Tomography: Recognizing Non-Transitive Evolutionary Games from One Range-Expansion Image

Model ReleasesDGX agent

arXiv:2607.00378v1 Announce Type: new Abstract: Colored sectors in a microbial range expansion encode more than lineage survival counts. We formulate a computer-vision inverse problem: from one endpoi

RareDxR1: Autonomous Medical Reasoning for Rare Disease Diagnosis Beyond Human Annotation

AgentsDGX agent

arXiv:2607.00147v1 Announce Type: new Abstract: Rare disease differential diagnosis is a critical yet arduous clinical task, requiring physicians to identify precise phenotypes from complex, unstructu

RC-GeoCP: Geometric Consensus for Radar-Camera Collaborative Perception

Model ReleasesDGX agent

arXiv:2603.00654v3 Announce Type: replace Abstract: Collaborative perception (CP) improves scene understanding through multi-agent information sharing, yet LiDAR-centric systems remain costly and vuln

Readable but Not Controllable: Neuron-Level Evidence for Medical LLM Hallucination

Local AiDGX agent

arXiv:2607.00158v1 Announce Type: new Abstract: Hallucination remains one of the central obstacles to deploying medical LLMs. Yet, even when hallucination can be detected, it is still unclear whether

Reading Order Inference for Complex Document Layouts

ResearchDGX agent

arXiv:2607.01018v1 Announce Type: cross Abstract: Reading order inference remains a critical bottleneck in the digitization of complex historical manuscripts, where pages contain multiple spatially in

Real-Time Hard Negative Sampling via LLM-based Clustering for Large-Scale Two-Tower Retrieval

SafetyDGX agent

arXiv:2607.00448v1 Announce Type: cross Abstract: The two-tower model has been widely used for large-scale recommendation systems, particularly in the retrieval stage. Industry standards for training

Reasoning Up the Instruction Ladder for Controllable Language Models

SafetyDGX agent

arXiv:2511.04694v5 Announce Type: replace-cross Abstract: As large language model (LLM) based systems take on high-stakes roles in real-world decision-making, they must reconcile competing instruction

Recovering Input Text from Hidden States: Study of Gradient-Based Inversion of Decoder-Only Language Models

Model ReleasesDGX agent

arXiv:2607.00852v1 Announce Type: cross Abstract: This work studies the hidden-state inversion problem: recovering the original input token sequence of a decoder-only language model from its last-laye

Recursively Enumerably Representable Classes and Computable Versions of the Fundamental Theorem of Statistical Learning

ResearchDGX agent

arXiv:2511.02644v2 Announce Type: replace Abstract: We study computable probably approximately correct (CPAC) learning, where learners are required to be computable functions. It had been previously o

Relation-Centric Open-Vocabulary 3D Gaussian Segmentation

ResearchDGX agent

arXiv:2607.01140v1 Announce Type: new Abstract: Open-vocabulary 3D Gaussian segmentation is challenging because it requires language understanding for diverse queries and accurate separation of Gaussi

Representation as a Bottleneck for Mechanistic Interpretability: The Manifestation Unit Protocol

ResearchDGX agent

arXiv:2607.00089v1 Announce Type: new Abstract: Mechanistic interpretability has produced a rich inventory of component-level analyses that characterise what neural-network components encode and how t

Restore3D: Breathing Life into Broken Objects with Shape and Texture Restoration

TutorialsDGX agent

arXiv:2607.00522v1 Announce Type: new Abstract: Restoring incomplete or damaged 3D objects is crucial for cultural heritage preservation, occluded object reconstruction, and artistic design. Existing

RetailSMV: Exocentric vs. Egocentric Adaptation of Foundation Video World Models in Retail

Model ReleasesDGX agent

arXiv:2607.00310v1 Announce Type: cross Abstract: Foundation video diffusion models are increasingly viewed as world simulators for embodied agents, yet their pretraining on internet-scale generic vid

Rethinking Multi-Label Image Classification With Deep Learning: Taxonomy, Challenge, and Outlook

AgentsDGX agent

arXiv:2607.00839v1 Announce Type: new Abstract: Multi-label image classification (MLIC), a fundamental task in computer vision, focuses on identifying multiple objects or concepts within an image, und

Rethinking Robust Adversarial Concept Erasure in Diffusion Models

ResearchDGX agent

arXiv:2510.27285v4 Announce Type: replace Abstract: Concept erasure methods aim to remove specific unsafe target concepts in diffusion models while preserving image generation utility. To address the

Rethinking Visual Privacy: A Compositional Privacy Risk Framework for Severity Assessment with VLMs

ResearchDGX agent

arXiv:2603.21573v2 Announce Type: replace Abstract: Existing visual privacy benchmarks largely treat privacy as a binary property, labeling images as private or non-private based on visible sensitive

Retrieved Images as Visual Thought: Training-Free Multimodal In-Context Learning for the Open-vs-Closed Gap

Model ReleasesDGX agent

arXiv:2607.00606v1 Announce Type: new Abstract: Recent work on Thinking with Images makes vision a dynamic part of reasoning, but does so through generation: the model invokes external tools, synthesi

Revisiting Autoregressive Models for Generative Image Classification

SafetyDGX agent

arXiv:2603.19122v2 Announce Type: replace Abstract: Class-conditional generative models have emerged as accurate and robust classifiers, with diffusion models demonstrating clear advantages over other

Reward function compression facilitates goal-dependent reinforcement learning

ResearchDGX agent

arXiv:2509.06810v3 Announce Type: replace-cross Abstract: Humans can uniquely assign value to novel, abstract outcomes to support reinforcement learning. However, this flexibility is cognitively costl

Right in the Right Way: LM Training with Verifiable Rewards and Human Demonstrations

Model ReleasesDGX agent

arXiv:2607.01181v1 Announce Type: cross Abstract: RL with verifiable rewards (RLVR) has emerged as a powerful paradigm for training LMs on tasks with well-defined success metrics, such as code generat

RoadBench: Benchmarking MLLMs on Fine-Grained Spatial Understanding and Reasoning under Urban Road Scenarios

Model ReleasesDGX agent

arXiv:2511.18011v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) have demonstrated powerful capabilities in general spatial understanding and reasoning. However, their fine

Robots Ask the Way: Communication-Enabled Social Navigation

SafetyDGX agent

arXiv:2607.01044v1 Announce Type: new Abstract: Assistive autonomous robots operating in multi-agent environments require efficient strategies to locate specific individuals among multiple residents.

RoboWorld: Fast and Reliable Neural Simulators for Generalist Robot Policy Evaluation

SafetyDGX agent

arXiv:2607.01060v1 Announce Type: new Abstract: Video world models are emerging as a scalable alternative for evaluating generalist robot policies, bypassing the physical constraints and engineering b

Robust 3D Alignment of Generative Reconstructions via Partial Monocular Observations

Model ReleasesDGX agent

arXiv:2607.00498v1 Announce Type: new Abstract: Aligning generative 3D reconstructions with partial monocular observations is a critical but under-explored challenge in computer vision. This task is i

Robust Operational Space Control with Conformal Disturbance Bounds for Safe Redundant Manipulation

SafetyDGX agent

arXiv:2607.00424v1 Announce Type: new Abstract: Redundant robotic manipulators operating in constrained and human-interactive environments require accurate task-space tracking together with rigorous s

Root Cause Analysis of Outliers in Unknown Cyclic Graphs

ResearchDGX agent

arXiv:2510.06995v2 Announce Type: replace-cross Abstract: We study the propagation of outliers in cyclic causal graphs with linear structural equations, tracing them back to one or several 'root cause

ROSA: A Robotics Foundation Model Serving System for Robot Factories

HardwareDGX agent

arXiv:2607.01088v1 Announce Type: new Abstract: Robotics foundation models (RFMs) are making general-purpose robots increasingly practical for factory deployments. While RFM serving systems are centra

Rosetta: Composable Native Multimodal Pretraining

ResearchDGX agent

arXiv:2607.00293v1 Announce Type: cross Abstract: Achieving true artificial general intelligence requires foundation models capable of integrating new modalities without forgetting prior knowledge. Ho

Safe Alone, Unsafe Together: Safeguarding Against Implicit Toxicity When Benign Images Combine

SafetyDGX agent

arXiv:2607.00576v1 Announce Type: new Abstract: Multi-image content has become an increasingly prevalent form of visual communication in social media, giving rise to a new safety issue, multi-image im

← Previous
1…292293294295296…1025
Next →