AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,965
  • Agents7,446
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,131
  • Local Ai4,857
  • Model Releases23,360
  • Research19,834
  • Safety13,174
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,965
  • Agents7,446
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,131
  • Local Ai4,857
  • Model Releases23,360
  • Research19,834
  • Safety13,174
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
Human
86,965Total entries
1Added by human
86,964Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
61,904 results
Research

VidSplat: Gaussian Splatting Reconstruction with Geometry-Guided Video Diffusion Priors

DGX agent

arXiv:2605.11424v1 Announce Type: new Abstract: Gaussian Splatting has achieved remarkable progress in multi-view surface reconstruction, yet it exhibits notable degradation when only few views are av

researcharxiv-cs-cv
13 May 2026
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

VIP: Visual-guided Prompt Evolution for Efficient Dense Vision-Language Inference

DGX agent

arXiv:2605.12325v1 Announce Type: new Abstract: Pursuing training-free open-vocabulary semantic segmentation in an efficient and generalizable manner remains challenging due to the deep-seated spatial

safetyarxiv-cs-cv
13 May 2026
Research

Vision-aligned Latent Reasoning for Multi-modal Large Language Model

DGX agent

arXiv:2602.04476v2 Announce Type: replace Abstract: Despite recent advancements in Multi-modal Large Language Models (MLLMs) on diverse understanding tasks, these models struggle to solve problems whi

researcharxiv-cs-cv
13 May 2026
Model Releases

Vision-Based Hand Shadowing for Robotic Manipulation via Inverse Kinematics

DGX agent

arXiv:2603.11383v2 Announce Type: replace Abstract: Teleoperation of low-cost robotic manipulators remains challenging due to the difficulty of retargeting human hand motion to robot joint commands. W

model-releasesarxiv-cs-ro
13 May 2026
Model Releases

Vision2Code: A Multi-Domain Benchmark for Evaluating Image-to-Code Generation

DGX agent

arXiv:2605.11307v1 Announce Type: new Abstract: Image-to-code generation tests whether a vision-language model (VLM) can recover the structure of an image enough to express it as executable code. Exis

model-releasesarxiv-cs-cv
13 May 2026
Safety

VNDUQE: Information-Theoretic Novelty Detection using Deep Variational Information Bottleneck

DGX agent

arXiv:2605.11551v1 Announce Type: cross Abstract: Detecting out-of-distribution (OOD) samples is critical for safe deployment of neural networks in safety-critical applications. While maximum softmax

safetyarxiv-cs-cv
13 May 2026
Model Releases

Weather-Robust Cross-View Geo-Localization via Prototype-Based Semantic Part Discovery

DGX agent

arXiv:2605.11654v1 Announce Type: new Abstract: Cross-view geo-localization (CVGL), which matches an oblique drone view to a geo-referenced satellite tile, has emerged as a key alternative for autonom

model-releasesarxiv-cs-cv
13 May 2026
Research

Welfare as a Guiding Principle for Machine Learning -- From Compass, to Lens, to Roadmap

DGX agent

arXiv:2502.11981v3 Announce Type: replace Abstract: Decades of research in machine learning have given us powerful tools for making accurate predictions. But when used in social settings and on human

researcharxiv-cs-lg
13 May 2026
Model Releases

What Does It Mean for a Medical AI System to Be Right?

DGX agent

arXiv:2605.11963v1 Announce Type: new Abstract: This paper examines what it means for a medical AI system to be right by grounding the question in a specific clinical context: the automatic classifica

model-releasesarxiv-cs-cv
13 May 2026
Tutorials

What makes a word hard to learn? Modeling L1 influence on English vocabulary difficulty

DGX agent

arXiv:2605.12281v1 Announce Type: new Abstract: What makes a word difficult to learn, and how does the difficulty depend on the learner's native language? We computationally model vocabulary difficult

tutorialsarxiv-cs-cl
13 May 2026
Safety

What-Where Transformer: A Slot-Centric Visual Backbone for Concurrent Representation and Localization

DGX agent

arXiv:2605.12021v1 Announce Type: new Abstract: Many image understanding tasks involve identifying what is present and where it appears. However, tasks that address where, such as object discovery, de

safetyarxiv-cs-cv
13 May 2026
Tutorials

When and How to Canonize: A Generalization Perspective

DGX agent

arXiv:2605.11008v1 Announce Type: new Abstract: While invariant architectures are standard for processing symmetric data, there is growing interest in achieving invariance by applying group averaging

tutorialsarxiv-cs-lg
13 May 2026
Research

When Brains Disagree: Biological Ambiguity Underlies the Challenge of Amyloid PET Synthesis from Structural MRI

DGX agent

arXiv:2605.11867v1 Announce Type: new Abstract: Structural MRI-to-amyloid PET synthesis has been proposed as a non-invasive alternative for amyloid assessment in Alzheimer's disease (AD). However, rep

researcharxiv-cs-cv
13 May 2026
Safety

When Does ell_2-Boosting Overfit Benignly? High-Dimensional Risk Asymptotics and the ell_1 Implicit Bias

DGX agent

arXiv:2605.06314v2 Announce Type: replace Abstract: Benign overfitting is well-characterized in ell_2 geometries, but its behavior under the ell_1 implicit bias of greedy ensembles remains challenging

safetyarxiv-cs-lg
13 May 2026
Research

When Emotion Becomes Trigger: Emotion-style dynamic Backdoor Attack Parasitising Large Language Models

DGX agent

arXiv:2605.11612v1 Announce Type: new Abstract: Backdoor vulnerabilities widely exist in the fine-tuning of large language models(LLMs). Most backdoor poisoning methods operate mainly at the token lev

researcharxiv-cs-cl
13 May 2026
Research

When Looking Is Not Enough: Visual Attention Structure Reveals Hallucination in MLLMs

DGX agent

arXiv:2605.11559v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have become a key interface for visual reasoning and grounded question answering, yet they remain vulnerable to

researcharxiv-cs-cv
13 May 2026
Safety

When Policy Entropy Constraint Fails: Preserving Diversity in Flow-based RLHF via Perceptual Entropy

DGX agent

arXiv:2605.12112v1 Announce Type: new Abstract: RLHF is widely used to align flow-matching text-to-image models with human preferences, but often leads to severe diversity collapse after fine-tuning.

safetyarxiv-cs-cv
13 May 2026
Research

When the Gold Standard Isn't Necessarily Standard: Challenges of Evaluating the Translation of User-Generated Content

DGX agent

arXiv:2512.17738v2 Announce Type: replace Abstract: User-generated content (UGC) is characterised by frequent use of non-standard language, from spelling errors to expressive choices such as slang, ch

researcharxiv-cs-cl
13 May 2026
Safety

When to Ask a Question: Understanding Communication Strategies in Generative AI Tools

DGX agent

arXiv:2605.11240v1 Announce Type: cross Abstract: Generative AI models differ from traditional machine learning tools in that they allow users to provide as much or as little information as they choos

safetyarxiv-cs-lg
13 May 2026
Applications

Why Conclusions Diverge from the Same Observations: Formalizing World-Model Non-Identifiability via an Inference

DGX agent

arXiv:2605.12255v1 Announce Type: cross Abstract: When people share the same documents and observations yet reach different conclusions, the disagreement often shifts into a judgment that the other pa

applicationsarxiv-cs-lg
13 May 2026
Model Releases

WildRelight: A Real-World Benchmark and Physics-Guided Adaptation for Single-Image Relighting

DGX agent

arXiv:2605.11696v1 Announce Type: new Abstract: Recent single-image relighting methods, powered by advanced generative models, have achieved impressive photorealism on synthetic benchmarks. However, t

model-releasesarxiv-cs-cv
13 May 2026
Safety

World Action Models: The Next Frontier in Embodied AI

DGX agent

arXiv:2605.12090v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have achieved strong semantic generalization for embodied policy learning, yet they learn reactive observation-to-

safetyarxiv-cs-cl
13 May 2026
Research

WorldComp2D: Spatio-semantic Representations of Object Identity and Location from Local Views

DGX agent

arXiv:2605.11743v1 Announce Type: new Abstract: Learning latent representations that capture both semantic and spatial information is central to efficient spatio-semantic reasoning. However, many exis

researcharxiv-cs-cv
13 May 2026
Applications

X-Imitator: Spatial-Aware Imitation Learning via Bidirectional Action-Pose Interaction

DGX agent

arXiv:2605.12162v1 Announce Type: new Abstract: Effectively handling the interplay between spatial perception and action generation remains a critical bottleneck in robotic manipulation. Existing meth

applicationsarxiv-cs-ro
13 May 2026
Research

xi-DPO: Direct Preference Optimization via Ratio Reward Margin

DGX agent

arXiv:2605.10981v1 Announce Type: new Abstract: Reference-free preference optimization has emerged as an efficient alternative to reinforcement learning from human feedback, with Simple Preference Opt

researcharxiv-cs-lg
13 May 2026
Model Releases

XWOD: A Real-World Benchmark for Object Detection under Extreme Weather Conditions

DGX agent

arXiv:2605.11521v1 Announce Type: new Abstract: Autonomous driving and intelligent transportation systems remain vulnerable under extreme weather. The U.S. Federal Highway Administration reports that

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

YFPO: A Preliminary Study of Yoked Feature Preference Optimization with Neuron-Guided Rewards for Mathematical Reasoning

DGX agent

arXiv:2605.11906v1 Announce Type: new Abstract: Preference optimization has become an important post-training paradigm for improving the reasoning abilities of large language models. Existing methods

model-releasesarxiv-cs-cl
13 May 2026
Research

Zero-shot Human Pose Estimation using Diffusion-based Inverse solvers

DGX agent

arXiv:2510.02043v2 Announce Type: replace Abstract: Pose estimation refers to tracking a human's full body posture, including their head, torso, arms, and legs. The problem is challenging in practical

researcharxiv-cs-cv
13 May 2026
Safety

ZeroIDIR: Zero-Reference Illumination Degradation Image Restoration with Perturbed Consistency Diffusion Models

DGX agent

arXiv:2605.11435v1 Announce Type: new Abstract: In this paper, we propose a zero-reference diffusion-based framework, named ZeroIDIR, for illumination degradation image restoration, which decouples th

safetyarxiv-cs-cv
13 May 2026
Model Releases

100,000+ Movie Reviews from Kazakhstan: Russian, Kazakh, and Code-Switched Texts

DGX agent

arXiv:2605.08600v1 Announce Type: new Abstract: We present a new publicly available corpus of 100,502 movie reviews from Kazakhstan collected from kino.kz, spanning 2001-2025 and covering 4,943 unique

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

3DReflecNet: A Large-Scale Dataset for 3D Reconstruction of Reflective, Transparent, and Low-Texture Objects

DGX agent

arXiv:2605.10204v1 Announce Type: new Abstract: Accurate 3D reconstruction of objects with reflective, transparent, or low-texture surfaces still remains notoriously challenging. Such materials often

model-releasesarxiv-cs-cv
12 May 2026
Research

4D Neural Voxel Splatting: Dynamic Scene Rendering with Voxelized Guassian Splatting

DGX agent

arXiv:2511.00560v2 Announce Type: replace Abstract: Although 3D Gaussian Splatting (3D-GS) achieves efficient rendering for novel view synthesis, extending it to dynamic scenes still results in substa

researcharxiv-cs-cv
12 May 2026
Applications

A Breast Vision Pathology Foundation Model for Real-world Clinical Utility

DGX agent

arXiv:2605.08207v1 Announce Type: new Abstract: Pathology foundation models have shown strong retrospective performance, but whether such systems can support clinically relevant use remains unclear. T

applicationsarxiv-cs-cv
12 May 2026
Research

A Call to Lagrangian Action: Learning Population Mechanics from Temporal Snapshots

DGX agent

arXiv:2605.08550v1 Announce Type: new Abstract: The population dynamics of molecules, cells, and organisms are governed by a number of unknown forces. In the last decade, population dynamics have pred

researcharxiv-cs-lg
12 May 2026
Research

A cell-decomposition based path planner for 3D navigation in constrained workspaces

DGX agent

arXiv:2605.10086v1 Announce Type: new Abstract: This paper proposes a cell decomposition algorithm for binary occupancy grids that ensures mutual complete visibility from each cell to at least one adj

researcharxiv-cs-ro
12 May 2026
Model Releases

A Cognitively Grounded Bayesian Framework for Misinformation Susceptibility

DGX agent

arXiv:2605.09483v1 Announce Type: cross Abstract: In this (work in progress) paper, we present Bounded Pragmatic Listener (or BPL), a cognitively grounded Bayesian framework for modelling susceptibili

model-releasesarxiv-cs-ai
12 May 2026
Applications

A Cold Diffusion Approach for Percussive Dereverberation

DGX agent

arXiv:2605.10256v1 Announce Type: cross Abstract: Most recent advances in audio dereverberation focus almost exclusively on speech, leaving percussive and drum signals largely unexplored despite their

applicationsarxiv-cs-ai
12 May 2026
Model Releases

A Communication-Theoretic Framework for LLM Agents: Cost-Aware Adaptive Reliability

DGX agent

arXiv:2605.09121v1 Announce Type: cross Abstract: Agents built on large language models (LLMs) rely on a range of reliability techniques, including retry, majority voting, and self-consistency, that h

model-releasesarxiv-cs-ai
12 May 2026
Applications

A Comparative Study of Machine Learning and Deep Learning for Out-of-Distribution Detection

DGX agent

arXiv:2605.10181v1 Announce Type: cross Abstract: Out-of-distribution (OOD) detection is essential for building reliable AI systems, as models that produce outputs for invalid inputs cannot be trusted

applicationsarxiv-cs-ai
12 May 2026
Research

A Computational Operationalisation of Competing Maturational Theories of Syntactic Development via Statistical Grammar Induction

DGX agent

arXiv:2605.08476v1 Announce Type: new Abstract: This paper is concerned with what intermediate syntactic categories children acquire during first language development, and in what order. Maturational

researcharxiv-cs-cl
12 May 2026
Research

A Controlled Diagnostic Study of Hardware-Induced Distortions in Hardware-Aware Training

DGX agent

arXiv:2605.09416v1 Announce Type: new Abstract: Hardware-aware training (HAT) is widely used to improve the robustness of neural networks on non-ideal AI accelerators, such as analog in-memory computi

researcharxiv-cs-lg
12 May 2026
Safety

A Cross-Layered Multi-Drone Coordination for Medical Supply Delivery during Disaster Response Management

DGX agent

arXiv:2605.09342v1 Announce Type: cross Abstract: Autonomous drone fleets have immense potential in medical supply delivery during disaster incident response. However, coordinating multiple drones in

safetyarxiv-cs-lg
12 May 2026
Model Releases

A Deep Risk Estimator for Known Operator Learning

DGX agent

arXiv:2605.08517v1 Announce Type: cross Abstract: We describe an approach for estimating the statistical risk of deep networks that contain a mix of learned and known operators. Building on the maxima

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

A Game Theoretic Free Energy Analysis of Higher Order Synergy in Attention Heads of Large Language Models

DGX agent

arXiv:2605.09515v1 Announce Type: new Abstract: Large language models rely on multihead attention, but interactions among heads remain poorly understood. We apply the Game Theoretic Free Energy Princi

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

A Geometric Perspective on Next-Token Prediction in Large Language Models: Three Emerging Phases

DGX agent

arXiv:2605.09011v1 Announce Type: cross Abstract: We investigate the geometry of predictive information across the layers of large language models (LLMs). We repurpose representation lenses-learned af

model-releasesarxiv-cs-ai
12 May 2026
Research

A low-cost mockup to simulate robotic laser cutting in nuclear decommissioning

DGX agent

arXiv:2605.08947v1 Announce Type: new Abstract: This paper introduces a low-cost experimental mockup to simulate the laser cutting process of containers in nuclear decommissioning. It is composed of a

researcharxiv-cs-ro
12 May 2026
Research

A Market-Rule-Informed Neural Network for Efficient Imbalance Electricity Price Forecasting

DGX agent

arXiv:2605.09061v1 Announce Type: cross Abstract: Accurate and efficient imbalance electricity price forecasting is critical for industrial energy trading systems, especially as battery assets and aut

researcharxiv-cs-lg
12 May 2026
Model Releases

A meshfree exterior calculus for generalizable and data-efficient learning of physics from point clouds

DGX agent

arXiv:2605.08436v1 Announce Type: cross Abstract: We introduce a meshfree exterior calculus (MEEC) for learning structure-preserving descriptions of physics on point clouds, and use it to build MEEC-N

model-releasesarxiv-cs-ai
12 May 2026
← Previous
1…912913914915916…1290
Next →