AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “papers”

GridTimelineEvolution
12,265 results
Tutorials

Balancing Uncertainty and Diversity of Samples: Leveraging Diversity of Least, High Confidence Samples for Effective Active Learning

DGX agent

arXiv:2605.22169v1 Announce Type: new Abstract: Deep learning models, including Convolutional Neural Networks (CNNs) and Vision Transformers (ViTs), have achieved state-of-the-art performance on vario

tutorialsarxiv-cs-cv
22 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Bounding-Box Trajectories Matter for Video Anomaly Detection

DGX agent

arXiv:2605.21957v1 Announce Type: new Abstract: Video anomaly detection is critical for public safety and security, yet remains highly challenging despite extensive research due to large variations in

safetyarxiv-cs-cv
22 May 2026
Local Ai

Broadening Access to Transportation Safety Data with Generative AI: A Schema-Grounded Framework for Spatial Natural Language Queries

DGX agent

arXiv:2605.21712v1 Announce Type: new Abstract: Transportation safety analysis requires integrating crash records, roadway attributes, and geospatial data through GIS-based workflows, but access remai

local-aiarxiv-cs-cl
22 May 2026
Model Releases

Cohesion-6K: An Arabic Dataset for Analyzing Social Cohesion and Conflict in Online Discourse

DGX agent

arXiv:2605.22447v1 Announce Type: new Abstract: The study of online discourse has become central to understanding societal polarization. While much research has focused on detecting overt toxicity, th

model-releasesarxiv-cs-cl
22 May 2026
Research

ConvNeXt-FD: A Fractal-Based Deep Model for Robust Biomedical Image Segmentation

DGX agent

arXiv:2605.22002v1 Announce Type: new Abstract: Biomedical image segmentation is a critical task in medical diagnosis and treatment planning, enabling precise delineation of anatomical structures and

researcharxiv-cs-cv
22 May 2026
Tutorials

Cross-Domain Human Action Recognition from Multiview Motion and Textual Descriptions

DGX agent

arXiv:2605.22697v1 Announce Type: new Abstract: Robustness to domain changes is a key capability for effective deployment of human action recognition systems in real-world scenarios, where action cate

tutorialsarxiv-cs-cv
22 May 2026
Research

Detection of Virus and Small Cell Patches in Foci Images Using Switchable Convolution and Feature Pyramid Networks

DGX agent

arXiv:2605.22290v1 Announce Type: new Abstract: Accurate detection and counting of virus patches in focus-forming unit (FFU) images, also known as foci images, are important for quantifying viral infe

researcharxiv-cs-cv
22 May 2026
Hardware

EasyVFX: Frequency-Driven Decoupling for Resource-Efficient VFX Generation

DGX agent

arXiv:2605.22051v1 Announce Type: new Abstract: Generating high-fidelity visual effects (VFX) typically demands massive datasets and prohibitive computational power due to the intricate coupling of sp

hardwarearxiv-cs-cv
22 May 2026
Agents

Enabling Regulatory Multi-Agent Collaboration: Architecture, Challenges, and Solutions

DGX agent

arXiv:2509.09215v2 Announce Type: replace Abstract: Large language models (LLMs)-empowered autonomous agents are transforming both digital and physical environments by enabling adaptive, multi-agent c

agentsarxiv-cs-ai
22 May 2026
Research

Enhancing Gaze Reasoning in Vision Foundation Models for Gaze Following

DGX agent

arXiv:2605.22607v1 Announce Type: new Abstract: Gaze following requires both scene understanding and gaze reasoning to localize the gaze target of an in-scene person. Recently, vision foundation model

researcharxiv-cs-cv
22 May 2026
Research

Entropy-Guided Self-Supervised Learning for Medical Image Classification

DGX agent

arXiv:2605.21970v1 Announce Type: cross Abstract: Accurate and robust medical image classification is paramount for early disease diagnosis and treatment planning. However, challenges such as limited

researcharxiv-cs-cv
22 May 2026
Research

From TF-IDF to Transformers: A Comparative and Ensemble Approach to Sentiment Classification

DGX agent

arXiv:2605.22003v1 Announce Type: new Abstract: Sentiment analysis, also referred to as opinion mining, primarily tries to extract opinion from any text-based data. In the context of movie reviews and

researcharxiv-cs-cl
22 May 2026
Applications

GenHAR: Generalizing Cross-domain Human Activity Recognition for Last-mile Delivery

DGX agent

arXiv:2605.22086v1 Announce Type: new Abstract: Human Activity Recognition (HAR) has shown remarkable effectiveness in various applications, such as smart healthcare and intelligent manufacturing. How

applicationsarxiv-cs-cv
22 May 2026
Research

High Quality Embeddings for Horn Logic Reasoning

DGX agent

arXiv:2605.20467v1 Announce Type: new Abstract: Neural networks can be trained to rank the choices made by logical reasoners, resulting in more efficient searches for answers. A key step in this proce

researcharxiv-cs-ai
22 May 2026
Research

Higher Order Reasoning for Collaborative Communicationless Mobile Robot Operations

DGX agent

arXiv:2605.21901v1 Announce Type: new Abstract: In communicationless environments, multi-robot systems must operate without the constant information exchange that many coordination strategies typicall

researcharxiv-cs-ro
22 May 2026
Tutorials

How to Build Marcus's Algebraic Mind: Algebro-Deterministic Substrate over Galois Fields

DGX agent

arXiv:2605.21379v2 Announce Type: cross Abstract: In The Algebraic Mind, Gary Marcus identified three components essential for any adequate cognitive architecture: operations over variables, recursive

tutorialsarxiv-cs-ai
22 May 2026
Tutorials

HUSKY: Humanoid Skateboarding System via Physics-Aware Whole-Body Control

DGX agent

arXiv:2602.03205v2 Announce Type: replace Abstract: While current humanoid whole-body control frameworks predominantly rely on the static environment assumptions, addressing tasks characterized by hig

tutorialsarxiv-cs-ro
22 May 2026
Local Ai

I built a free demo for Pixal3D (Tencent new image-to-3D model)

DGX agent

Pixal3D is a Tencent image-to-3D model that generates high-fidelity 3D assets from a single image by explicitly lifting pixel features into 3D through back-projection to establish direct pixel-to-3D c

local-air-stablediffusion
22 May 2026
Applications

Impact of Atmospheric Turbulence and Pointing Error on Earth Observation

DGX agent

arXiv:2605.22268v1 Announce Type: cross Abstract: Earth Observation (EO) imagery is often degraded by atmospheric turbulence and pointing jitter; yet, these effects are rarely considered in datasets u

applicationsarxiv-cs-cv
22 May 2026
Research

Improving Viewpoint-Invariance and Temporal Consistency for Action Detection

DGX agent

arXiv:2605.22695v1 Announce Type: new Abstract: Viewpoint change invariance and action temporal consistency are critical aspects for the effective deployment of human action detection of untrimmed vid

researcharxiv-cs-cv
22 May 2026
Research

Industrial Dual-Arm Box Handling via Online Inertial Estimation and Convex Wrench Optimization

DGX agent

arXiv:2605.22021v1 Announce Type: new Abstract: Industrial robotic object handling often involves boxes and packages whose mass and center of mass are not known in advance. These uncertainties affect

researcharxiv-cs-ro
22 May 2026
Research

Interpreting and Enhancing Emotional Circuits in Large Vision-Language Models via Cross-Modal Information Flow

DGX agent

arXiv:2605.21980v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) represent a significant leap towards empathetic agents, demonstrating remarkable capabilities in emotion understand

researcharxiv-cs-cv
22 May 2026
Model Releases

Ishigaki-IDS-Bench: A Benchmark for Generating Information Delivery Specification from BIM Information Requirements

DGX agent

arXiv:2605.22079v1 Announce Type: new Abstract: Large language models (LLMs) are widely used to generate structured outputs such as JSON, SQL, and code, yet public resources remain limited for evaluat

model-releasesarxiv-cs-cl
22 May 2026
Local Ai

Look-Closer-Then-Diagnose: Confidence-Aware Ultrasound VQA via Active Zooming

DGX agent

arXiv:2605.21652v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have significantly advanced medical visual question answering, yet their performance in ultrasound remains suboptimal. In

local-aiarxiv-cs-cv
22 May 2026
Model Releases

Maestro: Reinforcement Learning to Orchestrate Hierarchical Model-Skill Ensembles

DGX agent

arXiv:2605.22177v1 Announce Type: cross Abstract: The proliferation of large language models (LLMs) and modular skills has endowed autonomous agents with increasingly powerful capabilities. Existing f

model-releasesarxiv-cs-cl
22 May 2026
Applications

Moment-Reenacting: Inverse Motion Degradation with Cross-shutter Guidance

DGX agent

arXiv:2605.22423v1 Announce Type: new Abstract: Motion degradation, manifested as blur in global shutter (GS) images or rolling shutter (RS) distortion in RS counterparts, remains a fundamental challe

applicationsarxiv-cs-cv
22 May 2026
Research

Motion Design for Grasp-Based Dynamic Locomotion in Microgravity

DGX agent

arXiv:2605.21704v1 Announce Type: new Abstract: Locomotion in microgravity often relies on sparsely and irregularly arranged anchors, motivating grasp-based mobility with multiple limbs. In this setti

researcharxiv-cs-ro
22 May 2026
Research

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering

DGX agent

arXiv:2605.22269v1 Announce Type: new Abstract: Long streaming video QA remains challenging due to growing visual tokens and limited reasoning length of large language models (LLMs). KV-caching stores

researcharxiv-cs-cv
22 May 2026
Research

Multi-Stage Training for Abusive Comment Detection in Indic Languages

DGX agent

arXiv:2605.22380v1 Announce Type: new Abstract: In recent years social media has become an increasingly popular tool for communication. People use it to share their ideas, exchange information, and di

researcharxiv-cs-cl
22 May 2026
Safety

Noise-Space Attribution and Control of Chunk-Boundary Artifact

DGX agent

arXiv:2603.11642v2 Announce Type: replace Abstract: Action chunking is widely used in generative visuomotor policies, yet the recurring execution discontinuities at chunk boundaries still lack a mecha

safetyarxiv-cs-ro
22 May 2026
Model Releases

Open-World Evaluations for Measuring Frontier AI Capabilities

DGX agent

arXiv:2605.20520v1 Announce Type: new Abstract: Benchmark-based evaluation remains important for tracking frontier AI progress. But it can both overstate and understate deployed capability because it

model-releasesarxiv-cs-ai
22 May 2026
Research

Optical Quantum Mixed-State Reconstruction With Multiple Deep Learning Approaches

DGX agent

arXiv:2407.01734v4 Announce Type: replace-cross Abstract: Quantum state tomography is a crucial technique for characterizing the state of a quantum system, which is essential for many applications in

researcharxiv-cs-ai
22 May 2026
Model Releases

OSCToM: RL-Guided Adversarial Generation for High-Order Theory of Mind

DGX agent

arXiv:2605.20423v1 Announce Type: new Abstract: Large Language Models (LLMs) perform well on many language tasks, but their Theory of Mind (ToM) reasoning is still uneven in complex social settings. E

model-releasesarxiv-cs-ai
22 May 2026
Hardware

PALS: Power-Aware LLM Serving for Mixture-of-Experts Models

DGX agent

arXiv:2605.21427v1 Announce Type: new Abstract: Large language model (LLM) inference has become a dominant workload in modern data centers, driving significant GPU utilization and energy consumption.

hardwarearxiv-cs-ai
22 May 2026
Model Releases

PartCo: Part-Level Correspondence Priors Enhance Category Discovery

DGX agent

arXiv:2509.22769v2 Announce Type: replace Abstract: Generalized Category Discovery (GCD) aims to identify both known and novel categories within unlabeled data by leveraging a set of labeled examples

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

Physiology and Anatomy Aware Inverse Inference of Myocardial Infarction for Cardiac Digital Twin

DGX agent

arXiv:2605.22044v1 Announce Type: new Abstract: Accurate localization of myocardial infarction is essential for risk stratification. While LGE-MRI remains the gold standard, it is resource-intensive.

model-releasesarxiv-cs-cv
22 May 2026
Research

Planning in the LLM Era: Building for Reliability and Efficiency

DGX agent

arXiv:2605.21902v1 Announce Type: cross Abstract: Growing attention to intelligent agents has put a spotlight on one of their central capabilities: planning. Early attempts to leverage large language

researcharxiv-cs-cl
22 May 2026
Applications

PointLLM-R: Enhancing 3D Point Cloud Reasoning via Chain-of-Thought

DGX agent

arXiv:2605.22013v1 Announce Type: new Abstract: Understanding 3D point clouds through language remains a fundamental challenge in computer graphics and visual computing, due to the irregular structure

applicationsarxiv-cs-cv
22 May 2026
Research

PolycubeNet: A Dual-latent Diffusion Model for Polycube-Based Hexahedral Mesh Generation

DGX agent

arXiv:2605.20274v1 Announce Type: cross Abstract: Hexahedral meshes are widely used in simulation pipelines, yet automatic generation remains challenging for complex CAD geometries. Polycube-based hex

researcharxiv-cs-ai
22 May 2026
Applications

PrivacyAkinator: Articulating Key Privacy Design Decisions by Answering LLM-Generated Multiple-choice Questions

DGX agent

arXiv:2605.20206v1 Announce Type: cross Abstract: NIST's Privacy Risk Assessment Methodology (PRAM) provides a structured framework for privacy experts to assess privacy risks. However, its complexity

applicationsarxiv-cs-ai
22 May 2026
Agents

Psy-Chronicle:A Structured Pipeline for Synthesizing Long-Horizon Campus Psychological Counseling Dialogues

DGX agent

arXiv:2605.22140v1 Announce Type: new Abstract: In recent years, large language models have shown substantial potential in psychological support tasks. However, existing psychological counseling data

agentsarxiv-cs-cl
22 May 2026
Model Releases

SADGE: Structure and Appearance Domain Gap Estimation of Synthetic and Real Data

DGX agent

arXiv:2605.22467v1 Announce Type: new Abstract: We propose SADGE, a quantitative similarity metric that predicts the performance of synthetic image datasets for common computer vision tasks without do

model-releasesarxiv-cs-cv
22 May 2026
Applications

SE3Kit: A Lightweight Python Library for Specialized Geometric Primitives in Robotics

DGX agent

arXiv:2605.22633v1 Announce Type: new Abstract: The Python robotics ecosystem faces a challenge: while many libraries exist for rigid body transformations, few are both lightweight and mathematically

applicationsarxiv-cs-ro
22 May 2026
Model Releases

Seeing the Poem: Image-Semantic Detection of AI-Generated Modern Chinese Poetry with MLLMs

DGX agent

arXiv:2605.22654v1 Announce Type: new Abstract: Previous detection studies have shown that LLMs cannot be effectively used as detectors, but these studies have not addressed modern Chinese poetry. Mor

model-releasesarxiv-cs-cl
22 May 2026
Safety

Self-Policy Distillation via Capability-Selective Subspace Projection

DGX agent

arXiv:2605.22675v1 Announce Type: new Abstract: Self-distillation bootstraps large language models (LLMs) by training on their own generations. However, existing methods either rely on external signal

safetyarxiv-cs-cl
22 May 2026
Research

Sem-Detect: Semantic Level Detection of AI Generated Peer-Reviews

DGX agent

arXiv:2605.21713v1 Announce Type: new Abstract: How can we distinguish whether a peer review was written by a human or generated by an AI model? We argue that, in this setting, authorship should not b

researcharxiv-cs-cl
22 May 2026
Safety

SENIOR: Efficient Query Selection and Preference-Guided Exploration in Preference-based Reinforcement Learning

DGX agent

arXiv:2506.14648v2 Announce Type: replace Abstract: Preference-based Reinforcement Learning (PbRL) methods provide a solution to avoid reward engineering by learning reward models based on human prefe

safetyarxiv-cs-ro
22 May 2026
Applications

Spectral Tail Auxiliary Learning for AI-Generated Image Detection

DGX agent

arXiv:2605.22751v1 Announce Type: new Abstract: As generative image models evolve rapidly, the perceptual gap between generated and real images continues to narrow, making AI-generated image detection

applicationsarxiv-cs-cv
22 May 2026
← Previous
1…172173174175176…256
Next →