AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
Human
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,851 results
23 Apr 2026

Unveiling Uncertainty-Aware Autonomous Cooperative Learning Based Planning Strategy

AgentsDGX agent

arXiv:2510.11041v2 Announce Type: replace Abstract: In future intelligent transportation systems, autonomous cooperative planning (ACP), becomes a promising technique to increase the effectiveness and

Using Learning Theories to Evolve Human-Centered XAI: Future Perspectives and Challenges

ResearchDGX agent

arXiv:2604.19788v1 Announce Type: new Abstract: As Artificial Intelligence (AI) systems continue to grow in size and complexity, so does the difficulty of the quest for AI transparency. In a world of

Utterance-Level Methods for Identifying Reliable ASR-Output for Child Speech

ResearchDGX agent

arXiv:2604.19801v1 Announce Type: cross Abstract: Automatic Speech Recognition (ASR) is increasingly used in applications involving child speech, such as language learning and literacy acquisition. Ho

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

UVIO: An UWB-Aided Visual-Inertial Odometry Framework with Bias-Compensated Anchors Initialization

SafetyDGX agent

arXiv:2308.00513v2 Announce Type: replace Abstract: This paper introduces UVIO, a multi-sensor framework that leverages Ultra Wide Band (UWB) technology and Visual-Inertial Odometry (VIO) to provide r

V-tableR1: Process-Supervised Multimodal Table Reasoning with Critic-Guided Policy Optimization

SafetyDGX agent

arXiv:2604.20755v1 Announce Type: new Abstract: We introduce V-tableR1, a process-supervised reinforcement learning framework that elicits rigorous, verifiable reasoning from multimodal large language

VAN-AD: Visual Masked Autoencoder with Normalizing Flow For Time Series Anomaly Detection

ApplicationsDGX agent

arXiv:2603.26842v2 Announce Type: replace-cross Abstract: Time series anomaly detection (TSAD) is essential for maintaining the reliability and security of IoT-enabled service systems. Existing method

Variance Is Not Importance: Structural Analysis of Transformer Compressibility Across Model Scales

Model ReleasesDGX agent

arXiv:2604.20682v1 Announce Type: new Abstract: We present a systematic empirical study of transformer compression through over 40 experiments on GPT-2 (124M parameters) and Mistral 7B (7.24B paramete

Verification of Machine Unlearning is Fragile

SafetyDGX agent

arXiv:2408.00929v2 Announce Type: replace Abstract: As privacy concerns escalate in the realm of machine learning, data owners now have the option to utilize machine unlearning to remove their data fr

veScale-FSDP: Flexible and High-Performance FSDP at Scale

ResearchDGX agent

arXiv:2602.22437v3 Announce Type: replace-cross Abstract: Fully Sharded Data Parallel (FSDP), also known as Zero Redundancy Optimizer (ZeRO), is widely used for large-scale model training, because of

Vibrotactile Preference Learning: Uncertainty-Aware Preference Learning for Personalized Vibration Feedback

Model ReleasesDGX agent

arXiv:2604.20210v1 Announce Type: cross Abstract: Individual differences in vibrotactile perception underscore the growing importance of personalization as haptic feedback becomes more prevalent in in

Video-ToC: Video Tree-of-Cue Reasoning

Model ReleasesDGX agent

arXiv:2604.20473v1 Announce Type: new Abstract: Existing Video Large Language Models (Video LLMs) struggle with complex video understanding, exhibiting limited reasoning capabilities and potential hal

Visual Reasoning through Tool-supervised Reinforcement Learning

AgentsDGX agent

arXiv:2604.19945v1 Announce Type: new Abstract: In this paper, we investigate the problem of how to effectively master tool-use to solve complex visual reasoning tasks for Multimodal Large Language Mo

Visual-Tactile Peg-in-Hole Assembly Learning from Peg-out-of-Hole Disassembly

SafetyDGX agent

arXiv:2604.20712v1 Announce Type: new Abstract: Peg-in-hole (PiH) assembly is a fundamental yet challenging robotic manipulation task. While reinforcement learning (RL) has shown promise in tackling s

VTouch++: A Multimodal Dataset with Vision-Based Tactile Enhancement for Bimanual Manipulation

ApplicationsDGX agent

arXiv:2604.20444v1 Announce Type: cross Abstract: Embodied intelligence has advanced rapidly in recent years; however, bimanual manipulation-especially in contact-rich tasks remains challenging. This

Wan-Image: Pushing the Boundaries of Generative Visual Intelligence

ApplicationsDGX agent

arXiv:2604.19858v1 Announce Type: new Abstract: We present Wan-Image, a unified visual generation system explicitly engineered to paradigm-shift image generation models from casual synthesizers into p

WebGen-R1: Incentivizing Large Language Models to Generate Functional and Aesthetic Websites with Reinforcement Learning

Model ReleasesDGX agent

arXiv:2604.20398v1 Announce Type: new Abstract: While Large Language Models (LLMs) excel at function-level code generation, project-level tasks such as generating functional and visually aesthetic mul

Weighted Knowledge Distillation for Semi-Supervised Segmentation of Maxillary Sinus in Panoramic X-ray Images

ResearchDGX agent

arXiv:2604.20213v1 Announce Type: new Abstract: Accurate segmentation of maxillary sinus in panoramic X-ray images is essential for dental diagnosis and surgical planning; however, this task remains r

What Language Models Know But Don't Say: Non-Generative Prior Extraction for Generalization

ApplicationsDGX agent

arXiv:2601.17609v2 Announce Type: replace Abstract: In domains like medicine and finance, large-scale labeled data is costly and often unavailable, leading to models trained on small datasets that str

What Makes a Bacterial Model a Good Reservoir Computer? Predicting Performance from Separability and Similarity

ResearchDGX agent

arXiv:2604.19850v1 Announce Type: cross Abstract: Biological systems are promising substrates for computation because they naturally process environmental information through complex internal dynamics

What Makes a Good AI Review? Concern-Level Diagnostics for AI Peer Review

SafetyDGX agent

arXiv:2604.19998v1 Announce Type: new Abstract: Evaluating AI-generated reviews by verdict agreement is widely recognized as insufficient, yet current alternatives rarely audit which concerns a system

Where and What: Reasoning Dynamic and Implicit Preferences in Situated Conversational Recommendation

SafetyDGX agent

arXiv:2604.20749v1 Announce Type: new Abstract: Situated conversational recommendation (SCR), which utilizes visual scenes grounded in specific environments and natural language dialogue to deliver co

Where are they looking in the operating room?

ResearchDGX agent

arXiv:2604.20574v1 Announce Type: new Abstract: Purpose: Gaze-following, the task of inferring where individuals are looking, has been widely studied in computer vision, advancing research in visual a

Where Reasoning Breaks: Logic-Aware Path Selection by Controlling Logical Connectives in LLMs Reasoning Chains

Local AiDGX agent

arXiv:2604.20564v1 Announce Type: new Abstract: While LLMs demonstrate impressive reasoning capabilities, they remain fragile in multi-step logical deduction, where a single transition error can propa

Which Reasoning Trajectories Teach Students to Reason Better? A Simple Metric of Informative Alignment

SafetyDGX agent

arXiv:2601.14249v4 Announce Type: replace Abstract: Long chain-of-thought (CoT) trajectories provide rich supervision signals for distilling reasoning from teacher to student LLMs. However, both prior

White-Basilisk: A Hybrid Model for Code Vulnerability Detection

Model ReleasesDGX agent

arXiv:2507.08540v5 Announce Type: replace-cross Abstract: The proliferation of software vulnerabilities presents a significant challenge to cybersecurity, necessitating more effective detection method

Whose Story Gets Told? Positionality and Bias in LLM Summaries of Life Narratives

SafetyDGX agent

arXiv:2604.20131v1 Announce Type: new Abstract: Increasingly, studies are exploring using Large Language Models (LLMs) for accelerated or scaled qualitative analysis of text data. While we can compare

Why AI-Generated Text Detection Fails: Evidence from Explainable AI Beyond Benchmark Accuracy

Model ReleasesDGX agent

arXiv:2603.23146v2 Announce Type: replace-cross Abstract: The widespread adoption of Large Language Models (LLMs) has made the detection of AI-Generated text a pressing and complex challenge. Although

WildFireVQA: A Large-Scale Radiometric Thermal VQA Benchmark for Aerial Wildfire Monitoring

Model ReleasesDGX agent

arXiv:2604.20190v1 Announce Type: new Abstract: Wildfire monitoring requires timely, actionable situational awareness from airborne platforms, yet existing aerial visual question answering (VQA) bench

WISCA: A Lightweight Model Transition Method to Improve LLM Training via Weight Scaling

ResearchDGX agent

arXiv:2508.16676v2 Announce Type: replace-cross Abstract: Transformer architecture gradually dominates the LLM field. Recent advances in training optimization for Transformer-based large language mode

WorkflowGen:an adaptive workflow generation mechanism driven by trajectory experience

Model ReleasesDGX agent

arXiv:2604.19756v1 Announce Type: cross Abstract: Large language model (LLM) agents often suffer from high reasoning overhead, excessive token consumption, unstable execution, and inability to reuse p

Working Memory Constraints Scaffold Learning in Transformers under Data Scarcity

SafetyDGX agent

arXiv:2604.20789v1 Announce Type: cross Abstract: We investigate the integration of human-like working memory constraints into the Transformer architecture and implement several cognitively inspired a

X-Cache: Cross-Chunk Block Caching for Few-Step Autoregressive World Models Inference

AgentsDGX agent

arXiv:2604.20289v1 Announce Type: new Abstract: Real-time world simulation is becoming a key infrastructure for scalable evaluation and online reinforcement learning of autonomous driving systems. Rec

X-IONet: Cross-Platform Inertial Odometry Network for Pedestrian and Legged Robot

ResearchDGX agent

arXiv:2511.08277v2 Announce Type: replace-cross Abstract: Learning-based inertial odometry has achieved remarkable progress in pedestrian navigation. However, extending these methods to quadruped robo

X-PCR: A Benchmark for Cross-modality Progressive Clinical Reasoning in Ophthalmic Diagnosis

Model ReleasesDGX agent

arXiv:2604.20350v1 Announce Type: new Abstract: Despite significant progress in Multi-modal Large Language Models (MLLMs), their clinical reasoning capacity for multi-modal diagnosis remains largely u

22 Apr 2026

3D Foundation Model for Generalizable Disease Detection in Head Computed Tomography

Model ReleasesDGX agent

arXiv:2502.02779v3 Announce Type: replace-cross Abstract: Head computed tomography (CT) imaging is a widely-used imaging modality with multitudes of medical indications, particularly in assessing path

A Bolu: A Structured Dataset for the Computational Analysis of Sardinian Improvisational Poetry

ApplicationsDGX agent

arXiv:2604.19584v1 Announce Type: new Abstract: The growing interest of Natural Language Processing (NLP) in minority languages has not yet bridged the gap in the preservation of oral linguistic herit

A Controlled Benchmark of Visual State-Space Backbones with Domain-Shift and Boundary Analysis for Remote-Sensing Segmentation

Model ReleasesDGX agent

arXiv:2604.18721v1 Announce Type: cross Abstract: Visual state-space models (SSMs) are increasingly promoted as efficient alternatives to Vision Transformers, yet their practical advantages remain unc

A Dual Perspective on Synthetic Trajectory Generators: Utility Framework and Privacy Vulnerabilities

ResearchDGX agent

arXiv:2604.19653v1 Announce Type: new Abstract: Human mobility data are used in numerous applications, ranging from public health to urban planning. Human mobility is inherently sensitive, as it can c

A Functionality-Grounded Benchmark for Evaluating Web Agents in E-commerce Domains

Model ReleasesDGX agent

arXiv:2508.15832v2 Announce Type: replace-cross Abstract: Web agents have shown great promise in performing many tasks on ecommerce website. To assess their capabilities, several benchmarks have been

A Gesture-Based Visual Learning Model for Acoustophoretic Interactions using a Swarm of AcoustoBots

AgentsDGX agent

arXiv:2604.19643v1 Announce Type: new Abstract: AcoustoBots are mobile acoustophoretic robots capable of delivering mid-air haptics, directional audio, and acoustic levitation, but existing implementa

A Heterogeneous Long-Micro Scale Cascading Architecture for General Aviation Health Management

Local AiDGX agent

arXiv:2603.22885v4 Announce Type: replace Abstract: BACKGROUND: General aviation fleet expansion demands intelligent health monitoring under computational constraints. Real-world aircraft health diagn

A-MAR: Agent-based Multimodal Art Retrieval for Fine-Grained Artwork Understanding

Model ReleasesDGX agent

arXiv:2604.19689v1 Announce Type: new Abstract: Understanding artworks requires multi-step reasoning over visual content and cultural, historical, and stylistic context. While recent multimodal large

A Mechanism and Optimization Study on the Impact of Information Density on User-Generated Content Named Entity Recognition

ResearchDGX agent

arXiv:2604.18944v1 Announce Type: new Abstract: Named Entity Recognition (NER) models trained on clean, high-resource corpora exhibit catastrophic performance collapse when deployed on noisy, sparse U

A Multi-Agent Framework with Structured Reasoning and Reflective Refinement for Multimodal Empathetic Response Generation

AgentsDGX agent

arXiv:2604.18988v1 Announce Type: new Abstract: Multimodal empathetic response generation (MERG) aims to generate emotionally engaging and empathetic responses based on users' multimodal contexts. Exi

A Network-Aware Evaluation of Distributed Energy Resource Control in Smart Distribution Systems

ResearchDGX agent

arXiv:2604.19715v1 Announce Type: new Abstract: Distribution networks with high penetration of Distributed Energy Resources (DERs) increasingly rely on communication networks to coordinate grid-intera

A neural operator framework for data-driven discovery of stability and receptivity in physical systems

ResearchDGX agent

arXiv:2604.19465v1 Announce Type: cross Abstract: Understanding how complex systems respond to perturbations, such as whether they will remain stable or what their most sensitive patterns are, is a fu

A PPA-Driven 3D-IC Partitioning Selection Framework with Surrogate Models

ResearchDGX agent

arXiv:2604.18806v1 Announce Type: new Abstract: 3D-IC netlist partitioning is commonly optimized using proxy objectives, while final PPA is treated as a costly evaluation rather than an optimization s

A Proxy Consistency Loss for Grounded Fusion of Earth Observation and Location Encoders

TutorialsDGX agent

arXiv:2604.18881v1 Announce Type: cross Abstract: Supervised learning with Earth observation inputs is often limited by the sparsity of high-quality labeled or in-situ measured data to use as training

A Self-Evolving Framework for Efficient Terminal Agents via Observational Context Compression

AgentsDGX agent

arXiv:2604.19572v1 Announce Type: new Abstract: As model capabilities advance, research has increasingly shifted toward long-horizon, multi-turn terminal-centric agentic tasks, where raw environment f

A Survey on MLLM-based Visually Rich Document Understanding: Methods, Challenges, and Emerging Trends

AgentsDGX agent

arXiv:2507.09861v2 Announce Type: replace-cross Abstract: Visually Rich Document Understanding (VRDU) has become a pivotal area of research, driven by the need to automatically interpret documents tha

AblateCell: A Reproduce-then-Ablate Agent for Virtual Cell Repositories

AgentsDGX agent

arXiv:2604.19606v1 Announce Type: new Abstract: Systematic ablations are essential to attribute performance gains in AI Virtual Cells, yet they are rarely performed because biological repositories are

AC-SINDy: Compositional Sparse Identification of Nonlinear Dynamics

ResearchDGX agent

arXiv:2604.18889v1 Announce Type: new Abstract: We present AC-SINDy, a compositional extension of the Sparse Identification of Nonlinear Dynamics (SINDy) framework that replaces explicit feature libra

Accelerating Optimization and Machine Learning through Decentralization

TutorialsDGX agent

arXiv:2604.19518v1 Announce Type: new Abstract: Decentralized optimization enables multiple devices to learn a global machine learning model while each individual device only has access to its local d

Accelerating trajectory optimization with Sobolev-trained diffusion policies

Local AiDGX agent

arXiv:2604.19011v1 Announce Type: new Abstract: Trajectory Optimization (TO) solvers exploit known system dynamics to compute locally optimal trajectories through iterative improvements. A downside is

Achieving Interaction Fluidity in a Wizard-of-Oz Robotic System: A Prototype for Fluid Error-Correction

ResearchDGX agent

arXiv:2604.19374v1 Announce Type: new Abstract: Achieving truly fluid interaction with robots with speech interfaces remains a hard problem, and the experience of current Human-Robot Interaction (HRI)

AD-Copilot: A Vision-Language Assistant for Industrial Anomaly Detection via Visual In-context Comparison

Model ReleasesDGX agent

arXiv:2603.13779v2 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) have achieved impressive success in natural visual understanding, yet they consistently underperform

AdaGScale: Viewpoint-Adaptive Gaussian Scaling in 3D Gaussian Splatting to Reduce Gaussian-Tile Pairs

HardwareDGX agent

arXiv:2604.18980v1 Announce Type: new Abstract: Reducing the number of Gaussian-tile pairs is one of the most promising approaches to improve 3D Gaussian Splatting (3D-GS) rendering speed on GPUs. How

Adapting Dijkstra for Buffers and Unlimited Transfers

ResearchDGX agent

arXiv:2603.11729v2 Announce Type: replace-cross Abstract: In recent years, RAPTOR based algorithms have been considered the state-of-the-art for path-finding with unlimited transfers without preproces

Adapting Self-Supervised Representations as a Latent Space for Efficient Generation

ResearchDGX agent

arXiv:2510.14630v2 Announce Type: replace Abstract: We introduce Representation Tokenizer (RepTok), a generative modeling framework that represents an image using a single continuous latent token obta

Adaptive MSD-Splitting: Enhancing C4.5 and Random Forests for Skewed Continuous Attributes

ApplicationsDGX agent

arXiv:2604.19722v1 Announce Type: cross Abstract: The discretization of continuous numerical attributes remains a persistent computational bottleneck in the induction of decision trees, particularly a

← Previous
1…868869870871872…998
Next →