AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “papers”

GridTimelineEvolution
12,151 results
10 Apr 2026

Hybrid CNN-Transformer Architecture for Arabic Speech Emotion Recognition

ResearchDGX agent

arXiv:2604.07357v1 Announce Type: new Abstract: Recognizing emotions from speech using machine learning has become an active research area due to its importance in building human-centered applications

Inference-Time Code Selection via Symbolic Equivalence Partitioning

ResearchDGX agent

arXiv:2604.06485v1 Announce Type: cross Abstract: 'Best-of-N' selection is a popular inference-time scaling method for code generation using Large Language Models (LLMs). However, to reliably identify

Information as Structural Alignment: A Dynamical Theory of Continual Learning

Model ReleasesDGX agent

arXiv:2604.07108v1 Announce Type: cross Abstract: Catastrophic forgetting is not an engineering failure. It is a mathematical consequence of storing knowledge as global parameter superposition. Existi

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Karma Mechanisms for Decentralised, Cooperative Multi Agent Path Finding

SafetyDGX agent

arXiv:2604.07970v1 Announce Type: cross Abstract: Multi-Agent Path Finding (MAPF) is a fundamental coordination problem in large-scale robotic and cyber-physical systems, where multiple agents must co

Knowledge Graphs Generation from Cultural Heritage Texts: Combining LLMs and Ontological Engineering for Scholarly Debates

Model ReleasesDGX agent

arXiv:2511.10354v1 Announce Type: cross Abstract: Cultural Heritage texts contain rich knowledge that is difficult to query systematically due to the challenges of converting unstructured discourse in

Lang2Act: Fine-Grained Visual Reasoning through Self-Emergent Linguistic Toolchains

ResearchDGX agent

arXiv:2602.13235v2 Announce Type: replace-cross Abstract: Visual Retrieval-Augmented Generation (VRAG) enhances Vision-Language Models (VLMs) by incorporating external visual documents to address a gi

Latent Anomaly Knowledge Excavation: Unveiling Sparse Sensitive Neurons in Vision-Language Models

ResearchDGX agent

arXiv:2604.07802v1 Announce Type: new Abstract: Large-scale vision-language models (VLMs) exhibit remarkable zero-shot capabilities, yet the internal mechanisms driving their anomaly detection (AD) pe

LLM-based Schema-Guided Extraction and Validation of Missing-Person Intelligence from Heterogeneous Data Sources

SafetyDGX agent

arXiv:2604.06571v1 Announce Type: cross Abstract: Missing-person and child-safety investigations rely on heterogeneous case documents, including structured forms, bulletin-style posters, and narrative

Looking Beyond the Obvious: A Survey on Abstract Concept Recognition for Video Understanding

ResearchDGX agent

arXiv:2508.20765v2 Announce Type: replace-cross Abstract: The automatic understanding of video content is advancing rapidly. Empowered by deeper neural networks and large datasets, machines are increa

LoRA-DA: Data-Aware Initialization for Low-Rank Adaptation via Asymptotic Analysis

Model ReleasesDGX agent

arXiv:2510.24561v2 Announce Type: replace-cross Abstract: LoRA has become a widely adopted method for PEFT, and its initialization methods have attracted increasing attention. However, existing method

Lost in the Hype: Revealing and Dissecting the Performance Degradation of Medical Multimodal Large Language Models in Image Classification

ResearchDGX agent

arXiv:2604.08333v1 Announce Type: new Abstract: The rise of multimodal large language models (MLLMs) has sparked an unprecedented wave of applications in the field of medical imaging analysis. However

LumiCtrl : Learning Illuminant Prompts for Lighting Control in Personalized Text-to-Image Models

ResearchDGX agent

arXiv:2512.17489v2 Announce Type: replace Abstract: Text-to-image (T2I) models have demonstrated remarkable progress in creative image generation, yet they still lack precise control over scene illumi

Luwen Technical Report

ApplicationsDGX agent

arXiv:2604.06737v1 Announce Type: cross Abstract: Large language models have demonstrated remarkable capabilities across a wide range of natural language processing tasks, yet their application in the

MARCH: Evaluating the Intersection of Ambiguity Interpretation and Multi-hop Inference

Model ReleasesDGX agent

arXiv:2509.22750v3 Announce Type: replace Abstract: Real-world multi-hop QA is naturally linked with ambiguity, where a single query can trigger multiple reasoning paths that require independent resol

MedRoute: RL-Based Dynamic Specialist Routing in Multi-Agent Medical Diagnosis

AgentsDGX agent

arXiv:2604.06180v1 Announce Type: cross Abstract: Medical diagnosis using Large Multimodal Models (LMMs) has gained increasing attention due to capability of these models in providing precise diagnose

MemReader: From Passive to Active Extraction for Long-Term Agent Memory

SafetyDGX agent

arXiv:2604.07877v1 Announce Type: new Abstract: Long-term memory is fundamental for personalized and autonomous agents, yet populating it remains a bottleneck. Existing systems treat memory extraction

MF-GLaM: A multifidelity stochastic emulator using generalized lambda models

Model ReleasesDGX agent

arXiv:2507.10303v2 Announce Type: replace-cross Abstract: Stochastic simulators exhibit intrinsic stochasticity due to unobservable, uncontrollable, or unmodeled input variables, resulting in random o

Mixture Proportion Estimation and Weakly-supervised Kernel Test for Conditional Independence

SafetyDGX agent

arXiv:2604.07191v1 Announce Type: cross Abstract: Mixture proportion estimation (MPE) aims to estimate class priors from unlabeled data. This task is a critical component in weakly supervised learning

Modernizing Amdahl's Law: How AI Scaling Laws Shape Computer Architecture

ResearchDGX agent

arXiv:2603.20654v4 Announce Type: replace-cross Abstract: Classical Amdahl's Law conceptualized the limit of speedup for an era of fixed serial-parallel decomposition and homogeneous replication. Mode

Monocular Depth Estimation From the Perspective of Feature Restoration: A Diffusion Enhanced Depth Restoration Approach

Model ReleasesDGX agent

arXiv:2604.07664v1 Announce Type: new Abstract: Monocular Depth Estimation (MDE) is a fundamental computer vision task with important applications in 3D vision. The current mainstream MDE methods empl

Multi-modal user interface control detection using cross-attention

ResearchDGX agent

arXiv:2604.06934v1 Announce Type: cross Abstract: Detecting user interface (UI) controls from software screenshots is a critical task for automated testing, accessibility, and software analytics, yet

NaviSlim: Adaptive Context-Aware Navigation and Sensing via Dynamic Slimmable Networks

AgentsDGX agent

arXiv:2407.01563v2 Announce Type: replace Abstract: Small-scale autonomous airborne vehicles, such as micro-drones, are expected to be a central component of a broad spectrum of applications ranging f

NaviSplit: Dynamic Multi-Branch Split DNNs for Efficient Distributed Autonomous Navigation

AgentsDGX agent

arXiv:2406.13086v2 Announce Type: replace Abstract: Lightweight autonomous unmanned aerial vehicles (UAV) are emerging as a central component of a broad range of applications. However, autonomous navi

NestPipe: Large-Scale Recommendation Training on 1,500+ Accelerators via Nested Pipelining

HardwareDGX agent

arXiv:2604.06956v1 Announce Type: cross Abstract: Modern recommendation models have increased to trillions of parameters. As cluster scales expand to O(1k), distributed training bottlenecks shift from

Noise Immunity in In-Context Tabular Learning: An Empirical Robustness Analysis of TabPFN's Attention Mechanisms

Model ReleasesDGX agent

arXiv:2604.04868v2 Announce Type: replace-cross Abstract: Tabular foundation models (TFMs) such as TabPFN (Tabular Prior-Data Fitted Network) are designed to generalize across heterogeneous tabular da

OceanMAE: A Foundation Model for Ocean Remote Sensing

TutorialsDGX agent

arXiv:2604.08171v1 Announce Type: new Abstract: Accurate ocean mapping is essential for applications such as bathymetry estimation, seabed characterization, marine litter detection, and ecosystem moni

On-Policy Distillation of Language Models for Autonomous Vehicle Motion Planning

Model ReleasesDGX agent

arXiv:2604.07944v1 Announce Type: new Abstract: Large language models (LLMs) have recently demonstrated strong potential for autonomous vehicle motion planning by reformulating trajectory prediction a

On the Global Photometric Alignment for Low-Level Vision

SafetyDGX agent

arXiv:2604.08172v1 Announce Type: new Abstract: Supervised low-level vision models rely on pixel-wise losses against paired references, yet paired training sets exhibit per-pair photometric inconsiste

OVS-DINO: Open-Vocabulary Segmentation via Structure-Aligned SAM-DINO with Language Guidance

SafetyDGX agent

arXiv:2604.08461v1 Announce Type: new Abstract: Open-Vocabulary Segmentation (OVS) aims to segment image regions beyond predefined category sets by leveraging semantic descriptions. While CLIP based a

PD-SOVNet: A Physics-Driven Second-Order Vibration Operator Network for Estimating Wheel Polygonal Roughness from Axle-Box Vibrations

ApplicationsDGX agent

arXiv:2604.06620v1 Announce Type: new Abstract: Quantitative estimation of wheel polygonal roughness from axle-box vibration signals is a challenging yet practically relevant problem for rail-vehicle

Physical Adversarial Attacks on AI Surveillance Systems:Detection, Tracking, and Visible--Infrared Evasion

ResearchDGX agent

arXiv:2604.06865v1 Announce Type: cross Abstract: Physical adversarial attacks are increasingly studied in settings that resemble deployed surveillance systems rather than isolated image benchmarks. I

Pistachio: Towards Synthetic, Balanced, and Long-Form Video Anomaly Benchmarks

Model ReleasesDGX agent

arXiv:2511.19474v4 Announce Type: replace-cross Abstract: Automatically detecting abnormal events in videos is crucial for modern autonomous systems, yet existing Video Anomaly Detection (VAD) benchma

PixelCAM: Pixel Class Activation Mapping for Histology Image Classification and ROI Localization

Local AiDGX agent

arXiv:2503.24135v3 Announce Type: replace Abstract: Weakly supervised object localization (WSOL) methods allow training models to classify images and localize ROIs. WSOL only requires low-cost image-c

Planning Task Shielding: Detecting and Repairing Flaws in Planning Tasks through Turning them Unsolvable

ResearchDGX agent

arXiv:2604.07042v1 Announce Type: new Abstract: Most research in planning focuses on generating a plan to achieve a desired set of goals. However, a goal specification can also be used to encode a pro

Planning with Minimal Disruption

ResearchDGX agent

arXiv:2508.15358v2 Announce Type: replace Abstract: In many planning applications, we might be interested in finding plans that minimally modify the initial state to achieve the goals. We refer to thi

Predictive Representations for Skill Transfer in Reinforcement Learning

AgentsDGX agent

arXiv:2604.07016v1 Announce Type: new Abstract: A key challenge in scaling up Reinforcement Learning is generalizing learned behaviour. Without the ability to carry forward acquired knowledge an agent

Preventing Overfitting in Deep Image Prior for Hyperspectral Image Denoising

ResearchDGX agent

arXiv:2604.08272v1 Announce Type: new Abstract: Deep image prior (DIP) is an unsupervised deep learning framework that has been successfully applied to a variety of inverse imaging problems. However,

ProofSketcher: Hybrid LLM + Lightweight Proof Checker for Reliable Math/Logic Reasoning

ResearchDGX agent

arXiv:2604.06401v1 Announce Type: new Abstract: The large language models (LLMs) might produce a persuasive argument within mathematical and logical fields, although such argument often includes some

Quantitative Estimation of Target Task Performance from Unsupervised Pretext Task in Semi/Self-Supervised Learning

Model ReleasesDGX agent

arXiv:2508.07299v2 Announce Type: replace-cross Abstract: The effectiveness of unlabeled data in Semi/Self-Supervised Learning (SSL) depends on appropriate assumptions for specific scenarios, thereby

Quantum-Inspired Tensor Network Autoencoders for Anomaly Detection: A MERA-Based Approach

Model ReleasesDGX agent

arXiv:2604.06541v1 Announce Type: cross Abstract: We investigate whether a multiscale tensor-network architecture can provide a useful inductive bias for reconstruction-based anomaly detection in coll

Reading Recognition in the Wild

TutorialsDGX agent

arXiv:2505.24848v4 Announce Type: replace Abstract: To enable egocentric contextual AI in always-on smart glasses, it is crucial to be able to keep a record of the user's interactions with the world,

Reason in Chains, Learn in Trees: Self-Rectification and Grafting for Multi-turn Agent Policy Optimization

SafetyDGX agent

arXiv:2604.07165v1 Announce Type: new Abstract: Reinforcement learning for Large Language Model agents is often hindered by sparse rewards in multi-step reasoning tasks. Existing approaches like Group

RectifiedHR: Enable Efficient High-Resolution Synthesis via Energy Rectification

ResearchDGX agent

arXiv:2503.02537v4 Announce Type: replace Abstract: Diffusion models have achieved remarkable progress across various visual generation tasks. However, their performance significantly declines when ge

Rectifying LLM Thought from Lens of Optimization

ResearchDGX agent

arXiv:2512.01925v2 Announce Type: replace-cross Abstract: Recent advancements in large language models (LLMs) have been driven by their emergent reasoning capabilities, particularly through long chain

ReDAct: Uncertainty-Aware Deferral for LLM Agents

AgentsDGX agent

arXiv:2604.07036v1 Announce Type: cross Abstract: Recently, LLM-based agents have become increasingly popular across many applications, including complex sequential decision-making problems. However,

Resource-constrained Amazons chess decision framework integrating large language models and graph attention

ResearchDGX agent

arXiv:2603.10512v2 Announce Type: replace Abstract: Artificial intelligence has advanced significantly through the development of intelligent game-playing systems, providing rigorous testbeds for deci

Revisiting Radar Perception With Spectral Point Clouds

Model ReleasesDGX agent

arXiv:2604.08282v1 Announce Type: new Abstract: Radar perception models are trained with different inputs, from range-Doppler spectra to sparse point clouds. Dense spectra are assumed to outperform sp

RLBoost: Harvesting Preemptible Resources for Cost-Efficient Reinforcement Learning on LLMs

HardwareDGX agent

arXiv:2510.19225v3 Announce Type: replace-cross Abstract: Reinforcement learning (RL) has become essential for unlocking advanced reasoning capabilities in large language models (LLMs). RL workflows i

Robust Multi-Agent Target Tracking in Intermittent Communication Environments via Analytical Belief Merging

AgentsDGX agent

arXiv:2604.07575v1 Announce Type: new Abstract: Autonomous multi-agent target tracking in GPS-denied and communication-restricted environments (e.g., underwater exploration, subterranean search and re

Scientific Knowledge-driven Decoding Constraints Improving the Reliability of LLMs

ResearchDGX agent

arXiv:2604.06603v1 Announce Type: cross Abstract: Large language models (LLMs) have shown strong knowledge reserves and task-solving capabilities, but still face the challenge of severe hallucination,

SD-FSMIS: Adapting Stable Diffusion for Few-Shot Medical Image Segmentation

ResearchDGX agent

arXiv:2604.03134v2 Announce Type: replace Abstract: Few-Shot Medical Image Segmentation (FSMIS) aims to segment novel object classes in medical images using only minimal annotated examples, addressing

Self-Distilled RLVR

SafetyDGX agent

arXiv:2604.03128v2 Announce Type: replace Abstract: On-policy distillation (OPD) has become a popular training paradigm in the LLM community. This paradigm selects a larger model as the teacher to pro

ShadowNPU: System and Algorithm Co-design for NPU-Centric On-Device LLM Inference

Local AiDGX agent

arXiv:2508.16703v4 Announce Type: replace-cross Abstract: On-device running Large Language Models (LLMs) is nowadays a critical enabler towards preserving user privacy. We observe that the attention o

Smart Commander: A Hierarchical Reinforcement Learning Framework for Fleet-Level PHM Decision Optimization

ResearchDGX agent

arXiv:2604.07171v1 Announce Type: new Abstract: Decision-making in military aviation Prognostics and Health Management (PHM) faces significant challenges due to the 'curse of dimensionality' in large-

SpecQuant: Spectral Decomposition and Adaptive Truncation for Ultra-Low-Bit LLMs Quantization

Model ReleasesDGX agent

arXiv:2511.11663v2 Announce Type: replace-cross Abstract: The emergence of accurate open large language models (LLMs) has sparked a push for advanced quantization techniques to enable efficient deploy

State and Trajectory Estimation of Tensegrity Robots via Factor Graphs and Chebyshev Polynomials

ApplicationsDGX agent

arXiv:2604.08185v1 Announce Type: new Abstract: Tensegrity robots offer compliance and adaptability, but their nonlinear, and underconstrained dynamics make state estimation challenging. Reliable cont

Stop Listening to Me! How Multi-turn Conversations Can Degrade LLM Diagnostic Reasoning

ApplicationsDGX agent

arXiv:2603.11394v2 Announce Type: replace Abstract: Patients and clinicians are increasingly using chatbots powered by large language models (LLMs) for healthcare inquiries. While state-of-the-art LLM

Syntax Is Easy, Semantics Is Hard: Evaluating LLMs for LTL Translation

ResearchDGX agent

arXiv:2604.07321v1 Announce Type: cross Abstract: Propositional Linear Temporal Logic (LTL) is a popular formalism for specifying desirable requirements and security and privacy policies for software,

Tabular GANs for uneven distribution

Model ReleasesDGX agent

arXiv:2010.00638v2 Announce Type: replace-cross Abstract: Generative models for tabular data have evolved rapidly beyond Generative Adversarial Networks (GANs). While GANs pioneered synthetic tabular

TalkLoRA: Communication-Aware Mixture of Low-Rank Adaptation for Large Language Models

Model ReleasesDGX agent

arXiv:2604.06291v1 Announce Type: cross Abstract: Low-Rank Adaptation (LoRA) enables parameter-efficient fine-tuning of Large Language Models (LLMs), and recent Mixture-of-Experts (MoE) extensions fur

← Previous
1…200201202203
Next →