AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,429
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,935
  • Model Releases23,918
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,429
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,935
  • Model Releases23,918
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
88,429Total entries
1Added by human
88,428Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
51,935 results
Local Ai

VistaRef: Boosting Visual Spatial Orientation Awareness for Pointing-to-Object Detection

DGX agent

arXiv:2606.24498v1 Announce Type: new Abstract: Grounding deictic gestures in natural images is fundamental to AR and human-robot collaboration, providing a basis for seamless spatial interaction. Whi

local-aiarxiv-cs-cv
24 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

A Generalized Formalism of Auto-Regressive Decoding for Speech Processing

DGX agent

arXiv:2606.20714v1 Announce Type: cross Abstract: In speech processing, most state-of-the-art sequence prediction models rely on auto-regressive (AR) strategies to generate output sequences based on t

researcharxiv-cs-lg
23 Jun 2026
Model Releases

A Hybrid, Multi-Layered Pipeline for Phishing and Threat Classification: Independently Validated URL and NLP Engines with a Calibrated Multi-Channel Fusion Stage

DGX agent

arXiv:2606.21690v1 Announce Type: cross Abstract: Phishing is a multi-modal threat. We present a hybrid pipeline that scores each modality with its own engine and fuses the results. Three engines are

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

A Latent Representation Learning Framework for Hyperspectral Image Emulation in Remote Sensing

DGX agent

arXiv:2603.21911v2 Announce Type: replace Abstract: Synthetic hyperspectral image (HSI) generation is essential for large-scale simulation, algorithm development, and mission design, yet traditional r

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

A Smart Classroom Behavior Analysis Framework with a New Highly Congested Classroom Dataset

DGX agent

arXiv:2606.21568v1 Announce Type: new Abstract: Student behavior detection is important for intelligent classroom analysis but remains challenging in large-class scenarios due to dense instance co-occ

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

A Standard Processing Pipeline for High-accuracy Measurement of Few-shot Regression on Laser Induced Breakdown Spectroscopy

DGX agent

arXiv:2606.21960v1 Announce Type: new Abstract: Laser-induced breakdown spectroscopy (LIBS) faces challenges in high-accuracy quantitative measurement under few-shot scenarios due to spectral noise an

model-releasesarxiv-cs-lg
23 Jun 2026
Safety

A Stitch in Time Saves Nine: Preserving Policy Compatibility Under Perception Updates in End-to-End Autonomous Driving

DGX agent

arXiv:2606.21509v1 Announce Type: new Abstract: End-to-end autonomous driving systems tightly couple perception and decision-making through latent representations. Consequently, updates to perception

safetyarxiv-cs-ro
23 Jun 2026
Hardware

An Analysis of Untrained Deep Reservoir Networks for Audio Surveillance

DGX agent

arXiv:2606.22218v1 Announce Type: cross Abstract: In this paper, we investigate untrained recurrent models from the Reservoir Computing (RC) paradigm for audio surveillance, focusing on bidirectional

hardwarearxiv-cs-lg
23 Jun 2026
Applications

An Efficient and Effective Architecture for Large-Scale Traffic Prediction via Geometry-Adaptive Square Partitioning

DGX agent

arXiv:2606.21072v1 Announce Type: new Abstract: Traffic prediction is a core task in intelligent transportation systems and urban-scale decision making. Despite the effectiveness of mainstream neural-

applicationsarxiv-cs-lg
23 Jun 2026
Model Releases

Anticipating the Optimism Gap: Predicting Distribution-Shift Degradation of RF-Impairment Detectors from In-Distribution Statistics

DGX agent

arXiv:2606.22054v1 Announce Type: cross Abstract: Detectors for GNSS radio-frequency impairments (jamming, spoofing, multipath) are usually reported with a single AUC measured on the distribution they

model-releasesarxiv-cs-lg
23 Jun 2026
Research

Boundary-by-Mask: Few-Shot Instance Segmentation with Mask-Conditioned Boundary Learning for Texture-Poor Industrial Parts

DGX agent

arXiv:2606.21594v1 Announce Type: new Abstract: Recent advances in large pre-trained models have led to remarkable progress in instance segmentation on general images. However, industrial scenarios re

researcharxiv-cs-cv
23 Jun 2026
Model Releases

BranchShine: Compact Raw-Audio-to-IPA Transcription with a RoPE E-Branchformer Encoder

DGX agent

arXiv:2606.22824v1 Announce Type: new Abstract: Speech-to-IPA transcription is useful when the desired output is pronunciation rather than orthographic text, but competitive multilingual systems are o

model-releasesarxiv-cs-lg
23 Jun 2026
Research

Category-Adaptive Cross-Modal Semantic Refinement and Transfer for Open-Vocabulary Multi-Label Recognition

DGX agent

arXiv:2412.06190v2 Announce Type: replace Abstract: Benefiting from the generalization capability of CLIP, recent vision language pre-training (VLP) models have demonstrated the ability to capture a w

researcharxiv-cs-cv
23 Jun 2026
Safety

CFPO: Counterfactual Policy Optimization for Multimodal Reasoning

DGX agent

arXiv:2606.23206v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) have demonstrated remarkable capabilities in multimodal reasoning. However, prevailing reinforcement learning (RL)

safetyarxiv-cs-cv
23 Jun 2026
Research

Cloak: Zero-Shot Cross-Embodiment Manipulation by Masking the End-Effector from the VLA

DGX agent

arXiv:2606.22836v1 Announce Type: new Abstract: We present Cloak, a training recipe that endows a Vision-Language-Action (VLA) model with zero-shot cross-embodiment transfer by cloaking the end-effect

researcharxiv-cs-ro
23 Jun 2026
Model Releases

CodePercept: Code-Grounded Visual STEM Perception for MLLMs

DGX agent

arXiv:2603.10757v2 Announce Type: replace Abstract: When MLLMs fail at Science, Technology, Engineering, and Mathematics (STEM) visual reasoning, a fundamental question arises: is it due to perceptual

model-releasesarxiv-cs-cv
23 Jun 2026
Local Ai

CoDMD: Copula-aware Distribution Matching Distillation for Fast Video Generation

DGX agent

arXiv:2606.21982v1 Announce Type: new Abstract: Few-step distillation for video diffusion models has attracted significant attention, driven by the urgent demand for efficient deployment in real-world

local-aiarxiv-cs-cv
23 Jun 2026
Research

Coherence Under Commitment: Probing Generalization and Vacuous Memorization in LLM Logical Reasoning

DGX agent

arXiv:2606.21083v1 Announce Type: cross Abstract: Large language models (LLMs) deployed for logical reasoning in knowledge-intensive domains exhibit a subtle but critical failure: coherence can be vac

researcharxiv-cs-lg
23 Jun 2026
Research

Data Selection Through Iterative Self-Filtering for Vision-Language Settings

DGX agent

arXiv:2606.23611v1 Announce Type: new Abstract: The availability of large amounts of clean data is paramount to training neural networks. However, at large scales, manual oversight is impractical, res

researcharxiv-cs-cv
23 Jun 2026
Safety

Distribution-Aware Diffusion-LLM for Robust Ultra-Long-Term Time Series Forecasting

DGX agent

arXiv:2606.23391v1 Announce Type: new Abstract: Time series forecasting is a fundamental machine learning task. Recent work has explored Large Language Models (LLMs) for this purpose due to their stro

safetyarxiv-cs-lg
23 Jun 2026
Model Releases

Do Location Encoders Capture Spatial Effects? A GeoShapley Benchmark Across Scales

DGX agent

arXiv:2606.23453v1 Announce Type: new Abstract: Location encoders transform geographic coordinates into high dimensional embeddings for downstream machine learning, but it is unclear how well these re

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

DR-Mamba: Automatic Inference-Time Domain Adaptation for Document Image Binarization via Sample-Conditioned Detail-Background Suppression

DGX agent

arXiv:2606.22625v1 Announce Type: new Abstract: Degraded document image binarization is sensitive to domain shifts caused by paper aging, bleed-through, stains, shadows, and uneven illumination, and t

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

DrivingVoxels: Compositional Sparse Voxel Rasterization for Dynamic Driving Scene Reconstruction

DGX agent

arXiv:2606.23031v1 Announce Type: new Abstract: Reconstructing dynamic urban scenes remains challenging due to the unbounded nature of driving environments and the presence of multiple dynamic objects

model-releasesarxiv-cs-cv
23 Jun 2026
Local Ai

Enabling Cloud-Level Accuracy in Edge AI through IoT Data Preprocessing

DGX agent

arXiv:2606.22496v1 Announce Type: cross Abstract: Large language models (LLMs) offer a natural-language interface for interpreting Internet of Things (IoT) sensor data in smart environments; however,

local-aiarxiv-cs-lg
23 Jun 2026
Agents

Enhancing Creativity in 3D Generative Design via a TRIZ-Inspired Text-to-CAD Framework

DGX agent

arXiv:2606.21378v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) have demonstrated significant potential in supporting engineering design tasks, including computer-aided

agentsarxiv-cs-lg
23 Jun 2026
Model Releases

ENVS: Environment-Native Verified Search for Long-Horizon GUI Agents

DGX agent

arXiv:2606.22948v1 Announce Type: cross Abstract: As multimodal agents move from interface understanding to real software control, successful trajectory discovery in live desktop environments becomes

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Explore-Execute Chain: Towards an Efficient Structured Reasoning Paradigm

DGX agent

arXiv:2509.23946v3 Announce Type: replace Abstract: Many LLMs plan before they act, yet planning and execution are often still entangled in one long generation trace, enforced only through prompts, or

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Factor-Aware Mixture-of-Experts with Pretrained Encoder for Combinatorial Generalization

DGX agent

arXiv:2606.21100v1 Announce Type: new Abstract: The integration of pretrained encoders with diffusion policies has become a dominant paradigm for visual robotic manipulation. However, it still struggl

model-releasesarxiv-cs-ro
23 Jun 2026
Model Releases

Factored Gossip DiLoCo: Reducing Blocking Communication in DiLoCo

DGX agent

arXiv:2606.22768v1 Announce Type: new Abstract: To make large-scale distributed training practical outside high-bandwidth datacenters, we must reduce blocking, high-volume synchronization. While DiLoC

model-releasesarxiv-cs-lg
23 Jun 2026
Research

Faithful Grounded Visual Reasoning via Learned Proxy-Tokens

DGX agent

arXiv:2606.23354v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable success in Visual Question Answering (VQA), yet their 'black-box' nature hinders deplo

researcharxiv-cs-cv
23 Jun 2026
Model Releases

FirstPass: Grounding AI Scientific Judgment in Multi-Round Editorial Outcomes

DGX agent

arXiv:2606.20769v1 Announce Type: cross Abstract: AI systems for peer review fail on three fronts: they train on Computer Science and Machine Learning venues alone, ignore the iterative dialogue that

model-releasesarxiv-cs-lg
23 Jun 2026
Applications

FLFL: Federated Latent Factor Learning for Private Recovery of Spatio-Temporal Signals

DGX agent

arXiv:2606.23091v1 Announce Type: new Abstract: Wireless sensor network (WSNs) stands out as a burgeoning and promising domain in intelligent sensing. Owing to various factors such as sudden sensor ma

applicationsarxiv-cs-lg
23 Jun 2026
Safety

FOCA: Future-Oriented Conditioning for Data-Efficient Vision-Language-Action Adaptation

DGX agent

arXiv:2606.20867v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models enable general-purpose robotic control via large-scale multimodal pretraining, yet their effectiveness under few-sho

safetyarxiv-cs-cv
23 Jun 2026
Research

From Markov to Laplace: How Mamba In-Context Learns Markov Chains

DGX agent

arXiv:2502.10178v2 Announce Type: replace Abstract: While transformer-based language models have driven the AI revolution thus far, their computational complexity has spurred growing interest in viabl

researcharxiv-cs-lg
23 Jun 2026
Research

Full Nonlinear Nonholonomic Dynamics and Motion Analysis of a 3-DoF Underactuated Spherical Rolling Robot

DGX agent

arXiv:2606.22169v1 Announce Type: new Abstract: This paper presents a full nonlinear constrained dynamic model of MonoRollBot, a novel 3-DoF spherical rolling robot driven by a single motor, a lead-sc

researcharxiv-cs-ro
23 Jun 2026
Model Releases

GeoRouteNet: Geometry-Enhanced Non-Autoregressive Neural Solver for the Traveling Salesman Problem

DGX agent

arXiv:2606.22776v1 Announce Type: new Abstract: The traveling salesman problem (TSP) is a canonical NP-hard combinatorial optimization benchmark that tests the representational capacity and generaliza

model-releasesarxiv-cs-lg
23 Jun 2026
Safety

HilDA: Hierarchical Distillation with Diffusion for Advancing Self-Supervised LiDAR Pre-training

DGX agent

arXiv:2606.20189v2 Announce Type: replace Abstract: Leveraging Vision Foundation Models (VFMs) for camera-to-LiDAR knowledge distillation offers a promising solution to the scarcity of annotated data

safetyarxiv-cs-cv
23 Jun 2026
Model Releases

HUGE-Bench: A Benchmark for High-Level UAV Vision-Language-Action Tasks

DGX agent

arXiv:2603.19822v2 Announce Type: replace Abstract: Existing UAV vision-language navigation (VLN) benchmarks have enabled language-guided flight, but they largely focus on long, step-wise route descri

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

IMAGIN-4D: Image-Guided Controllable Interaction Generation

DGX agent

arXiv:2606.23675v1 Announce Type: new Abstract: Generating human-object interactions (HOI) is central to character animation, robotics, AR/VR, and embodied AI. Recent HOI generation methods synthesize

model-releasesarxiv-cs-cv
23 Jun 2026
Safety

Integrating Facial Generation into Full-Duplex Spoken Dialogue Systems

DGX agent

arXiv:2606.21970v1 Announce Type: cross Abstract: Full-duplex spoken dialogue models, such as Moshi, enable natural, low-latency voice conversations. However, they remain limited to the audio modality

safetyarxiv-cs-cv
23 Jun 2026
Research

Interpretable Kolmogorov-Arnold Network with Feature-Isolated Temporal Attention Mechanism for Electricity Load Forecasting

DGX agent

arXiv:2606.23425v1 Announce Type: new Abstract: Accurate electricity load forecasting is a crucial prerequisite for stable power system operations. While prevalent deep learning models present competi

researcharxiv-cs-lg
23 Jun 2026
Model Releases

Jury Duty: Calibration and Orientation Failures in MLLM-as-a-Judge Under Cultural Ambiguity

DGX agent

arXiv:2606.20676v1 Announce Type: new Abstract: MLLM-as-a-Judge is conventionally validated by agreement with human annotations, but this metric is undefined when the human pool is culturally heteroge

model-releasesarxiv-cs-cv
23 Jun 2026
Tutorials

Leveraging AutoML for Sustainable Deep Learning: A Multi-Objective HPO Approach on Deep Shift Neural Networks

DGX agent

arXiv:2606.23208v1 Announce Type: new Abstract: Deep Learning (DL) has advanced various fields by extracting complex patterns from large datasets. However, the computational demands of DL models pose

tutorialsarxiv-cs-lg
23 Jun 2026
Model Releases

LOGOS: LiDAR-Only Gaussian Elevation Splatting for Unified Tiny Obstacle Segmentation

DGX agent

arXiv:2606.21527v1 Announce Type: cross Abstract: Robust obstacle segmentation is essential for the safety of intelligent robots, where LiDAR-based perception systems play a fundamental role in the ro

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Mat-Pref: Verifiable-Reward Training Improves Compositional Reasoning in Inorganic Materials

DGX agent

arXiv:2606.21830v1 Announce Type: new Abstract: Reinforcement learning from verifiable rewards (RLVR) has driven rapid progress in mathematical and code reasoning, but when extended to science, existi

model-releasesarxiv-cs-lg
23 Jun 2026
Local Ai

MIRAGE: Stealthy Visual Prompt Injection for Vulnerability Detection in Web Agents

DGX agent

arXiv:2606.20717v1 Announce Type: new Abstract: Multimodal Large Language Model (MLLM)-based web agents provide practical, high-precision solutions for visual browser automation; however, they inheren

local-aiarxiv-cs-cv
23 Jun 2026
Research

Mixture-of-Experts Graph Transformers for Interpretable Particle Collision Detection

DGX agent

arXiv:2501.03432v3 Announce Type: replace Abstract: The Large Hadron Collider at CERN produces immense volumes of complex data from high-energy particle collisions, demanding sophisticated analytical

researcharxiv-cs-lg
23 Jun 2026
Model Releases

Morphology-Aware Multimodal Representation Learning for Insect Phylogenetic Reconstruction

DGX agent

arXiv:2606.22077v1 Announce Type: new Abstract: Morphological traits provide important evidence for phylogenetic reconstruction and evolutionary relationship analysis. Recent image-based approaches ha

model-releasesarxiv-cs-cv
23 Jun 2026
← Previous
1…556557558559560…1082
Next →