AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,042
  • Agents7,457
  • Applications5,325
  • Concepts5
  • Hardware1,802
  • Industry6,143
  • Local Ai4,863
  • Model Releases23,393
  • Research19,837
  • Safety13,180
  • Syntheses17
  • Tools1,671
  • Tutorials3,349

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,042
  • Agents7,457
  • Applications5,325
  • Concepts5
  • Hardware1,802
  • Industry6,143
  • Local Ai4,863
  • Model Releases23,393
  • Research19,837
  • Safety13,180
  • Syntheses17
  • Tools1,671
  • Tutorials3,349

Source
Human
87,042Total entries
1Added by human
87,041Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
61,904 results
29 Jun 2026

Drop-Then-Recovery: How Redundant Are Vision-Language-Action Models?

ResearchDGX agent

arXiv:2606.27755v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models enable instruction-driven robotic manipulation, but they inherit oversized language backbones from pretrained VLMs

Dual-Learning based Penalized Multi-Align Clustering for Multi-View Incomplete and Disorderly Data

SafetyDGX agent

arXiv:2606.27984v1 Announce Type: new Abstract: Multimodal feature fusion can effectively capture complex patterns in real-world data by integrating complementary information from different modalities

DysLexLens: A Low-Resource LLM Framework for Analysing Dyslexic Learners Insights from Online Forums

SafetyDGX agent

arXiv:2606.27619v1 Announce Type: new Abstract: Dyslexic learners increasingly use artificial intelligence (AI) tools to support reading, writing, organisation, and study-related tasks. However, their

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

EchoSonar-R: A Multi-View Reasoning-Enabled Model for Disease Classification and Report Generation in Echocardiography

SafetyDGX agent

arXiv:2606.28164v1 Announce Type: new Abstract: Echocardiography is the most widely used non-invasive cardiac imaging modality, providing essential information for cardiovascular diagnosis. Interpreti

Effects of relational graph modularity and depth on the learning performance of neural networks

ApplicationsDGX agent

arXiv:2507.10005v2 Announce Type: replace Abstract: In recent years, graph-based machine learning techniques, such as reinforcement learning and graph neural networks, have garnered significant attent

Efficient and Stable Multi-Dimensional Kolmogorov-Smirnov Distance

ResearchDGX agent

arXiv:2504.11299v2 Announce Type: replace-cross Abstract: We revisit extending the Kolmogorov-Smirnov distance between probability distributions to the multi-dimensional setting, and make new argument

Elastic Time: Dynamic Frame Rate Bottlenecks for Neural Audio Coding

ResearchDGX agent

arXiv:2606.27320v1 Announce Type: cross Abstract: Neural audio autoencoders have become a core component of compression, feature extraction, and generation. However, while existing systems support var

EMOSH: Expressive Motion and Shape Disentanglement for Human Animation

ResearchDGX agent

arXiv:2606.28026v1 Announce Type: new Abstract: High-fidelity and expressive controllable human animation is essential for content creation and digital avatar applications. However, existing methods f

End-to-End Dynamic Sparsity for Resource-Adaptive LLM Inference

Model ReleasesDGX agent

arXiv:2606.27743v1 Announce Type: cross Abstract: Large Language Models (LLMs) inference is typically deployed under a static resource assumption, where models execute a fixed computational graph rega

Enhanced Neural Video Representation Compression across Extreme Complexity and Quality Scales

ApplicationsDGX agent

arXiv:2606.28163v1 Announce Type: cross Abstract: Implicit neural representations (INRs) have recently emerged as a promising approach to video compression, delivering competitive rate-distortion perf

Enhancing Co-packaging Optics Enabled Silicon Photonics Security Assurance Hardware Fingerprinting

HardwareDGX agent

arXiv:2606.27612v1 Announce Type: cross Abstract: Silicon photonics enables integration of optical components using standard semiconductor processes, greatly improving data communication bandwidth and

Enhancing Numerical Prediction in LLMs via Smooth MMD Alignment

Local AiDGX agent

arXiv:2606.27731v1 Announce Type: cross Abstract: Despite their strong general capabilities, large language models (LLMs) often remain unreliable when outputs must be numerically precise. A key reason

EntMTP: Accelerating LLM Inference with Entropy Guided Multi Token Prediction

ResearchDGX agent

arXiv:2606.27550v1 Announce Type: new Abstract: Multi-token prediction has been shown to increase data density during training, improve downstream text-generation quality, and serves as the defacto ap

Estimation--Prediction Tradeoff in Causal Probabilistic Temporal Graphs

Model ReleasesDGX agent

arXiv:2606.28225v1 Announce Type: new Abstract: Temporal link prediction is usually evaluated by predictive performance on unseen edges, but in probabilistic temporal graphs this criterion can conflat

ET-SAM: Efficient Point Prompt Prediction in SAM for Unified Scene Text Detection and Layout Analysis

ResearchDGX agent

arXiv:2603.25168v2 Announce Type: replace Abstract: Previous works based on Segment Anything Model (SAM) have achieved promising performance in unified scene text detection and layout analysis. Howeve

Every Step of the Way: Video-based Parkinsonian Turning Step Counting

ApplicationsDGX agent

arXiv:2606.27918v1 Announce Type: cross Abstract: As a prominent symptom of Parkinson's disease (PD), turning impairment is evaluated through parameters such as turning angle, duration, and particular

Explainable AI for Biodiversity Monitoring and Ecological Image Analysis

TutorialsDGX agent

arXiv:2606.27667v1 Announce Type: cross Abstract: Artificial intelligence is transforming biodiversity monitoring by enabling automated analysis of ecological imagery collected from camera traps, dron

EXPLORE-Bench: Egocentric Scene Prediction with Long-Horizon Reasoning

Model ReleasesDGX agent

arXiv:2603.09731v3 Announce Type: replace-cross Abstract: Multimodal large language models (MLLMs) are increasingly considered as a foundation for embodied agents, yet it remains unclear whether they

Exposure Bias Can Alleviate Itself via Directional and Frequency Rectification in Flow Matching

SafetyDGX agent

arXiv:2606.28226v1 Announce Type: cross Abstract: Flow Matching (FM) has achieved remarkable generative performance, yet it suffers from exposure bias due to discrepancies between training and inferen

FailSafe: Reasoning and Recovery from Failures in Vision-Language-Action Models

ApplicationsDGX agent

arXiv:2510.01642v3 Announce Type: replace Abstract: Recent advances in robotic manipulation have integrated low-level robotic control into Vision-Language Models (VLMs), extending them into Vision-Lan

Fair Classification with Efficient and Post-hoc Controllable Fairness-Accuracy Trade-off

SafetyDGX agent

arXiv:2606.28097v1 Announce Type: new Abstract: Post-hoc controllability of fair machine learning models, the ability to control the trade-off between fairness and accuracy after training, is valuable

Fine-Grained Behavior and Lane Constraints Guided Trajectory Prediction Method

AgentsDGX agent

arXiv:2503.21477v3 Announce Type: replace Abstract: Trajectory prediction, as a critical component of autonomous driving systems, has attracted the attention of many researchers. Existing prediction a

Fine-tuning a multimodal large language model for clinician-grade autism behavioral scoring from short home videos

Model ReleasesDGX agent

arXiv:2606.27484v1 Announce Type: new Abstract: Autism spectrum disorder (ASD) affects 1 in 31 US children, yet median age at diagnosis exceeds four years. Artificial intelligence pipelines that provi

Flexformer: Flexible Linear Transformer with Learnable Attention Kernel

TutorialsDGX agent

arXiv:2606.27748v1 Announce Type: cross Abstract: Transformer models rely on attention mechanism to capture long-range dependencies but suffer from quadratic complexity, limiting their scalability to

FlexMoE: One-for-All Nested Intra-Expert Pruning for MoE Language Models

TutorialsDGX agent

arXiv:2606.27866v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) language models scale model ability with sparsely activated experts, making this architecture a standard recipe for modern larg

FoggyTrust: Robust Federated Learning with Hierarchical Trust Networks

Model ReleasesDGX agent

arXiv:2606.27622v1 Announce Type: new Abstract: Byzantine-robust federated learning seeks to protect distributed model training from malicious or corrupted clients without requiring access to their pr

Forecasting Technological Directions in Wireless Networks and Mobile Computing via AutoML Framework

ResearchDGX agent

arXiv:2606.27394v1 Announce Type: cross Abstract: The exponential increase in scientific publications has driven the emergence of new trends. Accurate forecasting of these developments is essential fo

Formalizing Latent Thoughts: Four Axioms of Thought Representation in LLMs

Model ReleasesDGX agent

arXiv:2606.27378v1 Announce Type: new Abstract: We introduce an axiomatic evaluation framework for latent thought representations in LLMs, comprising metrics that are independent of downstream benchma

Foundation vs. Specialized Models: Evaluating Catastrophic Forgetting in Continual Time Series Forecasting

ApplicationsDGX agent

arXiv:2510.00809v3 Announce Type: replace Abstract: While Time Series Foundation Models (TSFMs) excel in zero-shot tasks, their behavior under continual fine tuning is poorly understood. We present th

Freshness and the Limits of Heuristic Trend Detection in Temporal RAG

Model ReleasesDGX agent

arXiv:2509.19376v2 Announce Type: replace-cross Abstract: We present a lightweight, model-agnostic temporal layer for RAG and use cybersecurity data to separate two problems that are usually conflated

From Black-Box to Clinical Insight: A Multi-Stage Explainable Framework for Speech-Based Cognitive Impairment Detection

Model ReleasesDGX agent

arXiv:2606.27973v1 Announce Type: cross Abstract: Speech-based cognitive impairment detection offers a noninvasive, accessible alternative to costly biomarker assays, yet transformer-based models rema

From Detection to Action: Using LLM Agents for Fault-Tolerant Control

Model ReleasesDGX agent

arXiv:2606.28011v1 Announce Type: cross Abstract: We propose an agentic Large Language Model (LLM) framework for active Fault-Tolerant Control (FTC) that transforms fault detection outputs into constr

From General-Purpose Audio Tagging to Spatially Grounded Sound Event Localization and Detection

ResearchDGX agent

arXiv:2606.27751v1 Announce Type: cross Abstract: This report investigates the extension of pretrained General-Purpose Audio Tagging (GP-AT) models toward spatially grounded Sound Event Localization a

From Signals to Transfer: A Factorised Study of Probe-Based Uncertainty Estimation in Large Language Models

Model ReleasesDGX agent

arXiv:2606.27679v1 Announce Type: cross Abstract: Probe-based uncertainty estimation (UE) has emerged as a prominent approach to detect hallucinations in Large Language Models (LLMs) by learning uncer

From Tokens to States: LLMs as a Special Case of World Models and the Continuous Path Beyond

ResearchDGX agent

arXiv:2606.28127v1 Announce Type: cross Abstract: The AI community has framed the relationship between large language models (LLMs) and world models as a dichotomy: LLMs predict tokens; world models s

GAIA: A Data Flywheel System for Training GUI Test-Time Scaling Critic Models

Model ReleasesDGX agent

arXiv:2601.18197v2 Announce Type: replace Abstract: While Large Vision-Language Models (LVLMs) have significantly advanced GUI agents' capabilities in parsing textual instructions, interpreting screen

'Generate' the Future of Work through AI: Empirical Evidence from Online Labor Markets

ResearchDGX agent

arXiv:2308.05201v4 Announce Type: replace Abstract: Large Language Model (LLM)-based generative AI systems are general-purpose tools capable of augmenting or even automating a wide range of job functi

GeoFace: Consistent Multi-View Face Generation with Geometry-Constrained Diffusion

SafetyDGX agent

arXiv:2606.27659v1 Announce Type: new Abstract: We present GeoFace, a geometry-constrained multi-view diffusion framework for consistent face generation from a single input. % While recent multi-view

Global Explanations for Multivariate Time Series Forecasting Models via K-Order Markov Approximations

ApplicationsDGX agent

arXiv:2606.27599v1 Announce Type: cross Abstract: While many explainable AI (XAI) methods have been proposed, most are not designed for time-series forecasting models and often rely on the implicit as

GNBAN: Graph Neural Basis Attention Networks for Long-Horizon Forecasting over Large Entity Sets

ResearchDGX agent

arXiv:2606.27863v1 Announce Type: cross Abstract: Demand forecasting at the bottom of a retail hierarchy requires predicting tens of thousands of correlated long-horizon series across products, stores

Govern the Repository, Not the Agent: Measuring Ecosystem-Level Risk in AI-Native Software

Model ReleasesDGX agent

arXiv:2606.28235v1 Announce Type: cross Abstract: Autonomous coding agents now open and merge pull requests in shared repositories at scale, and the field evaluates them the way it has always evaluate

GRAFT: Biological Graph and Hypergraph Benchmarks for Linked Gene Expression and Phenotypic Trait Prediction in Arabidopsis thaliana

Model ReleasesDGX agent

arXiv:2606.27413v1 Announce Type: cross Abstract: Understanding which genes control which traits in an organism remains one of the central challenges in biology. Despite significant advances in data c

Graph Dimensionality Reduction for Contextual Bandits: Structure-Specific Regret Bounds under Approximate Smoothness and Noisy Eigenspaces

Model ReleasesDGX agent

arXiv:2606.27917v1 Announce Type: new Abstract: Contextual bandits with graph-structured arms arise in recommendation, citation retrieval, and social advertising, where arms connected on a graph tend

Graph Unfolding and Sampling for Transitory Video Keyframe Selection via Gershgorin Disc Alignment

SafetyDGX agent

arXiv:2408.01859v2 Announce Type: replace Abstract: User-generated videos (UGVs) uploaded from mobile phones to social media sites like YouTube and TikTok are short and non-repetitive. We summarize a

GraphPilot: Grounded Scene Graph Conditioning for Language-Based Autonomous Driving

AgentsDGX agent

arXiv:2511.11266v4 Announce Type: replace Abstract: Vision-language models have recently emerged as promising planners for autonomous driving, where success hinges on topology-aware reasoning over spa

Grounded Iterative Language Planning: How Parameterized World Models Reduce Hallucination Propagation in LLM Agents

AgentsDGX agent

arXiv:2606.27806v1 Announce Type: new Abstract: World models for language agents come in two useful forms. An agent-based world model calls an LLM API and reasons flexibly in language, but its errors

Halt Fast! Early Stopping for Certified Robustness

SafetyDGX agent

arXiv:2606.27694v1 Announce Type: cross Abstract: Randomized Smoothing (RS) provides rigorous robustness guarantees for neural networks without architectural constraints, yet its adoption is limited b

HAT-4D: Lifting Monocular Video for 4D Multi-Object Interactions via Human-Agent Collaboration

Model ReleasesDGX agent

arXiv:2606.28215v1 Announce Type: cross Abstract: Extracting dynamic 4D object interactions from massive, in-the-wild monocular videos offers a highly efficient data collection pathway for scaling Emb

Health-ORSC-Bench: A Benchmark for Measuring Over-Refusal and Safety Completion in Health Context

Model ReleasesDGX agent

arXiv:2601.17642v2 Announce Type: replace Abstract: Safety alignment in Large Language Models is critical for healthcare; however, reliance on binary refusal boundaries often results in over-refusal o

hia-gat: A Heterogeneous Interaction-Aware Graph Attention Network For Frame-Level Traffic Conflict Risk Prediction On Freeways

SafetyDGX agent

arXiv:2606.27577v1 Announce Type: cross Abstract: This paper formulates frame-level freeway risk assessment as a multi-agent scene graph-level binary classification problem, where each video or trajec

Hierarchical Control in Multi-Agent Games: LLM-based Planning and RL Execution

AgentsDGX agent

arXiv:2606.20014v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) has achieved strong performance in sequential decision-making, yet scaling to complex multi-agent environments rem

Higher-Order Fourier Neural Operator: Explicit Mode Mixer for Nonlinear PDEs

SafetyDGX agent

arXiv:2606.28122v1 Announce Type: cross Abstract: Neural operators provide deep neural networks for learning mappings between function spaces. Among them, the Fourier Neural Operator (FNO) is particul

HiMu: Hierarchical Multimodal Frame Selection for Long Video Question Answering

ResearchDGX agent

arXiv:2603.18558v2 Announce Type: replace-cross Abstract: Long-form video question answering requires reasoning over extended temporal contexts, making frame selection a critical bottleneck for multi-

Hippocampus-DETR: An Explicit Memory Object Detection Framework Based on Hippocampus Modeling

ResearchDGX agent

arXiv:2606.27831v1 Announce Type: cross Abstract: This paper addresses the lack of explicit memory mechanisms in current object detection models and proposes Hippocampus-DETR, a novel detection framew

Home3D 1.0: A High-Fidelity Image-to-3D Asset Generation System for Interior Design

ResearchDGX agent

arXiv:2606.27923v1 Announce Type: cross Abstract: We present Home3D 1.0, a modular image-to-3D generation system that produces high-quality 3D assets from a single reference image, targeting interior

How Width and Data Shape Generalization Scaling Laws in Quadratic Neural Networks

ResearchDGX agent

arXiv:2606.28242v1 Announce Type: cross Abstract: Understanding how performance scales jointly with model size and data is a central problem in modern machine learning. Existing theoretical works on s

HPRO: Hierarchical Progressive Reward Optimization via Preference Extraction for Emotional Text-to-Speech

TutorialsDGX agent

arXiv:2606.28249v1 Announce Type: cross Abstract: Recently, Large Language Model (LLM)-based Text-to-Speech (TTS) models have achieved remarkable naturalness. However, the standard Supervised Fine-Tun

HumanMoveVQA: Can Video MLLMs reason about human movement in videos?

Model ReleasesDGX agent

arXiv:2606.27999v1 Announce Type: new Abstract: Despite the rapid advance of Multimodal Large Language Models (MLLMs) in high-level video understanding, a fundamental bottleneck remains: these models

HunyuanImage 3.0 Technical Report

SafetyDGX agent

arXiv:2509.23951v3 Announce Type: replace Abstract: We present HunyuanImage 3.0, a native multimodal model that unifies multimodal understanding and generation within an autoregressive framework, with

Hybrid coupling with operator inference and the overlapping Schwarz alternating method

Local AiDGX agent

arXiv:2511.20687v3 Announce Type: replace-cross Abstract: This paper presents a novel hybrid approach for coupling subdomain-local non-intrusive Operator Inference (OpInf) reduced order models (ROMs)

← Previous
1…339340341342343…1032
Next →