AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “papers”

GridTimelineEvolution
12,158 results
5 Aug 2026

SP3O: Reinforcement Learning from Segment Preferences without Reward Modeling

SafetyDGX agent

arXiv:2608.02951v1 Announce Type: cross Abstract: Preference-based reinforcement learning (PbRL) for general stochastic MDPs often requires training a reward model. Existing reward-model-free methods

State Propagation Also Satisfies: A Complex-Valued State-Space Model for Deterministic State Tracking

ResearchDGX agent

arXiv:2608.03425v1 Announce Type: new Abstract: Transformer-based architectures have dominated sequence modeling, largely due to the expressive power of attention mechanisms. However, for a class of d

Staying on Spec: Real-Time Monitoring under Uncertainty with a Maritime Case Study

SafetyDGX agent

arXiv:2608.02811v1 Announce Type: new Abstract: Robotic systems must operate under uncertainty while satisfying complex task and safety specifications. Monitoring such specifications under uncertainty

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Stiffness Copilot: An Impedance Policy for Contact-Rich Teleoperation

SafetyDGX agent

arXiv:2603.14068v2 Announce Type: replace Abstract: In teleoperation of contact-rich manipulation tasks, selecting robot impedance is critical but difficult. The robot must be compliant to avoid damag

Stochastic Multiple Shooting Trajectory Optimization via Sequential Local Policy Evaluation

Local AiDGX agent

arXiv:2608.03978v1 Announce Type: new Abstract: Stochastic single shooting trajectory optimization methods such as Model Predictive Path Integral control (MPPI) have been widely adopted in robotics du

Stochastic Saddle Avoidance Beyond Unit Excitation and Smoothness: A Pathwise Lyapunov-Perron Framework

ResearchDGX agent

arXiv:2608.03001v1 Announce Type: cross Abstract: Unit excitation (UE) is a common assumption in stochastic saddle avoidance: the stochastic error must have a uniformly positive component along every

STREAM-VAE: Dual-Path Routing for Slow and Fast Dynamics in Vehicle Telemetry Anomaly Detection

Model ReleasesDGX agent

arXiv:2511.15339v3 Announce Type: replace-cross Abstract: Automotive telemetry data exhibits slow drifts and fast spikes, often within the same sequence, making reliable anomaly detection challenging.

T2VAttack: Adversarial Attack on Text-to-Video Diffusion Models

SafetyDGX agent

arXiv:2512.23953v2 Announce Type: replace Abstract: The rapid evolution of Text-to-Video (T2V) diffusion models has driven remarkable advancements in generating high-quality, temporally coherent video

Test-Time Scaling for Safe Text-Guided Image Generation via Intermediate Clean Estimates

SafetyDGX agent

arXiv:2608.03284v1 Announce Type: cross Abstract: Ensuring safety and policy compliance in text-to-image diffusion models remains a critical challenge, as benign or adversarial prompts can often elici

The Ignition Is Real, and It Lives at the Readout: Latent composition, difficulty-clocked ignition, and the interface-constituted commit in a recurrent-depth reasoner

Model ReleasesDGX agent

arXiv:2608.03263v1 Announce Type: cross Abstract: We test whether the 'compositional ignition' reported in latent-reasoning models is real computation, an instrument artifact, or inherited from verbal

Toward Visual Grounding: A Survey

ResearchDGX agent

arXiv:2412.20206v4 Announce Type: replace Abstract: Visual Grounding, also known as Referring Expression Comprehension and Phrase Grounding, aims to ground the specific region(s) within the image(s) b

Training Documents Reranker with Search Rubrics for Deep Research Agent

AgentsDGX agent

arXiv:2608.03527v1 Announce Type: cross Abstract: Retrieval systems help deep research agents generate high-quality answers by providing relevant documents. However, existing retrievers typically sele

UL-UNAS: Ultra-Lightweight U-Nets for Real-Time Speech Enhancement via Network Architecture Search

ResearchDGX agent

arXiv:2503.00340v2 Announce Type: cross Abstract: Lightweight models are essential for real-time speech enhancement applications. In recent years, there has been a growing trend toward developing incr

Xiaomi-Robotics-1: New robotics model released

Model ReleasesDGX agent

Xiaomi-Robotics-1 is a robot foundation model trained on over 100K hours of real-world manipulation trajectories. It is a Vision-Language-Action (VLA) model engineered for out-of-the-box mobile manipu

Your Agentic LLMs Secretly Encode Latent Signals of Indirect Prompt-Injection Exposure

Model ReleasesDGX agent

arXiv:2608.02657v1 Announce Type: cross Abstract: Agentic LLMs are vulnerable to indirect prompt injection (IPI) attacks, e.g., malicious side-tasks hidden in external tool results. While many efforts

4 Aug 2026

A Benchmark Dataset for MLLM-Generated Image Detection: GPT Image2 & Nano Banana2

Model ReleasesDGX agent

arXiv:2608.01258v1 Announce Type: new Abstract: The realism of images generated by multimodal large language models (MLLMs), such as GPT Image2 and Nano Banana2, has improved rapidly in recent years.

A Comparative Analysis of MLP and Kolmogorov-Arnold Networks (KAN) for Faster-than-Nyquist (FTN) Signaling Detection

Model ReleasesDGX agent

arXiv:2608.02062v1 Announce Type: cross Abstract: Faster-than-Nyquist signaling improves spectral ef- ficiency by deliberately introducing inter-symbol interference. Classical sequence detectors such

A drone that learns to efficiently find non-uniformly distributed objects in agricultural fields: from simulation to the real world

AgentsDGX agent

arXiv:2505.09278v2 Announce Type: replace Abstract: Drones are promising for data collection in precision agriculture but are limited by battery capacity. Drone paths are usually planned using full co

A Multi-Objective AutoML-based Efficient Intrusion Detection System for EV Charging Networks

ResearchDGX agent

arXiv:2608.02274v1 Announce Type: cross Abstract: Electric Vehicle Charging Systems (EVCSs) are increasingly connected with Internet of Things (IoT) devices, which improves charging intelligence but a

A Tilt-Rotor UAV with a Gripper for Stable Contact-Based Tasks via Environmental Anchoring

ResearchDGX agent

arXiv:2608.01736v1 Announce Type: new Abstract: Maintaining a stable pose during physical interaction is a significant challenge for aerial robots, often limiting their use in contact-based tasks. Thi

ABRA: Teleporting Fine-Tuned Knowledge Across Domains for Open-Vocabulary Object Detection

ResearchDGX agent

arXiv:2603.12409v2 Announce Type: replace Abstract: Although recent Open-Vocabulary Object Detection architectures, such as Grounding DINO, demonstrate strong zero-shot capabilities, their performance

Action Chunk Scheduling for Batched Robot Policy Serving

Local AiDGX agent

arXiv:2608.00337v1 Announce Type: new Abstract: Deploying robot foundation models at scale is the next step towards realizing the potential of general-purpose robots. However, Vision-Language-Action (

AdaHAT: Adaptive Hard Attention to the Task in Task-Incremental Learning

TutorialsDGX agent

arXiv:2608.01252v1 Announce Type: new Abstract: Catastrophic forgetting is a major problem in task-incremental learning, where neural networks tend to overwrite previously learned knowledge when train

AirSplat: Alignment and Rating for Robust Feed-Forward 3D Gaussian Splatting

Local AiDGX agent

arXiv:2603.25129v2 Announce Type: replace Abstract: While 3D Vision Foundation Models (3DVFMs) have demonstrated remarkable zero-shot capabilities in visual geometry estimation, their direct applicati

An Embedded RISC-V Evaluation of Kolmogorov--Arnold Networks in Hard-Constrained Recurrent Physics-Informed Models

Model ReleasesDGX agent

arXiv:2608.00737v1 Announce Type: new Abstract: Hard-constrained recurrent physics-informed networks (HRPINNs) embed known dynamics inside a recurrent numerical integrator and restrict a neural branch

An Uncertainty-Driven Hybrid Deep Learning Approach for Broad-Coverage RF Modulation Recognition

ResearchDGX agent

arXiv:2608.00796v1 Announce Type: cross Abstract: Automatic RF modulation recognition is of critical importance in spectrum monitoring, electronic warfare, and cognitive radio applications, where low

ArabicDialectSafety: A Dialect-Aware Benchmark for Arabic Content Safety Classification

Model ReleasesDGX agent

arXiv:2608.01291v1 Announce Type: new Abstract: We present ArabicDialectSafety, a human-curated Arabic safety dataset of 25,071 prompts covering six Arabic varieties: Modern Standard Arabic, Syrian, E

AROpt: An Optimization Method for Autoregressive Time Series Forecasting

ResearchDGX agent

arXiv:2602.02288v3 Announce Type: replace Abstract: Current time-series forecasting models are primarily based on transformer-style neural networks. These models achieve long-term forecasting mainly b

Assessing the Benefits of Combining Advanced Deep Learning Techniques for Post-Disaster Building Damage Assessment from UAV Imagery

ApplicationsDGX agent

arXiv:2608.01906v1 Announce Type: new Abstract: Rapid and accurate post-disaster building damage assessment is essential, yet remains a challenging task. Unmanned Aerial Vehicle (UAV) imagery offers a

Automatic Annotation of Ancient Greek Vowel Length

Model ReleasesDGX agent

arXiv:2608.01935v1 Announce Type: new Abstract: Prior work in Ancient Greek NLP relies on corpora that do not disambiguate the phonemic vowel length of alpha, iota, and ypsilon, together known as the

Beyond Magnitude and Shape: A Direction-Aware Loss for Time Series Forecasting

ResearchDGX agent

arXiv:2608.01857v1 Announce Type: new Abstract: The direction of change --- whether a series will move up or down --- is often as important as its exact value in decisiondriven applications such as ri

Breaking the Horizontal Prior: From Long-Tailed Orientation Bias to Roll-Robust Monocular Depth Estimation

Model ReleasesDGX agent

arXiv:2608.00678v1 Announce Type: new Abstract: Despite recent advances in Monocular Depth Estimation, state-of-the-art depth foundation models remain vulnerable to robustness issues. Particularly, ev

CARNet: Channel-Adaptive Receiver Network for Robust NextG Communications

ResearchDGX agent

arXiv:2608.02172v1 Announce Type: cross Abstract: Neural receivers have been recognized as a promising paradigm for the next-generation (NextG) communications. However, due to the reliance on a static

CAVE: Competence-Aware Visual Boundary Evidence Alignment for Video Temporal Grounding

SafetyDGX agent

arXiv:2608.02078v1 Announce Type: new Abstract: Large vision-language models (LVLMs) have achieved substantial performance gains in Video Temporal Grounding (VTG) through reinforcement learning (RL).

CDG-MAE: Cross-view Masked Modeling using Diffusion Generated Views

Local AiDGX agent

arXiv:2506.18164v2 Announce Type: replace Abstract: Cross-view masked autoencoding has emerged as a powerful pretext task for learning dense correspondences, which are essential for applications such

Credit Assignment and Efficient Exploration based on Influence Scope in Multi-agent Reinforcement Learning

AgentsDGX agent

arXiv:2505.08630v2 Announce Type: replace Abstract: Training cooperative agents in sparse-reward scenarios poses significant challenges for multi-agent reinforcement learning (MARL). Without clear fee

CRISP: Critical Step Perception for Training Efficient Deep Search Agents

SafetyDGX agent

arXiv:2608.01867v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly extended into deep search agents that solve complex questions through multi-step interaction with external

Crushing the Evidence: A Dual-Penalty Evasion Framework for Fooling White-Box Explainable AI Auditors

Model ReleasesDGX agent

arXiv:2608.00566v1 Announce Type: new Abstract: Post-hoc model explainers such as LIME, SHAP, and Integrated Gradients are widely deployed to audit models in high-stakes sensitive domains, including f

CTRAG: An In-Context Retrieval-based Framework for Automated Compliance Checking using LLMs

SafetyDGX agent

arXiv:2608.02472v1 Announce Type: new Abstract: Trust is fundamental in modern regulatory ecosystems, and compliance checking plays a critical role in fostering that trust. Regulatory compliance verif

DAG-FM: A Foundation Model for Causal Discovery under Heterogeneous Causal Mechanisms

ApplicationsDGX agent

arXiv:2607.11510v2 Announce Type: replace Abstract: Causal discovery from observational tabular data remains fundamentally challenging, primarily due to the heterogeneity of underlying causal mechanis

Deep Learning CNN and Recurrence Analysis for Alpha Gamma EEG Biomarkers in Fragile X Syndrome

TutorialsDGX agent

arXiv:2608.00835v1 Announce Type: cross Abstract: Fragile X Syndrome (FXS) is a neurodevelopmental disorder caused by reduced expression of fragile X mental retardation protein (FMRP), leading to disr

Deep Learning for Retinal Degeneration Assessment: A Comprehensive Analysis of the MARIO Challenge

Model ReleasesDGX agent

arXiv:2506.02976v4 Announce Type: replace Abstract: The MARIO challenge, held at MICCAI 2024, focused on advancing the automated detection and monitoring of age-related macular degeneration (AMD) thro

Deep Multimodal Fusion Detection through Spatial Mask and Channel Fusion

TutorialsDGX agent

arXiv:2608.02092v1 Announce Type: new Abstract: Deep multimodal fusion for object detection has demonstrated good performance through mining modal characteristics. However, existing feature-level fusi

Deep Research Pretraining via Predictive Navigation

SafetyDGX agent

arXiv:2608.00432v1 Announce Type: new Abstract: Deep research agents are often trained on expensive, environment-grounded tool-use trajectories that require repeated retrieval, document inspection, an

DeepDefense: Robust Learning via Layer-Wise Gradient-Feature Alignment

SafetyDGX agent

arXiv:2511.13749v2 Announce Type: replace Abstract: Deep neural networks are known to be vulnerable to adversarial perturbations, which are small, carefully crafted inputs that lead to incorrect predi

Deployment-Ready UWB Localization for Industrial Ground Robots with Automatic Anchor Calibration and Terrain-Aware Fusion

Model ReleasesDGX agent

arXiv:2607.15807v2 Announce Type: replace Abstract: Ultra-Wideband (UWB) ranging has become a viable option for industrial Autonomous Mobile Robot (AMR) localization due to improved accuracy and low c

DeVIT: Low-Power Vision Transformer Acceleration Using Delta Computation

ResearchDGX agent

arXiv:2608.01343v1 Announce Type: new Abstract: The emergence of transformer-based deep learning models has brought unprecedented performance across various domains, particularly in natural language p

DiffPhysCam: Differentiable Physics-Based Camera Simulation for Inverse Rendering and Embodied AI

AgentsDGX agent

arXiv:2508.08831v2 Announce Type: replace-cross Abstract: Generating synthetic images that closely mimic those from real cameras is instrumental in training visual models and enabling end-to-end visuo

DiffPrune: differentiable information throttling for token pruning in vision-language models

TutorialsDGX agent

arXiv:2608.01985v1 Announce Type: new Abstract: Visual token pruning reduces the computational cost of Vision-Language Models (VLMs) by removing redundant visual tokens. The key is to learn a score th

DODA: A Database of Datasets for Aesthetics Research

ResearchDGX agent

arXiv:2608.00089v1 Announce Type: new Abstract: With rapid growth in the fields of empirical and computational aesthetics we have seen a vast increase in large image datasets annotated for aesthetics.

Does Explainability Transfer? A Controlled Benchmark of Attribution Methods on Vision Transformers and CNNs

Model ReleasesDGX agent

arXiv:2608.02396v1 Announce Type: new Abstract: Most evidence on the effectiveness of explainable artificial intelligence (XAI) attribution methods has been established on convolutional neural network

Does the Competitive Component of Adversarial Self-Play Improve Legal Reasoning? A Controlled Negative Result

ApplicationsDGX agent

arXiv:2608.01559v1 Announce Type: cross Abstract: Adversarial self-play is an appealing recipe for legal reasoning: have a student model draft an argument, have an adversary attack it, and reward the

Domain-Specific Evaluation of Text-to-Speech Systems: A Multi-Metric Benchmarking Study

Model ReleasesDGX agent

arXiv:2608.02235v1 Announce Type: new Abstract: Recent advances in neural text-to-speech (TTS) systems have substantially improved speech naturalness and intelligibility across many languages. However

DriveCode: Domain Specific Numerical Encoding for LLM-Based Autonomous Driving

AgentsDGX agent

arXiv:2603.00919v3 Announce Type: replace Abstract: Large language models (LLMs) have shown great promise for autonomous driving. However, discretizing numbers into tokens limits precise numerical rea

DS@GT ARC at MEDIQA-CORE-Task-1 2026: Trimodal Model Fusion with Task-Specific Gates for Brain Tumor Subtype Classification

Model ReleasesDGX agent

arXiv:2608.00086v1 Announce Type: new Abstract: Brain tumor diagnosis is a time-sensitive process in which patients may wait weeks for a finalized pathology report. This problem motivates automated sy

DynamicManip: Enabling Dynamic Manipulation from a Single Static Demonstration

Model ReleasesDGX agent

arXiv:2608.01452v1 Announce Type: cross Abstract: Dynamic manipulation is a critical capability for robots operating in complex and dynamic environments, where robots must interact with objects that a

DynImmune-BERT: Dynamic Immune Repertoire Modeling with Neural ODE Driven Continuous Transformers

Model ReleasesDGX agent

arXiv:2607.17244v2 Announce Type: replace Abstract: Longitudinal T cell receptor repertoires contain signals of clonal expansion, contraction, disappearance, and reappearance after immune perturbation

EchoCache: Energy-Guided Cross-Modal Caching for Efficient Audio-Driven Video Generation

Model ReleasesDGX agent

arXiv:2608.02474v1 Announce Type: new Abstract: Audio-driven video generation (A2V) has achieved promising progress in synthesizing temporally coherent and audio-visually aligned videos, yet its infer

Empowering Credit Risk Detection in Weixin Pay with Billion-Scale Deep Graph Learning

Local AiDGX agent

arXiv:2608.02168v1 Announce Type: new Abstract: Credit risk detection, particularly mitigating individual fraud, is crucial for maintaining the stability of digital financial ecosystems. Accurately id

Enhancing Large Language Model Reasoning with Reward Models: An Analytical Survey

ResearchDGX agent

arXiv:2510.01925v3 Announce Type: replace Abstract: Reward models (RMs) play a critical role in enhancing the reasoning performance of LLMs. For example, they can provide training signals to finetune

← Previous
1…6162636465…203
Next →