AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “papers”

GridTimelineEvolution
11,164 results
Safety

From Alignment to Prediction: A Study of Self-Supervised Learning and Predictive Representation Learning

DGX agent

arXiv:2604.13518v1 Announce Type: new Abstract: Self-supervised learning has emerged as a major technique for the task of learning from unlabeled data, where the current methods mostly revolve around

safetyarxiv-cs-lg
16 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

From P(y|x) to P(y): Investigating Reinforcement Learning in Pre-train Space

DGX agent

arXiv:2604.14142v1 Announce Type: cross Abstract: While reinforcement learning with verifiable rewards (RLVR) significantly enhances LLM reasoning by optimizing the conditional distribution P(y|x), it

safetyarxiv-cs-cl
16 Apr 2026
Agents

GeoVision-Enabled Digital Twin for Hybrid Autonomous-Teleoperated Medical Responses

DGX agent

arXiv:2604.13248v1 Announce Type: new Abstract: Remote medical response systems are increasingly being deployed to support emergency care in disaster-affected and infrastructure-limited environments.

agentsarxiv-cs-ro
16 Apr 2026
Safety

HAMLET: Switch your Vision-Language-Action Model into a History-Aware Policy

DGX agent

arXiv:2510.00695v3 Announce Type: replace-cross Abstract: Inherently, robotic manipulation tasks are history-dependent: leveraging past context could be beneficial. However, most existing Vision-Langu

safetyarxiv-cs-cv
16 Apr 2026
Agents

Hierarchical DLO Routing with Reinforcement Learning and In-Context Vision-language Models

DGX agent

arXiv:2510.19268v2 Announce Type: replace-cross Abstract: Long-horizon routing tasks of deformable linear objects (DLOs), such as cables and ropes, are common in industrial assembly lines and everyday

agentsarxiv-cs-lg
16 Apr 2026
Research

Hierarchical Reinforcement Learning with Augmented Step-Level Transitions for LLM Agents

DGX agent

arXiv:2604.05808v2 Announce Type: replace-cross Abstract: Large language model (LLM) agents have demonstrated strong capabilities in complex interactive decision-making tasks. However, existing LLM ag

researcharxiv-cs-lg
16 Apr 2026
Model Releases

Hierarchical Reinforcement Learning with Runtime Safety Shielding for Power Grid Operation

DGX agent

arXiv:2604.14032v1 Announce Type: cross Abstract: Reinforcement learning has shown promise for automating power-grid operation tasks such as topology control and congestion management. However, its de

model-releasesarxiv-cs-lg
16 Apr 2026
Applications

IWLV-Ramayana: A Sarga-Aligned Parallel Corpus of Valmiki's Ramayana Across Indian Languages

DGX agent

arXiv:2604.13078v1 Announce Type: new Abstract: The Ramayana is among the most influential literary traditions of South and Southeast Asia, transmitted across numerous linguistic and cultural contexts

applicationsarxiv-cs-cl
16 Apr 2026
Safety

Jump-Start Reinforcement Learning with Vision-Language-Action Regularization

DGX agent

arXiv:2604.13733v1 Announce Type: new Abstract: Reinforcement learning (RL) enables high-frequency, closed-loop control for robotic manipulation, but scaling to long-horizon tasks with sparse or imper

safetyarxiv-cs-lg
16 Apr 2026
Model Releases

KV Packet: Recomputation-Free Context-Independent KV Caching for LLMs

DGX agent

arXiv:2604.13226v1 Announce Type: new Abstract: Large Language Models (LLMs) rely heavily on Key-Value (KV) caching to minimize inference latency. However, standard KV caches are context-dependent: re

model-releasesarxiv-cs-lg
16 Apr 2026
Applications

Learning Dynamics from Input-Output Data with Hamiltonian Gaussian Processes

DGX agent

arXiv:2511.05330v2 Announce Type: replace Abstract: Embedding non-restrictive prior knowledge, such as energy conservation laws, into learning methods is a key motive to construct physically consisten

applicationsarxiv-cs-lg
16 Apr 2026
Research

Learning Inference Concurrency in DynamicGate MLP Structural and Mathematical Justification

DGX agent

arXiv:2604.13546v1 Announce Type: new Abstract: Conventional neural networks strictly separate learning and inference because if parameters are updated during inference, outputs become unstable and ev

researcharxiv-cs-lg
16 Apr 2026
Applications

Lite Any Stereo: Efficient Zero-Shot Stereo Matching

DGX agent

arXiv:2511.16555v3 Announce Type: replace Abstract: Recent advances in stereo matching have focused on accuracy, often at the cost of significantly increased model size. Traditionally, the community h

applicationsarxiv-cs-cv
16 Apr 2026
Safety

Med-CAM: Minimal Evidence for Explaining Medical Decision Making

DGX agent

arXiv:2604.13695v1 Announce Type: new Abstract: Reliable and interpretable decision-making is essential in medical imaging, where diagnostic outcomes directly influence patient care. Despite advances

safetyarxiv-cs-cv
16 Apr 2026
Research

node2vec or triangle-biased random walks: stationarity, regularity & recurrence

DGX agent

arXiv:2604.13681v1 Announce Type: cross Abstract: The node2vec random walk is a non-Markovian random walk on the vertex set of a graph, widely used for network embedding and exploration. This random w

researcharxiv-cs-lg
16 Apr 2026
Agents

Numerical Instability and Chaos: Quantifying the Unpredictability of Large Language Models

DGX agent

arXiv:2604.13206v1 Announce Type: cross Abstract: As Large Language Models (LLMs) are increasingly integrated into agentic workflows, their unpredictability stemming from numerical instability has eme

agentsarxiv-cs-lg
16 Apr 2026
Agents

Optimized Human-Robot Co-Dispatch Planning for Petro-Site Surveillance under Varying Criticalities

DGX agent

arXiv:2602.07924v2 Announce Type: replace Abstract: Securing petroleum infrastructure requires balancing autonomous system efficiency with human judgment for threat escalation, a challenge unaddressed

agentsarxiv-cs-ro
16 Apr 2026
Research

Outperforming Self-Attention Mechanisms in Solar Irradiance Forecasting via Physics-Guided Neural Networks

DGX agent

arXiv:2604.13455v1 Announce Type: new Abstract: Accurate Global Horizontal Irradiance (GHI) forecasting is critical for grid stability, particularly in arid regions characterized by rapid aerosol fluc

researcharxiv-cs-lg
16 Apr 2026
Model Releases

Parameter-Free Non-Ergodic Extragradient Algorithms for Solving Monotone Variational Inequalities

DGX agent

arXiv:2604.07662v2 Announce Type: replace-cross Abstract: Monotone variational inequalities (VIs) provide a unifying framework for convex minimization, equilibrium computation, and convex-concave sadd

model-releasesarxiv-cs-lg
16 Apr 2026
Safety

Pareto-Optimal Offline Reinforcement Learning via Smooth Tchebysheff Scalarization

DGX agent

arXiv:2604.13175v1 Announce Type: new Abstract: Large language models can be aligned with human preferences through offline reinforcement learning (RL) on small labeled datasets. While single-objectiv

safetyarxiv-cs-lg
16 Apr 2026
Research

Person Re-Identification via Generalized Class Prototypes

DGX agent

arXiv:2510.17043v2 Announce Type: replace Abstract: Advanced feature extraction methods have significantly contributed to enhancing the task of person re-identification. In addition, modifications to

researcharxiv-cs-cv
16 Apr 2026
Model Releases

PersonaVLM: Long-Term Personalized Multimodal LLMs

DGX agent

arXiv:2604.13074v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) serve as daily assistants for millions. However, their ability to generate responses aligned with individual pr

model-releasesarxiv-cs-cl
16 Apr 2026
Research

Physics-Informed Neural Networks for Solving Derivative-Constrained PDEs

DGX agent

arXiv:2604.13723v1 Announce Type: new Abstract: Physics-Informed Neural Networks (PINNs) recast PDE solving as an optimisation problem in function space by minimising a residual-based objective, yet m

researcharxiv-cs-lg
16 Apr 2026
Applications

Power Transform Revisited: Numerically Stable, and Federated

DGX agent

arXiv:2510.04995v3 Announce Type: replace Abstract: Power transforms are popular parametric methods for making data more Gaussian-like, and are widely used as preprocessing steps in statistical analys

applicationsarxiv-cs-lg
16 Apr 2026
Model Releases

Red Skills or Blue Skills? A Dive Into Skills Published on ClawHub

DGX agent

arXiv:2604.13064v1 Announce Type: new Abstract: Skill ecosystems have emerged as an increasingly important layer in Large Language Model (LLM) agent systems, enabling reusable task packaging, public d

model-releasesarxiv-cs-cl
16 Apr 2026
Safety

Representation over Routing: Overcoming Surrogate Hacking in Multi-Timescale PPO

DGX agent

arXiv:2604.13517v1 Announce Type: new Abstract: Temporal credit assignment in reinforcement learning has long been a central challenge. Inspired by the multi-timescale encoding of the dopamine system

safetyarxiv-cs-lg
16 Apr 2026
Research

Saber: An Efficient Sampling with Adaptive Acceleration and Backtracking Enhanced Remasking for Diffusion Language Model

DGX agent

arXiv:2510.18165v2 Announce Type: replace-cross Abstract: Diffusion language models (DLMs) are emerging as a powerful and promising alternative to the dominant autoregressive paradigm, offering inhere

researcharxiv-cs-cl
16 Apr 2026
Hardware

Scalable Spatiotemporal Inference with Biased Scan Attention Transformer Neural Processes

DGX agent

arXiv:2506.09163v2 Announce Type: replace Abstract: Neural Processes (NPs) are a rapidly evolving class of models designed to directly model the posterior predictive distribution of stochastic process

hardwarearxiv-cs-lg
16 Apr 2026
Model Releases

Scaling Test-Time Compute to Achieve IOI Gold Medal with Open-Weight Models

DGX agent

arXiv:2510.14232v2 Announce Type: replace-cross Abstract: Competitive programming has become a rigorous benchmark for evaluating the reasoning and problem-solving capabilities of large language models

model-releasesarxiv-cs-cl
16 Apr 2026
Research

SceneGlue: Scene-Aware Transformer for Feature Matching without Scene-Level Annotation

DGX agent

arXiv:2604.13941v1 Announce Type: new Abstract: Local feature matching plays a critical role in understanding the correspondence between cross-view images. However, traditional methods are constrained

researcharxiv-cs-cv
16 Apr 2026
Safety

Self-adaptive Multi-Access Edge Architectures: A Robotics Case

DGX agent

arXiv:2604.13542v1 Announce Type: new Abstract: The growth of compute-intensive AI tasks highlights the need to mitigate the processing costs and improve performance and energy efficiency. This necess

safetyarxiv-cs-ro
16 Apr 2026
Model Releases

Self-Organizing Maps with Optimized Latent Positions

DGX agent

arXiv:2604.13622v1 Announce Type: new Abstract: Self-Organizing Maps (SOM) are a classical method for unsupervised learning, vector quantization, and topographic mapping of high-dimensional data. Howe

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

SemAttNet: Towards Attention-based Semantic Aware Guided Depth Completion

DGX agent

arXiv:2204.13635v2 Announce Type: replace Abstract: Depth completion involves recovering a dense depth map from a sparse map and an RGB image. Recent approaches focus on utilizing color images as guid

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

SocialMirror: Reconstructing 3D Human Interaction Behaviors from Monocular Videos with Semantic and Geometric Guidance

DGX agent

arXiv:2604.13581v1 Announce Type: new Abstract: Accurately reconstructing human behavior in close-interaction scenarios is crucial for enabling realistic virtual interactions in augmented reality, pre

model-releasesarxiv-cs-cv
16 Apr 2026
Applications

Spectral Thompson sampling

DGX agent

arXiv:2604.13739v1 Announce Type: new Abstract: Thompson Sampling (TS) has attracted a lot of interest due to its good empirical performance, in particular in the computational advertising. Though suc

applicationsarxiv-cs-lg
16 Apr 2026
Research

Stability Principle Underlying Passive Dynamic Walking of Rimless Wheel

DGX agent

arXiv:2604.13530v1 Announce Type: new Abstract: Rimless wheels are known as the simplest model for passive dynamic walking. It is known that the passive gait generated only by gravity effect always be

researcharxiv-cs-ro
16 Apr 2026
Research

Swap Regret Minimization Through Response-Based Approachability

DGX agent

arXiv:2602.06264v2 Announce Type: replace Abstract: We consider the problem of minimizing different notions of swap regret in online optimization. These forms of regret are tightly connected to correl

researcharxiv-cs-lg
16 Apr 2026
Model Releases

Syn-TurnTurk: A Synthetic Dataset for Turn-Taking Prediction in Turkish Dialogues

DGX agent

arXiv:2604.13620v1 Announce Type: new Abstract: Managing natural dialogue timing is a significant challenge for voice-based chatbots. Most current systems usually rely on simple silence detection, whi

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Synthesis and Deployment of Maximal Robust Control Barrier Functions through Adversarial Reinforcement Learning

DGX agent

arXiv:2604.13192v1 Announce Type: cross Abstract: Robust control barrier functions (CBFs) provide a principled mechanism for smooth safety enforcement under worst-case disturbances. However, existing

model-releasesarxiv-cs-ro
16 Apr 2026
Model Releases

TLoRA+: A Low-Rank Parameter-Efficient Fine-Tuning Method for Large Language Models

DGX agent

arXiv:2604.13368v1 Announce Type: new Abstract: Fine-tuning large language models (LLMs) aims to adapt pre-trained models to specific tasks using relatively small and domain-specific datasets. Among P

model-releasesarxiv-cs-cl
16 Apr 2026
Agents

Training-Free Test-Time Contrastive Learning for Large Language Models

DGX agent

arXiv:2604.13552v1 Announce Type: new Abstract: Large language models (LLMs) demonstrate strong reasoning capabilities, but their performance often degrades under distribution shift. Existing test-tim

agentsarxiv-cs-cl
16 Apr 2026
Model Releases

TREX: Automating LLM Fine-tuning via Agent-Driven Tree-based Exploration

DGX agent

arXiv:2604.14116v1 Announce Type: cross Abstract: While Large Language Models (LLMs) have empowered AI research agents to perform isolated scientific tasks, automating complex, real-world workflows, s

model-releasesarxiv-cs-cl
16 Apr 2026
Research

Two Pathways to Truthfulness: On the Intrinsic Encoding of LLM Hallucinations

DGX agent

arXiv:2601.07422v2 Announce Type: replace Abstract: Despite their impressive capabilities, large language models (LLMs) frequently generate hallucinations. Previous work shows that their internal stat

researcharxiv-cs-cl
16 Apr 2026
Research

UHR-BAT: Budget-Aware Token Compression Vision-Language model for Ultra-High-Resolution Remote Sensing

DGX agent

arXiv:2604.13565v1 Announce Type: new Abstract: Ultra-high-resolution (UHR) remote sensing imagery couples kilometer-scale context with query-critical evidence that may occupy only a few pixels. Such

researcharxiv-cs-cv
16 Apr 2026
Model Releases

UniBlendNet: Unified Global, Multi-Scale, and Region-Adaptive Modeling for Ambient Lighting Normalization

DGX agent

arXiv:2604.13383v1 Announce Type: new Abstract: Ambient Lighting Normalization (ALN) aims to restore images degraded by complex, spatially varying illumination conditions. Existing methods, such as IF

model-releasesarxiv-cs-cv
16 Apr 2026
Research

VibeFlow: Versatile Video Chroma-Lux Editing through Self-Supervised Learning

DGX agent

arXiv:2604.13425v1 Announce Type: new Abstract: Video chroma-lux editing, which aims to modify illumination and color while preserving structural and temporal fidelity, remains a significant challenge

researcharxiv-cs-cv
16 Apr 2026
Agents

Vision-and-Language Navigation for UAVs: Progress, Challenges, and a Research Roadmap

DGX agent

arXiv:2604.13654v1 Announce Type: new Abstract: Vision-and-Language Navigation for Unmanned Aerial Vehicles (UAV-VLN) represents a pivotal challenge in embodied artificial intelligence, focused on ena

agentsarxiv-cs-ro
16 Apr 2026
Model Releases

When 'YES' Meets 'BUT': Can Large Models Comprehend Contradictory Humor Through Comparative Reasoning?

DGX agent

arXiv:2503.23137v2 Announce Type: replace-cross Abstract: Understanding humor-particularly when it involves complex, contradictory narratives that require comparative reasoning-remains a significant c

model-releasesarxiv-cs-cl
16 Apr 2026
← Previous
1…218219220221222…233
Next →