AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
Human
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,295 results
16 Apr 2026

Golden Handcuffs make safer AI agents

SafetyDGX agent

arXiv:2604.13609v1 Announce Type: new Abstract: Reinforcement learners can attain high reward through novel unintended strategies. We study a Bayesian mitigation for general environments: we expand th

Gradient Descent's Last Iterate is Often (slightly) Suboptimal

ResearchDGX agent

arXiv:2604.13870v1 Announce Type: cross Abstract: We consider the well-studied setting of minimizing a convex Lipschitz function using either gradient descent (GD) or its stochastic variant (SGD), and

Granularity-Aware Transfer for Tree Instance Segmentation in Synthetic and Real Forests

ResearchDGX agent

arXiv:2604.13722v1 Announce Type: new Abstract: We address the challenge of synthetic-to-real transfer in forestry perception where real data have only coarse Tree labels while synthetic data provide

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Graph In-Context Operator Networks for Generalizable Spatiotemporal Prediction

ApplicationsDGX agent

arXiv:2603.12725v3 Announce Type: replace Abstract: In-context operator learning enables neural networks to infer solution operators from contextual examples without weight updates. While prior work h

Graph Propagated Projection Unlearning: A Unified Framework for Vision and Audio Discriminative Models

ResearchDGX agent

arXiv:2604.13127v1 Announce Type: new Abstract: The need to selectively and efficiently erase learned information from deep neural networks is becoming increasingly important for privacy, regulatory c

GRITS: A Spillage-Aware Guided Diffusion Policy for Robot Food Scooping Tasks

SafetyDGX agent

arXiv:2510.00573v2 Announce Type: replace Abstract: Robotic food scooping is a critical manipulation skill for food preparation and service robots. However, existing robot learning algorithms, especia

Guided Transfer Learning for Discrete Diffusion Models

ApplicationsDGX agent

arXiv:2512.10877v4 Announce Type: replace Abstract: Discrete diffusion models (DMs) have achieved strong performance in language and other discrete domains, offering a compelling alternative to autore

HAMLET: Switch your Vision-Language-Action Model into a History-Aware Policy

SafetyDGX agent

arXiv:2510.00695v3 Announce Type: replace-cross Abstract: Inherently, robotic manipulation tasks are history-dependent: leveraging past context could be beneficial. However, most existing Vision-Langu

Hardware-Efficient Neuro-Symbolic Networks with the Exp-Minus-Log Operator

SafetyDGX agent

arXiv:2604.13871v1 Announce Type: new Abstract: Deep neural networks (DNNs) deliver state-of-the-art accuracy on regression and classification tasks, yet two structural deficits persistently obstruct

Heavy-Tailed Class-Conditional Priors for Long-Tailed Generative Modeling

SafetyDGX agent

arXiv:2509.02154v2 Announce Type: replace-cross Abstract: Variational Autoencoders (VAEs) with global priors trained under an imbalanced empirical class distribution can lead to underrepresentation of

Hessian-Enhanced Token Attribution (HETA): Interpreting Autoregressive LLMs

Model ReleasesDGX agent

arXiv:2604.13258v1 Announce Type: new Abstract: Attribution methods seek to explain language model predictions by quantifying the contribution of input tokens to generated outputs. However, most exist

Heuristic Style Transfer for Real-Time, Efficient Weather Attribute Detection

ResearchDGX agent

arXiv:2604.13947v1 Announce Type: new Abstract: We present lightweight and efficient architectures to detect weather conditions from RGB images, predicting the weather type (sunny, rain, snow, fog) an

Hierarchical DLO Routing with Reinforcement Learning and In-Context Vision-language Models

AgentsDGX agent

arXiv:2510.19268v2 Announce Type: replace-cross Abstract: Long-horizon routing tasks of deformable linear objects (DLOs), such as cables and ropes, are common in industrial assembly lines and everyday

Hierarchical Reinforcement Learning with Augmented Step-Level Transitions for LLM Agents

ResearchDGX agent

arXiv:2604.05808v2 Announce Type: replace-cross Abstract: Large language model (LLM) agents have demonstrated strong capabilities in complex interactive decision-making tasks. However, existing LLM ag

Hierarchical Reinforcement Learning with Runtime Safety Shielding for Power Grid Operation

Model ReleasesDGX agent

arXiv:2604.14032v1 Announce Type: cross Abstract: Reinforcement learning has shown promise for automating power-grid operation tasks such as topology control and congestion management. However, its de

HINTBench: Horizon-agent Intrinsic Non-attack Trajectory Benchmark

Model ReleasesDGX agent

arXiv:2604.13954v1 Announce Type: new Abstract: Existing agent-safety evaluation has focused mainly on externally induced risks. Yet agents may still enter unsafe trajectories under benign conditions.

HiProto: Hierarchical Prototype Learning for Interpretable Object Detection Under Low-quality Conditions

ResearchDGX agent

arXiv:2604.13981v1 Announce Type: new Abstract: Interpretability is essential for deploying object detection systems in critical applications, especially under low-quality imaging conditions that degr

HiVLA: A Visual-Grounded-Centric Hierarchical Embodied Manipulation System

ApplicationsDGX agent

arXiv:2604.14125v1 Announce Type: new Abstract: While end-to-end Vision-Language-Action (VLA) models offer a promising paradigm for robotic manipulation, fine-tuning them on narrow control data often

How Can We Synthesize High-Quality Pretraining Data? A Systematic Study of Prompt Design, Generator Model, and Source Data

ResearchDGX agent

arXiv:2604.13977v1 Announce Type: new Abstract: Synthetic data is a standard component in training large language models, yet systematic comparisons across design dimensions, including rephrasing stra

(How) Learning Rates Regulate Catastrophic Overtraining

ResearchDGX agent

arXiv:2604.13627v1 Announce Type: cross Abstract: Supervised fine-tuning (SFT) is a common first stage of LLM post-training, teaching the model to follow instructions and shaping its behavior as a hel

HUANet: Hard-Constrained Unrolled ADMM for Constrained Convex Optimization

ResearchDGX agent

arXiv:2604.13179v1 Announce Type: cross Abstract: This paper presents HUANet, a constrained deep neural network architecture that unrolls the iterations of the Alternating Direction Method of Multipli

Hybrid Approach for Enhancing Lesion Segmentation in Fundus Images

ResearchDGX agent

arXiv:2509.25549v2 Announce Type: replace Abstract: Choroidal nevi are common benign pigmented lesions in the eye, with a small risk of transforming into melanoma. Early detection is critical to impro

Hybrid Attention Model Using Feature Decomposition and Knowledge Distillation for Glucose Forecasting

ApplicationsDGX agent

arXiv:2411.10703v3 Announce Type: replace Abstract: The availability of continuous glucose monitors as over-the-counter commodities have created a unique opportunity to monitor a person's blood glucos

Hybrid Retrieval for COVID-19 Literature: Comparing Rank Fusion and Projection Fusion with Diversity Reranking

Model ReleasesDGX agent

arXiv:2604.13728v1 Announce Type: cross Abstract: We present a hybrid retrieval system for COVID-19 scientific literature, evaluated on the TREC-COVID benchmark (171,332 papers, 50 expert queries). Th

Hydra: Unifying Document Retrieval and Generation in a Single Vision-Language Model

HardwareDGX agent

arXiv:2603.28554v2 Announce Type: replace Abstract: Visual document understanding typically requires separate retrieval and generation models, doubling memory and system complexity. We present Hydra,

ID and Graph View Contrastive Learning with Multi-View Attention Fusion for Sequential Recommendation

Model ReleasesDGX agent

arXiv:2604.14114v1 Announce Type: cross Abstract: Sequential recommendation has become increasingly prominent in both academia and industry, particularly in e-commerce. The primary goal is to extract

Identifiability of Potentially Degenerate Gaussian Mixture Models With Piecewise Affine Mixing

ResearchDGX agent

arXiv:2604.13218v1 Announce Type: cross Abstract: Causal representation learning (CRL) aims to identify the underlying latent variables from high-dimensional observations, even when variables are depe

IGen: Scalable Data Generation for Robot Learning from Open-World Images

SafetyDGX agent

arXiv:2512.01773v2 Announce Type: replace Abstract: The rise of generalist robotic policies has created an exponential demand for large-scale training data. However, on-robot data collection is labor-

Indexing Multimodal Language Models for Large-scale Image Retrieval

ResearchDGX agent

arXiv:2604.13268v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have demonstrated strong cross-modal reasoning capabilities, yet their potential for vision-only tasks remain

IndicDB -- Benchmarking Multilingual Text-to-SQL Capabilities in Indian Languages

Model ReleasesDGX agent

arXiv:2604.13686v1 Announce Type: new Abstract: While Large Language Models (LLMs) have significantly advanced Text-to-SQL performance, existing benchmarks predominantly focus on Western contexts and

InfiniteScienceGym: An Unbounded, Procedurally-Generated Benchmark for Scientific Analysis

Model ReleasesDGX agent

arXiv:2604.13201v1 Announce Type: new Abstract: Large language models are emerging as scientific assistants, but evaluating their ability to reason from empirical data remains challenging. Benchmarks

Interpretable Stylistic Variation in Human and LLM Writing Across Genres, Models, and Decoding Strategies

TutorialsDGX agent

arXiv:2604.14111v1 Announce Type: new Abstract: Large Language Models (LLMs) are now capable of generating highly fluent, human-like text. They enable many applications, but also raise concerns such a

Irregularly Sampled Time Series Interpolation for Binary Evolution Simulations Using Dynamic Time Warping

SafetyDGX agent

arXiv:2604.13604v1 Announce Type: cross Abstract: Binary stellar evolution simulations are computationally expensive. Stellar population synthesis relies on these detailed evolution models at a fundam

IWLV-Ramayana: A Sarga-Aligned Parallel Corpus of Valmiki's Ramayana Across Indian Languages

ApplicationsDGX agent

arXiv:2604.13078v1 Announce Type: new Abstract: The Ramayana is among the most influential literary traditions of South and Southeast Asia, transmitted across numerous linguistic and cultural contexts

Joint Representation Learning and Clustering via Gradient-Based Manifold Optimization

Model ReleasesDGX agent

arXiv:2604.13484v1 Announce Type: cross Abstract: Clustering and dimensionality reduction have been crucial topics in machine learning and computer vision. Clustering high-dimensional data has been ch

Jump-Start Reinforcement Learning with Vision-Language-Action Regularization

SafetyDGX agent

arXiv:2604.13733v1 Announce Type: new Abstract: Reinforcement learning (RL) enables high-frequency, closed-loop control for robotic manipulation, but scaling to long-horizon tasks with sparse or imper

Just Use XML: Revisiting Joint Translation and Label Projection

ResearchDGX agent

arXiv:2603.12021v2 Announce Type: replace Abstract: Label projection is an effective technique for cross-lingual transfer, extending span-annotated datasets from a high-resource language to low-resour

KMMMU: Evaluation of Massive Multi-discipline Multimodal Understanding in Korean Language and Context

Model ReleasesDGX agent

arXiv:2604.13058v1 Announce Type: new Abstract: We introduce KMMMU, a native Korean benchmark for evaluating multimodal understanding in Korean cultural and institutional settings. KMMMU contains 3,46

KV Packet: Recomputation-Free Context-Independent KV Caching for LLMs

Model ReleasesDGX agent

arXiv:2604.13226v1 Announce Type: new Abstract: Large Language Models (LLMs) rely heavily on Key-Value (KV) caching to minimize inference latency. However, standard KV caches are context-dependent: re

Kwame 2.0: Human-in-the-Loop Generative AI Teaching Assistant for Large Scale Online Coding Education in Africa

TutorialsDGX agent

arXiv:2603.29159v2 Announce Type: replace Abstract: Providing timely and accurate learning support in large-scale online coding courses is challenging, particularly in resource-constrained contexts. W

L2D-Clinical: Learning to Defer for Adaptive Model Selection in Clinical Text Classification

Model ReleasesDGX agent

arXiv:2604.13285v1 Announce Type: new Abstract: Clinical text classification requires choosing between specialized fine-tuned models (BERT variants) and general-purpose large language models (LLMs), y

Language steering in latent space to mitigate unintended code-switching

Model ReleasesDGX agent

arXiv:2510.13849v3 Announce Type: replace Abstract: Multilingual Large Language Models (LLMs) often exhibit hallucinations such as unintended code-switching, reducing reliability in downstream tasks.

LaoBench: A Large-Scale Multidimensional Lao Benchmark for Large Language Models

Model ReleasesDGX agent

arXiv:2511.11334v3 Announce Type: replace Abstract: The rapid advancement of large language models (LLMs) has not been matched by their evaluation in low-resource languages, especially Southeast Asian

Learning-Based Estimation of Spatially Resolved Scatter Radiation Fields in Interventional Radiology

ResearchDGX agent

arXiv:2512.17654v3 Announce Type: replace Abstract: We present three variants of a lightweight, fully connected artificial neural network, suited for interactive estimation of three-dimensional, spati

Learning Class Difficulty in Imbalanced Histopathology Segmentation via Dynamic Focal Attention

SafetyDGX agent

arXiv:2604.13479v1 Announce Type: cross Abstract: Semantic segmentation of histopathology images under class imbalance is typically addressed through frequency-based loss reweighting, which implicitly

Learning Dynamics from Input-Output Data with Hamiltonian Gaussian Processes

ApplicationsDGX agent

arXiv:2511.05330v2 Announce Type: replace Abstract: Embedding non-restrictive prior knowledge, such as energy conservation laws, into learning methods is a key motive to construct physically consisten

Learning from Change: Predictive Models for Incident Prevention in a Regulated IT Environment

ApplicationsDGX agent

arXiv:2604.13462v1 Announce Type: cross Abstract: Effective IT change management is important for businesses that depend on software and services, particularly in highly regulated sectors such as fina

Learning Inference Concurrency in DynamicGate MLP Structural and Mathematical Justification

ResearchDGX agent

arXiv:2604.13546v1 Announce Type: new Abstract: Conventional neural networks strictly separate learning and inference because if parameters are updated during inference, outputs become unstable and ev

Learning Probabilistic Responsibility Allocations for Multi-Agent Interactions

SafetyDGX agent

arXiv:2604.13128v1 Announce Type: cross Abstract: Human behavior in interactive settings is shaped not only by individual objectives but also by shared constraints with others, such as safety. Underst

Learning Sewing Patterns via Latent Flow Matching of Implicit Fields

ResearchDGX agent

arXiv:2601.17740v2 Announce Type: replace Abstract: Sewing patterns define the structural foundation of garments and are essential for applications such as fashion design, fabrication, and physical si

Learning the Cue or Learning the Word? Analyzing Generalization in Metaphor Detection for Verbs

Model ReleasesDGX agent

arXiv:2604.13713v1 Announce Type: new Abstract: Metaphor detection models achieve strong benchmark performance, yet it remains unclear whether this reflects transferable generalization or lexical memo

LEGO-MOF: Equivariant Latent Manipulation for Editable, Generative, and Optimizable MOF Design

ResearchDGX agent

arXiv:2604.13520v1 Announce Type: new Abstract: Metal-organic frameworks (MOFs) are highly promising for carbon capture, yet navigating their vast design space remains challenging. Recent deep generat

LEO-RobotAgent: A General-purpose Robotic Agent for Language-driven Embodied Operator

AgentsDGX agent

arXiv:2512.10605v2 Announce Type: replace Abstract: We propose LEO-RobotAgent, a general-purpose language-driven intelligent agent framework for robots. Under this framework, LLMs can operate differen

Leveraging LLM-GNN Integration for Open-World Question Answering over Knowledge Graphs

Model ReleasesDGX agent

arXiv:2604.13979v1 Announce Type: new Abstract: Open-world Question Answering (OW-QA) over knowledge graphs (KGs) aims to answer questions over incomplete or evolving KGs. Traditional KGQA assumes a c

Linear Probe Accuracy Scales with Model Size and Benefits from Multi-Layer Ensembling

ResearchDGX agent

arXiv:2604.13386v1 Announce Type: new Abstract: Linear probes can detect when language models produce outputs they 'know' are wrong, a capability relevant to both deception and reward hacking. However

Lite Any Stereo: Efficient Zero-Shot Stereo Matching

ApplicationsDGX agent

arXiv:2511.16555v3 Announce Type: replace Abstract: Recent advances in stereo matching have focused on accuracy, often at the cost of significantly increased model size. Traditionally, the community h

LiveClawBench: Benchmarking LLM Agents on Complex, Real-World Assistant Tasks

Model ReleasesDGX agent

arXiv:2604.13072v1 Announce Type: new Abstract: LLM-based agents are increasingly expected to handle real-world assistant tasks, yet existing benchmarks typically evaluate them under isolated sources

Logical Phase Transitions: Understanding Collapse in LLM Logical Reasoning

ApplicationsDGX agent

arXiv:2601.02902v2 Announce Type: replace-cross Abstract: Symbolic logical reasoning is a critical yet underexplored capability of large language models (LLMs), providing reliable and verifiable decis

LongCoT: Benchmarking Long-Horizon Chain-of-Thought Reasoning

Model ReleasesDGX agent

arXiv:2604.14140v1 Announce Type: new Abstract: As language models are increasingly deployed for complex autonomous tasks, their ability to reason accurately over longer horizons becomes critical. An

LongVideo-R1: Smart Navigation for Low-cost Long Video Understanding

Model ReleasesDGX agent

arXiv:2602.20913v2 Announce Type: replace Abstract: This paper addresses the critical and underexplored challenge of long video understanding with low computational budgets. We propose LongVideo-R1, a

← Previous
1…918919920921922…989
Next →