AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
Human
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,295 results
15 Apr 2026

Robust Optimization for Mitigating Reward Hacking with Correlated Proxies

SafetyDGX agent

arXiv:2604.12086v1 Announce Type: new Abstract: Designing robust reinforcement learning (RL) agents in the presence of imperfect reward signals remains a core challenge. In practice, agents are often

Robust Reasoning and Learning with Brain-Inspired Representations under Hardware-Induced Nonlinearities

ResearchDGX agent

arXiv:2604.12079v1 Announce Type: cross Abstract: Traditional machine learning depends on high-precision arithmetic and near-ideal hardware assumptions, which is increasingly challenged by variability

Robust Semi-Supervised Temporal Intrusion Detection for Adversarial Cloud Networks

ApplicationsDGX agent

arXiv:2604.12655v1 Announce Type: new Abstract: Cloud networks increasingly rely on machine learning based Network Intrusion Detection Systems to defend against evolving cyber threats. However, real-w

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

RoleMAG: Learning Neighbor Roles in Multimodal Graphs

ResearchDGX agent

arXiv:2604.12271v1 Announce Type: new Abstract: Multimodal attributed graphs (MAGs) combine multimodal node attributes with structured relations. However, existing methods usually perform shared messa

ROSE: An Intent-Centered Evaluation Metric for NL2SQL

ResearchDGX agent

arXiv:2604.12988v1 Announce Type: cross Abstract: Execution Accuracy (EX), the widely used metric for evaluating the effectiveness of Natural Language to SQL (NL2SQL) solutions, is becoming increasing

Round-Trip Translation Reveals What Frontier Multilingual Benchmarks Miss

Model ReleasesDGX agent

arXiv:2604.12911v1 Announce Type: cross Abstract: Multilingual benchmarks guide the development of frontier models. Yet multilingual evaluations reported by frontier models are structured similar to p

RPG-SAM: Reliability-Weighted Prototypes and Geometric Adaptive Threshold Selection for Training-Free One-Shot Polyp Segmentation

Model ReleasesDGX agent

arXiv:2603.07436v2 Announce Type: replace Abstract: Training-free one-shot segmentation offers a scalable alternative to expert annotations where knowledge is often transferred from support images and

RPRA: Predicting an LLM-Judge for Efficient but Performant Inference

TutorialsDGX agent

arXiv:2604.12634v1 Announce Type: new Abstract: Large language models (LLMs) face a fundamental trade-off between computational efficiency (e.g., number of parameters) and output quality, especially w

RSGMamba: Reliability-Aware Self-Gated State Space Model for Multimodal Semantic Segmentation

ResearchDGX agent

arXiv:2604.12319v1 Announce Type: new Abstract: Multimodal semantic segmentation has emerged as a powerful paradigm for enhancing scene understanding by leveraging complementary information from multi

Safe-FedLLM: Delving into the Safety of Federated Large Language Models

Local AiDGX agent

arXiv:2601.07177v2 Announce Type: replace-cross Abstract: Federated learning (FL) addresses privacy and data-silo issues in the training of large language models (LLMs). Most prior work focuses on imp

Safe reinforcement learning with online filtering for fatigue-predictive human-robot task planning and allocation in production

ApplicationsDGX agent

arXiv:2604.12667v1 Announce Type: new Abstract: Human-robot collaborative manufacturing, a core aspect of Industry 5.0, emphasizes ergonomics to enhance worker well-being. This paper addresses the dyn

Safe-SAIL: Towards a Fine-grained Safety Landscape of Large Language Models via Sparse Autoencoder Interpretation Framework

Model ReleasesDGX agent

arXiv:2509.18127v3 Announce Type: replace-cross Abstract: Sparse autoencoders (SAEs) enable interpretability research by decomposing entangled model activations into monosemantic features. However, un

Safety Training Modulates Harmful Misalignment Under On-Policy RL, But Direction Depends on Environment Design

SafetyDGX agent

arXiv:2604.12500v1 Announce Type: new Abstract: Specification gaming under Reinforcement Learning (RL) is known to cause LLMs to develop sycophantic, manipulative, or deceptive behavior, yet the condi

SAM3-I: Segment Anything with Instructions

SafetyDGX agent

arXiv:2512.04585v3 Announce Type: replace Abstract: Segment Anything Model 3 (SAM3) advances open-vocabulary segmentation through promptable concept segmentation, enabling users to segment all instanc

Sample Complexity of Autoregressive Reasoning: Chain-of-Thought vs. End-to-End

TutorialsDGX agent

arXiv:2604.12013v1 Announce Type: new Abstract: Modern large language models generate text autoregressively, producing tokens one at a time. To study the learnability of such systems, Joshi et al. (CO

Scaffold-Conditioned Preference Triplets for Controllable Molecular Optimization with Large Language Models

SafetyDGX agent

arXiv:2604.12350v1 Announce Type: cross Abstract: Molecular property optimization is central to drug discovery, yet many deep learning methods rely on black-box scoring and offer limited control over

Scalable and General Whole-Body Control for Cross-Humanoid Locomotion

SafetyDGX agent

arXiv:2602.05791v2 Announce Type: replace Abstract: Learning-based whole-body controllers have become a key driver for humanoid robots, yet most existing approaches require robot-specific training. In

Scalable Trajectory Generation for Whole-Body Mobile Manipulation

HardwareDGX agent

arXiv:2604.12565v1 Announce Type: cross Abstract: Robots deployed in unstructured environments must coordinate whole-body motion -- simultaneously moving a mobile base and arm -- to interact with the

Scalable Verification of Neural Control Barrier Functions Using Linear Bound Propagation

SafetyDGX agent

arXiv:2511.06341v2 Announce Type: replace Abstract: Control barrier functions (CBFs) are a popular tool for safety certification of nonlinear dynamical control systems. Recently, CBFs represented as n

Scale-aware Message Passing For Graph Node Classification

Model ReleasesDGX agent

arXiv:2411.19392v3 Announce Type: replace Abstract: Most Graph Neural Networks (GNNs) operate at the first-order scale, even though multi-scale representations are known to be crucial in domains such

Scaling Exposes the Trigger: Input-Level Backdoor Detection in Text-to-Image Diffusion Models via Cross-Attention Scaling

ResearchDGX agent

arXiv:2604.12446v1 Announce Type: cross Abstract: Text-to-image (T2I) diffusion models have achieved remarkable success in image synthesis, but their reliance on large-scale data and open ecosystems i

Scaling In-Context Segmentation with Hierarchical Supervision

ResearchDGX agent

arXiv:2604.12752v1 Announce Type: new Abstract: In-context learning (ICL) enables medical image segmentation models to adapt to new anatomical structures from limited examples, reducing the clinical a

SceneCritic: A Symbolic Evaluator for 3D Indoor Scene Synthesis

ResearchDGX agent

arXiv:2604.13035v1 Announce Type: cross Abstract: Large Language Models (LLMs) and Vision-Language Models (VLMs) increasingly generate indoor scenes through intermediate structures such as layouts and

Schema-Adaptive Tabular Representation Learning with LLMs for Generalizable Multimodal Clinical Reasoning

SafetyDGX agent

arXiv:2604.11835v1 Announce Type: cross Abstract: Machine learning for tabular data remains constrained by poor schema generalization, a challenge rooted in the lack of semantic understanding of struc

SCRIPT: A Subcharacter Compositional Representation Injection Module for Korean Pre-Trained Language Models

ResearchDGX agent

arXiv:2604.12377v1 Announce Type: cross Abstract: Korean is a morphologically rich language with a featural writing system in which each character is systematically composed of subcharacter units know

SDFed: Bridging Local Global Discrepancy via Subspace Refinement and Divergence Control in Federated Prompt Learning

TutorialsDGX agent

arXiv:2602.08590v4 Announce Type: replace Abstract: Vision-language pretrained models offer strong transferable representations, yet adapting them in privacy-sensitive multi-party settings is challeng

SEATrack: Simple, Efficient, and Adaptive Multimodal Tracker

Model ReleasesDGX agent

arXiv:2604.12502v1 Announce Type: cross Abstract: Parameter-efficient fine-tuning (PEFT) in multimodal tracking reveals a concerning trend where recent performance gains are often achieved at the cost

SecureWebArena: A Holistic Security Evaluation Benchmark for LVLM-based Web Agents

Model ReleasesDGX agent

arXiv:2510.10073v2 Announce Type: replace-cross Abstract: Large vision-language model (LVLM)-based web agents are emerging as powerful tools for automating complex online tasks. However, when deployed

Security and Resilience in Autonomous Vehicles: A Proactive Design Approach

AgentsDGX agent

arXiv:2604.12408v1 Announce Type: cross Abstract: Autonomous vehicles (AVs) promise efficient, clean and cost-effective transportation systems, but their reliance on sensors, wireless communications,

See, Point, Refine: Multi-Turn Approach to GUI Grounding with Visual Feedback

Model ReleasesDGX agent

arXiv:2604.13019v1 Announce Type: new Abstract: Computer Use Agents (CUAs) fundamentally rely on graphical user interface (GUI) grounding to translate language instructions into executable screen acti

SeedPrints: Fingerprints Can Even Tell Which Seed Your Large Language Model Was Trained From

Model ReleasesDGX agent

arXiv:2509.26404v2 Announce Type: replace-cross Abstract: Fingerprinting Large Language Models (LLMs)is essential for provenance verification and model attribution. Existing fingerprinting methods are

Self-Adversarial One Step Generation via Condition Shifting

Model ReleasesDGX agent

arXiv:2604.12322v1 Announce Type: new Abstract: The push for efficient text to image synthesis has moved the field toward one step sampling, yet existing methods still face a three way tradeoff among

Self-Distillation Zero: Self-Revision Turns Binary Rewards into Dense Supervision

SafetyDGX agent

arXiv:2604.12002v1 Announce Type: new Abstract: Current post-training methods in verifiable settings fall into two categories. Reinforcement learning (RLVR) relies on binary rewards, which are broadly

Self-Monitoring Benefits from Structural Integration: Lessons from Metacognition in Continuous-Time Multi-Timescale Agents

Model ReleasesDGX agent

arXiv:2604.11914v1 Announce Type: new Abstract: Self-monitoring capabilities -- metacognition, self-prediction, and subjective duration -- are often proposed as useful additions to reinforcement learn

SEW: Self-Evolving Agentic Workflows for Automated Code Generation

Model ReleasesDGX agent

arXiv:2505.18646v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have demonstrated effectiveness in code generation tasks. To enable LLMs to address more complex coding challenge

Should There be a Teacher In-the-Loop? A Study of Generative AI Personalized Tasks Middle School

ResearchDGX agent

arXiv:2602.15876v1 Announce Type: cross Abstract: Adapting instruction to the fine-grained needs of individual students is a powerful application of recent advances in large language models. These gen

Siamese Foundation Models for Crystal Structure Prediction

TutorialsDGX agent

arXiv:2503.10471v2 Announce Type: replace-cross Abstract: Predicting crystal structures from chemical compositions is a fundamental challenge in materials discovery, complicated by complex 3D geometri

Silo-Bench: A Scalable Environment for Evaluating Distributed Coordination in Multi-Agent LLM Systems

Model ReleasesDGX agent

arXiv:2603.01045v2 Announce Type: replace-cross Abstract: Large language models are increasingly deployed in multi-agent systems to overcome context limitations by distributing information across agen

Simulation as Supervision: Mechanistic Pretraining for Scientific Discovery

SafetyDGX agent

arXiv:2507.08977v4 Announce Type: replace-cross Abstract: Scientific modeling faces a tradeoff between the interpretability of mechanistic theory and the predictive power of machine learning. While ex

SinkSAM-Net: Knowledge-Driven Self-Supervised Sinkhole Segmentation Using Topographic Priors and Segment Anything Model

Model ReleasesDGX agent

arXiv:2410.01473v2 Announce Type: replace Abstract: Soil sinkholes significantly influence soil degradation, infrastructure vulnerability, and landscape evolution. However, their irregular shapes, com

SIR-Bench: Evaluating Investigation Depth in Security Incident Response Agents

Model ReleasesDGX agent

arXiv:2604.12040v1 Announce Type: cross Abstract: We present SIR-Bench, a benchmark of 794 test cases for evaluating autonomous security incident response agents that distinguishes genuine forensic in

SIRI-Bench: Challenging VLMs' Spatial Intelligence through Complex Reasoning Tasks

Model ReleasesDGX agent

arXiv:2506.14512v4 Announce Type: replace Abstract: Large Language Models (LLMs) have undergone rapid progress, largely attributed to reinforcement learning on complex reasoning tasks. In contrast, wh

Skill-informed Data-driven Haptic Nudges for High-dimensional Human Motor Learning

SafetyDGX agent

arXiv:2603.12583v2 Announce Type: replace Abstract: In this work, we propose a data-driven framework to design optimal haptic nudge feedback leveraging the learner's estimated skill to address the cha

SmellNet: A Large-scale Dataset for Real-world Smell Recognition

ApplicationsDGX agent

arXiv:2506.00239v5 Announce Type: replace Abstract: The ability of AI to sense and identify various substances based on their smell alone can have profound impacts on allergen detection (e.g. smelling

SOAR: Self-Correction for Optimal Alignment and Refinement in Diffusion Models

SafetyDGX agent

arXiv:2604.12617v1 Announce Type: cross Abstract: The post-training pipeline for diffusion models currently has two stages: supervised fine-tuning (SFT) on curated data and reinforcement learning (RL)

Social Learning Strategies for Evolved Virtual Soft Robots

TutorialsDGX agent

arXiv:2604.12482v1 Announce Type: cross Abstract: Optimizing the body and brain of a robot is a coupled challenge: the morphology determines what control strategies are effective, while the control pa

Socrates Loss: Unifying Confidence Calibration and Classification by Leveraging the Unknown

Model ReleasesDGX agent

arXiv:2604.12245v1 Announce Type: cross Abstract: Deep neural networks, despite their high accuracy, often exhibit poor confidence calibration, limiting their reliability in high-stakes applications.

SOLARIS: Speculative Offloading of Latent-bAsed Representation for Inference Scaling

ResearchDGX agent

arXiv:2604.12110v1 Announce Type: new Abstract: Recent advances in recommendation scaling laws have led to foundation models of unprecedented complexity. While these models offer superior performance,

SpanKey: Dynamic Key Space Conditioning for Neural Network Access Control

ResearchDGX agent

arXiv:2604.12254v1 Announce Type: cross Abstract: SpanKey is a lightweight way to gate inference without encrypting weights or chasing leaderboard accuracy on gated inference. The idea is to condition

Sparse Growing Transformer: Training-Time Sparse Depth Allocation via Progressive Attention Looping

Model ReleasesDGX agent

arXiv:2603.23998v2 Announce Type: replace Abstract: Existing approaches to increasing the effective depth of Transformers predominantly rely on parameter reuse, extending computation through recursive

SparseWorld-TC: Trajectory-Conditioned Sparse Occupancy World Model

Model ReleasesDGX agent

arXiv:2511.22039v3 Announce Type: replace Abstract: This paper introduces a novel architecture for trajectory-conditioned forecasting of future 3D scene occupancy. In contrast to methods that rely on

Spatial Atlas: Compute-Grounded Reasoning for Spatial-Aware Research Agent Benchmarks

Model ReleasesDGX agent

arXiv:2604.12102v1 Announce Type: new Abstract: We introduce compute-grounded reasoning (CGR), a design paradigm for spatial-aware research agents in which every answerable sub-problem is resolved by

Spatial-Spectral Adaptive Fidelity and Noise Prior Reduction Guided Hyperspectral Image Denoising

Local AiDGX agent

arXiv:2604.12600v1 Announce Type: new Abstract: The core challenge of hyperspectral image denoising is striking the right balance between data fidelity and noise prior modeling. Most existing methods

Speaker effects in language comprehension: An integrative model of language and speaker processing

ResearchDGX agent

arXiv:2412.07238v3 Announce Type: replace Abstract: The identity of a speaker influences language comprehension through modulating perception and expectation. This review explores speaker effects and

SpecBound: Adaptive Bounded Self-Speculation with Layer-wise Confidence Calibration

ResearchDGX agent

arXiv:2604.12247v1 Announce Type: cross Abstract: Speculative decoding has emerged as a promising approach to accelerate autoregressive inference in large language models (LLMs). Self-draft methods, w

SpecBranch: Speculative Decoding via Hybrid Drafting and Rollback-Aware Branch Parallelism

ApplicationsDGX agent

arXiv:2506.01979v4 Announce Type: replace-cross Abstract: Recently, speculative decoding (SD) has emerged as a promising technique to accelerate LLM inference by employing a small draft model to propo

StableSketcher: Enhancing Diffusion Model for Pixel-based Sketch Generation via Visual Question Answering Feedback

SafetyDGX agent

arXiv:2510.20093v2 Announce Type: replace-cross Abstract: Although recent advancements in diffusion models have significantly enriched the quality of generated images, challenges remain in synthesizin

StoryScope: Investigating idiosyncrasies in AI fiction

Model ReleasesDGX agent

arXiv:2604.03136v4 Announce Type: replace Abstract: As AI-generated fiction becomes increasingly prevalent, questions of authorship and originality are becoming central to how written work is evaluate

Stress Detection Using Wearable Physiological and Sociometric Sensors

ResearchDGX agent

arXiv:2604.12746v1 Announce Type: new Abstract: Stress remains a significant social problem for individuals in modern societies. This paper presents a machine learning approach for the automatic detec

StructDiff: A Structure-Preserving and Spatially Controllable Diffusion Model for Single-Image Generation

TutorialsDGX agent

arXiv:2604.12575v1 Announce Type: new Abstract: This paper introduces StructDiff, a generative framework based on a single-scale diffusion model for single-image generation. Single-image generation ai

← Previous
1…932933934935936…989
Next →