AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,403
  • Agents7,557
  • Applications5,411
  • Concepts5
  • Hardware1,836
  • Industry6,170
  • Local Ai4,931
  • Model Releases23,900
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,403
  • Agents7,557
  • Applications5,411
  • Concepts5
  • Hardware1,836
  • Industry6,170
  • Local Ai4,931
  • Model Releases23,900
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
Human
88,403Total entries
1Added by human
88,402Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
62,897 results
4 Jun 2026

SemBlock: Semantic Boundary Dynamic Blocks for Diffusion LLMs

Local AiDGX agent

arXiv:2606.04964v1 Announce Type: new Abstract: Diffusion language models (DLMs) generate text through iterative denoising, and blockwise decoding improves their practicality by committing tokens in l

Semiparametric Preference Optimization: Your Language Model is Secretly a Single-Index Model

SafetyDGX agent

arXiv:2512.21917v3 Announce Type: replace-cross Abstract: Policy alignment to preference data typically assumes a known link function between observed preferences and latent rewards (e.g., Bradley-Ter

SePO: Self-Evolving Prompt Agent for System Prompt Optimization

AgentsDGX agent

arXiv:2606.04465v1 Announce Type: cross Abstract: System prompt optimization improves agent behavior without modifying the underlying model, yielding human-readable, model-agnostic instructions. Exist

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Sequential Data Poisoning in LLM Post-Training

ResearchDGX agent

arXiv:2606.04929v1 Announce Type: new Abstract: LLM post-training proceeds through multiple stages, e.g., supervised fine-tuning (SFT) followed by reinforcement learning from human feedback (RLHF) or

SFMambaNet: Spectral-Frequency Enhanced Selective State Space Model for Correspondence Pruning

ResearchDGX agent

arXiv:2606.04493v1 Announce Type: cross Abstract: Correspondence pruning aims to identify inliers from an initial set of correspondences. Most existing Graph Neural Network (GNN)-based methods rely on

SFMP: Fine-Grained, Hardware-Friendly and Search-Free Mixed-Precision Quantization for Large Language Models

ResearchDGX agent

arXiv:2602.01027v2 Announce Type: replace Abstract: Mixed-precision quantization is a promising approach for compressing large language models under tight memory budgets. However, existing mixed-preci

SharedRequest: Privacy-Preserving Model-Agnostic Inference for Large Language Models

ResearchDGX agent

arXiv:2606.05004v1 Announce Type: cross Abstract: With the widespread deployment of public large language models (LLMs) such as ChatGPT, protecting user prompt privacy has become an increasingly criti

ShareVerse: Multi-Agent Consistent Video Generation for Shared World Modeling

AgentsDGX agent

arXiv:2603.02697v2 Announce Type: replace-cross Abstract: This paper presents ShareVerse, a video generation framework enabling multi-agent shared world modeling, addressing the gap in existing works

SharpNet: Enhancing MLPs to Represent Functions with Controlled Non-differentiability

ResearchDGX agent

arXiv:2601.19683v2 Announce Type: replace Abstract: Multi-layer perceptrons (MLPs) are a standard tool for learning and function approximation, but they inherently produce globally smooth outputs. Con

Shifting the Breaking Point of Flow Matching for Multi-Instance Editing

Model ReleasesDGX agent

arXiv:2602.08749v3 Announce Type: replace Abstract: Flow matching models have recently emerged as an efficient alternative to diffusion, especially for text-guided image generation and editing, offeri

Shortcomings and capacities of real-constrained neural networks in complex spaces

ResearchDGX agent

arXiv:2606.04390v1 Announce Type: new Abstract: We find the asymptotic ratio between the storage capacities when enforcing real pre-activations in a complex hypothesis class as opposed to complex ones

Signed Dual Attention: Capturing Signed Dependencies in Time Series Forecasting

Model ReleasesDGX agent

arXiv:2606.04833v1 Announce Type: cross Abstract: Initially developed for natural language processing, Transformer architectures and attention mechanisms are now central to a wide range of deep learni

Simplicial Embeddings Improve Sample Efficiency in Actor-Critic Agents

SafetyDGX agent

arXiv:2510.13704v2 Announce Type: replace-cross Abstract: Recent works have proposed accelerating the wall-clock training time of actor-critic methods via the use of large-scale environment paralleliz

Simulate, Reason, Decide: Scientific Reasoning with LLMs for Simulation-Driven Decision Making

ResearchDGX agent

arXiv:2606.04505v1 Announce Type: new Abstract: Scientific simulators are increasingly being integrated into LLM-driven systems for high-stakes simulation-driven decision-making. However, existing fra

SMAC-Talk: A Natural Language Extension of the StarCraft Multi-Agent Challenge for Large Language Models

Model ReleasesDGX agent

arXiv:2606.04202v1 Announce Type: new Abstract: As LLMs become more widely deployed, they are increasingly expected to work alongside other AI agents rather than operating in isolation. Effective coor

SMADE-IE: Sparse Multi-Agent Framework with Evidence-Driven Debate for Zero-Shot Information Extraction

Model ReleasesDGX agent

arXiv:2606.04691v1 Announce Type: new Abstract: Zero-shot information extraction (IE) with large language models (LLMs) has attracted increasing attention due to its flexibility in adapting to new sch

Smart Picks in the Dark: Towards Efficient RLVR for Reasoning via Tracing Metacognitive Pivots

ResearchDGX agent

arXiv:2606.04503v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has greatly advanced large reasoning models (LRMs), but it requires timely training on a huge fu

Smart Transportation Without Neurons -- Fair Metro Network Expansion with Tabular Reinforcement Learning

SafetyDGX agent

arXiv:2606.04167v1 Announce Type: cross Abstract: We tackle the Metro Network Expansion Problem (MNEP), a subset of the Transport Network Design Problem (TNDP), which focuses on expanding metro system

SocialCoach: Personalized Social Skill Learning with RL-based Agentic Tutoring and Practice

SafetyDGX agent

arXiv:2606.04155v1 Announce Type: cross Abstract: Social skills such as negotiation and leadership are crucial for personal and professional success in today's interconnected world. However, scalable

SoftPINCH: EMG-Driven Soft Exoskeleton Assistance for Finger Flexion and Grasping

ResearchDGX agent

arXiv:2606.04776v1 Announce Type: new Abstract: Surface electromyography (sEMG) provides a non-invasive interface for detecting hand-movement intention and controlling wearable assistive devices. Howe

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization

SafetyDGX agent

arXiv:2505.11166v3 Announce Type: replace-cross Abstract: Despite advances in pretraining with extended context sizes, large language models (LLMs) still face challenges in effectively utilizing real-

Solving Zebra Puzzles Using Constraint-Guided Multi-Agent Systems

Model ReleasesDGX agent

arXiv:2407.03956v3 Announce Type: replace-cross Abstract: Prior research has enhanced the ability of Large Language Models (LLMs) to solve logic puzzles using techniques such as chain-of-thought promp

SparDA: Sparse Decoupled Attention for Efficient Long-Context LLM Inference

HardwareDGX agent

arXiv:2606.04511v1 Announce Type: new Abstract: Sparse attention reduces compute and memory bandwidth for long-context LLM inference. However, two key challenges remain: (1) KV cache capacity still gr

Sparse Bayesian Deep Functional Learning with Structured Region Selection

ApplicationsDGX agent

arXiv:2602.20651v3 Announce Type: replace Abstract: In modern applications such as ECG monitoring, neuroimaging, wearable sensing, and industrial equipment diagnostics, complex and continuously struct

Sparse Mixture-of-Experts Reward Models Learn Interpretable and Specialized Experts for Personalized Preference Modeling

TutorialsDGX agent

arXiv:2606.04284v1 Announce Type: cross Abstract: Preference modeling plays a central role in reinforcement learning from human feedback (RLHF), enabling large language models (LLMs) to align with hum

Spatial Artifact Coherence Determines Codec Robustness in Patch-Based rPPG

ResearchDGX agent

arXiv:2606.04198v1 Announce Type: new Abstract: Remote photoplethysmography (rPPG) achieves low heart-rate error on uncompressed benchmarks yet is deployed over compressed video channels in telehealth

Spatial Transcriptomics as Images for Large-Scale Pretraining

ResearchDGX agent

arXiv:2603.13432v4 Announce Type: replace-cross Abstract: Spatial Transcriptomics (ST) profiles thousands of gene expression values at discrete spots with precise coordinates on tissue sections, prese

Spatially Grounded Concept Bottleneck Models via Part-Factorized Attention

ResearchDGX agent

arXiv:2606.04364v1 Announce Type: new Abstract: Concept bottleneck models (CBMs) predict a layer of human-named attributes before predicting a class, which makes their decisions auditable. On fine-gra

Spectral Scaling Laws of Muon

ResearchDGX agent

arXiv:2606.04058v1 Announce Type: cross Abstract: Orthonormalized update rules have rapidly become a leading choice of optimizer for training large language models, with recent open-source state-of-th

Speculative Thinking: Enhancing Small-Model Reasoning with Large Model Guidance at Inference Time

Model ReleasesDGX agent

arXiv:2504.12329v2 Announce Type: replace-cross Abstract: Recent advances leverage post-training to enhance model reasoning performance, which typically requires costly training pipelines and still su

SpliceBind: Isoform-Aware Prediction of Binding Pocket Druggability

ResearchDGX agent

arXiv:2606.04020v1 Announce Type: cross Abstract: Splice-mediated drug resistance occurs in up to 40% of patients on targeted kinase inhibitors, yet state-of-the-art druggability tools operate on sing

SPLIT-PINN: Separable Probability Learning Technique via Physics-Informed Neural Networks for High-Dimensional Probabilistic Modeling

Model ReleasesDGX agent

arXiv:2606.04000v1 Announce Type: cross Abstract: We present a probabilistic modeling framework for incorporating small-scale spatial heterogeneity into macroscopic descriptions of material behavior f

SSA: Sparse Sparse Attention by Aligning Full and Sparse Attention Outputs in Feature Space

SafetyDGX agent

arXiv:2511.20102v3 Announce Type: replace Abstract: Sparse attention reduces the quadratic complexity of full self-attention but faces two challenges: (1) an attention gap, where applying sparse atten

SSSD: Simply-Scalable Speculative Decoding

ApplicationsDGX agent

arXiv:2411.05894v3 Announce Type: replace-cross Abstract: Speculative Decoding has emerged as a popular technique for accelerating inference in Large Language Models. However, most existing approaches

StandardE2E: A Unified Framework for End-to-End Autonomous Driving Datasets

Model ReleasesDGX agent

arXiv:2606.04271v1 Announce Type: cross Abstract: Autonomous driving has shifted from modular perception-prediction-planning stacks toward end-to-end (E2E) models that map sensor inputs directly to ve

STaR-Quant: State-Time Consistent Post-Training Quantization for Diffusion Large Language Models

ResearchDGX agent

arXiv:2606.04945v1 Announce Type: new Abstract: Diffusion large language models (DLLMs) have recently emerged as a promising alternative to autoregressive LLMs by generating text through iterative mas

Stateful Visual Encoders for Vision-Language Models

AgentsDGX agent

arXiv:2606.04433v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly used in multi-image, multi-turn agentic settings where decisions depend on visual changes. However, in

Stationarity-Aware Retrieval-Augmented Time Series Forecasting

ApplicationsDGX agent

arXiv:2606.04135v1 Announce Type: new Abstract: Time series forecasting relies on historical patterns, but real-world series often exhibit non-stationarity and regime shifts that challenge fully param

Stein Kernelized Molecular Dynamics for Active Learning of Interatomic Potentials

TutorialsDGX agent

arXiv:2606.04100v1 Announce Type: new Abstract: Machine learning interatomic potentials (MLIPs) enable efficient and accurate atomistic simulations but depend critically on the quality and diversity o

StepPRM-RTL: Stepwise Process-Reward Guided LLM Fine-Tuning for Enhanced RTL Synthesis

Model ReleasesDGX agent

arXiv:2606.04246v1 Announce Type: new Abstract: Automatic generation of RTL code for digital hardware designs remains challenging due to long-horizon reasoning, multi-step dependencies, and strict cor

Stepwise Reasoning Enhancement for LLMs via External Subgraph Generation

Model ReleasesDGX agent

arXiv:2606.04454v1 Announce Type: new Abstract: Large language models have shown strong performance in natural language generation and downstream reasoning tasks, but they still struggle with logical

Strabo: Declarative Specification and Implementation of Agentic Interaction Protocols

AgentsDGX agent

arXiv:2606.05043v1 Announce Type: new Abstract: The last few years have witnessed major advances in the modeling and implementation of multiagent systems based on declarative interaction protocols. Ou

Streaming Communication in Multi-Agent Reasoning

Model ReleasesDGX agent

arXiv:2606.05158v1 Announce Type: cross Abstract: Multi-agent reasoning systems adopt a 'generate-then-transfer' paradigm that forces end-to-end latency to scale linearly with pipeline depth. We intro

STRIDE: Training Data Attribution via Sparse Recovery from Subset Perturbations

Model ReleasesDGX agent

arXiv:2606.05165v1 Announce Type: cross Abstract: Training Data Attribution (TDA) seeks to trace a model's predictions back to its training data. The gold standard for TDA relies on causal interventio

StrokeTimer: Robust Representation Learning for Ischemic Stroke Onset-Time Estimation from Non-contrast CT

ApplicationsDGX agent

arXiv:2606.04722v1 Announce Type: new Abstract: Ischemic stroke is a major global disease. Treatment decisions are highly time-sensitive, as eligibility for reperfusion therapies relies on the interva

Structure-Aware Prediction of PROTAC-Mediated Protein Degradability via Graph Neural Networks

Model ReleasesDGX agent

arXiv:2606.04021v1 Announce Type: cross Abstract: Proteolysis-targeting chimeras (PROTACs) can selectively degrade disease-causing proteins, yet predicting which targets are amenable to degradation re

Stumbling Into AI Emotional Dependence: How Routine AI Interactions Reshape Human Connection

SafetyDGX agent

arXiv:2606.04150v1 Announce Type: new Abstract: Public discourse and emerging policy typically assume that AI emotional support is a deliberate act: a lonely user consciously seeking comfort from a de

Success Conditioning as Policy Improvement: The Optimization Problem Solved by Imitating Success

SafetyDGX agent

arXiv:2601.18175v2 Announce Type: replace Abstract: A widely used technique for improving policies is success conditioning, in which one collects trajectories, identifies those that achieve a desired

Supportive Token Revealing for Fast Diffusion Language Model Decoding

ResearchDGX agent

arXiv:2606.04236v1 Announce Type: cross Abstract: Discrete diffusion language models can generate text efficiently by updating multiple masked positions in parallel, but this parallelism introduces a

SurvPFN: Towards Foundation Models for Survival Predictions

TutorialsDGX agent

arXiv:2606.04564v1 Announce Type: new Abstract: Tabular foundation models (TFMs) have made rapid progress in standard classification and regression, but time-to-event survival prediction tasks have re

SUSD: Structured Unsupervised Skill Discovery through State Factorization

AgentsDGX agent

arXiv:2602.01619v2 Announce Type: replace-cross Abstract: Unsupervised Skill Discovery (USD) aims to autonomously learn a diverse set of skills without relying on extrinsic rewards. One of the most co

Symbolic Regression for Shared Expressions: Introducing Partial Parameter Sharing

Model ReleasesDGX agent

arXiv:2601.04051v3 Announce Type: replace Abstract: Symbolic regression aims to find symbolic expressions that describe datasets. Due to its inherent interpretability, symbolic regression (SR) is a po

SymTRELLIS: Symmetry-Enforced Voxel Latents for 3D Generation

Model ReleasesDGX agent

arXiv:2606.04108v1 Announce Type: cross Abstract: Single-view 3D generative models have achieved impressive visual quality, yet they are not designed to satisfy structural or functional requirements,

Synthetic Personalities: How Well Can LLMs Mimic Individual Respondents Using Socio-Economic Microdata?

ResearchDGX agent

arXiv:2606.04592v1 Announce Type: cross Abstract: LLM-based digital twins promise to scale and accelerate market research, but most published twins are either coarse persona bots conditioned on a few

TaDA: Calibrated Probe Gating for Task-Domain LoRA Merging

Model ReleasesDGX agent

arXiv:2606.05016v1 Announce Type: new Abstract: Combining a task LoRA adapter with a domain LoRA adapter into a single unified model is a practical yet largely unexplored challenge. Existing methods t

Take a Peek: Efficient Encoder Adaptation for Few-Shot Semantic Segmentation via LoRA

ResearchDGX agent

arXiv:2512.10521v2 Announce Type: replace Abstract: Few-shot semantic segmentation (FSS) aims to segment novel classes in query images using only a small annotated support set. While prior research ha

TamperBench: Systematically Stress-Testing LLM Safety Under Fine-Tuning and Tampering

SafetyDGX agent

arXiv:2602.06911v2 Announce Type: replace-cross Abstract: As increasingly capable open-weight large language models (LLMs) are deployed, improving their tamper resistance against unsafe modifications,

TANDEM: Bi-Level Data Mixture Optimization with Twin Networks

ResearchDGX agent

arXiv:2606.04401v1 Announce Type: new Abstract: The capabilities of large language models (LLMs) significantly depend on training data drawn from various domains. Optimizing domain-specific mixture ra

Teaching Robots to Say 'I Don't Know' : SENTINEL for Uncertainty-Aware SLAM

ResearchDGX agent

arXiv:2606.04853v1 Announce Type: new Abstract: Low-cost 2D LiDARs lack the intensity channel that higher-end sensors use to diagnose measurement failures, yet they are widely used on educational and

Temporal Order Matters for Agentic Memory: Segment Trees for Long-Horizon Agents

Local AiDGX agent

arXiv:2606.04555v1 Announce Type: cross Abstract: Long-horizon conversational agents need to interact with users through evolving events, tasks, and goals. Such histories are naturally temporal, yet m

← Previous
1…487488489490491…1049
Next →