AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
Human
85,202Total entries
1Added by human
85,201Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
60,292 results
12 May 2026

SearchSkill: Teaching LLMs to Use Search Tools with Evolving Skill Banks

ResearchDGX agent

arXiv:2605.09038v1 Announce Type: new Abstract: Teaching language models to use search tools is not only a question of whether they search, but also of whether they issue good queries. This is especia

SeasonScapes: Learning Large-scale Re-lightable 3D Landscapes with Seasonal Variation from Sparse Webcams

ResearchDGX agent

arXiv:2605.09039v1 Announce Type: new Abstract: We introduce SeasonScapes framework and a the SeasonScapes dataset: Swiss Sparse-view Mountain Scenes with Seasonal Changes that covers over 50 km x 60

SeBA: Semi-supervised few-shot learning via Separated-at-Birth Alignment for tabular data

Model ReleasesDGX agent
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.08519v1 Announce Type: new Abstract: Learning from scarce labeled data with a larger pool of unlabeled samples, known as semi-supervised few-shot learning (SS-FSL), remains critical for app

SecureForge: Finding and Preventing Vulnerabilities in LLM-Generated Code via Prompt Optimization

AgentsDGX agent

arXiv:2605.08382v1 Announce Type: cross Abstract: LLM coding agents now generate code at an unprecedented scale, yet LLM-generated code introduces cybersecurity vulnerabilities into codebases without

Security Risks in Tool-Enabled AI Agents: A Systematic Analysis of Privileged Execution Environments

AgentsDGX agent

arXiv:2605.09721v1 Announce Type: cross Abstract: Tool-enabled AI agents are increasingly deployed in cloud-hosted environments and offered as services, where they perform side-effecting operations th

Seed Hijacking of LLM Sampling and Quantum Random Number Defense

Model ReleasesDGX agent

arXiv:2605.08313v1 Announce Type: cross Abstract: Large language models (LLMs) rely on deterministic pseudorandom number generators (PRNGs) for autoregressive sampling, creating a critical supply-chai

SeePhys Pro: Diagnosing Modality Transfer and Blind-Training Effects in Multimodal RLVR for Physics Reasoning

Model ReleasesDGX agent

arXiv:2605.09266v1 Announce Type: new Abstract: We introduce SeePhys Pro, a fine-grained modality transfer benchmark that studies whether models preserve the same reasoning capability when critical in

Segment Anything with Robust Uncertainty-Accuracy Correlation

SafetyDGX agent

arXiv:2605.10603v1 Announce Type: new Abstract: Despite strong zero-shot performance, SAM is unreliable under domain shift due to Mask-level Confidence Confusion (MCC), where a single IoU-based mask s

SegSTRONG-C: Segmenting Surgical Tools Robustly On Non-adversarial Generated Corruptions -- An EndoVis'24 Challenge

ResearchDGX agent

arXiv:2407.11906v3 Announce Type: replace Abstract: Surgical data science has seen rapid advancement with the excellent performance of end-to-end deep neural networks (DNNs). Despite their successes,

SEIS: Subspace-based Equivariance and Invariance Scores for Neural Representations

ResearchDGX agent

arXiv:2602.04054v2 Announce Type: replace-cross Abstract: Understanding how neural representations respond to geometric transformations is essential for evaluating whether learned features preserve me

Select-then-differentiate: Solving Bilevel Optimization with Manifold Lower-level Solution Sets

Local AiDGX agent

arXiv:2605.09209v1 Announce Type: cross Abstract: We study optimistic bilevel optimization when the lower-level problem has a non-isolated manifold of minimizers. In this setting, the hyper-objective

Selection of the Best Policy under Fairness Constraints for Subpopulations

SafetyDGX agent

arXiv:2605.09945v1 Announce Type: new Abstract: Many high-stakes decisions in health care, public policy, and clinical development require committing to a single policy that will be applied uniformly

Selection Plateau and a Sparsity-Dependent Hierarchy of Pruning Features

SafetyDGX agent

arXiv:2605.09345v1 Announce Type: new Abstract: We identify a Selection Plateau phenomenon in one-shot neural network pruning: all rank-monotone weight scorers converge to identical accuracy at fixed

Selective Deficits in LLM Mental Self-Modeling in a Behavior-Based Test of Theory of Mind

Model ReleasesDGX agent

arXiv:2603.26089v2 Announce Type: replace-cross Abstract: The ability to represent oneself and others as agents with knowledge, intentions, and belief states that guide their behavior - Theory of Mind

Selective LoRA for Visual Tokens and Attention Heads

Model ReleasesDGX agent

arXiv:2512.19219v2 Announce Type: replace-cross Abstract: Low-rank adaptation (LoRA) is widely used for parameter-efficient fine-tuning, but its standard all-token, all-head design ignores the heterog

Self-Attention as a Covariance Readout: A Unified View of In-Context Learning and Repetition

ResearchDGX agent

arXiv:2605.10466v1 Announce Type: new Abstract: Large language models (LLMs) exhibit two striking and ostensibly unrelated behaviours: in-context learning (ICL) and repetitive generation. In both, the

Self-Captioning Multimodal Interaction Tuning: Amplifying Exploitable Redundancies for Robust Vision Language Models

ResearchDGX agent

arXiv:2605.08145v1 Announce Type: cross Abstract: Current vision language models face hallucination and robustness issues against ambiguous or corrupted modalities. We hypothesize that these issues ca

Self-ReSET: Learning to Self-Recover from Unsafe Reasoning Trajectories

SafetyDGX agent

arXiv:2605.08936v1 Announce Type: new Abstract: Large Reasoning Models possess remarkable capabilities for self-correction in general domain; however, they frequently struggle to recover from unsafe r

Semantic Alignment in Hyperbolic Space for Open-Vocabulary Semantic Segmentation

SafetyDGX agent

arXiv:2605.08874v1 Announce Type: new Abstract: Open-vocabulary semantic segmentation requires adapting image-level vision-language models such as CLIP to dense pixel-level prediction, which is challe

Semantic Voting: Execution-Grounded Consensus for LLM Code Generation

ResearchDGX agent

arXiv:2605.08680v1 Announce Type: cross Abstract: LLM code-generation pipelines often sample multiple candidates and select one final answer without access to a complete oracle. Existing pipelines mix

SEMASIA: A Large-Scale Dataset of Semantically Structured Latent Representations

Model ReleasesDGX agent

arXiv:2605.09485v1 Announce Type: new Abstract: Latent representations learned by neural networks often exhibit semantic structure, where concept similarity is reflected by geometric proximity in embe

Semi-Supervised Neural Super-Resolution for Mesh-Based Simulations

Model ReleasesDGX agent

arXiv:2605.09284v1 Announce Type: cross Abstract: Mesh-based simulations provide high-fidelity solutions to partial differential equations (PDEs), but achieving such accuracy typically requires fine m

Sens-VisualNews: A Benchmark Dataset for Sensational Image Detection

Model ReleasesDGX agent

arXiv:2605.10394v1 Announce Type: new Abstract: The detection of sensational content in media items can be a critical filtering mechanism for identifying check-worthy content and flagging potential di

SenseBench: A Benchmark for Remote Sensing Low-Level Visual Perception and Description in Large Vision-Language Models

Model ReleasesDGX agent

arXiv:2605.10576v1 Announce Type: cross Abstract: Low-level visual perception underpins reliable remote sensing (RS) image analysis, yet current image quality assessment (IQA) methods output uninterpr

Separate First, Fuse Later: Mitigating Cross-Modal Interference in Audio-Visual LLMs Reasoning with Modality-Specific Chain-of-Thought

Model ReleasesDGX agent

arXiv:2605.09906v1 Announce Type: new Abstract: Audio and vision provide complementary evidence for audio-visual question answering, yet current audio-visual large language models may suffer from cros

Sequential Causal Discovery with Noisy Language Model Priors

Model ReleasesDGX agent

arXiv:2506.16234v2 Announce Type: replace Abstract: Causal discovery from observational data typically assumes access to complete data and availability of perfect domain experts. In practice, data oft

Sequential Feature Selection for Efficient Landslide Segmentation from Multi-Spectral Data

Model ReleasesDGX agent

arXiv:2605.09746v1 Announce Type: cross Abstract: Landslide detection from satellite imagery has advanced through deep learning, yet most models rely on large, highly correlated spectral-topographic i

Sequential Membership Inference Attacks

ResearchDGX agent

arXiv:2602.16596v2 Announce Type: replace Abstract: Modern AI models are not static. They go through multiple updates in their lifecycles. We propose to design Sequential Membership Inference (SeMI) a

Set-Based Groupwise Registration for Variable-Length, Variable-Contrast Cardiac MRI

ResearchDGX agent

arXiv:2605.10571v1 Announce Type: cross Abstract: Quantitative cardiac magnetic resonance imaging (MRI) enables non-invasive myocardial tissue characterization but relies on robust motion correction w

Set Prediction for Next-Day Active Fire Forecasting

Model ReleasesDGX agent

arXiv:2605.10298v1 Announce Type: new Abstract: Accurate next-day active fire forecasts can support early warning, disaster response, forest risk assessment, and downstream estimation of fire-related

SGC-RML: A reliable and interpretable longitudinal assessment for PD in real-world DNS

SafetyDGX agent

arXiv:2605.08302v1 Announce Type: cross Abstract: Real-world digital Parkinson's disease assessment faces challenges such as heterogeneous modalities, cross-device bias, and incomplete labeling. Exist

ShadowMerge: A Novel Poisoning Attack on Graph-Based Agent Memory via Relation-Channel Conflicts

AgentsDGX agent

arXiv:2605.09033v1 Announce Type: cross Abstract: Graph-based agent memory is increasingly used in LLM agents to support structured long-term recall and multi-hop reasoning, but it also creates a new

Shaping Schema via Language Representation as the Next Frontier for LLM Intelligence Expanding

ApplicationsDGX agent

arXiv:2605.09271v1 Announce Type: new Abstract: Although natural language is the default medium for Large Language Models (LLMs), its limited expressive capacity creates a profound bottleneck for comp

Shapley Regression for Rare Disease Diagnosis Support: a case study on APDS

ApplicationsDGX agent

arXiv:2605.08897v1 Announce Type: cross Abstract: Activated PI3K8 Syndrome (APDS) is a rare genetic immune disorder caused by variants in PIK3CD or PIK3R1, with highly heterogeneous symptoms that ofte

Sharp feature-learning transitions and Bayes-optimal neural scaling laws in extensive-width networks

ResearchDGX agent

arXiv:2605.10395v1 Announce Type: cross Abstract: We study the information-theoretic limits of learning a one-hidden-layer teacher network with hierarchical features from noisy queries, in the context

Shepherd: A Runtime Substrate Empowering Meta-Agents with a Formalized Execution Trace

AgentsDGX agent

arXiv:2605.10913v1 Announce Type: new Abstract: We introduce Shepherd, a functional programming model that formalizes meta-agent operations on target agents as functions, with core operations mechaniz

SHIELD: Scalable Optimal Control with Certification using Duality and Convexity

SafetyDGX agent

arXiv:2605.09171v1 Announce Type: new Abstract: We present SHIELD, a hierarchical algorithm that reduces both the decision-variable dimension and the constraint set in ell_1-regularized convex program

Shields to Guarantee Probabilistic Safety in MDPs

SafetyDGX agent

arXiv:2605.10888v1 Announce Type: cross Abstract: Shielding is a prominent model-based technique to ensure safety of autonomous agents. Classical shielding aims to ensure that nothing bad ever happens

ShifaMind: A Multiplicative Concept Bottleneck for Interpretable ICD-10 Coding

ResearchDGX agent

arXiv:2605.08482v1 Announce Type: cross Abstract: Automated ICD-10 coding from clinical discharge summaries requires models that are both accurate on long-tailed multi-label classification tasks and i

Signal from Structure: Exploiting Submodular Upper Bounds in Generative Flow Networks

TutorialsDGX agent

arXiv:2601.21061v2 Announce Type: replace Abstract: Generative Flow Networks (GFlowNets; GFNs) are a class of generative models that learn to sample compositional objects proportionally to their a pri

Signature Approach for Contextual Bandits with Nonlinear and Path-dependent Rewards

ResearchDGX agent

arXiv:2605.10313v1 Announce Type: new Abstract: We study contextual bandits with nonlinear and path-dependent rewards through a novel signature-transform-based approach. Leveraging the universal nonli

simpleposter: a simple baseline for product poster generation

Model ReleasesDGX agent

arXiv:2605.08784v1 Announce Type: new Abstract: Product poster generation poses distinct challenges beyond general poster design, requiring both faithful preservation of product appearance and precise

SimReg: Achieving Higher Performance in the Pretraining via Embedding Similarity Regularization

ResearchDGX agent

arXiv:2605.08809v1 Announce Type: cross Abstract: Pretraining large language models (LLMs) with next-token prediction has led to remarkable advances, yet the context-dependent nature of token embeddin

Simulating Complex Multi-Turn Tool Calling Interactions in Stateless Execution Environments

AgentsDGX agent

arXiv:2601.19914v2 Announce Type: replace-cross Abstract: Synthetic data has proven itself to be a valuable resource for tuning smaller, cost-effective language models to handle the complexities of mu

Simultaneous Long-tailed Recognition and Multi-modal Fusion for Highly Imbalanced Multi-modal Data

Model ReleasesDGX agent

arXiv:2605.10498v1 Announce Type: cross Abstract: Long-tailed distributions in class-imbalanced data present a fundamental challenge for deep learning models, which tend to be biased toward majority c

Simultaneous Monitoring of Shape and Surface Color via 4D Point Clouds: A Registration-free Approach

ApplicationsDGX agent

arXiv:2605.08753v1 Announce Type: new Abstract: Advanced manufacturing technologies allow for the production of intricate parts featuring high shape complexity and spatially-varying material compositi

Simulus: Combining Improvements in Sample-Efficient World Model Agents

AgentsDGX agent

arXiv:2502.11537v4 Announce Type: replace-cross Abstract: World models (WMs) represent the frontier of sample-efficient reinforcement learning, but their complexity leaves many promising improvements

SimWorld Studio: Automatic Environment Generation with Evolving Coding Agent for Embodied Agent Learning

AgentsDGX agent

arXiv:2605.09423v1 Announce Type: new Abstract: LLM/VLM-based digital agents have advanced rapidly thanks to scalable sandboxes for coding, web navigation, and computer use, which provide rich interac

Single-Configuration Attack Success Rate Is Not Enough: Jailbreak Evaluations Should Report Distributional Attack Success

Model ReleasesDGX agent

arXiv:2605.09070v1 Announce Type: cross Abstract: Many jailbreak attack research papers report attack success rates for a limited number of parameter settings, even though there are many combinations

Single-Thread JPEG Decoder Benchmarks Mis-Evaluate ML Data Loaders

Model ReleasesDGX agent

arXiv:2605.08731v1 Announce Type: cross Abstract: JPEG decode is routine ML infrastructure, but Python decoder choices are often justified by single-process, single-thread microbenchmarks. We audit th

Sink vs. diagonal patterns as mechanisms for attention switch and oversmoothing prevention

SafetyDGX agent

arXiv:2605.08453v1 Announce Type: cross Abstract: This paper studies the role of sinks and diagonal patterns as attention switch and anti-oversmoothing mechanisms. We analyze geometric conditions unde

Sinkhorn Treatment Effects: A Causal Optimal Transport Measure

Model ReleasesDGX agent

arXiv:2605.08485v1 Announce Type: cross Abstract: We introduce the Sinkhorn treatment effect, an entropic optimal transport measure of divergence between counterfactual distributions. Unlike classical

Sketch-and-Verify: Structured Inference-Time Scaling via Program Sketching

Model ReleasesDGX agent

arXiv:2605.08658v1 Announce Type: cross Abstract: SKETCHVERIFY is a within-tier cost-performance policy, not a universal accuracy improvement. The operational question: a practitioner stuck with a sma

SKG-VLA: Scene Knowledge Graph Priors for Structured Scene Semantics and Multimodal Reasoning for Decision Making

SafetyDGX agent

arXiv:2605.09343v1 Announce Type: new Abstract: Decision making in large-scale complaint handling systems increasingly relies on heterogeneous evidence, including complaint narratives, screenshots, or

Skill-CMIB: Multimodal Agent Skill for Consistent Action via Conditional Multimodal Information Bottleneck

AgentsDGX agent

arXiv:2605.08526v1 Announce Type: new Abstract: While LLM-based agents excel at planning and executing long action sequences, their execution often remains inconsistent across trials, limiting reliabi

Skill-R1: Agent Skill Evolution via Reinforcement Learning

SafetyDGX agent

arXiv:2605.09359v1 Announce Type: cross Abstract: Agentic large language models often rely on skills, reusable natural language procedures that guide planning, action, and tool use. In practice, skill

SkillEvolver: Skill Learning as a Meta-Skill

HardwareDGX agent

arXiv:2605.10500v1 Announce Type: new Abstract: Agent skills today are static artifact: authored once -- by human curation or one-shot generation from parametric knowledge -- and then consumed unchang

SkillLens: Adaptive Multi-Granularity Skill Reuse for Cost-Efficient LLM Agents

AgentsDGX agent

arXiv:2605.08386v1 Announce Type: new Abstract: Skill libraries have become a practical way for LLM agents to reuse procedural experience across tasks. However, existing systems typically treat skills

SkillMAS: Skill Co-Evolution with LLM-based Multi-Agent System

AgentsDGX agent

arXiv:2605.09341v1 Announce Type: cross Abstract: Large language model (LLM) agent systems are increasingly expected to improve after deployment, but existing work often decouples two adaptation targe

SkillMaster: Toward Autonomous Skill Mastery in LLM Agents

AgentsDGX agent

arXiv:2605.08693v1 Announce Type: new Abstract: Skills provide an effective mechanism for improving LLM agents on complex tasks, yet in existing agent frameworks, their creation, refinement, and selec

← Previous
1…730731732733734…1005
Next →