AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,914
  • Agents7,747
  • Applications5,533
  • Concepts5
  • Hardware1,920
  • Industry6,196
  • Local Ai5,094
  • Model Releases24,727
  • Research20,781
  • Safety13,739
  • Syntheses17
  • Tools1,678
  • Tutorials3,477

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,914
  • Agents7,747
  • Applications5,533
  • Concepts5
  • Hardware1,920
  • Industry6,196
  • Local Ai5,094
  • Model Releases24,727
  • Research20,781
  • Safety13,739
  • Syntheses17
  • Tools1,678
  • Tutorials3,477

Source
Human
90,914Total entries
1Added by human
90,913Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
90,913 results
2 Jun 2026

Safety Game: Inference-Time Alignment of Black-Box LLMs via Constrained Optimization

SafetyDGX agent

arXiv:2510.09330v3 Announce Type: replace Abstract: Ensuring that large language models (LLMs) comply with safety requirements is a central challenge in AI deployment. Existing alignment approaches pr

Safety Mirage: How Spurious Correlations Undermine VLM Safety Fine-Tuning and Can Be Mitigated by Machine Unlearning

SafetyDGX agent

arXiv:2503.11832v5 Announce Type: replace Abstract: Recent vision language models (VLMs) have made remarkable strides in generative modeling with multimodal inputs, particularly text and images. Howev

SafeVLA-Bench: A Benchmark for the Success-Safety Gap in Vision-Language-Action Models

Model ReleasesDGX agent
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.00773v1 Announce Type: new Abstract: Vision-language-action (VLA) benchmarks measure whether a policy completes a requested manipulation task, but binary success can hide safety-relevant tr

Saliency-Aware Model Merging

Model ReleasesDGX agent

arXiv:2606.00511v1 Announce Type: cross Abstract: Model merging aims to consolidate multiple task-specific models fine-tuned on different datasets into a unified architecture that performs cross-domai

SALSA: Speech Aware LLM Adaptation via Learned Steering Activation Vectors

ResearchDGX agent

arXiv:2606.00460v1 Announce Type: new Abstract: Speech-aware large language models often generalize poorly to out-of-domain settings. We propose SALSA (Speech-Aware LLM Adaptation via Learned Steering

Same Payload, Different Channel: Measuring Trust Asymmetry in Tool-Using Language Models

Model ReleasesDGX agent

arXiv:2606.00566v1 Announce Type: cross Abstract: As language models take on agentic roles that span calling external APIs, reading tool outputs, and acting on instructions embedded in third-party con

Sample Complexity and Decision-Theoretic Guarantees for Bayesian Model Averaging over Decision Trees with Catalan-Exponential Priors

ResearchDGX agent

arXiv:2606.01340v1 Announce Type: new Abstract: We ask: when do Bayesian model averaging (BMA) weights over decision trees carry sufficient epistemic information to justify committed exploitation of t

Sandboxed Coding Agents are Competitive Omni-modal Task Solvers

Model ReleasesDGX agent

arXiv:2606.00579v1 Announce Type: new Abstract: As multimodal LLMs increasingly target video and audio, it is often assumed that such tasks require native omnimodal models. We show that this is not al

SARA: Stress Test Reasoning in Audio Deepfake Detection

ResearchDGX agent

arXiv:2601.03615v2 Announce Type: replace Abstract: Audio Language Models (ALMs) offer a promising shift towards explainable audio deepfake detections (ADD), moving beyond extit{black-box} classifiers

SAVMap: Structure-Aided Visual Mapping of Large-Scale 2.5D Manhattan Wireframes from Panoramic Video

ApplicationsDGX agent

arXiv:2606.01939v1 Announce Type: new Abstract: Precise 3D representations of industrial environments enable tasks such as robot localization and digital twin generation. We propose SAVMap, a method f

Scalable Counterfactual Risk Estimation for Rare Events in Longitudinal Data

ResearchDGX agent

arXiv:2606.01539v1 Announce Type: cross Abstract: Estimating the causal effect of time-varying treatments on survival outcomes in large observational studies is computationally demanding, particularly

Scalable Ride-Sourcing Vehicle Rebalancing with Service Accessibility Guarantee: A Constrained Mean-Field Reinforcement Learning Approach

SafetyDGX agent

arXiv:2503.24183v3 Announce Type: replace Abstract: The expansion of ride-sourcing services such as Uber and Lyft has reshaped urban transportation by offering flexible, on-demand mobility via mobile

Scalar-Measurement Attitude Estimation on mathbf{SO}(3) with Bias Compensation

SafetyDGX agent

arXiv:2603.02478v2 Announce Type: replace-cross Abstract: Attitude estimation methods typically rely on full vector measurements from inertial sensors such as accelerometers and magnetometers. This pa

Scaling Agentic Capabilities via Grounded Interaction Synthesis

AgentsDGX agent

arXiv:2606.02001v1 Announce Type: new Abstract: General agentic intelligence hinges on the ability to interact with diverse real-world tools to complete complex tasks, a capability fundamentally tied

Scaling Behavior of Single LLM-Driven Multi-Agent Systems

AgentsDGX agent

arXiv:2606.00655v1 Announce Type: cross Abstract: The burgeoning field of LLM-based Multi-Agent Systems (MAS) promises to tackle complex tasks through collaborative intelligence, yet fundamental quest

// Scaling Behavior of Single LLM-Driven Multi-Agent Systems // Does adding more agents actually make a multi-agent system better? It's poss…

AgentsDGX agent

// Scaling Behavior of Single LLM-Driven Multi-Agent Systems // Does adding more agents actually make a multi-agent system better? It's possible that collective intelligence emerges from interaction d

Scaling depth capacity via zero/one-layer model expansion

ResearchDGX agent

arXiv:2511.04981v2 Announce Type: replace Abstract: Model depth is a double-edged sword in deep learning: deeper models achieve higher accuracy but require higher computational cost. To efficiently tr

Scaling Parallel Sequence Models to Foundation-Scale Vision Encoders

Model ReleasesDGX agent

arXiv:2606.00746v1 Announce Type: new Abstract: Vision foundation models are bottlenecked by the quadratic cost of self-attention, which limits usable resolution and increases the cost of large-scale

Scaling Pre-training to One Hundred Billion Data for Vision Language Models

ResearchDGX agent

arXiv:2502.07617v2 Announce Type: replace Abstract: We provide an empirical investigation of the potential of pre-training vision-language models on an unprecedented scale: 100 billion examples. We fi

Scaling Search Relevance: Augmenting App Store Ranking with LLM-Generated Judgments

ApplicationsDGX agent

arXiv:2602.23234v4 Announce Type: replace-cross Abstract: Large-scale commercial search systems optimize for relevance to drive successful sessions that help users find what they are looking for. To m

SCAPO: Self-Supervised Category-Level Articulated Pose Estimation from a Single 3D Observation

SafetyDGX agent

arXiv:2606.01940v1 Announce Type: new Abstract: Existing methods for category-level object articulation from a single 3D observation often rely on dense supervision, multi-frame inputs, or CAD templat

ScaRF-SLAM: Scale-Consistent Reconstruction with Feed-Forward Models and Classical Visual SLAM

ResearchDGX agent

arXiv:2606.00307v1 Announce Type: new Abstract: Recent works have explored unifying SLAM with geometric foundation models (GFMs). However, directly using GFM predictions for tracking is highly sensiti

SceneSmith: Agentic Generation of Simulation-Ready Indoor Scenes

SafetyDGX agent

arXiv:2602.09153v2 Announce Type: replace-cross Abstract: Simulation has become a key tool for training and evaluating home robots at scale, yet existing environments fail to capture the diversity and

SciAgentGym: Benchmarking Multi-Step Scientific Tool-use in LLM Agents

AgentsDGX agent

arXiv:2602.12984v2 Announce Type: replace Abstract: Scientific reasoning inherently demands integrating sophisticated toolkits to navigate domain-specific knowledge. Yet, current benchmarks largely ov

scicode-lint: Detecting Methodology Bugs in Scientific Python Code with LLM-Generated Patterns

Local AiDGX agent

arXiv:2603.17893v2 Announce Type: replace-cross Abstract: Methodology bugs in scientific Python code produce plausible but incorrect results that traditional linters and static analysis tools cannot d

Science Earth: Towards A Planet-Scale Operating System for AI-Native Scientific Discovery

Model ReleasesDGX agent

arXiv:2606.01316v1 Announce Type: new Abstract: Scientific discovery demands intelligence, perseverance, and serendipity across vast search spaces. Today, top scientific capabilities remain siloed--on

SCL: Towards Domain Generalization via Single-Temporal Multimodal Contrastive Learning for Remote Sensing Change Detection

TutorialsDGX agent

arXiv:2404.11326v5 Announce Type: replace Abstract: In recent years, change detection and anomaly detection models based on CNN and transformer have achieved remarkable success across various datasets

Score-Control for Hallucination Reduction in Diffusion Models

Model ReleasesDGX agent

arXiv:2606.00377v1 Announce Type: new Abstract: Diffusion models have emerged as the backbone of modern generative AI, powering advances in vision, language, audio and other modalities. Despite their

Score Function Gradient Estimation to Widen the Applicability of Decision-Focused Learning

ApplicationsDGX agent

arXiv:2307.05213v3 Announce Type: replace-cross Abstract: Many real-world optimization problems contain parameters that are unknown before deployment time, either due to stochasticity or to lack of in

Score imes Decoder: A Unified View of Unsupervised Inference-Time Scaling for Hallucination Mitigation

ResearchDGX agent

arXiv:2606.00739v1 Announce Type: new Abstract: Large language models hallucinate even when the answer lies within their parameters. While inference-time scaling can surface this latent knowledge, the

SDR: Set-Distance Rewards for Radiology Report Generation

Model ReleasesDGX agent

arXiv:2606.00440v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards has rapidly advanced reasoning in vision--language models. However, for chest X-ray report generation, th

Search-on-Graph: Iterative Informed Navigation for Large Language Model Reasoning on Knowledge Graphs

ResearchDGX agent

arXiv:2510.08825v2 Announce Type: replace Abstract: Large language models (LLMs) augmented with knowledge graphs (KGs) offer a promising approach for knowledge-intensive reasoning. Central to this app

SEArch: Optimistic Policy Selection Between Scene Noise and Drift for UAV Radar Search

SafetyDGX agent

arXiv:2606.01325v1 Announce Type: cross Abstract: Unmanned Aerial Vehicles (UAVs) equipped with radar sensors are deployed for target search missions in diverse environments, where targets exhibit cha

SeClaw: Spec-Driven Security Task Synthesis for Evaluating Autonomous Agents

Model ReleasesDGX agent

arXiv:2606.02302v1 Announce Type: cross Abstract: Autonomous LLM agents increasingly operate in stateful environments where they access tools, files, memory, and external services. While such capabili

SECUREVENT: Hybrid AI/ML Security Monitoring for Distributed Event-Based Systems

SafetyDGX agent

arXiv:2606.01741v1 Announce Type: cross Abstract: Distributed event-based systems have become a common substrate for Internet-scale publish/subscribe services, IoT telemetry, cloud-native microservice

See, Plan, Rewind: Progress-Aware Vision-Language-Action Models for Robust Robotic Manipulation

Model ReleasesDGX agent

arXiv:2603.09292v2 Announce Type: replace-cross Abstract: Measurement of task progress through explicit, actionable milestones is critical for robust robotic manipulation. This progress awareness enab

Seeing Martin Scorsese using FLUX for storyboarding and scene exploration was absolutely insane. Experiencing how one of the absolute master…

IndustryDGX agent

Seeing Martin Scorsese using FLUX for storyboarding and scene exploration was absolutely insane. Experiencing how one of the absolute masters of cinema & filmmaking uses the technology that we develop

Seeing Through the MiRAGE: Evaluating Multimodal Retrieval Augmented Generation

ResearchDGX agent

arXiv:2510.24870v2 Announce Type: replace Abstract: We introduce MiRAGE, an evaluation framework for retrieval-augmented generation (RAG) from multimodal sources. As audiovisual media becomes a preval

Seg-Zero: Reasoning-Chain Guided Segmentation via Cognitive Reinforcement

Model ReleasesDGX agent

arXiv:2503.06520v3 Announce Type: replace Abstract: Traditional methods for reasoning segmentation rely on supervised fine-tuning with categorical labels and simple descriptions, limiting its out-of-d

Segment-driven Structural Induction and Semantic Alignment for Heterogeneous Tabular Representation

Local AiDGX agent

arXiv:2606.01890v1 Announce Type: new Abstract: Real-world domains often contain heterogeneous tables whose headers vary while their underlying attribute semantics are shared, making it difficult to i

Segmentation-Guided Spatial Indexing for Generalizable and Explainable Deepfake Detection

ResearchDGX agent

arXiv:2606.00098v1 Announce Type: new Abstract: We introduce segmentation-guided spatial indexing for generalizable and explainable deepfake detection. The key idea reverses the standard design order:

Self-Conditioned Positional HNSW for Overlap-Aware Retrieval in Chunked-Document RAG Systems: Method and Industrial Evidence-Quality Audit

ResearchDGX agent

arXiv:2606.01542v1 Announce Type: cross Abstract: Chunked-document retrieval is a common component of retrieval-augmented generation (RAG) systems. Documents are split into overlapping chunks, embedde

Self-Evolving Hermes Agents: Enterprise AI That Gets Better With Use | Nemotron Labs https://x.com/i/broadcasts/1pJdRRyneOjKW

Model ReleasesDGX agent

This likely describes a framework or system for deploying AI agents that autonomously improve their performance over time through continuous learning and adaptation in enterprise environments. The sel

Self-Healing Agentic Orchestrators for Reliable Tool-Augmented Large Language Model Systems

Model ReleasesDGX agent

arXiv:2606.01416v1 Announce Type: new Abstract: Tool-augmented large language model (LLM) agents rely on orchestration layers that coordinate planning, retrieval, tool invocation, validation, memory,

Self-Imitated Diffusion Policy for Efficient and Robust Visual Navigation

Model ReleasesDGX agent

arXiv:2601.22965v2 Announce Type: replace Abstract: Diffusion policies (DP) have demonstrated significant potential in visual navigation by capturing diverse multi-modal trajectory distributions. Howe

Self-Improving Small Object Grounding in LVLMs

ResearchDGX agent

arXiv:2606.01612v1 Announce Type: new Abstract: Can internal attention patterns in Large Vision Language Models (LVLMs) identify reliable small-object boxes without fine-tuning? In this work, we provi

Self-Regulating Annealing in Heavy-Tailed Diffusion Models

ResearchDGX agent

arXiv:2606.01645v1 Announce Type: cross Abstract: Diffusion models have emerged as a leading framework for deep generative modeling. While the standard Gaussian formulation is theoretically convenient

Self-Revising Discovery Systems for Science: A Categorical Framework for Agentic Artificial Intelligence

AgentsDGX agent

arXiv:2606.01444v1 Announce Type: new Abstract: Scientific discovery is not only answer generation but revision of the representational regime in which evidence, artifacts, operations, and verifiers a

Self-supervised Monocular Depth and Pose Estimation for Endoscopy with Latent Priors

ResearchDGX agent

arXiv:2411.17790v3 Announce Type: replace-cross Abstract: Accurate 3D mapping in endoscopy enables quantitative, holistic lesion characterization within the gastrointestinal (GI) tract, requiring reli

Semantic-Geometric Task Representations for Bimanual Manipulation from Human Demonstrations to Robot Action Planning

ResearchDGX agent

arXiv:2601.11460v2 Announce Type: replace-cross Abstract: Learning structured task representations from human demonstrations is essential for bimanual manipulation, where action ordering, object invol

Semantic Retrieval for Product Search in E-Commerce

SafetyDGX agent

arXiv:2606.01504v1 Announce Type: cross Abstract: Semantic retrieval in e-commerce must handle short, noisy, and colloquial queries over large product catalogs with fine-grained attribute distinctions

SEMBridge: Tagless-Final Program Semantics with Weakest-Precondition and Bounded-Checking Interpretations

ResearchDGX agent

arXiv:2606.00220v1 Announce Type: cross Abstract: Formal methods provide rigorous accounts of program behavior, but practical software engineering often works through executable libraries, tests, and

Semi-Supervised Hyperbolic Hierarchical Clustering with Set-Level Structural Priors

Model ReleasesDGX agent

arXiv:2606.01525v1 Announce Type: new Abstract: Semi-supervised hierarchical clustering aims to learn a tree structure consistent with data patterns and user-provided supervision. Supervision is usual

Semi-Supervised Learning with Noisy Proxy Covariates: Generalization Bounds and Distribution Regression

ResearchDGX agent

arXiv:2606.00512v1 Announce Type: new Abstract: In many modern machine learning pipelines, abundant pretrained representations serve as noisy proxy covariates, while task-specific labels remain scarce

Semi-Supervised Noise Adaptation: Transferring Knowledge from Noise Domain

ResearchDGX agent

arXiv:2606.00558v1 Announce Type: new Abstract: Transfer learning aims to facilitate the learning of a target domain by transferring knowledge from a source domain. The source domain typically contain

Semimage: HSV-Based Semantic Image Encoding for Disentangled Text Representation

ResearchDGX agent

arXiv:2512.00088v2 Announce Type: replace Abstract: We propose SemImage, a novel method for representing a text document as a two-dimensional semantic image to be processed by convolutional neural net

SEMixer: Semantics Enhanced MLP-Mixer for Multiscale Mixing and Long-term Time Series Forecasting

SafetyDGX agent

arXiv:2602.16220v2 Announce Type: replace Abstract: Modeling multiscale patterns is crucial for long-term time series forecasting (TSF). However, redundancy and noise in time series, together with sem

Send the video to everyone you know showing how heinously Nowak was treated by the police in his dying moments and how the police cravenly k…

IndustryDGX agent

Send the video to everyone you know showing how heinously Nowak was treated by the police in his dying moments and how the police cravenly kowtowed to his murderer. Legacy mainstream media, same ones

SENSE: Semantic Embedding Navigation with Soft-gated Evaluation for Retrieval-based Speculative Decoding

Model ReleasesDGX agent

arXiv:2606.00021v1 Announce Type: cross Abstract: Speculative Decoding (SD) accelerates Large Language Model (LLM) inference by employing a lightweight draft model to propose candidate tokens, which a

Sensitivity as a Double-Edged Sword: A Trade-off Between Discriminability and Adversarial Robustness

ResearchDGX agent

arXiv:2606.01746v1 Announce Type: new Abstract: Modern neural networks are highly susceptible to adversarial perturbations. In this work, we identify that part of this vulnerability stems from the sen

← Previous
1…753754755756757…1516
Next →