AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,566
  • Agents7,790
  • Applications5,565
  • Concepts5
  • Hardware1,943
  • Industry6,214
  • Local Ai5,132
  • Model Releases24,957
  • Research20,928
  • Safety13,836
  • Syntheses17
  • Tools1,680
  • Tutorials3,499

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries91,566
  • Agents7,790
  • Applications5,565
  • Concepts5
  • Hardware1,943
  • Industry6,214
  • Local Ai5,132
  • Model Releases24,957
  • Research20,928
  • Safety13,836
  • Syntheses17
  • Tools1,680
  • Tutorials3,499

Source
HumanDGX agent

Content type
AllBlog
91,566Total entries
1Added by human
91,565Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
91,566 results
Research

Safeguarded Stochastic Polyak Step Sizes for Non-smooth Optimization: Robust Performance Without Small (Sub)Gradients

DGX agent

arXiv:2512.02342v3 Announce Type: replace-cross Abstract: The stochastic Polyak step size (SPS) has proven to be a promising choice for stochastic gradient descent (SGD), delivering competitive perfor

researcharxiv-cs-lg
2 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Safety

SafeMCP: Proactive Power Regulation for LLM Agent Defense via Environment-Grounded Look-Ahead Reasoning

DGX agent

arXiv:2606.01991v1 Announce Type: new Abstract: As Large Language Model (LLM) agents increasingly leverage the Model Context Protocol (MCP) to operate in complex environments, the expansion of their a

safetyarxiv-cs-ai
2 Jun 2026
Safety

SafeSteer: Localized On-Policy Distillation for Efficient Safety Alignment

DGX agent

arXiv:2606.02530v1 Announce Type: new Abstract: Aligning Large Language Models (LLMs) with human values often degrades their general capabilities, termed the alignment tax. Existing methods mitigate t

safetyarxiv-cs-ai
2 Jun 2026
Safety

Safety Alignment of LMs via Non-cooperative Games

DGX agent

arXiv:2512.20806v3 Announce Type: replace Abstract: Ensuring the safety of language models (LMs) while maintaining their usefulness remains a critical challenge in AI alignment. Current approaches rel

safetyarxiv-cs-ai
2 Jun 2026
Safety

Safety Game: Inference-Time Alignment of Black-Box LLMs via Constrained Optimization

DGX agent

arXiv:2510.09330v3 Announce Type: replace Abstract: Ensuring that large language models (LLMs) comply with safety requirements is a central challenge in AI deployment. Existing alignment approaches pr

safetyarxiv-cs-lg
2 Jun 2026
Safety

Safety Mirage: How Spurious Correlations Undermine VLM Safety Fine-Tuning and Can Be Mitigated by Machine Unlearning

DGX agent

arXiv:2503.11832v5 Announce Type: replace Abstract: Recent vision language models (VLMs) have made remarkable strides in generative modeling with multimodal inputs, particularly text and images. Howev

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

SafeVLA-Bench: A Benchmark for the Success-Safety Gap in Vision-Language-Action Models

DGX agent

arXiv:2606.00773v1 Announce Type: new Abstract: Vision-language-action (VLA) benchmarks measure whether a policy completes a requested manipulation task, but binary success can hide safety-relevant tr

model-releasesarxiv-cs-ro
2 Jun 2026
Model Releases

Saliency-Aware Model Merging

DGX agent

arXiv:2606.00511v1 Announce Type: cross Abstract: Model merging aims to consolidate multiple task-specific models fine-tuned on different datasets into a unified architecture that performs cross-domai

model-releasesarxiv-cs-cv
2 Jun 2026
Research

SALSA: Speech Aware LLM Adaptation via Learned Steering Activation Vectors

DGX agent

arXiv:2606.00460v1 Announce Type: new Abstract: Speech-aware large language models often generalize poorly to out-of-domain settings. We propose SALSA (Speech-Aware LLM Adaptation via Learned Steering

researcharxiv-cs-cl
2 Jun 2026
Model Releases

Same Payload, Different Channel: Measuring Trust Asymmetry in Tool-Using Language Models

DGX agent

arXiv:2606.00566v1 Announce Type: cross Abstract: As language models take on agentic roles that span calling external APIs, reading tool outputs, and acting on instructions embedded in third-party con

model-releasesarxiv-cs-cl
2 Jun 2026
Research

Sample Complexity and Decision-Theoretic Guarantees for Bayesian Model Averaging over Decision Trees with Catalan-Exponential Priors

DGX agent

arXiv:2606.01340v1 Announce Type: new Abstract: We ask: when do Bayesian model averaging (BMA) weights over decision trees carry sufficient epistemic information to justify committed exploitation of t

researcharxiv-cs-lg
2 Jun 2026
Model Releases

Sandboxed Coding Agents are Competitive Omni-modal Task Solvers

DGX agent

arXiv:2606.00579v1 Announce Type: new Abstract: As multimodal LLMs increasingly target video and audio, it is often assumed that such tasks require native omnimodal models. We show that this is not al

model-releasesarxiv-cs-cl
2 Jun 2026
Research

SARA: Stress Test Reasoning in Audio Deepfake Detection

DGX agent

arXiv:2601.03615v2 Announce Type: replace Abstract: Audio Language Models (ALMs) offer a promising shift towards explainable audio deepfake detections (ADD), moving beyond extit{black-box} classifiers

researcharxiv-cs-cl
2 Jun 2026
Applications

SAVMap: Structure-Aided Visual Mapping of Large-Scale 2.5D Manhattan Wireframes from Panoramic Video

DGX agent

arXiv:2606.01939v1 Announce Type: new Abstract: Precise 3D representations of industrial environments enable tasks such as robot localization and digital twin generation. We propose SAVMap, a method f

applicationsarxiv-cs-cv
2 Jun 2026
Research

Scalable Counterfactual Risk Estimation for Rare Events in Longitudinal Data

DGX agent

arXiv:2606.01539v1 Announce Type: cross Abstract: Estimating the causal effect of time-varying treatments on survival outcomes in large observational studies is computationally demanding, particularly

researcharxiv-cs-lg
2 Jun 2026
Safety

Scalable Ride-Sourcing Vehicle Rebalancing with Service Accessibility Guarantee: A Constrained Mean-Field Reinforcement Learning Approach

DGX agent

arXiv:2503.24183v3 Announce Type: replace Abstract: The expansion of ride-sourcing services such as Uber and Lyft has reshaped urban transportation by offering flexible, on-demand mobility via mobile

safetyarxiv-cs-lg
2 Jun 2026
Safety

Scalar-Measurement Attitude Estimation on mathbf{SO}(3) with Bias Compensation

DGX agent

arXiv:2603.02478v2 Announce Type: replace-cross Abstract: Attitude estimation methods typically rely on full vector measurements from inertial sensors such as accelerometers and magnetometers. This pa

safetyarxiv-cs-ro
2 Jun 2026
Agents

Scaling Agentic Capabilities via Grounded Interaction Synthesis

DGX agent

arXiv:2606.02001v1 Announce Type: new Abstract: General agentic intelligence hinges on the ability to interact with diverse real-world tools to complete complex tasks, a capability fundamentally tied

agentsarxiv-cs-cl
2 Jun 2026
Agents

Scaling Behavior of Single LLM-Driven Multi-Agent Systems

DGX agent

arXiv:2606.00655v1 Announce Type: cross Abstract: The burgeoning field of LLM-based Multi-Agent Systems (MAS) promises to tackle complex tasks through collaborative intelligence, yet fundamental quest

agentsarxiv-cs-ai
2 Jun 2026
Agents

// Scaling Behavior of Single LLM-Driven Multi-Agent Systems // Does adding more agents actually make a multi-agent system better? It's poss…

DGX agent

// Scaling Behavior of Single LLM-Driven Multi-Agent Systems // Does adding more agents actually make a multi-agent system better? It's possible that collective intelligence emerges from interaction d

agentsdair-ai--x
2 Jun 2026
Research

Scaling depth capacity via zero/one-layer model expansion

DGX agent

arXiv:2511.04981v2 Announce Type: replace Abstract: Model depth is a double-edged sword in deep learning: deeper models achieve higher accuracy but require higher computational cost. To efficiently tr

researcharxiv-cs-lg
2 Jun 2026
Model Releases

Scaling Parallel Sequence Models to Foundation-Scale Vision Encoders

DGX agent

arXiv:2606.00746v1 Announce Type: new Abstract: Vision foundation models are bottlenecked by the quadratic cost of self-attention, which limits usable resolution and increases the cost of large-scale

model-releasesarxiv-cs-cv
2 Jun 2026
Research

Scaling Pre-training to One Hundred Billion Data for Vision Language Models

DGX agent

arXiv:2502.07617v2 Announce Type: replace Abstract: We provide an empirical investigation of the potential of pre-training vision-language models on an unprecedented scale: 100 billion examples. We fi

researcharxiv-cs-cv
2 Jun 2026
Applications

Scaling Search Relevance: Augmenting App Store Ranking with LLM-Generated Judgments

DGX agent

arXiv:2602.23234v4 Announce Type: replace-cross Abstract: Large-scale commercial search systems optimize for relevance to drive successful sessions that help users find what they are looking for. To m

applicationsarxiv-cs-ai
2 Jun 2026
Safety

SCAPO: Self-Supervised Category-Level Articulated Pose Estimation from a Single 3D Observation

DGX agent

arXiv:2606.01940v1 Announce Type: new Abstract: Existing methods for category-level object articulation from a single 3D observation often rely on dense supervision, multi-frame inputs, or CAD templat

safetyarxiv-cs-cv
2 Jun 2026
Research

ScaRF-SLAM: Scale-Consistent Reconstruction with Feed-Forward Models and Classical Visual SLAM

DGX agent

arXiv:2606.00307v1 Announce Type: new Abstract: Recent works have explored unifying SLAM with geometric foundation models (GFMs). However, directly using GFM predictions for tracking is highly sensiti

researcharxiv-cs-ro
2 Jun 2026
Safety

SceneSmith: Agentic Generation of Simulation-Ready Indoor Scenes

DGX agent

arXiv:2602.09153v2 Announce Type: replace-cross Abstract: Simulation has become a key tool for training and evaluating home robots at scale, yet existing environments fail to capture the diversity and

safetyarxiv-cs-ai
2 Jun 2026
Agents

SciAgentGym: Benchmarking Multi-Step Scientific Tool-use in LLM Agents

DGX agent

arXiv:2602.12984v2 Announce Type: replace Abstract: Scientific reasoning inherently demands integrating sophisticated toolkits to navigate domain-specific knowledge. Yet, current benchmarks largely ov

agentsarxiv-cs-cl
2 Jun 2026
Local Ai

scicode-lint: Detecting Methodology Bugs in Scientific Python Code with LLM-Generated Patterns

DGX agent

arXiv:2603.17893v2 Announce Type: replace-cross Abstract: Methodology bugs in scientific Python code produce plausible but incorrect results that traditional linters and static analysis tools cannot d

local-aiarxiv-cs-ai
2 Jun 2026
Model Releases

Science Earth: Towards A Planet-Scale Operating System for AI-Native Scientific Discovery

DGX agent

arXiv:2606.01316v1 Announce Type: new Abstract: Scientific discovery demands intelligence, perseverance, and serendipity across vast search spaces. Today, top scientific capabilities remain siloed--on

model-releasesarxiv-cs-ai
2 Jun 2026
Tutorials

SCL: Towards Domain Generalization via Single-Temporal Multimodal Contrastive Learning for Remote Sensing Change Detection

DGX agent

arXiv:2404.11326v5 Announce Type: replace Abstract: In recent years, change detection and anomaly detection models based on CNN and transformer have achieved remarkable success across various datasets

tutorialsarxiv-cs-cv
2 Jun 2026
Model Releases

Score-Control for Hallucination Reduction in Diffusion Models

DGX agent

arXiv:2606.00377v1 Announce Type: new Abstract: Diffusion models have emerged as the backbone of modern generative AI, powering advances in vision, language, audio and other modalities. Despite their

model-releasesarxiv-cs-cv
2 Jun 2026
Applications

Score Function Gradient Estimation to Widen the Applicability of Decision-Focused Learning

DGX agent

arXiv:2307.05213v3 Announce Type: replace-cross Abstract: Many real-world optimization problems contain parameters that are unknown before deployment time, either due to stochasticity or to lack of in

applicationsarxiv-cs-ai
2 Jun 2026
Research

Score imes Decoder: A Unified View of Unsupervised Inference-Time Scaling for Hallucination Mitigation

DGX agent

arXiv:2606.00739v1 Announce Type: new Abstract: Large language models hallucinate even when the answer lies within their parameters. While inference-time scaling can surface this latent knowledge, the

researcharxiv-cs-lg
2 Jun 2026
Model Releases

SDR: Set-Distance Rewards for Radiology Report Generation

DGX agent

arXiv:2606.00440v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards has rapidly advanced reasoning in vision--language models. However, for chest X-ray report generation, th

model-releasesarxiv-cs-ai
2 Jun 2026
Research

Search-on-Graph: Iterative Informed Navigation for Large Language Model Reasoning on Knowledge Graphs

DGX agent

arXiv:2510.08825v2 Announce Type: replace Abstract: Large language models (LLMs) augmented with knowledge graphs (KGs) offer a promising approach for knowledge-intensive reasoning. Central to this app

researcharxiv-cs-cl
2 Jun 2026
Safety

SEArch: Optimistic Policy Selection Between Scene Noise and Drift for UAV Radar Search

DGX agent

arXiv:2606.01325v1 Announce Type: cross Abstract: Unmanned Aerial Vehicles (UAVs) equipped with radar sensors are deployed for target search missions in diverse environments, where targets exhibit cha

safetyarxiv-cs-lg
2 Jun 2026
Model Releases

SeClaw: Spec-Driven Security Task Synthesis for Evaluating Autonomous Agents

DGX agent

arXiv:2606.02302v1 Announce Type: cross Abstract: Autonomous LLM agents increasingly operate in stateful environments where they access tools, files, memory, and external services. While such capabili

model-releasesarxiv-cs-ai
2 Jun 2026
Safety

SECUREVENT: Hybrid AI/ML Security Monitoring for Distributed Event-Based Systems

DGX agent

arXiv:2606.01741v1 Announce Type: cross Abstract: Distributed event-based systems have become a common substrate for Internet-scale publish/subscribe services, IoT telemetry, cloud-native microservice

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

See, Plan, Rewind: Progress-Aware Vision-Language-Action Models for Robust Robotic Manipulation

DGX agent

arXiv:2603.09292v2 Announce Type: replace-cross Abstract: Measurement of task progress through explicit, actionable milestones is critical for robust robotic manipulation. This progress awareness enab

model-releasesarxiv-cs-cv
2 Jun 2026
Industry

Seeing Martin Scorsese using FLUX for storyboarding and scene exploration was absolutely insane. Experiencing how one of the absolute master…

DGX agent

Seeing Martin Scorsese using FLUX for storyboarding and scene exploration was absolutely insane. Experiencing how one of the absolute masters of cinema & filmmaking uses the technology that we develop

industryemad-mostaque--x
2 Jun 2026
Research

Seeing Through the MiRAGE: Evaluating Multimodal Retrieval Augmented Generation

DGX agent

arXiv:2510.24870v2 Announce Type: replace Abstract: We introduce MiRAGE, an evaluation framework for retrieval-augmented generation (RAG) from multimodal sources. As audiovisual media becomes a preval

researcharxiv-cs-cl
2 Jun 2026
Model Releases

Seg-Zero: Reasoning-Chain Guided Segmentation via Cognitive Reinforcement

DGX agent

arXiv:2503.06520v3 Announce Type: replace Abstract: Traditional methods for reasoning segmentation rely on supervised fine-tuning with categorical labels and simple descriptions, limiting its out-of-d

model-releasesarxiv-cs-cv
2 Jun 2026
Local Ai

Segment-driven Structural Induction and Semantic Alignment for Heterogeneous Tabular Representation

DGX agent

arXiv:2606.01890v1 Announce Type: new Abstract: Real-world domains often contain heterogeneous tables whose headers vary while their underlying attribute semantics are shared, making it difficult to i

local-aiarxiv-cs-lg
2 Jun 2026
Research

Segmentation-Guided Spatial Indexing for Generalizable and Explainable Deepfake Detection

DGX agent

arXiv:2606.00098v1 Announce Type: new Abstract: We introduce segmentation-guided spatial indexing for generalizable and explainable deepfake detection. The key idea reverses the standard design order:

researcharxiv-cs-cv
2 Jun 2026
Research

Self-Conditioned Positional HNSW for Overlap-Aware Retrieval in Chunked-Document RAG Systems: Method and Industrial Evidence-Quality Audit

DGX agent

arXiv:2606.01542v1 Announce Type: cross Abstract: Chunked-document retrieval is a common component of retrieval-augmented generation (RAG) systems. Documents are split into overlapping chunks, embedde

researcharxiv-cs-ai
2 Jun 2026
Model Releases

Self-Evolving Hermes Agents: Enterprise AI That Gets Better With Use | Nemotron Labs https://x.com/i/broadcasts/1pJdRRyneOjKW

DGX agent

This likely describes a framework or system for deploying AI agents that autonomously improve their performance over time through continuous learning and adaptation in enterprise environments. The sel

model-releasesnous-research--x
2 Jun 2026
Model Releases

Self-Healing Agentic Orchestrators for Reliable Tool-Augmented Large Language Model Systems

DGX agent

arXiv:2606.01416v1 Announce Type: new Abstract: Tool-augmented large language model (LLM) agents rely on orchestration layers that coordinate planning, retrieval, tool invocation, validation, memory,

model-releasesarxiv-cs-ai
2 Jun 2026
← Previous
1…955956957958959…1908
Next →