AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,343
  • Agents7,552
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,164
  • Local Ai4,928
  • Model Releases23,861
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,402

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries88,343
  • Agents7,552
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,164
  • Local Ai4,928
  • Model Releases23,861
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,402

Source
HumanDGX agent

88,343Total entries
1Added by human
88,342Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
88,343 results
19 May 2026

Single Image Reflection Removal with Patch Reflectance Prior

TutorialsDGX agent

arXiv:2312.03798v2 Announce Type: replace Abstract: Single Image Reflection Removal (SIRR) in real-world images is a challenging task due to diverse image degradations occurring on the glass surface d

Single-Sample Black-Box Membership Inference Attack against Vision-Language Models via Cross-modal Semantic Alignment

SafetyDGX agent

arXiv:2605.17341v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have achieved remarkable success, yet their reliance on massive datasets and unintended memorization of training data ra

SIPO: Stabilized and Improved Preference Optimization for Aligning Diffusion Models

Model ReleasesDGX agent

arXiv:2505.21893v3 Announce Type: replace-cross Abstract: Preference learning has garnered extensive attention as an effective technique for aligning diffusion models with human preferences in visual

Content type
AllBlogX PostPaperYouTubeRedditGitHub

SIREM: Speech-Informed MRI Reconstruction with Learned Sampling

Model ReleasesDGX agent

arXiv:2605.18221v1 Announce Type: cross Abstract: Real-time magnetic resonance imaging (rtMRI) of speech production enables non-invasive visualization of dynamic vocal-tract motion and is valuable for

Sketch Then Paint: Hierarchical Reinforcement Learning for Diffusion Multi-Modal Large Language Models

Local AiDGX agent

arXiv:2605.16842v1 Announce Type: new Abstract: Diffusion Multi-Modal Large Language Models (dMLLMs) are powerful for image generation, but optimizing them through reinforcement learning (RL) remains

SKG-Eval: Stateful Evaluation of Multi-Turn Dialogue via Incremental Semantic Knowledge Graphs

Local AiDGX agent

arXiv:2605.16650v1 Announce Type: cross Abstract: Evaluating multi-turn dialogue systems remains challenging because response quality depends not only on the current prompt, but also on previously est

SkillGenBench: Benchmarking Skill Generation Pipelines for LLM Agents

Model ReleasesDGX agent

arXiv:2605.18693v1 Announce Type: new Abstract: As LLM agents are increasingly built around reusable skills, a central challenge is no longer only whether agents can use provided skills, but whether t

SkillJect: Effectively Automating Skill-Based Prompt Injection for Skill-Enabled Agents

AgentsDGX agent

arXiv:2602.14211v2 Announce Type: replace-cross Abstract: Agent skills are increasingly used to extend LLM agents with task-specific instructions, executable scripts, and auxiliary resources. While im

Skills on the Fly: Test-Time Adaptive Skill Synthesis for LLM Agents

Model ReleasesDGX agent

arXiv:2605.16986v1 Announce Type: cross Abstract: LLM agents benefit from reusable skills, yet test-time tasks often require guidance more specific than a static skill library can provide. We propose

SkillsVote: Lifecycle Governance of Agent Skills from Collection, Recommendation to Evolution

Model ReleasesDGX agent

arXiv:2605.18401v1 Announce Type: cross Abstract: Long-horizon LLM agents leave traces that could become reusable experience, but raw trajectories are noisy and hard to govern. We treat Agent Skills a

Skim: Speculative Execution for Fast and Efficient Web Agents

AgentsDGX agent

arXiv:2605.16565v1 Announce Type: new Abstract: Skim is a speculative execution framework for web agents that exploits the predictable structure of purpose-built websites. Today's web-agent expense is

SkyNative: A Native Multimodal Framework for Remote Sensing Visual Evidence Reasoning

Model ReleasesDGX agent

arXiv:2605.17949v1 Announce Type: new Abstract: Remote sensing vision-language models commonly rely on pretrained visual encoders to convert images into semantic features before language-model reasoni

SLEIGHT-Bench: A Benchmark of Evasion Attacks Against Agent Monitors

Model ReleasesDGX agent

arXiv:2605.16626v1 Announce Type: cross Abstract: Since autonomous coding agents generate complex behaviors at high-volume, we may want to use other LLMs to monitor actions to reduce the risk from dan

Small-scale photonic Kolmogorov-Arnold networks using standard telecom nonlinear modules

Model ReleasesDGX agent

arXiv:2604.08432v2 Announce Type: replace-cross Abstract: Photonic neural networks promise ultrafast inference, yet most architectures rely on linear optical meshes with electronic nonlinearities, rei

Sneaky interventions, interactive installations, tools built for other artists to use. @kcimc's practice lives where machine learning, compu…

ToolsDGX agent

Sneaky interventions, interactive installations, tools built for other artists to use. @kcimc's practice lives where machine learning, computer vision, and social technology collide. See his new inter

SNLP: Layer-Parallel Inference via Structured Newton Corrections

SafetyDGX agent

arXiv:2605.17842v1 Announce Type: new Abstract: Autoregressive language models execute Transformer layers sequentially, creating a latency bottleneck that is not removed by conventional tensor or pipe

So google is replacing gemini-cli with agy (antigravity cli), but: 1. agy is not opensource 2. It no longer supports ACP Really unfortunate …

Model ReleasesDGX agent

Google is transitioning from the Gemini CLI tool to a new CLI called AGY (Antigravity CLI), but this change has drawbacks: AGY is not open source and no longer supports ACP functionality, representing

Soap2Soap: Long Cinematic Video Remaking via Multi-Agent Collaboration

SafetyDGX agent

arXiv:2605.17423v1 Announce Type: new Abstract: We study series-level cinematic remaking, a long-horizon video-to-video generation problem that localizes full episodes or films via stylization or acto

SocialMemBench: Are AI Memory Systems Ready for Social Group Settings?

Model ReleasesDGX agent

arXiv:2605.17789v1 Announce Type: cross Abstract: Memory systems for AI assistants were built for single-user dialogue and fail characteristically when applied to multi-party social group settings. Th

Sola Security launches Lumina to cut enterprise security alert noise with contextual AI

Model ReleasesDGX agent

Cybersecurity startup Sola Security Ltd. today announced the launch of Lumina, an autonomous risk intelligence platform that applies contextual artificial intelligence across cloud, identity, software

SomaliWeb v1: A Quality-Filtered Somali Web Corpus with a Matched Tokenizer and a Public Language-Identification Benchmark

Model ReleasesDGX agent

arXiv:2605.18232v1 Announce Type: cross Abstract: Somali is a Cushitic language of the Horn of Africa with ~25 million speakers, yet no documented dedicated Somali pretraining corpus with a companion

Some fun Gemini Omni use cases from the community👇🧵 (We’ll keep updating this thread throughout the day)

Model ReleasesDGX agent

This X thread from Google AI showcases practical and creative applications of Gemini Omni, Google's multimodal AI model, as demonstrated and shared by the user community. The thread appears to be a cu

Some[Body] Must Receive That Pain for Agent Accountability

AgentsDGX agent

arXiv:2605.16872v1 Announce Type: cross Abstract: AI agents increasingly act consequentially in the real world. This creates a problem we call consequence reception: harm occurs, the producing system

Sometin Beta Pass Notin (SBPN): Improving Multilingual ASR for Nigerian Languages via Knowledge Distillation

Model ReleasesDGX agent

arXiv:2605.17710v1 Announce Type: new Abstract: Although modern multilingual Automatic Speech Recognition (ASR) systems support several Nigerian languages, their performance consistently lags behind h

SonarSweep: Fusing Sonar and Vision for Robust 3D Reconstruction via Plane Sweeping

ApplicationsDGX agent

arXiv:2511.00392v2 Announce Type: replace-cross Abstract: Accurate 3D reconstruction in visually-degraded underwater environments remains a formidable challenge. Single-modality approaches are insuffi

Sources: Chinese AI startup Moonshot told investors it would revamp its corporate structure to pave the way for a Hong Kong IPO and comply with Beijing's rules (Bloomberg)

IndustryDGX agent

Bloomberg: Sources: Chinese AI startup Moonshot told investors it would revamp its corporate structure to pave the way for a Hong Kong IPO and comply with Beijing's rules — Moonshot AI has informed it

Sources: Demis Hassabis was an angel investor in Anthropic; PitchBook and Dealroom: former DeepMind researchers founded 12+ companies since 2021, raising $14B+ (Financial Times)

IndustryDGX agent

Financial Times: Sources: Demis Hassabis was an angel investor in Anthropic; PitchBook and Dealroom: former DeepMind researchers founded 12+ companies since 2021, raising $14B+ — Demis Hassabis, the f

Sources detail growing concerns inside SoftBank over Masayoshi Son's $60B+ bet on OpenAI, which some fear concentrates too much capital into a single company (Bloomberg)

IndustryDGX agent

Bloomberg: Sources detail growing concerns inside SoftBank over Masayoshi Son's 60B+ bet on OpenAI, which some fear concentrates too much capital into a single company — With more than 60 billion comm

Sources: Intel asks its leading PC partners, including those in the US, China, and Taiwan, to use more 18A CPUs, citing better supply than chips on older nodes (Nikkei Asia)

ApplicationsDGX agent

Nikkei Asia: Sources: Intel asks its leading PC partners, including those in the US, China, and Taiwan, to use more 18A CPUs, citing better supply than chips on older nodes — TAIPEI — Intel is urging

Sources: SpaceX expects to proceed with its acquisition of Cursor 30 days after its public trading debut, which is expected to occur on June 12 (Bloomberg)

IndustryDGX agent

Bloomberg: Sources: SpaceX expects to proceed with its acquisition of Cursor 30 days after its public trading debut, which is expected to occur on June 12 — SpaceX expects to proceed with its acquisit

Sources: Zyphra, which trains and runs inference for its open-weight models on AMD hardware, is raising a 500M Series B at a valuation of at least 5B (Anna Tong/Forbes)

HardwareDGX agent

Anna Tong / Forbes: Sources: Zyphra, which trains and runs inference for its open-weight models on AMD hardware, is raising a 500M Series B at a valuation of at least 5B — The Series B round, which ch

Sparse Autoencoders are Topic Models

TutorialsDGX agent

arXiv:2511.16309v2 Announce Type: replace Abstract: Sparse autoencoders (SAEs) are used to analyze embeddings, but their role and practical value are debated. We propose a new perspective on SAEs by d

Sparse Deep Additive Model with Interactions: Enhancing Interpretability and Predictability

TutorialsDGX agent

arXiv:2509.23068v2 Announce Type: replace-cross Abstract: Recent advances in deep learning highlight the need for personalized models that can learn from small samples, handle high-dimensional feature

Sparse Mamba Decoder for Quantum Error Correction: Efficient Defect-Centric Processing of Surface Code Syndromes

HardwareDGX agent

arXiv:2605.17156v1 Announce Type: cross Abstract: Quantum error correction (QEC) is essential for building fault-tolerant quantum computers, requiring decoders that are simultaneously accurate, fast,

Sparse-to-Dense: A Free Lunch for Lossless Acceleration of Video Understanding in LLMs

ResearchDGX agent

arXiv:2505.19155v2 Announce Type: replace-cross Abstract: Due to the auto-regressive nature of current video large language models (Video-LLMs), the inference latency increases as the input sequence l

Sparse Training of Neural Networks based on Multilevel Mirror Descent

Model ReleasesDGX agent

arXiv:2602.03535v2 Announce Type: replace Abstract: We introduce a dynamic sparse training algorithm based on linearized Bregman iterations / mirror descent that exploits the naturally incurred sparsi

SparseSAM: Structured Sparsification of Activations in Segment Anything Models

ResearchDGX agent

arXiv:2605.17633v1 Announce Type: cross Abstract: The Segment Anything Model (SAM) achieves strong open-vocabulary segmentation, but its ViT-based image encoders dominate inference latency and memory.

Spatial Blindness in Whole-Slide Multiple Instance Learning

ResearchDGX agent

arXiv:2605.17449v1 Announce Type: cross Abstract: Whole-slide MIL models are often called context-aware once graphs, Transform ers, or state-space modules are placed above patch embeddings. We show th

Spatially Aware Linear Transformer (SAL-T) for Particle Jet Tagging

Local AiDGX agent

arXiv:2510.23641v2 Announce Type: replace-cross Abstract: Transformers are very effective in capturing both global and local correlations within high-energy particle collisions, but they present deplo

SPATIOROUTE: Dynamic Prompt Routing for Zero-Shot Spatial Reasoning

Model ReleasesDGX agent

arXiv:2605.18209v1 Announce Type: cross Abstract: Spatial question answering over egocentric video is a challenging task that requires Vision-Language Models (VLMs) to reason about 3D object positions

Spatiotemporal Robustness of Temporal Logic Tasks using Multi-Objective Reasoning

AgentsDGX agent

arXiv:2603.29868v2 Announce Type: replace Abstract: The reliability of autonomous systems depends on their robustness, i.e., their ability to meet their objectives under uncertainty. In this paper, we

Speak Your Mind: The Speech Continuation Task as a Probe of Voice-Based Model Bias

SafetyDGX agent

arXiv:2509.22061v2 Announce Type: replace-cross Abstract: Speech Continuation (SC) is the task of generating a coherent extension of a spoken prompt while preserving both semantic context and speaker

SpecSem-Net: Integrating Spectral and Semantic Features for Robust AI-generated Video Detection

Model ReleasesDGX agent

arXiv:2605.17311v1 Announce Type: new Abstract: The remarkable visual fidelity of recent commercial video generative models, such as Sora and Veo, renders robust AI-generated video detection increasin

Spectral Progressive Diffusion for Efficient Image and Video Generation

ResearchDGX agent

arXiv:2605.18736v1 Announce Type: new Abstract: Diffusion models have been shown to implicitly generate visual content autoregressively in the frequency domain, where low-frequency components are gene

Speech-Guided Multimodal Learning for Vocal Tract Segmentation in Real-Time MRI

Local AiDGX agent

arXiv:2605.18466v1 Announce Type: new Abstract: Segmenting vocal tract articulators in real-time MRI (rtMRI) is a challenging dynamic image segmentation problem characterized by low contrast, rapid mo

Speech-Hands: A Self-Reflection Voice Agentic Approach to Speech Recognition and Audio Reasoning with Omni Perception

AgentsDGX agent

arXiv:2601.09413v2 Announce Type: replace-cross Abstract: We introduce a voice-agentic framework that learns one critical omni-understanding skill: knowing when to trust itself versus when to consult

Spherical Harmonic Optimal Transport: Application to Climate Models Comparisons

HardwareDGX agent

arXiv:2605.18389v1 Announce Type: new Abstract: Optimal transport provides a powerful framework for comparing measures while respecting the geometry of their support, but comes with an expensive compu

Spherical Steering: Geometry-Aware Activation Rotation for Language Models

ResearchDGX agent

arXiv:2602.08169v2 Announce Type: replace-cross Abstract: Inference-time steering offers a promising way to control language models (LMs) without retraining. However, standard approaches typically rel

Spherical VAE with Cluster-Aware Feasible Regions: Guaranteed Prevention of Posterior Collapse

Model ReleasesDGX agent

arXiv:2603.10935v4 Announce Type: replace-cross Abstract: Variational autoencoders (VAEs) frequently suffer from posterior collapse, where the latent variables become uninformative as the approximate

SPIKE: An Adaptive Dual Controller Framework for Cost-Efficient Long-Horizon Game Agents

ResearchDGX agent

arXiv:2605.18636v1 Announce Type: new Abstract: Long-horizon multimodal agents in open-world games must stay goal-directed across many low-level interactions under tight token and latency budgets. Exi

Spiker-LL: An Energy-Efficient FPGA Accelerator Enabling Adaptive Local Learning in Spiking Neural Networks

Local AiDGX agent

arXiv:2605.18003v1 Announce Type: cross Abstract: Deploying adaptive intelligence at the edge remains challenging due to the high computational and energy cost of training neural models. Spiking Neura

Spotify's Chief Architect just showed how they ship 4,5K deployments /day with Claude at Anthropic stage 27-minutes. free. By #1 music app d…

Model ReleasesDGX agent

Spotify's Chief Architect just showed how they ship 4,5K deployments /day with Claude at Anthropic stage 27-minutes. free. By #1 music app dev 'More than 99% of our engineers use AI coding tools. Adop

SRC-Flow: Compact Semantic Representations Enable Normalizing Flows for Image Generation

TutorialsDGX agent

arXiv:2605.18267v1 Announce Type: new Abstract: Normalizing flows (NFs) provide exact likelihoods and deterministic invertible sampling, but have historically lagged behind diffusion models for large-

SSL4RL: Revisiting Self-supervised Learning as Intrinsic Reward for Visual-Language Reasoning

SafetyDGX agent

arXiv:2510.16416v4 Announce Type: replace-cross Abstract: Vision-language models (VLMs) have shown remarkable abilities by integrating large language models with visual inputs. However, they often fai

SSTL: Self-Sensing Tendon Loop for Hysteresis Modeling and Compensation in Tendon-Sheath Mechanisms

ResearchDGX agent

arXiv:2605.16870v1 Announce Type: new Abstract: Flexible endoscopic robots enable minimally invasive access through natural orifices, but their control accuracy is limited by configuration-dependent h

ST-BCP: Tightening Coverage Bound for Backward Conformal Prediction via Non-Conformity Score Transformation

ResearchDGX agent

arXiv:2602.01733v2 Announce Type: replace-cross Abstract: Conformal Prediction (CP) provides a statistical framework for uncertainty quantification that constructs prediction sets with coverage guaran

Stabilizing, Scaling & Enhancing MeanFlow for Large-scale Diffusion Distillation

Model ReleasesDGX agent

arXiv:2605.17834v1 Announce Type: new Abstract: Diffusion models exhibit remarkable generative capability, but their high latency limits practical deployment. Many studies have attempted to reduce sam

Stabilizing Temporal Inference Dynamics for Online Surgical Phase Recognition

ResearchDGX agent

arXiv:2605.16387v1 Announce Type: cross Abstract: Online Surgical Phase Recognition (SPR) models can reach high frame-wise accuracy, yet their predictions often lack temporal stability, fragmenting wo

Stable and Near-Reversible Diffusion ODE Solvers for Image Editing

SafetyDGX agent

arXiv:2605.16399v1 Announce Type: new Abstract: The inversion of diffusion models plays a central role in image editing. Algebraically reversible ODE solvers provide an appealing approach to diffusion

Stable Audio 3

HardwareDGX agent

arXiv:2605.17991v1 Announce Type: cross Abstract: Stable Audio 3 is a family of fast latent diffusion models (small, medium, large) for variable-length audio generation and editing. Since our models c

← Previous
1…926927928929930…1473
Next →