AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
91,060Total entries
1Added by human
91,059Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
91,059 results
Research

Rethinking Stepwise Model Routing: A Cost-Efficient Table Reasoning Perspective

DGX agent

arXiv:2605.29319v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) achieve strong performance on table reasoning tasks but incur substantial inference cost due to long reasoning traces. Ste

researcharxiv-cs-cl
29 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Return-to-Go Is More Than a Number: Q-Guided Alignment for Return-Conditioned Supervised Learning

DGX agent

arXiv:2605.29028v1 Announce Type: cross Abstract: Conditioned Sequence Models (CSMs) learn policies by treating return-to-go (RTG) as a control signal. However, existing CSMs often treat the RTGs as s

model-releasesarxiv-cs-ai
29 May 2026
Safety

Review Arcade: On the Human Alignment and Gameability of LLM Reviews

DGX agent

arXiv:2605.28897v1 Announce Type: new Abstract: LLM-generated reviews for scientific papers are gaining considerable traction and are even being officially piloted by major conferences. We have to ass

safetyarxiv-cs-ai
29 May 2026
Agents

Revisiting Observation Reduction for Web Agents: Comprehensive Evaluation with a Lightweight Framework

DGX agent

arXiv:2605.29397v1 Announce Type: new Abstract: HTML observations in LLM-based web agents are extremely long, and while many reduction methods have been proposed, it remains unclear which methods redu

agentsarxiv-cs-cl
29 May 2026
Agents

RewardFlow: Topology-Aware Reward Propagation on State Graphs for Agentic RL with Large Language Models

DGX agent

arXiv:2603.18859v2 Announce Type: replace Abstract: Reinforcement learning (RL) shows promise for enhancing LLM agentic reasoning, yet sparse terminal rewards hinder fine-grained optimization. Process

agentsarxiv-cs-ai
29 May 2026
Model Releases

RHO: Robust Holistic OSM-Based Metric Cross-View Geo-Localization

DGX agent

arXiv:2603.27758v2 Announce Type: replace Abstract: Metric Cross-View Geo-Localization (MCVGL) aims to estimate the 3-DoF camera pose (position and heading) by matching ground and satellite images. In

model-releasesarxiv-cs-cv
29 May 2026
Research

Ridge Regression from Poisson Resetting: A Renewal Perspective on Spectral Regularization

DGX agent

arXiv:2605.30059v1 Announce Type: new Abstract: We connect stochastic resetting from non-equilibrium statistical physics with ridge regularization in statistical learning. For linear gradient flow, re

researcharxiv-cs-lg
29 May 2026
Research

Riemannian AmbientFlow: Towards Simultaneous Manifold Learning and Generative Modeling from Corrupted Data

DGX agent

arXiv:2601.18728v2 Announce Type: replace Abstract: Modern generative modeling methods have demonstrated strong performance in learning complex data distributions from clean samples. In many scientifi

researcharxiv-cs-lg
29 May 2026
Model Releases

RightNow-Arabic-0.5B-Turbo: An Open Sub-1B Arabic Language Model via Vocabulary Injection and Edge-First Deployment

DGX agent

arXiv:2605.28827v1 Announce Type: new Abstract: Open Arabic large language models split into two classes: sub-1B multilingual models that treat Arabic as an afterthought (Qwen2.5-0.5B, Falcon-H1-0.5B)

model-releasesarxiv-cs-cl
29 May 2026
Model Releases

Risk-averse Fair Multi-class Classification

DGX agent

arXiv:2509.05771v2 Announce Type: replace-cross Abstract: We develop a new classification framework based on the theory of coherent risk measures and systemic risk. The proposed approach is suitable f

model-releasesarxiv-cs-lg
29 May 2026
Safety

RL2ML: Finite-Rollout Surrogate Objectives from Reinforcement Learning to Maximum Likelihood

DGX agent

arXiv:2605.30154v1 Announce Type: new Abstract: Correctness-based Reinforcement Learning with Verifiable Rewards (RLVR) trains language models from binary feedback on sampled outputs, but the objectiv

safetyarxiv-cs-lg
29 May 2026
Model Releases

RoboWits: Unexpected Challenges for Robotic Creative Problem Solving

DGX agent

arXiv:2605.30326v1 Announce Type: cross Abstract: The ability to reason, adapt, and creatively solve problems under unexpected challenges is essential for robots operating in real-world environments.

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Robust and Efficient Guardrails with Latent Reasoning

DGX agent

arXiv:2605.29068v1 Announce Type: new Abstract: Maintaining the safety of large language models (LLMs) is crucial as they are increasingly deployed in real-world applications. Existing safety guardrai

model-releasesarxiv-cs-ai
29 May 2026
Applications

Robust and Efficient Writer-Independent IMU-Based Handwriting Recognition

DGX agent

arXiv:2502.20954v3 Announce Type: replace Abstract: Handwriting recognition (HWR) using inertial measurement unit (IMU) data remains challenging due to variations in writing styles and the limited ava

applicationsarxiv-cs-lg
29 May 2026
Safety

Robust and Generalizable Safety Steering for Text-to-Image Diffusion Transformers

DGX agent

arXiv:2605.30049v1 Announce Type: new Abstract: Diffusion Transformers have become a powerful backbone for text-to-image generation, but their layered and cross-modal generation process makes safety c

safetyarxiv-cs-ai
29 May 2026
Tutorials

Robust Cross-Domain Generalization Using Unlabeled Target Data with Source-Domain Supervision

DGX agent

arXiv:2605.29122v1 Announce Type: new Abstract: It is often desirable to generalize medical imaging AI models trained with dense annotations to data acquired from different ultrasound scanners or clin

tutorialsarxiv-cs-cv
29 May 2026
Research

Robust Frequency-Calibrated Virtual EEG Channel Generation from Four Frontal Electrodes for Wearable EEG Augmentation

DGX agent

arXiv:2605.29263v1 Announce Type: new Abstract: Low-channel wearable electroencephalography (EEG) is attractive for long-term monitoring, but four frontal electrodes provide only a sparse and spatiall

researcharxiv-cs-lg
29 May 2026
Industry

Rocket Report: A dark day for Blue Origin; Pentagon eyes new launch site

DGX agent

Blue Origin's New Glenn rocket exploded on the launch pad at Cape Canaveral on May 28 during a static fire engine test, destroying the 321-foot rocket . The explosion froze all 24 of Amazon's contract

industryars-technica
29 May 2026
Safety

Rooted Absorbed Prefix Trajectory Balance with Submodular Replay for GFlowNet Training

DGX agent

arXiv:2603.00454v2 Announce Type: replace-cross Abstract: Generative Flow Networks (GFlowNets) enable fine-tuning large language models to approximate reward-proportional posteriors, but they remain p

safetyarxiv-cs-ai
29 May 2026
Tutorials

Routing by Reaching: Composition of Pre-trained GFlowNets for Multi-Objective Generation

DGX agent

arXiv:2602.21565v2 Announce Type: replace Abstract: Generative Flow Networks (GFlowNets) learn to sample diverse candidates in proportion to a reward function, making them well-suited for scientific d

tutorialsarxiv-cs-lg
29 May 2026
Safety

RUBRIC-ARROW: Alternating Pointwise Rubric Reward Modeling for LLM Post-training in Non-verifiable Domains

DGX agent

arXiv:2605.29156v1 Announce Type: cross Abstract: Pointwise reward modeling offers critical signals for LLM post-training, yet struggles with absolute scoring in subjective, non-verifiable settings. R

safetyarxiv-cs-cl
29 May 2026
Safety

Rubric-Guided Process Reward for Stepwise Model Routing

DGX agent

arXiv:2605.29310v1 Announce Type: new Abstract: Stepwise model routing improves the efficiency of Large Reasoning Models (LRMs) by assigning each reasoning step to a suitable model. Recent methods for

safetyarxiv-cs-ai
29 May 2026
Industry

RUD (rapid unscheduled disassembly) events are not unusual in the rocket world

DGX agent

RUD (rapid unscheduled disassembly) is a term used in the aerospace industry to describe unexpected vehicle failures or explosions during rocket testing and launches. Elon Musk's statement acknowledge

industryelon-musk--x
29 May 2026
Tools

Run Docker containers inside Vercel Sandbox

DGX agent

Vercel announced the ability to run Docker containers directly within Vercel Sandbox, enabling developers to execute containerized applications and services as part of their development and testing wo

toolsvercel-blog
29 May 2026
Hardware

Run Step 3.7 Flash on NVIDIA GPUs with Enterprise-Ready Multimodal AI

DGX agent

Step 3.7 Flash is a 198B-parameter Mixture-of-Experts vision-language model designed for enterprise-scale production workloads, featuring native image and video input, a 256k context window, and confi

hardwarenvidia-developer
29 May 2026
Model Releases

S-MARC: Causal Streaming Reasoning for Full-Duplex Conversational Behavior Modeling

DGX agent

arXiv:2602.11065v2 Announce Type: replace-cross Abstract: Human conversation is organized by an implicit chain of thought and manifests as temporally structured conversational behaviors. Capturing thi

model-releasesarxiv-cs-ai
29 May 2026
Research

S2MDF: A Plug-And-Play Layer for Intersection-Free Multi-Object Signed Distance Fields

DGX agent

arXiv:2605.29761v1 Announce Type: new Abstract: Compositional implicit surface representations model scenes as collections of objects, each encoded by a Signed Distance Field (SDF). A fundamental limi

researcharxiv-cs-cv
29 May 2026
Local Ai

S3Mem: Structured Spatiotemporal Scene-Event Memory for Long-Horizon Interactive Question Answering

DGX agent

arXiv:2605.28831v1 Announce Type: cross Abstract: Long-horizon interactive agents often accumulate large trajectory histories yet still fail to answer questions about earlier events reliably. We argue

local-aiarxiv-cs-ai
29 May 2026
Model Releases

SAAS: Self-Aware Reinforcement Learning for Over-Search Mitigation in Agentic Search

DGX agent

arXiv:2605.29796v1 Announce Type: new Abstract: Agentic search enables LLMs to solve complex multi-hop questions through iterative reasoning and external search. Despite the effectiveness, these syste

model-releasesarxiv-cs-ai
29 May 2026
Research

SADA: Safe and Adaptive Aggregation of Multiple Black-Box Predictions in Semi-Supervised Learning

DGX agent

arXiv:2509.21707v3 Announce Type: replace-cross Abstract: Semi-supervised learning (SSL) arises in practice when labeled data are scarce or expensive to obtain, while large quantities of unlabeled dat

researcharxiv-cs-lg
29 May 2026
Applications

SAFE-Pruner: Semantic Attention-Guided Future-Aware Token Pruning for Efficient Vision-Language-Action Manipulation

DGX agent

arXiv:2605.29662v1 Announce Type: new Abstract: Real-time inference of vision-language-action (VLA) models is essential for robotic control. While visual token pruning has shown strong potential for a

applicationsarxiv-cs-cv
29 May 2026
Safety

SafeRx-Agent: A Knowledge-Grounded Multi-Agent Framework for Safe and Explainable Medication Recommendation

DGX agent

arXiv:2605.29146v1 Announce Type: cross Abstract: Medication recommendation predicts medications for patient visits, but existing methods still face two key challenges. At the model level, traditional

safetyarxiv-cs-ai
29 May 2026
Model Releases

SafeSearch: Automated Red-Teaming of LLM-Based Search Agents

DGX agent

arXiv:2509.23694v5 Announce Type: replace Abstract: Search agents connect LLMs to the Internet, enabling them to access broader and more up-to-date information. However, this also introduces a new thr

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

SAGE: Segment-Aware Gloss-Free Encoding for Token-Efficient Sign Language Translation

DGX agent

arXiv:2507.09266v2 Announce Type: replace Abstract: Gloss-free Sign Language Translation (SLT) has advanced rapidly, achieving strong performances without relying on gloss annotations. However, these

model-releasesarxiv-cs-cv
29 May 2026
Safety

SAHG: Sector-Anisotropic Hyperbolic Graph Model for Social Bot Detection

DGX agent

arXiv:2605.30166v1 Announce Type: cross Abstract: LLM-driven social bots can generate fluent, human-like text, reducing the discriminative advantage of content-based detection alone. However, coordina

safetyarxiv-cs-lg
29 May 2026
Research

Sakana AIは、一般社団法人DEEP DIVEとAIを活用した情報分析に関するパートナーシップを締結しました。 https://sakana.ai/deep-dive-partnership/ ■専門家の知見・オープンデータ × 独自AI技術 ■人手では難しかった規模・速度…

DGX agent

Sakana AIは、一般社団法人DEEP DIVEとAIを活用した情報分析に関するパートナーシップを締結しました。 https://sakana.ai/deep-dive-partnership/ ■専門家の知見・オープンデータ × 独自AI技術 ■人手では難しかった規模・速度・解像度での分析を実現 当社が注力する「防衛・インテリジェンス」領域への社会実装を 本格化し、我が国の安全保障環境の発展

researchdavid-ha--x
29 May 2026
Model Releases

Salesforce published a detailed writeup on going agentic with Claude Code. A couple things jumped out. A migration they'd scoped at 231 days…

DGX agent

Salesforce published a detailed writeup on going agentic with Claude Code. A couple things jumped out. A migration they'd scoped at 231 days shipped in 13. One PR delivered 21 endpoints at 100% test c

model-releasesboris-cherny--x
29 May 2026
Research

SalsaAgent: A multimodal embodied language model for interactive dance generation

DGX agent

arXiv:2605.29219v1 Announce Type: new Abstract: Interaction between humanoids involves bidirectional and nonverbal reactivity, coordination and synchrony. Toward socially aware robots and interactive

researcharxiv-cs-cv
29 May 2026
Applications

SAM3D-Phys: Towards Multi-Object Interactive Simulation in Real World

DGX agent

arXiv:2605.30239v1 Announce Type: new Abstract: This work addresses the problem of recovering complete, simulatable object geometry from reconstructed real-world scenes, enabling physics-based interac

applicationsarxiv-cs-cv
29 May 2026
Safety

Same Evidence, Different Answers: Canonical-Context On-Policy Distillation for Multi-Turn Language Models

DGX agent

arXiv:2605.30251v1 Announce Type: cross Abstract: Large language models (LLMs) often solve a task when all instructions are given in a single prompt, but fail when the same information is revealed gra

safetyarxiv-cs-ai
29 May 2026
Model Releases

Same Question, Different Source, Different Answer: Auditing Source-Dependence in Medical Multi-Source RAG

DGX agent

arXiv:2605.29084v1 Announce Type: cross Abstract: A retrieval-augmented generation (RAG) system deployed over a multi-author institutional corpus can give a different answer to the same question depen

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Sample-Efficient Diffusion-based Reinforcement Learning with Critic Guidance

DGX agent

arXiv:2605.30056v1 Announce Type: cross Abstract: Recent advances in reinforcement learning (RL) have achieved great successes by leveraging the multimodality and exploration capability of diffusion p

model-releasesarxiv-cs-lg
29 May 2026
Industry

Samsung says it has started shipping its first 12-layer HBM4E samples to major clients; SK Hynix said in April that it aimed to ship HBM4E samples in H2 2026 (Yoolim Lee/Bloomberg)

DGX agent

Yoolim Lee / Bloomberg: Samsung says it has started shipping its first 12-layer HBM4E samples to major clients; SK Hynix said in April that it aimed to ship HBM4E samples in H2 2026 — Samsung Electron

industrytechmeme
29 May 2026
Research

SAVAA: Mitigating Hallucinations in LVLMs via Step-wise Adaptive Visual Attention Amplification

DGX agent

arXiv:2602.13600v2 Announce Type: replace Abstract: A line of recent training-free methods for mitigating hallucinations in large vision-language models (LVLMs) operates by amplifying attention to vis

researcharxiv-cs-cv
29 May 2026
Applications

Scalable RF Simulation in Generative 4D Worlds

DGX agent

arXiv:2508.12176v2 Announce Type: replace-cross Abstract: Radio Frequency (RF) sensing has emerged as a powerful, privacy-preserving alternative to vision-based methods for various perception tasks. H

applicationsarxiv-cs-ai
29 May 2026
Agents

// Scaling Laws for Agent Harnesses // If you build agent harnesses, this one is worth your time. (bookmark it) Most harness tuning treats e…

DGX agent

// Scaling Laws for Agent Harnesses // If you build agent harnesses, this one is worth your time. (bookmark it) Most harness tuning treats every token and tool call as if volume is all that counts. Ne

agentsdair-ai--x
29 May 2026
Model Releases

Scaling Laws for Agent Harnesses via Effective Feedback Compute

DGX agent

arXiv:2605.29682v1 Announce Type: new Abstract: Agent harnesses increasingly determine the performance of language-model systems by deciding how models call tools, receive feedback, verify intermediat

model-releasesarxiv-cs-cl
29 May 2026
Model Releases

Scaling Monosemanticity: Extracting Interpretable Features from Claude 3 Sonnet

DGX agent

arXiv:2605.29358v1 Announce Type: new Abstract: We demonstrate that sparse autoencoders can extract interpretable features from Claude 3 Sonnet, a production-scale language model, addressing the open

model-releasesarxiv-cs-ai
29 May 2026
← Previous
1…10041005100610071008…1898
Next →