AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,488 results
Safety

Mixture-of-Gaussians-Guided Schedule Design for Brownian Bridge Diffusion Models

DGX agent

arXiv:2607.03517v1 Announce Type: cross Abstract: Brownian Bridge Diffusion Models (BBDM) offer an appealing framework for image restoration and inverse problems by constructing a stochastic bridge fr

safetyarxiv-cs-cv
7 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

More than half of Americans can’t afford food or gas. Meanwhile, the top .1% increased their wealth by $2.59T during Trump’s second term. Re…

DGX agent

More than half of Americans can’t afford food or gas. Meanwhile, the top .1% increased their wealth by $2.59T during Trump’s second term. Republicans ran on affordability, and then immediately passed

safetyyann-lecun--x
7 Jul 2026
Safety

MT-EditFlow: Reinforcement Learning for Multi-Turn Image Editing with Flow Matching

DGX agent

Recent breakthroughs in instruction-based image editing have captured significant attention, as models are now capable of handling real-world editing demands with the practicality required by everyday

safetyapple-ml-research
7 Jul 2026
Safety

Multi-Turn On-Policy Distillation with Prefix Replay

DGX agent

arXiv:2607.04763v1 Announce Type: cross Abstract: We study on-policy distillation (OPD) for agentic tasks, where an LLM agent interacts with an environment over multiple turns and a student imitates a

safetyarxiv-cs-ai
7 Jul 2026
Safety

Multi-Way Representation Alignment

DGX agent

arXiv:2602.06205v2 Announce Type: replace-cross Abstract: The Platonic Representation Hypothesis suggests that independently trained neural networks converge to increasingly similar latent spaces. How

safetyarxiv-cs-ai
7 Jul 2026
Safety

MV-Forcing: Long Multi-View Video Generation via 4D-Grounded Spatio-Temporal Self-Forcing

DGX agent

arXiv:2607.05376v1 Announce Type: new Abstract: Recent advances in video diffusion models have enabled either long single-view generation through temporal autoregression, or short multi-view synthesis

safetyarxiv-cs-cv
7 Jul 2026
Safety

NeuroOnline: Bridging Pretraining and Online Adaptation for EEG Foundation Models

DGX agent

arXiv:2607.03925v1 Announce Type: new Abstract: EEG foundation models have shown strong potential in learning generalized representations across subjects and tasks. However, most existing approaches f

safetyarxiv-cs-lg
7 Jul 2026
Safety

new @WSJ editorial board piece on 'The Socialist Temptation of Sam Altman' highlights dangers of politicization of AI, regulatory capture, &…

DGX agent

new @WSJ editorial board piece on 'The Socialist Temptation of Sam Altman' highlights dangers of politicization of AI, regulatory capture, & bailouts. 'The larger harm will be to the U.S. economy if i

safetygary-marcus--x
7 Jul 2026
Safety

No Time Like the Present: Agentic Test-Time Training for LLM Agents

DGX agent

arXiv:2607.03441v1 Announce Type: cross Abstract: LLM agents often degrade over long episodes: as trajectories grow, they revisit explored states, repeat failed actions, and lose strategies that previ

safetyarxiv-cs-ai
7 Jul 2026
Safety

Non-Asymptotic Error Bounds for SMC with Biased Proposals: Application to Conditional Diffusion Sampling

DGX agent

arXiv:2607.04780v1 Announce Type: cross Abstract: Sequential Monte Carlo (SMC) methods are a natural tool for post-hoc conditioning of pretrained generative models, but in many applications the mutati

safetyarxiv-cs-lg
7 Jul 2026
Safety

OmniDS: Dual-Stream Context Fusion for Omnidirectional Depth from Fisheye Cameras

DGX agent

arXiv:2607.03038v1 Announce Type: new Abstract: Omnidirectional depth estimation from multi-fisheye camera rigs is complicated by visibility conflicts: wide baselines cause different cameras to observ

safetyarxiv-cs-cv
7 Jul 2026
Safety

OmniTacTune: Policy-Agnostic Real-World RL for Tactile Residual Adaptation of Visual Policies

DGX agent

arXiv:2607.03723v1 Announce Type: cross Abstract: Visual policies learned from human videos, teleoperation, and robot demonstrations offer scalable motion priors, but often fail in contact-rich manipu

safetyarxiv-cs-ai
7 Jul 2026
Safety

Online Linear Programming for Multi-Objective Routing in LLM Serving

DGX agent

arXiv:2607.03948v1 Announce Type: new Abstract: We study the online routing problem in large language model serving, where requests arrive sequentially and must be dispatched to parallel decode worker

safetyarxiv-cs-ai
7 Jul 2026
Safety

OpenAI just asked the US government to take a 5% stake in the company. The pitch: give every citizen a share in the profits of AI. The probl…

DGX agent

OpenAI just asked the US government to take a 5% stake in the company. The pitch: give every citizen a share in the profits of AI. The problem: AI has no profits. So the real plan is to give every cit

safetygary-marcus--x
7 Jul 2026
Safety

OpenTinker: Separating Concerns in Agentic Reinforcement Learning

DGX agent

arXiv:2601.07376v2 Announce Type: replace Abstract: We introduce extsc{OpenTinker}, an open infrastructure for training large language model (LLM) agents with many LoRA-backed policies over shared exe

safetyarxiv-cs-ai
7 Jul 2026
Safety

Optimality-Informed Neural Networks for Lunar Landing Trajectory Optimization

DGX agent

arXiv:2607.02741v1 Announce Type: cross Abstract: This paper develops an Optimality-Informed Neural Network (OINN) approach for the energy-optimal, free-final-time powered descent of a lunar lander fr

safetyarxiv-cs-lg
7 Jul 2026
Safety

Parity-Aware Byte-Pair Encoding: Improving Cross-lingual Fairness in Tokenization

DGX agent

arXiv:2508.04796v3 Announce Type: replace-cross Abstract: Tokenization is the first -- and often least scrutinized -- step of most NLP pipelines. Standard algorithms for learning tokenizers rely on fr

safetyarxiv-cs-ai
7 Jul 2026
Safety

PIEFS: Physics-Informed Eigenfunction Features with Learnable Scaling

DGX agent

arXiv:2607.03692v1 Announce Type: new Abstract: Spectral methods are widely used to construct representations from the geometry of data, but they often rely on a fixed kernel, graph Laplacian, or manu

safetyarxiv-cs-lg
7 Jul 2026
Safety

PixelPilot: Scalable Vision-Language-Action Models for End-to-End Autonomous Driving

DGX agent

arXiv:2607.04637v1 Announce Type: new Abstract: Vision-Language-Action Models (VLAs), which leverage the advanced reasoning capabilities of Vision-Language Models (VLMs), show promising generalization

safetyarxiv-cs-cv
7 Jul 2026
Safety

Policy Improvement with Style-Specific Demonstrations

DGX agent

arXiv:2506.16995v4 Announce Type: replace Abstract: Proficient game agents with diverse play styles enrich the gaming experience and enhance the replay value of games. However, recent advancements in

safetyarxiv-cs-ai
7 Jul 2026
Safety

Post-Generation Curation of Synthetic Images via Homogeneous-Heterogeneous Splitting

DGX agent

arXiv:2607.02637v1 Announce Type: cross Abstract: Recent generative models can produce high-quality synthetic images, offering scalable training training data for data-hungry models. Existing approach

safetyarxiv-cs-ai
7 Jul 2026
Safety

PRIMA: Pre-training with Risk-integrated Image-Metadata Alignment for Medical Diagnosis via LLM

DGX agent

arXiv:2602.23297v2 Announce Type: replace Abstract: Medical diagnosis requires the effective synthesis of visual manifestations and clinical metadata. However, existing methods often treat metadata as

safetyarxiv-cs-cv
7 Jul 2026
Safety

PRISM: Personalized Robotic Dataset Generation via Image-based Scene and Motion Synthesis

DGX agent

arXiv:2607.04880v1 Announce Type: new Abstract: Recent advances in large-scale pretrained vision-language-action models have improved robot policy learning, but directly deploying such policies in use

safetyarxiv-cs-ro
7 Jul 2026
Safety

Progress- and Reliability-Oriented Group Policy Optimization for Agentic Reinforcement Learning

DGX agent

arXiv:2607.04242v1 Announce Type: new Abstract: Group-based reinforcement learning (RL) has become an effective paradigm for improving large language model agents on long-horizon interactive tasks. To

safetyarxiv-cs-ai
7 Jul 2026
Safety

Proportionally Representative Clustering

DGX agent

arXiv:2304.13917v4 Announce Type: replace Abstract: In recent years, there has been a surge in effort to formalize notions of fairness in machine learning. We focus on centroid clustering--one of the

safetyarxiv-cs-lg
7 Jul 2026
Safety

ProxyUp: Training-Free Proxy-Conditioned Video Generation for Controllable Dynamics

DGX agent

arXiv:2607.03732v1 Announce Type: new Abstract: Precise control over complex dynamics remains challenging for modern video generative models, as text prompts alone often cannot specify physically plau

safetyarxiv-cs-cv
7 Jul 2026
Safety

QSVideo: Query-Conditioned Semantic Temporal Retrieval for Video Understanding

DGX agent

arXiv:2607.04559v1 Announce Type: new Abstract: The performance of vision-language models (VLMs) in video understanding declines with increasing video duration, as video moments unrelated to the query

safetyarxiv-cs-cv
7 Jul 2026
Safety

R^2PO: Decoupling Rollout and Inference Policies for LLM Reasoning

DGX agent

arXiv:2601.11960v3 Announce Type: replace-cross Abstract: Existing reinforcement learning methods for LLM reasoning implicitly assume that the policy generating training trajectories should coincide w

safetyarxiv-cs-ai
7 Jul 2026
Safety

RADIANCE: Relative Adaptive Denoising with IP-Adapter for Novel Concept Enhancement

DGX agent

arXiv:2607.05088v1 Announce Type: new Abstract: Text-to-image (T2I) diffusion models have achieved striking progress but still struggle to synthesize rare concepts involving unusual attribute-object p

safetyarxiv-cs-cv
7 Jul 2026
Safety

RADIO1D: Elastic Representations for Condensed Vision Modeling

DGX agent

arXiv:2607.03624v1 Announce Type: cross Abstract: This paper challenges the assumption that vision-language models (VLMs) require fixed patch-based 2D vision features. Analyzing fine-tuned vision enco

safetyarxiv-cs-ai
7 Jul 2026
Safety

Reason, Reward, Refine: Step-Level Errors Corrections with Structured Feedback for Physics Reasoning in Small Language Models

DGX agent

arXiv:2607.05199v1 Announce Type: new Abstract: Physics reasoning fails structurally in small language models: an error at any step propagates forward, corrupting every inference that follows. Limited

safetyarxiv-cs-ai
7 Jul 2026
Safety

ReCal3R: Reliability-Calibrated Learning Rates for Streaming 3D Reconstruction

DGX agent

arXiv:2607.05356v1 Announce Type: new Abstract: Streaming 3D reconstruction relies on a compact recurrent scene state to process long image streams in linear time and bounded memory. However, repeated

safetyarxiv-cs-cv
7 Jul 2026
Safety

REDDIT: Correcting Model-Generated Timestamp Drift in ASR without Forgetting via Replay-Based Distribution Editing

DGX agent

arXiv:2607.05364v1 Announce Type: cross Abstract: Modern autoregressive ASR systems can emit timestamps as decoded tokens, enabling timestamped transcription without frame-level aligners or inference-

safetyarxiv-cs-ai
7 Jul 2026
Safety

Reference-Induced Consensus for Selective Posed-Reference Visual Localization

DGX agent

arXiv:2607.04722v1 Announce Type: new Abstract: We present RIC-Loc (Reference-Induced Consensus localization), a scene-training-free posed-reference localizer that is SfM-point-map-free in its main es

safetyarxiv-cs-cv
7 Jul 2026
Safety

Regime-Conditional Stabilisation of LLM-Augmented Cooperative Multi-Agent Reinforcement Learning

DGX agent

arXiv:2607.04470v1 Announce Type: cross Abstract: Large Language Models (LLMs) offer a natural interface for translating human objectives into reward signals for cooperative multi-agent reinforcement

safetyarxiv-cs-ai
7 Jul 2026
Safety

Reinforcement Learning for Data-Efficient Code-Switched ASR

DGX agent

arXiv:2607.02757v1 Announce Type: new Abstract: Audio-language models can be prompted for code-switched speech, but their decoding is not optimized for code-switching and often fails at language bound

safetyarxiv-cs-cl
7 Jul 2026
Safety

Relational Multi-Agent Reinforcement Learning for Dynamic Pricing in High-Speed Railway Markets

DGX agent

arXiv:2607.05179v1 Announce Type: cross Abstract: In liberalised railway systems, operators must set prices dynamically in an environment with partial observability, as they retain private information

safetyarxiv-cs-ai
7 Jul 2026
Safety

Reliability and Identifiability in Persona-Trained Monte Carlo: Variance Decomposition, Stability Bounds, and the Identifiability of Heterogeneous News Reaction

DGX agent

arXiv:2607.04627v1 Announce Type: new Abstract: Persona-Trained Monte Carlo (PTMC) estimates distributions of market-outcome functionals by repeatedly simulating limit-order-book interaction among K n

safetyarxiv-cs-lg
7 Jul 2026
Safety

Reliability-Aware CT-MRI Registration: A Quality Engineering Framework with Stability Analysis and Risk Classification

DGX agent

arXiv:2607.02585v1 Announce Type: new Abstract: Multimodal CT-MRI registration is central to image-guided radiotherapy, surgical navigation, and diagnostic workflows, but most pipelines report only ag

safetyarxiv-cs-cv
7 Jul 2026
Safety

Rethinking Brain Decoding with CLIP: The Role of Adversarial Robustness

DGX agent

arXiv:2607.03165v1 Announce Type: new Abstract: Brain decoding aims to uncover neural mechanisms by inferring stimulus-related representations from brain signals. In fMRI studies, this is typically ac

safetyarxiv-cs-cv
7 Jul 2026
Safety

Rethinking On-Policy Self-Distillation for Thinking Models

DGX agent

arXiv:2607.05184v1 Announce Type: new Abstract: Self-distillation is a promising recipe for self-improvement in language models. In this setting, a model can serve as its own teacher when given privil

safetyarxiv-cs-ai
7 Jul 2026
Safety

Reward-Gated On-Policy Distillation

DGX agent

arXiv:2607.04037v1 Announce Type: cross Abstract: On-policy distillation is a powerful way to transfer reasoning ability from a strong teacher to a smaller student: the student samples trajectories fr

safetyarxiv-cs-ai
7 Jul 2026
Safety

Reward Granularity in RLVR: Comparing Process and Outcome Reward Structures for Mathematical Reasoning in Small Language Models

DGX agent

arXiv:2607.02869v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has emerged as a promising paradigm for improving mathematical reasoning in language models. Yet m

safetyarxiv-cs-lg
7 Jul 2026
Safety

Reward Lightning: Fast Video Generation via Homologous Preference Distillation

DGX agent

arXiv:2607.03960v1 Announce Type: new Abstract: Achieving simultaneous preference alignment and distillation acceleration in video diffusion models remains an open challenge. Existing methods optimize

safetyarxiv-cs-cv
7 Jul 2026
Safety

RSPO: Reward-Swap Policy Optimization for Multi-Turn LLM Agents

DGX agent

arXiv:2607.04713v1 Announce Type: cross Abstract: Reinforcement learning holds significant potential for training large language models (LLMs) to handle multi-turn interactive tasks. However, in long-

safetyarxiv-cs-ai
7 Jul 2026
Safety

SABLE: An NDA-Safe Closed-Loop LLM Framework for Analog Circuit Optimization in Industrial EDA Flows

DGX agent

arXiv:2607.03701v1 Announce Type: cross Abstract: Large language models (LLMs) can propose circuit-optimization decisions, but industrial analog flows cannot expose foundry PDK content, proprietary sc

safetyarxiv-cs-lg
7 Jul 2026
Safety

Sample-Efficient Pareto Front Modeling for Energy-Aware Reinforcement Learning Using Bayesian Optimization

DGX agent

arXiv:2607.03140v1 Announce Type: new Abstract: Industrial automation increasingly demands control strategies that balance operational performance with strict energy efficiency requirements. A common

safetyarxiv-cs-lg
7 Jul 2026
Safety

Saving GPU Hours in LLM Inference System Development and Online Workloads with Simulation and DBMS-Inspired Cache Replacement Policies

DGX agent

arXiv:2411.07447v5 Announce Type: replace-cross Abstract: LLMs are increasingly used world-wide from daily tasks to agentic systems and data analytics, requiring significant GPU resources. While LLM i

safetyarxiv-cs-ai
7 Jul 2026
← Previous
1…109110111112113…302
Next →