AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlog
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,481 results
Safety

YouTube-Occ: Learning Indoor 3D Semantic Occupancy Prediction from YouTube Videos

DGX agent

arXiv:2506.18266v2 Announce Type: replace Abstract: 3D semantic occupancy prediction is crucial for fine-grained scene understanding, yet its advancement in privacy-sensitive indoor environments is fu

safetyarxiv-cs-cv
6 Aug 2026
Safety
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

A Hierarchical Approach to Imitation Learning for Manipulation Tasks Requiring Time Varying Forces

DGX agent

arXiv:2608.03103v1 Announce Type: cross Abstract: Diffusion policies have shown strong performance in learning complex, multi-modal behaviors for robotic manipulation. However, their application to co

safetyarxiv-cs-ai
5 Aug 2026
Safety

A Physics-Flavored Transformer Network for Parametrizing Contraction Dynamics of Engineered Skeletal Muscle Tissues

DGX agent

arXiv:2608.03927v1 Announce Type: new Abstract: Engineered Skeletal Muscle Tissues (ESMs) have become a key structure for biomedical disease modeling and pharmacological screening, yet their functiona

safetyarxiv-cs-lg
5 Aug 2026
Safety

A Security-Oriented Lifecycle Model for Large Language Model Systems

DGX agent

arXiv:2608.03626v1 Announce Type: cross Abstract: Large language models are being integrated into critical infrastructure and enterprise workflows at unprecedented scale,yet the lifecycle frameworks g

safetyarxiv-cs-ai
5 Aug 2026
Safety

Agentic Reinforcement Learning with Self-Distilled Reward Shaping

DGX agent

arXiv:2608.03223v1 Announce Type: cross Abstract: Agentic reinforcement learning enables LLM agents to learn through interaction, but sparse trajectory-level rewards reveal success without identifying

safetyarxiv-cs-ai
5 Aug 2026
Safety

AI Alignment and Fiduciary Obligation

DGX agent

arXiv:2608.02660v1 Announce Type: cross Abstract: Advanced AI assistants engage users in extended interactions across a widening range of roles, including advice, decision support, collaboration, lear

safetyarxiv-cs-ai
5 Aug 2026
Safety

AI-Assisted Peer Review Across Research Communities: From Reviewer AI Policies to LLM Review Quality

DGX agent

arXiv:2608.03581v1 Announce Type: cross Abstract: AI-assisted peer review is increasingly discussed and adopted as a tool to support the scientific publishing process, yet there is little systematic u

safetyarxiv-cs-ai
5 Aug 2026
Safety

BAP-SQL: Budget-Aware Observation Planning for Agentic Text-to-SQL

DGX agent

arXiv:2608.02876v1 Announce Type: new Abstract: Tool-using agents do not merely consume observations: their actions determine what arrives next. In agentic text-to-SQL, a broad query can spend context

safetyarxiv-cs-ai
5 Aug 2026
Safety

Bimanual Manipulation Within an 8 GB Budget: Zero-Copy Sensing and Quantized ACT on an Entry-Level Jetson

DGX agent

arXiv:2608.03938v1 Announce Type: new Abstract: Bimanual manipulation policies trained with imitation learning are typically evaluated on workstation or datacenter-class GPUs, leaving the cost of depl

safetyarxiv-cs-ro
5 Aug 2026
Safety

BODHI: Do LLMs Branch Out and Discover Heterogeneous Inferences?

DGX agent

arXiv:2608.02867v1 Announce Type: cross Abstract: Although reinforcement learning with verifiable rewards (RLVR) has improved the performance of large language models (LLMs) across a variety of reason

safetyarxiv-cs-ai
5 Aug 2026
Safety

BOW: Training Language Models to Reason Over Plausible Next Words

DGX agent

arXiv:2506.13502v3 Announce Type: replace Abstract: Next-word prediction (NWP) trains language models against a single observed continuation, even though many contexts admit multiple plausible next wo

safetyarxiv-cs-cl
5 Aug 2026
Safety

Bridging Prediction and Attribution: Identifying Forward and Backward Causal Influence Ranges Using Assimilative Causal Inference

DGX agent

arXiv:2510.21889v2 Announce Type: replace-cross Abstract: Causal inference identifies cause-and-effect relationships between variables. While traditional approaches rely on data to reveal causal links

safetyarxiv-cs-lg
5 Aug 2026
Safety

CAPE-T2V: Captioner-Anchored Prompt Enhancement toward Two-Sided Conditioning Alignment in Text-to-Video Generation

DGX agent

arXiv:2608.03046v1 Announce Type: new Abstract: Text-to-video (T2V) diffusion transformers (DiTs) are trained with detailed video captions, whereas inference often relies on user prompts rewritten by

safetyarxiv-cs-cv
5 Aug 2026
Safety

CausalOPD: First-Wrong-Step Supervision for Distilling Causal Chain Reasoning

DGX agent

arXiv:2608.03673v1 Announce Type: new Abstract: Many critical reasoning tasks, including clinical diagnosis, legal judgment, and industrial fault diagnosis, require step-dependent causal chains in whi

safetyarxiv-cs-lg
5 Aug 2026
Safety

Channel-wise Dynamic Knowledge Distillation via Adaptive Sample Generation for Action Recognition

DGX agent

arXiv:2608.03100v1 Announce Type: new Abstract: Knowledge Distillation (KD) offers a promising yet underexplored path for compressing large action recognition models. However, existing KD methods suff

safetyarxiv-cs-cv
5 Aug 2026
Safety

Clinically-Grounded Hierarchical Classification for Consistent Chest X-ray Interpretation

DGX agent

arXiv:2608.03016v1 Announce Type: new Abstract: Accurate chest X-ray interpretation is inherently hierarchical. Clinical decisions depend not only on what abnormality is present but where it is situat

safetyarxiv-cs-cv
5 Aug 2026
Safety

CLIP4VI-ReID: Learning Modality-shared Representations via CLIP Semantic Bridge for Visible-Infrared Person Re-identification

DGX agent

arXiv:2511.10309v2 Announce Type: replace Abstract: This paper proposes a novel CLIP-driven modality-shared representation learning network named CLIP4VI-ReID for VI-ReID task, which consists of Text

safetyarxiv-cs-cv
5 Aug 2026
Safety

⚠️⚠️⚠️ Constitutional AI is not working. Aligning LLMs is not working. We need a different approach. If society doesn’t place its bets diffe…

DGX agent

⚠️⚠️⚠️ Constitutional AI is not working. Aligning LLMs is not working. We need a different approach. If society doesn’t place its bets differently, we are screwed. Anthropic's Mythos created fake iden

safetygary-marcus--x
5 Aug 2026
Safety

Continue or Replan? Bernoulli-Continuation Policy Learning for Adaptive Horizon Execution

DGX agent

arXiv:2608.03483v1 Announce Type: cross Abstract: Existing chunk-based Vision-Language-Action (VLA) models execute a fixed number of actions (i.e., execution horizon) before replanning, turning replan

safetyarxiv-cs-ai
5 Aug 2026
Safety

CROSS: Cascaded Distillation and Dual-Constraint Grounding for Remote Sensing Referring Segmentation

DGX agent

arXiv:2608.03147v1 Announce Type: new Abstract: Referring Remote Sensing Image Segmentation (RRSIS) has achieved significant progress through the integration of VLMs and the Segment Anything Model (SA

safetyarxiv-cs-cv
5 Aug 2026
Safety

CT-HEG: A Bidirectional, Timestamp-Attributed Event Graph for ICU In-Hospital Mortality Prediction - An Architectural Ablation Study

DGX agent

arXiv:2608.02663v1 Announce Type: cross Abstract: Accurate ICU mortality prediction requires modeling irregular clinical observations across heterogeneous entity types. Existing sequence models handle

safetyarxiv-cs-ai
5 Aug 2026
Safety

CVPO: Enhancing LLM Reinforcement Learning Reasoning via Value-Variance Adaptation and Dynamic Curriculum Learning

DGX agent

arXiv:2608.03068v1 Announce Type: cross Abstract: Reinforcement learning (RL) has emerged as an effective method for enhancing the reasoning capabilities of large language models (LLMs). However, exis

safetyarxiv-cs-ai
5 Aug 2026
Safety

Deferred Exposure of Future Trajectories for Verifiable Reasoning in Autonomous Driving VLMs

DGX agent

arXiv:2608.01755v2 Announce Type: replace Abstract: Recent Vision-Language-Action (VLA) models for autonomous driving (AD) increasingly utilize chain-of-thought (CoT) supervision to enhance the reason

safetyarxiv-cs-ai
5 Aug 2026
Safety

Disentangling Language Modeling and Boundaries

DGX agent

arXiv:2608.03599v1 Announce Type: new Abstract: Byte-level language models are usually argued for on the grounds of robustness, multilingual fairness, and character-level skills. We point to a differe

safetyarxiv-cs-cl
5 Aug 2026
Safety

DiverseDiT++: Quantifying, Analyzing, and Promoting Representation Diversity in Diffusion Transformers

DGX agent

arXiv:2608.03082v1 Announce Type: new Abstract: Recent advances in Diffusion Transformers (DiTs) have enabled remarkable progress in visual synthesis, benefiting from their superior scalability. To fa

safetyarxiv-cs-cv
5 Aug 2026
Safety

DocTrace: Towards Traceable Long Document VQA via Hierarchical Evidence Graph Reasoning

DGX agent

arXiv:2608.03292v1 Announce Type: new Abstract: Long Document Visual Question Answering (LongDocVQA) requires Multimodal Large Language Models (MLLMs) to locate, integrate, and reason over heterogeneo

safetyarxiv-cs-ai
5 Aug 2026
Safety

Don't Peek at the Answer: Outcome-Masked Group Relative Policy Optimization for Label-Free RLVR

DGX agent

arXiv:2608.03119v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) improves LLM reasoning but typically relies on ground-truth (GT) answers, limiting scalability. Vo

safetyarxiv-cs-ai
5 Aug 2026
Model Releases

Don't Walk the Line: Boundary Guidance for Filtered Generation

DGX agent

arXiv:2510.11834v3 Announce Type: replace-cross Abstract: Generative models are increasingly paired with safety classifiers that filter harmful or undesirable outputs. A common strategy is to fine-tun

model-releasesarxiv-cs-cl
5 Aug 2026
Safety

DRIFT: Derailing Denoising Trajectories of Flow-Matching VLAs with Adversarial Patch Attack

DGX agent

arXiv:2608.03207v1 Announce Type: new Abstract: Flow-matching vision-language-action (VLA) models such as pi0 generate robot actions by integrating a learned denoising velocity field, and have been re

safetyarxiv-cs-cv
5 Aug 2026
Safety

DriftWorld: Fast World Modeling through Drifting

DGX agent

arXiv:2607.15065v2 Announce Type: replace-cross Abstract: Predictive world models enable robots to plan by imagining the outcomes of their actions, but their value for control hinges on generating man

safetyarxiv-cs-cv
5 Aug 2026
Safety

Emulate or Estimate? The Divergent Strengths of Base and Post-Trained Language Models for Opinion Simulation

DGX agent

arXiv:2608.03044v1 Announce Type: cross Abstract: Large language models are increasingly used to simulate human opinions, but prior work reports conflicting results: some studies find promising alignm

safetyarxiv-cs-ai
5 Aug 2026
Safety

Enhanced Polarization Locking in VCSELs

DGX agent

arXiv:2604.01857v2 Announce Type: replace-cross Abstract: While optical injection locking (OIL) of vertical-cavity surface-emitting lasers (VCSELs) has been widely studied in the past, the polarizatio

safetyarxiv-cs-cv
5 Aug 2026
Safety

Enhancing Q-Value Updates in Deep Q-Learning via Successor-State Prediction

DGX agent

arXiv:2511.03836v2 Announce Type: replace Abstract: Deep Q-Networks (DQNs) estimate future returns by learning from transitions sampled from a replay buffer. However, the target updates in DQN often r

safetyarxiv-cs-lg
5 Aug 2026
Safety

Enhancing VLM Reward Models Through Structure-Aware Fine-Tuning

DGX agent

arXiv:2608.03875v1 Announce Type: cross Abstract: Designing effective reward functions remains a major bottleneck in Reinforcement Learning (RL). Recent work uses large foundation Vision-Language Mode

safetyarxiv-cs-ai
5 Aug 2026
Safety

EvoHIL: Self-Evolving Reward and Flow-Matched Policy Optimization for Robust Human-in-the-Loop Reinforcement Learning

DGX agent

arXiv:2608.03872v1 Announce Type: new Abstract: Human-in-the-loop reinforcement learning (HIL-RL) enables robots to learn contact-rich manipulation from limited real-world interaction, but deployment

safetyarxiv-cs-ro
5 Aug 2026
Safety

Fields Medalist Jacob Tsimerman knows MUCH MUCH more about math than I do, or ever will, but I have been studying natural and artificial int…

DGX agent

Fields Medalist Jacob Tsimerman knows MUCH MUCH more about math than I do, or ever will, but I have been studying natural and artificial intelligence for 40 years, and I think his prediction here (“AI

safetygary-marcus--x
5 Aug 2026
Safety

From Routes to Steps: Separating Semantic Progress from Local Execution in Vision-and-Language Navigation

DGX agent

arXiv:2608.03143v1 Announce Type: new Abstract: Vision-and-Language Navigation (VLN) requires an agent to follow a route-level instruction by executing its constituent steps from egocentric visual obs

safetyarxiv-cs-cv
5 Aug 2026
Safety

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation

DGX agent

arXiv:2608.03753v1 Announce Type: new Abstract: Learning long-horizon manipulation skills with reinforcement learning remains challenging due to the complexity of reward design, the limited guidance o

safetyarxiv-cs-ro
5 Aug 2026
Safety

GROW: Group-Relative Advantage-Weighted On-Policy Reinforcement Learning of Autoregressive-Diffusion Text-to-Speech model

DGX agent

arXiv:2608.03215v1 Announce Type: cross Abstract: Reinforcement learning for flow-matching text-to-speech is complicated by deterministic ODE sampling: trajectory-level policy-gradient methods typical

safetyarxiv-cs-ai
5 Aug 2026
Safety

HERO: Hierarchical Evidential Reasoning Optimization for Radiology Report Generation via Reason-then-Summarize

DGX agent

arXiv:2601.03321v3 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) have substantially advanced Radiology Report Generation (RRG), yet aligning them through reinforcemen

safetyarxiv-cs-ai
5 Aug 2026
Safety

Hi-Token: Hierarchical Coordinate Tokenization for Generative Visual Grounding

DGX agent

arXiv:2608.03471v1 Announce Type: new Abstract: Generative Vision-Language Models (VLMs) commonly treat bounding-box coordinates as independent output symbols, leaving numerical order and axis semanti

safetyarxiv-cs-cv
5 Aug 2026
Safety

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning

DGX agent

arXiv:2608.03545v1 Announce Type: new Abstract: Test-time reinforcement learning (TTRL) improves the reasoning capabilities of large language models without labeled data by updating the policy with ps

safetyarxiv-cs-cl
5 Aug 2026
Safety

HomoEnsNER: Does Language Alignment Outperform Architectural Complexity in Gujarati Named Entity Recognition?

DGX agent

arXiv:2608.03105v1 Announce Type: new Abstract: Named Entity Recognition (NER) for Gujarati remains underexplored, hindered by the absence of capitalization cues, rich morphology, lexical ambiguity, a

safetyarxiv-cs-cl
5 Aug 2026
Safety

Implementing Causal Perception: Competing SCMs and Situated Fairness

DGX agent

arXiv:2608.03917v1 Announce Type: new Abstract: Causal perception occurs when agents with competing Structural Causal Models (SCMs) of the same system infer different probability distributions, includ

safetyarxiv-cs-ai
5 Aug 2026
Safety

Improved Quantum Algorithms for Reinforcement Learning Under a Generative Model

DGX agent

arXiv:2608.02826v1 Announce Type: cross Abstract: Reinforcement learning is a subfield of machine learning that studies how an agent interacts with an environment in order to extract as large a reward

safetyarxiv-cs-ai
5 Aug 2026
Safety

interesting (and completely opposed to @haider1’s take on the same graph) *the labs themselves* have only modestly changed their predictions…

DGX agent

interesting (and completely opposed to @haider1’s take on the same graph) *the labs themselves* have only modestly changed their predictions on AGI timelines over the last decade. note also that (on a

safetygary-marcus--x
5 Aug 2026
Safety

Interpretable Adaptive Sampling for LLM Test-Time Scaling

DGX agent

arXiv:2608.03961v1 Announce Type: new Abstract: Test-time scaling improves LLM reasoning by generating and aggregating multiple candidate answers, yet many pipelines use fixed per-query budgets that s

safetyarxiv-cs-ai
5 Aug 2026
Safety

is there a polite synonym for “circle jerk”?

DGX agent

The post contains two distinct snippets. First, user @GaryMarcus asks whether there is a more polite way to refer to “circle jerk.” Second, it shares a (likely satirical) claim that Microsoft’s AI rev

safetygary-marcus--x
5 Aug 2026
← Previous
1…7879808182…302
Next →