AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,600 results
Safety

Channel-wise Dynamic Knowledge Distillation via Adaptive Sample Generation for Action Recognition

DGX agent

arXiv:2608.03100v1 Announce Type: new Abstract: Knowledge Distillation (KD) offers a promising yet underexplored path for compressing large action recognition models. However, existing KD methods suff

safetyarxiv-cs-cv
5 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Clinically-Grounded Hierarchical Classification for Consistent Chest X-ray Interpretation

DGX agent

arXiv:2608.03016v1 Announce Type: new Abstract: Accurate chest X-ray interpretation is inherently hierarchical. Clinical decisions depend not only on what abnormality is present but where it is situat

safetyarxiv-cs-cv
5 Aug 2026
Safety

CLIP4VI-ReID: Learning Modality-shared Representations via CLIP Semantic Bridge for Visible-Infrared Person Re-identification

DGX agent

arXiv:2511.10309v2 Announce Type: replace Abstract: This paper proposes a novel CLIP-driven modality-shared representation learning network named CLIP4VI-ReID for VI-ReID task, which consists of Text

safetyarxiv-cs-cv
5 Aug 2026
Safety

⚠️⚠️⚠️ Constitutional AI is not working. Aligning LLMs is not working. We need a different approach. If society doesn’t place its bets diffe…

DGX agent

⚠️⚠️⚠️ Constitutional AI is not working. Aligning LLMs is not working. We need a different approach. If society doesn’t place its bets differently, we are screwed. Anthropic's Mythos created fake iden

safetygary-marcus--x
5 Aug 2026
Safety

Continue or Replan? Bernoulli-Continuation Policy Learning for Adaptive Horizon Execution

DGX agent

arXiv:2608.03483v1 Announce Type: cross Abstract: Existing chunk-based Vision-Language-Action (VLA) models execute a fixed number of actions (i.e., execution horizon) before replanning, turning replan

safetyarxiv-cs-ai
5 Aug 2026
Safety

Control Barrier Functions via Minkowski Operations for Safe Navigation among Polytopes

DGX agent

arXiv:2608.02886v1 Announce Type: new Abstract: Safely navigating polytopic environments while respecting the dynamics, control, and exact geometry of the underlying system is a challenge in robotics.

safetyarxiv-cs-ro
5 Aug 2026
Safety

CROSS: Cascaded Distillation and Dual-Constraint Grounding for Remote Sensing Referring Segmentation

DGX agent

arXiv:2608.03147v1 Announce Type: new Abstract: Referring Remote Sensing Image Segmentation (RRSIS) has achieved significant progress through the integration of VLMs and the Segment Anything Model (SA

safetyarxiv-cs-cv
5 Aug 2026
Safety

CT-HEG: A Bidirectional, Timestamp-Attributed Event Graph for ICU In-Hospital Mortality Prediction - An Architectural Ablation Study

DGX agent

arXiv:2608.02663v1 Announce Type: cross Abstract: Accurate ICU mortality prediction requires modeling irregular clinical observations across heterogeneous entity types. Existing sequence models handle

safetyarxiv-cs-ai
5 Aug 2026
Safety

CVPO: Enhancing LLM Reinforcement Learning Reasoning via Value-Variance Adaptation and Dynamic Curriculum Learning

DGX agent

arXiv:2608.03068v1 Announce Type: cross Abstract: Reinforcement learning (RL) has emerged as an effective method for enhancing the reasoning capabilities of large language models (LLMs). However, exis

safetyarxiv-cs-ai
5 Aug 2026
Safety

Deferred Exposure of Future Trajectories for Verifiable Reasoning in Autonomous Driving VLMs

DGX agent

arXiv:2608.01755v2 Announce Type: replace Abstract: Recent Vision-Language-Action (VLA) models for autonomous driving (AD) increasingly utilize chain-of-thought (CoT) supervision to enhance the reason

safetyarxiv-cs-ai
5 Aug 2026
Safety

Disentangling Language Modeling and Boundaries

DGX agent

arXiv:2608.03599v1 Announce Type: new Abstract: Byte-level language models are usually argued for on the grounds of robustness, multilingual fairness, and character-level skills. We point to a differe

safetyarxiv-cs-cl
5 Aug 2026
Safety

DiverseDiT++: Quantifying, Analyzing, and Promoting Representation Diversity in Diffusion Transformers

DGX agent

arXiv:2608.03082v1 Announce Type: new Abstract: Recent advances in Diffusion Transformers (DiTs) have enabled remarkable progress in visual synthesis, benefiting from their superior scalability. To fa

safetyarxiv-cs-cv
5 Aug 2026
Safety

DocTrace: Towards Traceable Long Document VQA via Hierarchical Evidence Graph Reasoning

DGX agent

arXiv:2608.03292v1 Announce Type: new Abstract: Long Document Visual Question Answering (LongDocVQA) requires Multimodal Large Language Models (MLLMs) to locate, integrate, and reason over heterogeneo

safetyarxiv-cs-ai
5 Aug 2026
Safety

Don't Peek at the Answer: Outcome-Masked Group Relative Policy Optimization for Label-Free RLVR

DGX agent

arXiv:2608.03119v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) improves LLM reasoning but typically relies on ground-truth (GT) answers, limiting scalability. Vo

safetyarxiv-cs-ai
5 Aug 2026
Safety

Double Down on Defense: Strengthening Deep Perceptual Hashes against Evasion Attacks without Retraining

DGX agent

arXiv:2608.03101v1 Announce Type: new Abstract: Near-duplicate image matching is crucial for trust and safety, provenance verification, copyright enforcement, and large-scale visual search. Modern pla

safetyarxiv-cs-cv
5 Aug 2026
Safety

DRIFT: Derailing Denoising Trajectories of Flow-Matching VLAs with Adversarial Patch Attack

DGX agent

arXiv:2608.03207v1 Announce Type: new Abstract: Flow-matching vision-language-action (VLA) models such as pi0 generate robot actions by integrating a learned denoising velocity field, and have been re

safetyarxiv-cs-cv
5 Aug 2026
Safety

DriftWorld: Fast World Modeling through Drifting

DGX agent

arXiv:2607.15065v2 Announce Type: replace-cross Abstract: Predictive world models enable robots to plan by imagining the outcomes of their actions, but their value for control hinges on generating man

safetyarxiv-cs-cv
5 Aug 2026
Safety

Emulate or Estimate? The Divergent Strengths of Base and Post-Trained Language Models for Opinion Simulation

DGX agent

arXiv:2608.03044v1 Announce Type: cross Abstract: Large language models are increasingly used to simulate human opinions, but prior work reports conflicting results: some studies find promising alignm

safetyarxiv-cs-ai
5 Aug 2026
Safety

Enhanced Polarization Locking in VCSELs

DGX agent

arXiv:2604.01857v2 Announce Type: replace-cross Abstract: While optical injection locking (OIL) of vertical-cavity surface-emitting lasers (VCSELs) has been widely studied in the past, the polarizatio

safetyarxiv-cs-cv
5 Aug 2026
Safety

Enhancing Q-Value Updates in Deep Q-Learning via Successor-State Prediction

DGX agent

arXiv:2511.03836v2 Announce Type: replace Abstract: Deep Q-Networks (DQNs) estimate future returns by learning from transitions sampled from a replay buffer. However, the target updates in DQN often r

safetyarxiv-cs-lg
5 Aug 2026
Safety

Enhancing VLM Reward Models Through Structure-Aware Fine-Tuning

DGX agent

arXiv:2608.03875v1 Announce Type: cross Abstract: Designing effective reward functions remains a major bottleneck in Reinforcement Learning (RL). Recent work uses large foundation Vision-Language Mode

safetyarxiv-cs-ai
5 Aug 2026
Safety

Evading Chain-of-Thought Monitoring Through Model Poisoning

DGX agent

arXiv:2608.02820v1 Announce Type: cross Abstract: Chain-of-thought (CoT) monitoring is an increasingly important component of AI safety stacks but relies on the assumption that a model's reasoning tra

safetyarxiv-cs-ai
5 Aug 2026
Safety

EvoHIL: Self-Evolving Reward and Flow-Matched Policy Optimization for Robust Human-in-the-Loop Reinforcement Learning

DGX agent

arXiv:2608.03872v1 Announce Type: new Abstract: Human-in-the-loop reinforcement learning (HIL-RL) enables robots to learn contact-rich manipulation from limited real-world interaction, but deployment

safetyarxiv-cs-ro
5 Aug 2026
Safety

Fast Object Removal Attacks on Safety-Critical Video-based Perception Systems

DGX agent

arXiv:2608.02806v1 Announce Type: cross Abstract: By leveraging data from video-based perception systems, intelligent transportation systems (ITS) support safety-critical applications that improve roa

safetyarxiv-cs-cv
5 Aug 2026
Safety

Fields Medalist Jacob Tsimerman knows MUCH MUCH more about math than I do, or ever will, but I have been studying natural and artificial int…

DGX agent

Fields Medalist Jacob Tsimerman knows MUCH MUCH more about math than I do, or ever will, but I have been studying natural and artificial intelligence for 40 years, and I think his prediction here (“AI

safetygary-marcus--x
5 Aug 2026
Safety

Flying over The Uncertain Nature (FORTUNE): Intelligent and Humanistic 3D Path Planning for Low-Altitude Collaboration

DGX agent

arXiv:2608.03408v1 Announce Type: new Abstract: The proliferation of low-altitude intelligent agents is increasing the demand for timely and socially responsible collaborative sensing in dynamic urban

safetyarxiv-cs-ro
5 Aug 2026
Safety

Forbidden Region Dynamic Active Constraints in Robot-Assisted Minimally Invasive Surgery

DGX agent

arXiv:2608.03010v1 Announce Type: new Abstract: In robot-assisted surgery, Forbidden Region Active Constraints (FRAC) represent a control strategy that helps maintain task safety by generating anisotr

safetyarxiv-cs-ro
5 Aug 2026
Safety

From Routes to Steps: Separating Semantic Progress from Local Execution in Vision-and-Language Navigation

DGX agent

arXiv:2608.03143v1 Announce Type: new Abstract: Vision-and-Language Navigation (VLN) requires an agent to follow a route-level instruction by executing its constituent steps from egocentric visual obs

safetyarxiv-cs-cv
5 Aug 2026
Safety

GenOS: Compositional Certificates for Semantic Robustness in AI Code Generation

DGX agent

arXiv:2608.03588v1 Announce Type: cross Abstract: AI coding agents are stochastic workflows: prompts are interpreted, artifacts are sampled, validators produce observations, and orchestrators commit o

safetyarxiv-cs-ai
5 Aug 2026
Safety

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation

DGX agent

arXiv:2608.03753v1 Announce Type: new Abstract: Learning long-horizon manipulation skills with reinforcement learning remains challenging due to the complexity of reward design, the limited guidance o

safetyarxiv-cs-ro
5 Aug 2026
Safety

GROW: Group-Relative Advantage-Weighted On-Policy Reinforcement Learning of Autoregressive-Diffusion Text-to-Speech model

DGX agent

arXiv:2608.03215v1 Announce Type: cross Abstract: Reinforcement learning for flow-matching text-to-speech is complicated by deterministic ODE sampling: trajectory-level policy-gradient methods typical

safetyarxiv-cs-ai
5 Aug 2026
Safety

HERO: Hierarchical Evidential Reasoning Optimization for Radiology Report Generation via Reason-then-Summarize

DGX agent

arXiv:2601.03321v3 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) have substantially advanced Radiology Report Generation (RRG), yet aligning them through reinforcemen

safetyarxiv-cs-ai
5 Aug 2026
Safety

Hi-Token: Hierarchical Coordinate Tokenization for Generative Visual Grounding

DGX agent

arXiv:2608.03471v1 Announce Type: new Abstract: Generative Vision-Language Models (VLMs) commonly treat bounding-box coordinates as independent output symbols, leaving numerical order and axis semanti

safetyarxiv-cs-cv
5 Aug 2026
Safety

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning

DGX agent

arXiv:2608.03545v1 Announce Type: new Abstract: Test-time reinforcement learning (TTRL) improves the reasoning capabilities of large language models without labeled data by updating the policy with ps

safetyarxiv-cs-cl
5 Aug 2026
Safety

HomoEnsNER: Does Language Alignment Outperform Architectural Complexity in Gujarati Named Entity Recognition?

DGX agent

arXiv:2608.03105v1 Announce Type: new Abstract: Named Entity Recognition (NER) for Gujarati remains underexplored, hindered by the absence of capitalization cues, rich morphology, lexical ambiguity, a

safetyarxiv-cs-cl
5 Aug 2026
Safety

ICO: Enhancing Semantic-Shift Jailbreaks via Iterative Context Optimization

DGX agent

arXiv:2608.03210v1 Announce Type: new Abstract: Foundation models have achieved remarkable success across diverse tasks, but they remain vulnerable. To investigate such vulnerabilities, semantic-shift

safetyarxiv-cs-cl
5 Aug 2026
Safety

Implementing Causal Perception: Competing SCMs and Situated Fairness

DGX agent

arXiv:2608.03917v1 Announce Type: new Abstract: Causal perception occurs when agents with competing Structural Causal Models (SCMs) of the same system infer different probability distributions, includ

safetyarxiv-cs-ai
5 Aug 2026
Safety

Improved Quantum Algorithms for Reinforcement Learning Under a Generative Model

DGX agent

arXiv:2608.02826v1 Announce Type: cross Abstract: Reinforcement learning is a subfield of machine learning that studies how an agent interacts with an environment in order to extract as large a reward

safetyarxiv-cs-ai
5 Aug 2026
Safety

interesting (and completely opposed to @haider1’s take on the same graph) *the labs themselves* have only modestly changed their predictions…

DGX agent

interesting (and completely opposed to @haider1’s take on the same graph) *the labs themselves* have only modestly changed their predictions on AGI timelines over the last decade. note also that (on a

safetygary-marcus--x
5 Aug 2026
Safety

Interpretable Adaptive Sampling for LLM Test-Time Scaling

DGX agent

arXiv:2608.03961v1 Announce Type: new Abstract: Test-time scaling improves LLM reasoning by generating and aggregating multiple candidate answers, yet many pipelines use fixed per-query budgets that s

safetyarxiv-cs-ai
5 Aug 2026
Safety

is there a polite synonym for “circle jerk”?

DGX agent

The post contains two distinct snippets. First, user @GaryMarcus asks whether there is a more polite way to refer to “circle jerk.” Second, it shares a (likely satirical) claim that Microsoft’s AI rev

safetygary-marcus--x
5 Aug 2026
Safety

Joint Affine Spectral Shaping: Coupling Weight and Bias Updates Beyond Weight-Only Muon

DGX agent

arXiv:2608.02991v1 Announce Type: new Abstract: Matrix spectral optimizers reshape weight-update spectra but usually delegate vector-valued biases to a separate optimizer. We study whether this separa

safetyarxiv-cs-lg
5 Aug 2026
Safety

Just had to create an 'accidental-cyberattacks' tag on my blog We're up to four now: the original OpenAI+Hugging Face one, Anthropic's me-to…

DGX agent

Just had to create an 'accidental-cyberattacks' tag on my blog We're up to four now: the original OpenAI+Hugging Face one, Anthropic's me-too attacks, then two new ones from the UK AI Safety Institute

safetysimon-willison--x
5 Aug 2026
Safety

Kernel weighted importance sampling for off-policy evaluation in contextual bandits

DGX agent

arXiv:2607.15067v2 Announce Type: replace Abstract: This article presents a novel estimator for performing off-policy evaluation using only offline data for contextual bandits. The proposed estimator,

safetyarxiv-cs-lg
5 Aug 2026
Safety

KernelBrain: Coarse-to-Fine, Budget-Aware Search for Agentic GPU Kernel Optimization

DGX agent

arXiv:2608.02611v1 Announce Type: cross Abstract: Automating GPU kernel optimization remains difficult in practice: generated variants can violate correctness constraints, runtime measurements are noi

safetyarxiv-cs-ai
5 Aug 2026
Safety

Language-Specialized Multi-Teacher On-Policy Distillation for Multilingual LLM-Based ASR

DGX agent

arXiv:2608.03610v1 Announce Type: new Abstract: Modern LLM-based ASR systems have established multilingual capability as a standard feature, leveraging large-scale multilingual corpora and LLMs' cross

safetyarxiv-cs-cl
5 Aug 2026
Safety

LatentGuard: Efficient and Inspectable Latent Reasoning for LLM Safeguards

DGX agent

arXiv:2608.03838v1 Announce Type: new Abstract: Reasoning-based guard models improve LLM safeguards, but decoding explicit rationales for every interaction makes them costly to deploy. Although latent

safetyarxiv-cs-ai
5 Aug 2026
Safety

Learning Attribute-aware Representations for Few-shot Scene Text Segmentation

DGX agent

arXiv:2504.11164v2 Announce Type: replace Abstract: Supervised scene text segmentation has achieved notable progress in recent years. However, its development is largely constrained by the scarcity of

safetyarxiv-cs-cv
5 Aug 2026
← Previous
1…1314151617…263
Next →