AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
All
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,809 results
Safety

Learning to Fold: prizewinning solution at LeHome Challenge 2026 (1st place online, 2nd offline)

DGX agent

arXiv:2606.27163v1 Announce Type: cross Abstract: I describe my solution to the LeHome Challenge 2026, an ICRA 2026 competition on bimanual garment folding. The system placed 1st of 62 teams in the on

safetyarxiv-cs-ai
26 Jun 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

Linguistics and Human Brain: A Perspective of Computational Neuroscience

DGX agent

arXiv:2602.08275v3 Announce Type: replace-cross Abstract: Elucidating the language-brain relationship requires bridging the methodological gap between the abstract theoretical frameworks of linguistic

safetyarxiv-cs-cl
26 Jun 2026
Safety

LISA: Likelihood Score Alignment for Visual-condition Controllable Generation

DGX agent

arXiv:2606.27192v1 Announce Type: new Abstract: The prevalent dual-branch paradigm, i.e., training a side network to encode visual conditions and fusing its intermediate-layer features to a frozen pre

safetyarxiv-cs-cv
26 Jun 2026
Safety

Listening Like a Judge: A Music-Aware Framework for Automatic Singing Performance Evaluation

DGX agent

arXiv:2606.26451v1 Announce Type: cross Abstract: Automatic singing quality assessment (SQA) requires evaluating lyrical correctness and musical fidelity while handling expressive variations. However,

safetyarxiv-cs-lg
26 Jun 2026
Safety

LLM-based Models for Detecting Emerging Topics in Service Feedback

DGX agent

arXiv:2606.26595v1 Announce Type: new Abstract: Enhancing the analysis of service feedback is essential for public sector organizations, particularly tax administrations, where trust and compliance de

safetyarxiv-cs-ai
26 Jun 2026
Safety

Metaphors are a Source of Cross-Domain Misalignment of Large Reasoning Models

DGX agent

arXiv:2601.03388v3 Announce Type: replace-cross Abstract: Earlier research has shown that metaphors influence human decision-making, raising the question of whether metaphors also influence large lang

safetyarxiv-cs-ai
26 Jun 2026
Safety

MinGram: A Minimalist Unigram Tokenizer with High Compression and Competitive Morphological Alignment

DGX agent

arXiv:2606.27019v1 Announce Type: new Abstract: The Unigram tokenizer uses an elegant representation which makes it straightforward to edit vocabularies, but its training is comparatively heavy and co

safetyarxiv-cs-cl
26 Jun 2026
Safety

NASimJax: A GPU-Accelerated Policy Learning Framework for Penetration Testing

DGX agent

arXiv:2603.19864v2 Announce Type: replace Abstract: Penetration testing, the practice of simulating cyberattacks to identify vulnerabilities, is a complex sequential decision-making task that is inher

safetyarxiv-cs-lg
26 Jun 2026
Safety

NaviCache: Test-Time Self-Calibration Caching for Video Generation

DGX agent

arXiv:2606.26795v1 Announce Type: cross Abstract: Video Diffusion Models (VDMs) is constrained by immense computational costs. While offline calibration-based acceleration suffers from calibration dat

safetyarxiv-cs-ai
26 Jun 2026
Safety

Neural Speaker Diarization via Multilingual Training: Evaluation on Low-Resource Nepali-Hindi Speech

DGX agent

arXiv:2606.26144v1 Announce Type: cross Abstract: Speaker diarization, the task of determining 'who spoke when' in a multi-speaker recording, is a critical component in applications such as meeting tr

safetyarxiv-cs-cl
26 Jun 2026
Safety

Normalizing Flows are Capable Models for Continuous Control

DGX agent

arXiv:2505.23527v4 Announce Type: replace Abstract: Modern reinforcement learning (RL) algorithms have found success by using powerful probabilistic models, such as transformers, energy-based models,

safetyarxiv-cs-lg
26 Jun 2026
Safety

OmniContact: Chaining Meta-Skills via Contact Flow for Generalizable Humanoid Loco-Manipulation

DGX agent

arXiv:2606.26201v1 Announce Type: new Abstract: Learning long-horizon humanoid loco-manipulation poses a dual challenge: it requires not only the robust execution of meta-skills but also their seamles

safetyarxiv-cs-ro
26 Jun 2026
Safety

OPID: On-Policy Skill Distillation for Agentic Reinforcement Learning

DGX agent

arXiv:2606.26790v1 Announce Type: new Abstract: Outcome-based reinforcement learning provides a stable optimization backbone for language agents, but its sparse trajectory-level rewards provide little

safetyarxiv-cs-cl
26 Jun 2026
Safety

Ordinal Neural Collapse as a Representation Prior for Visual Navigation

DGX agent

arXiv:2606.26839v1 Announce Type: cross Abstract: Learning robust navigation policies directly from visual observations remains a fundamental challenge in vision-based robotic navigation. In end-to-en

safetyarxiv-cs-cv
26 Jun 2026
Safety

orwellian bullshit

DGX agent

This post likely discusses how Orwellian tactics—such as propaganda, doublespeak, and manipulation of language—are being employed in contemporary contexts, possibly relating to AI, technology, or poli

safetygary-marcus--x
26 Jun 2026
Safety

Overcoming State Inertia: Minimally Invasive Temporal Alignment for Evolving Contexts

DGX agent

arXiv:2512.03704v3 Announce Type: replace Abstract: Long-context dialogue systems suffer from state inertia, where models over-attend to history and fail to adapt to evolving intents. We demonstrate t

safetyarxiv-cs-cl
26 Jun 2026
Safety

PAMAE: Phase-Aware-MoE Action Experts Towards Reliable Flow-Matching Vision-Language-Action Policies

DGX agent

arXiv:2606.27144v1 Announce Type: new Abstract: Reliable action generation for multi-stage robotic manipulation remains challenging for Vision-Language-Action (VLA) models. While existing flow-matchin

safetyarxiv-cs-ro
26 Jun 2026
Safety

PathFLIP: Fine-grained Language-Image Pretraining for Versatile Computational Pathology

DGX agent

arXiv:2512.17621v2 Announce Type: replace Abstract: While Vision-Language Models (VLMs) have achieved notable progress in computational pathology (CPath), the gigapixel scale and spatial heterogeneity

safetyarxiv-cs-cv
26 Jun 2026
Safety

Paved with True Intents: Intent-Aware Training Improves LLM Safety Classification Across Training Regimes

DGX agent

arXiv:2606.27210v1 Announce Type: new Abstract: We argue that safety classifiers should model user intent as an explicit signal between the prompt and the final label. To study this, we introduce AIMS

safetyarxiv-cs-cl
26 Jun 2026
Safety

Paying More Attention to Visual Tokens in Self-Evolving Large Multimodal Models

DGX agent

arXiv:2606.27373v1 Announce Type: new Abstract: Recently, self-evolving large multimodal models (LMMs) have received attention for improving visual reasoning in a purely unsupervised setting. However,

safetyarxiv-cs-cv
26 Jun 2026
Safety

People want jobs, not subsidies:

DGX agent

People want jobs, not subsidies: Taxing corporations to fund programs is usually controversial and partisan - but when it comes to AI job displacement voters are shockingly left wing. It's genuinely e

safetygary-marcus--x
26 Jun 2026
Safety

people with advanced degrees who can’t distinguish between pure LLMs (which is what I critiqued in 2022) and LLMs enhanced with neurosymboli…

DGX agent

people with advanced degrees who can’t distinguish between pure LLMs (which is what I critiqued in 2022) and LLMs enhanced with neurosymbolic techniques (which is what I championed in 2022) disappoint

safetygary-marcus--x
26 Jun 2026
Safety

PhysEditWorld: A Large-Scale Dataset Toward Physics-Editable World Models

DGX agent

arXiv:2606.26694v1 Announce Type: new Abstract: Recent game world models can synthesize visually plausible, action-conditioned rollouts. However, their interaction behaviors often remain limited to ex

safetyarxiv-cs-cv
26 Jun 2026
Safety

PlanRL: A Trajectory Planning Architecture for Reinforcement Learning-based Driving Experts

DGX agent

arXiv:2606.26858v1 Announce Type: new Abstract: Reinforcement learning (RL) has become a prominent framework for developing driving experts in autonomous vehicles. However, most existing RL-based expe

safetyarxiv-cs-ro
26 Jun 2026
Safety

Position: Align AI to Our Aspirations, Not Our Flaws

DGX agent

arXiv:2606.13755v2 Announce Type: replace-cross Abstract: We argue that aligning AI to aggregated human preferences is the wrong target. With current technology, one can train AIs to share the values

safetyarxiv-cs-ai
26 Jun 2026
Safety

PressMimic: Pressure-Guided Motion Capture and Control for Humanoid Robot Imitation

DGX agent

arXiv:2606.26741v1 Announce Type: cross Abstract: Humanoid motion imitation requires not only accurate perception of human kinematics but also faithful reproduction of physical interactions with the e

safetyarxiv-cs-cv
26 Jun 2026
Safety

Prompt Injection in Automated Resume Screening with Large Language Models: Single and Multi-Injection Settings

DGX agent

arXiv:2606.27287v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used to screen and rank job applicants, creating incentives for candidates to strategically manipulate alg

safetyarxiv-cs-ai
26 Jun 2026
Safety

Proposal-Conditioned Latent Diffusion for Closed-Loop Traffic Scenario Generation

DGX agent

arXiv:2606.27123v1 Announce Type: cross Abstract: Closed-loop traffic simulation remains challenging because it must generate interactive multi-agent behaviors that are scene-consistent and controllab

safetyarxiv-cs-cv
26 Jun 2026
Safety

Racing a Wheeled Quadruped: Active Load Transfer Mitigation via Model Predictive Control

DGX agent

arXiv:2606.26313v1 Announce Type: new Abstract: This paper presents a hierarchical control framework using model predictive control (MPC) and reinforcement learning (RL) for active roll control to man

safetyarxiv-cs-ro
26 Jun 2026
Safety

Radical AI Interpretability

DGX agent

arXiv:2606.26523v1 Announce Type: new Abstract: We develop a framework for interpreting AI systems as agents, drawing on the philosophical tradition of radical interpretation and the tools of mechanis

safetyarxiv-cs-ai
26 Jun 2026
Safety

RecallRisk-BERT: A Multi-Task Framework for Post-Report Medical Device Recall Triage

DGX agent

arXiv:2606.27174v1 Announce Type: new Abstract: Medical device recalls are a critical regulatory mechanism for protecting patient safety. The growing volume of FDA recall records presents challenges i

safetyarxiv-cs-lg
26 Jun 2026
Safety

Reconstruction Alignment Improves Unified Multimodal Models

DGX agent

arXiv:2509.07295v4 Announce Type: replace-cross Abstract: Unified multimodal models (UMMs) unify visual understanding and generation within a single architecture. However, conventional training relies

safetyarxiv-cs-ai
26 Jun 2026
Safety

Reducing Conversational Escalation in Large Language Model Dialogue with Nonviolent Communication Constraints

DGX agent

arXiv:2606.26106v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used in emotionally charged situations involving interpersonal conflict, frustration, and distress. Whil

safetyarxiv-cs-ai
26 Jun 2026
Safety

remember my botched bicycle examples? here’s an actual REI ad, h/t Oren Etzioni:

DGX agent

Gary Marcus shares an actual REI advertisement as a follow-up to previous discussion about flawed bicycle examples, crediting Oren Etzioni for the reference. The post appears to contrast a real-world

safetygary-marcus--x
26 Jun 2026
Safety

Residual RL-MPC for Robust Microrobotic Cell Pushing Under Time-Varying Flow

DGX agent

arXiv:2603.05448v2 Announce Type: replace-cross Abstract: Contact-rich micromanipulation in microfluidic flow is challenging because small disturbances can break pushing contact and induce large later

safetyarxiv-cs-ai
26 Jun 2026
Safety

Retrieval-Warmed Energy-Based Reasoning: A Five-Arm Ablation Methodology for Diffusion-as-Inference on Structured Reasoning Tasks

DGX agent

arXiv:2606.26476v1 Announce Type: cross Abstract: Warm-started diffusion samplers accelerate iterative inference, but it is rarely clear which part of the pipeline carries the gain. We study extbf{ret

safetyarxiv-cs-ai
26 Jun 2026
Safety

Risk-Aware Selective Multimodal Driver Monitoring with Driver-State World Modeling

DGX agent

arXiv:2606.26922v1 Announce Type: cross Abstract: Continuous driver monitoring in automated vehicles requires low-latency inference while avoiding unsafe decisions under uncertain driver states. Large

safetyarxiv-cs-ai
26 Jun 2026
Safety

RMTL: Reinforced Micro-task Learning for Long-Horizon Manipulation with VLM Rewards

DGX agent

arXiv:2606.26175v1 Announce Type: new Abstract: Reinforcement learning (RL) for robotic manipulation often requires manually designing a dense reward function, which is difficult to tune and often fra

safetyarxiv-cs-ro
26 Jun 2026
Safety

RobOralScan: Learning Active Intraoral Scanning for Robotic Dental Reconstruction

DGX agent

arXiv:2606.26955v1 Announce Type: new Abstract: Intraoral scanning is widely used for digital optical impressions in prosthodontic, implant, and orthodontic treatment, but full-arch and long-span scan

safetyarxiv-cs-ro
26 Jun 2026
Safety

RolloutPipe: Overlapping Pipelined Rollout and Training in Disaggregated On-Policy LLM Reinforcement Learning

DGX agent

arXiv:2606.26997v1 Announce Type: cross Abstract: Large language model (LLM) post-training for reasoning increasingly relies on reinforcement learning with verifiable rewards (RLVR), where models lear

safetyarxiv-cs-lg
26 Jun 2026
Safety

Rotary Position Encodings for Graphs

DGX agent

arXiv:2509.22259v4 Announce Type: replace-cross Abstract: We study the extent to which rotary position encodings (RoPE), a recent transformer position encoding algorithm broadly adopted in large langu

safetyarxiv-cs-ai
26 Jun 2026
Safety

RouterVLA: Turning Smoke Tests into Supervision for Heterogeneous VLA Selection

DGX agent

arXiv:2606.27355v1 Announce Type: new Abstract: We study whether pre-deployment evaluation rollouts can be reused to supervise policy selection. Robot teams routinely smoke test candidate vision-langu

safetyarxiv-cs-ro
26 Jun 2026
Safety

Safe Autoregressive Image Generation with Iterative Self-Improving Codebooks

DGX agent

arXiv:2606.27147v1 Announce Type: cross Abstract: Unlike diffusion-based models that operate in continuous latent spaces, autoregressive unified multimodal models produce images by sequentially predic

safetyarxiv-cs-ai
26 Jun 2026
Safety

Sample-efficient Transfer Reinforcement Learning via Adaptive Reward Shaping and Policy-Ratio Reweighting Strategy

DGX agent

arXiv:2606.26527v1 Announce Type: new Abstract: Transfer learning improves policy learning efficiency by reusing knowledge from source tasks, providing a feasible paradigm for safe and efficient auton

safetyarxiv-cs-lg
26 Jun 2026
Safety

Scale Robot Policy Evaluation with Ray

DGX agent

This article discusses using Ray and Anyscale's distributed computing platform to scale the evaluation of robot control policies across multiple simulations in parallel. It likely covers techniques fo

safetyanyscale-ray
26 Jun 2026
Safety

Scoring Is Not Enough: Addressing Gaps in Utility-fairness Trade-offs for Ranking

DGX agent

arXiv:2606.26369v1 Announce Type: cross Abstract: Scoring functions are used to represent the relevance of individual documents. In modern information retrieval or recommendation systems, they are oft

safetyarxiv-cs-lg
26 Jun 2026
Safety

Semantic Early-Stopping for Iterative LLM Agent Loops

DGX agent

arXiv:2606.27009v1 Announce Type: new Abstract: Multi-agent large language model (LLM) loops, for example a Writer that drafts and a Critic that revises, are almost always terminated by a fixed iterat

safetyarxiv-cs-ai
26 Jun 2026
Safety

Sketched Linear Contrastive Learning: Approximation, Optimization, and Statistical Scaling

DGX agent

arXiv:2606.26617v1 Announce Type: new Abstract: Scaling laws describe how learning performance varies with model size, data size, and compute. While recent theoretical work has established scaling law

safetyarxiv-cs-lg
26 Jun 2026
← Previous
1…7778798081…267
Next →