AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
All
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,813 results
Safety

CineDance: Towards Next-Generation Multi-Shot Long-Form Cinematic Audio-Video Generation

DGX agent

arXiv:2606.09639v1 Announce Type: new Abstract: The fidelity and structural diversity of training datasets fundamentally determine the capabilities of video generation models. While commercial systems

safetyarxiv-cs-cv
9 Jun 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

Claw-R1: A Step-Level Data Middleware System for Agentic Reinforcement Learning

DGX agent

arXiv:2606.09138v1 Announce Type: new Abstract: Agentic reinforcement learning (RL) has become an important post-training paradigm for turning LLMs from static chatbots into interactive agents, giving

safetyarxiv-cs-lg
9 Jun 2026
Safety

CLPO: Curriculum Learning meets Policy Optimization for LLM Reasoning

DGX agent

arXiv:2509.25004v2 Announce Type: replace Abstract: Online reinforcement learning with verifiable rewards (RLVR) has become an effective paradigm for improving the reasoning abilities of large languag

safetyarxiv-cs-ai
9 Jun 2026
Safety

Code Is More Than Text: Uncertainty Estimation for Code Generation

DGX agent

arXiv:2606.09577v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed as code generators, where silently wrong programs pose real safety and reliability risks. Relia

safetyarxiv-cs-lg
9 Jun 2026
Safety

Cold feet about coating the entire surface of the earth in data centers?

DGX agent

Gary Marcus likely discusses concerns about the environmental and practical implications of exponentially expanding data center infrastructure across the globe, questioning whether covering Earth's su

safetygary-marcus--x
9 Jun 2026
Safety

Comparative evaluation of training strategies using partially labelled datasets for segmentation of white matter hyperintensities and stroke lesions in FLAIR MRI

DGX agent

arXiv:2601.20503v2 Announce Type: replace-cross Abstract: White matter hyperintensities (WMH) and ischaemic stroke lesions (ISL) are key imaging biomarkers of cerebral small vessel disease (SVD) detec

safetyarxiv-cs-ai
9 Jun 2026
Safety

ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning

DGX agent

arXiv:2606.08088v1 Announce Type: new Abstract: Reinforcement Learning from Verifiable Rewards (RLVR) has recently become a key paradigm for improving the reasoning abilities of Large Language Models

safetyarxiv-cs-lg
9 Jun 2026
Safety

Constrained Paraphrase Consistency for LLM Hallucination Detection

DGX agent

arXiv:2606.08158v1 Announce Type: cross Abstract: Large language models (LLMs) can generate factually inconsistent claims, motivating accurate and scalable hallucination detectors. Prior work largely

safetyarxiv-cs-ai
9 Jun 2026
Safety

Constrained user-item allocation for e-commerce marketing campaigns

DGX agent

arXiv:2606.09623v1 Announce Type: new Abstract: When running marketing campaigns, retailers must decide which products to promote and which users to target. These decisions are inherently coupled: eff

safetyarxiv-cs-lg
9 Jun 2026
Safety

Constraint-Aware Optimization for Robust Protein Stability Prediction

DGX agent

arXiv:2606.08100v1 Announce Type: new Abstract: Multimodal DeltaDelta G predictors integrating protein language models with inverse-folding representations achieve strong in-distribution accuracy on t

safetyarxiv-cs-lg
9 Jun 2026
Safety

Context-Fractured Decomposition Attacks on Tool-Using LLM Agents: Exploiting Artifact Provenance Gaps

DGX agent

arXiv:2606.09084v1 Announce Type: cross Abstract: Tool-using LLM agents interact with the world through actions that persist state in artifacts (e.g., workspace files or logs). Consequently, jailbreak

safetyarxiv-cs-ai
9 Jun 2026
Safety

Context Over Compute Human-in-the-Loop Outperforms Iterative Chain-of-Thought Prompting in Interview Answer Quality

DGX agent

arXiv:2603.09995v2 Announce Type: replace-cross Abstract: Behavioral interview evaluation using large language models presents unique challenges that require structured assessment, realistic interview

safetyarxiv-cs-ai
9 Jun 2026
Safety

Contrast encodes inductive bias: separating slow noise from dynamics in predictive representation learning

DGX agent

arXiv:2606.07770v1 Announce Type: new Abstract: Self-supervised methods that learn representations and predict dynamics fully in the latent space, such as JEPA, have been shown to confuse slowly varyi

safetyarxiv-cs-lg
9 Jun 2026
Safety

Contribution Weights: A Geometrical Analysis of Self-Attention Transformers

DGX agent

arXiv:2606.07604v1 Announce Type: cross Abstract: Analyzing attention weights has become a standard approach for interpreting the information flow of Large Language Models (LLMs). However, this approa

safetyarxiv-cs-ai
9 Jun 2026
Safety

Cooperative Long Rope Skipping via Multi-Agent Reinforcement Learning

DGX agent

arXiv:2606.08064v1 Announce Type: new Abstract: Humans exhibit remarkable motor agility, enabling a wide range of dynamic skills such as running and jumping, which highlights the great potential of hu

safetyarxiv-cs-ro
9 Jun 2026
Safety

Correct Looks Better: Pairwise Comparisons Reveal Accuracy Rankings

DGX agent

arXiv:2606.09409v1 Announce Type: new Abstract: Pairwise comparisons combined with aggregation methods like Elo have become central to evaluating generative models, yet concerns remain that they rewar

safetyarxiv-cs-ai
9 Jun 2026
Safety

Cranio-Diff: Diffusion-based Cross-domain Craniofacial Reconstruction with 2D X-ray Skull Guidance and Structural Identity Constraints

DGX agent

arXiv:2606.09699v1 Announce Type: new Abstract: The state-of-the-art generative models, such as CycleGAN, Pix2Pix, and diffusion models have demonstrated remarkable performance in the face generation

safetyarxiv-cs-cv
9 Jun 2026
Safety

Crayotter: Traceable Multi-Agent Workflows for Long-Form Video Editing

DGX agent

arXiv:2606.07636v1 Announce Type: new Abstract: Editing a long-form video from heterogeneous footage requires more than selecting clips: an agent must preserve narrative intent across material prepara

safetyarxiv-cs-cv
9 Jun 2026
Safety

Culturally-Adapted Red-Teaming Across East and Southeast Asian Contexts: A Methodological and Comparative Analysis

DGX agent

arXiv:2606.09178v1 Announce Type: cross Abstract: Multilingual safety evaluation of large language models (LLMs) has predominantly relied on direct translation (DT) of English benchmarks into target l

safetyarxiv-cs-ai
9 Jun 2026
Safety

CURE: Curriculum-guided Multi-task Training for Reliable Anatomy Grounded Report Generation

DGX agent

arXiv:2601.15408v2 Announce Type: replace-cross Abstract: Medical vision-language models can automate the generation of radiology reports but struggle with accurate visual grounding and factual consis

safetyarxiv-cs-ai
9 Jun 2026
Safety

Data Agents Under Attack: Vulnerabilities in LLM-Driven Analytical Systems

DGX agent

arXiv:2606.08661v1 Announce Type: cross Abstract: Data agents integrate LLM-driven reasoning with relational data access, executable analytical tools, and multi-step workflow orchestration, making the

safetyarxiv-cs-ai
9 Jun 2026
Safety

Decentralized End-to-End Multi-AAV Pursuit Using Predictive Spatio-Temporal Observation via Deep Reinforcement Learning

DGX agent

arXiv:2603.24238v2 Announce Type: replace Abstract: Decentralized cooperative pursuit in cluttered environments is challenging for autonomous aerial swarms, especially under partial and noisy percepti

safetyarxiv-cs-ro
9 Jun 2026
Safety

Decoupling Semantics and Logic: A Training-Free Coarse-to-Fine Pipeline for Video Retrieval-Augmented Generation

DGX agent

arXiv:2606.07924v1 Announce Type: cross Abstract: This paper presents our system description for the 2nd Workshop on Multimodal Augmented Generation via MultimodAl Retrieval (MAGMaR). Addressing the c

safetyarxiv-cs-ai
9 Jun 2026
Safety

DexPIE: Stable Dexterous Policy Improvement from Real-World Experience

DGX agent

arXiv:2606.09615v1 Announce Type: cross Abstract: Dexterous manipulation presents substantial challenges for imitation learning due to its high-dimensional action space and complex contact-rich dynami

safetyarxiv-cs-cv
9 Jun 2026
Safety

Diffuse AI Control on Fuzzy Tasks

DGX agent

arXiv:2606.08892v1 Announce Type: new Abstract: AI models deployed in critical domains, such as AI safety research, may subtly sabotage our efforts due to misalignment. Diffuse AI Control is a subfiel

safetyarxiv-cs-lg
9 Jun 2026
Safety

Disentanglement with Holographic Reduced Representations

DGX agent

arXiv:2606.09725v1 Announce Type: new Abstract: Disentanglement, the separation of factors of variation in data using neural networks, remains a long-standing challenge in machine learning. Prior work

safetyarxiv-cs-lg
9 Jun 2026
Safety

Distant Object Localisation from Noisy Image Segmentation Sequences

DGX agent

arXiv:2509.20906v3 Announce Type: replace Abstract: 3D object localisation based on a sequence of camera measurements is essential for safety-critical surveillance tasks, such as drone-based wildfire

safetyarxiv-cs-cv
9 Jun 2026
Safety

Distilling LLM Reasoning into an Interpretable Policy Tree for Human-AI Collaboration

DGX agent

arXiv:2606.08596v1 Announce Type: new Abstract: Constructing efficient and reliable policies to assist humans is indispensable for human-AI collaboration. Existing methods mainly follow two lines of w

safetyarxiv-cs-ai
9 Jun 2026
Safety

DIVERGE: Diversity-Enhanced RAG for Open-Ended Information Seeking

DGX agent

arXiv:2602.00238v2 Announce Type: replace-cross Abstract: Existing retrieval-augmented generation (RAG) systems often assume that each query has a single correct answer. This assumption overlooks open

safetyarxiv-cs-ai
9 Jun 2026
Safety

Diverse Thinking Schemata Elicit Better Reasoning in Large Language Models

DGX agent

arXiv:2606.08974v1 Announce Type: new Abstract: Large reasoning models (LRMs) have attracted increasing attention for their ability to solve complex mathematical problems by generating extended reason

safetyarxiv-cs-ai
9 Jun 2026
Safety

Do VLMs See What Sensors Feel? A Scalable Expert-Guided Design for Wheelchair Accessibility Assessment from Street View

DGX agent

arXiv:2606.07642v1 Announce Type: new Abstract: Assessing built-environment interaction, such as wheelchair accessibility, is difficult because real-world mobility is shaped by distributed, context-de

safetyarxiv-cs-cv
9 Jun 2026
Safety

Does Persona Make LLMs K-pop Fans? A Pilot Study of LLM-Based Online Concert Audience Agents

DGX agent

arXiv:2606.07837v1 Announce Type: cross Abstract: A concert is a collective experience, but recorded performance videos are typically watched alone, stripping away the shared audience presence that ma

safetyarxiv-cs-ai
9 Jun 2026
Safety

DOG-DPO:Dynamic Optimization in Geometry for Safety Alignment

DGX agent

arXiv:2606.07678v1 Announce Type: cross Abstract: Safety alignment for large language models relies on preference data, but current pipelines often train on large, redundant datasets. Existing data se

safetyarxiv-cs-ai
9 Jun 2026
Safety

@dpetrou @karpathy Yes. Locking in a permanent status quo power structure. Incredibly unsafe, and damaging for humanity's prospects.

DGX agent

Gary Marcus argues that establishing a permanent, locked-in power structure is fundamentally unsafe and harmful to humanity's long-term prospects. The statement appears to be part of a discussion with

safetygary-marcus--x
9 Jun 2026
Safety

Dr. SHAP-AV: Decoding Relative Modality Contributions via Shapley Attribution in Audio-Visual Speech Recognition

DGX agent

arXiv:2603.12046v2 Announce Type: replace-cross Abstract: Audio-Visual Speech Recognition (AVSR) leverages both acoustic and visual information for robust recognition under noise. However, how models

safetyarxiv-cs-cv
9 Jun 2026
Safety

Dream-Tac: A Unified Tactile World Action Model for Contact-Rich Robot Manipulation

DGX agent

arXiv:2606.08737v1 Announce Type: new Abstract: World action models inherit the predictive capability of world models, enabling action generation to be guided by anticipated future observations. Howev

safetyarxiv-cs-ro
9 Jun 2026
Safety

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning

DGX agent

arXiv:2606.08035v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has emerged as a leading paradigm for enhancing visual reasoning in Multimodal Large Language Mode

safetyarxiv-cs-cv
9 Jun 2026
Safety

EgoAERO: Learning Dexterous Manipulation from a Single Egocentric Video without Object Assets

DGX agent

arXiv:2606.08057v1 Announce Type: cross Abstract: Egocentric RGB-D videos offer a natural source of human dexterous manipulation demonstrations, but existing data is difficult to use for robot learnin

safetyarxiv-cs-ai
9 Jun 2026
Safety

Emergent alignment and the projectability of ethical personas

DGX agent

arXiv:2606.09475v1 Announce Type: new Abstract: Work on `emergent misalignment' shows that finetuning LLMs on narrow tasks can induce broadly misaligned behavior. This supports the `persona selection'

safetyarxiv-cs-ai
9 Jun 2026
Safety

Enhancing AI Interpretability and Safety through Localised Architectures

DGX agent

arXiv:2606.07998v1 Announce Type: cross Abstract: Recent advances in generative AI, especially powerful Large Language Models (LLMs) and Large Reasoning Models (LRMs), raise concerns over the interpre

safetyarxiv-cs-ai
9 Jun 2026
Safety

Entropic Optimal Transport Eigenmaps for Nonlinear Alignment and Joint Embedding of High-Dimensional Datasets

DGX agent

arXiv:2407.01718v2 Announce Type: replace-cross Abstract: Embedding high-dimensional data into a low-dimensional space is an indispensable component of data analysis. In numerous applications, it is n

safetyarxiv-cs-lg
9 Jun 2026
Safety

Escaping the KL Agreement Trap in On-Policy Distillation

DGX agent

arXiv:2606.09471v1 Announce Type: new Abstract: On-policy distillation (OPD) provides dense token-level supervision by asking a teacher to score student-generated rollouts. However, when the student d

safetyarxiv-cs-lg
9 Jun 2026
Safety

Evaluating AI Investment Strategies

DGX agent

arXiv:2606.08791v1 Announce Type: cross Abstract: We study the problem of auditing a black-box algorithmic decision-maker from observable inputs and outputs alone. Our main result is an exact decompos

safetyarxiv-cs-ai
9 Jun 2026
Safety

Evaluation of ML Resource Utilization Requires Model Life Cycle Assessment

DGX agent

arXiv:2606.07632v1 Announce Type: new Abstract: Proper accounting of the energy requirements and environmental impact of artificial intelligence (AI) systems is necessary for researchers, developers,

safetyarxiv-cs-lg
9 Jun 2026
Safety

Exposing Hidden Biases in Text-to-Image Models via Automated Prompt Search

DGX agent

arXiv:2512.08724v3 Announce Type: replace Abstract: Text-to-image (TTI) diffusion models have achieved remarkable visual quality, yet they have been repeatedly shown to exhibit social biases across se

safetyarxiv-cs-lg
9 Jun 2026
Safety

fable’s safety guardrails broken within an hour. raise your hand if you are surprised.

DGX agent

fable’s safety guardrails broken within an hour. raise your hand if you are surprised. We tested Anthropic’s new @claudeai Fable 5. It did not fail like an ordinary jailbreak. It failed more quietly.

safetygary-marcus--x
9 Jun 2026
Safety

FADRW: A Feature-Aware Modulated and Dynamically Reweighted Loss for Few-Shot Linguistic Steganalysis

DGX agent

arXiv:2606.07655v1 Announce Type: cross Abstract: The ubiquity of social media platforms facilitates malicious linguistic steganography, posing significant security risks. However, detection is severe

safetyarxiv-cs-cv
9 Jun 2026
Safety

FADTI: Fourier and Attention Driven Diffusion for Multivariate Time Series Imputation

DGX agent

arXiv:2512.15116v2 Announce Type: replace-cross Abstract: Multivariate time series imputation is fundamental in applications such as healthcare, traffic forecasting, and biological modeling, where sen

safetyarxiv-cs-ai
9 Jun 2026
← Previous
1…99100101102103…267
Next →