AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,704 results
Safety

UVIO: An UWB-Aided Visual-Inertial Odometry Framework with Bias-Compensated Anchors Initialization

DGX agent

arXiv:2308.00513v2 Announce Type: replace Abstract: This paper introduces UVIO, a multi-sensor framework that leverages Ultra Wide Band (UWB) technology and Visual-Inertial Odometry (VIO) to provide r

safetyarxiv-cs-ro
23 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

V-tableR1: Process-Supervised Multimodal Table Reasoning with Critic-Guided Policy Optimization

DGX agent

arXiv:2604.20755v1 Announce Type: new Abstract: We introduce V-tableR1, a process-supervised reinforcement learning framework that elicits rigorous, verifiable reasoning from multimodal large language

safetyarxiv-cs-ai
23 Apr 2026
Safety

Verification of Machine Unlearning is Fragile

DGX agent

arXiv:2408.00929v2 Announce Type: replace Abstract: As privacy concerns escalate in the realm of machine learning, data owners now have the option to utilize machine unlearning to remove their data fr

safetyarxiv-cs-lg
23 Apr 2026
Safety

Visual-Tactile Peg-in-Hole Assembly Learning from Peg-out-of-Hole Disassembly

DGX agent

arXiv:2604.20712v1 Announce Type: new Abstract: Peg-in-hole (PiH) assembly is a fundamental yet challenging robotic manipulation task. While reinforcement learning (RL) has shown promise in tackling s

safetyarxiv-cs-ro
23 Apr 2026
Safety

What Makes a Good AI Review? Concern-Level Diagnostics for AI Peer Review

DGX agent

arXiv:2604.19998v1 Announce Type: new Abstract: Evaluating AI-generated reviews by verdict agreement is widely recognized as insufficient, yet current alternatives rarely audit which concerns a system

safetyarxiv-cs-ai
23 Apr 2026
Safety

Where and What: Reasoning Dynamic and Implicit Preferences in Situated Conversational Recommendation

DGX agent

arXiv:2604.20749v1 Announce Type: new Abstract: Situated conversational recommendation (SCR), which utilizes visual scenes grounded in specific environments and natural language dialogue to deliver co

safetyarxiv-cs-ai
23 Apr 2026
Safety

Which Reasoning Trajectories Teach Students to Reason Better? A Simple Metric of Informative Alignment

DGX agent

arXiv:2601.14249v4 Announce Type: replace Abstract: Long chain-of-thought (CoT) trajectories provide rich supervision signals for distilling reasoning from teacher to student LLMs. However, both prior

safetyarxiv-cs-cl
23 Apr 2026
Safety

Whose Story Gets Told? Positionality and Bias in LLM Summaries of Life Narratives

DGX agent

arXiv:2604.20131v1 Announce Type: new Abstract: Increasingly, studies are exploring using Large Language Models (LLMs) for accelerated or scaled qualitative analysis of text data. While we can compare

safetyarxiv-cs-cl
23 Apr 2026
Safety

Working Memory Constraints Scaffold Learning in Transformers under Data Scarcity

DGX agent

arXiv:2604.20789v1 Announce Type: cross Abstract: We investigate the integration of human-like working memory constraints into the Transformer architecture and implement several cognitively inspired a

safetyarxiv-cs-ai
23 Apr 2026
Safety

A UK tribunal rules Microsoft must face a lawsuit alleging it overcharged UK businesses to run Windows Server on cloud services from Amazon, Google, and Alibaba (Sam Tobin/Reuters)

DGX agent

Sam Tobin / Reuters: A UK tribunal rules Microsoft must face a lawsuit alleging it overcharged UK businesses to run Windows Server on cloud services from Amazon, Google, and Alibaba — Microsoft (MSFT.

safetytechmeme
22 Apr 2026
Safety

Adaptive Prompt Elicitation for Text-to-Image Generation

DGX agent

arXiv:2602.04713v2 Announce Type: replace-cross Abstract: Aligning text-to-image generation with user intent remains challenging, as users frequently provide ambiguous inputs and struggle with model i

safetyarxiv-cs-ai
22 Apr 2026
Safety

AeroBridge-TTA: Test-Time Adaptive Language-Conditioned Control for UAVs

DGX agent

arXiv:2604.19059v1 Announce Type: new Abstract: Language-guided unmanned aerial vehicles (UAVs) often fail not from bad reasoning or perception, but from execution mismatch: the gap between a planned

safetyarxiv-cs-ro
22 Apr 2026
Safety

AI failure could trigger the next financial crisis, warns Elizabeth Warren

DGX agent

'I know a bubble when I see one.' That's what Sen. Elizabeth Warren (D-MA), who led the push to create a new consumer financial regulator in the wake of the 2008 recession, told a crowd at a Vanderbil

safetythe-verge-ai
22 Apr 2026
Safety

AlignedCut: Visual Concepts Discovery on Brain-Guided Universal Feature Space

DGX agent

arXiv:2406.18344v2 Announce Type: replace Abstract: We study the intriguing connection between visual data, deep networks, and the brain. Our method creates a universal channel alignment by using brai

safetyarxiv-cs-cv
22 Apr 2026
Safety

Allo{SR}^2: Rectifying One-Step Super-Resolution to Stay Real via Allomorphic Generative Flows

DGX agent

arXiv:2604.19238v1 Announce Type: new Abstract: Real-world image super-resolution (Real-SR) has been revolutionized by leveraging the powerful generative priors of large-scale diffusion and flow-based

safetyarxiv-cs-cv
22 Apr 2026
Safety

Anthropic’s own internal security blows.

DGX agent

Anthropic’s own internal security blows. Anthropic said Mythos was too dangerous to release. Then four random guys in a Discord gained access on day one by guessing the URL... This is pretty insane: →

safetygary-marcus--x
22 Apr 2026
Safety

ARES: Adaptive Red-Teaming and End-to-End Repair of Policy-Reward System

DGX agent

arXiv:2604.18789v1 Announce Type: new Abstract: Reinforcement Learning from Human Feedback (RLHF) is central to aligning Large Language Models (LLMs), yet it introduces a critical vulnerability: an im

safetyarxiv-cs-ai
22 Apr 2026
Safety

ARM: Advantage Reward Modeling for Long-Horizon Manipulation

DGX agent

arXiv:2604.03037v2 Announce Type: replace-cross Abstract: Long-horizon robotic manipulation remains challenging for reinforcement learning (RL) because sparse rewards provide limited guidance for cred

safetyarxiv-cs-ai
22 Apr 2026
Safety

Assessing VLM-Driven Semantic-Affordance Inference for Non-Humanoid Robot Morphologies

DGX agent

arXiv:2604.19509v1 Announce Type: new Abstract: Vision-language models (VLMs) have demonstrated remarkable capabilities in understanding human-object interactions, but their application to robotic sys

safetyarxiv-cs-ro
22 Apr 2026
Safety

ASVSim (AirSim for Surface Vehicles): A High-Fidelity Simulation Framework for Autonomous Surface Vehicle Research

DGX agent

arXiv:2506.22174v2 Announce Type: replace-cross Abstract: The transport industry has recently shown significant interest in unmanned surface vehicles (USVs), specifically for port and inland waterway

safetyarxiv-cs-lg
22 Apr 2026
Safety

Attention-based Multi-modal Deep Learning Model of Spatio-temporal Crop Yield Prediction with Satellite, Soil and Climate Data

DGX agent

arXiv:2604.19217v1 Announce Type: cross Abstract: Crop yield prediction is one of the most important challenge, which is crucial to world food security and policy-making decisions. The conventional fo

safetyarxiv-cs-ai
22 Apr 2026
Safety

Auditing LLMs for Algorithmic Fairness in Casenote-Augmented Tabular Prediction

DGX agent

arXiv:2604.19204v1 Announce Type: cross Abstract: LLMs are increasingly being considered for prediction tasks in high-stakes social service settings, but their algorithmic fairness properties in this

safetyarxiv-cs-lg
22 Apr 2026
Safety

Australia's eSafety Commissioner issues transparency notices to Roblox, Microsoft's Minecraft, and other online gaming platforms to detail child safety measures (Renju Jose/Reuters)

DGX agent

Renju Jose / Reuters: Australia's eSafety Commissioner issues transparency notices to Roblox, Microsoft's Minecraft, and other online gaming platforms to detail child safety measures — Australia's int

safetytechmeme
22 Apr 2026
Safety

AutoAWG: Adverse Weather Generation with Adaptive Multi-Controls for Automotive Videos

DGX agent

arXiv:2604.18993v1 Announce Type: cross Abstract: Perception robustness under adverse weather remains a critical challenge for autonomous driving, with the core bottleneck being the scarcity of real-w

safetyarxiv-cs-ai
22 Apr 2026
Safety

BAPO: Boundary-Aware Policy Optimization for Reliable Agentic Search

DGX agent

arXiv:2601.11037v2 Announce Type: replace Abstract: RL-based agentic search enables LLMs to solve complex questions via dynamic planning and external search. While this approach significantly enhances

safetyarxiv-cs-ai
22 Apr 2026
Safety

Benchmarking Misuse Mitigation Against Covert Adversaries

DGX agent

arXiv:2506.06414v2 Announce Type: replace-cross Abstract: Existing language model safety evaluations focus on overt attacks and low-stakes tasks. In reality, an attacker can easily subvert existing sa

safetyarxiv-cs-ai
22 Apr 2026
Safety

Beyond Bellman: High-Order Generator Regression for Continuous-Time Policy Evaluation

DGX agent

arXiv:2604.18972v1 Announce Type: cross Abstract: We study finite-horizon continuous-time policy evaluation from discrete closed-loop trajectories under time-inhomogeneous dynamics. The target value s

safetyarxiv-cs-lg
22 Apr 2026
Safety

Beyond Linear Probes: Dynamic Safety Monitoring for Language Models

DGX agent

arXiv:2509.26238v4 Announce Type: replace Abstract: Monitoring large language models' (LLMs) activations is an effective way to detect harmful requests before they lead to unsafe outputs. However, tra

safetyarxiv-cs-lg
22 Apr 2026
Safety

Beyond Marginal Distributions: A Framework to Evaluate the Representativeness of Demographic-Aligned LLMs

DGX agent

arXiv:2601.15755v3 Announce Type: replace Abstract: Large language models are increasingly used to represent human opinions, values, or beliefs, and their steerability towards these ideals is an activ

safetyarxiv-cs-cl
22 Apr 2026
Safety

Beyond Semantic Similarity: A Component-Wise Evaluation Framework for Medical Question Answering Systems with Health Equity Implications

DGX agent

arXiv:2604.19281v1 Announce Type: cross Abstract: The use of Large Language Models (LLMs) to support patients in addressing medical questions is becoming increasingly prevalent. However, most of the m

safetyarxiv-cs-ai
22 Apr 2026
Safety

Beyond the Bellman Fixed Point: Geometry and Fast Policy Identification in Value Iteration

DGX agent

arXiv:2604.17457v2 Announce Type: replace-cross Abstract: Dynamic programming is one of the most fundamental methodologies for solving Markov decision problems. Among its many variants, Q-value iterat

safetyarxiv-cs-ai
22 Apr 2026
Safety

Breaking the Illusion: Consensus-Based Generative Mitigation of Adversarial Illusions in Multi-Modal Embeddings

DGX agent

arXiv:2511.21893v2 Announce Type: replace Abstract: Multi-modal foundation models align images, text, and other modalities in a shared embedding space but remain vulnerable to adversarial illusions [3

safetyarxiv-cs-lg
22 Apr 2026
Safety

Bridging Semantics and Geometry: A Decoupled LVLM-SAM Framework for Reasoning Segmentation in Optical Remote Sensing

DGX agent

arXiv:2512.19302v2 Announce Type: replace Abstract: Large Vision--Language Models (LVLMs) hold great promise for advancing optical remote sensing (RS) analysis, yet existing reasoning segmentation fra

safetyarxiv-cs-cv
22 Apr 2026
Safety

CAHAL: Clinically Applicable resolution enHAncement for Low-resolution MRI scans

DGX agent

arXiv:2604.18781v1 Announce Type: new Abstract: Large-scale automated morphometric analysis of brain MRI is limited by the thick-slice, anisotropic acquisitions prevalent in routine clinical practice.

safetyarxiv-cs-cv
22 Apr 2026
Safety

Capturing Classic Authorial Style in Long-Form Story Generation with GRPO Fine-Tuning

DGX agent

arXiv:2512.05747v3 Announce Type: replace Abstract: Evaluating and optimising authorial style in long-form story generation remains challenging because style is often assessed with ad hoc prompting an

safetyarxiv-cs-cl
22 Apr 2026
Safety

CentaurTA Studio: A Self-Improving Human-Agent Collaboration System for Thematic Analysis

DGX agent

arXiv:2604.18589v1 Announce Type: cross Abstract: Thematic analysis is difficult to scale: manual workflows are labor-intensive, while fully automated pipelines often lack controllability and transpar

safetyarxiv-cs-ai
22 Apr 2026
Safety

Chain-of-Thought as a Lens: Evaluating Structured Reasoning Alignment between Human Preferences and Large Language Models

DGX agent

arXiv:2511.06168v3 Announce Type: replace Abstract: This paper primarily demonstrates a method to quantitatively assess the alignment between multi-step, structured reasoning in large language models

safetyarxiv-cs-ai
22 Apr 2026
Safety

ChatGPT doesn’t know its whisk from its elbow

DGX agent

Gary Marcus critiques ChatGPT's lack of embodied understanding and spatial reasoning, arguing that the language model struggles with physical concepts that humans intuitively grasp through bodily expe

safetygary-marcus--x
22 Apr 2026
Safety

Cloning Deterministic Worlds: The Critical Role of Latent Geometry in Long-Horizon World Models

DGX agent

arXiv:2510.26782v3 Announce Type: replace-cross Abstract: A world model is an internal model that simulates how the world evolves. Given past observations and actions, it predicts the future physical

safetyarxiv-cs-ai
22 Apr 2026
Safety

Counting Worlds Branching Time Semantics for post-hoc Bias Mitigation in generative AI

DGX agent

arXiv:2604.19431v1 Announce Type: cross Abstract: Generative AI systems are known to amplify biases present in their training data. While several inference-time mitigation strategies have been propose

safetyarxiv-cs-ai
22 Apr 2026
Safety

CreatiParser: Generative Image Parsing of Raster Graphic Designs into Editable Layers

DGX agent

arXiv:2604.19632v1 Announce Type: new Abstract: Graphic design images consist of multiple editable layers, such as text, background, and decorative elements, while most generative models produce raste

safetyarxiv-cs-cv
22 Apr 2026
Safety

Curvature-Aware PCA with Geodesic Tangent Space Aggregation for Semi-Supervised Learning

DGX agent

arXiv:2604.18816v1 Announce Type: cross Abstract: Principal Component Analysis (PCA) is a fundamental tool for representation learning, but its global linear formulation fails to capture the structure

safetyarxiv-cs-ai
22 Apr 2026
Safety

Debiased neural operators for estimating functionals

DGX agent

arXiv:2604.19296v1 Announce Type: new Abstract: Neural operators are widely used to approximate solution maps of complex physical systems. In many applications, however, the goal is not to recover the

safetyarxiv-cs-lg
22 Apr 2026
Safety

Decomposed Trust: Privacy, Adversarial Robustness, Ethics, and Fairness in Low-Rank LLMs

DGX agent

arXiv:2511.22099v3 Announce Type: replace-cross Abstract: Large language models (LLMs) have driven major advances across domains, yet their massive size hinders deployment in resource-constrained sett

safetyarxiv-cs-ai
22 Apr 2026
Safety

Denoising, Fast and Slow: Difficulty-Aware Adaptive Sampling for Image Generation

DGX agent

arXiv:2604.19141v1 Announce Type: new Abstract: Diffusion- and flow-based models usually allocate compute uniformly across space, updating all patches with the same timestep and number of function eva

safetyarxiv-cs-cv
22 Apr 2026
Safety

Developing a Robotic Surgery Training System for Wide Accessibility and Research

DGX agent

arXiv:2505.20562v2 Announce Type: replace Abstract: Robotic surgery represents a major breakthrough in medical interventions, which has revolutionized surgical procedures. However, the high cost and l

safetyarxiv-cs-ro
22 Apr 2026
Safety

Diagnosable ColBERT: Debugging Late-Interaction Retrieval Models Using a Learned Latent Space as Reference

DGX agent

arXiv:2604.19566v1 Announce Type: cross Abstract: Reliable biomedical and clinical retrieval requires more than strong ranking performance: it requires a practical way to find systematic model failure

safetyarxiv-cs-cl
22 Apr 2026
Safety

Diamond Maps: Efficient Reward Alignment via Stochastic Flow Maps

DGX agent

arXiv:2602.05993v2 Announce Type: replace-cross Abstract: Flow and diffusion models produce high-quality samples, but adapting them to user preferences or constraints post-training remains costly and

safetyarxiv-cs-ai
22 Apr 2026
← Previous
1…230231232233234…265
Next →