AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
All
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,702 results
Safety

Symbolic Quantile Regression for the Interpretable Prediction of Conditional Quantiles

DGX agent

arXiv:2508.08080v2 Announce Type: replace Abstract: Symbolic Regression (SR) is a well-established framework for generating interpretable or white-box predictive models. Although SR has been successfu

safetyarxiv-cs-lg
22 Apr 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

Task-Adaptive Admittance Control for Human-Quadrotor Cooperative Load Transportation with Dynamic Cable-Length Regulation

DGX agent

arXiv:2604.18905v1 Announce Type: new Abstract: The collaboration between humans and robots is critical in many robotic applications, especially in those requiring physical human-robot interaction (pH

safetyarxiv-cs-ro
22 Apr 2026
Safety

TEMPO: Scaling Test-time Training for Large Reasoning Models

DGX agent

arXiv:2604.19295v1 Announce Type: new Abstract: Test-time training (TTT) adapts model parameters on unlabeled test instances during inference time, which continuously extends capabilities beyond the r

safetyarxiv-cs-lg
22 Apr 2026
Safety

Terrible

DGX agent

Terrible This is insane… The Virginia redistricting amendment on the ballot today is framed as a vote to 'restore fairness in the upcoming elections.' In reality, it turns a state that Kamala barely w

safetyelon-musk--x
22 Apr 2026
Safety

The Alignment Waltz: Jointly Training Agents to Collaborate for Safety

DGX agent

arXiv:2510.08240v2 Announce Type: replace Abstract: Harnessing the power of LLMs requires a delicate dance between being helpful and harmless. This creates a fundamental tension between two competing

safetyarxiv-cs-cl
22 Apr 2026
Safety

The Data-Driven Censored Newsvendor Problem

DGX agent

arXiv:2412.01763v3 Announce Type: replace-cross Abstract: We study a censored variant of the data-driven newsvendor problem, where the decision-maker must select an ordering quantity that minimizes ex

safetyarxiv-cs-lg
22 Apr 2026
Safety

The Essence of Balance for Self-Improving Agents in Vision-and-Language Navigation

DGX agent

arXiv:2604.19064v1 Announce Type: new Abstract: In vision-and-language navigation (VLN), self-improvement from policy-induced experience, using only standard VLN action supervision, critically depends

safetyarxiv-cs-cv
22 Apr 2026
Safety

The PROPER Approach to Proactivity: Benchmarking and Advancing Knowledge Gap Navigation

DGX agent

arXiv:2601.09926v3 Announce Type: replace Abstract: Current approaches to proactive assistance move beyond the ask-and-respond paradigm by anticipating user needs. In practice, they either burden user

safetyarxiv-cs-lg
22 Apr 2026
Safety

The signal is the ceiling: Measurement limits of LLM-predicted experience ratings from open-ended survey text

DGX agent

arXiv:2604.19645v1 Announce Type: new Abstract: An earlier paper (Hong, Potteiger, and Zapata 2026) established that an unoptimized GPT 4.1 prompt predicts fan-reported experience ratings within one p

safetyarxiv-cs-cl
22 Apr 2026
Safety

The Triadic Loop: A Framework for Negotiating Alignment in AI Co-hosted Livestreaming

DGX agent

arXiv:2604.18850v1 Announce Type: cross Abstract: AI systems are increasingly embedded in multi-user social environments, yet most alignment frameworks conceptualize interaction as a dyadic relationsh

safetyarxiv-cs-ai
22 Apr 2026
Safety

This actually happened. Smart second-graders know better. WATF. 🤯

DGX agent

Gary Marcus expresses skepticism or concern about an artificial intelligence claim or incident that he finds implausible, suggesting that even young children would recognize the flaw in the reasoning

safetygary-marcus--x
22 Apr 2026
Safety

Toward Clinically Acceptable Chest X-ray Report Generation: A Qualitative Retrospective Pilot Study of CXRMate-2

DGX agent

arXiv:2604.18967v1 Announce Type: new Abstract: Chest X-ray (CXR) radiology report generation (RRG) models have shown rapid progress, yet their clinical utility remains uncertain due to limited evalua

safetyarxiv-cs-cv
22 Apr 2026
Safety

TRN-R1-Zero: Text-rich Network Reasoning via LLMs with Reinforcement Learning Only

DGX agent

arXiv:2604.19070v1 Announce Type: new Abstract: Zero-shot reasoning on text-rich networks (TRNs) remains a challenging frontier, as models must integrate textual semantics with relational structure wi

safetyarxiv-cs-cl
22 Apr 2026
Safety

TROJail: Trajectory-Level Optimization for Multi-Turn Large Language Model Jailbreaks with Process Rewards

DGX agent

arXiv:2512.07761v3 Announce Type: replace Abstract: Large language models have seen widespread adoption, yet they remain vulnerable to multi-turn jailbreak attacks, threatening their safe deployment.

safetyarxiv-cs-ai
22 Apr 2026
Safety

User Simulation in the Era of Generative AI: User Modeling, Synthetic Data Generation, and System Evaluation

DGX agent

arXiv:2501.04410v2 Announce Type: replace Abstract: User simulation is an emerging interdisciplinary topic with multiple critical applications in the era of Generative AI. It involves creating an inte

safetyarxiv-cs-ai
22 Apr 2026
Safety

VIGIL: An Extensible System for Real-Time Detection and Mitigation of Cognitive Bias Triggers

DGX agent

arXiv:2604.03261v2 Announce Type: replace Abstract: The rise of generative AI is posing increasing risks to online information integrity and civic discourse. Most concretely, such risks can materialis

safetyarxiv-cs-cl
22 Apr 2026
Safety

VimRAG: Navigating Massive Visual Context in Retrieval-Augmented Generation via Multimodal Memory Graph

DGX agent

arXiv:2602.12735v2 Announce Type: replace-cross Abstract: Effectively retrieving, reasoning, and understanding multimodal information remains a critical challenge for agentic systems. Traditional Retr

safetyarxiv-cs-cl
22 Apr 2026
Safety

Vision-Based Human Awareness Estimation for Enhanced Safety and Efficiency of AMRs in Industrial Warehouses

DGX agent

arXiv:2604.18627v1 Announce Type: new Abstract: Ensuring human safety is of paramount importance in warehouse environments that feature mixed traffic of human workers and autonomous mobile robots (AMR

safetyarxiv-cs-cv
22 Apr 2026
Safety

Visual Adversarial Attack on Vision-Language Models for Autonomous Driving

DGX agent

arXiv:2411.18275v2 Announce Type: replace Abstract: Vision-language models (VLMs) have significantly advanced autonomous driving (AD) by enhancing reasoning capabilities. However, these models remain

safetyarxiv-cs-cv
22 Apr 2026
Safety

VoteGCL: Enhancing Graph-based Recommendations with Majority-Voting LLM-Rerank Augmentation

DGX agent

arXiv:2507.21563v4 Announce Type: replace-cross Abstract: Recommendation systems often suffer from data sparsity caused by limited user-item interactions, which degrade their performance and amplify p

safetyarxiv-cs-lg
22 Apr 2026
Safety

Weakly supervised framework for wildlife detection and counting in challenging Arctic environments: a case study on caribou (Rangifer tarandus)

DGX agent

arXiv:2601.18891v3 Announce Type: replace Abstract: Caribou across the Arctic has declined in recent decades, motivating scalable and accurate monitoring approaches to guide evidence-based conservatio

safetyarxiv-cs-cv
22 Apr 2026
Safety

When Can We Trust Deep Neural Networks? Towards Reliable Industrial Deployment with an Interpretability Guide

DGX agent

arXiv:2604.19206v1 Announce Type: new Abstract: The deployment of AI systems in safety-critical domains, such as industrial defect inspection, autonomous driving, and medical diagnosis, is severely ha

safetyarxiv-cs-cv
22 Apr 2026
Safety

3 new ways Ads Advisor is making Google Ads safer and faster

DGX agent

Ads Advisor is an agentic conversational experience built with Gemini in Google Ads designed to help maximize performance based on business goals. The tool can help users get personalized answers, und

safetygoogle-ai
21 Apr 2026
Safety

A Hamilton-Jacobi Reachability-Guided Search Framework for Efficient and Safe Indoor Planar Robot Navigation

DGX agent

arXiv:2604.17679v1 Announce Type: new Abstract: Autonomous navigation requires planning to reach a goal safely and efficiently in complex and potentially dynamic environments. Graph search-based algor

safetyarxiv-cs-ro
21 Apr 2026
Safety

A High-Accuracy Optical Music Recognition Method Based on Bottleneck Residual Convolutions

DGX agent

arXiv:2604.16446v1 Announce Type: new Abstract: Optical Music Recognition (OMR) aims to convert printed or handwritten music score images into editable symbolic representations. This paper presents an

safetyarxiv-cs-cv
21 Apr 2026
Safety

A Quasi-Experimental Developer Study of Security Training in LLM-Assisted Web Application Development

DGX agent

arXiv:2604.17763v1 Announce Type: cross Abstract: This paper presents a controlled quasi-experimental developer study examining whether a layer-based security training package is associated with impro

safetyarxiv-cs-lg
21 Apr 2026
Safety

A Real-Time Bike-Pedestrian Safety System with Wide-Angle Perception and Evaluation Testbed for Urban Intersections

DGX agent

arXiv:2604.17046v1 Announce Type: new Abstract: Collisions between cyclists and pedestrians at urban intersections remain a persistent source of injuries, yet few systems attempt real-time warnings to

safetyarxiv-cs-cv
21 Apr 2026
Safety

A Sensitivity Approach to Causal Inference Under Limited Overlap

DGX agent

arXiv:2511.22003v2 Announce Type: replace-cross Abstract: Limited overlap between treated and control groups is a key challenge in observational analysis. Standard approaches like trimming importance

safetyarxiv-cs-lg
21 Apr 2026
Safety

A Text-To-Text Alignment Algorithm for Better Evaluation of Modern Speech Recognition Systems

DGX agent

arXiv:2509.24478v2 Announce Type: replace Abstract: Modern neural networks have greatly improved performance across speech recognition benchmarks. However, gains are often driven by frequent words wit

safetyarxiv-cs-cl
21 Apr 2026
Safety

Adversarial Arena: Crowdsourcing Data Generation through Interactive Competition

DGX agent

arXiv:2604.17803v1 Announce Type: cross Abstract: Post-training Large Language Models requires diverse, high-quality data which is rare and costly to obtain, especially in low resource domains and for

safetyarxiv-cs-lg
21 Apr 2026
Safety

Agree, Disagree, Explain: Decomposing Human Label Variation in NLI through the Lens of Explanations

DGX agent

arXiv:2510.16458v2 Announce Type: replace Abstract: Natural Language Inference (NLI) datasets often exhibit human label variation. To better understand these variations, explanation-based approaches a

safetyarxiv-cs-cl
21 Apr 2026
Safety

Align Documents to Questions: Question-Oriented Document Rewriting for Retrieval-Augmented Generation

DGX agent

arXiv:2604.17325v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) enhances the factuality of Large Language Models (LLMs) by incorporating retrieved documents and/or generated conte

safetyarxiv-cs-cl
21 Apr 2026
Safety

Aligning Backchannel and Dialogue Context Representations via Contrastive LLM Fine-Tuning

DGX agent

arXiv:2604.16622v1 Announce Type: new Abstract: Backchannels (e.g., `yeah', `mhm', and `right') are short, non-interruptive feedback signals whose lexical form and prosody jointly convey pragmatic mea

safetyarxiv-cs-cl
21 Apr 2026
Safety

Aligning Language Models for Lyric-to-Melody Generation with Rule-Based Musical Constraints

DGX agent

arXiv:2604.18489v1 Announce Type: cross Abstract: Large Language Models (LLMs) show promise in lyric-to-melody generation, but models trained with Supervised Fine-Tuning (SFT) often produce musically

safetyarxiv-cs-cl
21 Apr 2026
Safety

Alignment Data Map for Efficient Preference Data Selection and Diagnosis

DGX agent

arXiv:2505.23114v3 Announce Type: replace Abstract: Human preference data is essential for aligning large language models (LLMs) with human values, but collecting such data is often costly and ineffic

safetyarxiv-cs-cl
21 Apr 2026
Safety

An `Inverse' Experimental Framework to Estimate Market Efficiency

DGX agent

arXiv:2604.18130v1 Announce Type: new Abstract: Digital marketplaces processing billions of dollars annually represent critical infrastructure in sociotechnical ecosystems, yet their performance optim

safetyarxiv-cs-lg
21 Apr 2026
Safety

Annotation-Assisted Learning of Treatment Policies From Multimodal Electronic Health Records

DGX agent

arXiv:2507.20993v3 Announce Type: replace Abstract: We study how to learn treatment policies from multimodal electronic health records (EHRs) that consist of tabular data and clinical text. These poli

safetyarxiv-cs-lg
21 Apr 2026
Safety

APIs and limited releases for AI models are not a safety policy, they’re a business model (which is totally ok as long as you’re transparent…

DGX agent

APIs and limited releases for AI models are not a safety policy, they’re a business model (which is totally ok as long as you’re transparent about it). Especially on cyber-security, they give a false

safetyclem-delangue--x
21 Apr 2026
Safety

Arch: An AI-Native Hardware Description Language for Register-Transfer Clocked Hardware Design

DGX agent

arXiv:2604.05983v2 Announce Type: replace-cross Abstract: We present Arch (AI-native Register-transfer Clocked Hardware), a hardware description language for micro-architecture specification and AI-as

safetyarxiv-cs-cl
21 Apr 2026
Safety

ARCS: Autoregressive Circuit Synthesis with Topology-Aware Graph Attention and Spec Conditioning

DGX agent

arXiv:2603.29068v3 Announce Type: replace Abstract: This paper presents ARCS (Autoregressive Circuit Synthesis), a system for amortized analog circuit generation. ARCS produces complete, SPICE-simulat

safetyarxiv-cs-lg
21 Apr 2026
Safety

Asset Harvester: Extracting 3D Assets from Autonomous Driving Logs for Simulation

DGX agent

arXiv:2604.18468v1 Announce Type: new Abstract: Closed-loop simulation is a core component of autonomous vehicle (AV) development, enabling scalable testing, training, and safety validation before rea

safetyarxiv-cs-cv
21 Apr 2026
Safety

ASTRA: An Automated Framework for Strategy Discovery, Retrieval, and Evolution for Jailbreaking LLMs

DGX agent

arXiv:2511.02356v2 Announce Type: replace-cross Abstract: Despite extensive safety alignment, Large Language Models (LLMs) remain vulnerable to jailbreak attacks. However, existing methods generally l

safetyarxiv-cs-lg
21 Apr 2026
Safety

Attention-space Contrastive Guidance for Efficient Hallucination Mitigation in LVLMs

DGX agent

arXiv:2601.13707v2 Announce Type: replace Abstract: Hallucinations in large vision--language models (LVLMs) often arise when language priors dominate over visual evidence, leading to object misidentif

safetyarxiv-cs-cv
21 Apr 2026
Safety

Audio-DeepThinker: Progressive Reasoning-Aware Reinforcement Learning for High-Quality Chain-of-Thought Emergence in Audio Language Models

DGX agent

arXiv:2604.18187v1 Announce Type: cross Abstract: Large Audio-Language Models (LALMs) have made significant progress in audio understanding, yet they primarily operate as perception-and-answer systems

safetyarxiv-cs-cl
21 Apr 2026
Safety

AutoGraph-R1: End-to-End Reinforcement Learning for Knowledge Graph Construction

DGX agent

arXiv:2510.15339v3 Announce Type: replace Abstract: Building effective knowledge graphs (KGs) for Retrieval-Augmented Generation (RAG) is pivotal for advancing question answering (QA) systems. However

safetyarxiv-cs-cl
21 Apr 2026
Safety

Autonomous Vehicle Collision Avoidance With Racing Parameterized Deep Reinforcement Learning

DGX agent

arXiv:2604.16702v1 Announce Type: new Abstract: Road traffic accidents are a leading cause of fatalities worldwide. In the US, human error causes 94% of crashes, resulting in excess of 7,000 pedestria

safetyarxiv-cs-ro
21 Apr 2026
Safety

Back into Plato's Cave: Examining Cross-modal Representational Convergence at Scale

DGX agent

arXiv:2604.18572v1 Announce Type: new Abstract: The Platonic Representation Hypothesis suggests that neural networks trained on different modalities (e.g., text and images) align and eventually conver

safetyarxiv-cs-cv
21 Apr 2026
Safety

BASIS: Balanced Activation Sketching with Invariant Scalars for 'Ghost Backpropagation'

DGX agent

arXiv:2604.16324v1 Announce Type: new Abstract: The activation memory required for exact backpropagation scales linearly with network depth, context length, and feature dimensionality, forming an O(L

safetyarxiv-cs-lg
21 Apr 2026
← Previous
1…233234235236237…265
Next →