AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
All
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,806 results
Safety

The Wittgensteinian Representation Hypothesis: Is Language the Attractor of Multimodal Convergence?

DGX agent

arXiv:2605.09352v1 Announce Type: new Abstract: Understanding why independently trained neural networks from different modalities converge toward shared representations, and where this convergence lea

safetyarxiv-cs-ai
12 May 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

There will be no AI jobpocalypse. The story that AI will lead to massive unemployment is stoking unnecessary fear. AI — like any other techn…

DGX agent

There will be no AI jobpocalypse. The story that AI will lead to massive unemployment is stoking unnecessary fear. AI — like any other technology — does affect jobs, but telling overblown stories of l

safetyandrew-ng--x
12 May 2026
Safety

Thermal-Det: Language-Guided Cross-Modal Distillation for Open-Vocabulary Thermal Object Detection

DGX agent

arXiv:2605.10130v1 Announce Type: new Abstract: Existing open-vocabulary detectors focus on RGB images and fail to generalize to thermal imagery, where low texture and emissivity variations challenge

safetyarxiv-cs-cv
12 May 2026
Safety

TIE: Time Interval Encoding for Video Generation over Events

DGX agent

arXiv:2605.10543v1 Announce Type: new Abstract: Director-style prompting, robotic action prediction, and interactive video agents demand temporal grounding over concurrent events -- a regime in which

safetyarxiv-cs-cv
12 May 2026
Safety

tl;dr - let's not just remove negative behavior from models, but also add positive ones 👍

DGX agent

tl;dr - let's not just remove negative behavior from models, but also add positive ones 👍 If anyone builds it, everyone thrives. Over the past decade, a lot of important work on AI alignment has focus

safetyyohei-nakajima--x
12 May 2026
Safety

TodyComm: Task-Oriented Dynamic Communication for Multi-Round LLM-based Multi-Agent System

DGX agent

arXiv:2602.03688v2 Announce Type: replace Abstract: Multi-round LLM-based multi-agent systems rely on effective communication structures to support collaboration across rounds. However, most existing

safetyarxiv-cs-ai
12 May 2026
Safety

Token Buncher: Shielding LLMs from Harmful Reinforcement Learning Fine-Tuning

DGX agent

arXiv:2508.20697v3 Announce Type: replace-cross Abstract: As large language models (LLMs) continue to grow in capability, so do the risks of harmful misuse through fine-tuning. While most prior studie

safetyarxiv-cs-cl
12 May 2026
Safety

Total Generalized Variation regularization closes the gap between neural-eld and classical methods in seismic travel-time tomography

DGX agent

arXiv:2605.09960v1 Announce Type: cross Abstract: Travel-time tomography forces a trade-off between mesh resolution and stability in which the regularizer choice dominates what can be recovered. We in

safetyarxiv-cs-lg
12 May 2026
Safety

Toward Reliable Sim-to-Real Predictability for MoE-based Robust Quadrupedal Locomotion

DGX agent

arXiv:2602.00678v4 Announce Type: replace Abstract: Reinforcement learning has shown strong promise for quadrupedal agile locomotion, even with proprioception-only sensing. In practice, however, sim-t

safetyarxiv-cs-ro
12 May 2026
Safety

Towards Customized Multimodal Role-Play

DGX agent

arXiv:2605.08129v1 Announce Type: new Abstract: Unified multimodal understanding and generation models enable richer human-AI interaction. Yet jointly customizing a character's persona, dialogue style

safetyarxiv-cs-lg
12 May 2026
Safety

TRACE: Distilling Where It Matters via Token-Routed Self On-Policy Alignment

DGX agent

arXiv:2605.10194v1 Announce Type: new Abstract: On-policy self-distillation (self-OPD) densifies reinforcement learning with verifiable rewards (RLVR) by letting a policy teach itself under privileged

safetyarxiv-cs-ai
12 May 2026
Safety

Training-Free Cultural Alignment of Large Language Models via Persona Disagreement

DGX agent

arXiv:2605.10843v1 Announce Type: cross Abstract: Large language models increasingly mediate decisions that turn on moral judgement, yet a growing body of evidence shows that their implicit preference

safetyarxiv-cs-ai
12 May 2026
Safety

Training with Harnesses: On-Policy Harness Self-Distillation for Complex Reasoning

DGX agent

arXiv:2605.08741v1 Announce Type: new Abstract: Inference-time harnesses substantially improve large language models on complex reasoning tasks. However, the intrinsic capabilities of the underlying m

safetyarxiv-cs-cl
12 May 2026
Safety

Trajectory-Consistent Flow Matching for Robust Visuomotor Policy Learning

DGX agent

arXiv:2605.08511v1 Announce Type: new Abstract: Flow matching policies learn continuous velocity fields that transport noise to actions, enabling fast deterministic inference for robot manipulation. H

safetyarxiv-cs-ro
12 May 2026
Safety

TrajTok: Learning Trajectory Tokens enables better Video Understanding

DGX agent

arXiv:2602.22779v2 Announce Type: replace Abstract: Tokenization in video models, typically through patchification, generates an excessive and redundant number of tokens. This severely limits video ef

safetyarxiv-cs-cv
12 May 2026
Safety

TripleWin: Fixed-Point Equilibrium Pricing for Data-Model Coupled Markets

DGX agent

arXiv:2511.03368v2 Announce Type: replace Abstract: The rise of the machine learning (ML) model economy has intertwined markets for training datasets and pre-trained models. However, most pricing appr

safetyarxiv-cs-lg
12 May 2026
Safety

Trustworthy AI: Ensuring Reliability and Accountability from Models to Agents

DGX agent

arXiv:2605.08964v1 Announce Type: new Abstract: In this thesis, we develop algorithms with theoretical guarantees for ensuring reliability and accountability of Machine Learning (ML) systems. As ML sy

safetyarxiv-cs-lg
12 May 2026
Safety

UAV-Assisted Scan-to-Simulation for Landslides Using Physics-Informed Gaussian Splatting

DGX agent

arXiv:2605.10715v1 Announce Type: new Abstract: Landslide monitoring and simulation play an important role in urban safety assessment and disaster prevention. Existing landslide simulation pipelines t

safetyarxiv-cs-cv
12 May 2026
Safety

Uncertainty-Aware and Decoder-Aligned Learning for Video Summarization

DGX agent

arXiv:2605.09507v1 Announce Type: new Abstract: Video summarization aims to produce a compact representation of a long video by selecting a subset of temporally important segments that best reflect hu

safetyarxiv-cs-cv
12 May 2026
Safety

Uni-Hand: Universal Hand Motion Forecasting in Egocentric Views

DGX agent

arXiv:2511.12878v4 Announce Type: replace Abstract: Forecasting how human hands move in egocentric views is critical for applications like augmented reality and human-robot policy transfer. Recently,

safetyarxiv-cs-cv
12 May 2026
Safety

Unified Noise Steering for Efficient Human-Guided VLA Adaptation

DGX agent

arXiv:2605.10821v1 Announce Type: new Abstract: Diffusion-based vision-language-action (VLA) models have emerged as strong priors for robotic manipulation, yet adapting them to real-world distribution

safetyarxiv-cs-ro
12 May 2026
Safety

Uniform Inductive Spatio-Temporal Kriging

DGX agent

arXiv:2603.05301v2 Announce Type: replace Abstract: Inductive spatio-temporal kriging infers signals at unobserved locations from observed sensors, but real-world observations are often incomplete and

safetyarxiv-cs-ai
12 May 2026
Safety

Unison: Harmonizing Motion, Speech, and Sound for Human-Centric Audio-Video Generation

DGX agent

arXiv:2605.08729v1 Announce Type: new Abstract: Motion, speech, and sound effects are fundamental elements of human-centric videos, yet their heterogeneous temporal characteristics make joint generati

safetyarxiv-cs-cv
12 May 2026
Safety

Unlearners Can Lie: Evaluating and Improving Honesty in LLM Unlearning

DGX agent

arXiv:2605.08765v1 Announce Type: cross Abstract: Unlearning in large language models (LLMs) aims to remove harmful training data while preserving overall utility. However, we find that existing metho

safetyarxiv-cs-ai
12 May 2026
Safety

Unsupervised Process Reward Models

DGX agent

arXiv:2605.10158v1 Announce Type: new Abstract: Process Reward Models (PRMs) are a powerful mechanism for steering large language model reasoning by providing fine-grained, step-level supervision. How

safetyarxiv-cs-lg
12 May 2026
Safety

Upholding Epistemic Agency: A Brouwerian Assertibility Constraint for Responsible AI

DGX agent

arXiv:2603.03971v2 Announce Type: replace-cross Abstract: Generative AI can convert uncertainty into hypersuasive, authoritative-seeming verdicts, displacing the justificatory work on which democratic

safetyarxiv-cs-ai
12 May 2026
Safety

Users as Annotators: LLM Preference Learning from Comparison Mode

DGX agent

arXiv:2510.13830v2 Announce Type: replace-cross Abstract: Pairwise preference data have played an important role in the alignment of large language models (LLMs). Each sample of such data consists of

safetyarxiv-cs-ai
12 May 2026
Safety

V-ABS: Action-Observer Driven Beam Search for Dynamic Visual Reasoning

DGX agent

arXiv:2605.10172v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) have achieved remarkable success in general perception, yet complex multi-step visual reasoning remains a per

safetyarxiv-cs-cl
12 May 2026
Safety

Value-Decomposed Reinforcement Learning Framework for Taxiway Routing with Hierarchical Conflict-Aware Observations

DGX agent

arXiv:2605.08754v1 Announce Type: new Abstract: Taxiway routing and on-surface conflict avoidance are coupled safety-critical decision problems in airport surface operations. Existing planning and opt

safetyarxiv-cs-ai
12 May 2026
Safety

Variational Inference for Levy Process-Driven SDEs via Neural Tilting

DGX agent

arXiv:2605.10934v1 Announce Type: cross Abstract: Modelling extreme events and heavy-tailed phenomena is central to building reliable predictive systems in domains such as finance, climate science, an

safetyarxiv-cs-ai
12 May 2026
Safety

Verification Mirage: Mapping the Reliability Boundary of Self-Verification in Medical VQA

DGX agent

arXiv:2605.10850v1 Announce Type: new Abstract: Self-verification, re-invoking the same vision language model (VLM) in a fresh context to check its own generated answer, is increasingly used as a defa

safetyarxiv-cs-cv
12 May 2026
Safety

Verifier-Free RL for LLMs via Intrinsic Gradient-Norm Reward

DGX agent

arXiv:2605.09920v1 Announce Type: cross Abstract: While Reinforcement Learning with Verifiable Rewards (RLVR) has recently emerged as a promising post-training paradigm for Large Language Models (LLMs

safetyarxiv-cs-ai
12 May 2026
Safety

VISTA: A Generative Egocentric Video Framework for Daily Assistance

DGX agent

arXiv:2605.10579v1 Announce Type: new Abstract: Training AI agents to proactively assist humans in daily activities, from routine household tasks to urgent safety situations, requires large-scale visu

safetyarxiv-cs-cl
12 May 2026
Safety

Wavelet Policy: Imitation Learning in the Scale Domain with World Prior Memory

DGX agent

arXiv:2504.04991v4 Announce Type: replace Abstract: Conventional visuomotor imitation learning usually predicts future robot actions directly in the time domain. Such formulations often have limited p

safetyarxiv-cs-ro
12 May 2026
Safety

'We're going to reach a point of diminishing returns from scaling. That's what actually happened. The industry knows this. But doesn't want …

DGX agent

'We're going to reach a point of diminishing returns from scaling. That's what actually happened. The industry knows this. But doesn't want to tell you.' @GaryMarcus, AI Expert, Scientist & Author at

safetygary-marcus--x
12 May 2026
Safety

What should post-training optimize? A test-time scaling law perspective

DGX agent

arXiv:2605.10716v1 Announce Type: new Abstract: Large language models are increasingly deployed with test-time strategies: sample N responses, score them with a reward model or verifier, and return th

safetyarxiv-cs-lg
12 May 2026
Safety

What Structural Inductive Bias Helps Transformers Reason Over Knowledge Graphs? A Study with Tabula RASA

DGX agent

arXiv:2602.02834v3 Announce Type: replace-cross Abstract: What structural inductive bias helps transformers reason over knowledge graphs? Through controlled ablations of a minimal transformer modifica

safetyarxiv-cs-ai
12 May 2026
Safety

When a Robot is More Capable than a Human: Learning from Constrained Demonstrators

DGX agent

arXiv:2510.09096v3 Announce Type: replace-cross Abstract: Learning from demonstrations enables experts to teach robots complex tasks using interfaces such as kinesthetic teaching, joystick control, an

safetyarxiv-cs-ai
12 May 2026
Safety

When Agents Overtrust Environmental Evidence: An Extensible Agentic Framework for Benchmarking Evidence-Grounding Defects in LLM Agents

DGX agent

arXiv:2605.08828v1 Announce Type: new Abstract: Large language model agents increasingly operate through environment-facing scaffolds that expose files, web pages, APIs, and logs. These observations i

safetyarxiv-cs-ai
12 May 2026
Safety

When Can Digital Personas Reliably Approximate Human Survey Findings?

DGX agent

arXiv:2605.10659v1 Announce Type: cross Abstract: Digital personas powered by Large Language Models (LLMs) are increasingly proposed as substitutes for human survey respondents, yet it remains unclear

safetyarxiv-cs-ai
12 May 2026
Safety

When Language Overwrites Vision: Over-Alignment and Geometric Debiasing in Vision-Language Models

DGX agent

arXiv:2605.08245v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) increasingly power high-stakes applications, from medical imaging to autonomous systems, yet they routinely hallucinate,

safetyarxiv-cs-ai
12 May 2026
Safety

When More Parameters Hurt: Foundation Model Priors Amplify Worst-Client Disparity Under Extreme Federated Heterogeneity

DGX agent

arXiv:2605.08992v1 Announce Type: new Abstract: Federated learning (FL) is increasingly used to fine-tune foundation models (FMs) on distributed private data. The community largely assumes that large-

safetyarxiv-cs-lg
12 May 2026
Safety

Where Do Flow Semantics Reside? A Protocol-Native Tabular Pretraining Paradigm for Encrypted Traffic Classification

DGX agent

arXiv:2603.10051v2 Announce Type: replace-cross Abstract: Self-supervised masked modeling shows promise for encrypted traffic classification by masking and reconstructing raw bytes. Yet recent work re

safetyarxiv-cs-ai
12 May 2026
Safety

White Circle raises $11M to help companies secure and monitor AI model behavior

DGX agent

Artificial intelligence guardrail and monitoring startup Pumpkin Intelligence Inc., which operates as White Circle, announced today it raised 11 million in seed funding from a who’s who of AI leadersh

safetysiliconangle
12 May 2026
Safety

Why Adam Works Better with eta_1 = eta_2: The Missing Gradient Scale Invariance Principle

DGX agent

arXiv:2601.21739v2 Announce Type: replace-cross Abstract: Adam has been at the core of large-scale training for almost a decade, yet a simple empirical fact remains unaccounted for: both validation sc

safetyarxiv-cs-ai
12 May 2026
Safety

Why Do DiT Editors Drift? Plug-and-Play Low Frequency Alignment in VAE Latent Space

DGX agent

arXiv:2605.08250v1 Announce Type: cross Abstract: Recent advances in diffusion transformers (DiTs) have enabled promising single-turn image editing capabilities. However, multi-turn editing often lead

safetyarxiv-cs-ai
12 May 2026
Safety

WISTERIA: Learning Clinical Representations from Noisy Supervision via Multi-View Consistency in Electronic Health Records

DGX agent

arXiv:2605.09765v1 Announce Type: cross Abstract: Representation learning in electronic health records (EHR) has largely followed paradigms inherited from natural language processing, relying on seque

safetyarxiv-cs-ai
12 May 2026
Safety

wow! “It is the [House] Committee’s understanding that the new board of directors at OpenAI tried to address these problems upon your return…

DGX agent

wow! “It is the [House] Committee’s understanding that the new board of directors at OpenAI tried to address these problems upon your return by creating an “audit committee to review potential conflic

safetygary-marcus--x
12 May 2026
← Previous
1…193194195196197…267
Next →