AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,357 results
Safety

Reoptimization Algorithms for Contextual Bandits with Knapsack Constraints

DGX agent

arXiv:2608.11383v1 Announce Type: new Abstract: We study new algorithms for Contextual Bandits with Knapsack. In these problems, there are finitely many types of customers, products, and resources. Ea

safetyarxiv-cs-lg
13 Aug 2026
Safety
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Robust Multi-Tier Infant-Centered Audio Understanding with Whisper via Structured Speaker Conditioning

DGX agent

arXiv:2608.11587v1 Announce Type: cross Abstract: Recent advances in model design and self-supervised audio representations have improved speech and audio understanding, yet infant-centered naturalist

safetyarxiv-cs-cl
13 Aug 2026
Safety

ScaleVid: Geometry-Aware Video Object Scaling with Mesh-Free Inference

DGX agent

arXiv:2608.12232v1 Announce Type: new Abstract: Geometry-aware video object scaling aims to anisotropically resize the object along object-centric axes while preserving geometric plausibility, tempora

safetyarxiv-cs-cv
13 Aug 2026
Safety

Small Data Explainer -- The impact of small data methods in everyday life

DGX agent

arXiv:2507.11773v2 Announce Type: replace-cross Abstract: The emergence of breakthrough artificial intelligence (AI) techniques has led to a renewed focus on how small data settings, i.e., settings wi

safetyarxiv-cs-ai
13 Aug 2026
Safety

Stolen LLM Reasoning: How come OpenAI, Anthrophic, Google have the same vulnerabilities?

DGX agent

If you haven't checked the paper: https://arxiv.org/abs/2608.09867 TLDR: the authors show that you can swap out the 'encrypted' reasoning of the biggest model, like Opus, Sol, and put them into weaker

safetyr-localllama
13 Aug 2026
Safety

Through Van Gogh's Eyes: Global Style Transfer with Diffusion Mod

DGX agent

arXiv:2608.11546v1 Announce Type: new Abstract: Artistic image synthesis aims to recreate the expressive visual identity of a target artist, yet existing methods often fail to capture an artist's glob

safetyarxiv-cs-cv
13 Aug 2026
Safety

ToolHazard: Scaling Adversarial Environments for Security Evaluation and Alignment of LLM-based Agents

DGX agent

arXiv:2608.11878v1 Announce Type: cross Abstract: Large language model (LLM) agents integrated with external tools are vulnerable to indirect prompt injections embedded in environmental states. Howeve

safetyarxiv-cs-cl
13 Aug 2026
Safety

Toward Meaningful Transparency for AI Chatbots: Disclosing Persuasive Intent Reduces Persuasion

DGX agent

arXiv:2608.11794v1 Announce Type: cross Abstract: The growing role of AI-generated content and AI-enabled systems in public communication has led regulators to demand clear disclosure of content prove

safetyarxiv-cs-ai
13 Aug 2026
Safety

Towards a Formal Definition of Agent Memory: Basis, Span, Optimality, and the Sequential Memory Problem

DGX agent

arXiv:2608.11654v1 Announce Type: new Abstract: Despite the wide deployment of memory in large-model agents, there is no unified formal account of what a memory is or when it is optimal. This paper ta

safetyarxiv-cs-lg
13 Aug 2026
Safety

Towards Human Motion World Models via Executable Behaviour Representations

DGX agent

arXiv:2604.18064v2 Announce Type: replace Abstract: Human motion world models should capture motion's intentionality by being executable: adaptable to different actions and capable of assessing motion

safetyarxiv-cs-ai
13 Aug 2026
Safety

Towards Understanding On-Policy Distillation through the Lens of Test-Time Scaling

DGX agent

arXiv:2608.11829v1 Announce Type: cross Abstract: On-policy distillation (OPD) has emerged as a promising post-training technique for enhancing LLM reasoning. It is commonly believed to enable the stu

safetyarxiv-cs-cl
13 Aug 2026
Safety

UniSwap: Streaming Audio-Visual Identity Swapping for Talking Videos

DGX agent

arXiv:2608.11752v1 Announce Type: new Abstract: Talking-video character replacement requires coordinated transfer of appearance and voice while preserving the source motion, scene, linguistic content,

safetyarxiv-cs-cv
13 Aug 2026
Safety

Variable Selection in the Context of AI Fairness

DGX agent

arXiv:2608.11251v1 Announce Type: cross Abstract: Fairness in AI systems has become more important with recent regulatory demands, such as the EU AI Act. Traditional approaches often do not take into

safetyarxiv-cs-ai
13 Aug 2026
Safety

Who Would You Vote For? Auditing Political Alignment in LLMs: An Italian Case-Study

DGX agent

arXiv:2608.11649v1 Announce Type: new Abstract: As users increasingly turn to Large Language Models (LLMs) for information and advice on political matters, particularly during election periods, the po

safetyarxiv-cs-cl
13 Aug 2026
Safety

Why AI Detection Fails for Academic Integrity

DGX agent

arXiv:2608.11256v1 Announce Type: new Abstract: Institutions use commercial AI detectors for academic integrity, yet detectors cannot distinguish AI editing from full LLM drafts and may treat both as

safetyarxiv-cs-lg
13 Aug 2026
Safety

A Joint-Distribution Route to Fair Representations with Continuous Sensitive Attributes

DGX agent

arXiv:2608.10470v1 Announce Type: new Abstract: Fair representation learning with a continuous sensitive attribute S requires a representation Z that is statistically independent of S. Existing criter

safetyarxiv-cs-lg
12 Aug 2026
Safety

Actions Speak Louder than Words: Measuring Cross-Lingual Policy Retention in Tool-Using Agents

DGX agent

arXiv:2608.11110v1 Announce Type: new Abstract: When a tool-using agent is given the same task in a different language, does it still take the same steps? Multilingual evaluation rarely asks: it compa

safetyarxiv-cs-cl
12 Aug 2026
Safety

// Actions Speak Louder Than Words // Multilingual agent evaluation compares final answers and throws the trajectory away. The trajectory fi…

DGX agent

// Actions Speak Louder Than Words // Multilingual agent evaluation compares final answers and throws the trajectory away. The trajectory fixes cost, latency, failure mode, and auditability. New resea

safetydair-ai--x
12 Aug 2026
Safety

AdvFD: Boosting Visual Generation via Adversarial Fr'echet Distance Loss

DGX agent

arXiv:2608.11205v1 Announce Type: new Abstract: Frechet distance has recently emerged as an effective distribution-level objective for generator post-training, complementing the conventional sample-le

safetyarxiv-cs-cv
12 Aug 2026
Safety

APCReg: Anatomical-Prior-Guided Coarse-to-Fine CBCT--IOS Registration via Multi-View Projection and Reliability-Controlled Residual Correction

DGX agent

arXiv:2608.09993v1 Announce Type: cross Abstract: Registration between cone-beam computed tomography (CBCT) and intraoral scans (IOS) is essential for patient-specific surgical planning. However, disp

safetyarxiv-cs-cv
12 Aug 2026
Safety

Beyond Forecasting: Recasting Volatility Control as a Routing Problem

DGX agent

arXiv:2608.10375v1 Announce Type: cross Abstract: Volatility control converts risk estimates into portfolio exposure, yet existing approaches often rely on a fixed volatility estimator or a pre-define

safetyarxiv-cs-ai
12 Aug 2026
Safety

Big Tech's reliance on OpenAI and Anthropic for growth is systemically pervasive. Both companies are massively unprofitable — and much of wh…

DGX agent

Big Tech's reliance on OpenAI and Anthropic for growth is systemically pervasive. Both companies are massively unprofitable — and much of what they spend is Big Tech's own money, counted right back as

safetygary-marcus--x
12 Aug 2026
Safety

BooST: Bridging Semantics and Motions for Efficient Skill Transfer

DGX agent

arXiv:2608.10600v1 Announce Type: cross Abstract: Skill abstraction---the process of learning reusable and temporally extended behaviors---has emerged as a key paradigm for improving sample efficiency

safetyarxiv-cs-cv
12 Aug 2026
Model Releases

Boundary-Seeking Policy Gradient for Safe Reinforcement Learning

DGX agent

arXiv:2608.10204v1 Announce Type: new Abstract: Safe reinforcement learning maximizes reward subject to safety constraints. For Constrained Markov Decision Processes, the linear-programming view over

model-releasesarxiv-cs-lg
12 Aug 2026
Safety

CARE: Confidence-Aware Reasoning for Reliable Medical VQA

DGX agent

arXiv:2608.10964v1 Announce Type: cross Abstract: Reinforcement Fine-Tuning (RFT) has enabled medical Multimodal Large Language Models (MLLMs) to produce Chain-of-Thought (CoT) reasoning for visual qu

safetyarxiv-cs-ai
12 Aug 2026
Safety

Carefully Considering Culture: Analyzing LLM Alignment in Single- and Multi-Cultural Settings using Cultural Consensus Theory

DGX agent

arXiv:2608.09937v1 Announce Type: new Abstract: Recent work in NLP has probed large language models for their understanding of cultural norms across countries. However, this work typically considers d

safetyarxiv-cs-cl
12 Aug 2026
Safety

ConRub-Med: Reinforcement Learning with Consensus Rubrics for Open-Ended Medical Question Answering

DGX agent

arXiv:2608.10996v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards has been especially effective in mathematics and coding, where answers can be checked automatically. Many

safetyarxiv-cs-cl
12 Aug 2026
Safety

Convergence of Sign-based Random Reshuffling Algorithms for Nonconvex Optimization

DGX agent

arXiv:2310.15976v4 Announce Type: replace Abstract: signSGD is attractive in nonconvex optimization because it communicates sign-valued rather than full-precision gradients. Several standard analyses

safetyarxiv-cs-lg
12 Aug 2026
Safety

Critic-Free Pretraining for Efficient Online Reinforcement Learning Fine-Tuning

DGX agent

arXiv:2608.10473v1 Announce Type: cross Abstract: Offline-to-online (O2O) reinforcement learning aims to leverage policies pretrained on static datasets while improving them through online interaction

safetyarxiv-cs-ai
12 Aug 2026
Safety

Detecting an Effect Is Not Learning to Act on It: A Reward-SNR Floor for LLM Acquisition Agents

DGX agent

arXiv:2608.10441v1 Announce Type: cross Abstract: Many pipelines can pay a per-example cost to acquire an auxiliary, model-derived observation -- an LLM's structured reasoning, a slow oracle, an expen

safetyarxiv-cs-cl
12 Aug 2026
Safety

DIMOS: Disentangling Instance-level Moving Object Segmentation

DGX agent

arXiv:2606.12826v2 Announce Type: replace-cross Abstract: Moving instance segmentation (MIS) attracts increasing attention due to its broad applications in traffic surveillance, autonomous driving, an

safetyarxiv-cs-ai
12 Aug 2026
Safety

Do AI weather models miss extremes?

DGX agent

arXiv:2608.09972v1 Announce Type: cross Abstract: First-generation AI weather models are often reported to underperform at extremes, mostly in reanalysis-based evaluations of deterministic regression

safetyarxiv-cs-ai
12 Aug 2026
Safety

Do Time-Series Forecasters Use the Right History: Recoverability, Recovery, and Functional Use of Temporal Delays

DGX agent

arXiv:2608.10433v1 Announce Type: new Abstract: Forecast accuracy does not tell us which past inputs produced a prediction. We separate three questions for time-series models with known delay structur

safetyarxiv-cs-lg
12 Aug 2026
Safety

Dual-Loop Self-Evolution via Verifiable Emotion Feedback for Multi-Turn Empathetic Dialogue

DGX agent

arXiv:2608.10626v1 Announce Type: new Abstract: Large language models have demonstrated conversational capabilities, yet empathetic competence remains challenging. Empathetic support is inherently mul

safetyarxiv-cs-cl
12 Aug 2026
Safety

Dual Space Preconditioning for Gradient Descent in the Overparameterized Regime

DGX agent

arXiv:2603.10485v3 Announce Type: replace-cross Abstract: In this work, we study the convergence properties of the Dual Space Preconditioned Gradient Descent, encompassing optimizers such as Normalize

safetyarxiv-cs-lg
12 Aug 2026
Safety

Efficient Hypergradient Descent for Inverse Reinforcement Learning

DGX agent

arXiv:2608.11052v1 Announce Type: new Abstract: Inverse reinforcement learning (IRL) aims to recover a reward function under which the resulting policy reproduces the behavior observed in expert demon

safetyarxiv-cs-lg
12 Aug 2026
Safety

Enhancing Automated Essay Scoring With Three Techniques: Two-Stage Fine-Tuning, Score Alignment, and Self-Training

DGX agent

arXiv:2602.01747v2 Announce Type: replace Abstract: Automated Essay Scoring (AES) plays a crucial role in education by providing scalable and efficient assessment tools. However, in real-world setting

safetyarxiv-cs-cl
12 Aug 2026
Safety

Evaluation-Conditioned Training: Teaching Models to Generalize to Stronger Oversight Regimes

DGX agent

arXiv:2608.10209v1 Announce Type: new Abstract: Feedback signals used to train Large Language Models (LLMs) are the primary driver of their behavior and our main lever for instilling alignment with hu

safetyarxiv-cs-ai
12 Aug 2026
Safety

Every Token Counts: Exact Likert-Scale Distributions for Measuring LLM Attitudes and Biases

DGX agent

arXiv:2608.10503v1 Announce Type: new Abstract: As Large Language Models (LLMs) are increasingly deployed as autonomous agents, accurately evaluating their latent values and biases is critical. The NL

safetyarxiv-cs-cl
12 Aug 2026
Safety

FedCGR: Federated Cross-Domain Generative Recommendation

DGX agent

arXiv:2608.10929v1 Announce Type: new Abstract: Cross-domain recommendation (CDR) transfers preference knowledge across related domains, but federated deployment makes cross-domain alignment difficult

safetyarxiv-cs-ai
12 Aug 2026
Safety

FoR-SALE: Frame of Reference-guided Spatial Adjustment in LLM-based Diffusion Editing

DGX agent

arXiv:2509.23452v2 Announce Type: replace-cross Abstract: Current text-to-image generation models, even state-of-the-art models, exhibit a significant performance gap when spatial expressions are desc

safetyarxiv-cs-cl
12 Aug 2026
Safety

From Prediction to Incrementality: Causal Optimization for Large-Scale Targeting and Recommendation

DGX agent

arXiv:2608.10182v1 Announce Type: cross Abstract: Large-scale targeting and recommendation systems are typically built around predictive scores fed into heuristic or local allocation. When the busines

safetyarxiv-cs-ai
12 Aug 2026
Safety

Generation-Step-Aware Framework for Cross-Modal Representation and Control in Multilingual Speech-Text Models

DGX agent

arXiv:2601.17387v3 Announce Type: replace Abstract: Multilingual speech-text models rely on cross-modal language alignment to transfer knowledge between speech and text, but it remains unclear whether

safetyarxiv-cs-cl
12 Aug 2026
Safety

Hierarchical Empirical-Bayes Naive Bayes: Minimax Smoothing and Calibration with AODE Extension

DGX agent

arXiv:2608.11162v1 Announce Type: new Abstract: The Naive Bayes (NB) classifier remains a standard choice for categorical data, yet its widely used smoothing rules, such as Laplace, Lidstone, Krichevs

safetyarxiv-cs-lg
12 Aug 2026
Safety

Hip Energized Monopedal Hopping

DGX agent

arXiv:2608.10387v1 Announce Type: new Abstract: We present a novel stepping strategy for pitch unlocked planar monopeds where the reaction torques from stabilizing pitch with a conventional PD + feedf

safetyarxiv-cs-ro
12 Aug 2026
Safety

IADD-TR: Intervention-Aware Dynamics Decoupling with Targeted Regularization for Model-Based Reinforcement Learning

DGX agent

arXiv:2608.10634v1 Announce Type: new Abstract: Model-based reinforcement learning (MBRL), which learns environment dynamics to generate synthetic experience, is a promising approach to sample-efficie

safetyarxiv-cs-lg
12 Aug 2026
Safety

INSIDE the Student's Mind: Jointly Modeling Latent Reasoning and Action in LLM Student Simulators

DGX agent

arXiv:2608.10492v1 Announce Type: new Abstract: Large Language Model (LLM)-based simulators often reproduce observable actions but fail to capture the underlying reasoning behind them. In education, w

safetyarxiv-cs-ai
12 Aug 2026
Safety

IO Factory: Simulating AI-Enabled Influence Campaigns at Scale

DGX agent

arXiv:2608.10920v1 Announce Type: new Abstract: We introduce IO Factory, an AI-driven framework for simulating information and influence campaigns as fully integrated, traceable processes. The threat

safetyarxiv-cs-ai
12 Aug 2026
← Previous
1…6566676869…300
Next →