AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,708 results
12 May 2026

Happy One Year Anniversary, AGI!* What an amazing, hallucination-free year it has been!* Glad we can totally trust you now!* *I was being sa…

SafetyDGX agent

Happy One Year Anniversary, AGI!* What an amazing, hallucination-free year it has been!* Glad we can totally trust you now!* *I was being sarcastic. @codeslubber tyler cowen literally said o3 was AGI.

HapticLDM: A Diffusion Model for Text-to-Vibrotactile Generation

SafetyDGX agent

arXiv:2605.09971v1 Announce Type: cross Abstract: Text-to-vibration generation converts natural language into haptic feedback, enabling vibration-effect designers to get scenarios-fitted vibrations mo

Hierarchical Causal Abduction: A Foundation Framework for Explainable Model Predictive Control

SafetyDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.10624v1 Announce Type: new Abstract: Model Predictive Control (MPC) is widely used to operate safety-critical infrastructure by predicting future trajectories and optimizing control actions

Hierarchical End-to-End Taylor Bounds for Complete Neural Network Verification

SafetyDGX agent

arXiv:2605.10621v1 Announce Type: new Abstract: Reachability analysis of neural networks, which seeks to compute or bound the set of outputs attainable over a given input domain, is central to certify

HLGFA: High-Low Resolution Guided Feature Alignment for Unsupervised Anomaly Detection

SafetyDGX agent

arXiv:2602.09524v3 Announce Type: replace Abstract: Unsupervised industrial anomaly detection (UAD) is essential for modern manufacturing inspection, where defect samples are scarce and reliable detec

How LLMs Are Persuaded: A Few Attention Heads, Rerouted

SafetyDGX agent

arXiv:2605.09314v1 Announce Type: new Abstract: Language models can be persuaded to abandon factual knowledge. This vulnerability is central to AI safety, but its internal mechanism remains poorly und

How Much is Brain Data Worth for Machine Learning?

SafetyDGX agent

arXiv:2605.09243v1 Announce Type: new Abstract: If a person can solve a task, can measuring their brain make it easier to train a model to solve that task too? Recent NeuroAI work suggests that supple

HTPO: Towards Exploration-Exploitation Balanced Policy Optimization via Hierarchical Token-level Objective Control

SafetyDGX agent

arXiv:2605.08283v1 Announce Type: cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has emerged as a pivotal technique for enhancing the reasoning capabilities of Large Language Mo

https://x.com/xiaochuan8688/status/2053007309459341699 'ByteDance has quietly shut down 30% of its AI projects' Even ByteDance faces the sam…

SafetyDGX agent

https://x.com/xiaochuan8688/status/2053007309459341699 'ByteDance has quietly shut down 30% of its AI projects' Even ByteDance faces the same fate as OpenAI's Sora. While the Chinese public enjoys 'AI

HY-Himmel Technical Report: Hierarchical Interleaved Multi-stream Motion Encoding for Long Video Understanding

SafetyDGX agent

arXiv:2605.08158v1 Announce Type: cross Abstract: Long-video understanding with multimodal language models suffers from three compounding bottlenecks: heavy decode cost to obtain dense RGB frames, qua

HyNeuralMap: Hyperbolic Mapping of Visual Semantics to Neural Hierarchies

SafetyDGX agent

arXiv:2605.09392v1 Announce Type: new Abstract: Understanding the intricate mappings between visual stimuli and neural responses is a fundamental challenge in cognitive neuroscience. While current app

HyPER: Bridging Exploration and Exploitation for Scalable LLM Reasoning with Hypothesis Path Expansion and Reduction

SafetyDGX agent

arXiv:2602.06527v2 Announce Type: replace Abstract: Scaling test-time compute with multi-path chain-of-thought improves reasoning accuracy, but its effectiveness depends critically on the exploration-

Hyperspherical Autoencoder for High-Fidelity Image Reconstruction and Generation

SafetyDGX agent

arXiv:2601.22904v2 Announce Type: replace-cross Abstract: Recent studies have explored using pretrained Vision Foundation Models (VFMs) such as DINO for generative autoencoders, showing strong generat

I sat down with @JonHernandezIA in Madrid to discuss the growing risks and impacts of AI and the urgent need to improve our social, politica…

SafetyDGX agent

I sat down with @JonHernandezIA in Madrid to discuss the growing risks and impacts of AI and the urgent need to improve our social, political, and technical safeguards. Thanks for an excellent convers

Improving Human Image Animation via Semantic Representation Alignment

SafetyDGX agent

arXiv:2605.10523v1 Announce Type: new Abstract: The field of image-to-video generation has made remarkable progress. However, challenges such as human limb twisting and facial distortion persist, espe

Improving Lexical Difficulty Prediction with Context-Aligned Contrastive Learning and Ridge Ensembling

SafetyDGX agent

arXiv:2605.08950v1 Announce Type: cross Abstract: Lexical difficulty prediction is a fundamental problem in language learning and readability assessment, requiring models to estimate word difficulty a

Improving Text-to-Image Generation with Intrinsic Self-Confidence Rewards

SafetyDGX agent

arXiv:2603.00918v3 Announce Type: replace-cross Abstract: Text-to-image generation powers content creation across design, media, and data augmentation. Post-training of text-to-image generative models

Instruction Anchor: Dissecting the Mechanistic Dynamics of Modality Arbitration

SafetyDGX agent

arXiv:2602.03677v2 Announce Type: replace Abstract: Modality following is the ability to selectively leverage multimodal contexts based on user instructions. It is fundamental to the safety and reliab

Intelligent Autonomous Orchestration for Distributed Cloud Resources using Complex-Stability Analysis

SafetyDGX agent

arXiv:2605.08139v1 Announce Type: cross Abstract: In modern distributed cloud environments, efficient resource allocation is required as traditional scaling mechanisms are often subject to cloud thras

Interactive Inverse Reinforcement Learning of Interaction Scenarios via Bi-level Optimization

SafetyDGX agent

arXiv:2605.08131v1 Announce Type: new Abstract: Inverse reinforcement learning (IRL) learns a reward function and a corresponding policy that best fit the demonstration data of an expert. However, in

Internalizing Safety Understanding in Large Reasoning Models via Verification

SafetyDGX agent

arXiv:2605.08930v1 Announce Type: new Abstract: While explicit Chain-of-Thought (CoT) empowers large reasoning models (LRMs), it enables the generation of riskier final answers. Current alignment para

Investigating Anisotropy in Visual Grounding under Controlled Counterfactual Perturbations

SafetyDGX agent

arXiv:2605.09090v1 Announce Type: cross Abstract: Visual Grounding benchmarks assume that the object described by a referring expression is always present in the image, and grounding models are theref

Is Class Signal Clustered or Routed in Task-Induced Implicit Neural Representation Weight Spaces?

SafetyDGX agent

arXiv:2605.08281v1 Announce Type: new Abstract: Implicit neural representations (INRs) encode images as neural-network weights, making image classification a problem of weight-space classifiability. A

Ister: Linear Transformer for Efficient Multivariate Time Series Forecasting

SafetyDGX agent

arXiv:2412.18798v3 Announce Type: replace-cross Abstract: Transformer-based models have achieved remarkable success in multivariate time series forecasting (MTSF) by capturing long-range dependencies.

Iterative Critique-and-Routing Controller for Multi-Agent Systems with Heterogeneous LLMs

SafetyDGX agent

arXiv:2605.08686v1 Announce Type: new Abstract: Multi-agent large language model (LLM) systems often rely on a controller to coordinate a pool of heterogeneous models, yet existing controllers are typ

KeyframeFace: Language-Driven Facial Animation via Semantic Keyframes

SafetyDGX agent

arXiv:2512.11321v3 Announce Type: replace Abstract: Facial animation is a core component for creating digital characters in Computer Graphics (CG) industry. A typical production workflow relies on spa

Language Conditioned Multi-Finger Dexterous Manipulation Enabled by Physical Compliance and Switching of Controllers

SafetyDGX agent

arXiv:2410.14022v2 Announce Type: replace-cross Abstract: Human dexterity arises from combining high-level task reasoning with finger-level dexterity control and physical compliance at the muscle and

LAQuant: A Simple Overhead-free Large Reasoning Model Quantization by Layer-wise Lookahead Loss

SafetyDGX agent

arXiv:2605.08755v1 Announce Type: new Abstract: Large reasoning models (LRMs) reach competition-level math and coding accuracy via long autoregressive decoding, making per-token decoding cost a primar

Large Language Models for Sequential Decision-Making: Improving In-Context Learning via Supervised Fine-Tuning

SafetyDGX agent

arXiv:2605.09009v1 Announce Type: cross Abstract: Large language models (LLMs) have shown remarkable in-context learning (ICL) capabilities, yet their potential for sequential decision-making remains

Latent Personality Alignment: Improving Harmlessness Without Mentioning Harms

SafetyDGX agent

arXiv:2605.08496v1 Announce Type: new Abstract: Current adversarial robustness methods for large language models require extensive datasets of harmful prompts (thousands to hundreds of thousands of ex

LaWM: Least Action World Models for Long-Horizon Physical Consistency from Visual Observations

SafetyDGX agent

arXiv:2605.08279v1 Announce Type: cross Abstract: Learning predictive world models from visual observations is a core problem in embodied AI, with applications to model-based reinforcement learning an

Learning Approximate Nash Equilibria in Cooperative Multi-Agent Reinforcement Learning via Mean-Field Subsampling

SafetyDGX agent

arXiv:2603.03759v2 Announce Type: replace-cross Abstract: Many large-scale platforms and networked control systems have a centralized decision maker interacting with a massive population of agents und

Learning the Preferences of a Learning Agent

SafetyDGX agent

arXiv:2605.09217v1 Announce Type: new Abstract: For AI systems to be useful to humans, they must understand and act in accordance with our values and preferences. Since specifying preferences is a har

Learning to Align Generative Appearance Priors for Fine-grained Image Retrieval

SafetyDGX agent

arXiv:2605.09859v1 Announce Type: new Abstract: Fine-grained image retrieval (FGIR) typically relies on supervision from seen categories to learn discriminative embeddings for retrieving unseen catego

Learning to Compress Time-to-Control: A Reinforcement Learning Framework for Chronic Disease Management

SafetyDGX agent

arXiv:2605.09818v1 Announce Type: new Abstract: Reinforcement learning (RL) in healthcare has had mixed results, with reward sparsity, unreliable off-policy evaluation, and deployment-simulation gap a

Learning to Explore: Scaling Agentic Reasoning via Exploration-Aware Policy Optimization

SafetyDGX agent

arXiv:2605.08978v1 Announce Type: new Abstract: Recent advancements in agentic test-time scaling allow models to gather environmental feedback before committing to final actions. A key limitation of e

Learning to Stay Safe: Adaptive Regularization Against Safety Degradation during Fine-Tuning

SafetyDGX agent

arXiv:2602.17546v2 Announce Type: replace Abstract: Instruction-following language models are trained to be helpful and safe, yet their safety behavior can deteriorate under benign fine-tuning and wor

Learning When to Jump for Off-road Navigation

SafetyDGX agent

arXiv:2602.00877v2 Announce Type: replace Abstract: Low speed does not always guarantee safety in off-road driving. For instance, crossing a ditch may be risky at a low speed due to the risk of gettin

Learning When to Stop: Selective Imitation Learning Under Arbitrary Dynamics Shift

SafetyDGX agent

arXiv:2605.09183v1 Announce Type: new Abstract: Behavior cloning provides strong imitation learning guarantees when training and test environments share the same dynamics. However, in many deployment

Let the Target Select for Itself: Data Selection via Target-Aligned Paths

SafetyDGX agent

arXiv:2605.09404v1 Announce Type: cross Abstract: Targeted data selection aims to identify training samples from a large candidate pool that improve performance on a specific downstream task. Many rec

LiLAW: Lightweight Learnable Adaptive Weighting to Learn Sample Difficulty & Improve Noisy Training

SafetyDGX agent

arXiv:2509.20786v3 Announce Type: replace Abstract: Training deep neural networks with noise and data heterogeneity is a major challenge. We introduce Lightweight Learnable Adaptive Weighting (LiLAW),

Liouville PDE-based sliced-Wasserstein flow

SafetyDGX agent

arXiv:2505.17204v3 Announce Type: replace-cross Abstract: The sliced Wasserstein flow (SWF), a nonparametric and implicit generative gradient flow, is transformed into a Liouville partial differential

LLM Advertisement based on Neuron Auctions

SafetyDGX agent

arXiv:2605.08326v1 Announce Type: cross Abstract: As Large Language Models (LLMs) transition into conversational agents, generative advertising emerges as a crucial monetization strategy. However, emb

LLM-Agnostic Semantic Representation Attack

SafetyDGX agent

arXiv:2605.08898v1 Announce Type: cross Abstract: Large Language Models (LLMs) increasingly employ alignment techniques to prevent harmful outputs. Despite these safeguards, attackers can circumvent t

LLM-Driven Performance-Space Augmentation for Meta-Learning-Based Algorithm Selection

SafetyDGX agent

arXiv:2605.09518v1 Announce Type: new Abstract: Meta-learning for algorithm selection relies on a meta-dataset in which each row corresponds to a supervised learning dataset described by meta-features

Long-Horizon Q-Learning: Accurate Value Learning via n-Step Inequalities

SafetyDGX agent

arXiv:2605.05812v2 Announce Type: replace Abstract: Off-policy, value-based reinforcement learning methods such as Q-learning are appealing because they can learn from arbitrary experience, including

LoopVLA: Learning Sufficiency in Recurrent Refinement for Vision-Language-Action Models

SafetyDGX agent

arXiv:2605.09948v1 Announce Type: new Abstract: Current Vision-Language-Action (VLA) models typically treat the deepest representation of a vision-language backbone as universally optimal for action p

M^3: Reframing Training Measures for Discretized Physical Simulations

SafetyDGX agent

arXiv:2605.08843v1 Announce Type: new Abstract: Neural surrogate models for physical simulations are trained on discretized samples of continuous domains, where the induced empirical measure leads to

Machine Unlearning on Pre-trained Models by Residual Feature Alignment Using LoRA

SafetyDGX agent

arXiv:2411.08443v2 Announce Type: replace-cross Abstract: Machine unlearning is an emerging technology that removes a subset of the training data from a trained model without significantly affecting t

MAG-VLAQ: Multi-modal Aerial-Ground Query Aggregation for Cross-View Place Recognition

SafetyDGX agent

arXiv:2605.09418v1 Announce Type: new Abstract: Multi-modal cross-view place recognition remains a fundamental challenge in computer vision and robotics due to the severe viewpoint, modality, and spat

Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction

SafetyDGX agent

arXiv:2605.09649v1 Announce Type: new Abstract: The key-value (KV) cache is a major bottleneck in long-context inference, where memory and computation grow with sequence length. Existing KV eviction m

MapFormer: Self-Supervised Learning of Cognitive Maps with Input-Dependent Positional Embeddings

SafetyDGX agent

arXiv:2511.19279v4 Announce Type: replace-cross Abstract: A cognitive map is an internal model which encodes the abstract relationships among entities in the world, giving humans and animals the flexi

“Marcus's repeated warnings about the 'wall of generalization' since 1998 have once again been proven true.”

SafetyDGX agent

“Marcus's repeated warnings about the 'wall of generalization' since 1998 have once again been proven true.” Marcus氏が1998年から繰り返す'汎化の壁'の警告。またも証明された。完璧なAIを待つより、今の限界を熟知して使いこなすチームが勝つ。少人数ゆえの意思決定の速さとリスク許容度が

MARLaaS: Multi-Tenant Asynchronous Reinforcement Learning as a Service

SafetyDGX agent

arXiv:2605.08527v1 Announce Type: cross Abstract: Reinforcement Learning from Verifiable Rewards (RLVR) has significantly improved the reasoning capabilities of large language models (LLMs), particula

MARS-SQL: A multi-agent reinforcement learning framework for Text-to-SQL

SafetyDGX agent

arXiv:2511.01008v2 Announce Type: replace Abstract: Large Language Models (LLMs) often struggle with the precise logic and schema alignment required for complex Text-to-SQL tasks. While current method

MASS-DPO: Multi-negative Active Sample Selection for Direct Policy Optimization

SafetyDGX agent

arXiv:2605.10784v1 Announce Type: new Abstract: Multi-negative preference optimization under the Plackett--Luce (PL) model extends Direct Preference Optimization (DPO) by leveraging comparative signal

Mechanism Design Is Not Enough: Prosocial Agents for Cooperative AI

SafetyDGX agent

arXiv:2605.08426v1 Announce Type: cross Abstract: Ensuring that AI agents behave safely and beneficially when interacting with other parties has emerged as one of the central challenges of modern AI s

MedFL-Stress: A Systematic Robustness Evaluation of Federated Brain Tumor Segmentation under Cross-Hospital MRI Appearance Shift

SafetyDGX agent

arXiv:2605.09025v1 Announce Type: new Abstract: Federated learning enables hospitals to collaboratively train segmentation models without sharing patient data. However, current evaluation protocols re

Mem-W: Latent Memory-Native GUI Agents

SafetyDGX agent

arXiv:2605.09317v1 Announce Type: new Abstract: GUI agents are beginning to operate the web, mobile, and desktop as interactive worlds, where successful control depends on carrying forward visual, pro

Memory Inception: Latent-Space KV Cache Manipulation for Steering LLMs

SafetyDGX agent

arXiv:2605.06225v2 Announce Type: replace-cross Abstract: Steering large language models (LLMs) is usually done by either instruction prompting or activation steering. Prompting often gives strong con

← Previous
1…149150151152153…212
Next →