AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
All
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,702 results
Safety

AIFIND: Artifact-Aware Interpreting Fine-Grained Alignment for Incremental Face Forgery Detection

DGX agent

arXiv:2604.16207v1 Announce Type: cross Abstract: As forgery types continue to emerge consistently, Incremental Face Forgery Detection (IFFD) has become a crucial paradigm. However, existing methods t

safetyarxiv-cs-ai
20 Apr 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

Amazon’s AI coding boom is creating a mess. Vibe coding their way to chaos.

DGX agent

Amazon’s AI coding boom is creating a mess. Vibe coding their way to chaos. @GaryMarcus 🦔The internal document was reported by Business Insider. Here’s the link: https://www.businessinsider.com/ai-spr

safetygary-marcus--x
20 Apr 2026
Safety

AutoDrive-R^2: Incentivizing Reasoning and Self-Reflection Capacity for VLA Model in Autonomous Driving

DGX agent

arXiv:2509.01944v3 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models in autonomous driving systems have recently demonstrated transformative potential by integrating multimoda

safetyarxiv-cs-cv
20 Apr 2026
Safety

C-Mining: Unsupervised Discovery of Seeds for Cultural Data Synthesis via Geometric Misalignment

DGX agent

arXiv:2604.15675v1 Announce Type: new Abstract: Achieving cultural alignment in Large Language Models (LLMs) increasingly depends on synthetic data generation. For such synthesis, the most vital initi

safetyarxiv-cs-cl
20 Apr 2026
Safety

CASR: A Robust Cyclic Framework for Arbitrary Large-Scale Super-Resolution with Distribution Alignment and Self-Similarity Awareness

DGX agent

arXiv:2602.22159v2 Announce Type: replace Abstract: Arbitrary-Scale SR (ASISR) remains fundamentally limited by cross-scale distribution shift: once the inference scale leaves the training range, nois

safetyarxiv-cs-cv
20 Apr 2026
Safety

Causal Bootstrapped Alignment for Unsupervised Video-Based Visible-Infrared Person Re-Identification

DGX agent

arXiv:2604.15631v1 Announce Type: new Abstract: VVI-ReID is a critical technique for all-day surveillance, where temporal information provides additional cues beyond static images. However, existing a

safetyarxiv-cs-cv
20 Apr 2026
Safety

Cognitive Agency Surrender: Defending Epistemic Sovereignty via Scaffolded AI Friction

DGX agent

arXiv:2603.21735v2 Announce Type: replace-cross Abstract: The proliferation of Generative Artificial Intelligence has transformed benign cognitive offloading into a systemic risk of cognitive agency s

safetyarxiv-cs-ai
20 Apr 2026
Safety

Comparing the latent features of universal machine-learning interatomic potentials

DGX agent

arXiv:2512.05717v3 Announce Type: replace-cross Abstract: The past few years have seen the development of ``universal'' machine-learning interatomic potentials (uMLIPs) capable of approximating the gr

safetyarxiv-cs-lg
20 Apr 2026
Safety

Concept-wise Attention for Fine-grained Concept Bottleneck Models

DGX agent

arXiv:2604.15748v1 Announce Type: new Abstract: Recently impressive performance has been achieved in Concept Bottleneck Models (CBM) by utilizing the image-text alignment learned by a large pre-traine

safetyarxiv-cs-cv
20 Apr 2026
Safety

Constant-Factor Approximations for Doubly Constrained Fair k-Center, k-Median and k-Means

DGX agent

arXiv:2604.16061v1 Announce Type: cross Abstract: We study discrete k-clustering problems in general metric spaces that are constrained by a combination of two different fairness conditions within the

safetyarxiv-cs-lg
20 Apr 2026
Safety

Contact-Aware Planning and Control of Continuum Robots in Highly Constrained Environments

DGX agent

arXiv:2604.15638v1 Announce Type: new Abstract: Continuum robots are well suited for navigating confined and fragile environments, such as vascular or endoluminal anatomy, where contact with surroundi

safetyarxiv-cs-ro
20 Apr 2026
Safety

Deliberative Searcher: Improving LLM Reliability via Reinforcement Learning with constraints

DGX agent

arXiv:2507.16727v3 Announce Type: replace Abstract: Improving the reliability of large language models (LLMs) is critical for deploying them in real-world scenarios. In this paper, we propose extbf{De

safetyarxiv-cs-ai
20 Apr 2026
Safety

Distribution Shift Alignment Helps LLMs Simulate Survey Response Distributions

DGX agent

arXiv:2510.21977v2 Announce Type: replace Abstract: Large language models (LLMs) offer a promising way to simulate human survey responses, potentially reducing the cost of large-scale data collection.

safetyarxiv-cs-ai
20 Apr 2026
Safety

Dynamic Sampling that Adapts: Self-Aware Iterative Data Persistent Optimization for Mathematical Reasoning

DGX agent

arXiv:2505.16176v2 Announce Type: replace Abstract: In mathematical reasoning, data selection strategies predominantly rely on static, externally defined metrics, which fail to adapt to the evolving c

safetyarxiv-cs-ai
20 Apr 2026
Safety

DyTact: Capturing Dynamic Contacts in Hand-Object Manipulation

DGX agent

arXiv:2506.03103v2 Announce Type: replace Abstract: Reconstructing dynamic hand-object contacts is essential for realistic manipulation in AI character animation, XR, and robotics, yet it remains chal

safetyarxiv-cs-cv
20 Apr 2026
Safety

Elucidating the SNR-t Bias of Diffusion Probabilistic Models

DGX agent

arXiv:2604.16044v1 Announce Type: new Abstract: Diffusion Probabilistic Models have demonstrated remarkable performance across a wide range of generative tasks. However, we have observed that these mo

safetyarxiv-cs-cv
20 Apr 2026
Safety

Enhancing AI and Dynamical Subseasonal Forecasts with Probabilistic Bias Correction

DGX agent

arXiv:2604.16238v1 Announce Type: new Abstract: Decision-makers rely on weather forecasts to plant crops, manage wildfires, allocate water and energy, and prepare for weather extremes. Today, such for

safetyarxiv-cs-lg
20 Apr 2026
Safety

Exploitation Over Exploration: Unmasking the Bias in Linear Bandit Recommender Offline Evaluation

DGX agent

arXiv:2507.18756v2 Announce Type: replace Abstract: Multi-Armed Bandit (MAB) algorithms are widely used in recommender systems that require continuous, incremental learning. A core aspect of MABs is t

safetyarxiv-cs-lg
20 Apr 2026
Safety

Find, Fix, Reason: Context Repair for Video Reasoning

DGX agent

arXiv:2604.16243v1 Announce Type: new Abstract: Reinforcement learning has advanced video reasoning in large multi-modal models, yet dominant pipelines either rely on on-policy self-exploration, which

safetyarxiv-cs-cv
20 Apr 2026
Safety

FineSteer: A Unified Framework for Fine-Grained Inference-Time Steering in Large Language Models

DGX agent

arXiv:2604.15488v1 Announce Type: cross Abstract: Large language models (LLMs) often exhibit undesirable behaviors, such as safety violations and hallucinations. Although inference-time steering offer

safetyarxiv-cs-ai
20 Apr 2026
Safety

Flexible Empowerment at Reasoning with Extended Best-of-N Sampling

DGX agent

arXiv:2604.15614v1 Announce Type: new Abstract: This paper proposes a novel method that incorporates empowerment when reasoning actions in reinforcement learning (RL), thereby achieving the flexibilit

safetyarxiv-cs-lg
20 Apr 2026
Safety

Follow the Flow: On Information Flow Across Textual Tokens in Text-to-Image Models

DGX agent

arXiv:2504.01137v3 Announce Type: replace Abstract: Text-to-image generation models suffer from alignment problems, where generated images fail to accurately capture the objects and relations in the t

safetyarxiv-cs-cl
20 Apr 2026
Safety

From Competition to Coopetition: Coopetitive Training-Free Image Editing Based on Text Guidance

DGX agent

arXiv:2604.15948v1 Announce Type: new Abstract: Text-guided image editing, a pivotal task in modern multimedia content creation, has seen remarkable progress with training-free methods that eliminate

safetyarxiv-cs-cv
20 Apr 2026
Safety

From Intention to Text: AI-Supported Goal Setting in Academic Writing

DGX agent

arXiv:2604.15800v1 Announce Type: cross Abstract: This study presents WriteFlow, an AI voice-based writing assistant designed to support reflective academic writing through goal-oriented interaction.

safetyarxiv-cs-ai
20 Apr 2026
Safety

Geometric regularization of autoencoders via observed stochastic dynamics

DGX agent

arXiv:2604.16282v1 Announce Type: new Abstract: Stochastic dynamical systems with slow or metastable behavior evolve, on long time scales, on an unknown low-dimensional manifold in high-dimensional am

safetyarxiv-cs-lg
20 Apr 2026
Safety

GroupDPO: Memory efficient Group-wise Direct Preference Optimization

DGX agent

arXiv:2604.15602v1 Announce Type: new Abstract: Preference optimization is widely used to align Large Language Models (LLMs) with preference feedback. However, most existing methods train on a single

safetyarxiv-cs-cl
20 Apr 2026
Safety

Hierarchical Codec Diffusion for Video-to-Speech Generation

DGX agent

arXiv:2604.15923v1 Announce Type: cross Abstract: Video-to-Speech (VTS) generation aims to synthesize speech from a silent video without auditory signals. However, existing VTS methods disregard the h

safetyarxiv-cs-cv
20 Apr 2026
Safety

How people use Copilot for Health

DGX agent

arXiv:2604.15331v1 Announce Type: cross Abstract: We analyze over 500,000 de-identified health-related conversations with Microsoft Copilot from January 2026 to characterize what people ask conversati

safetyarxiv-cs-ai
20 Apr 2026
Safety

Import AI 454: Automating alignment research; safety study of a Chinese model; HiFloat4

DGX agent

This newsletter covers three main topics: advances in automating alignment research to improve AI safety processes, a safety evaluation study of a Chinese AI model, and technical details about HiFloat

safetyimport-ai
20 Apr 2026
Safety

Importantly, the tweet by @itsolelehmann may be misleading. I stand by what I said but have some reason to think that @AmandaAskell’s views …

DGX agent

Importantly, the tweet by @itsolelehmann may be misleading. I stand by what I said but have some reason to think that @AmandaAskell’s views may have been misrepresented. Seeking clarification and will

safetygary-marcus--x
20 Apr 2026
Safety

Improving Reasoning Capabilities in Small Models through Mixture-of-Layers Distillation with Stepwise Attention on Key Information

DGX agent

arXiv:2604.15701v1 Announce Type: new Abstract: The significant computational demands of large language models have increased interest in distilling reasoning abilities into smaller models via Chain-o

safetyarxiv-cs-cl
20 Apr 2026
Safety

Information-Consistent Language Model Recommendations through Group Relative Policy Optimization

DGX agent

arXiv:2512.12858v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly deployed in business-critical domains such as finance, education, healthcare, and customer suppo

safetyarxiv-cs-ai
20 Apr 2026
Safety

Jailbreak Scaling Laws for Large Language Models: Polynomial-Exponential Crossover

DGX agent

arXiv:2603.11331v2 Announce Type: replace-cross Abstract: Adversarial attacks can reliably steer safety-aligned large language models toward unsafe behavior. Empirically, we find that strong adversari

safetyarxiv-cs-ai
20 Apr 2026
Safety

Je suis passé à Découverte de @CBCRadioCanada pour discuter des risques de l’IA, des raisons scientifiques qui expliquent certains des compo…

DGX agent

Je suis passé à Découverte de @CBCRadioCanada pour discuter des risques de l’IA, des raisons scientifiques qui expliquent certains des comportements inquiétants des modèles de pointe, et des solutions

safetyyoshua-bengio--x
20 Apr 2026
Safety

Joint-Centric Dual Contrastive Alignment with Structure-Preserving and Information-Balanced Regularization

DGX agent

arXiv:2604.16247v1 Announce Type: cross Abstract: We propose HILBERT (HIerarchical Long-sequence Balanced Embedding with Reciprocal contrastive Training), a cross-attentive multimodal framework for le

safetyarxiv-cs-ai
20 Apr 2026
Safety

Language Models as Semantic Teachers: Post-Training Alignment for Medical Audio Understanding

DGX agent

arXiv:2512.04847v2 Announce Type: replace-cross Abstract: Pre-trained audio models excel at detecting acoustic patterns in auscultation sounds but often fail to grasp their clinical significance, limi

safetyarxiv-cs-ai
20 Apr 2026
Safety

Language, Place, and Social Media: Geographic Dialect Alignment in New Zealand

DGX agent

arXiv:2604.15744v1 Announce Type: new Abstract: This thesis investigates geographic dialect alignment in place-informed social media communities, focussing on New Zealand-related Reddit communities. B

safetyarxiv-cs-cl
20 Apr 2026
Safety

Large Language Models for Market Research: A Data-augmentation Approach

DGX agent

arXiv:2412.19363v3 Announce Type: replace Abstract: Large Language Models (LLMs) have transformed artificial intelligence by excelling in complex natural language processing tasks. Their ability to ge

safetyarxiv-cs-ai
20 Apr 2026
Safety

Learning to Look before Learning to Like: Incorporating Human Visual Cognition into Aesthetic Quality Assessment

DGX agent

arXiv:2604.15853v1 Announce Type: new Abstract: Automated Aesthetic Quality Assessment (AQA) treats images primarily as static pixel vectors, aligning predictions with human-rating scores largely thro

safetyarxiv-cs-cv
20 Apr 2026
Safety

literally my basic model since 1998. crazy that some people still haven’t figured this out.

DGX agent

literally my basic model since 1998. crazy that some people still haven’t figured this out. My basic model of capabilities: LLMs are good at problems similar to those that appear in their training dat

safetygary-marcus--x
20 Apr 2026
Safety

LLMs are “ good at math style problems, where you tell them A, B and C are true, and then ask them to figure out D [but] extremely bad at an…

DGX agent

LLMs are “ good at math style problems, where you tell them A, B and C are true, and then ask them to figure out D [but] extremely bad at anything involving what I would call mature scholarship .. [so

safetygary-marcus--x
20 Apr 2026
Safety

Long-Term Memory for VLA-based Agents in Open-World Task Execution

DGX agent

arXiv:2604.15671v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have demonstrated significant potential for embodied decision-making; however, their application in complex chemical

safetyarxiv-cs-ro
20 Apr 2026
Safety

M3R: Localized Rainfall Nowcasting with Meteorology-Informed MultiModal Attention

DGX agent

arXiv:2604.15377v1 Announce Type: cross Abstract: Accurate and timely rainfall nowcasting is crucial for disaster mitigation and water resource management. Despite recent advances in deep learning, pr

safetyarxiv-cs-cv
20 Apr 2026
Safety

MFC-RFNet: A Multi-scale Guided Rectified Flow Network for Radar Sequence Prediction

DGX agent

arXiv:2601.03633v2 Announce Type: replace-cross Abstract: Accurate and high-resolution precipitation nowcasting from radar echo sequences is crucial for disaster mitigation and economic planning, yet

safetyarxiv-cs-ai
20 Apr 2026
Safety

Multi-objective Reinforcement Learning With Augmented States Requires Rewards After Deployment

DGX agent

arXiv:2604.15757v1 Announce Type: new Abstract: This research note identifies a previously overlooked distinction between multi-objective reinforcement learning (MORL), and more conventional single-ob

safetyarxiv-cs-lg
20 Apr 2026
Safety

On the Rejection Criterion for Proxy-based Test-time Alignment

DGX agent

arXiv:2604.16146v1 Announce Type: new Abstract: Recent works proposed test-time alignment methods that rely on a small aligned model as a proxy that guides the generation of a larger base (unaligned)

safetyarxiv-cs-cl
20 Apr 2026
Safety

On theCUBE Pod: IBM’s AI strategy, infrastructure bottlenecks and ecosystem partnerships reshape markets

DGX agent

Artificial intelligence infrastructure is now the deciding force behind enterprise competitiveness. What was once a backend concern has moved directly into the center of business strategy, shaping how

safetysiliconangle
20 Apr 2026
Safety

One-Shot Cross-Geometry Skill Transfer through Part Decomposition

DGX agent

arXiv:2604.15455v1 Announce Type: new Abstract: Given a demonstration, a robot should be able to generalize a skill to any object it encounters-but existing approaches to skill transfer often fail to

safetyarxiv-cs-ro
20 Apr 2026
← Previous
1…239240241242243…265
Next →