AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,351 results
Safety

Grok just hit its highest monthly traffic EVER: over 326 MILLION visits in March alone That’s a massive 61% jump YoY and up 9.3% just since …

DGX agent

Grok just hit its highest monthly traffic EVER: over 326 MILLION visits in March alone That’s a massive 61% jump YoY and up 9.3% just since February People love Grok because it's the only AI they can

safetyelon-musk--x
15 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

GUIDE: Guided Updates for In-context Decision Evolution in LLM-Driven Spacecraft Operations

DGX agent

arXiv:2603.27306v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have been proposed as supervisory agents for spacecraft operations, but existing approaches rely on static prompt

safetyarxiv-cs-ai
15 Apr 2026
Safety

Hail to the Thief: Exploring Attacks and Defenses in Decentralised GRPO

DGX agent

arXiv:2511.09780v2 Announce Type: replace Abstract: Group Relative Policy Optimization (GRPO) has demonstrated wide adoption in the post-training of Large Language Models (LLMs). In GRPO, prompts are

safetyarxiv-cs-lg
15 Apr 2026
Safety

HiCoLoRA: Addressing Context-Prompt Misalignment via Hierarchical Collaborative LoRA for Zero-Shot DST

DGX agent

arXiv:2509.19742v4 Announce Type: replace-cross Abstract: Zero-shot Dialog State Tracking (zs-DST) is essential for enabling Task-Oriented Dialog Systems (TODs) to generalize to new domains without co

safetyarxiv-cs-ai
15 Apr 2026
Safety

Human-Centric Topic Modeling with Goal-Prompted Contrastive Learning and Optimal Transport

DGX agent

arXiv:2604.12663v1 Announce Type: new Abstract: Existing topic modeling methods, from LDA to recent neural and LLM-based approaches, which focus mainly on statistical coherence, often produce redundan

safetyarxiv-cs-ai
15 Apr 2026
Safety

Hybrid-AIRL: Enhancing Inverse Reinforcement Learning with Supervised Expert Guidance

DGX agent

arXiv:2511.21356v2 Announce Type: replace-cross Abstract: Adversarial Inverse Reinforcement Learning (AIRL) has shown promise in addressing the sparse reward problem in reinforcement learning (RL) by

safetyarxiv-cs-ai
15 Apr 2026
Safety

I find it fascinating that OpenAI's Global Affairs team will denounce AI doomers for being too negative and then, simultaneously, advocate f…

DGX agent

I find it fascinating that OpenAI's Global Affairs team will denounce AI doomers for being too negative and then, simultaneously, advocate for an Illinois bill that would shield them from liability if

safetygary-marcus--x
15 Apr 2026
Safety

I spent some time trying to distill all the complex factors impacting open models -- economics, capabilities, distribution, policy, etc. -- …

DGX agent

I spent some time trying to distill all the complex factors impacting open models -- economics, capabilities, distribution, policy, etc. -- into a clear list of beliefs. Here they are in full. 1. It’s

safetyyann-lecun--x
15 Apr 2026
Safety

Incentivizing High-Quality Human Annotations with Golden Questions

DGX agent

arXiv:2505.19134v2 Announce Type: replace-cross Abstract: Human-annotated data plays a vital role in training large language models (LLMs), such as supervised fine-tuning and human preference alignmen

safetyarxiv-cs-lg
15 Apr 2026
Safety

Information-Geometric Decomposition of Generalization Error in Unsupervised Learning

DGX agent

arXiv:2604.12340v1 Announce Type: cross Abstract: We decompose the Kullback--Leibler generalization error (GE) -- the expected KL divergence from the data distribution to the trained model -- of unsup

safetyarxiv-cs-lg
15 Apr 2026
Safety

InsightFlow: LLM-Driven Synthesis of Patient Narratives for Mental Health into Causal Models

DGX agent

arXiv:2604.12721v1 Announce Type: new Abstract: Clinical case formulation organizes patient symptoms and psychosocial factors into causal models, often using the 5P framework. However, constructing su

safetyarxiv-cs-cl
15 Apr 2026
Safety

Labeled TrustSet Guided: Batch Active Learning with Reinforcement Learning

DGX agent

arXiv:2604.12303v1 Announce Type: new Abstract: Batch active learning (BAL) is a crucial technique for reducing labeling costs and improving data efficiency in training large-scale deep learning model

safetyarxiv-cs-lg
15 Apr 2026
Safety

Learning step-level dynamic soaring in shear flow

DGX agent

arXiv:2604.12413v1 Announce Type: cross Abstract: Dynamic soaring enables sustained flight by extracting energy from wind shear, yet it is commonly understood as a cycle-level maneuver that assumes st

safetyarxiv-cs-ro
15 Apr 2026
Safety

Learning Versatile Humanoid Manipulation with Touch Dreaming

DGX agent

arXiv:2604.13015v1 Announce Type: new Abstract: Humanoid robots promise general-purpose assistance, yet real-world humanoid loco-manipulation remains challenging because it requires whole-body stabili

safetyarxiv-cs-ro
15 Apr 2026
Safety

Lightning OPD: Efficient Post-Training for Large Reasoning Models with Offline On-Policy Distillation

DGX agent

arXiv:2604.13010v1 Announce Type: cross Abstract: On-policy distillation (OPD) has emerged as an efficient post-training paradigm for large language models. However, standard OPD requires a live teach

safetyarxiv-cs-ai
15 Apr 2026
Safety

LiveMoments: Reselected Key Photo Restoration in Live Photos via Reference-guided Diffusion

DGX agent

arXiv:2604.12286v1 Announce Type: new Abstract: Live Photo captures both a high-quality key photo and a short video clip to preserve the precious dynamics around the captured moment. While users may c

safetyarxiv-cs-cv
15 Apr 2026
Safety

Man and machine: artificial intelligence and judicial decision making

DGX agent

arXiv:2603.19042v4 Announce Type: replace Abstract: The integration of artificial intelligence (AI) technologies into judicial decision-making, particularly in pretrial, sentencing, and parole context

safetyarxiv-cs-ai
15 Apr 2026
Safety

Meet Dynamic Individual Preferences: Resolving Conflicting Human Value with Paired Fine-Tuning

DGX agent

arXiv:2604.12479v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) have significantly improved the alignment of models with general human preferences. However, a major cha

safetyarxiv-cs-cl
15 Apr 2026
Safety

Models Know Their Shortcuts: Deployment-Time Shortcut Mitigation

DGX agent

arXiv:2604.12277v1 Announce Type: new Abstract: Pretrained language models often rely on superficial features that appear predictive during training yet fail to generalize at test time, a phenomenon k

safetyarxiv-cs-lg
15 Apr 2026
Safety

MODIX: A Training-Free Multimodal Information-Driven Positional Index Scaling for Vision-Language Models

DGX agent

arXiv:2604.12537v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have achieved remarkable progress in multimodal understanding, yet their positional encoding mechanisms remain suboptima

safetyarxiv-cs-ai
15 Apr 2026
Safety

MolMem: Memory-Augmented Agentic Reinforcement Learning for Sample-Efficient Molecular Optimization

DGX agent

arXiv:2604.12237v1 Announce Type: cross Abstract: In drug discovery, molecular optimization aims to iteratively refine a lead compound to improve molecular properties while preserving structural simil

safetyarxiv-cs-ai
15 Apr 2026
Safety

Mutual Information Surprise: Rethinking Unexpectedness in Autonomous Systems

DGX agent

arXiv:2508.17403v3 Announce Type: replace Abstract: A community of researchers appears to think that a machine can be surprised and have introduced various surprise measures, principally the Shannon S

safetyarxiv-cs-lg
15 Apr 2026
Safety

No More Stale Feedback: Co-Evolving Critics for Open-World Agent Learning

DGX agent

arXiv:2601.06794v2 Announce Type: replace Abstract: Critique-guided reinforcement learning (RL) has emerged as a powerful paradigm for training LLM agents by augmenting sparse outcome rewards with nat

safetyarxiv-cs-ai
15 Apr 2026
Safety

Not All Turns Are Equally Hard: Adaptive Thinking Budgets For Efficient Multi-Turn Reasoning

DGX agent

arXiv:2604.05164v2 Announce Type: replace-cross Abstract: As LLM reasoning performance plateau, improving inference-time compute efficiency is crucial to mitigate overthinking and long thinking traces

safetyarxiv-cs-ai
15 Apr 2026
Safety

Offline-Online Reinforcement Learning for Linear Mixture MDPs

DGX agent

arXiv:2604.11994v1 Announce Type: new Abstract: We study offline-online reinforcement learning in linear mixture Markov decision processes (MDPs) under environment shift. In the offline phase, data ar

safetyarxiv-cs-lg
15 Apr 2026
Safety

PAINT: Partner-Agnostic Intent-Aware Cooperative Transport with Legged Robots

DGX agent

arXiv:2604.12852v1 Announce Type: new Abstract: Collaborative transport requires robots to infer partner intent through physical interaction while maintaining stable loco-manipulation. This becomes pa

safetyarxiv-cs-ro
15 Apr 2026
Safety

Perception-Aware Policy Optimization for Multimodal Reasoning

DGX agent

arXiv:2507.06448v5 Announce Type: replace Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has proven to be a highly effective strategy for endowing Large Language Models (LLMs) with ro

safetyarxiv-cs-cl
15 Apr 2026
Safety

PR-MaGIC: Prompt Refinement Via Mask Decoder Gradient Flow For In-Context Segmentation

DGX agent

arXiv:2604.12113v1 Announce Type: cross Abstract: Visual Foundation Models (VFMs) such as the Segment Anything Model (SAM) have significantly advanced broad use of image segmentation. However, SAM and

safetyarxiv-cs-ai
15 Apr 2026
Safety

Prediction from a year ago about AI backlash that looks to be on track:

DGX agent

Prediction from a year ago about AI backlash that looks to be on track: Public backlash against AI will so be strong by 2028 that anti-AI sentiment will probably be a major factor in the 2028 US Presi

safetygary-marcus--x
15 Apr 2026
Safety

Progress-Think: Semantic Progress Reasoning for Vision-Language Navigation

DGX agent

arXiv:2511.17097v2 Announce Type: replace Abstract: Vision-Language Navigation requires agents to act coherently over long horizons by understanding not only local visual context but also how far they

safetyarxiv-cs-ro
15 Apr 2026
Safety

PubSwap: Public-Data Off-Policy Coordination for Federated RLVR

DGX agent

arXiv:2604.12160v1 Announce Type: new Abstract: Reasoning post-training with reinforcement learning from verifiable rewards (RLVR) is typically studied in centralized settings, yet many realistic appl

safetyarxiv-cs-lg
15 Apr 2026
Safety

Redefining Quality Criteria and Distance-Aware Score Modeling for Image Editing Assessment

DGX agent

arXiv:2604.12175v1 Announce Type: new Abstract: Recent advances in image editing have heightened the need for reliable Image Editing Quality Assessment (IEQA). Unlike traditional methods, IEQA require

safetyarxiv-cs-cv
15 Apr 2026
Safety

Relaxing Anchor-Frame Dominance for Mitigating Hallucinations in Video Large Language Models

DGX agent

arXiv:2604.12582v1 Announce Type: new Abstract: Recent Video Large Language Models (Video-LLMs) have demonstrated strong capability in video understanding, yet they still suffer from hallucinations. E

safetyarxiv-cs-cv
15 Apr 2026
Safety

Rethinking On-Policy Distillation of Large Language Models: Phenomenology, Mechanism, and Recipe

DGX agent

arXiv:2604.13016v1 Announce Type: cross Abstract: On-policy distillation (OPD) has become a core technique in the post-training of large language models, yet its training dynamics remain poorly unders

safetyarxiv-cs-ai
15 Apr 2026
Safety

Retrieval as a Decision: Training-Free Adaptive Gating for Efficient RAG

DGX agent

arXiv:2511.09803v2 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) improves factuality but retrieving for every query often hurts quality while inflating tokens and latency. We p

safetyarxiv-cs-cl
15 Apr 2026
Safety

SAM3-I: Segment Anything with Instructions

DGX agent

arXiv:2512.04585v3 Announce Type: replace Abstract: Segment Anything Model 3 (SAM3) advances open-vocabulary segmentation through promptable concept segmentation, enabling users to segment all instanc

safetyarxiv-cs-cv
15 Apr 2026
Safety

Scaffold-Conditioned Preference Triplets for Controllable Molecular Optimization with Large Language Models

DGX agent

arXiv:2604.12350v1 Announce Type: cross Abstract: Molecular property optimization is central to drug discovery, yet many deep learning methods rely on black-box scoring and offer limited control over

safetyarxiv-cs-ai
15 Apr 2026
Safety

Scalable and General Whole-Body Control for Cross-Humanoid Locomotion

DGX agent

arXiv:2602.05791v2 Announce Type: replace Abstract: Learning-based whole-body controllers have become a key driver for humanoid robots, yet most existing approaches require robot-specific training. In

safetyarxiv-cs-ro
15 Apr 2026
Safety

Schema-Adaptive Tabular Representation Learning with LLMs for Generalizable Multimodal Clinical Reasoning

DGX agent

arXiv:2604.11835v1 Announce Type: cross Abstract: Machine learning for tabular data remains constrained by poor schema generalization, a challenge rooted in the lack of semantic understanding of struc

safetyarxiv-cs-ai
15 Apr 2026
Safety

Self-Distillation Zero: Self-Revision Turns Binary Rewards into Dense Supervision

DGX agent

arXiv:2604.12002v1 Announce Type: new Abstract: Current post-training methods in verifiable settings fall into two categories. Reinforcement learning (RLVR) relies on binary rewards, which are broadly

safetyarxiv-cs-cl
15 Apr 2026
Safety

Simulation as Supervision: Mechanistic Pretraining for Scientific Discovery

DGX agent

arXiv:2507.08977v4 Announce Type: replace-cross Abstract: Scientific modeling faces a tradeoff between the interpretability of mechanistic theory and the predictive power of machine learning. While ex

safetyarxiv-cs-ai
15 Apr 2026
Safety

So true. “Being right too soon is socially unacceptable”

DGX agent

So true. “Being right too soon is socially unacceptable” I remember reading this from Heinlein as a young man and it took experience for me to fully understand it. People and institutions will persist

safetygary-marcus--x
15 Apr 2026
Safety

SOAR: Self-Correction for Optimal Alignment and Refinement in Diffusion Models

DGX agent

arXiv:2604.12617v1 Announce Type: cross Abstract: The post-training pipeline for diffusion models currently has two stages: supervised fine-tuning (SFT) on curated data and reinforcement learning (RL)

safetyarxiv-cs-ai
15 Apr 2026
Safety

StableSketcher: Enhancing Diffusion Model for Pixel-based Sketch Generation via Visual Question Answering Feedback

DGX agent

arXiv:2510.20093v2 Announce Type: replace-cross Abstract: Although recent advancements in diffusion models have significantly enriched the quality of generated images, challenges remain in synthesizin

safetyarxiv-cs-ai
15 Apr 2026
Safety

Task Alignment: A simple and effective proxy for model merging in computer vision

DGX agent

arXiv:2604.12935v1 Announce Type: new Abstract: Efficiently merging several models fine-tuned for different tasks, but stemming from the same pretrained base model, is of great practical interest. Des

safetyarxiv-cs-cv
15 Apr 2026
Safety

Teaching LLMs Human-Like Editing of Inappropriate Argumentation via Reinforcement Learning

DGX agent

arXiv:2604.12770v1 Announce Type: new Abstract: Editing human-written text has become a standard use case of large language models (LLMs), for example, to make one's arguments more appropriate for a d

safetyarxiv-cs-cl
15 Apr 2026
Safety

The front door to the internet hasn't changed -- but the person going through that door has changed from a human browsing 5 blue links, to a…

DGX agent

The front door to the internet hasn't changed -- but the person going through that door has changed from a human browsing 5 blue links, to an agent browsing a similar index in vastly different ways. A

safetysonya-huang--x
15 Apr 2026
Safety

The role of System 1 and System 2 semantic memory structure in human and LLM biases

DGX agent

arXiv:2604.12816v1 Announce Type: new Abstract: Implicit biases in both humans and large language models (LLMs) pose significant societal risks. Dual process theories propose that biases arise primari

safetyarxiv-cs-cl
15 Apr 2026
← Previous
1…262263264265266…299
Next →