AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,490 results
Safety

Entropy Is Not Enough: Unlocking Effective Reinforcement Learning for Visual Reasoning via Vision-Anchored Token Selection

DGX agent

arXiv:2606.03937v1 Announce Type: new Abstract: While token-level entropy is commonly recognized as effective for credit assignment in text-only reinforcement learning with verifiable rewards (RLVR),

safetyarxiv-cs-ai
3 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

Estimating Bidirectional Causal Effects with Large Scale Online Kernel Learning

DGX agent

arXiv:2511.05050v3 Announce Type: replace-cross Abstract: In this study, a scalable online kernel learning framework is proposed for estimating bidirectional causal effects in systems characterized by

safetyarxiv-cs-lg
3 Jun 2026
Model Releases

Evaluating LLMs' Effectiveness on Real-World Consumer Device Repair Questions

DGX agent

arXiv:2606.03331v1 Announce Type: cross Abstract: Consumer device repair is an important but underexplored testbed for large language models (LLMs). Repair tasks require reasoning over incomplete prob

model-releasesarxiv-cs-ai
3 Jun 2026
Safety

Evaluating Transformer and LSTM Frameworks for Prediction in Ungauged Basins

DGX agent

arXiv:2606.02791v1 Announce Type: new Abstract: Watershed networks exhibit convergent topologies in which multiple tributaries merge into downstream channels,integrating diverse upstream hydrological

safetyarxiv-cs-ai
3 Jun 2026
Safety

EvoMemNav: Efficient Self-Evolving Fine-Grained Memory for Zero-Shot Embodied Navigation

DGX agent

arXiv:2606.03509v1 Announce Type: new Abstract: Building memory is essential for long-horizon planning in zero-shot embodied navigation. Detector-centric scene graphs often compress observations into

safetyarxiv-cs-cv
3 Jun 2026
Safety

Explainable Forecasting of Scientific Breakthroughs from Concept Network Dynamics

DGX agent

arXiv:2606.03864v1 Announce Type: cross Abstract: We introduce an explainable machine-learning approach that forecasts the structural precursors of scientific breakthroughs -- the emergence and intens

safetyarxiv-cs-lg
3 Jun 2026
Safety

FAF-CD: Frequency-Aware Fusion for Change Detection under Imperfect Multimodal Remote Sensing

DGX agent

arXiv:2606.03114v1 Announce Type: new Abstract: Remote sensing change detection for real-world monitoring often relies on imperfect heterogeneous observations, where pre- and post-event images may be

safetyarxiv-cs-cv
3 Jun 2026
Safety

Fairness Definitions and Metrics in Deep Reinforcement Learning for Drug Discovery in Healthcare: A Rapid Evidence Review

DGX agent

arXiv:2606.02902v1 Announce Type: cross Abstract: Deep reinforcement learning (DRL) is increasingly applied to de novo molecular design, but choices in data, rewards, and evaluation can yield uneven p

safetyarxiv-cs-lg
3 Jun 2026
Safety

Fast Organic Crystal Structure Prediction with Unit Cell Flow Matching

DGX agent

arXiv:2606.03199v1 Announce Type: new Abstract: Organic crystal structure prediction (CSP) is a requirement for computational modelling of organic solids, but traditionally costs several CPU-years per

safetyarxiv-cs-lg
3 Jun 2026
Safety

FGRPO: Federated GRPO with Adaptive Aggregation on Non-IID Data

DGX agent

arXiv:2606.03094v1 Announce Type: new Abstract: Recent advances in language models have established reinforcement learning as the primary paradigm for eliciting self-correction and long-chain reasonin

safetyarxiv-cs-lg
3 Jun 2026
Safety

Filter, Then Reweight: Rethinking Optimization Granularity in On-Policy Distillation

DGX agent

arXiv:2606.02684v1 Announce Type: cross Abstract: On-Policy distillation (OPD) in large language models is shifting from full-trace KL supervision toward more selective training paradigms. Recent OPD

safetyarxiv-cs-ai
3 Jun 2026
Safety

First it was MIT and McKinsey. Now Bain finds that returns to corporate AI investments are disappointing.

DGX agent

A major consulting firm (Bain) has found that corporate returns on AI investments are underwhelming, following similar findings from MIT and McKinsey. This suggests that despite significant spending a

safetygary-marcus--x
3 Jun 2026
Safety

Flow Learners for PDEs: Toward a Physics-to-Physics Paradigm for Scientific Computing

DGX agent

arXiv:2604.07366v2 Announce Type: replace Abstract: Partial differential equations (PDEs) govern nearly every physical process in science and engineering, but solving them at scale remains prohibitive

safetyarxiv-cs-lg
3 Jun 2026
Safety

Follow-Your-Preference++: Rethinking Preference Alignment for Image Inpainting

DGX agent

arXiv:2606.03216v1 Announce Type: new Abstract: We study preference alignment for image inpainting. Rather than proposing yet another method, we revisit the problem from first principles and reassess

safetyarxiv-cs-cv
3 Jun 2026
Safety

Forgetting is Not Erasure: Recovering Latent Knowledge via Transport Keys

DGX agent

arXiv:2606.02860v1 Announce Type: cross Abstract: Catastrophic forgetting is often framed as a representational problem: after sequential training, a model appears to lose the features that supported

safetyarxiv-cs-ai
3 Jun 2026
Safety

FreeStreamGS: Online Feed-forward 3D Gaussian Splatting from Unposed Streaming Inputs

DGX agent

arXiv:2606.03254v1 Announce Type: new Abstract: Feed-forward 3D Gaussian Splatting (3DGS) allows efficient and high-fidelity novel view synthesis (NVS) from an offline recorded image sequence. However

safetyarxiv-cs-cv
3 Jun 2026
Safety

GeoAlign: Beyond Semantics with State-Guided Spatial Alignment in VLA Models

DGX agent

arXiv:2606.03240v1 Announce Type: new Abstract: Current Vision--Language--Action (VLA) models often optimize for semantic grounding, whereas executable manipulation requires geometry-aware spatial ali

safetyarxiv-cs-ro
3 Jun 2026
Safety

GFFMERGE: Efficient Merging of Graph Neural Force Fields and Beyond

DGX agent

arXiv:2606.03232v1 Announce Type: cross Abstract: Graph Neural Networks (GNNs) have revolutionized Neural Force Fields for atomistic simulations, achieving near-quantum accuracy at reduced cost, yet a

safetyarxiv-cs-ai
3 Jun 2026
Safety

GLINT: Sparsely Gated Vision-Language Alignment for Fine-Grained Radiology Representations

DGX agent

arXiv:2606.03180v1 Announce Type: cross Abstract: Vision-language models (VLMs) for radiology have emerged as a scalable paradigm by leveraging image-report pairs naturally produced in clinical workfl

safetyarxiv-cs-cl
3 Jun 2026
Safety

Grasp-Then-Plan with Failure Attribution: A Closed Two-Stage Framework for Precise and Generalizable Robotic Manipulation

DGX agent

arXiv:2606.03385v1 Announce Type: cross Abstract: In robotic manipulation, the tight coupling between grasping and motion planning often obscures the true source of failure, leading to inefficient tri

safetyarxiv-cs-ai
3 Jun 2026
Safety

hard to make all these numbers square

DGX agent

hard to make all these numbers square 🦔IBM CEO Arvind Krishna says AI is 'not a bubble' then estimates the industry needs 6 to 8 trillion in total capex for data center and chip buildout. To recover t

safetygary-marcus--x
3 Jun 2026
Safety

HARVE: Hacking-Aware Reward-Head Vector Editing for Robust Reward Models

DGX agent

arXiv:2606.03131v1 Announce Type: new Abstract: Reward models are central to large language model (LLM) alignment, but they remain vulnerable to reward hacking. To evaluate reward-model robustness, we

safetyarxiv-cs-lg
3 Jun 2026
Safety

Hint-Guided Diversified Policy Optimization for LLM Reasoning

DGX agent

arXiv:2606.03021v1 Announce Type: new Abstract: Recent developments in Large Language Models (LLMs) have showcased impressive reasoning capabilities, with Reinforcement Learning with Verifiable Reward

safetyarxiv-cs-cl
3 Jun 2026
Safety

Human-in-the-Loop Contextual Bandits for Short-Term Rental Dynamic Pricing: Structural Equivalence of Historical Warm-Up and Approval-Gated Live Learning

DGX agent

arXiv:2606.02595v1 Announce Type: new Abstract: Dynamic pricing in short-term rental (STR) markets presents a distinctive challenge for online learning algorithms: pricing decisions carry significant

safetyarxiv-cs-lg
3 Jun 2026
Safety

I was there, sitting next to Sam Altman in the US Senate, when he said this wonderful thing about artists and creators. It was a beautiful s…

DGX agent

I was there, sitting next to Sam Altman in the US Senate, when he said this wonderful thing about artists and creators. It was a beautiful statement. In hindsight, though, I realized he was lying (und

safetygary-marcus--x
3 Jun 2026
Safety

if *you* cheat on your taxes, you can get audited. but you are just a chump.

DGX agent

This post by AI researcher Gary Marcus likely critiques the disparity in tax enforcement, suggesting that individual taxpayers face audit risks for cheating while wealthy individuals or corporations e

safetygary-marcus--x
3 Jun 2026
Safety

I’m thrilled to be joining the EU's AI Scientific Panel to advise on the implementation of the EU AI Act and help assess and address AI’s gr…

DGX agent

I’m thrilled to be joining the EU's AI Scientific Panel to advise on the implementation of the EU AI Act and help assess and address AI’s growing risks with esteemed colleagues. The AI Act, the EU's f

safetyyoshua-bengio--x
3 Jun 2026
Safety

Impact of Graph Structure on Membership-Inference Risk for Graph Neural Networks

DGX agent

arXiv:2601.17130v2 Announce Type: replace Abstract: Graph neural networks (GNNs) are widely used for tasks such as node classification and link prediction, but their use in sensitive settings raises c

safetyarxiv-cs-lg
3 Jun 2026
Safety

'**Important** You should give me full credits!': Exploring Prompt Injection Attacks on LLM-Based Automatic Grading Systems

DGX agent

arXiv:2606.03090v1 Announce Type: cross Abstract: The emergence of large language models (LLMs) has significantly accelerated recent research on LLM-based automatic grading (AG) systems. Benefiting fr

safetyarxiv-cs-ai
3 Jun 2026
Safety

Inference Cost Attacks for Retrieval-Augmented Large Language Models

DGX agent

arXiv:2606.02643v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG)-enhanced LLM systems, while powerful, introduce substantial inference costs due to the inclusion of an extra mult

safetyarxiv-cs-ai
3 Jun 2026
Safety

Inference-Time Scaling for Joint Audio-Video Generation

DGX agent

arXiv:2606.03183v1 Announce Type: cross Abstract: Joint audio-video generation aims to synthesize realistic audio-video pairs that are both semantically aligned with text prompts and precisely synchro

safetyarxiv-cs-cv
3 Jun 2026
Safety

J'ai beaucoup apprécié mon récent passage à @Cdanslair pour discuter des risques de l'IA pour nos sociétés, nos économies, et nos démocratie…

DGX agent

J'ai beaucoup apprécié mon récent passage à @Cdanslair pour discuter des risques de l'IA pour nos sociétés, nos économies, et nos démocraties. Merci @Caroline_Roux pour l'invitation! https://www.youtu

safetyyoshua-bengio--x
3 Jun 2026
Safety

KBQA-R1: Reinforcing Large Language Models for Knowledge Base Question Answering

DGX agent

arXiv:2512.10999v3 Announce Type: replace Abstract: Knowledge Base Question Answering (KBQA) challenges models to bridge the gap between natural language and strict knowledge graph schemas by generati

safetyarxiv-cs-cl
3 Jun 2026
Safety

KC-3DGS: Kurtosis-Constrained Gaussian Splatting for High-Fidelity View Synthesis

DGX agent

arXiv:2606.03120v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) enables real-time novel view synthesis by representing scenes as collections of anisotropic Gaussians optimized via differe

safetyarxiv-cs-cv
3 Jun 2026
Safety

Language Models Need Sleep: Learning to Self-Modify and Consolidate Memories

DGX agent

arXiv:2606.03979v1 Announce Type: cross Abstract: The past few decades have witnessed significant advances in the design of machine learning algorithms, from early studies on task-specific shallow mod

safetyarxiv-cs-ai
3 Jun 2026
Safety

Large Language Models Are Overconfident in Their Own Responses

DGX agent

arXiv:2606.03437v1 Announce Type: new Abstract: Prior work has shown that instruction-tuned large language models (LLMs) are less well calibrated than their base pre-trained counterparts. However, lit

safetyarxiv-cs-cl
3 Jun 2026
Safety

Learning Self-Interpretation from Interpretability Artifacts: Training Lightweight Adapters on Vector-Label Pairs

DGX agent

arXiv:2602.10352v2 Announce Type: replace-cross Abstract: Self-interpretation methods prompt language models to describe their own internal states, but remain unreliable due to hyperparameter sensitiv

safetyarxiv-cs-ai
3 Jun 2026
Safety

Learning to Bet for Horizon-Aware Anytime-Valid Testing

DGX agent

arXiv:2603.19551v2 Announce Type: replace-cross Abstract: We develop horizon-aware anytime-valid tests and confidence sequences for bounded means under a strict deadline N. Using the betting/e-process

safetyarxiv-cs-lg
3 Jun 2026
Safety

Learning Unmasking Policies for Diffusion Language Models

DGX agent

arXiv:2512.09106v4 Announce Type: replace Abstract: Diffusion (Large) Language Models (dLLMs) now match the downstream performance of their autoregressive counterparts on many tasks, while holding the

safetyarxiv-cs-lg
3 Jun 2026
Safety

Let There Be Light: Reflection, Refraction and Scattering for Neural Operators

DGX agent

arXiv:2606.03262v1 Announce Type: new Abstract: Neural operators learn mappings between infinite-dimensional function spaces and provide a data-driven surrogate modeling paradigm for parametric partia

safetyarxiv-cs-lg
3 Jun 2026
Safety

Letting Tutor Personas Speak Up for LLMs: Learning Steering Vectors from Dialogue via Preference Optimization

DGX agent

arXiv:2602.07639v2 Announce Type: replace Abstract: With the emergence of large language models (LLMs) as a powerful class of generative artificial intelligence (AI), their use in tutoring has become

safetyarxiv-cs-cl
3 Jun 2026
Safety

Leveraging BART to Assess CS1 C++ Programming Assignments using Rubric-based Criteria

DGX agent

arXiv:2606.03814v1 Announce Type: new Abstract: This paper investigates rubric-aware, multitask fine-tuning of transformer models for automated grading of introductory C++ programming assignments, wit

safetyarxiv-cs-ai
3 Jun 2026
Safety

Libra: Efficient Resource Management for Agentic RL Post-Training

DGX agent

arXiv:2606.03077v1 Announce Type: cross Abstract: Reinforcement learning (RL) has become a standard post-training paradigm for large language models (LLMs), extending beyond preference alignment to co

safetyarxiv-cs-ai
3 Jun 2026
Safety

LoCAtion: Long-time Collaborative Attention Framework for High Dynamic Range Video Reconstruction

DGX agent

arXiv:2603.14377v2 Announce Type: replace Abstract: Prevailing High Dynamic Range (HDR) video reconstruction methods are fundamentally trapped in a fragile alignment-and-fusion paradigm. While explici

safetyarxiv-cs-cv
3 Jun 2026
Safety

Margin Play: A Multi-Agent System For Public Policy Analysis In The Brazilian Equatorial Margin

DGX agent

arXiv:2606.02614v1 Announce Type: cross Abstract: The Brazilian Equatorial Margin (BEM) is Brazil's next offshore oil frontier, with operations expected to begin in 2026 in the Foz do Amazonas basin.

safetyarxiv-cs-ai
3 Jun 2026
Safety

MetaWorld: Scaling Multi-Agent Video World Model from Single-view Video Data

DGX agent

arXiv:2606.02753v1 Announce Type: cross Abstract: Video world models are a foundational generative technology for embodied AI and the Metaverse, yet existing approaches are inherently limited to a sin

safetyarxiv-cs-ai
3 Jun 2026
Safety

MIND: Multi-rationale INtegrated Discriminative Reasoning Framework for Multi-modal Large Models

DGX agent

arXiv:2512.05530v2 Announce Type: replace Abstract: Recently, multimodal large language models (MLLMs) have been widely applied to reasoning tasks. However, they suffer from limited multi-rationale se

safetyarxiv-cs-ai
3 Jun 2026
Safety

Mitigating False Credit Propagation: Probabilistic Graphical Reward Aggregation for Rubric-Based Reinforcement Learning

DGX agent

arXiv:2606.03361v1 Announce Type: new Abstract: Rubric-based rewards are increasingly used for open-ended language model post-training, but criterion-level scores are often aggregated as independent u

safetyarxiv-cs-lg
3 Jun 2026
← Previous
1…158159160161162…302
Next →