AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,488 results
25 Jun 2026

Fourier Multi-Component and Multi-Layer Neural Networks: Unlocking High-Frequency Potential

SafetyDGX agent

arXiv:2502.18959v4 Announce Type: replace Abstract: The architecture of a neural network and the choice of its activation function are both fundamental to its performance. Equally important is ensurin

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning

SafetyDGX agent

arXiv:2505.19532v2 Announce Type: replace Abstract: The current state-of-the-art backdoor attacks against Reinforcement Learning (RL) rely upon unrealistically permissive access models, that assume th

From Forecasting Leaderboards to Deployment Decisions: A Fail-Closed Certification Protocol

SafetyDGX agent

arXiv:2606.24996v1 Announce Type: new Abstract: Forecasting leaderboards rank models by predictive quality, but their winners are often read as deployment-ready top-1 advice. That reading can fail whe

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Fully Differentiable Neural Forced Alignment via Soft Dynamic Programming

SafetyDGX agent

arXiv:2606.25460v1 Announce Type: cross Abstract: Recent advances in sequence modeling have significantly improved ASR systems, bringing them close to human-level recognition accuracy and enhancing ro

GCT-MARL: Graph-Based Contrastive Transfer for Sample-Efficient Cooperative Multi-Agent Reinforcement Learning

SafetyDGX agent

arXiv:2606.25073v1 Announce Type: new Abstract: In cooperative multi-agent reinforcement learning (MARL), from a deployment perspective, it is challenging and expensive to train agents from scratch fo

Generalised Medical Phrase Grounding

SafetyDGX agent

arXiv:2512.01085v3 Announce Type: replace-cross Abstract: Medical phrase grounding (MPG) maps textual descriptions of radiological findings to corresponding image regions. These grounded reports are e

Geometry-Anchored Transport Framework for Exemplar-Free Class-Incremental Learning

SafetyDGX agent

arXiv:2606.25347v1 Announce Type: cross Abstract: Exemplar-free class-incremental learning (EFCIL) requires stable decision boundaries within a shifting feature space. While maintaining class-conditio

Grifters and many other salesman are like LLMs. They say things that aren’t always true with absolutely certainty.

SafetyDGX agent

Grifters and many other salesman are like LLMs. They say things that aren’t always true with absolutely certainty. Bitcoin has won. Global consensus is that $BTC is digital capital. The four-year cycl

Grok may not be the AGI that Elon promised, but it appears that the man has found his use case.

SafetyDGX agent

Grok may not be the AGI that Elon promised, but it appears that the man has found his use case. xAI is betting big on AI video and image generation — but the biggest driver of Grok Imagine may not be

Heuresis: Search Strategies for Autonomous AI Research Agents Across Quality, Diversity and Novelty

SafetyDGX agent

arXiv:2606.25198v1 Announce Type: new Abstract: Autonomous AI Research promises to accelerate the scientific progress of machine learning. To realise this goal, current Large Language Model (LLM)-base

Homogeneity Bias in Open-Weight LLMs Is Robust to Decoding Hyperparameters

SafetyDGX agent

arXiv:2501.02211v2 Announce Type: replace-cross Abstract: Large language models (LLMs) reproduce homogeneity bias -- the tendency to portray marginalized groups as more internally similar than dominan

Imagine if i had a business that gave you 20 bills for 10 bills. As word spread, my customer base would grow. My revenue (how many $10 bil…

SafetyDGX agent

Imagine if i had a business that gave you 20 bills for 10 bills. As word spread, my customer base would grow. My revenue (how many $10 bills I collected) would grow. And I would continue to lose money

In the end, will Elon be best known for his efforts to mainstream electric cars, for his rockets, …. or for his efforts to mainstream AI-gen…

SafetyDGX agent

In the end, will Elon be best known for his efforts to mainstream electric cars, for his rockets, …. or for his efforts to mainstream AI-generated porn? xAI is betting big on AI video and image genera

It is amazing to me how little the X-verse cares about the rise of authoritarianism in the United States. Is that because most of the “peopl…

SafetyDGX agent

It is amazing to me how little the X-verse cares about the rise of authoritarianism in the United States. Is that because most of the “people” here are actually Russian bots? 🤔 More Soviet-style decep

Learning Action Priors for Cross-embodiment Robot Manipulation

SafetyDGX agent

arXiv:2606.26095v1 Announce Type: cross Abstract: Most Vision-Language-Action (VLA) models build on a Vision-Language Model (VLM) backbone by attaching an action module and optimizing the full policy

Learning Asynchronous Upper-body Task-space Trajectory Tracking Policy for Humanoid Robots

SafetyDGX agent

arXiv:2606.25706v1 Announce Type: new Abstract: High-level humanoid planners often output sparse task-space, low-rate trajectories, whereas whole-body controllers run at high frequency. This creates t

Learning Optimization Proxies for Sequential Contextual Stochastic Programs: An Order Fulfillment Application

SafetyDGX agent

arXiv:2606.25362v1 Announce Type: cross Abstract: Sequential contextual stochastic programs model real-time decision systems in which each time epoch commits to an action under uncertainty whose conse

Learning Robot Visual Navigation in Crowds via Intention-Aware Scene Representations

SafetyDGX agent

arXiv:2606.26047v1 Announce Type: new Abstract: Robot crowd navigation requires the ability to infer human intentions while accounting for the structural constraints of the environment. Currently, dee

Learning Subset-Shared Invariances for Domain Generalization with Mixture-of-Experts

SafetyDGX agent

arXiv:2606.25665v1 Announce Type: new Abstract: Domain generalization (DG) aims to learn a model from one or more source domains that generalizes to an unseen target domain without accessing target da

Learning with a Single Rollout via Monte Carlo Pass@k Critic

SafetyDGX agent

arXiv:2606.25451v1 Announce Type: new Abstract: Estimating token-level advantages in reinforcement learning (RL) for language models remains challenging because scaling up episodic experience collecti

Lightweight PCGAE-Net: Parallel CrossGate Attention and Bottleneck AutoEncoder for Efficient 5G Channel Prediction

SafetyDGX agent

arXiv:2606.25401v1 Announce Type: cross Abstract: Accurate channel state information (CSI) prediction is essential for proactive beamforming and resource management in 5G massive MIMO systems, yet the

LLM-Based Scientific Peer Review: Methods, Benchmarks, and Reliability Challenges

SafetyDGX agent

arXiv:2606.25057v1 Announce Type: new Abstract: The rapid growth of scientific submissions has pushed traditional peer review toward its scalability limits, motivating the exploration of large languag

Looks like OpenAI’s IPO will be delayed until 2027. • might be a sign that their finances don’t look compelling yet. • gives Anthropic time …

SafetyDGX agent

Looks like OpenAI’s IPO will be delayed until 2027. • might be a sign that their finances don’t look compelling yet. • gives Anthropic time to go first If they can bump up the valuation by $999.3 bill

love that one of my readers made her own, explicit set of public predictions. 🙏

SafetyDGX agent

love that one of my readers made her own, explicit set of public predictions. 🙏 I want to post my predictions for 2ish years horizon. AI stock bubble collapses due to fundamentals not working out with

Low-Complexity Policy Tessellations in Structured Markov Decision Processes

SafetyDGX agent

arXiv:2606.25593v1 Announce Type: new Abstract: We study optimal-policy geometry in structured Markov decision processes. While approximate dynamic programming and reinforcement learning typically app

Low Variance Trust Region Optimization with Independent Actors and Sequential Updates in Cooperative Multi-agent Reinforcement Learning

SafetyDGX agent

arXiv:2606.25526v1 Announce Type: new Abstract: Cooperative multi-agent reinforcement learning assumes each agent shares the same reward function and can be trained effectively using the Trust Region

MAPL: Multi-Objective Preference Learning for Robot Locomotion

SafetyDGX agent

arXiv:2606.25398v1 Announce Type: new Abstract: Reward design remains a major bottleneck in reinforcement learning for robot locomotion, where successful policies often depend on carefully tuned, task

Memory Retrieval in Visuomotor Policies for Long-Horizon Robot Control

SafetyDGX agent

arXiv:2606.25136v1 Announce Type: new Abstract: General-purpose robots operating in partially observable environments, such as homes, require memory to support autonomy. They must recall diverse infor

Minimax PAC Bounds for Learning in Exogenous Contextual MDPs

SafetyDGX agent

arXiv:2606.25170v1 Announce Type: cross Abstract: We study PAC learning in tabular discounted Markov decision processes with exogenous i.i.d. contexts, with discount factor gamma, finite state space m

MiniOpt: Reasoning to Model and Solve General Optimization Problems with Limited Resources

SafetyDGX agent

arXiv:2606.25832v1 Announce Type: new Abstract: Achieving strong optimization generalization across diverse optimization problems while requiring limited training resources remains a challenging probl

More Soviet-style deceptions that I never expected to see in the United States 😢

SafetyDGX agent

Gary Marcus expresses concern about deceptive practices in the United States that he compares to Soviet-style tactics, suggesting he has observed propaganda, misinformation, or manipulative government

Neglected Free Lunch from Post-training: Progress Advantage for LLM Agents

SafetyDGX agent

arXiv:2606.26080v1 Announce Type: new Abstract: Process reward models enable fine-grained, step-level evaluation of LLMs, yet building them for agentic settings remains prohibitively difficult: long-h

Neural Machine Translation for Low-Resource Tangkhul--English

SafetyDGX agent

arXiv:2606.25365v1 Announce Type: new Abstract: We present a study on low-resource machine translation for the Tangkhul-English (nmf-en) language pair. Tangkhul is a severely under-resourced Tibeto-Bu

noah smith, ignoring that Anthropic’s rumored profitable quarter is a. during peak tokenmaxxing b, before recent Chinese advances c. in a qu…

SafetyDGX agent

noah smith, ignoring that Anthropic’s rumored profitable quarter is a. during peak tokenmaxxing b, before recent Chinese advances c. in a quarter in which Elon gave them a big one time subsidy. that’s

Offline Multi-agent Continual Cooperation via Skill Partition and Reuse

SafetyDGX agent

arXiv:2606.25389v1 Announce Type: new Abstract: Extracting skills from multi-agent offline dataset improves learning efficiency via sharing task-invariant coordination skills among tasks. In settings

On-Policy Self-Distillation with Sampled Demonstrations Reduces Output Diversity

SafetyDGX agent

arXiv:2606.26091v1 Announce Type: new Abstract: On-policy self-distillation achieves strong pass@1 accuracy by using a single model as both teacher and student, with the teacher conditioned on a corre

One Body, Two Minds: Variable Autonomy Approach for a Co-embodied Robotic Hand

SafetyDGX agent

arXiv:2606.25575v1 Announce Type: new Abstract: Assistive robotic systems face a fundamental trade-off: fully autonomous systems lack user agency, while fully user-controlled systems demand continuous

OPERA: Aligning Open-Ended Reasoning via Objective Perplexity-based Reinforcement Learning

SafetyDGX agent

arXiv:2606.25757v1 Announce Type: new Abstract: Reinforcement Learning (RL) has enabled LLMs to excel in objective reasoning tasks such as mathematics and code generation. However, applying RL to open

Paid Voices vs. Public Feeds: Interpretable Cross-Platform Theme-Based Analysis of Climate Discourse

SafetyDGX agent

arXiv:2601.13317v2 Announce Type: replace Abstract: Climate discourse online shapes public understanding of climate change and informs political and policy debate, yet it unfolds across structurally d

People are not getting how significant this news – The White House delaying GPT 5.6 – is. Here’s the context and a suggestion. Both @pmarca …

SafetyDGX agent

People are not getting how significant this news – The White House delaying GPT 5.6 – is. Here’s the context and a suggestion. Both @pmarca and @DavidSacks did everything in their power to keep the Wh

PERRY: Policy Evaluation with Confidence Intervals using Auxiliary Data

SafetyDGX agent

arXiv:2507.20068v3 Announce Type: replace Abstract: Off-policy evaluation (OPE) methods estimate the value of a new reinforcement learning (RL) policy prior to deployment. Recent advances have shown t

Point Cloud Diffusion with Global and Local Reconstruction for Instance-Level 3D Anomaly Detection

SafetyDGX agent

arXiv:2606.25740v1 Announce Type: new Abstract: 3D anomaly detection in point clouds is critical for high-precision industrial manufacturing. Reconstruction-based methods have laid a strong foundation

Position Spaces and Graphs

SafetyDGX agent

arXiv:2606.25719v1 Announce Type: new Abstract: In this paper, we introduce position graphs, a graph-based reasoning framework based on the formalization of position spaces. This framework utilizes tw

Power-Budgeted Underwater Vehicle Control via Constrained Reinforcement Learning

SafetyDGX agent

arXiv:2606.25680v1 Announce Type: new Abstract: Underwater vehicles operate from a fixed onboard energy budget that propulsion rapidly depletes, so a controller that completes its task while drawing l

Reflective VLA: In-Context Action Consequences Make VLAs Generalize

SafetyDGX agent

arXiv:2606.25215v1 Announce Type: new Abstract: Most vision-language-action (VLA) models are reactive: they predict the next action from the current instruction and observation, implicitly assuming th

RGB: RL Guided Whole-Body MPPI for Humanoid Control

SafetyDGX agent

arXiv:2606.25123v1 Announce Type: new Abstract: Humanoid robots require whole-body controllers that are both robust and precise in contact-rich environments. While deep reinforcement learning (RL) ach

RN-D: Discretized Categorical Actors for On-Policy Reinforcement Learning

SafetyDGX agent

arXiv:2601.23075v2 Announce Type: replace Abstract: On-policy Reinforcement Learning (RL) remains a dominant paradigm for continuous control, yet standard implementations rely on Gaussian actors and r

ROAD-VLA: Robust Online Adaptation via Self-Distillation for Vision-Language-Action Models

SafetyDGX agent

arXiv:2606.25800v1 Announce Type: new Abstract: Effective online adaptation of vision-language-action (VLA) models remains challenging, as sparse rewards provide weak supervision for high-dimensional

SAC^2-Net: Semantic Anchoring and Complementary-Consensus Fusion for Multimodal Micro-Expression Recognition

SafetyDGX agent

arXiv:2606.25542v1 Announce Type: new Abstract: Micro-expression recognition (MER) is challenging due to subtle facial movements, limited data, and the ambiguous relationship between Action Units (AUs

SAGE-Nav: Leveraging LLM Planning and Alignment Fusion for Hierarchical Scene Graph-Guided Navigation

SafetyDGX agent

arXiv:2606.25497v1 Announce Type: new Abstract: Object-Goal Navigation (ObjNav) requires embodied agents to autonomously locate specified targets using only egocentric visual observations. Existing mo

ScalingAR: Scaling Confidence for Autoregressive Image Generation

SafetyDGX agent

arXiv:2509.26376v3 Announce Type: replace Abstract: Test-time strategies have shown remarkable success in improving large language models, but their application to next-token prediction (NTP) autoregr

Semantic Consistency Policy Optimization for Reinforcement Learning of LLM Agents

SafetyDGX agent

arXiv:2606.25852v1 Announce Type: new Abstract: Group-based reinforcement learning effectively post-trains LLM agents for long-horizon, sparse-reward tasks by deriving step-level credit from trajector

Solving Markov Decision Processes with Future Information via MPC

SafetyDGX agent

arXiv:2606.24991v1 Announce Type: cross Abstract: Model Predictive Control (MPC) is widely used in industrial and robotic systems for enforcing constraints and embedding domain knowledge through finit

Specifying AI-SDLC Processes: A Protocol Language for Human-Agent Boundaries

SafetyDGX agent

arXiv:2606.20615v2 Announce Type: replace Abstract: AI agents now participate as first-class team members across the software development lifecycle, yet no specification language exists for expressing

StairMaster: Learning to Conquer Risky Hollow Stairs for Agile Quadrupedal Robots

SafetyDGX agent

arXiv:2606.25765v1 Announce Type: new Abstract: Climbing hollow stairs remains a challenging problem for quadruped robots due to the high risk of leg trapping, severe depth sparsity, and high-frequenc

STOCKSTAY Another Day: The Latest Addition to Turla’s Intelligence Gathering Apparatus

SafetyDGX agent

Written by: Jordan Jones Introduction Google Threat Intelligence Group (GTIG) has conducted an in-depth analysis of a .NET backdoor, tracked as STOCKSTAY, that has been continually developed and deplo

Supervised Reinforcement Learning for the Coordination of Distributed Energy Resources

SafetyDGX agent

arXiv:2606.24947v1 Announce Type: new Abstract: The increasing integration of distributed energy resources (DERs) is crucial for power system decarbonization, yet unlocking DERs' flexibility is challe

Taxonomy-aware deep learning for hierarchical marine species classification in underwater imagery

SafetyDGX agent

arXiv:2606.25989v1 Announce Type: new Abstract: Automated classification of marine species from underwater imagery is essential for scalable ocean biodiversity monitoring and conservation policy. Exis

The Clinician's Veto: Navigating Trust, Liability, and Uncertainty in Autonomous AI Prescribing

SafetyDGX agent

arXiv:2606.25108v1 Announce Type: new Abstract: Autonomous AI systems are transitioning from advisory to autonomous roles for medication prescriptions. Recent United States bill H.R. 238 and Utah's pr

The Hitchhiker's Guide to Agentic AI: From Foundations to Systems

SafetyDGX agent

arXiv:2606.24937v1 Announce Type: cross Abstract: The Hitchhiker's Guide to Agentic AI is a comprehensive practitioner's reference for building autonomous AI systems. The book covers the full stack fr

← Previous
1…102103104105106…242
Next →