AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
All
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,809 results
Safety

It would be very useful to understand more about the government safety concerns associated with frontier AI releases so we could (a) know wh…

DGX agent

It would be very useful to understand more about the government safety concerns associated with frontier AI releases so we could (a) know what risks everyone will face if/when open source reaches Myth

safetyethan-mollick--x
25 Jun 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

Learning Action Priors for Cross-embodiment Robot Manipulation

DGX agent

arXiv:2606.26095v1 Announce Type: cross Abstract: Most Vision-Language-Action (VLA) models build on a Vision-Language Model (VLM) backbone by attaching an action module and optimizing the full policy

safetyarxiv-cs-cv
25 Jun 2026
Safety

Learning Asynchronous Upper-body Task-space Trajectory Tracking Policy for Humanoid Robots

DGX agent

arXiv:2606.25706v1 Announce Type: new Abstract: High-level humanoid planners often output sparse task-space, low-rate trajectories, whereas whole-body controllers run at high frequency. This creates t

safetyarxiv-cs-ro
25 Jun 2026
Safety

Learning Optimization Proxies for Sequential Contextual Stochastic Programs: An Order Fulfillment Application

DGX agent

arXiv:2606.25362v1 Announce Type: cross Abstract: Sequential contextual stochastic programs model real-time decision systems in which each time epoch commits to an action under uncertainty whose conse

safetyarxiv-cs-lg
25 Jun 2026
Safety

Learning Robot Visual Navigation in Crowds via Intention-Aware Scene Representations

DGX agent

arXiv:2606.26047v1 Announce Type: new Abstract: Robot crowd navigation requires the ability to infer human intentions while accounting for the structural constraints of the environment. Currently, dee

safetyarxiv-cs-ro
25 Jun 2026
Safety

Learning Subset-Shared Invariances for Domain Generalization with Mixture-of-Experts

DGX agent

arXiv:2606.25665v1 Announce Type: new Abstract: Domain generalization (DG) aims to learn a model from one or more source domains that generalizes to an unseen target domain without accessing target da

safetyarxiv-cs-lg
25 Jun 2026
Safety

Learning with a Single Rollout via Monte Carlo Pass@k Critic

DGX agent

arXiv:2606.25451v1 Announce Type: new Abstract: Estimating token-level advantages in reinforcement learning (RL) for language models remains challenging because scaling up episodic experience collecti

safetyarxiv-cs-lg
25 Jun 2026
Safety

Lightweight PCGAE-Net: Parallel CrossGate Attention and Bottleneck AutoEncoder for Efficient 5G Channel Prediction

DGX agent

arXiv:2606.25401v1 Announce Type: cross Abstract: Accurate channel state information (CSI) prediction is essential for proactive beamforming and resource management in 5G massive MIMO systems, yet the

safetyarxiv-cs-ai
25 Jun 2026
Safety

LLM-Based Scientific Peer Review: Methods, Benchmarks, and Reliability Challenges

DGX agent

arXiv:2606.25057v1 Announce Type: new Abstract: The rapid growth of scientific submissions has pushed traditional peer review toward its scalability limits, motivating the exploration of large languag

safetyarxiv-cs-cl
25 Jun 2026
Safety

Long-Term Simulation Exposes Cognitive-Developmental Risks in AI Companions

DGX agent

arXiv:2606.25396v1 Announce Type: new Abstract: AI companions powered by large language models increasingly interact with cognition-developing users, including children and adolescents, creating risks

safetyarxiv-cs-ai
25 Jun 2026
Safety

Looks like OpenAI’s IPO will be delayed until 2027. • might be a sign that their finances don’t look compelling yet. • gives Anthropic time …

DGX agent

Looks like OpenAI’s IPO will be delayed until 2027. • might be a sign that their finances don’t look compelling yet. • gives Anthropic time to go first If they can bump up the valuation by $999.3 bill

safetygary-marcus--x
25 Jun 2026
Safety

love that one of my readers made her own, explicit set of public predictions. 🙏

DGX agent

love that one of my readers made her own, explicit set of public predictions. 🙏 I want to post my predictions for 2ish years horizon. AI stock bubble collapses due to fundamentals not working out with

safetygary-marcus--x
25 Jun 2026
Safety

Low-Complexity Policy Tessellations in Structured Markov Decision Processes

DGX agent

arXiv:2606.25593v1 Announce Type: new Abstract: We study optimal-policy geometry in structured Markov decision processes. While approximate dynamic programming and reinforcement learning typically app

safetyarxiv-cs-lg
25 Jun 2026
Safety

Low Variance Trust Region Optimization with Independent Actors and Sequential Updates in Cooperative Multi-agent Reinforcement Learning

DGX agent

arXiv:2606.25526v1 Announce Type: new Abstract: Cooperative multi-agent reinforcement learning assumes each agent shares the same reward function and can be trained effectively using the Trust Region

safetyarxiv-cs-lg
25 Jun 2026
Safety

MAPL: Multi-Objective Preference Learning for Robot Locomotion

DGX agent

arXiv:2606.25398v1 Announce Type: new Abstract: Reward design remains a major bottleneck in reinforcement learning for robot locomotion, where successful policies often depend on carefully tuned, task

safetyarxiv-cs-ro
25 Jun 2026
Safety

MedGuards: Multi-Agent System for Reliable Medical Error Detection and Correction

DGX agent

arXiv:2606.25651v1 Announce Type: new Abstract: As Large Language Models (LLMs) are increasingly deployed in healthcare settings, accurate error detection and correction in generated or existing text

safetyarxiv-cs-cl
25 Jun 2026
Safety

Memory Retrieval in Visuomotor Policies for Long-Horizon Robot Control

DGX agent

arXiv:2606.25136v1 Announce Type: new Abstract: General-purpose robots operating in partially observable environments, such as homes, require memory to support autonomy. They must recall diverse infor

safetyarxiv-cs-ro
25 Jun 2026
Safety

Minimax PAC Bounds for Learning in Exogenous Contextual MDPs

DGX agent

arXiv:2606.25170v1 Announce Type: cross Abstract: We study PAC learning in tabular discounted Markov decision processes with exogenous i.i.d. contexts, with discount factor gamma, finite state space m

safetyarxiv-cs-lg
25 Jun 2026
Safety

MiniOpt: Reasoning to Model and Solve General Optimization Problems with Limited Resources

DGX agent

arXiv:2606.25832v1 Announce Type: new Abstract: Achieving strong optimization generalization across diverse optimization problems while requiring limited training resources remains a challenging probl

safetyarxiv-cs-lg
25 Jun 2026
Safety

More Soviet-style deceptions that I never expected to see in the United States 😢

DGX agent

Gary Marcus expresses concern about deceptive practices in the United States that he compares to Soviet-style tactics, suggesting he has observed propaganda, misinformation, or manipulative government

safetygary-marcus--x
25 Jun 2026
Safety

Narrative Feature or Structured Feature? A Study of Large Language Models to Identify Cancer Patients at Risk of Heart Failure

DGX agent

arXiv:2403.11425v4 Announce Type: replace-cross Abstract: Cancer treatments are known to introduce cardiotoxicity, negatively impacting outcomes and survivorship. Identifying cancer patients at risk o

safetyarxiv-cs-cl
25 Jun 2026
Safety

Neglected Free Lunch from Post-training: Progress Advantage for LLM Agents

DGX agent

arXiv:2606.26080v1 Announce Type: new Abstract: Process reward models enable fine-grained, step-level evaluation of LLMs, yet building them for agentic settings remains prohibitively difficult: long-h

safetyarxiv-cs-lg
25 Jun 2026
Safety

Neural Machine Translation for Low-Resource Tangkhul--English

DGX agent

arXiv:2606.25365v1 Announce Type: new Abstract: We present a study on low-resource machine translation for the Tangkhul-English (nmf-en) language pair. Tangkhul is a severely under-resourced Tibeto-Bu

safetyarxiv-cs-cl
25 Jun 2026
Safety

noah smith, ignoring that Anthropic’s rumored profitable quarter is a. during peak tokenmaxxing b, before recent Chinese advances c. in a qu…

DGX agent

noah smith, ignoring that Anthropic’s rumored profitable quarter is a. during peak tokenmaxxing b, before recent Chinese advances c. in a quarter in which Elon gave them a big one time subsidy. that’s

safetygary-marcus--x
25 Jun 2026
Safety

Offline Multi-agent Continual Cooperation via Skill Partition and Reuse

DGX agent

arXiv:2606.25389v1 Announce Type: new Abstract: Extracting skills from multi-agent offline dataset improves learning efficiency via sharing task-invariant coordination skills among tasks. In settings

safetyarxiv-cs-ai
25 Jun 2026
Safety

On-Policy Self-Distillation with Sampled Demonstrations Reduces Output Diversity

DGX agent

arXiv:2606.26091v1 Announce Type: new Abstract: On-policy self-distillation achieves strong pass@1 accuracy by using a single model as both teacher and student, with the teacher conditioned on a corre

safetyarxiv-cs-lg
25 Jun 2026
Safety

One Body, Two Minds: Variable Autonomy Approach for a Co-embodied Robotic Hand

DGX agent

arXiv:2606.25575v1 Announce Type: new Abstract: Assistive robotic systems face a fundamental trade-off: fully autonomous systems lack user agency, while fully user-controlled systems demand continuous

safetyarxiv-cs-ro
25 Jun 2026
Safety

OPERA: Aligning Open-Ended Reasoning via Objective Perplexity-based Reinforcement Learning

DGX agent

arXiv:2606.25757v1 Announce Type: new Abstract: Reinforcement Learning (RL) has enabled LLMs to excel in objective reasoning tasks such as mathematics and code generation. However, applying RL to open

safetyarxiv-cs-cl
25 Jun 2026
Safety

Paid Voices vs. Public Feeds: Interpretable Cross-Platform Theme-Based Analysis of Climate Discourse

DGX agent

arXiv:2601.13317v2 Announce Type: replace Abstract: Climate discourse online shapes public understanding of climate change and informs political and policy debate, yet it unfolds across structurally d

safetyarxiv-cs-cl
25 Jun 2026
Safety

People are not getting how significant this news – The White House delaying GPT 5.6 – is. Here’s the context and a suggestion. Both @pmarca …

DGX agent

People are not getting how significant this news – The White House delaying GPT 5.6 – is. Here’s the context and a suggestion. Both @pmarca and @DavidSacks did everything in their power to keep the Wh

safetygary-marcus--x
25 Jun 2026
Safety

PERRY: Policy Evaluation with Confidence Intervals using Auxiliary Data

DGX agent

arXiv:2507.20068v3 Announce Type: replace Abstract: Off-policy evaluation (OPE) methods estimate the value of a new reinforcement learning (RL) policy prior to deployment. Recent advances have shown t

safetyarxiv-cs-lg
25 Jun 2026
Safety

Phoneme-Level Mispronunciation Screening in Polish-Speaking Children with an Explainable Assistant

DGX agent

arXiv:2606.25181v1 Announce Type: cross Abstract: Early identification of speech sound errors in children is often limited by access to specialists, motivating lightweight screening tools that can ope

safetyarxiv-cs-ai
25 Jun 2026
Safety

Point Cloud Diffusion with Global and Local Reconstruction for Instance-Level 3D Anomaly Detection

DGX agent

arXiv:2606.25740v1 Announce Type: new Abstract: 3D anomaly detection in point clouds is critical for high-precision industrial manufacturing. Reconstruction-based methods have laid a strong foundation

safetyarxiv-cs-cv
25 Jun 2026
Safety

Position Spaces and Graphs

DGX agent

arXiv:2606.25719v1 Announce Type: new Abstract: In this paper, we introduce position graphs, a graph-based reasoning framework based on the formalization of position spaces. This framework utilizes tw

safetyarxiv-cs-ai
25 Jun 2026
Safety

Power-Budgeted Underwater Vehicle Control via Constrained Reinforcement Learning

DGX agent

arXiv:2606.25680v1 Announce Type: new Abstract: Underwater vehicles operate from a fixed onboard energy budget that propulsion rapidly depletes, so a controller that completes its task while drawing l

safetyarxiv-cs-ro
25 Jun 2026
Safety

Reflective VLA: In-Context Action Consequences Make VLAs Generalize

DGX agent

arXiv:2606.25215v1 Announce Type: new Abstract: Most vision-language-action (VLA) models are reactive: they predict the next action from the current instruction and observation, implicitly assuming th

safetyarxiv-cs-cv
25 Jun 2026
Safety

Reliability-Asymmetric Spacecraft Autonomy: Co-Designing a Capable Learned GNC Stack with a Verified, Adaptation-Aware Runtime Shield

DGX agent

arXiv:2606.25366v1 Announce Type: new Abstract: Deep-space missions need onboard autonomy that is both capable and certifiable. Rule-based autonomy is certifiable but brittle, while learned autonomy i

safetyarxiv-cs-ro
25 Jun 2026
Safety

Reward-Conditioned Attention: How Reward Design Shapes What Autonomous Driving Agents See

DGX agent

arXiv:2606.25127v1 Announce Type: new Abstract: We investigate how reward design shapes the internal attention patterns of reinforcement learning agents trained for autonomous driving. Using three Per

safetyarxiv-cs-lg
25 Jun 2026
Safety

RGB: RL Guided Whole-Body MPPI for Humanoid Control

DGX agent

arXiv:2606.25123v1 Announce Type: new Abstract: Humanoid robots require whole-body controllers that are both robust and precise in contact-rich environments. While deep reinforcement learning (RL) ach

safetyarxiv-cs-ro
25 Jun 2026
Safety

RN-D: Discretized Categorical Actors for On-Policy Reinforcement Learning

DGX agent

arXiv:2601.23075v2 Announce Type: replace Abstract: On-policy Reinforcement Learning (RL) remains a dominant paradigm for continuous control, yet standard implementations rely on Gaussian actors and r

safetyarxiv-cs-lg
25 Jun 2026
Safety

ROAD-VLA: Robust Online Adaptation via Self-Distillation for Vision-Language-Action Models

DGX agent

arXiv:2606.25800v1 Announce Type: new Abstract: Effective online adaptation of vision-language-action (VLA) models remains challenging, as sparse rewards provide weak supervision for high-dimensional

safetyarxiv-cs-lg
25 Jun 2026
Safety

SAC^2-Net: Semantic Anchoring and Complementary-Consensus Fusion for Multimodal Micro-Expression Recognition

DGX agent

arXiv:2606.25542v1 Announce Type: new Abstract: Micro-expression recognition (MER) is challenging due to subtle facial movements, limited data, and the ambiguous relationship between Action Units (AUs

safetyarxiv-cs-cv
25 Jun 2026
Safety

Safe Learning Control with Optimality and Stability Guarantees

DGX agent

arXiv:2501.15373v2 Announce Type: replace-cross Abstract: Merely pursuing performance may adversely affect safety, while a conservative policy for safe exploration will degrade the performance. How to

safetyarxiv-cs-lg
25 Jun 2026
Safety

SAGE-Nav: Leveraging LLM Planning and Alignment Fusion for Hierarchical Scene Graph-Guided Navigation

DGX agent

arXiv:2606.25497v1 Announce Type: new Abstract: Object-Goal Navigation (ObjNav) requires embodied agents to autonomously locate specified targets using only egocentric visual observations. Existing mo

safetyarxiv-cs-ro
25 Jun 2026
Safety

ScalingAR: Scaling Confidence for Autoregressive Image Generation

DGX agent

arXiv:2509.26376v3 Announce Type: replace Abstract: Test-time strategies have shown remarkable success in improving large language models, but their application to next-token prediction (NTP) autoregr

safetyarxiv-cs-cv
25 Jun 2026
Safety

Semantic Consistency Policy Optimization for Reinforcement Learning of LLM Agents

DGX agent

arXiv:2606.25852v1 Announce Type: new Abstract: Group-based reinforcement learning effectively post-trains LLM agents for long-horizon, sparse-reward tasks by deriving step-level credit from trajector

safetyarxiv-cs-lg
25 Jun 2026
Safety

Solving Markov Decision Processes with Future Information via MPC

DGX agent

arXiv:2606.24991v1 Announce Type: cross Abstract: Model Predictive Control (MPC) is widely used in industrial and robotic systems for enforcing constraints and embedding domain knowledge through finit

safetyarxiv-cs-lg
25 Jun 2026
Safety

Specifying AI-SDLC Processes: A Protocol Language for Human-Agent Boundaries

DGX agent

arXiv:2606.20615v2 Announce Type: replace Abstract: AI agents now participate as first-class team members across the software development lifecycle, yet no specification language exists for expressing

safetyarxiv-cs-ai
25 Jun 2026
← Previous
1…8081828384…267
Next →