AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,494 results
Safety

RASFT: Rollout-Adaptive Supervised Fine-Tuning for Reasoning

DGX agent

arXiv:2606.07006v1 Announce Type: cross Abstract: Supervised fine-tuning (SFT) is a prevailing method for adapting large language models to reasoning tasks by imitating offline expert demonstrations,

safetyarxiv-cs-cl
8 Jun 2026
Safety
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

RAVEN: Retrieval-Augmented Vulnerability Exploration Network for Memory Corruption Analysis in User Code and Binary Programs

DGX agent

arXiv:2604.17948v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities across various cybersecurity tasks, including vulnerability classificat

safetyarxiv-cs-ai
8 Jun 2026
Safety

Robotic Policy Adaptation via Weight-Space Meta-Learning

DGX agent

arXiv:2606.07217v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models are emerging as a promising paradigm for robotic manipulation, enabling general-purpose policies trained from larg

safetyarxiv-cs-cv
8 Jun 2026
Safety

Robots Need More than VLA and World Models

DGX agent

arXiv:2606.06556v1 Announce Type: new Abstract: Generalist robot intelligence is often framed as a policy-scaling problem: collect more robot demonstrations, train larger Vision-Language-Action (VLA)

safetyarxiv-cs-ro
8 Jun 2026
Safety

Self-evolving LLM agents with in-distribution Optimization

DGX agent

arXiv:2606.07367v1 Announce Type: new Abstract: Large Language Models (LLMs) have recently emerged as powerful controllers for interactive agents in complex environments, yet training them to perform

safetyarxiv-cs-lg
8 Jun 2026
Safety

Semantic-Structural Alignment for Generative Pictorial Charts

DGX agent

arXiv:2606.06498v1 Announce Type: cross Abstract: Traditional statistical graphics are precise but often lack the visual appeal, memorability, and engagement of pictorial charts. We present a generati

safetyarxiv-cs-cv
8 Jun 2026
Safety

SERNF: Sample-Efficient Real-World Dexterous Policy Fine-Tuning via Action-Chunked Critics and Normalizing Flows

DGX agent

arXiv:2602.09580v4 Announce Type: replace-cross Abstract: Real-world fine-tuning of dexterous manipulation policies remains challenging due to limited real-world interaction budgets and highly multimo

safetyarxiv-cs-lg
8 Jun 2026
Safety

Silverfort brings runtime identity controls to Microsoft Copilot Studio agents

DGX agent

Identity security company Silverfort Inc. today launched an integration that applies its identity and access controls to artificial intelligence agents built into Microsoft Corp.’s Copilot Studio, enf

safetysiliconangle
8 Jun 2026
Safety

Simulation-Driven Imitation Learning for Biosignals-Free Shared-Autonomy Prosthetic Grasping

DGX agent

arXiv:2606.07389v1 Announce Type: new Abstract: Biosignals-free shared-autonomy control of upper-limb prosthetic hands aims to enable natural and low-effort manipulation without relying on EMG or othe

safetyarxiv-cs-ro
8 Jun 2026
Safety

SlimSearcher: Training Efficiency-Aware Web Agents via Adaptive Reward Gating

DGX agent

arXiv:2606.07074v1 Announce Type: cross Abstract: Deep research agents have demonstrated remarkable capabilities in complex information-seeking tasks, yet this power comes at a steep computational cos

safetyarxiv-cs-ai
8 Jun 2026
Safety

Socratic-SWE: Self-Evolving Coding Agents via Trace-Derived Agent Skills

DGX agent

arXiv:2606.07412v1 Announce Type: cross Abstract: LLM-driven software engineering agents have become a central testbed for real-world language-model capability, yet their training remains limited by t

safetyarxiv-cs-ai
8 Jun 2026
Safety

SpaceX goes public Friday at ~94x revenue. Across 45 years of data, IPOs that debut above 40x sales underperform the market by 58% over the …

DGX agent

SpaceX goes public Friday at ~94x revenue. Across 45 years of data, IPOs that debut above 40x sales underperform the market by 58% over the next 3 years, and by 76% style-adjusted. The golden rule of

safetygary-marcus--x
8 Jun 2026
Safety

Stable Reasoning, Unstable Responses: Mitigating LLM Deception via Stability Asymmetry

DGX agent

arXiv:2603.26846v2 Announce Type: replace-cross Abstract: As Large Language Models (LLMs) expand in capability and application scope, their trustworthiness becomes critical. A vital risk is intrinsic

safetyarxiv-cs-ai
8 Jun 2026
Safety

Step-Wise Refusal Dynamics in Autoregressive and Diffusion Language Models

DGX agent

arXiv:2602.02600v3 Announce Type: replace-cross Abstract: Diffusion language models (DLMs) have recently emerged as a competitive alternative to autoregressive (AR) models, offering parallel decoding,

safetyarxiv-cs-ai
8 Jun 2026
Safety

SV-Detect: AI-generated Text Detection with Steering Vectors

DGX agent

arXiv:2606.07313v1 Announce Type: cross Abstract: Detecting machine-generated text is especially difficult under distribution shift, such as transfer across domains, source models, and editing attacks

safetyarxiv-cs-ai
8 Jun 2026
Safety

Sycophantic Praise: Evaluating Excessive Praise in Language Models

DGX agent

arXiv:2606.07441v1 Announce Type: new Abstract: Sycophancy in language models is typically studied as excessive agreement or validation, while explicit praise and flattery have received comparatively

safetyarxiv-cs-cl
8 Jun 2026
Safety

T-GMP: Terrain-conditioned Generative Motion Priors for Versatile and Natural Humanoid Locomotion

DGX agent

arXiv:2606.06944v1 Announce Type: new Abstract: Achieving both anthropomorphic naturalness and robust terrain traversal remains a fundamental challenge in humanoid locomotion. Existing Reinforcement L

safetyarxiv-cs-ro
8 Jun 2026
Safety

Teach a Reward Model to Correct Itself: Reward Guided Adversarial Failure Discovery for Robust Reward Modeling

DGX agent

arXiv:2507.06419v3 Announce Type: replace Abstract: Reward modeling (RM), which captures human preferences to align large language models (LLMs), is increasingly employed in tasks such as model finetu

safetyarxiv-cs-cl
8 Jun 2026
Safety

Teaching the Way, Not the Answer: Privileged Tutoring Distillation for Multimodal Policy Optimization

DGX agent

arXiv:2606.07000v1 Announce Type: new Abstract: Recent post-training methods, particularly Reinforcement Learning with Verifiable Rewards (RLVR), have significantly enhanced the reasoning ability of L

safetyarxiv-cs-ai
8 Jun 2026
Safety

The discovery of the effects of women employment participation on the fertility of developing countries: A panel data approach

DGX agent

arXiv:2606.07093v1 Announce Type: new Abstract: The fertility trend in developing countries has experienced a significant decline in the last few decades; at the same time, the role of women in the wo

safetyarxiv-cs-lg
8 Jun 2026
Safety

The legendary investor Vinod Khosla is in the news because of questions about his integrity. I want to share my own experience: he lied vici…

DGX agent

The legendary investor Vinod Khosla is in the news because of questions about his integrity. I want to share my own experience: he lied viciously about me (a skeptic of a company he stands to make bil

safetygary-marcus--x
8 Jun 2026
Safety

The Trump administration relaunches efforts to block state AI laws; Sen. Blackburn is leading negotiations and pushing KOSA as part of an AI preemption package (Axios)

DGX agent

Axios: The Trump administration relaunches efforts to block state AI laws; Sen. Blackburn is leading negotiations and pushing KOSA as part of an AI preemption package — The White House is negotiating

safetytechmeme
8 Jun 2026
Safety

There is much AI in the news that I initally completely misread this headline 🤣

DGX agent

Gary Marcus shared a humorous post on X about misreading an AI-related headline due to the current saturation of AI news coverage. The post reflects on how the prevalence of AI stories in media can le

safetygary-marcus--x
8 Jun 2026
Safety

this is gonna go great

DGX agent

this is gonna go great xAI’s development of artificial intelligence is “a mess”, Co-Executive Editor @mvpeers says. “Elon has fired most of the people who he originally hired at xAI.' 'He has a tenden

safetygary-marcus--x
8 Jun 2026
Safety

Translate-R1: Cost-Aware Translation Tool Use via Reinforcement Learning

DGX agent

arXiv:2606.06835v1 Announce Type: new Abstract: The performance gap across languages in LLMs is well documented, and closing it natively requires pretraining or fine-tuning on corpora that, for most l

safetyarxiv-cs-cl
8 Jun 2026
Safety

TrioPose: Native Triple-Stream Diffusion Transformers for Pose-Guided Text-to-Image Generation

DGX agent

arXiv:2606.07053v1 Announce Type: new Abstract: Pose-guided text-to-image generation often suffers from limb distortions and feature crosstalk in complex multi-person scenarios. While existing UNet-ba

safetyarxiv-cs-cv
8 Jun 2026
Safety

update: the person who posted the original has (rare!) acknowledged the error and delete the original post.

DGX agent

Gary Marcus posted an update on X noting that the original poster of a viral claim acknowledged their error and deleted the post, highlighting a rare instance of public correction on social media. The

safetygary-marcus--x
8 Jun 2026
Safety

VALUEFLOW: Toward Pluralistic and Steerable Value-based Alignment in Large Language Models

DGX agent

arXiv:2602.03160v2 Announce Type: replace Abstract: Aligning Large Language Models (LLMs) with the diverse spectrum of human values remains a central challenge: preference-based methods often fail to

safetyarxiv-cs-ai
8 Jun 2026
Safety

Watch, Remember, Reason: Human-View Video Understanding with MLLMs

DGX agent

arXiv:2606.07433v1 Announce Type: cross Abstract: Video understanding is being rapidly transformed by multimodal large language models (MLLMs), as research moves from short clips to long, multimodal,

safetyarxiv-cs-ai
8 Jun 2026
Safety

What Do People Actually Want From AI? Mapping Preference Plurality

DGX agent

arXiv:2606.06674v1 Announce Type: new Abstract: Large Language Models (LLMs) are often fine-tuned through Reinforcement Learning from Human Feedback (RLHF) to align with people's preferences and value

safetyarxiv-cs-cl
8 Jun 2026
Safety

What if AI turns out to be a lot less profitable than we have been told? @kirtlenus explores the economic risks https://thecritic.co.uk/the-…

DGX agent

This article explores potential economic downsides to artificial intelligence, questioning assumptions about AI's profitability and examining financial risks that may have been underestimated or overl

safetygary-marcus--x
8 Jun 2026
Safety

What Is My Robot Thinking? Design Considerations for Transparent and Trustworthy Shared Autonomy

DGX agent

arXiv:2606.06870v1 Announce Type: new Abstract: Assistive robots operating under shared autonomy must balance user control with autonomous assistance. Because robot actions depend on internal intent i

safetyarxiv-cs-ro
8 Jun 2026
Safety

What Matters When Cotraining Robot Manipulation Policies on Everyday Human Videos?

DGX agent

arXiv:2606.06627v1 Announce Type: cross Abstract: Human video datasets used for cotraining robot manipulation policies largely consist of curated demonstrations where motions are orchestrated to resem

safetyarxiv-cs-ai
8 Jun 2026
Safety

Where to Touch, How to Contact: Hierarchical RL-MPC Framework for Geometry-Aware Long-Horizon Dexterous Manipulation

DGX agent

arXiv:2601.10930v3 Announce Type: replace Abstract: A key challenge in contact-rich dexterous manipulation is the need to jointly reason over global geometry and nonsmooth contact dynamics. End-to-end

safetyarxiv-cs-ro
8 Jun 2026
Safety

Would you buy SpaceX at the proposed price of $135 a share?

DGX agent

Gary Marcus poses a hypothetical investment question about SpaceX's valuation at $135 per share, likely exploring perspectives on the company's market value, growth prospects, and investment merit. Th

safetygary-marcus--x
8 Jun 2026
Safety

a riff on https://quoteinvestigator.com/2011/07/09/poker-patsy/

DGX agent

Gary Marcus likely references the 'poker patsy' concept, commonly attributed to Warren Buffett, which warns that if you've been in a poker game for a while and haven't identified the fool at the table

safetygary-marcus--x
7 Jun 2026
Safety

AI is mentioned more often in SpaceX’s S-1 than Jesus is mentioned in the Bible.

DGX agent

Gary Marcus compares the frequency of 'AI' mentions in SpaceX's SEC filing (S-1 document) to the frequency of 'Jesus' mentions in the Bible, suggesting that SpaceX emphasizes artificial intelligence e

safetygary-marcus--x
7 Jun 2026
Safety

another take on hard fork and IPOs (again i haven’t listened)

DGX agent

another take on hard fork and IPOs (again i haven’t listened) I just listened to it. In fairness, they never said anything along the lines of encouraging anyone to buy the IPOs. They also discuss how

safetygary-marcus--x
7 Jun 2026
Safety

as always beware of scammers; here’s a new one, a fake account trying to fool people into thinking i am recommending some garbage. please be…

DGX agent

Gary Marcus warned X users about a scam involving fake accounts impersonating him to fraudulently promote products or services by falsely claiming his endorsement. The post alerts followers to be caut

safetygary-marcus--x
7 Jun 2026
Safety

blast from the past 3.5 years ago; some things have changed (esp. coding and math, via neurosymbolic techniques) but many haven’t:

DGX agent

blast from the past 3.5 years ago; some things have changed (esp. coding and math, via neurosymbolic techniques) but many haven’t: Bottom line: From the outset Large Language Models like GPT-3 have gr

safetygary-marcus--x
7 Jun 2026
Safety

California Is Blocking a Federal Audit of Its Voter Rolls California allows first-time voters to register using forms of ID that most Americ…

DGX agent

California Is Blocking a Federal Audit of Its Voter Rolls California allows first-time voters to register using forms of ID that most Americans would find surprising, including: -Gym membership card -

safetyelon-musk--x
7 Jun 2026
Safety

From fastest growing company to “worst value among its peers” in 18 months. Why, oh why, stick with the CEO?

DGX agent

From fastest growing company to “worst value among its peers” in 18 months. Why, oh why, stick with the CEO? PitchBook's analysts just ranked OpenAI last for value among its AI peers. Not last for cap

safetygary-marcus--x
7 Jun 2026
Safety

If you aren’t one of the banks running the SpaceX IPO, you’re the mark:

DGX agent

This post likely critiques the SpaceX IPO process, suggesting that only major financial institutions acting as underwriters benefit substantially from the offering while retail investors and others ar

safetygary-marcus--x
7 Jun 2026
Safety

Imaginary conversations that might actually have happened Act I Sam: We missed all our metrics, Anthropic and Google have gained on us. Give…

DGX agent

Imaginary conversations that might actually have happened Act I Sam: We missed all our metrics, Anthropic and Google have gained on us. Give us 40 billion dollars. Masa: No way! Sam: If you don't, we

safetygary-marcus--x
7 Jun 2026
Safety

Indeed, I don’t think Trump has thought through the implications.

DGX agent

Indeed, I don’t think Trump has thought through the implications. This will simply provide even more incentive- as if any was needed - for UK/European players & states to head towards Sovereign AI, ai

safetygary-marcus--x
7 Jun 2026
Safety

Now that the public is waking up to the enormous costs to society of AI, the new dirty trick is to spin the costs of AI as if they were marg…

DGX agent

Gary Marcus argues that as public awareness grows regarding AI's societal costs, a rhetorical strategy is being employed to reframe or minimize these costs by presenting them as marginal or inevitable

safetygary-marcus--x
7 Jun 2026
Safety

One of the more interesting takes on positive alignment that have recently come out-it’s long and interesting, combining philosophy and trai…

DGX agent

One of the more interesting takes on positive alignment that have recently come out-it’s long and interesting, combining philosophy and training setups (eg reward proposals), and worth a read. What ha

safetydan-hendrycks--x
7 Jun 2026
Safety

Only in an America can an industry that has collectively lost over half a trillion dollars —at a pace of roughly a million dollars a minute …

DGX agent

Gary Marcus critiques the AI industry for its massive financial losses, noting that collectively the sector has lost over half a trillion dollars at an alarming rate of approximately one million dolla

safetygary-marcus--x
7 Jun 2026
← Previous
1…150151152153154…302
Next →