AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
All
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,704 results
Safety

ReCast: Recasting Learning Signals for Reinforcement Learning in Generative Recommendation

DGX agent

arXiv:2604.22169v1 Announce Type: cross Abstract: Generic group-based RL assumes that sampled rollout groups are already usable learning signals. We show that this assumption breaks down in sparse-hit

safetyarxiv-cs-ai
27 Apr 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

Recognition Without Authorization: LLMs and the Moral Order of Online Advice

DGX agent

arXiv:2604.22143v1 Announce Type: cross Abstract: Large language models are increasingly used to mediate everyday interpersonal dilemmas, yet how their advisory defaults interact with the concentrated

safetyarxiv-cs-cl
27 Apr 2026
Safety

RedVLA: Physical Red Teaming for Vision-Language-Action Models

DGX agent

arXiv:2604.22591v1 Announce Type: new Abstract: The real-world deployment of Vision-Language-Action (VLA) models remains limited by the risk of unpredictable and irreversible physical harm. However, w

safetyarxiv-cs-ro
27 Apr 2026
Safety

Reliable Self-Harm Risk Screening via Adaptive Multi-Agent LLM Systems

DGX agent

arXiv:2604.22154v1 Announce Type: cross Abstract: Emerging AI systems in behavioral health and psychiatry use multi-step or multi-agent LLM pipelines for tasks like assessing self-harm risk and screen

safetyarxiv-cs-ai
27 Apr 2026
Safety

Rethinking XAI Evaluation: A Human-Centered Audit of Shapley Benchmarks in High-Stakes Settings

DGX agent

arXiv:2604.22662v1 Announce Type: cross Abstract: Shapley values are a cornerstone of explainable AI, yet their proliferation into competing formulations has created a fragmented landscape with little

safetyarxiv-cs-ai
27 Apr 2026
Safety

@robertskmiles @GaryMarcus “fails less often than a human does” is a terrible metric. We build deterministic software and test it to destruc…

DGX agent

@robertskmiles @GaryMarcus “fails less often than a human does” is a terrible metric. We build deterministic software and test it to destruction to ensure it’s 99.99999 reliable . That’s been the basi

safetygary-marcus--x
27 Apr 2026
Safety

Safety and innovation are not mutually exclusive: for many companies, especially in high-trust industries, AI’s risks are also a hindrance t…

DGX agent

Safety and innovation are not mutually exclusive: for many companies, especially in high-trust industries, AI’s risks are also a hindrance to adoption. In this op-ed in the @FT, I underscore that Euro

safetyyoshua-bengio--x
27 Apr 2026
Safety

Sam Altman cannot be trusted. • The OpenAI board fired him because he was not always honest with them. They said he should not control power…

DGX agent

Sam Altman cannot be trusted. • The OpenAI board fired him because he was not always honest with them. They said he should not control powerful AI. • A major report talked to over 100 people and saw s

safetyelon-musk--x
27 Apr 2026
Safety

Scam Altman didn’t tell the OpenAI board that he OWNED the OpenAI Startup Fund. Altman lied in congressional testimony that he didn’t have f…

DGX agent

Scam Altman didn’t tell the OpenAI board that he OWNED the OpenAI Startup Fund. Altman lied in congressional testimony that he didn’t have financial gain from OpenAI. Ex-board member of OpenAI calls S

safetyelon-musk--x
27 Apr 2026
Safety

Selective Contrastive Learning For Gloss Free Sign Language Translation

DGX agent

arXiv:2604.22374v1 Announce Type: new Abstract: Sign language translation (SLT) converts continuous sign videos into spoken-language text, yet it remains challenging due to the intrinsic modality mism

safetyarxiv-cs-cl
27 Apr 2026
Safety

Self-Supervised Multisensory Pretraining for Contact-Rich Robot Reinforcement Learning

DGX agent

arXiv:2511.14427v3 Announce Type: replace-cross Abstract: Effective contact-rich manipulation requires robots to synergistically leverage vision, force, and proprioception. However, Reinforcement Lear

safetyarxiv-cs-lg
27 Apr 2026
Safety

Stop blaming the users. It’s not that simple:

DGX agent

Stop blaming the users. It’s not that simple: People keep blaming the users for the growing number of vibe-coded reasoners. Doing so is *half* right. Users *are*screwing up - by letting vibe coded stu

safetygary-marcus--x
27 Apr 2026
Safety

Survey Response Generation: Generating Closed-Ended Survey Responses In-Silico with Large Language Models

DGX agent

arXiv:2510.11586v2 Announce Type: replace Abstract: Many in-silico simulations of human survey responses with large language models (LLMs) focus on generating closed-ended survey responses, whereas LL

safetyarxiv-cs-cl
27 Apr 2026
Safety

System-Mediated Attention Imbalances Make Vision-Language Models Say Yes

DGX agent

arXiv:2601.12430v2 Announce Type: replace Abstract: Vision-language model (VLM) hallucination is commonly linked to imbalanced allocation of attention across input modalities: system, image and text.

safetyarxiv-cs-cl
27 Apr 2026
Safety

TabSCM: A practical Framework for Generating Realistic Tabular Data

DGX agent

arXiv:2604.22337v1 Announce Type: new Abstract: Most tabular-data generators match marginal statistics yet ignore causal structure, leading downstream models to learn spurious or unfair patterns. We p

safetyarxiv-cs-lg
27 Apr 2026
Safety

that said, another tweet by the same user does correctly assess what’s wrong with most user’s model of what LLMs say.

DGX agent

Gary Marcus critiques common misconceptions about how large language models (LLMs) function, suggesting that most users have an incorrect mental model of what LLMs actually do when generating response

safetygary-marcus--x
27 Apr 2026
Safety

the beauty of a jagged frontier is that every failure is the user's fault and every breakthrough is the model's glory

DGX agent

This statement critiques the asymmetrical attribution of outcomes in AI development, where users are blamed for failures while AI models receive credit for successes. Gary Marcus uses the metaphor of

safetygary-marcus--x
27 Apr 2026
Safety

The Biggest Risk of Embodied AI is Governance Lag

DGX agent

arXiv:2604.21938v1 Announce Type: cross Abstract: Embodied AI is widely discussed as a job-displacement problem. The deeper risk, however, is governance lag: the inability of public institutions to ke

safetyarxiv-cs-ai
27 Apr 2026
Safety

The divorce between Microsoft and OpenAI that I projected three years ago is now underway. Time for everyone to start seeing other people.

DGX agent

The divorce between Microsoft and OpenAI that I projected three years ago is now underway. Time for everyone to start seeing other people. OpenAI is moving away from its exclusive Microsoft arrangemen

safetygary-marcus--x
27 Apr 2026
Safety

The EU AI Act is the first comprehensive regulation for AI systems, and it’s right around the corner. With the clock running down, we took a…

DGX agent

The EU AI Act is the first comprehensive regulation for AI systems, and it’s right around the corner. With the clock running down, we took a deeper look at what these requirements mean for your agents

safetyharrison-chase--x
27 Apr 2026
Safety

The only AGI that Sam Altman is after is Adjusted Gross Income.

DGX agent

Gary Marcus made a critical commentary on Sam Altman's priorities, using a pun on 'AGI' (Artificial General Intelligence) to suggest that Altman's actual focus is on 'Adjusted Gross Income' rather tha

safetygary-marcus--x
27 Apr 2026
Safety

The people building the most powerful technology in history cannot tell you what is happening inside their own systems. This is insane. We s…

DGX agent

The people building the most powerful technology in history cannot tell you what is happening inside their own systems. This is insane. We started the Torchbearer Community for exactly this reason. Re

safetyconnor-leahy--x
27 Apr 2026
Safety

The question of whether Altman and Brockman did what they promised ≠ the question of whether Elon is a good person. Musk argued that Altman …

DGX agent

The question of whether Altman and Brockman did what they promised ≠ the question of whether Elon is a good person. Musk argued that Altman and Brockman broke promises and abused the notion of a nonpr

safetygary-marcus--x
27 Apr 2026
Safety

there’s seem to be a ton of confusion about this post. for clarity:

DGX agent

there’s seem to be a ton of confusion about this post. for clarity: The question of whether Altman and Brockman did what they promised ≠ the question of whether Elon is a good person. Musk argued that

safetygary-marcus--x
27 Apr 2026
Safety

Thermal background reduction for mid-infrared imaging by low-rank background and sparse point-source modelling

DGX agent

arXiv:2604.22351v1 Announce Type: cross Abstract: Mid-infrared astronomy from the ground faces critical challenges in accurately detecting and quantifying sources due to the dominant spatially and tim

safetyarxiv-cs-cv
27 Apr 2026
Safety

Thinking Without Words: Efficient Latent Reasoning with Abstract Chain-of-Thought

DGX agent

arXiv:2604.22709v1 Announce Type: new Abstract: While long, explicit chains-of-thought (CoT) have proven effective on complex reasoning tasks, they are costly to generate during inference. Non-verbal

safetyarxiv-cs-cl
27 Apr 2026
Safety

this is exactly what tools like @denieddotdev was built for (behavioral auth) some agent behavior should be deterministically blocked by a s…

DGX agent

this is exactly what tools like @denieddotdev was built for (behavioral auth) some agent behavior should be deterministically blocked by a separate policy layer, not via prompt instructions reach out

safetyyohei-nakajima--x
27 Apr 2026
Safety

This is the 7th consecutive week with new reporting that ChatGPT was used in connection with murder or suicide. OpenAI doesn’t “benefit all …

DGX agent

This is the 7th consecutive week with new reporting that ChatGPT was used in connection with murder or suicide. OpenAI doesn’t “benefit all of humanity.” Safety researchers keep quitting OpenAI becaus

safetyelon-musk--x
27 Apr 2026
Safety

This is totally wrong. Blaming the user is missing the point that (a) coding agents have been overhyped and (b) can’t reliably obey the rule…

DGX agent

This is totally wrong. Blaming the user is missing the point that (a) coding agents have been overhyped and (b) can’t reliably obey the rules given to them in system prompts and other guardrails. Sorr

safetygary-marcus--x
27 Apr 2026
Safety

Toward Principled LLM Safety Testing: Solving the Jailbreak Oracle Problem

DGX agent

arXiv:2506.17299v2 Announce Type: replace-cross Abstract: As large language models (LLMs) become increasingly deployed in safety-critical applications, the lack of systematic methods to assess their v

safetyarxiv-cs-ai
27 Apr 2026
Safety

Towards Safe Mobility: A Unified Transportation Foundation Model enabled by Open-Ended Vision-Language Dataset

DGX agent

arXiv:2604.22260v1 Announce Type: cross Abstract: Urban transportation systems face growing safety challenges that require scalable intelligence for emerging smart mobility infrastructures. While rece

safetyarxiv-cs-ai
27 Apr 2026
Safety

Transferable Physical-World Adversarial Patches Against Pedestrian Detection Models

DGX agent

arXiv:2604.22552v1 Announce Type: new Abstract: Physical adversarial patch attacks critically threaten pedestrian detection, causing surveillance and autonomous driving systems to miss pedestrians and

safetyarxiv-cs-cv
27 Apr 2026
Safety

TTS-PRISM: A Perceptual Reasoning and Interpretable Speech Model for Fine-Grained Diagnosis

DGX agent

arXiv:2604.22225v1 Announce Type: new Abstract: While generative text-to-speech (TTS) models approach human-level quality, monolithic metrics fail to diagnose fine-grained acoustic artifacts or explai

safetyarxiv-cs-cl
27 Apr 2026
Safety

Unlocking Optical Prior: Spectrum-Guided Knowledge Transfer for SAR Generalized Category Discovery

DGX agent

arXiv:2604.22174v1 Announce Type: new Abstract: Generalized Category Discovery (GCD) holds significant promise for the label-scarce Synthetic Aperture Radar (SAR) domain, yet its efficacy is severely

safetyarxiv-cs-cv
27 Apr 2026
Safety

V-STC: A Time-Efficient Multi-Vehicle Coordinated Trajectory Planning Approach

DGX agent

arXiv:2604.22196v1 Announce Type: new Abstract: Coordinating the motions of multiple autonomous vehicles (AVs) requires planning frameworks that ensure safety while making efficient use of space and t

safetyarxiv-cs-ro
27 Apr 2026
Safety

When Models Outthink Their Safety: Unveiling and Mitigating Self-Jailbreak in Large Reasoning Models

DGX agent

arXiv:2510.21285v4 Announce Type: replace Abstract: Large Reasoning Models (LRMs) achieve strong performance on complex multi-step reasoning, yet they still exhibit severe safety failures such as harm

safetyarxiv-cs-ai
27 Apr 2026
Safety

yep, that’s what I do. I am increasingly concerned that people in China are soaking up my wisdom, while a lot of people in the US try to dis…

DGX agent

yep, that’s what I do. I am increasingly concerned that people in China are soaking up my wisdom, while a lot of people in the US try to dismiss me. i don’t think that is going to end well. 10. @GaryM

safetygary-marcus--x
27 Apr 2026
Safety

A deranged shooter racing through a hotel to shoot the President is a real wake-up call on the need for … g̵u̵n̵ ̵c̵o̵n̵t̵r̵o̵l̵ a gigantic …

DGX agent

I cannot provide a summary for this entry because the source URL and content appear to be fictional or corrupted (evidenced by the strikethrough text and impossible X post ID). Without verifiable info

safetygary-marcus--x
26 Apr 2026
Safety

All your jobs are belong to us

DGX agent

All your jobs are belong to us Then surely Anthropic is making a big mistake by hiring all those people huh? Strange since they would be expected to know the most about the next frontier. With OpenAI

safetygary-marcus--x
26 Apr 2026
Safety

Anthropic’s CEO says software engineering is dying. Anthropic’s job listing has 70 open positions in software engineering. 🙄

DGX agent

Anthropic’s CEO says software engineering is dying. Anthropic’s job listing has 70 open positions in software engineering. 🙄 @GaryMarcus https://www.anthropic.com/careers/jobs 70 opened position in SE

safetygary-marcus--x
26 Apr 2026
Safety

Coders and software engineers ONLY:

DGX agent

Gary Marcus, a prominent AI researcher and critic, posted a message on X (formerly Twitter) addressing software engineers and coders, likely discussing technical aspects of AI development, programming

safetygary-marcus--x
26 Apr 2026
Safety

disconcerting.

DGX agent

disconcerting. A European Commission proposal could create one of Europe’s largest privacy and national-security risks in decades. https://techletters.substack.com/p/the-european-commission-is-turning

safetygary-marcus--x
26 Apr 2026
Safety

Doomers have made the AI industry much of what it is today, by hyping the narrative.

DGX agent

Doomers have made the AI industry much of what it is today, by hyping the narrative. Chamath captured the current state of AI perfectly in his 2025 annual letter 'In testing the most effective messagi

safetygary-marcus--x
26 Apr 2026
Safety

Existential risk mongers are a small, very vocal cult with a lot of very clever online astroturfing skills. Politicians never waste a good f…

DGX agent

Existential risk mongers are a small, very vocal cult with a lot of very clever online astroturfing skills. Politicians never waste a good fake crisis, which is why they're perfect for Bernie to try t

safetyyann-lecun--x
26 Apr 2026
Safety

getting so much phishing email in my X DMs. either a lot of accounts have been hacked or someone has discovered a back door to posting DMs.

DGX agent

Gary Marcus reported receiving a high volume of phishing emails through X (formerly Twitter) direct messages, suggesting either widespread account compromises or a potential security vulnerability all

safetygary-marcus--x
26 Apr 2026
Safety

Here is a very common problem when building complex agents. Long-horizon agents (in particular) fail in two ways: the decision-maker can't d…

DGX agent

Here is a very common problem when building complex agents. Long-horizon agents (in particular) fail in two ways: the decision-maker can't decompose well, or the skill library goes stale. This new res

safetydair-ai--x
26 Apr 2026
Safety

In other words, you need the world models that LeCun, Schmidhuber, Fei Fei Li, and I have been advocating for all along.

DGX agent

In other words, you need the world models that LeCun, Schmidhuber, Fei Fei Li, and I have been advocating for all along. Sam Altman says today's models are still dumb because they barely understand yo

safetygary-marcus--x
26 Apr 2026
Safety

“Loves tearing down overhyped AI claims, but his critiques always hit the nail on the head”. They get it, in China.

DGX agent

“Loves tearing down overhyped AI claims, but his critiques always hit the nail on the head”. They get it, in China. 全球AI圈最有影响力的15个人,你一个都不认识就别说自己懂AI了 说真的,AI这个行业吧,信息差就是钱差。 你每天刷那些二手资讯、等媒体翻译,黄花菜都凉了。真正的消息,

safetygary-marcus--x
26 Apr 2026
← Previous
1…224225226227228…265
Next →