AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,704 results
27 Apr 2026

ReCast: Recasting Learning Signals for Reinforcement Learning in Generative Recommendation

SafetyDGX agent

arXiv:2604.22169v1 Announce Type: cross Abstract: Generic group-based RL assumes that sampled rollout groups are already usable learning signals. We show that this assumption breaks down in sparse-hit

Recognition Without Authorization: LLMs and the Moral Order of Online Advice

SafetyDGX agent

arXiv:2604.22143v1 Announce Type: cross Abstract: Large language models are increasingly used to mediate everyday interpersonal dilemmas, yet how their advisory defaults interact with the concentrated

RedVLA: Physical Red Teaming for Vision-Language-Action Models

SafetyDGX agent

arXiv:2604.22591v1 Announce Type: new Abstract: The real-world deployment of Vision-Language-Action (VLA) models remains limited by the risk of unpredictable and irreversible physical harm. However, w


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Reliable Self-Harm Risk Screening via Adaptive Multi-Agent LLM Systems

SafetyDGX agent

arXiv:2604.22154v1 Announce Type: cross Abstract: Emerging AI systems in behavioral health and psychiatry use multi-step or multi-agent LLM pipelines for tasks like assessing self-harm risk and screen

Rethinking XAI Evaluation: A Human-Centered Audit of Shapley Benchmarks in High-Stakes Settings

SafetyDGX agent

arXiv:2604.22662v1 Announce Type: cross Abstract: Shapley values are a cornerstone of explainable AI, yet their proliferation into competing formulations has created a fragmented landscape with little

@robertskmiles @GaryMarcus “fails less often than a human does” is a terrible metric. We build deterministic software and test it to destruc…

SafetyDGX agent

@robertskmiles @GaryMarcus “fails less often than a human does” is a terrible metric. We build deterministic software and test it to destruction to ensure it’s 99.99999 reliable . That’s been the basi

Safety and innovation are not mutually exclusive: for many companies, especially in high-trust industries, AI’s risks are also a hindrance t…

SafetyDGX agent

Safety and innovation are not mutually exclusive: for many companies, especially in high-trust industries, AI’s risks are also a hindrance to adoption. In this op-ed in the @FT, I underscore that Euro

Sam Altman cannot be trusted. • The OpenAI board fired him because he was not always honest with them. They said he should not control power…

SafetyDGX agent

Sam Altman cannot be trusted. • The OpenAI board fired him because he was not always honest with them. They said he should not control powerful AI. • A major report talked to over 100 people and saw s

Scam Altman didn’t tell the OpenAI board that he OWNED the OpenAI Startup Fund. Altman lied in congressional testimony that he didn’t have f…

SafetyDGX agent

Scam Altman didn’t tell the OpenAI board that he OWNED the OpenAI Startup Fund. Altman lied in congressional testimony that he didn’t have financial gain from OpenAI. Ex-board member of OpenAI calls S

Selective Contrastive Learning For Gloss Free Sign Language Translation

SafetyDGX agent

arXiv:2604.22374v1 Announce Type: new Abstract: Sign language translation (SLT) converts continuous sign videos into spoken-language text, yet it remains challenging due to the intrinsic modality mism

Self-Supervised Multisensory Pretraining for Contact-Rich Robot Reinforcement Learning

SafetyDGX agent

arXiv:2511.14427v3 Announce Type: replace-cross Abstract: Effective contact-rich manipulation requires robots to synergistically leverage vision, force, and proprioception. However, Reinforcement Lear

Stop blaming the users. It’s not that simple:

SafetyDGX agent

Stop blaming the users. It’s not that simple: People keep blaming the users for the growing number of vibe-coded reasoners. Doing so is *half* right. Users *are*screwing up - by letting vibe coded stu

Survey Response Generation: Generating Closed-Ended Survey Responses In-Silico with Large Language Models

SafetyDGX agent

arXiv:2510.11586v2 Announce Type: replace Abstract: Many in-silico simulations of human survey responses with large language models (LLMs) focus on generating closed-ended survey responses, whereas LL

System-Mediated Attention Imbalances Make Vision-Language Models Say Yes

SafetyDGX agent

arXiv:2601.12430v2 Announce Type: replace Abstract: Vision-language model (VLM) hallucination is commonly linked to imbalanced allocation of attention across input modalities: system, image and text.

TabSCM: A practical Framework for Generating Realistic Tabular Data

SafetyDGX agent

arXiv:2604.22337v1 Announce Type: new Abstract: Most tabular-data generators match marginal statistics yet ignore causal structure, leading downstream models to learn spurious or unfair patterns. We p

that said, another tweet by the same user does correctly assess what’s wrong with most user’s model of what LLMs say.

SafetyDGX agent

Gary Marcus critiques common misconceptions about how large language models (LLMs) function, suggesting that most users have an incorrect mental model of what LLMs actually do when generating response

the beauty of a jagged frontier is that every failure is the user's fault and every breakthrough is the model's glory

SafetyDGX agent

This statement critiques the asymmetrical attribution of outcomes in AI development, where users are blamed for failures while AI models receive credit for successes. Gary Marcus uses the metaphor of

The Biggest Risk of Embodied AI is Governance Lag

SafetyDGX agent

arXiv:2604.21938v1 Announce Type: cross Abstract: Embodied AI is widely discussed as a job-displacement problem. The deeper risk, however, is governance lag: the inability of public institutions to ke

The divorce between Microsoft and OpenAI that I projected three years ago is now underway. Time for everyone to start seeing other people.

SafetyDGX agent

The divorce between Microsoft and OpenAI that I projected three years ago is now underway. Time for everyone to start seeing other people. OpenAI is moving away from its exclusive Microsoft arrangemen

The EU AI Act is the first comprehensive regulation for AI systems, and it’s right around the corner. With the clock running down, we took a…

SafetyDGX agent

The EU AI Act is the first comprehensive regulation for AI systems, and it’s right around the corner. With the clock running down, we took a deeper look at what these requirements mean for your agents

The only AGI that Sam Altman is after is Adjusted Gross Income.

SafetyDGX agent

Gary Marcus made a critical commentary on Sam Altman's priorities, using a pun on 'AGI' (Artificial General Intelligence) to suggest that Altman's actual focus is on 'Adjusted Gross Income' rather tha

The people building the most powerful technology in history cannot tell you what is happening inside their own systems. This is insane. We s…

SafetyDGX agent

The people building the most powerful technology in history cannot tell you what is happening inside their own systems. This is insane. We started the Torchbearer Community for exactly this reason. Re

The question of whether Altman and Brockman did what they promised ≠ the question of whether Elon is a good person. Musk argued that Altman …

SafetyDGX agent

The question of whether Altman and Brockman did what they promised ≠ the question of whether Elon is a good person. Musk argued that Altman and Brockman broke promises and abused the notion of a nonpr

there’s seem to be a ton of confusion about this post. for clarity:

SafetyDGX agent

there’s seem to be a ton of confusion about this post. for clarity: The question of whether Altman and Brockman did what they promised ≠ the question of whether Elon is a good person. Musk argued that

Thermal background reduction for mid-infrared imaging by low-rank background and sparse point-source modelling

SafetyDGX agent

arXiv:2604.22351v1 Announce Type: cross Abstract: Mid-infrared astronomy from the ground faces critical challenges in accurately detecting and quantifying sources due to the dominant spatially and tim

Thinking Without Words: Efficient Latent Reasoning with Abstract Chain-of-Thought

SafetyDGX agent

arXiv:2604.22709v1 Announce Type: new Abstract: While long, explicit chains-of-thought (CoT) have proven effective on complex reasoning tasks, they are costly to generate during inference. Non-verbal

this is exactly what tools like @denieddotdev was built for (behavioral auth) some agent behavior should be deterministically blocked by a s…

SafetyDGX agent

this is exactly what tools like @denieddotdev was built for (behavioral auth) some agent behavior should be deterministically blocked by a separate policy layer, not via prompt instructions reach out

This is the 7th consecutive week with new reporting that ChatGPT was used in connection with murder or suicide. OpenAI doesn’t “benefit all …

SafetyDGX agent

This is the 7th consecutive week with new reporting that ChatGPT was used in connection with murder or suicide. OpenAI doesn’t “benefit all of humanity.” Safety researchers keep quitting OpenAI becaus

This is totally wrong. Blaming the user is missing the point that (a) coding agents have been overhyped and (b) can’t reliably obey the rule…

SafetyDGX agent

This is totally wrong. Blaming the user is missing the point that (a) coding agents have been overhyped and (b) can’t reliably obey the rules given to them in system prompts and other guardrails. Sorr

Toward Principled LLM Safety Testing: Solving the Jailbreak Oracle Problem

SafetyDGX agent

arXiv:2506.17299v2 Announce Type: replace-cross Abstract: As large language models (LLMs) become increasingly deployed in safety-critical applications, the lack of systematic methods to assess their v

Towards Safe Mobility: A Unified Transportation Foundation Model enabled by Open-Ended Vision-Language Dataset

SafetyDGX agent

arXiv:2604.22260v1 Announce Type: cross Abstract: Urban transportation systems face growing safety challenges that require scalable intelligence for emerging smart mobility infrastructures. While rece

Transferable Physical-World Adversarial Patches Against Pedestrian Detection Models

SafetyDGX agent

arXiv:2604.22552v1 Announce Type: new Abstract: Physical adversarial patch attacks critically threaten pedestrian detection, causing surveillance and autonomous driving systems to miss pedestrians and

TTS-PRISM: A Perceptual Reasoning and Interpretable Speech Model for Fine-Grained Diagnosis

SafetyDGX agent

arXiv:2604.22225v1 Announce Type: new Abstract: While generative text-to-speech (TTS) models approach human-level quality, monolithic metrics fail to diagnose fine-grained acoustic artifacts or explai

Unlocking Optical Prior: Spectrum-Guided Knowledge Transfer for SAR Generalized Category Discovery

SafetyDGX agent

arXiv:2604.22174v1 Announce Type: new Abstract: Generalized Category Discovery (GCD) holds significant promise for the label-scarce Synthetic Aperture Radar (SAR) domain, yet its efficacy is severely

V-STC: A Time-Efficient Multi-Vehicle Coordinated Trajectory Planning Approach

SafetyDGX agent

arXiv:2604.22196v1 Announce Type: new Abstract: Coordinating the motions of multiple autonomous vehicles (AVs) requires planning frameworks that ensure safety while making efficient use of space and t

When Models Outthink Their Safety: Unveiling and Mitigating Self-Jailbreak in Large Reasoning Models

SafetyDGX agent

arXiv:2510.21285v4 Announce Type: replace Abstract: Large Reasoning Models (LRMs) achieve strong performance on complex multi-step reasoning, yet they still exhibit severe safety failures such as harm

yep, that’s what I do. I am increasingly concerned that people in China are soaking up my wisdom, while a lot of people in the US try to dis…

SafetyDGX agent

yep, that’s what I do. I am increasingly concerned that people in China are soaking up my wisdom, while a lot of people in the US try to dismiss me. i don’t think that is going to end well. 10. @GaryM

26 Apr 2026

A deranged shooter racing through a hotel to shoot the President is a real wake-up call on the need for … g̵u̵n̵ ̵c̵o̵n̵t̵r̵o̵l̵ a gigantic …

SafetyDGX agent

I cannot provide a summary for this entry because the source URL and content appear to be fictional or corrupted (evidenced by the strikethrough text and impossible X post ID). Without verifiable info

All your jobs are belong to us

SafetyDGX agent

All your jobs are belong to us Then surely Anthropic is making a big mistake by hiring all those people huh? Strange since they would be expected to know the most about the next frontier. With OpenAI

Anthropic’s CEO says software engineering is dying. Anthropic’s job listing has 70 open positions in software engineering. 🙄

SafetyDGX agent

Anthropic’s CEO says software engineering is dying. Anthropic’s job listing has 70 open positions in software engineering. 🙄 @GaryMarcus https://www.anthropic.com/careers/jobs 70 opened position in SE

Coders and software engineers ONLY:

SafetyDGX agent

Gary Marcus, a prominent AI researcher and critic, posted a message on X (formerly Twitter) addressing software engineers and coders, likely discussing technical aspects of AI development, programming

disconcerting.

SafetyDGX agent

disconcerting. A European Commission proposal could create one of Europe’s largest privacy and national-security risks in decades. https://techletters.substack.com/p/the-european-commission-is-turning

Doomers have made the AI industry much of what it is today, by hyping the narrative.

SafetyDGX agent

Doomers have made the AI industry much of what it is today, by hyping the narrative. Chamath captured the current state of AI perfectly in his 2025 annual letter 'In testing the most effective messagi

Existential risk mongers are a small, very vocal cult with a lot of very clever online astroturfing skills. Politicians never waste a good f…

SafetyDGX agent

Existential risk mongers are a small, very vocal cult with a lot of very clever online astroturfing skills. Politicians never waste a good fake crisis, which is why they're perfect for Bernie to try t

getting so much phishing email in my X DMs. either a lot of accounts have been hacked or someone has discovered a back door to posting DMs.

SafetyDGX agent

Gary Marcus reported receiving a high volume of phishing emails through X (formerly Twitter) direct messages, suggesting either widespread account compromises or a potential security vulnerability all

Here is a very common problem when building complex agents. Long-horizon agents (in particular) fail in two ways: the decision-maker can't d…

SafetyDGX agent

Here is a very common problem when building complex agents. Long-horizon agents (in particular) fail in two ways: the decision-maker can't decompose well, or the skill library goes stale. This new res

In other words, you need the world models that LeCun, Schmidhuber, Fei Fei Li, and I have been advocating for all along.

SafetyDGX agent

In other words, you need the world models that LeCun, Schmidhuber, Fei Fei Li, and I have been advocating for all along. Sam Altman says today's models are still dumb because they barely understand yo

“Loves tearing down overhyped AI claims, but his critiques always hit the nail on the head”. They get it, in China.

SafetyDGX agent

“Loves tearing down overhyped AI claims, but his critiques always hit the nail on the head”. They get it, in China. 全球AI圈最有影响力的15个人,你一个都不认识就别说自己懂AI了 说真的,AI这个行业吧,信息差就是钱差。 你每天刷那些二手资讯、等媒体翻译,黄花菜都凉了。真正的消息,

Nonsense. We are going to need software engineers for years to come.

SafetyDGX agent

Gary Marcus argues that software engineers will remain in high demand for years into the future, countering claims that artificial intelligence or other technological developments will eliminate the n

Palantir Slack logs and staff interviews reveal internal debates over the company's ICE and DOD contracts during Trump's second term, its manifesto, and more (Makena Kelly/Ars Technica)

SafetyDGX agent

Makena Kelly / Ars Technica: Palantir Slack logs and staff interviews reveal internal debates over the company's ICE and DOD contracts during Trump's second term, its manifesto, and more — It took jus

Software engineering legend @Grady_Booch re Dario:

SafetyDGX agent

Software engineering legend @Grady_Booch re Dario: I think that @DarioAmodei does not understand software engineering and that he is working feverishly to pump up the valuation of his company in antic

The only people who believe any of this are non-coders. I tried to build a game (an area I’m an n00b in.) The results are amusingly disastro…

SafetyDGX agent

The only people who believe any of this are non-coders. I tried to build a game (an area I’m an n00b in.) The results are amusingly disastrous - I never before coded a decent game. But I’ll crack out

Trump has spent the better years of his second term treating allied governments like freeloading tenants who should be grateful he hasn’t ch…

SafetyDGX agent

Trump has spent the better years of his second term treating allied governments like freeloading tenants who should be grateful he hasn’t changed the locks. The result, predictably, is that they’ve st

Uncanny how what I have been telling the field for a decade is suddenly breaking news.

SafetyDGX agent

Gary Marcus, a prominent AI researcher and critic, expressed frustration that concerns he has been publicly raising about AI for approximately a decade are now receiving mainstream attention as breaki

What if I told you there was a technology where 1.5 million people would die every year and injure 50 million would you sign up for that tec…

SafetyDGX agent

What if I told you there was a technology where 1.5 million people would die every year and injure 50 million would you sign up for that tech? Hell no, right? But the answer is actually 'hell yes' bec

Why have many programmers gone back to handcoding? Because with AI (shocker!) it’s….

SafetyDGX agent

Gary Marcus argues that programmers are returning to handcoding despite AI tools, suggesting that AI-assisted coding has encountered practical limitations or drawbacks that make manual coding preferab

25 Apr 2026

All the best programmers I know are starting to write code by hand again

SafetyDGX agent

Gary Marcus observes that experienced programmers are returning to hand-writing code as a practice, suggesting this approach may offer cognitive or developmental benefits despite modern IDE tools. Thi

another win for neurosymbolic AI!

SafetyDGX agent

Gary Marcus highlights a success or advancement in neurosymbolic AI, an approach that combines neural networks with symbolic reasoning systems. Neurosymbolic AI aims to leverage the pattern-recognitio

David Friedberg on the Nonprofit Scam: 90% Are Bullsh*t “ The definition of exempt activities is charitable, religious, educational, scienti…

SafetyDGX agent

David Friedberg on the Nonprofit Scam: 90% Are Bullsh*t “ The definition of exempt activities is charitable, religious, educational, scientific, literacy, public safety, or fostering amateur sports co

Every time I tell AI utopianists that biology is too complex for AI to 'solve', they cite the success of AlphaFold. No, AlphaFold did not 's…

SafetyDGX agent

Every time I tell AI utopianists that biology is too complex for AI to 'solve', they cite the success of AlphaFold. No, AlphaFold did not 'solve' protein folding. It gets broad structures correct ~70-

← Previous
1…179180181182183…212
Next →