AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,598 results
Safety

The Defense Trilemma: Why Prompt Injection Defense Wrappers Fail?

DGX agent

arXiv:2604.06436v2 Announce Type: cross Abstract: We prove that no continuous, utility-preserving wrapper defense-a function D: Xo X that preprocesses inputs before the model sees them-can make al

safetyarxiv-cs-ai
10 Apr 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

The End of the Foundation Model Era: Open-Weight Models, Sovereign AI, and Inference as Infrastructure

DGX agent

arXiv:2604.06217v1 Announce Type: cross Abstract: The foundation model era -- roughly 2020 to 2025 -- is over. The forces that defined it have inverted. Open source models have reached frontier perfor

safetyarxiv-cs-ai
10 Apr 2026
Safety

The Master Key Hypothesis: Unlocking Cross-Model Capability Transfer via Linear Subspace Alignment

DGX agent

arXiv:2604.06377v1 Announce Type: cross Abstract: We investigate whether post-trained capabilities can be transferred across models without retraining, with a focus on transfer across different model

safetyarxiv-cs-ai
10 Apr 2026
Safety

The reaction people are having to AIs that can find bugs in code is fascinating. Finally, we have the capacity to fix the crisis in computer…

DGX agent

The reaction people are having to AIs that can find bugs in code is fascinating. Finally, we have the capacity to fix the crisis in computer security we’ve had for decades, and everyone is treating it

safetyclem-delangue--x
10 Apr 2026
Safety

The Sustainability Gap in Robotics: A Large-Scale Survey of Sustainability Awareness in 50,000 Research Articles

DGX agent

arXiv:2604.07921v1 Announce Type: new Abstract: We present a large-scale survey of sustainability communication and motivation in robotics research. Our analysis covers nearly 50,000 open-access paper

safetyarxiv-cs-ro
10 Apr 2026
Safety

The Theorems of Dr. David Blackwell and Their Contributions to Artificial Intelligence

DGX agent

arXiv:2604.06621v1 Announce Type: cross Abstract: Dr. David Blackwell was a mathematician and statistician of the first rank, whose contributions to statistical theory, game theory, and decision theor

safetyarxiv-cs-lg
10 Apr 2026
Safety

this from @Kasparov63 applies to the AI autocrats as well. “he would never sell my personal conversations to the government.” oh, yes, he wo…

DGX agent

this from @Kasparov63 applies to the AI autocrats as well. “he would never sell my personal conversations to the government.” oh, yes, he would. I will point out preemptively that one of the autocrat'

safetygary-marcus--x
10 Apr 2026
Safety

This is despicable. Satanic maybe. So AI can talk your kids into suicide and there's nothing you can do about it. Join MAMA and be a part of…

DGX agent

This is despicable. Satanic maybe. So AI can talk your kids into suicide and there's nothing you can do about it. Join MAMA and be a part of the future we all want and deserve. Mothers Against Media A

safetygary-marcus--x
10 Apr 2026
Safety

To a degree that may surprise some people, I agree with much of this* from @deanwball and would only add that you don’t have to believe that…

DGX agent

To a degree that may surprise some people, I agree with much of this* from @deanwball and would only add that you don’t have to believe that AGI is remotely close to want to find—ASAP—a regulatory reg

safetygary-marcus--x
10 Apr 2026
Safety

Today we visited Japan's hottest AI startup @SakanaAILabs🎏🇯🇵! We met their research scientists and discussed the implications and impact …

DGX agent

Today we visited Japan's hottest AI startup @SakanaAILabs🎏🇯🇵! We met their research scientists and discussed the implications and impact of some their works like 'The AI scientist' and 'Continous Thou

safetydavid-ha--x
10 Apr 2026
Safety

Tool-MCoT: Tool Augmented Multimodal Chain-of-Thought for Content Safety Moderation

DGX agent

arXiv:2604.06205v1 Announce Type: cross Abstract: The growth of online platforms and user content requires strong content moderation systems that can handle complex inputs from various media types. Wh

safetyarxiv-cs-ai
10 Apr 2026
Safety

Towards Privacy-Preserving Large Language Model: Text-free Inference Through Alignment and Adaptation

DGX agent

arXiv:2604.06831v1 Announce Type: cross Abstract: Current LLM-based services typically require users to submit raw text regardless of its sensitivity. While intuitive, such practice introduces substan

safetyarxiv-cs-ai
10 Apr 2026
Safety

Towards provable probabilistic safety for scalable embodied AI systems

DGX agent

arXiv:2506.05171v3 Announce Type: replace-cross Abstract: Embodied AI systems, comprising AI models and physical plants, are increasingly prevalent across various applications. Due to the rarity of sy

safetyarxiv-cs-ai
10 Apr 2026
Safety

Towards the Development of an LLM-Based Methodology for Automated Security Profiling in Compliance with Ukrainian Cybersecurity Regulations

DGX agent

arXiv:2604.06274v1 Announce Type: cross Abstract: In recent years, the pace of development of information technology in various areas has increased drastically, forcing cybersecurity specialists to co

safetyarxiv-cs-ai
10 Apr 2026
Safety

TwinLoop: Simulation-in-the-Loop Digital Twins for Online Multi-Agent Reinforcement Learning

DGX agent

arXiv:2604.06610v1 Announce Type: cross Abstract: Decentralised online learning enables runtime adaptation in cyber-physical multi-agent systems, but when operating conditions change, learned policies

safetyarxiv-cs-ai
10 Apr 2026
Safety

URMF: Uncertainty-aware Robust Multimodal Fusion for Multimodal Sarcasm Detection

DGX agent

arXiv:2604.06728v1 Announce Type: cross Abstract: Multimodal sarcasm detection (MSD) aims to identify sarcastic intent from semantic incongruity between text and image. Although recent methods have im

safetyarxiv-cs-ai
10 Apr 2026
Safety

ViVa: A Video-Generative Value Model for Robot Reinforcement Learning

DGX agent

arXiv:2604.08168v1 Announce Type: new Abstract: Vision-language-action (VLA) models have advanced robot manipulation through large-scale pretraining, but real-world deployment remains challenging due

safetyarxiv-cs-ro
10 Apr 2026
Safety

VLMShield: Efficient and Robust Defense of Vision-Language Models against Malicious Prompts

DGX agent

arXiv:2604.06502v1 Announce Type: new Abstract: Vision-Language Models (VLMs) face significant safety vulnerabilities from malicious prompt attacks due to weakened alignment during visual integration.

safetyarxiv-cs-lg
10 Apr 2026
Safety

We are not getting to the G in Artificial General Intelligence; we are getting to (impressive) advances in particular areas where particular…

DGX agent

We are not getting to the G in Artificial General Intelligence; we are getting to (impressive) advances in particular areas where particular (verifiable) techniques can be used, on problems with advan

safetygary-marcus--x
10 Apr 2026
Safety

WebExpert: domain-aware web agents with critic-guided expert experience for high-precision search

DGX agent

arXiv:2604.06177v1 Announce Type: cross Abstract: Specialized web tasks in finance, biomedicine, and pharmaceuticals remain challenging due to missing domain priors: queries drift, evidence is noisy,

safetyarxiv-cs-ai
10 Apr 2026
Safety

WebSP-Eval: Evaluating Web Agents on Website Security and Privacy Tasks

DGX agent

arXiv:2604.06367v1 Announce Type: cross Abstract: Web agents automate browser tasks, ranging from simple form completion to complex workflows like ordering groceries. While current benchmarks evaluate

safetyarxiv-cs-ai
10 Apr 2026
Safety

What Drives Representation Steering? A Mechanistic Case Study on Steering Refusal

DGX agent

arXiv:2604.08524v1 Announce Type: cross Abstract: Applying steering vectors to large language models (LLMs) is an efficient and effective model alignment technique, but we lack an interpretable explan

safetyarxiv-cs-cl
10 Apr 2026
Safety

What Makes an Ideal Quote? Recommending 'Unexpected yet Rational' Quotations via Novelty

DGX agent

arXiv:2602.22220v2 Announce Type: replace-cross Abstract: Quotation recommendation aims to enrich writing by suggesting quotes that complement a given context, yet existing systems mostly optimize sur

safetyarxiv-cs-ai
10 Apr 2026
Safety

What They Saw, Not Just Where They Looked: Semantic Scanpath Similarity via VLMs and NLP metric

DGX agent

arXiv:2604.08494v1 Announce Type: cross Abstract: Scanpath similarity metrics are central to eye-movement research, yet existing methods predominantly evaluate spatial and temporal alignment while neg

safetyarxiv-cs-cl
10 Apr 2026
Safety

When Numbers Speak: Aligning Textual Numerals and Visual Instances in Text-to-Video Diffusion Models

DGX agent

arXiv:2604.08546v1 Announce Type: new Abstract: Text-to-video diffusion models have enabled open-ended video synthesis, but often struggle with generating the correct number of objects specified in a

safetyarxiv-cs-cv
10 Apr 2026
Safety

A different sense of the word “bubble”

DGX agent

A different sense of the word “bubble” @GaryMarcus It's actually kinda hilarious that @OpenAI thinks that buying a niche tech podcast followed mostly by Bay Area insiders is going to solve their colos

safetygary-marcus--x
9 Apr 2026
Safety

All In’s @davidsacks liking one of my tweets was not in my 2026 bingo card.

DGX agent

The specific tweet from Gary Marcus (X post ID 2042361570253217968) is not publicly accessible without authentication, and the search results do not surface its specific content. Based on available...

safetygary-marcus--x
9 Apr 2026
Safety

Also not a sign of someone who has any clue about the realities of current neuroscience.

DGX agent

Also not a sign of someone who has any clue about the realities of current neuroscience. Sam Altman has admitted he is on a waitlist for a procedure that would digitize his brain. The procedure would

safetygary-marcus--x
9 Apr 2026
Safety

Folks, you can relax. Mythos is not some off-trend exponential gain. And the gains weren’t about recursive self-improvement. Good thread fro…

DGX agent

Folks, you can relax. Mythos is not some off-trend exponential gain. And the gains weren’t about recursive self-improvement. Good thread from @ramez. Anthropic's Mythos does not appear to show any acc

safetygary-marcus--x
9 Apr 2026
Safety

GenAI’s popularity has hit a wall. Hard to see how OpenAI is going to make its numbers, and easy to see why they bought a media company. The…

DGX agent

AI critic Gary Marcus argues that GenAI's popularity has plateaued, with a landmark MIT study finding that only 5% of enterprise AI pilots generate revenue, while most deliver little to no measura...

safetygary-marcus--x
9 Apr 2026
Safety

Google's AI Overviews spew out millions of false answers per hour, bombshell study reveals https://trib.al/1ao7qB1

DGX agent

A study commissioned by *The New York Times* and conducted by AI startup Oumi tested 4,326 Google searches using the SimpleQA benchmark, finding that Google's AI Overviews were accurate 85% of the...

safetygary-marcus--x
9 Apr 2026
Safety

Guardrails at the gateway: Securing AI inference on GKE with Model Armor

DGX agent

Enterprises are rapidly moving AI workloads from experimentation to production on Google Kubernetes Engine (GKE), using its scalability to serve powerful inference endpoints. However, as these models

safetygoogle-cloud-ai
9 Apr 2026
Safety

I will never get used to the sheer number of people who lie about me.

DGX agent

I was unable to retrieve results for that specific tweet URL. The tweet ID referenced (2042037797503299600) does not appear in any indexed web search results, and X (formerly Twitter) posts are gen...

safetygary-marcus--x
9 Apr 2026
Safety

My considered opinion is that The Mythos stuff was mostly a myth. Take it as a serious warning sign that we need to get our act together wit…

DGX agent

My considered opinion is that The Mythos stuff was mostly a myth. Take it as a serious warning sign that we need to get our act together with respect to cybersecurity. But don’t take the details serio

safetygary-marcus--x
9 Apr 2026
Safety

Perhaps because professional investors would look more carefully at the numbers? Wouldn’t want that!

DGX agent

Perhaps because professional investors would look more carefully at the numbers? Wouldn’t want that! OpenAI intends to set aside a share allocation for retail investors when the company goes public, t

safetygary-marcus--x
9 Apr 2026
Safety

Read The Story of Civilization by Durant

DGX agent

Read The Story of Civilization by Durant “But the curse of every ancient civilization was that its men in the end became unable to fight. Materialism, luxury, safety, even sometimes an almost modern s

safetyelon-musk--x
9 Apr 2026
Safety

Stargate never made sense.

DGX agent

AI critic and cognitive scientist Gary Marcus publicly challenged the economic rationale behind Project Stargate, the Trump-announced $500 billion AI infrastructure initiative, arguing that the mat...

safetygary-marcus--x
9 Apr 2026
Safety

Tesla V14.3 self-driving review. The point releases will bring polish. V15 will far exceed human levels of safety, even in completely unsupe…

DGX agent

Tesla V14.3 self-driving review. The point releases will bring polish. V15 will far exceed human levels of safety, even in completely unsupervised and complex situations. 600 miles in with FSD v14.3 a

safetyelon-musk--x
9 Apr 2026
Safety

The AI industry’s race for profits is now existential

DGX agent

Today on Decoder, let’s talk about the looming AI monetization cliff, and whether some of the biggest companies in the space can become real, profitable businesses before they careen right off it. My

safetythe-verge-ai
9 Apr 2026
Safety

There are plenty of people who are awed by AI who are not coders, I think the argument that AI impresses programmers most is, in part, selec…

DGX agent

There are plenty of people who are awed by AI who are not coders, I think the argument that AI impresses programmers most is, in part, selection bias on X, which is heavy on coders and people making f

safetyethan-mollick--x
9 Apr 2026
Safety

this map is the story of the last 25 years of us foreign policy in its own backyard. and that was before the tariffs.

DGX agent

I was unable to retrieve the specific content of the linked X (Twitter) post by Ian Bremmer, as the URL points to a social media post that is not directly accessible or indexed with its full conten...

safetyyann-lecun--x
9 Apr 2026
Safety

This SHOULD be an obvious point.

DGX agent

I was unable to retrieve the specific tweet at the URL provided (https://x.com/GaryMarcus/status/2042242813384142965). The tweet ID (2042242813384142965) appears to be from the future relative to c...

safetygary-marcus--x
9 Apr 2026
Safety

This – “The CEO of Google DeepMind (@demishassabis) just admitted that if the decision had been his, we would've cured cancer before anyone …

DGX agent

This – “The CEO of Google DeepMind (@demishassabis) just admitted that if the decision had been his, we would've cured cancer before anyone ever used ChatGPT.” is exactly what i am trying to say in my

safetygary-marcus--x
9 Apr 2026
Safety

Tragically I am continuing to find that the most effective guardrail against slop is extremely talented engineers doing very thoughtful, hum…

DGX agent

I was unable to retrieve the specific content from the X (Twitter) post at the URL provided, and my web search did not surface the original post or any reliable secondary sources quoting or summari...

safetyjeremy-howard--x
9 Apr 2026
Safety

Trump's emergency orders pushing coal power are 'illegal' as well as dumb

DGX agent

The Trump administration's Department of Energy has invoked Section 202(c) of the Federal Power Act — a provision granting broad emergency authority over the electricity system that had previously...

safetyars-technica
9 Apr 2026
Safety

1. If true (and it does fit with my perceptions FWIW), this is an amazing and incredibly damning graph 2. Can anyone find the source on whic…

DGX agent

1. If true (and it does fit with my perceptions FWIW), this is an amazing and incredibly damning graph 2. Can anyone find the source on which it is based? 'Anthropic, OpenAl and Google release their n

safetygary-marcus--x
8 Apr 2026
Safety

A few weeks ago I had a conversation with an American who genuinely believed Europe and Canada would help the United States in its war with …

DGX agent

A few weeks ago I had a conversation with an American who genuinely believed Europe and Canada would help the United States in its war with Iran. I asked him why he thought that, given that Trump had

safetyyann-lecun--x
8 Apr 2026
Safety

Again, if you care about computer security, read the red team report: https://red.anthropic.com/2026/mythos-preview/

DGX agent

Anthropic's Frontier Red Team report (April 2026) details the cybersecurity capabilities of Claude Mythos Preview, a general-purpose frontier model that performs strongly across the board but is s...

safetyethan-mollick--x
8 Apr 2026
← Previous
1…260261262263
Next →