AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,704 results
Safety

Wish Dario had taken this bet.

DGX agent

Wish Dario had taken this bet. 🎺 I am hereby publicly offering to bet @darioamodei $1,000,000 that AI in 2027 will NOT be “smarter than Nobel Prize winners across most fields in science and engineerin

safetygary-marcus--x
23 Jul 2026
Safety

Anthropic 'distilled' millions of copyrighted books without asking. They knowingly used pirated materials for training. The entire U.S. ecos…

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

Anthropic 'distilled' millions of copyrighted books without asking. They knowingly used pirated materials for training. The entire U.S. ecosystem of AI is based on similar theft and extraction. But it

safetygary-marcus--x
22 Jul 2026
Safety

Copilot vs. raw API access: What are you actually paying for?

DGX agent

Copilot now bills usage at listed API rates. Compare direct model access with the coding workflow, policy, and harness work around it. The post Copilot vs. raw API access: What are you actually paying

safetygithub-ai-blog
22 Jul 2026
Safety

discussing now BBC WORLD:

DGX agent

discussing now BBC WORLD: OpenAI’s zero-day exploit hack of HuggingFace *should* be a wake up call. Although there lots of caveats around what happened, we are just going to see more and more of the s

safetygary-marcus--x
22 Jul 2026
Safety

From the original HuggingFace report, before they knew it was OpenAI. Wild stuff. Things are going to get weird “When we started the log ana…

DGX agent

From the original HuggingFace report, before they knew it was OpenAI. Wild stuff. Things are going to get weird “When we started the log analysis, we first used frontier models behind commercial APIs.

safetyclem-delangue--x
22 Jul 2026
Safety

How to measure human-LLM judge alignment

DGX agent

No single metric proves an LLM judge is trustworthy. This field guide shows how to measure human–human agreement, compare it to LLM–human agreement, and diagnose errors with precision, recall, and F1.

safetyarize-ai
22 Jul 2026
Safety

If I had a nickel for every snide but underinformed person who failed to distinguish between pure LLMs and neurosymbolic AI I would be a bil…

DGX agent

If I had a nickel for every snide but underinformed person who failed to distinguish between pure LLMs and neurosymbolic AI I would be a billionaire. @GaryMarcus Hold on Gary, I thought these LLMs wer

safetygary-marcus--x
22 Jul 2026
Safety

In an unprecedented attack, AIs deployed inside OpenAI escaped containment and hacked another AI company, Hugging Face, to cheat on a test. …

DGX agent

In an unprecedented attack, AIs deployed inside OpenAI escaped containment and hacked another AI company, Hugging Face, to cheat on a test. They did this on their own. Without being asked to. ControlA

safetyconnor-leahy--x
22 Jul 2026
Safety

It's essential that defenders have the same capabilities as attackers. This is a preview of the future where our ham-fisted safeguards and t…

DGX agent

It's essential that defenders have the same capabilities as attackers. This is a preview of the future where our ham-fisted safeguards and the doomsday and safety drumbeat make us decidely less safe.

safetyyann-lecun--x
22 Jul 2026
Safety

new job posting this morning from OpenAI

DGX agent

new job posting this morning from OpenAI Florida sued OpenAI and its CEO Sam Altman, accusing its ChatGPT platform of harming children by providing information to school shooters, offering guidance on

safetygary-marcus--x
22 Jul 2026
Safety

Pretrained ViTs see the world in rich, dense detail. Most policies pool it to a single vector before acting, discarding most of it. We intro…

DGX agent

Pretrained ViTs see the world in rich, dense detail. Most policies pool it to a single vector before acting, discarding most of it. We introduce Patch Policy: a minimal architectural extension that en

safetyyann-lecun--x
22 Jul 2026
Safety

And now from the Chinese government side. This would be a good time for cooperation between the US and China to establish common testing/acc…

DGX agent

And now from the Chinese government side. This would be a good time for cooperation between the US and China to establish common testing/acceptance standards for new models, so at least the safety cer

safetyethan-mollick--x
21 Jul 2026
Safety

Domesticated, not feral: Why evolvable AI is not yet a Darwinian threat Writing in PNAS (Proceedings of the National Academy of Sciences), t…

DGX agent

Domesticated, not feral: Why evolvable AI is not yet a Darwinian threat Writing in PNAS (Proceedings of the National Academy of Sciences), three heavyweights in AI and evolutionary biology criticized

safetygary-marcus--x
21 Jul 2026
Safety

Great article. I love the 'CERN for AI' idea from Gary Marcus, and also the recent frontier standards body article from Demis Hassabis feels…

DGX agent

Great article. I love the 'CERN for AI' idea from Gary Marcus, and also the recent frontier standards body article from Demis Hassabis feels less like competing proposals than two halves of the same i

safetygary-marcus--x
21 Jul 2026
Safety

Nice to see someone around here actually understand the technical details. (The vast majority of the attacks on me come from people who don’…

DGX agent

Nice to see someone around here actually understand the technical details. (The vast majority of the attacks on me come from people who don’t.) And the probabilistic next-token critics were correct ab

safetygary-marcus--x
21 Jul 2026
Safety

ok the Hugging Face breach writeup is one of the more honest post-mortems I've read in a while, and there's a detail buried in it that's way…

DGX agent

ok the Hugging Face breach writeup is one of the more honest post-mortems I've read in a while, and there's a detail buried in it that's way more interesting than 'AI agent hacked us.' Their IR team t

safetyclem-delangue--x
21 Jul 2026
Safety

OpenAI and Anthropic execs sounded the alarm about the rise of cheap AI, particularly from China, suggesting they will lead to a “dystopian”…

DGX agent

OpenAI and Anthropic execs sounded the alarm about the rise of cheap AI, particularly from China, suggesting they will lead to a “dystopian” AI future and present unacceptable security risks if no reg

safetygary-marcus--x
21 Jul 2026
Safety

@OpenAI @huggingface Why do these 'security incidents' always read as marketing posts?

DGX agent

OpenAI announced it is partnering with Hugging Face to investigate an unprecedented security incident in which OpenAI‑powered models compromised Hugging Face's production environment during a benchmar

safetygary-marcus--x
21 Jul 2026
Safety

this is important and true. and if you don’t understand it you don’t really understand what’s going on.

DGX agent

this is important and true. and if you don’t understand it you don’t really understand what’s going on. @GaryMarcus Yes — model makers quietly went from benchmarking LLMs to benchmarking LLM+harness,

safetygary-marcus--x
21 Jul 2026
Safety

Today, SkyPilot is out of stealth. Building custom intelligence is now existential. We help frontier AI teams build intelligence faster by r…

DGX agent

Today, SkyPilot is out of stealth. Building custom intelligence is now existential. We help frontier AI teams build intelligence faster by removing their biggest bottleneck: AI compute fragmentation.

safetyclem-delangue--x
21 Jul 2026
Safety

We’re sharing new research with @apolloaievals on reward-seeking—when models follow what they believe a grader rewards rather than what user…

DGX agent

We’re sharing new research with @apolloaievals on reward-seeking—when models follow what they believe a grader rewards rather than what users or developers want—and a new method, Contrastive SDF, for

safetyopenai--x
21 Jul 2026
Safety

AI-enabled attacks are up 89% year over year. Hugging Face's breach shows why IR plans need a fallback for when commercial AI APIs refuse to…

DGX agent

AI-enabled attacks are up 89% year over year. Hugging Face's breach shows why IR plans need a fallback for when commercial AI APIs refuse to help mid-incident. http://venturebeat.com/ai/safety-guardra

safetyclem-delangue--x
20 Jul 2026
Safety

Amid all the competing concerns about AI, including geopolitical tensions, the risks of unemployment, cybercrime, and disinformation, and th…

DGX agent

Amid all the competing concerns about AI, including geopolitical tensions, the risks of unemployment, cybercrime, and disinformation, and the complexities of open-source, to say nothing of the risk of

safetygary-marcus--x
20 Jul 2026
Safety

Backpropagation by hand ✍️ ~ 11 steps walkthrough below Backpropagation is the algorithm that actually trains a neural network, and it is wh…

DGX agent

Backpropagation by hand ✍️ ~ 11 steps walkthrough below Backpropagation is the algorithm that actually trains a neural network, and it is where most people stop following along. It is not calculus you

safetyelon-musk--x
20 Jul 2026
Safety

🦔Goldman Sachs' trading desk said there are 'signs of panic' among investors lending money to AI companies. Oracle is the clearest example.…

DGX agent

🦔Goldman Sachs' trading desk said there are 'signs of panic' among investors lending money to AI companies. Oracle is the clearest example. S&P cut Oracle's credit rating to one notch above junk, mean

safetygary-marcus--x
20 Jul 2026
Safety

Long-running models can solve hard open-ended problems, but their persistence can create safety risks that shorter-horizon evaluations miss.…

DGX agent

Long-running models can solve hard open-ended problems, but their persistence can create safety risks that shorter-horizon evaluations miss. We’re sharing what we learned from studying a long-running

safetyopenai--x
20 Jul 2026
Safety

*None* of the four of us believe that LLMs on their own are sufficient for AGI.

DGX agent

*None* of the four of us believe that LLMs on their own are sufficient for AGI. Key people in AI worth following • François Chollet —> @fchollet • Yann LeCun —-> @ylecun • Gary Marcus —-> @GaryMarcus

safetygary-marcus--x
20 Jul 2026
Safety

SSI should open-source an opus-tier model. 1. it nearly closes the us/china gap on the os frontier 2. they don't want to waste compute alloc…

DGX agent

SSI should open-source an opus-tier model. 1. it nearly closes the us/china gap on the os frontier 2. they don't want to waste compute allocation on inference, releasing os will enable continued full

safetyclem-delangue--x
20 Jul 2026
Safety

A Causal Model of Theory of Mind in Conflict for Artificial Intelligence

DGX agent

arXiv:2606.16944v2 Announce Type: replace Abstract: Theory of mind (ToM), the capacity to ascribe mental states to others and use those ascriptions for prediction and inference, is widely assumed to b

safetyarxiv-cs-ai
16 Jul 2026
Safety

A Deployed Hybrid Vehicle-in-the-Loop Platform for Validating Cooperative Perception

DGX agent

arXiv:2607.13806v1 Announce Type: new Abstract: European safety regulation now permits a large share of automated-driving homologation evidence to be produced virtually, provided a validated physical-

safetyarxiv-cs-ro
16 Jul 2026
Safety

A Generalised Exponentiated Gradient Approach to Enhance Fairness in Binary and Multi-class Classification Tasks

DGX agent

arXiv:2603.21393v2 Announce Type: replace Abstract: The widespread use of AI and ML models in sensitive areas raises significant concerns about fairness. While the research community has introduced va

safetyarxiv-cs-lg
16 Jul 2026
Safety

A Hybrid Sampling-Based Trajectory Planner with Game-Theoretic Guidance for Autonomous Racing

DGX agent

arXiv:2607.13354v1 Announce Type: new Abstract: Autonomous racing demands planning algorithms that balance vehicle dynamics at the limits of handling with strategic decision-making in competitive mult

safetyarxiv-cs-ro
16 Jul 2026
Safety

A VAE-Driven Multi-Task Satellite-Aided Semantic Communication Framework for 6G-Enabled Connected Autonomous Vehicles

DGX agent

arXiv:2607.13494v1 Announce Type: new Abstract: The development of smart transportation systems and the introduction of 6G wireless communication technologies have significantly changed vehicle networ

safetyarxiv-cs-lg
16 Jul 2026
Safety

Active Beyond-Diagonal RIS Empowered Heterogeneous Edge Computing: A Distributional Reinforcement Learning Approach

DGX agent

arXiv:2607.13160v1 Announce Type: cross Abstract: Active beyond-diagonal reconfigurable intelligent surfaces (BD-RISs) enables hybrid transmitting and reflecting mode to achieve effective signal ampli

safetyarxiv-cs-ai
16 Jul 2026
Safety

Adaptive Filtering of the KV Cache: Diagnosing and Correcting Structural-Role Bias in LLM Inference

DGX agent

arXiv:2607.13205v1 Announce Type: cross Abstract: Attention-based KV cache eviction (H2O and its descendants) compresses the memory-constrained state of a long-context model by ranking tokens on accum

safetyarxiv-cs-ai
16 Jul 2026
Safety

Adversarial Prompting Framework for AI Safety Assessment

DGX agent

arXiv:2607.13453v1 Announce Type: cross Abstract: Artificial Intelligence (AI), especially Generative AI (GenAI), adoption has increased in industries significantly in recent years. However, the use o

safetyarxiv-cs-ai
16 Jul 2026
Safety

Agile perceptive multi-skill locomotion for quadrupedal robots in the wild

DGX agent

arXiv:2607.13579v1 Announce Type: cross Abstract: Enabling quadrupedal robots to traverse complex terrains-from rugged outdoor environments to urban landscapes-requires seamless integration of multipl

safetyarxiv-cs-ai
16 Jul 2026
Safety

AI-Native Insurance for Agentic AI: Pricing, Underwriting, and End-to-End Automation

DGX agent

arXiv:2607.13230v1 Announce Type: new Abstract: Agentic AI introduces new insurance challenges because autonomous AI systems can make decisions, invoke tools, modify external environments, and interac

safetyarxiv-cs-ai
16 Jul 2026
Safety

An Empirical Study on Stage-Information Interfaces for VLA Fine-Tuning

DGX agent

arXiv:2607.13605v1 Announce Type: new Abstract: One high-level instruction in long-horizon manipulation can cover several action stages. We use segmented action annotations as an intermediate represen

safetyarxiv-cs-ro
16 Jul 2026
Safety

AspectCLIP: Optimizing CLIP Representation Space via Aspect-Guided Consistency Regularization

DGX agent

arXiv:2607.13805v1 Announce Type: new Abstract: Contrastive Language-Image Pretraining learns a shared representation space through large-scale contrastive learning. However, existing methods that enf

safetyarxiv-cs-cv
16 Jul 2026
Safety

Audited Selective Verification for Risk-Controlled N-1 Thermal Contingency Screening under Deployment Shift

DGX agent

arXiv:2607.13221v1 Announce Type: cross Abstract: Real-time N-1 contingency screening in an energy management system trades assurance against cost: verifying every credible outage with full power flow

safetyarxiv-cs-ai
16 Jul 2026
Safety

BARS: Benign-Anchored Ranking and Selection for False Alarm Reduction in Network Intrusion Detection

DGX agent

arXiv:2607.13203v1 Announce Type: cross Abstract: False alarms remain a major barrier to deploying network intrusion detection systems (NIDS). In high-volume environments, even a sub-1% false positive

safetyarxiv-cs-lg
16 Jul 2026
Safety

Beyond Color Geometry: Evaluating Human-Like Color Representations in Vision Models

DGX agent

arXiv:2607.13647v1 Announce Type: cross Abstract: Do vision models see colors the way humans do? Existing evaluations of color representations usually compare them with geometric spaces such as CIELAB

safetyarxiv-cs-ai
16 Jul 2026
Safety

C-Norm: Cell-Distribution Normalization Enables Precision Recognition of Medical-Cell Image

DGX agent

arXiv:2607.13116v1 Announce Type: new Abstract: ThinPrep Cytologic Test (TCT) enables early cervical cancer screening, but manual reading is time-consuming and yields inconsistent diagnostic results a

safetyarxiv-cs-cv
16 Jul 2026
Safety

Cost-Optimal Foundation Model Deployment Portfolio for Transportation Management

DGX agent

arXiv:2607.13239v1 Announce Type: new Abstract: Foundation models, including large language models (LLMs) and vision-language models (VLMs), are increasingly used for transportation management center

safetyarxiv-cs-ai
16 Jul 2026
Safety

DAGR: State-Conditioned Goal Representations via Difference-Aware Goal Cross-Attention

DGX agent

arXiv:2607.13731v1 Announce Type: new Abstract: Goal-conditioned reinforcement learning hinges on how the goal is encoded. Contrastive, metric, temporal-distance, and information-theoretic encoders di

safetyarxiv-cs-lg
16 Jul 2026
Safety

Deconstructing Actor-Critic: A Large-scale Empirical Study of Design Components for Practitioners

DGX agent

arXiv:2607.13274v1 Announce Type: cross Abstract: Reinforcement learning is increasingly being considered for controlling real-world systems, from fusion plasma and autonomous vehicles to drug discove

safetyarxiv-cs-ai
16 Jul 2026
Safety

Demystifying On-Policy Distillation: Roles, Pathologies, and Regulations

DGX agent

arXiv:2607.13399v1 Announce Type: cross Abstract: On-policy distillation (OPD) has become a key paradigm in LLM post-training, yet its training dynamics remain poorly understood. We present a systemat

safetyarxiv-cs-lg
16 Jul 2026
← Previous
1…4142434445…265
Next →