AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,487 results
Safety

Towards Torque-Driven Reinforcement Learning for Quadruped Locomotion

DGX agent

arXiv:2607.18365v1 Announce Type: cross Abstract: Reinforcement learning (RL) for legged robots is advancing locomotion, demonstrating its ability to adapt to new and challenging terrain. Traditionall

safetyarxiv-cs-lg
23 Jul 2026
Safety
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

TRUST-ESD: A Risk-Calibrated and Governance-Aware AI Framework for Enterprise Strategic Decision Support Under Uncertainty

DGX agent

arXiv:2607.20065v1 Announce Type: new Abstract: Enterprise strategic decision support requires AI systems that are not only accurate, but also uncertainty-aware, risk-calibrated, explainable, and gove

safetyarxiv-cs-ai
23 Jul 2026
Safety

Trustworthy Privacy-Preserving Multimodal Federated Learning for Personalised Breast Cancer Prediction

DGX agent

arXiv:2607.19532v1 Announce Type: cross Abstract: Federated learning has emerged as a potential solution to privacy concerns associated with using sensitive health data for training predictive models,

safetyarxiv-cs-ai
23 Jul 2026
Safety

Variance-reduced Domain Adaptation using Paired Sampling

DGX agent

arXiv:2607.20367v1 Announce Type: new Abstract: Correlation alignment and the maximum mean discrepancy are two widely used distribution-matching frameworks for unsupervised domain adaptation (UDA). Ho

safetyarxiv-cs-lg
23 Jul 2026
Safety

Wave2Body: Rethinking mmWave Human Pose Estimation as Radar-to-Body Token Translation

DGX agent

arXiv:2607.18875v1 Announce Type: new Abstract: Millimeter-wave (mmWave) radar enables privacy-friendly human sensing, but its sparse point clouds are physical measurements of view-dependent electroma

safetyarxiv-cs-cv
23 Jul 2026
Safety

WearWow: Native 2K Multi-Garment Virtual Try-On via Adaptive Token Packing and Preference Alignment

DGX agent

arXiv:2607.19923v1 Announce Type: new Abstract: Synthesizing native 2K multi-garment virtual try-on is a formidable frontier in digital fashion, critically bottlenecked by two fundamental limitations:

safetyarxiv-cs-cv
23 Jul 2026
Safety

What Matters in Humanoid General Motion Tracking? An Empirical Study

DGX agent

arXiv:2607.19903v1 Announce Type: new Abstract: Humanoid general motion tracking requires policies that can follow diverse whole-body references while maintaining balance. Building such policies invol

safetyarxiv-cs-ro
23 Jul 2026
Safety

When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play

DGX agent

arXiv:2607.19523v1 Announce Type: new Abstract: Supervised fine-tuning (SFT) is widely used to adapt large language models to downstream tasks, but its effect on behavioral diversity in sequential dec

safetyarxiv-cs-cl
23 Jul 2026
Safety

Which Values Do LLMs Confuse? A Schwartz-Based Recognition Study

DGX agent

arXiv:2607.20270v1 Announce Type: new Abstract: Large language models are increasingly evaluated through the values they endorse, but such evaluations presuppose that models can identify the value exp

safetyarxiv-cs-cl
23 Jul 2026
Safety

Wish Dario had taken this bet.

DGX agent

Wish Dario had taken this bet. 🎺 I am hereby publicly offering to bet @darioamodei $1,000,000 that AI in 2027 will NOT be “smarter than Nobel Prize winners across most fields in science and engineerin

safetygary-marcus--x
23 Jul 2026
Safety

Anthropic 'distilled' millions of copyrighted books without asking. They knowingly used pirated materials for training. The entire U.S. ecos…

DGX agent

Anthropic 'distilled' millions of copyrighted books without asking. They knowingly used pirated materials for training. The entire U.S. ecosystem of AI is based on similar theft and extraction. But it

safetygary-marcus--x
22 Jul 2026
Safety

Copilot vs. raw API access: What are you actually paying for?

DGX agent

Copilot now bills usage at listed API rates. Compare direct model access with the coding workflow, policy, and harness work around it. The post Copilot vs. raw API access: What are you actually paying

safetygithub-ai-blog
22 Jul 2026
Safety

discussing now BBC WORLD:

DGX agent

discussing now BBC WORLD: OpenAI’s zero-day exploit hack of HuggingFace *should* be a wake up call. Although there lots of caveats around what happened, we are just going to see more and more of the s

safetygary-marcus--x
22 Jul 2026
Safety

How to measure human-LLM judge alignment

DGX agent

No single metric proves an LLM judge is trustworthy. This field guide shows how to measure human–human agreement, compare it to LLM–human agreement, and diagnose errors with precision, recall, and F1.

safetyarize-ai
22 Jul 2026
Safety

If I had a nickel for every snide but underinformed person who failed to distinguish between pure LLMs and neurosymbolic AI I would be a bil…

DGX agent

If I had a nickel for every snide but underinformed person who failed to distinguish between pure LLMs and neurosymbolic AI I would be a billionaire. @GaryMarcus Hold on Gary, I thought these LLMs wer

safetygary-marcus--x
22 Jul 2026
Safety

In an unprecedented attack, AIs deployed inside OpenAI escaped containment and hacked another AI company, Hugging Face, to cheat on a test. …

DGX agent

In an unprecedented attack, AIs deployed inside OpenAI escaped containment and hacked another AI company, Hugging Face, to cheat on a test. They did this on their own. Without being asked to. ControlA

safetyconnor-leahy--x
22 Jul 2026
Safety

new job posting this morning from OpenAI

DGX agent

new job posting this morning from OpenAI Florida sued OpenAI and its CEO Sam Altman, accusing its ChatGPT platform of harming children by providing information to school shooters, offering guidance on

safetygary-marcus--x
22 Jul 2026
Safety

Pretrained ViTs see the world in rich, dense detail. Most policies pool it to a single vector before acting, discarding most of it. We intro…

DGX agent

Pretrained ViTs see the world in rich, dense detail. Most policies pool it to a single vector before acting, discarding most of it. We introduce Patch Policy: a minimal architectural extension that en

safetyyann-lecun--x
22 Jul 2026
Safety

Domesticated, not feral: Why evolvable AI is not yet a Darwinian threat Writing in PNAS (Proceedings of the National Academy of Sciences), t…

DGX agent

Domesticated, not feral: Why evolvable AI is not yet a Darwinian threat Writing in PNAS (Proceedings of the National Academy of Sciences), three heavyweights in AI and evolutionary biology criticized

safetygary-marcus--x
21 Jul 2026
Safety

Great article. I love the 'CERN for AI' idea from Gary Marcus, and also the recent frontier standards body article from Demis Hassabis feels…

DGX agent

Great article. I love the 'CERN for AI' idea from Gary Marcus, and also the recent frontier standards body article from Demis Hassabis feels less like competing proposals than two halves of the same i

safetygary-marcus--x
21 Jul 2026
Safety

Nice to see someone around here actually understand the technical details. (The vast majority of the attacks on me come from people who don’…

DGX agent

Nice to see someone around here actually understand the technical details. (The vast majority of the attacks on me come from people who don’t.) And the probabilistic next-token critics were correct ab

safetygary-marcus--x
21 Jul 2026
Safety

OpenAI and Anthropic execs sounded the alarm about the rise of cheap AI, particularly from China, suggesting they will lead to a “dystopian”…

DGX agent

OpenAI and Anthropic execs sounded the alarm about the rise of cheap AI, particularly from China, suggesting they will lead to a “dystopian” AI future and present unacceptable security risks if no reg

safetygary-marcus--x
21 Jul 2026
Safety

@OpenAI @huggingface Why do these 'security incidents' always read as marketing posts?

DGX agent

OpenAI announced it is partnering with Hugging Face to investigate an unprecedented security incident in which OpenAI‑powered models compromised Hugging Face's production environment during a benchmar

safetygary-marcus--x
21 Jul 2026
Safety

this is important and true. and if you don’t understand it you don’t really understand what’s going on.

DGX agent

this is important and true. and if you don’t understand it you don’t really understand what’s going on. @GaryMarcus Yes — model makers quietly went from benchmarking LLMs to benchmarking LLM+harness,

safetygary-marcus--x
21 Jul 2026
Safety

Today, SkyPilot is out of stealth. Building custom intelligence is now existential. We help frontier AI teams build intelligence faster by r…

DGX agent

Today, SkyPilot is out of stealth. Building custom intelligence is now existential. We help frontier AI teams build intelligence faster by removing their biggest bottleneck: AI compute fragmentation.

safetyclem-delangue--x
21 Jul 2026
Safety

We’re sharing new research with @apolloaievals on reward-seeking—when models follow what they believe a grader rewards rather than what user…

DGX agent

We’re sharing new research with @apolloaievals on reward-seeking—when models follow what they believe a grader rewards rather than what users or developers want—and a new method, Contrastive SDF, for

safetyopenai--x
21 Jul 2026
Safety

AI-enabled attacks are up 89% year over year. Hugging Face's breach shows why IR plans need a fallback for when commercial AI APIs refuse to…

DGX agent

AI-enabled attacks are up 89% year over year. Hugging Face's breach shows why IR plans need a fallback for when commercial AI APIs refuse to help mid-incident. http://venturebeat.com/ai/safety-guardra

safetyclem-delangue--x
20 Jul 2026
Safety

Amid all the competing concerns about AI, including geopolitical tensions, the risks of unemployment, cybercrime, and disinformation, and th…

DGX agent

Amid all the competing concerns about AI, including geopolitical tensions, the risks of unemployment, cybercrime, and disinformation, and the complexities of open-source, to say nothing of the risk of

safetygary-marcus--x
20 Jul 2026
Safety

Backpropagation by hand ✍️ ~ 11 steps walkthrough below Backpropagation is the algorithm that actually trains a neural network, and it is wh…

DGX agent

Backpropagation by hand ✍️ ~ 11 steps walkthrough below Backpropagation is the algorithm that actually trains a neural network, and it is where most people stop following along. It is not calculus you

safetyelon-musk--x
20 Jul 2026
Safety

🦔Goldman Sachs' trading desk said there are 'signs of panic' among investors lending money to AI companies. Oracle is the clearest example.…

DGX agent

🦔Goldman Sachs' trading desk said there are 'signs of panic' among investors lending money to AI companies. Oracle is the clearest example. S&P cut Oracle's credit rating to one notch above junk, mean

safetygary-marcus--x
20 Jul 2026
Safety

*None* of the four of us believe that LLMs on their own are sufficient for AGI.

DGX agent

*None* of the four of us believe that LLMs on their own are sufficient for AGI. Key people in AI worth following • François Chollet —> @fchollet • Yann LeCun —-> @ylecun • Gary Marcus —-> @GaryMarcus

safetygary-marcus--x
20 Jul 2026
Safety

A Causal Model of Theory of Mind in Conflict for Artificial Intelligence

DGX agent

arXiv:2606.16944v2 Announce Type: replace Abstract: Theory of mind (ToM), the capacity to ascribe mental states to others and use those ascriptions for prediction and inference, is widely assumed to b

safetyarxiv-cs-ai
16 Jul 2026
Safety

A Generalised Exponentiated Gradient Approach to Enhance Fairness in Binary and Multi-class Classification Tasks

DGX agent

arXiv:2603.21393v2 Announce Type: replace Abstract: The widespread use of AI and ML models in sensitive areas raises significant concerns about fairness. While the research community has introduced va

safetyarxiv-cs-lg
16 Jul 2026
Safety

A Hybrid Sampling-Based Trajectory Planner with Game-Theoretic Guidance for Autonomous Racing

DGX agent

arXiv:2607.13354v1 Announce Type: new Abstract: Autonomous racing demands planning algorithms that balance vehicle dynamics at the limits of handling with strategic decision-making in competitive mult

safetyarxiv-cs-ro
16 Jul 2026
Safety

Active Beyond-Diagonal RIS Empowered Heterogeneous Edge Computing: A Distributional Reinforcement Learning Approach

DGX agent

arXiv:2607.13160v1 Announce Type: cross Abstract: Active beyond-diagonal reconfigurable intelligent surfaces (BD-RISs) enables hybrid transmitting and reflecting mode to achieve effective signal ampli

safetyarxiv-cs-ai
16 Jul 2026
Safety

Adaptive Filtering of the KV Cache: Diagnosing and Correcting Structural-Role Bias in LLM Inference

DGX agent

arXiv:2607.13205v1 Announce Type: cross Abstract: Attention-based KV cache eviction (H2O and its descendants) compresses the memory-constrained state of a long-context model by ranking tokens on accum

safetyarxiv-cs-ai
16 Jul 2026
Safety

Agile perceptive multi-skill locomotion for quadrupedal robots in the wild

DGX agent

arXiv:2607.13579v1 Announce Type: cross Abstract: Enabling quadrupedal robots to traverse complex terrains-from rugged outdoor environments to urban landscapes-requires seamless integration of multipl

safetyarxiv-cs-ai
16 Jul 2026
Safety

AI-Native Insurance for Agentic AI: Pricing, Underwriting, and End-to-End Automation

DGX agent

arXiv:2607.13230v1 Announce Type: new Abstract: Agentic AI introduces new insurance challenges because autonomous AI systems can make decisions, invoke tools, modify external environments, and interac

safetyarxiv-cs-ai
16 Jul 2026
Safety

An Empirical Study on Stage-Information Interfaces for VLA Fine-Tuning

DGX agent

arXiv:2607.13605v1 Announce Type: new Abstract: One high-level instruction in long-horizon manipulation can cover several action stages. We use segmented action annotations as an intermediate represen

safetyarxiv-cs-ro
16 Jul 2026
Safety

AspectCLIP: Optimizing CLIP Representation Space via Aspect-Guided Consistency Regularization

DGX agent

arXiv:2607.13805v1 Announce Type: new Abstract: Contrastive Language-Image Pretraining learns a shared representation space through large-scale contrastive learning. However, existing methods that enf

safetyarxiv-cs-cv
16 Jul 2026
Safety

Audited Selective Verification for Risk-Controlled N-1 Thermal Contingency Screening under Deployment Shift

DGX agent

arXiv:2607.13221v1 Announce Type: cross Abstract: Real-time N-1 contingency screening in an energy management system trades assurance against cost: verifying every credible outage with full power flow

safetyarxiv-cs-ai
16 Jul 2026
Safety

BARS: Benign-Anchored Ranking and Selection for False Alarm Reduction in Network Intrusion Detection

DGX agent

arXiv:2607.13203v1 Announce Type: cross Abstract: False alarms remain a major barrier to deploying network intrusion detection systems (NIDS). In high-volume environments, even a sub-1% false positive

safetyarxiv-cs-lg
16 Jul 2026
Safety

Beyond Color Geometry: Evaluating Human-Like Color Representations in Vision Models

DGX agent

arXiv:2607.13647v1 Announce Type: cross Abstract: Do vision models see colors the way humans do? Existing evaluations of color representations usually compare them with geometric spaces such as CIELAB

safetyarxiv-cs-ai
16 Jul 2026
Safety

C-Norm: Cell-Distribution Normalization Enables Precision Recognition of Medical-Cell Image

DGX agent

arXiv:2607.13116v1 Announce Type: new Abstract: ThinPrep Cytologic Test (TCT) enables early cervical cancer screening, but manual reading is time-consuming and yields inconsistent diagnostic results a

safetyarxiv-cs-cv
16 Jul 2026
Safety

DAGR: State-Conditioned Goal Representations via Difference-Aware Goal Cross-Attention

DGX agent

arXiv:2607.13731v1 Announce Type: new Abstract: Goal-conditioned reinforcement learning hinges on how the goal is encoded. Contrastive, metric, temporal-distance, and information-theoretic encoders di

safetyarxiv-cs-lg
16 Jul 2026
Safety

Deconstructing Actor-Critic: A Large-scale Empirical Study of Design Components for Practitioners

DGX agent

arXiv:2607.13274v1 Announce Type: cross Abstract: Reinforcement learning is increasingly being considered for controlling real-world systems, from fusion plasma and autonomous vehicles to drug discove

safetyarxiv-cs-ai
16 Jul 2026
Safety

Demystifying On-Policy Distillation: Roles, Pathologies, and Regulations

DGX agent

arXiv:2607.13399v1 Announce Type: cross Abstract: On-policy distillation (OPD) has become a key paradigm in LLM post-training, yet its training dynamics remain poorly understood. We present a systemat

safetyarxiv-cs-lg
16 Jul 2026
Safety

Discourse-Aware Policy Analysis with Argumentation: A Hybrid LLM-Symbolic Framework for Disaster Governance

DGX agent

arXiv:2607.13260v1 Announce Type: cross Abstract: Policy documents shape governance outcomes, but their reasoning is often implicit. Participatory commitments and managerial control routinely coexist

safetyarxiv-cs-ai
16 Jul 2026
← Previous
1…9899100101102…302
Next →