AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,813 results
Safety

100%, hallucinations are baked in (as I have been saying since 2001) and that is why basically no LLM company afford to operate in Germany n…

DGX agent

100%, hallucinations are baked in (as I have been saying since 2001) and that is why basically no LLM company afford to operate in Germany now. We need a better technology. @GaryMarcus Lol, and Google

safetygary-marcus--x
10 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

3SPO: State-Score-Supervised Policy Optimization for LLM Agents

DGX agent

arXiv:2606.09961v1 Announce Type: cross Abstract: Training large language models (LLMs) as autonomous agents via reinforcement learning (RL) has enabled frontier models to achieve superhuman performan

safetyarxiv-cs-ai
10 Jun 2026
Safety

A Comprehensive Survey of Direct Preference Optimization: Datasets, Theories, Variants, and Applications

DGX agent

arXiv:2410.15595v4 Announce Type: replace Abstract: With the rapid advancement of large language models (LLMs), aligning policy models with human preferences has become increasingly critical. Direct P

safetyarxiv-cs-ai
10 Jun 2026
Safety

A fine-grained attention and geometric correspondence model for musculoskeletal risk classification in athletes using multimodal visual and skeletal features

DGX agent

arXiv:2509.05913v3 Announce Type: replace Abstract: Musculoskeletal disorders pose significant risks to athletes, and early risk assessment is essential for prevention. However, most existing methods

safetyarxiv-cs-cv
10 Jun 2026
Safety

A Practical Recipe Towards Improving Sim-and-Real Correlation for VLA Evaluation

DGX agent

arXiv:2606.10366v1 Announce Type: cross Abstract: Simulation has become an essential tool for evaluating and improving vision-language-action (VLA) policies, offering scalable, reproducible, and contr

safetyarxiv-cs-ai
10 Jun 2026
Safety

A Reliable Fault Diagnosis Method Based on Belief Rule Base Consider Robustness Analysis

DGX agent

arXiv:2606.10500v1 Announce Type: new Abstract: In equipment operation, the implementation of fault diagnosis is essential to ensure the continuity and safety of production equipment, improve operatio

safetyarxiv-cs-ai
10 Jun 2026
Safety

A Source Domain is All You Need: Source-Only Cross-OS Transfer Learning for APT Anomaly Detection via Semantic Alignment and Optimal Transport

DGX agent

arXiv:2606.10216v1 Announce Type: cross Abstract: Advanced Persistent Threats (APTs) are stealthy, multi-stage cyberattacks whose detection is difficult due to scarce labeled traces, severe class imba

safetyarxiv-cs-ai
10 Jun 2026
Safety

A Unified Multi-Modal Framework for Intelligent Financial Systems: Integrating Reinforcement Learning, High-Frequency Trading, and Game-Theoretic Approaches with Cross-Modal Sentiment Analysis

DGX agent

arXiv:2606.10412v1 Announce Type: new Abstract: The rapid evolution of financial technology demands sophisticated artificial intelligence systems capable of handling diverse challenges across multiple

safetyarxiv-cs-ai
10 Jun 2026
Safety

Act on What You See: Unlocking Safe Social Navigation in Vision-Language-Action Models

DGX agent

arXiv:2606.10495v1 Announce Type: new Abstract: Safe social navigation requires robots to distinguish people from ordinary obstacles and to react before danger becomes imminent. We show that pretraine

safetyarxiv-cs-ro
10 Jun 2026
Safety

Adoption of Generative Artificial Intelligence in the German Software Engineering Industry: An Empirical Study

DGX agent

arXiv:2601.16700v2 Announce Type: replace-cross Abstract: Generative artificial intelligence (GenAI) tools have seen rapid adoption among software developers. While adoption rates in the industry are

safetyarxiv-cs-ai
10 Jun 2026
Safety

AI oligarchs are trying to buy elections (again)

DGX agent

AI oligarchs are trying to buy elections (again) AI titans are spending 1.3M to shape Utah primaries. Defending Our Values, run by Chris Stewart, spent 880K to support Rep. Celeste Maloy. DOV operates

safetygary-marcus--x
10 Jun 2026
Safety

Alignment Defends LLMs from Property Inference Attacks

DGX agent

arXiv:2606.10217v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly fine-tuned on domain-specific datasets that may contain sensitive, dataset-level properties. Recent work h

safetyarxiv-cs-lg
10 Jun 2026
Safety

Alongside it, Anthropic is releasing a proposal for how governments can address the risks posed by frontier AI and a policy framework for jo…

DGX agent

Alongside it, Anthropic is releasing a proposal for how governments can address the risks posed by frontier AI and a policy framework for job displacement, for which we intend to provide substantial f

safetydario-amodei--x
10 Jun 2026
Safety

An essay on policy responses to AI's exponential progress across regulation and public safety, macroeconomics and taxes, science, civil liberties, geopolitics (Dario Amodei)

DGX agent

Dario Amodei: An essay on policy responses to AI's exponential progress across regulation and public safety, macroeconomics and taxes, science, civil liberties, geopolitics — In one of the side plots

safetytechmeme
10 Jun 2026
Safety

An LLM-Native Psychometric Instrument Does Not Predict LLM Behavior: Evidence Across 25 Models

DGX agent

arXiv:2606.09843v1 Announce Type: cross Abstract: Large language models (LLMs) produce stable self-reports on personality inventories, but these self-reports do not predict observed behavior. Whether

safetyarxiv-cs-ai
10 Jun 2026
Safety

and see for deeper analysis https://open.substack.com/pub/garymarcus/p/breaking-news-and-how-the-end-might?r=8tdk6&utm_medium=ios

DGX agent

Gary Marcus discusses breaking news and potential existential risks or significant developments related to AI systems, likely examining how current AI capabilities and limitations might lead to critic

safetygary-marcus--x
10 Jun 2026
Safety

AnimaSpark: A Feed-Forward Method for Animating Arbitrary 3D Objects

DGX agent

arXiv:2606.10988v1 Announce Type: new Abstract: While recent advancements in generative AI have substantially accelerated static 3D model creation workflows, the synthesis of category-agnostic 3D anim

safetyarxiv-cs-cv
10 Jun 2026
Safety

Anthropic and OpenAI did not call for a pause. Read the wording: 'good for the world to have the option', 'possible' to slow down 'when need…

DGX agent

Anthropic and OpenAI did not call for a pause. Read the wording: 'good for the world to have the option', 'possible' to slow down 'when needed'. This is how they signal safety to one audience, acceler

safetyconnor-leahy--x
10 Jun 2026
Safety

Anthropic has long advocated for transparency requirements for frontier AI, because the risks weren't yet clear enough to regulate precisely…

DGX agent

Anthropic's leadership, through CEO Dario Amodei, has advocated for transparency requirements in frontier AI development as a regulatory approach, arguing that the risks were insufficiently understood

safetydario-amodei--x
10 Jun 2026
Safety

Anthropic, if they really believe what they say, should show some leadership:

DGX agent

Anthropic, if they really believe what they say, should show some leadership: 🔔 IF Anthropic is for real about safety, the should pause, for one month, and show leadership. If they won’t pause, even b

safetygary-marcus--x
10 Jun 2026
Safety

Anthropic releases two policy proposals on how governments should address catastrophic risks and manage labor market disruption from advanced AI systems (Anthropic)

DGX agent

Anthropic: Anthropic releases two policy proposals on how governments should address catastrophic risks and manage labor market disruption from advanced AI systems — AI is advancing at exponential spe

safetytechmeme
10 Jun 2026
Safety

Architect-Ant: Editable Automatic Furnishing of Architectural Floor Plans

DGX agent

arXiv:2606.10953v1 Announce Type: new Abstract: Furnished floor plans are fundamental to real estate visualization, interior design, and architectural workflows. However, progress in automatic furnitu

safetyarxiv-cs-ai
10 Jun 2026
Safety

are you kidding me

DGX agent

are you kidding me Today I'm publishing a new essay, Policy on the AI Exponential. AI is progressing extremely fast—much faster than the policy process was built to handle. The essay lays out where I

safetyjeremy-howard--x
10 Jun 2026
Safety

ARM: An AutoRegressive Large Multimodal Model with Unified Discrete Representations

DGX agent

arXiv:2606.11188v1 Announce Type: new Abstract: This paper introduces ARM, a discrete representation-based AutoRegressive Model that unifies image understanding, generation, and editing within a next-

safetyarxiv-cs-cv
10 Jun 2026
Safety

As vertically integrated platforms start to dominate they lock out third party access to the most valuable portions of the platform. Of cour…

DGX agent

As vertically integrated platforms start to dominate they lock out third party access to the most valuable portions of the platform. Of course, Anthropic is has the right to implement whatever policy

safetytogether-ai--x
10 Jun 2026
Safety

AsyncWebRL: Efficient Multi-Step RL for Visual Web Agents

DGX agent

arXiv:2606.05597v2 Announce Type: replace Abstract: Training vision-language web agents with multi-step RL is compute-intensive, with two dominant forms of inefficiency: idle GPUs in synchronous RL, a

safetyarxiv-cs-lg
10 Jun 2026
Safety

Automated Alignment between Elicitation Interviews and Requirements

DGX agent

arXiv:2510.08622v2 Announce Type: replace Abstract: Software requirements are derived from a variety of elicitation techniques, many of which have a conversational nature, like interviews. However, ev

safetyarxiv-cs-cl
10 Jun 2026
Safety

Automated Scoring of Arabic Text Using Large Language Models: A Literature Review

DGX agent

arXiv:2606.09830v1 Announce Type: new Abstract: In modern educational systems, Automatic Text Scoring (ATS) plays a central role by enabling scalable and consistent evaluation of learner responses wit

safetyarxiv-cs-cl
10 Jun 2026
Safety

Automatic Labelling for Low-Light Pedestrian Detection

DGX agent

arXiv:2507.02513v4 Announce Type: replace Abstract: Pedestrian detection in RGB images is a key task in pedestrian safety, as the most common sensor in autonomous vehicles and advanced driver assistan

safetyarxiv-cs-cv
10 Jun 2026
Safety

Banger post from @testdrivenzen There is no good outcome for the world unless countries coordinate to stop ASI development.

DGX agent

Banger post from @testdrivenzen There is no good outcome for the world unless countries coordinate to stop ASI development. Any plan for surviving superintelligent AI that doesn't go through strong in

safetyconnor-leahy--x
10 Jun 2026
Safety

Baseline-Free Policy Optimization for Neural Combinatorial Optimization

DGX agent

arXiv:2606.10321v1 Announce Type: cross Abstract: Neural combinatorial optimization (NCO) trains autoregressive policies to solve routing problems. The standard training algorithm, REINFORCE with a ro

safetyarxiv-cs-ai
10 Jun 2026
Safety

Bellman-Taylor Score Decoding for Markov Decision Processes with State-Dependent Feasible Action Sets

DGX agent

arXiv:2606.10979v1 Announce Type: new Abstract: Many Markov decision processes (MDPs) in operations research have feasible actions that are state dependent and defined implicitly by various operationa

safetyarxiv-cs-ai
10 Jun 2026
Safety

Beyond Uniform Token-Level Trust Region in LLM Reinforcement Learning

DGX agent

arXiv:2606.10968v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has become standard for improving LLM reasoning. However, existing PPO-style trust-region mechan

safetyarxiv-cs-ai
10 Jun 2026
Safety

🚨Breaking news that could be huge, and enormously bad for GenAI, if other countries make similar decisions. https://the-decoder.com/landmar…

DGX agent

🚨Breaking news that could be huge, and enormously bad for GenAI, if other countries make similar decisions. https://the-decoder.com/landmark-german-ruling-declares-googles-ai-overviews-are-googles-own

safetygary-marcus--x
10 Jun 2026
Safety

BREAKING! This may well be the beginning of the end.

DGX agent

BREAKING! This may well be the beginning of the end. 🚨BREAKING: SoftBank tried to borrow 6 billion against its 13% OpenAI stock… to keep funding OpenAI Banks said NO They don’t believe OpenAI is worth

safetygary-marcus--x
10 Jun 2026
Safety

Bypassing Copyright Protection in Diffusion-based Customization via Two-Stage Latent Feature Optimization

DGX agent

arXiv:2606.09909v1 Announce Type: cross Abstract: With the growing concerns over copyright infringement in diffusion-based customization, adversarial attacks have emerged as a prominent defense strate

safetyarxiv-cs-ai
10 Jun 2026
Safety

CameraMatics, which uses AI to help fleet operators improve safety, reduce operational risk, and lower carbon emissions, raised €49M (Joe Brennan/The Irish Times)

DGX agent

Joe Brennan / The Irish Times: CameraMatics, which uses AI to help fleet operators improve safety, reduce operational risk, and lower carbon emissions, raised €49M — Tech focuses on accident preventio

safetytechmeme
10 Jun 2026
Safety

Canada introduces the Safe Social Media Act, a bill that would ban social media for children under 16 and establish safety standards for AI chatbots (Maria Cheng/Reuters)

DGX agent

Maria Cheng / Reuters: Canada introduces the Safe Social Media Act, a bill that would ban social media for children under 16 and establish safety standards for AI chatbots — The Canadian government in

safetytechmeme
10 Jun 2026
Safety

Causal Ensemble Agent: Hierarchical Causal Discovery with LLM-guided Expert Reweighting

DGX agent

arXiv:2606.10607v1 Announce Type: cross Abstract: Causal discovery aims to uncover causal structures from observational data, which is crucial for real-world decision-making. However, different causal

safetyarxiv-cs-ai
10 Jun 2026
Safety

Closing the Modality Gap in Zero-Shot HAR: Contrastive Training and Separability-Optimized Prototypes on IMU Data

DGX agent

arXiv:2606.10789v1 Announce Type: new Abstract: Zero-shot learning (ZSL) for inertial measurement unit (IMU)-based human activity recognition (HAR) faces a central challenge: bridging the gap between

safetyarxiv-cs-lg
10 Jun 2026
Safety

Conditional Vendi Score: Prompt-Aware Diversity Evaluation for Generative AI Models and LLMs

DGX agent

arXiv:2411.02817v2 Announce Type: replace-cross Abstract: Generative models guided by text prompts are widely evaluated for fidelity and prompt alignment, yet their ability to produce outputs remains

safetyarxiv-cs-ai
10 Jun 2026
Safety

Conformal Prediction for Neural Operators: Distribution-Free Uncertainty Quantification in Physics Simulation

DGX agent

arXiv:2606.09923v1 Announce Type: cross Abstract: Neural operators such as the Fourier Neural Operator (FNO) have emerged as powerful surrogates for solving partial differential equations (PDEs), achi

safetyarxiv-cs-ai
10 Jun 2026
Safety

Contrastive Spectral Rectification: Test-Time Defense towards Zero-shot Adversarial Robustness of CLIP

DGX agent

arXiv:2601.19210v2 Announce Type: replace Abstract: Vision-language models (VLMs) such as CLIP have demonstrated remarkable zero-shot generalization, yet remain highly vulnerable to adversarial exampl

safetyarxiv-cs-cv
10 Jun 2026
Safety

Convergence of Monte Carlo Optimistic Policy Iteration: Beyond Uniform State-Action Updates

DGX agent

arXiv:2606.10580v1 Announce Type: cross Abstract: The asymptotic behaviour of Monte Carlo optimistic policy iteration (MC-O-PI) is a long-standing open question. When the model of the environment is u

safetyarxiv-cs-ai
10 Jun 2026
Safety

Cross-Modal Knowledge Distillation without Paired Data: Theoretical Foundation and Algorithm

DGX agent

arXiv:2606.10504v1 Announce Type: new Abstract: Cross-modal knowledge distillation (CMKD) studies how a (large) teacher model trained on one type of data (e.g., images) can guide a (smaller) student m

safetyarxiv-cs-ai
10 Jun 2026
Safety

Decision-Calibrated Conformal Uncertainty for Pacing Decisions in Streaming Advertising

DGX agent

arXiv:2606.10187v1 Announce Type: cross Abstract: We develop a decision-calibrated conformal framework for pacing decisions in streaming advertising. Pacing depends on uncertain future inventory, dema

safetyarxiv-cs-lg
10 Jun 2026
Safety

Decoupling Thought from Speech: Knowledge-Grounded Counterfactual Reasoning for Resilient Multi-Agent Argumentation

DGX agent

arXiv:2606.10475v1 Announce Type: cross Abstract: Multi-agent debate frameworks have been shown to improve large language model performance in convergent tasks, but they are currently optimized in a w

safetyarxiv-cs-ai
10 Jun 2026
Safety

DeRA-MOS: Optimizing Text-to-Music Evaluation via Decoupled Listwise Ranking and Modality Alignment

DGX agent

arXiv:2606.10010v1 Announce Type: cross Abstract: Evaluating text-to-music (TTM) systems remains expensive because music impression (MI) and text alignment (TA) scores rely on human mean opinion score

safetyarxiv-cs-ai
10 Jun 2026
← Previous
1…9495969798…267
Next →