AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,490 results
8 Jun 2026

further evidence that people shoveling money into AI are either bad at math or blind to risk

SafetyDGX agent

further evidence that people shoveling money into AI are either bad at math or blind to risk Striking paper from Wharton. The big conclusion: AI must increase productivity 2.7x -- and quickly -- or te

@GaryMarcus Marcus isn't wrong. Everyone bolts tools onto transformers because they hit walls fast. Neurosymbolic AI is rising for a reason:…

SafetyDGX agent

Gary Marcus argues that large language models based on transformers quickly reach their limitations, prompting developers to add external tools as workarounds, while neurosymbolic AI—which combines ne

Generative Modeling of Discrete Latent Structures via Dynamic Policy Gradients

SafetyDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.07400v1 Announce Type: new Abstract: Many scientific problems require inferring unobserved mechanistic latent states from indirect observations. While classical approaches, including expect

Generative Models Erode Human Temporal Learning Through Market Selection

SafetyDGX agent

arXiv:2606.06572v1 Announce Type: cross Abstract: We argue that modern generative models create structural risks for knowledge and cultural production at current, sub-AGI capability levels. We define

GenPO++: Generative Policy Optimization with Jacobian-free Likelihood Ratios

SafetyDGX agent

arXiv:2606.06967v1 Announce Type: new Abstract: Generative policies provide expressive and multimodal action distributions, making them attractive for reinforcement learning (RL) in complex continuous

GRASP: Geometry-aware Residual Alignment for Scalable Pretraining Data Attribution

SafetyDGX agent

arXiv:2606.06892v1 Announce Type: new Abstract: Scalable data attribution methods typically assign isolated utility scores to individual training examples. This prevalent additive assumption fundament

Heterogeneous Effects of Green Finance on Urban Decarbonization: Evidence from 285 Cities in China

SafetyDGX agent

arXiv:2606.06986v1 Announce Type: new Abstract: While green finance has become a key instrument for low-carbon city transitions, its actual decarbonization effects and transmission mechanisms remain u

Historic: when the President of the United States can’t provide a single bit of specific concrete evidence to back up his serious charges, a…

SafetyDGX agent

Historic: when the President of the United States can’t provide a single bit of specific concrete evidence to back up his serious charges, and has a temper tantrum when pressed. Trump has a meltdown a

How reliable are LLMs when it comes to playing dice?

SafetyDGX agent

arXiv:2606.07515v1 Announce Type: cross Abstract: We investigate the probabilistic reasoning capabilities of large language models through a controlled benchmarking study on discrete probability probl

Hypers got hype but this is an embarrassingly bad piece of analysis that sets up a false dichotomy and ignores the fact that Terence Tao als…

SafetyDGX agent

Hypers got hype but this is an embarrassingly bad piece of analysis that sets up a false dichotomy and ignores the fact that Terence Tao also signed the Leiden Declaration. 🤦‍♂️ This post is unavailab

I am (seriously) toying with writing a book called Seven Lies About AI, and the title of a sequel just came to me: Stupid Nonsense, Said by …

SafetyDGX agent

Gary Marcus, a prominent AI researcher and critic, is considering writing a book titled 'Seven Lies About AI' that would debunk common misconceptions about artificial intelligence, with a planned sequ

i expect value for money to increase over time as competitive pressures force providers to lower prices. but in the short time, looks like t…

SafetyDGX agent

i expect value for money to increase over time as competitive pressures force providers to lower prices. but in the short time, looks like the opposite may be happening: Does a token buy you more or l

🤣 I have to think that anyone who buys the SpaceX IPO just isn’t that good at math. Or history.

SafetyDGX agent

🤣 I have to think that anyone who buys the SpaceX IPO just isn’t that good at math. Or history. “Garnering that almost 600x increase means hiking sales 50% a year, on average, for a decade.” 👇🏼 https:

I will start to take Geoff Hinton seriously again (I used to and don’t anymore) when - radiologists actually start losing jobs - he acknowle…

SafetyDGX agent

I will start to take Geoff Hinton seriously again (I used to and don’t anymore) when - radiologists actually start losing jobs - he acknowledges that LLMs do sometimes effectively memorize things - he

I would trust @eastdakota over Vinod Khosla any day of the week; it’s not even close. One is a man of integrity, with compassion for other h…

SafetyDGX agent

I would trust @eastdakota over Vinod Khosla any day of the week; it’s not even close. One is a man of integrity, with compassion for other human beings; the other isn’t. Uhhh… your memory is failing,

in hindsight, this was chump change

SafetyDGX agent

in hindsight, this was chump change New: Microsoft and OpenAI plotting new 'Stargate' supercomputer that could cost $100 billion. An incredible look at what it will take to build the next generation o

Interpreting Brain Responses to Language with Sparse Features from Language Models

SafetyDGX agent

arXiv:2606.06857v1 Announce Type: new Abstract: A central goal of cognitive neuroscience is to characterize the features that are represented by human language cortex. Artificial language models (LMs)

Interpreting Learning Under Competing Models: Joint and Stepwise Approaches for Dynamic Cognitive Diagnosis

SafetyDGX agent

arXiv:2606.06804v1 Announce Type: new Abstract: Digital learning environments record learners' responses to individual items, making it possible to study the development of specific skills rather than

✅ July 2024: “public sentiment towards AI will turn sharply negative. In the end, AI leaders will be viewed like cigarette industry leaders.…

SafetyDGX agent

✅ July 2024: “public sentiment towards AI will turn sharply negative. In the end, AI leaders will be viewed like cigarette industry leaders.” Prediction: @a16z’s meddling will backfire massively. US w

Just-In-Time Reinforcement Learning: Continual Learning in LLM Agents Without Gradient Updates

SafetyDGX agent

arXiv:2601.18510v2 Announce Type: replace-cross Abstract: While Large Language Model (LLM) agents excel at general tasks, they inherently struggle with continual adaptation due to the frozen weights a

Korean Culture into LLM Alignment: Toward Cultural Coherence

SafetyDGX agent

arXiv:2606.06797v1 Announce Type: new Abstract: Cultural-aspect work on large language models is dominated by a negative target: which outputs to suppress. We argue that a constructive counterpart is

LARA: Latent Action Representation Alignment for Vision-Language-Action Models

SafetyDGX agent

arXiv:2606.07100v1 Announce Type: new Abstract: Visual-language action (VLA) models enable robots to predict actions directly from observations and language instructions, but their performance depends

Learning All-Terrain Locomotion for a Planetary Rover with Actively Articulated Suspension

SafetyDGX agent

arXiv:2606.06790v1 Announce Type: cross Abstract: This paper presents ERNEST, a four-wheeled planetary rover concept equipped with a two-degree-of-freedom Active Gimbal Suspension that combines yaw an

Learning Fair Demand Models

SafetyDGX agent

arXiv:2606.06830v1 Announce Type: cross Abstract: Data-driven pricing is increasingly prevalent in sectors such as airlines, lending, insurance, and retail. By learning demand models from customer fea

LLM-Augmented Digital Twin for Policy Evaluation in Short-Video Platforms

SafetyDGX agent

arXiv:2603.11333v2 Announce Type: replace Abstract: Short-video platforms are closed-loop, human-in-the-loop ecosystems where platform policy, creator incentives, and user behavior co-evolve. This fee

MADRAG: Multi-Agent Debate with Retrieval-Augmented Generation for Training-Free Analytic Essay Scoring

SafetyDGX agent

arXiv:2606.06754v1 Announce Type: cross Abstract: We present MADRAG, a training-free framework for analytic essay scoring that combines multi-agent reasoning with retrieval-augmented grounding. Unlike

many young researchers think it might be good wipe out humanity. terrifying article. who are ceding control of our planet to?

SafetyDGX agent

many young researchers think it might be good wipe out humanity. terrifying article. who are ceding control of our planet to? Part 1 of my 'Pro-Human Manifesto' is now out! I'd love to know what you t

Mining Useful General Data for Low-Resource Domain Adaptation

SafetyDGX agent

arXiv:2511.07380v2 Announce Type: replace Abstract: Adapting large language models (LLMs) to low-resource domains remains challenging due to the scarcity of domain-specific data. While in-domain data

More people need to listen to what Gary is saying here. Many, MANY of these talking heads who suck the oxygen out of the room with just thei…

SafetyDGX agent

More people need to listen to what Gary is saying here. Many, MANY of these talking heads who suck the oxygen out of the room with just their ego, say things that are so wildly wrong, that its misguid

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion

SafetyDGX agent

arXiv:2509.17446v3 Announce Type: replace-cross Abstract: Multimodal intent recognition (MMIR) suffers from weak semantic grounding and poor robustness under noisy or rare-class conditions. We propose

My own personal experience w Khosla is described here:

SafetyDGX agent

My own personal experience w Khosla is described here: The legendary investor Vinod Khosla is in the news because of questions about his integrity. I want to share my own experience: he lied viciously

Native3D: End-to-End 3D Scene Generation via Unified Mesh-Texture Modeling and Semantic Alignment

SafetyDGX agent

arXiv:2606.07117v1 Announce Type: cross Abstract: This paper presents Native3D, the first end-to-end 3D scene generation framework that completely bypasses 2D intermediate representations. Traditional

Network Recovery from Cascade Data: A Debiased Jacobian-Based Machine Learning Approach

SafetyDGX agent

arXiv:2606.07483v1 Announce Type: new Abstract: Many important outcomes unfold as dynamic cascades, including product adoption, disease spread, financial distress, and information diffusion. A central

Neuro-Symbolic Learning for Long-Horizon Task Planning Under Complex Logical Constraints

SafetyDGX agent

arXiv:2606.06877v1 Announce Type: cross Abstract: Task planning often suffers from severe efficiency bottlenecks when robots must reason over long-horizon action sequences under complex logical constr

No. Not by itself. Sergey Brin is absolutely wrong. Transformers by themselves are not “sufficient” for AGI. Nobody uses transformers on the…

SafetyDGX agent

No. Not by itself. Sergey Brin is absolutely wrong. Transformers by themselves are not “sufficient” for AGI. Nobody uses transformers on their own anymore. Everybody is supplementing them tools and ha

Nobody sane would pay list price for SpaceX.

SafetyDGX agent

Gary Marcus comments on SpaceX's valuation, suggesting that rational investors would negotiate below the asking price for the company, likely critiquing either the company's financial metrics, market

Online Pandora's Box for Contextual LLM Cascading

SafetyDGX agent

arXiv:2606.07392v1 Announce Type: new Abstract: Motivated by Large Language Model (LLM) cascading, we propose an online contextual Pandora's Box model for adaptively querying and selecting LLM APIs. I

≈ “our financials make no sense but we want to see how far we can push things”

SafetyDGX agent

≈ “our financials make no sense but we want to see how far we can push things” OAI files confidentially for IPO: “We have not decided on timing yet; it may be a while because there are things we want

Planning-aligned Token Compression for Long-Context Autonomous Driving

SafetyDGX agent

arXiv:2606.07464v1 Announce Type: cross Abstract: Monolithic vision-action models represent an emerging paradigm in autonomous driving. However, this architecture produces token sequences that quickly

Predictive Statistics Shape Emergent World Representations of Grid Walkers

SafetyDGX agent

arXiv:2603.16689v2 Announce Type: replace Abstract: Next-token predictors often appear to develop internal representations of the latent world and its rules. The probabilistic nature of these models s

pretty much like i speculated yesterday

SafetyDGX agent

pretty much like i speculated yesterday 🚨 TRUMP’S AI NATIONALIZATION WAS SAM ALTMAN’S IDEA Trump: “I have spoken to ALL of them” >ai companies had no idea this was happening >they learned about this f

Progress-SQL: Improving Reinforcement Learning for Text-to-SQL via Progressive Rewards

SafetyDGX agent

arXiv:2606.06825v1 Announce Type: cross Abstract: Reinforcement learning has recently shown promise in improving large language models for Text-to-SQL generation, yet existing methods typically optimi

QuadVerse: An Integrated Framework Aligning Visual-Physical Reality for Quadruped Simulation

SafetyDGX agent

arXiv:2606.07118v1 Announce Type: new Abstract: Simulation is central to robot learning, yet the sim-to-real gap remains a major bottleneck.Existing approaches often tackle visual or dynamic gaps sepa

Rapid co-design of Buoyancy-assisted robots for Challenging Locomotion using Gaussian Evolutionary Specialists

SafetyDGX agent

arXiv:2606.07424v1 Announce Type: new Abstract: Designing high-performance legged robots requires jointly optimizing morphology and control. Model-free Reinforcement Learning (RL) offers an alternativ

RASFT: Rollout-Adaptive Supervised Fine-Tuning for Reasoning

SafetyDGX agent

arXiv:2606.07006v1 Announce Type: cross Abstract: Supervised fine-tuning (SFT) is a prevailing method for adapting large language models to reasoning tasks by imitating offline expert demonstrations,

RAVEN: Retrieval-Augmented Vulnerability Exploration Network for Memory Corruption Analysis in User Code and Binary Programs

SafetyDGX agent

arXiv:2604.17948v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities across various cybersecurity tasks, including vulnerability classificat

Robotic Policy Adaptation via Weight-Space Meta-Learning

SafetyDGX agent

arXiv:2606.07217v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models are emerging as a promising paradigm for robotic manipulation, enabling general-purpose policies trained from larg

Robots Need More than VLA and World Models

SafetyDGX agent

arXiv:2606.06556v1 Announce Type: new Abstract: Generalist robot intelligence is often framed as a policy-scaling problem: collect more robot demonstrations, train larger Vision-Language-Action (VLA)

Self-evolving LLM agents with in-distribution Optimization

SafetyDGX agent

arXiv:2606.07367v1 Announce Type: new Abstract: Large Language Models (LLMs) have recently emerged as powerful controllers for interactive agents in complex environments, yet training them to perform

Semantic-Structural Alignment for Generative Pictorial Charts

SafetyDGX agent

arXiv:2606.06498v1 Announce Type: cross Abstract: Traditional statistical graphics are precise but often lack the visual appeal, memorability, and engagement of pictorial charts. We present a generati

SERNF: Sample-Efficient Real-World Dexterous Policy Fine-Tuning via Action-Chunked Critics and Normalizing Flows

SafetyDGX agent

arXiv:2602.09580v4 Announce Type: replace-cross Abstract: Real-world fine-tuning of dexterous manipulation policies remains challenging due to limited real-world interaction budgets and highly multimo

Silverfort brings runtime identity controls to Microsoft Copilot Studio agents

SafetyDGX agent

Identity security company Silverfort Inc. today launched an integration that applies its identity and access controls to artificial intelligence agents built into Microsoft Corp.’s Copilot Studio, enf

Simulation-Driven Imitation Learning for Biosignals-Free Shared-Autonomy Prosthetic Grasping

SafetyDGX agent

arXiv:2606.07389v1 Announce Type: new Abstract: Biosignals-free shared-autonomy control of upper-limb prosthetic hands aims to enable natural and low-effort manipulation without relying on EMG or othe

SlimSearcher: Training Efficiency-Aware Web Agents via Adaptive Reward Gating

SafetyDGX agent

arXiv:2606.07074v1 Announce Type: cross Abstract: Deep research agents have demonstrated remarkable capabilities in complex information-seeking tasks, yet this power comes at a steep computational cos

Socratic-SWE: Self-Evolving Coding Agents via Trace-Derived Agent Skills

SafetyDGX agent

arXiv:2606.07412v1 Announce Type: cross Abstract: LLM-driven software engineering agents have become a central testbed for real-world language-model capability, yet their training remains limited by t

SpaceX goes public Friday at ~94x revenue. Across 45 years of data, IPOs that debut above 40x sales underperform the market by 58% over the …

SafetyDGX agent

SpaceX goes public Friday at ~94x revenue. Across 45 years of data, IPOs that debut above 40x sales underperform the market by 58% over the next 3 years, and by 76% style-adjusted. The golden rule of

Stable Reasoning, Unstable Responses: Mitigating LLM Deception via Stability Asymmetry

SafetyDGX agent

arXiv:2603.26846v2 Announce Type: replace-cross Abstract: As Large Language Models (LLMs) expand in capability and application scope, their trustworthiness becomes critical. A vital risk is intrinsic

Step-Wise Refusal Dynamics in Autoregressive and Diffusion Language Models

SafetyDGX agent

arXiv:2602.02600v3 Announce Type: replace-cross Abstract: Diffusion language models (DLMs) have recently emerged as a competitive alternative to autoregressive (AR) models, offering parallel decoding,

SV-Detect: AI-generated Text Detection with Steering Vectors

SafetyDGX agent

arXiv:2606.07313v1 Announce Type: cross Abstract: Detecting machine-generated text is especially difficult under distribution shift, such as transfer across domains, source models, and editing attacks

Sycophantic Praise: Evaluating Excessive Praise in Language Models

SafetyDGX agent

arXiv:2606.07441v1 Announce Type: new Abstract: Sycophancy in language models is typically studied as excessive agreement or validation, while explicit praise and flattery have received comparatively

← Previous
1…119120121122123…242
Next →