AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,812 results
8 Jun 2026

EVA: Evolving Semantic Adversaries for Red-Teaming GUI Agents Against Environmental Injection Attacks

SafetyDGX agent

arXiv:2505.14289v2 Announce Type: replace Abstract: Graphical User Interface (GUI) agents powered by Multimodal Large Language Models (MLLMs) are increasingly deployed yet vulnerable to Environmental

Evaluating AI-based Scientific Knowledge Synthesis with Epidemiological Systematic Reviews

SafetyDGX agent

arXiv:2603.22327v2 Announce Type: replace-cross Abstract: Systematic literature reviews (SLRs) are a demanding and high-stakes form of scientific knowledge synthesis that remains underspecified as an

Evidence-Grounded Ensemble Diagnosis of 802.11 Packet Captures: A Multi-Stage Pipeline with Deterministic Reliability Scoring

SafetyDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.06871v1 Announce Type: new Abstract: Diagnosing 802.11 packet captures requires expert protocol knowledge, is slow, inconsistent across engineers, and unscalable. LLM-based approaches sound

Explicit Evidence Grounding via Structured Inline Citation Generation

SafetyDGX agent

arXiv:2606.07130v1 Announce Type: new Abstract: As AI systems become more widely adopted, the demand for factual and faithful generation grows. Properly attributing information through citations becom

Extending Responsibility-Sensitive Safety for the Assessment of Offloaded Autonomous Driving Services

SafetyDGX agent

arXiv:2606.07067v1 Announce Type: new Abstract: Safety is a fundamental requirement in the development of autonomous driving (AD) systems. While function offloading has demonstrated significant benefi

FAIR-Calib: Frontier-Aware Instability-Reweighted Calibration for Post-Training Quantization of Diffusion Large Language Models

SafetyDGX agent

arXiv:2606.06547v1 Announce Type: cross Abstract: Diffusion Large Language Models (dLLMs) refine tokens iteratively but commit them irreversibly, leading to a 'stability lag' where early decisions rem

Fast and Robust Convergence Rate for TD(0) with Linear Function Approximation, Universal Learning Steps and I.I.D. Samples

SafetyDGX agent

arXiv:2606.05967v2 Announce Type: replace-cross Abstract: In this paper, we study the finite-time behavior of the TD(0) temporal-difference method with linear function approximation (LFA). We consider

FIGMA: Towards FIne-Grained Music retrievAl

SafetyDGX agent

arXiv:2606.06615v1 Announce Type: cross Abstract: Retrieving music using natural language descriptions has improved with contrastive audio-text models such as CLAP, but current systems remain limited

ForensicConcept: Transferable Forensic Concepts for AIGI Detection

SafetyDGX agent

arXiv:2606.07034v1 Announce Type: new Abstract: AI-generated image detectors achieve high accuracy on in-distribution data but often fail on unseen generators. A key obstacle to understanding this fai

FreeAnimate: Training-Free Human Image Animation with Preview-Guided Denoising

SafetyDGX agent

arXiv:2606.06885v1 Announce Type: cross Abstract: Human Image Animation has seen significant advancements, primarily driven by diffusion models. However, existing methods typically demand substantial

further evidence that people shoveling money into AI are either bad at math or blind to risk

SafetyDGX agent

further evidence that people shoveling money into AI are either bad at math or blind to risk Striking paper from Wharton. The big conclusion: AI must increase productivity 2.7x -- and quickly -- or te

@GaryMarcus Marcus isn't wrong. Everyone bolts tools onto transformers because they hit walls fast. Neurosymbolic AI is rising for a reason:…

SafetyDGX agent

Gary Marcus argues that large language models based on transformers quickly reach their limitations, prompting developers to add external tools as workarounds, while neurosymbolic AI—which combines ne

Generative Modeling of Discrete Latent Structures via Dynamic Policy Gradients

SafetyDGX agent

arXiv:2606.07400v1 Announce Type: new Abstract: Many scientific problems require inferring unobserved mechanistic latent states from indirect observations. While classical approaches, including expect

Generative Models Erode Human Temporal Learning Through Market Selection

SafetyDGX agent

arXiv:2606.06572v1 Announce Type: cross Abstract: We argue that modern generative models create structural risks for knowledge and cultural production at current, sub-AGI capability levels. We define

GenPO++: Generative Policy Optimization with Jacobian-free Likelihood Ratios

SafetyDGX agent

arXiv:2606.06967v1 Announce Type: new Abstract: Generative policies provide expressive and multimodal action distributions, making them attractive for reinforcement learning (RL) in complex continuous

GRASP: Geometry-aware Residual Alignment for Scalable Pretraining Data Attribution

SafetyDGX agent

arXiv:2606.06892v1 Announce Type: new Abstract: Scalable data attribution methods typically assign isolated utility scores to individual training examples. This prevalent additive assumption fundament

Heterogeneous Effects of Green Finance on Urban Decarbonization: Evidence from 285 Cities in China

SafetyDGX agent

arXiv:2606.06986v1 Announce Type: new Abstract: While green finance has become a key instrument for low-carbon city transitions, its actual decarbonization effects and transmission mechanisms remain u

Historic: when the President of the United States can’t provide a single bit of specific concrete evidence to back up his serious charges, a…

SafetyDGX agent

Historic: when the President of the United States can’t provide a single bit of specific concrete evidence to back up his serious charges, and has a temper tantrum when pressed. Trump has a meltdown a

How reliable are LLMs when it comes to playing dice?

SafetyDGX agent

arXiv:2606.07515v1 Announce Type: cross Abstract: We investigate the probabilistic reasoning capabilities of large language models through a controlled benchmarking study on discrete probability probl

Hypers got hype but this is an embarrassingly bad piece of analysis that sets up a false dichotomy and ignores the fact that Terence Tao als…

SafetyDGX agent

Hypers got hype but this is an embarrassingly bad piece of analysis that sets up a false dichotomy and ignores the fact that Terence Tao also signed the Leiden Declaration. 🤦‍♂️ This post is unavailab

I am (seriously) toying with writing a book called Seven Lies About AI, and the title of a sequel just came to me: Stupid Nonsense, Said by …

SafetyDGX agent

Gary Marcus, a prominent AI researcher and critic, is considering writing a book titled 'Seven Lies About AI' that would debunk common misconceptions about artificial intelligence, with a planned sequ

i expect value for money to increase over time as competitive pressures force providers to lower prices. but in the short time, looks like t…

SafetyDGX agent

i expect value for money to increase over time as competitive pressures force providers to lower prices. but in the short time, looks like the opposite may be happening: Does a token buy you more or l

🤣 I have to think that anyone who buys the SpaceX IPO just isn’t that good at math. Or history.

SafetyDGX agent

🤣 I have to think that anyone who buys the SpaceX IPO just isn’t that good at math. Or history. “Garnering that almost 600x increase means hiking sales 50% a year, on average, for a decade.” 👇🏼 https:

I will start to take Geoff Hinton seriously again (I used to and don’t anymore) when - radiologists actually start losing jobs - he acknowle…

SafetyDGX agent

I will start to take Geoff Hinton seriously again (I used to and don’t anymore) when - radiologists actually start losing jobs - he acknowledges that LLMs do sometimes effectively memorize things - he

I would trust @eastdakota over Vinod Khosla any day of the week; it’s not even close. One is a man of integrity, with compassion for other h…

SafetyDGX agent

I would trust @eastdakota over Vinod Khosla any day of the week; it’s not even close. One is a man of integrity, with compassion for other human beings; the other isn’t. Uhhh… your memory is failing,

Import AI 460: Reward hacking society, RSI data from Anthropic; and RL-based quadcopter racing

SafetyDGX agent

This Import AI newsletter issue covers three main topics: societal implications of reward hacking (optimizing for measurable metrics at the expense of intended goals), new reinforcement learning data

in hindsight, this was chump change

SafetyDGX agent

in hindsight, this was chump change New: Microsoft and OpenAI plotting new 'Stargate' supercomputer that could cost $100 billion. An incredible look at what it will take to build the next generation o

Interpreting Brain Responses to Language with Sparse Features from Language Models

SafetyDGX agent

arXiv:2606.06857v1 Announce Type: new Abstract: A central goal of cognitive neuroscience is to characterize the features that are represented by human language cortex. Artificial language models (LMs)

Interpreting Learning Under Competing Models: Joint and Stepwise Approaches for Dynamic Cognitive Diagnosis

SafetyDGX agent

arXiv:2606.06804v1 Announce Type: new Abstract: Digital learning environments record learners' responses to individual items, making it possible to study the development of specific skills rather than

✅ July 2024: “public sentiment towards AI will turn sharply negative. In the end, AI leaders will be viewed like cigarette industry leaders.…

SafetyDGX agent

✅ July 2024: “public sentiment towards AI will turn sharply negative. In the end, AI leaders will be viewed like cigarette industry leaders.” Prediction: @a16z’s meddling will backfire massively. US w

Just-In-Time Reinforcement Learning: Continual Learning in LLM Agents Without Gradient Updates

SafetyDGX agent

arXiv:2601.18510v2 Announce Type: replace-cross Abstract: While Large Language Model (LLM) agents excel at general tasks, they inherently struggle with continual adaptation due to the frozen weights a

Korean Culture into LLM Alignment: Toward Cultural Coherence

SafetyDGX agent

arXiv:2606.06797v1 Announce Type: new Abstract: Cultural-aspect work on large language models is dominated by a negative target: which outputs to suppress. We argue that a constructive counterpart is

LARA: Latent Action Representation Alignment for Vision-Language-Action Models

SafetyDGX agent

arXiv:2606.07100v1 Announce Type: new Abstract: Visual-language action (VLA) models enable robots to predict actions directly from observations and language instructions, but their performance depends

Latent-space Attacks for Refusal Evasion in Language Models

SafetyDGX agent

arXiv:2605.21706v2 Announce Type: replace Abstract: Safety-aligned language models are trained to refuse harmful requests, yet refusal behavior can be suppressed by steering their internal representat

Learning All-Terrain Locomotion for a Planetary Rover with Actively Articulated Suspension

SafetyDGX agent

arXiv:2606.06790v1 Announce Type: cross Abstract: This paper presents ERNEST, a four-wheeled planetary rover concept equipped with a two-degree-of-freedom Active Gimbal Suspension that combines yaw an

Learning Fair Demand Models

SafetyDGX agent

arXiv:2606.06830v1 Announce Type: cross Abstract: Data-driven pricing is increasingly prevalent in sectors such as airlines, lending, insurance, and retail. By learning demand models from customer fea

LLM-Augmented Digital Twin for Policy Evaluation in Short-Video Platforms

SafetyDGX agent

arXiv:2603.11333v2 Announce Type: replace Abstract: Short-video platforms are closed-loop, human-in-the-loop ecosystems where platform policy, creator incentives, and user behavior co-evolve. This fee

MADRAG: Multi-Agent Debate with Retrieval-Augmented Generation for Training-Free Analytic Essay Scoring

SafetyDGX agent

arXiv:2606.06754v1 Announce Type: cross Abstract: We present MADRAG, a training-free framework for analytic essay scoring that combines multi-agent reasoning with retrieval-augmented grounding. Unlike

many young researchers think it might be good wipe out humanity. terrifying article. who are ceding control of our planet to?

SafetyDGX agent

many young researchers think it might be good wipe out humanity. terrifying article. who are ceding control of our planet to? Part 1 of my 'Pro-Human Manifesto' is now out! I'd love to know what you t

Mining Useful General Data for Low-Resource Domain Adaptation

SafetyDGX agent

arXiv:2511.07380v2 Announce Type: replace Abstract: Adapting large language models (LLMs) to low-resource domains remains challenging due to the scarcity of domain-specific data. While in-domain data

Mission-Level Runtime Assurance Framework for Autonomous Driving

SafetyDGX agent

arXiv:2606.06996v1 Announce Type: new Abstract: This paper studies runtime safety for autonomous driving when high-level driving commands become faulty or unreliable. Unlike conventional runtime-safet

More people need to listen to what Gary is saying here. Many, MANY of these talking heads who suck the oxygen out of the room with just thei…

SafetyDGX agent

More people need to listen to what Gary is saying here. Many, MANY of these talking heads who suck the oxygen out of the room with just their ego, say things that are so wildly wrong, that its misguid

Multi-Scale Feature Attention Network for Polymer Classification using THz Dual-Comb Spectroscopy

SafetyDGX agent

arXiv:2606.06554v1 Announce Type: cross Abstract: Reliable polymer identification is essential for ensuring the quality and safety of recycled plastics, yet conventional sorting and spectroscopic tech

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion

SafetyDGX agent

arXiv:2509.17446v3 Announce Type: replace-cross Abstract: Multimodal intent recognition (MMIR) suffers from weak semantic grounding and poor robustness under noisy or rare-class conditions. We propose

My own personal experience w Khosla is described here:

SafetyDGX agent

My own personal experience w Khosla is described here: The legendary investor Vinod Khosla is in the news because of questions about his integrity. I want to share my own experience: he lied viciously

Native3D: End-to-End 3D Scene Generation via Unified Mesh-Texture Modeling and Semantic Alignment

SafetyDGX agent

arXiv:2606.07117v1 Announce Type: cross Abstract: This paper presents Native3D, the first end-to-end 3D scene generation framework that completely bypasses 2D intermediate representations. Traditional

Network Recovery from Cascade Data: A Debiased Jacobian-Based Machine Learning Approach

SafetyDGX agent

arXiv:2606.07483v1 Announce Type: new Abstract: Many important outcomes unfold as dynamic cascades, including product adoption, disease spread, financial distress, and information diffusion. A central

Neuro-Symbolic Learning for Long-Horizon Task Planning Under Complex Logical Constraints

SafetyDGX agent

arXiv:2606.06877v1 Announce Type: cross Abstract: Task planning often suffers from severe efficiency bottlenecks when robots must reason over long-horizon action sequences under complex logical constr

No. Not by itself. Sergey Brin is absolutely wrong. Transformers by themselves are not “sufficient” for AGI. Nobody uses transformers on the…

SafetyDGX agent

No. Not by itself. Sergey Brin is absolutely wrong. Transformers by themselves are not “sufficient” for AGI. Nobody uses transformers on their own anymore. Everybody is supplementing them tools and ha

Nobody sane would pay list price for SpaceX.

SafetyDGX agent

Gary Marcus comments on SpaceX's valuation, suggesting that rational investors would negotiate below the asking price for the company, likely critiquing either the company's financial metrics, market

Online Pandora's Box for Contextual LLM Cascading

SafetyDGX agent

arXiv:2606.07392v1 Announce Type: new Abstract: Motivated by Large Language Model (LLM) cascading, we propose an online contextual Pandora's Box model for adaptively querying and selecting LLM APIs. I

≈ “our financials make no sense but we want to see how far we can push things”

SafetyDGX agent

≈ “our financials make no sense but we want to see how far we can push things” OAI files confidentially for IPO: “We have not decided on timing yet; it may be a while because there are things we want

Planning-aligned Token Compression for Long-Context Autonomous Driving

SafetyDGX agent

arXiv:2606.07464v1 Announce Type: cross Abstract: Monolithic vision-action models represent an emerging paradigm in autonomous driving. However, this architecture produces token sequences that quickly

Position: Don't Just 'Fix it in Post': A Science of AI Must Study Training Dynamics

SafetyDGX agent

arXiv:2606.06533v1 Announce Type: new Abstract: What would it mean to have a scientific understanding of AI? Models are not static objects: they are snapshots of time-evolving processes shaped by data

Predictive Statistics Shape Emergent World Representations of Grid Walkers

SafetyDGX agent

arXiv:2603.16689v2 Announce Type: replace Abstract: Next-token predictors often appear to develop internal representations of the latent world and its rules. The probabilistic nature of these models s

pretty much like i speculated yesterday

SafetyDGX agent

pretty much like i speculated yesterday 🚨 TRUMP’S AI NATIONALIZATION WAS SAM ALTMAN’S IDEA Trump: “I have spoken to ALL of them” >ai companies had no idea this was happening >they learned about this f

Progress-SQL: Improving Reinforcement Learning for Text-to-SQL via Progressive Rewards

SafetyDGX agent

arXiv:2606.06825v1 Announce Type: cross Abstract: Reinforcement learning has recently shown promise in improving large language models for Text-to-SQL generation, yet existing methods typically optimi

QuadVerse: An Integrated Framework Aligning Visual-Physical Reality for Quadruped Simulation

SafetyDGX agent

arXiv:2606.07118v1 Announce Type: new Abstract: Simulation is central to robot learning, yet the sim-to-real gap remains a major bottleneck.Existing approaches often tackle visual or dynamic gaps sepa

Rapid co-design of Buoyancy-assisted robots for Challenging Locomotion using Gaussian Evolutionary Specialists

SafetyDGX agent

arXiv:2606.07424v1 Announce Type: new Abstract: Designing high-performance legged robots requires jointly optimizing morphology and control. Model-free Reinforcement Learning (RL) offers an alternativ

RASFT: Rollout-Adaptive Supervised Fine-Tuning for Reasoning

SafetyDGX agent

arXiv:2606.07006v1 Announce Type: cross Abstract: Supervised fine-tuning (SFT) is a prevailing method for adapting large language models to reasoning tasks by imitating offline expert demonstrations,

← Previous
1…8485868788…214
Next →