AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,813 results
Safety

check out Marcus on AI

DGX agent

Gary Marcus, a prominent AI researcher and critic, shared a post on X directing followers to learn more about his perspectives on artificial intelligence. The post likely promotes his work, writings,

safetygary-marcus--x
8 Jun 2026
Safety

Chunking the Critic: A Transformer-based Soft Actor-Critic with N-Step Returns

DGX agent

arXiv:2503.03660v4 Announce Type: replace Abstract: We introduce a sequence-conditioned critic for Soft Actor-Critic (SAC) that models trajectory context with a lightweight Transformer and trains on a

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
safetyarxiv-cs-lg
8 Jun 2026
Safety

Coarse-to-Control: Action-Token Planning for Vision-Language-Action Models

DGX agent

arXiv:2606.07107v1 Announce Type: new Abstract: Most vision-language-action (VLA) models map observations directly to actions without explicit intermediate planning, which limits performance on long-h

safetyarxiv-cs-ro
8 Jun 2026
Safety

Consistent-Inversion: Reverse Consistency Guidance for Structure-Preserving Visual Editing

DGX agent

arXiv:2606.07145v1 Announce Type: new Abstract: Text-guided diffusion models have become effective tools for real-image visual editing, where the edited image must follow a target instruction while pr

safetyarxiv-cs-cv
8 Jun 2026
Safety

Covariance Shrinkage via Stochastic Interpolation

DGX agent

arXiv:2606.07382v1 Announce Type: new Abstract: We recast classical shrinkage of high-dimensional covariance estimators as empirical risk minimization over a parametric stochastic interpolant between

safetyarxiv-cs-lg
8 Jun 2026
Safety

Data-Efficient Autoregressive-to-Diffusion Language Models via On-Policy Distillation

DGX agent

arXiv:2606.06712v1 Announce Type: cross Abstract: We study the transformation of autoregressive models (ARLMs) into diffusion language models (DLMs). Rather than pretraining from scratch, prior work r

safetyarxiv-cs-ai
8 Jun 2026
Safety

DEFINED: A Data-Efficient Computational Framework for Fine-Grained Creativity Assessment in Debate Scenarios

DGX agent

arXiv:2606.07226v1 Announce Type: cross Abstract: Human creativity has emerged as a critical competency in the era of large language models. Assessing creativity in complex, open-ended environments is

safetyarxiv-cs-ai
8 Jun 2026
Safety

Depth over Fidelity in Fixed-Budget Noisy Evolution Strategies

DGX agent

arXiv:2606.06555v1 Announce Type: cross Abstract: Noisy evolution strategies under fixed evaluation budgets face a depth-fidelity trade-off: spending evaluations to denoise intra-generation rankings r

safetyarxiv-cs-lg
8 Jun 2026
Safety

Detecting and Mitigating Bias by Treating Fairness as a Symmetry Operation

DGX agent

arXiv:2606.06514v1 Announce Type: new Abstract: Machine learning systems deployed in high stakes socioeconomic settings routinely display bias. We formalize bias as a symmetry breaking operation: a cl

safetyarxiv-cs-ai
8 Jun 2026
Safety

Didact: A Cross-Domain Capability Discovery System for Defence

DGX agent

arXiv:2606.06942v1 Announce Type: cross Abstract: Policymakers in defence and defence-aligned sectors must monitor rapidly evolving research alongside sector priorities relevant to operational and str

safetyarxiv-cs-ai
8 Jun 2026
Safety

DOPPLER: Dual-Policy Learning for Device Assignment in Asynchronous Dataflow Graphs

DGX agent

arXiv:2505.23131v2 Announce Type: replace Abstract: We study the problem of assigning operations in a dataflow graph to devices to minimize execution time in a work-conserving system, with emphasis on

safetyarxiv-cs-lg
8 Jun 2026
Safety

Elmes*: Automated Construction of Fine-Grained Evaluation Rubrics for Large Language Models in Long-Tail Educational Scenarios

DGX agent

arXiv:2606.06546v1 Announce Type: new Abstract: Evaluating large language models (LLMs) for education requires measuring how models teach, not only what they know. Existing benchmarks emphasize domain

safetyarxiv-cs-lg
8 Jun 2026
Safety

Empirical Transfer Operators and Finite-Sample Change Detection for Noisy Expanding Interval Maps

DGX agent

arXiv:2606.06785v1 Announce Type: cross Abstract: We study finite-sample change detection for one-dimensional noisy dynamical systems using partition-based empirical approximations of stationary behav

safetyarxiv-cs-lg
8 Jun 2026
Safety

EVA: Evolving Semantic Adversaries for Red-Teaming GUI Agents Against Environmental Injection Attacks

DGX agent

arXiv:2505.14289v2 Announce Type: replace Abstract: Graphical User Interface (GUI) agents powered by Multimodal Large Language Models (MLLMs) are increasingly deployed yet vulnerable to Environmental

safetyarxiv-cs-ai
8 Jun 2026
Safety

Evaluating AI-based Scientific Knowledge Synthesis with Epidemiological Systematic Reviews

DGX agent

arXiv:2603.22327v2 Announce Type: replace-cross Abstract: Systematic literature reviews (SLRs) are a demanding and high-stakes form of scientific knowledge synthesis that remains underspecified as an

safetyarxiv-cs-ai
8 Jun 2026
Safety

Evidence-Grounded Ensemble Diagnosis of 802.11 Packet Captures: A Multi-Stage Pipeline with Deterministic Reliability Scoring

DGX agent

arXiv:2606.06871v1 Announce Type: new Abstract: Diagnosing 802.11 packet captures requires expert protocol knowledge, is slow, inconsistent across engineers, and unscalable. LLM-based approaches sound

safetyarxiv-cs-lg
8 Jun 2026
Safety

Explicit Evidence Grounding via Structured Inline Citation Generation

DGX agent

arXiv:2606.07130v1 Announce Type: new Abstract: As AI systems become more widely adopted, the demand for factual and faithful generation grows. Properly attributing information through citations becom

safetyarxiv-cs-cl
8 Jun 2026
Safety

Extending Responsibility-Sensitive Safety for the Assessment of Offloaded Autonomous Driving Services

DGX agent

arXiv:2606.07067v1 Announce Type: new Abstract: Safety is a fundamental requirement in the development of autonomous driving (AD) systems. While function offloading has demonstrated significant benefi

safetyarxiv-cs-ro
8 Jun 2026
Safety

FAIR-Calib: Frontier-Aware Instability-Reweighted Calibration for Post-Training Quantization of Diffusion Large Language Models

DGX agent

arXiv:2606.06547v1 Announce Type: cross Abstract: Diffusion Large Language Models (dLLMs) refine tokens iteratively but commit them irreversibly, leading to a 'stability lag' where early decisions rem

safetyarxiv-cs-ai
8 Jun 2026
Safety

Fast and Robust Convergence Rate for TD(0) with Linear Function Approximation, Universal Learning Steps and I.I.D. Samples

DGX agent

arXiv:2606.05967v2 Announce Type: replace-cross Abstract: In this paper, we study the finite-time behavior of the TD(0) temporal-difference method with linear function approximation (LFA). We consider

safetyarxiv-cs-lg
8 Jun 2026
Safety

FIGMA: Towards FIne-Grained Music retrievAl

DGX agent

arXiv:2606.06615v1 Announce Type: cross Abstract: Retrieving music using natural language descriptions has improved with contrastive audio-text models such as CLAP, but current systems remain limited

safetyarxiv-cs-ai
8 Jun 2026
Safety

ForensicConcept: Transferable Forensic Concepts for AIGI Detection

DGX agent

arXiv:2606.07034v1 Announce Type: new Abstract: AI-generated image detectors achieve high accuracy on in-distribution data but often fail on unseen generators. A key obstacle to understanding this fai

safetyarxiv-cs-cv
8 Jun 2026
Safety

FreeAnimate: Training-Free Human Image Animation with Preview-Guided Denoising

DGX agent

arXiv:2606.06885v1 Announce Type: cross Abstract: Human Image Animation has seen significant advancements, primarily driven by diffusion models. However, existing methods typically demand substantial

safetyarxiv-cs-ai
8 Jun 2026
Safety

further evidence that people shoveling money into AI are either bad at math or blind to risk

DGX agent

further evidence that people shoveling money into AI are either bad at math or blind to risk Striking paper from Wharton. The big conclusion: AI must increase productivity 2.7x -- and quickly -- or te

safetygary-marcus--x
8 Jun 2026
Safety

@GaryMarcus Marcus isn't wrong. Everyone bolts tools onto transformers because they hit walls fast. Neurosymbolic AI is rising for a reason:…

DGX agent

Gary Marcus argues that large language models based on transformers quickly reach their limitations, prompting developers to add external tools as workarounds, while neurosymbolic AI—which combines ne

safetygary-marcus--x
8 Jun 2026
Safety

Generative Modeling of Discrete Latent Structures via Dynamic Policy Gradients

DGX agent

arXiv:2606.07400v1 Announce Type: new Abstract: Many scientific problems require inferring unobserved mechanistic latent states from indirect observations. While classical approaches, including expect

safetyarxiv-cs-lg
8 Jun 2026
Safety

Generative Models Erode Human Temporal Learning Through Market Selection

DGX agent

arXiv:2606.06572v1 Announce Type: cross Abstract: We argue that modern generative models create structural risks for knowledge and cultural production at current, sub-AGI capability levels. We define

safetyarxiv-cs-ai
8 Jun 2026
Safety

GenPO++: Generative Policy Optimization with Jacobian-free Likelihood Ratios

DGX agent

arXiv:2606.06967v1 Announce Type: new Abstract: Generative policies provide expressive and multimodal action distributions, making them attractive for reinforcement learning (RL) in complex continuous

safetyarxiv-cs-lg
8 Jun 2026
Safety

GRASP: Geometry-aware Residual Alignment for Scalable Pretraining Data Attribution

DGX agent

arXiv:2606.06892v1 Announce Type: new Abstract: Scalable data attribution methods typically assign isolated utility scores to individual training examples. This prevalent additive assumption fundament

safetyarxiv-cs-lg
8 Jun 2026
Safety

Heterogeneous Effects of Green Finance on Urban Decarbonization: Evidence from 285 Cities in China

DGX agent

arXiv:2606.06986v1 Announce Type: new Abstract: While green finance has become a key instrument for low-carbon city transitions, its actual decarbonization effects and transmission mechanisms remain u

safetyarxiv-cs-lg
8 Jun 2026
Safety

Historic: when the President of the United States can’t provide a single bit of specific concrete evidence to back up his serious charges, a…

DGX agent

Historic: when the President of the United States can’t provide a single bit of specific concrete evidence to back up his serious charges, and has a temper tantrum when pressed. Trump has a meltdown a

safetygary-marcus--x
8 Jun 2026
Safety

How reliable are LLMs when it comes to playing dice?

DGX agent

arXiv:2606.07515v1 Announce Type: cross Abstract: We investigate the probabilistic reasoning capabilities of large language models through a controlled benchmarking study on discrete probability probl

safetyarxiv-cs-ai
8 Jun 2026
Safety

Hypers got hype but this is an embarrassingly bad piece of analysis that sets up a false dichotomy and ignores the fact that Terence Tao als…

DGX agent

Hypers got hype but this is an embarrassingly bad piece of analysis that sets up a false dichotomy and ignores the fact that Terence Tao also signed the Leiden Declaration. 🤦‍♂️ This post is unavailab

safetygary-marcus--x
8 Jun 2026
Safety

I am (seriously) toying with writing a book called Seven Lies About AI, and the title of a sequel just came to me: Stupid Nonsense, Said by …

DGX agent

Gary Marcus, a prominent AI researcher and critic, is considering writing a book titled 'Seven Lies About AI' that would debunk common misconceptions about artificial intelligence, with a planned sequ

safetygary-marcus--x
8 Jun 2026
Safety

i expect value for money to increase over time as competitive pressures force providers to lower prices. but in the short time, looks like t…

DGX agent

i expect value for money to increase over time as competitive pressures force providers to lower prices. but in the short time, looks like the opposite may be happening: Does a token buy you more or l

safetygary-marcus--x
8 Jun 2026
Safety

🤣 I have to think that anyone who buys the SpaceX IPO just isn’t that good at math. Or history.

DGX agent

🤣 I have to think that anyone who buys the SpaceX IPO just isn’t that good at math. Or history. “Garnering that almost 600x increase means hiking sales 50% a year, on average, for a decade.” 👇🏼 https:

safetygary-marcus--x
8 Jun 2026
Safety

I will start to take Geoff Hinton seriously again (I used to and don’t anymore) when - radiologists actually start losing jobs - he acknowle…

DGX agent

I will start to take Geoff Hinton seriously again (I used to and don’t anymore) when - radiologists actually start losing jobs - he acknowledges that LLMs do sometimes effectively memorize things - he

safetygary-marcus--x
8 Jun 2026
Safety

I would trust @eastdakota over Vinod Khosla any day of the week; it’s not even close. One is a man of integrity, with compassion for other h…

DGX agent

I would trust @eastdakota over Vinod Khosla any day of the week; it’s not even close. One is a man of integrity, with compassion for other human beings; the other isn’t. Uhhh… your memory is failing,

safetygary-marcus--x
8 Jun 2026
Safety

Import AI 460: Reward hacking society, RSI data from Anthropic; and RL-based quadcopter racing

DGX agent

This Import AI newsletter issue covers three main topics: societal implications of reward hacking (optimizing for measurable metrics at the expense of intended goals), new reinforcement learning data

safetyimport-ai
8 Jun 2026
Safety

in hindsight, this was chump change

DGX agent

in hindsight, this was chump change New: Microsoft and OpenAI plotting new 'Stargate' supercomputer that could cost $100 billion. An incredible look at what it will take to build the next generation o

safetygary-marcus--x
8 Jun 2026
Safety

Interpreting Brain Responses to Language with Sparse Features from Language Models

DGX agent

arXiv:2606.06857v1 Announce Type: new Abstract: A central goal of cognitive neuroscience is to characterize the features that are represented by human language cortex. Artificial language models (LMs)

safetyarxiv-cs-cl
8 Jun 2026
Safety

Interpreting Learning Under Competing Models: Joint and Stepwise Approaches for Dynamic Cognitive Diagnosis

DGX agent

arXiv:2606.06804v1 Announce Type: new Abstract: Digital learning environments record learners' responses to individual items, making it possible to study the development of specific skills rather than

safetyarxiv-cs-lg
8 Jun 2026
Safety

✅ July 2024: “public sentiment towards AI will turn sharply negative. In the end, AI leaders will be viewed like cigarette industry leaders.…

DGX agent

✅ July 2024: “public sentiment towards AI will turn sharply negative. In the end, AI leaders will be viewed like cigarette industry leaders.” Prediction: @a16z’s meddling will backfire massively. US w

safetygary-marcus--x
8 Jun 2026
Safety

Just-In-Time Reinforcement Learning: Continual Learning in LLM Agents Without Gradient Updates

DGX agent

arXiv:2601.18510v2 Announce Type: replace-cross Abstract: While Large Language Model (LLM) agents excel at general tasks, they inherently struggle with continual adaptation due to the frozen weights a

safetyarxiv-cs-ai
8 Jun 2026
Safety

Korean Culture into LLM Alignment: Toward Cultural Coherence

DGX agent

arXiv:2606.06797v1 Announce Type: new Abstract: Cultural-aspect work on large language models is dominated by a negative target: which outputs to suppress. We argue that a constructive counterpart is

safetyarxiv-cs-cl
8 Jun 2026
Safety

LARA: Latent Action Representation Alignment for Vision-Language-Action Models

DGX agent

arXiv:2606.07100v1 Announce Type: new Abstract: Visual-language action (VLA) models enable robots to predict actions directly from observations and language instructions, but their performance depends

safetyarxiv-cs-cv
8 Jun 2026
Safety

Latent-space Attacks for Refusal Evasion in Language Models

DGX agent

arXiv:2605.21706v2 Announce Type: replace Abstract: Safety-aligned language models are trained to refuse harmful requests, yet refusal behavior can be suppressed by steering their internal representat

safetyarxiv-cs-ai
8 Jun 2026
Safety

Learning All-Terrain Locomotion for a Planetary Rover with Actively Articulated Suspension

DGX agent

arXiv:2606.06790v1 Announce Type: cross Abstract: This paper presents ERNEST, a four-wheeled planetary rover concept equipped with a two-degree-of-freedom Active Gimbal Suspension that combines yaw an

safetyarxiv-cs-lg
8 Jun 2026
← Previous
1…105106107108109…267
Next →