AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,809 results
16 May 2026

tesla robotaxis going about as well you might expect.

SafetyDGX agent

Gary Marcus, a prominent AI researcher and critic, comments on Tesla's robotaxi development progress, likely offering skeptical or cautionary observations about the practical challenges and timeline o

Truly an all-star cast, on one of the most important questions in AI. Thrilled to see some many people finally willing to confront the hard …

SafetyDGX agent

Truly an all-star cast, on one of the most important questions in AI. Thrilled to see some many people finally willing to confront the hard questions of how we can move beyond LLMs, and into what worl

💯. Way too much focus on language models.

SafetyDGX agent

💯. Way too much focus on language models. Fei-Fei Li warns that AI may be staring too hard at language models. The world is not just text on a screen. It is physical, visual, spatial, and always chang


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
15 May 2026

A cross-species neural foundation model for end-to-end speech decoding

SafetyDGX agent

arXiv:2511.21740v5 Announce Type: replace-cross Abstract: Speech brain-computer interfaces (BCIs) aim to restore communication for people with paralysis by translating neural activity into text. Most

A Regret Perspective on Online Multiple Testing

SafetyDGX agent

arXiv:2605.13916v1 Announce Type: cross Abstract: Online Multiple Testing (OMT), a fundamental pillar of sequential statistical inference, traditionally evaluates the False Discovery Rate (FDR) and st

A Security Analysis of the OpenClaw AI Agent Framework

SafetyDGX agent

arXiv:2603.27517v3 Announce Type: replace-cross Abstract: AI agent frameworks connecting large language model (LLM) reasoning to host execution surfaces -- shell, filesystem, containers, and messaging

Achieving Approximate Symmetry Is Exponentially Easier than Exact Symmetry

SafetyDGX agent

arXiv:2512.11855v2 Announce Type: replace-cross Abstract: Enforcing exact symmetry in machine learning models often yields significant gains in scientific applications, serving as a powerful inductive

Action-Conditioned Risk Gating for Safety-Critical Control under Partial Observability

SafetyDGX agent

arXiv:2605.14246v1 Announce Type: cross Abstract: Many safety-critical control problems are modeled as risk-sensitive partially observable Markov decision processes, where the controller must make dec

Active Learners as Efficient PRP Rerankers

SafetyDGX agent

arXiv:2605.14236v1 Announce Type: cross Abstract: Pairwise Ranking Prompting (PRP) elicits pairwise preference judgments from an LLM, which are then aggregated into a ranking, usually via classical so

ActivePusher: Active Learning and Planning with Residual Physics for Nonprehensile Manipulation

SafetyDGX agent

arXiv:2506.04646v4 Announce Type: replace-cross Abstract: Planning with learned dynamics models offers a promising approach toward versatile real-world manipulation, particularly in nonprehensile sett

Agentifying Patient Dynamics within LLMs through Interacting with Clinical World Model

SafetyDGX agent

arXiv:2605.14723v1 Announce Type: new Abstract: Sepsis management in the ICU requires sequential treatment decisions under rapidly evolving patient physiology. Although large language models (LLMs) en

AIS: Adaptive Importance Sampling for Quantized RL

SafetyDGX agent

arXiv:2605.13907v1 Announce Type: cross Abstract: Reinforcement learning (RL) for large language models (LLMs) is dominated by the cost of rollout generation, which has motivated the use of low-precis

Aligning Latent Geometry for Spherical Flow Matching in Image Generation

SafetyDGX agent

arXiv:2605.15193v1 Announce Type: new Abstract: Latent flow matching for image generation usually transports Gaussian noise to variational autoencoder latents along linear paths. Both endpoints, howev

AMiD: Knowledge Distillation for LLMs with alpha-mixture Assistant Distribution

SafetyDGX agent

arXiv:2510.15982v3 Announce Type: replace-cross Abstract: Autoregressive large language models (LLMs) have achieved remarkable improvement across many tasks but incur high computational and memory cos

Anti-Length Shift: Dynamic Outlier Truncation for Training Efficient Reasoning Models

SafetyDGX agent

arXiv:2601.03969v2 Announce Type: replace Abstract: Large reasoning models enhanced by reinforcement learning with verifiable rewards have achieved significant performance gains by extending their cha

AQKA: Active Quantum Kernel Acquisition Under a Shot Budget

SafetyDGX agent

arXiv:2605.14672v1 Announce Type: new Abstract: Estimating an N imes N quantum kernel from circuit fidelities requires Theta(N^2 S) measurement shots, the dominant bottleneck for deployment on near-te

ASH: Agents that Self-Hone via Embodied Learning

SafetyDGX agent

arXiv:2605.14211v1 Announce Type: new Abstract: Long-horizon embodied tasks remain a fundamental challenge in AI, as current methods rely on hand-engineered rewards or action-labeled demonstrations, n

AutoMoT: A Unified Vision-Language-Action Model with Asynchronous Mixture-of-Transformers for End-to-End Autonomous Driving

SafetyDGX agent

arXiv:2603.14851v3 Announce Type: replace Abstract: Integrating vision-language models (VLMs) into end-to-end (E2E) autonomous driving (AD) systems has shown promise in improving scene understanding.

Before the Body Moves: Learning Anticipatory Joint Intent for Language-Conditioned Humanoid Control

SafetyDGX agent

arXiv:2605.14417v1 Announce Type: cross Abstract: Natural language is an intuitive interface for humanoid robots, yet streaming whole-body control requires control representations that are executable

Behavioral Data-Driven Optimal Trajectory Generation for Rotary Cranes

SafetyDGX agent

arXiv:2605.14944v1 Announce Type: new Abstract: With the growth of the construction industry and the global shortage of skilled labor, the automation of crane control has become increasingly important

Bellman Value Decomposition for Task Logic in Safe Optimal Control

SafetyDGX agent

arXiv:2602.19532v2 Announce Type: replace Abstract: Real-world tasks involve nuanced combinations of goal and safety specifications. In high dimensions, the challenge is exacerbated: formal automata b

Beyond Instance-Level Self-Supervision in 3D Multi-Modal Medical Imaging

SafetyDGX agent

arXiv:2605.14654v1 Announce Type: new Abstract: Self-supervised pre-training methods in medical imaging typically treat each individual as an isolated instance, learning representations through augmen

Big move from arXiv: a one-year ban for authors who submit AI-generated content without proper checking. This is not about banning AI from a…

SafetyDGX agent

Big move from arXiv: a one-year ban for authors who submit AI-generated content without proper checking. This is not about banning AI from academia. AI can be extremely useful. It can help us write be

BiSpikCLM: A Spiking Language Model integrating Softmax-Free Spiking Attention and Spike-Aware Alignment Distillation

SafetyDGX agent

arXiv:2605.13859v1 Announce Type: cross Abstract: Spiking Neural Networks (SNNs) offer promising energy-efficient alternatives to large language models (LLMs) due to their event-driven nature and ultr

Boosting Reinforcement Learning with Verifiable Rewards via Randomly Selected Few-Shot Guidance

SafetyDGX agent

arXiv:2605.15012v1 Announce Type: cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has achieved great success in developing Large Language Models (LLMs) with chain-of-thought roll

Break-the-Beat! Controllable MIDI-to-Drum Audio Synthesis

SafetyDGX agent

arXiv:2605.14555v1 Announce Type: cross Abstract: Current methods for creating drum loop audio in digital music production, such as using one-shot samples or resampling, often demand non-trivial effor

CHASM: Cross-frequency Harmonized Axis-Separable Mixing for Spectral Token Operators

SafetyDGX agent

arXiv:2605.14727v1 Announce Type: new Abstract: Spectral token mixers based on Fourier transforms provide an efficient way to model global interactions in visual feature maps. Existing designs often e

COAL: Counterfactual and Observation-Enhanced Alignment Learning for Discriminative Referring Multi-Object Tracking

SafetyDGX agent

arXiv:2605.14795v1 Announce Type: new Abstract: Referring Multi-Object Tracking (RMOT) faces a fundamental structural contradiction between the high-discriminability demand and the sparse semantic sup

Collaborative Yet Personalized Policy Training: Single-Timescale Federated Actor-Critic

SafetyDGX agent

arXiv:2605.14423v1 Announce Type: cross Abstract: Despite the popularity of the actor-critic method and the practical needs of collaborative policy training, existing works typically either overlook e

company that steals IP urges US government not to allow others to steal their IP

SafetyDGX agent

company that steals IP urges US government not to allow others to steal their IP Anthropic drops a paper on the US-China AI race They believe the US and its allies may be able to lock in a 12-24 month

Comparing Developer and LLM Biases in Code Evaluation

SafetyDGX agent

arXiv:2603.24586v2 Announce Type: replace-cross Abstract: As LLMs are increasingly used as judges in code applications, they should be evaluated in realistic interactive settings that capture partial

Complacent, Not Sycophantic: Reframing Large Language Models and Designing AI Literacy for Complacent Machines

SafetyDGX agent

arXiv:2605.14544v1 Announce Type: new Abstract: Large language models are often described as sycophantic, in the sense that they appear to flatter users or mirror their beliefs. We argue that this lab

Compositional Sparsity as an Inductive Bias for Neural Architecture Design

SafetyDGX agent

arXiv:2605.14764v1 Announce Type: cross Abstract: Identifying the structural priors that enable Deep Neural Networks (DNNs) to overcome the curse of dimensionality is a fundamental challenge in machin

CrystalReasoner: Reasoning and RL for Property-Conditioned Crystal Structure Generation

SafetyDGX agent

arXiv:2605.14344v1 Announce Type: new Abstract: Generative modeling has emerged as a promising approach for crystal structure discovery. However, existing LLM-based generative models struggle with low

CuSearch: Curriculum Rollout Sampling via Search Depth for Agentic RAG

SafetyDGX agent

arXiv:2605.11611v2 Announce Type: replace Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has emerged as a promising paradigm for training agentic retrieval-augmented generation (RAG)

DAPL: Integration of Positive and Negative Descriptions in Text-Based Person Search

SafetyDGX agent

arXiv:2405.07459v3 Announce Type: replace Abstract: Text-based person search (TBPS) aims to retrieve specific images of individuals from large datasets using textual descriptions. Existing TBPS method

Deciphering Neural Reparameterized Full-Waveform Inversion with Neural Sensitivity Kernel and Wave Tangent Kernel

SafetyDGX agent

arXiv:2605.14370v1 Announce Type: cross Abstract: Full-waveform inversion (FWI) estimates unknown parameters in the wave equation from limited boundary measurements. Recent advances in neural reparame

Deepchecks: Evaluating Retrieval-Augmented Generation (RAG)

SafetyDGX agent

arXiv:2605.14488v1 Announce Type: new Abstract: Large Language Models (LLMs) augmented with Retrieval-Augmented Generation (RAG) techniques are revolutionizing applications across multiple domains, su

Delta Forcing: Trust Region Steering for Interactive Autoregressive Video Generation

SafetyDGX agent

arXiv:2605.14382v1 Announce Type: new Abstract: Interactive real-time autoregressive video generation is essential for applications such as content creation and world modeling, where visual content mu

Diagnosing Training Inference Mismatch in LLM Reinforcement Learning

SafetyDGX agent

arXiv:2605.14220v1 Announce Type: cross Abstract: Modern LLM RL systems separate rollout generation from policy optimization. These two stages are expected to produce token probabilities that match ex

did you know that Queen Elizabeth II wrote a Python graduate textbook?

SafetyDGX agent

did you know that Queen Elizabeth II wrote a Python graduate textbook? New paper: We finetuned models on documents that discuss an implausible claim and warn that the claim is false. Models ended up b

DiffusionOPD: A Unified Perspective of On-Policy Distillation in Diffusion Models

SafetyDGX agent

arXiv:2605.15055v1 Announce Type: cross Abstract: Reinforcement learning has emerged as a powerful tool for improving diffusion-based text-to-image models, but existing methods are largely limited to

Dimension-Level Intent Fidelity Evaluation for Large Language Models: Evidence from Structured Prompt Ablation

SafetyDGX agent

arXiv:2605.14517v1 Announce Type: cross Abstract: Holistic evaluation scores capture overall output quality but do not distinguish whether a model reproduced the structural form of a user's request fr

Distance-Matrix Wasserstein Statistics for Scalable Gromov--Wasserstein Learning

SafetyDGX agent

arXiv:2605.14981v1 Announce Type: new Abstract: Gromov--Wasserstein (GW) distances compare graphs, shapes, and point clouds through internal distances, without requiring a common coordinate system. Th

Distribution Corrected Offline Data Distillation for Large Language Models

SafetyDGX agent

arXiv:2605.14071v1 Announce Type: new Abstract: Distilling reasoning traces from strong large language models into smaller ones is a promising route to improve intelligence in resource-constrained set

Distributions as Actions: A Unified Framework for Diverse Action Spaces

SafetyDGX agent

arXiv:2506.16608v3 Announce Type: replace-cross Abstract: We introduce a novel reinforcement learning (RL) framework that treats parameterized action distributions as actions, redefining the boundary

DIVER: Reinforced Diffusion Breaks Imitation Bottlenecks in End-to-End Autonomous Driving

SafetyDGX agent

arXiv:2507.04049v4 Announce Type: replace Abstract: Most end-to-end autonomous driving methods rely on imitation learning from single expert demonstrations, often leading to conservative and homogeneo

Do Language Models Align with Brains? Prediction Scores Are Not Enough

SafetyDGX agent

arXiv:2605.14025v1 Announce Type: cross Abstract: Brain-language model comparisons often interpret neural prediction scores as evidence that model representations capture brain-relevant language compu

Do Reasoning LLMs Refuse What They Infer in Long Contexts?

SafetyDGX agent

arXiv:2602.08874v2 Announce Type: replace Abstract: Long-context LLMs can infer objectives that are not stated explicitly. This capability is useful for reasoning over documents, code, retrieved evide

DSSP: Diffusion State Space Policy with Full-History Encoding

SafetyDGX agent

arXiv:2605.14598v1 Announce Type: new Abstract: Diffusion-based imitation learning has shown strong promise for robot manipulation. However, most existing policies condition only on the current observ

Dual Hierarchical Dialogue Policy Learning for Legal Inquisitive Conversational Agents

SafetyDGX agent

arXiv:2605.14057v1 Announce Type: new Abstract: Most existing dialogue systems are user-driven, primarily designed to fulfill user requests. However, in many critical real-world scenarios, a conversat

Dynamic Mixed-Precision Routing for Efficient Multi-step LLM Interaction

SafetyDGX agent

arXiv:2602.02711v2 Announce Type: replace Abstract: Large language models (LLMs) achieve strong performance in long-horizon decision-making tasks through multi-step interaction and reasoning at test t

Efficient Generative Retrieval for E-commerce Search with Semantic Cluster IDs and Expert-Guided RL

SafetyDGX agent

arXiv:2605.14434v1 Announce Type: cross Abstract: Generative retrieval offers a promising alternative by unifying the fragmented multi-stage retrieval process into a single end-to-end model. However,

EponaV2: Driving World Model with Comprehensive Future Reasoning

SafetyDGX agent

arXiv:2605.14696v1 Announce Type: new Abstract: Data scaling plays a pivotal role in the pursuit of general intelligence. However, the prevailing perception-planning paradigm in autonomous driving rel

Every Subtlety Counts: Fine-grained Person Independence Micro-Action Recognition via Distributionally Robust Optimization

SafetyDGX agent

arXiv:2509.21261v3 Announce Type: replace Abstract: Micro-action Recognition is vital for psychological assessment and human-computer interaction. However, existing methods often fail in real-world sc

Evo-Depth: A Lightweight Depth-Enhanced Vision-Language-Action Model

SafetyDGX agent

arXiv:2605.14950v1 Announce Type: new Abstract: Vision-Language-Action models have emerged as a promising paradigm for robotic manipulation by unifying perception, language grounding, and action gener

Evolving Layer-Specific Scalar Functions for Hardware-Aware Transformer Adaptation

SafetyDGX agent

arXiv:2605.14047v1 Announce Type: new Abstract: Vision Transformers (ViTs) achieve state-of-the-art performance on challenging vision tasks, but their deployment on edge devices is severely hindered b

Exploring Geographic Relative Space in Large Language Models through Activation Patching

SafetyDGX agent

arXiv:2605.14535v1 Announce Type: new Abstract: The increased use of Large Language Models (LLMs) in geography raises substantial questions about the safety of integrating these tools across a wide ra

Fair and Calibrated Toxicity Detection with Robust Training and Abstention

SafetyDGX agent

arXiv:2605.14074v1 Announce Type: new Abstract: Fairness in toxicity classification involves three integrated axes: ranking, calibration, and abstention. Training-time interventions and post-hoc safet

FlowSteer: Towards Agents Designing Agentic Workflows via Reinforced Progressive Canvas Editing

SafetyDGX agent

arXiv:2602.01664v4 Announce Type: replace Abstract: In recent years, agentic workflows have been widely applied to solve complex human tasks. However, existing workflow construction still faces key ch

← Previous
1…139140141142143…214
Next →