AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
All
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,816 results
Safety

From Pixels to Words -- Towards Native One-Vision Models at Scale

DGX agent

arXiv:2605.28820v1 Announce Type: new Abstract: Current vision-language models (VLMs) typically stitch together separate image encoders and language decoders via multi-stage alignment, a modular frame

safetyarxiv-cs-cv
28 May 2026
Safety

fun thread on consciousness with @Grimezsz:

Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

Gary Marcus shares a discussion thread about consciousness with user @Grimezsz on X (formerly Twitter), likely exploring philosophical, scientific, or technical perspectives on consciousness and relat

safetygary-marcus--x
28 May 2026
Safety

GE-Sim 2.0: A Roadmap Towards Comprehensive Closed-loop Video World Simulators for Robotic Manipulation

DGX agent

arXiv:2605.27491v1 Announce Type: new Abstract: We introduce GE-Sim 2.0 (Genie Envisioner World Simulator 2.0), a closed-loop video world simulator for robotic manipulation. Building on the action-con

safetyarxiv-cs-ro
28 May 2026
Safety

GeneralThinker: Domain-General Reasoning through Likelihood-Guided Answer-Conditioned Optimization

DGX agent

arXiv:2605.27934v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards improves language model reasoning, but its reliance on domain-specific verifiers, sparse outcome rewards,

safetyarxiv-cs-cl
28 May 2026
Safety

Geometry of Human Perceptual Domains Emerges Transiently in LLM Representations

DGX agent

arXiv:2605.27970v1 Announce Type: new Abstract: While large language models (LLMs) are trained purely on textual data, prior work has shown that their internal representations can exhibit rich geometr

safetyarxiv-cs-ai
28 May 2026
Safety

Geometry of Relaxed Fair Regression: A Unified Framework for Aware and Unaware Settings

DGX agent

arXiv:2605.28233v1 Announce Type: cross Abstract: Fairness-accuracy trade-offs are a central concern in the deployment of fairness-aware machine learning methods. When sensitive attributes are unavail

safetyarxiv-cs-lg
28 May 2026
Safety

Got a Secret? LLM Agents Can't Keep It: Evaluating Privacy in Multi-Agent Systems

DGX agent

arXiv:2605.27766v1 Announce Type: new Abstract: LLM safety evaluations predominantly test models in isolation, yet deployed AI agents increasingly operate within persistent social environments alongsi

safetyarxiv-cs-ai
28 May 2026
Safety

Grimlock: Guarding High-Agency Systems with eBPF and Attested Channels

DGX agent

arXiv:2605.27488v1 Announce Type: cross Abstract: Agentic systems increasingly run user-authored orchestration code that invokes tools, spawns subtasks, and delegates work across machines and clouds.

safetyarxiv-cs-ai
28 May 2026
Safety

Grounded Cache Routing for Retrieval-Augmented Generation: When Is It Safe to Reuse an Answer?

DGX agent

arXiv:2605.27494v1 Announce Type: cross Abstract: Modern retrieval-augmented generation(RAG) deployments increasingly rely on caching to reduce token cost and time-to-first-token(TTFT). Prefix-level K

safetyarxiv-cs-ai
28 May 2026
Safety

GS-FUSE: Granger-Supervised Gated Fusion and Multi-Granularity Alignment for Event-Driven Financial Forecasting

DGX agent

arXiv:2605.28520v1 Announce Type: new Abstract: Accurately forecasting the impact of salient financial events on markets is critical for investors and policymakers. However, existing multimodal time-s

safetyarxiv-cs-ai
28 May 2026
Safety

Guaranteed Optimal Compositional Explanations for Neurons

DGX agent

arXiv:2511.20934v2 Announce Type: replace Abstract: Compositional explanations are a family of methods that aim to describe the spatial alignment between neurons' receptive field activations and conce

safetyarxiv-cs-ai
28 May 2026
Safety

Heterogeneous Causal Discovery of Repeated Undesirable Health Outcomes

DGX agent

arXiv:2503.11477v2 Announce Type: replace Abstract: Understanding the factors that trigger or prevent undesirable health outcomes across patient subpopulations is essential for designing targeted inte

safetyarxiv-cs-ai
28 May 2026
Safety

Hierarchical Synthetic Tabular Data Generation: A Hybrid Top-Down and Bottom-Up Framework

DGX agent

arXiv:2605.28198v1 Announce Type: new Abstract: Existing approaches for synthetic tabular data generation are based on either purely generative models or LLMs, both of which struggle with data heterog

safetyarxiv-cs-lg
28 May 2026
Safety

High-Fidelity Industrial Crash Dynamics Prediction via Geometry-Aware Operator Learning with Memory-Efficient Low-Rank Attention

DGX agent

arXiv:2605.27758v1 Announce Type: cross Abstract: Automotive crashworthiness optimization remains a safety-critical challenge, requiring the management of large-scale nonlinear structural deformations

safetyarxiv-cs-ai
28 May 2026
Safety

HiRQA: Hierarchical Ranking and Quality Alignment for Opinion-Unaware Image Quality Assessment

DGX agent

arXiv:2508.15130v2 Announce Type: replace Abstract: Despite significant progress in no-reference image quality assessment (NR-IQA), dataset biases and reliance on subjective labels continue to hinder

safetyarxiv-cs-cv
28 May 2026
Safety

Holy shit! They changed the rules for Elon again... They waved the profitability rule & are adding SpaceX to indices only 5 days after IPO..…

DGX agent

Holy shit! They changed the rules for Elon again... They waved the profitability rule & are adding SpaceX to indices only 5 days after IPO... normally it's 90 This forces 401k retirement & passive fun

safetygary-marcus--x
28 May 2026
Safety

How the Optimizer Shapes Learned Solutions in Equivariant Neural Networks

DGX agent

arXiv:2605.27662v1 Announce Type: cross Abstract: Equivariant neural networks encode geometric symmetries by construction, yet they are often difficult to optimize and can underperform less constraine

safetyarxiv-cs-ai
28 May 2026
Safety

How VLAs Fail Differently: Black-Box Action Monitoring Reveals Architecture-Specific Failure Signatures

DGX agent

arXiv:2605.28726v1 Announce Type: cross Abstract: We discover that VLA architectures fail in fundamentally different, predictable ways at the motor-command level. Running VQ-BeT, Diffusion Policy, and

safetyarxiv-cs-lg
28 May 2026
Safety

Human-like in-group bias in instruction-tuned language model agents

DGX agent

arXiv:2605.28114v1 Announce Type: new Abstract: As autonomous AI agents are deployed in persistent, interacting networks -- coordinating tasks, routing resources, and accumulating reputational histori

safetyarxiv-cs-ai
28 May 2026
Safety

I read this garbage (in a big UK newspaper) when Google can’t even count reliably and wonder why people don’t spend more time learning about…

DGX agent

I read this garbage (in a big UK newspaper) when Google can’t even count reliably and wonder why people don’t spend more time learning about AI’s actual strengths and weakness before running their mou

safetygary-marcus--x
28 May 2026
Safety

ICAN-Deploy: Identity-Stable Canary Deployment for Safety-Critical Embodied Agents

DGX agent

arXiv:2605.28097v1 Announce Type: new Abstract: Canary deployment routes a fraction of traffic to a new software version, monitors metrics, and rolls back on regression. Mainstream controllers (Argo R

safetyarxiv-cs-ro
28 May 2026
Safety

ICG: Improving Cover Image Generation via MLLM-based Prompting and Personalized Preference Alignment

DGX agent

arXiv:2605.27374v1 Announce Type: new Abstract: Recent advances in multimodal large language models (MLLMs) and diffusion models (DMs) have opened new possibilities for AI-generated content. Yet, pers

safetyarxiv-cs-cl
28 May 2026
Safety

Identifying and Mitigating Bottlenecks in Role-Playing Agents: A Systematic Study of Disentangling Character Profile Axes

DGX agent

arXiv:2601.04716v3 Announce Type: replace Abstract: While Large Language Model (LLM) role-playing agents have advanced rapidly, it remains unclear which profile elements genuinely drive role-playing q

safetyarxiv-cs-cl
28 May 2026
Safety

Imitating and Finetuning Model Predictive Control for Robust and Symmetric Quadrupedal Locomotion

DGX agent

arXiv:2311.02304v3 Announce Type: replace Abstract: Control of legged robots is a challenging problem that has been investigated by different approaches, such as model-based control and learning algor

safetyarxiv-cs-ro
28 May 2026
Safety

Important context for latest OpenAI announcement. Especially (5:30): 'The model spit out a long transcript of an answer. Then a team of expe…

DGX agent

Important context for latest OpenAI announcement. Especially (5:30): 'The model spit out a long transcript of an answer. Then a team of expert mathematicians poured over this [transcript] and identifi

safetygary-marcus--x
28 May 2026
Safety

IMU Propagation as Preintegration

DGX agent

arXiv:2605.28279v1 Announce Type: new Abstract: IMU preintegration is widely used in factor-graph-based visual--inertial, lidar--inertial, and radar--inertial state estimation, yet it is often treated

safetyarxiv-cs-ro
28 May 2026
Safety

Information-theoretic Multimodal Representation Learning for Electrocardiogram Signals

DGX agent

arXiv:2605.27583v1 Announce Type: new Abstract: Electrocardiograms (ECGs) are widely used non-invasive measurements of cardiac activity and play a central role in clinical diagnosis. Recent multimodal

safetyarxiv-cs-lg
28 May 2026
Safety

Informing AI Policy Assessment using Large-Scale Simulation of Interventions

DGX agent

arXiv:2605.27395v1 Announce Type: cross Abstract: As the rapid proliferation of AI systems and harms spurs efforts in AI governance around the world, prioritizing among competing policy options has be

safetyarxiv-cs-ai
28 May 2026
Safety

Insurance Pricing Optimization via Off-Policy Evaluation

DGX agent

arXiv:2605.28327v1 Announce Type: cross Abstract: Traditional insurance pricing relies on risk-based principles that ensure actuarial fairness and solvency but do not explicitly account for policyhold

safetyarxiv-cs-lg
28 May 2026
Safety

Intelligence as Managed Autonomy: Failure, Escalation, and Governance for Agentic AI Systems

DGX agent

arXiv:2605.27628v1 Announce Type: new Abstract: As autonomous and agentic AI systems scale in robotic and human-machine environments, managing hallucination and persistent but unjustified action remai

safetyarxiv-cs-ai
28 May 2026
Safety

JECA^2: Judgment-Explanation Consistent Adversarial Attack against Forensic Vision-Language Models

DGX agent

arXiv:2605.28609v1 Announce Type: new Abstract: Forensic vision-language models (VLMs) have recently been developed to detect image tampering and provide natural-language explanations. However, their

safetyarxiv-cs-cv
28 May 2026
Safety

Joint Training of Multi-Token Prediction in Reinforcement Learning via Optimal Coefficient Calibration

DGX agent

arXiv:2605.28184v1 Announce Type: new Abstract: Reinforcement Learning from Verifiable Rewards (RLVR) has emerged as the standard paradigm for improving reasoning capability of large language models,

safetyarxiv-cs-lg
28 May 2026
Safety

just imagine what will happen to the economy and people’s retirement funds if these projections from FT are correct. brace for bailouts.

DGX agent

just imagine what will happen to the economy and people’s retirement funds if these projections from FT are correct. brace for bailouts. The AI numbers are starting to look very ugly. Even under 'best

safetygary-marcus--x
28 May 2026
Safety

LACUNA: Safe Agents as Recursive Program Holes

DGX agent

arXiv:2605.28617v1 Announce Type: new Abstract: LLM agents increasingly act by writing code, yet a split persists between the runtime that drives the agent and the code the model writes. The runtime o

safetyarxiv-cs-ai
28 May 2026
Safety

Learning Deliberately, Acting Intuitively: Unlocking Test-Time Reasoning in Multimodal LLMs

DGX agent

arXiv:2507.06999v2 Announce Type: replace-cross Abstract: Reasoning is essential for large language models (LLMs), especially in complex tasks such as mathematical problem solving. However, multimodal

safetyarxiv-cs-cl
28 May 2026
Safety

Learning High-Dimensional Parity Functions with Product Networks using Gradient Descent

DGX agent

arXiv:2605.28612v1 Announce Type: new Abstract: Parity functions are fundamental Boolean operations with critical applications across machine learning, cryptography, and error correction. Yet, learnin

safetyarxiv-cs-lg
28 May 2026
Safety

Learning to Assign Prediction Tasks to Agents with Capacity Constraints

DGX agent

arXiv:2605.27999v1 Announce Type: cross Abstract: We address the problem of learning to assign prediction tasks to one agent from a set of available human or AI agents. In particular, we focus on the

safetyarxiv-cs-ai
28 May 2026
Safety

Learning to Bid in Repeated Second-Price Auctions with Dynamic Values and Aggregated Feedback

DGX agent

arXiv:2605.28133v1 Announce Type: new Abstract: We study the problem of learning to bid when the bidder's value is dynamic, i.e., when the current value depends on past outcomes. Specifically, we cons

safetyarxiv-cs-lg
28 May 2026
Safety

Learning with Importance Weighted Variational Inference

DGX agent

arXiv:2410.12035v2 Announce Type: replace-cross Abstract: Several variational bounds involving importance weighting ideas generalize the Evidence Lower BOund (ELBO) for marginal likelihood optimizatio

safetyarxiv-cs-lg
28 May 2026
Safety

Let Relations Speak: An End-to-End LLM-GNN Soft Prompt Framework for Fraud Detection

DGX agent

arXiv:2605.28524v1 Announce Type: new Abstract: In recent years, Large Language Models (LLMs) have shown great capability in processing graph tasks such as fraud detection. However, most existing meth

safetyarxiv-cs-ai
28 May 2026
Safety

LLM Watermark Evasion via Bias Inversion

DGX agent

arXiv:2509.23019v5 Announce Type: replace-cross Abstract: Watermarking offers a promising solution for detecting LLM-generated content, yet its robustness under realistic query-free (black-box) evasio

safetyarxiv-cs-ai
28 May 2026
Safety

Long Live The Balance: Information Bottleneck Driven Tree-based Policy Optimization

DGX agent

arXiv:2605.28109v1 Announce Type: new Abstract: Recent advances in online reinforcement learning (RL) for large language models (LLMs) have demonstrated promising performance in complex reasoning task

safetyarxiv-cs-lg
28 May 2026
Safety

Mag-VLA: Vision-Language-Action Model for Bimanual Magnetically Actuated Microrobot Manipulation

DGX agent

arXiv:2605.28486v1 Announce Type: new Abstract: Magnetically actuated microrobots have been used as wireless, non-contact manipulation tools at microscales, making them promising for minimally invasiv

safetyarxiv-cs-ro
28 May 2026
Safety

Mathematical Modelling of Ethical AI Use in Higher Education: A Coordination Game Framework for Future-Facing Learning

DGX agent

arXiv:2605.27400v1 Announce Type: cross Abstract: The rapid uptake of generative artificial intelligence (AI) in higher education is reshaping assessment practices and intensifying concerns around aca

safetyarxiv-cs-ai
28 May 2026
Safety

Mining Multi-Modality Spatio-Temporal Cues for Video Important Person Identification

DGX agent

arXiv:2605.28604v1 Announce Type: cross Abstract: Identifying key individuals in video scenes is essential for applications such as automated video editing and intelligent surveillance. Current method

safetyarxiv-cs-ai
28 May 2026
Safety

Mitigating Adaptive Attacks against Reasoning Models with Activation Consistency Training

DGX agent

arXiv:2605.28467v1 Announce Type: new Abstract: As LLMs gain stronger reasoning capabilities, their extended chain-of-thought introduces new degrees of complexity for defending against adversarial jai

safetyarxiv-cs-lg
28 May 2026
Safety

Mitigating Cross-Lingual Cultural Inconsistencies in LLMs via Consensus-Driven Preference Optimisation

DGX agent

arXiv:2605.12515v2 Announce Type: replace Abstract: Despite their impressive capabilities, multilingual large language models (MLLMs) frequently exhibit inconsistent behaviour when the prompt's langua

safetyarxiv-cs-cl
28 May 2026
Safety

Mobile-Aptus: Confidence-Driven Proactive and Robust Interaction in MLLM-based Mobile-Using Agents

DGX agent

arXiv:2605.28629v1 Announce Type: new Abstract: Recent advancements in multimodal large language models (MLLMs) have shown exceptional potential in enabling mobile-using agents to autonomously execute

safetyarxiv-cs-cl
28 May 2026
← Previous
1…137138139140141…267
Next →