AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,435 results
Safety

GenPO++: Generative Policy Optimization with Jacobian-free Likelihood Ratios

DGX agent

arXiv:2606.06967v1 Announce Type: new Abstract: Generative policies provide expressive and multimodal action distributions, making them attractive for reinforcement learning (RL) in complex continuous

safetyarxiv-cs-lg
8 Jun 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

GRASP: Geometry-aware Residual Alignment for Scalable Pretraining Data Attribution

DGX agent

arXiv:2606.06892v1 Announce Type: new Abstract: Scalable data attribution methods typically assign isolated utility scores to individual training examples. This prevalent additive assumption fundament

safetyarxiv-cs-lg
8 Jun 2026
Safety

Heterogeneous Effects of Green Finance on Urban Decarbonization: Evidence from 285 Cities in China

DGX agent

arXiv:2606.06986v1 Announce Type: new Abstract: While green finance has become a key instrument for low-carbon city transitions, its actual decarbonization effects and transmission mechanisms remain u

safetyarxiv-cs-lg
8 Jun 2026
Safety

How reliable are LLMs when it comes to playing dice?

DGX agent

arXiv:2606.07515v1 Announce Type: cross Abstract: We investigate the probabilistic reasoning capabilities of large language models through a controlled benchmarking study on discrete probability probl

safetyarxiv-cs-ai
8 Jun 2026
Safety

Interpreting Brain Responses to Language with Sparse Features from Language Models

DGX agent

arXiv:2606.06857v1 Announce Type: new Abstract: A central goal of cognitive neuroscience is to characterize the features that are represented by human language cortex. Artificial language models (LMs)

safetyarxiv-cs-cl
8 Jun 2026
Safety

Interpreting Learning Under Competing Models: Joint and Stepwise Approaches for Dynamic Cognitive Diagnosis

DGX agent

arXiv:2606.06804v1 Announce Type: new Abstract: Digital learning environments record learners' responses to individual items, making it possible to study the development of specific skills rather than

safetyarxiv-cs-lg
8 Jun 2026
Safety

Just-In-Time Reinforcement Learning: Continual Learning in LLM Agents Without Gradient Updates

DGX agent

arXiv:2601.18510v2 Announce Type: replace-cross Abstract: While Large Language Model (LLM) agents excel at general tasks, they inherently struggle with continual adaptation due to the frozen weights a

safetyarxiv-cs-ai
8 Jun 2026
Safety

Korean Culture into LLM Alignment: Toward Cultural Coherence

DGX agent

arXiv:2606.06797v1 Announce Type: new Abstract: Cultural-aspect work on large language models is dominated by a negative target: which outputs to suppress. We argue that a constructive counterpart is

safetyarxiv-cs-cl
8 Jun 2026
Safety

LARA: Latent Action Representation Alignment for Vision-Language-Action Models

DGX agent

arXiv:2606.07100v1 Announce Type: new Abstract: Visual-language action (VLA) models enable robots to predict actions directly from observations and language instructions, but their performance depends

safetyarxiv-cs-cv
8 Jun 2026
Safety

Learning All-Terrain Locomotion for a Planetary Rover with Actively Articulated Suspension

DGX agent

arXiv:2606.06790v1 Announce Type: cross Abstract: This paper presents ERNEST, a four-wheeled planetary rover concept equipped with a two-degree-of-freedom Active Gimbal Suspension that combines yaw an

safetyarxiv-cs-lg
8 Jun 2026
Safety

Learning Fair Demand Models

DGX agent

arXiv:2606.06830v1 Announce Type: cross Abstract: Data-driven pricing is increasingly prevalent in sectors such as airlines, lending, insurance, and retail. By learning demand models from customer fea

safetyarxiv-cs-lg
8 Jun 2026
Safety

LLM-Augmented Digital Twin for Policy Evaluation in Short-Video Platforms

DGX agent

arXiv:2603.11333v2 Announce Type: replace Abstract: Short-video platforms are closed-loop, human-in-the-loop ecosystems where platform policy, creator incentives, and user behavior co-evolve. This fee

safetyarxiv-cs-ai
8 Jun 2026
Safety

MADRAG: Multi-Agent Debate with Retrieval-Augmented Generation for Training-Free Analytic Essay Scoring

DGX agent

arXiv:2606.06754v1 Announce Type: cross Abstract: We present MADRAG, a training-free framework for analytic essay scoring that combines multi-agent reasoning with retrieval-augmented grounding. Unlike

safetyarxiv-cs-cl
8 Jun 2026
Safety

Mining Useful General Data for Low-Resource Domain Adaptation

DGX agent

arXiv:2511.07380v2 Announce Type: replace Abstract: Adapting large language models (LLMs) to low-resource domains remains challenging due to the scarcity of domain-specific data. While in-domain data

safetyarxiv-cs-cl
8 Jun 2026
Safety

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion

DGX agent

arXiv:2509.17446v3 Announce Type: replace-cross Abstract: Multimodal intent recognition (MMIR) suffers from weak semantic grounding and poor robustness under noisy or rare-class conditions. We propose

safetyarxiv-cs-ai
8 Jun 2026
Safety

Native3D: End-to-End 3D Scene Generation via Unified Mesh-Texture Modeling and Semantic Alignment

DGX agent

arXiv:2606.07117v1 Announce Type: cross Abstract: This paper presents Native3D, the first end-to-end 3D scene generation framework that completely bypasses 2D intermediate representations. Traditional

safetyarxiv-cs-ai
8 Jun 2026
Safety

Network Recovery from Cascade Data: A Debiased Jacobian-Based Machine Learning Approach

DGX agent

arXiv:2606.07483v1 Announce Type: new Abstract: Many important outcomes unfold as dynamic cascades, including product adoption, disease spread, financial distress, and information diffusion. A central

safetyarxiv-cs-lg
8 Jun 2026
Safety

Neuro-Symbolic Learning for Long-Horizon Task Planning Under Complex Logical Constraints

DGX agent

arXiv:2606.06877v1 Announce Type: cross Abstract: Task planning often suffers from severe efficiency bottlenecks when robots must reason over long-horizon action sequences under complex logical constr

safetyarxiv-cs-ai
8 Jun 2026
Safety

Online Pandora's Box for Contextual LLM Cascading

DGX agent

arXiv:2606.07392v1 Announce Type: new Abstract: Motivated by Large Language Model (LLM) cascading, we propose an online contextual Pandora's Box model for adaptively querying and selecting LLM APIs. I

safetyarxiv-cs-ai
8 Jun 2026
Safety

Planning-aligned Token Compression for Long-Context Autonomous Driving

DGX agent

arXiv:2606.07464v1 Announce Type: cross Abstract: Monolithic vision-action models represent an emerging paradigm in autonomous driving. However, this architecture produces token sequences that quickly

safetyarxiv-cs-ai
8 Jun 2026
Safety

Predictive Statistics Shape Emergent World Representations of Grid Walkers

DGX agent

arXiv:2603.16689v2 Announce Type: replace Abstract: Next-token predictors often appear to develop internal representations of the latent world and its rules. The probabilistic nature of these models s

safetyarxiv-cs-lg
8 Jun 2026
Safety

Progress-SQL: Improving Reinforcement Learning for Text-to-SQL via Progressive Rewards

DGX agent

arXiv:2606.06825v1 Announce Type: cross Abstract: Reinforcement learning has recently shown promise in improving large language models for Text-to-SQL generation, yet existing methods typically optimi

safetyarxiv-cs-ai
8 Jun 2026
Safety

QuadVerse: An Integrated Framework Aligning Visual-Physical Reality for Quadruped Simulation

DGX agent

arXiv:2606.07118v1 Announce Type: new Abstract: Simulation is central to robot learning, yet the sim-to-real gap remains a major bottleneck.Existing approaches often tackle visual or dynamic gaps sepa

safetyarxiv-cs-ro
8 Jun 2026
Safety

Rapid co-design of Buoyancy-assisted robots for Challenging Locomotion using Gaussian Evolutionary Specialists

DGX agent

arXiv:2606.07424v1 Announce Type: new Abstract: Designing high-performance legged robots requires jointly optimizing morphology and control. Model-free Reinforcement Learning (RL) offers an alternativ

safetyarxiv-cs-ro
8 Jun 2026
Safety

RASFT: Rollout-Adaptive Supervised Fine-Tuning for Reasoning

DGX agent

arXiv:2606.07006v1 Announce Type: cross Abstract: Supervised fine-tuning (SFT) is a prevailing method for adapting large language models to reasoning tasks by imitating offline expert demonstrations,

safetyarxiv-cs-cl
8 Jun 2026
Safety

RAVEN: Retrieval-Augmented Vulnerability Exploration Network for Memory Corruption Analysis in User Code and Binary Programs

DGX agent

arXiv:2604.17948v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities across various cybersecurity tasks, including vulnerability classificat

safetyarxiv-cs-ai
8 Jun 2026
Safety

Robotic Policy Adaptation via Weight-Space Meta-Learning

DGX agent

arXiv:2606.07217v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models are emerging as a promising paradigm for robotic manipulation, enabling general-purpose policies trained from larg

safetyarxiv-cs-cv
8 Jun 2026
Safety

Robots Need More than VLA and World Models

DGX agent

arXiv:2606.06556v1 Announce Type: new Abstract: Generalist robot intelligence is often framed as a policy-scaling problem: collect more robot demonstrations, train larger Vision-Language-Action (VLA)

safetyarxiv-cs-ro
8 Jun 2026
Safety

Self-evolving LLM agents with in-distribution Optimization

DGX agent

arXiv:2606.07367v1 Announce Type: new Abstract: Large Language Models (LLMs) have recently emerged as powerful controllers for interactive agents in complex environments, yet training them to perform

safetyarxiv-cs-lg
8 Jun 2026
Safety

Semantic-Structural Alignment for Generative Pictorial Charts

DGX agent

arXiv:2606.06498v1 Announce Type: cross Abstract: Traditional statistical graphics are precise but often lack the visual appeal, memorability, and engagement of pictorial charts. We present a generati

safetyarxiv-cs-cv
8 Jun 2026
Safety

SERNF: Sample-Efficient Real-World Dexterous Policy Fine-Tuning via Action-Chunked Critics and Normalizing Flows

DGX agent

arXiv:2602.09580v4 Announce Type: replace-cross Abstract: Real-world fine-tuning of dexterous manipulation policies remains challenging due to limited real-world interaction budgets and highly multimo

safetyarxiv-cs-lg
8 Jun 2026
Safety

Simulation-Driven Imitation Learning for Biosignals-Free Shared-Autonomy Prosthetic Grasping

DGX agent

arXiv:2606.07389v1 Announce Type: new Abstract: Biosignals-free shared-autonomy control of upper-limb prosthetic hands aims to enable natural and low-effort manipulation without relying on EMG or othe

safetyarxiv-cs-ro
8 Jun 2026
Safety

SlimSearcher: Training Efficiency-Aware Web Agents via Adaptive Reward Gating

DGX agent

arXiv:2606.07074v1 Announce Type: cross Abstract: Deep research agents have demonstrated remarkable capabilities in complex information-seeking tasks, yet this power comes at a steep computational cos

safetyarxiv-cs-ai
8 Jun 2026
Safety

Socratic-SWE: Self-Evolving Coding Agents via Trace-Derived Agent Skills

DGX agent

arXiv:2606.07412v1 Announce Type: cross Abstract: LLM-driven software engineering agents have become a central testbed for real-world language-model capability, yet their training remains limited by t

safetyarxiv-cs-ai
8 Jun 2026
Safety

Stable Reasoning, Unstable Responses: Mitigating LLM Deception via Stability Asymmetry

DGX agent

arXiv:2603.26846v2 Announce Type: replace-cross Abstract: As Large Language Models (LLMs) expand in capability and application scope, their trustworthiness becomes critical. A vital risk is intrinsic

safetyarxiv-cs-ai
8 Jun 2026
Safety

Step-Wise Refusal Dynamics in Autoregressive and Diffusion Language Models

DGX agent

arXiv:2602.02600v3 Announce Type: replace-cross Abstract: Diffusion language models (DLMs) have recently emerged as a competitive alternative to autoregressive (AR) models, offering parallel decoding,

safetyarxiv-cs-ai
8 Jun 2026
Safety

SV-Detect: AI-generated Text Detection with Steering Vectors

DGX agent

arXiv:2606.07313v1 Announce Type: cross Abstract: Detecting machine-generated text is especially difficult under distribution shift, such as transfer across domains, source models, and editing attacks

safetyarxiv-cs-ai
8 Jun 2026
Safety

Sycophantic Praise: Evaluating Excessive Praise in Language Models

DGX agent

arXiv:2606.07441v1 Announce Type: new Abstract: Sycophancy in language models is typically studied as excessive agreement or validation, while explicit praise and flattery have received comparatively

safetyarxiv-cs-cl
8 Jun 2026
Safety

T-GMP: Terrain-conditioned Generative Motion Priors for Versatile and Natural Humanoid Locomotion

DGX agent

arXiv:2606.06944v1 Announce Type: new Abstract: Achieving both anthropomorphic naturalness and robust terrain traversal remains a fundamental challenge in humanoid locomotion. Existing Reinforcement L

safetyarxiv-cs-ro
8 Jun 2026
Safety

Teach a Reward Model to Correct Itself: Reward Guided Adversarial Failure Discovery for Robust Reward Modeling

DGX agent

arXiv:2507.06419v3 Announce Type: replace Abstract: Reward modeling (RM), which captures human preferences to align large language models (LLMs), is increasingly employed in tasks such as model finetu

safetyarxiv-cs-cl
8 Jun 2026
Safety

Teaching the Way, Not the Answer: Privileged Tutoring Distillation for Multimodal Policy Optimization

DGX agent

arXiv:2606.07000v1 Announce Type: new Abstract: Recent post-training methods, particularly Reinforcement Learning with Verifiable Rewards (RLVR), have significantly enhanced the reasoning ability of L

safetyarxiv-cs-ai
8 Jun 2026
Safety

The discovery of the effects of women employment participation on the fertility of developing countries: A panel data approach

DGX agent

arXiv:2606.07093v1 Announce Type: new Abstract: The fertility trend in developing countries has experienced a significant decline in the last few decades; at the same time, the role of women in the wo

safetyarxiv-cs-lg
8 Jun 2026
Safety

Translate-R1: Cost-Aware Translation Tool Use via Reinforcement Learning

DGX agent

arXiv:2606.06835v1 Announce Type: new Abstract: The performance gap across languages in LLMs is well documented, and closing it natively requires pretraining or fine-tuning on corpora that, for most l

safetyarxiv-cs-cl
8 Jun 2026
Safety

TrioPose: Native Triple-Stream Diffusion Transformers for Pose-Guided Text-to-Image Generation

DGX agent

arXiv:2606.07053v1 Announce Type: new Abstract: Pose-guided text-to-image generation often suffers from limb distortions and feature crosstalk in complex multi-person scenarios. While existing UNet-ba

safetyarxiv-cs-cv
8 Jun 2026
Safety

VALUEFLOW: Toward Pluralistic and Steerable Value-based Alignment in Large Language Models

DGX agent

arXiv:2602.03160v2 Announce Type: replace Abstract: Aligning Large Language Models (LLMs) with the diverse spectrum of human values remains a central challenge: preference-based methods often fail to

safetyarxiv-cs-ai
8 Jun 2026
Safety

Watch, Remember, Reason: Human-View Video Understanding with MLLMs

DGX agent

arXiv:2606.07433v1 Announce Type: cross Abstract: Video understanding is being rapidly transformed by multimodal large language models (MLLMs), as research moves from short clips to long, multimodal,

safetyarxiv-cs-ai
8 Jun 2026
Safety

What Do People Actually Want From AI? Mapping Preference Plurality

DGX agent

arXiv:2606.06674v1 Announce Type: new Abstract: Large Language Models (LLMs) are often fine-tuned through Reinforcement Learning from Human Feedback (RLHF) to align with people's preferences and value

safetyarxiv-cs-cl
8 Jun 2026
Safety

What Is My Robot Thinking? Design Considerations for Transparent and Trustworthy Shared Autonomy

DGX agent

arXiv:2606.06870v1 Announce Type: new Abstract: Assistive robots operating under shared autonomy must balance user control with autonomous assistance. Because robot actions depend on internal intent i

safetyarxiv-cs-ro
8 Jun 2026
← Previous
1…132133134135136…260
Next →