AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
All
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,708 results
Safety

Behavior Prompting Policy: Demonstrations as Prompts for Manipulation

DGX agent

arXiv:2606.30457v1 Announce Type: new Abstract: We study behavior prompting, a paradigm that enables robots to perform new tasks at inference time given a single human demonstration, which we call a b

safetyarxiv-cs-ro
30 Jun 2026
Safety
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Behavior Uncloning: Distilling Mode Redirection into Policy Weights without Inference-Time Steering

DGX agent

arXiv:2606.29201v1 Announce Type: cross Abstract: Behavior-cloned policies often learn multiple behavior modes from demonstration datasets, including modes that are unsafe or otherwise undesired at de

safetyarxiv-cs-ai
30 Jun 2026
Safety

Being at the launch of Cybercab two years ago was magic. A decade from now there will be millions of these things rolling around the world. …

DGX agent

Being at the launch of Cybercab two years ago was magic. A decade from now there will be millions of these things rolling around the world. The war over autonomous vehicles is on, as everyone in San F

safetyelon-musk--x
30 Jun 2026
Safety

Beyond Backscatter: AlphaEarth Land-Cover Priors for Rapid SAR Flood Segmentation Across Foundation Backbones

DGX agent

arXiv:2606.29134v1 Announce Type: new Abstract: Rapid flood mapping is critical for emergency response, yet optical imagery is often unusable during major flooding and single-temporal SAR is ambiguous

safetyarxiv-cs-cv
30 Jun 2026
Safety

BrainJanus: A Unified Model for Understanding and Generation across Brain, Vision, and Language

DGX agent

arXiv:2606.30319v1 Announce Type: new Abstract: Modeling the bidirectional correspondence between external sensory stimuli and internal neural activity has emerged as a critical frontier in neuroscien

safetyarxiv-cs-cv
30 Jun 2026
Safety

Bridgewater, one of the worlds largest hedge funds, a Tinker customer talks through how they've carefully fine-tuned a model focused on what…

DGX agent

Bridgewater, one of the worlds largest hedge funds, a Tinker customer talks through how they've carefully fine-tuned a model focused on what makes interesting financial news. Their fine-tuned model is

safetysoumith-chintala--x
30 Jun 2026
Safety

BTI-Net: Bidirectional Decoder-Level Task Interaction via Uncertainty-Aware Gating for Multi-Task Medical Image Analysis

DGX agent

arXiv:2606.29102v1 Announce Type: cross Abstract: Jointly learning to segment and classify medical images demands cross-task synergy, yet encoder-sharing architectures limit decoder reconstruction to

safetyarxiv-cs-ai
30 Jun 2026
Safety

Budgeted Act-or-Defer Multi-Agent LLM Deliberation with Local Reliability Bounds

DGX agent

arXiv:2606.29654v1 Announce Type: new Abstract: Multi-agent deliberation among LLMs can improve reasoning, but deployment requires deciding when the current answer is reliable enough to act on and whe

safetyarxiv-cs-ai
30 Jun 2026
Safety

Building Multi-Task Agentic LLMs via Two-Phase Distillation

DGX agent

arXiv:2606.30044v1 Announce Type: new Abstract: A key step toward artificial general intelligence is to train models that can perform multiple tasks. In this paper, we study how to build such models b

safetyarxiv-cs-lg
30 Jun 2026
Safety

BV-Blend: Uncertainty-Weighted Historical Baselines for Stable Critic-Free RL with Verifiable Rewards

DGX agent

arXiv:2606.28707v1 Announce Type: new Abstract: Critic-free reinforcement learning with verifiable rewards (RLVR), exemplified by Group Relative Policy Optimization (GRPO), avoids training a value fun

safetyarxiv-cs-ai
30 Jun 2026
Safety

CAMI: Cost-Aware Agent-Guided Multi-Indexing for Semantic Retrieval

DGX agent

arXiv:2606.28365v1 Announce Type: cross Abstract: RAG ingestion pipelines frequently augment search corpus index with semantic enrichment indices (e.g., synthetic queries or summaries generated from c

safetyarxiv-cs-ai
30 Jun 2026
Safety

Can LLMs Reliably Self-Report Adversarial Prefills, and How?

DGX agent

arXiv:2606.23671v2 Announce Type: replace Abstract: Prior work shows that large language models (LLMs) exhibit introspective capability on benign tasks. We extend the question to safety contexts and e

safetyarxiv-cs-cl
30 Jun 2026
Safety

Can MLLMs Critique Like Humans? Evaluating Open-Ended Aesthetic Reasoning in Multimodal Large Language Models

DGX agent

arXiv:2606.29689v1 Announce Type: new Abstract: Open-ended aesthetic critique is a challenge for multimodal large language models (MLLMs): unlike multiple-choice aesthetic benchmarks, it has no single

safetyarxiv-cs-cl
30 Jun 2026
Safety

CaresAI at CT-DEB26: Detecting Dosing Errors In Clinical Trials Using Domain-Specific Transformer Embeddings and Classification Models

DGX agent

arXiv:2606.30236v1 Announce Type: new Abstract: Medication errors, particularly dosing errors in clinical trials (CT), can lead to patient harm, adverse drug events and worse patient outcomes. Dosing

safetyarxiv-cs-cl
30 Jun 2026
Safety

Carolina Guide: A Multi-Agent RAG System with Institutional Guardrails for Academic Policy Assistance

DGX agent

arXiv:2606.28360v1 Announce Type: cross Abstract: University students often struggle to navigate complex academic policies, leading to advising bottlenecks and delayed access to critical information.

safetyarxiv-cs-ai
30 Jun 2026
Safety

Categorizing Mathematical Concepts with LLM Voting Ensembles in Mathswitch

DGX agent

arXiv:2606.28815v1 Announce Type: cross Abstract: Mathswitch is an open-source project that imports mathematical concept records from sources such as Wikidata, Wikipedia, MathWorld, Encyclopedia of Ma

safetyarxiv-cs-ai
30 Jun 2026
Safety

Characterizing Large Language Model Agentic Workflows: A Study on N8n Ecosystem

DGX agent

arXiv:2606.29116v1 Announce Type: new Abstract: Large Language Models (LLMs) are rapidly being adopted in low-code and no-code automation platforms, where non-expert users design workflows that combin

safetyarxiv-cs-ai
30 Jun 2026
Safety

Chronos: A Physics-Informed Full-History Framework for Non-Markovian Long-Horizon Manipulation

DGX agent

arXiv:2606.30318v1 Announce Type: new Abstract: General-purpose robot policies should be modeled as dynamical systems, yet many VLA and generative imitation policies still rely on present observations

safetyarxiv-cs-ro
30 Jun 2026
Safety

Clearer Sight, Fewer Lies: Oriented Pickup Preference Optimization for Multimodal Hallucination Mitigation

DGX agent

arXiv:2606.29805v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) are prone to hallucination as their generation preferences are insufficiently calibrated to visual evidence, ca

safetyarxiv-cs-cv
30 Jun 2026
Safety

CMTFormer: Marrying Transformer with Hierarchical Information Interaction for RGB-Event Object Detection

DGX agent

arXiv:2606.29136v1 Announce Type: cross Abstract: Event cameras capture sparse brightness changes with high temporal resolution and high dynamic range, compensating for the deficiencies of the convent

safetyarxiv-cs-ai
30 Jun 2026
Safety

Complementary RL: Towards Efficient Experience-Driven Agent Learning

DGX agent

arXiv:2603.17621v2 Announce Type: replace-cross Abstract: Reinforcement Learning (RL) has emerged as a powerful paradigm for training LLM-based agents, yet remains limited by low sample efficiency, st

safetyarxiv-cs-cl
30 Jun 2026
Safety

ConCent: Contact-Centric Real-to-Sim-to-Real Learning from One Demonstration

DGX agent

arXiv:2606.30268v1 Announce Type: new Abstract: Sim-to-real policy transfer -- deploying policies trained in simulation in the real world -- is a promising paradigm for scaling robot manipulation with

safetyarxiv-cs-ro
30 Jun 2026
Safety

Concept Removal Guidance: Evidence-Calibrated Negative Guidance for Safe Diffusion Sampling

DGX agent

arXiv:2606.29801v1 Announce Type: new Abstract: Text-to-image diffusion models remain vulnerable to adversarial prompts that elicit disallowed content, motivating reliable inference-time controls. A p

safetyarxiv-cs-cv
30 Jun 2026
Safety

Consistency as Inductive Bias: Learning Cross-View Invariance for Robust Multimodal Reasoning

DGX agent

arXiv:2606.29812v1 Announce Type: new Abstract: Inductive biases steer learning toward generalizable solutions by encoding task structure. In this work, we identify a crucial missing bias in MLLMs: cr

safetyarxiv-cs-cv
30 Jun 2026
Safety

CORE: Common Outcome Regularities from Action-Free Visual Demonstrations for Robot Manipulation

DGX agent

arXiv:2606.29517v1 Announce Type: new Abstract: Robot imitation learning often relies on costly robot demonstrations, while abundant action-free visual demonstrations, such as human videos, are diffic

safetyarxiv-cs-ro
30 Jun 2026
Safety

CRAFT: Counterfactual Credit Assignment from Free Sibling Rollouts for Self-Distilled Agentic Reinforcement Learning

DGX agent

arXiv:2606.29476v1 Announce Type: cross Abstract: Self-distilled agentic reinforcement learning augments trajectory-level reward with a token-level distillation loss, using as its teacher the same pol

safetyarxiv-cs-ai
30 Jun 2026
Safety

Critical Interval MSE: Toward Reliable Offline Validation for Robot Manipulation Policies

DGX agent

arXiv:2606.29898v1 Announce Type: cross Abstract: Real-world evaluation is the gold standard for robot policies because it tests them against the physical conditions and deployment challenges they are

safetyarxiv-cs-ai
30 Jun 2026
Safety

Data-Efficient Multimodal Alignment for Histopathology-based Molecular Prediction

DGX agent

arXiv:2606.29949v1 Announce Type: cross Abstract: H&E-stained whole-slide images offer cohort-scale availability and rich spatial context but lack molecular specificity, whereas bulk RNA-seq provides

safetyarxiv-cs-ai
30 Jun 2026
Safety

Delayed Bidirectional Alignment via Disentangled Audio Semantics for Audio-Visual Segmentation

DGX agent

arXiv:2512.20117v2 Announce Type: replace Abstract: Audio-Visual Segmentation (AVS) aims to localize sound-producing objects at the pixel level by integrating auditory and visual cues. However, existi

safetyarxiv-cs-cv
30 Jun 2026
Safety

Deterministic Decisions for High-Stakes AI. A Zero-Egress Pipeline with the Deployability of RAG and the Accuracy of Machine Learning

DGX agent

arXiv:2606.29280v1 Announce Type: cross Abstract: We identify intervention bias as a previously unquantified failure mode of zero-shot large-language-model (LLM) educational advisory agents: without t

safetyarxiv-cs-ai
30 Jun 2026
Safety

Diffusion Fine-tuning with Rewarded Moment Matching Distillation

DGX agent

arXiv:2606.30414v1 Announce Type: new Abstract: Distillation and Reinforcement Learning (RL) fine-tuning are the primary pillars of diffusion post-training. While traditionally studied in isolation, t

safetyarxiv-cs-lg
30 Jun 2026
Safety

Distribution Matching Variational AutoEncoder

DGX agent

arXiv:2512.07778v2 Announce Type: replace Abstract: Most visual generative models compress images into a latent space before applying diffusion or autoregressive modelling. Yet, existing approaches su

safetyarxiv-cs-cv
30 Jun 2026
Safety

Distributionally Robust Reinforcement Learning with Human Feedback

DGX agent

arXiv:2503.00539v2 Announce Type: replace-cross Abstract: Reinforcement learning from human feedback (RLHF) has evolved to be one of the main methods for fine-tuning large language models (LLMs). Howe

safetyarxiv-cs-ai
30 Jun 2026
Safety

Do Models Read What They Write? Causal Registers in Scratchpad Reasoning

DGX agent

arXiv:2606.29522v1 Announce Type: cross Abstract: A central hope behind process supervision is that models can expose intermediate variables that matter for their later behavior. For this to help with

safetyarxiv-cs-cl
30 Jun 2026
Safety

DOPD: Dual On-policy Distillation

DGX agent

arXiv:2606.30626v1 Announce Type: new Abstract: On-policy distillation (OPD) offers superior capacity transfer by supervising student-sampled trajectories with dense token-level signals. To furnish hi

safetyarxiv-cs-ai
30 Jun 2026
Safety

DRIFT: Difficulty Routing Self-DIstillation with Rhythm-Gated Exploration and Success BuFfer Training

DGX agent

arXiv:2606.30345v1 Announce Type: cross Abstract: Enabling large language models to achieve stable self-improvement without external expert supervision remains a central challenge in complex reasoning

safetyarxiv-cs-ai
30 Jun 2026
Safety

DrivenMorph: Bridging Attention Mechanism and Variational Image Registration via Difference Modeling

DGX agent

arXiv:2606.30183v1 Announce Type: new Abstract: Medical image registration benefits significantly from deep learning, yet existing approaches often lack physical explainability and fine-grained deform

safetyarxiv-cs-cv
30 Jun 2026
Safety

DTI: Dynamic Trajectory Initialization for Generative Face Video Super-Resolution

DGX agent

arXiv:2606.29198v1 Announce Type: new Abstract: As the most perceptually powerful Face Video Super-Resolution (FVSR) method, existing works in Generative FVSR (GFVSR) mainly exploit the generative pri

safetyarxiv-cs-cv
30 Jun 2026
Safety

Dual-Flow Reinforcement Learning with State-Aware Exploration

DGX agent

arXiv:2606.29820v1 Announce Type: cross Abstract: In complex continuous-control reinforcement learning tasks, multimodal optimal actions often coincide with uncertain, multimodal return distributions,

safetyarxiv-cs-ai
30 Jun 2026
Safety

DyGnROLE: Asymmetric Pretraining for Edge Classification on Dynamic Graphs

DGX agent

arXiv:2602.23135v2 Announce Type: replace-cross Abstract: Edge classification on directed dynamic graphs requires modeling interactions between source and destination nodes exhibiting asymmetrical beh

safetyarxiv-cs-ai
30 Jun 2026
Safety

Entity Binding Failures in Tool-Augmented Agents

DGX agent

arXiv:2606.30531v1 Announce Type: new Abstract: Tool-augmented language-model agents are often evaluated by whether they select the correct tool, produce valid API arguments, and complete the requeste

safetyarxiv-cs-ai
30 Jun 2026
Safety

EntroRouter: Learning Efficient Model Routing via Entropy Regulation

DGX agent

arXiv:2606.29424v1 Announce Type: new Abstract: Model routing balances solution accuracy and computational cost by selecting among models of varying capabilities. While recent multi-round frameworks i

safetyarxiv-cs-cl
30 Jun 2026
Safety

EPIC-EuroParl-UdS: Information-Theoretic Perspectives on Translation and Interpreting

DGX agent

arXiv:2603.09785v3 Announce Type: replace Abstract: This paper introduces an updated and combined version of the bidirectional English-German EPIC-UdS (spoken) and EuroParl-UdS (written) corpora conta

safetyarxiv-cs-cl
30 Jun 2026
Safety

EraseLoRA: MLLM-Driven Foreground Exclusion and Background Subtype Aggregation for Dataset-Free Object Removal

DGX agent

arXiv:2512.21545v2 Announce Type: replace Abstract: Object removal must prevent the masked target from reappearing and reconstruct the occluded background with structural and contextual fidelity, rath

safetyarxiv-cs-cv
30 Jun 2026
Safety

Estimating Grammatical Gender Directions in Contextual Embeddings under Controlled and Natural Contexts

DGX agent

arXiv:2606.30152v1 Announce Type: cross Abstract: Contextual language models conflate grammatical gender and social semantic bias in gendered languages such as Spanish. Existing gender debiasing appro

safetyarxiv-cs-ai
30 Jun 2026
Safety

Evidence-Based Text-Conditioned 3D CT Synthesis for Ovarian Cancer

DGX agent

arXiv:2606.28980v1 Announce Type: cross Abstract: Ovarian cancer is frequently diagnosed at an advanced stage, making preoperative contrast-enhanced computed tomography (CT) central to staging and sur

safetyarxiv-cs-ai
30 Jun 2026
Safety

Experience-Evolving Multi-Turn Tool-Use Agent with Hybrid Episodic-Procedural Memory

DGX agent

arXiv:2512.07287v3 Announce Type: replace-cross Abstract: As intents unfold and environments change, multi-turn agents face continuously shifting decision contexts. Although reusing past experience is

safetyarxiv-cs-ai
30 Jun 2026
Safety

Expert-guided Clinical Text Augmentation via Query-Based Model Collaboration

DGX agent

arXiv:2509.21530v2 Announce Type: replace Abstract: Data augmentation is a widely used strategy to improve model robustness and generalization by enriching training datasets with synthetic examples. W

safetyarxiv-cs-lg
30 Jun 2026
← Previous
1…6667686970…265
Next →