AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,806 results
Safety

Bandwidth Selection in Kernel Density Estimation for Model Calibration

DGX agent

arXiv:2606.29925v1 Announce Type: new Abstract: As deep learning models are increasingly deployed in high-stakes applications, providing well-calibrated uncertainty estimates has become as critical as

safetyarxiv-cs-lg
30 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Be Faithful When Response: Returning Fluent and Grounded Answers for Vision-Language Models Reinforcement Learning

DGX agent

arXiv:2606.29984v1 Announce Type: new Abstract: Reinforcement Learning (RL) is an important paradigm for improving the reasoning capabilities of Vision-Language Models (VLMs). However, directly applyi

safetyarxiv-cs-ai
30 Jun 2026
Safety

Behavior Prompting Policy: Demonstrations as Prompts for Manipulation

DGX agent

arXiv:2606.30457v1 Announce Type: new Abstract: We study behavior prompting, a paradigm that enables robots to perform new tasks at inference time given a single human demonstration, which we call a b

safetyarxiv-cs-ro
30 Jun 2026
Safety

Behavior Uncloning: Distilling Mode Redirection into Policy Weights without Inference-Time Steering

DGX agent

arXiv:2606.29201v1 Announce Type: cross Abstract: Behavior-cloned policies often learn multiple behavior modes from demonstration datasets, including modes that are unsafe or otherwise undesired at de

safetyarxiv-cs-ai
30 Jun 2026
Safety

Being at the launch of Cybercab two years ago was magic. A decade from now there will be millions of these things rolling around the world. …

DGX agent

Being at the launch of Cybercab two years ago was magic. A decade from now there will be millions of these things rolling around the world. The war over autonomous vehicles is on, as everyone in San F

safetyelon-musk--x
30 Jun 2026
Safety

Beyond Backscatter: AlphaEarth Land-Cover Priors for Rapid SAR Flood Segmentation Across Foundation Backbones

DGX agent

arXiv:2606.29134v1 Announce Type: new Abstract: Rapid flood mapping is critical for emergency response, yet optical imagery is often unusable during major flooding and single-temporal SAR is ambiguous

safetyarxiv-cs-cv
30 Jun 2026
Safety

BrainJanus: A Unified Model for Understanding and Generation across Brain, Vision, and Language

DGX agent

arXiv:2606.30319v1 Announce Type: new Abstract: Modeling the bidirectional correspondence between external sensory stimuli and internal neural activity has emerged as a critical frontier in neuroscien

safetyarxiv-cs-cv
30 Jun 2026
Safety

Bridgewater, one of the worlds largest hedge funds, a Tinker customer talks through how they've carefully fine-tuned a model focused on what…

DGX agent

Bridgewater, one of the worlds largest hedge funds, a Tinker customer talks through how they've carefully fine-tuned a model focused on what makes interesting financial news. Their fine-tuned model is

safetysoumith-chintala--x
30 Jun 2026
Safety

BTI-Net: Bidirectional Decoder-Level Task Interaction via Uncertainty-Aware Gating for Multi-Task Medical Image Analysis

DGX agent

arXiv:2606.29102v1 Announce Type: cross Abstract: Jointly learning to segment and classify medical images demands cross-task synergy, yet encoder-sharing architectures limit decoder reconstruction to

safetyarxiv-cs-ai
30 Jun 2026
Safety

Budgeted Act-or-Defer Multi-Agent LLM Deliberation with Local Reliability Bounds

DGX agent

arXiv:2606.29654v1 Announce Type: new Abstract: Multi-agent deliberation among LLMs can improve reasoning, but deployment requires deciding when the current answer is reliable enough to act on and whe

safetyarxiv-cs-ai
30 Jun 2026
Safety

Building Multi-Task Agentic LLMs via Two-Phase Distillation

DGX agent

arXiv:2606.30044v1 Announce Type: new Abstract: A key step toward artificial general intelligence is to train models that can perform multiple tasks. In this paper, we study how to build such models b

safetyarxiv-cs-lg
30 Jun 2026
Safety

BV-Blend: Uncertainty-Weighted Historical Baselines for Stable Critic-Free RL with Verifiable Rewards

DGX agent

arXiv:2606.28707v1 Announce Type: new Abstract: Critic-free reinforcement learning with verifiable rewards (RLVR), exemplified by Group Relative Policy Optimization (GRPO), avoids training a value fun

safetyarxiv-cs-ai
30 Jun 2026
Safety

CAMI: Cost-Aware Agent-Guided Multi-Indexing for Semantic Retrieval

DGX agent

arXiv:2606.28365v1 Announce Type: cross Abstract: RAG ingestion pipelines frequently augment search corpus index with semantic enrichment indices (e.g., synthetic queries or summaries generated from c

safetyarxiv-cs-ai
30 Jun 2026
Safety

Can LLMs Reliably Self-Report Adversarial Prefills, and How?

DGX agent

arXiv:2606.23671v2 Announce Type: replace Abstract: Prior work shows that large language models (LLMs) exhibit introspective capability on benign tasks. We extend the question to safety contexts and e

safetyarxiv-cs-cl
30 Jun 2026
Safety

Can MLLMs Critique Like Humans? Evaluating Open-Ended Aesthetic Reasoning in Multimodal Large Language Models

DGX agent

arXiv:2606.29689v1 Announce Type: new Abstract: Open-ended aesthetic critique is a challenge for multimodal large language models (MLLMs): unlike multiple-choice aesthetic benchmarks, it has no single

safetyarxiv-cs-cl
30 Jun 2026
Safety

CaresAI at CT-DEB26: Detecting Dosing Errors In Clinical Trials Using Domain-Specific Transformer Embeddings and Classification Models

DGX agent

arXiv:2606.30236v1 Announce Type: new Abstract: Medication errors, particularly dosing errors in clinical trials (CT), can lead to patient harm, adverse drug events and worse patient outcomes. Dosing

safetyarxiv-cs-cl
30 Jun 2026
Safety

Carolina Guide: A Multi-Agent RAG System with Institutional Guardrails for Academic Policy Assistance

DGX agent

arXiv:2606.28360v1 Announce Type: cross Abstract: University students often struggle to navigate complex academic policies, leading to advising bottlenecks and delayed access to critical information.

safetyarxiv-cs-ai
30 Jun 2026
Safety

Categorizing Mathematical Concepts with LLM Voting Ensembles in Mathswitch

DGX agent

arXiv:2606.28815v1 Announce Type: cross Abstract: Mathswitch is an open-source project that imports mathematical concept records from sources such as Wikidata, Wikipedia, MathWorld, Encyclopedia of Ma

safetyarxiv-cs-ai
30 Jun 2026
Safety

Characterizing Large Language Model Agentic Workflows: A Study on N8n Ecosystem

DGX agent

arXiv:2606.29116v1 Announce Type: new Abstract: Large Language Models (LLMs) are rapidly being adopted in low-code and no-code automation platforms, where non-expert users design workflows that combin

safetyarxiv-cs-ai
30 Jun 2026
Safety

Chronos: A Physics-Informed Full-History Framework for Non-Markovian Long-Horizon Manipulation

DGX agent

arXiv:2606.30318v1 Announce Type: new Abstract: General-purpose robot policies should be modeled as dynamical systems, yet many VLA and generative imitation policies still rely on present observations

safetyarxiv-cs-ro
30 Jun 2026
Safety

Clearer Sight, Fewer Lies: Oriented Pickup Preference Optimization for Multimodal Hallucination Mitigation

DGX agent

arXiv:2606.29805v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) are prone to hallucination as their generation preferences are insufficiently calibrated to visual evidence, ca

safetyarxiv-cs-cv
30 Jun 2026
Safety

CMTFormer: Marrying Transformer with Hierarchical Information Interaction for RGB-Event Object Detection

DGX agent

arXiv:2606.29136v1 Announce Type: cross Abstract: Event cameras capture sparse brightness changes with high temporal resolution and high dynamic range, compensating for the deficiencies of the convent

safetyarxiv-cs-ai
30 Jun 2026
Safety

Complementary RL: Towards Efficient Experience-Driven Agent Learning

DGX agent

arXiv:2603.17621v2 Announce Type: replace-cross Abstract: Reinforcement Learning (RL) has emerged as a powerful paradigm for training LLM-based agents, yet remains limited by low sample efficiency, st

safetyarxiv-cs-cl
30 Jun 2026
Safety

ConCent: Contact-Centric Real-to-Sim-to-Real Learning from One Demonstration

DGX agent

arXiv:2606.30268v1 Announce Type: new Abstract: Sim-to-real policy transfer -- deploying policies trained in simulation in the real world -- is a promising paradigm for scaling robot manipulation with

safetyarxiv-cs-ro
30 Jun 2026
Safety

Concept Removal Guidance: Evidence-Calibrated Negative Guidance for Safe Diffusion Sampling

DGX agent

arXiv:2606.29801v1 Announce Type: new Abstract: Text-to-image diffusion models remain vulnerable to adversarial prompts that elicit disallowed content, motivating reliable inference-time controls. A p

safetyarxiv-cs-cv
30 Jun 2026
Safety

Consistency as Inductive Bias: Learning Cross-View Invariance for Robust Multimodal Reasoning

DGX agent

arXiv:2606.29812v1 Announce Type: new Abstract: Inductive biases steer learning toward generalizable solutions by encoding task structure. In this work, we identify a crucial missing bias in MLLMs: cr

safetyarxiv-cs-cv
30 Jun 2026
Safety

CORE: Common Outcome Regularities from Action-Free Visual Demonstrations for Robot Manipulation

DGX agent

arXiv:2606.29517v1 Announce Type: new Abstract: Robot imitation learning often relies on costly robot demonstrations, while abundant action-free visual demonstrations, such as human videos, are diffic

safetyarxiv-cs-ro
30 Jun 2026
Safety

CRAFT: Counterfactual Credit Assignment from Free Sibling Rollouts for Self-Distilled Agentic Reinforcement Learning

DGX agent

arXiv:2606.29476v1 Announce Type: cross Abstract: Self-distilled agentic reinforcement learning augments trajectory-level reward with a token-level distillation loss, using as its teacher the same pol

safetyarxiv-cs-ai
30 Jun 2026
Safety

Critical Interval MSE: Toward Reliable Offline Validation for Robot Manipulation Policies

DGX agent

arXiv:2606.29898v1 Announce Type: cross Abstract: Real-world evaluation is the gold standard for robot policies because it tests them against the physical conditions and deployment challenges they are

safetyarxiv-cs-ai
30 Jun 2026
Safety

Data-Efficient Multimodal Alignment for Histopathology-based Molecular Prediction

DGX agent

arXiv:2606.29949v1 Announce Type: cross Abstract: H&E-stained whole-slide images offer cohort-scale availability and rich spatial context but lack molecular specificity, whereas bulk RNA-seq provides

safetyarxiv-cs-ai
30 Jun 2026
Safety

Delayed Bidirectional Alignment via Disentangled Audio Semantics for Audio-Visual Segmentation

DGX agent

arXiv:2512.20117v2 Announce Type: replace Abstract: Audio-Visual Segmentation (AVS) aims to localize sound-producing objects at the pixel level by integrating auditory and visual cues. However, existi

safetyarxiv-cs-cv
30 Jun 2026
Safety

Deterministic Decisions for High-Stakes AI. A Zero-Egress Pipeline with the Deployability of RAG and the Accuracy of Machine Learning

DGX agent

arXiv:2606.29280v1 Announce Type: cross Abstract: We identify intervention bias as a previously unquantified failure mode of zero-shot large-language-model (LLM) educational advisory agents: without t

safetyarxiv-cs-ai
30 Jun 2026
Safety

Diffusion Fine-tuning with Rewarded Moment Matching Distillation

DGX agent

arXiv:2606.30414v1 Announce Type: new Abstract: Distillation and Reinforcement Learning (RL) fine-tuning are the primary pillars of diffusion post-training. While traditionally studied in isolation, t

safetyarxiv-cs-lg
30 Jun 2026
Safety

Distribution Matching Variational AutoEncoder

DGX agent

arXiv:2512.07778v2 Announce Type: replace Abstract: Most visual generative models compress images into a latent space before applying diffusion or autoregressive modelling. Yet, existing approaches su

safetyarxiv-cs-cv
30 Jun 2026
Safety

Distributionally Robust Reinforcement Learning with Human Feedback

DGX agent

arXiv:2503.00539v2 Announce Type: replace-cross Abstract: Reinforcement learning from human feedback (RLHF) has evolved to be one of the main methods for fine-tuning large language models (LLMs). Howe

safetyarxiv-cs-ai
30 Jun 2026
Safety

Do Models Read What They Write? Causal Registers in Scratchpad Reasoning

DGX agent

arXiv:2606.29522v1 Announce Type: cross Abstract: A central hope behind process supervision is that models can expose intermediate variables that matter for their later behavior. For this to help with

safetyarxiv-cs-cl
30 Jun 2026
Safety

DOPD: Dual On-policy Distillation

DGX agent

arXiv:2606.30626v1 Announce Type: new Abstract: On-policy distillation (OPD) offers superior capacity transfer by supervising student-sampled trajectories with dense token-level signals. To furnish hi

safetyarxiv-cs-ai
30 Jun 2026
Safety

DRIFT: Difficulty Routing Self-DIstillation with Rhythm-Gated Exploration and Success BuFfer Training

DGX agent

arXiv:2606.30345v1 Announce Type: cross Abstract: Enabling large language models to achieve stable self-improvement without external expert supervision remains a central challenge in complex reasoning

safetyarxiv-cs-ai
30 Jun 2026
Safety

DrivenMorph: Bridging Attention Mechanism and Variational Image Registration via Difference Modeling

DGX agent

arXiv:2606.30183v1 Announce Type: new Abstract: Medical image registration benefits significantly from deep learning, yet existing approaches often lack physical explainability and fine-grained deform

safetyarxiv-cs-cv
30 Jun 2026
Safety

DTI: Dynamic Trajectory Initialization for Generative Face Video Super-Resolution

DGX agent

arXiv:2606.29198v1 Announce Type: new Abstract: As the most perceptually powerful Face Video Super-Resolution (FVSR) method, existing works in Generative FVSR (GFVSR) mainly exploit the generative pri

safetyarxiv-cs-cv
30 Jun 2026
Safety

Dual-Flow Reinforcement Learning with State-Aware Exploration

DGX agent

arXiv:2606.29820v1 Announce Type: cross Abstract: In complex continuous-control reinforcement learning tasks, multimodal optimal actions often coincide with uncertain, multimodal return distributions,

safetyarxiv-cs-ai
30 Jun 2026
Safety

DyGnROLE: Asymmetric Pretraining for Edge Classification on Dynamic Graphs

DGX agent

arXiv:2602.23135v2 Announce Type: replace-cross Abstract: Edge classification on directed dynamic graphs requires modeling interactions between source and destination nodes exhibiting asymmetrical beh

safetyarxiv-cs-ai
30 Jun 2026
Safety

Entity Binding Failures in Tool-Augmented Agents

DGX agent

arXiv:2606.30531v1 Announce Type: new Abstract: Tool-augmented language-model agents are often evaluated by whether they select the correct tool, produce valid API arguments, and complete the requeste

safetyarxiv-cs-ai
30 Jun 2026
Safety

EntroRouter: Learning Efficient Model Routing via Entropy Regulation

DGX agent

arXiv:2606.29424v1 Announce Type: new Abstract: Model routing balances solution accuracy and computational cost by selecting among models of varying capabilities. While recent multi-round frameworks i

safetyarxiv-cs-cl
30 Jun 2026
Safety

EPIC-EuroParl-UdS: Information-Theoretic Perspectives on Translation and Interpreting

DGX agent

arXiv:2603.09785v3 Announce Type: replace Abstract: This paper introduces an updated and combined version of the bidirectional English-German EPIC-UdS (spoken) and EuroParl-UdS (written) corpora conta

safetyarxiv-cs-cl
30 Jun 2026
Safety

EraseLoRA: MLLM-Driven Foreground Exclusion and Background Subtype Aggregation for Dataset-Free Object Removal

DGX agent

arXiv:2512.21545v2 Announce Type: replace Abstract: Object removal must prevent the masked target from reappearing and reconstruct the occluded background with structural and contextual fidelity, rath

safetyarxiv-cs-cv
30 Jun 2026
Safety

Estimating Grammatical Gender Directions in Contextual Embeddings under Controlled and Natural Contexts

DGX agent

arXiv:2606.30152v1 Announce Type: cross Abstract: Contextual language models conflate grammatical gender and social semantic bias in gendered languages such as Spanish. Existing gender debiasing appro

safetyarxiv-cs-ai
30 Jun 2026
Safety

Evidence-Based Text-Conditioned 3D CT Synthesis for Ovarian Cancer

DGX agent

arXiv:2606.28980v1 Announce Type: cross Abstract: Ovarian cancer is frequently diagnosed at an advanced stage, making preoperative contrast-enhanced computed tomography (CT) central to staging and sur

safetyarxiv-cs-ai
30 Jun 2026
← Previous
1…6869707172…267
Next →