AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,481 results
Safety

Test-Time Scaling in Reasoning Models Is Not Effective for Knowledge-Intensive Tasks Yet

DGX agent

arXiv:2509.06861v3 Announce Type: replace Abstract: Test-time scaling increases inference-time computation through longer reasoning chains and has shown strong performance gains across many domains. H

safetyarxiv-cs-ai
7 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Text Generation: A Systematic Literature Review of Tasks, Evaluation, and Challenges

DGX agent

arXiv:2405.15604v4 Announce Type: replace Abstract: Text generation has become more accessible than ever, and the growing interest in these systems, especially those using large language models, has s

safetyarxiv-cs-cl
7 Aug 2026
Safety

The Closing Window: How Governments Could Lose Their Ability to Restrain Advanced AI

DGX agent

arXiv:2608.05173v1 Announce Type: cross Abstract: As AI capabilities advance, AI systems will pose greater risks to national security and potentially humanity as a whole. Governments may eventually co

safetyarxiv-cs-ai
7 Aug 2026
Safety

The Illusion of Visual Tool-Use: A Causal Audit of Thinking with Images

DGX agent

arXiv:2608.06270v1 Announce Type: new Abstract: The 'thinking-with-images' paradigm equips multimodal LLMs with active visual operations such as crop-and-zoom. However, models using these operations o

safetyarxiv-cs-ai
7 Aug 2026
Safety

TRACE: Learned Proprioceptive Odometry for Legged Robots under Unreliable Contact Conditions

DGX agent

arXiv:2608.05975v1 Announce Type: cross Abstract: In this paper, we present TRACE (Tokenized Robust Attention for Contact-Aware Estimation), an end-to-end learned proprioceptive odometry estimator for

safetyarxiv-cs-ai
7 Aug 2026
Safety

Training a Conditioned Video Game Agent on a VLM Annotated Dataset

DGX agent

arXiv:2608.05954v1 Announce Type: new Abstract: Reinforcement Learning (RL) is a powerful but far from easy-to-use technique for policy learning. In the specific case of video games, access to the gam

safetyarxiv-cs-ai
7 Aug 2026
Safety

UniVVT: A Unified End-to-End Framework for High-Fidelity Video Virtual Try-on

DGX agent

arXiv:2608.05745v1 Announce Type: cross Abstract: Video Virtual Try-On (VVT) synthesizes a video of a person wearing a target garment while preserving identity, motion, and scene dynamics. Dominant ap

safetyarxiv-cs-ai
7 Aug 2026
Safety

VIDP: Variable Impedance Diffusion Policy for Compliant Robot Manipulation from Diverse Demonstrations

DGX agent

arXiv:2608.06210v1 Announce Type: new Abstract: Contact-rich manipulation requires precise tracking and mechanical compliance, where variable impedance control can improve robustness in task success,

safetyarxiv-cs-ro
7 Aug 2026
Safety

Visual Grounding in Zero-Shot Vision-Language Control

DGX agent

arXiv:2608.06154v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly used as zero-shot controllers, but successful trajectories do not necessarily show that decisions are g

safetyarxiv-cs-ai
7 Aug 2026
Safety

When Privileged Guidance Misaligns: State-Matched Routing and Contextualized Self-Distillation for Multi-Turn Agents

DGX agent

arXiv:2608.05219v1 Announce Type: new Abstract: Privileged on-policy distillation provides dense supervision for multi-turn agents by allowing a synchronized teacher to re-score the student's response

safetyarxiv-cs-ai
7 Aug 2026
Safety

XEWorld: Can Action-Conditioned World Models Generalize to Unseen Robot Embodiments?

DGX agent

arXiv:2608.05799v1 Announce Type: cross Abstract: Action-conditioned world models are promising learned simulators for robotic manipulation, yet evaluating them exclusively on training robots fails to

safetyarxiv-cs-cv
7 Aug 2026
Safety

A-SR: Self-Evolving Agentic LLMs for Symbolic Regression via Hierarchical Coordination

DGX agent

arXiv:2608.04872v1 Announce Type: cross Abstract: Symbolic regression aims to discover closed-form equations from data, but existing LLM-guided methods often rely on a unified proposal loop that compr

safetyarxiv-cs-ai
6 Aug 2026
Safety

A/B Agent: A Self-Evolving Agent for Strategy Iteration in Industrial A/B Testing

DGX agent

arXiv:2608.04625v1 Announce Type: new Abstract: Industrial recommendation strategy iteration heavily relies on large-scale A/B experimentation. Traditional tuning requires experts to repeatedly design

safetyarxiv-cs-ai
6 Aug 2026
Safety

Agent Skills for Automated Reasoning policies in Amazon Bedrock

DGX agent

Learn how to run the full Amazon Bedrock Automated Reasoning policy lifecycle from your coding agent. A suite of open source Agent Skills builds, reviews, tests, debugs, deploys, and validates a custo

safetyaws-ml-blog
6 Aug 2026
Safety

Agentic Reinforcement Learning with Observation-Calibrated Self-Distillation

DGX agent

arXiv:2608.04788v1 Announce Type: cross Abstract: Large language model agents are commonly trained through reinforcement learning with sparse trajectory-level rewards, which offer limited guidance on

safetyarxiv-cs-ai
6 Aug 2026
Safety

Among proponents of neurosymbolic architectures, there had been some debate over the years about whether the outer level would be symbolic (…

DGX agent

Among proponents of neurosymbolic architectures, there had been some debate over the years about whether the outer level would be symbolic (i.e. a harness that calls neural models) or whether the oute

safetygary-marcus--x
6 Aug 2026
Safety

Arnold: A multi-task, multi-embodiment muscle transformer policy

DGX agent

arXiv:2508.18066v2 Announce Type: replace-cross Abstract: Controlling high-dimensional and nonlinear musculoskeletal models of the human body is a foundational scientific challenge. Recent machine lea

safetyarxiv-cs-ai
6 Aug 2026
Safety

ATLAS: Adaptive Topological Learning with Abstract Successors for Continual Learning

DGX agent

arXiv:2608.04334v1 Announce Type: cross Abstract: Contemporary model-free reinforcement learning algorithms can achieve very high performance, but have low sample efficiency and are not robust to chan

safetyarxiv-cs-ai
6 Aug 2026
Safety

Attention Fusion for Bridge Deck Delamination Detection

DGX agent

arXiv:2512.20113v4 Announce Type: replace Abstract: Subsurface delaminations in reinforced concrete bridge decks escape conventional visual inspection, and the two principal sensing techniques used to

safetyarxiv-cs-cv
6 Aug 2026
Safety

Beyond the Dirac Delta: Mitigating Diversity Collapse in Reinforcement Fine-Tuning for Versatile Image Generation

DGX agent

arXiv:2601.12401v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) has emerged as a powerful paradigm for fine-tuning large-scale generative models, such as diffusion and flow model

safetyarxiv-cs-ai
6 Aug 2026
Safety

BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation

DGX agent

arXiv:2608.05042v1 Announce Type: new Abstract: Leveraging pre-trained vision-language models (VLMs) to construct vision-language-action (VLA) models has emerged as a promising paradigm for 3D robot m

safetyarxiv-cs-ro
6 Aug 2026
Safety

Calibrating Artificial Guilt: Neurally Grounded Reward Shaping for Prosocial Multi-Agent Reinforcement Learning

DGX agent

arXiv:2608.04663v1 Announce Type: new Abstract: Cooperative multi-agent reinforcement learning often adds social terms to individual rewards, yet the scale of those terms is usually chosen by hand. We

safetyarxiv-cs-ai
6 Aug 2026
Safety

Calibrating Transformer Attention via Task-Space Sensitivity Feedback

DGX agent

arXiv:2512.20661v2 Announce Type: replace Abstract: Transformer-based pre-trained language models (PLMs) excel in text classification but suffer from attention dilution and attention sink effects, for

safetyarxiv-cs-ai
6 Aug 2026
Safety

CARGO-VL: Counterfactual Arbitration with Risk-Constrained Group Optimization for Vision-Language Models

DGX agent

arXiv:2608.04509v1 Announce Type: new Abstract: Vision-language systems combine images with retrieved text, but these sources can disagree or jointly fail to support an answer. Reliable models must id

safetyarxiv-cs-ai
6 Aug 2026
Safety

CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Models

DGX agent

arXiv:2608.04302v1 Announce Type: new Abstract: Benchmarking video-language models has largely focused on short clips and single-sentence metrics, leaving open whether current systems can generate acc

safetyarxiv-cs-cv
6 Aug 2026
Safety

C’mon, it’s not game over for Google Seven reasons why not, excerpted from my newsletter:

DGX agent

C’mon, it’s not game over for Google Seven reasons why not, excerpted from my newsletter: Jeff Dean and Demis Hassabis are the two most important AI executives at Google. Jeff is leaving and Demis is

safetygary-marcus--x
6 Aug 2026
Safety

CofactVLA: Deconfounding Vision-Language-Action Models via Counterfactual Intervention

DGX agent

arXiv:2608.04396v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have driven significant progress in robotic manipulation, yet they fundamentally struggle with the vision-override p

safetyarxiv-cs-cv
6 Aug 2026
Safety

Compass: Continuously Aligning Social Media Feeds via In-Situ Reflections

DGX agent

arXiv:2608.04274v1 Announce Type: cross Abstract: Social media recommendation feeds often optimize for users' immediate impulses rather than preferences they would hold after deeper reflection. Some s

safetyarxiv-cs-ai
6 Aug 2026
Safety

Contrastive Diffusion Alignment: Learning Structured Latents for Controllable Generation

DGX agent

arXiv:2510.14190v3 Announce Type: replace Abstract: Diffusion models excel at generation, but their latent spaces are high dimensional and not explicitly organized for interpretation or control. We in

safetyarxiv-cs-lg
6 Aug 2026
Safety

Control agent behaviors and cost beyond a single action: new capabilities in Amazon Bedrock AgentCore

DGX agent

Learn about new capabilities in Amazon Bedrock AgentCore: temporal policies powered by Dogwood, a new open source policy language for AI agents, and rate limiting on the gateway. These features give y

safetyaws-ml-blog
6 Aug 2026
Safety

COSMO: Consensus-Driven Shift Modulation for Source-Free Domain Adaptation

DGX agent

arXiv:2608.04604v1 Announce Type: new Abstract: Source-free domain adaptation (SFDA) adapts a source-trained model to an unlabeled target domain without source data, a practical setting under privacy

safetyarxiv-cs-cv
6 Aug 2026
Safety

critical history and context that a lot of people have conveniently forgotten

DGX agent

critical history and context that a lot of people have conveniently forgotten For a very long time most high-performing AI models were end-to-end neural models; vector input -> vector output, with onl

safetygary-marcus--x
6 Aug 2026
Safety

DAC-Pose: Dual-Agent Collaborative Framework for Pose-Guided Human Generation

DGX agent

arXiv:2608.04622v1 Announce Type: new Abstract: AI agents have emerged as a powerful new paradigm in generative image synthesis, enabling systems to perform complex semantic reasoning rather than pass

safetyarxiv-cs-cv
6 Aug 2026
Safety

Data-Aware and Scalable Sensitivity Analysis for Decision Tree Ensembles

DGX agent

arXiv:2602.07453v2 Announce Type: replace Abstract: Decision tree ensembles are widely used in critical domains, making robustness and sensitivity analysis essential to their trustworthiness. We study

safetyarxiv-cs-lg
6 Aug 2026
Safety

Differentiating Through Dual Prices: End-to-End Policy Learning Under Capacity Constraints

DGX agent

arXiv:2608.04669v1 Announce Type: new Abstract: Many social services assign scarce resources, such as housing assistance or hospital interventions, to people who arrive one at a time: each arrival mus

safetyarxiv-cs-lg
6 Aug 2026
Safety

DXC partners with Primary on zero-trust security for enterprise AI

DGX agent

DXC Technology Co. today announced a partnership with security startup Primary that makes the information technology services company the exclusive managed services partner for Primary’s zero-trust pl

safetysiliconangle
6 Aug 2026
Safety

Enabling Urgency-aware Robot Swarm Intralogistics using Smart IoT Tags

DGX agent

arXiv:2608.04721v1 Announce Type: new Abstract: Warehouse items differ in how urgently they must be moved: perishable goods, pharmaceutical shipments, and just-in-time production materials must be del

safetyarxiv-cs-ro
6 Aug 2026
Safety

EndoVLM: An Endoscopy Vision-Language Pre-training Model via Anatomy-Guided Sparsity and Progressive Alignment

DGX agent

arXiv:2608.04472v1 Announce Type: cross Abstract: The development of foundation models (FMs) is crucial for advancing endoscopic image analysis. However, existing endoscopy FMs mainly rely on self-sup

safetyarxiv-cs-ai
6 Aug 2026
Safety

Exact Model-Free Policy Iteration for Co-safe LTL Planning

DGX agent

arXiv:2608.05047v1 Announce Type: cross Abstract: This work studies model-free reinforcement learning for co-safe linear temporal logic (sc-LTL) objectives in finite Markov decision processes, which c

safetyarxiv-cs-ro
6 Aug 2026
Safety

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning

DGX agent

arXiv:2608.04771v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) excel on complex tasks through long chain-of-thought (CoT) reasoning, but their lengthy intermediate steps cause severe ov

safetyarxiv-cs-ai
6 Aug 2026
Safety

Flash-VAED: Plug-and-Play VAE Decoders for Efficient Video Generation

DGX agent

arXiv:2602.19161v2 Announce Type: replace Abstract: Latent diffusion models have enabled high-quality video synthesis, yet their inference remains costly and time-consuming. As diffusion transformers

safetyarxiv-cs-cv
6 Aug 2026
Safety

FocusMem: Factorizing Content, Readout, and Trust in Latent GUI Memory

DGX agent

arXiv:2608.04530v1 Announce Type: new Abstract: GUI agents must remember both useful experience from earlier tasks and unfinished progress in the current interaction. Latent memory offers a compact so

safetyarxiv-cs-cv
6 Aug 2026
Safety

From Brute Force to Semantic Insight: Performance-Guided Data Transformation Design with LLMs

DGX agent

arXiv:2601.03808v2 Announce Type: replace Abstract: Large language models (LLMs) have achieved notable performance in code synthesis; however, data-aware augmentation remains a limiting factor, handle

safetyarxiv-cs-cv
6 Aug 2026
Safety

Gary Marcus won! @GaryMarcus

DGX agent

Gary Marcus won! @GaryMarcus I would have assumed it was fairly obvious, but in case it's not: a million-line codebase (also known as a 'harness'), running at inference time, orchestrating thousands o

safetygary-marcus--x
6 Aug 2026
Safety

Generative Optimization for Incentivized Advertising with Global Level Constraints

DGX agent

arXiv:2608.04421v1 Announce Type: cross Abstract: Incentivized advertising allocates monetary or virtual rewards to drive user engagement, where a key challenge is optimizing continuous incentive magn

safetyarxiv-cs-ai
6 Aug 2026
Safety

GeoReward: Mitigating Contextual Variable Overestimation in Vision-Language Models for Cross-Market Preference Prediction

DGX agent

arXiv:2608.04504v1 Announce Type: cross Abstract: Vision-language models excel in many multimodal tasks but remain prone to a subtle yet impactful failure mode: they tend to overestimate dominant visu

safetyarxiv-cs-ai
6 Aug 2026
Safety

GFlowNet Training by Policy Gradients

DGX agent

arXiv:2408.05885v3 Announce Type: replace Abstract: Generative Flow Networks (GFlowNets) have been shown effective to generate combinatorial objects with desired properties. We here propose a new GFlo

safetyarxiv-cs-lg
6 Aug 2026
Safety

Governing Execution Risk in Agentic AI Systems: A Trajectory-Guided Framework for Red Teaming

DGX agent

arXiv:2608.04018v1 Announce Type: cross Abstract: AI agents are increasingly embedded in organizational workflows, where they interact with external information sources and invoke digital tools to per

safetyarxiv-cs-ai
6 Aug 2026
← Previous
1…7677787980…302
Next →