AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,487 results
Safety

Quantifying Political Partisanship for Cross-Platform Analyses

DGX agent

arXiv:2607.21842v1 Announce Type: cross Abstract: Research on political polarization on social media depends on the ability to reliably measure partisanship in user-generated content. However, existin

safetyarxiv-cs-lg
27 Jul 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Reliability Scales Inversely: Bigger Language Models Compound Mistakes Faster

DGX agent

arXiv:2607.18292v2 Announce Type: replace-cross Abstract: As language models scale, answers start truer but degrade faster: scaling buys capability but erodes reliability. The knowledge-gap account --

safetyarxiv-cs-cl
27 Jul 2026
Safety

Self-Guided Process Reward Optimization with Redefined Step-wise Advantage for Process Reinforcement Learning

DGX agent

arXiv:2507.01551v3 Announce Type: replace-cross Abstract: Process Reinforcement Learning~(PRL) has demonstrated considerable potential in enhancing the reasoning capabilities of Large Language Models~

safetyarxiv-cs-cl
27 Jul 2026
Safety

Spectral Prior for Reducing Exposure Bias in Diffusion Models

DGX agent

arXiv:2607.22091v1 Announce Type: new Abstract: Diffusion models typically suffer from error accumulation during iterative sampling, commonly referred to as exposure bias. We reveal systematic frequen

safetyarxiv-cs-cv
27 Jul 2026
Safety

Stocks should be ripping today given the move in oil. But they're not thanks to NVDA and it's unbelievable 250 Billion roundtrip with Open…

DGX agent

Stocks should be ripping today given the move in oil. But they're not thanks to NVDA and it's unbelievable 250 Billion roundtrip with OpenAI. There's no hiding anymore that the whole AI bubble is real

safetygary-marcus--x
27 Jul 2026
Safety

TextSLIP: Text Self-Supervised CLIP for Medical Report Generation

DGX agent

arXiv:2607.21970v1 Announce Type: new Abstract: Automating radiology report generation is important for improving reporting consistency and clinical workflows . While Contrastive Language--Image Pretr

safetyarxiv-cs-cv
27 Jul 2026
Safety

The Coordination Gap: Multi-Agent Alternation Metrics for Temporal Fairness in Repeated Games

DGX agent

arXiv:2603.05789v5 Announce Type: replace-cross Abstract: Repeated multi-agent interactions require evaluation metrics that capture not only payoff distributions but also their temporal organization.

safetyarxiv-cs-lg
27 Jul 2026
Safety

TRACE-ROUTER: Task-Consistent and Adaptive Online Routing for Agentic AI

DGX agent

arXiv:2607.22465v1 Announce Type: cross Abstract: Routing to select large language models (LLMs) with different cost-quality trade-offs has become a fundamental deployment feature of enterprise AI. Ex

safetyarxiv-cs-lg
27 Jul 2026
Safety

Twins: Learn to Predict Unified Representations with Focal Loss

DGX agent

arXiv:2607.22531v1 Announce Type: new Abstract: Unified multimodal models seek a shared visual token space that supports both multimodal understanding and image generation. Discrete methods unify the

safetyarxiv-cs-cv
27 Jul 2026
Safety

Very cool paper from Microsoft. The idea is to train agents on replayed teacher trajectories instead of live environment rollouts. On-policy…

DGX agent

Very cool paper from Microsoft. The idea is to train agents on replayed teacher trajectories instead of live environment rollouts. On-policy distillation for agentic tasks is expensive because every u

safetydair-ai--x
27 Jul 2026
Safety

ViTacWorld: Scaling Visuo-Tactile World Models for Contact-Rich Robot Manipulation

DGX agent

arXiv:2607.22530v1 Announce Type: new Abstract: Contact-rich robot manipulation requires physical interaction cues that are often invisible to cameras, making tactile sensing essential for robust cont

safetyarxiv-cs-ro
27 Jul 2026
Model Releases

WHBench: Evaluating Frontier LLMs with Expert-in-the-Loop Validation on Women's Health Topics

DGX agent

arXiv:2604.00024v2 Announce Type: replace Abstract: Large language models are increasingly used for medical guidance, but women's health remains under-evaluated in benchmark design. We present the Wom

model-releasesarxiv-cs-cl
27 Jul 2026
Safety

When Ethics and Payoffs Diverge: LLM Agents in Morally Charged Social Dilemmas

DGX agent

arXiv:2505.19212v2 Announce Type: replace Abstract: Recent advances in LLMs have enabled their use in complex agentic roles, involving decision-making with humans or other agents, making ethical align

safetyarxiv-cs-cl
27 Jul 2026
Safety

Why Large Language Models and Humans Converge and Diverge in Evaluating Creativity

DGX agent

arXiv:2607.22218v1 Announce Type: new Abstract: Despite the growing use of large language models (LLMs) as creativity evaluators, evidence of their alignment with human evaluations remains mixed, rais

safetyarxiv-cs-cl
27 Jul 2026
Safety

important thread; don’t read just the first tweet (which has a caveat in the second).

DGX agent

important thread; don’t read just the first tweet (which has a caveat in the second). In case you think the problem of science slop is hypothetical: Here's the president of OpenAI retweeting wrong sci

safetygary-marcus--x
26 Jul 2026
Safety

training to benchmarks ≠ getting to AGI

DGX agent

training to benchmarks ≠ getting to AGI This suggests that Opus 5 's gain on ARC-AGI-3 was the result of specific training to improve on that eval, and not a generalized increase in abstract reasoning

safetygary-marcus--x
26 Jul 2026
Safety

All anyone has to do is go back 15 years and look at all the ridiculous predictions Elmo has made about Mars, self driving cars, his ridicul…

DGX agent

All anyone has to do is go back 15 years and look at all the ridiculous predictions Elmo has made about Mars, self driving cars, his ridiculous Boring Company. No one should take anything he says seri

safetygary-marcus--x
25 Jul 2026
Safety

I am surprised how few people are aware that the reasoning for OpenAI/Anthropic models is all encrypted. The 'reasoning' you see in the UI i…

DGX agent

OpenAI and Anthropic’s language models keep their internal reasoning encrypted; what users see in the UI is only a filtered summary of that reasoning. This practice was highlighted in a tweet by Sarah

safetygary-marcus--x
25 Jul 2026
Safety

The two giant generative AI startups were treated like gods for a couple years. Both are now facing massive pushback.

DGX agent

The two giant generative AI startups were treated like gods for a couple years. Both are now facing massive pushback. Anthropic employees go nuclear on the popularity of open source AI. It is getting

safetygary-marcus--x
25 Jul 2026
Safety

When a normal company does a bad thing, they take responsibility + apologize + outline how it won't happen again. OpenAI instead is like 'we…

DGX agent

When a normal company does a bad thing, they take responsibility + apologize + outline how it won't happen again. OpenAI instead is like 'we are entering a new era. this will happen again. no one amon

safetygary-marcus--x
25 Jul 2026
Safety

A Framework for Reputation Aware Uninorm-driven Consensus Algorithms for Blockchain Networks

DGX agent

arXiv:2607.20700v1 Announce Type: cross Abstract: The operation of blockchain is governed by consensus algorithms (CA). Several consensus mechanisms require significant computational power, while othe

safetyarxiv-cs-ai
24 Jul 2026
Safety

A Self-Evolving Default Action for Cooperative Tasks with Continuous Action Space

DGX agent

arXiv:2607.18597v2 Announce Type: replace Abstract: Counterfactual credit assignment has proven effective in multi-agent reinforcement learning (MARL) for discrete action spaces, yet its extension to

safetyarxiv-cs-lg
24 Jul 2026
Safety

A Unified Moral-Value Dataset for Instruction Tuning

DGX agent

arXiv:2607.21279v1 Announce Type: new Abstract: Large language models (LLMs) have developed rapidly and become valuable tools in everyday life. However, how to align LLMs to a particular set of human

safetyarxiv-cs-cl
24 Jul 2026
Safety

Adaptive Depth Sparse Framework: Similarity-Driven Resource Allocation for Pre-Trained LLMs

DGX agent

arXiv:2607.21291v1 Announce Type: new Abstract: Large language models (LLMs) achieve strong generation and reasoning performance, but the Transformer architecture incurs high inference cost. Existing

safetyarxiv-cs-cl
24 Jul 2026
Safety

Anticipate Before Acting: Future-State-Conditioned Vision-Language Navigation

DGX agent

arXiv:2607.18042v2 Announce Type: replace-cross Abstract: End-to-end vision-language navigation (VLN) with causal vision-language models maps instructions and egocentric observations directly to actio

safetyarxiv-cs-ai
24 Jul 2026
Safety

Approximate Quantum State Preparation Through Proximal Policy Optimization

DGX agent

arXiv:2607.21121v1 Announce Type: cross Abstract: In this work, a quantum architecture search framework for approximate quantum state preparation (QSP) is proposed. QSP is a challenging task, since th

safetyarxiv-cs-lg
24 Jul 2026
Safety

ARCO: Adaptive Rubrics with Co-Evolution for Multi-Step LLM-Based Agents

DGX agent

arXiv:2606.21262v2 Announce Type: replace Abstract: Reinforcement learning for multi-step LLM agents often relies on scalar rewards that indicate success but cannot explain why a trajectory is good or

safetyarxiv-cs-ai
24 Jul 2026
Safety

ASTRA-Net: Anatomy-Specific Transfer and Representation Alignment for Drug-Induced Sleep Endoscopy Segmentation

DGX agent

arXiv:2607.21370v1 Announce Type: new Abstract: Quantitative drug-induced sleep endoscopy (DISE) requires reliable airway boundaries at specific anatomical levels. Pixel-level DISE annotations are sca

safetyarxiv-cs-cv
24 Jul 2026
Safety

AttriMem: Attribution-Guided Process Feedback for Agent Memory Learning

DGX agent

arXiv:2607.21106v1 Announce Type: new Abstract: Effective memory is crucial for LLM agents, yet constructing it effectively remains challenging. A memory-construction policy decides what information t

safetyarxiv-cs-ai
24 Jul 2026
Safety

Axolotl3D: a Unified Framework for Faithful 3D Shape Completion

DGX agent

arXiv:2607.20660v1 Announce Type: new Abstract: Recent 3D generative models produce high-quality geometry from a single image using large-scale priors and diffusion architectures. However, they assume

safetyarxiv-cs-cv
24 Jul 2026
Safety

Belief Propagation in LLM World Models: Measuring Strategic Information Bias with Prediction Markets

DGX agent

arXiv:2607.20441v1 Announce Type: new Abstract: Every information ecosystem produces beliefs that shape strategic decisions. Both human analysts and AI systems inherit the blind spots of their informa

safetyarxiv-cs-cl
24 Jul 2026
Safety

Beyond Sycophancy: Structured Resistance and Compliance in LLM Moral Reasoning

DGX agent

arXiv:2607.21558v1 Announce Type: new Abstract: Building socially calibrated large language models, which can learn from others without simply yielding to them, requires more than reducing sycophancy

safetyarxiv-cs-ai
24 Jul 2026
Safety

Climate-resilient electric vehicle charging infrastructure for sustainable cities: An interpretable causal-ensemble framework for preventive maintenance and low-carbon mobility

DGX agent

arXiv:2607.21444v1 Announce Type: cross Abstract: Reliable electric vehicle (EV) charging infrastructure is a cornerstone of sustainable, low-carbon cities, yet urban climate stress such as extreme he

safetyarxiv-cs-lg
24 Jul 2026
Safety

Confidently Deceptive: How Confidence Amplifies the Risk of LLM Deception

DGX agent

arXiv:2607.20444v1 Announce Type: cross Abstract: Large language models (LLMs) can produce deceptive responses: outputs that mislead users in service of a contextually or experimentally induced goal.

safetyarxiv-cs-ai
24 Jul 2026
Safety

CSPF: A Constrained Shared-Private Fusion Method for Non-Verifiable Preference Evaluation

DGX agent

arXiv:2607.20862v1 Announce Type: new Abstract: At present, reliable evaluation of non-verifiable tasks remains challenging. Existing approaches often fail to adequately capture the diverse evaluative

safetyarxiv-cs-cl
24 Jul 2026
Safety

DINOde: Continuous Vision-Text Alignment for Open-Vocabulary Semantic Segmentation

DGX agent

arXiv:2607.21371v1 Announce Type: cross Abstract: Open-vocabulary semantic segmentation (OVSS) leverages textual semantics to segment objects beyond predefined categories. While the self-supervised mo

safetyarxiv-cs-ai
24 Jul 2026
Safety

Distribution-Alignment Bridge for Uncertainty-Aware Text-to-Video Retrieval

DGX agent

arXiv:2607.20984v1 Announce Type: new Abstract: This paper proposes the Distribution-Alignment Bridge (DAB), a framework that reconceptualizes text-to-video retrieval as a distribution alignment task

safetyarxiv-cs-cv
24 Jul 2026
Safety

DTIF: Robust Loop Closure Detection via Delaunay Triangle Topology in Complex Forests

DGX agent

arXiv:2607.21138v1 Announce Type: new Abstract: Accurate forest inventory and large-scale mapping are essential for ecosystem monitoring and sustainable forest management. Multiple low-cost edge platf

safetyarxiv-cs-cv
24 Jul 2026
Safety

DynaMark: A Reinforcement Learning Framework for Dynamic Watermarking in Industrial Machine Tool Controllers

DGX agent

arXiv:2508.21797v2 Announce Type: replace-cross Abstract: Industry 4.0's highly networked Machine Tool Controllers (MTCs) are prime targets for replay attacks that use outdated sensor data to manipula

safetyarxiv-cs-ai
24 Jul 2026
Safety

EmoAgent-R1: Towards Multimodal Emotion Understanding with Reinforcement Learning-based Dynamic Agent Specialization

DGX agent

arXiv:2607.21013v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have achieved impressive performance in multimodal emotion recognition (MER) tasks and lifted MER to a new leve

safetyarxiv-cs-ai
24 Jul 2026
Safety

EmoSpace: Immersive Affective Image Generation Guided by Fine-Grained Emotion Prototypes

DGX agent

arXiv:2602.11658v2 Announce Type: replace Abstract: Immersive affective content generation aims to create visually compelling VR imagery with controllable emotional nuance, yet existing methods typica

safetyarxiv-cs-cv
24 Jul 2026
Safety

Enhancing Explainable Cardiac Diagnosis with Guide-Grounded Multimodal LLMs

DGX agent

arXiv:2607.20814v1 Announce Type: new Abstract: The electrocardiogram (ECG) is a cornerstone of cardiac as- sessment, yet clinical deployment of deep learning models remains con- strained by limited i

safetyarxiv-cs-ai
24 Jul 2026
Safety

Environment-Aware Channel Inference via Cross-Modal Flow: From Multimodal Sensing to Wireless Channels

DGX agent

arXiv:2512.04966v2 Announce Type: replace-cross Abstract: Accurate channel state information (CSI) underpins reliable and efficient wireless communication. However, acquiring CSI via pilot estimation

safetyarxiv-cs-lg
24 Jul 2026
Safety

Evaluating Risks in Weak-to-Strong Alignment: A Bias-Variance Perspective

DGX agent

arXiv:2604.25077v2 Announce Type: replace Abstract: Weak-to-strong alignment offers a promising route to scalable supervision, but it can fail when a strong model becomes confidently wrong on examples

safetyarxiv-cs-ai
24 Jul 2026
Safety

EvoSQL: Memory-Augmented Critic-Generator Co-Evolution for Text-to-SQL

DGX agent

arXiv:2607.20489v1 Announce Type: new Abstract: Text-to-SQL has advanced rapidly with large language models, but complex database queries still require reasoning beyond one-shot generation, including

safetyarxiv-cs-ai
24 Jul 2026
Safety

Expert Behavior Prior Reinforcement Learning

DGX agent

arXiv:2607.21302v1 Announce Type: new Abstract: Behavior prior reinforcement learning (BPRL) has emerged as a promising paradigm to improve sample efficiency in online reinforcement learning (RL) by l

safetyarxiv-cs-ai
24 Jul 2026
Safety

Explainability Framework for Policy-Aware Autonomous Agents

DGX agent

arXiv:2607.21209v1 Announce Type: cross Abstract: In the field of Artificial Intelligence, an agent is a system which is able to autonomously make decisions in order to reach a desired goal. As these

safetyarxiv-cs-ai
24 Jul 2026
Safety

Explainable graph attention network for stress recognition (StressGAT) via differential action units

DGX agent

arXiv:2607.20819v1 Announce Type: new Abstract: Stress is a dynamic process characterized by significant individual variability in facial expression. Traditional architectures, such as Recurrent Neura

safetyarxiv-cs-cv
24 Jul 2026
← Previous
1…9495969798…302
Next →