AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,478 results
12 May 2026

Beyond Position Bias: Shifting Context Compression from Position-Driven to Semantic-Driven

SafetyDGX agent

arXiv:2605.09463v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated exceptional performance across diverse tasks. However, their deployment in long-context scenarios faces h

Beyond Static Bias: Adaptive Multi-Fidelity Bandits with Improving Proxies

SafetyDGX agent

arXiv:2605.08558v1 Announce Type: new Abstract: As an extension of the classical multi-armed bandit problem, multi-fidelity multi-armed bandits (MF-MAB) enable individual arms to be evaluated using di

Bi-CoG: Bi-Consistency-Guided Self-Training for Vision-Language Models

SafetyDGX agent

arXiv:2510.20477v2 Announce Type: replace Abstract: Exploiting unlabeled data through semi-supervised learning (SSL) or leveraging pre-trained models via fine-tuning are two prevailing paradigms for a

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Bias by Necessity: Impossibility Theorems for Sequential Processing with Convergent AI and Human Validation

SafetyDGX agent

arXiv:2605.08716v1 Announce Type: new Abstract: Are certain cognitive biases mathematically inevitable consequences of sequential information processing? We prove that primacy effects, anchoring, and

Big AI is accelerating the metacrisis: What can we do?

SafetyDGX agent

arXiv:2512.24863v2 Announce Type: replace-cross Abstract: The world is in the grip of ecological, meaning, and language crises that are converging into a metacrisis. Big AI is accelerating them all. L

Biological Plausibility and Representational Alignment of Feedback Alignment in Convolutional Networks

SafetyDGX agent

arXiv:2605.08564v1 Announce Type: new Abstract: The feedback alignment (FA) algorithm offers a biologically plausible alternative to backpropagation (BP) for training neural networks yet notably fails

Block-Wise Differentiable Sinkhorn Attention: Tail-Refinement Gradients with a Gap-Aware Dustbin Bridge

SafetyDGX agent

arXiv:2605.08123v1 Announce Type: cross Abstract: We study long-context balanced entropic optimal transport (OT) attention on TPU hardware through a stopped-base, fixed-depth tail-refinement surrogate

BONSAI: Bayesian Optimization with Natural Simplicity and Interpretability

SafetyDGX agent

arXiv:2602.07144v2 Announce Type: replace-cross Abstract: Bayesian optimization (BO) is a popular technique for sample-efficient optimization of black-box functions. In many applications, the paramete

Breaking. Sam Altman himself finally confirms what I was the first to point out publicly: his indirect equity stake in OpenAI. … which he fa…

SafetyDGX agent

Breaking. Sam Altman himself finally confirms what I was the first to point out publicly: his indirect equity stake in OpenAI. … which he failed to acknowledge when asked about his financial interest

Breaking the Grid: Distance-Guided Reinforcement Learning in Large Discrete Action Spaces

SafetyDGX agent

arXiv:2602.08616v2 Announce Type: replace-cross Abstract: Reinforcement Learning (RL) is increasingly applied to large-scale decision-making problems like logistics, scheduling, and recommender system

Breaking the Impasse: Dual-Scale Evolutionary Policy Training for Social Language Agents

SafetyDGX agent

arXiv:2605.08721v1 Announce Type: new Abstract: While Reinforcement Learning with Verifiable Rewards (RLVR) has proven effective for closed-ended tasks, extending it to open-ended social language game

BRIDGE: Building Representations In Domain Guided Program Synthesis

SafetyDGX agent

arXiv:2511.21104v3 Announce Type: replace Abstract: Large language models can generate plausible code, but remain brittle for formal verification in proof assistants such as Lean. A central scalabilit

BROS: Bias-Corrected Randomized Subspaces for Memory-Efficient Single-Loop Bilevel Optimization

SafetyDGX agent

arXiv:2605.10288v1 Announce Type: new Abstract: Stochastic bilevel optimization (SBO) has become a standard framework for hyperparameter learning, data reweighting, representation learning, and data-m

Budget-Efficient Automatic Algorithm Design via Code Graph

SafetyDGX agent

arXiv:2605.10598v1 Announce Type: new Abstract: Large language models (LLMs) have emerged as powerful tools for automatic algorithm design (AAD). However, existing pipelines remain inefficient. They o

CalBench: Evaluating Coordination-Privacy Trade-offs in Multi-Agent LLMs

SafetyDGX agent

arXiv:2605.09823v1 Announce Type: cross Abstract: We introduce CalBench, a controlled evaluation environment for studying multi-agent coordination through calendar scheduling. In CalBench, N agents ea

CAMAL: Improving Attention Alignment and Faithfulness with Segmentation Masks

SafetyDGX agent

arXiv:2605.08325v1 Announce Type: cross Abstract: Many vision datasets now provide segmentation masks in addition to annotated images to support a wide range of tasks. In this work, we propose Class A

Can LLMs Estimate Student Struggles? Human-AI Difficulty Alignment with Proficiency Simulation for Item Difficulty Prediction

SafetyDGX agent

arXiv:2512.18880v2 Announce Type: replace-cross Abstract: Accurate estimation of item (question or task) difficulty is critical for educational assessment but suffers from the cold start problem. Whil

Can Revealed Preferences Clarify LLM Alignment and Steering?

SafetyDGX agent

arXiv:2605.08556v1 Announce Type: new Abstract: LLMs are increasingly used to make or support high-stakes decisions under uncertainty, where alignment depends not only on factual accuracy but on how m

CapCLIP: A Vision-Language Representation Alignment Approach for Wireless Capsule Endoscopy Analysis

SafetyDGX agent

arXiv:2605.08493v1 Announce Type: new Abstract: Wireless capsule endoscopy (WCE) enables non-invasive visual assessment of the small bowel, but its clinical utility is constrained by the large volume

CARL: Criticality-Aware Agentic Reinforcement Learning

SafetyDGX agent

arXiv:2512.04949v3 Announce Type: replace-cross Abstract: Agents capable of accomplishing complex tasks through multiple interactions with the environment have emerged as a popular research direction.

CFSPMNet: Cross-subject Fourier-guided Spatial-Patch Mamba Network for EEG Motor Imagery Decoding in Stroke Patients

SafetyDGX agent

arXiv:2605.10111v1 Announce Type: cross Abstract: Motor imagery electroencephalography (MI-EEG) decoding offers a non-invasive route for post-stroke rehabilitation, but cross-patient use remains diffi

Change My View? The Dynamics of Persuasion and Polarization in Online Discourse

SafetyDGX agent

arXiv:2605.08383v1 Announce Type: new Abstract: Philosophical accounts of persuasion often assume that shared evidence and rational argumentation should lead to a convergence of views between peers, y

CoLA-Flow Policy: Temporally Coherent Imitation Learning via Continuous Latent Action Flow Matching for Robotic Manipulation

SafetyDGX agent

arXiv:2601.23087v3 Announce Type: replace Abstract: Learning long-horizon robotic manipulation requires jointly achieving expressive behavior modeling, real-time inference, and stable execution, which

Composing Policy Gradients and Prompt Optimization for Language Model Programs

SafetyDGX agent

arXiv:2508.04660v2 Announce Type: replace Abstract: Group Relative Policy Optimization (GRPO) has proven to be an effective tool for post-training language models (LMs). However, AI systems are increa

Compute as Teacher: Turning Inference Compute Into Reference-Free Supervision

SafetyDGX agent

arXiv:2509.14234v3 Announce Type: replace Abstract: Where do learning signals come from when there is no ground truth in post-training? We show that inference compute itself can serve as supervision.

Compute Where it Counts: Self Optimizing Language Models

SafetyDGX agent

arXiv:2605.10875v1 Announce Type: cross Abstract: Efficient LLM inference research has largely focused on reducing the cost of each decoding step (e.g., using quantization, pruning, or sparse attentio

Continuity Laws for Sequential Models

SafetyDGX agent

arXiv:2605.08539v1 Announce Type: cross Abstract: Inductive biases influence the behavior and performance of sequential models. In this work, we study an underexplored inductive bias in sequential mod

Control-Augmented Autoregressive Diffusion for Data Assimilation

SafetyDGX agent

arXiv:2510.06637v3 Announce Type: replace-cross Abstract: Despite advances in test-time scaling and diffusion finetuning, guidance for Auto-Regressive Diffusion Models (ARDMs) remains underexplored. W

Core-Halo Decomposition: Decentralizing Large-Scale Fixed-Point Problems

SafetyDGX agent

arXiv:2605.08681v1 Announce Type: cross Abstract: We study solving large-scale fixed-point equation (x^star=ar F(x^star)) with decomposition. Standard strict decomposition assigns each agent a disjoin

Cornerstones or Stumbling Blocks? Deciphering the Rock Tokens in On-Policy Distillation

SafetyDGX agent

arXiv:2605.09253v1 Announce Type: cross Abstract: While recent work in Reinforcement Learning with Verifiable Rewards (RLVR) has shown that a small subset of critical tokens disproportionately drives

Crosslingual On-Policy Self-Distillation for Multilingual Reasoning

SafetyDGX agent

arXiv:2605.09548v1 Announce Type: new Abstract: Large language models (LLMs) have achieved remarkable progress in mathematical reasoning, but this ability is not equally accessible across languages. E

Crowding Out The Noise: Algorithmic Collective Action Under Differential Privacy

SafetyDGX agent

arXiv:2505.05707v2 Announce Type: replace Abstract: The integration of AI into daily life has generated considerable attention and excitement, while also raising concerns about automating algorithmic

DAP: Doppler-aware Point Network for Heterogeneous mmWave Action Recognition

SafetyDGX agent

arXiv:2605.09604v1 Announce Type: new Abstract: Millimeter-wave (mmWave) radar provides privacy-preserving sensing and is valuable for human action recognition (HAR). Existing mmWave point cloud datas

DAPE: Dynamic Non-uniform Alignment and Progressive Detail Enhancement Techniques for Improving the Performance of Efficient Visual Language Models

SafetyDGX agent

arXiv:2605.08902v1 Announce Type: cross Abstract: In recent years, pre-trained visual-linguistic models have demonstrated tremendous potential, becoming a crucial foundational framework for numerous d

DARE: Difficulty-Adaptive Reinforcement Learning with Co-Evolved Difficulty Estimation

SafetyDGX agent

arXiv:2605.09188v1 Announce Type: cross Abstract: Reinforcement learning improves the reasoning ability of large language models but remains costly and sample-inefficient, as many rollouts provide wea

Data-driven transport modelling without overfit

SafetyDGX agent

arXiv:2605.08801v1 Announce Type: new Abstract: Macroscopic transport modelling aims to predict traffic flows after proposed public policy interventions, such as a new road or railway section or a tem

Debugging the Debuggers: Failure-Anchored Structured Recovery for Software Engineering Agents

SafetyDGX agent

arXiv:2605.08717v1 Announce Type: cross Abstract: Software engineering agents are increasingly deployed in evaluable engineering environments, yet post-failure recovery remains costly, manual, and ad

Decoupling Endpoint and Semantic Transition Learning for Zero-Shot Composed Image Retrieval

SafetyDGX agent

arXiv:2605.08389v1 Announce Type: cross Abstract: Zero-shot composed image retrieval (ZS-CIR) retrieves a target image from a reference image and a text modification without human-annotated CIR triple

Dependency-Aware Discrete Diffusion for Scene Graph Generation

SafetyDGX agent

arXiv:2605.09065v1 Announce Type: new Abstract: Scene graphs (SGs) represent objects and their relationships as structured graphs, enabling applications in image generation, robotics, and 3D understan

DexWrist: A Robotic Wrist for Constrained and Dynamic Manipulation

SafetyDGX agent

arXiv:2507.01008v3 Announce Type: replace Abstract: Development of dexterous manipulation hardware has primarily focused on hands and grippers. However, these end-effectors are often paired with bulky

dFlowGRPO: Rate-Aware Policy Optimization for Discrete Flow Models

SafetyDGX agent

arXiv:2605.09291v1 Announce Type: new Abstract: Discrete flow models (DFMs) are a class of flexible generative models for generating discrete data, and diffusion large language models (dLLMs) can be v

DGPO: Beyond Pairwise Preferences with Directional Consistent Groupwise Optimization

SafetyDGX agent

arXiv:2605.10863v1 Announce Type: new Abstract: Although Large Language Models (LLMs) have made remarkable progress, current preference optimization methods still struggle to align directional consist

Disentangled Representation Learning via Flow Matching

SafetyDGX agent

arXiv:2602.05214v2 Announce Type: replace Abstract: Disentangled representation learning aims to capture the underlying explanatory factors of observed data, enabling a principled understanding of the

Do Linear Probes Generalize Better in Persona Coordinates?

SafetyDGX agent

arXiv:2605.09391v1 Announce Type: new Abstract: It is becoming increasingly necessary to have monitors check for harmful behaviors during language model interactions, but text-only monitoring has not

Drum Synthesis from Expressive Drum Grids via Neural Audio Codecs

SafetyDGX agent

arXiv:2605.10281v1 Announce Type: cross Abstract: Generating realistic drum audio directly from symbolic representations is a challenging task at the intersection of music perception and machine learn

DuetFair: Coupling Inter- and Intra-Subgroup Robustness for Fair Medical Image Segmentation

SafetyDGX agent

arXiv:2605.10521v1 Announce Type: cross Abstract: Medical image segmentation models can perform unevenly across subgroups. Most existing fairness methods focus on improving average subgroup performanc

Dynamic Skill Lifecycle Management for Agentic Reinforcement Learning

SafetyDGX agent

arXiv:2605.10923v1 Announce Type: cross Abstract: Large language model agents increasingly rely on external skills to solve complex tasks, where skills act as modular units that extend their capabilit

E-TCAV: Formalizing Penultimate Proxies for Efficient Concept Based Interpretability

SafetyDGX agent

arXiv:2605.10261v1 Announce Type: new Abstract: TCAV (Testing with Concept Activation Vectors) is an interpretability method that assesses the alignment between the internal representations of a train

Effective Explanations Support Planning Under Uncertainty

SafetyDGX agent

arXiv:2605.08406v1 Announce Type: cross Abstract: Explaining how to get from A to B can be challenging. It requires mentally simulating what the listener will do based on what they are told. To captur

EGL-SCA: Structural Credit Assignment for Co-Evolving Instructions and Tools in Graph Reasoning Agents

SafetyDGX agent

arXiv:2605.10366v1 Announce Type: new Abstract: Graph reasoning agents operating from natural-language inputs must solve a coupled problem: they must reconstruct a structured graph instance from text,

ElasticFlow: One-Step Physics-Consistent Policy with Elastic Time Horizons for Language-Guided Manipulation

SafetyDGX agent

arXiv:2605.08799v1 Announce Type: new Abstract: Diffusion policies have demonstrated exceptional performance in embodied AI. However, their iterative denoising process results in high latency, and exi

Emergence of Physical Intelligence via Controllable Information Production

SafetyDGX agent

arXiv:2601.22449v2 Announce Type: replace Abstract: Intrinsic Motivation (IM) aims to train agents without external rewards, enabling useful behavior to emerge from the agent's interaction with its en

Empowering VLMs for Few-Shot Multimodal Time Series Classification via Tailored Agentic Reasoning

SafetyDGX agent

arXiv:2605.09395v1 Announce Type: new Abstract: In this paper, we propose the first VLnderline{extbf{M}} nderline{extbf{a}}gentic nderline{extbf{r}}easoning framework for few-nderline{extbf{s}}hot mul

Equivariant Reinforcement Learning for Clifford Quantum Circuit Synthesis

SafetyDGX agent

arXiv:2605.10910v1 Announce Type: cross Abstract: We consider the problem of synthesizing Clifford quantum circuits for devices with all-to-all qubit connectivity. We approach this task as a reinforce

ETS: Energy-Guided Test-Time Scaling for Training-Free RL Alignment

SafetyDGX agent

arXiv:2601.21484v2 Announce Type: replace Abstract: Reinforcement Learning (RL) post-training alignment for language models is effective, but also costly and unstable in practice, owing to its complic

EvoMAS: Learning Execution-Time Workflows for Multi-Agent Systems

SafetyDGX agent

arXiv:2605.08769v1 Announce Type: new Abstract: Large language model (LLM)-based multi-agent systems have shown strong potential on complex tasks through agent specialization, tool use, and collaborat

EvoPref: Multi-Objective Evolutionary Optimization Discovers Diverse LLM Alignments Beyond Gradient Descent

SafetyDGX agent

arXiv:2605.09777v1 Announce Type: cross Abstract: Gradient-based preference optimization methods for large language model (LLM) alignment suffer from preference collapse, converging to narrow behavior

EvoStreaming: Your Offline Video Model Is a Natively Streaming Assistant

SafetyDGX agent

arXiv:2605.10343v1 Announce Type: cross Abstract: Streaming video understanding demands more than watching longer videos: assistants must decide when to speak in real time, balancing responsiveness ag

Execution Envelopes: A Shared Admission Contract for Backend AI Execution Requests

SafetyDGX agent

arXiv:2605.08267v1 Announce Type: cross Abstract: Enterprise AI backends increasingly admit heterogeneous execution requests across model deployment, inference, evaluation, data movement, and agentic

Explanation-Aware Learning for Enhanced Interpretability in Biomedical Imaging

SafetyDGX agent

arXiv:2605.10054v1 Announce Type: new Abstract: Deep neural networks for medical image diagnosis often achieve high predictive accuracy while relying on spurious or clinically irrelevant visual cues,

← Previous
1…173174175176177…242
Next →