AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
Human
86,428Total entries
1Added by human
86,427Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
61,498 results
13 May 2026

Adaptive TD-Lambda for Cooperative Multi-agent Reinforcement Learning

SafetyDGX agent

arXiv:2605.11880v1 Announce Type: new Abstract: TD(lambda) in value-based MARL algorithms or the Temporal Difference critic learning in Actor-Critic-based (AC-based) algorithms synergistically integra

Adaptive Teacher Exposure for Self-Distillation in LLM Reasoning

SafetyDGX agent

arXiv:2605.11458v1 Announce Type: cross Abstract: On-policy self-distillation has become a strong recipe for LLM reasoning, where a privileged teacher supervises the student's own rollouts while condi

ADMM-Q: An Improved Hessian-based Weight Quantizer for Post-Training Quantization of Large Language Models

Local AiDGX agent

arXiv:2605.11222v1 Announce Type: new Abstract: Quantization is an effective strategy to reduce the storage and computation footprint of large language models (LLMs). Post-training quantization (PTQ)

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Adversarial Causal Tuning for Realistic Time-series Generation

ResearchDGX agent

arXiv:2506.02084v2 Announce Type: replace Abstract: We address the problem of generating simulated, yet realistic, time-series data from a causal model with the same observational and interventional d

Adversarial Effects on Expressibility and Trainability in Distributed Variational Quantum Algorithms

ResearchDGX agent

arXiv:2605.03629v1 Announce Type: cross Abstract: Distributed quantum algorithms offer a promising pathway to scale variational quantum algorithms beyond the constraints of noisy intermediate-scale qu

AESOP: Adversarial Execution-path Selection to Overload Deep Learning Pipelines

ApplicationsDGX agent

arXiv:2605.10987v1 Announce Type: new Abstract: Modern machine learning deployments increasingly compose specialized models into dynamic inference pipelines, where upstream components produce intermed

Agent-Based Post-Hoc Correction of Agricultural Yield Forecasts

Model ReleasesDGX agent

arXiv:2605.12375v1 Announce Type: new Abstract: Accurate crop yield forecasting in commercial soft fruit production is constrained by the data available in typical commercial farm records, which lack

Agent-BRACE: Decoupling Beliefs from Actions in Long-Horizon Tasks via Verbalized State Uncertainty

Model ReleasesDGX agent

arXiv:2605.11436v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed on long-horizon tasks in partially observable environments, where they must act while inferring a

AgentDisCo: Towards Disentanglement and Collaboration in Open-ended Deep Research Agents

Model ReleasesDGX agent

arXiv:2605.11732v1 Announce Type: cross Abstract: In this paper, we present AgentDisCo, a novel Disentangled and Collaborative agentic architecture that formulates deep research as an adversarial opti

AgentShield: Deception-based Compromise Detection for Tool-using LLM Agents

AgentsDGX agent

arXiv:2605.11026v1 Announce Type: cross Abstract: Defenses against indirect prompt injection (IPI) in tool-using LLM agents share two structural weaknesses. First, they all attempt to prevent attacks

AIA: Rethinking Architecture Decoupling Strategy In Unified Multimodal Model

SafetyDGX agent

arXiv:2511.22663v5 Announce Type: replace Abstract: Unified multimodal models for image generation and understanding represent a significant step toward AGI and have attracted widespread attention fro

Aligning Flow Map Policies with Optimal Q-Guidance

SafetyDGX agent

arXiv:2605.12416v1 Announce Type: new Abstract: Generative policies based on expressive model classes, such as diffusion and flow matching, are well-suited to complex control problems with highly mult

Allegory of the Cave: Measurement-Grounded Vision-Language Learning

Model ReleasesDGX agent

arXiv:2605.11727v1 Announce Type: cross Abstract: Vision-language models typically reason over post-ISP RGB images, although RGB rendering can clip, suppress, or quantize sensor evidence before infere

AlphaEarth Satellite Embeddings for Modelling Climate Sensitive Diseases Towards Global Health Resilience

ResearchDGX agent

arXiv:2605.10949v1 Announce Type: cross Abstract: Malaria, childhood acute respiratory infection, and child undernutrition together account for over two million deaths annually in children under five,

AlphaGRPO: Unlocking Self-Reflective Multimodal Generation in UMMs via Decompositional Verifiable Reward

SafetyDGX agent

arXiv:2605.12495v1 Announce Type: new Abstract: In this paper, we propose AlphaGRPO, a novel framework that applies Group Relative Policy Optimization (GRPO) to AR-Diffusion Unified Multimodal Models

An Empirical Study of Automating Agent Evaluation

Model ReleasesDGX agent

arXiv:2605.11378v1 Announce Type: new Abstract: Agent evaluation requires assessing complex multi-step behaviors involving tool use and intermediate reasoning, making it costly and expertise-intensive

Analytical Provisioning for Attention-FFN Disaggregated LLM Serving under Stochastic Workloads

ResearchDGX agent

arXiv:2601.21351v3 Announce Type: replace Abstract: Attentio-FFN disaggregation (AFD) is an emerging architecture for LLM decoding that separates state-heavy, KV-cache-dominated Attention computation

Anomaly-Aware Vision-Language Adapters for Zero-Shot Anomaly Detection

ResearchDGX agent

arXiv:2605.12069v1 Announce Type: new Abstract: Zero-shot anomaly detection aims to identify defects in unseen categories without target-specific training. Existing methods usually apply the same feat

Anti-Self-Distillation for Reasoning RL via Pointwise Mutual Information

SafetyDGX agent

arXiv:2605.11609v1 Announce Type: cross Abstract: On-policy self-distillation, where a student is pulled toward a copy of itself conditioned on privileged context (e.g., a verified solution or feedbac

AOI-SSL: Self-Supervised Framework for Efficient Segmentation of Wire-bonded Semiconductors In Optical Inspection

ResearchDGX agent

arXiv:2605.12430v1 Announce Type: new Abstract: Segmentation models in automated optical inspection of wire-bonded semiconductors are typically device-specific and must be re-trained when new devices

Approximating Simple ReLU Networks based on Spectral Decomposition of Fisher Information

ResearchDGX agent

arXiv:2505.17907v2 Announce Type: replace-cross Abstract: Properties of Fisher information matrices of 2-layer neural ReLU networks with random hidden weights are studied. For these networks, it is kn

Approximation of Maximally Monotone Operators : A Graph Convergence Perspective

ResearchDGX agent

arXiv:2605.12301v1 Announce Type: new Abstract: Operator learning has been highly successful for continuous mappings between infinite-dimensional spaces, such as PDE solution operators. However, many

Approximation Theory of Laplacian-Based Neural Operators for Reaction-Diffusion System

Model ReleasesDGX agent

arXiv:2605.12025v1 Announce Type: new Abstract: Neural operators provide a framework for learning solution operators of partial differential equations (PDEs), enabling efficient surrogate modeling for

Arbitrated Indirect Treatment Comparisons

ResearchDGX agent

arXiv:2510.18071v2 Announce Type: replace-cross Abstract: Matching-adjusted indirect comparison (MAIC) has been increasingly employed in health technology assessments (HTA). By reweighting subjects fr

arepsilon-Good Action Identification in Fixed-Budget Monte Carlo Tree Search

ResearchDGX agent

arXiv:2605.11324v1 Announce Type: new Abstract: We study the fixed-budget max-min action identification problem in depth-2 max-min trees, an important special case of Monte Carlo Tree Search. A learne

ASD-Bench: A Four-Axis Comprehensive Benchmark of AI Models for Autism Spectrum Disorder

Model ReleasesDGX agent

arXiv:2605.11091v1 Announce Type: new Abstract: Automated ASD screening tools remain limited by single-architecture evaluations, axis-restricted assessment, and near-exclusive focus on adult cohorts,

ASIP-Planner: Adaptive Planning for UAV Surface Inspection in Partially Known Indoor Environments

ApplicationsDGX agent

arXiv:2605.11119v1 Announce Type: new Abstract: Indoor infrastructure inspection, such as tunnels and industrial facilities, requires systematic surface coverage to ensure that all inspection targets

Assessment of cloud and associated radiation fields from a GAN stochastic cloud subcolumn generator

SafetyDGX agent

arXiv:2605.11968v1 Announce Type: cross Abstract: Modern Earth System Models (ESMs) operate on horizontal scales far larger than typical cloud features, requiring stochastic subcolumn generators to re

Asymmetric Advantage Modulation Calibrates Entropy Dynamics in RLVR

SafetyDGX agent

arXiv:2604.04894v2 Announce Type: replace Abstract: Reinforcement learning with verifiable rewards (RLVR) has substantially improved the reasoning ability of large language models (LLMs), but it often

Attacks and Mitigations for Distributed Governance of Agentic AI under Byzantine Adversaries

AgentsDGX agent

arXiv:2605.12364v1 Announce Type: cross Abstract: Agentic AI governance is a critical component of agentic AI infrastructure ensuring that agents follow their owner's communication and interaction pol

Audio-Visual Camera Pose Estimation with Passive Scene Sounds and In-the-Wild Video

ApplicationsDGX agent

arXiv:2512.12165v3 Announce Type: replace Abstract: Understanding camera motion is a fundamental problem in embodied perception and 3D scene understanding. While visual methods have advanced rapidly,

Augmented Lagrangian Method for Last-Iterate Convergence for Constrained MDPs

SafetyDGX agent

arXiv:2605.11694v1 Announce Type: new Abstract: We study policy optimization for infinite-horizon, discounted constrained Markov decision processes (CMDPs). While existing theoretical guarantees typic

AutoLLMResearch: Training Research Agents for Automating LLM Experiment Configuration -- Learning from Cheap, Optimizing Expensive

HardwareDGX agent

arXiv:2605.11518v1 Announce Type: cross Abstract: Effectively configuring scalable large language model (LLM) experiments, spanning architecture design, hyperparameter tuning, and beyond, is crucial f

AutoMonitor-Bench: Evaluating the Reliability of LLM-Based Misbehavior Monitor

Model ReleasesDGX agent

arXiv:2601.05752v3 Announce Type: replace Abstract: We introduce AutoMonitor-Bench, the first benchmark designed to systematically evaluate the reliability of LLM-based misbehavior monitors across div

Autoregressive Learning in Joint KL: Sharp Oracle Bounds and Lower Bounds

SafetyDGX agent

arXiv:2605.12316v1 Announce Type: new Abstract: We study the fundamental and timely problem of learning long sequences in autoregressive modeling and next-token prediction under model misspecification

B4DL: A Benchmark for 4D LiDAR LLM in Spatio-Temporal Understanding

Model ReleasesDGX agent

arXiv:2508.05269v2 Announce Type: replace Abstract: Understanding dynamic outdoor environments requires capturing complex object interactions and their evolution over time. LiDAR-based 4D point clouds

Backbone-Equated Diffusion OOD via Sparse Internal Snapshots

Model ReleasesDGX agent

arXiv:2605.11014v1 Announce Type: new Abstract: Fair comparison between diffusion-based OOD detectors is challenging, as conclusions can vary with backbone choice, corruption parameterization, and tes

BARISTA: A Multi-Task Egocentric Benchmark for Compositional Visual Understanding

Model ReleasesDGX agent

arXiv:2605.12074v1 Announce Type: new Abstract: Scene understanding is central to general physical intelligence, and video is a primary modality for capturing both state and temporal dynamics of a sce

Bayesian Surrogate Training on Multiple Data Sources: A Hybrid Modeling Strategy

ApplicationsDGX agent

arXiv:2412.11875v3 Announce Type: replace-cross Abstract: Surrogate models are often used as computationally efficient approximations to complex simulation models, enabling tasks such as solving inver

BEExformer: A Fast Inferencing Binarized Transformer with Early Exits

Model ReleasesDGX agent

arXiv:2412.05225v3 Announce Type: replace Abstract: Large Language Models (LLMs) based on transformers achieve cutting-edge results on a variety of applications. However, their enormous size and proce

Behavioral Mode Discovery for Fine-tuning Multimodal Generative Policies

ResearchDGX agent

arXiv:2605.11387v1 Announce Type: new Abstract: We address the problem of fine-tuning pre-trained generative policies with reinforcement learning (RL) while preserving the multimodality of their actio

Beyond GRPO and On-Policy Distillation: An Empirical Sparse-to-Dense Reward Principle for Language-Model Post-Training

Model ReleasesDGX agent

arXiv:2605.12483v1 Announce Type: new Abstract: In settings where labeled verifiable training data is the binding constraint, each checked example should be allocated carefully. The standard practice

Beyond Localization: A Comprehensive Diagnosis of Perspective-Conditioned Spatial Reasoning in MLLMs from Omnidirectional Images

Model ReleasesDGX agent

arXiv:2605.12413v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) show strong visual perception, yet remain limited in reasoning about space under changing viewpoints. We study

Beyond Manual Curation: Augmenting Targeted Protein Degradation Databases via Agentic Literature Extraction Workflows

AgentsDGX agent

arXiv:2605.11221v1 Announce Type: cross Abstract: Predictive models in biomedicine depend on structured assay data locked in the text, tables, and supplements of primary publications. This bottleneck

Beyond Masks: The Case for Medical Image Parsing

ResearchDGX agent

arXiv:2605.11438v1 Announce Type: new Abstract: Medical imaging research has spent a decade getting very good at one thing: producing per-voxel masks. Masks tell us size, volume, and location, and a d

Beyond Parameter Aggregation: Semantic Consensus for Federated Fine-Tuning of LLMs

Model ReleasesDGX agent

arXiv:2605.11857v1 Announce Type: new Abstract: Federated fine-tuning of large language models is commonly formulated as a parameter aggregation problem. However, even parameter-efficient methods requ

Beyond Point Estimates: Distributional Uncertainty in Machine Learning Performance Evaluation

ResearchDGX agent

arXiv:2501.16931v2 Announce Type: replace Abstract: Machine learning models are often evaluated using point estimates of performance metrics such as accuracy, F1 score, or mean squared error. Such sum

Beyond Point-wise Neural Collapse: A Topology-Aware Hierarchical Classifier for Class-Incremental Learning

SafetyDGX agent

arXiv:2605.11904v1 Announce Type: new Abstract: The Nearest Class Mean (NCM) classifier is widely favored in Class-Incremental Learning (CIL) for its superior resistance to catastrophic forgetting com

Beyond Prediction: Interval Neural Networks for Uncertainty-Aware System Identification

Model ReleasesDGX agent

arXiv:2605.11460v1 Announce Type: new Abstract: System identification (SysID) is critical for modeling dynamical systems from experimental data, yet traditional approaches often fail to capture nonlin

Beyond Similarity: Temporal Operator Attention for Time Series Analysis

ResearchDGX agent

arXiv:2605.11287v1 Announce Type: new Abstract: A persistent paradox in time-series forecasting is that structurally simple MLP and linear models often outperform high-capacity Transformers. We argue

Beyond Text Prompts: Visual-to-Visual Generation as A Unified Paradigm

Model ReleasesDGX agent

arXiv:2605.12271v1 Announce Type: new Abstract: Humans often specify and create through visual artifacts: typography sheets, sketches, reference images, and annotated scenes. Yet modern visual generat

Bin Latent Transformer (BiLT): A shift-invariant autoencoder for calibration-free spectral unmixing of turbid media

Model ReleasesDGX agent

arXiv:2605.11829v1 Announce Type: cross Abstract: The accurate recovery of constituent-level optical properties from integrating sphere measurements is a central analytical challenge in pharmaceutical

Birds of a Feather Flock Together: Background-Invariant Representations via Linear Structure in VLMs

ApplicationsDGX agent

arXiv:2605.11107v1 Announce Type: new Abstract: Vision-language models (VLMs), such as CLIP and SigLIP 2, are widely used for image classification, yet their vision encoders remain vulnerable to syste

BitLM: Unlocking Multi-Token Language Generation with Bitwise Continuous Diffusion

ResearchDGX agent

arXiv:2605.11577v1 Announce Type: new Abstract: Autoregressive language models generate text one token at a time, yet natural language is inherently structured in multi-token units, including phrases,

BLOCK-EM: Preventing Emergent Misalignment via Latent Blocking

ResearchDGX agent

arXiv:2602.00767v2 Announce Type: replace Abstract: Emergent misalignment can arise when a language model is fine-tuned on a narrowly scoped supervised objective: the model learns the target behavior,

Block-R1: Rethinking the Role of Block Size in Multi-domain Reinforcement Learning for Diffusion Large Language Models

Model ReleasesDGX agent

arXiv:2605.11726v1 Announce Type: new Abstract: Recently, reinforcement learning (RL) has been widely applied during post-training for diffusion large language models (dLLMs) to enhance reasoning with

BOOST: BOttleneck-Optimized Scalable Training Framework for Low-Rank Large Language Models

HardwareDGX agent

arXiv:2512.12131v2 Announce Type: replace Abstract: The scale of transformer model pre-training is constrained by the increasing computation and communication cost. Low-rank bottleneck architectures o

Boosting Omni-Modal Language Models: Staged Post-Training with Visually Debiased Evaluation

Model ReleasesDGX agent

arXiv:2605.12034v1 Announce Type: cross Abstract: Omni-modal language models are intended to jointly understand audio, visual inputs, and language, but benchmark gains can be inflated when visual evid

Breaking Down and Building Up: Mixture of Skill-Based Vision-and-Language Navigation Agents

Model ReleasesDGX agent

arXiv:2508.07642v3 Announce Type: replace-cross Abstract: Vision-and-Language Navigation (VLN) poses significant challenges for agents to interpret natural language instructions and navigate complex 3

Breaking extit{Winner-Takes-All}: Cooperative Policy Optimization Improves Diverse LLM Reasoning

Model ReleasesDGX agent

arXiv:2605.11461v1 Announce Type: cross Abstract: Reinforcement learning with verifiers (RLVR) has become a central paradigm for improving LLM reasoning, yet popular group-based optimization algorithm

← Previous
1…710711712713714…1025
Next →