AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
Human
85,188Total entries
1Added by human
85,187Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
60,292 results
11 May 2026

Relay Buffer Independent Communication over Pooled HBM for Efficient MoE Inference on Ascend

ResearchDGX agent

arXiv:2605.06055v2 Announce Type: replace-cross Abstract: Mixture-of-Experts (MoE) inference requires large-scale token exchange across devices, making dispatch and combine major bottlenecks in both p

Reliable Chain-of-Thought via Prefix Consistency

ResearchDGX agent

arXiv:2605.07654v1 Announce Type: cross Abstract: Large Language Models often improve accuracy on reasoning tasks by sampling multiple Chain-of-Thought (CoT) traces and aggregating them with majority

RELO: Reinforcement Learning to Localize for Visual Object Tracking

SafetyDGX agent

arXiv:2605.07379v1 Announce Type: cross Abstract: Conventional visual object trackers localize targets using handcrafted spatial priors, often in the form of heatmaps. Such priors provide only surroga

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Render, Don't Decode: Weight-Space World Models with Latent Structural Disentanglement

HardwareDGX agent

arXiv:2605.06298v2 Announce Type: replace-cross Abstract: Training world models on vast quantities of unlabelled videos is a critical step toward fully autonomous intelligence. However, the prevailing

Rep2Text: Decoding Full Text from a Single LLM Token Representation

Model ReleasesDGX agent

arXiv:2511.06571v3 Announce Type: replace-cross Abstract: Large language models (LLMs) have achieved remarkable progress across diverse tasks, yet their internal mechanisms remain largely opaque. In t

Repeated Deceptive Path Planning against Learnable Observer

SafetyDGX agent

arXiv:2605.07174v1 Announce Type: new Abstract: We study the problem of deceptive path planning (DPP), where an agent aims to conceal its true destination from external observers. While existing work

Replicating Human Motivated Reasoning Studies with LLMs

ResearchDGX agent

arXiv:2601.16130v2 Announce Type: replace-cross Abstract: Motivated reasoning - the idea that individuals processing information may be motivated to either arrive at accurate beliefs or arrive at desi

ReSeek: A Self-Correcting Framework for Search Agents with Instructive Rewards

Model ReleasesDGX agent

arXiv:2510.00568v3 Announce Type: replace Abstract: Search agents powered by Large Language Models (LLMs) have demonstrated significant potential in tackling knowledge-intensive tasks. Reinforcement l

Resource-Element Energy Difference for Noncoherent Over-the-Air Federated Learning

SafetyDGX agent

arXiv:2605.07263v1 Announce Type: cross Abstract: Over-the-air federated learning (OTA-FL) reduces uplink latency by exploiting waveform superposition, but conventional analog aggregation schemes typi

Response-G1: Explicit Scene Graph Modeling for Proactive Streaming Video Understanding

SafetyDGX agent

arXiv:2605.07575v1 Announce Type: cross Abstract: Proactive streaming video understanding requires Video-LLMs to decide when to respond as a video unfolds, a task where existing methods often fall sho

Response Time Enhances Alignment with Heterogeneous Preferences

SafetyDGX agent

arXiv:2605.06987v1 Announce Type: new Abstract: Aligning large language models (LLMs) to human preferences typically relies on aggregating pooled feedback into a single reward model. However, this sta

Rethinking Dense Optical Flow without Test-Time Scaling

Model ReleasesDGX agent

arXiv:2605.08000v1 Announce Type: new Abstract: Recent progress in dense optical flow has been driven by increasingly complex architectures and multi-step refinement for test-time scaling. While these

Rethinking Dense Sequential Chains: Reasoning Language Models Can Extract Answers from Sparse, Order-Shuffling Chain-of-Thoughts

ResearchDGX agent

arXiv:2605.07307v1 Announce Type: new Abstract: Modern reasoning language models generate dense, sequential chain-of-thought traces implicitly assuming that every token contributes and that steps must

Rethinking Experience Utilization in Self-Evolving Language Model Agents

ResearchDGX agent

arXiv:2605.07164v1 Announce Type: new Abstract: Self-evolving agents improve by accumulating and reusing experience from past interactions. Existing work has largely focused on how experience is const

Rethinking Importance Sampling in LLM Policy Optimization: A Cumulative Token Perspective

SafetyDGX agent

arXiv:2605.07331v1 Announce Type: cross Abstract: Reinforcement learning, including reinforcement learning with verifiable rewards (RLVR), has emerged as a powerful approach for LLM post-training. Cen

Rethinking State Tracking in Recurrent Models Through Error Control Dynamics

TutorialsDGX agent

arXiv:2605.07755v1 Announce Type: cross Abstract: The theory of state tracking in recurrent architectures has predominantly focused on expressive capacity: whether a fixed architecture can theoretical

Rethinking Weight Tying: Pseudo-Inverse Tying for LM Stable Training and Updates

Model ReleasesDGX agent

arXiv:2602.04556v2 Announce Type: replace Abstract: Weight tying is widely used in compact language models to reduce parameters by sharing the token table between the input embedding and the output pr

Retina-RAG: Retrieval-Augmented Vision-Language Modeling for Joint Retinal Diagnosis and Clinical Report Generation

Model ReleasesDGX agent

arXiv:2605.06173v2 Announce Type: replace-cross Abstract: Diabetic Retinopathy (DR) is a leading cause of preventable blindness among working-age adults worldwide, yet most automated screening systems

Retrieval from Within: An Intrinsic Capability of Attention-Based Models

ResearchDGX agent

arXiv:2605.05806v2 Announce Type: replace Abstract: Retrieval-augmented generation (RAG) typically treats retrieval and generation as separate systems. We ask whether an attention-based encoder-decode

Retrieval Heads are Dynamic

ResearchDGX agent

arXiv:2602.11162v2 Announce Type: replace Abstract: Recent studies have identified 'retrieval heads' in Large Language Models (LLMs) responsible for extracting information from input contexts. However

Retrieve, Integrate, and Synthesize: Spatial-Semantic Grounded Latent Visual Reasoning

ResearchDGX agent

arXiv:2605.07106v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have made remarkable progress on vision-language reasoning, yet most methods still compress visual evidence int

Revisiting Adam for Streaming Reinforcement Learning

ResearchDGX agent

arXiv:2605.06764v1 Announce Type: cross Abstract: Learning from a sequence of interactions, as soon as observations are perceived and acted upon, without explicitly storing them, holds the promise of

Revisiting Transformer Layer Parameterization Through Causal Energy Minimization

Model ReleasesDGX agent

arXiv:2605.07588v1 Announce Type: cross Abstract: Transformer blocks typically combine multi-head attention (MHA) for token mixing with gated MLPs for token-wise feature transformation, yet many choic

RIDER: 3D RNA Inverse Design with Reinforcement Learning-Guided Diffusion

SafetyDGX agent

arXiv:2602.16548v2 Announce Type: replace Abstract: The inverse design of RNA three-dimensional (3D) structures is crucial for engineering functional RNAs in synthetic biology and therapeutics. While

Risk-Consistent Multiclass Learning from Random Label-Subset Membership Queries

SafetyDGX agent

arXiv:2605.07413v1 Announce Type: new Abstract: Obtaining accurate class labels is often costly or unreliable, and may also be limited by privacy or other practical conditions. Compared with asking an

RL-RIG: A Generative Spatial Reasoner via Intrinsic Reflection

ResearchDGX agent

arXiv:2602.19974v2 Announce Type: replace Abstract: Recent advancements in image generation have achieved impressive results in producing high-quality images. However, existing image generation models

RNAGenScape: Property-Guided, Optimized Generation of mRNA Sequences with Manifold Langevin Dynamics

Local AiDGX agent

arXiv:2510.24736v3 Announce Type: replace-cross Abstract: Generating property-optimized mRNA sequences is central to applications such as vaccine design and protein replacement therapy, but remains ch

Robust and Reliable AI for Predictive Quality in Semiconductor Materials Manufacturing with MLOps and Uncertainty Quantification

ApplicationsDGX agent

arXiv:2605.07752v1 Announce Type: new Abstract: Semiconductor materials manufacturing presents unique challenges for machine learning deployment due to evolving process conditions, equipment degradati

Robust stochastic first order methods in heavy-tailed noise via medoid mini-batch gradient sampling

ResearchDGX agent

arXiv:2605.07634v1 Announce Type: cross Abstract: We consider a first order stochastic optimization framework where, at each iteration, K independent identically distributed (i.i.d.) data point sample

Robust Sublinear Convergence Rates for Iterative Bregman Projections

Model ReleasesDGX agent

arXiv:2602.01372v2 Announce Type: replace-cross Abstract: Entropic regularization provides a simple way to approximate linear programs whose constraints split into two or more tractable blocks. The re

Robustness of Refugee-Matching Gains to Off-Policy Evaluation Choices

SafetyDGX agent

arXiv:2605.06686v1 Announce Type: new Abstract: Previous research has investigated the potential of refugee matching for boosting refugee outcomes, first considered by Bansak et al. (2018). This paper

Rollback-Free Stable Brick Structures Generation

SafetyDGX agent

arXiv:2605.06947v1 Announce Type: new Abstract: While autoregressive models have advanced 3D generation, creating physically stable brick structures remains a challenge due to the strict requirements

RRCM: Ranking-Driven Retrieval over Collaborative and Meta Memories for LLM Recommendation

Model ReleasesDGX agent

arXiv:2605.07129v1 Announce Type: cross Abstract: Large Language Models (LLMs) have emerged as a promising paradigm for next-generation recommender systems, offering strong semantic understanding and

Rubric-based On-policy Distillation

SafetyDGX agent

arXiv:2605.07396v1 Announce Type: cross Abstract: On-policy distillation (OPD) is a powerful paradigm for model alignment, yet its reliance on teacher logits restricts its application to white-box sce

Rubric-Grounded RL: Structured Judge Rewards for Generalizable Reasoning

Model ReleasesDGX agent

arXiv:2605.08061v1 Announce Type: new Abstract: We argue that decomposing reward into weighted, verifiable criteria and using an LLM judge to score them provides a partial-credit optimization signal:

RuleSafe-VL: Evaluating Rule-Conditioned Decision Reasoning in Vision-Language Content Moderation

Model ReleasesDGX agent

arXiv:2605.07760v1 Announce Type: new Abstract: Platform content moderation applies explicit policy rules and context-dependent conditions to decide whether user content is allowed, restricted, or rem

S2M-Net: Spectral-Spatial Mixing for Medical Image Segmentation with Morphology-Aware Adaptive Loss

Model ReleasesDGX agent

arXiv:2601.01285v2 Announce Type: replace Abstract: Medical image segmentation requires balancing local precision for boundary-critical clinical applications, global context for anatomical coherence,

S2S-Arena: Evaluating Paralinguistic Instruction Following in Speech-to-Speech Models

Model ReleasesDGX agent

arXiv:2503.05085v2 Announce Type: replace Abstract: Recent advances in large language models (LLMs) have fundamentally reshaped speech-to-speech (S2S) systems, enabling increasingly natural spoken int

Safactory: A Scalable Agentic Infrastructure for Training Trustworthy Autonomous Intelligence

SafetyDGX agent

arXiv:2605.06230v2 Announce Type: replace Abstract: As large models evolve from conversational assistants into autonomous agents, challenges increasingly arise from long-horizon decision making, tool

Safe, or Simply Incapable? Rethinking Safety Evaluation for Phone-Use Agents

Model ReleasesDGX agent

arXiv:2605.07630v1 Announce Type: cross Abstract: When a phone-use agent avoids harm, does that show safety, or simply inability to act? Existing evaluations often cannot tell. A harmful outcome may b

Safety Anchor: Defending Harmful Fine-tuning via Geometric Bottlenecks

Model ReleasesDGX agent

arXiv:2605.05995v2 Announce Type: replace-cross Abstract: The safety alignment of Large Language Models (LLMs) remains vulnerable to Harmful Fine-tuning (HFT). While existing defenses impose constrain

SAGE: Hierarchical LLM-Based Literary Evaluation through Ontology-Grounded Interpretive Dimensions

SafetyDGX agent

arXiv:2605.07102v1 Announce Type: new Abstract: Evaluating literary quality requires assessing interpretive dimensions such as cultural representation, emotional depth, and philosophical sophisticatio

Saliency-Aware Regularized Quantization Calibration for Large Language Models

ResearchDGX agent

arXiv:2605.05693v2 Announce Type: replace Abstract: Post-training quantization (PTQ) is an effective approach for deploying large language models (LLMs) under memory and latency constraints. Most exis

SAM 3D Animal: Promptable Animal 3D Reconstruction from Images in the Wild

ResearchDGX agent

arXiv:2605.07604v1 Announce Type: cross Abstract: 3D animal reconstruction in the wild remains challenging due to large species variation, frequent occlusions, and the prevalence of multi-animal scene

Same Brain, Different Prediction: How Preprocessing Choices Undermine EEG Decoding Reliability

TutorialsDGX agent

arXiv:2605.07212v1 Announce Type: cross Abstract: Electroencephalography (EEG) is a cornerstone of brain-computer interfaces and clinical neuroscience, yet deep learning models are typically trained a

Same Signal, Opposite Meaning: Direction-Informed Adaptive Learning for LLM Agents

SafetyDGX agent

arXiv:2605.06908v1 Announce Type: cross Abstract: Adaptive test-time compute for LLM agents aims to invoke extra computation only when it improves performance. Existing methods typically use confidenc

Sample Complexity of Stochastic Optimization with Integer Variables

ResearchDGX agent

arXiv:2605.07239v1 Announce Type: new Abstract: We establish sample complexity results for stochastic optimization over the integers, especially with a view to understand the complexity with respect t

SARA: Semantically Adaptive Relational Alignment for Video Diffusion Models

SafetyDGX agent

arXiv:2605.07800v1 Announce Type: new Abstract: Recent video diffusion models (VDMs) synthesize visually convincing clips, yet still drop entities, mis-bind attributes, and weaken the interactions spe

Sat3R: Satellite DSM Reconstruction via RPC-Aware Depth Fine-tuning

Model ReleasesDGX agent

arXiv:2605.07264v1 Announce Type: new Abstract: Accurate Digital Surface Model (DSM) reconstruction from satellite imagery is critical for applications such as disaster response, urban planning, and l

SatSurfGS: Generalizable 2D Gaussian Splatting for Sparse-View Satellite Surface Reconstruction

Model ReleasesDGX agent

arXiv:2605.07181v1 Announce Type: new Abstract: Sparse-view satellite image surface reconstruction remains highly challenging, fundamentally because the reliability of multi-view matching under satell

Saving Foundation Flow-Matching Priors for Inverse Problems

ResearchDGX agent

arXiv:2511.16520v2 Announce Type: replace-cross Abstract: Foundation flow-matching (FM) models promise a universal prior for solving inverse problems (IPs), yet today they trail behind domain-specific

SB-TRPO: Towards Safe Reinforcement Learning with Hard Constraints

SafetyDGX agent

arXiv:2512.23770v3 Announce Type: replace-cross Abstract: In safety-critical domains, reinforcement learning (RL) agents must often satisfy strict, zero-cost safety constraints while accomplishing tas

Scalable Equilibrium Propagation via Intermediate Error Signals for Deep Convolutional CRNNs

ResearchDGX agent

arXiv:2508.15989v2 Announce Type: replace Abstract: Equilibrium Propagation (EP) is a biologically inspired local learning rule first proposed for convergent recurrent neural networks (CRNNs), in whic

Scalable Option Learning in High-Throughput Environments

ResearchDGX agent

arXiv:2509.00338v3 Announce Type: replace-cross Abstract: Hierarchical reinforcement learning (RL) has the potential to enable effective decision-making over long timescales. Existing approaches, whil

Scaling Categorical Flow Maps

Model ReleasesDGX agent

arXiv:2605.07820v1 Announce Type: new Abstract: Continuous diffusion and flow matching models could represent a powerful alternative to autoregressive approaches for language modelling (LM), as they u

Scaling Continual Learning to 300+ Tasks with Bi-Level Routing Mixture-of-Experts

Model ReleasesDGX agent

arXiv:2602.03473v2 Announce Type: replace-cross Abstract: Continual learning, especially class-incremental learning (CIL), on the basis of a pre-trained model (PTM) has garnered substantial research i

SCENE: Recognizing Social Norms and Sanctioning in Group Chats

Model ReleasesDGX agent

arXiv:2605.07823v1 Announce Type: new Abstract: Online group chats are social spaces with implicit behavior patterns that, when broken, are often met with social sanctioning from the group. The abilit

SCOPE: Structured Decomposition and Conditional Skill Orchestration for Complex Image Generation

Model ReleasesDGX agent

arXiv:2605.08043v1 Announce Type: cross Abstract: While text-to-image models have made strong progress in visual fidelity, faithfully realizing complex visual intents remains challenging because many

SCOUT: Closed-Loop in-vivo System for Continuous Methane Concentration Monitoring in Cattle

AgentsDGX agent

arXiv:2508.04056v2 Announce Type: replace Abstract: Enteric methane measurement from ruminant livestock faces fundamental trade-offs between accuracy and operational feasibility. Existing methods quan

ScrapeGraphAI-100k: Dataset for Schema-Constrained LLM Generation

Model ReleasesDGX agent

arXiv:2602.15189v2 Announce Type: replace-cross Abstract: Producing output that conforms to a specified JSON schema underlies tool use, structured extraction, and knowledge base construction in modern

← Previous
1…750751752753754…1005
Next →