AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,171
  • Agents7,461
  • Applications5,337
  • Concepts5
  • Hardware1,806
  • Industry6,146
  • Local Ai4,871
  • Model Releases23,435
  • Research19,874
  • Safety13,191
  • Syntheses17
  • Tools1,673
  • Tutorials3,355

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,171
  • Agents7,461
  • Applications5,337
  • Concepts5
  • Hardware1,806
  • Industry6,146
  • Local Ai4,871
  • Model Releases23,435
  • Research19,874
  • Safety13,191
  • Syntheses17
  • Tools1,673
  • Tutorials3,355

Source
HumanDGX agent

Content type
87,171Total entries
1Added by human
87,170Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
62,006 results
Research

Process Rewards with Learned Reliability

DGX agent

arXiv:2605.15529v1 Announce Type: cross Abstract: Process Reward Models (PRMs) provide step-level feedback for reasoning, but current PRMs usually output only a single reward score for each step. Down

researcharxiv-cs-ai
18 May 2026
Research

Prompt Stability Scoring for Text Annotation with Large Language Models

DGX agent
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2407.02039v3 Announce Type: replace Abstract: Researchers are increasingly using language models (LMs) for text annotation. These approaches rely only on a prompt telling the model to return a g

researcharxiv-cs-cl
18 May 2026
Safety

Propagating Unsafe Actions in LLM Controlled Multi-Robot Collaboration via Single Robot Compromise

DGX agent

arXiv:2605.15641v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used as general planners in embodied intelligence, enabling high level coordination and low level task pla

safetyarxiv-cs-ro
18 May 2026
Tutorials

Property-Guided LLM Program Synthesis for Planning

DGX agent

arXiv:2605.16142v1 Announce Type: new Abstract: LLMs have shown impressive success in program synthesis, discovering programs that surpass prior solutions. However, these approaches rely on simple num

tutorialsarxiv-cs-ai
18 May 2026
Agents

Prospective multi-pathogen disease forecasting using autonomous LLM-guided tree search

DGX agent

arXiv:2605.16238v1 Announce Type: new Abstract: Probabilistic forecasting of infectious diseases is crucial for public health but relies on labor-intensive manual model curation by expert modeling tea

agentsarxiv-cs-ai
18 May 2026
Safety

PSD: Pushing the Pareto Frontier of Diffusion LLMs via Parallel Speculative Decoding

DGX agent

arXiv:2605.15609v1 Announce Type: new Abstract: Diffusion large language models (dLLMs) generate text by iteratively denoising masked token sequences. Although dLLMs can predict all masked positions i

safetyarxiv-cs-cl
18 May 2026
Model Releases

Quantization Undoes Alignment: Bias Emergence in Compressed LLMs Across Models and Precision Levels

DGX agent

arXiv:2605.15208v1 Announce Type: cross Abstract: Large Language Models are routinely compressed via post-training quantization to reduce inference costs and memory footprint for cloud and edge deploy

model-releasesarxiv-cs-ai
18 May 2026
Safety

Quantum Artificial Intelligence for Mission-Critical Systems: Foundations, Architectural Elements, and Future Directions

DGX agent

arXiv:2511.09884v2 Announce Type: replace Abstract: Mission critical (MC) applications such as defense operations, energy management, cybersecurity, and aerospace control require reliable, determinist

safetyarxiv-cs-ai
18 May 2026
Model Releases

Quantum Feature Pyramid Gating for Seismic Image Segmentation

DGX agent

arXiv:2605.15370v1 Announce Type: cross Abstract: Accurate salt-body delineation is essential for seismic interpretation because salt structures distort wave propagation, complicate velocity-model bui

model-releasesarxiv-cs-lg
18 May 2026
Safety

RanSOM: Second-Order Momentum with Randomized Scaling for Constrained and Unconstrained Optimization

DGX agent

arXiv:2602.06824v2 Announce Type: replace-cross Abstract: Momentum methods, such as Polyak's Heavy Ball, are the standard for training deep networks but suffer from curvature-induced bias in stochasti

safetyarxiv-cs-lg
18 May 2026
Research

RaPD: Resolution-Agnostic Pixel Diffusion via Semantics-Enriched Implicit Representations

DGX agent

arXiv:2605.15908v1 Announce Type: cross Abstract: Natural images are continuous, yet most generative models synthesize them on discrete grids, limiting resolution-flexible generation. Continuous neura

researcharxiv-cs-ai
18 May 2026
Model Releases

RapidUn: Influence-Driven Parameter Reweighting for Efficient Large Language Model Unlearning

DGX agent

arXiv:2512.04457v2 Announce Type: replace Abstract: Removing specific data influence from large language models (LLMs) remains challenging, as retraining is costly and existing approximate unlearning

model-releasesarxiv-cs-cl
18 May 2026
Model Releases

RAR: Retrieving And Ranking Augmented MLLMs for Visual Recognition

DGX agent

arXiv:2403.13805v2 Announce Type: replace-cross Abstract: CLIP (Contrastive Language-Image Pre-training) uses contrastive learning from noise image-text pairs to excel at recognizing a wide array of c

model-releasesarxiv-cs-ai
18 May 2026
Safety

RE-SAC: Disentangling aleatoric and epistemic risks in bus fleet control: A stable and robust ensemble DRL approach

DGX agent

arXiv:2603.18396v3 Announce Type: replace Abstract: Bus holding control is challenging due to stochastic traffic and passenger demand. While deep reinforcement learning (DRL) shows promise, standard a

safetyarxiv-cs-lg
18 May 2026
Safety

Reactive Robot-Centric Safety for Autonomous Navigation in Constrained and Dynamic Environments

DGX agent

arXiv:2605.15782v1 Announce Type: new Abstract: In this work, we address the problem of ensuring real-time safety in autonomous robot navigation, in spatially constrained dynamic environments, by util

safetyarxiv-cs-ro
18 May 2026
Safety

ReactiveGWM: Steering NPC in Reactive Game World Models

DGX agent

arXiv:2605.15256v1 Announce Type: new Abstract: Current game world models simulate environments from a subjective, player-centric perspective. However, by treating the Non-Player Character (NPC) merel

safetyarxiv-cs-cv
18 May 2026
Research

Reading the Cell, Designing the Cure: Perturbation-Conditioned Molecular Diffusion for Function-Oriented Drug Design

DGX agent

arXiv:2605.15243v1 Announce Type: cross Abstract: When reliable target structures are unavailable at scale or phenotypes arise from dysregulated pathways, transcriptomic perturbations provide a system

researcharxiv-cs-ai
18 May 2026
Safety

ReAlign: Generalizable Image Forgery Detection via Reasoning-Aligned Representation

DGX agent

arXiv:2605.16080v1 Announce Type: new Abstract: The rise of AI-generated images (AIGIs) poses growing challenges for digital authenticity, prompting the need for efficient, generalizable image forgery

safetyarxiv-cs-cv
18 May 2026
Applications

RealRep: Generalized SDR-to-HDR Conversion via Attribute-Disentangled Representation Learning

DGX agent

arXiv:2505.07322v4 Announce Type: replace Abstract: High-Dynamic-Range Wide-Color-Gamut (HDR-WCG) technology is becoming increasingly widespread, driving a growing need for converting Standard Dynamic

applicationsarxiv-cs-cv
18 May 2026
Applications

Reasoners or Translators? Contamination-aware Evaluation and Neuro-Symbolic Robustness in Tax Law

DGX agent

arXiv:2605.16052v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) have significantly enhanced automated legal reasoning. Yet, it remains unclear whether their performance

applicationsarxiv-cs-ai
18 May 2026
Research

Reasoning Models Don't Just Think Longer, They Move Differently

DGX agent

arXiv:2605.15454v1 Announce Type: new Abstract: Reasoning-trained language models often spend more tokens on harder problems, but longer chains of thought do not show whether a model is merely computi

researcharxiv-cs-cl
18 May 2026
Agents

RecMem: Recurrence-based Memory Consolidation for Efficient and Effective Long-Running LLM Agents

DGX agent

arXiv:2605.16045v1 Announce Type: cross Abstract: Memory systems often organize user-agent interactions as retrievable external memory and are crucial for long-running agents by overcoming the limited

agentsarxiv-cs-ai
18 May 2026
Model Releases

Reducing the Safety Tax in LLM Safety Alignment with On-Policy Self-Distillation

DGX agent

arXiv:2605.15239v1 Announce Type: new Abstract: Safety alignment often improves robustness to harmful queries at the cost of reasoning ability, a tradeoff known as the safety tax. A common cause is di

model-releasesarxiv-cs-lg
18 May 2026
Safety

Reference-Free Reinforcement Learning Fine-Tuning for MT: A Seq2Seq Perspective

DGX agent

arXiv:2605.15976v1 Announce Type: cross Abstract: Production machine translation relies overwhelmingly on encoder-decoder Seq2Seq models, yet reinforcement learning approaches to MT fine-tuning have l

safetyarxiv-cs-ai
18 May 2026
Safety

Reference Games as a Testbed for the Alignment of Model Uncertainty and Clarification Requests

DGX agent

arXiv:2601.07820v2 Announce Type: replace Abstract: In human conversation, both interlocutors play an active role in maintaining mutual understanding. When listeners are uncertain about what speakers

safetyarxiv-cs-cl
18 May 2026
Model Releases

Registers Matter for Pixel-Space Diffusion Transformers

DGX agent

arXiv:2605.16147v1 Announce Type: new Abstract: Vision Transformers (ViTs) are known to exhibit high-norm patch-token outliers that degrade feature map quality, a problem effectively mitigated by exti

model-releasesarxiv-cs-cv
18 May 2026
Research

Regret and Sample Complexity of Online Q-Learning via Concentration of Stochastic Approximation with Time-Inhomogeneous Markov Chains

DGX agent

arXiv:2602.16274v2 Announce Type: replace Abstract: We present the first regret bound for classical online Q-learning in infinite-horizon discounted Markov decision processes (MDPs), without relying o

researcharxiv-cs-lg
18 May 2026
Model Releases

Reinforcement learning for adaptive interior point methods in convex quadratic programming

DGX agent

arXiv:2509.07404v2 Announce Type: replace-cross Abstract: Quadratic programming is a workhorse of modern nonlinear optimization, control, and data science. Although regularized methods offer convergen

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

Representation Without Reward: A JEPA Audit for LLM Fine-Tuning

DGX agent

arXiv:2605.15394v1 Announce Type: cross Abstract: Joint-embedding predictive architectures (JEPAs) propose that a model should learn more useful abstractions when trained to predict latent representat

model-releasesarxiv-cs-ai
18 May 2026
Safety

Res^2CLIP: Few-Shot Generalist Anomaly Detection with Residual-to-Residual Alignment

DGX agent

arXiv:2605.16171v1 Announce Type: new Abstract: Few-shot Generalist Anomaly Detection requires models to generalize to novel categories without retraining, posing significant challenges in real-world

safetyarxiv-cs-cv
18 May 2026
Safety

Residual Reinforcement Learning for Robot Teleoperation under Stochastic Delays

DGX agent

arXiv:2605.15480v1 Announce Type: cross Abstract: Stochastic communication delays in teleoperation introduce signal discontinuities that undermine control stability and degrade control performance. Co

safetyarxiv-cs-ai
18 May 2026
Safety

Response-Conditioned Parallel-to-Sequential Orchestration for Multi-Agent Systems

DGX agent

arXiv:2605.15573v1 Announce Type: new Abstract: Multi-agent systems can solve complex tasks through collaboration between multiple Large Language Model agents. Existing collaboration frameworks typica

safetyarxiv-cs-cl
18 May 2026
Research

Rethinking Neural Network Learning Rates: A Stackelberg Perspective

DGX agent

arXiv:2605.15530v1 Announce Type: new Abstract: Neural networks are typically trained with a single learning rate across all layers. While recent empirical evidence suggests that assigning layer-speci

researcharxiv-cs-lg
18 May 2026
Local Ai

Rethinking Predictive Modeling for LLM Routing: When Simple kNN Beats Complex Learned Routers

DGX agent

arXiv:2505.12601v2 Announce Type: replace Abstract: As large language models (LLMs) grow in scale and specialization, routing--selecting the best model for a given input--has become essential for effi

local-aiarxiv-cs-lg
18 May 2026
Model Releases

Retrieval-Augmented Large Language Models for Schema-Constrained Clinical Information Extraction

DGX agent

arXiv:2605.15467v1 Announce Type: cross Abstract: Conversational nurse-patient transcripts contain actionable observations, but converting these transcripts into structured representations at scale re

model-releasesarxiv-cs-ai
18 May 2026
Safety

Reversing the Flow: Generation-to-Understanding Synergy in Large Multimodal Models

DGX agent

arXiv:2605.15792v1 Announce Type: new Abstract: The long-standing goal of multimodal AI is to build unified models in which visual understanding and visual generation mutually enhance one another. Des

safetyarxiv-cs-cv
18 May 2026
Applications

Reweighting free energy profiles between universal machine learning interatomic potentials for fast consensus building

DGX agent

arXiv:2605.15630v1 Announce Type: cross Abstract: Free energy profiles serve as a fundamental bridge between microscopic atomic fluctuations and macroscopic thermodynamic observables. Estimating the f

applicationsarxiv-cs-lg
18 May 2026
Research

RIDE: Retinex-Informed Decoupling for Exposing Concealed Objects

DGX agent

arXiv:2605.15450v1 Announce Type: cross Abstract: Concealed Object Segmentation (COS) encompasses a family of dense-prediction tasks, including camouflaged object detection, polyp segmentation, transp

researcharxiv-cs-ai
18 May 2026
Model Releases

RoadmapBench: Evaluating Long-Horizon Agentic Software Development Across Version Upgrades

DGX agent

arXiv:2605.15846v1 Announce Type: cross Abstract: Coding agents are increasingly deployed in real software development, where a single version iteration requires months of coordinated work across many

model-releasesarxiv-cs-ai
18 May 2026
Research

Robust Prior-Guided Segmentation for Editable 3D Gaussian Splatting

DGX agent

arXiv:2605.16065v1 Announce Type: cross Abstract: 3D Gaussian Splatting (3D-GS) enables real-time 3D scene reconstruction but lacks robust segmentation for editing tasks such as object removal, extrac

researcharxiv-cs-ai
18 May 2026
Research

RoiMAM: Region-of-Interest Medical Attention Model for Efficient Vision-Language Understanding

DGX agent

arXiv:2605.15561v1 Announce Type: new Abstract: Vision-Language Models (VLMs) facilitate medical visual question answering (MedVQA) by jointly interpreting images and text. However, existing models ty

researcharxiv-cs-cv
18 May 2026
Safety

RoPE Distinguishes Neither Positions Nor Tokens in Long Contexts, Provably

DGX agent

arXiv:2605.15514v1 Announce Type: cross Abstract: We identify intrinsic limitations of Rotary Positional Embeddings (RoPE) in Transformer-based long-context language models. Our theoretical analysis a

safetyarxiv-cs-ai
18 May 2026
Model Releases

RTL-BenchMT: Dynamic Maintenance of RTL Generation Benchmark Through Agent-Assisted Analysis and Revision

DGX agent

arXiv:2605.15537v1 Announce Type: new Abstract: This paper introduces RTL-BenchMT, an agentic framework for dynamically maintaining RTL generation benchmarks. Large Language Models (LLMs) assisted aut

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Rule2DRC: Benchmarking LLM Agents for DRC Script Synthesis with Execution-Guided Test Generation

DGX agent

arXiv:2605.15669v1 Announce Type: new Abstract: Manufacturable chip layouts must satisfy thousands of geometry-based design rules, and design rule checking (DRC) enforces them by running executable DR

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

Runtime-Orchestrated Second-Order Optimization for Scalable LLM Training

DGX agent

arXiv:2605.16184v1 Announce Type: cross Abstract: Second-order methods offer an attractive path toward more sample-efficient LLM training, but their practical use is often blocked by the systems cost

model-releasesarxiv-cs-lg
18 May 2026
Agents

Runtime-Structured Task Decomposition for Agentic Coding Systems

DGX agent

arXiv:2605.15425v1 Announce Type: cross Abstract: Agentic coding systems increasingly use large language models (LLMs) for software engineering tasks such as debugging, root cause analysis, and code r

agentsarxiv-cs-ai
18 May 2026
Model Releases

SaaS-Bench: Can Computer-Use Agents Leverage Real-World SaaS to Solve Professional Workflows?

DGX agent

arXiv:2605.15777v1 Announce Type: new Abstract: Computer-Using Agents (CUAs) are rapidly extending large language models (LLMs) beyond text-based reasoning toward action execution in more complex envi

model-releasesarxiv-cs-ai
18 May 2026
Research

SAE-RNA: A Sparse Autoencoder Model for Interpreting RNA Language Model Representations

DGX agent

arXiv:2510.02734v2 Announce Type: replace-cross Abstract: Deep learning, particularly with the advancement of Large Language Models, has transformed biomolecular modeling, with protein language models

researcharxiv-cs-ai
18 May 2026
← Previous
1…859860861862863…1292
Next →