AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,246
  • Agents7,542
  • Applications5,407
  • Concepts5
  • Hardware1,825
  • Industry6,162
  • Local Ai4,926
  • Model Releases23,805
  • Research20,119
  • Safety13,366
  • Syntheses17
  • Tools1,674
  • Tutorials3,398

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,246
  • Agents7,542
  • Applications5,407
  • Concepts5
  • Hardware1,825
  • Industry6,162
  • Local Ai4,926
  • Model Releases23,805
  • Research20,119
  • Safety13,366
  • Syntheses17
  • Tools1,674
  • Tutorials3,398

Source
Human
88,246Total entries
1Added by human
88,245Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
62,897 results
29 May 2026

Rethinking Post-Training Recipes for Multimodal Time-Series Forecasting

Model ReleasesDGX agent

arXiv:2605.29401v1 Announce Type: new Abstract: Time-Series Foundation Models (TSFMs) excel at zero-shot unimodal forecasting using numerical data, but unlike LLMs they cannot consume multimodal, non-

Rethinking Stepwise Model Routing: A Cost-Efficient Table Reasoning Perspective

ResearchDGX agent

arXiv:2605.29319v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) achieve strong performance on table reasoning tasks but incur substantial inference cost due to long reasoning traces. Ste

Return-to-Go Is More Than a Number: Q-Guided Alignment for Return-Conditioned Supervised Learning

Model ReleasesDGX agent
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.29028v1 Announce Type: cross Abstract: Conditioned Sequence Models (CSMs) learn policies by treating return-to-go (RTG) as a control signal. However, existing CSMs often treat the RTGs as s

Review Arcade: On the Human Alignment and Gameability of LLM Reviews

SafetyDGX agent

arXiv:2605.28897v1 Announce Type: new Abstract: LLM-generated reviews for scientific papers are gaining considerable traction and are even being officially piloted by major conferences. We have to ass

Revisiting Observation Reduction for Web Agents: Comprehensive Evaluation with a Lightweight Framework

AgentsDGX agent

arXiv:2605.29397v1 Announce Type: new Abstract: HTML observations in LLM-based web agents are extremely long, and while many reduction methods have been proposed, it remains unclear which methods redu

RewardFlow: Topology-Aware Reward Propagation on State Graphs for Agentic RL with Large Language Models

AgentsDGX agent

arXiv:2603.18859v2 Announce Type: replace Abstract: Reinforcement learning (RL) shows promise for enhancing LLM agentic reasoning, yet sparse terminal rewards hinder fine-grained optimization. Process

RHO: Robust Holistic OSM-Based Metric Cross-View Geo-Localization

Model ReleasesDGX agent

arXiv:2603.27758v2 Announce Type: replace Abstract: Metric Cross-View Geo-Localization (MCVGL) aims to estimate the 3-DoF camera pose (position and heading) by matching ground and satellite images. In

Ridge Regression from Poisson Resetting: A Renewal Perspective on Spectral Regularization

ResearchDGX agent

arXiv:2605.30059v1 Announce Type: new Abstract: We connect stochastic resetting from non-equilibrium statistical physics with ridge regularization in statistical learning. For linear gradient flow, re

Riemannian AmbientFlow: Towards Simultaneous Manifold Learning and Generative Modeling from Corrupted Data

ResearchDGX agent

arXiv:2601.18728v2 Announce Type: replace Abstract: Modern generative modeling methods have demonstrated strong performance in learning complex data distributions from clean samples. In many scientifi

RightNow-Arabic-0.5B-Turbo: An Open Sub-1B Arabic Language Model via Vocabulary Injection and Edge-First Deployment

Model ReleasesDGX agent

arXiv:2605.28827v1 Announce Type: new Abstract: Open Arabic large language models split into two classes: sub-1B multilingual models that treat Arabic as an afterthought (Qwen2.5-0.5B, Falcon-H1-0.5B)

Risk-averse Fair Multi-class Classification

Model ReleasesDGX agent

arXiv:2509.05771v2 Announce Type: replace-cross Abstract: We develop a new classification framework based on the theory of coherent risk measures and systemic risk. The proposed approach is suitable f

RL2ML: Finite-Rollout Surrogate Objectives from Reinforcement Learning to Maximum Likelihood

SafetyDGX agent

arXiv:2605.30154v1 Announce Type: new Abstract: Correctness-based Reinforcement Learning with Verifiable Rewards (RLVR) trains language models from binary feedback on sampled outputs, but the objectiv

RoboWits: Unexpected Challenges for Robotic Creative Problem Solving

Model ReleasesDGX agent

arXiv:2605.30326v1 Announce Type: cross Abstract: The ability to reason, adapt, and creatively solve problems under unexpected challenges is essential for robots operating in real-world environments.

Robust and Efficient Guardrails with Latent Reasoning

Model ReleasesDGX agent

arXiv:2605.29068v1 Announce Type: new Abstract: Maintaining the safety of large language models (LLMs) is crucial as they are increasingly deployed in real-world applications. Existing safety guardrai

Robust and Efficient Writer-Independent IMU-Based Handwriting Recognition

ApplicationsDGX agent

arXiv:2502.20954v3 Announce Type: replace Abstract: Handwriting recognition (HWR) using inertial measurement unit (IMU) data remains challenging due to variations in writing styles and the limited ava

Robust and Generalizable Safety Steering for Text-to-Image Diffusion Transformers

SafetyDGX agent

arXiv:2605.30049v1 Announce Type: new Abstract: Diffusion Transformers have become a powerful backbone for text-to-image generation, but their layered and cross-modal generation process makes safety c

Robust Cross-Domain Generalization Using Unlabeled Target Data with Source-Domain Supervision

TutorialsDGX agent

arXiv:2605.29122v1 Announce Type: new Abstract: It is often desirable to generalize medical imaging AI models trained with dense annotations to data acquired from different ultrasound scanners or clin

Robust Frequency-Calibrated Virtual EEG Channel Generation from Four Frontal Electrodes for Wearable EEG Augmentation

ResearchDGX agent

arXiv:2605.29263v1 Announce Type: new Abstract: Low-channel wearable electroencephalography (EEG) is attractive for long-term monitoring, but four frontal electrodes provide only a sparse and spatiall

Rooted Absorbed Prefix Trajectory Balance with Submodular Replay for GFlowNet Training

SafetyDGX agent

arXiv:2603.00454v2 Announce Type: replace-cross Abstract: Generative Flow Networks (GFlowNets) enable fine-tuning large language models to approximate reward-proportional posteriors, but they remain p

Routing by Reaching: Composition of Pre-trained GFlowNets for Multi-Objective Generation

TutorialsDGX agent

arXiv:2602.21565v2 Announce Type: replace Abstract: Generative Flow Networks (GFlowNets) learn to sample diverse candidates in proportion to a reward function, making them well-suited for scientific d

RUBRIC-ARROW: Alternating Pointwise Rubric Reward Modeling for LLM Post-training in Non-verifiable Domains

SafetyDGX agent

arXiv:2605.29156v1 Announce Type: cross Abstract: Pointwise reward modeling offers critical signals for LLM post-training, yet struggles with absolute scoring in subjective, non-verifiable settings. R

Rubric-Guided Process Reward for Stepwise Model Routing

SafetyDGX agent

arXiv:2605.29310v1 Announce Type: new Abstract: Stepwise model routing improves the efficiency of Large Reasoning Models (LRMs) by assigning each reasoning step to a suitable model. Recent methods for

S-MARC: Causal Streaming Reasoning for Full-Duplex Conversational Behavior Modeling

Model ReleasesDGX agent

arXiv:2602.11065v2 Announce Type: replace-cross Abstract: Human conversation is organized by an implicit chain of thought and manifests as temporally structured conversational behaviors. Capturing thi

S2MDF: A Plug-And-Play Layer for Intersection-Free Multi-Object Signed Distance Fields

ResearchDGX agent

arXiv:2605.29761v1 Announce Type: new Abstract: Compositional implicit surface representations model scenes as collections of objects, each encoded by a Signed Distance Field (SDF). A fundamental limi

S3Mem: Structured Spatiotemporal Scene-Event Memory for Long-Horizon Interactive Question Answering

Local AiDGX agent

arXiv:2605.28831v1 Announce Type: cross Abstract: Long-horizon interactive agents often accumulate large trajectory histories yet still fail to answer questions about earlier events reliably. We argue

SAAS: Self-Aware Reinforcement Learning for Over-Search Mitigation in Agentic Search

Model ReleasesDGX agent

arXiv:2605.29796v1 Announce Type: new Abstract: Agentic search enables LLMs to solve complex multi-hop questions through iterative reasoning and external search. Despite the effectiveness, these syste

SADA: Safe and Adaptive Aggregation of Multiple Black-Box Predictions in Semi-Supervised Learning

ResearchDGX agent

arXiv:2509.21707v3 Announce Type: replace-cross Abstract: Semi-supervised learning (SSL) arises in practice when labeled data are scarce or expensive to obtain, while large quantities of unlabeled dat

SAFE-Pruner: Semantic Attention-Guided Future-Aware Token Pruning for Efficient Vision-Language-Action Manipulation

ApplicationsDGX agent

arXiv:2605.29662v1 Announce Type: new Abstract: Real-time inference of vision-language-action (VLA) models is essential for robotic control. While visual token pruning has shown strong potential for a

SafeRx-Agent: A Knowledge-Grounded Multi-Agent Framework for Safe and Explainable Medication Recommendation

SafetyDGX agent

arXiv:2605.29146v1 Announce Type: cross Abstract: Medication recommendation predicts medications for patient visits, but existing methods still face two key challenges. At the model level, traditional

SafeSearch: Automated Red-Teaming of LLM-Based Search Agents

Model ReleasesDGX agent

arXiv:2509.23694v5 Announce Type: replace Abstract: Search agents connect LLMs to the Internet, enabling them to access broader and more up-to-date information. However, this also introduces a new thr

SAGE: Segment-Aware Gloss-Free Encoding for Token-Efficient Sign Language Translation

Model ReleasesDGX agent

arXiv:2507.09266v2 Announce Type: replace Abstract: Gloss-free Sign Language Translation (SLT) has advanced rapidly, achieving strong performances without relying on gloss annotations. However, these

SAHG: Sector-Anisotropic Hyperbolic Graph Model for Social Bot Detection

SafetyDGX agent

arXiv:2605.30166v1 Announce Type: cross Abstract: LLM-driven social bots can generate fluent, human-like text, reducing the discriminative advantage of content-based detection alone. However, coordina

SalsaAgent: A multimodal embodied language model for interactive dance generation

ResearchDGX agent

arXiv:2605.29219v1 Announce Type: new Abstract: Interaction between humanoids involves bidirectional and nonverbal reactivity, coordination and synchrony. Toward socially aware robots and interactive

SAM3D-Phys: Towards Multi-Object Interactive Simulation in Real World

ApplicationsDGX agent

arXiv:2605.30239v1 Announce Type: new Abstract: This work addresses the problem of recovering complete, simulatable object geometry from reconstructed real-world scenes, enabling physics-based interac

Same Evidence, Different Answers: Canonical-Context On-Policy Distillation for Multi-Turn Language Models

SafetyDGX agent

arXiv:2605.30251v1 Announce Type: cross Abstract: Large language models (LLMs) often solve a task when all instructions are given in a single prompt, but fail when the same information is revealed gra

Same Question, Different Source, Different Answer: Auditing Source-Dependence in Medical Multi-Source RAG

Model ReleasesDGX agent

arXiv:2605.29084v1 Announce Type: cross Abstract: A retrieval-augmented generation (RAG) system deployed over a multi-author institutional corpus can give a different answer to the same question depen

Sample-Efficient Diffusion-based Reinforcement Learning with Critic Guidance

Model ReleasesDGX agent

arXiv:2605.30056v1 Announce Type: cross Abstract: Recent advances in reinforcement learning (RL) have achieved great successes by leveraging the multimodality and exploration capability of diffusion p

SAVAA: Mitigating Hallucinations in LVLMs via Step-wise Adaptive Visual Attention Amplification

ResearchDGX agent

arXiv:2602.13600v2 Announce Type: replace Abstract: A line of recent training-free methods for mitigating hallucinations in large vision-language models (LVLMs) operates by amplifying attention to vis

Scalable RF Simulation in Generative 4D Worlds

ApplicationsDGX agent

arXiv:2508.12176v2 Announce Type: replace-cross Abstract: Radio Frequency (RF) sensing has emerged as a powerful, privacy-preserving alternative to vision-based methods for various perception tasks. H

Scaling Laws for Agent Harnesses via Effective Feedback Compute

Model ReleasesDGX agent

arXiv:2605.29682v1 Announce Type: new Abstract: Agent harnesses increasingly determine the performance of language-model systems by deciding how models call tools, receive feedback, verify intermediat

Scaling Monosemanticity: Extracting Interpretable Features from Claude 3 Sonnet

Model ReleasesDGX agent

arXiv:2605.29358v1 Announce Type: new Abstract: We demonstrate that sparse autoencoders can extract interpretable features from Claude 3 Sonnet, a production-scale language model, addressing the open

Scaling Small Agents Through Strategy Auctions

AgentsDGX agent

arXiv:2602.02751v2 Announce Type: replace-cross Abstract: Small language models are increasingly viewed as a promising, cost-effective approach to agentic AI, with proponents claiming they are suffici

SCDBench: A Benchmark for LLM-Based Smart Contract Decompilers

Model ReleasesDGX agent

arXiv:2605.29059v1 Announce Type: cross Abstract: Smart contract decompilation aims to recover high-level source code from bytecode, but evaluating decompilers remains difficult because existing studi

ScheduleStream: Temporal Planning with Samplers for GPU-Accelerated Multi-Arm Task and Motion Planning & Scheduling

HardwareDGX agent

arXiv:2511.04758v2 Announce Type: replace-cross Abstract: Bimanual and humanoid robots are appealing because of their human-like ability to leverage multiple arms to efficiently complete tasks. Howeve

SchGen: PCB Schematic Generation with Semantic-Grounded Code Representations

AgentsDGX agent

arXiv:2605.30345v1 Announce Type: new Abstract: Printed circuit board (PCB) schematic design defines nearly all electronic hardware, but it remains manual and expertise-intensive. While generative AI

SciIntBench: Measuring LLM Compliance with Research Integrity Norms Under Adversarial Framing

Model ReleasesDGX agent

arXiv:2605.29468v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used to support scientific work, but it is unclear whether they uphold responsible conduct of research (

SCoOP: Semantic Consistent Opinion Pooling for Uncertainty Quantification in Multiple Vision-Language Model Systems

ResearchDGX agent

arXiv:2603.23853v3 Announce Type: replace Abstract: Combining multiple Vision-Language Models (VLMs) can enhance multimodal reasoning and robustness, but aggregating heterogeneous models' outputs ampl

SCOPE: A Lightweight-training LLM Framework for Air Traffic Control Readback Monitoring

ResearchDGX agent

arXiv:2605.29543v1 Announce Type: cross Abstract: Pilot readback of Air Traffic Control (ATC) voice instructions is a primary safeguard against miscommunication in air transportation. However, readbac

SCOPE: Prompt Evolution for Enhancing Agent Effectiveness

Model ReleasesDGX agent

arXiv:2512.15374v2 Announce Type: replace Abstract: Large Language Model (LLM) agents are increasingly deployed in environments that generate massive, dynamic contexts. However, a critical bottleneck

SDF-Net: Structure-Aware Disentangled Feature Learning for Opticall-SAR Ship Re-identification

Model ReleasesDGX agent

arXiv:2603.12588v2 Announce Type: replace Abstract: Cross-modal ship re-identification (ReID) between optical and synthetic aperture radar (SAR) imagery is fundamentally challenged by the severe radio

SEAL: Can Saturated Benchmarks Be Revived by LLM-as-a-Meta-Judge?

AgentsDGX agent

arXiv:2605.30104v1 Announce Type: new Abstract: Widely used language-model benchmarks are increasingly saturated, with frontier systems often receiving near-tied scores that standard metrics cannot re

Securing SIM-Assisted Wireless Networks via Quantum Reinforcement Learning

SafetyDGX agent

arXiv:2602.13238v2 Announce Type: replace-cross Abstract: Stacked intelligent metasurfaces (SIMs) have recently emerged as a powerful wave-domain technology that enables multi-stage manipulation of el

Seeing through boxes: Non-Line-of-Sight 3D Reconstruction from Radar Signals

TutorialsDGX agent

arXiv:2605.29098v1 Announce Type: new Abstract: Reconstructing object geometry from radio frequency (RF) signals is fundamentally challenging due to the lensless imaging nature of RF sensing, which le

Selecting Hyperparameters for Tree-Boosting

ResearchDGX agent

arXiv:2602.05786v2 Announce Type: replace Abstract: Tree-boosting is a widely used machine learning technique for tabular data. However, its out-of-sample accuracy is critically dependent on multiple

Selection Hyper-heuristics Can Automatically Adjust the Learning Period to Optimally Solve Pseudo-Boolean Problems

Model ReleasesDGX agent

arXiv:2605.29916v1 Announce Type: cross Abstract: The Random Gradient hyper-heuristic was recently shown to be able to learn the optimal neighbourhood size when optimizing the LeadingOnes benchmark vi

Selective QA over Conflicting Multi-Source Personal Memory: A Diagnostic Testbed and Method Comparison

Model ReleasesDGX agent

arXiv:2605.30087v1 Announce Type: new Abstract: Emerging personal AI agents are moving toward persistent, multi-source memory. This creates an evaluation problem: systems must decide how to use confli

Self-Play Reinforcement Learning under Imperfect Information in Big 2

SafetyDGX agent

arXiv:2605.28863v1 Announce Type: cross Abstract: Imperfect-information multiplayer games test whether agents can act under hidden information, sparse rewards, and non-stationary opponents. We study t

Self-Trained Verification for Training- and Test-Time Self-Improvement

ResearchDGX agent

arXiv:2605.30290v1 Announce Type: cross Abstract: Self-improvement at scale has been a longstanding goal for reasoning models, and there are two natural places to do it: at test time, through verifica

Semantic and Visual Evidence for Efficient Long-Video Reasoning: A Solution for the HD-EPIC VQA Challenge

Model ReleasesDGX agent

arXiv:2605.29402v1 Announce Type: cross Abstract: Understanding long-form egocentric videos remains challenging for multimodal large language models (MLLMs) due to limited context length and insuffici

Sequential Physics-Constrained Neural Operator Forward Modeling for the extit{Norne} Reservoir System

Model ReleasesDGX agent

arXiv:2605.28909v1 Announce Type: new Abstract: We develop a comprehensive mathematical and computational framework for sequential surrogate modeling of three-phase black-oil reservoir dynamics using

← Previous
1…562563564565566…1049
Next →