AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Applications

QuantFlow: A Federated Mamba-Based Post-Transformer Foundation Model for Time-Series Forecasting

DGX agent

arXiv:2607.02632v1 Announce Type: cross Abstract: Time-series forecasting supports decisions in finance, en-ergy, transportation, public health, and industrial monitoring. Recent foundation models imp

applicationsarxiv-cs-ai
7 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Quantum-Inspired Harmonic Decision Models: A Computational Framework for Music Generation

DGX agent

arXiv:2607.05007v1 Announce Type: new Abstract: This paper introduces a quantum-inspired computational framework for harmonic decision-making in music. The proposed approach formulates harmonization a

researcharxiv-cs-ai
7 Jul 2026
Research

Quick ViTs: Speeding up Vision Transformers through Equivariance

DGX agent

arXiv:2505.15441v5 Announce Type: replace-cross Abstract: Natural images exhibit strong geometric regularities: local structures, such as edges, corners, and textures, appear in many orientations and

researcharxiv-cs-ai
7 Jul 2026
Research

Quickest Detection of Hallucination Onset: Delay Bounds and Learned CUSUM Statistics

DGX agent

arXiv:2606.12476v3 Announce Type: replace-cross Abstract: Token-level hallucination detectors are evaluated as classifiers, by AUC over all tokens, yet a streaming monitor is judged by its reaction ti

researcharxiv-cs-ai
7 Jul 2026
Safety

R^2PO: Decoupling Rollout and Inference Policies for LLM Reasoning

DGX agent

arXiv:2601.11960v3 Announce Type: replace-cross Abstract: Existing reinforcement learning methods for LLM reasoning implicitly assume that the policy generating training trajectories should coincide w

safetyarxiv-cs-ai
7 Jul 2026
Model Releases

R3D: Quantitative 3D Spatial Reasoning for Egocentric Wearables

DGX agent

arXiv:2607.02921v1 Announce Type: cross Abstract: Quantitative 3D spatial reasoning from egocentric RGB-D video is a critical capability for next-generation wearable assistants. Yet existing benchmark

model-releasesarxiv-cs-ai
7 Jul 2026
Safety

RADIO1D: Elastic Representations for Condensed Vision Modeling

DGX agent

arXiv:2607.03624v1 Announce Type: cross Abstract: This paper challenges the assumption that vision-language models (VLMs) require fixed patch-based 2D vision features. Analyzing fine-tuned vision enco

safetyarxiv-cs-ai
7 Jul 2026
Model Releases

Rational Inverse Reasoning: Few-Shot Imitation by Inferring Intent through Planning

DGX agent

arXiv:2508.08983v2 Announce Type: replace-cross Abstract: Humans can learn a new manipulation task from one or two demonstrations and then perform it in a new room, with new objects, under new constra

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Reading Between the Dots: Decoding Hidden Computation across Filler Tokens

DGX agent

arXiv:2607.03502v1 Announce Type: cross Abstract: Frontier LLMs can perform multi-step reasoning over content-free filler tokens like dots or counting sequences, producing correct answers with no visi

model-releasesarxiv-cs-ai
7 Jul 2026
Safety

Reason, Reward, Refine: Step-Level Errors Corrections with Structured Feedback for Physics Reasoning in Small Language Models

DGX agent

arXiv:2607.05199v1 Announce Type: new Abstract: Physics reasoning fails structurally in small language models: an error at any step propagates forward, corrupting every inference that follows. Limited

safetyarxiv-cs-ai
7 Jul 2026
Safety

REDDIT: Correcting Model-Generated Timestamp Drift in ASR without Forgetting via Replay-Based Distribution Editing

DGX agent

arXiv:2607.05364v1 Announce Type: cross Abstract: Modern autoregressive ASR systems can emit timestamps as decoded tokens, enabling timestamped transcription without frame-level aligners or inference-

safetyarxiv-cs-ai
7 Jul 2026
Model Releases

Reducing the Complexity of Deep Learning Models for EEG Analysis on Wearable Devices

DGX agent

arXiv:2606.12742v3 Announce Type: replace Abstract: Wearable healthcare devices are the fastest-growing Internet of Things (IoT) sector. Many automated healthcare services rely on two crucial biologic

model-releasesarxiv-cs-ai
7 Jul 2026
Tutorials

Reflective Dialogue or Prompt Refinement? Effects of Tutor Scaffolding on Students' Independent LLM Use for Programming

DGX agent

arXiv:2607.03303v1 Announce Type: new Abstract: While Large Language Models (LLMs) can provide personalized support in learning, several studies have raised concerns regarding their use in education.

tutorialsarxiv-cs-ai
7 Jul 2026
Model Releases

Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents

DGX agent

arXiv:2607.03968v1 Announce Type: cross Abstract: Large language models are increasingly deployed as IDE-integrated coding agents that decompose tasks, generate and edit files, run code, and refine ou

model-releasesarxiv-cs-ai
7 Jul 2026
Safety

Regime-Conditional Stabilisation of LLM-Augmented Cooperative Multi-Agent Reinforcement Learning

DGX agent

arXiv:2607.04470v1 Announce Type: cross Abstract: Large Language Models (LLMs) offer a natural interface for translating human objectives into reward signals for cooperative multi-agent reinforcement

safetyarxiv-cs-ai
7 Jul 2026
Agents

Reinforcement Learning for Evidence-Seeking Diagnostic Reasoning with Large Language Models

DGX agent

arXiv:2607.02983v1 Announce Type: new Abstract: Recent reasoning-centric Large Language Models (LLMs) have made significant strides, yet they predominantly operate on a passive-inference pattern that

agentsarxiv-cs-ai
7 Jul 2026
Safety

Relational Multi-Agent Reinforcement Learning for Dynamic Pricing in High-Speed Railway Markets

DGX agent

arXiv:2607.05179v1 Announce Type: cross Abstract: In liberalised railway systems, operators must set prices dynamically in an environment with partial observability, as they retain private information

safetyarxiv-cs-ai
7 Jul 2026
Research

ResearchStudio-Idea: An Evidence-Grounded Research-Ideation Skill Suite from ML Conference Outcomes

DGX agent

arXiv:2607.04439v1 Announce Type: new Abstract: Large language models have made research ideation increasingly accessible, yet effective idea development requires more than generating candidate direct

researcharxiv-cs-ai
7 Jul 2026
Model Releases

ResearchStudio-Reel: Automate the Last Mile of Research from Paper to Poster, Video, and Blog

DGX agent

arXiv:2607.04438v1 Announce Type: cross Abstract: Research dissemination, turning a paper into a poster, a talk video, and a blog post, is still a manual last mile. Prior automation treats each artifa

model-releasesarxiv-cs-ai
7 Jul 2026
Agents

Resilient by Design -- Active Inference for Distributed Continuum Intelligence

DGX agent

arXiv:2511.07202v3 Announce Type: replace-cross Abstract: Failures are the norm in highly complex and heterogeneous devices spanning the distributed computing continuum (DCC), from resource-constraine

agentsarxiv-cs-ai
7 Jul 2026
Applications

Resource-constrained Project Scheduling with Time-of-Use Energy Tariffs and Machine States: A Logic-based Benders Decomposition Approach

DGX agent

arXiv:2601.06542v2 Announce Type: replace-cross Abstract: In this paper, we investigate the Resource-Constrained Project Scheduling Problem (RCPSP) with Time-of-Use (TOU) energy tariffs and machine st

applicationsarxiv-cs-ai
7 Jul 2026
Model Releases

Responsibility Distribution Estimation in Ego-View Accident Videos with Multimodal Large Language Models

DGX agent

arXiv:2607.03591v1 Announce Type: cross Abstract: Recent studies on multimodal traffic accident understanding have mainly relied on infrastructure-camera footage, satellite imagery, or structured cras

model-releasesarxiv-cs-ai
7 Jul 2026
Research

Restricted Bernoulli Matrix Factorization: Balancing the trade-off between prediction accuracy and coverage in classification based collaborative filtering

DGX agent

arXiv:2210.10619v3 Announce Type: replace-cross Abstract: Reliability measures associated with the prediction of the machine learning models are critical to strengthening user confidence in artificial

researcharxiv-cs-ai
7 Jul 2026
Research

Rethinking Depth Pruning for Vision Transformers: A Heterogeneity-Aware Perspective

DGX agent

arXiv:2607.03784v1 Announce Type: cross Abstract: While prior studies have successfully compressed vision Transformers (ViTs) through various pruning techniques, most have concentrated on width prunin

researcharxiv-cs-ai
7 Jul 2026
Research

Rethinking Neural Nonlinearity as Gating

DGX agent

arXiv:2607.03148v1 Announce Type: cross Abstract: Activation functions are considered an essential primitive for neural nonlinearity, i.e., they enable neural networks to serve as universal approximat

researcharxiv-cs-ai
7 Jul 2026
Safety

Rethinking On-Policy Self-Distillation for Thinking Models

DGX agent

arXiv:2607.05184v1 Announce Type: new Abstract: Self-distillation is a promising recipe for self-improvement in language models. In this setting, a model can serve as its own teacher when given privil

safetyarxiv-cs-ai
7 Jul 2026
Model Releases

Retroactive Chain-of-Thought (RetroCoT): Forensic Reconstruction Prompts as a Safety Diagnostic Across Model Generations

DGX agent

arXiv:2607.04645v1 Announce Type: cross Abstract: Safety alignment in large language models is typically evaluated against direct, imperative harmful requests. We show that this alignment is highly co

model-releasesarxiv-cs-ai
7 Jul 2026
Tutorials

Revealing Hidden Model Behaviors with Task-Specific Self-Reports

DGX agent

arXiv:2607.03640v1 Announce Type: cross Abstract: Fine-tuning can give a language model a hidden behavior--it may give false answers under a narrow condition, or give harmful advice only when a prompt

tutorialsarxiv-cs-ai
7 Jul 2026
Safety

Reward-Gated On-Policy Distillation

DGX agent

arXiv:2607.04037v1 Announce Type: cross Abstract: On-policy distillation is a powerful way to transfer reasoning ability from a strong teacher to a smaller student: the student samples trajectories fr

safetyarxiv-cs-ai
7 Jul 2026
Model Releases

Risk-Constrained Freshness-Aware Semantic Caching for Open-Web Retrieval-Augmented LLMs

DGX agent

arXiv:2607.04281v1 Announce Type: cross Abstract: Semantic caching reduces the latency and cost of retrieval-augmented generation (RAG) by serving cached answers to semantically similar queries, but m

model-releasesarxiv-cs-ai
7 Jul 2026
Tutorials

RLIE: Rule Generation with Logistic Regression, Iterative Refinement, and Evaluation for Large Language Models

DGX agent

arXiv:2510.19698v3 Announce Type: replace Abstract: Large Language Models (LLMs) can propose rules in natural language, sidestepping the need for a predefined predicate space in traditional rule learn

tutorialsarxiv-cs-ai
7 Jul 2026
Model Releases

RoboDojo: A Unified Sim-and-Real Benchmark for Comprehensive Evaluation of Generalist Robot Manipulation Policies

DGX agent

arXiv:2607.04434v1 Announce Type: cross Abstract: Generalist robot manipulation policies have advanced rapidly, yet existing benchmarks remain limited in systematically evaluating their capabilities.

model-releasesarxiv-cs-ai
7 Jul 2026
Research

Robust Counterfactual Explanations under Model Multiplicity Using Multi-Objective Optimization

DGX agent

arXiv:2501.05795v4 Announce Type: replace-cross Abstract: In recent years, explainability in machine learning has gained importance. In this context, counterfactual explanation (CE), which is an expla

researcharxiv-cs-ai
7 Jul 2026
Model Releases

Robust Feasible Route Construction through Collaborative Partition Optimization

DGX agent

arXiv:2607.03694v1 Announce Type: new Abstract: Large-scale Capacitated Vehicle Routing Problems (CVRPs) are commonly solved by partitioning customers into smaller routing problems that can be optimiz

model-releasesarxiv-cs-ai
7 Jul 2026
Agents

Robustness Verification of an Autonomous Underwater Vehicle-based Plankton Classifier

DGX agent

arXiv:2607.04453v1 Announce Type: cross Abstract: The assessment of planktonic standing stocks and microorganism structures is critical for understanding upper ocean biological processes. Currently, a

agentsarxiv-cs-ai
7 Jul 2026
Safety

RSPO: Reward-Swap Policy Optimization for Multi-Turn LLM Agents

DGX agent

arXiv:2607.04713v1 Announce Type: cross Abstract: Reinforcement learning holds significant potential for training large language models (LLMs) to handle multi-turn interactive tasks. However, in long-

safetyarxiv-cs-ai
7 Jul 2026
Research

RUFNet: Query-Guided Support Mask Refinement and Uncertainty Fusion based on Hybrid Mamba for Few-Shot Brain Tumor Segmentation

DGX agent

arXiv:2607.05035v1 Announce Type: cross Abstract: Few-shot brain tumor segmentation remains challenging due to noisy support masks, inter-patient variations between support and query images, and the l

researcharxiv-cs-ai
7 Jul 2026
Model Releases

RustMizan: A Compilable, Contamination-Aware Benchmarking Framework for Rust Vulnerabilities

DGX agent

arXiv:2607.04729v1 Announce Type: cross Abstract: LLM agents are increasingly applied to vulnerability analysis, but existing benchmarks have not kept pace. They typically rely on small non-compilable

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

S-EMBER: A Large-Scale Benchmark for Streaming Egocentric Memory Retrieval

DGX agent

arXiv:2607.02689v1 Announce Type: cross Abstract: As wearable devices enable continuous first-person recording, AI assistants must reason across long time horizons to recall past experiences-a capabil

model-releasesarxiv-cs-ai
7 Jul 2026
Safety

Safe Inference-Time Alignment via Lagrangian Reward Augmentation

DGX agent

arXiv:2607.02781v1 Announce Type: cross Abstract: Inference-time alignment steers a frozen language model during decoding using auxiliary reward signals, avoiding the cost of repeated weight updates.

safetyarxiv-cs-ai
7 Jul 2026
Safety

Safe RLHF Beyond Expectation: Stochastic Dominance for Universal Spectral Risk Control

DGX agent

arXiv:2603.10938v2 Announce Type: replace-cross Abstract: Safe Reinforcement Learning from Human Feedback (RLHF) typically enforces safety through expected cost constraints, but the expectation captur

safetyarxiv-cs-ai
7 Jul 2026
Safety

Saving GPU Hours in LLM Inference System Development and Online Workloads with Simulation and DBMS-Inspired Cache Replacement Policies

DGX agent

arXiv:2411.07447v5 Announce Type: replace-cross Abstract: LLMs are increasingly used world-wide from daily tasks to agentic systems and data analytics, requiring significant GPU resources. While LLM i

safetyarxiv-cs-ai
7 Jul 2026
Research

Scalable Maximal Frequent Episode Mining with Desbordante

DGX agent

arXiv:2607.03188v1 Announce Type: cross Abstract: Episode mining aims to extract subsequences of events that possess certain distinctive properties and constitute facts valuable to the user. Maximal f

researcharxiv-cs-ai
7 Jul 2026
Safety

Scalable Semantic Steering of Embedding Projections

DGX agent

arXiv:2607.03978v1 Announce Type: cross Abstract: Low-dimensional projections support interactive visual analysis of high-dimensional data embeddings, but their structure often does not align with ana

safetyarxiv-cs-ai
7 Jul 2026
Research

Score-Regularized Joint Sampling with Importance Weights for Flow Matching

DGX agent

arXiv:2511.17812v3 Announce Type: replace-cross Abstract: Flow matching models effectively represent complex distributions, yet estimating expectations of functions of their outputs remains challengin

researcharxiv-cs-ai
7 Jul 2026
Agents

Search Beyond What Can Be Taught: Evolving the Knowledge Boundary in Agentic Visual Generation

DGX agent

arXiv:2607.05382v1 Announce Type: cross Abstract: Visual generators excel at rendering, but they confidently fabricate what they do not know. User requests are unbounded, evolving, and deeply long-tai

agentsarxiv-cs-ai
7 Jul 2026
Safety

Securing Multi-Tool AI Agent Chains With Dynamic, Real-Time Compositional Policies

DGX agent

arXiv:2607.03423v1 Announce Type: cross Abstract: Modern AI agent implementations such as frontier coding agents chain multiple tools at runtime that create a security surface that per-tool guardrails

safetyarxiv-cs-ai
7 Jul 2026
Model Releases

Seduced by the Narrative: Assessing Rule Adherence in Semi-Open Textual Sandboxes

DGX agent

arXiv:2607.02802v1 Announce Type: cross Abstract: As LLMs are increasingly deployed as autonomous adjudicators in semi-open textual game environments, robust rule adherence becomes critical when user

model-releasesarxiv-cs-ai
7 Jul 2026
← Previous
1…113114115116117…448
Next →