AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
7 Jul 2026

Prima.cpp: Fast 30-70B LLM Inference on Heterogeneous and Low-Resource Home Clusters

Model ReleasesDGX agent

arXiv:2504.08791v3 Announce Type: replace-cross Abstract: On-device inference offers privacy, offline use, and instant response, but consumer hardware restricts large language models (LLMs) to low thr

Privacy-Preserving Robustness Verification for Neural Networks

ResearchDGX agent

arXiv:2607.05251v1 Announce Type: cross Abstract: Neural network verification and data privacy are inherently in tension: verification demands full access to model parameters and input data, yet both

Probe, Don't Prompt: A Hidden-State Probe for Metadata Filtering in Multi-Meta-RAG

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2607.03929v1 Announce Type: cross Abstract: Multi-Meta-RAG improves retrieval for multi-hop question answering by filtering a vector store on metadata (the news source) that it extracts from eac

Probing Low-Level Acoustic Attribute Encoding in CLAP Audio Embeddings

ResearchDGX agent

arXiv:2607.03806v1 Announce Type: cross Abstract: Audio foundation models are widely adopted as general-purpose feature extractors, yet the internal structure of their learned representations remains

Programming over Thinking: Efficient and Robust Multi-Constraint Planning

ResearchDGX agent

arXiv:2601.09097v3 Announce Type: replace Abstract: Multi-constraint planning involves identifying, evaluating, and refining candidate plans while satisfying multiple, potentially conflicting constrai

Progress- and Reliability-Oriented Group Policy Optimization for Agentic Reinforcement Learning

SafetyDGX agent

arXiv:2607.04242v1 Announce Type: new Abstract: Group-based reinforcement learning (RL) has become an effective paradigm for improving large language model agents on long-horizon interactive tasks. To

PromptPET: Privacy-Utility Optimized Prompt Obfuscation

AgentsDGX agent

arXiv:2607.02932v1 Announce Type: cross Abstract: Privacy is an important challenge when users interact with AI chatbots, since users may share sensitive information, explicitly or implicitly, and AI

ProPS: Prompted Profile Synthesis for Natural Language-Conditioned Speaker Embedding Distributions

ResearchDGX agent

arXiv:2607.05276v1 Announce Type: cross Abstract: Speaker embeddings, or x-vectors, are widely used to represent speaker identity and speaker-related attributes, but existing embedding extractors are

PulmoSight-XAI: An Explainable Multi-View Attention Ensemble with Gradient Boosting Meta-Learning for Multi-Label Chest X-Ray Classification

Local AiDGX agent

arXiv:2607.04478v1 Announce Type: cross Abstract: Automated chest X-ray classification remains challenging due to severe class imbalance, co-occurring pathologies, and the loss of localized features i

Punching Above Their Weight: Classification-Head Fine-Tuning of Tiny Language Models (TLMs) for Verifiable Multiple-Choice Tasks

Model ReleasesDGX agent

arXiv:2607.03801v1 Announce Type: cross Abstract: We define Tiny Language Models (TLMs) as models below roughly 3B parameters that fit on mainstream consumer devices. We study how to adapt them for an

PuzzleMoE: Efficient Compression of Large Mixture-of-Experts Models via Sparse Expert Merging and Bit-packed inference

ResearchDGX agent

arXiv:2511.04805v2 Announce Type: replace-cross Abstract: Mixture-of-Experts (MoE) models have shown strong potential in scaling language models efficiently by activating only a small subset of expert

Q-TriM: Question-Guided Tri-Modal Attention for Audio-Visual Question Answering

ResearchDGX agent

arXiv:2607.03825v1 Announce Type: cross Abstract: Audio-Visual Question Answering (AVQA) extends classical VQA by requiring joint reasoning over video and synchronized audio. However, many AVQA system

QuantFlow: A Federated Mamba-Based Post-Transformer Foundation Model for Time-Series Forecasting

ApplicationsDGX agent

arXiv:2607.02632v1 Announce Type: cross Abstract: Time-series forecasting supports decisions in finance, en-ergy, transportation, public health, and industrial monitoring. Recent foundation models imp

Quantum-Inspired Harmonic Decision Models: A Computational Framework for Music Generation

ResearchDGX agent

arXiv:2607.05007v1 Announce Type: new Abstract: This paper introduces a quantum-inspired computational framework for harmonic decision-making in music. The proposed approach formulates harmonization a

Quick ViTs: Speeding up Vision Transformers through Equivariance

ResearchDGX agent

arXiv:2505.15441v5 Announce Type: replace-cross Abstract: Natural images exhibit strong geometric regularities: local structures, such as edges, corners, and textures, appear in many orientations and

Quickest Detection of Hallucination Onset: Delay Bounds and Learned CUSUM Statistics

ResearchDGX agent

arXiv:2606.12476v3 Announce Type: replace-cross Abstract: Token-level hallucination detectors are evaluated as classifiers, by AUC over all tokens, yet a streaming monitor is judged by its reaction ti

R^2PO: Decoupling Rollout and Inference Policies for LLM Reasoning

SafetyDGX agent

arXiv:2601.11960v3 Announce Type: replace-cross Abstract: Existing reinforcement learning methods for LLM reasoning implicitly assume that the policy generating training trajectories should coincide w

R3D: Quantitative 3D Spatial Reasoning for Egocentric Wearables

Model ReleasesDGX agent

arXiv:2607.02921v1 Announce Type: cross Abstract: Quantitative 3D spatial reasoning from egocentric RGB-D video is a critical capability for next-generation wearable assistants. Yet existing benchmark

RADIO1D: Elastic Representations for Condensed Vision Modeling

SafetyDGX agent

arXiv:2607.03624v1 Announce Type: cross Abstract: This paper challenges the assumption that vision-language models (VLMs) require fixed patch-based 2D vision features. Analyzing fine-tuned vision enco

Rational Inverse Reasoning: Few-Shot Imitation by Inferring Intent through Planning

Model ReleasesDGX agent

arXiv:2508.08983v2 Announce Type: replace-cross Abstract: Humans can learn a new manipulation task from one or two demonstrations and then perform it in a new room, with new objects, under new constra

Reading Between the Dots: Decoding Hidden Computation across Filler Tokens

Model ReleasesDGX agent

arXiv:2607.03502v1 Announce Type: cross Abstract: Frontier LLMs can perform multi-step reasoning over content-free filler tokens like dots or counting sequences, producing correct answers with no visi

Reason, Reward, Refine: Step-Level Errors Corrections with Structured Feedback for Physics Reasoning in Small Language Models

SafetyDGX agent

arXiv:2607.05199v1 Announce Type: new Abstract: Physics reasoning fails structurally in small language models: an error at any step propagates forward, corrupting every inference that follows. Limited

REDDIT: Correcting Model-Generated Timestamp Drift in ASR without Forgetting via Replay-Based Distribution Editing

SafetyDGX agent

arXiv:2607.05364v1 Announce Type: cross Abstract: Modern autoregressive ASR systems can emit timestamps as decoded tokens, enabling timestamped transcription without frame-level aligners or inference-

Reducing the Complexity of Deep Learning Models for EEG Analysis on Wearable Devices

Model ReleasesDGX agent

arXiv:2606.12742v3 Announce Type: replace Abstract: Wearable healthcare devices are the fastest-growing Internet of Things (IoT) sector. Many automated healthcare services rely on two crucial biologic

Reflective Dialogue or Prompt Refinement? Effects of Tutor Scaffolding on Students' Independent LLM Use for Programming

TutorialsDGX agent

arXiv:2607.03303v1 Announce Type: new Abstract: While Large Language Models (LLMs) can provide personalized support in learning, several studies have raised concerns regarding their use in education.

Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents

Model ReleasesDGX agent

arXiv:2607.03968v1 Announce Type: cross Abstract: Large language models are increasingly deployed as IDE-integrated coding agents that decompose tasks, generate and edit files, run code, and refine ou

Regime-Conditional Stabilisation of LLM-Augmented Cooperative Multi-Agent Reinforcement Learning

SafetyDGX agent

arXiv:2607.04470v1 Announce Type: cross Abstract: Large Language Models (LLMs) offer a natural interface for translating human objectives into reward signals for cooperative multi-agent reinforcement

Reinforcement Learning for Evidence-Seeking Diagnostic Reasoning with Large Language Models

AgentsDGX agent

arXiv:2607.02983v1 Announce Type: new Abstract: Recent reasoning-centric Large Language Models (LLMs) have made significant strides, yet they predominantly operate on a passive-inference pattern that

Relational Multi-Agent Reinforcement Learning for Dynamic Pricing in High-Speed Railway Markets

SafetyDGX agent

arXiv:2607.05179v1 Announce Type: cross Abstract: In liberalised railway systems, operators must set prices dynamically in an environment with partial observability, as they retain private information

ResearchStudio-Idea: An Evidence-Grounded Research-Ideation Skill Suite from ML Conference Outcomes

ResearchDGX agent

arXiv:2607.04439v1 Announce Type: new Abstract: Large language models have made research ideation increasingly accessible, yet effective idea development requires more than generating candidate direct

ResearchStudio-Reel: Automate the Last Mile of Research from Paper to Poster, Video, and Blog

Model ReleasesDGX agent

arXiv:2607.04438v1 Announce Type: cross Abstract: Research dissemination, turning a paper into a poster, a talk video, and a blog post, is still a manual last mile. Prior automation treats each artifa

Resilient by Design -- Active Inference for Distributed Continuum Intelligence

AgentsDGX agent

arXiv:2511.07202v3 Announce Type: replace-cross Abstract: Failures are the norm in highly complex and heterogeneous devices spanning the distributed computing continuum (DCC), from resource-constraine

Resource-constrained Project Scheduling with Time-of-Use Energy Tariffs and Machine States: A Logic-based Benders Decomposition Approach

ApplicationsDGX agent

arXiv:2601.06542v2 Announce Type: replace-cross Abstract: In this paper, we investigate the Resource-Constrained Project Scheduling Problem (RCPSP) with Time-of-Use (TOU) energy tariffs and machine st

Responsibility Distribution Estimation in Ego-View Accident Videos with Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2607.03591v1 Announce Type: cross Abstract: Recent studies on multimodal traffic accident understanding have mainly relied on infrastructure-camera footage, satellite imagery, or structured cras

Restricted Bernoulli Matrix Factorization: Balancing the trade-off between prediction accuracy and coverage in classification based collaborative filtering

ResearchDGX agent

arXiv:2210.10619v3 Announce Type: replace-cross Abstract: Reliability measures associated with the prediction of the machine learning models are critical to strengthening user confidence in artificial

Rethinking Depth Pruning for Vision Transformers: A Heterogeneity-Aware Perspective

ResearchDGX agent

arXiv:2607.03784v1 Announce Type: cross Abstract: While prior studies have successfully compressed vision Transformers (ViTs) through various pruning techniques, most have concentrated on width prunin

Rethinking Neural Nonlinearity as Gating

ResearchDGX agent

arXiv:2607.03148v1 Announce Type: cross Abstract: Activation functions are considered an essential primitive for neural nonlinearity, i.e., they enable neural networks to serve as universal approximat

Rethinking On-Policy Self-Distillation for Thinking Models

SafetyDGX agent

arXiv:2607.05184v1 Announce Type: new Abstract: Self-distillation is a promising recipe for self-improvement in language models. In this setting, a model can serve as its own teacher when given privil

Retroactive Chain-of-Thought (RetroCoT): Forensic Reconstruction Prompts as a Safety Diagnostic Across Model Generations

Model ReleasesDGX agent

arXiv:2607.04645v1 Announce Type: cross Abstract: Safety alignment in large language models is typically evaluated against direct, imperative harmful requests. We show that this alignment is highly co

Revealing Hidden Model Behaviors with Task-Specific Self-Reports

TutorialsDGX agent

arXiv:2607.03640v1 Announce Type: cross Abstract: Fine-tuning can give a language model a hidden behavior--it may give false answers under a narrow condition, or give harmful advice only when a prompt

Reward-Gated On-Policy Distillation

SafetyDGX agent

arXiv:2607.04037v1 Announce Type: cross Abstract: On-policy distillation is a powerful way to transfer reasoning ability from a strong teacher to a smaller student: the student samples trajectories fr

Risk-Constrained Freshness-Aware Semantic Caching for Open-Web Retrieval-Augmented LLMs

Model ReleasesDGX agent

arXiv:2607.04281v1 Announce Type: cross Abstract: Semantic caching reduces the latency and cost of retrieval-augmented generation (RAG) by serving cached answers to semantically similar queries, but m

RLIE: Rule Generation with Logistic Regression, Iterative Refinement, and Evaluation for Large Language Models

TutorialsDGX agent

arXiv:2510.19698v3 Announce Type: replace Abstract: Large Language Models (LLMs) can propose rules in natural language, sidestepping the need for a predefined predicate space in traditional rule learn

RoboDojo: A Unified Sim-and-Real Benchmark for Comprehensive Evaluation of Generalist Robot Manipulation Policies

Model ReleasesDGX agent

arXiv:2607.04434v1 Announce Type: cross Abstract: Generalist robot manipulation policies have advanced rapidly, yet existing benchmarks remain limited in systematically evaluating their capabilities.

Robust Counterfactual Explanations under Model Multiplicity Using Multi-Objective Optimization

ResearchDGX agent

arXiv:2501.05795v4 Announce Type: replace-cross Abstract: In recent years, explainability in machine learning has gained importance. In this context, counterfactual explanation (CE), which is an expla

Robust Feasible Route Construction through Collaborative Partition Optimization

Model ReleasesDGX agent

arXiv:2607.03694v1 Announce Type: new Abstract: Large-scale Capacitated Vehicle Routing Problems (CVRPs) are commonly solved by partitioning customers into smaller routing problems that can be optimiz

Robustness Verification of an Autonomous Underwater Vehicle-based Plankton Classifier

AgentsDGX agent

arXiv:2607.04453v1 Announce Type: cross Abstract: The assessment of planktonic standing stocks and microorganism structures is critical for understanding upper ocean biological processes. Currently, a

RSPO: Reward-Swap Policy Optimization for Multi-Turn LLM Agents

SafetyDGX agent

arXiv:2607.04713v1 Announce Type: cross Abstract: Reinforcement learning holds significant potential for training large language models (LLMs) to handle multi-turn interactive tasks. However, in long-

RUFNet: Query-Guided Support Mask Refinement and Uncertainty Fusion based on Hybrid Mamba for Few-Shot Brain Tumor Segmentation

ResearchDGX agent

arXiv:2607.05035v1 Announce Type: cross Abstract: Few-shot brain tumor segmentation remains challenging due to noisy support masks, inter-patient variations between support and query images, and the l

RustMizan: A Compilable, Contamination-Aware Benchmarking Framework for Rust Vulnerabilities

Model ReleasesDGX agent

arXiv:2607.04729v1 Announce Type: cross Abstract: LLM agents are increasingly applied to vulnerability analysis, but existing benchmarks have not kept pace. They typically rely on small non-compilable

S-EMBER: A Large-Scale Benchmark for Streaming Egocentric Memory Retrieval

Model ReleasesDGX agent

arXiv:2607.02689v1 Announce Type: cross Abstract: As wearable devices enable continuous first-person recording, AI assistants must reason across long time horizons to recall past experiences-a capabil

Safe Inference-Time Alignment via Lagrangian Reward Augmentation

SafetyDGX agent

arXiv:2607.02781v1 Announce Type: cross Abstract: Inference-time alignment steers a frozen language model during decoding using auxiliary reward signals, avoiding the cost of repeated weight updates.

Safe RLHF Beyond Expectation: Stochastic Dominance for Universal Spectral Risk Control

SafetyDGX agent

arXiv:2603.10938v2 Announce Type: replace-cross Abstract: Safe Reinforcement Learning from Human Feedback (RLHF) typically enforces safety through expected cost constraints, but the expectation captur

Saving GPU Hours in LLM Inference System Development and Online Workloads with Simulation and DBMS-Inspired Cache Replacement Policies

SafetyDGX agent

arXiv:2411.07447v5 Announce Type: replace-cross Abstract: LLMs are increasingly used world-wide from daily tasks to agentic systems and data analytics, requiring significant GPU resources. While LLM i

Scalable Maximal Frequent Episode Mining with Desbordante

ResearchDGX agent

arXiv:2607.03188v1 Announce Type: cross Abstract: Episode mining aims to extract subsequences of events that possess certain distinctive properties and constitute facts valuable to the user. Maximal f

Scalable Semantic Steering of Embedding Projections

SafetyDGX agent

arXiv:2607.03978v1 Announce Type: cross Abstract: Low-dimensional projections support interactive visual analysis of high-dimensional data embeddings, but their structure often does not align with ana

Score-Regularized Joint Sampling with Importance Weights for Flow Matching

ResearchDGX agent

arXiv:2511.17812v3 Announce Type: replace-cross Abstract: Flow matching models effectively represent complex distributions, yet estimating expectations of functions of their outputs remains challengin

Search Beyond What Can Be Taught: Evolving the Knowledge Boundary in Agentic Visual Generation

AgentsDGX agent

arXiv:2607.05382v1 Announce Type: cross Abstract: Visual generators excel at rendering, but they confidently fabricate what they do not know. User requests are unbounded, evolving, and deeply long-tai

Securing Multi-Tool AI Agent Chains With Dynamic, Real-Time Compositional Policies

SafetyDGX agent

arXiv:2607.03423v1 Announce Type: cross Abstract: Modern AI agent implementations such as frontier coding agents chain multiple tools at runtime that create a security surface that per-tool guardrails

Seduced by the Narrative: Assessing Rule Adherence in Semi-Open Textual Sandboxes

Model ReleasesDGX agent

arXiv:2607.02802v1 Announce Type: cross Abstract: As LLMs are increasingly deployed as autonomous adjudicators in semi-open textual game environments, robust rule adherence becomes critical when user

← Previous
1…9091929394…358
Next →