AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
9 Jul 2026

Non-contact, Real-time, Heart-rate Measurement using Image Processing with Commodity Cameras and AI Agents

AgentsDGX agent

arXiv:2607.06598v1 Announce Type: cross Abstract: Heart rate measurement is one of the key requirements for real-time health monitoring, in particular for health caring of elderly people. Traditional

NonTextual Target Attack

SafetyDGX agent

arXiv:2510.02999v5 Announce Type: replace-cross Abstract: Existing gradient-based jailbreak attacks on Large Language Models (LLMs) typically optimize adversarial suffixes to align the LLM output with

Object Search in Partially-Known Environments via LLM-informed Model-based Planning and Prompt Selection

ResearchDGX agent

arXiv:2603.23800v2 Announce Type: replace-cross Abstract: We present a novel LLM-informed model-based planning framework, and a novel prompt selection method, for object search in partially-known envi


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

On Adversarial Vulnerability of Vision-Language Models through the Lens of Intermediate Spectral Subspaces

ResearchDGX agent

arXiv:2607.07375v1 Announce Type: cross Abstract: Adversarial vulnerability in deep neural networks (DNNs) has been studied from the perspectives of decision-boundary geometry, feature robustness, inp

On the Principles of Deep Feedforward ReLU Networks

TutorialsDGX agent

arXiv:2607.07035v1 Announce Type: cross Abstract: The architecture of deep feedforward neural networks is ubiquitous in deep learning, either as a whole system or as a subnetwork of other architecture

Open-Ended Scenario Reasoning for Specialist Model Adaptation

SafetyDGX agent

arXiv:2607.06625v1 Announce Type: cross Abstract: Process industries have accumulated validated specialist models, yet sensor drift, feedstock variation, and regime switching cause these models to deg

Operational Reframing and Approval-Framed Delegation in Multi-Agent LLM Safety

Model ReleasesDGX agent

arXiv:2607.07097v1 Announce Type: new Abstract: Safety evaluations of multi-agent LLM systems often compare a direct prompt with a planner-executor pipeline and report the difference as a single 'pipe

ORCAID: Oblique Rule-Based Continuous-Action Interpretation for Deep RL Policies

SafetyDGX agent

arXiv:2607.07235v1 Announce Type: cross Abstract: Explainability remains a key issue in reinforcement learning (RL). Distilling an interpretable policy from an agent trained in a complex environment i

Overview of the NLPCC 2026 Shared Task 1: Difficulty-Aware Multilingual and Multimodal Medical Instructional Video Understanding Evaluation

Model ReleasesDGX agent

arXiv:2607.06618v1 Announce Type: cross Abstract: Following the CMIVQA, MMI-VQA, and M4IVQA challenges in NLPCC 2023--2025, we introduce the Difficulty-Aware Medical Instructional Video Question Answe

Physics-Audited Agentic Discovery in Scientific Machine Learning

AgentsDGX agent

arXiv:2607.07379v1 Announce Type: new Abstract: In agentic scientific machine learning (SciML), large language model (LLM) agents can discover surrogate models and select one by an automated score, ty

Physics-guided spatiotemporal neural models for fuel density prediction

ResearchDGX agent

arXiv:2607.06999v1 Announce Type: cross Abstract: This paper presents a physics-guided machine learning (PGML) framework for fuel density prediction, integrating physics constraints and domain knowled

POO-LPSP: Parallel Osprey Optimized Least Penalty-Squared Prioritization Methods for Priority Derivation in the Analytic Hierarchy Process

ResearchDGX agent

arXiv:2607.07313v1 Announce Type: cross Abstract: Pairwise comparison (PC) via pairwise reciprocal matrices (PRMs) is central to the Analytic Hierarchy Process (AHP). Although the traditional eigenvec

Power and Limitations of Aggregation in Compound AI Systems

AgentsDGX agent

arXiv:2602.21556v2 Announce Type: replace Abstract: When designing compound AI systems, a common approach is to query multiple copies of the same model and aggregate the responses to produce a synthes

Predicting LLM Safety Before Release by Simulating Deployment

Model ReleasesDGX agent

arXiv:2607.07184v1 Announce Type: cross Abstract: Pre-deployment safety evaluations aim to inform the downstream risks of releasing a new AI model. Yet most evaluations provide limited evidence about

Progressive Crystallization: Turning Agent Exploration into Deterministic, Lower-Cost Workflows in Production

SafetyDGX agent

arXiv:2607.07052v1 Announce Type: cross Abstract: AI agents deployed for IT operations are typically permanent cost centers because every execution requires full LLM inference, even for previously sol

ProMoE-FL: Prototype-conditioned Mixture of Experts for Multimodal Federated Learning with Missing Modalities

ResearchDGX agent

arXiv:2607.06633v1 Announce Type: cross Abstract: In this paper, we address the problem of multimodal federated learning with missing modality. Existing methods utilize an additional public dataset or

PRoVeFL: Private Robust and Verifiable Aggregation in Federated Learning

Local AiDGX agent

arXiv:2607.06612v1 Announce Type: cross Abstract: Federated Learning (FL) enables multiple clients to collaboratively train machine learning models while retaining data locality, thereby enhancing use

QANTIS: Hardware-Calibrated Sequential POMDP Belief Updates on IBM Heron

AgentsDGX agent

arXiv:2607.06760v1 Announce Type: new Abstract: Autonomous systems under partial observability act on beliefs, not raw sensor events. QANTIS treats the quantum processor as a calibrated belief-update

QCNN with Rough Path Signature Kernels

ResearchDGX agent

arXiv:2607.07634v1 Announce Type: cross Abstract: Time series analysis plays a vital role across a wide range of scientific and engineering domains but poses substantial computational challenges. A ma

Quantum simulation of real-world nonlinear dynamics via Koopman method

ApplicationsDGX agent

arXiv:2607.07338v1 Announce Type: cross Abstract: Nonlinear dynamics is ubiquitous in nature, ranging from chemical pattern formation to ocean circulation, yet its simulation on quantum computers is f

Rapidly Learning Soft Robot Control via Implicit Time-Stepping

SafetyDGX agent

arXiv:2511.06667v2 Announce Type: replace-cross Abstract: With the explosive growth of rigid-body simulators, policy learning in simulation has become the de facto standard for most rigid morphologies

Reason Less, Verify More: Deterministic Gates Recover a Silent Policy-Violation Failure Mode in Tool-Using LLM Agents

Model ReleasesDGX agent

arXiv:2607.07405v1 Announce Type: new Abstract: Tool-using LLM agents can violate the very policies they are deployed to enforce while appearing to complete the task successfully. In policy-permissive

Reasoning Consistency Scanning: A Framework for Auditing Chain-of-Thought Validity in AI Safety Evaluations

Model ReleasesDGX agent

arXiv:2607.07229v1 Announce Type: new Abstract: Prior work has shown that chain-of-thought (CoT) reasoning is often unfaithful: a model's stated reasoning does not reliably reflect the process that pr

Recursive Self-Improvement in AI: From Bounded Self-Refinement to Autonomous Research Loops

SafetyDGX agent

arXiv:2607.07663v1 Announce Type: new Abstract: AI systems increasingly participate in their own improvement: revising their outputs, adapting their own harnesses during deployment, training on data t

Refine Thought: A Test-Time Inference Method for Embedding Model Reasoning

Model ReleasesDGX agent

arXiv:2511.13726v2 Announce Type: replace-cross Abstract: We propose RT (Refine Thought), a method that can enhance the semantic reasoning ability of text embedding models. The method obtains the fina

Reliable and Developer-Aligned Evaluation of Agents for Software Engineering

AgentsDGX agent

arXiv:2607.06713v1 Announce Type: cross Abstract: Large language models are rapidly moving towards closing the development cycle, transitioning from simple assistive companions to autonomous contribut

ReMoDEx: A Local-to-Global Relevance-Based Model Decision Explainability Framework for large-Scale Image Datasets

Local AiDGX agent

arXiv:2607.06889v1 Announce Type: cross Abstract: Deep learning image classifiers achieve strong predictive performance yet remain opaque in how decisions are formed. A model may predict correctly whi

Reward-Adaptive Iterative Discovery: A Case Study on Automated Game Testing for NHL26

ApplicationsDGX agent

arXiv:2607.07498v1 Announce Type: cross Abstract: Testing is a major effort for the gaming industry, requiring a significant part of development budget and people power. We present a case study on a d

Riemannian Geometry for Pre-trained Language Model Embeddings

Model ReleasesDGX agent

arXiv:2607.07047v1 Announce Type: cross Abstract: Understanding the geometric structure of pre-trained language model embeddings matters for interpretability and safety. We ask whether sentence-level

RL Post-Training Builds Compositional Reasoning Strategies

ResearchDGX agent

arXiv:2607.07646v1 Announce Type: new Abstract: Does RL post-training merely amplify primitive skills already latent in a base model, or can it compose primitive skills into new higher-level strategie

RLVP: Penalize the Path, Reward the Outcome

AgentsDGX agent

arXiv:2607.07435v1 Announce Type: cross Abstract: Agents acting on our behalf in the real world (e.g. placing phone calls) must learn online from costly, often irreversible interactions rather than ch

Search, Fail, Recover: A Training Framework for Correction-Aware Reasoning

ResearchDGX agent

arXiv:2607.07492v1 Announce Type: new Abstract: Many reasoning tasks are not well described by a single left-to-right chain: a solver may need to pursue a plausible branch, observe delayed failure, an

Security and Privacy in Agentic AI: Grand Challenges and Future Directions

AgentsDGX agent

arXiv:2607.06608v1 Announce Type: cross Abstract: We present key challenges and future research directions in the security and privacy of agentic AI, based on a horizon-scanning exercise that brought

Selective Timestep Weighting and Advantage-Based Replay for Sample-Efficient Diffusion RLHF

SafetyDGX agent

arXiv:2607.07693v1 Announce Type: cross Abstract: Reinforcement learning from human feedback (RLHF) has emerged as a powerful paradigm for aligning generative models with human preferences. However, a

Self-Supervised Pretraining Improves Cross-Site and Cross-Scale Robustness of Point Cloud Leaf-Wood Segmentation

Model ReleasesDGX agent

arXiv:2607.06948v1 Announce Type: cross Abstract: The accuracy of existing leaf-wood segmentation methods for tree point clouds varies across forest types and sites. Self-supervised learning (SSL) on

Shared Modular Recurrence in Contextual MDPs for Universal Morphology Control

AgentsDGX agent

arXiv:2506.08630v3 Announce Type: replace Abstract: A universal controller for any robot morphology would greatly improve computational and data efficiency. Steps have been made towards such multi-rob

Single-Rollout Asynchronous Optimization for Agentic Reinforcement Learning

SafetyDGX agent

arXiv:2607.07508v1 Announce Type: cross Abstract: Reinforcement learning (RL) is becoming increasingly important for post-training large language models (LLMs). Previous RL pipelines for LLMs were mos

SkillCenter: A Large-Scale Source-Grounded Skill Library for Autonomous AI Agents

AgentsDGX agent

arXiv:2607.07676v1 Announce Type: new Abstract: Autonomous AI agents can execute complex tasks with limited human review, yet they often lack the grounded operational knowledge to make their outputs n

SmartHomeSecure: Automated Detection and Repair of Smart Home Configuration Errors Using Large Language Models

Model ReleasesDGX agent

arXiv:2607.06748v1 Announce Type: cross Abstract: Smart home automation platforms increasingly rely on user-authored YAML configuration files to define device behaviors, but these files are prone to s

SOMtime the World Ain't Fair: Violating Fairness Using Self-Organizing Maps

SafetyDGX agent

arXiv:2602.18201v2 Announce Type: replace Abstract: Unsupervised representations are widely assumed to be neutral with respect to sensitive attributes when those attributes are withheld from training.

SpaCellAgent: A Self-Evolving LLM-Based Multi-Agent Framework for Trajectory Analysis

AgentsDGX agent

arXiv:2607.07467v1 Announce Type: new Abstract: Spatial and Single-cell transcriptomics are transformative in deciphering cellular dynamics. As the fundamental paradigm for reconstructing cell develop

SpaR3D-MoE: Adaptive 3D Spatial Reasoning from Sparse Views Meets Geometry-Inductive Mixture-of-Experts

ResearchDGX agent

arXiv:2607.06620v1 Announce Type: cross Abstract: Recent Multimodal Large Language Models (MLLMs) struggle to bridge the representational gap between 2D semantic understanding and 3D spatial geometry.

Spatiotemporal Semantic V2X Framework for Cooperative Collision Prediction

SafetyDGX agent

arXiv:2601.17216v3 Announce Type: replace-cross Abstract: Intelligent Transportation Systems (ITS) demand real-time collision prediction to ensure road safety and reduce accident severity. Conventiona

SPEAR: A Simulator for Photorealistic Embodied AI Research

ResearchDGX agent

arXiv:2607.06701v1 Announce Type: cross Abstract: Interactive simulators have become powerful tools for training embodied agents and generating synthetic visual data, but existing photorealistic simul

Specification Grounding Drives Test Effectiveness for LLM Code

Model ReleasesDGX agent

arXiv:2607.06636v1 Announce Type: cross Abstract: Large language models frequently generate code that appears correct on typical inputs yet fails on edge cases, invalid inputs, and other specification

Stability of Flow Models for Graph Signals

ApplicationsDGX agent

arXiv:2607.07510v1 Announce Type: cross Abstract: Generating signals on graphs requires permutation-equivariant models that exhibit stability with respect to relative structural perturbations. While f

STAGformer: A Spatio-temporal Agent Graph Transformer for Micro Mobility Demand Forecasting

AgentsDGX agent

arXiv:2607.06614v1 Announce Type: cross Abstract: Accurate station-level demand forecasting is essential for the efficient operation of bike-sharing systems, yet it remains challenging due to complex

Successor-Generator Planning with LLM-generated Heuristics

ResearchDGX agent

arXiv:2501.18784v5 Announce Type: replace Abstract: Heuristics are a central component of deterministic planning, particularly in domain-independent settings where general applicability is prioritized

SynthAVE: Scalable Synthetic Labeling for E-Commerce with LLM-Arena Validation

Model ReleasesDGX agent

arXiv:2607.07469v1 Announce Type: cross Abstract: Fine-tuning large language models (LLMs) for e-commerce attribute extraction requires labeled data representative across thousands of product types, a

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs

SafetyDGX agent

arXiv:2511.16107v3 Announce Type: replace-cross Abstract: Visual in-context learning (VICL) solves visual tasks by conditioning on a few input-output demonstrations without any model training. Recent

The Blind Curator: How a Biased Judge Silently Disables Skill Retirement in Self-Evolving Agents

SafetyDGX agent

arXiv:2607.07436v1 Announce Type: new Abstract: A self-evolving agent retires its bad skills by watching them fail, so what happens when the judge cannot see the failures? Skill retirement is the stru

The Harness Effect: How Orchestration Design Sets the Token Economics of Enterprise Agentic AI

Model ReleasesDGX agent

arXiv:2607.06906v1 Announce Type: new Abstract: Agentic AI development today runs on token maxing: buying capability with tokens -- longer reasoning traces, more turns, wider tool payloads, bigger rep

The Rank-One Corner: How Much Value Equivalence Does a Task Need from a World Model?

ResearchDGX agent

arXiv:2607.06640v1 Announce Type: cross Abstract: A learned world model is usually judged by how faithfully it reconstructs its observations or predicts reward, as though quality were something the mo

The Signs Were Always There: Training-Free Concept Detection and Steering in Raw Transformer Dimensions

TutorialsDGX agent

arXiv:2606.12629v3 Announce Type: replace-cross Abstract: The standard basis of transformer hidden states is a training-free, architecture-general feature basis for detecting concepts and, in language

Thinking Ahead: Foresight Intelligence in MLLMs and World Model

Model ReleasesDGX agent

arXiv:2511.18735v3 Announce Type: replace-cross Abstract: In this work, we define Foresight Intelligence as the capability to anticipate and interpret future events-an ability essential for applicatio

TimEE: End-to-end Time Series Classification via In-Context Learning

Model ReleasesDGX agent

arXiv:2607.07500v1 Announce Type: cross Abstract: Time series classification (TSC) is dominated by a two-stage paradigm: train a feature encoder -- either from scratch on the target dataset or via pre

Towards Agentic AI Governance: A Preliminary Assessment

AgentsDGX agent

arXiv:2607.07612v1 Announce Type: cross Abstract: Artificial intelligence is rapidly evolving from generative systems to agentic AI capable of autonomously planning and executing tasks. Widely charact

Tree-of-Thoughts Reasoning for Text-to-Image In-Context Learning

Model ReleasesDGX agent

arXiv:2607.07117v1 Announce Type: cross Abstract: In text-to-image in-context learning (T2I-ICL), a model has to infer a latent compositional pattern from fewshot demonstrations for generating a query

TriRoute: Unified Learned Routing for Joint Adaptive Attention, Experts, and KV-Cache Allocation

Local AiDGX agent

arXiv:2607.06601v1 Announce Type: cross Abstract: Conditional computation can decouple language model quality from per-token inference cost, yet leading techniques act on a single axis in isolation: M

tsbootstrap: Distribution-Free Uncertainty Quantification and Conformal Prediction for Time Series

ApplicationsDGX agent

arXiv:2607.06690v1 Announce Type: cross Abstract: Finance, sensing, and demand streams violate the exchangeability that IID conformal prediction and the IID bootstrap assume, and existing libraries im

← Previous
1…7879808182…358
Next →