AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
20,981 results
14 Apr 2026

Process-Centric Analysis of Agentic Software Systems

AgentsDGX agent

arXiv:2512.02393v3 Announce Type: replace-cross Abstract: Agentic systems are modern software systems: they consist of orchestrated modules, expose interfaces, and are deployed in software pipelines.

Product Review Based on Optimized Facial Expression Detection

ResearchDGX agent

arXiv:2604.10885v1 Announce Type: cross Abstract: This paper proposes a method to review public acceptance of products based on their brand by analyzing the facial expression of the customer intending

Progressive Multimodal Interaction Network for Reliable Quantification of Fish Feeding Intensity in Aquaculture

Model ReleasesDGX agent

arXiv:2506.14170v3 Announce Type: replace-cross Abstract: Accurate quantification of fish feeding intensity is crucial for precision feeding in aquaculture, as it directly affects feed utilization and


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Prompt Injection as Role Confusion

SafetyDGX agent

arXiv:2603.12277v3 Announce Type: replace-cross Abstract: Language models remain vulnerable to prompt injection attacks despite extensive safety training. We trace this failure to role confusion: mode

Prosociality by Coupling, Not Mere Observation: Homeostatic Sharing in an Inspectable Recurrent Artificial Life Agent

AgentsDGX agent

arXiv:2604.10760v1 Announce Type: cross Abstract: Artificial agents can be made to 'help' for many reasons, including explicit social reward, hard-coded prosocial bonuses, or direct access to another

Proximal Supervised Fine-Tuning

SafetyDGX agent

arXiv:2508.17784v2 Announce Type: replace-cross Abstract: Supervised fine-tuning (SFT) of foundation models often leads to poor generalization, where prior capabilities deteriorate after tuning on new

Pseudo-Unification: Entropy Probing Reveals Divergent Information Patterns in Unified Multimodal Models

ResearchDGX agent

arXiv:2604.10949v1 Announce Type: cross Abstract: Unified multimodal models (UMMs) were designed to combine the reasoning ability of large language models (LLMs) with the generation capability of visi

Putting the Value Back in RL: Better Test-Time Scaling by Unifying LLM Reasoners With Verifiers

ResearchDGX agent

arXiv:2505.04842v2 Announce Type: replace-cross Abstract: Prevalent reinforcement learning~(RL) methods for fine-tuning LLM reasoners, such as GRPO or Leave-one-out PPO, abandon the learned value func

Pyramid MoA: A Probabilistic Framework for Cost-Optimized Anytime Inference

ResearchDGX agent

arXiv:2602.19509v3 Announce Type: replace-cross Abstract: We observe that LLM cascading and routing implicitly solves an anytime computation problem -- a class of algorithms, well-studied in classical

QShield: Securing Neural Networks Against Adversarial Attacks using Quantum Circuits

SafetyDGX agent

arXiv:2604.10933v1 Announce Type: cross Abstract: Deep neural networks remain highly vulnerable to adversarial perturbations, limiting their reliability in security- and safety-critical applications.

Quantitative Introspection in Language Models: Tracking Emotive States Across Conversation

Model ReleasesDGX agent

arXiv:2603.18893v2 Announce Type: replace Abstract: Tracking the internal states of large language models across conversations is important for safety, interpretability, and model welfare, yet current

Quantization Dominates Rank Reduction for KV-Cache Compression

Model ReleasesDGX agent

arXiv:2604.11501v1 Announce Type: cross Abstract: We compare two strategies for compressing the KV cache in transformer inference: rank reduction (discard dimensions) and quantization (keep all dimens

Query Lower Bounds for Diffusion Sampling

ResearchDGX agent

arXiv:2604.10857v1 Announce Type: cross Abstract: Diffusion models generate samples by iteratively querying learned score estimates. A rapidly growing literature focuses on accelerating sampling by mi

RAG-KT: Cross-platform Explainable Knowledge Tracing with Multi-view Fusion Retrieval Generation

SafetyDGX agent

arXiv:2604.10960v1 Announce Type: new Abstract: Knowledge Tracing (KT) infers a student's knowledge state from past interactions to predict future performance. Conventional Deep Learning (DL)-based KT

RationalRewards: Reasoning Rewards Scale Visual Generation Both Training and Test Time

Model ReleasesDGX agent

arXiv:2604.11626v1 Announce Type: new Abstract: Most reward models for visual generation reduce rich human judgments to a single unexplained score, discarding the reasoning that underlies preference.

Real-Time Voicemail Detection in Telephony Audio Using Temporal Speech Activity Features

HardwareDGX agent

arXiv:2604.09675v1 Announce Type: cross Abstract: Outbound AI calling systems must distinguish voicemail greetings from live human answers in real time to avoid wasted agent interactions and dropped c

Reasoning as Data: Representation-Computation Unity and Its Implementation in a Domain-Algebraic Inference Engine

ResearchDGX agent

arXiv:2604.10908v1 Announce Type: new Abstract: Every existing knowledge system separates storage from computation. We show this separation is unnecessary and eliminate it. In a standard triple is_a(A

Reasoning as Gradient: Scaling MLE Agents Beyond Tree Search

Model ReleasesDGX agent

arXiv:2603.01692v3 Announce Type: replace-cross Abstract: LLM-based agents for machine learning engineering (MLE) predominantly rely on tree search, a form of gradient-free optimization that uses scal

Rebooting Microreboot: Architectural Support for Safe, Parallel Recovery in Microservice Systems

SafetyDGX agent

arXiv:2604.09963v1 Announce Type: cross Abstract: Microreboot enables fast recovery by restarting only the failing component, but in modern microservices naive restarts are unsafe: dense dependencies

RECIPER: A Dual-View Retrieval Pipeline for Procedure-Oriented Materials Question Answering

ResearchDGX agent

arXiv:2604.11229v1 Announce Type: cross Abstract: Retrieving procedure-oriented evidence from materials science papers is difficult because key synthesis details are often scattered across long, conte

Record-Remix-Replay: Hierarchical GPU Kernel Optimization using Evolutionary Search

HardwareDGX agent

arXiv:2604.11109v1 Announce Type: cross Abstract: As high-performance computing and AI workloads become increasingly dependent on GPUs, maintaining high performance across rapidly evolving hardware ge

ReFEree: Reference-Free and Fine-Grained Method for Evaluating Factual Consistency in Real-World Code Summarization

Model ReleasesDGX agent

arXiv:2604.10520v1 Announce Type: cross Abstract: As Large Language Models (LLMs) have become capable of generating long and descriptive code summaries, accurate and reliable evaluation of factual con

Regional Explanations: Bridging Local and Global Variable Importance

Local AiDGX agent

arXiv:2604.11223v1 Announce Type: cross Abstract: We analyze two widely used local attribution methods, Local Shapley Values and LIME, which aim to quantify the contribution of a feature value x_i to

Reinforced Generation of Combinatorial Structures: Ramsey Numbers

AgentsDGX agent

arXiv:2603.09172v4 Announce Type: replace-cross Abstract: We present improved lower bounds for seven classical Ramsey numbers: mathbf{R}(3, 13) is increased from 60 to 61, mathbf{R}(3, 18) from 99 to

Relational Preference Encoding in Looped Transformer Internal States

Model ReleasesDGX agent

arXiv:2604.09870v1 Announce Type: cross Abstract: We investigate how looped transformers encode human preference in their internal iteration states. Using Ouro-2.6B-Thinking, a 2.6B-parameter looped t

Reliable Evaluation Protocol for Low-Precision Retrieval

SafetyDGX agent

arXiv:2508.03306v4 Announce Type: replace-cross Abstract: Lowering the numerical precision of model parameters and computations is widely adopted to improve the efficiency of retrieval systems. Howeve

Resilient Write: A Six-Layer Durable Write Surface for LLM Coding Agents

SafetyDGX agent

arXiv:2604.10842v1 Announce Type: cross Abstract: LLM-powered coding agents increasingly rely on tool-use protocols such as the Model Context Protocol~(MCP) to read and write files on a developer's wo

Resisting Humanization: Ethical Front-End Design Choices in AI for Sensitive Contexts

ApplicationsDGX agent

arXiv:2603.24853v2 Announce Type: replace Abstract: Ethical debates in AI have primarily focused on back-end issues such as data governance, model training, and algorithmic decision-making. Less atten

Resource Consumption Threats in Large Language Models

ResearchDGX agent

arXiv:2603.16068v3 Announce Type: replace-cross Abstract: Given limited and costly computational infrastructure, resource efficiency is a key requirement for large language models (LLMs). Efficient LL

ReSpinQuant: Efficient Layer-Wise LLM Quantization via Subspace Residual Rotation Approximation

Local AiDGX agent

arXiv:2604.11080v1 Announce Type: cross Abstract: Rotation-based Post-Training Quantization (PTQ) has emerged as a promising solution for mitigating activation outliers in the quantization of Large La

Rethinking the Diffusion Model from a Langevin Perspective

ResearchDGX agent

arXiv:2604.10465v1 Announce Type: cross Abstract: Diffusion models are often introduced from multiple perspectives, such as VAEs, score matching, or flow matching, accompanied by dense and technically

Rethinking Token-Level Credit Assignment in RLVR: A Polarity-Entropy Analysis

SafetyDGX agent

arXiv:2604.11056v1 Announce Type: cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has substantially improved the reasoning ability of Large Language Models (LLMs). However, its s

Rethinking Video Human-Object Interaction: Set Prediction over Time for Unified Detection and Anticipation

Model ReleasesDGX agent

arXiv:2604.10397v1 Announce Type: cross Abstract: Video-based human-object interaction (HOI) understanding requires both detecting ongoing interactions and anticipating their future evolution. However

Retinal Cyst Detection from Optical Coherence Tomography Images

ResearchDGX agent

arXiv:2604.10843v1 Announce Type: cross Abstract: Retinal Cysts are formed by leakage and accumulation of fluid in the retina due to the incompetence of retinal vasculature. These cystic spaces have s

Retrieval as Generation: A Unified Framework with Self-Triggered Information Planning

TutorialsDGX agent

arXiv:2604.11407v1 Announce Type: cross Abstract: We revisit retrieval-augmented generation (RAG) by embedding retrieval control directly into generation. Instead of treating retrieval as an external

Retrieval-Augmented Large Language Models for Evidence-Informed Guidance on Cannabidiol Use in Older Adults

Model ReleasesDGX agent

arXiv:2604.09548v1 Announce Type: cross Abstract: Older adults commonly experience chronic conditions such as pain and sleep disturbances and may consider cannabidiol for symptom management. Safe use

Retrieval Is Not Enough: Why Organizational AI Needs Epistemic Infrastructure

ResearchDGX agent

arXiv:2604.11759v1 Announce Type: new Abstract: Organizational knowledge used by AI agents typically lacks epistemic structure: retrieval systems surface semantically relevant content without distingu

ReXSonoVQA: A Video QA Benchmark for Procedure-Centric Ultrasound Understanding

Model ReleasesDGX agent

arXiv:2604.10916v1 Announce Type: cross Abstract: Ultrasound acquisition requires skilled probe manipulation and real-time adjustments. Vision-language models (VLMs) could enable autonomous ultrasound

Rhizome OS-1: Rhizome's Semi-Autonomous Operating System for Small Molecule Drug Discovery

Model ReleasesDGX agent

arXiv:2604.07512v2 Announce Type: replace Abstract: We present Rhizome OS-1, a semi-autonomous operating system for small molecule drug discovery in which multi-modal AI agents operate as a full multi

RISK: A Framework for GUI Agents in E-commerce Risk Management

Model ReleasesDGX agent

arXiv:2509.21982v2 Announce Type: replace Abstract: E-commerce risk management requires aggregating diverse, deeply embedded web data through multi-step, stateful interactions, which traditional scrap

Risk Awareness Injection: Calibrating Vision-Language Models for Safety without Compromising Utility

SafetyDGX agent

arXiv:2602.03402v3 Announce Type: replace Abstract: Vision language models (VLMs) extend the reasoning capabilities of large language models (LLMs) to cross-modal settings, yet remain highly vulnerabl

RL-Driven Sustainable Land-Use Allocation for the Lake Malawi Basin

Model ReleasesDGX agent

arXiv:2604.03768v2 Announce Type: replace Abstract: Unsustainable land-use practices in ecologically sensitive regions threaten biodiversity, water resources, and the livelihoods of millions. This pap

RoboLab: A High-Fidelity Simulation Benchmark for Analysis of Task Generalist Policies

Model ReleasesDGX agent

arXiv:2604.09860v1 Announce Type: cross Abstract: The pursuit of general-purpose robotics has yielded impressive foundation models, yet simulation-based benchmarking remains a bottleneck due to rapid

RPA-Check: A Multi-Stage Automated Framework for Evaluating Dynamic LLM-based Role-Playing Agents

Local AiDGX agent

arXiv:2604.11655v1 Announce Type: cross Abstract: The rapid adoption of Large Language Models (LLMs) in interactive systems has enabled the creation of dynamic, open-ended Role-Playing Agents (RPAs).

RTMC: Step-Level Credit Assignment via Rollout Trees

AgentsDGX agent

arXiv:2604.11037v1 Announce Type: cross Abstract: Multi-step agentic reinforcement learning benefits from fine-grained credit assignment, yet existing approaches offer limited options: critic-free met

S^3: Structured Sparsity Specification

ResearchDGX agent

arXiv:2604.11315v1 Announce Type: cross Abstract: We introduce the Structured Sparsity Specification (S^3), an algebraic framework for defining, composing, and implementing structured sparse patterns.

Safety Guarantees in Zero-Shot Reinforcement Learning for Cascade Dynamical Systems

SafetyDGX agent

arXiv:2604.10429v1 Announce Type: new Abstract: This paper considers the problem of zero-shot safety guarantees for cascade dynamical systems. These are systems where a subset of the states (the inner

Sanity Checks for Agentic Data Science

AgentsDGX agent

arXiv:2604.11003v1 Announce Type: new Abstract: Agentic data science (ADS) pipelines have grown rapidly in both capability and adoption, with systems such as OpenAI Codex now able to directly analyze

Sat2Sound: A Unified Framework for Zero-Shot Soundscape Mapping

ResearchDGX agent

arXiv:2505.13777v2 Announce Type: replace-cross Abstract: We present Sat2Sound, a unified multimodal framework for geospatial soundscape understanding, designed to predict and map the distribution of

Scalable Stewardship of an LLM-Assisted Clinical Benchmark with Physician Oversight

Model ReleasesDGX agent

arXiv:2512.19691v3 Announce Type: replace Abstract: Reference labels for machine-learning benchmarks are increasingly synthesized with LLM assistance, but their reliability remains underexamined. We a

SciPredict: Can LLMs Predict the Outcomes of Scientific Experiments in Natural Sciences?

Model ReleasesDGX agent

arXiv:2604.10718v1 Announce Type: new Abstract: Accelerating scientific discovery requires the identification of which experiments would yield the best outcomes before committing resources to costly p

SCITUNE: Aligning Large Language Models with Human-Curated Scientific Multimodal Instructions

Model ReleasesDGX agent

arXiv:2307.01139v2 Announce Type: replace-cross Abstract: Instruction finetuning is a popular paradigm to align large language models (LLM) with human intent. Despite its popularity, this idea is less

SCMAPR: Self-Correcting Multi-Agent Prompt Refinement for Complex-Scenario Text-to-Video Generation

Model ReleasesDGX agent

arXiv:2604.05489v3 Announce Type: replace Abstract: Text-to-Video (T2V) generation has benefited from recent advances in diffusion models, yet current systems still struggle under complex scenarios, w

SCNO: Spiking Compositional Neural Operator -- Towards a Neuromorphic Foundation Model for Nuclear PDE Solving

HardwareDGX agent

arXiv:2604.11625v1 Announce Type: cross Abstract: Neural operators have emerged as powerful surrogates for partial differential equation (PDE) solvers, yet they are typically trained as monolithic mod

Scone: Bridging Composition and Distinction in Subject-Driven Image Generation via Unified Understanding-Generation Modeling

Model ReleasesDGX agent

arXiv:2512.12675v2 Announce Type: replace-cross Abstract: Subject-driven image generation has advanced from single- to multi-subject composition, while neglecting distinction, the ability to distingui

SCOPE: Signal-Calibrated On-Policy Distillation Enhancement with Dual-Path Adaptive Weighting

SafetyDGX agent

arXiv:2604.10688v1 Announce Type: cross Abstract: On-policy reinforcement learning has become the dominant paradigm for reasoning alignment in large language models, yet its sparse, outcome-level rewa

SEARL: Joint Optimization of Policy and Tool Graph Memory for Self-Evolving Agents

SafetyDGX agent

arXiv:2604.07791v2 Announce Type: replace Abstract: Recent advances in Reinforcement Learning with Verifiable Rewards (RLVR) have demonstrated significant potential in single-turn reasoning tasks. Wit

SecureVibeBench: Evaluating Secure Coding Capabilities of Code Agents with Realistic Vulnerability Scenarios

Model ReleasesDGX agent

arXiv:2509.22097v3 Announce Type: replace-cross Abstract: Large language model-powered code agents are rapidly transforming software engineering, yet the security risks of their generated code have be

Select Smarter, Not More: Prompt-Aware Evaluation Scheduling with Submodular Guarantees

Model ReleasesDGX agent

arXiv:2604.11328v1 Announce Type: new Abstract: Automatic prompt optimization (APO) hinges on the quality of its evaluation signal, yet scoring every prompt candidate on the full training set is prohi

Self-Certifying Primal-Dual Optimization Proxies for Large-Scale Batch Economic Dispatch

ResearchDGX agent

arXiv:2510.15850v2 Announce Type: replace-cross Abstract: Recent research has shown that optimization proxies can be trained to high fidelity, achieving average optimality gaps under 1% for large-scal

← Previous
1…335336337338339…350
Next →