AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
Research

Retinal Cyst Detection from Optical Coherence Tomography Images

DGX agent

arXiv:2604.10843v1 Announce Type: cross Abstract: Retinal Cysts are formed by leakage and accumulation of fluid in the retina due to the incompetence of retinal vasculature. These cystic spaces have s

researcharxiv-cs-ai
14 Apr 2026
Tutorials
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Retrieval as Generation: A Unified Framework with Self-Triggered Information Planning

DGX agent

arXiv:2604.11407v1 Announce Type: cross Abstract: We revisit retrieval-augmented generation (RAG) by embedding retrieval control directly into generation. Instead of treating retrieval as an external

tutorialsarxiv-cs-ai
14 Apr 2026
Model Releases

Retrieval-Augmented Large Language Models for Evidence-Informed Guidance on Cannabidiol Use in Older Adults

DGX agent

arXiv:2604.09548v1 Announce Type: cross Abstract: Older adults commonly experience chronic conditions such as pain and sleep disturbances and may consider cannabidiol for symptom management. Safe use

model-releasesarxiv-cs-ai
14 Apr 2026
Research

Retrieval Is Not Enough: Why Organizational AI Needs Epistemic Infrastructure

DGX agent

arXiv:2604.11759v1 Announce Type: new Abstract: Organizational knowledge used by AI agents typically lacks epistemic structure: retrieval systems surface semantically relevant content without distingu

researcharxiv-cs-ai
14 Apr 2026
Model Releases

ReXSonoVQA: A Video QA Benchmark for Procedure-Centric Ultrasound Understanding

DGX agent

arXiv:2604.10916v1 Announce Type: cross Abstract: Ultrasound acquisition requires skilled probe manipulation and real-time adjustments. Vision-language models (VLMs) could enable autonomous ultrasound

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Rhizome OS-1: Rhizome's Semi-Autonomous Operating System for Small Molecule Drug Discovery

DGX agent

arXiv:2604.07512v2 Announce Type: replace Abstract: We present Rhizome OS-1, a semi-autonomous operating system for small molecule drug discovery in which multi-modal AI agents operate as a full multi

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

RISK: A Framework for GUI Agents in E-commerce Risk Management

DGX agent

arXiv:2509.21982v2 Announce Type: replace Abstract: E-commerce risk management requires aggregating diverse, deeply embedded web data through multi-step, stateful interactions, which traditional scrap

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

Risk Awareness Injection: Calibrating Vision-Language Models for Safety without Compromising Utility

DGX agent

arXiv:2602.03402v3 Announce Type: replace Abstract: Vision language models (VLMs) extend the reasoning capabilities of large language models (LLMs) to cross-modal settings, yet remain highly vulnerabl

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

RL-Driven Sustainable Land-Use Allocation for the Lake Malawi Basin

DGX agent

arXiv:2604.03768v2 Announce Type: replace Abstract: Unsustainable land-use practices in ecologically sensitive regions threaten biodiversity, water resources, and the livelihoods of millions. This pap

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

RoboLab: A High-Fidelity Simulation Benchmark for Analysis of Task Generalist Policies

DGX agent

arXiv:2604.09860v1 Announce Type: cross Abstract: The pursuit of general-purpose robotics has yielded impressive foundation models, yet simulation-based benchmarking remains a bottleneck due to rapid

model-releasesarxiv-cs-ai
14 Apr 2026
Local Ai

RPA-Check: A Multi-Stage Automated Framework for Evaluating Dynamic LLM-based Role-Playing Agents

DGX agent

arXiv:2604.11655v1 Announce Type: cross Abstract: The rapid adoption of Large Language Models (LLMs) in interactive systems has enabled the creation of dynamic, open-ended Role-Playing Agents (RPAs).

local-aiarxiv-cs-ai
14 Apr 2026
Agents

RTMC: Step-Level Credit Assignment via Rollout Trees

DGX agent

arXiv:2604.11037v1 Announce Type: cross Abstract: Multi-step agentic reinforcement learning benefits from fine-grained credit assignment, yet existing approaches offer limited options: critic-free met

agentsarxiv-cs-ai
14 Apr 2026
Research

S^3: Structured Sparsity Specification

DGX agent

arXiv:2604.11315v1 Announce Type: cross Abstract: We introduce the Structured Sparsity Specification (S^3), an algebraic framework for defining, composing, and implementing structured sparse patterns.

researcharxiv-cs-ai
14 Apr 2026
Safety

Safety Guarantees in Zero-Shot Reinforcement Learning for Cascade Dynamical Systems

DGX agent

arXiv:2604.10429v1 Announce Type: new Abstract: This paper considers the problem of zero-shot safety guarantees for cascade dynamical systems. These are systems where a subset of the states (the inner

safetyarxiv-cs-ai
14 Apr 2026
Agents

Sanity Checks for Agentic Data Science

DGX agent

arXiv:2604.11003v1 Announce Type: new Abstract: Agentic data science (ADS) pipelines have grown rapidly in both capability and adoption, with systems such as OpenAI Codex now able to directly analyze

agentsarxiv-cs-ai
14 Apr 2026
Research

Sat2Sound: A Unified Framework for Zero-Shot Soundscape Mapping

DGX agent

arXiv:2505.13777v2 Announce Type: replace-cross Abstract: We present Sat2Sound, a unified multimodal framework for geospatial soundscape understanding, designed to predict and map the distribution of

researcharxiv-cs-ai
14 Apr 2026
Model Releases

Scalable Stewardship of an LLM-Assisted Clinical Benchmark with Physician Oversight

DGX agent

arXiv:2512.19691v3 Announce Type: replace Abstract: Reference labels for machine-learning benchmarks are increasingly synthesized with LLM assistance, but their reliability remains underexamined. We a

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

SciPredict: Can LLMs Predict the Outcomes of Scientific Experiments in Natural Sciences?

DGX agent

arXiv:2604.10718v1 Announce Type: new Abstract: Accelerating scientific discovery requires the identification of which experiments would yield the best outcomes before committing resources to costly p

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

SCITUNE: Aligning Large Language Models with Human-Curated Scientific Multimodal Instructions

DGX agent

arXiv:2307.01139v2 Announce Type: replace-cross Abstract: Instruction finetuning is a popular paradigm to align large language models (LLM) with human intent. Despite its popularity, this idea is less

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

SCMAPR: Self-Correcting Multi-Agent Prompt Refinement for Complex-Scenario Text-to-Video Generation

DGX agent

arXiv:2604.05489v3 Announce Type: replace Abstract: Text-to-Video (T2V) generation has benefited from recent advances in diffusion models, yet current systems still struggle under complex scenarios, w

model-releasesarxiv-cs-ai
14 Apr 2026
Hardware

SCNO: Spiking Compositional Neural Operator -- Towards a Neuromorphic Foundation Model for Nuclear PDE Solving

DGX agent

arXiv:2604.11625v1 Announce Type: cross Abstract: Neural operators have emerged as powerful surrogates for partial differential equation (PDE) solvers, yet they are typically trained as monolithic mod

hardwarearxiv-cs-ai
14 Apr 2026
Model Releases

Scone: Bridging Composition and Distinction in Subject-Driven Image Generation via Unified Understanding-Generation Modeling

DGX agent

arXiv:2512.12675v2 Announce Type: replace-cross Abstract: Subject-driven image generation has advanced from single- to multi-subject composition, while neglecting distinction, the ability to distingui

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

SCOPE: Signal-Calibrated On-Policy Distillation Enhancement with Dual-Path Adaptive Weighting

DGX agent

arXiv:2604.10688v1 Announce Type: cross Abstract: On-policy reinforcement learning has become the dominant paradigm for reasoning alignment in large language models, yet its sparse, outcome-level rewa

safetyarxiv-cs-ai
14 Apr 2026
Safety

SEARL: Joint Optimization of Policy and Tool Graph Memory for Self-Evolving Agents

DGX agent

arXiv:2604.07791v2 Announce Type: replace Abstract: Recent advances in Reinforcement Learning with Verifiable Rewards (RLVR) have demonstrated significant potential in single-turn reasoning tasks. Wit

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

SecureVibeBench: Evaluating Secure Coding Capabilities of Code Agents with Realistic Vulnerability Scenarios

DGX agent

arXiv:2509.22097v3 Announce Type: replace-cross Abstract: Large language model-powered code agents are rapidly transforming software engineering, yet the security risks of their generated code have be

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Select Smarter, Not More: Prompt-Aware Evaluation Scheduling with Submodular Guarantees

DGX agent

arXiv:2604.11328v1 Announce Type: new Abstract: Automatic prompt optimization (APO) hinges on the quality of its evaluation signal, yet scoring every prompt candidate on the full training set is prohi

model-releasesarxiv-cs-ai
14 Apr 2026
Research

Self-Certifying Primal-Dual Optimization Proxies for Large-Scale Batch Economic Dispatch

DGX agent

arXiv:2510.15850v2 Announce Type: replace-cross Abstract: Recent research has shown that optimization proxies can be trained to high fidelity, achieving average optimality gaps under 1% for large-scal

researcharxiv-cs-ai
14 Apr 2026
Safety

Self-Organizing Dual-Buffer Adaptive Clustering Experience Replay (SODACER) for Safe Reinforcement Learning in Optimal Control

DGX agent

arXiv:2601.06540v2 Announce Type: replace-cross Abstract: This paper proposes a novel reinforcement learning framework, named Self-Organizing Dual-buffer Adaptive Clustering Experience Replay (SODACER

safetyarxiv-cs-ai
14 Apr 2026
Applications

SemaCDR: LLM-Powered Transferable Semantics for Cross-Domain Sequential Recommendation

DGX agent

arXiv:2604.09551v1 Announce Type: cross Abstract: Cross-domain recommendation (CDR) addresses the data sparsity and cold-start problems in the target domain by leveraging knowledge from data-rich sour

applicationsarxiv-cs-ai
14 Apr 2026
Safety

SemaClaw: A Step Towards General-Purpose Personal AI Agents through Harness Engineering

DGX agent

arXiv:2604.11548v1 Announce Type: new Abstract: The rise of OpenClaw in early 2026 marks the moment when millions of users began deploying personal AI agents into their daily lives, delegating tasks r

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

Semantic-Geometric Dual Compression: Training-Free Visual Token Reduction for Ultra-High-Resolution Remote Sensing Understanding

DGX agent

arXiv:2604.11122v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have demonstrated immense potential in Earth observation. However, the massive visual tokens generated when p

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Semantic Manipulation Localization

DGX agent

arXiv:2604.10132v1 Announce Type: cross Abstract: Image Manipulation Localization (IML) aims to identify edited regions in an image. However, with the increasing use of modern image editing and genera

model-releasesarxiv-cs-ai
14 Apr 2026
Agents

Semantic Segmentation Algorithm Based on Light Field and LiDAR Fusion

DGX agent

arXiv:2510.06687v2 Announce Type: replace-cross Abstract: Semantic segmentation serves as a cornerstone of scene understanding in autonomous driving but continues to face significant challenges under

agentsarxiv-cs-ai
14 Apr 2026
Research

Seven simple steps for log analysis in AI systems

DGX agent

arXiv:2604.09563v1 Announce Type: new Abstract: AI systems produce large volumes of logs as they interact with tools and users. Analysing these logs can help understand model capabilities, propensitie

researcharxiv-cs-ai
14 Apr 2026
Research

ShapShift: Explaining Model Prediction Shifts with Subgroup Conditional Shapley Values

DGX agent

arXiv:2604.11200v1 Announce Type: cross Abstract: Changes in input distribution can induce shifts in the average predictions of machine learning models. Such prediction shifts may impact downstream bu

researcharxiv-cs-ai
14 Apr 2026
Model Releases

Shared Emotion Geometry Across Small Language Models: A Cross-Architecture Study of Representation, Behavior, and Methodological Confounds

DGX agent

arXiv:2604.11050v1 Announce Type: cross Abstract: We extract 21-emotion vector sets from twelve small language models (six architectures x base/instruct, 1B-8B parameters) under a unified comprehensio

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

SHE: Stepwise Hybrid Examination Reinforcement Learning Framework for E-commerce Search Relevance

DGX agent

arXiv:2510.07972v3 Announce Type: replace Abstract: Query-product relevance prediction is vital for AI-driven e-commerce, yet current LLM-based approaches face a dilemma: SFT and DPO struggle with lon

safetyarxiv-cs-ai
14 Apr 2026
Research

Should We be Pedantic About Reasoning Errors in Machine Translation?

DGX agent

arXiv:2604.09890v1 Announce Type: cross Abstract: Across multiple language pairings (English o {Spanish, French, German, Mandarin, Japanese, Urdu, Cantonese}), we find reasoning errors in translation.

researcharxiv-cs-ai
14 Apr 2026
Model Releases

SimBench: Benchmarking the Ability of Large Language Models to Simulate Human Behaviors

DGX agent

arXiv:2510.17516v4 Announce Type: replace-cross Abstract: Large language model (LLM) simulations of human behavior have the potential to revolutionize the social and behavioral sciences, if and only i

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

Skill-SD: Skill-Conditioned Self-Distillation for Multi-turn LLM Agents

DGX agent

arXiv:2604.10674v1 Announce Type: cross Abstract: Reinforcement learning (RL) has been widely used to train LLM agents for multi-turn interactive tasks, but its sample efficiency is severely limited b

safetyarxiv-cs-ai
14 Apr 2026
Safety

SLALOM: Simulation Lifecycle Analysis via Longitudinal Observation Metrics for Social Simulation

DGX agent

arXiv:2604.11466v1 Announce Type: cross Abstract: Large Language Model (LLM) agents offer a potentially-transformative path forward for generative social science but face a critical crisis of validity

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

SMART: When is it Actually Worth Expanding a Speculative Tree?

DGX agent

arXiv:2604.09731v1 Announce Type: cross Abstract: Tree-based speculative decoding accelerates autoregressive generation by verifying a branching tree of draft tokens in a single target-model forward p

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Solving Physics Olympiad via Reinforcement Learning on Physics Simulators

DGX agent

arXiv:2604.11805v1 Announce Type: cross Abstract: We have witnessed remarkable advances in LLM reasoning capabilities with the advent of DeepSeek-R1. However, much of this progress has been fueled by

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Spatial Competence Benchmark

DGX agent

arXiv:2604.09594v1 Announce Type: new Abstract: Spatial competence is the quality of maintaining a consistent internal representation of an environment and using it to infer discrete structure and pla

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

SpatialScore: Towards Comprehensive Evaluation for Spatial Intelligence

DGX agent

arXiv:2505.17012v3 Announce Type: replace-cross Abstract: Existing evaluations of multimodal large language models (MLLMs) on spatial intelligence are typically fragmented and limited in scope. In thi

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

Speaking to No One: Ontological Dissonance and the Double Bind of Conversational AI

DGX agent

arXiv:2604.10833v1 Announce Type: cross Abstract: Recent reports indicate that sustained interaction with conversational artificial intelligence (AI) systems can, in a small subset of users, contribut

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

SpecMoE: A Fast and Efficient Mixture-of-Experts Inference via Self-Assisted Speculative Decoding

DGX agent

arXiv:2604.10152v1 Announce Type: new Abstract: The Mixture-of-Experts (MoE) architecture has emerged as a promising approach to mitigate the rising computational costs of large language models (LLMs)

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

SPEED-Bench: A Unified and Diverse Benchmark for Speculative Decoding

DGX agent

arXiv:2604.09557v1 Announce Type: cross Abstract: Speculative Decoding (SD) has emerged as a critical technique for accelerating Large Language Model (LLM) inference. Unlike deterministic system optim

model-releasesarxiv-cs-ai
14 Apr 2026
← Previous
1…425426427428429…443
Next →