AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
19 May 2026

Pairwise Preference Reward and Group-Based Diversity Enhancement for Superior Open-Ended Generation

SafetyDGX agent

arXiv:2605.18191v1 Announce Type: new Abstract: Current reinforcement learning(RL) methods are broadly applicable and powerful in verifiable settings where scalar rewards can be provided. However, in

PARALLAX: Separating Genuine Hallucination Detection from Benchmark Construction Artifacts

Model ReleasesDGX agent

arXiv:2605.17028v1 Announce Type: cross Abstract: Large language models (LLMs) hallucinate with confidence: their outputs can be fluent, authoritative, and simply wrong. In medical, legal, and scienti

Parameterized 4-Qubit EWL Quantum Game Circuits with Dirac-Solow-Swan Hamiltonian Integration for Quadruple Helix Disruptive Innovation Recommender Systems

Safety

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2605.18080v1 Announce Type: cross Abstract: We present a novel parameterized 4-qubit Eisert-Wilkens-Lewenstein (EWL) quantum game circuit for recommender systems in quadruple helix innovation ec

PAREDA: A Multi-Accent Speech Dataset of Natural Language Processing Research Discussions

Model ReleasesDGX agent

arXiv:2605.17860v1 Announce Type: cross Abstract: While modern Automatic Speech Recognition (ASR) systems achieve high accuracy on benchmark corpora, their performance often degrades when there is rea

Patients Speak, AI Listens: LLM-based Analysis of Online Reviews Uncovers Key Drivers for Urgent Care Satisfaction

ApplicationsDGX agent

arXiv:2503.20981v2 Announce Type: replace-cross Abstract: Investigating the public experience of urgent care facilities is essential for promoting community healthcare development. Traditional survey

Peak-Detector: Explainable Peak Detection via Instruction-Tuned Large Language Models in Physiological Sign

SafetyDGX agent

arXiv:2605.16452v1 Announce Type: cross Abstract: Accurate peak detection across diverse cardiac physiological signals, including the Electrocardiogram (ECG), Photoplethysmogram (PPG), Ballistocardiog

Pedestrian-Aware LLM-Driven Behavioral Planning for Autonomous Vehicles

SafetyDGX agent

arXiv:2605.16858v1 Announce Type: cross Abstract: Autonomous Vehicles (AVs) must make reliable decisions in dense urban environments where pedestrian behavior is variable, sometimes abnormal, and ofte

PEIRA: Learning Predictive Encoders through Inter-View Regressor Alignment

SafetyDGX agent

arXiv:2605.17671v1 Announce Type: cross Abstract: Non-contrastive self-supervised learning (SSL) is an effective framework for predictive representation learning, but popular (and in practice effectiv

Perception-based Image Denoising via Generative Compression

ResearchDGX agent

arXiv:2602.11553v2 Announce Type: replace-cross Abstract: Image denoising aims to remove noise while preserving structural details and perceptual realism, yet distortion-driven methods often produce o

Perceptual implications of automatic anonymization in pathological speech

SafetyDGX agent

arXiv:2505.00409v3 Announce Type: replace-cross Abstract: Automatic anonymization is increasingly used to enable ethical sharing of clinical speech, yet its perceptual and clinical consequences remain

PERMA: Benchmarking Personalized Memory Agents via Event-Driven Preference and Realistic Task Environments

Model ReleasesDGX agent

arXiv:2603.23231v2 Announce Type: replace Abstract: Empowering large language models with long-term memory is crucial for building agents that adapt to users' evolving needs. Existing evaluations of t

Perovskite-R1: a domain-specialized large language model for intelligent discovery of precursor additives and experimental design

ApplicationsDGX agent

arXiv:2507.16307v2 Announce Type: replace-cross Abstract: Perovskite solar cells (PSCs) have rapidly emerged as a leading contender in next-generation photovoltaic technologies, owing to their excepti

PersonaArena: Dynamic Simulation for Evaluating and Enhancing Persona-Level Role-Playing in Large Language Models

AgentsDGX agent

arXiv:2605.17044v1 Announce Type: new Abstract: Large language models (LLMs) increasingly serve as interactive social agents, yet their ability to maintain coherent and authentic persona-level role-pl

PersonaDual: Balancing Personalization and Objectivity via Adaptive Reasoning

TutorialsDGX agent

arXiv:2601.08679v3 Announce Type: replace Abstract: As users increasingly expect LLMs to align with their preferences, personalized information becomes valuable. However, personalized information can

PESD-TSF: A Period-Aware and Explicit Structured Decomposition Framework for Long-Term Time Series Forecasting

Model ReleasesDGX agent

arXiv:2605.16449v1 Announce Type: cross Abstract: Deep forecasting models often suffer from attenuated periodic perception and entangled trend-noise representations as network depth increases. Moreove

PH-Dreamer: A Physics-Driven World Model via Port-Hamiltonian Generative Dynamics

SafetyDGX agent

arXiv:2605.18303v1 Announce Type: cross Abstract: World models built on recurrent state space architectures enable efficient latent imagination, yet remain physically unstructured, producing dynamics

Phase Transitions in Driven Informational Systems: A Two-Field Perspective on Learning Theory and Non-Equilibrium Chemistry

SafetyDGX agent

arXiv:2605.16325v1 Announce Type: cross Abstract: Phase-transition phenomena in deep learning (grokking, emergent capabilities, and ontological reorganization under context shift) have been studied th

Physics-Guided Geometric Diffusion for Macro Placement Generation

ResearchDGX agent

arXiv:2605.16451v1 Announce Type: cross Abstract: Macro placement is a pivotal stage in VLSI physical design, fundamentally determining the overall chip performance. Recent data-driven placement metho

PhysioSeq2Seq: A Hybrid Physiological Digital Twin and Sequence-to-Sequence LSTM for Long-Horizon Glucose Forecasting in Type 1 Diabetes

SafetyDGX agent

arXiv:2605.16860v1 Announce Type: cross Abstract: Accurate long-horizon glucose forecasting is critical for automated insulin delivery systems, which help people with type 1 diabetes (T1D) manage thei

PIMSM: Physics-Informed Multi-Scale Mamba for Stable Neural Representations under Distribution Shift

SafetyDGX agent

arXiv:2605.16351v1 Announce Type: cross Abstract: Scientific foundation models are expected to reuse representations under changes in dataset, acquisition protocol, and deployment domain, yet many seq

PIPER: Content-Based Table Search via profiling and LLM-Generated Pseudoqueries

ResearchDGX agent

arXiv:2605.18199v1 Announce Type: cross Abstract: The rapid growth of tabular datasets in data lakes, data spaces, and open data portals makes effective dataset search essential for reuse and analysis

Plan First, Diffuse Later: Extrinsic Graph Guidance for Long-Horizon Diffusion Planning

Local AiDGX agent

arXiv:2605.16863v1 Announce Type: cross Abstract: Compositional diffusion models offer a promising route to long-horizon planning by denoising multiple overlapping sub-trajectories while ensuring that

PluRule: A Benchmark for Moderating Pluralistic Communities on Social Media

Model ReleasesDGX agent

arXiv:2605.17187v1 Announce Type: cross Abstract: Social media are shifting towards pluralism -- community-governed platforms where groups define their own norms. What violates rules in one community

Pocket Foundation Models: Distilling TFMs into CPU-Ready Gradient-Boosted Trees

HardwareDGX agent

arXiv:2605.18654v1 Announce Type: cross Abstract: A fraud scorer needs to answer in under 2 ms. The best tabular foundation models (TFMs) take 151-1,275 ms on GPU. We close this gap by distilling the

Policy-Grounded Dynamic Facet Suggestions for Job Search

SafetyDGX agent

arXiv:2605.16479v1 Announce Type: cross Abstract: Job seekers often initiate search with short, underspecified queries. At LinkedIn, over 80% of job-related queries contain three or fewer keywords, ma

PopPy: Opportunistically Exploiting Parallelism in Python Compound AI Applications

ApplicationsDGX agent

arXiv:2605.18697v1 Announce Type: cross Abstract: Compound AI applications, which compose calls to ML models using a general-purpose programming language like Python, are widely used for a variety of

PopuLoRA: Co-Evolving LLM Populations for Reasoning Self-Play

AgentsDGX agent

arXiv:2605.16727v1 Announce Type: new Abstract: We introduce PopuLoRA, a population-based asymmetric self-play framework for reinforcement learning with verifiable rewards (RLVR) post-training of LLMs

Position: A Three-Layer Probabilistic Assume-Guarantee Architecture Is Structurally Required for Safe LLM Agent Deployment

SafetyDGX agent

arXiv:2605.18672v1 Announce Type: new Abstract: This position paper argues that enforcing LLM agent safety within a single abstraction layer is not merely suboptimal but categorically insufficient for

Position: AI Evaluations Should be Grounded on a Theory of Capability

Model ReleasesDGX agent

arXiv:2509.19590v2 Announce Type: replace Abstract: Evaluations of generative models are now ubiquitous, and their outcomes critically shape public and scientific expectations of AI's capabilities. Ye

Position: Universal Time Series Foundation Models Rest on a Category Error

AgentsDGX agent

arXiv:2602.05287v2 Announce Type: replace Abstract: This position paper argues that the pursuit of 'Universal Foundation Models for Time Series' rests on a fundamental category error, mistaking a stru

Position: Weight Space Should Be a First-Class Generative AI Modality

ResearchDGX agent

arXiv:2605.18632v1 Announce Type: cross Abstract: Neural network checkpoints have quietly become a large-scale data resource: millions of trained weight vectors now exist, each encoding task-, domain-

POST: Prior-Observation Adversarial Learning of Spatio-Temporal Associations for Multivariate Time Series Anomaly Detection

Model ReleasesDGX agent

arXiv:2605.18128v1 Announce Type: new Abstract: Existing Multivariate Time Series Anomaly Detection (MTSAD) frameworks increasingly rely on integrating Graph Neural Networks (GNNs) with sequence model

Post-Trained MoE Can Skip Half Experts via Self-Distillation

Model ReleasesDGX agent

arXiv:2605.18643v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) scales language models efficiently through sparse expert activation, and its dynamic variant further reduces computation by a

Predictable Confabulations: Factual Recall by LLMs Scales with Model Size and Topic Frequency

Model ReleasesDGX agent

arXiv:2605.18732v1 Announce Type: cross Abstract: While scaling laws govern aggregate large language model performance, no scaling law has linked factual recall to both model size and training-data co

Prediction-Intervention Games and Invariant Sets

ApplicationsDGX agent

arXiv:2605.16828v1 Announce Type: cross Abstract: We consider the following two-player game: using observational data, the leader chooses a prediction function for a response variable Y from given cov

Prediction of Challenging Behaviors Associated with Profound Autism in a Classroom Setting Using Wearable Sensors

SafetyDGX agent

arXiv:2605.17618v1 Announce Type: new Abstract: Autism Spectrum Disorder (ASD) is characterized by challenges with social interaction and communication and by restricted or repetitive patterns of thou

Predictive Prefetching for Retrieval-Augmented Generation

ResearchDGX agent

arXiv:2605.17989v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) improves factual grounding in large language models but suffers from substantial latency due to synchronous retri

Prefix-Adaptive Block Diffusion for Efficient Document Recognition

ResearchDGX agent

arXiv:2605.16861v1 Announce Type: cross Abstract: Block Diffusion Models (BDMs) support parallel generation, flexible-length output, and KV caching, making them promising for efficient document parsin

PriHA: A RAG-Enhanced LLM Framework for Primary Healthcare Assistant in Hong Kong

Model ReleasesDGX agent

arXiv:2604.14215v2 Announce Type: replace-cross Abstract: To address the unsustainable rise in public health expenditures, the Hong Kong SAR Government is shifting its strategic focus to primary healt

Principles of frugal inference and control

Local AiDGX agent

arXiv:2406.14427v4 Announce Type: replace Abstract: A central challenge for intelligent agents in an uncertain world is striking the right balance between utility maximization and resource use, not on

Prior Knowledge Makes It Possible: From Sublinear Graph Algorithms to LLM Test-Time Methods

AgentsDGX agent

arXiv:2510.16609v3 Announce Type: replace-cross Abstract: Test-time augmentation, such as Retrieval-Augmented Generation (RAG) or tool use, critically depends on an interplay between a model's paramet

PRISMat: Policy-Driven, Permutation-Invariant Autoregressive Material Generation

Model ReleasesDGX agent

arXiv:2605.16612v1 Announce Type: new Abstract: Rapid identification of candidate materials with target properties has become a key task in materials science. Machine learning has emerged as an altern

Privacy Policy Enforcement Guardrails for Data-Sensitive Retrieval-Augmented Generation

Model ReleasesDGX agent

arXiv:2605.17034v1 Announce Type: cross Abstract: Standard PII filters often miss contextual data leakage in RAG systems, such as non-regulated attribute clusters that collectively identify individual

Privacy Preserving Reinforcement Learning with One-Sided Feedback

Model ReleasesDGX agent

arXiv:2605.18246v1 Announce Type: cross Abstract: We study reinforcement learning (RL) in multi-dimensional continuous state and action spaces with one-sided feedback, where the agent receives partial

Probing for Representation Manifolds in Superposition

Model ReleasesDGX agent

arXiv:2605.18537v1 Announce Type: cross Abstract: This paper introduces the Manifold Probe, a supervised method for discovering representation manifolds in superposition. The method generalizes linear

Probing SMEFT Operators through tar{t}tar{t} Production with Hyper-Graph Neural Networks at the LHC

Model ReleasesDGX agent

arXiv:2605.18382v1 Announce Type: cross Abstract: We present a phenomenological study of tar{t}tar{t} production in proton-proton collisions at sqrt{s} = 13~TeV, using a Hyper-Graph Neural Network (H-

ProfBench: Multi-Domain Rubrics requiring Professional Knowledge to Answer and Judge

Model ReleasesDGX agent

arXiv:2510.18941v2 Announce Type: replace-cross Abstract: Evaluating progress in large language models (LLMs) is often constrained by the challenge of verifying responses, limiting assessments to task

Progressive Generalization Augmentation with Deeply Coupled RND-PPO and Domain-Prioritized Noise Injection for Robust Crop Management Reinforcement Learning

HardwareDGX agent

arXiv:2605.17428v1 Announce Type: cross Abstract: Our preliminary experiments on gym-DSSAT maize irrigation tasks revealed that +/-2 degrees C temperature noise causes an 11.9% reduction in economic r

Prompt Compression in Diffusion Large Language Models: Evaluating LLMLingua-2 on LLaDA

Model ReleasesDGX agent

arXiv:2605.17932v1 Announce Type: cross Abstract: Prompt compression reduces inference cost and context length in large language models, but prior evaluations focus primarily on autoregressive archite

Prompt2Fingerprint: Plug-and-Play LLM Fingerprinting via Text-to-Weight Generation

Model ReleasesDGX agent

arXiv:2605.18474v1 Announce Type: cross Abstract: The widespread deployment and redistribution of large language models (LLMs) have made model provenance tracking a critical challenge. While existing

PromptDecipher: Supporting AI Tutor Authoring Through Editable Simulated Interactions

TutorialsDGX agent

arXiv:2605.16605v1 Announce Type: cross Abstract: Chatbots have long been explored as tools to support learning, and recent advances in large language models have significantly expanded the availabili

Prompts Don't Protect: Architectural Enforcement via MCP Proxy for LLM Tool Access Control

Model ReleasesDGX agent

arXiv:2605.18414v1 Announce Type: cross Abstract: Large language models increasingly operate as autonomous agents that select and invoke tools from large registries. We identify a critical gap: when u

PropGuard: Safeguarding LLM-MAS via Propagation-Aware Exploration and Remediation

Local AiDGX agent

arXiv:2605.16346v1 Announce Type: cross Abstract: LLM-based multi-agent systems (LLM-MAS) have become a promising paradigm for solving complex tasks through role specialization, tool use, memory, and

PROTEA: Offline Evaluation and Iterative Refinement for Multi-Agent LLM Workflows

Local AiDGX agent

arXiv:2605.18032v1 Announce Type: cross Abstract: Multi-agent LLM workflows -- systems composed of multiple role-specific LLM calls -- often outperform single-prompt baselines, but they remain difficu

ProtoSiTex: Learning Semi-Interpretable Prototypes for Multi-label Text Classification

Model ReleasesDGX agent

arXiv:2510.12534v4 Announce Type: replace Abstract: The rapid growth of user-generated text across digital platforms has intensified the need for interpretable models capable of fine-grained text clas

ProxyKV: Cross-Model Proxy Pruning for Efficient Long-Context LLM Inference

Model ReleasesDGX agent

arXiv:2605.16360v1 Announce Type: cross Abstract: Efficient long-context inference in Large Language Models (LLMs) is severely constrained by the Key-Value (KV) cache memory wall, yet existing pruning

PULSE: Agentic Investigation with Passive Sensing for Proactive Intervention in Cancer Survivorship

AgentsDGX agent

arXiv:2605.17679v1 Announce Type: cross Abstract: Cancer survivors face elevated rates of depression, anxiety, and general emotional distress, yet the precise moments they most need support are often

PyHealth 2.0: A Comprehensive Open-Source Toolkit for Accessible and Reproducible Clinical Deep Learning

ApplicationsDGX agent

arXiv:2601.16414v2 Announce Type: replace-cross Abstract: Difficulty replicating baselines, high computational costs, and required domain expertise create persistent barriers to clinical AI research.

QQJ: Quantifying Qualitative Judgment for Scalable and Human-Aligned Evaluation of Generative AI

SafetyDGX agent

arXiv:2605.17382v1 Announce Type: new Abstract: The rapid progress of generative artificial intelligence has exposed fundamental limitations in existing evaluation methodologies, particularly for open

QSTRBench: a New Benchmark to Evaluate the Ability of Language Models to Reason with Qualitative Spatial and Temporal Calculi

Model ReleasesDGX agent

arXiv:2605.18380v1 Announce Type: new Abstract: We introduce an extensive qualitative spatial and temporal reasoning (QSTR) benchmark for evaluating large language models (LLMs). We pose questions con

← Previous
1…240241242243244…358
Next →