AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlog
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Research

Partition-of-Unity Gaussian Kolmogorov-Arnold Networks

DGX agent

arXiv:2604.23599v1 Announce Type: cross Abstract: Gaussian basis functions provide an efficient and flexible alternative to spline activations in KANs. In this work, we introduce the partition-of-unit

researcharxiv-cs-ai
28 Apr 2026
Model Releases
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Patching LLM Like Software: A Lightweight Method for Improving Safety Policy in Large Language Models

DGX agent

arXiv:2511.08484v2 Announce Type: replace Abstract: We propose patching for large language models (LLMs) like software versions, a lightweight and modular approach for addressing safety vulnerabilitie

model-releasesarxiv-cs-ai
28 Apr 2026
Research

PathMoG: A Pathway-Centric Modular Graph Neural Network for Multi-Omics Survival Prediction

DGX agent

arXiv:2604.24371v1 Announce Type: cross Abstract: Cancer survival prediction from multi-omics data remains challenging because prognostic signals are high-dimensional, heterogeneous, and distributed a

researcharxiv-cs-ai
28 Apr 2026
Model Releases

Patterns vs. Patients: Evaluating LLMs against Mental Health Professionals on Personality Disorder Diagnosis through First-Person Narratives

DGX agent

arXiv:2512.20298v2 Announce Type: replace-cross Abstract: Growing reliance on LLMs for psychiatric self-assessment raises questions about their ability to interpret qualitative patient narratives. Thi

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

PDF-WuKong: A Large Multimodal Model for Efficient Long PDF Reading with End-to-End Sparse Sampling

DGX agent

arXiv:2410.05970v3 Announce Type: replace-cross Abstract: Multimodal document understanding is a challenging task to process and comprehend large amounts of textual and visual information. Recent adva

model-releasesarxiv-cs-ai
28 Apr 2026
Safety

Peer Identity Bias in Multi-Agent LLM Evaluation: An Empirical Study Using the TRUST Democratic Discourse Analysis Pipeline

DGX agent

arXiv:2604.22971v1 Announce Type: cross Abstract: The TRUST democratic discourse analysis pipeline exposes its large language model (LLM) components to peer model identity through multiple structural

safetyarxiv-cs-ai
28 Apr 2026
Research

Personalized Worked Example Generation from Student Code Submissions using Pattern-based Knowledge Components

DGX agent

arXiv:2604.24758v1 Announce Type: cross Abstract: Adaptive programming practice often relies on fixed libraries of worked examples and practice problems, which require substantial authoring effort and

researcharxiv-cs-ai
28 Apr 2026
Model Releases

PExA: Parallel Exploration Agent for Complex Text-to-SQL

DGX agent

arXiv:2604.22934v1 Announce Type: new Abstract: LLM-based agents for text-to-SQL often struggle with latency-performance trade-off, where performance improvements come at the cost of latency or vice v

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

PhysCodeBench: Benchmarking Physics-Aware Symbolic Simulation of 3D Scenes via Self-Corrective Multi-Agent Refinement

DGX agent

arXiv:2604.23580v1 Announce Type: cross Abstract: Physics-aware symbolic simulation of 3D scenes is critical for robotics, embodied AI, and scientific computing, requiring models to understand natural

model-releasesarxiv-cs-ai
28 Apr 2026
Agents

PhySE: A Psychological Framework for Real-Time AR-LLM Social Engineering Attacks

DGX agent

arXiv:2604.23148v1 Announce Type: new Abstract: The emerging threat of AR-LLM-based Social Engineering (AR-LLM-SE) attacks (e.g. SEAR) poses a significant risk to real-world social interactions. In su

agentsarxiv-cs-ai
28 Apr 2026
Agents

PhysNote: Self-Knowledge Notes for Evolvable Physical Reasoning in Vision-Language Model

DGX agent

arXiv:2604.24443v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have demonstrated strong performance on textbook-style physics problems, yet they frequently fail when confronted with dyn

agentsarxiv-cs-ai
28 Apr 2026
Model Releases

PivotMerge: Bridging Heterogeneous Multimodal Pre-training via Post-Alignment Model Merging

DGX agent

arXiv:2604.22823v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) rely on multimodal pre-training over diverse data sources, where different datasets often induce complementar

model-releasesarxiv-cs-ai
28 Apr 2026
Research

Planning Under Observation Mismatch for Traffic Signal Control via Adaptive Modular World Models

DGX agent

arXiv:2501.02548v2 Announce Type: replace-cross Abstract: Deploying learned decision-making systems often requires transferring to new sites where the sensing pipeline differs. In such cases, observat

researcharxiv-cs-ai
28 Apr 2026
Safety

Polychromic Objectives for Reinforcement Learning

DGX agent

arXiv:2509.25424v5 Announce Type: replace-cross Abstract: Reinforcement learning fine-tuning (RLFT) is a dominant paradigm for improving pretrained policies for downstream tasks. These pretrained poli

safetyarxiv-cs-ai
28 Apr 2026
Research

POPI: Personalizing LLMs via Optimized Natural Language Preference Inference

DGX agent

arXiv:2510.17881v3 Announce Type: replace-cross Abstract: Large language models (LLMs) are typically aligned with population-level preferences, despite substantial variation across individual users. W

researcharxiv-cs-ai
28 Apr 2026
Agents

Poster: ClawdGo: Endogenous Security Awareness Training for Autonomous AI Agents

DGX agent

arXiv:2604.24020v1 Announce Type: cross Abstract: Autonomous AI agents deployed on platforms such as OpenClaw face prompt injection, memory poisoning, supply-chain attacks, and social engineering, yet

agentsarxiv-cs-ai
28 Apr 2026
Model Releases

PRAXIS: Integrating Program Analysis with Observability for Root-Cause Analysis

DGX agent

arXiv:2512.22113v2 Announce Type: replace-cross Abstract: Unresolved production cloud incidents cost an average of over $2M per hour. This paper introduces PRAXIS, an orchestrator that manages and dep

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Predicting one-year clinical instability and mortality in heart failure patients using sequence modeling

DGX agent

arXiv:2511.16839v3 Announce Type: replace-cross Abstract: Heart failure (HF) discharge planning depends on identifying patients at risk of deterioration or death, yet accurate prediction from routinel

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Pref-CTRL: Preference Driven LLM Alignment using Representation Editing

DGX agent

arXiv:2604.23543v1 Announce Type: cross Abstract: Test-time alignment methods offer a promising alternative to fine-tuning by steering the outputs of large language models (LLMs) at inference time wit

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Probe-Based Data Attribution: Discovering and Mitigating Undesirable Behaviors in LLM Post-Training

DGX agent

arXiv:2602.11079v3 Announce Type: replace-cross Abstract: We propose probe-based data attribution, a method that traces behavioral changes in post-trained language models to responsible training datap

model-releasesarxiv-cs-ai
28 Apr 2026
Tutorials

Probing Visual Planning in Image Editing Models

DGX agent

arXiv:2604.22868v1 Announce Type: cross Abstract: Visual planning represents a crucial facet of human intelligence, especially in tasks that require complex spatial reasoning and navigation. Yet, in m

tutorialsarxiv-cs-ai
28 Apr 2026
Safety

ProEval: Proactive Failure Discovery and Efficient Performance Estimation for Generative AI Evaluation

DGX agent

arXiv:2604.23099v1 Announce Type: cross Abstract: Evaluating generative AI models is increasingly resource-intensive due to slow inference, expensive raters, and a rapidly growing landscape of models

safetyarxiv-cs-ai
28 Apr 2026
Model Releases

Progressive Approximation in Deep Residual Networks: Theory and Validation

DGX agent

arXiv:2604.24154v1 Announce Type: cross Abstract: The Universal Approximation Theorem (UAT) guarantees universal function approximation but does not explain how residual models distribute approximatio

model-releasesarxiv-cs-ai
28 Apr 2026
Safety

Protecting the Trace: A Principled Black-Box Approach Against Distillation Attacks

DGX agent

arXiv:2604.23238v1 Announce Type: cross Abstract: Frontier models push the boundaries of what is learnable at extreme computational costs, yet distillation via sampling reasoning traces exposes closed

safetyarxiv-cs-ai
28 Apr 2026
Research

PushupBench: Your VLM is not good at counting pushups

DGX agent

arXiv:2604.23407v1 Announce Type: cross Abstract: Large vision-language models (VLMs) can recognize extit{what} happens in video but fail to count extit{how many} times. We introduce extbf{PushupBench

researcharxiv-cs-ai
28 Apr 2026
Model Releases

QED: An Open-Source Multi-Agent System for Generating Mathematical Proofs on Open Problems

DGX agent

arXiv:2604.24021v1 Announce Type: new Abstract: We explore a central question in AI for mathematics: can AI systems produce original, nontrivial proofs for open research problems? Despite strong bench

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

QEVA: A Reference-Free Evaluation Metric for Narrative Video Summarization with Multimodal Question Answering

DGX agent

arXiv:2604.24052v1 Announce Type: cross Abstract: Video-to-text summarization remains underexplored in terms of comprehensive evaluation methods. Traditional n-gram overlap-based metrics and recent la

model-releasesarxiv-cs-ai
28 Apr 2026
Applications

Quantifying and Improving the Robustness of Retrieval-Augmented Language Models Against Spurious Features in Grounding Data

DGX agent

arXiv:2503.05587v3 Announce Type: replace-cross Abstract: Robustness has become a critical attribute for the deployment of RAG systems in real-world applications. Existing research focuses on robustne

applicationsarxiv-cs-ai
28 Apr 2026
Safety

Quantifying and Mitigating Self-Preference Bias of LLM Judges

DGX agent

arXiv:2604.22891v1 Announce Type: cross Abstract: LLM-as-a-Judge has become a dominant approach in automated evaluation systems, playing critical roles in model alignment, leaderboard construction, qu

safetyarxiv-cs-ai
28 Apr 2026
Safety

Quantifying Divergence in Inter-LLM Communication Through API Retrieval and Ranking

DGX agent

arXiv:2604.22760v1 Announce Type: cross Abstract: Large language models (LLMs) increasingly operate as autonomous agents that reason over external APIs to perform complex tasks. However, their reliabi

safetyarxiv-cs-ai
28 Apr 2026
Research

Quantum Kernel Advantage over Classical Collapse in Medical Foundation Model Embeddings

DGX agent

arXiv:2604.24597v1 Announce Type: cross Abstract: We provide evidence of quantum kernel advantage under noiseless simulation in binary insurance classification on MIMIC-CXR chest radiographs using qua

researcharxiv-cs-ai
28 Apr 2026
Model Releases

Quantum Knowledge Graph: Modeling Context-Dependent Triplet Validity

DGX agent

arXiv:2604.23972v1 Announce Type: cross Abstract: Knowledge graphs (KGs) are increasingly used to support large lan guage model (LLM) reasoning, but standard triplet-based KGs treat each relation as g

model-releasesarxiv-cs-ai
28 Apr 2026
Research

Quasi-Quadratic Gradient: A New Direction for Accelerating the BFGS Method in Quasi-Newton Optimization

DGX agent

arXiv:2604.23922v1 Announce Type: cross Abstract: In this paper, we introduce the Quasi-Quadratic Gradient (QQG), a novel search direction designed to accelerate the BFGS method within the quasi-Newto

researcharxiv-cs-ai
28 Apr 2026
Research

Query2Diagram: Answering Developer Queries with UML Diagrams

DGX agent

arXiv:2604.23816v1 Announce Type: cross Abstract: Software documentation frequently becomes outdated or fails to exist entirely, yet developers need focused views of their codebase to understand compl

researcharxiv-cs-ai
28 Apr 2026
Tutorials

Question-Adaptive Graph Learning for Multi-hop Retrieval Augmented Generation

DGX agent

arXiv:2510.11541v2 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) has demonstrated its ability to enhance Large Language Models (LLMs) by integrating external knowledge so

tutorialsarxiv-cs-ai
28 Apr 2026
Local Ai

RADIANT-LLM: an Agentic Retrieval Augmented Generation Framework for Reliable Decision Support in Safety-Critical Nuclear Engineering

DGX agent

arXiv:2604.22755v1 Announce Type: cross Abstract: Reliable decision support in nuclear engineering requires traceable, domain-grounded knowledge retrieval, yet safety and risk analysis workflows remai

local-aiarxiv-cs-ai
28 Apr 2026
Model Releases

RAS: a Reliability Oriented Metric for Automatic Speech Recognition

DGX agent

arXiv:2604.24278v1 Announce Type: cross Abstract: Automatic speech recognition systems often produce confident yet incorrect transcriptions under noisy or ambiguous conditions, which can be misleading

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

RAT: RunAnyThing via Fully Automated Environment Configuration

DGX agent

arXiv:2604.23190v1 Announce Type: cross Abstract: Automating repository-level software engineering tasks is a foundational challenge for autonomous code agents, largely due to the difficulty of config

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

RaV-IDP: A Reconstruction-as-Validation Framework for Faithful Intelligent Document Processing

DGX agent

arXiv:2604.23644v1 Announce Type: cross Abstract: Intelligent document processing pipelines extract structured entities (tables, images, and text) from documents for use in downstream systems such as

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

RCSB PDB AI Help Desk: retrieval-augmented generation for protein structure deposition support

DGX agent

arXiv:2604.22800v1 Announce Type: cross Abstract: Motivation: Structural Biologists have contributed more than 245,000 experimentally determined three-dimensional structures of biological macromolecul

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

RealFin: How Well Do LLMs Reason About Finance When Users Leave Things Unsaid?

DGX agent

arXiv:2602.07096v2 Announce Type: replace-cross Abstract: Reliable financial reasoning requires knowing not only how to answer, but also when an answer cannot be justified. In real financial practice,

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Reasonably reasoning AI agents can avoid game-theoretic failures in zero-shot, provably

DGX agent

arXiv:2603.18563v2 Announce Type: replace Abstract: As autonomous AI agents increasingly mediate online platform markets, a fundamental question emerges: do these markets generate stable strategic out

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Reclaiming Residual Knowledge: A Novel Paradigm to Low-Bit Quantization

DGX agent

arXiv:2408.00923v2 Announce Type: replace-cross Abstract: This paper explores a novel paradigm in low-bit (i.e. 4-bits or lower) quantization, differing from existing state-of-the-art methods, by fram

model-releasesarxiv-cs-ai
28 Apr 2026
Agents

Reconstructive Authority Model: Runtime Execution Validity Under Partial Observability

DGX agent

arXiv:2604.22898v1 Announce Type: cross Abstract: Autonomous systems increasingly operate under partial observability where execution-relevant state is never fully accessible. Existing governance mech

agentsarxiv-cs-ai
28 Apr 2026
Applications

RedParrot: Accelerating NL-to-DSL for Business Analytics via Query Semantic Caching

DGX agent

arXiv:2604.22758v1 Announce Type: cross Abstract: Recently, at Xiaohongshu, the rapid expansion of e-commerce and advertising demands real-time business analytics with high accuracy and low latency. T

applicationsarxiv-cs-ai
28 Apr 2026
Model Releases

RefEvo: Agentic Design with Co-Evolutionary Verification for Agile Reference Model Generation

DGX agent

arXiv:2604.24218v1 Announce Type: cross Abstract: As the complexity of System-on-Chip (SoC) designs grows, the shift-left paradigm necessitates the rapid development of high-fidelity reference models

model-releasesarxiv-cs-ai
28 Apr 2026
Tutorials

ReFinE: Streamlining UI Mockup Iteration with Research Findings

DGX agent

arXiv:2604.04353v2 Announce Type: replace-cross Abstract: Although HCI research papers offer valuable design insights, designers often struggle to apply them in design workflows due to difficulties in

tutorialsarxiv-cs-ai
28 Apr 2026
Safety

Reflective Flow Sampling Enhancement

DGX agent

arXiv:2603.06165v2 Announce Type: replace-cross Abstract: The growing demand for text-to-image generation has led to rapid advances in generative modeling. Recently, text-to-image diffusion models tra

safetyarxiv-cs-ai
28 Apr 2026
← Previous
1…381382383384385…448
Next →