AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
1 May 2026

Theory Under Construction: Orchestrating Language Models for Research Software Where the Specification Evolves

Model ReleasesDGX agent

arXiv:2604.27209v1 Announce Type: cross Abstract: Large language models can now generate substantial code and draft research text, but research-software projects require more than either artifact alon

Think it, Run it: Autonomous ML pipeline generation via self-healing multi-agent AI

AgentsDGX agent

arXiv:2604.27096v1 Announce Type: new Abstract: The purpose of our paper is to develop a unified multi-agent architecture that automates end-to-end machine learning (ML) pipeline generation from datas

TiMem: Temporal-Hierarchical Memory Consolidation for Long-Horizon Conversational Agents

ResearchDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2601.02845v2 Announce Type: replace-cross Abstract: Long-horizon conversational agents have to manage ever-growing interaction histories that quickly exceed the finite context windows of large l

TIO-SHACL: Comprehensive SHACL validation for TMF Intent Ontologies

ResearchDGX agent

arXiv:2604.27359v1 Announce Type: new Abstract: Intent-based networking promises to revolutionize telecommunications network management by enabling operators to specify high-level goals rather than lo

To Build or Not to Build? Factors that Lead to Non-Development or Abandonment of AI Systems

Model ReleasesDGX agent

arXiv:2604.28053v1 Announce Type: cross Abstract: Responsible AI research typically focuses on examining the use and impacts of deployed AI systems. Yet, there is currently limited visibility into the

TopBench: A Benchmark for Implicit Prediction and Reasoning over Tabular Question Answering

Model ReleasesDGX agent

arXiv:2604.28076v1 Announce Type: cross Abstract: Large Language Models (LLMs) have advanced Table Question Answering, where most queries can be answered by extracting information or simple aggregatio

Toward Autonomous SOC Operations: End-to-End LLM Framework for Threat Detection, Query Generation, and Resolution in Security Operations

AgentsDGX agent

arXiv:2604.27321v1 Announce Type: cross Abstract: Security Operations Centers (SOCs) face mounting operational challenges. These challenges come from increasing threat volumes, heterogeneous SIEM plat

Toward Personalized Digital Twins for Cognitive Decline Assessment: A Multimodal, Uncertainty-Aware Framework

ResearchDGX agent

arXiv:2604.27217v1 Announce Type: new Abstract: Cognitive decline is highly heterogeneous across individuals, which complicates prognosis, trial design, and treatment planning. We present the Personal

Towards Accelerated SCF Workflows with Equivariant Density-Matrix Learning and Analytic Refinement

ResearchDGX agent

arXiv:2604.27256v1 Announce Type: cross Abstract: We present extsc{dm-PhiSNet}, a physically constrained extsc{PhiSNet}-based equivariant model that predicts one-electron reduced density matrices (1-R

Towards Neuro-symbolic Causal Rule Synthesis, Verification, and Evaluation Grounded in Legal and Safety Principles

SafetyDGX agent

arXiv:2604.28087v1 Announce Type: cross Abstract: Rule-based systems remain central in safety-critical domains but often struggle with scalability, brittleness, and goal misspecification. These limita

Towards single-shot coherent imaging via overlap-free ptychography

Model ReleasesDGX agent

arXiv:2602.21361v3 Announce Type: replace-cross Abstract: Ptychographic imaging at synchrotron and XFEL sources requires dense overlapping scans, limiting throughput and increasing dose. Extending coh

Trace-Level Analysis of Information Contamination in Multi-Agent Systems

AgentsDGX agent

arXiv:2604.27586v1 Announce Type: new Abstract: Reasoning over heterogeneous artifacts (PDFs, spreadsheets, slide decks, etc.) increasingly occurs within structured agent workflows that iteratively ex

Training-Free Reward-Guided Image Editing via Trajectory Optimal Control

ResearchDGX agent

arXiv:2509.25845v3 Announce Type: replace-cross Abstract: Recent advancements in diffusion and flow-matching models have demonstrated remarkable capabilities in high-fidelity image synthesis. A promin

Training-Free Tunnel Defect Inspection and Engineering Interpretation via Visual Recalibration and Entity Reconstruction

Local AiDGX agent

arXiv:2604.27928v1 Announce Type: cross Abstract: Tunnel inspection requires outputs that can support defect localization, measurement, severity grading, and engineering documentation. Existing traini

Transformer-Empowered Actor-Critic Reinforcement Learning for Sequence-Aware Service Function Chain Partitioning

ResearchDGX agent

arXiv:2504.18902v2 Announce Type: replace-cross Abstract: In the forthcoming era of 6G networks, characterized by unprecedented data rates, ultra-low latency, and ubiquitous connectivity, effective ma

TransVLM: A Vision-Language Framework and Benchmark for Detecting Any Shot Transitions

Model ReleasesDGX agent

arXiv:2604.27975v1 Announce Type: cross Abstract: Traditional Shot Boundary Detection (SBD) inherently struggles with complex transitions by formulating the task around isolated cut points, frequently

TRUST: A Framework for Decentralized AI Service v.0.1

SafetyDGX agent

arXiv:2604.27132v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) and Multi-Agent Systems (MAS) in high-stakes domains demand reliable verification, yet centralized approaches suffer four

TypeBandit: Type-Level Context Allocation and Reweighting for Effective Attribute Completion in Heterogeneous Graph Neural Networks

ResearchDGX agent

arXiv:2604.27356v1 Announce Type: cross Abstract: Heterogeneous graphs are widely used to model multi-relational systems, but missing node attributes remain a major bottleneck for downstream learning.

Understanding and Improving Length Generalization in Hierarchical Sparse Attention Models

ResearchDGX agent

arXiv:2510.17196v3 Announce Type: replace-cross Abstract: Effectively processing long contexts is a critical challenge for language models. While standard Transformers are limited by quadratic complex

Unpacking Vibe Coding: Help-Seeking Processes in Student-AI Interactions While Programming

ApplicationsDGX agent

arXiv:2604.27134v1 Announce Type: new Abstract: Generative AI is reshaping higher education programming through vibe coding, where students collaborate with AI via natural language rather than writing

Unsupervised Electrofacies Classification and Porosity Characterization in the Offshore Keta Basin Using Wireline Logs

ResearchDGX agent

arXiv:2604.27126v1 Announce Type: new Abstract: This study presents an unsupervised machine learning workflow for electrofacies analysis in the offshore Keta Basin, Ghana, where core data are scarce.

Upskilling with Generative AI: Practices and Challenges for Freelance Knowledge Workers

TutorialsDGX agent

arXiv:2604.27231v1 Announce Type: cross Abstract: Freelance workers must continually acquire new skills to remain competitive in online labor markets, yet they lack the organizational training, mentor

Useless but Safe? Benchmarking Utility Recovery with User Intent Clarification in Multi-Turn Conversations

Model ReleasesDGX agent

arXiv:2604.27093v1 Announce Type: cross Abstract: Current LLM safety alignment techniques improve model robustness against adversarial attacks, but overlook whether and how LLMs can recover helpfulnes

Vanishing Contributions: A Unified Framework for Smooth and Iterative Model Compression

ResearchDGX agent

arXiv:2510.09696v2 Announce Type: replace-cross Abstract: The increasing scale of Deep Neural Networks (DNNs) increases the need for compression techniques such as pruning, quantization, and low-rank

VeriTaS: The First Dynamic Benchmark for Multimodal Automated Fact-Checking

Model ReleasesDGX agent

arXiv:2601.08611v2 Announce Type: replace-cross Abstract: The growing scale of online misinformation urgently demands Automated Fact-Checking (AFC). Existing benchmarks for evaluating AFC systems, how

VibroML: an automated toolkit for high-throughput vibrational analysis and dynamic instability remediation of crystalline materials using machine-learned potentials

ResearchDGX agent

arXiv:2604.27685v1 Announce Type: cross Abstract: While machine-learned interatomic potentials (MLIPs) accelerate phonon dispersion calculations, merely identifying dynamical instabilities in computat

VIPaint: Image Inpainting with Pre-Trained Diffusion Models via Variational Inference

TutorialsDGX agent

arXiv:2411.18929v2 Announce Type: replace-cross Abstract: Diffusion probabilistic models learn to remove noise added during training, generating novel data (e.g., images) from Gaussian noise through s

WaferSAGE: Large Language Model-Powered Wafer Defect Analysis via Synthetic Data Generation and Rubric-Guided Reinforcement Learning

Model ReleasesDGX agent

arXiv:2604.27629v1 Announce Type: new Abstract: We present WaferSAGE, a framework for wafer defect visual question answering using small vision-language models. To address data scarcity in semiconduct

Web2BigTable: A Bi-Level Multi-Agent LLM System for Internet-Scale Information Search and Extraction

AgentsDGX agent

arXiv:2604.27221v1 Announce Type: new Abstract: Agentic web search increasingly faces two distinct demands: deep reasoning over a single target, and structured aggregation across many entities and het

What Makes a Good Terminal-Agent Benchmark Task: A Guideline for Adversarial, Difficult, and Legible Evaluation Design

Model ReleasesDGX agent

arXiv:2604.28093v1 Announce Type: new Abstract: Terminal-agent benchmarks have become a primary signal for measuring the coding and system-administration capabilities of large language models. As the

What Suppresses Nash Equilibrium Play in Large Language Models? Mechanistic Evidence and Causal Control

Model ReleasesDGX agent

arXiv:2604.27167v1 Announce Type: cross Abstract: LLM agents are known to deviate from Nash equilibria in strategic interactions, but nobody has looked inside the model to understand why, or asked whe

When 2D Tasks Meet 1D Serialization: On Serialization Friction in Structured Tasks

Local AiDGX agent

arXiv:2604.27272v1 Announce Type: cross Abstract: Large language models (LLMs) conventionally process structured inputs as 1D token sequences. While natural for prose, such linearization may introduce

When Agents Evolve, Institutions Follow

AgentsDGX agent

arXiv:2604.27691v1 Announce Type: new Abstract: Across millennia, complex societies have faced the same coordination problem of how to organize collective action among cognitively bounded and informat

When Continual Learning Moves to Memory: A Study of Experience Reuse in LLM Agents

Model ReleasesDGX agent

arXiv:2604.27003v1 Announce Type: cross Abstract: Memory-augmented LLM agents offer an appealing shortcut to continual learning: rather than updating model parameters, they accumulate experience in ex

When Does Structure Matter in Continual Learning? Dimensionality Controls When Modularity Shapes Representational Geometry

SafetyDGX agent

arXiv:2604.27656v1 Announce Type: cross Abstract: To preserve previously learned representations, continual learning systems must strike a balance between plasticity, the ability to acquire new knowle

When Roles Fail: Epistemic Constraints on Advocate Role Fidelity in LLM-Based Political Statement Analysis

Model ReleasesDGX agent

arXiv:2604.27228v1 Announce Type: new Abstract: Democratic discourse analysis systems increasingly rely on multi-agent LLM pipelines in which distinct evaluator models are assigned adversarial roles t

When Your LLM Reaches End-of-Life: A Framework for Confident Model Migration in Production Systems

ApplicationsDGX agent

arXiv:2604.27082v1 Announce Type: new Abstract: We present a framework for migrating production Large Language Model (LLM) based systems when the underlying model reaches end-of-life or requires repla

Why Self-Supervised Encoders Want to Be Normal

ResearchDGX agent

arXiv:2604.27743v1 Announce Type: cross Abstract: We develop a geometric and information-theoretic framework for encoder-decoder learning built on the Information Bottleneck (IB) principle. Recasting

WindowsWorld: A Process-Centric Benchmark of Autonomous GUI Agents in Professional Cross-Application Environments

Model ReleasesDGX agent

arXiv:2604.27776v1 Announce Type: new Abstract: While GUI agents have shown impressive capabilities in common computer-use tasks such as OSWorld, current benchmarks mainly focus on isolated and single

ZAYAN: Disentangled Contrastive Transformer for Tabular Remote Sensing Data

ResearchDGX agent

arXiv:2604.27606v1 Announce Type: cross Abstract: Learning informative representations from tabular data in remote sensing and environmental science is challenging due to heterogeneity, scarce labels,

30 Apr 2026

A Data-Centric Framework for Intraoperative Fluorescence Lifetime Imaging for Glioma Surgical Guidance

TutorialsDGX agent

arXiv:2604.26147v1 Announce Type: cross Abstract: Accurate intraoperative assessment of glioma infiltration is essential for maximizing tumor resection while preserving functional brain tissue. Fluore

A Decision-Theoretic Formalisation of Steganography With Applications to LLM Monitoring

ResearchDGX agent

arXiv:2602.23163v3 Announce Type: replace Abstract: Large language models are beginning to show steganographic capabilities. Such capabilities could allow misaligned models to evade oversight mechanis

A Practice of Post-Training on Llama-3 70B with Optimal Selection of Additional Language Mixture Ratio

Model ReleasesDGX agent

arXiv:2409.06624v4 Announce Type: replace-cross Abstract: Large Language Models (LLM) often need to be Continual Pre-Trained (CPT) to obtain unfamiliar language skills or adapt to new domains. The hug

A Randomized PDE Energy driven Iterative Framework for Efficient and Stable PDE Solutions

ResearchDGX agent

arXiv:2604.25943v1 Announce Type: cross Abstract: Efficient and stable solution of partial differential equations (PDEs) is central to scientific and engineering applications, yet existing numerical s

A Scoping Review of LLM-as-a-Judge in Healthcare and the MedJUDGE Framework

SafetyDGX agent

arXiv:2604.25933v1 Announce Type: cross Abstract: As large language models (LLMs) increasingly generate and process clinical text, scalable evaluation has become critical. LLM-as-a-Judge (LaaJ), which

A Self-Calibrating Framework for Analog Circuit Sizing Using LLM-Derived Analytical Equations

ResearchDGX agent

arXiv:2604.07387v2 Announce Type: replace-cross Abstract: We present a design automation framework for analog circuit sizing that produces calibrated, topology-specific analytical equations from raw c

A self-evolving agent for explainable diagnosis of DFT-experiment band-gap mismatch

Model ReleasesDGX agent

arXiv:2604.26703v1 Announce Type: cross Abstract: Standard density functional theory (DFT) routinely misclassifies the electronic ground state of correlated and structurally complex compounds, predict

A Survey of Multi-Agent Deep Reinforcement Learning with Graph Neural Network-Based Communication

AgentsDGX agent

arXiv:2604.25972v1 Announce Type: cross Abstract: In multi-agent reinforcement learning (MARL), the integration of a communication mechanism, allowing agents to better learn to coordinate their action

A Survey of Process Reward Models: From Outcome Signals to Process Supervisions for Large Language Models

SafetyDGX agent

arXiv:2510.08049v3 Announce Type: replace-cross Abstract: Although Large Language Models (LLMs) exhibit advanced reasoning ability, conventional alignment remains largely dominated by outcome reward m

A Survey on the Safety and Security Threats of Computer-Using Agents: JARVIS or Ultron?

SafetyDGX agent

arXiv:2505.10924v4 Announce Type: replace-cross Abstract: Recently, AI-driven interactions with computing devices have advanced from basic prototype tools to sophisticated, LLM-based systems that emul

A Toolkit for Detecting Spurious Correlations in Speech Datasets

ResearchDGX agent

arXiv:2604.26676v1 Announce Type: cross Abstract: We introduce a toolkit for uncovering spurious correlations between recording characteristics and target class in speech datasets. Spurious correlatio

ACPO: Anchor-Constrained Perceptual Optimization for Diffusion Models with No-Reference Quality Guidance

ResearchDGX agent

arXiv:2604.26348v1 Announce Type: cross Abstract: Diffusion models have achieved remarkable success in image generation, yet their training is predominantly driven by full-reference objectives that en

AdaFRUGAL: Adaptive Memory-Efficient Training with Dynamic Control

HardwareDGX agent

arXiv:2601.11568v2 Announce Type: replace-cross Abstract: Training Large Language Models (LLMs) is highly memory-intensive due to optimizer state overhead. The FRUGAL framework mitigates this with gra

Adaptive Layerwise Perturbation: Unifying Off-Policy Corrections for LLM RL

Local AiDGX agent

arXiv:2603.19470v2 Announce Type: replace-cross Abstract: Off-policy problems such as policy staleness and training--inference mismatch have become a major bottleneck for training stability and furthe

Affective Flow Language Model for Emotional Support Conversation

Model ReleasesDGX agent

arXiv:2602.08826v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have been widely applied to emotional support conversation (ESC). However, complex multi-turn support remains cha

AGEL-Comp: A Neuro-Symbolic Framework for Compositional Generalization in Interactive Agents

AgentsDGX agent

arXiv:2604.26522v1 Announce Type: new Abstract: Large Language Model (LLM)-based agents exhibit systemic failures in compositional generalization, limiting their robustness in interactive environments

AHASD: Asynchronous Heterogeneous Architecture for LLM Adaptive Drafting Speculative Decoding on Mobile Devices

HardwareDGX agent

arXiv:2604.25326v2 Announce Type: replace-cross Abstract: Speculative decoding enhances the inference efficiency of large language models (LLMs) by generating drafts using a small draft language model

AMMA: A Multi-Chiplet Memory-Centric Architecture for Low-Latency 1M Context Attention Serving

HardwareDGX agent

arXiv:2604.26103v1 Announce Type: cross Abstract: All current LLM serving systems place the GPU at the center, from production-level attention-FFN disaggregation to NVIDIA's Rubin GPU-LPU heterogeneou

Analysing Lightweight Large Language Models for Biomedical Named Entity Recognition on Diverse Ouput Formats

ApplicationsDGX agent

arXiv:2604.25920v1 Announce Type: cross Abstract: Despite their strong linguistic capabilities, Large Language Models (LLMs) are computationally demanding and require substantial resources for fine-tu

Apriori-based Analysis of Learned Helplessness in Mathematics Tutoring: Behavioral Patterns by Level, Intervention, and Outcome

ResearchDGX agent

arXiv:2604.26237v1 Announce Type: new Abstract: This study applied the Apriori algorithm to analyze behavioral interaction patterns associated with learned helplessness (LH) in mathematics tutoring sy

← Previous
1…291292293294295…354
Next →