AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
Model Releases

ThermoQA: A Three-Tier Benchmark for Evaluating Thermodynamic Reasoning in Large Language Models

DGX agent

arXiv:2604.19758v1 Announce Type: new Abstract: We present ThermoQA, a benchmark of 293 open-ended engineering thermodynamics problems in three tiers: property lookups (110 Q), component analysis (101

model-releasesarxiv-cs-ai
23 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Research

Thinking While Listening: Fast-Slow Recurrence for Long-Horizon Sequential Modeling

DGX agent

arXiv:2604.01577v2 Announce Type: replace-cross Abstract: We extend the recent latent recurrent modeling to sequential input streams. By interleaving fast, recurrent latent updates with self-organizat

researcharxiv-cs-ai
23 Apr 2026
Model Releases

Tokenised Flow Matching for Hierarchical Simulation Based Inference

DGX agent

arXiv:2604.20723v1 Announce Type: cross Abstract: The cost of simulator evaluations is a key practical bottleneck for Simulation Based Inference (SBI). In hierarchical settings with shared global para

model-releasesarxiv-cs-ai
23 Apr 2026
Research

Toward Cross-Lingual Quality Classifiers for Multilingual Pretraining Data Selection

DGX agent

arXiv:2604.20549v1 Announce Type: cross Abstract: As Large Language Models (LLMs) scale, data curation has shifted from maximizing volume to optimizing the signal-to-noise ratio by performing quality

researcharxiv-cs-ai
23 Apr 2026
Model Releases

Towards Secure Logging: Characterizing and Benchmarking Logging Code Security Issues with LLMs

DGX agent

arXiv:2604.20211v1 Announce Type: cross Abstract: Logging code plays an important role in software systems by recording key events and behaviors, which are essential for debugging and monitoring. Howe

model-releasesarxiv-cs-ai
23 Apr 2026
Tutorials

Transformers Can Learn Connectivity in Some Graphs but Not Others

DGX agent

arXiv:2509.22343v2 Announce Type: replace-cross Abstract: Reasoning capability is essential to ensure the factual correctness of the responses of transformer-based Large Language Models (LLMs), and ro

tutorialsarxiv-cs-ai
23 Apr 2026
Research

Transparent Screening for LLM Inference and Training Impacts

DGX agent

arXiv:2604.19757v1 Announce Type: cross Abstract: This paper presents a transparent screening framework for estimating inference and training impacts of current large language models under limited obs

researcharxiv-cs-ai
23 Apr 2026
Research

Treatment, evidence, imitation, and chat

DGX agent

arXiv:2506.23040v4 Announce Type: replace-cross Abstract: Large language models are thought to have the potential to aid in medical decision making. This work investigates the degree to which this mig

researcharxiv-cs-ai
23 Apr 2026
Agents

TriEx: A Game-based Tri-View Framework for Explaining Internal Reasoning in Multi-Agent LLMs

DGX agent

arXiv:2604.20043v1 Announce Type: cross Abstract: Explainability for Large Language Model (LLM) agents is especially challenging in interactive, partially observable settings, where decisions depend o

agentsarxiv-cs-ai
23 Apr 2026
Agents

Trust, Lies, and Long Memories: Emergent Social Dynamics and Reputation in Multi-Round Avalon with LLM Agents

DGX agent

arXiv:2604.20582v1 Announce Type: cross Abstract: We study emergent social dynamics in LLM agents playing The Resistance: Avalon, a hidden-role deception game. Unlike prior work on single-game perform

agentsarxiv-cs-ai
23 Apr 2026
Research

TTKV: Temporal-Tiered KV Cache for Long-Context LLM Inference

DGX agent

arXiv:2604.19769v1 Announce Type: cross Abstract: Key-value (KV) caching is critical for efficient inference in large language models (LLMs), yet its memory footprint scales linearly with context leng

researcharxiv-cs-ai
23 Apr 2026
Model Releases

UCCL-Zip: Lossless Compression Supercharged GPU Communication

DGX agent

arXiv:2604.17172v2 Announce Type: replace-cross Abstract: The rapid growth of large language models (LLMs) has made GPU communication a critical bottleneck. While prior work reduces communication volu

model-releasesarxiv-cs-ai
23 Apr 2026
Research

uLEAD-TabPFN: Uncertainty-aware Dependency-based Anomaly Detection with TabPFN

DGX agent

arXiv:2604.20255v1 Announce Type: cross Abstract: Anomaly detection in tabular data is challenging due to high dimensionality, complex feature dependencies, and heterogeneous noise. Many existing meth

researcharxiv-cs-ai
23 Apr 2026
Research

Using Learning Theories to Evolve Human-Centered XAI: Future Perspectives and Challenges

DGX agent

arXiv:2604.19788v1 Announce Type: new Abstract: As Artificial Intelligence (AI) systems continue to grow in size and complexity, so does the difficulty of the quest for AI transparency. In a world of

researcharxiv-cs-ai
23 Apr 2026
Research

Utterance-Level Methods for Identifying Reliable ASR-Output for Child Speech

DGX agent

arXiv:2604.19801v1 Announce Type: cross Abstract: Automatic Speech Recognition (ASR) is increasingly used in applications involving child speech, such as language learning and literacy acquisition. Ho

researcharxiv-cs-ai
23 Apr 2026
Safety

V-tableR1: Process-Supervised Multimodal Table Reasoning with Critic-Guided Policy Optimization

DGX agent

arXiv:2604.20755v1 Announce Type: new Abstract: We introduce V-tableR1, a process-supervised reinforcement learning framework that elicits rigorous, verifiable reasoning from multimodal large language

safetyarxiv-cs-ai
23 Apr 2026
Applications

VAN-AD: Visual Masked Autoencoder with Normalizing Flow For Time Series Anomaly Detection

DGX agent

arXiv:2603.26842v2 Announce Type: replace-cross Abstract: Time series anomaly detection (TSAD) is essential for maintaining the reliability and security of IoT-enabled service systems. Existing method

applicationsarxiv-cs-ai
23 Apr 2026
Research

veScale-FSDP: Flexible and High-Performance FSDP at Scale

DGX agent

arXiv:2602.22437v3 Announce Type: replace-cross Abstract: Fully Sharded Data Parallel (FSDP), also known as Zero Redundancy Optimizer (ZeRO), is widely used for large-scale model training, because of

researcharxiv-cs-ai
23 Apr 2026
Model Releases

Vibrotactile Preference Learning: Uncertainty-Aware Preference Learning for Personalized Vibration Feedback

DGX agent

arXiv:2604.20210v1 Announce Type: cross Abstract: Individual differences in vibrotactile perception underscore the growing importance of personalization as haptic feedback becomes more prevalent in in

model-releasesarxiv-cs-ai
23 Apr 2026
Applications

VTouch++: A Multimodal Dataset with Vision-Based Tactile Enhancement for Bimanual Manipulation

DGX agent

arXiv:2604.20444v1 Announce Type: cross Abstract: Embodied intelligence has advanced rapidly in recent years; however, bimanual manipulation-especially in contact-rich tasks remains challenging. This

applicationsarxiv-cs-ai
23 Apr 2026
Safety

What Makes a Good AI Review? Concern-Level Diagnostics for AI Peer Review

DGX agent

arXiv:2604.19998v1 Announce Type: new Abstract: Evaluating AI-generated reviews by verdict agreement is widely recognized as insufficient, yet current alternatives rarely audit which concerns a system

safetyarxiv-cs-ai
23 Apr 2026
Safety

Where and What: Reasoning Dynamic and Implicit Preferences in Situated Conversational Recommendation

DGX agent

arXiv:2604.20749v1 Announce Type: new Abstract: Situated conversational recommendation (SCR), which utilizes visual scenes grounded in specific environments and natural language dialogue to deliver co

safetyarxiv-cs-ai
23 Apr 2026
Model Releases

White-Basilisk: A Hybrid Model for Code Vulnerability Detection

DGX agent

arXiv:2507.08540v5 Announce Type: replace-cross Abstract: The proliferation of software vulnerabilities presents a significant challenge to cybersecurity, necessitating more effective detection method

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Why AI-Generated Text Detection Fails: Evidence from Explainable AI Beyond Benchmark Accuracy

DGX agent

arXiv:2603.23146v2 Announce Type: replace-cross Abstract: The widespread adoption of Large Language Models (LLMs) has made the detection of AI-Generated text a pressing and complex challenge. Although

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

WorkflowGen:an adaptive workflow generation mechanism driven by trajectory experience

DGX agent

arXiv:2604.19756v1 Announce Type: cross Abstract: Large language model (LLM) agents often suffer from high reasoning overhead, excessive token consumption, unstable execution, and inability to reuse p

model-releasesarxiv-cs-ai
23 Apr 2026
Safety

Working Memory Constraints Scaffold Learning in Transformers under Data Scarcity

DGX agent

arXiv:2604.20789v1 Announce Type: cross Abstract: We investigate the integration of human-like working memory constraints into the Transformer architecture and implement several cognitively inspired a

safetyarxiv-cs-ai
23 Apr 2026
Model Releases

3D Foundation Model for Generalizable Disease Detection in Head Computed Tomography

DGX agent

arXiv:2502.02779v3 Announce Type: replace-cross Abstract: Head computed tomography (CT) imaging is a widely-used imaging modality with multitudes of medical indications, particularly in assessing path

model-releasesarxiv-cs-ai
22 Apr 2026
Research

A Dual Perspective on Synthetic Trajectory Generators: Utility Framework and Privacy Vulnerabilities

DGX agent

arXiv:2604.19653v1 Announce Type: new Abstract: Human mobility data are used in numerous applications, ranging from public health to urban planning. Human mobility is inherently sensitive, as it can c

researcharxiv-cs-ai
22 Apr 2026
Model Releases

A Functionality-Grounded Benchmark for Evaluating Web Agents in E-commerce Domains

DGX agent

arXiv:2508.15832v2 Announce Type: replace-cross Abstract: Web agents have shown great promise in performing many tasks on ecommerce website. To assess their capabilities, several benchmarks have been

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

A-MAR: Agent-based Multimodal Art Retrieval for Fine-Grained Artwork Understanding

DGX agent

arXiv:2604.19689v1 Announce Type: new Abstract: Understanding artworks requires multi-step reasoning over visual content and cultural, historical, and stylistic context. While recent multimodal large

model-releasesarxiv-cs-ai
22 Apr 2026
Research

A neural operator framework for data-driven discovery of stability and receptivity in physical systems

DGX agent

arXiv:2604.19465v1 Announce Type: cross Abstract: Understanding how complex systems respond to perturbations, such as whether they will remain stable or what their most sensitive patterns are, is a fu

researcharxiv-cs-ai
22 Apr 2026
Tutorials

A Proxy Consistency Loss for Grounded Fusion of Earth Observation and Location Encoders

DGX agent

arXiv:2604.18881v1 Announce Type: cross Abstract: Supervised learning with Earth observation inputs is often limited by the sparsity of high-quality labeled or in-situ measured data to use as training

tutorialsarxiv-cs-ai
22 Apr 2026
Agents

A Survey on MLLM-based Visually Rich Document Understanding: Methods, Challenges, and Emerging Trends

DGX agent

arXiv:2507.09861v2 Announce Type: replace-cross Abstract: Visually Rich Document Understanding (VRDU) has become a pivotal area of research, driven by the need to automatically interpret documents tha

agentsarxiv-cs-ai
22 Apr 2026
Agents

AblateCell: A Reproduce-then-Ablate Agent for Virtual Cell Repositories

DGX agent

arXiv:2604.19606v1 Announce Type: new Abstract: Systematic ablations are essential to attribute performance gains in AI Virtual Cells, yet they are rarely performed because biological repositories are

agentsarxiv-cs-ai
22 Apr 2026
Model Releases

AD-Copilot: A Vision-Language Assistant for Industrial Anomaly Detection via Visual In-context Comparison

DGX agent

arXiv:2603.13779v2 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) have achieved impressive success in natural visual understanding, yet they consistently underperform

model-releasesarxiv-cs-ai
22 Apr 2026
Research

Adapting Dijkstra for Buffers and Unlimited Transfers

DGX agent

arXiv:2603.11729v2 Announce Type: replace-cross Abstract: In recent years, RAPTOR based algorithms have been considered the state-of-the-art for path-finding with unlimited transfers without preproces

researcharxiv-cs-ai
22 Apr 2026
Applications

Adaptive MSD-Splitting: Enhancing C4.5 and Random Forests for Skewed Continuous Attributes

DGX agent

arXiv:2604.19722v1 Announce Type: cross Abstract: The discretization of continuous numerical attributes remains a persistent computational bottleneck in the induction of decision trees, particularly a

applicationsarxiv-cs-ai
22 Apr 2026
Safety

Adaptive Prompt Elicitation for Text-to-Image Generation

DGX agent

arXiv:2602.04713v2 Announce Type: replace-cross Abstract: Aligning text-to-image generation with user intent remains challenging, as users frequently provide ambiguous inputs and struggle with model i

safetyarxiv-cs-ai
22 Apr 2026
Research

Adversarial Attacks on Medical Hyperspectral Imaging Exploiting Spectral-Spatial Dependencies and Multiscale Features

DGX agent

arXiv:2601.07056v2 Announce Type: replace-cross Abstract: Medical hyperspectral imaging (MHSI) has shown strong potential for disease diagnosis by capturing spectral-spatial information of tissues. Wh

researcharxiv-cs-ai
22 Apr 2026
Model Releases

Agent-GWO: Collaborative Agents for Dynamic Prompt Optimization in Large Language Models

DGX agent

arXiv:2604.18612v1 Announce Type: cross Abstract: Large Language Models (LLMs) have demonstrated strong capabilities in complex reasoning tasks, while recent prompting strategies such as Chain-of-Thou

model-releasesarxiv-cs-ai
22 Apr 2026
Agents

AgentDynEx: Nudging the Mechanics and Dynamics of Multi-Agent Simulations

DGX agent

arXiv:2504.09662v3 Announce Type: replace-cross Abstract: Multi-agent large language model simulations have the potential to model complex human behaviors and interactions. If the mechanics are set up

agentsarxiv-cs-ai
22 Apr 2026
Model Releases

Agentic Forecasting using Sequential Bayesian Updating of Linguistic Beliefs

DGX agent

arXiv:2604.18576v2 Announce Type: replace Abstract: We present BLF (Bayesian Linguistic Forecaster), an agentic system for binary forecasting that achieves state-of-the-art performance on the Forecast

model-releasesarxiv-cs-ai
22 Apr 2026
Research

AI-Based Detection of Temporal Changes in MR-Linac Images Acquired During Routine Prostate Radiotherapy

DGX agent

arXiv:2602.04983v2 Announce Type: replace-cross Abstract: Purpose: To investigate whether an AI-based method can detect subtle inter-fraction changes in MR-Linac images acquired during radiotherapy an

researcharxiv-cs-ai
22 Apr 2026
Agents

AI scientists produce results without reasoning scientifically

DGX agent

arXiv:2604.18805v1 Announce Type: new Abstract: Large language model (LLM)-based systems are increasingly deployed to conduct scientific research autonomously, yet whether their reasoning adheres to t

agentsarxiv-cs-ai
22 Apr 2026
Agents

An AI Agent Execution Environment to Safeguard User Data

DGX agent

arXiv:2604.19657v1 Announce Type: cross Abstract: AI agents promise to serve as general-purpose personal assistants for their users, which requires them to have access to private user data (e.g., pers

agentsarxiv-cs-ai
22 Apr 2026
Safety

ARES: Adaptive Red-Teaming and End-to-End Repair of Policy-Reward System

DGX agent

arXiv:2604.18789v1 Announce Type: new Abstract: Reinforcement Learning from Human Feedback (RLHF) is central to aligning Large Language Models (LLMs), yet it introduces a critical vulnerability: an im

safetyarxiv-cs-ai
22 Apr 2026
Hardware

ARGUS: Agentic GPU Optimization Guided by Data-Flow Invariants

DGX agent

arXiv:2604.18616v1 Announce Type: cross Abstract: LLM-based coding agents can generate functionally correct GPU kernels, yet their performance remains far below hand-optimized libraries on critical co

hardwarearxiv-cs-ai
22 Apr 2026
Safety

ARM: Advantage Reward Modeling for Long-Horizon Manipulation

DGX agent

arXiv:2604.03037v2 Announce Type: replace-cross Abstract: Long-horizon robotic manipulation remains challenging for reinforcement learning (RL) because sparse rewards provide limited guidance for cred

safetyarxiv-cs-ai
22 Apr 2026
← Previous
1…394395396397398…443
Next →