AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
Research

VisionFoundry: Teaching VLMs Visual Perception with Synthetic Images

DGX agent

arXiv:2604.09531v1 Announce Type: cross Abstract: Vision-language models (VLMs) still struggle with visual perception tasks such as spatial understanding and viewpoint recognition. One plausible contr

researcharxiv-cs-ai
13 Apr 2026
Safety
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

VISOR: Agentic Visual Retrieval-Augmented Generation via Iterative Search and Over-horizon Reasoning

DGX agent

arXiv:2604.09508v1 Announce Type: cross Abstract: Visual Retrieval-Augmented Generation (VRAG) empowers Vision-Language Models to retrieve and reason over visually rich documents. To tackle complex qu

safetyarxiv-cs-ai
13 Apr 2026
Safety

Visually-Guided Policy Optimization for Multimodal Reasoning

DGX agent

arXiv:2604.09349v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has significantly advanced the reasoning ability of vision-language models (VLMs). However, the

safetyarxiv-cs-ai
13 Apr 2026
Research

VL-Calibration: Decoupled Confidence Calibration for Large Vision-Language Models Reasoning

DGX agent

arXiv:2604.09529v1 Announce Type: cross Abstract: Large Vision Language Models (LVLMs) achieve strong multimodal reasoning but frequently exhibit hallucinations and incorrect responses with high certa

researcharxiv-cs-ai
13 Apr 2026
Model Releases

VOLTA: The Surprising Ineffectiveness of Auxiliary Losses for Calibrated Deep Learning

DGX agent

arXiv:2604.08639v1 Announce Type: cross Abstract: Uncertainty quantification (UQ) is essential for deploying deep learning models in safety critical applications, yet no consensus exists on which UQ m

model-releasesarxiv-cs-ai
13 Apr 2026
Research

VSI: Visual Subtitle Integration for Keyframe Selection to enhance Long Video Understanding

DGX agent

arXiv:2508.06869v4 Announce Type: replace-cross Abstract: Multimodal large language models (MLLMs) demonstrate exceptional performance in vision-language tasks, yet their processing of long videos is

researcharxiv-cs-ai
13 Apr 2026
Research

WAND: Windowed Attention and Knowledge Distillation for Efficient Autoregressive Text-to-Speech Models

DGX agent

arXiv:2604.08558v1 Announce Type: cross Abstract: Recent decoder-only autoregressive text-to-speech (AR-TTS) models produce high-fidelity speech, but their memory and compute costs scale quadratically

researcharxiv-cs-ai
13 Apr 2026
Model Releases

Watt Counts: Energy-Aware Benchmark for Sustainable LLM Inference on Heterogeneous GPU Architectures

DGX agent

arXiv:2604.09048v1 Announce Type: cross Abstract: While the large energy consumption of Large Language Models (LLMs) is recognized by the community, system operators lack guidance for energy-efficient

model-releasesarxiv-cs-ai
13 Apr 2026
Research

Webscale-RL: Automated Data Pipeline for Scaling RL Data to Pretraining Levels

DGX agent

arXiv:2510.06499v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have achieved remarkable success through imitation learning on vast text corpora, but this paradigm creates a tra

researcharxiv-cs-ai
13 Apr 2026
Model Releases

When Identity Skews Debate: Anonymization for Bias-Reduced Multi-Agent Reasoning

DGX agent

arXiv:2510.07517v5 Announce Type: replace Abstract: Multi-agent debate (MAD) aims to improve large language model (LLM) reasoning by letting multiple agents exchange answers and then aggregate their o

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Why Adam Can Beat SGD: Second-Moment Normalization Yields Sharper Tails

DGX agent

arXiv:2603.03099v5 Announce Type: replace-cross Abstract: Despite Adam demonstrating faster empirical convergence than SGD in many applications, much of the existing theory yields guarantees essential

model-releasesarxiv-cs-ai
13 Apr 2026
Tutorials

WOMBET: World Model-based Experience Transfer for Robust and Sample-efficient Reinforcement Learning

DGX agent

arXiv:2604.08958v1 Announce Type: cross Abstract: Reinforcement learning (RL) in robotics is often limited by the cost and risk of data collection, motivating experience transfer from a source task to

tutorialsarxiv-cs-ai
13 Apr 2026
Model Releases

XFED: Non-Collusive Model Poisoning Attack Against Byzantine-Robust Federated Classifiers

DGX agent

arXiv:2604.09489v1 Announce Type: cross Abstract: Model poisoning attacks pose a significant security threat to Federated Learning (FL). Most existing model poisoning attacks rely on collusion, requir

model-releasesarxiv-cs-ai
13 Apr 2026
Agents

Yes, But Not Always. Generative AI Needs Nuanced Opt-in

DGX agent

arXiv:2604.09413v1 Announce Type: cross Abstract: This paper argues that a one-size-fits-all approach to specifying consent for the use of creative works in generative AI is insufficient. Real-world o

agentsarxiv-cs-ai
13 Apr 2026
Safety

You've Got a Golden Ticket: Improving Generative Robot Policies With A Single Noise Vector

DGX agent

arXiv:2603.15757v2 Announce Type: replace-cross Abstract: What happens when a pretrained generative robot policy is provided a constant initial noise as input, rather than repeatedly sampling it from

safetyarxiv-cs-ai
13 Apr 2026
Safety

A Clinical Point Cloud Paradigm for In-Hospital Mortality Prediction from Multi-Level Incomplete Multimodal EHRs

DGX agent

arXiv:2604.04614v2 Announce Type: replace-cross Abstract: Deep learning-based modeling of multimodal Electronic Health Records (EHRs) has become an important approach for clinical diagnosis and risk p

safetyarxiv-cs-ai
10 Apr 2026
Applications

A Comparative Study of Demonstration Selection for Practical Large Language Models-based Next POI Prediction

DGX agent

arXiv:2604.06207v1 Announce Type: cross Abstract: This paper investigates demonstration selection strategies for predicting a user's next point-of-interest (POI) using large language models (LLMs), ai

applicationsarxiv-cs-ai
10 Apr 2026
Agents

A Comprehensive Survey of Agents for Computer Use: Foundations, Challenges, and Future Directions

DGX agent

arXiv:2501.16150v3 Announce Type: replace Abstract: Agents for computer use (ACUs) are an emerging class of systems capable of executing complex tasks on digital devices -- such as desktops, mobile ph

agentsarxiv-cs-ai
10 Apr 2026
Safety

A First Guess is Rarely the Final Answer: Learning to Search in the Travelling Salesperson Problem

DGX agent

arXiv:2604.06940v1 Announce Type: cross Abstract: Most neural solvers for the Traveling Salesperson Problem (TSP) are trained to output a single solution, even though practitioners rarely stop there:

safetyarxiv-cs-ai
10 Apr 2026
Research

A Goal-Oriented Chatbot for Engaging the Elderly Through Family Photo Conversations

DGX agent

arXiv:2604.06184v1 Announce Type: cross Abstract: We propose a personalized chatbot designed for elderly individuals. The chatbot initiates discussions based on family photos, encouraging users to int

researcharxiv-cs-ai
10 Apr 2026
Research

A Graph-Enhanced Defense Framework for Explainable Fake News Detection with LLM

DGX agent

arXiv:2604.06666v1 Announce Type: cross Abstract: Explainable fake news detection aims to assess the veracity of news claims while providing human-friendly explanations. Existing methods incorporating

researcharxiv-cs-ai
10 Apr 2026
Hardware

A Lightweight Library for Energy-Based Joint-Embedding Predictive Architectures

DGX agent

arXiv:2602.03604v3 Announce Type: replace-cross Abstract: We present EB-JEPA, an open-source library for learning representations and world models using Joint-Embedding Predictive Architectures (JEPAs

hardwarearxiv-cs-ai
10 Apr 2026
Model Releases

A-MBER: Affective Memory Benchmark for Emotion Recognition

DGX agent

arXiv:2604.07017v1 Announce Type: new Abstract: AI assistants that interact with users over time need to interpret the user's current emotional state in order to respond appropriately and personally.

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

A Novel Automatic Framework for Speaker Drift Detection in Synthesized Speech

DGX agent

arXiv:2604.06327v1 Announce Type: cross Abstract: Recent diffusion-based text-to-speech (TTS) models achieve high naturalness and expressiveness, yet often suffer from speaker drift, a subtle, gradual

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

A Parameter-Efficient Transfer Learning Approach through Multitask Prompt Distillation and Decomposition for Clinical NLP

DGX agent

arXiv:2604.06650v1 Announce Type: cross Abstract: Existing prompt-based fine-tuning methods typically learn task-specific prompts independently, imposing significant computing and storage overhead at

model-releasesarxiv-cs-ai
10 Apr 2026
Tutorials

A Severity-Based Curriculum Learning Strategy for Arabic Medical Text Generation

DGX agent

arXiv:2604.06365v1 Announce Type: cross Abstract: Arabic medical text generation is increasingly needed to help users interpret symptoms and access general health guidance in their native language. Ne

tutorialsarxiv-cs-ai
10 Apr 2026
Research

A Study of LLMs' Preferences for Libraries and Programming Languages

DGX agent

arXiv:2503.17181v3 Announce Type: replace-cross Abstract: Despite the rapid progress of large language models (LLMs) in code generation, existing evaluations focus on functional correctness or syntact

researcharxiv-cs-ai
10 Apr 2026
Model Releases

A Systematic Study of Retrieval Pipeline Design for Retrieval-Augmented Medical Question Answering

DGX agent

arXiv:2604.07274v1 Announce Type: cross Abstract: Large language models (LLMs) have demonstrated strong capabilities in medical question answering; however, purely parametric models often suffer from

model-releasesarxiv-cs-ai
10 Apr 2026
Research

AdaProb: Efficient Machine Unlearning via Adaptive Probability

DGX agent

arXiv:2411.02622v3 Announce Type: replace-cross Abstract: Machine unlearning, enabling a trained model to forget specific data, is crucial for addressing erroneous data and adhering to privacy regulat

researcharxiv-cs-ai
10 Apr 2026
Applications

Adaptive Differential Privacy for Federated Medical Image Segmentation Across Diverse Modalities

DGX agent

arXiv:2604.06518v1 Announce Type: cross Abstract: Large volumes of medical data remain underutilized because centralizing distributed data is often infeasible due to strict privacy regulations and ins

applicationsarxiv-cs-ai
10 Apr 2026
Safety

Adaptive Replay Buffer for Offline-to-Online Reinforcement Learning

DGX agent

arXiv:2512.10510v2 Announce Type: replace-cross Abstract: Offline-to-Online Reinforcement Learning (O2O RL) faces a critical dilemma in balancing the use of a fixed offline dataset with newly collecte

safetyarxiv-cs-ai
10 Apr 2026
Safety

AEROS: A Single-Agent Operating Architecture with Embodied Capability Modules

DGX agent

arXiv:2604.07039v1 Announce Type: cross Abstract: Robotic systems lack a principled abstraction for organizing intelligence, capabilities, and execution in a unified manner. Existing approaches either

safetyarxiv-cs-ai
10 Apr 2026
Safety

AgentCity: Constitutional Governance for Autonomous Agent Economies via Separation of Power

DGX agent

arXiv:2604.07007v1 Announce Type: cross Abstract: Autonomous AI agents are beginning to operate across organizational boundaries on the open internet -- discovering, transacting with, and delegating t

safetyarxiv-cs-ai
10 Apr 2026
Model Releases

AgentGate: A Lightweight Structured Routing Engine for the Internet of Agents

DGX agent

arXiv:2604.06696v1 Announce Type: new Abstract: The rapid development of AI agent systems is leading to an emerging Internet of Agents, where specialized agents operate across local devices, edge node

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

AgentOpt v0.1 Technical Report: Client-Side Optimization for LLM-Based Agent

DGX agent

arXiv:2604.06296v1 Announce Type: cross Abstract: AI agents are increasingly deployed in real-world applications, including systems such as Manus, OpenClaw, and coding agents. Existing research has pr

model-releasesarxiv-cs-ai
10 Apr 2026
Safety

AI-Driven Research for Databases

DGX agent

arXiv:2604.06566v1 Announce Type: cross Abstract: As the complexity of modern workloads and hardware increasingly outpaces human research and engineering capacity, existing methods for database perfor

safetyarxiv-cs-ai
10 Apr 2026
Model Releases

An Automated Survey of Generative Artificial Intelligence: Large Language Models, Architectures, Protocols, and Applications

DGX agent

arXiv:2306.02781v4 Announce Type: replace-cross Abstract: Generative artificial intelligence, and large language models in particular, have emerged as one of the most transformative paradigms in moder

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

An empirical study of LoRA-based fine-tuning of large language models for automated test case generation

DGX agent

arXiv:2604.06946v1 Announce Type: cross Abstract: Automated test case generation from natural language requirements remains a challenging problem in software engineering due to the ambiguity of requir

model-releasesarxiv-cs-ai
10 Apr 2026
Tutorials

Analyzing Multimodal Interaction Strategies for LLM-Assisted Manipulation of 3D Scenes

DGX agent

arXiv:2410.22177v2 Announce Type: replace-cross Abstract: As more applications of large language models (LLMs) for 3D content for immersive environments emerge, it is crucial to study user behaviour t

tutorialsarxiv-cs-ai
10 Apr 2026
Safety

Android Coach: Improve Online Agentic Training Efficiency with Single State Multiple Actions

DGX agent

arXiv:2604.07277v1 Announce Type: cross Abstract: Online reinforcement learning (RL) serves as an effective method for enhancing the capabilities of Android agents. However, guiding agents to learn th

safetyarxiv-cs-ai
10 Apr 2026
Model Releases

Asking like Socrates: Socrates helps VLMs understand remote sensing images

DGX agent

arXiv:2511.22396v2 Announce Type: replace-cross Abstract: Recent multimodal reasoning models, inspired by DeepSeek-R1, have significantly advanced vision-language systems. However, in remote sensing (

model-releasesarxiv-cs-ai
10 Apr 2026
Agents

Assessing the Added Value of Onboard Earth Observation Processing with the IRIDE HEO Service Segment

DGX agent

arXiv:2604.07120v1 Announce Type: cross Abstract: Current operational Earth Observation (EO) services, including the Copernicus Emergency Management Service (CEMS), the European Forest Fire Informatio

agentsarxiv-cs-ai
10 Apr 2026
Model Releases

ATANT: An Evaluation Framework for AI Continuity

DGX agent

arXiv:2604.06710v1 Announce Type: new Abstract: We present ATANT (Automated Test for Acceptance of Narrative Truth), an open evaluation framework for measuring continuity in AI systems: the ability to

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

ATBench: A Diverse and Realistic Agent Trajectory Benchmark for Safety Evaluation and Diagnosis

DGX agent

arXiv:2604.02022v2 Announce Type: replace Abstract: Evaluating the safety of LLM-based agents is increasingly important because risks in realistic deployments often emerge over multi-step interactions

model-releasesarxiv-cs-ai
10 Apr 2026
Safety

Attention Flows: Tracing LLM Conceptual Engagement via Story Summaries

DGX agent

arXiv:2604.06416v1 Announce Type: cross Abstract: Although LLM context lengths have grown, there is evidence that their ability to integrate information across long-form texts has not kept pace. We ev

safetyarxiv-cs-ai
10 Apr 2026
Tutorials

Attribution-Driven Explainable Intrusion Detection with Encoder-Based Large Language Models

DGX agent

arXiv:2604.06266v1 Announce Type: cross Abstract: Software-Defined Networking (SDN) improves network flexibility but also increases the need for reliable and interpretable intrusion detection. Large L

tutorialsarxiv-cs-ai
10 Apr 2026
Safety

AudioRole: An Audio Dataset for Character Role-Playing in Large Language Models

DGX agent

arXiv:2509.23435v2 Announce Type: replace-cross Abstract: The creation of high-quality multimodal datasets remains fundamental for advancing role-playing capabilities in large language models (LLMs).

safetyarxiv-cs-ai
10 Apr 2026
Model Releases

Automating Database-Native Function Code Synthesis with LLMs

DGX agent

arXiv:2604.06231v1 Announce Type: cross Abstract: Database systems incorporate an ever-growing number of functions in their kernels (a.k.a., database native functions) for scenarios like new applicati

model-releasesarxiv-cs-ai
10 Apr 2026
← Previous
1…434435436437438…443
Next →