AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
20,981 results
13 Apr 2026

Generalization and Scaling Laws for Mixture-of-Experts Transformers

Model ReleasesDGX agent

arXiv:2604.09175v1 Announce Type: cross Abstract: We develop a theory of generalization and scaling for Mixture-of-Experts (MoE) Transformers that cleanly separates active per-input capacity from rout

GNN-as-Judge: Unleashing the Power of LLMs for Graph Learning with GNN Feedback

SafetyDGX agent

arXiv:2604.08553v1 Announce Type: cross Abstract: Large Language Models (LLMs) have shown strong performance on text-attributed graphs (TAGs) due to their superior semantic understanding ability on te

GRM: Utility-Aware Jailbreak Attacks on Audio LLMs via Gradient-Ratio Masking

SafetyDGX agent

arXiv:2604.09222v1 Announce Type: cross Abstract: Audio large language models (ALLMs) enable rich speech-text interaction, but they also introduce jailbreak vulnerabilities in the audio modality. Exis


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Group-Aware Coordination Graph for Multi-Agent Reinforcement Learning

AgentsDGX agent

arXiv:2404.10976v4 Announce Type: replace-cross Abstract: Cooperative Multi-Agent Reinforcement Learning (MARL) necessitates seamless collaboration among agents, often represented by an underlying rel

H-AdminSim: A Multi-Agent Simulator for Realistic Hospital Administrative Workflows with FHIR Integration

AgentsDGX agent

arXiv:2602.05407v2 Announce Type: replace Abstract: Hospital administration departments handle a wide range of operational tasks and, in large hospitals, process over 10,000 requests per day, driving

Hidden in Plain Sight: Visual-to-Symbolic Analytical Solution Inference from Field Visualizations

Model ReleasesDGX agent

arXiv:2604.08863v1 Announce Type: new Abstract: Recovering analytical solutions of physical fields from visual observations is a fundamental yet underexplored capability for AI-assisted scientific rea

HiFloat4 Format for Language Model Pre-training on Ascend NPUs

Model ReleasesDGX agent

arXiv:2604.08826v1 Announce Type: cross Abstract: Large foundation models have become central to modern machine learning, with performance scaling predictably with model size and data. However, traini

HiL-Bench (Human-in-Loop Benchmark): Do Agents Know When to Ask for Help?

Model ReleasesDGX agent

arXiv:2604.09408v1 Announce Type: new Abstract: Frontier coding agents solve complex tasks when given complete context but collapse when specifications are incomplete or ambiguous. The bottleneck is n

HM-Bench: A Comprehensive Benchmark for Multimodal Large Language Models in Hyperspectral Remote Sensing

Model ReleasesDGX agent

arXiv:2604.08884v1 Announce Type: cross Abstract: While multimodal large language models (MLLMs) have made significant strides in natural image understanding, their ability to perceive and reason over

How Similar Are Grokipedia and Wikipedia? A Multi-Dimensional Textual and Structural Comparison

SafetyDGX agent

arXiv:2510.26899v5 Announce Type: replace-cross Abstract: The launch of Grokipedia, an AI-generated encyclopedia developed by Elon Musk's xAI, was presented as a response to perceived ideological and

HTNav: A Hybrid Navigation Framework with Tiered Structure for Urban Aerial Vision-and-Language Navigation

Model ReleasesDGX agent

arXiv:2604.08883v1 Announce Type: cross Abstract: Inspired by the general Vision-and-Language Navigation (VLN) task, aerial VLN has attracted widespread attention, owing to its significant practical v

Hypergraph Neural Networks Accelerate MUS Enumeration

AgentsDGX agent

arXiv:2604.09001v1 Announce Type: new Abstract: Enumerating Minimal Unsatisfiable Subsets (MUSes) is a fundamental task in constraint satisfaction problems (CSPs). Its major challenge is the exponenti

Identification and Anonymization of Named Entities in Unstructured Information Sources for Use in Social Engineering Detection

ApplicationsDGX agent

arXiv:2604.09016v1 Announce Type: cross Abstract: This study addresses the challenge of creating datasets for cybercrime analysis while complying with the requirements of regulations such as the Gener

Improving Automatic Summarization of Radiology Reports through Mid-Training of Large Language Models

Model ReleasesDGX agent

arXiv:2603.19275v2 Announce Type: replace-cross Abstract: Automatic summarization of radiology reports is an essential application to reduce the burden on physicians. Previous studies have widely used

InstrAct: Towards Action-Centric Understanding in Instructional Videos

SafetyDGX agent

arXiv:2604.08762v1 Announce Type: cross Abstract: Understanding instructional videos requires recognizing fine-grained actions and modeling their temporal relations, which remains challenging for curr

Interactive ASR: Towards Human-Like Interaction and Semantic Coherence Evaluation for Agentic Speech Recognition

AgentsDGX agent

arXiv:2604.09121v1 Announce Type: cross Abstract: Recent years have witnessed remarkable progress in automatic speech recognition (ASR), driven by advances in model architectures and large-scale train

Interactive Program Synthesis for Modeling Collaborative Physical Activities from Narrated Demonstrations

ResearchDGX agent

arXiv:2509.24250v3 Announce Type: replace Abstract: Teaching systems physical tasks is a long standing goal in HCI, yet most prior work has focused on non collaborative physical activities. Collaborat

Investigating Multimodal Large Language Models to Support Usability Evaluation

ApplicationsDGX agent

arXiv:2508.16165v2 Announce Type: replace-cross Abstract: Usability evaluation is an essential method to support the design of effective and intuitive user interfaces (UIs). However, it commonly relie

Joint Interference Detection and Identification via Adversarial Multi-task Learning

ResearchDGX agent

arXiv:2604.08607v1 Announce Type: cross Abstract: Precise interference detection and identification are crucial for enhancing the survivability of communication systems in non-cooperative wireless env

Kill-Chain Canaries: Stage-Level Tracking of Prompt Injection Across Attack Surfaces and Model Safety Tiers

Model ReleasesDGX agent

arXiv:2603.28013v3 Announce Type: replace-cross Abstract: Multi-agent LLM systems are entering production -- processing documents, managing workflows, acting on behalf of users -- yet their resilience

Large Language Models Generate Harmful Content Using a Distinct, Unified Mechanism

SafetyDGX agent

arXiv:2604.09544v1 Announce Type: cross Abstract: Large language models (LLMs) undergo alignment training to avoid harmful behaviors, yet the resulting safeguards remain brittle: jailbreaks routinely

Large-Scale Universal Defect Generation: Foundation Models and Datasets

ResearchDGX agent

arXiv:2604.08915v1 Announce Type: cross Abstract: Existing defect/anomaly generation methods often rely on few-shot learning, which overfits to specific defect categories due to the lack of large-scal

Learning General Representation of 12-Lead Electrocardiogram with a Joint-Embedding Predictive Architecture

TutorialsDGX agent

arXiv:2410.08559v5 Announce Type: replace-cross Abstract: Electrocardiogram (ECG) captures the heart's electrical signals, offering valuable information for diagnosing cardiac conditions. However, the

Learning Vision-Language-Action World Models for Autonomous Driving

SafetyDGX agent

arXiv:2604.09059v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have recently achieved notable progress in end-to-end autonomous driving by integrating perception, reasoning, and

Leave My Images Alone: Preventing Multi-Modal Large Language Models from Analyzing Images via Visual Prompt Injection

SafetyDGX agent

arXiv:2604.09024v1 Announce Type: cross Abstract: Multi-modal large language models (MLLMs) have emerged as powerful tools for analyzing Internet-scale image data, offering significant benefits but al

LEGO: Latent-space Exploration for Geometry-aware Optimization of Humanoid Kinematic Design

TutorialsDGX agent

arXiv:2604.08636v1 Announce Type: cross Abstract: Designing robot morphologies and kinematics has traditionally relied on human intuition, with little systematic foundation. Motion-design co-optimizat

Lessons Without Borders? Evaluating Cultural Alignment of LLMs Using Multilingual Story Moral Generation

Model ReleasesDGX agent

arXiv:2604.08797v1 Announce Type: cross Abstract: Stories are key to transmitting values across cultures, but their interpretation varies across linguistic and cultural contexts. Thus, we introduce mu

Listener-Rewarded Thinking in VLMs for Image Preferences

Model ReleasesDGX agent

arXiv:2506.22832v3 Announce Type: replace-cross Abstract: Training robust and generalizable reward models for human visual preferences is essential for aligning text-to-image and text-to-video generat

Litmus (Re)Agent: A Benchmark and Agentic System for Predictive Evaluation of Multilingual Models

Model ReleasesDGX agent

arXiv:2604.08970v1 Announce Type: cross Abstract: We study predictive multilingual evaluation: estimating how well a model will perform on a task in a target language when direct benchmark results are

LLM-Rosetta: A Hub-and-Spoke Intermediate Representation for Cross-Provider LLM API Translation

ApplicationsDGX agent

arXiv:2604.09360v1 Announce Type: cross Abstract: The rapid proliferation of Large Language Model (LLM) providers--each exposing proprietary API formats--has created a fragmented ecosystem where appli

LLM4Delay: Flight Delay Prediction via Cross-Modality Adaptation of Large Language Models and Aircraft Trajectory Representation

ResearchDGX agent

arXiv:2510.23636v3 Announce Type: replace-cross Abstract: Flight delay prediction has become a key focus in air traffic management (ATM), as delays reflect inefficiencies in the system. This paper pro

LLMs Underperform Graph-Based Parsers on Supervised Relation Extraction for Complex Graphs

ResearchDGX agent

arXiv:2604.08752v1 Announce Type: cross Abstract: Relation extraction represents a fundamental component in the process of creating knowledge graphs, among other applications. Large language models (L

LMGenDrive: Bridging Multimodal Understanding and Generative World Modeling for End-to-End Driving

SafetyDGX agent

arXiv:2604.08719v1 Announce Type: cross Abstract: Recent years have seen remarkable progress in autonomous driving, yet generalization to long-tail and open-world scenarios remains a major bottleneck

Many Preferences, Few Policies: Towards Scalable Language Model Personalization

SafetyDGX agent

arXiv:2604.04144v2 Announce Type: replace-cross Abstract: The holy grail of LLM personalization is a single LLM for each user, perfectly aligned with that user's preferences. However, maintaining a se

Many-Tier Instruction Hierarchy in LLM Agents

Model ReleasesDGX agent

arXiv:2604.09443v1 Announce Type: cross Abstract: Large language model agents receive instructions from many sources-system messages, user prompts, tool outputs, and more-each carrying different level

Mapping generative AI use in the human brain: divergent neural, academic, and mental health profiles of functional versus socio emotional AI use

Local AiDGX agent

arXiv:2604.08594v1 Announce Type: cross Abstract: The widespread adoption of generative artificial intelligence conversational agents (AICAs) among university students constitutes a novel cognitive so

MARINER: A 3E-Driven Benchmark for Fine-Grained Perception and Complex Reasoning in Open-Water Environments

Model ReleasesDGX agent

arXiv:2604.08615v1 Announce Type: cross Abstract: Fine-grained visual understanding and high-level reasoning in real-world open-water environments remain under-explored due to the lack of dedicated be

MedFormer-UR: Uncertainty-Routed Transformer for Medical Image Classification

Local AiDGX agent

arXiv:2604.08868v1 Announce Type: cross Abstract: To ensure safe clinical integration, deep learning models must provide more than just high accuracy; they require dependable uncertainty quantificatio

Medical Reasoning with Large Language Models: A Survey and MR-Bench

Model ReleasesDGX agent

arXiv:2604.08559v1 Announce Type: cross Abstract: Large language models (LLMs) have achieved strong performance on medical exam-style tasks, motivating growing interest in their deployment in real-wor

Memory-efficient Continual Learning with Prototypical Exemplar Condensation

Model ReleasesDGX agent

arXiv:2603.13804v2 Announce Type: replace-cross Abstract: Rehearsal-based continual learning (CL) mitigates catastrophic forgetting by maintaining a subset of samples from previous tasks for replay. E

Memory Intelligence Agent

AgentsDGX agent

arXiv:2604.04503v3 Announce Type: replace Abstract: Deep research agents (DRAs) integrate LLM reasoning with external tools. Memory systems enable DRAs to leverage historical experiences, which are es

Mind the Gap Between Spatial Reasoning and Acting! Step-by-Step Evaluation of Agents With Spatial-Gym

TutorialsDGX agent

arXiv:2604.09338v1 Announce Type: new Abstract: Spatial reasoning is central to navigation and robotics, yet measuring model capabilities on these tasks remains difficult. Existing benchmarks evaluate

Mitigating Extrinsic Gender Bias for Bangla Classification Tasks

Model ReleasesDGX agent

arXiv:2411.10636v2 Announce Type: replace-cross Abstract: In this study, we investigate extrinsic gender bias in Bangla pretrained language models, a largely underexplored area in low-resource languag

Model Space Reasoning as Search in Feedback Space for Planning Domain Generation

AgentsDGX agent

arXiv:2604.08712v1 Announce Type: new Abstract: The generation of planning domains from natural language descriptions remains an open problem even with the advent of large language models and reasonin

MolPaQ: Modular Quantum-Classical Patch Learning for Interpretable Molecular Generation

Model ReleasesDGX agent

arXiv:2604.08575v1 Announce Type: cross Abstract: Molecular generative models must jointly ensure validity, diversity, and property control, yet existing approaches typically trade off among these obj

MONETA: Multimodal Industry Classification through Geographic Information with Multi Agent Systems

Model ReleasesDGX agent

arXiv:2604.07956v2 Announce Type: replace Abstract: Industry classification schemes are integral parts of public and corporate databases as they classify businesses based on economic activity. Due to

Mosaic: Multimodal Jailbreak against Closed-Source VLMs via Multi-View Ensemble Optimization

SafetyDGX agent

arXiv:2604.09253v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) are powerful but remain vulnerable to multimodal jailbreak attacks. Existing attacks mainly rely on either explicit visu

Multi-agent Adaptive Mechanism Design

AgentsDGX agent

arXiv:2512.21794v3 Announce Type: replace-cross Abstract: We study a sequential mechanism design problem in which a principal seeks to elicit truthful reports from multiple rational agents while start

Multivariate Time Series Anomaly Detection via Dual-Branch Reconstruction and Autoregressive Flow-based Residual Density Estimation

Model ReleasesDGX agent

arXiv:2604.08582v1 Announce Type: cross Abstract: Multivariate Time Series Anomaly Detection (MTSAD) is critical for real-world monitoring scenarios such as industrial control and aerospace systems. M

MuTSE: A Human-in-the-Loop Multi-use Text Simplification Evaluator

SafetyDGX agent

arXiv:2604.08947v1 Announce Type: cross Abstract: As Large Language Models (LLMs) become increasingly prevalent in text simplification, systematically evaluating their outputs across diverse prompting

Neighbourhood Transformer: Switchable Attention for Monophily-Aware Graph Learning

ApplicationsDGX agent

arXiv:2604.08980v1 Announce Type: cross Abstract: Graph neural networks (GNNs) have been widely adopted in engineering applications such as social network analysis, chemical research and computer visi

Neural Distribution Prior for LiDAR Out-of-Distribution Detection

SafetyDGX agent

arXiv:2604.09232v1 Announce Type: cross Abstract: LiDAR-based perception is critical for autonomous driving due to its robustness to poor lighting and visibility conditions. Yet, current models operat

Neural networks for Text-to-Speech evaluation

Model ReleasesDGX agent

arXiv:2604.08562v1 Announce Type: cross Abstract: Ensuring that Text-to-Speech (TTS) systems deliver human-perceived quality at scale is a central challenge for modern speech technologies. Human subje

Neurons Speak in Ranges: Breaking Free from Discrete Neuronal Attribution

Local AiDGX agent

arXiv:2502.06809v3 Announce Type: replace-cross Abstract: Pervasive polysemanticity in large language models (LLMs) undermines discrete neuron-concept attribution, posing a significant challenge for m

Noise-Aware In-Context Learning for Hallucination Mitigation in ALLMs

Model ReleasesDGX agent

arXiv:2604.09021v1 Announce Type: cross Abstract: Auditory large language models (ALLMs) have demonstrated strong general capabilities in audio understanding and reasoning tasks. However, their reliab

NyayaMind- A Framework for Transparent Legal Reasoning and Judgment Prediction in the Indian Legal System

SafetyDGX agent

arXiv:2604.09069v1 Announce Type: cross Abstract: Court Judgment Prediction and Explanation (CJPE) aims to predict a judicial decision and provide a legally grounded explanation for a given case based

OmniPrism: Learning Disentangled Visual Concept for Image Generation

TutorialsDGX agent

arXiv:2412.12242v2 Announce Type: replace-cross Abstract: Creative visual concept generation often draws inspiration from specific concepts in a reference image to produce relevant outcomes. However,

On Divergence Measures for Training GFlowNets

SafetyDGX agent

arXiv:2410.09355v2 Announce Type: cross Abstract: Generative Flow Networks (GFlowNets) are amortized inference models designed to sample from unnormalized distributions over composable objects, with a

On Semiotic-Grounded Interpretive Evaluation of Generative Art

Model ReleasesDGX agent

arXiv:2604.08641v1 Announce Type: cross Abstract: Interpretation is essential to deciphering the language of art: audiences communicate with artists by recovering meaning from visual artifacts. Howeve

On-the-Fly Adaptation to Quantization: Configuration-Aware LoRA for Efficient Fine-Tuning of Quantized LLMs

Model ReleasesDGX agent

arXiv:2509.25214v3 Announce Type: replace-cross Abstract: As increasingly large pre-trained models are released, deploying them on edge devices for privacy-preserving applications requires effective c

← Previous
1…340341342343344…350
Next →